Posted today · Greenhouse · cloudflare✓ Direct employer / ATS application

Hardware Systems Engineer

Cloudflare

otherremoteIn-OfficeSource verified
Role details

What you’ll be doing

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations: Bengaluru About the department Cloudflare’s Infrastructure group is responsible for building our global network. Our Hardware Engineering team helps research, develop, test, and deploy new equipment enabling 20% of the world’s internet traffic to be served smoothly. Deployed across 330 cities in 120+ countries, the hardware we select helps improve the security, reliability, and performance of the Internet. About the Role We need to make thoughtful infrastructure choices affecting a significant portion of the Internet. Hardware we work with includes servers and components, as well as PDUs and network hardware. . As a Hardware Systems Engineer, you will work with colleagues on the Hardware Engineering, Product teams, and Hardware Sourcing teams to troubleshoot and maintain Cloudflare’s worldwide fleet of storage and compute servers. What you'll do Exp: 3-5yrs Work with software teams to validate bug fixes and assess performance of new firmware revisions Validate and deploy firmware updates to the fleet, monitoring the progress of the rollout for compliance and reliability Work with server and component vendors to obtain, debug, and maintain the latest updates Work with our Site Reliability Engineering teams to triage hardware problem reports Support our Data Centre Engineering teams in resolving hardware issues Develop and maintain automation tools to update firmware on servers and components in Cloudflare’s fleet Communicate your results and updates through blog posts, internal talks, and tickets Examples of desirable skills, knowledge and experience Bachelor’s degree in Computer Engineering, Electrical Engineering, or Computer Science Desire to learn about the Cloudflare hardware used by 20% of all web sites Desire to learn how a diverse server fleet is managed at scale Desire to learn the tools Cloudflare uses to maintain and monitor our hardware Knowledge of bash and python and basic Linux task automation Knowledge of x86 server hardware including motherboards, CPUs, memory, storage and firmware updates. Knowledge of other platforms such as arm is a bonus. Knowledge of configuration management principals, in particular we use salt to manage our fleet Knowledge of Redfish, IPMI and server remote management protocols Knowledge of running production mission critical systems Bonus Points Familiarity with server hardware architecture Knowledge of debugging server hardware faults and the ability to engage with our sourcing team and vendors to improve quality Experience of managing large fleets comprising of thousands of servers Experience of observability and monitoring tools such as Prometheus and Grafana, and the ability to observe trends over time Experience with software development tools and processes such as git, Bitbucket and TeamCity and Jira Fraud Alert: Do not fall victim to recruitment fraud. Cloudflare never charges application fees or requires candidates to purchase third-pa