Staff Software Engineer, Data Center and AI Infrastructure
Our team's mission is to deliver high-performance, resilient switch software powering Google's data center and AI infrastructure.
Our team is responsible for developing and maintaining the switch operation system deployed in every Google data center. You will be directly contributing in supporting the growth of Gemini and Google cloud business. You will also be engaging in the technology in hardware (switching ASIC and optics) and software to shape the future of the global network.
In this role, you will be building software for switches that power the world's largest AI infrastructure, from data center fabrics, to wide area networks, to the peering edge of our infrastructure.
The AI and Infrastructure team is redefining what’s possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.
We're the driving force behind Google's groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.
Individual pay is determined by factors including job-related skills, experience, and relevant education or training.US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits
Learn more about benefits at Google.
Responsibilities
- Provide technical leadership on projects.
- Influence and coach a distributed team of engineers.
- Scale and performance tuning of a large-scale network.
- Design, develop, test, deploy, maintain and improve switch software.
- Work on projects enabling AI networking infrastructure, enabling high performing fabrics for Tensor Processing Units (TPUs) and Graphics Processing Units (GPUs).
Minimum qualifications:
- Bachelor's degree or equivalent practical experience.
- 8 years of experience programming in C++.
- 5 years of experience testing, and launching software products.
- 5 years of experience building and developing large-scale infrastructure, distributed systems or networks, or experience with compute technologies, storage, or hardware architecture.
- Experience in problem management, problem-solving, network architecture, and network design.
Preferred qualifications:
- Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
- 8 years of experience with data structures and algorithms.
- Knowledge of Linux user space development and multi-threading development.
- Knowledge of networking protocol, embedded software development and Software Defined Networking (SDN).