Staff Software Engineer, AI/ML Infrastructure, TPU Supercomputers

Google

Sunnyvale, CA, USA

Description

Minimum qualifications:

  • Bachelor's degree or equivalent practical experience.
  • 8 years of software development experience in C, C++, Go, or Python.
  • 5 years of experience testing, and launching software products.
  • 5 years of experience building and developing large-scale infrastructure, distributed systems or networks, or experience with compute technologies, storage, or hardware architecture.
  • 3 years of experience designing, building, and operating large-scale distributed systems, high-performance networking stacks, or operating system internals.

Preferred qualifications:

  • Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
  • 8 years of experience with data structures/algorithms.
  • 3 years of experience in a technical leadership role leading project teams and setting technical direction.
  • 3 years of experience working in a complex, matrixed organization involving cross-functional, or cross-business projects.
  • Experience with lower-half server architectures, hardware-adjacent orchestration, and low-level security implementations.
  • Experience with Kubernetes and Google-internal cluster systems, alongside a proven ability to build telemetry pipelines and monitoring systems for distributed hardware.