Resource Optimization and Fleet Logic Technical Lead

Socket.dev

Sunnyvale (CA)

On-site

USD 262,000 - 364,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Google in Sunnyvale, CA is seeking a senior software engineer to help scale AI infrastructure across Cloud and enterprise ecosystems. You will design and implement large-scale distributed systems, drive architecture decisions, and mentor teams while coordinating with product and cross-functional partners.

Requirements include strong C++ expertise, 8+ years in software development, and experience integrating AI tools into complex workflows. Expect a fast-paced, highly impactful role.

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience programming in C++.
  • 5 years of experience with design and architecture; and testing/launching software products.
  • Experience integrating generative AI tools or LLM interfaces into workflows.

Responsibilities

  • Set the technical goals and strategy for delivering the above requirements and lead a multi-year technical roadmap, balancing short- and long-term technology investments.
  • Shape the technical strategy and goal for optimal placement and reconfiguration of data center resources including machines, racks, space, power, network, cooling to ensure industry leading utilization efficiency.
  • Drive the roadmap and partner with Product Management on the definition and development of new features.
  • Engage with partner teams to align on product requirements, solution architecture, and emerging requirements and issues.
  • Help teams find the right short-term and long-term balance between development velocity and generalization of software solutions. Uphold a culture of respect, innovation, growth, engineering excellence, landings, and customer focus.

Skills

C++ programming
System design
AI integration

Education

Bachelor's degree or equivalent practical experience

Job description

Minimum qualifications:
  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience programming in C++.
  • 5 years of experience with design and architecture; and testing/launching software products.
  • Experience integrating generative AI tools or LLM interfaces into workflows.
Preferred qualifications:
  • Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
  • 12 years of professional software development experience or 8 years of experience with an advanced degree.
  • 5 years of experience in a technical leadership role leading project teams and setting technical direction.
  • 3 years of experience working in a complex, matrixed organization involving cross-functional, or cross-business projects.
  • Experience in distributed systems architecture and building scalable systems for an enterprise business and with capacity planning or networking engineering.
About the job:

Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to Google’s needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.

In this role, you will focus on solving a host of issues as we expand our product suite to support the full portfolio of Cloud products, develop optimization solutions for Google-scale capacity curation, and develop sublinear-scale automation for deploying Google’s ML hardware. You will help us deliver a customer-focused software ecosystem that outpaces the growth of Google.

You will lead the architecture of Google’s Fleet planning and optimization, establishing the roadmap, and leading the execution and delivery of the ecosystem that underpins Google’s internal/external Cloud. You will deliver a customer-focused, resource ecosystem that supports the growth of Google. We are a team, working on some of Google infrastructure's critical programs.

The AI and Infrastructure team is redefining what’s possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.

We're the driving force behind Google's groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $262000 - $364000 (USD) + 25% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities:
  • Set the technical goals and strategy for delivering the above requirements and lead a multi-year technical roadmap, balancing short- and long-term technology investments.
  • Shape the technical strategy and goal for optimal placement and reconfiguration of data center resources including machines, racks, space, power, network, cooling to ensure industry leading utilization efficiency.
  • Drive the roadmap and partner with Product Management on the definition and development of new features.
  • Engage with partner teams to align on product requirements, solution architecture, and emerging requirements and issues.
  • Help teams find the right short-term and long-term balance between development velocity and generalization of software solutions. Uphold a culture of respect, innovation, growth, engineering excellence, landings, and customer focus.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Resource Optimization and Fleet Logic Technical Lead
Resource Optimization and Fleet Logic Technical Lead

Google • Sunnyvale (CA)

On-site
USD 262,000 - 365,000
Equity grants
Bonus target
Staff Software Engineer, GCE Fleet Deployment Platform
Staff Software Engineer, GCE Fleet Deployment Platform

Socket.dev • Seattle (WA)

On-site
USD 207,000 - 300,000
Bonus target (20%)
Equity
Benefits
Software Engineer III, Infrastructure, Google Cloud Compute
Software Engineer III, Infrastructure, Google Cloud Compute

Socket.dev • Kirkland (WA)

On-site
USD 147,000 - 210,000
Technical Lead Manager, PIE Core Systems and Capacity
Technical Lead Manager, PIE Core Systems and Capacity

Socket.dev • Austin (TX)

On-site
USD 207,000 - 300,000
Staff Software Engineer, Data Center Resource Modeling
Staff Software Engineer, Data Center Resource Modeling

Google Inc. • Sunnyvale (CA)

On-site
USD 210,000 - 300,000
Tech Lead, TPU AI Infrastructure
Tech Lead, TPU AI Infrastructure

Google Inc. • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Senior Staff Software Engineer, AI/ML, Google Cloud
Senior Staff Software Engineer, AI/ML, Google Cloud

Socket.dev • Seattle (WA)

On-site
USD 262,000 - 364,000
Equity
Bonus target 25%
Staff Software Engineer, Network Health
Staff Software Engineer, Network Health

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Software Engineer, Google Cloud Platform, Fault Management
Software Engineer, Google Cloud Platform, Fault Management

Google • Sunnyvale (CA)

On-site
USD 180,000 - 300,000
Technical Program Manager, ML Fleet Capacity, Systems Enablement
Technical Program Manager, ML Fleet Capacity, Systems Enablement

Socket.dev • Kirkland (WA)

On-site
USD 199,000 - 270,000