Prepin
Log in
Google

engineering opportunity

Staff Site Reliability Engineer, AI Foundations, F1 Query

The role involves designing and implementing solutions to enhance the reliability and scalability of the F1 Query distributed system. You will collaborate across teams to drive reliability improvements while engaging in software engineering tasks using Java, C++, and Go.

San Jose, California, United StatesonsiteFULL_TIME

Posted

About the role

What will you do at Google?

MINIMUM QUALIFICATIONS:

* Bachelor's degree in Computer Science, a related technical field, or

equivalent practical experience.

* 8 years of experience building and developing infrastructure or distributed

systems.

* 5 years of experience programming in C++ or Go.

* 5 years of experience with reliability or quality engineering.

* 5 years of experience working in distributed systems and distributed

computing.

* Experience with cross-functional collaboration and stakeholder management.

PREFERRED QUALIFICATIONS:

* Master's degree in Computer Science or a related technical field.

* Experience in a Site Reliability Engineering role.

* Experience in designing, analyzing and troubleshooting large-scale

distributed systems.

* Experience with database systems, query language design, and optimization.

* Ability to create clean and scalable designs.

* Ability to apply a systematic problem-solving approach, with excellent

communication skills and a sense of ownership and motivation.

ABOUT THE JOB:

Site Reliability Engineering (SRE) combines software and systems engineering to

build and run large-scale, massively distributed, fault-tolerant systems. SRE

ensures that Google Cloud's services—both our internally critical and our

externally-visible systems—have reliability, uptime appropriate to customer's

needs and a fast rate of improvement. Additionally SRE’s will keep an

ever-watchful eye on our systems capacity and performance.

Much of our software development focuses on optimizing existing systems,

building infrastructure and eliminating work through automation. On the SRE

team, you’ll have the opportunity to manage the complex challenges of scale

which are unique to Google Cloud, while using your expertise in coding,

algorithms, complexity analysis and large-scale system design. SRE's culture of

intellectual curiosity, problem solving and openness is key to its success. Our

organization brings together people with a wide variety of backgrounds,

experiences and perspectives. We encourage them to collaborate, think big and

take risks in a blame-free environment. We promote self-direction to work on

meaningful projects, while we also strive to create an environment that provides

the support and mentorship needed to learn and grow.

F1 Query is a team under AI Foundations organization. F1 Query is a planet-scale

high-performance distributed GoogleSQL-compliant federated query engine. It

supports querying nearly any type of data source in common use at Google. F1

Query powers over 400 production systems across Ads, Finance, Play, Cloud,

YouTube, Google DeepMind and Enterprise AI, and helps users address various

ad-hoc and low-latency data processing, serving, investigative, and batch use

cases.

Behind everything our users see online is the architecture built by the

Technical Infrastructure team to keep it running. From developing and

maintaining our data centers to building the next generation of Google

platforms, we make Google's product portfolio possible. We're proud to be our

engineers' engineers and love voiding warranties by taking things apart so we

can rebuild them. We keep our networks up and running, ensuring our users have

the best and fastest experience possible.Individual pay is determined by factors

including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google

[https://www.google.com/about/careers/applications/benefits/].

RESPONSIBILITIES:

* Identify opportunities, and design and lead the implementation of solutions

to enhance the reliability of systems that support F1.

* Scale systems sustainably through mechanisms like automation, and evolve

systems by pushing for changes that improve reliability and velocity.

* Collaborate across multiple teams to provide in-team leadership for Product

Area and to drive the adoption of solutions that improve reliability for

multiple teams.

* Engage in software engineering on services written in Java, C++, and Go

(including instrumenting client software development kita (SDKs), server-side

performance changes, capacity planning, experiments, monitoring, and more),

and performance enhancement of existing systems.

* Advise and influence our developer partners on next-generation architecture

and implementation for new systems.

Which skills does this role require?

C++GoJavaInfrastructure DesignCapacity PlanningPerformance OptimizationSystem ArchitectureStakeholder ManagementCross-functional CollaborationF1 QueryPerformance EnhancementGoogleSQLFederated Query EngineData CentersSDKsMonitoringScalabilitySoftware EngineeringTechnical InfrastructureQuery OptimizationGCP

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.