Senior SRE, Compute Orchestration

Posted Yesterday
Be an Early Applicant
San Mateo, CA
Hybrid
234K-284K Annually
Senior level
Computer Vision • Gaming • Software • Virtual Reality • Web3
Roblox is an immersive platform for connection & communication.
The Role
As a Senior Site Reliability Engineer, you will build and support the infrastructure for Roblox's private cloud, focusing on orchestration systems, service discovery, and performance monitoring. You will automate processes, create fault-tolerant systems, and analyze system designs to ensure reliability and production readiness.
Summary Generated by Built In

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. 

At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. 

A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone.

What You’ll Do:

As a Site Reliability Engineer (SRE) on the Infra Compute Orchestration (ICO) team, you will create, support, and evolve the infrastructure at Roblox as we build out Roblox's private cloud. ICO's mission is to own and manage our underlying orchestration systems along with elements of service discovery, secrets management and related software layers. 

You Will:

  • Create systems & libraries that promote fault-tolerance and resilience– like retries, circuit breakers, and adaptive concurrency limits.
  • Build, automate and standardize process automation to create a "golden path" of tooling and platform support that powers the fundamental Roblox ecosystem.
  • Create tooling that provides production guardrails, for example evaluating release candidate capacity with load testing tooling before deploying to production.
  • Create performance monitoring services and observability towards understanding capacity issues and platform degradations.
  • Create tooling that monitors production services and their changes, like generalized canarying services with alerting.
  • Analyze systems and system designs for production readiness

You Have:

  • Experience: you have a BS degree (or equivalent professional experience) in Computer Science or related engineering field with proven track record including at least 6 years as an SRE or Software Engineer.
  • Passion for systems: You have experience and good habits around building software and tools and getting them adopted. Your system's focus advises a view of code needing to be deeply reliable.

You Are:

  • A Partner: You know that the best tools integrate broadly with the tooling ecosystem. You approach partners and processes with curiosity and seek to understand a problem deeply before you start coding.
  • A Coder: you have experience writing common programming languages (e.g., Go, Java, C#, Rust).
  • Passionate about problem-solving, finding creative work solutions, and addressing unexpected challenges as part of a team.
  • Problem Solver: you ask the right questions to tackle issues within your expertise and you use data to test your theories.
  • Planner: You have experience in large project lifecycles. You have experience working in sprints, breaking down complex tasks into achievements, and reporting status to keep project scheduling accurate.

For roles that are based at our headquarters in San Mateo, CA: The starting base pay for this position is as shown below. The actual base pay is dependent upon a variety of job-related factors such as professional background, training, work experience, location, business needs and market demand. Therefore, in some circumstances, the actual salary could fall outside of this expected range. This pay range is subject to change and may be modified in the future. All full-time employees are also eligible for equity compensation and for benefits.

Annual Salary Range

$233,840$283,780 USD

Roles that are based in our San Mateo, CA Headquarters are in-office Tuesday, Wednesday, and Thursday, with optional in-office on Monday and Friday (unless otherwise noted).

You’ll Love: 

  • Industry-leading compensation package
  • Excellent medical, dental, and vision coverage
  • A rewarding 401k program
  • Flexible vacation policy (varies by exemption status)
  • Roflex - Flexible and supportive work policy 
  • Roblox Admin badge for your avatar
  • At Roblox HQ: 
    • Free catered lunches five times a week and several fully stocked kitchens with unlimited snacks
    • Onsite fitness center and fitness program credit
    • Annual CalTrain Go Pass

Roblox provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. Roblox also provides reasonable accommodations for all candidates during the interview process.

Top Skills

C#
Go
Java
Rust

What the Team is Saying

Claus
Ying
Denise
Daniel
Andrea
The Company
San Mateo, CA
2,500 Employees
Hybrid Workplace
Year Founded: 2004

What We Do

Roblox is reimagining the way people come together by enabling them to create, connect, and express themselves in immersive 3D experiences built by a global community. Our mission is to connect billions of users with optimism and civility, which starts by fostering a safe and inclusive environment—one that inspires creativity and empowers positive relationships between people around the world.

Why Work With Us

At Roblox, we foster a culture of innovation and continuous learning. We are building the tools and technology that empower our global community of millions of creators to bring to life any experience they can imagine.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

Roblox Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

We are requiring employees to be in the office three days a week – with core days being Tuesday, Wednesday, and Thursday. On Mondays and Fridays, employees may choose to work remotely, although the office will still be open.

Typical time on-site: 3 days a week
San Mateo, CA

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account