Staff Software Engineer - Deployment Platform
About the Role and Team
Uber's Deployment Engine team is the backbone of how the world's largest mobility and delivery platform safely ships innovation at scale, impacting over 150 million consumers. Our deployment platform enforces policies and gathers real-time signals to ensure that every change to Uber's production systems rolls out incrementally and automatically reverts if those signals hint at trouble.
Our mission is simple: make production deployments stress-free for every engineer at Uber. We achieve this by building intelligent guardrails that continuously evolve alongside new workload types, deployment patterns, and scale demands. Orchestrating changes safely across ~1M rollout operations per week is a complex problem we are deeply passionate about solving. The stakes are high: our systems directly power experiences for 150M+ monthly active platform consumers, 7M+ active drivers and couriers, and services spanning 70+ countries and 10,000+ cities globally.
Our platform makes high-precision, data-driven decisions to detect potential incidents and regressions early, dramatically reducing outage risks. Key capabilities include:
- Automatic Rollbacks: Real-time reversal of problematic deployments based on active system signals.
- Emergency Policies & Lockdowns: Automated and manual policy enforcement during active incidents.
- Incident Mitigation Tooling: Safeguards that ensure Uber's global services remain available and reliable.
If you are energized by tackling distributed systems edge cases, leveraging cutting-edge infrastructure, and solving high-stakes challenges that keep a global business running 24/7, you will thrive here. You will work on a critical platform that enables Uber to innovate rapidly without compromising safety, reliability, or developer velocity.
Impact & Scope
The Deployment Engine orchestrates the safe rollout of changes across Uber's entire production fleet. This role carries broad, high-visibility ownership across every corner of Uber's technology stack:
- 5,000+ stateless microservices, deployed ~1M times per week.
- 40+ stateful storage technologies (MySQL, Schemaless, Cassandra, etc.) spanning 25 storage systems, 3.8M containers, and 100,000+ hosts.
- Open-source Kubernetes operators, configuration management systems, and every other change type powering Uber's global product suite.
Reliability is not just a nice-to-have—it is a first-class requirement. Operational excellence, developer velocity, and global system reliability all flow directly through the platforms we build.
We are looking for a proven, large-scale systems thinker who excels in high-stakes environments. You will navigate the messy reality of complex infrastructure spanning global data centers and multi-cloud environments while supporting a diverse range of workloads. You must be comfortable making sound decisions with imperfect information and taking ownership of global architectural outcomes. If you lead with clarity, build with heart, and are motivated by the challenge of peak efficiency at an unparalleled scale, this is where you will make your mark.
What You’ll Do
- Architect Safe Rollout Systems: Design resilient systems that distribute, deploy, observe, and recover changes at a massive global scale, ensuring predictable and safe behavior even during underlying infrastructure outages.
- Empower Internal Engineering: Serve as a foundational layer for hundreds of engineers. Collaborate closely with platform teams to build reliable, safe, and efficient deployment pipelines that adapt to evolving business and scale requirements.
- Provide Technical Leadership & Mentorship: Offer deep technical mentorship across the stack to ensure architectures and implementations are robust, extensible, and secure.
- Set Engineering Culture: Review code and designs while fostering a culture of continuous learning, humility, and high quality. Guide the team in leveraging AI-first development tools without sacrificing code rigor or reliability.
Basic Qualifications
- 10+ years of professional experience building scalable, fault-tolerant products or platforms, with a proven track record of solving strategically critical problems.
- Polyglot Engineering Proficiency: Deep expertise in one or more core programming languages (e.g., Go, C/C++, Java, Python) and a demonstrated ability to define organization-wide technical standards.
- Lifecycle Ownership: Proven experience leading large-scale engineering initiatives from inception through to global production rollout.
- Strategic Trade-off Analysis: Expertise in evaluating complex engineering and business trade-offs, such as "buy vs. build" or "incremental evolution vs. complete rewrite."
- Architectural Leadership: Demonstrated skill in identifying systemic architectural gaps and serving as a force multiplier to eliminate technical debt across organizations.
- Education: Bachelor’s degree in Computer Science, Computer Engineering, Mathematics, or equivalent practical experience.
Preferred Qualifications
- Advanced degree (Master’s or PhD) focused on distributed computing, system performance, or operating systems.
- Hands-on experience with Linux internals, container orchestration (Kubernetes internals/operators), or modern cloud-native security architecture patterns.
- Active contributions to open-source software (OSS) projects or a history of leading industry-level technical discussions.
- Proven ability to anticipate long-term business needs and design extensible, modular infrastructure capable of handling future scale targets.
Perks & Benefits:
- Monthly Uber Credits: Credits to use on Uber Rides and Uber Eats every month.
- Equity Compensation: Opportunity to be awarded stock options (RSUs) to ensure you own a piece of the mission you’re building.
- Culture & Socials: Frequent local social events and office clubs (chess, board games, running, crossfit, creative club, and more).
- Tech Community: We host regular local tech meetups to stay connected with the Aarhus engineering scene, and give our engineers the opportunity to sometimes share what they’re working on with the community.
- Well-being & Fertility: Global support programs for mental health, wellness, and family planning/fertility.
- Parental Leave: Generous, gender-neutral parental leave to support your life outside of work.
- Modern Aarhus Hub: Work in a center of technical excellence featuring catered lunches and top-tier collaboration spaces.