
Dropbox
Staff Site Reliability Engineer, Production Engineering
Remote (Worldwide)RemotePosted Today$223,400 – $302,200
Full TimeSeniorRemote
See how this job matches your profile
Sign in for an AI-powered fit score, breakdown, and a tailored resume.
Job Description
Role DescriptionAs a Site Reliability Engineer focused on company-wide reliability strategy, you will play a crucial role in advancing Dropbox’s stability, observability, incident response, and operat
Key Highlights
- Define and evolve Dropbox’s company-wide technical reliability strategy to support the changing engineering environment created by AI-assisted and agentic software development.
- Set multi-year reliability goals, standards, and roadmaps across observability, debugging, incident management, service health, and operational readiness.
- Lead cross-team initiatives that reduce reliability risk as software delivery velocity, pull request volume, service complexity, and incident volume increase.
- Partner with engineering leaders and platform teams to improve monitoring, alerting, debugging, SLOs, SLAs, and incident response systems at company scale.
- Identify emerging reliability risks introduced by AI-enabled development workflows and design scalable systems, processes, and guardrails to mitigate them.
Qualifications
Required Qualifications
- BS degree in Computer Science or related technical field involving coding(e.g., physics or mathematics), or equivalent technical experience.
- 12+ years of experience in software engineering, site reliability engineering, infrastructure engineering, or related technical roles.
- Proven ability to define and deliver multi-year, multi-team reliability, infrastructure, or platform strategies with measurable business and customer impact.
- Deep experience with distributed systems, production operations, observability, incident response, SLOs/SLAs, debugging, and reliability risk management.
- Demonstrated ability to diagnose complex technical problems, debug production systems, automate operational workflows, and design resilient software components.
- Experience influencing engineering roadmaps across multiple teams and making technical decisions that optimize for the broader engineering organization.
- Strong communication and collaboration skills, with the ability to align cross-functional stakeholders through ambiguity and drive execution across teams.
- Experience adapting reliability strategies, developer tooling, or operational processes for AI-assisted software development workflows.
- Experience building or scaling observability, debugging, incident management, or developer productivity platforms for large engineering organizations.
- Experience leading reliability improvements in environments with high deployment velocity, complex service dependencies, and large-scale production systems.
- Track record of mentoring senior engineers, setting technical standards, and spreading reliability best practices through documentation, reviews, talks, or architecture guidance.
- Familiarity with AI-enabled tooling, agentic development workflows, or operational risks introduced by rapid automation in the software development lifecycle.
About the Company
Dropbox
View company profile →
Interested in this role?
Sign in or create a free account to see how this job matches your skills, apply with one click, and let our AI tailor your resume.
Sign in to applyAI-powered resume optimization
Save and track your applications
Job Details
Employment Type
Full Time
Experience Level
Senior
Salary Range
$223,400 – $302,200
Location
Remote (Worldwide)
Work Mode
Remote
Posted
Today