Staff Software Engineer Robinhood Command Center

Robinhood helps users invest in stocks, ETFs, options, and cryptocurrencies through commission-free trading with no minimum account requirements.

Series D0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/25/2026

United States
About Robinhood

Robinhood helps retail investors access financial markets through commission-free trading of stocks, ETFs, options, and cryptocurrencies. With it, users can invest with no account minimums, earn rewards through retirement accounts with matching contributions, and access advanced trading tools. Robinhood democratizes investing by making financial markets accessible to everyone, not just wealthy investors.

View jobs by Robinhood

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead incident mitigation and coordinate service owners during active incidents. You will develop incident-management processes, dashboards, alerts, observability frameworks, and reliability tooling. You will drive post-incident learning, report on service quality, and mentor engineers.

Requirements

  • 8+ years of software engineering experience operating production systems
  • 3+ years of reliability engineering, infrastructure, distributed systems, or production operations experience
  • Incident leadership experience
  • Communication and cross-functional collaboration skills during high-severity incidents
  • Knowledge of systems reliability, observability frameworks, and fault-tolerant architecture design
  • Experience with multi-region or multi-cluster architectures, capacity planning, and failover strategies
  • Familiarity with OpenTelemetry, Prometheus, and Grafana
  • Ability to improve MTTD, MTTR, availability, or customer impact

Responsibilities

  • Drive the long-term reliability and observability strategy across infrastructure
  • Partner with engineers to improve operational excellence and incident response
  • Lead incident mitigation by coordinating service owners and time-sensitive decisions
  • Develop and maintain incident-management processes and procedures
  • Define and maintain global dashboards and alerts for critical user journeys and business-impact metrics
  • Own and evolve incident-response tooling, education, adoption, and MTTD/MTTR measurement
  • Drive post-incident governance, postmortem standards, SEV reviews, and follow-up tracking
  • Design and implement failure-mitigation strategies
  • Build frameworks for monitoring, alerting, and observability
  • Own the roadmap for observability across critical user journeys
  • Deliver reliability insights and executive-level reporting
  • Mentor engineers and contribute to hiring and engineering culture

Benefits

  • Bonus opportunities
  • Equity ownership
  • 401(k) matching
  • Health insurance with 100% employee coverage and 90% dependent coverage
  • Lifestyle wallet for wellness and learning
  • Employer-paid life and disability insurance
  • Fertility benefits
  • Mental health benefits
  • Company holidays
  • Paid time off
  • Sick time
  • Parental leave
  • Catered meals
  • Office events
Staff Software Engineer Robinhood Command Center at Robinhood | JobStash