Site Reliability Engineer
Jump Trading is a global trading firm where traders, engineers, and researchers develop trading strategies, models, infrastructure, and systems across asset classes and time horizons.
About Jump Trading
Jump Trading is a global trading firm focused on research-driven trading and the engineering of scalable models, tools, infrastructure, and execution systems. Its operations combine trading, technology, AI/ML, and quantitative research, and it also runs research and talent programs including conference travel grants and a fellowship program.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will develop expertise in your assigned technology area, own production deployment and release processes, and improve performance, reliability, and operability. You will build production tooling, define observability metrics and SLOs, lead incident response and post-mortems, manage operational risk, document procedures, and mentor peers.
Requirements
- Degree in Computer Science or a related field, or equivalent professional experience
- At least 5+ years of relevant IT operations experience
- Expert-level proficiency in C++
- Linux operating system knowledge
- Knowledge of network and system configuration
- Knowledge of kernel internals, scheduling, and performance tuning
- Networking knowledge including routing, multicast, LLDP, VLANs, and Ethernet
- Ability to handle shared operational and periodic on-call duties
- Reliable and predictable availability
Responsibilities
- Develop technical expertise in the assigned product area
- Own production deployment, configuration, and release processes
- Build and maintain production tooling
- Define observability, SLI, SLO, and performance metrics
- Use metrics and capacity planning to support scalability and uptime
- Troubleshoot and resolve production incidents
- Lead incident response, root cause analysis, and post-mortems
- Align with global SRE teams on architecture and best practices
- Document processes and procedures
- Provide mentorship and cross-training
- Manage operational risk for production changes
