Staff Infrastructure Engineer
Figure is an AI robotics company developing general-purpose humanoid robots and its Helix vision-language-action AI system.
Funding history
About Figure
Figure develops and deploys general-purpose humanoid robots for commercial and household tasks. Its current Figure 03 robot is powered by Helix, an onboard generalist vision-language-action model; the company also operates the Index data-collection service.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own internal systems infrastructure and build highly available, reliable, automated cloud and on-premises systems. You will support critical operational services, migrate suitable SaaS tools to self-hosted solutions, automate deployment and scaling, implement monitoring and incident-response practices, define service objectives, and apply security updates.
Requirements
- Linux/Unix systems administration
- Programming and scripting
- Cloud platforms including Azure, AWS, and GCP
- On-premises hardware architectures
- High-availability, fault-tolerant, and distributed systems
- Infrastructure as code including Terraform, CloudFormation, and Ansible
- Monitoring, logging, and alerting tools including Prometheus, Grafana, and Datadog
- Networking fundamentals including TCP/IP, DNS, HTTP, load balancers, and firewalls
- Service-level objectives, runbooks, incident-response plans, post-mortems, and systems asset management
- Cross-functional collaboration
- Verbal and written communication
Responsibilities
- Manage mission-critical infrastructure supporting source configuration management, CI/CD, software distribution, supplier portals, and manufacturing
- Migrate SaaS tools to self-hosted solutions
- Implement monitoring and alerting systems and define incident-response plans and runbooks
- Automate deployment and scaling to reduce manual work
- Identify infrastructure needs and establish service-level objectives with stakeholders
- Track service robustness and optimization work using data
- Apply security remediations and updates with the security team
