Supercomputing Intern

Etched is an AI-hardware company building rack-scale frontier inference clusters.

San Jose, United States
About Etched

Etched co-designs chips, racks, software, and manufacturing systems for efficient inference of frontier AI models, targeting throughput, latency, cost, and power efficiency across prefill and decode workloads.

View jobs by Etched

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will contribute to the design, development, and deployment of ML system software for rack-scale systems. Your work will cover network performance, telemetry pipelines, system health and performance analysis, software deployment and provisioning, hardware validation, and secure data-center-scale ML workloads.

Requirements

  • Computer science
  • Engineering
  • C
  • C++
  • Rust
  • Python
  • Data structure
  • Algorithm
  • Low-level software engineering
  • Hardware/software co-design
  • Communication
  • Collaboration

Responsibilities

  • Design, develop, and deploy ML system software for rack-scale systems
  • Analyze network performance and system health
  • Create and process telemetry pipelines
  • Deploy and provision software frameworks
  • Validate hardware
  • Maintain secure and performant systems for data-center-scale ML workloads

Benefits

  • 12-week paid internship
  • Housing support for relocating employees
  • Daily office lunch and dinner
  • Direct mentorship
Supercomputing Intern at Etched | JobStash