Senior Technical Program Manager Infrastructure
Glean is an enterprise AI platform that connects company knowledge and systems to provide permission-aware search, AI assistance, and agents.
Funding history
About Glean
Glean develops enterprise Work AI software. Its platform provides enterprise search, a conversational AI assistant, and tools to build, govern, and orchestrate AI agents using a company’s connected data and permissions.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will lead large-scale infrastructure programs across compute, networking, storage, orchestration, and AI workloads. You will coordinate engineering delivery, establish reliability and cost metrics, improve deployment automation and runtime health, and communicate technical risks to stakeholders.
Requirements
- BS/MS in Computer Science, Engineering, or a related technical field
- 8–10+ years of experience in technical program management, infrastructure, or SRE
- 3–5 years managing infrastructure or platform-scale programs
- Experience delivering cross-functional infrastructure programs in enterprise environments
- Experience working with Infrastructure, SRE, and ML/AI teams on distributed systems or data infrastructure
- Understanding of cloud infrastructure, including compute, networking, storage, and orchestration
- Ability to structure multi-quarter infrastructure programs with measurable impact
- Written and verbal communication skills
Responsibilities
- Drive the infrastructure roadmap across deployment, runtime, storage, and AI infrastructure
- Lead programs that improve scalability, reliability, cost efficiency, and developer velocity
- Define deployment, upgrade, and monitoring orchestration at scale
- Partner with AI and Data teams on ML pipelines, training infrastructure, and LLM serving
- Improve observability, configuration management, and resource utilization
- Coordinate capacity planning, infrastructure migrations, and performance optimization
- Develop frameworks for runtime health, scaling, and disaster recovery
- Establish reliability, performance, and cost-efficiency metrics
- Drive automation across deployment orchestration systems
- Communicate program status and technical risks to stakeholders
Benefits
- Medical coverage
- Vision coverage
- Dental coverage
- Generous time-off policy
- 401k plan
- Home office improvement stipend
- Annual education stipend
- Annual wellness stipend
- Regular events
- Daily healthy lunches
Hiring Process
Brief AI-focused exercise or discussion.
