Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 3
AWS
Azure
Communication @ 3
Distributed Systems @ 6
GCP
Kubernetes
Observability @ 3
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Glean is seeking a Software Engineer, Compute Infrastructure to help design, build, and operate the core compute and runtime platform powering AI search, assistant, and agentic workloads. Within the Platforms organization, this role focuses on Kubernetes-based runtime systems, multi-cloud infrastructure, and cost-efficient, low-latency execution for production services and pipelines serving customers at scale.
Responsibilities
- Design, build, and own backend and platform services focused on reliability, scalability, and performance for AI and search workloads.
- Develop and evolve Kubernetes-based runtime primitives, including service orchestration, scheduling integrations, and autoscaling patterns, across GCP, AWS, and Azure.
- Collaborate with platform, data, and product engineering teams to create safe and efficient deployment, configuration, and runtime-operation paths for services and batch workloads.
- Improve latency, resource utilization, and cost for core platform services, including multitenant runtime environments and experimental AI workloads.
- Implement and harden infrastructure-as-code patterns, observability, and operational guardrails, including SLOs, dashboards, alerts, and safe rollout and rollback procedures.
- Partner with the Costs and Runtime teams on attribution, guardrails, and automation mechanisms to improve runtime efficiency as customer and traffic volumes grow.
- Participate in an on-call rotation for critical platform services, lead incident response when needed, and improve reliability, tooling, and documentation based on operational learnings.
- Contribute to technical direction and roadmaps covering multitenancy, autoscaling, capacity and placement, and platformized patterns.
Requirements
- Experience as a backend or platform engineer working across application behavior, infrastructure, and cost optimization.
- Strong distributed systems fundamentals and experience operating high-throughput, low-latency services or batch pipelines in production.
- Ability to own systems end-to-end, including design, implementation, testing, deployment, observability, and ongoing operations.
- Experience with reliability practices such as SLOs, incident response, safe deployment strategies, and operational runbooks.
- Pragmatic, execution-oriented approach suited to a fast-moving startup environment.
- Clear communication and effective collaboration with infrastructure and product engineering teams.
- Interest in multi-cloud, multi-tenant environments and running AI workloads efficiently at scale.
The interview process includes a brief AI-focused exercise or discussion covering how candidates think about, design, and use AI. Prior Glean experience is not required.
Location
- Hybrid schedule requiring four days per week in the Mountain View office.
Compensation & Benefits
- Standard base salary: $140,000–$220,000 annually.
- Certain roles may be eligible for variable compensation, equity, and benefits.
- Medical, vision, and dental coverage.
- Generous time-off policy and 401(k) contribution opportunity.
- Home office improvement stipend.
- Annual education and wellness stipends.
- Regular company events and healthy lunches daily.
- Inclusive workplace committed to diversity and non-discrimination.
More jobs at Glean
Machine Learning Engineer, Assistant Quality
Glean · San Francisco, United States
USD 180,000-205,000 per year
Resident Solutions Architect
Glean · United States
USD 170,000-240,000 per year
Software Engineer
Glean · Mountain View, United States
USD 215,000-278,900 per year
Software Engineer
Glean · Mountain View, United States
USD 187,700-234,000 per year
Software Engineer, Cloud Deployment Infrastructure
Glean · San Francisco, United States
USD 200,000-270,000 per year
Similar jobs
Technical Program Manager, Infrastructure
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States
USD 290,000-365,000 per year
Forward Deployed Engineer - Physical AI Cloud Platform
Nebius · United States, Austin, United States
USD 179,500-224,300 per year
Senior Software Engineer, DGX Cloud Orchestration
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Full-Stack Lead Engineer
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Software Engineer
SentinelOne · United States
USD 132,000-182,000 per year
Senior Software Engineer, Attestation Services – DGX Cloud
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Software Engineer - HPC
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year