Software Engineer, Compute Infrastructure

at Glean
USD 140,000-220,000 per year
MIDDLE
✅ Hybrid

Tech Stack

AI @ 3 AWS Azure Communication @ 3 Distributed Systems @ 6 GCP Kubernetes Observability @ 3

Details

Glean is seeking a Software Engineer, Compute Infrastructure to help design, build, and operate the core compute and runtime platform powering AI search, assistant, and agentic workloads. Within the Platforms organization, this role focuses on Kubernetes-based runtime systems, multi-cloud infrastructure, and cost-efficient, low-latency execution for production services and pipelines serving customers at scale.

Responsibilities

  • Design, build, and own backend and platform services focused on reliability, scalability, and performance for AI and search workloads.
  • Develop and evolve Kubernetes-based runtime primitives, including service orchestration, scheduling integrations, and autoscaling patterns, across GCP, AWS, and Azure.
  • Collaborate with platform, data, and product engineering teams to create safe and efficient deployment, configuration, and runtime-operation paths for services and batch workloads.
  • Improve latency, resource utilization, and cost for core platform services, including multitenant runtime environments and experimental AI workloads.
  • Implement and harden infrastructure-as-code patterns, observability, and operational guardrails, including SLOs, dashboards, alerts, and safe rollout and rollback procedures.
  • Partner with the Costs and Runtime teams on attribution, guardrails, and automation mechanisms to improve runtime efficiency as customer and traffic volumes grow.
  • Participate in an on-call rotation for critical platform services, lead incident response when needed, and improve reliability, tooling, and documentation based on operational learnings.
  • Contribute to technical direction and roadmaps covering multitenancy, autoscaling, capacity and placement, and platformized patterns.

Requirements

  • Experience as a backend or platform engineer working across application behavior, infrastructure, and cost optimization.
  • Strong distributed systems fundamentals and experience operating high-throughput, low-latency services or batch pipelines in production.
  • Ability to own systems end-to-end, including design, implementation, testing, deployment, observability, and ongoing operations.
  • Experience with reliability practices such as SLOs, incident response, safe deployment strategies, and operational runbooks.
  • Pragmatic, execution-oriented approach suited to a fast-moving startup environment.
  • Clear communication and effective collaboration with infrastructure and product engineering teams.
  • Interest in multi-cloud, multi-tenant environments and running AI workloads efficiently at scale.

The interview process includes a brief AI-focused exercise or discussion covering how candidates think about, design, and use AI. Prior Glean experience is not required.

Location

  • Hybrid schedule requiring four days per week in the Mountain View office.

Compensation & Benefits

  • Standard base salary: $140,000–$220,000 annually.
  • Certain roles may be eligible for variable compensation, equity, and benefits.
  • Medical, vision, and dental coverage.
  • Generous time-off policy and 401(k) contribution opportunity.
  • Home office improvement stipend.
  • Annual education and wellness stipends.
  • Regular company events and healthy lunches daily.
  • Inclusive workplace committed to diversity and non-discrimination.

More jobs at Glean

Similar jobs