Site Reliability Engineer II

EUR 55-110 per hour
MIDDLE
✅ Hybrid ✅ On-site
✅ Contract / Freelance

Tech Stack

AWS Communication @ 3 Design Patterns Experimentation Kubernetes Observability @ 3 SRE Security

Details

Site Reliability Engineer II specialists treat operations as a software problem, focusing on the availability, performance, scalability, latency, observability, and efficiency of systems and services. The role aims to reduce operational toil and complexity through automation and improve system reliability.

SRE II engineers implement technical solutions based on business requirements, estimate effort and impact, deliver high-quality work, collaborate with partner teams, and participate in incident response. Depending on the business area, they may act as part of a business service owner team, infrastructure owner, or consultant to product development teams.

Responsibilities

Building Software Applications

  • Build software applications using relevant development languages and business-area systems, services, and tools.
  • Refactor and simplify code using design patterns where appropriate.
  • Ensure application quality by following testing techniques and the test strategy.
  • Write readable and reusable code using standard patterns and libraries.
  • Maintain data security, integrity, and quality according to company standards and best practices.

Software Systems Design

  • Evaluate architecture solutions considering cost, business requirements, technology requirements, and emerging technologies.
  • Understand and describe the implications of changing or adding systems within the infrastructure and architecture.
  • Apply engineering techniques such as prototyping, spiking, and vendor evaluation.
  • Design adaptable solutions that meet current and future business requirements.

End-to-End System Ownership

  • Own services end to end by monitoring application health and performance and acting on relevant metrics.
  • Reduce business continuity risks and bus factor through appropriate practices, tools, documentation, runbooks, and OpDocs.
  • Use continuous delivery and experimentation frameworks to reduce risk and obtain customer feedback.
  • Independently manage applications or services through deployment and production operations.
  • Maintain data security, integrity, and quality.

Incident Management, Automation, and Observability

  • Resolve live production issues and mitigate customer impact within SLA.
  • Improve system reliability through root cause analysis, long-term solutions, postmortems, and incident tracking.
  • Reduce technical debt, identify bottlenecks, and prepare infrastructure for scaling.
  • Reduce operational costs and human labor through new technologies, automation, and software features addressing availability, scalability, latency, and efficiency.
  • Monitor production systems and network infrastructure using observability metrics, business KPIs, and capacity planning.
  • Partner with development teams to establish appropriate observability metrics.

Additional Responsibilities

  • Apply critical thinking to identify underlying issues and develop logical solutions.
  • Identify and implement continuous quality, process, system, and structural improvements.
  • Communicate clearly with different audiences and use active listening to reach mutually agreeable solutions.
  • Advise product teams on technical solutions meeting functional, nonfunctional, and architectural requirements.
  • Help define technical capability direction by evaluating target architecture improvements and aligning architectural decisions.

Requirements

Knowledge and skills in:

  • Building software applications
  • Software system design
  • End-to-end system ownership
  • Technical incident management
  • Operations, automation, and toil reduction
  • Observability, monitoring, and alerting
  • Critical thinking
  • Continuous quality and process improvement
  • Effective communication
  • Architectural guidance

Tech Stack

  • Kubernetes
  • AWS
  • Any coding language

Contract Details

  • Independent contractor engagement
  • Project duration: 3 months
  • Start date: September 21, 2026
  • End date: December 20, 2026

More jobs at Booking.com

Similar jobs