Back to search
EngineeringHybrid

Site Reliability Engineer

Kong · Hybrid

Apply
Last seen by MeritLog September 11, 2026Source: AshbySource version: ashby-public-job-posting-v1

MeritLog read this listing from Kong's Ashby job board and last checked it on September 11, 2026.

Source: the employer's Ashby job board. Open the original listing for current details.

Job details

Work model
Hybrid
Salary
Not listed by source
Location
Milan, Italy
Occupation
Software Developers(O*NET 15-1252.00)
Company website
konghq.com

What the role asks for

What you'd do

  • Build and maintain our core infrastructure as code using tools like Terraform and Ansible.
  • Implement robust monitoring, logging, and alerting systems to ensure our services meet and exceed 99.99% uptime.
  • Resolve production incidents through systematic debugging, and drive the blameless post-mortem process to prevent recurrence.
  • Write automation to reduce operational toil, improve system efficiency, and enable self-service for engineering teams.
  • Collaborate with developers to embed reliability and scalability best practices directly into the application lifecycle.
  • Contribute to our capacity planning, disaster recovery drills, and security hardening processes.
  • Participate in a fair and sustainable on-call rotation to ensure our platform is always available.

What they're asking for

  • Experience operating production workloads on a major cloud provider (AWS, GCP, Azure).Skill
  • Proficiency in at least one programming or scripting language, such as Golang, Python, or Bash.Skill
  • Hands-on experience with containerization and orchestration technologies (Docker, Kubernetes).Skill
  • Knowledge of Infrastructure as Code principles and tools (Terraform is a plus).SkillPreferred
  • Familiarity with CI/CD concepts and pipeline tools (e.g., GitLab CI, Jenkins).Skill
  • An understanding of modern observability stacks (e.g., Prometheus, Grafana, ELK).Skill

Parsed by MeritLog from the employer’s own posting. The full description follows below.

Job description

Are you ready to unlock intelligence? If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others. ABOUT THE ROLE: The Site Reliability Engineering team is the backbone of Kong's cloud services, responsible for architecting and operating the large-scale infrastructure that powers our customers' most critical applications. Our mission is to achieve world-class reliability and performance, enabling our product engineering teams to ship features with velocity and confidence. We are the guardians of uptime and the champions of developer delight. WHAT YOU’LL DO: - Build and maintain our core infrastructure as code using tools like Terraform and Ansible. - Implement robust monitoring, logging, and alerting systems to ensure our services meet and exceed 99.99% uptime. - Resolve production incidents through systematic debugging, and drive the blameless post-mortem process to prevent recurrence. - Write automation to reduce operational toil, improve system efficiency, and enable self-service for engineering teams. - Collaborate with developers to embed reliability and scalability best practices directly into the application lifecycle. - Contribute to our capacity planning, disaster recovery drills, and security hardening processes. - Participate in a fair and sustainable on-call rotation to ensure our platform is always available. WHAT YOU’LL BRING: - Experience operating production workloads on a major cloud provider (AWS, GCP, Azure). - Proficiency in at least one programming or scripting language, such as Golang, Python, or Bash. - Hands-on experience with containerization and orchestration technologies (Docker, Kubernetes). - Knowledge of Infrastructure as Code principles and tools (Terraform is a plus). - Familiarity with CI/CD concepts and pipeline tools (e.g., GitLab CI, Jenkins). - An understanding of modern observability stacks (e.g., Prometheus, Grafana, ELK). #LI-BR2 About Kong: Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500® and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic - on any model, any cloud. For more information, visit www.konghq.com http://www.konghq.com.

Keep exploring

More Engineering roles

Search all jobs

Privacy choices

Analytics and advertising stay off unless you allow them. Private data stays out.

Read the privacy notice