Tech Job Finder - Find Software, Tech Sales and Product Manager Jobs.
Sign In
OR continue with e-mail and password
E-mail address
Password
Don't have an account?
Reset password
Join Tech Job Finder
OR continue with e-mail and password
Username
E-mail address
Password
Confirm Password
How did you hear about us?
By signing up, you agree to our Terms & Conditions and Privacy Policy.

Site Reliability Engineer

at Kong Inc.

Back to all Cloud & DevOps jobs
Kong Inc. logo
Industry not specified

Site Reliability Engineer

at Kong Inc.

Mid LevelNo visa sponsorshipAWS/GCP/Azure DevOps

Posted 19 hours ago

No clicks

Compensation
Not specified

Currency: Not specified

City
Milan
Country
Italy

**Site Reliability Engineer:** Build and maintain core infrastructure as code, ensuring high uptime and performance. Implement robust monitoring, logging, and alerting systems using tools like Terraform, Ansible, Prometheus, and ELK Stack. Collaborate with cross-functional teams, drive on-call rotation, and promote a culture of reliability. Requires 4+ years of experience in Site Reliability Engineering or similar roles, along with proficiency in Linux, cloud platforms (AWS, GCP), and containerization tools (Docker, Kubernetes).

Department: All Cost Center

Team: ENG

Location: Milan, Italy

Workplace Type: Hybrid

Employment Type: FullTime

Are you ready to unlock intelligence?

If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.

About the Role:

The Site Reliability Engineering team is the backbone of Kong's cloud services, responsible for architecting and operating the large-scale infrastructure that powers our customers' most critical applications. Our mission is to achieve world-class reliability and performance, enabling our product engineering teams to ship features with velocity and confidence. We are the guardians of uptime and the champions of developer delight.

What You’ll Do:

  • Build and maintain our core infrastructure as code using tools like Terraform and Ansible.

  • Implement robust monitoring, logging, and alerting systems to ensure our services meet and exceed 99.99% uptime.

  • Resolve production incidents through systematic debugging, and drive the blameless post-mortem process to prevent recurrence.

  • Write automation to reduce operational toil, improve system efficiency, and enable self-service for engineering teams.

  • Collaborate with developers to embed reliability and scalability best practices directly into the application lifecycle.

  • Contribute to our capacity planning, disaster recovery drills, and security hardening processes.

  • Participate in a fair and sustainable on-call rotation to ensure our platform is always available.

What You’ll Bring:

  • Experience operating production workloads on a major cloud provider (AWS, GCP, Azure).

  • Proficiency in at least one programming or scripting language, such as Golang, Python, or Bash.

  • Hands-on experience with containerization and orchestration technologies (Docker, Kubernetes).

  • Knowledge of Infrastructure as Code principles and tools (Terraform is a plus).

  • Familiarity with CI/CD concepts and pipeline tools (e.g., GitLab CI, Jenkins).

  • An understanding of modern observability stacks (e.g., Prometheus, Grafana, ELK).

#LI-BR2

About Kong:

Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit www.konghq.com.

Site Reliability Engineer

at Kong Inc.

Back to all Cloud & DevOps jobs
Kong Inc. logo
Industry not specified

Site Reliability Engineer

at Kong Inc.

Mid LevelNo visa sponsorshipAWS/GCP/Azure DevOps

Posted 19 hours ago

No clicks

Compensation
Not specified

Currency: Not specified

City
Milan
Country
Italy

**Site Reliability Engineer:** Build and maintain core infrastructure as code, ensuring high uptime and performance. Implement robust monitoring, logging, and alerting systems using tools like Terraform, Ansible, Prometheus, and ELK Stack. Collaborate with cross-functional teams, drive on-call rotation, and promote a culture of reliability. Requires 4+ years of experience in Site Reliability Engineering or similar roles, along with proficiency in Linux, cloud platforms (AWS, GCP), and containerization tools (Docker, Kubernetes).

Department: All Cost Center

Team: ENG

Location: Milan, Italy

Workplace Type: Hybrid

Employment Type: FullTime

Are you ready to unlock intelligence?

If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.

About the Role:

The Site Reliability Engineering team is the backbone of Kong's cloud services, responsible for architecting and operating the large-scale infrastructure that powers our customers' most critical applications. Our mission is to achieve world-class reliability and performance, enabling our product engineering teams to ship features with velocity and confidence. We are the guardians of uptime and the champions of developer delight.

What You’ll Do:

  • Build and maintain our core infrastructure as code using tools like Terraform and Ansible.

  • Implement robust monitoring, logging, and alerting systems to ensure our services meet and exceed 99.99% uptime.

  • Resolve production incidents through systematic debugging, and drive the blameless post-mortem process to prevent recurrence.

  • Write automation to reduce operational toil, improve system efficiency, and enable self-service for engineering teams.

  • Collaborate with developers to embed reliability and scalability best practices directly into the application lifecycle.

  • Contribute to our capacity planning, disaster recovery drills, and security hardening processes.

  • Participate in a fair and sustainable on-call rotation to ensure our platform is always available.

What You’ll Bring:

  • Experience operating production workloads on a major cloud provider (AWS, GCP, Azure).

  • Proficiency in at least one programming or scripting language, such as Golang, Python, or Bash.

  • Hands-on experience with containerization and orchestration technologies (Docker, Kubernetes).

  • Knowledge of Infrastructure as Code principles and tools (Terraform is a plus).

  • Familiarity with CI/CD concepts and pipeline tools (e.g., GitLab CI, Jenkins).

  • An understanding of modern observability stacks (e.g., Prometheus, Grafana, ELK).

#LI-BR2

About Kong:

Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit www.konghq.com.

SIMILAR OPPORTUNITIES

No similar jobs available at the moment.