Site Reliability Engineer (Europe)
Arango • Competitive / DOE
Site Reliability Engineer (Europe)
> ROLE OVERVIEW
At ArangoDB, we are building a robust, cloud-native infrastructure to support our distributed database systems, which power mission-critical applications for a wide range of industries. We are searching for a Site Reliability Engineer (SRE) to ensure the reliability, scalability, and performance of our infrastructure and applications, with a focus on automation, monitoring, and optimizing cloud environments. As a Site Reliability Engineer (SRE), you will be responsible for maintaining and improving the reliability of our distributed database systems running on Kubernetes and cloud environments (AWS, Google Cloud).
> CORE RESPONSIBILITIES
- Design, implement, and maintain cloud infrastructure on AWS and Google Cloud platforms.
- Ensure the scalability, performance, and reliability of our Kubernetes-based distributed database systems.
- Collaborate with developers to write efficient, production-grade code in Golang to automate infrastructure management and improve system operations.
- Optimize and automate CI/CD pipelines, deployment processes, and monitoring systems to support our production environment.
- Develop strategies for disaster recovery, high availability, and fault tolerance.
- Proactively identify system bottlenecks, troubleshoot, and resolve issues across the stack (network, OS, cloud infrastructure)
> HARD REQUIREMENTS & SPECS
- Golang programming skills
- Experience with cloud infrastructure on AWS and Google Cloud platforms
- Knowledge of Kubernetes and containerization
- Experience with automation tools such as Ansible, Terraform, or CloudFormation
Is the AI extraction inaccurate? Report an issue