BairesDev is a leading technology company delivering innovative solutions to major clients and startups. The Senior Site Reliability Engineer (SRE) will ensure systems remain reliable and resilient, focusing on infrastructure expertise and data-driven reliability practices.
Responsibilities:
- Manage and scale Kubernetes environments across multiple clusters
- Build and maintain CI/CD pipelines and Infrastructure as Code
- Implement and maintain observability across systems for proactive issue detection
- Define and track reliability standards, driving continuous improvement through incident learnings
Requirements:
- 5+ years of experience in Site Reliability Engineering or infrastructure engineering
- Strong expertise in Kubernetes, including operators, autoscaling, and multi-cluster management
- Experience with CI/CD pipeline engineering
- Proficiency in Infrastructure as Code using Terraform and Helm
- Hands-on experience with observability stacks such as Prometheus, Grafana, Datadog, or OpenTelemetry
- Experience defining SLOs/SLIs, managing error budgets, and leading post-incident reviews
- Background in cloud security tooling and IAM hardening
- Advanced proficiency in English