Site Reliability Engineer
at Encora10
- Work model
- Remote
- Location
- Bolivia; Colombia; Costa Rica; Peru
- Posted
- 5d ago
at Encora10
<p></p> <div> <p><strong>Job Title:</strong> Site Reliability Engineer (SRE)<br><strong>Key Skills:</strong> Kubernetes, AWS/Azure/GCP, Terraform, Python, Observability, CI/CD<br><strong>Experience:</strong> +6 YOE.<br><strong>Location:</strong> Costa Rica, Peru, Colombia, and Bolivia.<br><strong>Mode:</strong> Remote.</p> <p>We at Coforge are hiring <strong>Site Reliability Engineer (SRE) (#22323)</strong> with the following skill set.</p> <p><strong>Key Responsibilities</strong><br>· Design, build, and operate scalable and highly available cloud platforms.<br>· Ensure reliability, performance, and stability of distributed production systems.<br>· Implement and maintain Infrastructure as Code using Terraform or similar tools.<br>· Manage Kubernetes-based and containerized environments.<br>· Define and operate SLOs, SLIs, error budgets, dashboards, runbooks, and alerting standards.<br>· Implement observability, monitoring, and incident response practices.<br>· Participate in on-call rotations and respond to production incidents.<br>· Collaborate with engineering teams to improve automation, scalability, and platform resilience.<br>· Conduct postmortem reviews and drive continuous reliability improvements.</p> <p><strong>Required Skills & Qualifications</strong><br>· Bachelor’s degree in Computer Science, Engineering, Information Systems, Software Engineering, or a related technical field, or equivalent practical experience.<br>· 6+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, DevOps Engineering, Backend Engineering, or Production Engineering.<br>· Strong software engineering skills in at least one language such as Python, Go, Java, TypeScript, or C#.<br>· Strong understanding of distributed systems, microservices, APIs, asynchronous processing, queues, databases, caching, retries, idempotency, and failure modes.<br>· Experience with cloud infrastructure on AWS, Azure, or GCP.<br>· Experience with Kubernetes, containers, Terraform or similar IaC tooling, CI/CD pipelines, and Linux-based systems.<br>· Experience with observability tools such as Datadog, Prometheus, Grafana, OpenTelemetry, CloudWatch, New Relic, Splunk, or Sentry.<br>· Experience defining and operating SLOs, SLIs, error budgets, alerting standards, dashboards, runbooks, and incident response practices.<br>· Strong communication skills and experience working across cross-functional teams.</p> <p><strong>Preferred Skills</strong><br>· Cloud, Kubernetes, Infrastructure, Reliability Engineering, Security, or DevOps certifications.<br>· Experience in logistics, transportation, final-mile delivery, field-service software, routing, dispatch, or fleet operations.<br>· Experience working with operational SaaS or marketplace platforms.<br>· Experience driving automation, platform reliability, and operational excellence initiatives.</p> <p><strong>Posted On:</strong> 14-08-2026</p> <p>At Coforge, we hire professionals based solely on their skills and qualifications and do not discriminate based on age, disability, religion, gender, sexual orientation, socioeconomic status, or nationality.</p> </div>