Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Minor Hotels Europe and Americas

SRE Software Engineer

Minor Hotels Europe and Americas

SRE Software Engineer focusing on maintaining production environments in Bogota. Collaborating with product teams to improve system reliability and performance with a focus on incident management.

Posted 7/24/2026full-timeBogota • 🇨🇴 ColombiaMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering and DevOps practices, with a strong focus on managing Kubernetes clusters, cloud infrastructure, and automation through Infrastructure as Code. Proficient in monitoring and observability tools to ensure high availability and system reliability.

Highest-signal resume keywords
Site Reliability EngineeringKubernetes ManagementCloud Infrastructure (AWS, Azure, GCP)Infrastructure as Code (Terraform)Monitoring and Observability Tools (Prometheus, Grafana, Datadog)

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Linux AdministrationTroubleshootingRoot Cause Analysis (RCA)Automation and Scripting (Bash, Python, Go)CI/CD Pipelines
Tools & Technologies
KubernetesDockerOpenShiftPrometheusGrafanaDatadogSplunkELKCloudWatch
Industry Keywords
DevOpsCloud OperationsInfrastructure EngineeringProduction SupportContinuous Improvement

Tech Stack

Tools & technologies
AWSAzureCloudDockerGoGoogle Cloud PlatformGrafanaKubernetesLinuxOpenShiftPrometheusPythonSplunkTerraform

About the role

Key responsibilities & impact
  • Support and maintain business-critical production environments, ensuring high availability and system reliability.
  • Monitor infrastructure and applications, proactively identifying and resolving issues before they impact users.
  • Participate in incident response activities, troubleshooting production outages and coordinating recovery efforts.
  • Perform RCA and contribute to postmortems, corrective actions, and continuous improvement initiatives.
  • Manage and optimize Kubernetes clusters and cloud infrastructure.
  • Develop and maintain monitoring dashboards, alerts, and observability solutions.
  • Automate operational processes and infrastructure deployments using IaC and scripting.
  • Collaborate with engineering and product teams to improve scalability, performance, and operational excellence.
  • Support and enhance CI/CD pipelines to ensure reliable and efficient software delivery.

Requirements

What you’ll need
  • 4+ years of experience in Site Reliability Engineering, DevOps, Cloud Operations, or Infrastructure Engineering.
  • Strong hands-on experience with Linux administration, troubleshooting, and production support.
  • Experience managing and supporting Kubernetes and containerized workloads (Docker/OpenShift is a plus).
  • Solid knowledge of AWS, Azure, or GCP cloud environments.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or CloudWatch.
  • Experience with Infrastructure as Code (Terraform preferred) and CI/CD pipelines.
  • Ability to troubleshoot complex production issues, perform Root Cause Analysis (RCA), and drive preventive improvements.
  • Working knowledge of automation and scripting using Bash, Python, or Go.
  • Intermediate to advanced English (B2+).

Benefits

Comp & perks
  • Competitive salary
  • Flexible working hours