FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing cloud infrastructure on GCP and AWS, with strong capabilities in Kubernetes operations, CI/CD pipeline support, and automation using Terraform and scripting languages. Proficient in monitoring and optimizing infrastructure health and costs while ensuring compliance and operational excellence.
Highest-signal resume keywords
GCP Cloud ManagementKubernetes OperationsTerraform AutomationCI/CD Pipeline SupportCost Optimization
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
KubernetesTerraformPythonDockerHelmArgoCDELK/EFK StackPrometheusGrafanaService Mesh
Tools & Technologies
GitHub ActionsAlertForgeMySQLPostgresMongoDB
Industry Keywords
Cloud InfrastructureComplianceObservabilityCost MonitoringRoot Cause Analysis
Tech Stack
Tools & technologiesAWSCloudDNSDockerGoGoogle Cloud PlatformGrafanaGroovyKubernetesMongoDBMySQLNode.jsPostgresPrometheusPythonTerraform
About the role
Key responsibilities & impact- Manage and maintain cloud infrastructure on GCP (primary), with exposure to AWS
- Perform day-to-day Kubernetes (GKS) operations, pod troubleshooting, node management, scaling, and resource tuning
- Execute compliance infrastructure tasks: instance provisioning, security group updates, DNS changes, and certificate renewals
- Support and maintain CI/CD pipelines built on ArgoCD and GitHub Actions
- Assist engineering teams with onboarding applications to the CI/CD platform
- Troubleshoot build failures, deployment issues, and rollback scenarios
- Ensure deployment hygiene, proper tagging, versioning, and environment promotion
- Monitor infrastructure health using Prometheus, Grafana, and AlertForge
- Respond to alerts, triage production issues, and execute runbooks
- Perform initial RCA (Root Cause Analysis) for production incidents and escalate when needed
- Maintain and update runbooks and operational documentation
- Write and maintain Terraform modules for infrastructure provisioning
- Automate repetitive operational tasks using shell scripts, Python, or Go
- Identify and implement improvements to reduce toil and manual effort
- Contribute to internal tools and utilities that improve developer experience
- Own infrastructure cost as a first-class metric drive continuous cost optimisation across compute, network, and storage
Requirements
What you’ll need- Experience with Helm, orchestrating containerized infrastructure using docker and Kubernetes.
- Experience with service mesh (Linkerd, Istio)
- Exposure to ArgoCD and GitOps workflows
- Experience with log aggregation ELK/EFK stack
- Experience with cost monitoring and right-sizing recommendations
- Experience with GCP cloud environments
- Exposure on observability - Graphana, Prometheus, APM, Alter Manager
- Working knowledge of Python, Groovy or other programming languages.
- Strong prior experience using automation tools like Terraform (must)
- Familiarity to database operations: for MySQL, Postgres, MongoDB
Benefits
Comp & perks- Remote work options
- Dynamic work culture
- Professional development opportunities
