FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Customer Reliability Engineer – Infrastructure
AstronomerSenior infrastructure specialist focusing on reliability of cloud infrastructure for Astronomer's products. Directly interacting with customers and enhancing their experience for operational success.
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
KubernetesLinuxPython scriptingDevOpsCI/CDcloud infrastructuredistributed systemsmonitoringautomationtroubleshooting
Soft Skills
communicationcustomer experienceproblem-solvingcollaborationprioritizationguidancedocumentation enhancementactive triaging
Tools & Technologies
AWSGCPAzuremonitoring systemsalerting systems
Industry Keywords
multi-cloud implementationsproduction systemscustomer solutionsSLAs
Tech Stack
Tools & technologiesAWSAzureCloudDistributed SystemsGoogle Cloud PlatformKubernetesLinuxPython
About the role
Key responsibilities & impact- Provide solutions to customers to make them successful using our products.
- Troubleshoot customer environments and engage in active triaging with customers
- Provide feedback to the product development teams on customer needs and pain points.
- Build out our monitoring and alerting systems.
- Build and maintain automation to ensure daily operational tasks are handled as efficiently as possible.
- Help direct the architecture of the products and contribute where possible.
- Own the customer experience, working directly with customers to prioritize and solve issues, meet SLAs, and provide “white glove” guidance on the path to production.
- Participate remotely within a fully distributed team.
- Enhance and enrich customer documentation
- Work with the latest technology and multi-cloud implementations
Requirements
What you’ll need- 5 years of experience, preferably with large, complex cloud infrastructures operating at scale
- 3 years of experience with Kubernetes
- Experience managing a Production distributed system with at least one major cloud provider (one or all: AWS, GCP, Azure)
- Strong Linux experience
- Knowledge of how to operate and monitor issues for distributed systems
- Previous experience in handling customers issues (internal or external)
- Strong communication skills
- DevOps or CI/CD experience
- Python scripting
- Good troubleshooting Skills
Benefits
Comp & perks- Participate in on-call rotation for weekend coverage
- Work with the latest technology and multi-cloud implementations