FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Site Reliability Engineer, CCIP
Chainlink LabsSenior Site Reliability Engineer ensuring operational excellence of Chainlink's CCIP platform. Focused on enhancing reliability, scalability, and operational standards across engineering teams.
Posted 6/29/2026full-timeRemote • North Carolina • 🇺🇸 United StatesSenior💰 $129,000 - $244,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Expertise in Site Reliability Engineering and Production Engineering with a focus on improving system reliability, scalability, and operational excellence. Proficient in implementing SLOs, SLIs, and error budgets to enhance service health and engineering efficiency.
Highest-signal resume keywords
Site Reliability EngineeringProduction EngineeringSLOs, SLIs, And Error BudgetsKubernetes EnvironmentsOpenTelemetry
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Distributed SystemsOperational ExcellenceDeployment SafetyProduction Engineering PracticesAutomationObservabilityIncident InvestigationEngineering EfficiencyPlatform ScalabilityOperational Readiness
Tools & Technologies
KubernetesOpenTelemetry
Industry Keywords
Cross-Chain Interoperability ProtocolReliability PracticesOperational StandardsProduction Infrastructure
Tech Stack
Tools & technologiesDistributed SystemsKubernetes
About the role
Key responsibilities & impact- Ensure the reliability, scalability, and operational excellence of the systems powering Chainlink's Cross-Chain Interoperability Protocol (CCIP).
- Influence reliability practices across the platform and help establish operational standards that scale with the business.
- Improve deployment safety and increase delivery velocity by advancing production engineering practices.
- Establish distributed tracing across the platform to improve observability and accelerate incident investigation.
- Eliminate operational toil through automation that increases engineering efficiency and platform reliability.
- Drive adoption of meaningful SLOs, SLIs, and error budgets that guide engineering decisions and improve service health.
- Increase platform scalability and operational readiness as CCIP continues to grow.
Requirements
What you’ll need- Demonstrated experience in Site Reliability Engineering, Production Engineering, or a similar role operating large-scale distributed systems.
- Deep expertise defining, implementing, and driving adoption of SLOs, SLIs, and error budgets across engineering organizations.
- Built and operated production Kubernetes environments supporting critical services.
- Applied OpenTelemetry to improve observability across distributed systems.
- Experience improving the reliability, scalability, and operability of production infrastructure.
Benefits
Comp & perks- Long-term incentives
- Comprehensive benefits