Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Chainlink Labs

Senior Site Reliability Engineer, CCIP

Chainlink Labs

Senior Site Reliability Engineer ensuring operational excellence of Chainlink's CCIP platform. Focused on enhancing reliability, scalability, and operational standards across engineering teams.

Posted 6/29/2026full-timeRemote • North Carolina • 🇺🇸 United StatesSenior💰 $129,000 - $244,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Expertise in Site Reliability Engineering and Production Engineering with a focus on improving system reliability, scalability, and operational excellence. Proficient in implementing SLOs, SLIs, and error budgets to enhance service health and engineering efficiency.

Highest-signal resume keywords
Site Reliability EngineeringProduction EngineeringSLOs, SLIs, And Error BudgetsKubernetes EnvironmentsOpenTelemetry

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Distributed SystemsOperational ExcellenceDeployment SafetyProduction Engineering PracticesAutomationObservabilityIncident InvestigationEngineering EfficiencyPlatform ScalabilityOperational Readiness
Tools & Technologies
KubernetesOpenTelemetry
Industry Keywords
Cross-Chain Interoperability ProtocolReliability PracticesOperational StandardsProduction Infrastructure

Tech Stack

Tools & technologies
Distributed SystemsKubernetes

About the role

Key responsibilities & impact
  • Ensure the reliability, scalability, and operational excellence of the systems powering Chainlink's Cross-Chain Interoperability Protocol (CCIP).
  • Influence reliability practices across the platform and help establish operational standards that scale with the business.
  • Improve deployment safety and increase delivery velocity by advancing production engineering practices.
  • Establish distributed tracing across the platform to improve observability and accelerate incident investigation.
  • Eliminate operational toil through automation that increases engineering efficiency and platform reliability.
  • Drive adoption of meaningful SLOs, SLIs, and error budgets that guide engineering decisions and improve service health.
  • Increase platform scalability and operational readiness as CCIP continues to grow.

Requirements

What you’ll need
  • Demonstrated experience in Site Reliability Engineering, Production Engineering, or a similar role operating large-scale distributed systems.
  • Deep expertise defining, implementing, and driving adoption of SLOs, SLIs, and error budgets across engineering organizations.
  • Built and operated production Kubernetes environments supporting critical services.
  • Applied OpenTelemetry to improve observability across distributed systems.
  • Experience improving the reliability, scalability, and operability of production infrastructure.

Benefits

Comp & perks
  • Long-term incentives
  • Comprehensive benefits