Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
nilo

Senior Site Reliability Engineer, SRE, Backend

nilo

Senior Site Reliability Engineer owning AWS infrastructure and backend development for nilo's mental well-being platform. Ensuring reliability and security while supporting sensitive mental health data.

Posted 7/25/2026full-timeBerlin • 🇩🇪 GermanySeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates extensive expertise in AWS infrastructure management, including serverless and containerized environments, while ensuring compliance with GDPR and security best practices. Proficient in Terraform for infrastructure as code, alongside strong backend development capabilities in Python, Node.js, or Go.

Highest-signal resume keywords
AWS Infrastructure ManagementTerraform Module DesignEvent-Driven ArchitecturePostgreSQL Performance OptimizationIncident Response Leadership

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
AWS LambdaECS FargateTerraformPostgreSQLSQSSNSEventBridgeDatadogPythonNode.js
Soft Skills
Effective CommunicationIncident Response
Tools & Technologies
DatadogTerraformAWSSentryCI/CD Pipelines
Industry Keywords
GDPR ComplianceSecurity FundamentalsOperational ExcellenceEvent-Driven ServicesProduction Ownership

Tech Stack

Tools & technologies
AWSDynamoDBGoGrafanaJavaScriptNode.jsPostgresPythonTerraform

About the role

Key responsibilities & impact
  • Own our AWS infrastructure end to end — Lambda, ECS Fargate, SQS, SNS, EventBridge, SES, Cognito, DynamoDB, RDS Postgres, and DMS
  • Manage everything as code in Terraform, with well-designed modules, clean state management, and a solid review workflow
  • Build and maintain CI/CD pipelines with safe rollout and rollback across web, mobile backends, and infrastructure
  • Design our event-driven services for resilience: retries, dead-letter queues, idempotency, graceful degradation
  • Own our Datadog and Sentry setup — define SLOs, build dashboards, and keep alerting actionable instead of noisy
  • Lead incident response and run blameless postmortems that actually change how we build
  • Harden our security posture: IAM, secrets management, network boundaries, Cognito auth flows, and vulnerability remediation
  • Protect sensitive health data and support our GDPR and compliance requirements
  • Monitor and optimize AWS spend without compromising reliability
  • Contribute to backend development — APIs, event consumers, data pipelines, and Postgres and DynamoDB performance
  • Participate in architectural discussions and mentor engineers on operational excellence

Requirements

What you’ll need
  • 5+ years in SRE, DevOps, platform, or backend engineering, with real production ownership
  • Deep AWS experience across serverless and containers — Lambda, ECS Fargate, and debugging both under pressure
  • Strong Terraform skills, including module design and managing state across multiple environments
  • Hands-on experience with event-driven architecture (SQS, SNS, EventBridge) and a healthy respect for its failure modes
  • Solid PostgreSQL: query tuning, indexing, connection management, and zero-downtime migrations
  • Production experience with Datadog or a comparable observability platform (Grafana, New Relic, Honeycomb)
  • Comfortable writing production backend code in [Python / Node.js / Go]
  • Genuine on-call and incident response experience — you've led an incident and written the postmortem
  • Strong security fundamentals: IAM, least privilege, secrets, network isolation, common web vulnerabilities
  • Pragmatic about complexity — you reach for the simplest thing that meets the reliability bar
  • Effective communicator who can explain a technical tradeoff without jargon

Benefits

Comp & perks
  • Real ownership: a small team, short feedback loops, and no layers of approval between you and production
  • Work that matters: the reliability you build directly affects people reaching for mental health support
  • Free access to the nilo app (incl. family support)
  • A dedicated learning budget for your personal and professional development
  • Work abroad for up to 90 days per year (within the EU)
  • Hybrid working model: 2 days/week from the office, 3 days from home
  • Urban Sports Club membership at a discounted price
  • Equity options: you benefit from any increase in nilo's valuation that you've helped to create
  • Regular team and company events
  • Bring your dog to work: we have 4 office dogs