Site Reliability Engineer (f/m/d) – Observability & Internal Tools at Apply now! | Germany | Rezi

Site Reliability Engineer (f/m/d) – Observability & Internal Tools at Apply now!

Site Reliability Engineer (f/m/d) – Observability & Internal Tools

Apply now! · Germany

3 weeks ago

Site Reliability Engineer (f/m/d) – Observability & Internal Tools

Apply now! · Germany

22 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

Take full ownership of smartclip’s internal utility and platform tooling, focusing on observability, automation, and developer infrastructure. Evolve existing systems, research open-source alternatives, and implement them, prioritizing in-house expertise over expensive enterprise SaaS.

Responsibilities

  • Operate and advance our observability stack (including Prometheus, Grafana, and Forgejo).
  • Replace "buy" decisions with robust "build & maintain" strategies.
  • Design observability as a platform capability.
  • Define SLOs and create actionable alerting to stop incidents before they start.
  • Embed security engineering into the delivery process.
  • Navigate Linux systems and distributed tooling.
  • Balance bold exploration with production stability.

Requirements

  • Be motivated by systems thinking and deep technical curiosity.
  • Implement a clear strategy for metrics, logs, and traces.
  • Transform "noisy alerts" into "actionable insights."
  • Live the "you build it, you run it" philosophy.
  • Stop the ticket ping-pong and end the excuses.
  • Design and evolve production-grade setups on GCP or AWS.
  • Show contributions to open-source projects.
  • Turn passion for root-cause analysis into blameless post-mortems.
  • Provide a portfolio, side project, or demo repo showcasing shipped, production-ready, thought-through work.

Skills

  • Observability Mindset
  • Ownership
  • Systems Thinking
  • Technical Curiosity
  • Linux Systems
  • Distributed Tooling
  • Security Engineering
  • GCP
  • AWS
  • Open-Source Contributions
  • Root-Cause Analysis
  • Blameless Post-Mortems

Location

  • Remote

Work Type

  • Remote
  • On-site (for specific events)

Experience Level

  • Mid-level

Benefits

  • 30 days of vacation + Dec 24 & 31 off
  • Smart Fridays (4 days week possible)
  • Mobility (Germany ticket & JobRad)
  • Sports & health offerings
  • Mental health support
  • Corporate benefits
  • RTL+ access

About the Company

  • We use AI to accelerate – not to replace thinking.
  • We design the system, steer the output, and take responsibility for what we ship.
  • Fast where it makes sense. Careful where it matters.

Equal Opportunity

  • smartclip is committed to creating a diverse and inclusive environment. All qualified applicants will receive consideration for employment without regard to race, ethnicity, nationality, age, gender, gender identity, religion, sexual orientation, disability, or any other diverse characteristics.