Nomado24 Logo

Remote job

Verified remote

Site Reliability Engineer (f/m/d) – Observability & Internal Tools

Bertelsmann-Jobs

Source: smartrecruitersLocation: Germany onlyPublished: Jul 31, 2026Confirmed active: Aug 04, 2026
Full-timeTechnology

via smartrecruiters

Similar jobs by email. Free, unsubscribe anytime.

Free, unsubscribe anytime. Double opt-in.

At a glance

Tech stack
PrometheusGrafanaForgejoLinuxGCPAWS
Seniority:
Mid-level
Benefits
  • 30 days of vacation
  • Dec 24 & 31 off
  • Smart Fridays (4 days week possible)
  • Mobility (Germany ticket & JobRad)
  • Sports & health offerings
  • Mental health support
  • Corporate benefits
  • RTL+ access
Languages:
English

Site Reliability Engineer focused on observability and internal tools. You will own and evolve monitoring, alerting, automation and open-source platform components to improve reliability and security. Suited for engineers who prefer building, automating and taking full ownership of systems.

  • Own the observability stack (Prometheus, Grafana, Forgejo).
  • Define SLOs and actionable alerting to reduce firefighting.
  • Embed security engineering into delivery pipelines.
  • Remote-first with on-site collaboration in Berlin.

Automatically generated

Job description

We don't just subscribe to services; we engineer platforms. We believe that deep in-house expertise and a strong commitment to open source provide the flexibility and performance that enterprise SaaS cannot match. We design, we steer, and we own our stack end-to-end.

Your Mission

You will be the guardian and architect of smartclip’s internal infrastructure. Your goal is to turn observability and automation from "tools we use" into a "platform capability" that empowers every other engineer in the company.

  • Own the Observability Stack: Take the lead on Prometheus, Grafana, and Forgejo. You won't just maintain them; you will evolve them into a world-class monitoring ecosystem.
  • Engineer for Reliability: Design actionable alerting and define SLOs that move us from reactive firefighting to proactive stability.
  • Champion Open Source: Evaluate and implement cutting-edge open-source alternatives to proprietary software. You decide what enters our stack and how it integrates.
  • Secure the Pipeline: Integrate security engineering directly into our delivery process. You find the vulnerabilities before the pen tests do.
  • Master the Metal: Navigate the depths of Linux systems and distributed tooling, balancing bold experimentation with rock-solid production stability.

Your Skills

We are looking for a "builder" who is bored by simple configuration and thrives on systems thinking.

Core Competencies:

  • Observability Expertise: You have a proven track record of implementing metrics, logs, and traces. You know how to turn noisy data into actionable insights.
  • Linux & Systems Engineering: You are comfortable in the terminal and understand how distributed systems communicate and fail.
  • Automation Mindset: You hate doing the same thing twice. You use code to automate infrastructure and eliminate toil.
  • Ownership Culture: You embrace the "you build it, you run it" philosophy and take pride in the stability of the systems you touch.

The Power-Ups (Nice-to-haves):

  • Hands-on experience with GCP or AWS at a production scale.
  • Active contributions to the open-source community or a portfolio of self-hosted projects.
  • Experience in conducting blameless post-mortems and driving root-cause analysis.

Why you’ll love working with us

  • Ownership over tickets: You’re trusted with real responsibility, not just tasks. No unnecessary bureaucracy, no micromanagement – we rely on you to take things forward.
  • Build > Talk: We test what works – not what sounds good. Fail fast, learn faster.
  • High standards, low ego: We take our work seriously, but not ourselves. Direct feedback, honest collaboration, no drama.
  • Stay sharp: Hackathons, conferences, community – we invest in your growth and keep you at the cutting edge.
  • And yes – the fundamentals are covered too: 30 days of vacation + Dec 24 & 31 off, Smart Fridays (4 days week possible), mobility (Germany ticket & JobRad), sports & health offerings, mental health support, corporate benefits, RTL+ access, and more.

Your CV is just the starting point.

What matters more to us than your resume: a portfolio, a side project, a demo repo – anything that shows you don’t just talk about code, you ship it. Production-ready. Thought through. Done.

smartclip is committed to creating a diverse and inclusive environment. All qualified applicants will receive consideration for employment without regard to race, ethnicity, nationality, age, gender, gender identity, religion, sexual orientation, disability, or any other diverse characteristics.

Remote in our day-to-day work. On-site when it matters.

We work remote by default – focused, efficient, and with full ownership. For larger features, architectural decisions, and real brainstorming sessions, we come together in Berlin – fast, hands-on, and without unnecessary meeting overhead.

We use AI to accelerate – not to replace thinking.

We design the system, steer the output, and take responsibility for what we ship.

Fast where it makes sense. Careful where it matters.

smartrecruitersEngineeringFull timeAssociate

This role is provided by an external source. Applications are handled on the source website.

Prepare your application with AI

Have the AI advisor draft a cover letter, analyze how your profile fits this role, and prep you for the interview.

Sign in & prepare with AI

Free, you'll continue right where you left off after signing in.

Not the right job?

New jobs by email

We'll email you new matching remote jobs as soon as they appear. Free, unsubscribe anytime.

You confirm your subscription by email (double opt-in) and can unsubscribe at any time via the link in every email. Our Privacy Policy.

Matching job categories