KNIME

(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz

Stellenbeschreibung:

Overview

In this role you will help design, build and operate KNIME’s next-generation cloud platform with a focus on reliability, security and cost efficiency. You’ll collaborate with multiple product teams to bring KNIME SaaS to production at scale, using infrastructure-as-code and automation. The role emphasizes on-call ownership, incident response, and setting operational standards across teams. You will shape platform stability and observability to support a multi-tenant, enterprise-grade product. This is a chance to influence cloud-native foundations and drive scalable, secure deployments.

Leistungen / Benefits
  • Hybrid working
  • Flexible hours
  • Subsidised sports or yoga courses
  • Physiotherapy
  • Flu shots at select locations
  • Opportunity to impact cloud strategy
Verantwortungsbereiche
  • Automate deployment and operations of large-scale SaaS systems using code
  • Build infrastructure as code with Helm, Terraform, CloudFormation and Azure ARM
  • Participate in on-call rotations, triage incidents, perform root-cause analysis and resolution
  • Set deployment standards for reliability, scalability, traceability and monitoring, and drive adoption with product teams
  • Instrument systems for performance, reliability and cost effectiveness
  • Collaborate with product and engineering to plan, manage dependency risks, and promote reliability standards across teams
Zentrale Anforderungen
  • One or more cloud platform certifications (AWS, Kubernetes, Linux or similar)
  • Strong cloud experience with AWS or Azure (ideally both) and knowledge of VPC, IAM, EKS, ECR, EC2, S3, RDS, CloudWatch and equivalents in Azure
  • Experience deploying software to Kubernetes and understanding of Kubernetes patterns
  • Scripting in Python and Shell; knowledge of Go or Java is a plus
  • Linux systems knowledge with networking expertise (security, routing, load balancers, firewalls)
  • Experience with service telemetry: metrics, logging, tracing in distributed systems
  • Knowledge of OAuth/OIDC providers such as Keycloak
  • Knowledge of relational databases, especially Postgres
  • Ability to work independently and in a distributed, cross-cultural team
  • strong communication across distributed teams
  • demonstrated on-call and incident management capabilities
  • Clear and concise communication
  • Cross-functional collaboration
  • Problem-solving mindset
  • Kubernetes
  • Helm
  • Terraform
NOTE / HINWEIS:
EnglishEN: Please refer to Fuchsjobs for the source of your application
DeutschDE: Bitte erwähne Fuchsjobs, als Quelle Deiner Bewerbung

Stelleninformationen

  • Veröffentlichungsdatum:

    14 Sep 2026
  • Standort:

    Konstanz
  • Typ:

    Vollzeit
  • Arbeitsmodell:

    Vor Ort
  • Kategorie:

  • Erfahrung:

    2+ years
  • Arbeitsverhältnis:

    Angestellt

KI Suchagent

AI job search

Möchtest über ähnliche Jobs informiert werden? Dann beauftrage jetzt den Fuchsjobs KI Suchagenten!