Job
- Level
- Erfahren
- Ort
- Berlin
- Arbeitsmodell
- Hybrid, Onsite
- Job Feld
- IT, System
- Anstellung
- Vollzeit
- Vertragsart
- Unbefristetes Dienstverhältnis
Job Zusammenfassung
In dieser Rolle entwickelst und optimierst du die Plattform für Managed Nextcloud und weitere Dienste. Du integrierst neue Webservices in Kubernetes und automatisierst den Betrieb mit Tools wie Terraform und GitLab CI/CD.
Job Technologien
Deine Rolle im Team
- As a Site Reliability Engineer (SRE) in our Application Hosting Team, you form the technical backbone of our product platform for Managed Nextcloud, Nextcloud Workspace, IONOS GPT, as well as other web services that we operate on our Kubernetes platform.
- Together with experienced colleagues, you design new services and products that remain performant and resilient even under the highest load.
- Your main area of focus is the further development of the infrastructure/platform for our products, as well as the integration of new products/web services into our Kubernetes and Cloud infrastructure.
- You are responsible for the stable and secure operation of our product platform.
- Your expertise is in demand when it comes to in-depth analysis and optimization of our primarily containerized and Kubernetes-based application infrastructure.
- You live automation.
- Using tools like Terraform, GitLab CI/CD, and ArgoCD, you provision and manage our entire infrastructure declaratively and reproducibly.
- You analyze and resolve complex issues in a distributed system landscape and work on the continuous improvement of our platform.
- You develop and maintain our monitoring, logging, and alerting solutions (e.g., using Prometheus, Grafana, ELK stack) to proactively identify bottlenecks and sources of error.
Unsere Erwartungen an dich
Qualifikationen
- You have a proactive, solution-oriented, independent way of working, and the ability to systematically analyze and sustainably solve complex technical problems.
- Good German and English language skills are required.
Erfahrung
- You have several years of experience as a Site Reliability Engineer or in a related role (Linux System Administrator, Platform Engineer, DevOps Engineer, Full Stack Developer) in a Linux and Kubernetes environment.
- Strong knowledge and several years of experience using the Linux operating system, container technologies, and specifically Kubernetes.
- You have experience with Infrastructure as Code (preferably Terraform), CI/CD pipelines (e.g., GitLab CI/CD or GitHub Actions), and using Helm Charts.
- You can confidently develop in at least one programming or scripting language (e.g., Go, Python, Bash) to solve automation and monitoring tasks, and you may already have initial experience building Operators.
- Experience operating and troubleshooting highly available and distributed production environments, including monitoring, alerting, and log analysis of distributed applications (e.g., Prometheus, Grafana, FluentD, ELK, VictoriaMetrics, Icinga).
Unser Angebot
- Hybrid working model.
- Flexible working hours through trust-based working hours.
- At some locations a subsidized canteen and various free drinks.
- Modern office space with very good transport connections.
- Various employee discounts for activities and products.
- Employee events such as summer and winter parties, as well as workshops.
- Numerous training and development opportunities.
- Various health offers, such as sports and health courses.
Themen mit denen du dich im Job beschäftigst
Job Standorte
Das ist dein Arbeitgeber
Berufliche Schulen Potsdam der ASG - Anerkannten Schulgesellschaft mbH
Die Beruflichen Schulen Potsdam der ASG ermöglichen eine praxisorientierte Ausbildung zum Erzieher, die schulische Theorie und berufliche Praxis eng miteinander verbindet. Ein kompetentes Dozententeam sowie regelmäßige Kooperationen mit Praxispartnern fördern die Vorbereitung auf pädagogische Berufe in der Kinder- und Jugendhilfe.
Description
- Unternehmenstyp
- Etablierte Firma
- Branche
- Bildungswesen