Talent.com
AP Solutions GmbH
Site Reliability Engineer (m/f/d)AP Solutions GmbH • Saint-Ouen, 93, FR
Site Reliability Engineer (m/f/d)

Site Reliability Engineer (m/f/d)

AP Solutions GmbH • Saint-Ouen, 93, FR
Il y a 9 jours
Type de contrat
  • Temps plein
Description de poste
Site Reliability Engineer (m/f/d)

Key responsibilities


As a Site Reliability Engineer within Advanced Analytics (DA3) in the Chief Data & AI Office at Allianz Partners, you will join the platform engineering team to own the reliability and operational health of the central engineering platform.

You will define and maintain service level objectives, drive incident response at the infrastructure layer, and systematically eliminate operational toil through automation.

You will work closely with Platform Engineers, Security Engineers, and incident-response leads to ensure the platform meets its reliability commitments across production workloads spanning AI services, Java APIs, and frontend applications.


Through this role, you will have the main following responsibilities:

  • Define, instrument, and maintain SLOs and SLIs for platform components; own error budget tracking and produce regular reliability reports for senior leadership.
  • Serve on the on-call rotation as the infrastructure escalation tier; lead incident response for cluster-level, network-level, and storage failures; chair blameless post-incident reviews.
  • Implement and operate Kubernetes infrastructure (AKS): cluster lifecycle management, networking, resource quotas, autoscaling configuration, and multi-tenancy patterns across product team namespaces.
  • Develop Infrastructure as Code (Terraform) to provision and manage Azure resources with consistency, auditability, and repeatable rollback capability.
  • Build and maintain observability infrastructure: Prometheus, Grafana, Azure Monitor, and Application Insights; own alerting rules, dashboards, and distributed tracing coverage across platform components.
  • Perform capacity planning and cost-aware resource management: right-size node pools, tune vertical and horizontal pod autoscalers, and identify resource waste across namespaces.
  • Identify and eliminate toil: automate repetitive operational tasks through scripting and tooling; measure and track toil reduction over time.
  • Maintain platform reliability procedures: rolling upgrades, backup and recovery testing, disaster recovery runbooks, and change freeze coordination.
  • Contribute to CI/CD pipelines and GitOps tooling (GitHub Actions, ArgoCD) from a reliability and deployment safety perspective; work with platform engineering on release gates and rollback mechanisms.
  • Collaborate with incident-response leads on incident SLA targets and operational procedures; work with Security Engineers on infrastructure hardening and vulnerability remediation.

What you bring

  • 5+ years professional experience in site reliability engineering, DevOps, or platform engineering roles.
  • Strong Kubernetes experience: cluster operations, networking (Ingress, network policies), storage, autoscaling, and hands-on troubleshooting across production environments.
  • Solid Infrastructure as Code experience with Terraform; familiarity with Bicep or ARM templates is a plus.
  • Production experience with Azure cloud services: AKS, ACR, Key Vault, Azure Monitor, Application Insights, Virtual Networks, and Private Endpoints.
  • Strong observability experience: Prometheus, Grafana, centralized logging, alerting configuration, and distributed tracing instrumentation.
  • Working knowledge of SLO/SLI methodology: error budget principles, reliability target setting, and capacity planning.
  • Structured incident management experience: on-call ownership, blameless post-incident review, and runbook authorship.
  • Scripting and automation proficiency in Python or bash for toil elimination and operational tooling.
  • Strong CI/CD experience: GitHub Actions and ArgoCD or equivalent GitOps tooling.


Ways of Working

  • Comfortable in agile, iterative delivery environments with personal ownership and accountability for platform reliability.
  • Clear communicator across global, cross-functional stakeholders; able to translate technical reliability metrics into business impact for non-technical audiences.
  • Proactive learner with pragmatic adoption of AI-assisted developer tools (e.g., GitHub Copilot, Claude Code) to improve automation coverage and delivery velocity.


Nice to Have

  • Kubernetes certifications: CKA or CKAD.
  • Experience supporting AI or ML infrastructure workloads: GPU scheduling, model serving platforms, or inference pipeline operations.
  • Exposure to chaos engineering practices and fault injection testing.
  • FinOps experience: reserved capacity planning, resource right-sizing programs, and cost attribution per team or workload.
  • Service mesh experience (Istio, Linkerd) for traffic management and reliability patterns.
  • Experience in regulated industries (insurance, finance, healthcare) where auditability, change traceability, and secure-by-default operations are standard practice.


How We Hire

Allianz Partners does not accept unsolicited CV’s or approaches from agencies. We only work with partners on our approved supplier lists, under contract. Any unsolicited submission will not be considered.

What we offer


Our employees play an integral part in our success as a business. We appreciate that each of our employees are unique and have unique needs, ambitions and we enjoy being a part of their journey. We are there to empower and encourage you with your personal and professional development ensuring that you take control by offering a large variety of courses and targeted development programs.


All that in a global environment where international mobility and career progression are encouraged. Caring for your health and wellbeing is key priority for us. This is why we build Work Well programs to providing you with peace of mind and give the flexibility in planning and arranging for a better work-life balance.


90377 | Data & AI | Professional | Allianz Partners | Full-Time | Permanent

Allianz Group is one of the most trusted insurance and asset management companies in the world. Caring for our employees, their ambitions, dreams and challenges, is what makes us a unique employer. Together we can build an environment where everyone feels empowered and has the confidence to explore, to grow and to shape a better future for our customers and the world around us.

At Allianz, we stand for unity: we believe that a united world is a more prosperous world, and we are dedicated to consistently advocating for equal opportunities for all. And the foundation for this is our inclusive workplace, where people and performance both matter, and nurtures a culture grounded in integrity, fairness, inclusion and trust.

We therefore welcome applications regardless of ethnicity or cultural background, age, gender, nationality, religion, social class, disability or sexual orientation, or any other characteristics protected under applicable local laws and regulations.

Join us. Let's care for tomorrow.

Créer une alerte emploi pour cette recherche

Site Reliability Engineer (m/f/d) • Saint-Ouen, 93, FR

Offres similaires

Site Reliability Engineer - €65.000 - €70.000 Par An

Aston robinsonMontreuil, France, FR
Temps plein

Rejoins une équipe qui joue en Ligue 1 de l'infra - semaine de 4 jours.Divertissement en ligne | Fort trafic | AWS | Observabilité | Performance EngineeringEt si ton plus grand défi, c'étai... Voir plus

 • Offre sponsorisée

Service Reliability Engineer - €60.000 - €65.000 Par An

ITS ServicesNanterre, France, FR
Temps plein

Ingénieur SRE recherché pour mission longue chez un client grand compte.Objectifs communs de disponibilité et performance, automatisation des opérations, pilotage de l'activité, et garantie de ... Voir plus

 • Offre sponsorisée

Site Reliability Engineer - €72.200 - €83.100 Par An

Memo BankParis, France, FR
Temps plein

Ce poste de SRE vise à garantir la fiabilité, la sécurité et la performance de l'infrastructure de production de Memo Bank, tout en développant des outils pour améliorer la productivité des équ... Voir plus

 • Offre sponsorisée

Site Reliability Engineer Freelance H/F - €150 - €960 Par Jour

KickloxParis, France, FR
Temps plein

Ingénieur Site Reliability recherché pour analyser les dysfonctionnements, développer le code Infra As Code, assurer le monitoring et collaborer avec les équipes. Voir plus

 • Offre sponsorisée

Principal Site Reliability Engineer - Remote

A leading tech platform companyGentilly, France, FR
Temps plein

A leading AI technology firm is seeking experienced Site Reliability Engineers to enhance the reliability and scalability of its platform.The role involves collaboration with software engineers to ... Voir plus

 • Offre sponsorisée

Customer Reliability Engineer - €60.000 - €80.000 Par An

HirestoneParis, France, FR
Temps plein

Ingénieur SRE pour une scale-up en cybersécurité, responsable de la création de l'équipe CRE, de l'accompagnement technique des clients, du support, et de la rédaction de documentation. Voir plus

 • Offre sponsorisée

Senior Site Reliability Engineer – Devops & Automation - €60.000 - €65.000 Par An

Entreprise De Services InformatiquesParis, France, FR
Temps plein

Ingénieur SRE expérimenté recherché pour une entreprise de services informatiques à Montreuil.Responsable de la disponibilité et de la performance des services de production, avec 7 ans d'expér... Voir plus

 • Offre sponsorisée

Site Reliability Engineer – Cdi – Montpellier / Paris - €40.000 - €50.000 Par An

ACELYS SERVICES NUMERIQUESParis, France, FR
CDI

Le Site Reliability Engineer gérera l'automatisation, l'observabilité des solutions, l'architecture et l'industrialisation des systèmes. Voir plus

 • Offre sponsorisée

Site Reliability Engineer Confirmé·e - €48.000 - €58.000 Par An

ValeuriadParis, France, FR
Temps plein

Ingénieur SRE confirmé recherché pour déployer, administrer et monitorer des plateformes (Kubernetes, OpenShift).Automatisation des infrastructures (Terraform, Ansible) et maintien des outils d&#39... Voir plus

 • Offre sponsorisée

Senior Site Reliability Engineer - Remote - €52.000 - €65.000 A Year - Remote

RAISE FranceParis, France, FR
Temps plein

Experienced SRE sought to implement and maintain reliable, scalable systems.Collaboration with development and security teams is key. Voir plus

 • Offre sponsorisée

Alco2 Site Manager

AL CO2 EuropeMontrouge, France, FR
Temps plein

ALCO2 Site managerApplylocations:France, Ottmarsheimtime type:Full timeposted on:Posted Todayjob requisition id:R10095183**Entity Description**AL CO2 Europe SA, a 100% subsidiary of the Air Liquide... Voir plus

 • Offre sponsorisée

LAMINEUR/OPERATEUR LAMINAGE (H/F)

LEBRONZE ALLOYS - Ets BornelBornel, Oise, FR
Temps plein

LEBRONZE ALLOYS (LBA) est un leader de la production intégrée de matériaux de performance pour applications critiques.Nous investissons durablement depuis plus de 90 ans dans l'innovation pour répo... Voir plus

 • Offre sponsorisée

Site Reliability Engineer — Nantes | Flexible Hours - €45.000 - €60.000 Par An

Neo-Entreprise De Services NumériquesParis, France, FR
Temps plein

Une entreprise de services numériques cherche un Ingénieur Fiabilité de Site pour son agence à Nantes.Le poste requiert une expertise en gestion d'infrastructure, automatisation et optimisation... Voir plus

 • Offre sponsorisée

Systems Engineer - €6.717 A Month

OecdParis, France, FR
Temps plein

Design, build, operate, and support multi-service systems infrastructure, including Windows and Linux virtual machines, and Microsoft cloud services. Voir plus

 • Offre sponsorisée

Site Reliability Engineer - €45.000 - €60.000 Par An

YounupParis, France, FR
Temps plein

Ingénieur SRE recherché pour concevoir, déployer et automatiser des infrastructures CI/CD sur des environnements Linux, Windows et Cloud (AWS, Azure, GCP).Le candidat idéal possède au moins 5 ans d... Voir plus

 • Offre sponsorisée

Director, Site Reliability Engineering (X/F/M)

DoctolibMontrouge, France, FR
Temps plein

Since 2013, Doctolib has been on a mission to reinvent the daily life of health professionals and help people live healthier.With a team of 2,900 across France, Germany, Italy, and the Netherlands,... Voir plus

 • Offre sponsorisée

Senior Site Reliability Engineer - €60.000 - €80.000 A Year

LedgerParis, France, FR
Temps plein

The Senior Site Reliability Engineer will build and maintain digital asset platforms, automate processes, and improve system availability and efficiency. Voir plus

 • Offre sponsorisée

SRE - Site Reliability Engineer

NumberlyParis, Île-de-France, FR
Temps plein

Numberly est reconnu comme l’un des meilleurs spécialistes mondiaux du Data Marketing avec près de 500 collaborateurs et 11 bureaux dans le monde au service de plus de 300 clients de premier plan (... Voir plus

Founding Deployment Engineer - €50.000 - €100.000 A Year

LyskParis, France, FR
Temps plein

The Deployment Engineer will work with the founders to deploy prototypes with early customers, focusing on customer experience and technical enablement. Voir plus

 • Offre sponsorisée

Technicien Méthodes H/F/D

ESARIS INDUSTRIESCreil, France, FR
Temps plein

GROUPE ESARIS INDUSTRIES :Groupe français, nous concevons et fabriquons des composants, des systèmes électromécaniques, de connectique et de câblage.Notre spécialité est de conduire l’énergie élect... Voir plus