Filter by
By date
By section

0 results for ""

    House image

    09/03/2026

    KERING Cloud Engineer SRE

    Kering - Regular
    Paris - France

    Groupe de luxe mondial né d’une histoire familiale et entrepreneuriale, Kering réunit un ensemble de Maisons reconnues pour leur créativité en matière de couture et de prêt-à porter, de maroquinerie, de joaillerie, de lunetterie et de beauté : Gucci, Saint Laurent, Bottega Veneta, Balenciaga, McQueen, Brioni, Boucheron, Pomellato, Dodo, Qeelin, Ginori 1735, ainsi que Kering Eyewear. Inspirées par leur histoire et par leur patrimoine, les Maisons du Groupe conçoivent et façonnent des produits et des expériences d'exception qui reflètent l’engagement de Kering pour l’excellence, le développement durable et la culture. Cette vision s’incarne dans la signature du Groupe : Creativity is our Legacy. Comptant 44 000 collaborateurs, Kering a réalisé un chiffre d’affaires de 14,7 milliards d’euros en 2025.

    Au sein de la Direction Tech & Digital de Kering, le DevSecOps Studio conçoit, opère et fiabilise les plateformes et produits digitaux du Groupe. Le Studio recrute un Cloud Engineer (SRE) afin de renforcer la fiabilité de ses environnements de production et d’accompagner le passage à l’échelle de ses pratiques d’ingénierie augmentée par l’IA.

    Nous sommes actuellement à la recherche d’un(e) Cloud Engineer (SRE – Site Reliability Engineering) (H/F) pour intégrer l'équipe.

    Votre opportunité


    Au sein du DevSecOps Studio, vous contribuez au développement et à la résilience de produits digitaux à grande échelle selon un modèle AI-SDLC, où l’ingénierie augmentée par l’IA constitue la méthode de travail de référence et non un outil optionnel : de la discovery technique à la mise en production, en passant par la génération de code, les tests et la documentation.

    Comment vous allez contribuer


    Infrastructure Cloud et Kubernetes. Vous exploitez et optimisez l’infrastructure Cloud, administrez et maintenez les clusters Kubernetes de production. Vous définissez et suivez les SLOs et error budgets des services critiques, et contribuez à l’amélioration de la conception et de l’architecture de l’infrastructure.

    IaC, FinOps et automatisation. Vous concevez, construisez et maintenez l’infrastructure avec Terraform et Packer, et automatisez les workflows via des pipelines CI/CD (GitHub Actions). Vous pilotez l’optimisation des coûts cloud en collaboration avec l’équipe FinOps : suivi des dépenses, identification des gaspillages et mise en œuvre de recommandations mesurables. Vous utilisez les outils d’IA assistée (GitHub Copilot, Claude Code) comme pratique standard dans la production de code Terraform, de runbooks et de pipelines CI/CD, conformément à la politique d’usage des outils IA du Studio.

    Observabilité, résilience et chaos engineering. Vous maintenez et améliorez la stack d’observabilité (ELK, Grafana, APM, distributed tracing) et garantissez la qualité de la surveillance, des alertes et de la visibilité système. Vous implémentez des pratiques de chaos engineering pour tester et renforcer la résilience de la plateforme de manière proactive. Vous transformez les incidents en améliorations durables via des post-mortems structurés et pilotez la réduction du toil par l’automatisation.

    Fiabilité des workloads IA. Vous définissez et suivez des SLOs spécifiques aux services IA : latence LLM, taux d’erreur des agents, disponibilité du LLM Gateway (Galassia). Vous implémentez l’observabilité des pipelines agentiques, du tracing des appels MCP au monitoring de la consommation de tokens et à la détection des dérives de comportement. Vous contribuez à la résilience des Discovery et Run Squads en fournissant les outils de monitoring nécessaires à chaque incrément AI-SDLC.

    IA et excellence ingénierie. Vous intégrez des outils d’IA et d’automatisation intelligente dans les workflows SRE : détection d’anomalies, réponse aux incidents assistée par IA, génération de runbooks. Vous concevez et opérez des agents SRE dédiés au triage d’incidents, à la génération automatisée de runbooks et à l’analyse de root cause. Vous promouvez l’IaC, l’automatisation et les bonnes pratiques DevSecOps, et accompagnez les équipes dans l’adoption des outils et pratiques de la plateforme.

    Qui êtes-vous


    Vous justifiez de 5 à 8 ans d’expérience en ingénierie infrastructure ou plateforme, dont au moins 2 à 3 ans sur des environnements de production à fort trafic, en tant que SRE, Platform Engineer ou DevOps Engineer.

    Vous maîtrisez AWS (API Gateway, CloudFront, ECS, EKS, ALB), Terraform à un niveau avancé et Kubernetes en production, ainsi qu’Ansible et GitHub Actions. Vous êtes à l’aise avec la stack d’observabilité ELK et Grafana, avec Linux et les fondamentaux réseau (TCP/IP), et avec les concepts de SLO, SLI, error budget et FinOps. Une connaissance multi-cloud est appréciée.

    Vous pratiquez effectivement les outils AI-SDLC (Copilot, Claude Code ou équivalent) dans votre quotidien d’ingénierie et savez concevoir des automatisations SRE assistées par LLM, du triage à la génération de runbooks et de post-mortems, sur une plateforme LLM Gateway.

    Vous accompagnez les équipes de développement, expliquez vos choix techniques et vulgarisez des problématiques complexes. Vous influencez sans imposer et partagez les bonnes pratiques de fiabilité à l’échelle.

    Vous raisonnez en métriques d’impact : disponibilité, MTTR, taux d’erreur, économies FinOps, réduction du toil. Vous avez mené des transformations de fiabilité avec des améliorations quantifiées et transformez les incidents en apprentissages durables.

    Vous êtes rigoureux(se), autonome et curieux(se) techniquement. Vous savez intervenir sous pression et progresser dans des environnements en évolution rapide.

    Vous maîtrisez le français et l’anglais, à l’écrit comme à l’oral.

    Pourquoi nous rejoindre ?


    Vous intégrerez un Studio qui expérimente en production les pratiques AI-SDLC les plus avancées du secteur : LLM Gateway self-hosted, Discovery Squads, agents MCP. Pour un ingénieur fiabilité avec une appétence pour l’IA, c’est un terrain d’expérimentation rare dans un contexte Groupe exigeant.

    Kering s'engage en faveur de la diversité. Nous croyons que la diversité sous toutes ses formes - genre, âge, nationalité, culture, handicap, croyances religieuses et orientation sexuelle - enrichit le lieu de travail. Nos collaborateurs ont ainsi des opportunités pour exprimer leurs talents, à la fois individuellement et collectivement, et cela contribue à renforcer notre capacité d'adaptation à un monde en mutation. En tant qu'employeur de l'égalité des chances, nous accueillons et considérons les candidatures de tous les candidats qualifiés, indépendamment de leurs antécédents.
     

    English Version

    Kering is a global, family-led luxury group, home to people whose passion and expertise nurture creative Houses across couture and ready-to-wear, leather goods, jewelry, eyewear and beauty: Gucci, Saint Laurent, Bottega Veneta, Balenciaga, McQueen, Brioni, Boucheron, Pomellato, Dodo, Qeelin, Ginori 1735, as well as Kering Eyewear. Inspired by their creative heritage, Kering Houses design and craft exceptional products and experiences that reflect the Group’s commitment to excellence, sustainability and culture. This vision is expressed in our signature: Creativity is our Legacy. In 2025, Kering employed 44,000 people and generated revenue of €14.7 billion. 

    Within Kering’s Tech & Digital organization, the DevSecOps Studio designs, operates and safeguards the reliability of the Group’s digital platforms and products. The Studio is recruiting a Cloud Engineer (SRE) to strengthen the reliability of its production environments and to support the scaling of its AI-augmented engineering practices.

    We are currently seeking a Cloud Engineer (SRE – Site Reliability Engineering) (M/F) to join our team.

    Your opportunity


    Within the DevSecOps Studio, you contribute to the development and the resilience of large-scale digital products under an AI-SDLC model, where AI-augmented engineering is the standard way of working rather than an optional tool: from technical discovery to production release, including code generation, testing and documentation.

    How you will contribute


    Cloud infrastructure and Kubernetes. You operate and optimize the Cloud infrastructure, and administer and maintain production Kubernetes clusters. You define and track SLOs and error budgets for critical services, and contribute to improving infrastructure design and architecture.

    IaC, FinOps and automation. You design, build and maintain infrastructure with Terraform and Packer, and automate workflows through CI/CD pipelines (GitHub Actions). You drive cloud cost optimization together with the FinOps team: spend tracking, identification of waste and implementation of measurable recommendations. You use AI-assisted tooling (GitHub Copilot, Claude Code) as standard practice when producing Terraform code, runbooks and CI/CD pipelines, in line with the Studio’s AI tool usage policy.

    Observability, resilience and chaos engineering. You maintain and improve the observability stack (ELK, Grafana, APM, distributed tracing) and ensure effective monitoring, alerting and system visibility. You implement chaos engineering practices to proactively test and strengthen platform resilience. You turn incidents into lasting improvements through structured post-mortems and drive toil reduction through automation.

    Reliability of AI workloads. You define and track SLOs specific to AI services: LLM latency, agent error rate, availability of the LLM Gateway (Galassia). You implement observability for agentic pipelines, from tracing MCP calls to token consumption monitoring and behavioral drift detection. You contribute to the resilience of the Discovery and Run Squads by providing the monitoring tools required at each AI-SDLC increment.

    AI and engineering excellence. You embed AI and intelligent automation tools into SRE workflows: anomaly detection, AI-assisted incident response, runbook generation. You design and operate SRE agents for incident triage, automated runbook generation and assisted root cause analysis. You promote IaC, automation and DevSecOps best practices, and support teams in adopting the platform’s tools and practices.

    Who you are


    You have 5 to 8 years of experience in infrastructure or platform engineering, including at least 2 to 3 years on high-traffic production environments, as an SRE, Platform Engineer or DevOps Engineer.

    You have a strong command of AWS (API Gateway, CloudFront, ECS, EKS, ALB), advanced Terraform and Kubernetes in production, as well as Ansible and GitHub Actions. You are comfortable with the ELK and Grafana observability stack, with Linux and networking fundamentals (TCP/IP), and with SLO, SLI, error budget and FinOps concepts. Multi-cloud knowledge is a plus.

    You actively use AI-SDLC tooling (Copilot, Claude Code or equivalent) in your day-to-day engineering work and are able to design LLM-assisted SRE automations, from triage to runbook and post-mortem generation, on an LLM Gateway platform.

    You support development teams, explain your technical choices and make complex topics accessible. You influence without imposing and share reliability best practices at scale.

    You think in impact metrics: availability, MTTR, error rate, FinOps savings, toil reduction. You have led reliability transformations with quantified improvements and turn incidents into lasting learnings.

    You are rigorous, autonomous and technically curious. You can perform under pressure and thrive in fast-moving environments.

    You have full proficiency in French and English, both written and spoken.

    Why work with us?


    You will join a Studio that runs some of the most advanced AI-SDLC practices in the industry in production: a self-hosted LLM Gateway, Discovery Squads and MCP agents. For a reliability engineer with an appetite for AI, this is a rare playing field within a demanding Group context.

    Kering is committed to diversity and inclusion and to providing equal opportunities in employment. We believe diversity in all its forms – disability, color, ancestry, sex, national origin, sexual orientation, age, citizenship, marital status, gender identity, religion – enriches the workplace. It opens opportunities for people to express their talent, both individually and collectively and it helps foster our ability to adapt to a changing world. As an Equal Opportunity Employer, we welcome and consider applications from all qualified candidates, regardless of their background.

    Similar jobs