Faster chat, better deals — Get the App

Site Reliability Engineer

Indeed

Company

Job typeFull-time
Workplace typeOnsite
Experience levelNo experience limit
Education levelNo degree limit

Description

Summary: Seeking a passionate and skilled SRE to enhance our infrastructure platform, driving transformation from a legacy service to a secure, self-service model while balancing operational integrity and automation. Highlights: 1. Drive infrastructure transformation and innovation at a digital-first company. 2. Balance operational stability with proactive automation to eliminate tech debt. 3. Collaborate on cutting-edge cloud and containerized solutions. ### **Do you want a job that matches your ambition and your lifestyle?** ### **Read on to discover the profile we're looking for!** From its beginnings as a wool mill, La Redoute is a French brand that has forged unique relationships with its customers for the past 180 years – yesterday through its catalogue, today on the web. We are a mainly digital player, selling fashion as well as home decoration. Since 2019, we have been modernizing our legacy systems and website, focusing on transitioning to cloud and containerized solutions. This transformation aims to enhance our scalability, flexibility, and overall performance, ensuring we stay at the forefront of technological advancements. This allows us to focus on flexible deployments and shortening the software development lifecycle, enabling us to deliver new features to production more quickly and efficiently. La Redoute is currently seeking a passionate and skilled SRE to contribute to and enhance our infrastructure platform, playing a pivotal role in our journey towards innovation and excellence. **About the Position** As part of La Redoute's Infrastructure Platform team, you will help drive our transformation from a legacy ticket\-queue service into a secure, self\-service platform model. This position requires a balance between maintaining operational integrity ("keeping the lights on") and driving proactive automation to eliminate tech debt, team friction and pain points. Your day\-to\-day will combine keeping critical production systems stable with driving meaningful improvements through automation. You will own the lifecycle of our cloud and on\-prem virtual machines, maintain, configure and provision computing resources, and enforce security posture by deploying endpoint controls (EDR) and managing user permissions, SSO integrations, and identity access across a multi\-cloud environment. You will work shoulder\-to\-shoulder with the security team to enforce compliance and drive down vulnerability remediation SLAs, while turning repetitive operational pain points into automated platform capabilities. **Responsibilities** * Collaborate closely with internal and external stakeholders, developers, and managers to ensure timely deliverables. * Develop, maintain, and scale configuration management and automated provisioning using Ansible across cloud virtual machines and hybrid environments. * Build, secure, and operate scalable infrastructure on public cloud platforms (predominantly Azure), managing compute instances, networking, and storage. * Partner tightly with the security team to identify vulnerabilities, execute mitigation strategies, install security components like EDR, and ensure compliance standards are met across all managed systems. * Administer and optimize identity management systems, including Azure Active Directory / EntraID, user permissions, and SSO configurations, ensuring secure and seamless access. * Write, review, and refactor IaC modules using Terraform within a structured Git repository model. * Automate provisioning, configuration and operational recovery with a reliability\-first mindset * Integrate robust monitoring and logging stacks to ensure high availability, troubleshoot complex infrastructure incidents, and participate in continuous service reliability operations. * Maintain cost\-awareness and resource tagging policies to support platform financial governance (FinOps). * Participate in the 24/7 rotational hotline to maintain system resilience. **Requirements** * Solid background as a Site Reliability Engineer, DevOps Engineer, or Systems Engineer managing production environments in the cloud. * Strong operational experience with public cloud platforms, particularly Microsoft Azure. * Experience with identity management frameworks (Azure Active Directory / EntraID, SSO, user permissions) and implementing security tools (EDR, vulnerability scanners). * Proficiency with **Terraform** for state management and modular infrastructure deployments and **Ansible** for automated configuration, patching, and server management. * Strong autonomy and critical thinking skills, with a focus on collaboration and ability to work in cross\-functional teams. * Excellent communication skills (written, listening, and speaking), a collaborative mindset, and a proactive attitude. * Fluency in English (mandatory). **Additional/Preferable Skills** * Experience with container orchestration platforms (Kubernetes / AKS). * Proficiency in CI/CD pipeline management using GitLab and GitLab CI/CD. * Familiarity with the Atlassian suite, particularly JSM (Jira Service Management) for incident, request, and workflow tracking. * Background in observability stacks (Prometheus, Grafana, Loki). * Relevant certifications such as Certified Azure Administrator (AZ\-104\).

Posted by

João Santos

Indeed · HR

Location

João Santos

Indeed · HR

Similar jobs

Site Reliability Engineer job by Indeed in 2026 | ok.com