🇫🇷 Levallois-Perret, France · 1d ago

SRE /Platform Automation Engineer

ThoughtLabs Belgium

LinkedInseniorEnglish-friendly
Role: SRE / Platform Automation EngineerLocation: Levallois, France 🇫🇷Start Date: ASAPDuration: 3 years / Long-term assignmentEligibility: 🇫🇷 French Citizens Only🌍 Language RequirementsEnglish: Professional proficiency – MandatoryFrench: Professional proficiency – Preferred🎯 Role Overview:We are looking for an experienced SRE / Platform Automation Engineer to improve platform reliability, reduce operational workload through automation, and strengthen observability across large-scale cloud and infrastructure environments.The ideal candidate will work across cloud, Kubernetes, OpenStack, Linux, storage, networking, monitoring, security, and automation while driving operational excellence.🔧 Key Responsibilities:Automate recurring operational tasks, health checks, validation and remediation workflowsImprove platform reliability, resilience and operational efficiencyDesign and maintain monitoring, alerting and observability solutionsSupport production incidents, troubleshooting and root cause analysisDevelop automated remediation mechanismsAutomate KPI, SLA and service reportingIntegrate security and compliance controls into automationStandardize operational processes and change validationMaintain runbooks, SOPs and technical documentationDrive continuous improvement and reduce operational toilCollaborate with cloud, infrastructure, storage, network and security teams💻 Core Technical SkillsSRE & ReliabilitySite Reliability Engineering (SRE)Reliability Engineering PrinciplesSLOs & Error BudgetsPlatform ResilienceIncident Reduction & Service Reliability OptimizationAutomation & ScriptingPythonBash / Shell ScriptingAutomation FrameworksRemediation WorkflowsOperational Tooling DevelopmentInfrastructure as CodeTerraformAnsibleInfrastructure AutomationConfiguration ManagementCI/CD & DevOpsGitHubJenkinsCI/CD PipelinesDeployment ValidationVersion Control📊 Monitoring & ObservabilityPrometheusGrafanaOpenSearchOpenTelemetryLog CorrelationAlerting SystemsDashboards & Metrics Analysis☁️ Cloud & Platform TechnologiesOpenStackKubernetesGardenerCloud Infrastructure PlatformsPlatform Lifecycle ManagementLinux AdministrationGarden LinuxPerformance TroubleshootingHost-Level DiagnosticsSystem Automation🔐 Security & ComplianceSecurity AutomationSecrets ManagementCompliance ControlsSecure Automation PracticesAccess Control Principles🏢 IT Service ManagementServiceNowITIL FrameworkIncident ManagementChange ManagementProblem Management🎓 Required ExperienceMinimum 5+ years of experience in Infrastructure, Platform Engineering or Automation Engineering7+ years preferred in cloud operations, managed services, telecommunications, financial services or large enterprise environmentsStrong hands-on production operations experienceExperience in SLA-driven environmentsExperience with Incident, Change and Problem ManagementProven experience in automation, monitoring and operational improvementsComfortable with on-call and operational support activities🏆 CertificationsStrongly RecommendedITIL FoundationLinux Administration Certification or equivalent practical experiencePreferredCertified Kubernetes Administrator (CKA)HashiCorp Terraform AssociateRed Hat Certified Engineer (RHCE)OpenStack CertificationsPrometheus / Grafana Training

Sourced from LinkedIn. Relocantly aggregates public job postings; apply on the original site.