Skip to content
DevOps / InfrastructureActive

SRE — Site Reliability Engineer

Ensure the reliability, availability and performance of our infrastructure and solutions.

100 % Remote
CDI
Confirmé
DevOps / Infrastructure
The context

OFFZONE operates critical solutions for industry and healthcare. The SRE works at the boundary between development and operations to ensure every service remains reliable and performant as our systems grow.

The mission

You will implement SRE practices on our projects: monitoring, alerting, incident management, automation. You will work directly with development and DevOps teams.

Your responsibilities
  • Define and track SLO/SLIs for critical services
  • Set up proactive monitoring and alerting
  • Automate responses to recurring incidents
  • Conduct post-mortem analysis and propose improvements
  • Optimize service performance and scalability
  • Collaborate with developers to improve code resilience
Tech environment
  • Docker
  • Linux
  • Node.js
  • Git
Required skills
  • Experience managing production infrastructure
  • Proficiency in monitoring and observability (metrics, logs, traces)
  • Knowledge of Docker and containerization
  • Advanced Linux system administration
  • Ability to automate and script (Bash, Python)
Profile we're looking for
  • 3 to 6 years of SRE, DevOps or system administration experience
  • Methodical and responsive during incidents
  • Focused on continuous improvement
  • Comfortable collaborating with development teams
Nice to have
  • Experience with a cloud provider (AWS, GCP or Azure)
  • Knowledge of Kubernetes
  • Familiarity with chaos engineering tools
  • Cloud certifications appreciated
Why join us

High-impact projects in industry, healthcare and digital transformation

Autonomy and ownership on your deliverables

Demanding and constantly evolving technical environment

Multidisciplinary team where everyone genuinely contributes

Continuous skill growth and knowledge sharing

Multi-sector, multi-country context (Morocco, France, Malaysia)

Hiring process
  1. Application received and reviewed by the team

  2. Discovery call (30 min, remote or on-site)

  3. Technical assessment tailored to the role

  4. Interview with the project team

  5. Offer and onboarding

Apply to this position

SRE — Site Reliability Engineer

Apply

A few details so we can handle your application properly.