Here's the job offer description formatted in Markdown, adhering to your specifications:
🇪🇺 Our Story
Join Scaleway and shape the sovereign cloud of tomorrow!
Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.
Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.
With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.
Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customers across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.
📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
🚀 Why We Need You?
Our growth is driving us to strengthen our
Platform SRE team to build, standardize, and reliable the infrastructure hosting Scaleway's products.
Your mission will be to ensure the operational readiness of our infrastructures and the onboarding of product teams in order to maintain high-performance standards, ensure continuous improvement, and support the deployment of product stacks across new regions.
🤝 Your Future Team
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together.
You will be part of a team of 5 people, including your manager. The team operates in a stabilized environment with a fresh dynamic, focusing on onboarding multiple products and uniformizing engineering practices. We use a mix of Scrum and Kanban methodologies and rotate product referents to keep the work engaging and mitigate "bus factor" risks.
🗓️ Your Daily Routine
Tasks* Build, standardize, and enhance the reliability of the platform infrastructure.
* Onboard product teams and facilitate the deployment of product stacks in new regions.
* Implement and manage observability tools (Grafana, Thanos, Alertmanager).
* Automate infrastructure deployment and management using Gitops processes.
* Ensure configuration consistency through tools like Ansible or Salt.
* Manage operational maintenance (MCO) and handle production incidents.
* Define and monitor reliability metrics such as SLAs and SLOs.
* Maintain clear and accurate technical documentation.
* Manage security components and secrets (e.g., HashiCorp Vault).
* Participate in a weekly on-call rotati