Senior Site Reliability Engineer
1 мес. назад
FranceEuropeSeniorHybrid
google cloud platformautomationsystem administrationmonitoringinfrastructure as code
Senior Site Reliability Engineer role responsible for reliability, automation, and security of GCP platform in a quantum computing startup.
О компании
- Alice & Bob is developing the first universal, fault-tolerant quantum computer to solve the world’s hardest problems. The quantum computer we envision building is based on a new kind of superconducting qubit: the Schrödinger cat qubit 🐈⬛ . In comparison to other superconducting platforms, cat qubits have the astonishing ability to implement quantum error correction autonomously! We're a diverse team of 250+ brilliant minds from over 35 countries united by a single goal: to revolutionise computing with a practical fault-tolerant quantum machine. Are you ready to take on unprecedented challenges and contribute to revolutionising technology? Join us, and let's shape the future of quantum computing together! About the role We're a fast-moving scale-up and our infrastructure needs
Обязанности
- Own the reliability of our production systems on GCP: define SLOs/SLIs, drive down incidents, and lead blameless postmortems.
- Build the observability stack (Prometheus, Grafana, and related tooling) so we catch problems before customers do.
- Design and maintain infrastructure as code with Terraform, no click-ops.
- Operate and scale our Kubernetes / GKE workloads.
- Build and harden CI/CD pipelines so teams can ship safely and often.
- Relentlessly automate manual work; if it's done twice by hand, it's a candidate for automation.
- Bake security into the platform: IAM, secrets management, network policy, and least-privilege by default.
- Help us meet compliance requirements as we mature, and make the secure path the easy path for other engineers.
- Own cloud cost visibility and efficiency: right-sizing, capacity planning, and scaling strategy as the company grows.
Требования
- 5+ years in infrastructure, SRE, DevOps, or platform engineering, including production ownership of cloud systems.
- Strong hands-on experience with GCP.
- Deep experience with Terraform.
- Production experience running Kubernetes / GKE.
- Proven track record building and maintaining CI/CD pipelines.
- Solid grasp of observability practices and tooling (Prometheus, Grafana, etc.).
- A bias toward automation and a security-conscious mindset.
- The communication skills to mentor others and influence how the team works.
- Experience scaling infrastructure in a startup or high-growth environment.
- Scripting/programming beyond config (Go, Python, etc.).
- Service mesh, GitOps (e.g., ArgoCD/Flux), or progressive delivery experience.
Как откликнуться
- Screening call with Doriane (30 min)
- Hiring Manager interview (45 min)
- Technical onsite Interview - System Design (60 min)
- Technical Interview - Problem Solving (60 min)
- Leadership Interview (30 min)
- Fit Interview (30 min)