Proton · Barcelona, Geneva · по договорённости
Join Proton and build a better internet where privacy is the default
Proton was founded in 2014 by scientists from CERN on a simple truth: privacy is a fundamental human right . Since then, we’ve built the world’s largest encrypted email service (Proton Mail) and expanded into Proton VPN, Proton Drive, Proton Pass, and Proton Calendar—tools used by millions globally to protect their freedom, fight censorship, and keep their data safe. In some situations, Proton has literally helped save lives!
We are profitable, independent (no VC control), and selectively hire from the top ~1% of applicants. Our 700+ team members across 50+ countries come from leading organizations and elite academic backgrounds. We move fast, keep hierarchy light, and prioritize impact over optics. If you want to do meaningful work with exceptionally high-caliber people, this is it. Join us and do work you can truly be proud of. Check our open-source projects here !
The Team:
Proton VPN gives millions of people a private connection to the internet, including people in countries where the internet is censored. It runs on thousands of bare-metal servers. We choose the hardware, the datacenters, and the networks ourselves, and we operate all of them.
The VPN Site Reliability Engineering team builds and runs this fleet. It has eight engineers: six site reliability engineers and two engineers who look after day-to-day fleet health. The fleet runs Debian, is automated with Ansible and Python, and is monitored with Prometheus, VictoriaMetrics, and Grafana. The team has kept the service stable through several years of fast growth.
The role:
We are hiring an engineering manager to lead this team. You will report to the VPN Engineering Director and work from our Geneva or Barcelona office.
The fleet keeps growing as more people use Proton VPN. To stay ahead of that growth, the team is investing in two areas, and you will lead both.
1. Automation. Engineers still do much of the provisioning, deployment, and repair work by hand. We want software to do that work, so that a fix or a new location reaches the fleet quickly and safely.
2. Observability. We know when a server is down. We do not yet know when users in one country are getting poor throughput. We want to find a problem that affects one server or one country while it is still small.
We chose a simple stack on purpose, and we plan to keep it and extend it.
This is a management role, so you will not write production code. You do need to be technical enough to read the team's code, question its designs, and examine its postmortems in detail.
What you will do:
Manage and develop a team of eight engineers.
Decide what the team automates and measures first, and deliver that plan.
Own the reliability of the Proton VPN fleet, including SLOs, error budgets, incident response, and postmortems.
Build monitoring that shows what users experience in each country, starting with throughput.
Replace manual provisioning, deployment, and repair work with tooling that ships changes to the fleet quickly and safely.
Set the engineering standards for the team, and coach engineers to solve operational problems with software.
Plan capacity and new locations with VPN product engineering, Proton's infrastructure teams, and our network and datacenter providers.
What we are looking for:
Several years of experience managing site reliability, infrastructure, or platform teams.
Experience leading a team from manual operations to automated ones, with results you can describe.
Strong technical judgment on large Linux fleets. You can review a design, read Python and Ansible, and find the gaps in a postmortem.
Experience owning the reliability of a production service, including SLOs, error budgets, incident response, and postmortems.
Experience building or rebuilding the monitoring of a production service, and getting engineers to rely on it.
Experience running physical infrastructure, or a clea
Отклик ведёт на сайт работодателя. Бесплатная регистрация открывает отклик и разбор резюме.