Anthropic2 дня назад
Staff+ Site Reliability Engineer, Safeguards ML Infra
33 750–40 417 ₽в месяц · до вычета
Remote-Friendly (Travel-Requ…
Обязанности
- 01Launch captain model releases: stand up, configure, and verify safeguards for every new model, and serve as the safeguards point of contact in the launch room during release windows.
- 02Own the off-cycle deployment of new safety classifiers as they ship from research — canarying rollouts, running post-deploy validations, and investigating discrepancies when something looks wrong.
- 03Verify that the right safeguards are provably live on the right models across every deployment platform (1P, AWS Bedrock, GCP Vertex, etc.), and detect and eliminate configuration drift between them.
- 04Automate yourself out of last quarter's work: turn launch runbooks into tooling, hand-built checks into continuous validation, and one-off deploys into a repeatable pipeline.
- 05Build and maintain a safeguards registry with full provenance — what is running in production, on which model, on which platform, and when and by whom it was deployed.
- 06Participate in on-call and operational-duty rotations covering service incidents, model provisioning, and time-sensitive research and safety launches.
Требования
- 01Have owned production change management at scale — deploy pipelines, config management systems, canary analysis — and have strong opinions about what "verified" means.
- 02Have run high-stakes releases: served as a launch captain, incident commander, or release owner for systems where a bad deploy has real consequences, and are energized rather than drained by being in the critical path.
- 03Have meaningful on-call experience for production systems, including incident response and postmortem-driven improvements — and a track record of turning (and fixing!) postmortem action items into process and tooling changes.
- 04Have a desire to close the gap where nobody has yet raised their hand, even if it requires manually hand-holding processes until automation and tooling can be built.
- 05Have hands-on experience deploying and operating on cloud platforms (AWS, GCP) at scale.
- 06Are proficient in Python; experience with Rust is a plus but not required.
- 07Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
- 08Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
- 09Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Условия
- 01Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
- 02Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.