Together AI logoT

Lead/Manager Site Reliability Engineering Team (Amsterdam)

Together AI

Amsterdam, North Holland, NetherlandsFull Time

Lead/Manager Site Reliability Engineering Team (Amsterdam) at Together AI is a full time role based in Amsterdam, North Holland, Netherlands. It was published on 1 April 2026 and was open at last check.

Lead/Manager Site Reliability Engineering Team (Amsterdam) at Together AI — key details
RoleLead/Manager Site Reliability Engineering Team (Amsterdam)
CompanyTogether AI
LocationAmsterdam, North Holland, Netherlands
Employment typeFull Time
Published1 April 2026
StatusOpen at last check

About the Role

Lead a team of Site Reliability Engineer (SRE) at Together based out of our office in Amsterdam, you and the SRE team are responsible for keeping all user-facing services and production systems running smoothly. You are a blend of a pragmatic operator and a software engineer that applies sound engineering principles, operational discipline, and mature automation to our operating environments and codebase.

You specialize in systems (operating systems, storage subsystems, networking), while implementing best practices for availability, reliability and scalability, with varied interests in algorithms and distributed systems.

Responsibilities

  • Be on an on-call (PagerDuty) rotation to respond to incidents that impact availability
  • Manage, develop and coach the SRE Team.
  • Build and run our infrastructure with Ansible, Terraform, and Kubernetes to enable scaling to a massive number of concurrent users
  • Build monitoring systems to ensure the highest quality service for our customers
  • Design and implement operational processes (such as deployments and upgrades)
  • Debug production issues across all services and levels of the stack
  • Identify improvements for the product architecture from the reliability, performance and availability perspectives
  • Plan the growth of Together AI’s infrastructure

Requirements

  • 7+ years of professional SRE or related experience
  • Ideally 2 years as a Lead SRE
  • Bachelor's degree in Computer Science or a related field or equivalent work experience
  • Expert knowledge of Ansible (roles, playbooks), Terraform, and Kubernetes
  • Proficiency in programming/scripting languages
  • Direct experience in monitoring and observability practices
  • Advanced knowledge of cloud services
  • Ability to thrive in a collaborative environment involving different stakeholders and subject matter experts

About Together AI

Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

Please see our privacy policy at https://www.together.ai/privacy

Share:WhatsAppLinkedIn

Create your free OnJob profile to apply — we'll take you to Together AI's application after sign-up. · Posted 1 Apr 2026.

Lead/Manager Site Reliability Engineering Team (Amsterdam) at Together AI — questions answered

What does the Lead/Manager Site Reliability Engineering Team (Amsterdam) role at Together AI pay?

Together AI does not publish a salary on this Lead/Manager Site Reliability Engineering Team (Amsterdam) listing, so OnJob shows no figure for it rather than an estimate. For what this role pays across the market, the OnJob salary guides aggregate the live listings that do disclose pay.

Where is the Lead/Manager Site Reliability Engineering Team (Amsterdam) role at Together AI based?

Together AI lists this Lead/Manager Site Reliability Engineering Team (Amsterdam) role in Amsterdam, North Holland, Netherlands, advertised as full time work at that location. Larger employers sometimes cover several sites under one city name, so confirm the exact office with Together AI before you apply.

Is the Lead/Manager Site Reliability Engineering Team (Amsterdam) role at Together AI still open?

The Lead/Manager Site Reliability Engineering Team (Amsterdam) posting at Together AI was open at OnJob's last check of the employer's careers page, having been published on 1 April 2026. OnJob re-checks source listings on each build and marks a role closed once it disappears, but listings can close without notice, so the employer's own page is the final word.

How do you apply for the Lead/Manager Site Reliability Engineering Team (Amsterdam) role at Together AI?

Apply to the Lead/Manager Site Reliability Engineering Team (Amsterdam) at Together AI role through OnJob with a free profile: OnJob scores your fit against the listing, shows the skills lowering that score, and submits an ATS-ready profile to Together AI's own application page. Creating a profile is free and needs no card.

Related Engineering jobs

Hand-picked roles that match this listing on skills, category and location — each scored to your profile inside OnJob.

Explore more on OnJob

Hiring for a role like this?

Post a job on OnJob and reach AI-matched candidates.

Post a Job