Skit.ai logoS

Site Reliability Engineer — Multi-Cloud Infrastructure

Skit.ai

Bangalore, KA, IndiaFull Time

Site Reliability Engineer — Multi-Cloud Infrastructure at Skit.ai is a full time role based in Bangalore, KA, India. It was published on 5 July 2026 and was open at last check.

Site Reliability Engineer — Multi-Cloud Infrastructure at Skit.ai — key details
RoleSite Reliability Engineer — Multi-Cloud Infrastructure
CompanySkit.ai
LocationBangalore, KA, India
Employment typeFull Time
Published5 July 2026
StatusOpen at last check

About the Role

Skit.ai is the pioneer Conversational AI company transforming collections with omnichannel GenAI-powered assistants. Skit.ai’s Collection Orchestration Platform, the world’s first solution, streamlines collection conversations by syncing channels and accounts. Skit.ai’s Large Collection Model (LCM), a collection LLM, powers the strategy engine to optimize interactions, enhance customer experiences, and boost bottom lines for enterprises. Skit.ai has received several awards and recognitions, including the BIG AI Excellence Award 2024, Stevie Gold Winner 2023 for Most Innovative Company by The International Business Awards, and Disruptive Technology of the Year 2022 by CCW. Skit.ai is headquartered in New York City, NY. Visit https://skit.ai/

Job Title: Site Reliability Engineer — Multi-Cloud Infrastructure

Type: Full-time

Location: Bangalore

Why this role exists:

We run a voice AI platform for regulated enterprises in banking, telecom, and collections, spread across AWS, GCP, and Azure — for resilience, for cost, and because client data-residency rules leave us no choice. That's a lot of surface area: compute, networking, storage, identity, clusters, pipelines, and supporting services, all needing to stay healthy across three providers.

This role owns the day-to-day reliability and operations of that estate. It's the generalist counterpart to our real-time-platform SRE: where they go deep on the latency-critical call path, you go broad — keeping the whole infrastructure dependable, well-automated, and cost-sane, and sharing the on-call load. If you like knowing how everything fits together and making the boring parts reliable and self-serve, this is a good seat.

What you'll own:

  • Multi-cloud operations. Provision, operate, and keep healthy compute, networking, storage, and identity across AWS, GCP, and Azure — with sensible consistency instead of three snowflakes.
  • Infrastructure as code. Manage the estate through Terraform (or equivalent) and version control — reproducible environments, reviewed changes, no undocumented hand-tweaks.
  • CI/CD and delivery. Keep build and deploy pipelines fast and reliable so engineers ship safely and often.
  • Clusters and workloads. Run Kubernetes/container platforms and the supporting services (databases, queues, caches, internal tooling) that everything depends on.
  • Monitoring and on-call. Maintain monitoring and alerting for infrastructure health, take a turn in the rotation, and respond to and mitigate incidents with clear communication and blameless follow-up.
  • Cost and hygiene. Keep an eye on cloud spend, rightsizing, and waste; own the unglamorous but essential hygiene — patching, backups, secrets, and access.
  • Automation and toil reduction. Replace manual, repetitive operations with automation and self-service so the team scales without headcount scaling with it.

What the first year looks like:

  • First 90 days. Learn the estate across all three clouds. Take a turn on call. Close the most obvious gaps in monitoring, backups, and access hygiene.
  • By 6 months. More of the estate under consistent infrastructure-as-code. Reliable, reviewed CI/CD. A clearer, quieter alerting setup and documented runbooks for the common incidents.
  • By 12 months. Measurably less manual toil through automation and self-service. Sensible cost controls in place. Provisioning and environment setup that's repeatable rather than tribal knowledge.

What we're looking for:

Must-have

  • A few years in SRE, DevOps, or infrastructure operations for production systems, including on-call.
  • Hands-on experience across at least two of AWS, GCP, and Azure (all three is a strong plus).
  • Kubernetes and containers in production.
  • Infrastructure-as-code (Terraform or similar) and CI/CD pipelines.
  • Monitoring and alerting practice (e.g. Prometheus/Grafana) and structured incident handling.
  • A scripting/programming language for automation (Python, Go, or Bash beyond one-liners).
  • Solid Linux systems and networking fundamentals.

Nice-to-have

  • All three clouds run in production, and comfort designing for consistency across them.
  • Cost optimization / FinOps.
  • Secrets management, security hardening, and compliance/data-residency contexts.
  • PostgreSQL and other stateful-service operations at scale.
  • Some exposure to real-time or voice infrastructure — enough to back up the platform SRE on call.

Our Stack:

Representative — you'll help shape it. Multi-cloud across AWS, GCP, and Azure; Kubernetes/containers; Terraform and GitHub Actions CI/CD; PostgreSQL; Grafana/Tempo for monitoring; Modal for ML deployment; LiveKit/SIP telephony on the platform side.

How you'll know you're succeeding:

The infrastructure just works, across all three clouds, and when it doesn't it's caught early and fixed cleanly. Engineers provision what they need without filing tickets. Cloud spend is understood, not surprising. And the on-call rotation trends calmer because the estate is increasingly automated and self-healing.

We're an equal-opportunity employer and evaluate every candidate on merit. [Add benefits, compensation band, and application instructions before posting.]

Share:WhatsAppLinkedIn

Create your free OnJob profile to apply — we'll take you to Skit.ai's application after sign-up. · Posted 5 Jul 2026.

Site Reliability Engineer — Multi-Cloud Infrastructure at Skit.ai — questions answered

What does the Site Reliability Engineer — Multi-Cloud Infrastructure role at Skit.ai pay?

Skit.ai does not publish a salary on this Site Reliability Engineer — Multi-Cloud Infrastructure listing, so OnJob shows no figure for it rather than an estimate. For what this role pays across the market, the OnJob salary guides aggregate the live listings that do disclose pay.

Where is the Site Reliability Engineer — Multi-Cloud Infrastructure role at Skit.ai based?

Skit.ai lists this Site Reliability Engineer — Multi-Cloud Infrastructure role in Bangalore, KA, India, advertised as full time work at that location. Larger employers sometimes cover several sites under one city name, so confirm the exact office with Skit.ai before you apply.

Is the Site Reliability Engineer — Multi-Cloud Infrastructure role at Skit.ai still open?

The Site Reliability Engineer — Multi-Cloud Infrastructure posting at Skit.ai was open at OnJob's last check of the employer's careers page, having been published on 5 July 2026. OnJob re-checks source listings on each build and marks a role closed once it disappears, but listings can close without notice, so the employer's own page is the final word.

How do you apply for the Site Reliability Engineer — Multi-Cloud Infrastructure role at Skit.ai?

Apply to the Site Reliability Engineer — Multi-Cloud Infrastructure at Skit.ai role through OnJob with a free profile: OnJob scores your fit against the listing, shows the skills lowering that score, and submits an ATS-ready profile to Skit.ai's own application page. Creating a profile is free and needs no card.

Related Engineering jobs

Hand-picked roles that match this listing on skills, category and location — each scored to your profile inside OnJob.

Explore more on OnJob

Hiring for a role like this?

Post a job on OnJob and reach AI-matched candidates.

Post a Job