RapidAI is hiring for the profile of Senior Site Reliability Engineer . Graduates are eligible to apply for this position. This profile is open for the location of Bengaluru . Complete information about the hiring is mentioned below:
Job Overview
| Company Name | RapidAI |
|---|---|
| Profile Hiring for | Senior Site Reliability Engineer |
| Salary | As per Market Standards |
| Work Profile | Work From Office |
| Eligibility | Experienced Jobs |
| Location | Bengaluru |
| Job Category | Software Engineering & QA |
| Sub Category | DevOps |
| Job ID | NX11515 |
RapidAI is the trusted leader in deep clinical AI, helping hospitals deliver faster, more informed care through intelligent imaging and integrated workflows. The Rapid Enterprise™ Platform supports disease states across the care spectrum, but it’s our clinical depth that drives the most meaningful impact — improving decision-making, patient outcomes, and health-system performance. Used by more than 2,500 hospitals in over 100 countries and backed by 700+ clinical studies, including research that helped expand national stroke-treatment guidelines, RapidAI is the most clinically validated AI platform in healthcare.
- Own the availability, performance, and incident response for Rapid’s production EKS clusters
- Design and operate the full observability stack — metrics, logs, traces — with
Open Telemetry as the foundation - Define and track SLOs/SLIs/error budgets; lead post-mortems and drive blameless culture
- Build and maintain infrastructure-as-code using Terraform, Helm, and GitOps patterns
- Partner with engineering to bake reliability in early — capacity planning, load testing, chaos engineering
- Tune autoscaling, networking, and cost efficiency across AWS workloads
- On-call rotation with the expectation you’ll also fix the underlying cause, not just the alert
- 10+ years in SRE, DevOps, or infrastructure engineering roles
- Deep AWS expertise — EKS, EC2, VPC, IAM, RDS, S3, CloudWatch, and the
surrounding ecosystem - Production Kubernetes experience at scale: multi-cluster, multi-tenant, real traffic
- Hands-on Open Telemetry instrumentation and pipeline ownership (collectors, exporters, backends)
- Strong foundation in Linux, networking, and distributed systems fundamentals
- Experience with observability platforms (Prometheus, Grafana, Jaeger, or equivalents)
Comfortable writing automation in Go, Python, or Bash — you reach for code when the GUI runs out - Startup mindset: you make decisions with incomplete information and iterate quickly
What You Do:
What We Looking For:
RapidAI is committed to creating an inclusive and diverse workplace. We provide equal employment opportunities to all employees and applicants and prohibit discrimination and harassment of any type in regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
To apply for this job please visit jobs.lever.co.
How to Apply for RapidAI Recruitment
To apply for this job, interested candidates must follow the procedure outlined below:
Click Apply Here or Employer Contact Details. Both buttons will take you to the official application page.
1. Never pay any amount for getting a job. 2. Apply before the hiring closes or the employer stops accepting applications.
Explore More Jobs
By Role
By Location & Experience
More Options
Similar Jobs
Explore other jobs that may match your profile.
More Jobs From RapidAI
Explore other current openings from this company.
Nexpro247 shares job information for informational purposes only and is not affiliated with or endorsed by the companies mentioned unless explicitly stated. Job details may change without notice. Candidates should verify the vacancy, eligibility and application details on the employer's official website before applying. Nexpro247 does not charge candidates any fee for job applications.
