Open role
Site Reliability Engineer
NVIDIA
India, BengaluruPosted Aug 20, 2026 · 3h ago
Full-timeEntry LevelOn-siteSaaS
About this role
NVIDIA is seeking a Site Reliability Engineer to support and enhance the reliability, scalability, and efficiency of their enterprise systems. You will contribute to building and maintaining distributed systems, automating database operations, and improving system observability. This role involves participating in incident response and collaborating with various teams to implement SRE best practices, with opportunities to explore AI-assisted engineering.
What we are looking for
6- Support SRE initiatives for reliability and scalability
- Build and maintain distributed systems for AI enterprise products
- Automate database operations for relational and vector databases
- Enhance observability through dashboards, alerts, and automation
- Participate in incident response and contribute to MTTR reduction
- Collaborate with Cloud, Platform, Security, and AI/ML teams
Skills mentioned
13PythonJavaScriptTypeScriptSQLAWSAzureGCPDockerKubernetesLinuxPostgreSQLTerraformCI/CD
