Skip to main content
Global Expansion: US is live now! Expanding to UK, Canada, and Europe soon
SWE.Job
Zeta
ZetaHyderabadSenior Site Reliability Engineer

Senior Site Reliability Engineer

Hyderabad4+ years

About the role

The role involves managing large-scale distributed system software deployments in cloud or on-premise environments with a strong foundation in cloud management. The candidate should have experience working with observability tools like Prometheus and Grafana, as well as incident handling and debugging skills. Experience with AWS Cloud (preferred) including hands-on experience with AWS-CLI, orchestration and containerization like Kubernetes, containers, CI/CD practices, networking, Linux OS, and shell/python scripting is required.

Candidates will join an innovative team pushing engineering boundaries. Depending on the domains associated with this job, you will be expected to design clean APIs, write automated test cases, and participate in peer code reviews. We value developers who focus on performance optimization, fast load times, and simple, maintainable architectures.

Core Competency Domains

Sys AdminCloud ManagementUnix ShellPython & Go ProgrammingDatabase (MySQL or PostgreSQL)Team CollaborationDebugging and TroubleshootingDefining Standard Operating ProceduresObservability Tools (Prometheus, Grafana)Incident HandlingAWS Cloud Experience (preferred) including hands-on experience with AWS-CLIOrchestration and Containerization like Kubernetes, ContainersCI/CD (e.g., Jenkins, Argo CD)Networking (firewall, connectivity, routing, iptables, subnet config, etc.)Linux OS and Shell/Python ScriptingProgramming with Python, GoAPI Gateway like Kong, Nginx-based systemsSecurity Best Practices and Technologies
Source: Career PortalPosted on: Job ID: lev-6c5a8ebc-94b9-41f7-9a49-1659956343eb