Java Backend Developer | Site Reliability Engineer (SRE) | Spring Boot | Docker | Kubernetes | Observability
I'm a Java Backend Developer working across Site Reliability Engineering (SRE) and production support β monitoring system health, resolving incidents, and building observability into scalable backend systems.
I enjoy designing clean backend architectures, keeping systems reliable in production, and solving real-world engineering problems at the intersection of development and operations.
- π Monitoring system health and proactively resolving incidents across Kubernetes clusters and virtual servers
- ποΈ Monitoring and clearing WAL (Write-Ahead Log) directories for services to prevent disk pressure and maintain uptime
- π‘ Monitoring application and infrastructure health using Grafana, Prometheus, and the LGTM stack across Docker and Kubernetes
- π‘οΈ Managing production incidents through ServiceNow, ensuring SLA compliance
- π₯ Onboarding new applications into monitoring platforms, tracked via Jira
- π Building and maintaining alerting rules across Email, Webhook, and ServiceNow integrations
- π©Ί Implementing synthetic monitoring checks for proactive outage detection
- π Performing Root Cause Analysis (RCA) and collaborating directly with clients to optimize PromQL queries causing pod failures or system breakdowns from long-running queries
- π Hyderabad, India


