Ruthvik profile
available · open to roles

Ruthvik Garlapati

Site Reliability  ·  AIOps  ·  Infrastructure

SRE who has managed 50+ Kubernetes clusters at 99% uptime across AWS, Azure, and GCP — cut MTTR by 40% through observability automation, and built LLM-powered AIOps tools that resolve incidents before users notice.

TELEMETRY // SYSTEMLOG
50+ K8s Clusters Managed
99% Production Uptime
40% MTTR Reduction
26 Projects Built
projects github →
experience
Nestle Software Engineer · Contract Jan 2026 – Present · Seattle, WA
ArgoCDAzure DevOpsACR
SAP America Inc DevOps Engineer · Intern Jul 2024 – Sep 2025 · Chicago, IL
KubernetesGoOpenTelemetry
Northeastern University Graduate TA · Cloud Computing & RDBMS Sep 2023 – Apr 2024 · Boston, MA
AWSKubernetesTerraform
United Online Inc Software Quality Engineer Dec 2020 – Jul 2022 · Hyderabad, IN
JenkinsFastAPIPython
GreyCampus Edutech Software Engineer · Intern Jul 2019 – Apr 2020 · Hyderabad, IN
EC2CloudWatchAutoscaling
education
Northeastern University MS, Information Systems GPA 3.9 Boston, MA · May 2024
JNTU Hyderabad BTech, Computer Science   Hyderabad, India · Oct 2020
certifications
Certified Kubernetes Administrator Linux Foundation · CKA
AI Infrastructure & Operations NVIDIA Certified Associate
Claude Code in Action Anthropic
skills & expertise
Cloud & Infrastructure
Kubernetes Terraform AWS ArgoCD Docker Azure GCP Helm Istio Kyverno
Observability & SRE
Prometheus Grafana Incident Management OpenTelemetry ELK Stack SumoLogic DCGM
AI / LLM Engineering
Claude API RAG AWS Bedrock vLLM LangGraph MCP Triton Inference Google ADK
Languages
Python Bash Go Node.js Java C++
DevOps Practices
GitOps CI/CD Pipelines On-call SRE Canary Releases Chaos Engineering Runbook Automation
blog

latest writing

notes on software engineering, distributed systems, and real-time operations telemetry updates.

Connecting to Firestore log databases...