Runs on-prem, not just the cloud
I build systems that live in a data center or at the edge — reproducible provisioning, containers, and high availability, air-gapped when it has to be.
Site Reliability & Platform Engineer · Open to Junior SRE roles · UPES '27
I build and operate systems that stay up — and when they don't, I find out why.
I operate a production system where downtime is a safety defect — designing for failure, monitoring it, and root-causing incidents under pressure. Comfortable across Linux, cloud, containers, and the networking beneath them.
My work maps to how infrastructure actually runs — on-prem and hybrid, operated for uptime, over enterprise networks.
I build systems that live in a data center or at the edge — reproducible provisioning, containers, and high availability, air-gapped when it has to be.
Health checks, load testing, incident root-cause analysis, and runbooks. I treat reliability as the job, not an afterthought — and I harden the pipeline (OWASP, SAST/DAST) on the way.
I built UPES-ECS on Juniper SRX, EX, and Mist — now part of HPE Networking — so I've configured the fabric, not just consumed it.
I'm a systems & platform engineer (UPES CS '27) happiest close to the infrastructure — Linux, containers, the deploy pipeline, and the unglamorous reliability work that keeps a system up. I like owning things in production: health checks, deploys, monitoring, and the runbook that gets you through the incident.
That started on my own hardware — Linux compiled from source, Proxmox, a home lab I actually run — and now extends from a Software Engineer internship (FastAPI on AWS, security audits, CI/CD) to a resilient, offline-first emergency communication system I designed and operate on enterprise networking gear. Along the way: a HACKIITK finalist project at IIT Kanpur and leading the technical wing of my campus developer community.
Weighted toward the systems & reliability layer, with real breadth across cloud, backend, networking, and security.
A selection from 74 public repositories — a production emergency system, infrastructure tooling, backends, and security/AI work.
Resilient, offline-first campus emergency communication system on an HPE + Juniper networking backbone — designed, deployed, and operated end-to-end: active/standby HA, health checks, load testing, and incident response. Showcased live to Team HPE.
Scripted, unattended Proxmox provisioning that cut setup time by ~80%; Dockerized so the Debian-only installer runs on Arch and other non-Debian hosts — reproducible infrastructure from a config file.
Offline, AI-powered policy gap analyzer mapping security policies to NIST CSF 2.0 & CIS. LangChain pipeline, Electron desktop, 702-test suite (97.7% pass). Top ~0.4% of 1,300+ teams at IIT Kanpur.
Containerized personal-knowledge backend: secure REST APIs with OAuth2 + RBAC, Redis caching, and optimized schemas — Dockerized for consistent CI/CD deployments.
Declarative, reproducible machine configuration with Nix flakes + home-manager — infrastructure-as-code for my own systems, rebuildable from scratch.
Android KernelSU build automation across devices — low-level system configuration with SDLC practices for versioning and testing.
Building is table stakes. Operating UPES-ECS end to end — where a dropped call is a safety defect, not a bug — taught me to find the real root cause fast. Three from the logbook.
Building Python/FastAPI backend services on AWS (EC2, ECS, S3) and wiring automated security checks into CI/CD. Running application security audits — manual review plus SAST/DAST against the OWASP Top 10 — and contributing to WCAG 2.0 accessibility.
Led and delivered technical workshops for 100+ participants, managed cross-functional teams on production-grade initiatives, and mentored students in cloud, system design, and full-stack development.
Onboarded 10+ students to Linux (distro selection, dual-boot, driver/hardware troubleshooting), wrote shell scripts to automate post-install configuration, and mentored peers moving from Windows to Linux as a daily driver.
Delivered digital-safety and social-media-safety sessions across three official welfare programs, adapting technical security concepts for non-technical audiences. Recognized with three Certificates of Appreciation.
UPES-ECS — a resilient, offline-first campus emergency communication system on an HPE + Juniper networking backbone — was presented live to Team HPE. Full call flow, operations dashboard, and offline-first architecture.
Explore the collaboration →Open to SRE · Platform · DevOps internships now and new-grad roles from 2027 — based in India, happy to relocate to Bengaluru. If you're hiring for reliability and infrastructure, I'd love to talk.