About Santosh Parsa
Systems thinker, infrastructure engineer, and open-source builder
Background & Journey
I am a Senior Staff Systems & Infrastructure Engineer based in Hyderabad, India, with over 13 years of experience operating, automating, and scaling high-availability enterprise platforms.
Over the past 7+ years at ServiceNow, I have worked across the systems engineering and platform reliability spectrum—advancing from Senior Linux System Administrator to Staff Systems Design Engineer, and currently serving as Sr Staff Linux Systems Engineer. In this capacity, I have served as Primary On-Call Incident Commander for critical outages, driven automated remediation frameworks that slashed MTTR by 42%, and architected multi-region telemetry systems ingesting billions of daily events.
Prior to ServiceNow, I spent 6 years at CtrlS Datacenters managing Tier-4 datacenter server fleets, virtualization clusters, SAN storage arrays, and network appliances.
Engineering Philosophy
- Automate Toil: If an operational task must be performed more than twice, write an idempotent playbook or self-healing daemon for it.
- Shift Reliability Left: Production stability starts in architecture readiness reviews, CI/CD validation, and non-prod environment parity—not in 3 AM incident triage.
- Zero-Trust AI Integration: AI agents belong in operations, but only inside capability-bounded, read-only security sandboxes that strictly neutralize prompt injection and mask credentials.
- Psychological Safety in Postmortems: Systems fail because of systemic gaps and edge conditions, never individuals. Convert every outage into automated regression checks.
Get in Touch
I am always open to discussing distributed systems reliability, Linux internals, cloud architecture, and opportunities for engineering leadership.