Production Engineer at Meta Platforms | Former Apple SRE DRI
Meta PlatformsApple Inc. (Former)KubernetesApache Solr (3.5 PB)
GLOBAL UPTIME
99.999%
DATA CAPACITY
3.5 PB
ACTIVE ENDPOINTS
300+
INCIDENT STATUS
0 Active
Systems Topology & Scale Visualizer
Interactive Map
Interactive SRE deployment topology showing Anoop's engineering scope. Click on standard node points to query availability metrics, active capacity, and scaling logs.
SRE DRI
Meta PE
Content Integrity SRE (Remote) Ecosystem: Large-Scale Data Pipelines Scope: Compute Infrastructure & Health
Core Site Reliability Engineering DRI (Apple Inc.): Controlling availability, capacity optimization, and incident automation for mission-critical Apple Maps datasets across hundreds of global Kubernetes microservices. Click any peripheral node above to load its real-time telemetry.
Experience Timeline
Professional Deployments
20 years of operational experience designing high-availability SRE solutions, release management systems, and high-scale infrastructure migrations.
Production Engineer
Meta Platforms
Mar 2026 – Present
Content Integrity Production Engineering: Provide comprehensive Production Engineering (PE) and Site Reliability Engineering (SRE) support to the Content Integrity team, ensuring stability and performance of large data pipelines and compute infrastructure.
Operational Health: Manage reliability and operational health of high-scale distributed systems, applying SRE best practices to improve availability, latency, and incident response.
Senior Site Reliability Engineer & Tech Lead
Apple Inc.
Aug 2017 – Mar 2026
Core Maps Data SRE: Serving as the SRE DRI for critical global platforms, owning end-to-end performance, incident response, and SLA goals across **Places**, **Base Map**, and **BusinessConnect**.
3.5 Petabyte Solr Platform: Technical Lead managing 35+ high-performance search and indexing clusters processing massive telemetry datasets.
Kubernetes Inception & Operations: Designed, built, and operationalized the customized Kubernetes cluster environment for the Apple Maps Places org, heavily reducing server overhead.
Release Management Architecture: Created the unified automated Release Management platform for Apple Maps Data Ecosystem supporting 300+ active microservices.
Autonomic AI Agent SRE: Contributed to building a Langgraph-based intelligent AI agent for automated, RAG-driven pager alerts diagnosis and incident routing.
Distributed Team Management: Manages a team of 6 cross-geographical SRE and Cloud Engineers, running standups and capacity planning.
DevOps Engineer & Java Developer
State Street IMS
Mar 2011 – Aug 2017
CI/CD Implementation: Set up the first automated Jenkins continuous integration server infrastructure, optimizing nightly deployments.
Release Validation & Automation: Wrote complex automation scripts in Bash, Ant, and Windows batch to deploy multi-tiered enterprise servers.
Sybase Database Migration: Orchestrated major database migrations (Sybase to Oracle) using Talend Open Studio ETL pipelines.
Java Developer & Technical Lead
Wellpoint Anthem
Jul 2010 – Mar 2011
Enterprise Web Services: Wrote high-load Java web services for the Anthem Online Store, facilitating robust healthcare selection routing.
Pharmacy Website Upgrades: Optimized administrative portal backends, reducing runtime database constraints via Spring Batch automation.
SRE Skills Inventory
Stack Status
Tier-structured technical capabilities and toolchains classified by active cloud-native and system reliability domains.
AI Agents & Orchestration
LanggraphLangchainLarge Language Models (LLMs)RAG SystemsKnowledge Graphs