Asutosh Panda

Skills
Professional Experience
Socure Chennai, India
Senior Site Reliability Engineer (Remote) August 2026 - Present
  • Own the Fravity to Socure platform migration end to end: architected the target state, sequenced it across 5 Socure and Fravity teams, and negotiated the account-boundary and security sign-offs each phase depended on before a single workload moved.
  • Built the GPU inference platform on EKS: Karpenter-scaled scale-from-zero L40S running vLLM on Ray Serve, Envoy AI Gateway plus Gateway API Inference Extension dispatching to llm-d, APISIX auth and rate limits, benchmarked to ~527K requests/day at FP8.
  • Made inference performance measurable: guidellm precision and traffic sweeps to p999 against chat-tier SLOs, DCGM telemetry, an OOM crashloop root-caused at 1,400+ restarts, and GPUs halved by correcting an fp32 assumption to bf16.
  • Delivered both services across 18 merged PRs in 5 repos: Terragrunt units for ECR, Aurora PostgreSQL and Valkey, CI wrappers publishing to ECR, Argo CD Helm charts with ExternalSecrets and Istio policy; migrated 70 repos to GHE in < 5 days.
  • Cleared the security and compliance gates the migration ran on: provisioned 6 Socure security engineers into GCP, GitHub and Workspace, onboarded their audit team to GCP IAM, and bridged Socure laptops to 4 VNETs over Cloudflare Zero Trust.
  • Deployed the 3-environment AWS EKS platform on /21 IPAM VPCs with TGW hub-and-spoke and centralized ECR, then set the golden path on it: keyless GHE OIDC to ECR behind 4 scanners, Argo CD auto-deploy, Istio mTLS, FIPS golden bases.
  • Centralised release engineering into 10 reusable GitHub Actions across 4 repos: trunk-based dev/sandbox/prod promotion, hotfix lanes that block normal releases, image-only prod rollback leaving database state intact, actor allowlists on privileged workflows.
One2N Pune, Maharashtra, India
SRE Lead · Client: Fravity AI (Remote) May 2026 - July 2026
  • Led a licence-driven observability migration, AGPL to Apache-2.0, off Grafana Mimir/Loki/Tempo onto VictoriaMetrics/Logs/Traces with Perses: a zero-gap Grafana Alloy dual-run scraping ~305K active series across 192 targets in 4 clusters, dashboards as code.
  • Built fravity-spine, the internal Kubernetes platform other teams consume: private regional GKE, CMEK etcd, Workload Identity, CloudNativePG, OpenTofu across 6 state layers, and identity-based kubectl retiring bastions and shared SSH.
  • Led cost + access governance across 5 GCP + 5 AWS accounts as lead of a two-person SRE team: surfaced 19 live SA keys and ~$26,520/yr of list-price exposure, then shipped an org-wide FinOps board and IAM as Terraform on keyless CI.
Luxor Technology Seattle, WA, US
DevOps Engineer (Remote) June 2023 - April 2026
  • Operated a multi-cluster Kubernetes platform for 20+ teams and 3,000+ enterprise customers: 8+ clusters across Hetzner, OVH, and GCP with 24/7 uptime for L1/L2 validator and RPC workloads, including 2 Hetzner bare-metal HA clusters (KubeOne, OpenEBS) built from scratch. Standardised on a 7-layer FluxCD GitOps pipeline with Linkerd mTLS, migrating 70+ secrets to Workload Identity.
  • Owned CI/CD for 50+ repositories on self-hosted ARC runners. Ported a failing Playwright E2E suite from Mac hosted runners ($5,000/mo) to Linux ARC, cutting runtime 40 to 15 min and cost to $600/mo. Built code-signed Windows + Mac release pipelines and automated FluxCD deployments to GCP Artifact Registry.
Tata Consultancy Services Hyderabad, Telangana, India
Systems Engineer (Remote) October 2020 - June 2023
  • Migrated 2,500+ repositories (GitLab, Bitbucket, SVN) to GitHub Enterprise for 120 teams and built their CI/CD pipelines, resolving self-hosted-API pagination limits across the platforms and saving 15 weeks of manual effort.
  • Built PowerShell + Python automation (server monitoring, alerting, log rotation, Oracle DB replication, MS SQL updates) across 22 RHEL/WebLogic servers and 3 projects, saving 380+ man-hours/month.
  • Shipped a Golang CLI that cut HP storage provisioning from 30 man-hours across 30 monthly tickets to a single command, adopted by the storage team as their primary provisioning interface.
Open-Source Projects
Teaching Experience
Education
SASTRA Deemed University Master of Computer Applications (MCA) - Chennai, India - 2021 - 2023