Sai Chandu MachavarapuSai Chandu Machavarapu

$ whoami

Sai Chandu MachavarapuMLOps Engineer

DevOps Engineer | SRE | Cloud Engineer

DevOps Engineer / Site Reliability Engineer with 5+ years of experience architecting and operating secure, highly available cloud infrastructure across financial services and healthcare environments. Expertise in AWS, Azure, and GCP, with strong proficiency in Kubernetes, Docker, Terraform, and CI/CD automation (Jenkins, GitHub Actions, ArgoCD).

Tyler, TX

Illustration of Sai Chandu pointing upward while holding a tablet showing a successful CI/CD deploy
Illustration of Sai Chandu jumping forward holding a notebook with a circuit diagram

$ cat profile.md

Cloud platforms, reliability, and model-serving infrastructure.

I build observability and incident response frameworks using Prometheus, Grafana, and DataDog, and support MLOps and AI infrastructure for model-serving and RAG-based retrieval pipelines across cross-functional engineering and data teams.

M.S. Computer and Information Science, University of Texas at Tyler (Aug 2024 – May 2026) — DevOps, SRE, and platform engineering.

0+

daily transactions supported

0%

uptime on Kubernetes releases

0%

faster incident response

  • Koch Industries

    DevOps Engineer

    10+ microservices on AWS & GCP, 99.9% uptime

  • Cerner Corporation

    Cloud Platform Engineer

    4 production environments on Docker & GKE

  • Citibank

    Associate DevOps Engineer

    250,000+ daily transactions supported on AWS

$ cat experience.log

Five+ years keeping cloud platforms fast, secure, and up.

Multi-cloud infrastructure and reliability work across financial services and healthcare — from banking analytics at Citibank to healthcare data platforms at Cerner and analytics infrastructure at Koch Industries.

Core platforms

  • AWS
  • Microsoft Azure
  • Google Cloud Platform
  • Kubernetes
  • Terraform
  • Prometheus & Grafana
koch.log

DevOps Engineer

Koch Industries · NYC, NY · Mar 2026 – Present

  • Architected AWS and GCP cloud-native infrastructure, including virtual private networks, subnets, and Kubernetes-based container orchestration, to support analytics platforms powering 10+ microservices.
  • Deployed zero-downtime release strategies for containerized microservices on Kubernetes using rolling updates and automated CI/CD pipelines, increasing deployment frequency by roughly 2x while maintaining 99.9% uptime.
  • Established predictive monitoring and alerting across AWS and GCP using Prometheus, Grafana, and CloudWatch dashboards, cutting incident response time by roughly 40%.
  • Secured multi-cloud infrastructure by administering secrets management, IAM policies, and access controls through AWS and GCP key vault services.
cerner.log

Cloud Platform Engineer

Cerner Corporation · North Kansas City, MO · Mar 2025 – Feb 2026

  • Built containerized infrastructure using Docker and Kubernetes (GKE) to support distributed healthcare data-processing services across 4 production environments.
  • Implemented CI/CD pipelines using Jenkins and Gradle, improving release velocity by roughly 20%.
  • Designed observability dashboards using Prometheus, Grafana, and Splunk, cutting mean detection time for production issues by roughly 25%.
  • Partnered with security and engineering teams to harden infrastructure configurations and container registry controls, reducing vulnerability findings by roughly 30%.

Associate DevOps Engineer

Citibank · Bangalore, India · Jun 2021 – Jul 2024

Supported AWS cloud infrastructure (ECS, OpenShift) hosting distributed banking analytics services that processed over 250,000 daily transactions, deployed containerized workloads with Docker and CI/CD pipelines, and automated recurring operational workflows in Python — saving 15+ hours of manual effort per week.

$ ls skills/

The stack I build and operate with.

01

Programming & Scripting

PythonBashJavaScala
02

Cloud Platforms

AWSMicrosoft AzureGoogle Cloud PlatformVirtual NetworksCloud Infrastructure Design
03

Containerization & Orchestration

DockerKubernetesGKEEKSECSOpenShiftAWS FargateContainer RegistriesMicroservices Architecture
04

CI/CD & Infrastructure Automation

JenkinsGitHub ActionsArgoCDGradleTerraformBackstageGolden-Path CI/CD TemplatesInfrastructure as Code
05

Monitoring, Security & AI/MLOps

PrometheusGrafanaCloudWatchSplunkIAMCloud Key Vault ServicesSecrets ManagementMLflowFAISSVector DatabasesRAGModel-Serving InfrastructureNoSQL Databases
Illustration of Sai Chandu working beside a GPU server rack with a training-loss monitor
Illustration of Sai Chandu standing beside a small four-legged inference pod robot

$ ls projects/

Selected work — developer platforms and MLOps pipelines.

01
Platform

Self-Service Internal Developer Platform

BackstageTerraformCI/CDGolden Paths

Built a self-service developer platform using Backstage with a service catalog and scaffolding templates.

Integrated Terraform-based golden-path CI/CD workflows to cut new-service bootstrap time from a full day to under 20 minutes.

02
MLOps

MLOps Pipeline for Model Deployment & RAG-Based Retrieval

MLflowKubernetesFAISSRAGCI/CD

Built an automated CI/CD pipeline integrating an MLflow model registry with Kubernetes-based serving infrastructure and a FAISS vector database for RAG-based retrieval.

Cut model deployment time from a manual multi-hour process to under 10 minutes and achieved sub-200ms query response time.