USA · University of Houston–Clear Lake

Nandeeswar Badugu

Infrastructure Engineer. Cloud, platforms, and the cluster underneath the model.

I build reliable, automated infrastructure across DevOps, SRE, and platform engineering — and I am expanding into HPC, GPU clusters, and MLOps.

01 / About

I work at the intersection of cloud platforms, containers, and operational reliability. The goal is the same whether the surface is Kubernetes, Terraform, or a CI pipeline: systems that are observable, repeatable, and boring in the best way.

The public work below is taken from my own GitHub repositories — HPC notes, a geophysics capstone, application code, and on-chain projects. Forks of other people’s repos stay off this page.

02 / Stack

Tools I use to keep infrastructure repeatable and visible.

Cloud

  • AWS
  • Azure
  • GCP
  • Multi-cloud

Platforms

  • Kubernetes
  • Docker
  • Helm
  • EKS / AKS / GKE

DevOps & IaC

  • Terraform
  • Ansible
  • CloudFormation
  • Argo CD

Delivery

  • GitHub Actions
  • GitLab CI
  • Jenkins

Observability

  • Prometheus
  • Grafana
  • OpenTelemetry
  • Splunk

Ops

  • PagerDuty
  • ServiceNow
  • IAM
  • DNS
  • VPN

Systems

  • Linux
  • Networking
  • Load balancing

Automation

  • Python
  • Bash
  • PowerShell

Data

  • RDS
  • PostgreSQL
  • Redis
  • MongoDB

03 / Work

Original repositories — infrastructure, software, and on-chain work.

Repositories with a real system behind them, drawn from my own GitHub account.

04 / Focus

HPC & AI infrastructure

I am building toward the compute, networking, storage, and orchestration layer that ML workloads need at scale — not just training notebooks, but the cluster they run on.

  • SLURM
  • MPI
  • OpenMP
  • CUDA
  • Parallel computing
  • GPU workloads
  • Apptainer
  • Kubernetes for AI/ML
  • Model deployment
  • MLOps

05 / Contact

Let’s talk infrastructure.

Open to platform, SRE, and HPC/AI infrastructure conversations. The fastest way to reach me is GitHub.