Redox Inc.
Senior Software Engineer — Platform
02/2022 – Now · Remote
-
Designed and operated multi-cloud Kubernetes infrastructure in
AWS and GCP supporting 75+ production services and reducing
deployment time by 15%.
-
Designed and implemented a GitOps-based continuous delivery
platform using Crossplane and Argo CD, reducing deployment time
by 80% across production services.
-
Led implementation of a centralized observability and telemetry
platform using Prometheus, Grafana, and OpenTelemetry, enabling
proactive alerting for 75+ production services and delivering
20+ multi-cloud, multi-environment dashboards.
-
Served as primary on-call engineer and incident commander for
production incidents and maintenance events, improving system
reliability and reducing operational response times.
-
Authored 15+ architecture proposals and drove cross-team
platform initiatives impacting platform reliability, deployment
workflows, and observability practices.
-
Built internal tooling and automation leveraging LLM-based
workflows to reduce dependency management and platform
maintenance effort.
Georgia Tech Research Institute
Research Faculty — Cloud Engineer
09/2021 – 02/2022 · Atlanta, GA
-
Built standardized Kubernetes-based application platforms
supporting research workloads and production services.
-
Led multi-cloud infrastructure initiatives in Amazon Web
Services and Microsoft Azure.
-
Led development of a scalable remote learning lab platform used
by 20+ students per class across 3 curricula.
-
Designed and deployed a high-availability on-prem Kubernetes
cluster integrated with enterprise identity systems and VMware
infrastructure.
Georgia Tech Research Institute
Research Faculty — Systems Administrator
06/2019 – 09/2021
-
Automated large-scale bare-metal and VM deployments using
Foreman, Red Hat Linux, Python, Packer, and Ansible.
-
Designed and maintained enterprise infrastructure spanning
virtualization, storage fabrics, networking, and operating
systems.
-
Implemented centralized log aggregation and observability
tooling, reducing time to query system logs from 5 minutes to
under 30 seconds.
-
Led migration of critical file storage infrastructure to a
distributed SAN architecture, improving uptime to 99.99%.
-
Developed disaster recovery and business continuity strategies
reducing MTTR to under one day.