sai-charan@ops
Open to DevOps, Senior DevOps and SRE roles

Sai Charan G

Senior DevOps Engineer, 5+ years

I design, automate and run cloud infrastructure and CI/CD pipelines for high-availability enterprise platforms on AWS and Azure. My work centres on Kubernetes, Docker and Terraform, with GitOps delivery that raises release frequency and cuts operational overhead. I have supported platforms in media, banking and financial services, and transport and logistics.

Pipeline: commit to productionSelect a stage

Experience

Where I have shipped

ANTS IT Solutions Pvt Ltd

Jan 2023 to present Senior DevOps Engineer

Recast

Media · Dec 2024 to present · Senior DevOps Engineer

  • Designed and maintained CI/CD pipelines on Jenkins and GitHub Actions for automated builds, testing and multi-environment deployments.
  • Provisioned and managed AWS infrastructure with Terraform: EC2, VPC, RDS, IAM, Auto Scaling, load balancers, S3, CloudFront and Route 53.
  • Deployed and scaled containerised microservices on Amazon EKS using Docker, ECR and Helm, with GitOps delivery through Argo CD.
  • Set up Prometheus and Grafana for dashboards, alerting and high-availability support.
  • Designed secure multi-account AWS networking with Transit Gateway, VPC Peering and hardened security groups.
  • Automated server configuration, patching and application deployments with Ansible, reducing manual intervention by 65%.
  • Worked with development, QA and security teams to improve the SDLC, scalability, reliability and cloud operations.
JenkinsGitHub ActionsTerraformEKSHelmArgo CDPrometheusGrafanaAnsible
# delivery flow
git push
  └─ Jenkins / GitHub Actions
      └─ Docker image → ECR
          └─ Helm chart
              └─ Argo CD sync
                  └─ EKS
                      └─ Prometheus + Grafana

Airwallex

Banking and financial services · Jan 2023 to Nov 2024 · DevOps Engineer

Global payments and financial platform (FinTech).

  • Built and maintained Jenkins CI/CD pipelines with Git and Maven for application builds, testing and multi-environment deployments.
  • Provisioned highly available AWS infrastructure: EC2, Auto Scaling Groups, Elastic Load Balancing, IAM, Amazon Aurora, DB subnet groups and Multi-AZ deployments.
  • Scaled containerised AI inference workloads and data pipelines on Amazon EKS, orchestrating model service endpoints, MongoDB caching and Auto Scaling for low-latency batch and real-time processing.
  • Containerised applications with Docker and orchestrated workloads on Kubernetes to improve scalability and resource use.
  • Automated server configuration, software installation and compliance work with Ansible roles, playbooks and Ansible Tower.
  • Implemented S3 lifecycle policies, Intelligent-Tiering and versioning to optimise storage cost.
  • Configured CloudWatch dashboards, alarms and log groups, and supported disaster-recovery planning with RTO/RPO testing.
JenkinsMavenEC2AuroraEKSAnsible TowerS3CloudWatch
# high availability layout
ELB
  └─ Auto Scaling Group (EC2)
      ├─ AZ a
      └─ AZ b
  └─ Aurora
      └─ DB subnet group
          └─ Multi-AZ

# recovery
RTO / RPO tested

Cognizant Technology Solutions Pvt Ltd

Jul 2021 to Dec 2022 DevOps Engineer

WorkWave

Transportation and logistics · AWS and Azure · Jul 2021 to Dec 2022

  • Worked with client teams to understand deployment requests, and coordinated with development, DBA, QA and IT operations to prevent resource conflicts.
  • Defined code and configuration release scope, deployment requirements and success criteria with project managers.
  • Kept work visible by creating and maintaining Jira issues to prioritise tasks and report progress.
  • Built, managed and continuously improved the build and deployment infrastructure for global software teams.
  • Implemented CI/CD with Jenkins, Git and Maven, including automated build, test and deployment scripts across environments.
  • Administered Jenkins pipelines for weekly build, test and deploy cycles using Dev, Test and Production Git branching strategies.
  • Managed AWS services: S3 bucket policies, Glacier storage classes, SNS notifications, CloudWatch monitoring, RDS access and EBS volume alarms.
JenkinsGitMavenS3GlacierSNSCloudWatchRDSEBS
# weekly release cycle
dev   → build, test
test  → build, test, deploy
prod  → deploy

# tracked in
Jira

Key achievement

Intermittent 502 errors on EKS

An application on EKS, exposed through an AWS load balancer, returned intermittent 502 errors while scaling. The cause was in how pods left the service, not in the application.

  1. Investigate

    Looked at load balancer traffic, Kubernetes events, service endpoints, pod logs and pod behaviour during scaling events.

  2. Narrow down

    Errors appeared during scale-down only. Scale-up stayed clean.

  3. Root cause

    The load balancer was still sending traffic to pods that were terminating.

  4. Fix

    Added health probes and graceful pod termination so pods deregister cleanly. The 502 errors stopped.

Hands-on projects

Infrastructure and delivery work

On-prem to AWS server migration

Team project · my part: target infrastructure

Part of a team that moved application servers from on-prem VMs to AWS. I built the landing infrastructure in Terraform and validated it before cutover.

  • Mapped on-prem server specs to EC2 instance types.
  • Wrote the VPC, subnets, route tables and security groups as code.
  • Configured IAM roles and security group rules.
  • Checked networking and application startup before cutover.
TerraformAWSEC2VPCIAM
$ terraform plan
+ aws_vpc.target
+ aws_subnet.public[*]
+ aws_subnet.private[*]
+ aws_route_table.main
+ aws_security_group.app
+ aws_iam_role.app
+ aws_instance.app[*]

Jenkins CI/CD for Java applications

Pipeline as code · dev, staging, production

Built Jenkins pipelines that compile with Maven, analyse with SonarQube, publish to Nexus and deploy to Tomcat. Explored branch-based flows so each branch maps to an environment.

  • Pipelines written in Groovy, including the plugin and syntax problems that come with it.
  • Fixed artifact path and pipeline execution failures end to end.
JenkinsMavenSonarQubeNexusTomcat
// Jenkinsfile stages
✓ Checkout
✓ Build (Maven)
✓ Code analysis (SonarQube)
✓ Publish artifact (Nexus)
✓ Deploy (Tomcat)

AWS networking and container hosts with Terraform

Personal lab · modules and remote state

Built public and private network layouts on AWS, then bootstrapped Docker hosts on EC2 with user data so a container comes up without anyone logging in.

  • Public and private subnets, NAT patterns and ALB-related networking.
  • Reusable Terraform modules with remote state.
  • Dockerfiles, Docker Compose and images pushed to registries.
TerraformDockerComposeEC2 user dataALB
# layout
internet
  └─ ALB
      └─ public subnets
          └─ NAT
              └─ private subnets
                  └─ EC2 + Docker

Kubernetes lab clusters

kubeadm and Minikube

Stood up clusters by hand to understand what the managed services hide, then worked with pods, services, NodePort access and multi-container pods.

  • Diagnosed kubelet failures and cluster networking problems.
  • Traced DNS issues through CoreDNS and kube-proxy.
KuberneteskubeadmMinikubeCoreDNSkube-proxy
$ kubectl get nodes
$ kubectl get pods -A
$ kubectl -n kube-system \
    logs -l k8s-app=kube-dns
$ journalctl -u kubelet

Technical skills

Skills and tools

Cloud platforms

Azure, AWS: EC2, S3, RDS, IAM, VPC, ELB, Route 53, CloudWatch, CloudTrail, Auto Scaling, Lambda, Transit Gateway, VPC Peering, KMS, WAF, Secrets Manager, CloudFront, Aurora, SNS, EBS, Glacier

Containers and orchestration

Docker, Docker Compose, Kubernetes, Amazon EKS, ECR, Helm, OpenShift, kubeadm, Minikube

CI/CD

Jenkins, GitHub Actions, Argo CD (GitOps), Azure DevOps (ADO) Pipelines

Build and artifacts

Maven, Nexus Artifact Repository

Infrastructure as code

Terraform, modules, remote state

Configuration management

Ansible roles, playbooks, Ansible Tower

Monitoring and observability

Prometheus, Grafana, Datadog, AWS CloudWatch

Scripting and automation

Bash, shell scripting, Python

Version control

Git, GitHub, Bitbucket

Web servers

Nginx, Apache, Apache Tomcat

Databases

MySQL, AWS RDS, MariaDB, MongoDB (NoSQL)

Operating systems

Linux (Ubuntu, CentOS), Windows Server

Project tools

Jira, Confluence

Security and compliance

IAM roles and policies, VPC security groups, AWS CloudTrail auditing

DevSecOps

SonarQube, secrets scanning

Practices

GitOps, multi-environment releases, disaster recovery (RTO/RPO), cost optimisation, security hardening, incident troubleshooting

Currently learning

Kafka, Airflow, microservices, DevSecOps

Generative AI in DevOps work

Amazon Bedrock and GitHub Copilot for CI/CD pipeline analysis, Terraform and Ansible code generation, Linux troubleshooting, log analysis and root-cause identification

Field notes

What I have debugged

Terraform

Credential errors, state locking, provider installation failures and resource configuration mistakes.

Jenkins

Plugin problems, Groovy syntax, wrong artifact paths and pipelines that fail partway through.

Docker

Image push and authentication failures, crashing containers, port conflicts and storage limits.

Kubernetes

Kubelet failures, pod networking, CoreDNS and kube-proxy behaviour, and load balancer 502s during scale-down.

Git

Detached HEAD, lost commits recovered with reflog, .gitignore cleanup and Git LFS trouble.

Environment

Version mismatches, missing paths and packages, and VM resource limits on local labs.

About

Education and approach

I learn by building and breaking things. When something fails I want the reason it works the way it does, not only the command that clears the error.

I am looking for a DevOps, Senior DevOps or SRE role where I can own infrastructure as code, delivery pipelines and Kubernetes platforms.

Also studying: Kafka, Airflow, microservices and DevSecOps.

Full name
G Sai Charan Reddy
Education
B.Tech, graduated 2019
Experience
5+ years
Domains
Media, FinTech, Logistics
Cloud infrastructure managementDevOps automationCI/CD pipeline designKubernetes cluster administrationDocker containerisationGitOpsTerraform IaCAnsible configuration managementAWS cloud servicesSite reliability engineering (SRE)Agile/ScrumLinux administrationSystem monitoringHigh availability architectureDisaster recoveryMicroservicesServerless computingCost optimisationSecurity hardeningRelease management

Contact

Let's talk