// Dublin, Ireland

Hi, I'm Veera Maddula.

> Senior Site Reliability Engineer▮

Senior DevOps and Site Reliability Engineer with 5+ years of experience building, automating, and operating large-scale multi-cloud platforms across AWS, Azure, and GCP. Proven track record resolving 300+ high-priority incidents, reducing MTTR by ~25%, and driving observability and infrastructure-as-code across 30+ production services serving millions of daily transactions globally. Now leveraging AIOps for AI-assisted anomaly detection, automated remediation, and intelligent incident resolution across hybrid and multi-cloud environments.

veera@cloud: ~ --:--:-- UTC
$whoami
veera-maddula
$cat role.txt
Site Reliability Engineer
DevOps · AWS · Kubernetes · Observability
$uptime --career
~5 years · 30+ services · 300+ P1s resolved
$kubectl get impact
MTTR reduced ~25% · 4 global regions
$terraform providers
aws · azurerm · google · kubernetes · helm
$aiops status
anomaly detection · auto-remediation · active
$▮
5+
Years Experience
30+
Digital Products
300+
P1 Tickets Resolved
AWS
Certified Architect

Building Reliable Systems at Scale

I'm a Senior DevOps and Senior Site Reliability Engineer with 5+ years of experience building, automating, and operating cloud platforms at scale across AWS, Azure, and GCP. I own the full delivery path: CI/CD pipelines in Jenkins, GitHub Actions, Azure DevOps, and Argo CD, infrastructure as code in Terraform and Ansible, and container platforms on Kubernetes, EKS, and ECS. At Disney Interactive I ran the shared AWS infrastructure behind ecommerce platforms serving millions of daily transactions across four global regions, supporting 30+ digital products, each an independent microservice written in Java, Python, Go, or Node.js, so operating comfortably across very different runtimes, build systems, and failure modes became second nature.

I've led response to 300+ priority one production incidents and cut mean time to recovery by roughly 25% through automation and proactive monitoring. Reliability is the core of how I work: defining SLIs and SLOs, owning error budgets, running blameless postmortems, and building observability in Datadog, Grafana, Prometheus, Splunk, and the ELK stack. Alongside that I've tuned Kafka and SQS pipelines, written Python self healing automation, and driven vulnerability remediation across container images and cloud networking. More recently I've worked in AIOps, applying anomaly detection and automated remediation to reduce toil and shorten incident response in hybrid environments. I hold the AWS Certified Solutions Architect credential and completed an M.Sc. in Computing at South East Technological University in Dublin, and I'm open to Senior DevOps, Site Reliability, Platform, and Cloud Engineering roles across Ireland.

⚡
Automation First
Terraform, Ansible, CI/CD, GitOps
30+ services provisioned as code
🔭
Observability
Datadog, Grafana, Prometheus, ELK
SLI dashboards across 4 regions
🛡️
Reliability
SLIs, SLOs, error budgets, on call
300+ P1 incidents led
☁️
Cloud Native
Kubernetes, EKS, ECS, Docker, Helm
30+ microservices deployed
🌐
Multi Cloud
AWS, Azure, GCP
4 global regions on AWS
🧠
AIOps
Anomaly detection, auto remediation
~25% faster incident recovery
Veera Maddula
🧑‍💻 Senior DevOps / SRE 📍 Dublin, Ireland ⏱️ 5+ Years Experience 🎓 M.Sc. Computing (Completed) ☁️ AWS Certified

Skills & Technologies

Cloud & Infrastructure
AWS
Azure
GCP
🖥️ EC2
🚢 ECS
⎈ EKS
📦 S3
🗄️ RDS
🌐 Route 53
⚡ CloudFront
⚖️ ALB
📈 Auto Scaling
🔐 IAM
📊 CloudWatch
🕸️ VPC
Google Cloud Platform (GCP)
🖥️ Compute Engine
⎈ GKE
⚡ Cloud Functions
🚀 Cloud Run
📦 Cloud Storage
🗄️ Cloud SQL
📊 Cloud Monitoring
📝 Cloud Logging
🌐 Cloud DNS
⚖️ Cloud Load Balancing
🔐 Cloud IAM
🐳 Artifact Registry
🔄 Cloud Build
📬 Pub/Sub
🌐 Cloud CDN
🔒 VPC Networks
📈 BigQuery
🔑 Secret Manager
Containers & Orchestration
Docker
Kubernetes
Helm
⎈ Amazon EKS
☸️ Azure AKS
🚢 Amazon ECS
🚪 Ingress Controllers
🔌 CNI Networking
IaC & Config Management
TFTerraform
AAnsible
📝 CloudFormation
🧱 AWS CDK
💪 Bicep
🎭 Puppet
CI/CD & GitOps
Jenkins
Git
🐙 GitHub Actions
🔷 Azure DevOps
🐙 Argo CD
🔄 GitOps Workflows
🔧 Maven
Observability
🔥 Prometheus
📈 Grafana
🔭 OpenTelemetry
🐶 Datadog
📊 CloudWatch
🔍 Splunk
📊 EFK / ELK Stack
🚨 PagerDuty
🛰️ Synthetic Monitoring
🔔 Alerting & Dashboards
Reliability Practices
🎯 SLIs / SLOs
💸 Error Budgets
🚒 Incident Management
📋 Blameless Post-mortems
📟 On-Call
⏱️ MTTR Reduction
📐 Capacity Planning
🤖 Toil Reduction
💥 Chaos / Failure Injection
AIOps
🧠 AI-Assisted Operations
📡 Anomaly Detection
🔧 Automated Remediation
Event Streaming & Databases
📨 Apache Kafka
📬 AWS SQS
⚡ Event-Driven Architecture
🐘 Aurora Postgres
🗃️ Oracle RDS
🐬 MySQL
Systems, Networking & Security
🐧 Linux
🔗 TCP/IP
🌐 DNS
⚖️ Load Balancing
🔒 TLS
🛡️ Security Groups
🔑 IAM Least-Privilege
🧱 Hardening Baselines
🩹 Vulnerability Remediation
ITSM & Project Management
SNServiceNow
JIRA
Languages & Scripting
PyPython
$_Bash
GoGo

Featured Projects

Professional Experience

Feb 2026 – Present
Freelance Senior Systems Engineer
Booking & Staff Operations Platform, Dublin, Ireland
Client: Garcijo Investments Ltd

Designed and delivered a full-stack reservation and staff-operations platform for a hospitality business, owning the architecture, data model, role hierarchy, security posture, and third-party integrations end to end.

  • Defined system architecture and service boundaries, then directed AI-assisted implementation across front end and back end
  • Built a conflict-aware booking engine with best-fit table assignment, double-booking prevention, and manager approval workflow
  • Implemented role-based access control across 7+ staff roles with granular per-feature permissions
  • Delivered a staff kiosk system (camera clock-in/out, PIN auth with brute-force lockout, duty-window scheduling) and two-way Google Calendar sync
  • Led a full security review covering CSRF protection, rate limiting, session hardening, and dependency CVE remediation
  • Evaluated and integrated third-party providers (Resend for email, Sendmode for SMS) on cost and feature trade-offs
🏆 Refactored a 1,400-line monolithic data layer into 15 domain modules with zero regressions, verified by automated public-API diffing
Jul 2023 – Jul 2024
Site Reliability Engineer
NetEnrich Technologies, Hyderabad, India
Resolution Intelligence Cloud (AIOps / SecOps Platform)

Drove reliability, scalability, and observability for a multi-tenant AIOps/SecOps platform on AWS, built from microservices and event-driven pipelines serving enterprise customers.

  • Operated and improved reliability of the platform underpinning AI-assisted anomaly detection and automated incident-resolution workflows
  • Built Terraform modules and Ansible playbooks provisioning AWS infrastructure (EKS, EC2, S3, RDS, IAM, CloudWatch) under GitOps
  • Instrumented services with EFK/ELK and APM; defined SLIs and alerting thresholds protecting customer-facing SLOs
  • Developed Python self-healing automation, Kafka consumer-lag monitors, and scheduled pipeline health checks
  • Tuned Kafka and AWS SQS (partitions, consumer groups, retry logic) to sustain throughput during load spikes and backlog recovery
  • Partnered on vulnerability management and secure configuration of container images, Kubernetes workloads, and cloud networking
Jan 2020 – Sep 2023
Site Reliability Engineer / Cloud Operations
NetEnrich Technologies, Hyderabad, India
Client: Disney Interactive

Supported large-scale, multi-tenant SaaS production systems on AWS across four global regions, driving reliability, automation, and observability for microservice-based e-commerce platforms.

  • Managed shared multi-tenant AWS infrastructure for four Disney e-commerce platforms (EU, US, Japan, APAC)
  • Built comprehensive observability using Datadog, CloudWatch, Splunk, Grafana, ELK, and AppDynamics
  • Automated infrastructure provisioning with Terraform and Ansible, reducing configuration drift
  • Deployed containerized microservice workloads via Amazon ECS and EKS (Docker, Helm)
  • Led incident troubleshooting in on-call rotations with PagerDuty, driving blameless postmortems
  • Performed pre-scale capacity checks for major launch events (Black Friday, holiday sales)
  • Implemented Kafka and AWS SQS event-driven messaging for inter-service communication
  • Defined SLIs, SLOs, and error budgets for customer-facing services, driving error-budget-based release decisions
  • Led response to 300+ P1 incidents across 30+ production services as part of the Rescue Rangers on-call team
🏆 Reduced average incident resolution time by ~25% through workflow automation and proactive observability

Education & Certifications

AWS

AWS Certified Solutions Architect

Amazon Web Services • 2024

M.Sc. Computing (Information System Processes)
South East Technological University
📍 Dublin, Ireland
Sep 2024 – Oct 2025
B.Tech in Electrical & Electronics Engineering
Aditya University
📍 Surampalem, India
May 2018 – May 2021

Get in Touch

Interested in working together? I'd love to hear from you.
Click below to send me your details.

// Contact

Let's Build
Something Great

Fill in your details below and I'll get back to you as soon as possible. Looking forward to connecting!

✉️ jagannadham.ireland.edu@gmail.com
📞 +353 894 338 657
📍 Dublin, Ireland