Vadivelan M P

Vadivelan M P

Software Technical Ops Manager

Chennai, India

About

I am an SRE & Technology Operations Leader with nearly 15 years of experience managing large-scale enterprise platforms across telecom and cloud ecosystems. Most recently at Verizon, I have led end-to-end production operations for mission-critical billing systems, ensuring high availability, SLA compliance, and performance stability across AWS environments. My work spans automation-led operational transformation using Python and Shell scripting, improving monitoring, alerting, and recovery workflows. I have driven incident reduction initiatives through structured problem management and trend analysis, and played a key role in billing platform modernization from mainframe to microservices architecture using Spring Boot. I manage AWS infrastructure components such as EC2, ELB, and S3, and utilize ITSM tools like ServiceNow and Jira alongside DevOps tools including Jenkins, Git, and UDeploy.

Experience

Verizon

Software Technical Ops Manager

Verizon

Jun 2017 – Present

Directed end-to-end production operations for high-volume billing platforms, ensuring system availability, resilience, and SLA adherence across distributed environments. Functioned as the single point of ownership (SPOC) for multiple applications, driving operational accountability, escalation management, and service continuity. Led cross-functional teams comprising 5–6 internal engineers and 20+ vendor resources, ensuring execution excellence, governance, and performance alignment. Drove enterprise-scale incident reduction initiatives through trend analysis, structured problem management, and preventive engineering practices. Designed and implemented automation frameworks using Python and Shell scripting, improving monitoring accuracy, alerting efficiency, and recovery turnaround. Managed and optimized AWS infrastructure (EC2, ELB, S3) to enable scalable, fault-tolerant, and highly available systems. Strengthened observability frameworks, enhancing system visibility, early anomaly detection, and incident response effectiveness. Led operational support for large-scale billing platform transformation from legacy mainframe systems to microservices (Spring Boot-based architecture). Governed vendor delivery and SLA frameworks, ensuring accountability, cost efficiency, and consistent service delivery across distributed teams. Key Achievements: Improved MTTR by ~35–45% through automation-led recovery and monitoring optimization Reduced critical/high-severity incidents by ~30–40% via observability & alerting transformation Improved platform performance by ~20–30% post modernization via operational stabilization Enhanced vendor SLA adherence by ~15–25% via governance and performance frameworks Increased operational productivity by ~30–50% via automation and workflow standardization Delivered 25–30% reduction in incident inflow through proactive problem management Recognized with multiple organizational awards (6+) for operational excellence and modernization contributions. Successfully enabled operational transition for enterprise billing modernization impacting large- scale customer base. Architecture & Engineering Practices: Microservices Architecture (Spring Boot), Monitoring, Alerting & Observability Systems ITSM & Workflow Tools: ServiceNow, Jira DevOps & Release Engineering: Jenkins, Git, UDeploy Data & Performance: SQL Systems Engineering: Unix / Linux Programming & Automation: Python, Shell Scripting Cloud & Distributed Systems: AWS (EC2, ELB, S3), Apache Kafka

PythonShell ScriptingAWS
Accenture

Senior Software Engineer

Accenture

Aug 2011 – Jun 2017

Delivered L3 production support for telecom-grade applications, ensuring stability and continuity across high-availability environments. Led major incident management as Incident Commander, driving end-to-end resolution, stakeholder communication, and rapid service restoration. Designed and deployed automation solutions using Shell scripting, SQL, and Jenkins, reducing manual dependencies and improving operational efficiency. Conducted deep-dive root cause analysis (RCA), eliminating recurring issues and strengthening system reliability. Improved system performance through SQL optimization and backend tuning, enhancing responsiveness and processing efficiency. Key Achievements: Reduced incident recurrence by ~20–25% through structured RCA and permanent fixes Improved system performance by ~15–20% through database and backend optimization Additional Exposure: Worked across Bangalore (initial tenure) and Chennai (long-term base). International exposure: o Singapore – 2 months o Sydney – 2 visits (~1 month each)

Shell ScriptingSQLJenkinsJavaApplication MonitoringRelease Management

Education

Anna University

Anna University

B.E.

2011

Certifications

Google AI Essentials

Coursera

Apr 2026 – No expiry

Google AI Professional Certification

Coursera

Skills

Spring BootService DeliveryVendor ManagementShell ScriptingPythonAutomationAWSCloud OperationsMonitoringObservabilityError Budget ManagementSLO ManagementSLA ManagementITILChange ManagementProblem ManagementIncident ManagementProduction Operations

Languages

English (Full professional proficiency)Tamil (Full professional proficiency)