Prodapt is the largest specialized player in the Connectedness industry. As an AI-first strategic technology partner, Prodapt provides consulting, business reengineering, and managed services for the largest telecom and tech enterprises building networks and digital experiences of tomorrow. A ServiceNow-invested company, Prodapt has been recognized by Gartner as a Large, Telecom-Native, Regional IT Service Provider. A “Great Place To Work® Certified™” company, Prodapt employs over 5,000 technology and domain experts across the Americas, Europe, India, Africa, & Japan. Prodapt is part of the 130-year-old business conglomerate The Jhaver Group, which employs over 32,000 people across 80+ locations globally.
Prodapt is seeking a highly skilled and experienced Senior DevOps Automation Engineer to lead and drive strategic DevOps initiatives across multiple complex projects within a wireless network infrastructure environment. This role is critical to ensuring the successful deployment, scalability, security, and automation of solutions that support next-generation telecommunications services.
Responsibilities
- Lead the design, implementation, and continuous improvement of DevOps best practices, including Continuous Integration (CI), Continuous Deployment (CD), automated testing, and Test-Driven Development (TDD).
- Design and maintain highly resilient, scalable, secure, and software-defined infrastructure platforms supporting telecom-grade availability and performance requirements.
- Automate operational processes to improve deployment speed, consistency, and platform reliability.
- Drive continuous improvement initiatives across infrastructure, deployments, monitoring, and incident response processes.
- Design and implement end-to-end automated deployment pipelines using GitHub, GitHub Actions, and Microsoft SQL Server (MSSQL).
- Build and manage environment promotion workflows across Development, Staging, and Production environments with minimal to no manual intervention.
- Optimize deployment processes to enable rapid and reliable software releases.
- Install, configure, and maintain Windows Server-based GitHub Actions self-hosted runners.
- Ensure high availability (99.9% uptime), secure network connectivity, performance monitoring, and operational health checks.
- Develop and maintain production-grade PowerShell automation scripts for:
- Deployment
- Backup
- Rollback
- Health checks
- Validation
- Implement robust error handling, transaction management, auditing, and logging capabilities.
- Integrate GitHub repositories with Jira to automatically link code branches and deployments to Jira tickets.
- Automate ticket lifecycle management, including status updates throughout the deployment pipeline.
- Enrich Jira tickets with deployment logs, build data, commit hashes, and deployment URLs.
- Design and support cloud-native solutions leveraging AWS, DynamoDB, Amazon Aurora, and Amazon Kinesis.
- Support database deployment automation and governance for enterprise applications.
- Implement and maintain monitoring, observability, and alerting solutions using Prometheus, Grafana, Kubernetes Monitoring, OpenTelemetry, Splunk, Zabbix, and Dynatrace.
- Proactively monitor production systems and rapidly troubleshoot performance, infrastructure, and deployment issues.
- Configure real-time deployment notifications via Slack and email.
- Integrate critical alerts with PagerDuty for rapid incident response and escalation management.
- Participate in on-call support rotations and production incident management.
- Implement DevOps security best practices, including branch protection policies, environment approval gates, GitHub Secrets management, GPG-signed commits, and role-based access controls.
- Ensure compliance with enterprise security and audit requirements.
- Develop deployment auditing mechanisms, including deployment tracking tables within MSSQL.
- Build dashboards and reports to monitor deployment frequency, success rates, failure trends, and operational KPIs.
- Maintain comprehensive documentation for infrastructure, automation solutions, and deployment processes.
- Design and implement automated backup and rollback strategies, achieving a Recovery Time Objective (RTO) of under 15 minutes, a Recovery Point Objective (RPO) of under 5 minutes, and a 30-day backup retention policy.
- Develop disaster recovery runbooks and automate restoration procedures.
- Optimize deployment processes to reduce deployment times from approximately 30 minutes to under 5 minutes through automation and parallel execution strategies.
- Improve storage efficiency by implementing incremental backup solutions and reducing backup costs.
- Document operational runbooks, conduct knowledge transfer sessions, and lead continuous improvement reviews.
Requirements
- Bachelor's degree in Computer Science, Engineering, Information Technology, or a related technical field; equivalent industry experience will also be considered.
- 10+ years of hands-on experience in DevOps, Infrastructure Automation, Site Reliability Engineering (SRE), or Platform Engineering.
- Proven expertise designing, implementing, and managing enterprise-scale CI/CD pipelines.
- Strong scripting and automation experience with PowerShell, Python, Bash, and Groovy.
- Experience with configuration management tools such as Ansible, Puppet, or Chef.
- Hands-on experience with GitHub and GitHub Actions, Windows Server administration, Microsoft SQL Server, AWS cloud services, Docker, and Kubernetes.
- Strong knowledge of monitoring, observability, and logging platforms.
- Experience integrating DevOps solutions with Jira, Slack, PagerDuty, and other enterprise collaboration tools.
- Understanding of database deployment automation and release management methodologies.
- Experience supporting highly available production environments with strict uptime requirements.
- Excellent troubleshooting, analytical, and problem-solving skills.
- Strong communication and stakeholder management abilities.
Learn more about this Employer on their Career Site
