Customer Success & DevOps Engineer (Noida)
Job Description
The role
Were evolving live production and playout operations across newsrooms, sports, and 24/7 channels by fusing traditional broadcast engineering with modern DevOps and cloud/SaaS operations. Youll be a handson expert who can design, deploy, automate and support missioncritical workflows, from studio to edge to cloud, with a relentless focus on uptime, performance and customer success. This role bridges broadcast engineering expertise with deep platform operations for AMPP.local deployments. You will support the Kubernetes-based control layer, its lifecycle, and its resilience. Youll ensure live news, sports, and playout workflows run flawlessly while maintaining the underlying infrastructure that powers them.
The role requires regular local travelling to our clients' sites.
What youll do
Broadcast and Devops engineering (news, sport, playout)
- Implement, and support live news and sports production workflows (ingest, studio, replay, graphics, playout automation) across SDI and IP (SMPTE 2110/2022-6).
- Configure and troubleshoot playout systems, SCTE triggers, branding, captions, and loudness compliance.
- Produce deployment documentation and operational runbooks for hybrid and on-prem AMPP environments.
DevOps & Platform Ownership (AMPP.local)
- You are the operator of the control plane and responsible for its health, security, and availability:
Linux & OS-Level Responsibilities
- Install and harden Linux OS (Ubuntu or equivalent) on control-plane and node systems.
- Manage disk partitioning, LVM volumes, file systems, and log rotation to prevent disk saturation.
- Maintain DNS, NTP, hostname configs, and secure SSH key-based access for automation.
Kubernetes (K3s) Control Plane
- Deploy and initialize the AMPP.local K3s cluster (control plane setup).
- Perform AMPP.local upgrades, including Kubernetes component updates.
- Manage and rotate platform certificates to prevent expiration or security failures.
- Monitor and recover the etcd datastore; maintain kubelet and containerd services.
- Respond to pod crashes, failed scheduling, and manage persistent storage provisioning.
- Administer network policies, service routing, and RBAC for both Kubernetes and OS-level accounts.
- Test failover between cluster nodes for high availability.
