← Back to job listings
GR
Customer Success & DevOps Engineer
grassvalley · Mumbai, India
About The Role
The role
- We’re evolving live production and playout operations across
- newsrooms, sports, and 24/7 channels
- by fusing traditional broadcast engineering with modern DevOps and cloud/SaaS operations. You’ll be a hands‑on expert who can
- design, deploy, automate and support
- mission‑critical workflows, from studio to edge to cloud, with a relentless focus on uptime, performance and customer success. This role bridges
- broadcast engineering expertise
with
- deep platform operations
- for AMPP.local deployments. You will support the Kubernetes-based control layer, its lifecycle, and its resilience. You’ll ensure
- live news, sports, and playout workflows
- run flawlessly while maintaining the underlying infrastructure that powers them.
- The role requires regular local travelling to our clients' sites.
- What you’ll do
- Broadcast and Devops engineering (news, sport, playout)
- Implement, and support
- live news and sports production
- workflows (ingest, studio, replay, graphics, playout automation) across SDI and IP (SMPTE 2110/2022-6).
- Configure and troubleshoot
- playout systems
- , SCTE triggers, branding, captions, and loudness compliance.
- Produce deployment documentation and operational runbooks for hybrid and on-prem AMPP environments.
- DevOps & Platform Ownership (AMPP.local)
- You are the
- operator of the control plane
and responsible for its health, security, and availability
Linux & OS-Level Responsibilities
- Install and harden Linux OS (Ubuntu or equivalent) on control-plane and node systems.
- Manage disk partitioning, LVM volumes, file systems, and log rotation to prevent disk saturation.
- Maintain DNS, NTP, hostname configs, and secure SSH key-based access for automation.
Kubernetes (K3s) Control Plane
- Deploy and initialize the AMPP.local K3s cluster (control plane setup).
- Perform AMPP.local upgrades, including Kubernetes component updates.
- Manage and rotate platform certificates to prevent expiration or security failures.
- Monitor and recover the
- etcd datastore
- ; maintain kubelet and containerd services.
- Respond to pod crashes, failed scheduling, and manage persistent storage provisioning.
- Administer network policies, service routing, and RBAC for both Kubernetes and OS-level accounts.
- Test failover between cluster nodes for high availability.
Monitoring & Observability
Operate and maintain the
Prometheus/Grafana
and
Kibana
- stacks (provided but unsupported by GV).
- Ensure metrics collection, alerting, and log integration into broader observability pipelines.
Backup & Recovery
- Script and schedule full system backups using
- dump
and
- LVM snapshots
- .
- Securely offload backups and regularly test restoration procedures, including GRUB reinstall and partition recovery.
Security & Updates
- Apply OS-level patching, rebuild initramfs, and reinstall bootloaders as needed.
- Enforce secure separation of user accounts and access policies.
AMPP Node Responsibilities
- Maintain Linux OS and runtime environments on media-processing nodes.
- Monitor CPU, disk, network, and temperature for low-latency, high-availability performance.
- Troubleshoot connectivity and system-level issues.
- What you’ll bring
- Strong Linux administration and troubleshooting skills.
- Kubernetes (K3s) expertise: cluster lifecycle, etcd recovery, RBAC, persistent volumes.
Familiarity with
RabbitMQ
,
MongoDB
- , scripting (Bash, Python), and automation tools (Ansible, Terraform).
- Networking knowledge (Cisco): multicast, QoS, VPN, routing.
- Broadcast domain knowledge: newsroom workflows, sports live production, playout systems.
Qualifications
- Bachelor’s degree in Computer Science, IT, or equivalent experience.
- 5–8+ years in broadcast engineering and/or DevOps for live production.
- Ability to participate in on-call rotation and travel within APAC.
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring