Home/Job List/Senior Site Reliability Engineer Senior Devops Engineer - AWS
Sony Pictures Networks India

Senior Site Reliability Engineer Senior Devops Engineer - AWS

Sony Pictures Networks India

Mumbai, Maharashtra, India
Full-Time
Posted 2 months ago

Job Description & Responsibilities

Purpose As a Principal Platform Engineer / Principal SRE Engineer, you will be responsible for building and operating scalable, reliable, and highly automated infrastructure powering SonyLIVs digital streaming platform. This role requires deep hands-on expertise across Kubernetes, cloud infrastructure, CI/CD, observability, production operations, and automation. You will work closely with engineering, video, CDN, security, and platform teams to improve system reliability, deployment efficiency, scalability, and operational excellence. This is a deeply technical individual contributor role focused on execution, debugging, automation, and solving large-scale infrastructure and reliability challenges.

Key Responsibilities

Infrastructure & Platform Engineering

Design, build, and operate scalable cloud-native infrastructure on AWS

Build reusable Infrastructure-as-Code modules and automation frameworks using Terraform

Manage and optimize Kubernetes-based production platforms and containerized workloads

Improve infrastructure scalability, resiliency, reliability, and operational efficiency

Reliability Engineering & Production Operations

Take ownership of production reliability, system uptime, and operational excellence

Participate in incident response, production troubleshooting, root cause analysis, and permanent fix implementation

Support critical production systems during live events and high traffic situations

Define and improve operational standards around SLOs, SLIs, error budgets, and MTTR reduction

Observability & Monitoring

Build and maintain observability platforms across metrics, logs, traces, dashboards, and alerting

Implement and improve monitoring solutions using Prometheus, Grafana, OpenTelemetry, ELK/Loki and distributed tracing systems

Improve proactive detection, alert quality, debugging workflows, and operational visibility

CI/CD & Automation

Build and improve CI/CD pipelines and deployment automation workflows

Implement GitOps and progressive delivery practices including canary and blue-green deployments

Automate repetitive operational and infrastructure tasks using Python, Go, or Bash

Security & Platform Governance

Implement security-as-code practices across infrastructure and CI/CD systems

Work on vulnerability scanning, secrets management, SBOM tooling, and container/image security

Contribute to platform governance, policy enforcement, and infrastructure standardization

Engineering Collaboration

Work closely with application, video, data, CDN, and security teams to improve platform stability and scalability

Help improve engineering productivity through internal tooling and platform capabilitie

Drive operational best practices across production systems and engineering workflows

Technical Competencies

~8+ years of hands-on experience in Platform Engineering, SRE, DevOps, Infrastructure Engineering, or Cloud Engineering

Strong hands-on expertise with

Terraform

Jenkins

Bitbucket Pipelines

Kubernetes

Docker

Strong experience with

AWS cloud platform

Cloud-native architectures

Distributed systems

Production-scale infrastructure operations

Strong understanding of

CI/CD pipelines

GitOps workflows

Progressive delivery patterns

Reliability engineering principles

Hands-on experience with observability technologies

Prometheus

Grafana

OpenTelemetry

ELK / Loki

Distributed tracing systems

Experience defining and operating

SLOs / SLIs

Error Budgets

DORA Metrics

Strong scripting/programming skills in

Python

Go

Bash

Hands-on development experience with

React

Python

Golang

Strong debugging, troubleshooting, and production incident handling skills

Experience with OPA/Rego or Conftest for Policy-as-Code

Familiarity with Chaos Engineering tools such as

Litmus

Gremlin

Chaos Monkey

Experience building Internal Developer Platforms (IDP) or developer portals like Backstage

Experience with

Secrets management systems (Vault or equivalent)

Vulnerability scanning

SBOM tooling

Container/image signin

Experience working on

OTT platforms

Video streaming systems

CDN infrastructure

Live sports streaming platforms

Understanding of video delivery workflows, adaptive bitrate streaming, and large-scale traffic management

Why join us

Sony Pictures Networks India Private Limited (SPNI) (formerly known as Culver Max Entertainment Private Limited) is an indirect wholly owned subsidiary of Sony Group Corporation, Japan. A leading media and entertainment conglomerate, SPNI comprises 29 Premium Channels in both SD and HD formats, including leading Hindi General Entertainment Television Channels - Sony Entertainment Television; Sony SAB and Sony PAL; Marathi General Entertainment Channel - Sony Marathi; Bangla General Entertainment Channel - Sony AATH; Hindi Movie Channels - Sony MAX, Sony MAX 1, Sony MAX 2 and Sony WAH; renowned destination for sports fans - Sony Sports Network comprising Sony Sports Ten 1, Sony Sports Ten 2, Sony Sports Ten 3, Sony Sports Ten 4, Sony Sports Ten 5; Kids Entertainment Channel - Sony Purpo

Required Skills

ReactPythonGoAWSDockerKubernetesCI/CDTerraform

Job Details

Employment TypeFull-Time
Work ModeOn-Site
Experience813 years
Positions1

Posted by

N/A

Posted on:

30 Jun 2026

About Sony Pictures Networks India

More open roles

Browse all jobs →