Home/Job List/Senior AI Platform / MLOps Engineer
Ampera Technologies

Senior AI Platform / MLOps Engineer

Ampera Technologies

India
Full-Time
Posted 14 days ago

Job Description & Responsibilities

Title : Senior AI Platform / MLOps Engineer

Experience : 6+ years

Work type : Chennai - Work from Office/other locations - Remote

Employment Type : Full Time

Notice Period : Immediate

Work Day :Mon to Fri

Key Responsibilities

  • Install, configure and operate OpenShift, NVIDIA GPU operator, OpenShift AI, and NIM microservices on 12× RTX PRO 6000 across two servers; single-node and HA control-plane topologies
  • Serving configuration and tuning: quantized model deployment (FP8/FP4), replica balancing, batching, KV-cache and context management
  • Azure GPU build environments: provisioning, cost control, parity with the on-prem stack via pinned container/model versions; cloud-to-factory migration with parity regression
  • GitOps CI/CD, container registry, artifact/model versioning, environment promotion; observability and audit wiring (Splunk, Prometheus/Grafana)
  • Benchmark automation: load harness, p50/p95/p99 latency, tokens/sec, GPU utilization; the capacity report data pipeline
  • Platform upgrade procedure with evaluation-regression gates; deployment runbook as a first-class deliverable

Technical Skills

  • 6+ years infrastructure/platform engineering with 3+ years production Kubernetes; OpenShift experience strongly preferred
  • Hands-on GPU inference serving in production: NIM, Triton, vLLM, or TensorRT-LLM — you have sized, deployed, and tuned LLM serving on real GPUs and can talk memory-bandwidth trade-offs
  • GitOps fluency (ArgoCD/Flux), infrastructure-as-code, container internals; comfortable in air-gapped/proxy-restricted enterprise networks
  • Observability depth: metrics, traces, log pipelines; has built performance test harnesses, not just run them
  • Azure or AWS GPU compute operations experience

Strongly preferred

  • NVIDIA GPU operator and AI Enterprise stack specifics; KServe; Milvus or pgvector operations; VAST/NFS/S3 storage integration; banking or other regulated-environment delivery

About Ampera

Ampera Technologies, a purpose driven Digital IT Services with primary focus on supporting our client with their Data, AI / ML, Accessibility and other Digital IT needs. We also ensure that equal opportunities are provided to Persons with Disabilities Talent. Ampera Technologies has its Global Headquarters in Chicago, USA and its Global Delivery Center is based out of Chennai, India. We are actively expanding our Tech Delivery team in Chennai and across India. We offer exciting benefits for our teams, such as 1) Hybrid and Remote work options available, 2) Opportunity to work directly with our Global Enterprise Clients, 3) Opportunity to learn and implement evolving Technologies, 4) Comprehensive healthcare, and 5) Conducive environment for Persons with Disability Talent meeting Physical and Digital Accessibility standards

Required Skills

AWSAzureKubernetesCI/CDMachine Learning

Job Details

Employment TypeFull-Time
Work ModeRemote
Experience6 – 11 years
Positions1

Posted by

N/A

Posted on:

24 Sept 2026

About Ampera Technologies

More open roles

Browse all jobs →