openqareer

Senior/Lead DevOps + MLOps Engineer

Luxoft · Karnataka

# Senior/Lead DevOps + MLOps Engineer **Luxoft** · Karnataka · `On-site` 🕒 **Статус:** *Опубликовано: сегодня* · *Источник: Indeed* --- ### About the Role Project description We are seeking a Senior DevOps / Platform Engineer with deep AWS expertise to evolve and operate our core cloud platform, while enabling Enterprise-grade Agentic AI capabilities on top of established infrastructure. This role is focused on platform stability, scalability, and developer enablement, ensuring that traditional services and emerging agentic systems coexist securely and reliably. You will play a key role in transforming our platform into a foundation that supports AI-augmented and agentic SDLC workflows, without compromising operational excellence. Responsibilities Design, build, and operate a shared AWS cloud platform that supports both traditional services and agentic AI workloads. Own and evolve Infrastructure as Code (IaC) using Terraform, ensuring consistency, security, and repeatability across environments. Extend existing infrastructure to support Enterprise-grade Agentic AI systems, including: Execution runtimes for autonomous and semi-autonomous agents Secure access to data, services, and APIs Platform-level guardrails for safety, governance, and cost control Build platform abstractions, templates, and tooling that enable teams to safely consume agentic capabilities. Support and integrate agentic SDLC tools and processes, including AI-assisted development, testing, and release automation. Develop and maintain CI/CD pipelines for both traditional applications and AI-driven components. Implement platform-level observability (metrics, logs, traces) across services and agent workloads. Enforce security and compliance best practices (IAM, secrets management, encryption, least privilege). Collaborate closely with application, AI/ML, and security teams to improve developer experience and platform reliability. Act as a technical leader in architecture discussions, platform standards, and operational readiness. Skills Must have 8+ years of experience in DevOps, Platform Engineering, or Site Reliability Engineering roles. Deep, hands-on AWS expertise, including: EC2, EKS/ECS, VPC, IAM, S3, RDS/DynamoDB, CloudWatch, Lambda Strong understanding of VPC design, Security Groups, Route Tables, NACLs, and hybrid networking concepts. Ability to troubleshoot network connectivity issues, ingress/egress rules, and cloud security configurations. Strong production experience with Terraform, including: - Designing modular Terraform architectures - Managing state, environments, and multi-account setups Strong Linux administration skills: - User and permissions management - Process and service management - Package management - System troubleshooting Solid understanding of Linux commands and privilege escalation concepts (sudo, sudo su, sudoers). Hands-on scripting experience: - Python - Bash/Shell scripting Ability to automate operational tasks and troubleshooting workflows. Experience with container technologies: - Docker - Kubernetes (EKS preferred) - Container image lifecycle management and versioning/tagging strategies. Experience with artifact repositories: - JFrog Artifactory, Nexus or equivalent Understanding of image tagging, versioning, and promotion strategies. Proven experience rolling out Enterprise-grade Agentic AI infrastructure on top of existing platforms, including: - Supporting agent execution within established networking, security, and compliance boundaries - Enabling scalability, observability, and governance for agent behavior Hands-on experience supporting agentic SDLC tools and processes: - AI-assisted coding, testing, and deployment workflows - Agent-based automation within CI/CD and operational processes Experience with AWS AI services such as: Amazon Bedrock, SageMaker, Related AI/ML platform services. Solid understanding of Linux, networking, and cloud security fundamentals. Familiarity with MLOps or AI platform components Model serving, Vector databases, Feature stores Nice to have Experience designing internal developer platforms (IDPs). Knowledge of policy-as-code and governance frameworks (OPA, SCPs, tagging strategies). AWS certifications (Solutions Architect, DevOps Engineer). Experience operating platforms at enterprise scale or in regulated environments. Other Languages English: B2 Upper Intermediate Seniority Senior Bengaluru, India Req. VR-122728 DevOps BCM Industry 06/10/2026 Req. VR-122728

Наблюдалась 2026-10-06, впервые 2026-10-06, найдена на 2 площадках, источник — Indeed.

Открыть у работодателя