Staff DevOps Engineer

hmhco

Pune 8 Years Exp Posted 11d ago

Job Description

You will constantly be asking, what are the most important platform and infrastructure challenges we need to solve today that will improve the reliability, scalability, security, and developer experience of HMH's engineering organization.

  • Define the technical vision and architecture for HMH's DevOps and platform engineering capabilities.
  • Lead cross-functional initiatives spanning cloud infrastructure, CI/CD, observability, security, and developer platforms.
  • Drive AI-enhanced DevOps automation through autonomous agents and intelligent operational workflows.
  • Partner with Engineering, Architecture, Security, SRE, and Product teams to establish reusable platform capabilities and engineering standards.
  • Evaluate emerging technologies and influence long-term technology strategy.

 

Key Responsibilities

  • Lead the design and evolution of cloud infrastructure using Infrastructure as Code (Terraform, CloudFormation).
  • Define reusable platform patterns, golden paths, and self-service developer capabilities.
  • Own architecture decisions for Infrastructure patterns, CI/CD, Kubernetes platforms, observability, secrets management, and cloud governance.
  • Drive improvements in reliability, scalability, resiliency, security, and cloud cost optimization.
  • Lead major production incident reviews and drive long-term architectural improvements.
  • Mentor Senior Engineers and provide technical leadership across multiple engineering teams.
  • Establish engineering standards, best practices, and reference architectures.
  • Design and implement AI-enabled operational workflows using AWS Bedrock and agentic technologies.

 

Skills & Experience

  • 8+ years of hands-on DevOps, SRE, or Platform Engineering experience in Agile environments.
  • Expert knowledge of AWS architecture and cloud-native design patterns.
  • Deep expertise with Terraform, CloudFormation, Kubernetes, Docker, and CI/CD platforms.
  • Experience designing highly available, large-scale distributed systems.
  • Strong expertise with observability platforms including Datadog, Grafana, and Prometheus.
  • Strong understanding of networking, IAM, cloud security, secrets management, and governance.
  • Proven ability to influence engineering organizations without formal authority.
  • Excellent communication, architecture, and mentoring skills.

 

Similar Openings for You