Associate – ML Engineer

bain

New Delhi 3 Years Exp Posted 36d ago

Job Description

ML Pipeline Support

  • Support deployment and day-to-day operations of Pyxis ML microservices; including classifiers, clustering, LLM pipelines, and rule engines
  • Manage training data curation across large datasets and own model development and deployment end to end
  • Run and validate training and prediction pipelines for supervised and unsupervised models
  • Monitor async job queues and help troubleshoot failures across Redis Queue (RQ), EC2 machines, and other routers
  • Assist with model artifact management powering MLOps visibility and RCA
  • Assist with Snowflake-based data ingestion, SQL rule resolution, and data management for model development
  • Support LLM pipeline operations, including OpenAI/Anthropic batch jobs, token monitoring, and output validation

LLM & GenAI Support

  • Help manage OpenAI/Anthropic batch and synchronous completions pipelines
  • Monitor token usage, job statuses, and flag anomalies
  • Support prompt template updates and output validation

Infrastructure & Deployment

  • Support CI/CD workflows for ML services using GitHub Actions – Docker and Terraform builds, ECS deployments
  • Help maintain containerized ML services on AWS 
  • Assist with S3 integrations, IAM configurations, and queue management

Collaboration

  • Work closely with Data Scientists and DevOps on pipeline onboarding
  • Maintain documentation and runbooks for operational processes
  • Partner directly with Client Success and Operations on model outputs for clients

 

About you

About you

  • 3-5 years of experience in a data, ML, or software engineering role
  • Working knowledge of Python - exposure to FastAPI or similar frameworks is a plus
  • Basic understanding of AWS services - S3, ECS, IAM
  • Familiarity with Docker and containerization concepts
  • Exposure to ML frameworks - TensorFlow/Keras or Scikit-learn
  • Intermediate SQL skills and comfort working with cloud data warehouses like Snowflake

Good to have

  • Exposure to CI/CD pipelines and GitHub Actions
  • Familiarity with LLM APIs or prompt engineering
  • Understanding of NLP/text vectorization concepts
  • Interest in retail data, product taxonomy, or consumer analytics

 

Similar Openings for You