Forward Deploy AI Engineer
avathon
Job Description
- Build and fine-tune SLMs/LLMs for production use cases such as semantic search, forecasting, contract intelligence, and conversational insights. Optimize models for performance, cost, and latency using PEFT, quantization, and efficient inference techniques
- Embed supply chain knowledge into models by working closely with domain experts and product teams
- Enable agentic workflows, allowing AI Assistants to execute planning, optimization, and decision-support tasks
- Integrate models into products using custom Lambdas, and a Graph-based microservices architecture
- Apply vector search and semantic reasoning to help customers navigate complex supply chain relationships
- Work with real enterprise data—ERP, logistics, inventory, supplier, and contract datasets
- Deploy and operate models using strong MLOps practices on platforms like AWS, GCP or AZURE
- Measure and improve quality, reducing hallucinations and improving reliability through continuous evaluation
- Partner with engineering, product, and operations teams to ship impactful AI features end-to-end