Service Line Specialist

cognizant

Chennai, India 2 Years Exp Posted 69d ago

Job Description

  • Architect, build, and optimize multi-agent AI solutions using frameworks like LangChain, LlamaIndex, and AutoGen.

  • Design and implement complex data ingestion and processing pipelines, ensuring robust data semantics for Retrieval-Augmented Generation (RAG) architectures.

  • Develop and fine-tune Large Language Models (LLMs) and other foundational models using the Nvidia NeMo framework.

  • Deploy and manage high-throughput, low-latency model inference services using Nvidia Triton Inference Server.

  • Conduct performance profiling and optimization of AI workloads on Nvidia A100 and H100 Tensor Core GPUs.

  • Integrate and manage specialized vector databases such as Milvus, Pinecone, and Weaviate for high-dimensional data indexing and search.

  • Leverage Cognizant's ATK platform to orchestrate complex agentic workflows and ensure seamless integration with enterprise systems.

    • Collaborate with infrastructure teams to ensure optimal configuration of Kubernetes clusters for GPU-accelerated workloads.

Similar Openings for You