Service Line Specialist
cognizant
Job Description
-
Architect, build, and optimize multi-agent AI solutions using frameworks like LangChain, LlamaIndex, and AutoGen.
-
Design and implement complex data ingestion and processing pipelines, ensuring robust data semantics for Retrieval-Augmented Generation (RAG) architectures.
-
Develop and fine-tune Large Language Models (LLMs) and other foundational models using the Nvidia NeMo framework.
-
Deploy and manage high-throughput, low-latency model inference services using Nvidia Triton Inference Server.
-
Conduct performance profiling and optimization of AI workloads on Nvidia A100 and H100 Tensor Core GPUs.
-
Integrate and manage specialized vector databases such as Milvus, Pinecone, and Weaviate for high-dimensional data indexing and search.
-
Leverage Cognizant's ATK platform to orchestrate complex agentic workflows and ensure seamless integration with enterprise systems.
-
Collaborate with infrastructure teams to ensure optimal configuration of Kubernetes clusters for GPU-accelerated workloads.
-