Description
Lead the design and delivery of Generative AI solutions, owning the full lifecycle from RAG pipeline development to production deployment.
This role is on-site.
Responsibilities
- Build and enhance Generative AI applications using Azure OpenAI, LangChain, and Semantic Kernel.
- Develop RAG pipelines including chunking, embeddings, vector indexing, and retrieval logic.
- Implement LLM-based services, APIs, and Copilot extensions using FastAPI and React.
- Mentor engineering teams, conduct code reviews, and establish quality and security best practices.
- Establish evaluation frameworks (RAGAS/DeepEval) to monitor LLM accuracy, hallucinations, and safety.
Required Skills
- 7–12 years of software engineering experience, including 2–3 years in AI/LLM projects.
- Hands-on expertise with Azure OpenAI, Cognitive Search, Azure ML, and related Azure services.
- Strong proficiency in Python, FastAPI, React, and TypeScript.
- Deep understanding of LLMs, RAG pipelines, embeddings, vector databases, and prompt engineering.
- Experience with SQL, MySQL, Cosmos DB, and handling both structured and unstructured data.
- Familiarity with LangChain, Semantic Kernel, and vector databases like Pinecone or Redis Vector.
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
Preferred Skills
- Experience with Copilot extensibility, including plugins and Graph Connectors.
- Exposure to MLOps on Azure ML and ETL/data preparation for RAG.