Description
Design and implement production-ready Generative AI solutions using LLMs, RAG, and multimodal approaches.
This role is hybrid.
Responsibilities
- Develop and fine-tune transformer-based (BERT, GPT, Llama) and diffuser-based (Stable Diffusion) models.
- Architect communication flows for AI systems, incorporating RAG and multimodal capabilities.
- Build production-ready code using Python, C, C++, C#, or Java.
- Implement prompt engineering strategies and context-based prompt management.
- Deploy and manage AI infrastructure on Azure, AWS, or GCP.
Required Skills
- 2+ years of experience in AI/ML products or systems.
- Proficiency in Python, C, C++, C#, or Java.
- Hands-on experience with Generative AI models, specifically LLMs and RAG architectures.
- Experience with open-source models from Hugging Face.
- Knowledge of LangChain and Vector Databases.
- Cloud platform experience with Microsoft Azure, AWS, or GCP.
- Familiarity with transformer and diffuser-based model architectures.