Description
Design, develop, and optimize AI solutions using LLMs and generative models to solve complex business challenges.
This role is hybrid.
Responsibilities
- Design and implement scalable APIs and AI pipelines using Python, PyTorch, or TensorFlow.
- Optimize SDLC workflows for development and QA to improve productivity and delivery velocity.
- Engineer effective prompts and optimization strategies to maximize model accuracy and relevance.
- Integrate proprietary AI platforms with external data sources to deliver custom MCP agents.
- Monitor and evaluate model performance, ensuring robustness, fairness, and responsible AI standards.
Required Skills
- Strong proficiency in Python and AI/ML libraries (PyTorch, TensorFlow, Hugging Face Transformers).
- 5+ years of experience in software engineering or AI/ML roles.
- Solid understanding of LLM architectures, training paradigms, and prompt engineering techniques.
- Hands-on experience designing scalable APIs and AI pipelines.
- Familiarity with cloud platforms (AWS preferred, Azure or GCP).
- Experience working with Docker and Kubernetes.
- Good understanding of the Software Development Life Cycle.
- Professional level of English and effective communication skills.
Preferred Skills
- Strong proficiency in Java, including hands-on experience building MCP agents.
- Experience with LangChain, LlamaIndex, or similar LLM orchestration frameworks.
- Familiarity with vector databases (OpenSearch, Pinecone) and RAG architectures.