Description
You will design, develop, and deploy generative AI models, including fine-tuning pre-trained LLMs for specific tasks. You own the full AI development lifecycle, from data pre-processing to evaluation, optimization, and deployment.
This role is on-site.
Responsibilities
- Design and implement architecture patterns for generative AI applications like text generation and code synthesis.
- Fine-tune open-source and pre-trained models, ensuring performance and alignment with business goals.
- Manage the end-to-end AI lifecycle, including data preparation, training, and production deployment.
- Collaborate with data scientists and engineers to define requirements and integrate AI solutions seamlessly.
- Conduct code reviews and POCs to maintain high-quality, performant code and evaluate new technologies.
Required Skills
- 5+ years of core development and architecture experience.
- Strong proficiency in Python and deep learning frameworks (TensorFlow or PyTorch).
- Deep understanding of machine learning fundamentals, deep learning architectures, and NLP.
- Experience implementing design patterns and developing generative AI models.
- Ability to write high-quality, performant code and conduct rigorous code reviews.
- Experience with cloud platforms (AWS, GCP, Azure) and OpenShift.
- Champion responsible AI practices, focusing on bias mitigation and explainability.
Preferred Skills
- Master’s degree in Computer Science, AI, or a related field.
- Experience fine-tuning open-source AI models and domain-specific industry standards.
- Active participation in technology forums or open-source communities.