Description
You design and build scalable AI-driven applications, owning the lifecycle from API development to production deployment.
This role is on-site.
Responsibilities
- Design and develop AI applications using Python and FastAPI.
- Integrate and optimize Large Language Models via OpenAI APIs.
- Implement retrieval-augmented generation (RAG) pipelines, embeddings, and vector databases.
- Leverage LangChain for orchestration and prompt engineering.
- Collaborate on deployment, performance tuning, and observability for production-ready solutions.
Required Skills
- 9+ years of professional software engineering experience.
- Expert proficiency in Python.
- Strong experience with FastAPI for building API services.
- Hands-on experience integrating and optimizing LLMs.
- Working knowledge of OpenAI APIs.
- Proficiency with LangChain for orchestration.
- Experience implementing RAG pipelines and vector databases.
Preferred Skills
- Experience with production-grade observability and performance tuning.