Lead the design and development of an AI voice agent platform. Build architecture, integrate speech processing, and optimize model performance.
Responsibilities
- Design and build the architecture for an AI voice agent platform.
- Customize open-source LLMs using Hugging Face to meet specific business requirements.
- Integrate TTS and STT functionalities for audio and speech processing.
- Implement embeddings and vector databases to optimize data storage and retrieval.
- Use call recordings and transcripts to re-train models for improved intelligence.
Required Skills
- Hands-on experience in AI and Machine Learning.
- Expertise in audio and speech processing, including TTS and STT.
- Experience with LLMs and Hugging Face for model training and customization.
- Proficiency with Eleven Labs, Play.ht, and Deepgram.
- Practical knowledge of embeddings and vector databases.
- Experience with PyTorch, RAG, and AI/ML workflows.
- Ability to perform hands-on coding and architectural design.
Preferred Skills
- Familiarity with GPTCache or similar technologies for caching answers and audio files to reduce latency.
- Knowledge of GPU-based hosting companies.