Description
You will ensure the quality and validity of AI/ML systems by testing models, LLM features, and AI-driven workflows.
This role is remote.
Responsibilities
- Develop and execute test plans, test cases, and automation scripts for AI systems.
- Perform functional, regression, and performance testing on LLM features and AI-driven workflows.
- Evaluate model accuracy, bias, reliability, and edge cases to ensure system integrity.
- Collaborate with ML engineers and data scientists to document defects and track fixes.
Required Skills
- 5+ years of experience in QA/QE and software testing practices.
- Strong knowledge of Python for test automation.
- Experience with PyTest and Selenium frameworks.
- Understanding of AI/ML concepts, LLMs, and data workflows.
- Ability to design test cases for AI outputs, prompts, and model behavior.
- Familiarity with REST APIs and version control (Git).
- Knowledge of CI/CD pipelines.
Preferred Skills
- Experience testing complex AI-driven workflows and non-deterministic outputs.