Description
You design, build, and deploy machine learning models using Databricks, Python, and Spark to solve business problems.
This role is on-site.
Responsibilities
- Design and deploy ML models using Databricks (MLflow, Spark ML, Python) and implement end-to-end pipelines from ingestion to monitoring.
- Leverage Databricks Lakehouse (Delta Lake, Unity Catalog) for scalable analytics and optimize Spark jobs for performance and cost.
- Configure Databricks Genie for self-service analytics, designing semantic layers optimized for natural language queries.
- Partner with product owners and business leaders to identify high-value use cases and translate complex results into clear insights.
Required Skills
- Bachelor’s or Master’s degree in Data Science, Computer Science, Statistics, Engineering, or a related field.
- 4+ years of experience in data science or advanced analytics.
- Hands-on experience with Databricks, Apache Spark, and Python (PySpark, Pandas, NumPy, Scikit-learn).
- Strong programming skills in Python and solid understanding of SQL and data modeling.
- Experience building and deploying ML models in production, including MLflow and model lifecycle management.
- Knowledge of data governance, lineage, and security best practices.
Preferred Skills
- Experience configuring Genie for natural language analytics.
- Background in data engineering collaboration for reliable dataset quality.