Skills : SQL, HIVE, ETL-Big Data / Data Warehousing, GCP, Big Data, PySpark
Job description
Knowledge of Big Data / NoSQL design and development with variety of data stores (document, column family, graph, etc.)
Knowledge of database management system products and ecosystem - e.g., storage formats, access algorithms, data management processes, administration, data maintenance, replication, high availability, encryption, etc.
Knowledge of distributed (multi-tiered) systems, algorithms, and relational & non-relational databases
Knowledge of XML/JSON and schema development/reuse, relational database management systems, Open Source
Data engineering and development experience or a related job
Experience in SQL and Python (or any other programming language like Pyspak/Java)
Experience in design and development across one or more database management systems (e.g., HBase, Hive, GCP tools, NoSQL, BigTable, BigQuery, ) as appropriate
Knowledge of data modeling tools (e.g., ErStudio, ErWin)
Experience in performance tuning of SQL queries
Experience in distributed (multi-tiered) systems, algorithms, and application development
Experience in database development and support of online transaction processing and/or online analytical processing systems