About this role
<h3>Primary & Must have:</h3><p> Python/PySpark & GCP Cloud Must with experience </p><p> Hadoop Ecosystem Exposure (knowledge on HDFS, Hive, Big data) </p><p> SQL (Strong in SQL as this is the base to whatever we do in HQL) </p><p> CLOUD working experience- GCP preferred </p><h3>Job Description</h3><h3> Total IT exp of 5-12 years </h3><h3> Knowledge of Cloud (GCP). </h3><h3> Analyze and organize raw data </h3><h3> Build data systems and pipelines </h3><p> Evaluate business needs and objectives </p><h3> Interpret trends and patterns </h3><p> Conduct complex data analysis and report on results </p><p> Prepare data for prescriptive and predictive modeling </p><h3> Build algorithms and prototypes </h3><p> Combine raw information from different sources </p><p> Explore ways to enhance data quality and reliability </p><p> Identify opportunities for data acquisition </p><p> Develop analytical tools and programs </p><p> Collaborate with data scientists and architects on several projects </p><p> Technical expertise with data models, data mining, and segmentation techniques </p><p> Knowledge of programming languages (e.g. Java and Python) </p><p> Hands-on experience with SQL database design </p><p> Great numerical and analytical skills </p><h3>Soft Skills:</h3><h3> Good communication skills </h3><p> Flexible to work and learn on new technologies </p><p>Originally posted on <a href="https://himalayas.app">Himalayas</a></p>
big-data-developerdata-engineergcp-developercloud-data-engineerdata-pipeline-engineerremote-data-engineerdata-developer