The client puts data at the very heart of their business and are excited to be building a new enterprise data platform which provides powerful analytic capabilities for them internally and subsequently their customers. Therefore, this is a key role within their data engineering team working on this new data platform.
The role will involve loading data from multiple operational systems into the client’s vast data lake. Due to the large amount of data, they use HiveQL as well as Hadoop’s inbuilt data loading tools for much of the processing.
While working within the client’s dynamic team, you will be responsible for building and data loading using the Hortonworks toolset. This will also include creating data migration scripts to take all data stores in the cluster from one release to the next. This is ultimately a great opportunity to gain skills in the Hortonworks Data Platform and related tools in an innovative and passionate environment.
The ideal candidate will have experience with the following:
Please note, other relevant skills and experience are also considered.
Job details are sourced from the employer's original posting.
Open job postingAbout the company
The Client provides expert advisory and implementation services for open source big data solutions. As the first and only pure-play big data services firm, their Data Scientists and Engineers are trusted advisors to the world's most innovative companies. Their experienced teams combine a distinctive methodology and a proven framework that includes tested design patterns and pre-built components, to help clients build applications faster. The Client helps Customers leverage Big Data analytics by integrating open source platforms, such as Hadoop, NoSQL and Streaming Engines, with best-of-breed data warehousing environments. Service offers include: a Big Data roadmap, Data Engineering, Data Lake and Analytic Operations, Training and ongoing Big Data Solution Support.