BigData / Hadoop Engineer with Azure - Fulltime

  • Full-time

Job Description

Role: BigData / Hadoop Engineer with Azure

Location: Pittsburgh or Houston - Remote

Mode: Full-time / Permanent

 

It's basically a Hadoop Engineer with Azure

 

Key Responsibilities:

• Design, develop and maintain an optimal data pipeline architecture
 using both structured data sources and big data for both on-premise and
 cloud-based environments in both streaming and real time.

• Develop and automate ETL code using scripting languages, ETL tools
 and job scheduling software to support all reporting and analytical 
data needs.

• Design and build dimensional data models to support the data warehouse initiatives.

• Assemble large, complex data sets that meet the analytical needs of the data science team.

• Assess new data sources to better understand availability and quality of data.

• Identify, design, and implement internal process improvements: 
automating manual processes, optimizing data pipeline performance, 
re-designing infrastructure for greater scalability and access to 
information.

• Participate in requirements gathering sessions to distill technical requirements from business requests.

• Collaborate with business partners to productionize, optimize, and scale enterprise analytics.

• Collaborate with data architects and modelers on data store designs and best practices


Education/Certifications:

• Bachelor’s degree in Computer Science, Engineering, Information Science, Math or related discipline

• Data engineering, data management or cloud certification is a plus

Experience/Minimum Requirements:

• Five (5)+ years’ experience in traditional and modern Big Data 
technologies (HDFS, Hadoop, Hive, Pig, Sqoop, Kafka, Apache Spark, 
hBase, Oozie, No SQL databases, PostgreSQL, GIT, Python, REST API, 
Snowflake, etc.)

• Two (2)+ years’ experience building data platforms using Azure stack (Azure Data Factory, Azure DataBricks, etc.)

• Experience with object-oriented/object function scripting languages: Python, Java, C++, Scala

• Experience extracting/querying/joining large data sets at scale

• Experience utilizing Snowflake to build data marts with the data residing in Azure storage is a plus

Other Skills/Abilities:

• Thorough understanding of relational, columnar and NoSQL database architectures and industry best practices for development

• Understanding of dimensional data modeling for designing and building data warehouses

• Excellent advanced SQL coding and performance tuning skills

• Experience with parsing data formats such as XML/JSON and leveraging external APIs

• Understanding of agile development methodologies

• Ability to work in a team-oriented, collaborative environment; good interpersonal skills

• Strong analytical and problem-solving skills; ability to weigh 
various suggested technical solutions against the original business 
needs and choose the most cost-effective solution

• Keen attention to detail and ability to access impact of design changes prior to implementation

• Self-driven, highly motivated and ability to learn quick

• Ability to effectively prioritize and execute tasks in a high-pressure environment

• Strong customer service orientation

• Ability to present and explain technical information to diverse 
types of audiences in a way that establishes rapport and gains 
understanding

• Work experience with geospatial data and spatial analytics is preferred
 

Qualifications

This is Fulltime position and looking for only USC/GC

No H1B or EAD's

Additional Information

All your information will be kept confidential according to EEO guidelines.

By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply