Job Description
Data Engineer
Position Overview:
The successful Engineer will be responsible for expanding and optimizing our
Telematic data and data pipeline architecture, as well as optimizing data flow
and collection for cross functional teams. The ideal candidate is an
experienced data pipeline builder who enjoys optimizing data systems and
building them from the ground up. The Engineer will support our application
developers, database architects, data analysts, and data scientists on company
initiatives and will ensure optimal data delivery architecture is consistent
throughout ongoing projects. They must be self-directed and comfortable
supporting the data needs of multiple teams, systems, and products. The right
candidate will be excited by the prospect of optimizing or even reshaping our
company’s data architecture to support our next generation of products and
data initiatives.
Responsibilities:
- Create and maintain data pipelines and code pipelines.
-
Assemble large, complex data sets that meet functional / non-functional
business requirements.
-
Identify, design, and implement internal process improvements: automating
manual processes, optimizing data delivery, re-designing infrastructure for
greater scalability, etc.
-
Define the infrastructure required for optimal extraction, transformation,
and loading of data from a wide variety of data sources using SQL and AWS
technologies and services.
-
Support data sets in analytics tools to provide actionable insights into
operational efficiency and other key business performance metrics.
-
Work with stakeholders to assist with data-related technical issues and
support their data needs.
-
Work with data and analytics experts to strive for greater functionality in
our data systems.
Required Qualifications:
-
Bachelor’s Degree with 3-5 years relevant work experience as a data
engineer.
- Strong problem-solving skills with an emphasis on data value.
- Experience working with and creating data architectures.
-
Advanced working SQL knowledge and experience working with relational
databases, query authoring (SQL) as well as working familiarity with a
variety of databases.
-
Experience performing root cause analysis on internal and external data and
processes to answer specific business questions and identify opportunities
for improvement.
- Strong analytic skills related to working with unstructured datasets.
-
Build processes supporting data transformation, data structures, metadata,
dependency and workload management.
-
A successful history of manipulating, processing and extracting value from
large, disconnected datasets including business systems like PLM, ERP, and
MES.
-
Experience supporting and working with cross-functional teams in a dynamic
environment.
-
Experience using the following software/tools:
- Experience with big data tools: Kafka.
-
Experience with relational SQL and NoSQL databases, including Redshift,
Oracle.
-
Experience using AWS cloud services: Redshift, S3, Data Pipeline,
Elasticsearch, Glue, EMR, Athena.
Preferred Qualifications:
- Strong project management and organizational skills.
- Experience with Agile Scrum teams and associated tools.
- 5-7 years of experience manipulating data sets and data pipelines.
-
Coding knowledge and experience with several languages and environments.
-
Experience preparing data for visualizations for stakeholders using:
Qlikview, QuickSight, Tableau, Business Objects, etc.
-
Working knowledge of message queuing, stream processing, and highly scalable
‘big data’ data stores.
Language:
Fluent English Required