Create and maintain optimal data processing pipeline architecture
Build the infrastructure required for optimal extraction, transformation, and loading of data using SQL and big data technologies
Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability
Build analytics tools that utilize the data pipeline to provide actionable insights into customer acquisition, operational efficiency and other key business performance metrics
Interpret data, analyze results using statistical techniques and provide ongoing reports
Identify, analyze, and interpret trends or patterns in complex data sets
Work with stakeholders including the Executive, Engineering, and Operation teams to assist with data-related technical issues and support their data infrastructure, analysis, and reporting needs
Keep our data secure through multiple data centers and AWS regions
Requirements
MS in Computer Science, Electrical Engineering, or related field
1+ years of software engineering experience
Expertise in Python, SQL, microservices, databases, data pipelines
Technical expertise regarding data models, database design and development, data mining and segmentation techniques
Experience with large scale data processing, e.g. MapReduce, Spark, Kafka
Experience with AWS core technologies such as S3, EC2, RDS
Strong knowledge of and experience with reporting packages.
Tech Stack
AWS
EC2
Kafka
MapReduce
Microservices
Python
Spark
SQL
Benefits
Catered free lunch
Unlimited snacks and beverages
Highly competitive salary and benefits package, including 401(k) plan.