Job Description
Responsibilities
Create and maintain optimal data pipeline architecture, assemble large, complex data setsthat meet functional / non-functional business requirements using Python and SQL / AWS / Snowflakes.
Identify, design, and implement internal process improvements through: automating manual processes using Python, optimizing data delivery, re-designing infrastructurefor greater scalability, etc.
Build the infrastructure required for optimal extraction, transformation, and loading ofdata from a wide variety of data sources using SQL / AWS / Snowflakes technologies.
Build analytics tools that utilize the data pipeline to provide actionable insights into customer acquisition, operational efficiency and other key business performance metrics.
Work with stakeholders including the Executive, Product, Data and Design teams toassist with data-related technical issues and support their data infrastructure needs.
Keep our data separated and secure across national boundaries through multiple datacenters and AWS regions.
Work with data and analytics experts to strive for greater functionality in our data systems.
Experience :-6+ years of experience in a Python Scripting and Data specific role, with bachelor s degree.
Experience with data processing and cleaning libraries e.g. Pandas, numpy, etc., web scraping/ web crawling for automation of processes, API s and how they work.
Debugging code if it fails and find the solution. Should have basic knowledge of SQLserver job activity monitoring and also of Snowflake.
Experience with relational SQL and NoSQL databases, including PostgreSQL and Cassandra.
Experience with most or all of the following cloud services: AWS, Azure, Snowflake, Google