Experience
12 - 14 yrs
Job Location
Bengaluru / Bangalore, India
Vacancy
1
Designation
Senior Data Architect
Job Type
ONSITE
Job Description
Senior Architect - Databricks
Designation: Senior Manager (L5)
Profile Summary
We are seeking an experienced professional who apart from the required mathematical and statistical expertise also possesses the natural curiosity and creative mind to ask questions, connect the dots, and uncover opportunities that lie hidden with the ultimate goal of realizing the data's full potential.
Key Responsibilities
- Developing Modern Data Warehouse solutions using Databricks and AWS/ Azure Stack
- Ability to provide solutions that are forward-thinking in data engineering and analytics space
- Collaboration with DW/BI leads to understanding new ETL pipeline development requirements.
- Triage issues to find gaps in existing pipelines and fix the issues
- Work with business to understand the need in reporting layer and develop data model to fulfill reporting needs
- Help joiner team members to resolve issues and technical challenges.
- Drive technical discussion with client architect and team members
- Orchestrate the data pipelines in scheduler via Airflow
Requirements
- Bachelor's and/or master's degree in computer science or equivalent experience.
- Must have total 12+ yrs. of IT experience and 6+ years experience in Data warehouse/ETL projects.
- Deep understanding of Star and Snowflake dimensional modelling.
- Strong knowledge of Data Management principles
- Good understanding of Databricks Data & AI platform and Databricks Delta Lake Architecture
- Should have hands-on experience in SQL, Python and Spark (PySpark)
- Candidate must have experience in AWS/ Azure stack
- Desirable to have ETL with batch and streaming (Kinesis).
- Experience in building ETL / data warehouse transformation processes
- Experience with Apache Kafka for use with streaming data / event-based data
- Experience with other Open-Source big data products Hadoop (incl. Hive, Pig, Impala)
- Experience with Open Source non-relational / NoSQL data repositories (incl. MongoDB, Cassandra, Neo4J)
- Experience working with structured and unstructured data including imaging & geospatial data.
- Experience working in a Dev/Ops environment with tools such as Terraform, CircleCI, GIT.
- Proficiency in RDBMS, complex SQL, PL/SQL, Unix Shell Scripting, performance tuning and troubleshoot
- Databricks Certified Data Engineer Associate/Professional Certification (Desirable).
- Should have experience working in Agile methodology.
- 12+ Years of experience in Data & Analytics with Good Communication and presentations skills.
- At least 4 years experience in Databricks implementations, 2 large scale data warehouse end-to- end implementation experience.
- Must have Databricks certified architect.
- Proficiency in SQL and experience with scripting languages (e.g., Python, spark, Py spark) for data manipulation and automation.
- Solid understanding of cloud platforms (AWS, Azure, GCP) and their integration with Databricks.
- Familiarity with data governance and data management practices. exposure to Data sharing, unity catalog, DBT, replication tools, performance tuning will be added advantage / must have skills.
Keywords
CircleCI
No Referrers Available
There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.
