Job Description
Job Description:
Must-Have**
(Ideally should not be
more than 3-5)
• Strong experience with MongoDB for data lake management. • Proficiency in Scala programming language.
• Hands-on experience with GCP Dataflow, Pub/Sub, and other GCP services (BigQuery, Cloud Storage, etc.).
• Solid understanding of distributed data processing and streaming architectures.
• Experience with data modeling, ETL processes, and performance tuning.
• Familiarity with CI/CD pipelines and version control (Git).
• Excellent problem-solving and communication skills.
• candidate will design, develop, and optimize large-scale data processing pipelines and ensure efficient data storage and retrieval for analytics and business intelligence.
Good-to-Have
• Experience with Apache Beam.
• Knowledge of Kubernetes or containerized deployments.
• Exposure to data governance and security best practices on cloud platforms.
