Live opening · Posted 16 days ago

GCP Lead Data Engineer

Halcer · Bengaluru, Karnataka, India (On-site)
Linkedin No
You are 16 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 16 days ago
CompanyHalcer
LocationBengaluru, Karnataka, India (On-site)
Work modeNo
SourceLinkedin
Listed16 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
21 min from Linkedin publishing this role to us finding it
12 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
64,960 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Job Title: GCP Lead Data Engineer
Total Experience: 8–13 years
Work Location: Bangalore
Notice Period: Immediate to 30 Days Joiners Only
Role Overview
We are seeking a hands-on Lead Data Engineer with strong technical expertise in Google Cloud Platform (GCP), Scala, PySpark, and Kafka. In this role, you will design, build, and optimize enterprise-grade batch and streaming data processing frameworks. You will lead technical implementations to enable advanced analytics, machine learning, and AI solutions for global clients while adhering to cloud security, performance, and governance best practices.
Key Responsibilities
Pipeline Engineering: Build, maintain, and optimize robust batch and real-time streaming data pipelines using Scala, PySpark, Apache Spark, and Kafka/PubSub.
Architecture & Leadership: Drive technical architecture discussions, code reviews, and DevOps standards across the data engineering team.
Cross-Functional Collaboration: Partner with Data Scientists, Business Analysts, and stakeholders to deliver data platforms that support analytics and machine learning workflows.
Operations & Quality: Ensure high data quality, pipeline reliability, and optimal performance across dynamic distributed systems.
Production Support & Optimization: Perform root cause analysis on production issues, optimize resource costs, and implement strict cloud security and governance standards.
Required Core Skills
Mandatory Technical Stack: Scala, PySpark, and Apache Kafka integrated with core Data Engineering workflows.
Data Engineering Expertise: Proven track record in designing high-throughput ETL/ELT pipelines and processing large-scale datasets.
Distributed Systems: Deep understanding of distributed systems engineering, data modeling, performance tuning, and Apache Beam.
Cloud & Warehousing: Hands-on experience with SQL, NoSQL systems, and cloud data platform ecosystems (BigQuery, GCP).
Software Development Practices: Strong experience with CI/CD automation, version control (Git), and Agile delivery methodologies.
Acceptable Alternative Skill Profiles
Candidates matching any of the following technical skill combinations are also eligible:
GCP + Scala (with or without Kafka)
GCP + Apache Spark
Scala + Apache Spark + Kafka
GCP (BigQuery) + PySpark + SQL (Kafka is a plus)
Preferred (Good-to-Have) Skills
Exposure to building real-time streaming architectures using Google Cloud Pub/Sub alongside Kafka.
Python proficiency combined with Scala background.
Familiarity with machine learning data pipelines, feature stores, and MLOps integrations.

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App