Live opening · Posted 10 hours ago

Data Engineer - Gen AI & LLM

Talentgigs · Bengaluru, Karnataka, India (Remote)
Linkedin No
You are 10 hours behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 10 hours ago
CompanyTalentgigs
LocationBengaluru, Karnataka, India (Remote)
Work modeNo
SourceLinkedin
Listed10 hours ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
18 min from Linkedin publishing this role to us finding it
8 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,610 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

About the Role
We are looking for a passionate Data Engineer with 5+ years of experience in designing,
developing, and maintaining scalable data platforms and ETL/ELT pipelines. The ideal candidate
should possess strong expertise in Python, SQL, cloud data services, and modern data
engineering frameworks. You will play a key role in building reliable, high-performance data
solutions that support analytics, reporting, and AI/ML initiatives.
Key Responsibilities
• Design, develop, and maintain scalable data pipelines and ETL/ELT workflows.
• Build and optimize data ingestion processes from multiple structured and unstructured
data sources.
• Develop robust data models and data warehouses for analytics and reporting.
• Design and optimize SQL queries for high-performance data processing.
• Build and maintain data lakes using cloud storage solutions.
• Implement data validation, cleansing, transformation, and quality checks.
• Integrate data solutions with cloud platforms such as AWS, Azure, or Google Cloud
Platform.
• Develop batch and real-time data processing pipelines using modern data processing
frameworks.
• Collaborate with data scientists, analysts, and application teams to deliver reliable data
solutions.
• Follow software engineering best practices, including Git, CI/CD pipelines, testing,
monitoring, and technical documentation.
Required Skills
• 5+ years of experience in Data Engineering.
• Strong programming experience in Python.
• Hands-on experience with PySpark and Apache Spark.
• Advanced SQL skills, including query optimization and performance tuning.
• Experience with AWS services/ Azure/ GCP.
• Experience building scalable ETL/ELT pipelines.
• Strong exp in Generative AI and Large Language Models (LLMs)
• Knowledge of data warehousing concepts and dimensional modeling.
• Experience working with large-scale distributed datasets.
• Familiarity with Git and CI/CD pipelines.
• Understanding Agile/Scrum methodologies.
Preferred Qualifications
• Experience with Apache Airflow or similar workflow orchestration tools.
• Knowledge of Kafka and real-time data streaming.
• Familiarity with Docker, Kubernetes, and Terraform (IaC).
• Exposure to Databricks and modern data lake technologies (Delta Lake, Apache Iceberg,
or Apache Hudi).
• Understanding data governance, metadata management, and CI/CD practices.
• Strong exp in Generative AI and Large Language Models (LLMs

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App