Live opening · Posted 1 day ago

Databricks

Infosys · Bengaluru East, Karnataka, India (On-site)
Linkedin No
You are 1 day behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 1 day ago
CompanyInfosys
LocationBengaluru East, Karnataka, India (On-site)
Work modeNo
SourceLinkedin
Listed1 day ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
8 min from Linkedin publishing this role to us finding it
20 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
73,883 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Primary skills:Technology->Data Engineering->Databricks
Key Responsibilities:
Develop and maintain scalable ETL pipelines using Databricks and PySpark to process large datasets efficiently.
Implement data transformations, cleansing, and enrichment logic aligned to business and analytics requirements.
Optimize Spark jobs for performance and cost by tuning partitions, caching, and cluster configurations where applicable.
Build reusable notebooks/jobs and support scheduling/orchestration of workloads within the Databricks environment.
Perform data validation, reconciliation, and quality checks to ensure accuracy and reliability of curated datasets.
Troubleshoot pipeline failures, analyze logs, and resolve issues to maintain stable production operations.
Collaborate with cross-functional teams to gather requirements, provide estimates, and deliver enhancements iteratively.
Maintain clear technical documentation for pipelines, transformations, and operational runbooks. Minimum Qualifications:
Bachelor’s degree (or equivalent) in Engineering/Technology/Computer Science or related field (BTech/BE/MSc or equivalent).
3–5 years of experience in data engineering or related roles with hands-on Databricks experience.
Strong hands-on development experience with PySpark for distributed data processing.
Proven experience building and supporting ETL pipelines in production environments.
Ability to analyze data issues, debug Spark/ETL jobs, and implement reliable fixes. Preferred Qualifications:
Experience designing end-to-end data workflows on Databricks including job scheduling, monitoring, and operational support.
Strong understanding of data modeling concepts and building curated datasets for analytics use cases.
Familiarity with Delta Lake concepts (ACID tables, incremental processing, upserts/merges) and best practices for lakehouse implementations.
Experience with performance tuning techniques for Spark workloads and handling large-scale datasets efficiently.
Exposure to CI/CD practices for data pipelines and maintaining code quality through reviews and standards. Good to have skills: Delta Lake, Spark SQL, Data Modeling, Workflow Orchestration, Performance Tuning

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App