Live opening · Posted 7 days ago

Pyspark developer

Tata Consultancy Services (TCS) · Chennai, Tamil Nadu, India (On-site)
Linkedin No
You are 7 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 7 days ago
CompanyTata Consultancy Services (TCS)
LocationChennai, Tamil Nadu, India (On-site)
Work modeNo
SourceLinkedin
Listed7 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
190 min from Linkedin publishing this role to us finding it
11 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
73,859 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Role- PySpark Developer
Year of Experience- 4 to 15 years
Location -Pune, Hyderabad, Chennai, Bangalore
Technical Skills:
Pyspark
Python concepts and Framework
Spark Architecture
Big Data
SQL
Job Description:
Job Requirements*
Good work experience on Big Data Platforms like Hadoop, PySpark, Scala, Hive, Impala, SQL, Python
Good Python, Pyspark, Big Data experience
Spark UI/Optimization/debugging techniques
Good python scripting skills
Intermediate SQL exposure – Subquery, Joins, CTE’s
Database technologies
Key Responsibilities
*Design scalable PySpark-based test architectures for ETL/data pipelines, including modular frameworks for batch processing
.Architect end-to-end data validation systems in Hadoop environment for lineage, schema evolutio
nLead system design for Hadoop/Hive test environments, including YARN resource management, dynamic partitioning
.Exposure to Zephyr-Jira-ServiceNow integrated test management systems with experience on API-driven automation
.Design CI/CD test pipelines for PySpark/Hadoop jobs, incorporating artifact management, parallel execution, and blue-green deployments
.Create data quality system designs using PySpark integrated with Hive metadata services
.Design testing platforms, test data generator
sMentor juniors on PySpark testing basics, contribute to testing strategy discussion
sSpark session configurations for memory and core allocations for both local and cluster manager setting
sData handling with distributed file systems like HDFS and writing back to hive table
sImplementation of Partitioning, caching techniques in organizing code for transformation pipeline
sPerformance tuning implementation like salting, minimizing shufflin
g

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App