Live opening · Posted 8 hours ago

PySpark Developer

Tata Consultancy Services Limited · Pune District, Maharashtra, India (On-site)
Linkedin No
JobBeeper subscribers received an alert for this role.

At a glance

The key details from the original listing.

Posted 8 hours ago
CompanyTata Consultancy Services Limited
LocationPune District, Maharashtra, India (On-site)
Work modeNo
SkillsPython, AWS, PostgreSQL, TensorFlow, PyTorch
SourceLinkedin
ListedPosted 8 hours ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
10 min from Linkedin publishing this role to us finding it
13 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
63,029 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Greetings from TCS,
We are Hiring for PySpark Developers
About the Role
We are looking for skilled PySpark Developers to join a global banking team. The role involves designing and developing large-scale data pipelines, implementing machine learning models, and building robust data processing solutions on Big Data platforms.
The ideal candidate will have expertise in PySpark, Scala, Hadoop ecosystem technologies, and experience within Credit Risk, Regulatory Risk, Banking, or Financial Services domains.
Key Responsibilities
Design, develop, and optimize scalable ETL/ELT pipelines using PySpark and Scala.
Build ingestion frameworks for structured, semi-structured, and unstructured data sources.
Develop and enhance machine learning model pipelines in partnership with data scientists.
Implement feature engineering, data transformation, model scoring, and retraining processes.
Build reusable Spark-based frameworks for batch and near real-time data processing.
Optimize Big Data workloads running on Spark, Hadoop, EMR, Databricks, or similar platforms.
Ensure data quality, lineage, governance, and documentation standards.
Collaborate with Risk Analytics, Data Science, and Business teams to deliver data solutions.
Support deployment, version control, CI/CD practices, and production support activities.
Mandatory Skills
5+ years of experience in Data Engineering, Big Data Development, or Software Development.
Strong hands-on experience in:
PySpark
Scala
Python
Spark
Hive
Hadoop Ecosystem
MapReduce
Unix Shell Scripting
SQL
Experience building enterprise-grade data pipelines and ingestion frameworks.
Experience with distributed computing environments such as Hadoop, Spark, Databricks, or AWS EMR.
Strong understanding of Data Lake and Data Warehouse concepts.
Experience working with Banking, Financial Services, Credit Risk, or Regulatory Risk projects.
Knowledge of relational databases and platforms such as PostgreSQL, Redshift, or equivalent.
Preferred Skills
Spark Structured Streaming
Kafka
Delta Lake
AWS Glue
Machine Learning frameworks (Spark ML, Scikit-Learn, TensorFlow, PyTorch)
Jenkins, Git, CI/CD pipelines
Data Governance and Data Lineage tools
Domain Preference
Candidates with experience in:
Credit Risk
Regulatory Risk
Risk Modeling
Consumer Banking
Wholesale Banking
Financial Services
Education
Master's or bachelor's Degree in:
Engineering
Statistics
Mathematics
Finance
Computer Science

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App