Live opening · Posted 14 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Data Engineer – Job Description
Experience: 5+ Years
Location: Hyderabad / Chennai
Work Mode: On-site
Employment Type: Full-time
About the Role
We are looking for a passionate and experienced Data Engineer with 5+ years of experience in designing, developing, and maintaining scalable data platforms and ETL/ELT pipelines.
The ideal candidate should have strong expertise in Python, SQL, PySpark, Apache Spark, cloud data services, and modern data engineering frameworks, along with hands-on experience in Generative AI and Large Language Models (LLMs).
The candidate will play a key role in building reliable, scalable, and high-performance data solutions that support analytics, reporting, and AI/ML initiatives.
This is an on-site opportunity in Hyderabad or Chennai. Preference will be given to candidates currently based in these locations or willing to relocate.
Key Responsibilities
Design, develop, and maintain scalable data pipelines and ETL/ELT workflows.
Build and optimize data ingestion processes from multiple structured and unstructured data sources.
Develop robust data models and data warehouses for analytics and reporting.
Design and optimize SQL queries for high-performance data processing.
Build and maintain data lakes using cloud storage solutions.
Implement data validation, cleansing, transformation, and quality checks.
Integrate data solutions with cloud platforms such as AWS, Azure, or Google Cloud Platform (GCP).
Develop batch and real-time data processing pipelines using modern data processing frameworks.
Collaborate with data scientists, analysts, and application teams to deliver reliable data solutions.
Follow software engineering best practices, including Git, CI/CD, testing, monitoring, and technical documentation.
Required Skills
5+ years of experience in Data Engineering.
Strong programming experience in Python.
Hands-on experience with PySpark and Apache Spark.
Advanced SQL skills, including query optimization and performance tuning.
Experience with AWS, Azure, or GCP.
Strong experience in building scalable ETL/ELT pipelines.
Strong experience in Generative AI and Large Language Models (LLMs).
Good understanding of data warehousing concepts and dimensional modelling.
Experience working with large-scale distributed datasets.
Familiarity with Git and CI/CD pipelines.
Understanding of Agile/Scrum methodologies.
Preferred Qualifications
Experience with Apache Airflow or similar workflow orchestration tools.
Knowledge of Kafka and real-time data streaming.
Familiarity with Docker, Kubernetes, and Terraform (IaC).
Exposure to Databricks and modern data lake technologies such as Delta Lake, Apache Iceberg, or Apache Hudi.
Understanding of data governance and metadata management.
Strong exposure to Generative AI and LLM-based solutions.
Strong analytical, problem-solving, and communication skills
Educational Qualification
Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline.
Location Requirement
📍 Hyderabad / Chennai
🏢 On-site only
💼 5+ Years of Experience
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.