Live opening · Posted 2 days ago

Lead Data Architect (Databricks AWS)

Haparz · Pune Division, Maharashtra, India (Hybrid)
Linkedin No
You are 2 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 2 days ago
CompanyHaparz
LocationPune Division, Maharashtra, India (Hybrid)
Work modeNo
SourceLinkedin
Listed2 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
45 min from Linkedin publishing this role to us finding it
3 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,165 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

NO FRESHER WILL BE PREFERRED
We are looking for Lead Data Architect (Databricks AWS) as a Full Time for our IT Company.
Location- Pune
Experience- 8+ years
Budget- 2lac per month
Notice Period- Immediate Joiner
JD for Architect role :
Responsibilities
• Define the target data architecture for the Horizon MVP and its future evolution.
• Design ingestion, storage, processing, serving, reporting, and downstream API layers.
• Establish canonical schemas for travel signals, corridors, sources, evidence, scores, and generated insights.
• Design the normalization approach for structured, semi-structured, and unstructured data from multiple external sources.
• Define Databricks architecture, Unity Catalog configuration, data schemas, access controls, and governance standards.
• Design data-retention, lineage, quality, security, and privacy controls.
• Define integration patterns between Databricks, AWS PRODIGY, SharePoint, external APIs, and downstream applications.
• Design scalable processing for 30-day signal windows, convergence/divergence scoring, corridor ranking, and spike detection.
• Ensure the platform is ready for future vector storage, LLM integration, and multi-year analytics.
• Review technical designs and guide backend, data engineering, DevOps, and QA teams.
Requirements
• 8+ years of experience in data engineering or architecture, including 3+ years in a Data Architect role.
• Strong experience designing cloud-based data platforms, lakehouses, or analytical platforms.
• Advanced knowledge of Databricks, Apache Spark/PySpark, Delta Lake, and Unity Catalog.
• Experience with AWS data services, identity and access management, networking, and secure platform integration.
• Strong understanding of data modelling, metadata management, lineage, data quality, retention, and governance.
• Experience designing batch and API-driven ingestion pipelines for structured and unstructured data.
• Understanding of ML/AI workloads, feature pipelines, vector databases, embeddings, and LLM integration patterns.
• Experience designing APIs and downstream data-serving patterns.
• Knowledge of PII handling, data privacy, RBAC, and enterprise security controls.
• Ability to communicate architecture decisions to both technical and business stakeholders.

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App