Live opening · Posted 2 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
NO FRESHER WILL BE PREFERRED
We are looking for Lead Data Architect (Databricks AWS) as a Full Time for our IT Company.
Location- Pune
Experience- 8+ years
Budget- 2lac per month
Notice Period- Immediate Joiner
JD for Architect role :
Responsibilities
• Define the target data architecture for the Horizon MVP and its future evolution.
• Design ingestion, storage, processing, serving, reporting, and downstream API layers.
• Establish canonical schemas for travel signals, corridors, sources, evidence, scores, and generated insights.
• Design the normalization approach for structured, semi-structured, and unstructured data from multiple external sources.
• Define Databricks architecture, Unity Catalog configuration, data schemas, access controls, and governance standards.
• Design data-retention, lineage, quality, security, and privacy controls.
• Define integration patterns between Databricks, AWS PRODIGY, SharePoint, external APIs, and downstream applications.
• Design scalable processing for 30-day signal windows, convergence/divergence scoring, corridor ranking, and spike detection.
• Ensure the platform is ready for future vector storage, LLM integration, and multi-year analytics.
• Review technical designs and guide backend, data engineering, DevOps, and QA teams.
Requirements
• 8+ years of experience in data engineering or architecture, including 3+ years in a Data Architect role.
• Strong experience designing cloud-based data platforms, lakehouses, or analytical platforms.
• Advanced knowledge of Databricks, Apache Spark/PySpark, Delta Lake, and Unity Catalog.
• Experience with AWS data services, identity and access management, networking, and secure platform integration.
• Strong understanding of data modelling, metadata management, lineage, data quality, retention, and governance.
• Experience designing batch and API-driven ingestion pipelines for structured and unstructured data.
• Understanding of ML/AI workloads, feature pipelines, vector databases, embeddings, and LLM integration patterns.
• Experience designing APIs and downstream data-serving patterns.
• Knowledge of PII handling, data privacy, RBAC, and enterprise security controls.
• Ability to communicate architecture decisions to both technical and business stakeholders.
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.