About the Role

We are seeking a hands-on Databricks Architect / Databricks Engineer to lead the solution design, end-to-end pipeline implementation, and governance of enterprise Lakehouse platforms.

In this dual architecture and delivery role, you will translate complex business and commercial analytics requirements into scalable Lakehouse architectures. You will build governed data ingestion and transformation pipelines using Auto Loader, Delta Live Tables (DLT), PySpark, and Unity Catalog, while helping build reusable solution accelerators and mentoring delivery teams.

Key Responsibilities

  • Lakehouse Architecture & Solutioning: Design end-to-end Databricks Lakehouse architectures, covering ingestion strategies, medallion layouts (Bronze/Silver/Gold), serving layers, and unified governance.
  • Hands-on Pipeline Engineering: Build and maintain scalable, production-grade pipelines utilizing Auto Loader, Delta Live Tables (DLT), PySpark/SQL, and dbt.
  • Governance & Serving Layers: Implement enterprise-wide data governance, access controls, and data sharing models using Unity Catalog and Databricks SQL.
  • Analytics & AI/BI Integration: Configure and fine-tune Genie Spaces, AI/BI dashboards, and visualization layers (Tableau / Databricks SQL) to support business analytics use cases.
  • Commercial Dataset Modeling: Ingest and structure complex commercial datasets (e.g., IQVIA, Symphony, CRM, Claims, and Hub data) into clean, governed Delta Lake models.
  • Platform Hygiene & Cost Optimization: Manage workspace standards, cluster policies, job scheduling via Databricks Workflows, cost tracking, and performance tuning.
  • Practice Building & Pre-Sales: Develop reusable solution accelerators, coding frameworks, and reference architectures; contribute technical effort estimation and architecture design for pre-sales proposals.

Required Skills & Qualifications

Mandatory Requirements:

  • Experience: 5 to 9 years of dedicated Data Engineering / Data Architecture experience, with at least the last 3+ years hands-on with Databricks.
  • End-to-End Build & Design: Proven ability to architect Lakehouse solutions and execute the hands-on engineering build independently.
  • Core Databricks Stack: Deep technical proficiency across Delta Lake, Delta Live Tables (DLT), Unity Catalog, Auto Loader, Databricks SQL, and Workflows.
  • Data Engineering & Querying: Advanced proficiency in PySpark, SQL, and pipeline orchestration.
  • Mandatory Certification: Must hold at least one active Databricks Professional-level Certification (Databricks Certified Data Engineer Professional preferred).
  • Consulting Background: Must have prior delivery experience in an IT Services or Consulting firm, managing direct technical deliveries for US or global clients.
  • Client-Facing Communication: Ability to serve as the lead technical POC, articulate architectural trade-offs, and present to non-technical business stakeholders.

Preferred Qualifications:

  • Domain experience within Life Sciences, Pharmaceuticals, or Healthcare Analytics (working with IQVIA, Symphony, or Claims data).
  • Experience with machine learning frameworks (PyTorch) and BI visualization tools (Tableau).
  • Hands-on exposure to modern Databricks features such as Genie Spaces and Metric Views.

Work Arrangement & Compensation Details

  • Work Model: Hybrid (3 Days In-Office / 2 Days Remote per week; flexible work-from-home policy up to 6 days per month).
  • Location: Bengaluru, Karnataka, India.
  • Compensation Structure: CTC package includes a 20% variable component.
Job Category: Technology
Job Type: Hybrid
Job Location: Bengaluru - Karnataka

Apply for this position

Allowed Type(s): .pdf, .doc, .docx