Padmi
Honeywell Technologies logo
Honeywell Technologies

building automation · industrial automation

Advanced Data Engineer - Databricks

IN · OnsitePosted 2 months ago
Software engineeringSenior
Apply at Honeywell Technologies

Opens the source posting on ibqbjb.fa.ocs.oraclecloud.com

Source description

About the role

View original

Skip to main content.

View More Jobs

Advanced Data Engineer - Databricks

Bengaluru, Karnataka, India

Trending

Job Description

Ddvanced Data Engineer - Value Engineering & Component Engineering COE

Location: Bangalore, IN (Hybrid)

Role Overview: Honeywell's VECE COE is building a next-generation, AI-Ready data platform to power advanced analytics, predictive insights, and data science at enterprise scale. As a Senior Data Engineer, you will be a founding technical pillar of this platform: designing and building the data infrastructure that transforms raw, multi-source data into governed, high-quality, analytics-ready assets.

This is not a maintenance role. You will architect, build, and own end-to-end data pipelines using Azure Databricks as the primary platform, following Medallion Architecture principles, and delivering trusted data to downstream consumers in Google Cloud Platform (GCP). You will directly shape how Honeywell's VECE organization transitions from traditional descriptive analytics to proactive, AI-driven decision-making.

Responsibilities

What you will build?

Data Pipelines & Ingestion

  • Implement end-to-end ingestion pipelines from heterogeneous sources (i.e. Snowflake, SQL Server, Excel, REST APIs, and unstructured files) into Azure Databricks following defined architecture patterns
  • Build and maintain Bronze → Silver → Gold Medallion layers, applying transformation logic, business rules, and quality checks at each stage
  • Implement incremental loading pattern (i.e. CDC, watermarking, Delta Lake MERGE/UPSERT) to ensure efficient, scalable, and reliable data delivery
  • Develop pipelines for structured and unstructured data (i.e. documents, JSON, Parquet, Excel) supporting AI and ML consumption downstream

Data Modeling & Semantic Layer

  • Implement and extend data models (i.e. fact/dimension tables, domain data marts) following designs defined by the Senior DE and AI team.
  • Write clean, modular, reusable PySpark and SQL transformation logic that is testable, documented, and deployable via CI/CD
  • Contribute to the semantic layer that powers Power BI dashboards and GCP-connected analytics consumers
  • Maintain and improve existing models as business requirements evolve

Orchestration and Data Ops

  • Build and manage Databricks Workflows: configuring task dependencies, retry policies, and failure alerting
  • Follow and contribute to CI/CD practices: version control, pull requests, automated testing, and deployment to Dev/QA/Prod environments using Azure DevOps or GitHub Actions
  • Package and deploy reusable logic as Python libraries following team standards
  • Monitor pipeline health, investigate failures, and resolve data issues within SLA

Data Governance & Quality

  • Apply data quality rules (i.e. validation, deduplication, null checks, reconciliation) within pipelines to ensure data arrives fit for purpose
  • Operate within the Unity Catalog governance framework respecting RBAC, namespace structure, and tagging standards defined by platform leads
  • Ensure data delivered to GCP is schema-consistent, validated, and documented
  • Flag and escalate data quality issues proactively not reactively

FinOps Awareness

  • Write cost-conscious PySpark avoiding unnecessary full scans, optimizing joins, using appropriate cluster types
  • Apply Delta table best practices (i.e. VACUUM, OPTIMIZE, compaction) to manage storage costs
  • Follow cluster policies defined by platform leads and flag unusual resource consumption

Must Have

  • Databricks: 2+ years hands-on: PySpark, Delta Lake, Workflows, Unity Catalog.
  • Demonstrate expertise in data strategy, for example: Medallion Architecture, Domain Data Modeling and Functional Data Architecture.
  • Data Quality Frameworks (i.e. rule-based validation, anomaly detection)
  • Data Pipelines: incremental loading, CDC, CI/CD, Observability
  • Advanced Python/Pyspark and Advanced SQL
  • Strongly preferred: DLT, UC, GCP, Azure, Kafka.
  • Highly value Databricks Certified Professional

Qualifications

Experience

  • 4-6+ years of overall data engineering experience
  • 2+ years of hands-on Azure Databricks experience in production environments
  • Demonstrated ability to build and deliver pipelines — not just maintain or support them
  • Experience working within a defined architecture and contributing to its improvement
  • Comfortable working with multiple data source types — relational, file-based, API

About Honeywell: Honeywell Industrial Automation enhances process industry operations, creates sensor technologies, automates supply chains, and improves worker safety. The VECE COE focuses on optimizing operational processes and driving sustainable growth

Required Skills

  • Data Pipelines
  • Data Warehousing
  • Google Cloud
  • Microsoft Azure
  • Python and Pyspark
  • Structured Query Language

About Us

Honeywell helps organizations solve the world's most complex challenges in automation, the future of aviation and energy transition. As a trusted partner, we provide actionable solutions and innovation through our Aerospace Technologies, Building Automation, Energy and Sustainability Solutions, and Industrial Automation business segments – powered by our Honeywell Forge software – that help make the world smarter, safer and more sustainable.

Apply Now

Job Info

  • Job Identification149046
  • Job CategoryEngineering
  • Posting Date05/22/2026, 12:36 AM
  • Job ScheduleFull time
  • Locations Devarabisanahalli Village, KR Varturhobli,, Bangalore, KA, 560103, IN
  • Hire EligibilityInternal and External
  • Relocation PackageNone

Similar Jobs

See More Jobs

American English

I am an employee

We use cookies to improve website performance, facilitate information sharing on social media and offer advertising tailored to your interests. For more information, see our Cookie Notice. You can also customize your browser’s cookie settings. Please note that if you refuse cookies, it may affect site functionality and performance.

Accept AllReject All

Manage Preferences

Are You Still With Us?

It seems you've been gone for a while. For security reasons we will end your session automatically in 03:00 unless you would like to continue working.

End SessionContinue Working

Work Summary

This summary is generated by AI Assist. Click inside the summary text box to make changes as necessary.

DiscardAdd Summary

Page Advanced Data Engineer - Databricks - Honeywell Careers loaded

More at Honeywell Technologies

Related open roles

View all roles