ASAI job platform logo
  1. Home
  2. /Jobs
  3. /Data Engineer Jobs in Gurgaon
  4. /Senior Data Engineer - Gurugram
ClanX

Senior Data Engineer - Gurugram

Gurgaon · Senior

Older listing - lower visibility likelyVerified listingPosted 94d ago

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

What we know about this role

Hiring pulse

MEDIUM

ClanX is reviewing applications at a steady pace. Expect a standard response time as they evaluate the current pool.

Apply window

First 72 hours

Window passed - posted 94d ago

Early applicants get seen before the pile builds.

Not a repost

The first time we've seen this listing - it hasn't been closed and reopened.

Skills required

10 listed
Machine LearningPredictive MaintenanceQuery OptimizationQuery PerformanceSQL (Programming Language)Python (Programming Language)PysparkExtract Transform Load (ETL)+2 more
Machine LearningPredictive MaintenanceQuery OptimizationQuery PerformanceSQL (Programming Language)

You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.

Job description

Mechademy is hiring a Senior Data Engineer to build and scale reliable data platforms, pipelines, and models that power enterprise AI, machine learning, and analytics for industrial asset monitoring and predictive maintenance.

Company Details

Mechademy is an enterprise AI company building real-time monitoring, diagnostics, and predictive maintenance solutions for industrial equipment. The company serves clients across oil & gas, power generation, and LNG sectors through production-grade AI and physics-informed machine learning systems.

Website: https://mechademy.com/

Responsibilities

What You’ll Own

  1. Lakehouse Pipelines & Ingestion (35%)

Design and own batch ETL/ELT and CDC pipelines that bring sensor and operational data into the lakehouse, orchestrated in Dagster

Build for reliability: idempotent, incremental, backfill-safe pipelines with sane retry and failure handling, that still produce correct output when a worker is killed mid-run or a message is delivered twice

Onboard new client data sources: schema and tag mapping, time-series normalization, resampling, gap handling at scale

2. Modeling & Serving (25%)

Model raw data into well-structured, documented tables that downstream ML and analytics can trust

Build and maintain the datasets behind ML feature pipelines and the lakehouse layer powering self-serve analytics

Write performant Spark/PySpark and SQL; optimize partitioning, storage formats, and query cost

3. Data Quality & Reliability (10%)

Own data quality: validation, freshness/SLA monitoring, and observability so bad data is caught before it reaches consumers

Make the data layer debuggable: lineage, tests, and alerting that tell you what broke and where

Reason about failure modes across the whole path (queue, worker, orchestrator, database, object store) and design so that a partial failure leaves the system in a state you can recover from

4. Relational & Operational Data (30%)

Contribute to the schema, indexing, and query performance of the relational database the product runs on

Design tables and constraints so that correctness is enforced at the database layer, and diagnose slow queries from their plans

Own retention and the boundary between the operational database and the lakehouse: what stays, what moves, and how it gets there

What Success Looks Like

First 30 days: Productive in the codebase and orchestration layer. First pipeline change merged.

First 90 days: Independently shipping and owning pipelines. Onboarded at least one new data source end-to-end.

First 6 months: Owning a lakehouse data domain, its ingestion, models, and quality, that ML and analytics teams rely on you to drive.

Requirement

Must-Have

  1. 4+ years building production data pipelines: real systems with real consumers, not just one-off scripts

  2. Strong data engineering fundamentals: data modeling, batch vs. streaming, idempotency, incremental processing, partitioning.

  3. Expert SQL and strong Python: query optimization, window functions, clean production-quality code

  4. Relational database depth: you’ve designed schemas for a production PostgreSQL (or equivalent) system and understand normalization and when to break it, indexing strategies, transactions and isolation levels, locking, and how to read a query plan and fix the query

  5. Distributed systems fundamentals: at-least-once delivery and idempotent consumers, partitioning and its effect on ordering, consistency and durability trade-offs, retries, timeouts, and backpressure. You can explain what happens to in-flight work when a worker or a database node dies

  6. Hands-on with a distributed processing engine (Spark/PySpark or equivalent) on non-trivial data volumes

  7. Experience with an orchestrator (Dagster, Airflow, Prefect, or equivalent) and a cloud platform (AWS/Azure)

  8. Data-quality mindset: you build validation and monitoring into pipelines, not after something breaks

Strong Signals (Nice-to-Have)

  1. Time-series or high-frequency sensor data at scale

  2. TimescaleDB or another time-series database (hypertables, continuous aggregates, compression, retention policies)

  3. Warehouse/lakehouse modeling (Delta/Iceberg/Snowflake/Redshift or equivalent) and file-format/partition tuning (Parquet)

  4. CDC / database-replication pipelines

  5. Message brokers or task queues in production (Kafka, RabbitMQ, or equivalent)

  6. Building data for ML: feature pipelines, training datasets, serving consistency

  7. dbt or similar transformation/modeling frameworks

  8. Docker, Terraform/IaC, CI/CD for data

  9. IoT, energy, or industrial sector experience. Not required, but it compresses your ramp

Job Details

Gurugram - Hybrid (2–3 days on-site)

Interview Process

  • Technical Round (Python & SQL)

  • System Design Round

  • Culture Fit Round

Important Note

ClanX is a recruitment partner, helping Mechademy hire a Senior Data Engineer.

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

More Data Engineer roles in Gurgaon

See all

Sr Analyst II Data Engineering

Dxctechnology · Gurgaon

Technical Lead - Data and platform engineering

Affle · Gurgaon

Data Engineer

Saarthee · Gurgaon

Senior Data Engineer

Anaplan · Gurgaon

Keep browsing

All open roles at ClanXAll Data Engineer jobs in Gurgaon

Two ways in

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

ASAI job platform logo

A job platform finally, balanced in your favour.

Jobs by CityJobs by CompanyGuidesHow We VerifyAboutPrivacy PolicyTerms of Service

Built in India 🇮🇳

© 2026 ASAI. All rights reserved.