Skip to main content
Posted 27 August, 2026

Apache Doris Developer

Diverse Lynx
bengaluru,560063 Full Time
Reference: 365_569689_26-03670

Description:

Job Description:
We are looking for an Apache Doris Developer who can own the MPP query engine, lead migration of
existing pipelines and warehouses, and help establish a scalable, low-latency real-time analytics (OLAP)
platform.
You will work at the intersection of query engineering and data modeling - translating legacy
Snowflake/Spark logic into performant Doris workloads (both native tables and Iceberg external
catalogs) while ensuring correctness, parity, and reliability.
Key Responsibilities
MIGRATION & DELIVERY
Migrate SQL workloads, transformations, and data models from Snowflake and Spark to Apache
Doris.
Translate Snowflake-specific SQL (semi-structured VARIANT / OBJECT , stored procedures, tasks,
streams) and Spark jobs into equivalent, validated Doris SQL.
Design Doris table models appropriately - Duplicate, Aggregate, and Unique key models -
mapping source schemas to the right model for each workload.
Integrate Apache Iceberg via Doris Multi-Catalog for federated lakehouse querying, and design
ingestion paths from Iceberg into native Doris tables where low latency is required.
Design and implement data-parity validation- schema, row-count, and value-level reconciliation
between source (Snowflake/Spark) and target (Doris/Iceberg).
QUERY ENGINEERING & PERFORMANCE
Write, optimize, and troubleshoot complex analytical SQL on Doris (MySQL-protocol compatible).
Tune query performance: bucketing/partitioning strategy, colocate joins, runtime filters, and the
cost-based optimizer.
Design and maintain materialized views , rollups, and indexes (inverted, bitmap, bloom filter,
N-gram) to accelerate queries..
Analyze query plans (EXPLAIN, profile) and resolve memory pressure, tablet skew, and slow scans.
Required Skills & Qualifications
Hands-on production experience with Apache Doris (or a comparable MPP OLAP engine -
StarRocks, ClickHouse, Greenplum).
Strong, deep SQL expertise - complex analytical queries, window functions, CTEs, query
optimization
Solid understanding of Doris architecture (FE/BE), the three data models
(Duplicate/Aggregate/Unique), partitioning, bucketing, and tablet/replica management.
Experience with Apache Iceberg (or comparable open table formats - Delta Lake, Hudi) and
lakehouse / external-catalog federation.
Practical experience with Snowflake and/or Apache Spark- enough to read, understand, and
migrate existing workloads.
Understanding of distributed / MPP query execution: join distribution, runtime filters, memory
management, and data skew.
Experience with cloud object storage and columnar file formats (Parquet, ORC).Proficiency in at least one programming language (Python, Java, or Scala) for tooling, UDFs, and
automation.
Version control (Git) and CI/CD for data pipelines.
Preferred / Nice-to-Have
Experience leading a Snowflake -> Doris or Spark -> Doris migration at scale.
Experience building data-validation / reconciliation frameworks for migration parity.
Knowledge of Doris internals or connector development (custom UDFs/UDAFs).
Workflow orchestration (Airflow, Dagster) and dbt (with a Doris adapter).
Data governance, lineage, and cost-optimization tooling.
Soft Skills
Strong analytical and problem-solving mindset for debugging correctness and performance issues.
Clear communication - able to document migration decisions and work with analytics, platform, and
business teams.
Ownership mentality with attention to data correctness and reliability.
Project Location 1
ORS | BHUBANESWAR
Project Location 2
TELANGANA | HYDERABAD
Project Location 3
KRNK | BANGALORE
Project Location 4
WB | Kolkata
Relevant Experience
2+ Years
Mandatory skills
Apache Doris + Iceberg
Desired skills
SQL , Snowflake/ Apache Spark
Domain (Industry)
NA
Total Experience (Ex. 5-7 Years)
5-8 Years
Mode of Interview
Face to Face
WFO / WFH / Hybrid
Hybrid
Please enter shift timings
09:00 AM to 06:00 PM IST
Shift Timings
General

Sign up for Job Alerts