inDrive logo

Data Engineer, Data Platform

inDrive

Hybrid
Almaty, Kazakhstan
Full-time
Mid Level
Salary not listedPosted 3w ago

Real job — pulled straight from inDrive’s careers page · Verified July 21, 2026 · No reposts.

Job description

inDrive is hiring a Data Engineer, Data Platform — a full-time, based in Almaty, Kazakhstan role. Apply directly on inDrive's careers page below.

Data Engineer

Department: Data Platform

Employment Type: Full Time

Location: Almaty, Almaty Special District, Kazakhstan, Hybrid

We are looking for a Data Engineer to join one of the Data Platform teams that works with the Marketing, Growth, partner, and financial data domains.
You will be working with cutting edge cloud technologies (GCP, AWS, BigQuery, Databricks, K8s) and building a large scale data infrastructure for analytics, machine learning, and streaming/CDC data delivery.

Key Responsibilities

  • Build and operate batch and streaming ingestion into a layered BigQuery DWH (raw → ODS → data marts) using Airflow, Debezium CDC over Kafka with protobuf, Pub/Sub, and Dataflow
  • Integrate external data sources end-to-end — marketing platforms (GA4, AppsFlyer, TikTok/Meta/Google Ads), payment providers, S3 buckets, and third-party APIs — including schema contracts, backfills, and reconciliation
  • Engineer the data platform itself in Python: custom Airflow operators and connectors in a shared ETL framework, Kafka Connect on Strimzi (K8s), Cloud Functions, and API integrations with external providers
  • Build CI/CD and change-management tooling for BigQuery: GitHub-based test-and-approval flows, SQL migration engines (Liquibase/Flyway/Bytebase), sandbox validation, backup and rollback
  • Own reliability and correctness of pipelines: idempotency, deduplication, late-data handling, backfill and replay, freshness monitoring and alerting; write integration and unit tests
  • Drive data governance and compliance: ITGC-compliant change management for BigQuery, IAM and least-privilege access, PII policy tags and DLP, Unity Catalog on Databricks, column-level lineage (OpenMetadata/Dataplex), and disaster-recovery planning
  • Build internal data tools and platform services for agentic workflows with data — Streamlit apps, Slack bots, LLM-based agents and MCP servers that help teams find and use data
  • Support analysts and business teams with data requests, fostering data-driven decision-making across the company
  • Contribute to system design and architecture with the development team

Skills, Knowledge and Expertise

  • Strong practical Python: clean, well-structured, and tested code for services, tooling, and data pipelines
  • Solid software design skills (OOP, modularity, design patterns) — we build platform tools for agentic workflows with data and plan to develop data-related backend services, so well-designed code is highly valued
  • Experience building and operating services in a cloud environment (GCP, AWS or similar): CI/CD, containerization, monitoring and alerting
  • Familiarity with Kubernetes and Terraform — our infrastructure runs on GCP/K8s
  • Hands-on experience with DWH-related tasks (BigQuery or another cloud warehouse) and confident working SQL
  • Clear communication with non-engineering stakeholders — a meaningful share of the work is data requests from analysts and business teams
  • Demonstrated ability to take ownership of technologies or services and proactively contribute ideas to the team
Nice to have
  • Advanced SQL: complex queries, window functions, partitioning, clustering, and cost optimization
  • Experience building reliable pipelines around CDC (e.g., Debezium): idempotency, schema evolution, backfills, and reconciliation
  • Analytical data modeling skills: table grain, facts vs dimensions, slowly changing dimensions, and metric definitions
  • Experience with stream processing frameworks such as Flink or Apache Beam/Dataflow
  • Exposure to data governance and audit compliance (ITGC/SOX), Databricks Unity Catalog, or lineage/catalog tooling (OpenMetadata, Dataplex)
  • Interest in building LLM-based agents and AI tooling for data

Why join us

  • Help us challenge injustice by creating fair choices for millions of people across 48 countries.
  • Develop your professional skills with access to mentoring, career consulting, and learning programs.
  • Collaborate with teams around the world and gain international experience through our Global Talent Exchange Program.
  • Engage in company-wide challenges, awards, sports activities, employee-led social impact and volunteering projects.
  • Work alongside people who take initiative, speak openly, and challenge themselves to grow.
  • Improve your language skills through co-financed courses and internal speaking clubs.
Final benefits may vary depending on the location.


Get Data Engineer jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Arterys logo

Senior Software Engineer

$110k–$175kChicago, IL
✓ From careers page· 18m ago
Blue River Technology logo

Systems Engineer, Autonomous Systems

$131k–$228kRemote · US-eligible
✓ From careers page· 22m ago
Waystar logo

Senior Data Scientist

Lehi, UT
✓ From careers page· 23m ago
TetraScience logo

Lead Software Platform Engineer, MLOps (Remote)

$200k–$270kRemote · US-eligible
✓ From careers page· 23m ago

Frequently asked questions

What skills are required for Data Engineer, Data Platform at inDrive?

The required skills for Data Engineer, Data Platform at inDrive include: Python, SQL, GCP, AWS, BigQuery, Databricks, Kubernetes, Terraform, Kafka, Airflow, CI/CD, LLM, API.

What is the seniority level for Data Engineer, Data Platform at inDrive?

Data Engineer, Data Platform at inDrive is a Mid Level level position.

How do I apply for Data Engineer, Data Platform at inDrive?

You can view the full description and apply for Data Engineer, Data Platform at inDrive on EchoJobs: https://echojobs.io/job/indrive-data-engineer-data-platform-t0nu8.