
Real job — pulled straight from SkanAI’s careers page · Verified July 12, 2026 · No reposts.
Job description
SkanAI is hiring a Senior Data Engineer — a full-time, based in Bengaluru, India role. Apply directly on SkanAI's careers page below.
Senior Data Engineer
Location: Bengaluru, India
Department: Engineering
Experience: 5
- Design, build, and maintain high-performance Apache Flink ETL/ELT pipelines for real-time data synchronisation from PostgreSQL to StarRocks.
- Architect robust CDC (Change Data Capture) solutions using Flink CDC connectors and Debezium for PostgreSQL source ingestion.
- Implement Flink SQL and DataStream API pipelines for complex transformation logic, aggregations, and data enrichment.
- Develop and maintain custom Flink connectors and sinks for StarRocks integration using the StarRocks Flink Connector.
- Design fault-tolerant, exactly-once or at-least-once pipelines with appropriate checkpointing and state management strategies.
- Evaluate and implement schema evolution strategies to handle upstream PostgreSQL schema changes gracefully.
- Profile and tune Flink job performance: parallelism settings, task manager memory, operator chaining, and back-pressure management.
- Optimise StarRocks loading strategies (Stream Load vs. Routine Load) for high-throughput ingestion.
- Monitor pipeline latency and throughput SLAs; proactively identify and resolve bottlenecks.
- Implement efficient watermarking and windowing strategies for time-sensitive data flows.
- Manage Flink state backends (RocksDB / heap) and configure appropriate TTLs to control state size.
- Own end-to-end pipeline reliability: diagnose and resolve issues including data lag, job failures, checkpoint timeouts, and OOM errors.
- Establish alerting and observability for pipeline health using Flink metrics, Prometheus, and Grafana (or equivalent).
- Define and implement data quality checks, reconciliation processes, and dead-letter queue (DLQ) strategies.
- Perform root-cause analysis on data discrepancies between PostgreSQL source and StarRocks target.
- Maintain comprehensive runbooks for common failure scenarios and recovery procedures.
- Lead Data Engineers: assign tasks, conduct code reviews, and ensure delivery against sprint goals.
- Mentor engineers on Flink internals, best practices, and performance considerations.
- Collaborate with data consumers (analysts, BI teams) to understand requirements and translate them into pipeline specifications.
- Drive technical decisions on tooling, frameworks, and deployment strategies (Flink on Kubernetes on GCP).
- Maintain technical documentation including architecture diagrams, data flow documentation, and operational guides.
- Manage Flink cluster deployment and configuration on Kubernetes (GKE) using Helm charts and Harness CI/CD pipelines.
- Build and maintain CI/CD pipelines using Harness for Flink job packaging, testing, and deployment to GKE.
- Manage Flink cluster configuration using Helm charts; maintain Helm values and chart templates for environment-specific configurations.
- Coordinate with infrastructure and DBA teams for PostgreSQL slot management and StarRocks table design.
- 5+ years of experience in data engineering with a focus on ETL/ELT pipeline development.
- 2+ years of hands-on production experience with Apache Flink (Flink SQL and/or DataStream API).
- Strong proficiency in Java, Scala, or Python for Flink job development.
- Experience with Change Data Capture (CDC) patterns: Flink CDC, Debezium, or equivalent.
- Practical experience with PostgreSQL as a data source, including replication slots and WAL configuration.
- Demonstrated experience with columnar OLAP databases (StarRocks, Doris, ClickHouse, or similar).
- Solid understanding of distributed systems concepts: fault tolerance, exactly-once semantics, state management, and watermarking.
- Experience with Kafka or similar message brokers as part of streaming architectures.
- Proven ability to lead engineering teams, including task planning, code review, and mentoring.
- Strong analytical and problem-solving skills with the ability to debug complex distributed pipeline issues.
- Excellent written and verbal communication skills with the ability to document technical decisions clearly.
- Self-driven with the ability to work autonomously and manage priorities in a fast-paced environment.
- Hands-on experience with StarRocks Flink Connector and StarRocks primary-key table designs.
- Hands-on experience with Flink on Kubernetes (GKE) using Flink Kubernetes Operator or native K8s mode.
- Experience with Harness CI/CD and Helm chart management for data workloads.
- Experience with Flink Table API and Flink SQL for unified batch and stream processing.
- Knowledge of Apache Iceberg, Delta Lake, or Hudi for lakehouse architectures.
- Experience with orchestration tools such as Apache Airflow or Prefect for hybrid batch/streaming workflows.
- Familiarity with observability stacks: Prometheus, Grafana, and Flink metrics reporters.
- Prior experience in a technical lead or senior engineer role with delivery accountability.
- Understanding of data modelling best practices for analytical workloads in StarRocks.
- Experience in data-intensive domains such as data mining, large-scale data processing, or analytical platform engineering.
- Understanding of end-to-end data process flows: from source system ingestion through transformation, aggregation, and consumption layers.
- Familiarity with data governance, lineage tracking, and metadata management.
- Exposure to real-time analytics use cases such as dashboards, operational reporting, or data exploration pipelines.
Get Data Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What skills are required for Senior Data Engineer at SkanAI?
The required skills for Senior Data Engineer at SkanAI include: ETL, PostgreSQL, Java, Scala, Python, Kafka, Kubernetes, GCP, Helm, Prometheus, Grafana, Airflow.
What is the seniority level for Senior Data Engineer at SkanAI?
Senior Data Engineer at SkanAI is a Senior / Manager level position.
How do I apply for Senior Data Engineer at SkanAI?
You can view the full description and apply for Senior Data Engineer at SkanAI on EchoJobs: https://echojobs.io/job/skan-ai-senior-data-engineer-zw3wo.