Sign up to access all features of our service
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Skan.ai - Senior Data Engineer - Apache Flink

Skan ai

Be at the Forefront of the Agentic AI Revolution :At Skan AI, you'll be part of the team pioneering the context engine for human and agentic execution, bringing context from enterprise operators, systems, and processes to power how the world's largest organizations execute their most complex, mission-critical work.Why Join Skan AI :We're in hyper-growth mode at exactly the right moment in history. As enterprises race to adopt agentic AI, we're uniquely positioned to deliver the clear signal they desperately need: a platform that trains and grounds AI Agents in trillions of real execution signals, enabling reliable, compliant automation of their most complex processes. Backed by Dell Technologies Capital and other leading investors, we're the only company that can bridge the gap between AI's promise and enterprise reality, making us perfectly positioned to define the agentic era for modern enterprises. Our diverse, collaborative team of 250+ innovators is solving category-defining challenges at the intersection of AI, process intelligence, and enterprise work. Diverse perspectives fuel breakthrough thinking, cross-functional collaboration is the norm, and our work directly transforms how Fortune 500 companies operate. We are shaping the future of work itself.Role Overview :We are looking for an experienced Apache Flink ETL Lead to own the design, development, and delivery of our real-time data pipeline infrastructure. You will lead data engineering efforts responsible for synchronising data from PostgreSQL to StarRocks using Apache Flink's CDC and streaming pipeline capabilities. You will be hands-on in development while providing technical leadership and driving best practices across the engineering organisation. This is an individual contributor and lead role. We value hands-on engineers who can also mentor and drive delivery. If you love debugging Flink checkpoints as much as you enjoy growing a team, we want to hear from you.Key Responsibilities :1. Pipeline Development and Architecture :- Design, build, and maintain high-performance Apache Flink ETL/ELT pipelines for real-time data synchronisation from PostgreSQL to StarRocks.- Architect robust CDC (Change Data Capture) solutions using Flink CDC connectors and Debezium for PostgreSQL source ingestion.- Implement Flink SQL and DataStream API pipelines for complex transformation logic, aggregations, and data enrichment.- Develop and maintain custom Flink connectors and sinks for StarRocks integration using the StarRocks Flink Connector.- Design fault-tolerant, exactly-once or at-least-once pipelines with appropriate checkpointing and state management strategies.- Evaluate and implement schema evolution strategies to handle upstream PostgreSQL schema changes gracefully.2. Performance Optimisation :- Profile and tune Flink job performance: parallelism settings, task manager memory, operator chaining, and back-pressure management.- Optimise StarRocks loading strategies (Stream Load vs. Routine Load) for high-throughput ingestion.- Monitor pipeline latency and throughput SLAs; proactively identify and resolve bottlenecks.- Implement efficient watermarking and windowing strategies for time-sensitive data flows.- Manage Flink state backends (RocksDB / heap) and configure appropriate TTLs to control state size.3. Troubleshooting and Reliability :- Own end-to-end pipeline reliability : diagnose and resolve issues including data lag, job failures, checkpoint timeouts, and OOM errors.- Establish alerting and observability for pipeline health using Flink metrics, Prometheus, and Grafana (or equivalent).- Define and implement data quality checks, reconciliation processes, and dead-letter queue (DLQ) strategies.- Perform root-cause analysis on data discrepancies between PostgreSQL source and StarRocks target.- Maintain comprehensive runbooks for common failure scenarios and recovery procedures.4. Team Leadership and Delivery :- Lead Data Engineers : assign tasks, conduct code reviews, and ensure delivery against sprint goals.- Mentor engineers on Flink internals, best practices, and performance considerations.- Collaborate with data consumers (analysts, BI teams) to understand requirements and translate them into pipeline specifications.- Drive technical decisions on tooling, frameworks, and deployment strategies (Flink on Kubernetes on GCP).- Maintain technical documentation including architecture diagrams, data flow documentation, and operational guides.5. Deployment and DevOps : - Manage Flink cluster deployment and configuration on Kubernetes (GKE) using Helm charts and Harness CI/CD pipelines.- Build and maintain CI/CD pipelines using Harness for Flink job packaging, testing, and deployment to GKE.- Manage Flink cluster configuration using Helm charts; maintain Helm values and chart templates for environment-specific configurations.- Coordinate with infrastructure and DBA teams for PostgreSQL slot management and StarRocks table design.Qualifications and Experience :Technical Skills (Must Have) :- 5+ years of experience in data engineering with a focus on ETL/ELT pipeline development.- 2+ years of hands-on production experience with Apache Flink (Flink SQL and/or DataStream API).- Strong proficiency in Java, Scala, or Python for Flink job development.- Experience with Change Data Capture (CDC) patterns : Flink CDC, Debezium, or equivalent.- Practical experience with PostgreSQL as a data source, including replication slots and WAL configuration.- Demonstrated experience with columnar OLAP databases (StarRocks, Doris, ClickHouse, or similar).- Solid understanding of distributed systems concepts: fault tolerance, exactly-once semantics, state management, and watermarking.- Experience with Kafka or similar message brokers as part of streaming architectures.Leadership and Soft Skills :- Proven ability to lead engineering teams, including task planning, code review, and mentoring.- Strong analytical and problem-solving skills with the ability to debug complex distributed pipeline issues.- Excellent written and verbal communication skills with the ability to document technical decisions clearly.- Self-driven with the ability to work autonomously and manage priorities in a fast-paced environment.Preferred Qualifications :- Hands-on experience with StarRocks Flink Connector and StarRocks primary-key table designs.- Hands-on experience with Flink on Kubernetes (GKE) using Flink Kubernetes Operator or native K8s mode.- Experience with Harness CI/CD and Helm chart management for data workloads.- Experience with Flink Table API and Flink SQL for unified batch and stream processing.- Knowledge of Apache Iceberg, Delta Lake, or Hudi for lakehouse architectures.- Experience with orchestration tools such as Apache Airflow or Prefect for hybrid batch/streaming workflows.- Familiarity with observability stacks : Prometheus, Grafana, and Flink metrics reporters.- Prior experience in a technical lead or senior engineer role with delivery accountability.- Understanding of data modelling best practices for analytical workloads in StarRocks.Domain Knowledge (Nice to Have) :- Experience in data-intensive domains such as data mining, large-scale data processing, or analytical platform engineering.- Understanding of end-to-end data process flows: from source system ingestion through transformation, aggregation, and consumption layers.- Familiarity with data governance, lineage tracking, and metadata management.- Exposure to real-time analytics use cases such as dashboards, operational reporting, or data exploration pipelines.Skan AI is an equal opportunity employer committed to building a diverse, inclusive, and respectful workplace around the world. We do not discriminate based on race, color, religion or belief, sex (including pregnancy, sexual orientation, gender identity, or gender expression), national origin, ancestry, age, disability, medical condition, genetic information, marital or family status, military or veteran status, or any other characteristic protected by applicable laws in the locations where we operate. We welcome people from all backgrounds and provide reasonable accommodations throughout the hiring process. (ref:hirist.tech)

Vacancy posted 20 days ago
Similar jobs that could be interesting for youBased on the Skan.ai - Senior Data Engineer - Apache Flink in Bangalore vacancy
  •  ...platform uses cutting-edge technology, big data, machine learning, and AI to seamlessly bring together OEMs,...  ...role goes beyond traditional data engineering and offers the opportunity to shape the...  ...data processing frameworks such as Apache Spark.   Strong understanding of modern... 
    Senior

    Tekion

    Bangalore
    16 days ago
  • Senior Software Engineer -Data Engineer Position Description Founded in 1976, CGI is among the largest independent IT and business consulting services...  .... . Experience with workflow orchestration tools like Apache Airflow, Prefect, or Dagster. . Strong experience in... 
    Senior
    Full time
    Local area
    Shift work
    Bangalore
    21 days ago
  •  ...As a Data Integration Engineer, you will have an opportunity to build critical Integration infrastructure that surfaces data and insights...  ...with streaming data processing frameworks (e.g., Apache Kafka, Apache Flink). Familiarity with containerization technologies (e.... 
    Senior
    Work at office

    Talkdesk

    Bangalore
    a month ago
  • Be at the Forefront of the Agentic AI Revolution :At Skan AI, we are pioneering the context engine for human and agentic execution, bringing context from enterprise operators...  ...shaping the future of work itself.We are seeking a Senior DevOps Engineer to design, build, and operate... 
    Senior
    Hybrid work
    Shift work

    Skan ai

    Bangalore
    20 days ago
  • About the Role:We are seeking an experienced Senior Data Engineer with strong expertise in Databricks, PySpark, Python, and Cloud Data Platforms...  ...scalable and reliable data pipelines using Databricks, Apache Spark, and PySpark.- Build and optimize ETL/ELT workflows, data... 
    Senior
    Permanent employment
    Full time

    PRI INDIA IT SERVICES PRIVATE LIMITED.

    Bangalore
    27 days ago
  • Senior Data EngineerHumyn Labs & KGen - Full-timeMultimodal PipelinesAWS Data LakeGPU ValidationAI...  ...multi-modal data signals for physical AI. Operating across 20+ countries in India,...  ...:We are looking for a Senior Data Engineer to own, extend, and harden the multimodal... 
    Senior
    Full time

    KGeN

    Bangalore
    12 days ago
  • About the Role :Radancys Data Engineering team is seeking a Senior Data Engineer to join our Bangalore Product Engineering team and help build the next generation of our data platform, insights products, and AI-powered data agents. This role is part of our evolution from serving... 
    Senior

    Associates in Advertising Ltd T/A Radancy

    Bangalore
    2 days ago
  •  ...on less than 10% of available data. We built Orbital to change that...  ...energy that lets companies use AI at scale, harnessing all of their...  ...The Role As our Data Engineer, you’ll architect and maintain...  ...of streaming frameworks (Kafka, Flink, Spark Streaming) or MLOps stacks... 
    Senior
    Hybrid work

    Applied Computing

    Bangalore
    19 days ago
  •  ...innovative platform providing comprehensive data extraction, monitoring, and valuation...  ...private capital industry. The company's AI-powered platform streamlines middle-office...  ...About the Role: We are looking for a Senior Data Engineer to build and own the data pipelines, integrations... 
    Senior
    Full time
    Work at office

    73 Strings

    Bangalore
    13 days ago
  •  ...About the team The mission of Roku’s Data Engineering team is to develop a world-class big...  ...of new and existing initiatives. As a Senior Data Engineer on Roku's   Content Data Engineering...  ..., and Presto/Trino. Deep expertise in Apache Spark is required, including performance... 
    Senior
    Hybrid work
    Work at office
    Local area
    Remote job
    Worldwide
    Monday to Thursday
    Flexible hours

    Roku

    Bangalore
    21 days ago
  •  ...are proud of:     The role   As a Data Engineer , you are passionate about experience innovation...  ...with data processing frameworks:   ~ Apache Spark, Hadoop ecosystem   ~ Apache...  ...world.   At the intersection of data, AI, creativity, and technology, we drive... 
    Senior
    Full time
    Hybrid work
    Remote job

    Valtech

    Bangalore
    a month ago
  •  ...seeking a visionary and execution-focused Senior Data Science Enginee r to spearhead the...  ...the intersection of Advanced Generative AI (RAG & Multi-Agent Systems), High-Performance...  ...Data Architectures, and Enterprise Cloud Engineering. You will own the optimization of our... 
    Senior
    Hybrid work
    Work at office
    Remote job
    Worldwide
    Flexible hours

    Evertz Microsystems Limited

    Bangalore
    a month ago
  • We are seeking a highly skilled Senior Data Engineer / Lead Data Engineer to join our growing Data Engineering team. This role is ideal for professionals...  ...that support analytics, business intelligence, reporting, and AI/ML initiatives.You will work closely with Data Architects,... 
    Senior
    Hybrid work

    Squareroot Consulting Pvt Ltd

    Bangalore
    5 days ago
  • Role : Big Data EngineerJob Description :The Big Data Engineer at Draup is responsible for building scalable techniques and...  ...in Python.- Proficiency in Apache Spark (PySpark) is a must.- Good work...  ...solutions with Spark/HDFS/MapReduce/Flink.- Worked with different types of file... 

    Draup

    Bangalore
    27 days ago
  • Be at the Forefront of the Agentic AI RevolutionAt Skan AI, we are pioneering the context engine for human and agentic execution, bringing context from enterprise operators...  ...future of work itself.The RoleWe are looking for a Senior Data Scientist who loves to get their hands dirty.... 
    Senior
    Full time
    Local area
    Day shift

    Skan ai

    Bangalore
    20 days ago
  •  ...3 days/ weekStart date : ASAPPosition Overview :NTT DATA Americas is seeking a highly skilled Senior Data Engineer to support a Teradata Utilization Analysis engagement...  ...-on Databricks engineering capability and applied AI/ML skills. The analysis will be performed primarily... 
    Senior

    ZINGMIND PRIVATE LIMITED

    Bangalore
    14 days ago
  •  ...About the Role As a Finance Data and AI Specialist, you will be a hands-on contributor building...  ...the Finance org. You will report to the Senior Manager, Finance Data and AI and work as a core member of a team that combines engineering rigour with Finance domain knowledge.... 
    Senior
    Worldwide

    Databricks

    Bangalore
    20 days ago
  • Be at the Forefront of the Agentic AI Revolution :At Skan AI, we are pioneering the context engine for human and agentic execution, bringing context from enterprise operators...  ...application performance and ensure efficient data storage and retrieval with SQL and/or MongoDB... 
    Senior

    Skan ai

    Bangalore
    20 days ago
  • We need a Senior Pyspark Developer to work for a leading investment bank...  ...6+ years of experience in data intensive Pyspark development...  ...Basic experience in Gen/Agentic AI Experience...  ...Apache Airflow... 
    Senior

    Luxoft

    Bangalore
    16 days ago
  •  ...Employee Groups and Celebration. Position Overview The Senior Data Engineering Manager is responsible for leading the delivery and operational...  ...engineering capabilities that underpin Data Products, analytics, and AI enablement. This role manages a team of data engineers and... 
    Senior

    Ralph Lauren

    Bangalore
    more than 2 months ago
  •  ...small group of highly skilled engineers, that own significant responsibility...  ...large-scale backend systems, data pipelines, storage, and...  ...Role We are looking for a Senior Software Engineer with vast experience...  .... ~ Experience with Apache Spark and Apache Flink. ~ Experience... 
    Senior
    Hybrid work
    Work at office
    Local area
    Remote job
    Monday to Thursday
    Flexible hours

    Roku

    Bangalore
    a month ago
  • Job Description :We are looking for an experienced Senior Snowflake Data Engineer to design, build, and scale modern data platforms that power analytics, conversational analytics, and AI-driven initiatives.The ideal candidate should have strong expertise in developing scalable... 
    Senior

    Volto Consulting

    Bangalore
    29 days ago
  • Job Title : Senior ML Platform EngineerExperience : 6 - 10 YearsLocation...  ...skilled Senior ML Platform Engineer to design, build, and manage scalable...  ...platforms and cloud-native AI infrastructure. The ideal...  ...will play a key role in enabling data scientists and ML engineers by... 
    Senior

    TIGER ANALYTICS INDIA CONSULTING PRIVATE LIMITED

    Bangalore
    8 days ago
  •  ...pipelines on AWS.- Build cloud-native data solutions using AWS Glue, S3,...  ...orchestrate workflows using Apache Airflow (DAGs).- Work with...  ...to support machine learning and AI-driven data solutions.- Ensure...  ...data solutions.- Mentor junior engineers, conduct code reviews, and drive... 
    Senior

    https://vysystems.com/

    Bangalore
    6 days ago
  •  ...seeking a highly skilled and experienced Cloud Data Engineer to join our Digital Technologies team....  ...team.- Experience in AIML and Generative AI in Areas of Data engineeringPerformance...  ...computing and big data frameworks like Apache Spark and Hadoop.- Knowledge of Cloud data... 
    Senior

    Creative Synergies Group

    Bangalore
    18 days ago
  •  ...Glance AI is an AI commerce platform shaping the next wave of e-...  ...Through Glance AI’s rich first-party data and unparalleled consumer...  ...Mithril Capital. SDE-3 – Data Engineering and Capabilities Team   About...  ...data pipelines using Spark, Flink, Kafka, and Airflow.   ~ Develop... 
    Full time
    Worldwide

    Glance

    Bangalore
    3 days ago
  •  ...About the Role We're building the data infrastructure that powers decisions across...  ...large-scale batch computation. As a Senior Data Platform Engineer, you'll own the systems that process...  ...Tech-Stack | Area | Tools | Compute | Apache Spark, Databricks, PySpark, Scala |... 
    Senior
    Long term contract
    Full time
    Local area
    Worldwide

    Aspora

    Bangalore
    14 hours ago
  •  ...Roku runs one of the largest data lakes in the world. We store over...  ...Scribe, Kafka, Hive, Presto, Spark, Flink, Pinot, and others. The team is...  ...~ Strong familiarity with the Apache Hadoop ecosystem: Spark, Kafka,...  ...degree in CS or equivalent ~ AI Literacy /AI growth mindset... 
    Senior
    Hybrid work
    Work at office
    Local area
    Remote job
    Monday to Thursday
    Flexible hours

    Roku

    Bangalore
    a month ago
  • Job Title : Data Engineer (GenAI)Location : Gurugram / Bangalore (Hybrid)Experience : 5+ YearsEmployment Type : Full-TimeMandatory Skills :- Python...  ...GenAI applications using LLMs and/or RAG frameworks.- Integrate AI solutions with enterprise data platforms.- Collaborate with... 
    Senior
    Full time
    Hybrid work

    Aspyra

    Bangalore
    6 days ago
  •  ...Role Overview We are looking for a Data Engineer to support and enhance the CS-AT platform, which processes large-scale machine log data (...  ...data pipelines • Improved processing efficiency • Reliable log ingestion and analysis • Smooth integration with AI systems... 
    Senior
    Bangalore
    26 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Skan.ai - Senior Data Engineer - Apache Flink. Be the first to apply!