Data Engineer - ETL/PySpark
Worksconsultancy
Data Pipeline Development & Operations :- Design, build, and operate scalable and reliable data pipelines on the Databricks platform- Develop end-to-end data workflows from ingestion through transformation to consumption- Implement robust error handling, monitoring, and alerting mechanisms- Ensure data pipeline reliability, performance, and maintainability- Optimize pipeline performance through efficient Spark job design and cluster configuration- Manage and orchestrate complex data workflows using Databricks Jobs and workflowsLegacy Code Modernization :- Refactor legacy code and data pipelines to PySpark for improved performance and scalability- Migrate traditional ETL processes to modern ELT patterns on Databricks- Assess existing codebases and identify opportunities for optimization and modernization- Ensure backward compatibility and data integrity during migration processes- Document refactoring approaches and create migration playbooks- Collaborate with stakeholders to minimize disruption during code transitionsData Engineering Excellence :- Implement data quality checks and validation frameworks- Design and maintain Delta Lake tables with appropriate optimization strategies- Develop reusable code libraries and frameworks for common data engineering tasks- Follow software engineering best practices including version control, testing, and CI/CD- Participate in code reviews and provide constructive feedback to team members- Troubleshoot and resolve data pipeline issues in production environmentsCollaboration & Knowledge Sharing :- Work closely with data architects, analysts, and business stakeholders- Collaborate with Infrastructure (Infra), Applications (Apps), and Cyber teams- Share knowledge and best practices with Team NCS- Mentor junior data engineers on PySpark and Databricks technologies- Document technical solutions and maintain comprehensive documentationEssential Technical Skills :- Data Engineering: Strong foundation in data engineering principles, ETL/ELT processes, and data pipeline design patterns- PySpark: Proven hands-on experience developing data pipelines using PySpark, including DataFrames API, Spark SQL, and performance optimization- Databricks Platform: Practical experience with Databricks workspace, cluster management, notebooks, and job orchestration- Workspace AI Agent: Knowledge of Databricks Workspace AI Agent capabilities and integration- Data Modelling: Experience implementing data models including dimensional modeling, data vault, or lakehouse architectures- Delta Lake: Understanding of Delta Lake features including ACID transactions, schema evolution, and optimization techniques- Python: Strong Python programming skills for data processing and automation 5+ years of relevant experience- Strong foundation in data engineering principles, ETL/ELT processes, and data pipeline design patterns- Proven hands-on experience developing data pipelines using PySpark, including DataFrames API, Spark SQL, and performance optimization- Min 2 to 3 yrs exp in Databricks Platform: Practical experience with Databricks workspace, cluster management, notebooks, and job orchestration- Experience implementing data models including dimensional modeling, data vault, or lakehouse architectures- Understanding of Delta Lake features including ACID transactions, schema evolution, and optimization techniques- Strong Python programming skills for data processing and automation- Experience with cloud platforms (Azure, AWS, or GCP) mandatory to have at least one certification- Databricks Certified Data Engineer Associate OR Databricks Certified Data Engineer ProfessionalNotice Period - Immediate to 30 days (ref:hirist.tech)
- ...Skillsets : - 3+ years of experience in analytics, Pyspark, Python, Spark, SQL and associated data engineering jobs.- Must have experience with managing and transforming... ...: - Ability to understand data models and identify ETL optimization opportunities. - Exposure to ETL tools...DataImmediate start
- ...PRIVATE LIMITED is a technology services firm specializing in data engineering, cloud transformation, and advanced analytics. We partner... ....Key Responsibilities:- Develop and optimize scalable ETL pipelines using PySpark and Apache Spark to process large volumes of structured...DataHybrid work
- Role Overview :As a PySpark Developer, you will serve as a critical architect within our data engineering team, responsible for designing, developing, and maintaining high-performance... ...decision-making. By optimizing complex ETL workflows and ensuring the scalability of our...DataHybrid work
- ...about building and optimizing cloud-based data platforms that support modern analytics initiatives... ..., and supporting migration from legacy ETL systems to dbt-driven ELT architectures.... ...Platform : - Collaborate with Snowflake engineering teams to optimize execution strategy,...DataFull time
- Requirements : - At least 5 years of experience with data projects, strong in Data warehousing, building data lakes, AWS, Apache airflow, PySpark, SQL, Metadata management, Big Data implementation experience.- ETL/ELT, Python, Pyspark, AWS, Airflow, strong in SQL/Hive, metadata...DataImmediate start
- Role Summary:We are seeking an experienced Data Engineer with strong expertise in Informatica Intelligent Cloud Services (IICS) and Snowflake,... ...use cases.Key Responsibilities:- Design, develop, and maintain ETL/ELT pipelines using IICS.- Build scalable data models and transformations...DataLong term contract
- ...possible before the AI era. Our software engineering teams are highly valued by customers, whether... ...our journey!About the role : As a Senior Data Engineer BI Analytics & DWH, you will... ...experience, youll apply your solid foundation in ETL/ELT development, data modeling, and...DataFull time
- Job Description:Jash Data Sciences: Letting Data Speak!Do you love... ...-edge Data Sciences and Data Engineering startup based in Pune, India.... ...the target database.- Implement ETL/ ELT processes in the cloud... ...with Big Data technologies like PySpark/ Hadoop.- A good team player with...Data
- Job Summary : We are hiring a Python PySpark Data Engineer to develop scalable data processing applications and high-performance data pipelines. The... ...Build batch and real-time data processing solutions.- Design ETL workflows for large datasets.- Optimize Spark jobs for...Data
- ...ideal candidate should have hands-on experience in Reltio Master Data Management (MDM), data integration, data quality, and customer/domain... ...to Have :- Experience with Informatica, Talend, or other ETL tools- Exposure to cloud platforms like AWS/Azure/GCP- Knowledge of...Data
- ...client is looking for a highly skilled Senior Data Engineer to design, build, and optimize scalable... ...-scale data processing- Databricks & PySpark ecosystems- Real-time streaming systems-... ...Responsibilities :- Design and build scalable ETL/ELT pipelines using PySpark, SQL, and...Data
- Data Engineer - Pyspark ETL Hadoop Python SQLWe are seeking a skilled PySpark Developer to join our Data Engineering team. The ideal candidate should possess strong hands-on experience in PySpark, Python, and SQL, with expertise in developing scalable data processing solutions...Data
- Job Title : Python Data TestingTotal Exp : 5 to 8 yrsLocation : Hyderabad, Chennai, Pune, Mumbai, Bangalore, Noida, KolkataNotice Period... ...to validate data accuracy and integrity.- Collaborate with data engineering and analytics teams to understand data flows and testing...DataImmediate start
- ...Responsibilities :- Design, develop, and maintain large-scale data pipelines using Snowflake SQL, ETL processes, and Snowpipe.- Develop complex data models... ...GraduateKey Skills :- Snowflake, ETL Pipelines, Data Engineering, Snowsql, Snowflake Modeling, Snowpipe, ELT, SQL,...Data
- ...We are looking for a skilled SQL Developer/Engineer to design, develop, and maintain robust... ...and analytics. You will work closely with data analysts, application developers, and stakeholders... ...extraction, transformation, and loading (ETL).- Ensure data integrity, accuracy, and...Data
- Role Overview :As a Senior Data Engineer, you will serve as a technical cornerstone in building... ...scalable data processing pipelines using PySpark to handle large-scale datasets, ensuring... ...extraction, transformation, and loading (ETL) processes that support critical business...DataLong term contractHybrid work
- Senior Data Engineer (PySpark) :Location :Pune (Work from Office - 5 Days a Week).Job Summary :We are looking for a highly skilled and experienced... ...pipelines using Python, PySpark, and SQL.- Build and optimize ETL/ELT processes for large-scale data ingestion, transformation,...DataWork at office
- We are hiring for Data Engineer - 5-12 years experienceResponsibilities :- Build and Optimize ELT/ETL Pipelines using Big Query, GCS, Dataflow, Pub Sub and Orchestration services Composer/Airflow.- Hands-On experience in building ETL/ETL Pipelines with developing software code...Data
- ...Job Overview : We are looking for a Software Development Engineer in Test (SDET) to join our Data Engineering team. SDET has a dual role : They are both Data... ...the quality and reliability of complex data pipelines, ETL processes, and large-scale data systems. The role requires...Data
- Role Overview :We are seeking a seasoned Data Engineering Lead to spearhead our data integration and warehousing initiatives. In this role, you will be responsible for architecting and maintaining robust ETL pipelines using Matillion to power our Snowflake data ecosystem. You...DataHybrid work
- Description : Data Pipeline Development & Operations : - Design,... ...legacy code and data pipelines to PySpark for improved performance and... ....- Migrate traditional ETL processes to modern ELT patterns... ...during code transitions.Data Engineering Excellence : - Implement data...Data
- Description :As a Data Engineer, you will be an important part of the Enterprise Business Unit (... ...experience in data engineering- Strong Python ETL scripting and pipeline development-... ...partitioned data- Familiarity with Spark or PySpark for large-scale data ingestion- GitHub...DataFull timeWork from homeFlexible hours
- ...Script and have 6+ years of experience with big data frameworks like Spark, Hadoop or Kafka.- You excel in Apache Spark/PySpark fundamentals.- You excel in using stream and... ...data governance practices.- Data Engineering Expertise: You have advanced skills in designing...DataFull timeWork from homeFlexible hours
- ...Responsibilities : - Design, develop, and maintain large-scale data pipelines using Snowflake SQL, ETL processes, and Snowpipe.- Develop complex data models... ...field, reflecting a strong foundation in data engineering principles.- A track record of delivering high-quality...DataHybrid work
- Job Title : Databricks on AWS and PySpark EngineerJob SummaryWe're seeking an experienced Databricks on AWS and PySpark Engineer to join our team. The ideal candidate will have a strong... ..., building, and maintaining large-scale data pipelines and architectures using...Data
- Job Title : Azure Data Engineer - PySpark - RemoteJob Description : The engineer will be working within an established Microsoft Fabric data platform... .../ Databricks)- Strong data pipeline development experience (ETL / ELT)- Experience working within Medallion architecture (...DataFull timeWork at officeRemote jobWorking Monday to Friday
- ...- Design, develop, and maintain ETL/ELT pipelines using Microsoft Fabric.- Integrate data from multiple enterprise systems... ...stakeholders, data architects, and engineering teams to deliver end-to-end data... ...& ETL/ELT- SQL, Python & PySpark- Power BI- Azure AI Services & Azure...Data
- Job Title : Azure Data EngineerLocation : PuneExperience : 6+ YearsJob Summary :We are seeking... ...a skilled and experienced Azure Data Engineer to design, develop, and optimize scalable... ...have strong expertise in Azure Databricks, PySpark, Azure Data Factory (ADF), Azure Data Lake...Data
- Job Description : SAS ETL Developer (Banking Domain)Job Title : SAS ETL DeveloperLocation... ...strong expertise in SAS ETL development, data integration, testing, deployment, and production... ...Computer Science, Information Technology, Engineering, Statistics, Mathematics, or related field...DataFull timeHybrid work
$ 1.5 per day
...billion. We are looking for a Mid-level Business Intelligence Engineer to join our global team. If you are highly intelligent,... ...your opportunity! The individual in this role will perform data analysis, ELT/ETL design and support functions to deliver on strategic initiatives...DataRemote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer - ETL/PySpark. Be the first to apply!
