Job Description
We are currently looking to hire a Data Engineer. This is an exciting opportunity to expand your skill set, achieve job satisfaction and work-life balance. More details as below.
Location : Gurgaon (Hybrid)
Experience: 6+ years
Type : Contract (12 months, extendable)
Key Responsibilities:
• Design, develop, and optimize data pipelines using PySpark for large-scale data processing.
• Implement and maintain data integration solutions, ensuring efficient data movement across platforms.
• Work with SQL to write complex queries, optimize performance, and manage relational databases.
• Utilize Apache Airflow for workflow orchestration, scheduling, and monitoring of ETL processes.
• Work with Kafka for real-time data streaming and processing.
• Ensure data quality, integrity, and security in cloud or on-premise environments.
• Collaborate with cross-functional teams, including Data Scientists, Analysts, and DevOps, to support data-driven decision-making.
• Troubleshoot and resolve data pipeline and performance issues.
Required Skills & Qualifications:
• 6+ years of experience as a Data Engineer or in a similar role.
• Expertise in PySpark for distributed data processing.
• Strong proficiency in SQL for data querying and optimization.
• Experience in Apache Airflow for task scheduling and workflow orchestration.
• Intermediate knowledge of Kafka for real-time data processing.
• Cloud expertise (preferably AWS) with experience in S3, EMR, or similar cloud-based big data services.
• Hands-on experience in Python for scripting and automation.
• Understanding of data modeling, ETL processes, and data warehousing concepts.
• Experience working with structured and unstructured data in big data ecosystems.
• Strong analytical and problem-solving skills with the ability to optimize performance.
• Excellent communication and collaboration skills.
Preferred Qualifications:
• Experience with AWS Glue, Redshift, or Snowflake.
• Knowledge of containerization tools like Docker and orchestration with Kubernetes.
• Familiarity with CI/CD pipelines for data engineering.
• Experience with NoSQL databases like MongoDB or Cassandra.
WHAT’S ON OFFER:
You will be remunerated with an excellent base salary and entitled to attractive company benefits. Additionally, you will get the opportunity to enjoy a fun and collaborative work environment, alongside a strong career progression.
To submit your application, please apply online or email your UPDATED CV in Microsoft Word format to Swati.J@aven-sys.com Your interest will be treated with strict confidentiality.
CONSULTANT DETAILS:
Consultant Name : Swati Jaiswal
Avensys Consulting Pte Ltd
Email : Swati.J@aven-sys.com
Whatsapp : +65 6761 +826
Privacy Statement: Data collected will be used for recruitment purposes only. Personal data provided will be used strictly in accordance with the relevant data protection law and Avensys
Upon submission of your CV, you grant Avensys Consulting permission to retain your personal information in our electronic database, unless you specify otherwise. This data will be used to evaluate your suitability for current and potential job openings within our organization. Should you wish to have your personal data removed at any point, a simple notification to us will suffice.
Rest assured, we will not disclose your personal information to any third parties, and we remain steadfast in our commitment to providing equal opportunities to all applicants.
💡 Quick Summary
Seeking a career-building opportunity? The Data Engineer - AWS & Pyspark position is now open for candidates interested in the IT Engineer & Developer Jobs sector. This role in Gurgaon offers a professional environment and growth potential.
Requirement Snapshot: Candidates should possess basic communication skills, a proactive attitude, and the ability to work in a team. Experience in IT Engineer & Developer Jobs is a plus.
