Python pyspark developer

Apply now

Location

onsite, Hyderabad, India

Commitment

Full Time

Level

Senior (5+ years)

Required skills

PythonPySparkData EngineeringETL PipelinesELT PipelinesAWSSQLREST APIsScalaData IntegrationCloud TechnologiesDistributed Data ProcessingDatabase ManagementData Analytics

Job Description

Python & PySpark Developer

Summary

We are seeking an experienced Python and PySpark Developer to join our IT Services team. In this pivotal role, you will architect and implement robust data solutions, focusing on the design, development, and optimization of complex data pipelines. You will play a critical part in transforming raw data into actionable insights by leveraging cloud-native technologies and advanced processing frameworks. This position requires a deep understanding of data engineering principles to ensure high availability, scalability, and efficiency across our data infrastructure.

Responsibilities

  • Design, build, and maintain scalable ETL and ELT pipelines to facilitate seamless data movement and transformation.
  • Develop high-performance data processing applications using Python and PySpark to handle large-scale datasets.
  • Integrate and manage various AWS data services to support cloud-based data operations and storage requirements.
  • Create and optimize SQL queries for data extraction, manipulation, and reporting within complex database environments.
  • Implement and maintain REST APIs to enable efficient data integration between disparate systems and applications.
  • Utilize cloud technologies to deploy and manage data solutions, ensuring security and compliance with industry standards.
  • Collaborate with cross-functional teams to define data requirements and deliver data engineering solutions that meet business objectives.

Requirements

Requirements:

  • Experience: 6 to 9 years of professional experience in data engineering or related software development roles.
  • Core Programming: Proficiency in Python and Scala for developing data-intensive applications.
  • Big Data Technologies: Strong expertise in PySpark for distributed data processing and analytics.
  • Data Engineering: Demonstrated ability to design and implement ETL/ELT pipelines and data integration strategies.
  • Cloud & Infrastructure: Hands-on experience with AWS data services and general cloud technologies.
  • Database Skills: Advanced knowledge of SQL and relational database management systems.
  • API Development: Competence in building and consuming REST APIs for data connectivity.
  • Industry Context: Familiarity with the IT Services sector and its specific data handling challenges.

Ready to join the team?

Apply now