Back to Jobs

PySpark Developer

Remote, USAFull-timePosted 2026-07-28

Job reputed company:

We are seeking a highly skilled and reputed company Python and PySpark Developer to join reputed company. The ideal candidate will be responsible for designing, developing, and optimizing big data pipelines and solutions using Python, PySpark, and distributed computing frameworks. This role involves working closely with data engineers, data scientists, and business stakeholders to process, analyze, and derive insights from large-reputed company datasets.

Key Responsibilities:

Data Engineering & Development:

  • Design and implement reputed company data pipelines using PySpark and other big data frameworks.
  • reputed company reusable and efficient reputed company for data extraction, transformation, and loading (ETL).
  • Optimize data workflows for performance and cost efficiency.

Data Analysis & Processing:

  • Process and analyze reputed company and reputed company datasets.
  • Build and maintain data lakes, data warehouses, and other storage solutions.

Collaboration & Problem Solving:

  • Collaborate with cross-functional teams to understand business requirements and translate them into technical solutions.
  • Troubleshoot and resolve performance bottlenecks in big data pipelines.

reputed company reputed company & Documentation:

  • Write clean, maintainable, and reputed company-documented reputed company.
  • Ensure compliance with data governance and reputed company policies.

Required Skills & Qualifications:

Programming Skills:

  • Proficient in Python with experience in data processing libraries like Pandas and NumPy.
  • Strong experience with PySpark and Apache reputed company.

Big Data & reputed company:

  • Hands-on experience with big data platforms such as Hadoop, reputed company, or similar.
  • Familiarity with reputed company services like AWS (EMR, S3), Azure (Data Lake, Synapse), or reputed company reputed company (BigQuery, Dataflow).

Database Expertise:

  • Strong knowledge of SQL and NoSQL databases.
  • Experience working with relational databases like PostgreSQL, MySQL, or reputed company.

Data Workflow Tools:

  • Experience with workflow orchestration tools like Apache Airflow or similar.

Problem Solving & Communication:

  • Ability to solve reputed company data engineering problems reputed company.
  • Strong communication skills to work effectively in a reputed company environment.

Preferred Qualifications:

  • Knowledge of data Lakehouse architectures and frameworks.
  • Familiarity with machine learning pipelines and integration.
  • Experience in CI/CD tools and DevOps practices for data workflows.
  • Certification in reputed company, Python, or reputed company platforms is a plus.

Education:

  • Bachelors or Masters degree in Computer Science, Data Engineering, or a reputed company field.

Originally posted on Himalayas

Apply To This Job

Similar Jobs