AWS Python Developer with Pyspark

Polar IT Services

  • Newark, New Jersey
  • 30+ days ago

    Highlights

    The ideal candidate will have hands-on experience developing scalable data pipelines, processing large data sets, and integrating with cloud-based environments. Key Responsibilities: Design, develop, and maintain data pipelines and ETL workflows using Python, PySpark, and AWS services.

    Numbers & Facts

    LocationNewark, New Jersey
    Websitepolarits.com&d=DwMGaQ&c=euGZstcaTDllvimEN8b7jXrwqOf-v5A_CdpgnVfiiMM&r=LQkEbWLPKa65oL-CeXFMcJuP1fDvqOXNh5leLUyUOpk&m=Vg2j9rr4ny_lISjowRH4f8_a48N2SG63UYwmniphkt1NO5AbixjCLAaIRULrwIEe&s=mWRi-fYc6tRcu6NWXQbo7YfhnaFpoOlUQcd0YZRWZo0&e=

    Description

    Hello Folks,
    Hope you are doing good!

    Please find the below requirement and let me know your interest?
    Title: Python Developer with AWS & Pyspark(It’s backfill Role)
    Location: New Jersey
    Mode of work: Hybrid
    Visa:GC, GC EAD, USC, H4 EAD only
    Looking for Local candidates only with 10+years of experience and final round interview will be In-person interview

    Job Summary:
    We are seeking an experienced Python Developer with strong expertise in AWS and PySpark to join our data engineering team. The ideal candidate will have hands-on experience developing scalable data pipelines, processing large data sets, and integrating with cloud-based environments. This role requires excellent problem-solving skills and a strong understanding of distributed data processing frameworks.
    Key Responsibilities:
    • Design, develop, and maintain data pipelines and ETL workflows using Python, PySpark, and AWS services.
    • Build and optimize large-scale data processing and data transformation solutions.
    • Integrate various data sources and ensure data quality, performance, and reliability.
    • Collaborate with data engineers, analysts, and architects to deliver end-to-end data solutions.
    • Implement best practices for code optimization, error handling, and data validation.
    • Participate in code reviews, documentation, and deployment automation.
    • Ensure adherence to data security and compliance standards.
    Required Skills & Qualifications:
    • Bachelor’s degree in Computer Science, Data Engineering, or a related field.
    • 10+ years of experience in software development with a strong focus on Python.
    • Hands-on experience with PySpark for distributed data processing.
    • Solid understanding of AWS cloud services such as S3, Glue, Lambda, EMR, Redshift, and Athena.
    • Strong experience in ETL development and data pipeline orchestration.
    • Familiarity with SQL and relational/non-relational databases.
    • Excellent analytical, debugging, and communication skills.
    Preferred Skills:
    • Experience with Airflow, Databricks, or other workflow management tools.
    • Knowledge of CI/CD pipelines and version control tools like Git.
    • Exposure to data lake or data warehouse architectures.
    • Familiarity with Docker or Kubernetes for deployment.


    Thanks & Regards
    Jagdish
    Manager – IT Staffing
    Email ID:

    jagdish@polarits.com

     |  Phone: +1 443 489 4433
     

    6095 Marshalee Dr, Suite 250, Elkridge, MD 21075
    www.polarits.com          
    Consulting | Technology | Development
    Offices: Maryland | Georgia | India
     
    Note: If you have received this message in error, please notify us immediately by reply e-mail so that we may correct our internal records. Please then delete the original message (including any attachments) in its entirety. Thank you


     

    Similar Jobs

    See more jobs