AWS Python Developer with Pyspark

Polar IT Services

Newark, New Jersey

JOB DETAILS
SKILLS
AWS Lambda, Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Analysis Skills, Automation, Best Practices, Cloud Computing, Code Reviews, Communication Skills, Computer Science, Consulting, Continuous Deployment/Delivery, Continuous Integration, Data Analysis, Data Lake, Data Management, Data Processing, Data Quality, Data Sets, Data Warehousing, Database Extract Transform and Load (ETL), Debugging Skills, Docker, Documentation, Electronic Medical Records, Error Handling, Git, Information/Data Security (InfoSec), Maintain Compliance, Management of Information Systems/Technology (MIS), Problem Solving Skills, Python Programming/Scripting Language, Regulatory Compliance, Relational Databases (RDBMS), SQL (Structured Query Language), Sales Pipeline, Scalable System Development, Software Development, Software Engineering, Source Code/Configuration Management (SCM), Technical Recruiting
LOCATION
Newark, New Jersey
POSTED
30+ days ago
Hello Folks,
Hope you are doing good!

Please find the below requirement and let me know your interest?
Title: Python Developer with AWS & Pyspark(It’s backfill Role)
Location: New Jersey
Mode of work: Hybrid
Visa:GC, GC EAD, USC, H4 EAD only
Looking for Local candidates only with 10+years of experience and final round interview will be In-person interview

Job Summary:
We are seeking an experienced Python Developer with strong expertise in AWS and PySpark to join our data engineering team. The ideal candidate will have hands-on experience developing scalable data pipelines, processing large data sets, and integrating with cloud-based environments. This role requires excellent problem-solving skills and a strong understanding of distributed data processing frameworks.
Key Responsibilities:
  • Design, develop, and maintain data pipelines and ETL workflows using Python, PySpark, and AWS services.
  • Build and optimize large-scale data processing and data transformation solutions.
  • Integrate various data sources and ensure data quality, performance, and reliability.
  • Collaborate with data engineers, analysts, and architects to deliver end-to-end data solutions.
  • Implement best practices for code optimization, error handling, and data validation.
  • Participate in code reviews, documentation, and deployment automation.
  • Ensure adherence to data security and compliance standards.
Required Skills & Qualifications:
  • Bachelor’s degree in Computer Science, Data Engineering, or a related field.
  • 10+ years of experience in software development with a strong focus on Python.
  • Hands-on experience with PySpark for distributed data processing.
  • Solid understanding of AWS cloud services such as S3, Glue, Lambda, EMR, Redshift, and Athena.
  • Strong experience in ETL development and data pipeline orchestration.
  • Familiarity with SQL and relational/non-relational databases.
  • Excellent analytical, debugging, and communication skills.
Preferred Skills:
  • Experience with Airflow, Databricks, or other workflow management tools.
  • Knowledge of CI/CD pipelines and version control tools like Git.
  • Exposure to data lake or data warehouse architectures.
  • Familiarity with Docker or Kubernetes for deployment.


Thanks & Regards
Jagdish
Manager – IT Staffing
Email ID:

jagdish@polarits.com

 |  Phone: +1 443 489 4433
 

6095 Marshalee Dr, Suite 250, Elkridge, MD 21075
www.polarits.com          
Consulting | Technology | Development
Offices: Maryland | Georgia | India
 
Note: If you have received this message in error, please notify us immediately by reply e-mail so that we may correct our internal records. Please then delete the original message (including any attachments) in its entirety. Thank you


 

About the Company

P

Polar IT Services