hirly

Apply with hirly

Pyspark Developer

Infosys · Pune, India

Upload your resume to see how well you match this job — free, in seconds, no account needed.

Your resume is used only to score it against this job. If you don't create an account, it is deleted within 24 hours.

Already have an account? Sign in to see your saved application

We are seeking a skilled Pyspark Developer with 3-5 years of experience in building scalable solutions on Pyspark. The ideal candidate should have strong expertise in Pyspark with Spark, Hadoop, SQL exposure to Hive, Sqoop, UNIX shell scripting. Responsibilities

  • At least 3+ years of experience in designing and developing large scale, distributed data processing pipelines using PySpark, Hadoop and related technologies.
  • Having expertise in Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark Streaming
  • Experience with Hadoop, HDFS, Hive and other BigData technologies.
  • Familiarity with Data warehousing and ETL concepts and techniques
  • UNIX shell scripting will be an added advantage in scheduling/running application jobs.
  • At least 3 years of experience in Project development life cycle activities and development/maintenance projects
  • Work with business stakeholders and other SMEs to understand high level business requirements.
  • Work with the Solution Designers and contribute to the development of project plans by participating in the scoping and estimating of proposed project.
  • Apply technical background understanding, business knowledge, system knowledge in the elicitation of Systems Requirements for projects.
  • Possess good knowledge on Spark architecture and transformations using Spark and PySpark.
  • Work in an Agile environment and participation in scrum daily standups, sprint planning reviews and retrospectives.
  • Understand project requirements and translate them into technical solutions which meets the project quality standards
  • Ability to work in team in diverse/multiple stakeholder environment and collaborate with upstream/downstream functional teams to identify, troubleshoot and resolve data issues. Technical requirements Primary skills: Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark Streaming Hadoop, HDFS, Hive and other BigData technologies. Strong problem solving and Good Analytical skills.
  • Excellent verbal and written communication skills.
  • Experience and desire to work in a Global delivery environment.
  • Stay up to date with new technologies and industry trends in Development. Additional responsibilities Infosys is a global leader in next-generation digital services and consulting. Within the Data & Analytics unit, you will work on cutting-edge data engineering and analytics initiatives for global clients across industries. The role offers opportunities to work with cloud-native data platforms, modern analytics ecosystems, AI-driven solutions, and large-scale enterprise data modernization programs. Employees benefit from continuous learning programs, certifications, and career growth opportunities within Infosys' data and analytics practice Education Master Of Comp. Applications,Master Of Technology,Bachelor Of Comp. Applications,Bachelor Of Science,Bachelor of Engineering,Bachelor Of Technology