Alp Consulting Ltd.
Pyspark Developer
₹ Check with seller / month
✓ Actively Hiring
📍 Haridwar
💼 Full Time
🛡️ Verified Listing
⚡ Direct Apply — No Agent
🔒 Your Data is Safe
⭐ Trusted by 5 Lakh+ Jobseekers
Job at a Glance
- Category
- Software Developer
- Location
- Haridwar, Uttarakhand, India
- Salary
- Check with seller
- Job Type
- Full Time
- Company
- Alp Consulting Ltd.
- Status
- Open & Active
Job Description
JD:
Strong hands on experience on PySpark
• Good experience in AWS services, Oozie, Airflow
• Good Understanding experience in Hadoop, Hive, Oozie, HDFS, YARN, Sqoop
• Should have experience/understanding of AWS design and architectural concepts
• Clear in communication, ability to understand and articulates solution clearlyBuilding ETL/ELT jobs for batch data with HiveQL and Spark (Scala/Java/Pyspark), Scheduling Jobs using oozie, Data loads/extraction using Sqoop.
• Building Real-Time ingestion using Kafka and Spark Streaming, Data flow pipelines using Ni-Fi
• Building data pipelines within Big Data Eco-Systems with large structured/unstructured data from multiple sources.
Implementing large scale data platforms, Ingestion Automation Frameworks, Data Ops with Industry standards, reusable data products and business ready data by utilizing modern and open source technologies, Building automating end to end data lifecycle through CI/CD processes/tools, using Docker & Github.
Strong hands on experience on PySpark
• Good experience in AWS services, Oozie, Airflow
• Good Understanding experience in Hadoop, Hive, Oozie, HDFS, YARN, Sqoop
• Should have experience/understanding of AWS design and architectural concepts
• Clear in communication, ability to understand and articulates solution clearlyBuilding ETL/ELT jobs for batch data with HiveQL and Spark (Scala/Java/Pyspark), Scheduling Jobs using oozie, Data loads/extraction using Sqoop.
• Building Real-Time ingestion using Kafka and Spark Streaming, Data flow pipelines using Ni-Fi
• Building data pipelines within Big Data Eco-Systems with large structured/unstructured data from multiple sources.
Implementing large scale data platforms, Ingestion Automation Frameworks, Data Ops with Industry standards, reusable data products and business ready data by utilizing modern and open source technologies, Building automating end to end data lifecycle through CI/CD processes/tools, using Docker & Github.
Job Safety Alert
Real jobs on Jobsiya are always free. Never pay for an interview and never share bank or OTP details.
Report this job →
Similar Jobs:
Pyspark Developer Jobs in Haridwar
—
Software Developer Jobs Near You