
SPARK Data Onboarding Engineer
4 weeks ago
We are seeking a skilled PySpark Data Engineer to join our team and drive the development of robust data processing and transformation solutions within our data platform. You will be responsible for designing, implementing, and maintaining PySpark-based applications to handle complex data processing tasks, ensure data quality, and integrate with diverse data sources. The ideal candidate possesses strong PySpark development skills, experience with big data technologies, and the ability to work in a fast-paced, data-driven environment.
Key Responsibilities:Data Engineering Development:
- Design, develop, and test PySpark-based applications to process, transform, and analyze large-scale datasets from various sources, including relational databases, NoSQL databases, batch files, and real-time data streams.
- Implement efficient data transformation and aggregation using PySpark and relevant big data frameworks.
- Develop robust error handling and exception management mechanisms to ensure data integrity and system resilience within Spark jobs.
- Optimize PySpark jobs for performance, including partitioning, caching, and tuning of Spark configurations.
Data Analysis and Transformation:
- Collaborate with data analysts, data scientists, and data architects to understand data processing requirements and deliver high-quality data solutions.
- Analyze and interpret data structures, formats, and relationships to implement effective data transformations using PySpark.
- Work with distributed datasets in Spark, ensuring optimal performance for large-scale data processing and analytics.
Data Integration and ETL:
- Design and implement ETL (Extract, Transform, Load) processes to ingest and integrate data from various sources, ensuring consistency, accuracy, and performance.
- Integrate PySpark applications with data sources such as SQL databases, NoSQL databases, data lakes, and streaming platforms
Qualifications and Skills:
- Bachelors degreein Computer Science, Information Technology, or a related field.
- 5+ yearsof hands-on experience in big data development, preferably with exposure to data-intensive applications.
- Strong understanding ofdata processing principles, techniques, and best practices in a big data environment.
- Proficiency in PySpark, Apache Spark, and related big data technologiesfor data processing, analysis, and integration.
- Experience withETL developmentand data pipeline orchestration tools (e.g., Apache Airflow, Luigi).
- Strong analytical and problem-solving skills, with the ability to translate business requirements into technical solutions.
- Excellent communication and collaboration skills to work effectively with data analysts, data architects, and other team members.
Role:Data Science & Machine Learning - Other
Industry Type:IT Services & Consulting
Department:Data Science & Analytics
Employment Type:Full Time, Permanent
Role Category:Data Science & Machine Learning
Education
UG:Any Graduate
PG:Any Postgraduate
-
TechVerito - Data Engineer - Python/Spark
3 days ago
Pune, Maharashtra, India TechVerito Software Solutions LLP Full time ₹ 6,00,000 - ₹ 18,00,000 per yearAbout the Role : We are looking for a Data Engineer with strong experience in Spark (PySpark), SQL, and data pipeline architecture. You will play a critical role in designing, building, and optimizing data workflows that enable scalable analytics and real-time insights. The ideal candidate is hands-on, detail-oriented, and passionate about...
-
Java Spark Developer
1 day ago
Pune, Maharashtra, India Wipro Full time ₹ 15,00,000 - ₹ 25,00,000 per yearJava+SparkPrimary skill - Apache Spark Secondary skill - JavaStrong knowledge in Apache Spark framework Core Spark, Spark Data Frames, Spark streamingHands-on experience in any one of the programming languages (Java)Good understanding of distributed programming concepts.Experience in optimizing Spark DAG, and Hive queries on TezExperience using tools like...
-
SaaS Onboarding Engineer
3 days ago
Pune, Maharashtra, India NGDATA Full time ₹ 5,00,000 - ₹ 10,00,000 per yearThe SaaS Onboarding Engineer ensures new customers are successfully set up and configured on our SaaS platform, enabling a smooth transition to operational use. They guide customers through technical onboarding, integrations, and best practices to ensure a quick and effective go-live.Your Main ResponsibilitiesCoordinate and manage SaaS onboarding activities...
-
Data Engineer
6 days ago
Pune, Maharashtra, India Tata Consultancy Services Full time ₹ 15,00,000 - ₹ 25,00,000 per yearJob Title :- Data Engineer - PysparkExperience: 5 to 8 YearsLocation: Pune/HyderabadJob DescriptionRequired Skills:5+ years of experience in Big data and pysparkMust-HaveGood work experience on Big Data Platforms like Hadoop, Spark, Scala, Hive, Impala, SQLGood-to-HaveGood Spark, Pyspark,Big Data experienceSpark UI/Optimization/debugging techniquesGood...
-
Hadoop,Spark Data Engineer/Consultant Specialist
4 weeks ago
Pune, Maharashtra, India HSBC Full timeJob DescriptionJob descriptionSome careers shine brighter than others.If you're looking for a career that will help you stand out, join HSBC and fulfil your potential. Whether you want a career that could take you to the top, or simply take you in an exciting new direction, HSBC offers opportunities, support and rewards that will take you further.HSBC is one...
-
Data Engineer
5 days ago
Pune, Maharashtra, India Techverito Software Solutions LLP Full time ₹ 15,00,000 - ₹ 25,00,000 per yearJob DescriptionWe are looking for 4-5years Data Engineer with strong experience in Spark (PySpark), SQL, and data pipeline architecture. You will play a critical role in designing, building, and optimizing data workflows that enable scalable analytics and real-time insights. The ideal candidate is hands-on, detail-oriented, and passionate about crafting...
-
Spark/Scala Developer
3 days ago
Pune, Maharashtra, India Cyanous Software Private Limited Full time US$ 1,50,000 - US$ 2,00,000 per yearDetailed JD (Roles And Responsibilities)At least 8+ years of experience and strong knowledge in Scala programming language.Able to write clean, maintainable and efficient Scala code following best practices.Good knowledge on the fundamental Data Structures and their usageAt least 8+ years of experience in designing and developing large scale, distributed...
-
Data Engineer
4 days ago
Pune, Maharashtra, India Talent21 Management Shared Services Pvt. ltd. Full time ₹ 15,00,000 - ₹ 25,00,000 per yearData Engineer should have extensive knowledge in different programming or scripting languages [Python]Good understanding in streaming (Kafka/Storm/Kinesis)Good understanding on Big data components (HDFS, YARN, Map Reduce, Spark, Oozie)Good understanding on Azure components (ADF, ADB, ADLS)Good understanding Version controlling (Git, GitHub, azure...
-
Big Data Engineer
6 days ago
Pune, Maharashtra, India Leinex Consulting Full time ₹ 12,00,000 - ₹ 36,00,000 per yearRole & responsibilities : - Evaluate domain, financial and technical feasibility of solution ideas with help of all key stakeholders - Design, develop, and maintain highly scalable data processing applications - Write efficient, reusable and well documented code - Deliver big data projects using Spark, Scala , Python, SQL - Maintain and tune existing...
-
Senior Data Engineer
1 day ago
Pune, Maharashtra, India Simplify Healthcare Full time ₹ 20,00,000 - ₹ 25,00,000 per yearReady to take your career to the next level. Join a team where innovation meets impact.Job Title: Senior Developer (Big Data)Location: Pune Magarpatta (Hybrid)Company: Simplify HealthcareExperience: 5-10 YearsIndustry: Healthcare Employment Type: Full TimeAbout Simplify HealthcareSimplify Healthcare is a rapidly growing healthcare technology company offering...