- Experience
- 5–7 yrs
- Salary
- —
- Openings
- 1
- Posted
- 8 گھنٹے قبل
- Work mode
- In office
- Education
- Any graduate
- Eligibility
- Any Graduate, B.Tech / B.E. in any specialization
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Company Overview
Valuelabs thrives on technology, innovation, and problem-solving, specializing in digital enablement, software product creation, and data technologies. Partnering with over 300 clients worldwide and operating from 26 offices, they use a unified engagement method. Originating from a charitable initiative focused on free educational software for students, their team has expanded from 3 to over 7,500 members while preserving the spirit of service.
Role Summary and Responsibilities
- Design and sustain scalable, optimized ETL pipelines utilizing PySpark on the Cloudera Data Platform, ensuring data accuracy and integrity.
- Manage data ingestion from diverse sources such as relational databases, APIs, and file systems into data lakes or warehouses on CDP.
- Process, cleanse, and transform large data volumes using PySpark to deliver useful formats for analysis aligned with business needs.
- Perform performance enhancement and resource optimization of PySpark code and Cloudera components to reduce ETL execution times.
- Implement and monitor data quality controls to maintain reliable and accurate data within pipelines.
- Automate workflow orchestration employing tools like Apache Oozie or Airflow within the Cloudera ecosystem.
- Oversee ongoing pipeline performance, troubleshoot, and conduct routine platform maintenance.
- Collaborate cross-functionally with data engineers, analysts, product managers, and stakeholders to meet data-driven objectives.
- Thoroughly document data engineering workflows, codebases, and pipeline configurations.
Required Skills and Experience
- At least 5 to 7 years of hands-on experience developing data visualizations, particularly with Power BI and SQL Server Reporting Services (SSRS).
- Strong proficiency in SQL and analytical problem-solving.
- Experience in the banking sector is advantageous.
- Familiarity with Agile development methodologies is a plus.
- Technical capabilities include database design, data modeling, mining techniques, and segmentations.
- Proven track record in visualizing analytical insights through Tableau and Power BI.
- Experience handling large datasets using Hadoop and SQL platforms.
- ETL development skills and knowledge of PySpark preferred.
- Ability to spot process improvements, suggest system changes, and implement governance frameworks.
- Effective in independently learning and contributing within teams, plus critically evaluating stakeholder requirements.
Eligibility and Education
Individuals holding either any graduate degree or a B.Tech/B.E. in any discipline are eligible to apply.
Minimum education
Bachelor's Degree
Skills
Tools & software
Tableau
required
PySpark
required
How they work
Teamwork & Collaboration
Problem Solving
Independence
Learning Agility