Job Overview
Role: Associate – Platform Services / Data Engineer Location: Pune / Gurgaon Experience: 1–2 Years Qualification: Bachelor's / Master's Degree (Computer Science, MIS, Information Technology, or related) Key Skills: ETL, SQL, Python, Data Modeling, Data Warehousing, Analytics-Ready Data Products, Agentic AI-Ready Data Products, Cloud Platforms
Job Description
ZS is hiring for the position of Associate – Platform Services / Data Engineer in Pune and Gurgaon. This role focuses on developing scalable, analytics-ready, and Agentic AI-ready data products on the ZAIDYN platform. Candidates will work on ETL pipelines, SQL-based data transformation, Python scripting, data modeling, data warehousing, KPI development, analytics data products, and cloud technologies. The position is suitable for early-career data engineers who have hands-on development-project experience and want to work at the intersection of data engineering, analytics, cloud platforms, and AI-ready data infrastructure.
Roles and Responsibilities
- ZAIDYN Platform: Configure and develop solutions on the ZAIDYN platform to create scalable data products and pipelines for analytics and AI use cases.
- ETL Development: Build and maintain pipelines that perform data extraction, validation, transformation, loading, and data-quality checks.
- SQL and Data Transformation: Utilize SQL for data transformation, validation, troubleshooting, and performance tuning (SELECT, JOINs, Subqueries, CTEs, Window functions, GROUP BY, aggregations, data validation, query optimization).
- Python for Data Engineering: Apply Python scripting for data processing, automation, transformation logic, and pipeline utilities, focusing on functions, modules, file handling, data structures, exception handling, APIs, and data processing libraries.
- Data Modeling: Develop conceptual, logical, and physical data models to define how information is represented, related, stored, and accessed.
- Data Warehousing: Implement strong knowledge of dimensional modeling, fact tables, dimension tables, star/snowflake schemas, measures, dimensions, surrogate keys, historical data handling, and performance optimization.
- Analytics-Ready Data Products: Deliver reliable, well-structured, and documented analytics-ready data products for BI, reporting, and analytical models.
- Agentic AI-Ready Data Products: Structure data with useful features, metadata, and documentation to enable AI systems and intelligent automation to work with trusted enterprise information.
- Business Rules and Data Quality: Apply business rules to ensure data quality, consistency, and alignment with business definitions (completeness, accuracy, consistency, validity, uniqueness, timeliness).
- KPI Development: Define and compute Key Performance Indicators (KPIs).
- End-to-End Delivery to Technical Design: Translate business requirements into data models, pipelines, KPIs, and technical implementation plans.
- Agile, Version Control and Code Reviews: Work within an Agile framework, use version control, and participate in code reviews.
- Global Team Collaboration: Collaborate effectively within global teams.
- Client and Project Responsibilities: Manage client and project responsibilities.
- Mentoring and Coordination: Engage in mentoring and coordination activities.
Skills and Eligibility Criteria
Educational Background: Bachelor's or Master's degree specializing in Computer Science, MIS, Information Technology, or a related discipline.
Experience: 1–2 years of relevant development experience. Experience working on medium- to large-scale technology solution delivery engagements is preferred. Experience in Pharma or Life Sciences organizations is preferred.
Mandatory Technical Skills:
- Strong experience in ETL development and data-pipeline development.
- Proficiency in SQL for data transformation, validation, and performance tuning.
- Hands-on experience with Python scripting for data processing, automation, and transformation logic.
- Understanding of conceptual, logical, and physical data modeling.
- Strong knowledge of data warehousing (dimensional modeling, fact tables, dimension tables, performance optimization).
- Experience designing and delivering data products for analytics and reporting.
- Exposure to AWS/Azure and big-data or ETL tools (Hadoop, Spark/PySpark, Informatica, Talend, or SSIS).
Competencies:
- Strong analytical, problem-solving, and communication skills.
- Fluency in English.
- Client-first mentality and collaborative approach.