ActiveJobs
Merck Careers

Senior Specialist, Process Data Engineering

Merck Careers · IND - Telangana - Hyderabad (Hitec City Raidurg)

Full-timeOn-sitePosted 18 September 2026
Apply on Company Site →

Job description

Job Description Senior Specialist, Process Data Engineering The Opportunity Join a global biopharma company with a 130-year legacy and mission to achieve new milestones in healthcare. Be part of technology driven, data-led organization supporting a diversified portfolio of medicines, vaccines and animal health products. Work alongside passionate teams that use data, analytics and insights to drive decisions, and tackle some of the world’s greatest health threats. Our Technology Centers are globally distributed hubs that enable our digital transformation and business outcomes across IT. They bring together diverse teams to collaborate, share best practices, and deliver solutions that save and improve lives. This role is based at our Hyderabad Tech Center and follows a hybrid working model (3 days per week in the office and 2 days remote). Candidates are expected to reside within a commuting distance of the Hyderabad office. Role Overview As a Process Data Engineer, you are the technical architect of our global Process Intelligence and AI initiatives. Your primary focus is constructing robust data pipelines that ingest transactional data from diverse enterprise systems (SAP, Veeva, Appian, ServiceNow) into an advanced, object-centric process data model. Operating within an Agile framework, you will manage the end-to-end Software Development Life Cycle (SDLC) of these models, leveraging them to build predictive machine learning models and deploy process-aware AI agents that enable autonomous execution management across the enterprise. What will you do in this role Data Pipeline Engineering Architect and manage highly scalable ETL/ELT pipelines to extract raw event logs and transactional data from core organisation systems (SAP, Veeva Vault/CRM, Appian, Oracle). Process Data Architecture Design and implement Object-Centric Data Models. Define custom entity types and relationships rather than relying on flat, single-case event logs to build a dynamic Process Digital Twin. Agile Delivery & SDLC Operate within an Agile Scrum methodology. Translate business requirements into technical user stories, participate in sprint rituals, and manage strict SDLC code promotion (Dev, QA, Prod) for all data models and AI applications. AI & Agentic Orchestration Develop and deploy process-aware AI agents and automated action flows. Integrate Large Language Models (LLMs) to create "Process Copilots" that allow business users to query operational data using natural language. Predictive Process Monitoring Implement Machine Learning models (using Python/R) on top of the process data graph to predict process delays, compliance risks, and operational bottlenecks before they occur. IT Governance & AI Ethics Ensure all pipelines, AI integrations, and deployments strictly comply with our Information Technology policies, GxP standards, and corporate AI ethics guidelines using robust version control and CI/CD pipelines. What should you have Bachelor’s degree in information technology, Computer Science or quantitative field (Statistics, Econometrics, Computer Science, Mathematics, Engineering). 7 to 11 years of experience in Data Engineering. Required Skills: Database & Cloud Engineering Expert-level SQL and data modeling. Deep understanding of relational databases, data warehousing, and handling multi-terabyte datasets. Process Platform Expertise Extensive hands-on experience with leading Enterprise Process Mining platforms, specifically dealing with multi-object graphing and execution management. Agile & SDLC Methodologies Proven experience working in Agile/Scrum environments. Proficiency with Jira/Confluence for sprint tracking, and Git/Bitbucket for version control and managing release branches. Applied AI & Machine Learning Experience building predictive models applied to time-series or event log data. Proficiency in Python (pandas, scikit-learn, TensorFlow/PyTorch) and LLM API integrations. Source System Knowledge Strong understanding of underlying database schemas and APIs for Veeva (Vault/CRM), Appian, and SAP. Primary Skills (Must have): SQL Data Modeling Process Mining Spark SQL Agile SDLC Git Python Agentic Orchestration Process Data Architecture Process Modeling graph Secondary Skills (Nice to have): SDLC Architecture & Agile Leadership Proven ability to design robust CI/CD pipelines specifically for process data models and AI agents. Experience participating in backlog grooming, sprint planning, and cross-functional Agile delivery. AI-Driven Digital Twin Architecture Ability to structure the object-centric data model specifically so it can serve as a highly accurate, real-time knowledge base (RAG foundation) for enterprise AI agents. Business-to-Technical Data Mapping Ability to translate functional business workflows (e.g., Order-to-Cash, IT Service Management, Clinical Operations) directly into backend table schemas. SAP ECC & SAP HANA Mastery Deep architectural understanding of core SAP data structures (e.g., VBAK/VBAP, EKKO/EKPO, BKPF/BSEG) and how they evolve between ECC and S/4HANA environments. ServiceNow & Veeva Architecture Thorough understanding of the ServiceNow Task table architecture and Veeva Vault/CRM object models. Who we are: We are known as well-known org Inc., Rahway, New Jersey, USA in the United States and Canada and MSD everywhere else. For more than a century, bringing forward medicines and vaccines for many of the world's most challenging diseases. Today, our company continues to be at the forefront of research to deliver innovative health solutions and advance the prevention and treatment of diseases that threaten people and animals around the world. What we look for: Imagine getting up in the morning for a job as important as helping to save and improve lives around the world. Here, you have that opportunity. You can put your empathy, creativity, digital mastery, or scientific genius to work in collaboration with a diverse group of colleagues who pursue and bring hope to countless people who are battling some of the most challenging diseases of our time. Our team is constantly evolving, so if you are among the intellectually curious, join us—and start making your impact today. Required Skills: Agentic Orchestration, Agile SDLC, Data Architecture, Data Modeling, Git, Process Architecture, Process Mining, Process Modeling, Python (Programming Language), Spark SQL, Structured Query Language (SQL) Preferred Skills: Current Employees apply HERE Current Contingent Workers apply HERE Secondary Language(s) Job Description: #MSDHYDIT Search Firm Representatives Please Read Carefully Merck & Co., Inc., Rahway, NJ, USA, also known as Merck Sharp & Dohme LLC, Rahway, NJ, USA, does not accept unsolicited assistance from search firms for employment opportunities. All CVs / resumes submitted by search firms to any employee at our company without a valid written search agreement in place for this position will be deemed the sole property of our company. No fee will be paid in the event a candidate is hired by our company as a result of an agency referral where no pre-existing agreement is in place. Where agency agreements are in place, introductions are position specific. Please, no phone calls or emails. Employee Status: Regular Relocation: Domestic VISA Sponsorship: No Travel Requirements: No Travel Required Flexible Work Arrangements: Hybrid Shift: Not Indicated Valid Driving License: No Hazardous Material(s): n/a Job Posting End Date: 09/25/2026*A job posting is effective until 11:59:59PM on the day BEFORE the listed job posting end date. Please ensure you apply to a job posting no later than the day BEFORE the job posting end date.

Verified and listed by ActiveJobs. Applications are made directly on Merck Careers's own career page — we never sit in the middle.