Data Engineer - Python Developer
SAP
- Location
- Gurgaon, IN, 122002
- Work model
- On-Site
- Level
- Senior
Skills
About this role
We help the world run better At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. The work is challenging – but it matters. You'll find a place where you can be yourself, prioritize your wellbeing, and truly belong. What's in it for you? Constant learning, skill growth, great benefits, and a team that wants you to grow and succeed. Data Onboarding & Platform Business Enablement, Enterprise Knowledge Graph What you'll build: The SAP Enterprise Knowledge Graph (EKG) integrates SAP's reference data — business data such as products, processes, and business objects, as well as system data such as ERP transaction codes — into a unified semantic layer. To create and maintain the EKG, we develop an internal platform that exposes the knowledge graph via data access APIs and AI-powered services, and powers a neuro-symbolic AI application that enables intelligent search, reasoning, and knowledge exploration across SAP's enterprise data. As a consultant focused on data onboarding and platform orchestration in the Cross Solution Adoption & Content team, you will join a structured 6–9 month onboarding phase to deeply learn about the EKG, AI application, and the platform — and then transition into building a dedicated consulting practice around EKG data onboarding and service provisioning: - Immerse yourself in the platform during a structured onboarding period: learn the data architecture, ontology standards, ingestion patterns, and platform capabilities hands-on alongside the core engineering team. - You will onboard new datasets that are relevant for the Customer Value Group to grow the EKG into an actionable “Engagement Graph” for CVG. - Own the end-to-end onboarding of new data sources into the EKG — from requirements gathering and data modeling to pipeline implementation, validation, and production rollout. - Design and build scalable, automated data ingestion and transformation pipelines in Python, running on Kubernetes/Kyma, that reliably integrate diverse SAP data sources into the knowledge graph. - Orchestrate complex, multi-step data workflows across the platform — coordinating dependencies, monitoring pipeline health, and ensuring data quality and consistency at scale. - Act as a technical consultant to internal SAP teams seeking to contribute their data to the EKG: guide stakeholders through onboarding processes, translate business data requirements into semantic models, and build repeatable patterns that others can follow. - Develop reusable frameworks, tooling, and playbooks that lower the barrier for future data onboarding — laying the foundation for a scalable, self-service onboarding capability. - Build up a consulting business unit around EKG data onboarding and customer specific knowledge graphs: define service models, engagement patterns, delivery standards, and grow the practice. What you bring: - 5+ years of professional experience in software engineering or data engineering, with a strong focus on data pipelines, or data platform integration. - Previous consulting expertise with completed projects - Strong Python programming skills and hands-on experience building production-grade ETL/ELT pipelines and workflow. - Experience with Kubernetes-based cloud-native development and deployment is a plus. - Interest in and aptitude for knowledge graph or semantic technologies (RDF, OWL, SPARQL) — prior experience is a plus but not required. - Modeling skills with ontologies or UML object models - A consultant mindset: you are comfortable engaging stakeholders, scoping requirements, structuring solutions, and communicating technical concepts to non-technical audiences. - Excellent English communication skills — written and spoken — with the ability to present, document, and persuade across organizational