Skip to main content Skip to footer

Data Engineer

Data Eng, Mgmt & Governance Team Lead/Consultant | Full time | Experience: 5-10 years
Job No. ATCI-5747696-S2067150 | Hyderabad | Required Skill: Data Engineering
Apply for this job
Project Role : Data Engineer
Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems.
Must have skills : Data Engineering
Good to have skills : NA
Minimum 5 year(s) of experience is required
Educational Qualification : 15 years full time education

Data Engineer – Azure Data Platform & Graph
Experience: 5–8 years
Location: India
Role Summary
A hands-on, build-and-run role: writing and operating the ETL/ELT pipelines, lakehouse tables, and graph data models that bring data from many enterprise source systems into a governed, analytics- and AI-ready platform. Day to day, this means writing transformation code, debugging failed pipeline runs, and modeling connected data directly in both property-graph (Cosmos DB Gremlin) and RDF/semantic-graph (SPARQL, Turtle) stores — not just designing on a whiteboard.
Key Responsibilities
Build, schedule, and maintain ETL/ELT pipelines — extracting from source systems, transforming with PySpark/Spark SQL, and loading into bronze/silver/gold lakehouse tables — using Fabric Data Factory pipelines and Dataflow Gen2 (or ADF/equivalent)
Write and maintain transformation logic for incremental loads, change data capture (CDC), deduplication, slowly changing dimensions (SCD), and schema evolution
Hands-on troubleshooting of failed or delayed pipeline runs — diagnosing root cause, fixing transformation bugs, and rebuilding/backfilling data as needed
Model and query graph data (vertices, edges, properties) directly in Azure Cosmos DB for Apache Gremlin, writing and optimizing Gremlin traversals for relationship-heavy and knowledge-graph use cases
Write SPARQL queries and author Turtle (.ttl) files to populate and query RDF-based knowledge graphs, working with a triplestore such as Graphwise GraphDB
Build and publish data products aligned to data mesh principles — clear ownership, documented contracts, and discoverability for consuming teams
Register, tag, and maintain lineage for data assets in an enterprise data catalog to support governed, self-service discovery
Write data quality checks, schema validation rules, and pipeline monitoring/alerting across the ingestion-to-consumption flow
Tune pipeline performance, partitioning strategy, and cost/throughput trade-offs across relational, NoSQL, and graph stores
Collaborate with data architects, analytics engineers, and AI/ML teams to expose curated, trustworthy data for downstream consumption (BI, RAG/agentic AI, ML)
Required Skills & Experience
5+ years hands-on building and operating ETL/ELT pipelines in production, across relational, NoSQL, and graph data stores
Strong Python for pipeline and transformation development (PySpark, pandas, or equivalent) — comfortable writing and debugging transformation code daily
Hands-on experience building pipelines on a modern Azure data platform (e.g., Microsoft Fabric, or Azure Synapse/Databricks-equivalent) — Data Factory/Dataflow Gen2, Spark notebooks, Delta Lake tables
Practical experience implementing medallion (bronze/silver/gold) architecture, including incremental loads, CDC, deduplication, SCD, and schema evolution handling
Hands-on with Azure Cosmos DB, including the Gremlin (graph) API — writing vertex/edge data models, partition key design, and Gremlin queries/traversals
Working knowledge of RDF/semantic graph technologies — writing SPARQL queries and authoring Turtle (.ttl) files, using a triplestore such as Graphwise GraphDB (or equivalent, e.g., Amazon Neptune, Stardog)
Strong SQL — writing and optimizing complex transformation and analytical queries
Experience building integrations across multiple heterogeneous source systems (databases, SaaS applications, APIs, files)
Working knowledge of data catalog / metadata management practices (lineage, classification, glossary)
Understanding of data mesh concepts — data as a product, domain ownership, and self-serve data platforms
Preferred
Exposure to the broader semantic web stack — RDF/RDFS, OWL, SHACL, SKOS — for ontology-driven knowledge graph work
Experience with pipeline orchestration tools (e.g., Apache Airflow, or Fabric's Airflow-based orchestration)
Exposure to enterprise data mesh implementations, including federated governance models
Familiarity with Microsoft Purview (or equivalent) for enterprise-wide metadata, data map, and unified catalog capabilities
Experience with real-time/streaming ingestion (e.g., Eventstream, KQL, or equivalent)
Exposure to data preparation for vector stores / RAG-style AI consumption
Relevant platform certification (e.g., Microsoft Certified: Fabric Data Engineer Associate)
Production experience across multiple major cloud platforms (AWS, Azure, and GCP)
Education
Bachelor's/Master's in Computer Science, Data Engineering, or related field
15 years full time education

Hyderabad

Equal Employment Opportunity Statement

All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by federal, state, or local law.

Please read Accenture’s Recruiting and Hiring Statement for more information on how we process your data during the Recruiting and Hiring process.

We work with one shared purpose: to deliver on the promise of technology and human ingenuity. Every day, more than 775,000 of us help our stakeholders continuously reinvent. Together, we drive positive change and deliver value to our clients, partners, shareholders, communities, and each other.

We believe that delivering value requires innovation, and innovation thrives in an inclusive and diverse environment. We actively foster a workplace free from bias, where everyone feels a sense of belonging and is respected and empowered to do their best work.

At Accenture, we see well-being holistically, supporting our people’s physical, mental, and financial health. We also provide opportunities to keep skills relevant through certifications, learning, and diverse work experiences. We’re proud to be consistently recognized as one of the World’s Best Workplaces™.

Join Accenture to work at the heart of change. Visit us at www.accenture.com.

We have been alerted to the existence of fraudulent messages asking job seekers to set up payment to cover various costs associated with establishing employment at Accenture. No one is ever required to pay for employment at Accenture. If you are contacted by someone asking for payment, please do not respond, and contact us at india.fc.check@accenture.com immediately.

Discover where this job fits at Accenture

Artificial Intelligence

AI and data science jobs: Uncover new possibilities

Unlock the power of AI and data to reinvent all facets of business–responsibly.

Learn more