Source description
About the role
As a Data Engineer III at Expedia Group, you will have the opportunity to leverage your technical expertise to design solutions that enrich the Data and Intelligent service Metric Enablement platform with new features and functionality to run the business. This role goes beyond traditional data engineering and you will be designing, building, deploying, and operating data pipelines, embeddings workflows to power Agentic AI applications in production. You will be expected to own architecture decisions, drive AI platform evolution, and ensure enterprise-grade reliability, governance, and scalability. Additionally, you will work alongside junior developers as a coach/mentor and collaborate with Sr. Devs on various tech teams to come up with solutions. Key Responsibilities: - Design and develop scalable cloud-native solutions that are scalable, responsive & resilient. - Build scalable ingestion pipelines for structured and unstructured data (documents, logs, knowledge bases, transactional data). - Design semantic layers and context-building strategies for LLM consumption. - Architect and build production-ready RAG systems (retrieval pipelines, embeddings, vector indexing, ranking strategies) and work with vector databases and retrieval systems. - Develop embedding pipelines and manage vector databases at scale. - Develop, test, own, and deliver Sprint tasks and help drive the team forward. - Collaborate with teams and individuals to complete your team assignment on time, with quality. - Be a coach/mentor to junior developers on the team. - Work across multiple layers of the stack as the problem demands. - Have a strong sense of ownership of all technical issues. - Identify risks and issues and drive them to mitigation/resolution as required in the scope of your work. - Prototype ideas, execute and learn from them, and enrich the overall team experience. Qualifications Required: - 6+ years of development experience in an enterprise-level engineering environment with increasing levels of technical expertise. - 4+ years of hands-on backend Data Engineering application development experience with an excellent understanding of products with microservice architecture. - Proven hands-on experience designing, building, and operating data pipelines that enable LLM-based agentic AI systems, including support for embeddings, retrieval layers, and orchestration workflows. - Expert-level SQL and strong Python proficiency (Java is a plus). - Experience with distributed processing frameworks (Spark, Databricks, Flink, etc.). - Experience building data pipelines in cloud-native environments (AWS/GCP/Azure). - Experience building scalable, fault-tolerant, observable systems. - Good knowledge of Data Structures and Algorithm. - Strong understanding of data modeling and semantic layer design. - Understanding of embeddings, chunking strategies, retrieval optimization, and re-ranking. Please note that this role is only available in the following locations: Gurgaon or Bangalore, in alignment with our flexible work model which requires employees to be in-office at least three days a week. Relocation assistance will be considered for candidates relocating to these locations for this role. Expedia Group is committed to fostering an inclusive work environment where diversity is celebrated and all individuals are valued. Join us in shaping the future of travel and be a part of our innovative and diverse community. As a Data Engineer III at Expedia Group, you will have the opportunity to leverage your technical expertise to design solutions that enrich the Data and Intelligent service Metric Enablement platform with new features and functionality to run the business. This role goes beyond traditional data engineering and you will be designing, building, deploying, and operating data pipelines, embeddings workflows to power Agentic AI applications in production. You will be expected to own architecture decisions, drive AI platform evolution, and ensure enterprise-grade reliability, governance, and scalability. Additionally, you will work alongside junior developers as a coach/mentor and collaborate with Sr. Devs on various tech teams to come up with solutions. Key Responsibilities: - Design and develop scalable cloud-native solutions that are scalable, responsive & resilient. - Build scalable ingestion pipelines for structured and unstructured data (documents, logs, knowledge bases, transactional data). - Design semantic layers and context-building strategies for LLM consumption. - Architect and build production-ready RAG systems (retrieval pipelines, embeddings, vector indexing, ranking strategies) and work with vector databases and retrieval systems. - Develop embedding pipelines and manage vector databases at scale. - Develop, test, own, and deliver Sprint tasks and help drive the team forward. - Collaborate with teams and individuals to complete your team assignment on time, with quality. - Be a coach/mentor to junior
More at Expedia Group