Source description
About the role
Location: Bangalore Job Summary The Enterprise Data Architect is the technical system owner for Avance Clinical's enterprise data lake. This role oversees the design, implementation, and maintenance of data pipelines and integrations on the data lake. Core Responsibilities Act as the technical owner and subject matter expert for the enterprise data lake and associate pipelines. Design, develop, and maintain ETL processes for ingestion, transformation, and quality control of multi-source data (e.g., CTMS, EDC and other business systems). Administer and monitor the BigQuery environment, including performance optimization, access management, and compliance with data governance policies. Govern the implementation of the FHIR data model and enterprise data standards. Where relevant, support mapping of FHIR to other data standards as relevant (e.g. CDISC, OMOP) Collaborate with solution architects and system owners (EDC, Salesforce, Veeva, etc.) to design and implement new data integrations and automation workflows. Act as data steward for data lake, collaborating with business data owners to define standards, resolve data quality issues, and enforce agreed governance policies within the data pipelines and warehouse. Maintain roadmap of use cases in partnership with business stakeholders. Maintain documentation and metadata for all datasets, pipelines, and schema mappings. Support reporting and analytics teams through the creation of validated, standardized data views and sources. Lead root cause analysis and resolution of data quality or pipeline issues. Line Management of Enterprise Insights Analyst. Qualifications and Skills Bachelor's degree in Computer Science, Information Systems, or related discipline. 8-10+ years of experience in data science, data engineering or data management Strong experience in designing, mapping or implementing FHIR data model for healthcare interoperability Demonstrated experience designing, operating, or technically owning enterprise data platforms and analytics, preferably in a clinical research, biotechnology/pharma or healthcare environment Strong experience in managing clinical data, from source to submission, in clinical trials or biotech environment Strong proficiency in SQL, including data modelling, transformation and query optimization, with experience with BigQuery or equivalent cloud data warehouses Strong proficiency with Python or similar scripting languages for data transformation and automation in Jupyter Notebooks Demonstrated experience in ETL toolsets or orchestration frameworks. Working knowledge of data governance, metadata management, lineage, security, and compliance requirements (e.g., GxP, 21 CFR Part 11). Ability to negotiate with senior stakeholders to achieve enterprise-wide data definitions/agreements and workflow transformation Highly collaborative, with strong communication and documentation skills Established ability for effective, cross-functional teamwork with strong ability to communicate technical concepts to business stakeholders