Source description
About the role
Location : Hyderabad / Bengaluru - India Duration : 6 Months + Mid/Senior Data Engineer AI Data Pipelines (HYD/BLR) Were hiring a strong Data Engineer to turn messy customer support data into clean, AIready intelligence. Youll build highscale pipelines, vector DBs, and reliable data systems that power our nextgen AI agents. Role Highlights - Clean, process & structure huge volumes of unstructured support data (Salesforce, Sprinklr, Bliss, JIRA). - Build AI memory systems using vector databases (Pinecone, Milvus, Weaviate, pgvector). - Create a central metrics hub for analytics & AI insights. - Own highreliability pipelines with monitoring & zerodowntime standards. - Work closely with product/ops and push for better data logging at the source. Tech Stack - Python, PySpark - Spark, Kafka, Flink, Hadoop, Hudi, Presto, Pinot - Vector DBs (Pinecone/Milvus/Weaviate/pgvector) - Cloud data warehouses - AI tools (Claude, Codex, etc.) for automation & parsing Ideal Candidate - 5+ years in data engineering with big data systems - Solid Python + PySpark - Experience building AIready data layers - Proactive, independent, businessfocused - Comfortable working in ambiguity and driving clarity - Honest communicator who flags risks early Anam Chaudhry 847-201-6116 Location : Hyderabad / Bengaluru - India Duration : 6 Months + Mid/Senior Data Engineer AI Data Pipelines (HYD/BLR) Were hiring a strong Data Engineer to turn messy customer support data into clean, AIready intelligence. Youll build highscale pipelines, vector DBs, and reliable data systems that power our nextgen AI agents. Role Highlights - Clean, process & structure huge volumes of unstructured support data (Salesforce, Sprinklr, Bliss, JIRA). - Build AI memory systems using vector databases (Pinecone, Milvus, Weaviate, pgvector). - Create a central metrics hub for analytics & AI insights. - Own highreliability pipelines with monitoring & zerodowntime standards. - Work closely with product/ops and push for better data logging at the source. Tech Stack - Python, PySpark - Spark, Kafka, Flink, Hadoop, Hudi, Presto, Pinot - Vector DBs (Pinecone/Milvus/Weaviate/pgvector) - Cloud data warehouses - AI tools (Claude, Codex, etc.) for automation & parsing Ideal Candidate - 5+ years in data engineering with big data systems - Solid Python + PySpark - Experience building AIready data layers - Proactive, independent, businessfocused - Comfortable working in ambiguity and driving clarity - Honest communicator who flags risks early Anam Chaudhry 847-201-6116
More at PEGASUS KNOWLEDGE SOLUTIONS INC