Source description
About the role
3+ years of experience in tech/product roles with strong execution ownership
Direct experience building and running eval frameworks for AI products — golden datasets, rubrics, human-in-the-loop validation, and regression tracking across releases.
Technical fluency in LLMs, RAG, prompt engineering, embeddings, and agentic patterns — enough to have substantive conversations with engineers and know when a product problem is a model problem.
Strong track record of building and shipping (0→1 projects, internal tools, automations, or similar)
High ownership mindset—you don’t wait for instructions, you figure things out
Proven ability to operate in ambiguous, fast-changing environments
Excellent at structuring chaos into actionable plans
Strong stakeholder management and cross-functional coordination skills
Ability to drive outcomes without direct authority
Experience working closely with founders, senior leadership, or lean teams is a strong signal
Experience in startups or high-growth environments, operating with limited structure and wearing multiple hats is a strong plus
A self-learner who can quickly pick up new domains, tools, and contexts
More at Caseware