Padmi
Datalab logo
Datalab

document AI · OCR models

Senior Software Engineer

New York · Onsite$225k–$300k/yrPosted 14 days ago
Software engineeringSeniorFull Time
Apply at Datalab

Opens the source posting on jobs.ashbyhq.com

Source description

About the role

View original

Salary range: $225k - $300k | Equity: 0.15% - 0.35% | In-Person: NYC

About Datalab

Datalab is building the core infrastructure for how enterprises process and understand documents at scale. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal.

Our models - Chandra, Surya, Marker, and Lift - have significant adoption, with 60,000+ GitHub stars and broad developer mindshare.

We’re backed by founding members of OpenAI, FAIR, and Hugging Face. We move fast, ship often, and we're hiring builders who do the same.

Role Overview

We’re looking for a fullstack engineer who wants to build the interfaces, tools, and infrastructure that help developers and enterprises use our models. You’ll work across the stack to shape how people interact with OCR, extraction, and document-understanding systems. That includes building core inference workflows, creating intuitive UI for complex parsing tasks, and improving the developer experience across our open-source repos and API.

This is a high-ownership role that blends engineering, product thinking, and community engagement. You will work closely with the founders and the rest of the team to ship features, improve performance, and make our technology accessible to a global community of builders.

As a small and fast-moving team, roles are fluid. You should enjoy working across backend, frontend, performance, and user-facing surfaces. Your work will directly influence how teams evaluate and deploy our models.

DAY TO DAY, YOU WILL:

  • Ship features to our open source repos, API, and internal tooling.

  • Design and build frontend features that make document parsing more interactive and understandable.

  • Optimize inference performance and improve the reliability of our frameworks.

  • Engage with the community on Github and Discord by collecting feedback, debugging issues, and identifying opportunities for improvement.

  • Support customer implementations with technical guidance and troubleshooting when needed.

IDEAL CANDIDATE

You thrive at the intersection of engineering, product, and user experience. You like working close to real users and making technical systems feel simple and intuitive. You operate with autonomy and ownership and enjoy moving quickly in an environment where you can directly influence outcomes.

We’re eager to work with someone who has:

  • 5+ years of fullstack development experience building APIs and/or developer-focused products.

  • Experience building in early-stage startup environments.

  • Shipped and maintained production systems serving high-volume traffic.

  • Experience building with LLMs and agentic workflows.

Bonus points

  • Experience with Python ML frameworks (PyTorch, Transformers, etc.) or interest in training models.

  • Familiarity with document processing, computer vision, or OCR systems.

  • Maintained or contributed to open-source projects with active communities

  • Experience writing technical content or demos to showcase your work.

  • INTERVIEW PROCESS

  1. 30-minute video call to evaluate fit

  2. 45-minute second screening to go deeper on the role

  3. Paid take-home project (~3 hours, $300) - and yes, we actually do pay!

  4. Culture fit interviews with the team

  • At this stage of the company, every interview is somewhat custom, so these phases may be rearranged slightly.

  • WE CAN’T WAIT TO HEAR FROM YOU!

  • Apply here with your resume and references to past work to be considered.

More at Datalab

Related open roles

View all roles