About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're hiring our founding GTM - someone who can run the full cycle of sales; sourcing leads, managing the sales process, and closing deals, all while building the playbook that future hires will run on. Datalab makes document AI infrastructure that powers extraction at scale. We're at 8-figure revenue, and have grown revenue >5x YoY, with a team of 7. Anthropic uses Datalab. So do hundreds of other companies across FAANG, frontier AI labs, financial services, insurance, logistics, healthcare, and government. Our open-source projects (Marker, Surya, Chandra) have 60k+ stars and wide community adoption. Sales today is founder-led. The goal of this role is to build a real sales motion on top of that foundation - outbound, ICP definition, enterprise process, and the playbook itself. You won’t be selling alone - engineers and the founder are heavily involved in the sales process, and the whole company pitches in to unblock deals and support customers. This is a high-ownership, high-ambiguity role. We have standard pricing in some areas and open questions in others. We have strong signals about who our best customers are, but ICP isn't fully defined. You'll work alongside the founder and our GTM team to figure those th
Jobs in United States
Data Scientist in United States
2,696 active opportunities · Updated October 2026
Showing
15 jobs
Explore current data scientist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
$250K – $350K/yr
Salary range - $250k - $350k | Equity - up to 0.5% | In-person NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for a Research Engineer to own problems end to end across our models, inference service, and product. You won't just train a model and hand it off. You'll take it from training through benchmarking, into our inference stack, and work with the team to integrate it into our products. We're a small team that has shipped the current state of the art OCR model, Chandra. Our models collectively have 70k+ Github stars. Our tools are used internally at frontier AI labs like Anthropic, and Fortune 500 enterprises like Siemens. Our team focuses on training small, efficient models that outperform much larger LLMs on domain-specific tasks (like OCR, structured extraction, tables). We move fast, prioritize practical results, and build tools that are open, reproducible, and built to last. You'll test hypotheses quickly, iterate on results, and balance experimental rigor with shipping to customers. Day to day: A typical project might look like: identify a gap in extraction quality on long documents, train and benchmark a new model, optimize it for inference, and work with the team to ship it to users. Concretely: Train and evaluate models: Train task-
$300K – $350K/yr
Salary range: $300k - $350k | Equity: 0.4% - 0.6% | In-Person: NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for an engineering lead to guide our team while staying hands-on in the code. You'll set the technical direction and standards for how we build the interfaces, tools, and infrastructure behind our OCR, extraction, and document-understanding systems. This includes everything from optimizing agent loops and interfaces to helping to speed up inference. This is a player-coach role. You'll manage and grow a team of three engineers, own engineering delivery and quality, and spend a large share of your time writing code - focused on architecture, infrastructure, and the hard problems rather than routine feature work. You’ll partner closely with the research team to define the handoff between experimentation and production. As a small and fast-moving team, roles are fluid and ownership is high. You'll work directly with the founder to set priorities, ship features, and make our technology accessible to a global community of builders. Day to day, you will: Manage and grow a team of three engineers - 1:1s, prioritization, feedback, and hiring as we scale. Own engineering delivery, quality, and technical standards across code, testing, infrastruct
From $225K/yr
Founding Engineer, Open Source Salary range — $225k – $300k | Equity — .15%-.35% | In-person NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, chandra, surya, marker, and lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for a Founding Engineer to develop and evangelize our open source repos. This includes chandra, surya, marker, pdftext, and lift, which collectively have over 70k Github stars. It also includes new tools we have yet to build and launch. As you work on our repos, you’ll also become the credible person to evangelize them. You’ll turn your own work into demos, benchmarks, tutorials, and launches. You’ll also support other launches across Datalab, especially when they touch open source components like the SDK. Our projects have real reach: 70k+ GitHub stars and users everywhere from frontier AI labs to Fortune 500s. Your job is to turn that reach into a thriving, engaged developer community through content, code, and showing up where developers already are. If you're the kind of engineer who’s energized by both building things and helping other people build, this is the role for you. Day to day: Own our open source repos and SDK: Drive chandra, surya, marker, lift, and our Python SDK forward as a hands-on contributor. Ship new features and improvements that the market cares about, help plan new versions, and ow
$225K – $300K/yr
Salary range: $225k - $300k | Equity: 0.15% - 0.35% | In-Person: NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We’re looking for a fullstack engineer who wants to build the interfaces, tools, and infrastructure that help developers and enterprises use our models. You’ll work across the stack to shape how people interact with OCR, extraction, and document-understanding systems. That includes building core inference workflows, creating intuitive UI for complex parsing tasks, and improving the developer experience across our open-source repos and API. This is a high-ownership role that blends engineering, product thinking, and community engagement. You will work closely with the founders and the rest of the team to ship features, improve performance, and make our technology accessible to a global community of builders. As a small and fast-moving team, roles are fluid. You should enjoy working across backend, frontend, performance, and user-facing surfaces. Your work will directly influence how teams evaluate and deploy our models. Day to day, you will: Ship features to our open source repos, API, and internal tooling. Design and build frontend features that make document parsing more interactive and understandable. Optimize inference performance and improve th
From $160K/yr
Base — $160k – $180k | OTE — ~$230k – $255k | Equity — 0.3% | In-person NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We're at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, chandra, surya, marker, and lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for a Founding Customer Success Manager to own the relationship with our enterprise customers after they sign. You'll be accountable for adoption, retention, and growth — making sure customers get real value from our models, stay for the long run, and expand their usage over time. You'll be the face of Datalab for every account you own: their trusted advisor, their first call when something matters, and the person keeping their goals moving forward. You will be the commercial and relationship manager who owns the account strategy after a deal has been signed. You'll orchestrate the right people internally so the customer always feels progress. You'll know the product and the customer's architecture well enough to lead most conversations yourself, and you'll pull in Engineering when an account needs it. We also have a long tail of self-serve customers using our API. Many are strong candidates to expand into enterprise contracts, and you'll own that self-serve → enterprise motion end to end — from spotting high-potential accounts to closing the upgraded contract. This role is ideal for someone who thrives at the intersection of c
From $137K/yr
About Mixpanel Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how people use the products they build. By combining powerful analytics with AI that knows your business, Mixpanel helps teams see what’s working, diagnose what’s not, and decide what to build next. Learn more at mixpanel.com . About the Delivery Engineer Team As a member of the Delivery Engineer team, you will own the post-sales onboarding and data health for customers. Your goal will be to drive data trust, help embed Mixpanel into our customers’ data stack, and deliver on emerging AI integrations — delivering on the full value of Mixpanel as a self-serve analytics platform. You will work and consult with GTM team members and a diverse array of customers to successfully roll out product analytics to their organization and execute on technical projects and services that delight our customers. About the Role As a Delivery Engineer III, you will be on the front lines with our clients as they integrate Mixpanel into their core product development processes. You will lay the foundation for customers to adopt product analytics and get value from Mixpanel by delivering a world-class, on-time, and value-oriented onboarding experience. With your comprehensive project management knowledge, consultative approach, expertise in the analytics space, and technical knowledge of the modern data stack, you will lead our clients' first experience with Mixpanel as they incorporate product analytics into their data ecosystem as a foundational element. With your deep technical breadth and expertise in the analytics space, you’ll be the expert consultant on all things data and ecosystem. As Mixpanel and our customers continue to iterate on agentic integrations, you’ll own ensuring customers are able to build, assemble context, and integrate AI workflows with Mixpanel. Responsibilities Own onboarding and data health for our strategic and high-value
About the Team OpenAI’s People team is committed to hiring, engaging, and supporting world-class talent to help safely build and deploy universally beneficial Artificial General Intelligence (AGI). The People Operations team is responsible for the systems, data, and operational rigor that underpin every stage of the employee lifecycle. Our work sits at the intersection of employee experience, regulatory compliance, and business velocity — ensuring that as OpenAI scales rapidly and globally, our foundations remain accurate, auditable, and human-centered. This is a high-trust, high-impact function. We partner closely with Finance, Legal, IT, Security, and the broader People team to support growth in a dynamic environment with significant internal and external scrutiny. Operational excellence here is not about process for process’s sake — it is about building durable infrastructure that enables fairness, transparency, and responsible scale. About the Role As Director of People Operations, you will own the core operational infrastructure that enables OpenAI to scale responsibly. This role serves as the operational backbone of the People function, with accountability for data integrity, operational controls, and audit readiness across the systems that support a complex, global workforce, including the operational foundations of recruiting and hiring at scale. You will lead a team responsible for the end-to-end workforce lifecycle, from candidate and hiring operations through onboarding and offboarding, ensuring our people systems, processes, and data are accurate, resilient, and built to scale.In close partnership with Employment Legal and the People Compliance team, you will translate policy and regulatory requirements into scalable, well-controlled operational processes. This role operates with high visibility and close partnership with senior leaders across Finance, Legal, IT, and People, including ownership of SOX-relevant controls and workforce data relied upon by e
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Ramp is building a scaled Private Equity channel that makes Ramp the trusted financial-operations partner for PE firms and their portfolio companies. The Manager, Channel Partner Management — Private Equity will lead an established five-person team of Channel Partner Managers and raise the quality, rigor, and efficiency of an existing motion. The team includes one Upmarket PE CPM, focused on the most strategic and complex sponsor relationships, and four Core / Downmarket PE CPMs, focused on scaled coverage, activation, and portfolio-company pipeline generation. This leader is not being hired to redesign the team or recategorize the program. They will make the current team more effective: clarifying expectations, improving performance, and ensuring each CPM has the tools, coaching, and cross-functional support to execute at a high level. This is a people-leadership and operating-excellence role—not a player-coach role defined by personally owning every senior relationship or stepping into every deal. The manager will diagnose performan
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Emerging Talent has changed drastically. Ramp is building it into one of the most important recruiting functions at the company: a way to meet exceptional people earlier than everyone else and turn that connection into the next generation of Ramp talent. We are looking for a true builder to lead the full Emerging Talent function. Reporting to Ramp’s Head of Talent, you will set the strategy and lead the team responsible for the internship programme, early-career recruiting, campus and community relationships, events, candidate programming and conversion of top interns into full-time employees. You will have the backing of the executive team and a real mandate to build something category-defining. What You'll Do Own the strategy, operating model and results for Ramp’s Emerging Talent function, from early identification through full-time conversion. Build a world-class internship experience, including seasonal programmes, events, programming, manager partnership and intern-to-full-time conversion. Find original ways to identify and attr
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid Protect is a real-time fraud intelligence product built on a unique advantage: Plaid’s network-level visibility across bank accounts, devices, identities, sessions, institutions, applications, and financial behavior. Protect helps customers detect first-party fraud, synthetic identities, account takeovers, and coordinated attacks that are difficult to see from a single application, account, or transaction. Trust Index turns that fraud intelligence into real-time fraud scores and actionable attributes. This team builds the systems that make this intelligence possible: low-latency inference, new data and model integrations, customer-facing APIs and attributes, safe rollouts, and feedback loops. Ti3 expanded Plaid’s fraud graph nearly 10x and, in early testing, detected up to 41% more fraud at the same false-positive rate. Learn more about Ti2 and Ti3 . We are a small, high-agency team working closely with Product, Data Science, and Machine Learning. We value demos over docs, conviction over consensus/alignment, builder schedule over meeting-heavy calendars. We’re scrappy and a talent-dense team that has high agency and high ownership. As a Staff Software Engineer on the Protect Core team, you wi
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. We believe Plaid has the power to be the next-gen Credit Bureau - supporting large scale adoption of cash flow into the credit underwriting process. The Credit Decisioning platform team is responsible for building best-in-class cashflow based insights products that enable lenders to make more holistic lending decisions and empower broader access to Credit products for prospective borrowers. We own the systems and tooling that form the platform to build and serve these insights at huge scale, partnering with our Data partners to release new products yearly. You will be defining the future architecture of Credit insights products and executing against an ambitious product roadmap. You will partner with our Product, Data Science, and Machine Learning team to iterate on and productionize new insights that enable our customers to make more holistic lending decisions. Responsibilities: Leading technical architecture and execution across credit insights products: everything from data fetching and online feature serving for API requests, to offline production pipelines and tooling for model training. Scaling and evolving the architecture through an expected ~100x increase in load from deterministic factors
$106K – $142K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! This role operates on a hybrid schedule requiring two days of in-office collaboration per week. The position must be based in the San Francisco Bay Area. About the Role We're hiring a Software Engineer II within our Fulfillment organization — the backend systems that get the right job to the right Tasker and see it through to completion. You'll join Fulfillment Lifecycle, the team that decides how jobs are matched to Taskers for our partner and marketplace business, increasingly using unstructured data and experimentation to make matching smarter and fulfillment more reliable. The team is part of a company-wide platform modernization effort, breaking a legacy monolith into well-bounded, API-first services. We're hiring for a strong backend engineer who thrives on complex, data-intensive problems, is comfortable with ambiguity, and takes pride in well-tested, observable, production-ready code. What You'll Work On B
About the Team OpenAI’s acquisition of io marks our entry into consumer hardware and our ambition to define the next human–computer interface. Success in hardware requires strong financial stewardship across the full product cost stack—from early design and sourcing decisions through manufacturing, logistics, inventory, returns, and warranty. Hardware Finance works across Product, Supply Chain, Operations, Accounting, Systems/Data, and Finance to connect business decisions to product cost, inventory, cash, COGS, and margin. About the Role We are seeking a Hardware Finance Manager to own an assigned area of hardware COGS and inventory end to end. The initial assignment will depend on business priorities and the successful candidate’s expertise. It may include BOM and product cost, manufacturing variance analysis, inventory planning, logistics, returns and warranty, customer support, or another connected set of hardware-finance responsibilities. This is an individual-contributor role with broad scope. Prior hardware experience and deep, hands-on expertise in at least two relevant domains are required. The person will be expected to operate independently, build reusable processes and analytical workflows, and remain accountable for the analysis, judgment, and recommendations. In this role, you will: Own an assigned area of hardware COGS and inventory end to end. Own forecasting, close, and business variance analysis for the assigned scope. Provide hardware leadership with clear variance explanations, trend analysis, and forward-looking signals that connect business and supplier decisions to inventory, cash, COGS, and margin. Partner with business teams and Finance Platforms to establish the financial data, systems, and dashboards needed to support analysis. Ensure data integrity and governance through clear definitions, ownership, validation checks, controls, and review processes. Improve forecasting, reporting, systems, and finance processes so they remain reliable an
About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role We are seeking a Corporate Recruiter who can independently own full-cycle recruiting across a range of corporate functions. You are a strong partner to hiring managers, comfortable managing multiple searches at once, and able to move from intake through offer with a high degree of autonomy. You bring sound recruiting judgment, a thoughtful candidate experience, and a practical, data-informed approach to finding and closing strong talent. You will partner across corporate teams to understand hiring needs, build effective search strategies, and keep processes moving with clarity and consistency. Responsibilities: Own full-cycle recruiting independently. You’ll manage the end-to-end process—from intake and job kickoff through sourcing, screening, interview coordination, candidate management, and close—across a varied portfolio of roles. Build strong partnerships with hiring managers. You’ll set clear expectations, bring market context and recruiting expertise to the table, and keep searches moving with proactive communication and thoughtful recommendations. Build high-quality talent pipelines. You’ll use a mix of sourcing strategies, netw
Higher-paying openings
Jobs with higher listed pay
Staff Data Scientist, Forecasting
Pinterest · San Francisco, US; Remote, US
From $2M/yr
Data Scientist, Safety
OpenAI · San Francisco, California, United States
$230K – $325K/yr
Data Scientist, Core Experimentation
OpenAI · Seattle, Washington, United States
$293K – $325K/yr
Principal Data Scientist - Safety
Roblox · San Mateo, CA, United States
From $321.2K/yr
Senior / Principal Data Scientist - Discovery
Roblox · San Mateo, CA, United States
From $307.4K/yr
Senior Data Scientist, Consumer Apps
Roblox · San Mateo, CA, United States
From $263.7K/yr
Related career options
Similar roles with stronger pay
Demand 46/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 43/100 · 8 jobs
$300K – $300K/yr
Salary →Demand 42/100 · 7 jobs
$300K – $300K/yr
Salary →Demand 38/100 · 30 jobs
$278.9K – $278.9K/yr
Salary →Demand 30/100 · 11 jobs
$255.7K – $255.7K/yr
Salary →Other cities to consider
More places hiring for this role
Get new data scientist jobs in United States by email
Daily job updates · Unsubscribe anytime