Jobs in United States

Quality Management Supervisor in New York

174 active opportunities · Updated October 2026

Explore current quality management supervisor jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.

H
📍 New York, NY, United States
✓ Quality checkedCompany trend +310%

Become a part of our caring community Most AI engineering jobs are a thin wrapper around a model API. This role is different. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to the source document, and route ambiguous cases to human experts for review. Our users make decisions that impact real healthcare outcomes, so “good enough” is not good enough. Building AI systems that are accurate, reliable, auditable, and scalable is at the core of this role. As a Senior AI Applied Engineer, you will design, build, deploy, and operate production AI systems used at scale within one of the largest health insurers in the United States. You will own solutions end-to-end, from user experience and APIs to model orchestration, evaluation frameworks, infrastructure, and production operations. Why Join Us Build production AI systems where LLMs are in the critical path, not just demos or proofs of concept. Work on extraction, retrieval, agentic workflows, and human-review systems that process real healthcare data at scale. Own projects end-to-end across frontend, backend, AI orchestration, infrastructure, deployment, and operations. Solve challenging problems around accuracy, explainability, traceability, and reliability in regulated environments. Ship quickly in a small, high-impact team that embraces AI-assisted development and rigorous quality standards. Build systems that continuously improve through expert feedback, evaluations, and human-in-the-loop workflows. Key Responsibilities Design, develop, and deploy full-stack AI-powered application

JavaScriptTypeScriptPythonReact
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. We’re looking for an Implementation Engineer to help enterprise customers successfully deploy, configure, and operationalize Observe. This is a hands-on, post-sales technical role focused on delivering strong first outcomes, accelerating time-to-value, and establishing a solid foundation for long-term customer success. Implementation Engineers are deeply technical, customer-facing practitioners who work closely with customer platform, SRE, DevOps, and application teams during onboarding and early adoption. In this role, you’ll translate existing observability architectures (including OpenTelemetry-based pipelines, Splunk, ELK, and other monitoring solutions) into scalable, production-ready implementations on Observe—using best practices while balancing speed, quality, and customer enablement. Implementation Engineers focus on initia

AWSAzureGCPKubernetes
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -100%

$250K – $350K/yr

Quick readStrong listing-quality and freshness signals

Salary range - $250k - $350k | Equity - up to 0.5% | In-person NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for a Research Engineer to own problems end to end across our models, inference service, and product. You won't just train a model and hand it off. You'll take it from training through benchmarking, into our inference stack, and work with the team to integrate it into our products. We're a small team that has shipped the current state of the art OCR model, Chandra. Our models collectively have 70k+ Github stars. Our tools are used internally at frontier AI labs like Anthropic, and Fortune 500 enterprises like Siemens. Our team focuses on training small, efficient models that outperform much larger LLMs on domain-specific tasks (like OCR, structured extraction, tables). We move fast, prioritize practical results, and build tools that are open, reproducible, and built to last. You'll test hypotheses quickly, iterate on results, and balance experimental rigor with shipping to customers. Day to day: A typical project might look like: identify a gap in extraction quality on long documents, train and benchmark a new model, optimize it for inference, and work with the team to ship it to users. Concretely: Train and evaluate models: Train task-

PythonGitAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -100%

$300K – $350K/yr

Quick readStrong listing-quality and freshness signals

Salary range: $300k - $350k | Equity: 0.4% - 0.6% | In-Person: NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, Chandra, Surya, Marker, and Lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for an engineering lead to guide our team while staying hands-on in the code. You'll set the technical direction and standards for how we build the interfaces, tools, and infrastructure behind our OCR, extraction, and document-understanding systems. This includes everything from optimizing agent loops and interfaces to helping to speed up inference. This is a player-coach role. You'll manage and grow a team of three engineers, own engineering delivery and quality, and spend a large share of your time writing code - focused on architecture, infrastructure, and the hard problems rather than routine feature work. You’ll partner closely with the research team to define the handoff between experimentation and production. As a small and fast-moving team, roles are fluid and ownership is high. You'll work directly with the founder to set priorities, ship features, and make our technology accessible to a global community of builders. Day to day, you will: Manage and grow a team of three engineers - 1:1s, prioritization, feedback, and hiring as we scale. Own engineering delivery, quality, and technical standards across code, testing, infrastruct

M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our NYC or SF office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.

LinuxRestAIGo
B
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking an experienced and proactive Recruiter to help us grow our team. You will focus on hiring across our Sales team, collaborating closely with hiring managers and Sales leadership. This is a unique opportunity to build and scale the go-to-market recruiting function from the ground up—shaping strategies, processes, and candidate experience as we grow. Every hire you bring on board will play a direct role in building the future of ML infrastructure at Baseten. RESPONSIBILITIES Full-cycle recruiting: Own the hiring goals and recruiting process, from role kickoff through offer acceptance Sourcing excellence: Work closely with hiring managers to define what "excellent" looks like for a given role. Develop and execute sourcing strategies to build pipelines of highly qualified candidates, leveraging tools and creative outreach. Candidate experience: Ensure a smooth experience for every candidate, with clear communication and timely updates throughout the process Process improvements: Continuously refine and scale recruiting processes to increase efficiency, reduce time-to-fill, and improve quality of hire Data-driven insights: Track and analyze recruiting metrics (e.g., pipeline health, time-to-fill, conversion rates, acceptance rates) to inform strategies REQUIREMENTS 3+ years of full-cycle recruiting experience, preferably in a rapidly growing startup environment with big headcount goals Proven success

Machine LearningAIGo
N
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -40%

Company Description Novo is a venture-backed fintech that simplifies banking for small businesses. Since our Fall 2018 beta launch, we've expanded offerings from free checking accounts and debit cards to lending products, business tools (invoices, bookkeeping, and more), and integrations.Today , Novo is a powerfully simple banking platform that serves over 250,000 small businesses. In addition to providing smart tools built for entrepreneurs to better run and grow their businesses, Novo has processed billions in transactions in partnership with a number of established banking partners. Our vision is to be the go-to platform serving small businesses – from launch to everyday – so business owners can focus on growing, while Novo provides seamless money movement, money storage, and access to capital. We'd like to look back 5–10 years from now and know that we helped new generations of small businesses succeed because of the work we did at Novo. Novo raised $170 million in venture capital and is backed by leading investors, including Stripes, Valar Ventures, Crosslink Capital, and Notable Capital (formerly GGV). Learn more at https://www.novo.co . Role Description Novo is seeking a highly experienced and motivated product manager to own and scale our core banking experience — the accounts, money movement, cards, and program infrastructure that power banking for over 250,000 small businesses. This role is critical to ensuring the reliability, usability, and depth of the products at the heart of Novo, supporting our rapidly expanding customer base and business operations. The ideal candidate is a hands-on product leader who can set product direction for complex, regulated financial products, partner closely with Engineering, Design, Risk, and Compliance, and work hand-in-hand with banking partners to ship reliable, secure, high-quality experiences within a dynamic, regulated fintech environment. Responsibilities Product Strategy & Leadership Own the strategy, roadmap,

C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of committed researchers and engineers. The mission of the team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will work on advancing core audio model serving metrics, including latency, throughput, and quality by diving deep into our systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You’ll collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. You may

PythonGitRestMachine Learning
C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e

PythonGitRestAI
C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit

GitRestMachine LearningAI
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s mission is to unlock financial freedom for everyone by making money movement and access to financial data simple and secure. As a Software Engineer, you will design and build the systems that power how millions of people connect to their finances. You will work across the stack, from reliable backend services and APIs to intuitive applications that bring those systems to life. You will collaborate with engineers, product managers, and designers to ship products that make financial services more accessible and transparent. At Plaid, engineers take ownership early, grow quickly, and see their work reach millions of users. Responsibilities: Design & Development: Build and maintain backend services with a focus on performance, reliability and scalability. Collaboration: Work closely with product managers and other stakeholders to define and implement new features that meet product and customer needs. Code Quality: Write clean, maintainable and efficient code. Testing & Debugging: Develop automated tests to ensure the quality and reliability of the codebase. Troubleshoot and resolve issues. Engage in hands-on coding and architectural design, setting and maintaining high technical standard

SQLMySQLAWSMicroservices
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Credit team at Plaid builds lending solutions across the full lender lifecycle — including underwriting, verification, and servicing. We help lenders make faster, more informed decisions using consumer-permissioned financial data. Within Credit, the Insights team develops data-driven products including LendScore, Income Insights, CashFlow Insights, and Network Insights. These products power smarter credit decisioning for lenders of all sizes. As a Product Manager on the Insights team, you will drive execution across our portfolio of credit decisioning products. You’ll work closely with engineering and data science teams to ship high-quality product experiences, and partner with sales and marketing on go-to-market motions. You’ll report to the Insights product lead and play a critical role in ensuring we deliver on our roadmap commitments while maintaining the flexibility to pursue new product opportunities as they emerge. Responsibilities Own end-to-end execution of the product roadmap for LendScore, Income Insights, CashFlow Insights, Network Insights, and new Insights products Write product specs, define requirements, and work with engineering and data science to deliver features on schedule R

P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s mission is to unlock financial freedom for everyone by making money movement and access to financial data simple and secure. As a Software Engineer, you will design and build the systems that power how millions of people connect to their finances. You will work across the stack, from reliable backend services and APIs to intuitive applications that bring those systems to life. You will collaborate with engineers, product managers, and designers to ship products that make financial services more accessible and transparent. At Plaid, engineers take ownership early, grow quickly, and see their work reach millions of users. Responsibilities: System Design & Development: Build and maintain scalable, reliable backend or fullstack systems and APIs that power Plaid’s products. Collaboration: Work closely with product managers, designers, and other engineers to define and deliver features that solve real customer problems. Code Quality: Write clean, efficient, and well-tested code. Participate in reviews to maintain high engineering standards. Testing & Debugging: Build automated tests, monitor system performance, and troubleshoot issues in production environments. Continuous Improvement: Con

P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s mission is to unlock financial freedom for everyone by making money movement and access to financial data simple and secure. As a Fullstack Software Engineer, you will design and build the systems and experiences that power how millions of people connect to their finances. You will work across the stack, building scalable backend services and APIs while also crafting intuitive, high-quality frontend experiences that bring those systems to life. This role is ideal for engineers who enjoy switching between backend problem-solving and frontend user experience work, and who are excited to grow their impact across both. You will collaborate closely with product managers, designers, and other engineers to ship products that are reliable, secure, and delightful to use. At Plaid, engineers take ownership early, contribute to architectural decisions, and see their work reach millions of users. Responsibilities: Build across the stack. Design, develop, and maintain scalable backend services and APIs, as well as intuitive, high-quality frontend applications that bring those systems to life. Collaborate cross-functionally. Partner closely with product managers and designers to define requirements and de

JavaScriptJavaSQLMySQL
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Our Fraud Intelligence team's mission is to turn fraud signals into insight that our EPD teams transform into improvements across Protect, IDV, Signal, and Guaranteed Payments. We believe transaction patterns, device signals, identity linkages, and behavioral data are dramatically underleveraged tools in fraud prevention, and we ground our products in what adversaries are actually doing right now. As the Fraud Intelligence Lead, you will build and run a small, high-leverage team of Fraud Intelligence Analysts (and eventually a Staff Researcher) responsible for live casework across Protect, IDV, and Payments/ACH. You'll operate as a player-coach, hiring and coaching your team while staying close enough to the work to personally pick up casework and SEV response when needed. Responsibilities Team Building & People Leadership Set the casework quality bar: define what rigorous investigation, triage, and reporting look like for the team Coach analysts on investigation technique, pattern synthesis, and translating findings into product/model input Operating Model & Cross-PA Partnership Own coverage allocation across the Protect/IDV and Payments/ACH pods, including flexing assignments as volume shi

PythonSQLAWSAI
🔔

Get new quality management supervisor jobs in New York, United States by email

Daily job updates · Unsubscribe anytime