We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. The Data team within Plaid’s Fraud organization builds the machine learning systems that power Plaid’s fraud detection products, leveraging Plaid’s unique network data to identify and stop fraud before it happens. The team owns the full ML lifecycle—from feature pipelines and model training to production serving and monitoring—building reliable, scalable systems that deliver high-quality fraud detection as we grow to support hundreds of customers. As a Senior Machine Learning Engineer, you will own the development of high-performance feature computation and online inference pipelines that power production machine learning systems at scale. You’ll build robust observability, monitoring, and automated debugging capabilities, while leveraging AI-assisted tools to investigate complex system behavior and maintain high reliability. You’ll partner closely with ML Infrastructure, Data Science, and Product teams to execute critical technical initiatives and deliver scalable, high-impact ML solutions. Responsibilities: Build and scale machine learning systems that power a rapidly growing fraud detection product in a fast-paced environment. Solve complex technical challenges at the intersect
Jobs in United States
Inference Technical Lead in New York
99 active opportunities · Updated October 2026
Showing
15 jobs
Explore current inference technical lead jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.
$325K – $500K/yr
CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for a Staff Software Engineer to help build the next generation of CLEAR's enterprise identity platform, CLEAR1. We’re creating frictionless identity solutions for some of the most sophisticated companies in the world and pushing the boundaries everywhere we go. As a Staff Engineer, you’ll own high-impact, cross-functional products and integrations end to end from problem framing and architecture through implementation, rollout, adoption, and measurable outcomes. This is an ideal role for someone who thrives in scrappy, ambiguous environments, while bringing the technical rigor and operating discipline developed at larger companies. You’ll partner across teams, establish reusable patterns and standards, improve reliability and execution, influence technical direction, and mentor other engineers. A brief highlight of our tech stack: Python / Java / React / Typescript What you’ll do: Design, build, test, and deploy scalable applications that power CLEAR’s identity platform. Own projects end to end—from technical discovery and architecture through implementation, rollout, adoption, and operational support. Partner closely with Product, Design, Data, Security, and Operations to translate business problems into simple, scalable technical solutions. Design for reality by understanding failure modes, system dependencies, degradation strategies, recovery paths, and the points where systems may break under scale. Build quality and operability into the design from the start, including test strategy, observability, alerting, sa
From $184K/yr
TPMs at Datadog see the problems hiding between teams, engineer away the work that shouldn’t require humans, and drive the company’s most technically complex and consequential bets to completion. Technical Program Management at Datadog operates at the intersection of engineering depth and organizational reach by driving high priority, cross-functional programs that are too complex and consequential for any single team to own. We partner with engineering on solving deeply technical problems at scale by connecting the people, decisions, and context to move Datadog's most important work forward. We build the systems and automation that make entire classes of program work self-executing. We are in the architecture conversation early, earning trust through technical judgment. We use AI to surface risks earlier, accelerate program execution plans, and find cross-team patterns that would otherwise stay hidden. The faster teams move, the more essential it is to have someone who can operate across them. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What We Expect: These are the expectations we hold for every TPM at Datadog. Technical depth, product domain expertise, and AI systems literacy; knowing how AI solutions work, where they fail, and the scope and impact of those failures. AI brings more complexity into the picture - the technical bar is higher, not lower. Build the systems that reduce the need for coordination Identify what matters before anyone asks, and automate the rest Engineer program lifecycles end-to-end See what no single team can see and own the solution Drive the company's most technically complex and consequential bets through cross-functional agreement, organizational visibility, and influence Build AI powered automation tools and
From $145.6K/yr
TPMs at Datadog see the problems hiding between teams, engineer away the work that shouldn’t require humans, and drive the company’s most technically complex and consequential bets to completion. About the Role Technical Program Management at Datadog operates at the intersection of engineering depth and organizational reach by driving high priority, cross-functional programs that are too complex and consequential for any single team to own. We partner with engineering on solving deeply technical problems at scale by connecting the people, decisions, and context to move Datadog's most important work forward. We build the systems and automation that make entire classes of program work self-executing. We are in the architecture conversation early, earning trust through technical judgment. We use AI to surface risks earlier, accelerate program execution plans, and find cross-team patterns that would otherwise stay hidden. The faster teams move, the more essential it is to have someone who can operate across them. What we expect These are the expectations we hold for every TPM at Datadog. Technical depth, product domain expertise, and AI systems literacy; knowing how AI solutions work, where they fail, and the scope and impact of those failures. AI brings more complexity into the picture - the technical bar is higher, not lower. Build the systems that reduce the need for coordination Identify what matters before anyone asks, and automate the rest Engineer program lifecycles end-to-end See what no single team can see and own the solution Drive the company's most technically complex and consequential bets through cross-functional agreement, organizational visibility, and influence Build AI powered automation tools and deploy them at every stage of program execution What you will do See Across: Identify What No Single Team Can See Own large cross-functional programs spanning multiple engineering orgs, product, and business functions Proactively surface systemic risks, cross
From $156K/yr
The Team: As a Security Engineer 2 on the Cyber Threat Intelligence team, you will help Datadog stay ahead of evolving threats by identifying, analyzing, and operationalizing intelligence on threat actors, campaigns, and emerging threats. Working within Security Engineering, you will partner closely with security teams to translate intelligence into actionable security improvements across the company. You will serve as a subject matter expert on how the cyber threat landscape intersects with Datadog and contribute to intelligence-led decision making during both steady-state operations and active security incidents. This role provides opportunities to influence detection, response, and security strategy through technical analysis, collaboration, and intelligence-driven initiatives. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Develop and maintain tooling that automates the collection, processing, analysis, and dissemination of threat intelligence. Assess emerging vulnerabilities, threat activity, and security events to help stakeholders understand potential impact to Datadog. Conduct threat hunting and infrastructure analysis to identify adversary activity relevant to Datadog and improve defensive controls. Partner with security teams to operationalize intelligence into detections, investigations, and response workflows. Coordinate with information-sharing communities to gather, evaluate, and disseminate actionable intelligence. Produce technical briefings, threat reports, and intelligence products for security and engineering stakeholders. Who You Are: Experienced in writing and presenting operational and technical intelligence for threat detection, response, and security stakeholders. Skilled in partnering with detection and response te
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our Solution Engineering organization is seeking an AI Specialist who can provide hands-on expertise and support while working with technical decision makers and data scientists to design and architect AI solutions built on the Snowflake AI Data Cloud. This is a strategic role that works closely with cross-functional teams, including product, engineering, and the broader field organization to ensure successful execution and customer adoption of Snowflake’s AI & ML solutions. IN THIS ROLE YOU WILL GET TO: Be the technical expert in the room that positions Snowflake’s AI and ML features and value to technical stakeholders at Snowflake’s customers across the Americas. Partner with Snowflake account team teams and customer champions to scope and drive POCs to success and technical wins that prove the value of Snowflake’s capabilities, including executive readouts and business value cases. Collaborate with Snowflake’s product and engineering teams to influence Snowflake’s AI and ML roadmaps based on customer feedback. Publish content that helps the team and company scale beyond your individual efforts, like blog posts, presentations at conferences, or technical collateral like notebooks and demos. Influence, tailor and maintain Sales Engineering AI and ML selling assets, inc
$170K – $225K/yr
Lithic is the modern card issuing and processing platform empowering ambitious financial companies to build the future of payments. Our infrastructure powers card programs for 100+ innovative clients, from fintechs reimagining credit and digital banking to platforms transforming disbursements and spend management. Companies like Mercury, Flex, and Novo rely on Lithic's developer-friendly APIs, direct network connections, and flawless reconciliation to launch and scale card programs in weeks, not years. We're building a future where access to better financial products materially improves people's lives, free from the constraints of 30-year-old mainframes and legacy processors. We're proud to be backed by world-class investors who share that vision, including Bessemer Venture Partners, Index Ventures, Spark Capital, Stripes, and Mastercard, along with many others. We're a team of 170+ across 26 states and 7 countries, headquartered in New York City. Lithic is hiring a Solutions Engineer to design and deliver technical product solutions that meet customer needs and highlight the value of our platform. In this role, you'll be the technical and strategic bridge between Lithic and our clients. You'll own complex client relationships from pre-sale through implementation, helping partners design integrations, navigate the Lithic platform, and unlock the full potential of card issuing infrastructure. This is a high-impact, highly visible role that requires equal parts technical credibility, client empathy, and cross-functional influence. You’ll be responsible to collaborate with internal teams to develop tailored solutions that clearly demonstrate the benefits of working with Lithic. If you’re passionate about technology, problem-solving, and creating exceptional customer experiences this role is for you. What You'll Do Client Engagement Serve as a trusted technical advisor to a portfolio of strategic clients, from early-stage fintechs to enterprise partne
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of committed researchers and engineers. The mission of the team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will work on advancing core audio model serving metrics, including latency, throughput, and quality by diving deep into our systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You’ll collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. You may
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's GTM org is in hyper growth. As it grows and matures, the needs of the GTM stack get more sophisticated with scale — and this team exists to stay ahead of those needs. Our GTM Engineering team plays a critical role in building and maintaining the connective tissue across Baseten’s GTM tools, processes, and user experience for the field. The GTM tooling landscape is changing fast, and the teams that win are the ones that adapt and iterate the fastest. This role exists to make sure Baseten is one of them. You'll design, build, and ship AI-powered workflows that scale our GTM functions as a competitive advantage. We want someone who can walk in, audit what we have, identify what we're missing, and start shipping fast. You know when to reach for Clay and when to build something custom in Claude Code. You think two to three steps ahead about how the thing you build today fits into the broader systems architecture tomorrow. And you bring a point of view — on our stack, on what we should be building, and on where AI can do something low-code tooling simply can't. RESPONSIBILITIES Ship AI-powered workflows for the field — build the agents and automations that give reps and managers real leverage, off-loading the manual and repetitive work. Reach for AI where it does something low-code can't. Get insights in front of reps — turn Salesforce, warehouse, and usage data into the dashboards, scores, and alerts reps
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. ABOUT THE ROLE As a Event Marketing Manager, you will lead the development and execution of regional marketing strategies to support sales goals, drive demand, and enhance brand awareness. This role requires a strategic thinker with a builder mentality that has a strong understanding of the New York tech ecosystem. RESPONSIBILITIES Develop and execute regional field marketing plans and programs to align with sales priorities and pipeline goals. Plan and manage events to drive demand, accelerate pipeline, and engage customers. Own program management, including event logistics, budget oversight, and vendor coordination. Analyze and report on campaign effectiveness, providing insights to optimize future initiatives. Collaborate with product marketing to align messaging and go-to-market strategies with field initiatives. Prioritize seller requests by expected revenue. Represent Baseten at industry events and build brand presence in-market. QUALIFICATIONS Bachelor’s degree in Marketing, Business, or related field. 5-7 years of experience in field marketing in the technology or SaaS industry. Proven track record in delivering and optimizing demand generation programs to meet pipeline goals. Strong understanding of sales processes. Proficiency with marketing tools such as Salesforce & Hubspot. Excellent project management skills with the ability to prioritize and execute multiple initiatives simultaneously. Collaborativel
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking an experienced and proactive Recruiter to help us grow our team. You will focus on hiring across our Sales team, collaborating closely with hiring managers and Sales leadership. This is a unique opportunity to build and scale the go-to-market recruiting function from the ground up—shaping strategies, processes, and candidate experience as we grow. Every hire you bring on board will play a direct role in building the future of ML infrastructure at Baseten. RESPONSIBILITIES Full-cycle recruiting: Own the hiring goals and recruiting process, from role kickoff through offer acceptance Sourcing excellence: Work closely with hiring managers to define what "excellent" looks like for a given role. Develop and execute sourcing strategies to build pipelines of highly qualified candidates, leveraging tools and creative outreach. Candidate experience: Ensure a smooth experience for every candidate, with clear communication and timely updates throughout the process Process improvements: Continuously refine and scale recruiting processes to increase efficiency, reduce time-to-fill, and improve quality of hire Data-driven insights: Track and analyze recruiting metrics (e.g., pipeline health, time-to-fill, conversion rates, acceptance rates) to inform strategies REQUIREMENTS 3+ years of full-cycle recruiting experience, preferably in a rapidly growing startup environment with big headcount goals Proven success
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Marketing Operations Manager who can own and harden the systems layer of Baseten's Marketing engine. Marketing at Baseten is scaling fast — more spend, more campaigns, more model launches, more inbound. The systems underneath (Our ESP, CRMs, forms, tracking, routing, alerting) need an owner who treats them like production infrastructure that cannot go down. When something breaks, it costs us time, pipeline, and trust in the data. This role exists so it doesn't break. You’ll simultaneously build for the future and re-think assumptions about our tech stack in the age of agents. This is an offensive play that gives the rest of the team leverage and superpowers to hit our ambitious goals. This is NOT an IT or service role. This is a core member of the marketing team who implements technology to achieve outcomes. RESPONSIBILTIES Own the marketing tech stack end-to-end: ad platforms, email systems, tracking, pixels, forms, connectors. Build defense-in-depth on inbound: spam/bot protection, rate limiting, email/domain validation, sync gating — and the alerting to catch anomalies before they hit sales or leadership dashboards. Enforce data integrity: UTM governance, campaign membership, lifecycle stages, lead scoring and routing logic, field-level hygiene, canonical metric definitions. Operationalize the web request pipeline with our dev agency: structured briefs, tickets, SLAs, and launch-day runb
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit
Datadog's integrations are the connective tissue between our platform and the technologies our customers run in the real world. As a Sr. PM on the Agent Integrations team, you will own the vision, prioritization, and execution for 100+ integrations that run directly inside the Datadog Agent from foundational infrastructure (MySQL, Kafka, Kubernetes) to the rapidly growing landscape of self-hosted AI and on-premise enterprise technologies. This is a high-impact, breadth-first role at the intersection of infrastructure observability and the frontier of AI-native workloads. At Datadog, we place value in our office culture; the relationships it builds, the creativity it brings, and the collaboration of being together. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own the Agent Integrations roadmap. Determine which new integrations to build and which existing ones to improve, balancing customer demand, business impact, and engineering capacity across a catalog of 100+ technologies. Drive the expanding AI integration surface. Lead product strategy for self-hosted AI workloads, including LLM inference frameworks (e.g., Hugging Face TGI, BentoML), AI agents, MCP servers, and model orchestration tools, so Datadog customers can monitor every layer of their AI stack. Expand on-prem and hybrid coverage. Prioritize and execute new integrations for on-prem technologies including storage systems, HPC schedulers, network devices, and legacy enterprise platforms where customers run critical workloads. Build observability for ERP systems. Define and drive Datadog's strategy for monitoring enterprise ERP platforms (SAP, Oracle EBS/Fusion, Microsoft Dynamics) covering performance, job execution health, and integration layer telemetry so enterprise customers can observe their ERP stack alongside the rest of their infrastructure. Analyze adoption and customer feedback at scale. Use data from multiple sources to
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. About Modal Data: We’re growing our Data team and are looking for our first few key hires to build self-serve data tools and drive business strategy in the right direction. The mission of the Modal Data team is to make it easy to track company goals, make evidence-backed decisions, and prioritize the right work. We do this via: Self-serve AI analytics tools (Hex, Snowflake) Embedding with teams as a “data adviser”, providing strategic analysis and consulting What You'll Do: Contribute to building the most modern analytics stack in Data today to support AI-driven self-serve analysis, key metrics tracking, and external customer reporting Influence work on new products like LLM Inference Endpoints through product analytics tracking Identify millions of dollars of cost savings and optimization across our tools and financial operations Write data pipelines that power the operatio
Other cities to consider
More places hiring for this role
Get new inference technical lead jobs in New York, United States by email
Daily job updates · Unsubscribe anytime