At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. You will own the data Vanta's EPD organization actually runs on — bringing new signal sources online, standardizing them at the point of ingestion, and building the systems that make the underlying data trustworthy rather than merely stored. The EPD Systems team is building the infrastructure Vanta's engineering, product, and design organization depends on to understand itself. Not by asking teams to be more diligent but by going to the source, processing it, standardizing it, and pushing value back out so each source of truth earns its own adoption. What you’ll do as a Operations Manager, Signal Systems at Vanta: Bring new signal sources online end to end — from discovery and scoping through ingestion, standardization, and live operation Identify the specific failure modes in each information source and build systems that mitigate them at the point of ingestion Go to the teams that produce and consume a signal, understand what the information actually means and what they need back from it, and build accordingly Replace human-diligence dependencies with engineering solutions: derive fields, pull from source systems, validate at write time Build alongside teammates who are growing into building — raise their technical ceiling, not just your own output Partner with the inference and systems layers to ensure what you produce is queryable, trustworthy, and ready to build on How to be successful in this role: You look at a data source and see its failure modes before its contents: where it lies, where it goes stale, where it's duplicated, where the schema won't hold at 10x Your first move on an adherence problem is an engineering an
Jobs in United States
Inference Technical Lead in United States
672 active opportunities · Updated October 2026
Showing
15 jobs
Explore current inference technical lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We’re looking for a Senior Applied Scientist to help drive the future of credit applied science at Ramp. In this role, you will design, build, and optimize the models that power our credit risk systems, helping us make faster, smarter, and more scalable risk decisions for our customers. You’ll work at the intersection of machine learning, statistics, economics, and product strategy. This role requires strong technical depth as well as close collaboration with business, product, data, and engineering partners. You will help identify high-impact opportunities, translate ambiguous business problems into rigorous modeling work, and ship models that operate reliably in production. Applied scientists at Ramp focus on solving quantitative problems across credit, fraud, growth, and our core product by applying the right mix of machine learning, causal inference, structural modeling, and optimization. What You'll Do Design, build, and optimize machine learning models that support credit risk decisioning and portfolio management at Ramp Own the
From $114.3K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . This role focuses on advancing the science and systems behind ML measurement, feature understanding, and causal inference at scale. The work spans areas such as production feature importance platforms, observational causal estimation in Pytorch, large-scale proxy metric development, and data-driven approaches to ML infrastructure efficiency. We're looking for an enthusiastic individual contributor to perform high-impact technical work across this space. This person will drive foundational innovations, own the end-to-end design of production ML systems, establish rigorous methodological standards, and partner cross-functionally to turn successful research into durable platform capabilities that raise the ceiling for the entire ML organization. What you’ll do: We are looking for an experienced and highly capable Data & Applied Scientist
About the Team The Core Models team helps shape how OpenAI’s frontier models are built, measured, and launched. We work across Research, Engineering, Model Design, Data Science, and Product to turn advances in model capabilities into reliable, useful experiences for people. Our scope includes model planning and launches as well as building data flywheels, evaluations and measurement systems to ensure our models have strong capabilities and behavior. About the Role As a Product Manager for the Core Models team, you'll be at the forefront of defining and guiding the future of how our AI models work in real-world applications. You will connect user needs to model and systems decisions: how prompts are understood; how information is aggregated and made useful for training and evaluation data; and how capabilities move from research prototypes into the mainline model and launch stack. You will operate comfortably across research, infrastructure, and consumer product surfaces, creating clarity where ownership and technical boundaries are still emerging. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate user and product goals into clear model requirements, system architecture choices, and research priorities across query understanding, indexing, retrieval, ranking, tool boundaries, data, training, inference, and evaluation. Build closed learning loops that turn product usage, explicit feedback, and other user signals into datasets, evaluations, experiments, training priorities, and launch decisions. Define success across offline evaluations and online product metrics, balancing model quality, usefulness, latency, safety, reliability, and cost. Partner closely with post-training research, applied product engineering, Model Design, and Data Science to integrate capabilities into the mainline model stack. Create reusable platforms and operatin
About the Team Join the engineering teams that bring OpenAI’s ideas safely to the world! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role We’re seeking Software Engineers who can solve complex, high-impact problems across our stack. In this role, you’ll join a nimble team driving the deployment of OpenAI’s technology into new environments and infrastructure that power critical missions in the public sector. You’ll work cross-functionally with product, security, and compliance teams to build the functionality needed to deliver a scalable, reliable platform. You’ll also partner directly with customers to design and build new products and features that create real-world impact. From launching net-new capabilities to optimizing how we serve inference in unique, high-stakes environments, this role offers both breadth and technical depth—giving you the opportunity to shape the future of OpenAI’s technology where it matters most. This role is based in Washington D.C., San Francisco, CA or Seattle, WA. Occasional travel to customer sites is required for this role. In this role, you will: Own the development of new customer-facing ChatGPT and OpenAI API features end-to-end, both on-premises and in the cloud, for our public sector customers. Partner and directly embed with teams across the business, including engineering, security, and compliance, to enable our products to work within the unique constraints of new environments. Talk to users to understand their problems and design solutions to address them Work with the research team to get relevant feedback and iterate on their latest models, developing solutions specific for public sector customers at both the model & data
About the Team At OpenAI, we’re building safe and beneficial artificial general intelligence. We deploy our models through ChatGPT, our APIs, and other cutting-edge products. Behind the scenes, making these systems fast, reliable, and cost-efficient requires world-class infrastructure. The Caching Infrastructure team is responsible for building a caching layer that powers many critical use cases at OpenAI. We aim to provide a high-availability, multi-tenant cache platform that scales automatically with workload, minimizes tail latency, and supports a diverse range of use cases. We’re looking for an experienced engineer to help design and scale this critical infrastructure. The ideal candidate has deep experience in distributed caching systems (e.g., Redis, Memcached), networking fundamentals, and Kubernetes-based service orchestration. In This Role, You Will: Design, build, and operate OpenAI’s multi-tenant caching platform used across inference, identity, quota, and product experiences. Define the long-term vision and roadmap for caching as a core infra capability, balancing performance, durability, and cost. Collaborate with other infra teams (e.g., networking, observability, databases) and product teams to ensure our caching platform meets their needs. You Might Thrive In This Role If You: Have 5+ years of experience building and scaling distributed systems, with a strong focus on caching, load balancing, or storage systems. Have deep expertise with Redis, Memcached, or similar solutions, including clustering, durability configurations, client-side connection patterns, and performance tuning. Have production experience with Kubernetes, service meshes (e.g., Envoy), and autoscaling systems. Think rigorously about latency, reliability, throughput, and cost in designing platform capabilities. Thrive in a fast-paced environment and enjoy balancing pragmatic engineering with long-term technical excellence. About OpenAI OpenAI is an AI research and deployment company d
$325K – $500K/yr
CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for a Staff Software Engineer to help build the next generation of CLEAR's enterprise identity platform, CLEAR1. We’re creating frictionless identity solutions for some of the most sophisticated companies in the world and pushing the boundaries everywhere we go. As a Staff Engineer, you’ll own high-impact, cross-functional products and integrations end to end from problem framing and architecture through implementation, rollout, adoption, and measurable outcomes. This is an ideal role for someone who thrives in scrappy, ambiguous environments, while bringing the technical rigor and operating discipline developed at larger companies. You’ll partner across teams, establish reusable patterns and standards, improve reliability and execution, influence technical direction, and mentor other engineers. A brief highlight of our tech stack: Python / Java / React / Typescript What you’ll do: Design, build, test, and deploy scalable applications that power CLEAR’s identity platform. Own projects end to end—from technical discovery and architecture through implementation, rollout, adoption, and operational support. Partner closely with Product, Design, Data, Security, and Operations to translate business problems into simple, scalable technical solutions. Design for reality by understanding failure modes, system dependencies, degradation strategies, recovery paths, and the points where systems may break under scale. Build quality and operability into the design from the start, including test strategy, observability, alerting, sa
About the Team The Compute Strategy team works across research, engineering, product, finance, legal, and go-to-market teams to develop the partnerships, infrastructure capacity, and commercial models needed to advance AI infrastructure. About the Role As a member of the Compute Strategy team, you will develop commercial strategies for AI infrastructure partnerships and offerings. You’ll translate technical infrastructure opportunities into partnerships, transactions, and revenue. We’re looking for a commercially minded strategist who combines knowledge of semiconductors and AI infrastructure with strong financial judgment and the ability to execute complex partnerships. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Develop strategies for compute partnerships, vendor access, and infrastructure capacity. Structure and execute transactions with chipmakers, compute providers, and other infrastructure partners. Develop pricing frameworks and business cases for infrastructure-related partnerships. Evaluate partner technologies, strategic fit, commercial terms, and execution risks. Coordinate work across research, engineering, product, finance, legal, and go-to-market teams. Turn partnership learnings into repeatable operating models that can scale. You might thrive in this role if you: Have experience in strategy, corporate development, partnerships, or infrastructure transactions. Understand semiconductors and AI infrastructure. Can evaluate complex technical and commercial opportunities. Bring strong financial, analytical, and strategic judgment. Can influence and align technical and business stakeholders. Have negotiated or executed complex partnerships. Are comfortable operating in a fast-paced environment. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence ben
From $243.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Backend Engineer on the User Communities team, you’ll own the backend systems that power social engagement features (Announcements, Forums, Polls, Roles & Permissions systems) for millions of players and creators on Roblox. Your day-to-day will focus on building and scaling the infrastructure that makes both in-experience and on-platform communication and player engagement fast, safe, and reliable. That means driving architectural decisions, reasoning through trade-offs, influencing stakeholders, championing user-first safety standards, and defining developer-facing APIs so creators can build richer experiences. You'll work closely with frontend and backend engineers, product, design, and data scientists across teams, and you'll have real influence over our technical and product direction. If you're an experienced engineer who gets excited about large-scale distributed systems and wants to shape the way millions of people connect inside virtual worlds, we'd love to talk. You Will: Shape the backend engineering culture across the Communities team. Your influence extends beyond your immediate pod to improve how we build, review, and ship software across the entire org. Act as
From $243.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on the Communications team, you'll own the backend systems that power chat and messaging for millions of players on Roblox and support billions of messages per day. Your day-to-day will focus on building and scaling the infrastructure that makes in-game communication fast, safe, and reliable. That means driving architectural decisions, creating communication modes that don't exist yet, and defining developer-facing APIs so creators can build richer experiences. You'll work closely with frontend engineers, platform, product, and design, and you'll have real influence over the technical and product direction of the team. If you're an experienced engineer who gets excited about large-scale distributed systems and wants to shape the way millions of people connect inside virtual worlds, we'd love to talk. You Will Build and scale backend infrastructure for millions of concurrent users, architecting high-throughput distributed services, integrating ML models, and enabling rapid product experimentation Equip creators with the APIs and tools to build deeply integrated social experiences in their games Own projects end-to-end: from design and architecture through produc
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. YOU MAY BE A GOOD FIT FOR THE TEAM IF YOU HAVE: An understanding of Snowflake's product in order to be able to provide guidance to navigate complex technical evaluations Drive to enable, coach, develop and motivate a robust inside sales team Developed territories by planning strategically and effectively 2+ years guiding Account Executives through complex legal and commercial negotiations A customer first mindset to ensure success Consistently executed what is forecasted to senior management Understand how to sell business value, using a repeatable process Coach reps on how to create urgency and deal control IN THIS ROLE YOU WILL GET TO: Recruit and build successful sales professionals into high-performing revenue drivers for Snowflake Build cross-functional relationships with internal and external support teams & technical partners Influence go-to-market strategy for venture backed/high growth startups. Develop a team culture around Snowflake Core Values Sell an industry changing and market defining product ON DAY ONE WE WILL EXPECT YOU TO HAVE: BA/BS required 3+ years selling Cloud Solutions, Infrastructure Software, Databases, Analytic Tools, or Applications Software 2+ years managing a team of sales teams Sold to CEOs, CTOs, Heads of Product, Engineering, or Data Sc
Worldwide Technical Regulation Compliance Engineer Description - At HP, we create technology that makes life better for everyone, everywhere. As a Worldwide Technical Regulations (WTR) Compliance Engineer, you will play a critical role in ensuring HP products meet global regulatory, safety, environmental, cybersecurity, telecommunications, and market access requirements. Working as part of the WTR , you will collaborate with product development teams, manufacturing partners, certification laboratories, and regulatory experts worldwide to enable compliant and successful product launches. You will help drive compliance strategies throughout the product development lifecycle while balancing technical, business, and schedule objectives. This role provides the opportunity to influence products sold globally and work with experts across engineering, cybersecurity, sustainability, and regulatory compliance disciplines. Success in this Role Successful candidates are curious, collaborative, and technically strong engineers who enjoy solving complex compliance challenges. They are capable of balancing detailed technical requirements with business priorities and can influence diverse stakeholders to achieve compliant, on-time product launches. Responsibilities: Ensure HP products comply with applicable global regulatory and technical requirements, including Safety, EMC, Wireless/Telecom, Energy Efficiency and Cybersecurity regulations. Partner with product development teams early in the design cycle to identify regulatory requirements and compliance risks. Manage compliance activities for multiple product development programs from concept through commercialization. Develop and execute certification strategies that support product launch schedules and worldwide market access o
Technical Consultant / Sales Engineer - Print Level 3 Description - Overview HP is seeking talented Technical Consultants (Sales Engineers) to join our Presales organization. Technical Consultants serve as trusted advisors to customers, partners, and sales teams by providing technical expertise, solution design, business value articulation, and sales support throughout a customer's buying journey. This Sales Engineer - Print 3 posting supports ongoing hiring needs across multiple Technical Consultant job levels and technology disciplines. Multiple positions are open that support either Enterprise or Public Sector markets within the US, and focused on one of the following HP solution stacks: Personal Systems (PS), Print, or Collaboration Solutions. Successful candidates combine strong technical acumen, customer engagement skills, business understanding, and the ability to influence customer outcomes while helping drive revenue growth. Key Responsibilities Collaborate with customers, partners, and account teams to understand business objectives and technical requirements. Design and present solution architectures aligned to customer goals. Deliver demonstrations, workshops, proof-of-concepts, and technical presentations. Translate technical capabilities into business outcomes and customer value. Support opportunity qualification and sales strategy development. Respond to technical requirements, proposals, and customer inquiries. Develop trusted advisor relationships with customer decision makers and technical stakeholders. Partner closely with sales, product management, engineering, marketing, and services organizations. Maintain expertise in HP solutions, industry trends, c
From $184K/yr
TPMs at Datadog see the problems hiding between teams, engineer away the work that shouldn’t require humans, and drive the company’s most technically complex and consequential bets to completion. Technical Program Management at Datadog operates at the intersection of engineering depth and organizational reach by driving high priority, cross-functional programs that are too complex and consequential for any single team to own. We partner with engineering on solving deeply technical problems at scale by connecting the people, decisions, and context to move Datadog's most important work forward. We build the systems and automation that make entire classes of program work self-executing. We are in the architecture conversation early, earning trust through technical judgment. We use AI to surface risks earlier, accelerate program execution plans, and find cross-team patterns that would otherwise stay hidden. The faster teams move, the more essential it is to have someone who can operate across them. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What We Expect: These are the expectations we hold for every TPM at Datadog. Technical depth, product domain expertise, and AI systems literacy; knowing how AI solutions work, where they fail, and the scope and impact of those failures. AI brings more complexity into the picture - the technical bar is higher, not lower. Build the systems that reduce the need for coordination Identify what matters before anyone asks, and automate the rest Engineer program lifecycles end-to-end See what no single team can see and own the solution Drive the company's most technically complex and consequential bets through cross-functional agreement, organizational visibility, and influence Build AI powered automation tools and
From $145.6K/yr
TPMs at Datadog see the problems hiding between teams, engineer away the work that shouldn’t require humans, and drive the company’s most technically complex and consequential bets to completion. About the Role Technical Program Management at Datadog operates at the intersection of engineering depth and organizational reach by driving high priority, cross-functional programs that are too complex and consequential for any single team to own. We partner with engineering on solving deeply technical problems at scale by connecting the people, decisions, and context to move Datadog's most important work forward. We build the systems and automation that make entire classes of program work self-executing. We are in the architecture conversation early, earning trust through technical judgment. We use AI to surface risks earlier, accelerate program execution plans, and find cross-team patterns that would otherwise stay hidden. The faster teams move, the more essential it is to have someone who can operate across them. What we expect These are the expectations we hold for every TPM at Datadog. Technical depth, product domain expertise, and AI systems literacy; knowing how AI solutions work, where they fail, and the scope and impact of those failures. AI brings more complexity into the picture - the technical bar is higher, not lower. Build the systems that reduce the need for coordination Identify what matters before anyone asks, and automate the rest Engineer program lifecycles end-to-end See what no single team can see and own the solution Drive the company's most technically complex and consequential bets through cross-functional agreement, organizational visibility, and influence Build AI powered automation tools and deploy them at every stage of program execution What you will do See Across: Identify What No Single Team Can See Own large cross-functional programs spanning multiple engineering orgs, product, and business functions Proactively surface systemic risks, cross
Other cities to consider
More places hiring for this role
Get new inference technical lead jobs in United States by email
Daily job updates · Unsubscribe anytime