Jobiba hiring network

Inference Technical Lead Jobs

1,448 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current inference technical lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Core Models team helps shape how OpenAI’s frontier models are built, measured, and launched. We work across Research, Engineering, Model Design, Data Science, and Product to turn advances in model capabilities into reliable, useful experiences for people. Our scope includes model planning and launches as well as building data flywheels, evaluations and measurement systems to ensure our models have strong capabilities and behavior. About the Role As a Product Manager for the Core Models team, you'll be at the forefront of defining and guiding the future of how our AI models work in real-world applications. You will connect user needs to model and systems decisions: how prompts are understood; how information is aggregated and made useful for training and evaluation data; and how capabilities move from research prototypes into the mainline model and launch stack. You will operate comfortably across research, infrastructure, and consumer product surfaces, creating clarity where ownership and technical boundaries are still emerging. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate user and product goals into clear model requirements, system architecture choices, and research priorities across query understanding, indexing, retrieval, ranking, tool boundaries, data, training, inference, and evaluation. Build closed learning loops that turn product usage, explicit feedback, and other user signals into datasets, evaluations, experiments, training priorities, and launch decisions. Define success across offline evaluations and online product metrics, balancing model quality, usefulness, latency, safety, reliability, and cost. Partner closely with post-training research, applied product engineering, Model Design, and Data Science to integrate capabilities into the mainline model stack. Create reusable platforms and operatin

awsrestai
View job →
O
OpenAI
📍 Washington• Full-time
1mo ago

About the Team Join the engineering teams that bring OpenAI’s ideas safely to the world! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role We’re seeking Software Engineers who can solve complex, high-impact problems across our stack. In this role, you’ll join a nimble team driving the deployment of OpenAI’s technology into new environments and infrastructure that power critical missions in the public sector. You’ll work cross-functionally with product, security, and compliance teams to build the functionality needed to deliver a scalable, reliable platform. You’ll also partner directly with customers to design and build new products and features that create real-world impact. From launching net-new capabilities to optimizing how we serve inference in unique, high-stakes environments, this role offers both breadth and technical depth—giving you the opportunity to shape the future of OpenAI’s technology where it matters most. This role is based in Washington D.C., San Francisco, CA or Seattle, WA. Occasional travel to customer sites is required for this role. In this role, you will: Own the development of new customer-facing ChatGPT and OpenAI API features end-to-end, both on-premises and in the cloud, for our public sector customers. Partner and directly embed with teams across the business, including engineering, security, and compliance, to enable our products to work within the unique constraints of new environments. Talk to users to understand their problems and design solutions to address them Work with the research team to get relevant feedback and iterate on their latest models, developing solutions specific for public sector customers at both the model & data

javascriptpythonjava
View job →
O
1mo ago

About the Team At OpenAI, we’re building safe and beneficial artificial general intelligence. We deploy our models through ChatGPT, our APIs, and other cutting-edge products. Behind the scenes, making these systems fast, reliable, and cost-efficient requires world-class infrastructure. The Caching Infrastructure team is responsible for building a caching layer that powers many critical use cases at OpenAI. We aim to provide a high-availability, multi-tenant cache platform that scales automatically with workload, minimizes tail latency, and supports a diverse range of use cases. We’re looking for an experienced engineer to help design and scale this critical infrastructure. The ideal candidate has deep experience in distributed caching systems (e.g., Redis, Memcached), networking fundamentals, and Kubernetes-based service orchestration. In This Role, You Will: Design, build, and operate OpenAI’s multi-tenant caching platform used across inference, identity, quota, and product experiences. Define the long-term vision and roadmap for caching as a core infra capability, balancing performance, durability, and cost. Collaborate with other infra teams (e.g., networking, observability, databases) and product teams to ensure our caching platform meets their needs. You Might Thrive In This Role If You: Have 5+ years of experience building and scaling distributed systems, with a strong focus on caching, load balancing, or storage systems. Have deep expertise with Redis, Memcached, or similar solutions, including clustering, durability configurations, client-side connection patterns, and performance tuning. Have production experience with Kubernetes, service meshes (e.g., Envoy), and autoscaling systems. Think rigorously about latency, reliability, throughput, and cost in designing platform capabilities. Thrive in a fast-paced environment and enjoy balancing pragmatic engineering with long-term technical excellence. About OpenAI OpenAI is an AI research and deployment company d

redisawskubernetes
View job →

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We are seeking a Principal Product Manager, Okta Identity Threat Detection and Response Products , an individual contributor role, to drive the evolution of Okta’s cutting-edge identity security offerings. This position requires a strong technical foundation, keen product intuition, cybersecurity knowledge, and exceptional collaboration skills to shape the future of our strategic security products. You'll leverage significant cross-functional influence to execute key initiatives in this critical domain. Built upon the robust Okta Identity Cloud, our mission is to ensure Okta provides the most comprehensive, user-friendly, admin-friendly, and secure Identity Security product suite available. This team is responsible for products including Okta Identity Threat Protection with Okta AI , Okta ThreatInsight , and Okta Network Zones , among other vital security functionalities. Our customers depend on these solutions to safeguard their digital assets and ensure secure access for their users. This role is pivotal in ensuring that security is seamlessly integrated into every facet of our products. Success in this role demands a deep understanding of customer security requirements, a strong empathy for both end-user and administrator experiences, and the ability to strategically prioritize amidst competing demands from various Okta teams. Job Duties and Responsibilities: Roadmap and Vision Define the product strategy and roadmap for Okta Identity Security functional

machine learningartificial intelligenceai
View job →
DC
Diligent Corporation
📍 Netherlands• Full-time
17 days ago

Role Overview You’re a hands-on backend engineer who enjoys owning features end to end and working on real products that customers rely on every day. In this Software Engineer II role, you’ll help build and evolve a Third Party Risk Management SaaS platform using Laravel and PHP, designing scalable APIs and services that keep performance and reliability front and center. You’ll work in a product-focused team that owns its services from architecture and implementation through deployment, monitoring, and continuous improvement. You’ll mentor junior engineers, influence technical decisions, and use modern AI-powered tools thoughtfully to ship better code faster. If you’re looking for a mid-level role with real ownership, modern tooling, and the chance to grow your impact, this is for you. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Design, build, and maintain backend features and RESTful APIs in Laravel within a modern TALL stack environment. Own well-defined stories from implementation through deployment, monitoring, and iteration, ensuring performance and reliability. Contribute to architectural discussions and technical decisions that shape the Third Party Risk Management platform. Review code, improve test coverage, and strengthen CI/CD and engineering standards across the team. Mentor Software Engineer I colleagues through code reviews, pairing, and knowledge sharing. Use AI tools (e.g. GitHub Copilot, ChatGPT) to accelerate coding, debugging, testing, and documentation—while critically validating outputs and ensuring safe, responsible use. These are the essentials you’ll need to get an interview 3–5 years of professional software engineering experience in an agile, fast-paced environment. Strong experience with PHP and Laravel, ideally within the TALL stack (Tailwind, Alpine.js, Laravel, Livewire). Solid understanding of relational databases (MySQL or MariaDB), including data modelling and query optimisation. Experience designin

reactvuesql
View job →
DC
17 days ago

Here's a summary of the role: Do you love building scalable cloud platforms and solving complex engineering problems with modern technologies? As a Senior Software Engineer at Diligent, you'll design and deliver high-performing , serverless applications that power our global SaaS platform. You'll work extensively with TypeScript, Node.js, AWS, and event-driven microservices, owning services from design to deployment and production monitoring. This is an opportunity to influence technical decisions, mentor engineers, and explore how AI can transform software development and engineering productivity. If you're passionate about cloud-native architectures, distributed systems, and building software that scales to millions of users, we'd love to meet you. Here's a breakdown of what you'll do (not all of it, just the important stuff): Design and build scalable backend services and event-driven microservices using TypeScript and AWS. Develop secure APIs and integrations that power reporting, analytics, and dashboard experiences. Build and maintain serverless solutions using AWS services such as Lambda, EventBridge , SQS, and DynamoDB. Drive engineering excellence through testing, observability, automation, and production readiness practices. Contribute to infrastructure-as-code and CI/CD pipelines using AWS CDK and modern DevOps practices. Mentor engineers, participate in architecture discussions, and champion the use of AI tools to improve development efficiency. These are the essentials you'll need to get an interview: 6-8 years of professional software engineering experience. Strong experience with TypeScript, Node.js, and modern backend development patterns. Hands-on experience building cloud-native applications on AWS. Strong understanding of serverless architectures and event-driven microserv

typescriptreactnode.js
View job →
C-
CLEAR - Corporate
📍 New York• $325K – $500K/yr
12 days ago

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for a Staff Software Engineer to help build the next generation of CLEAR's enterprise identity platform, CLEAR1. We’re creating frictionless identity solutions for some of the most sophisticated companies in the world and pushing the boundaries everywhere we go. As a Staff Engineer, you’ll own high-impact, cross-functional products and integrations end to end from problem framing and architecture through implementation, rollout, adoption, and measurable outcomes. This is an ideal role for someone who thrives in scrappy, ambiguous environments, while bringing the technical rigor and operating discipline developed at larger companies. You’ll partner across teams, establish reusable patterns and standards, improve reliability and execution, influence technical direction, and mentor other engineers. A brief highlight of our tech stack: Python / Java / React / Typescript What you’ll do: Design, build, test, and deploy scalable applications that power CLEAR’s identity platform. Own projects end to end—from technical discovery and architecture through implementation, rollout, adoption, and operational support. Partner closely with Product, Design, Data, Security, and Operations to translate business problems into simple, scalable technical solutions. Design for reality by understanding failure modes, system dependencies, degradation strategies, recovery paths, and the points where systems may break under scale. Build quality and operability into the design from the start, including test strategy, observability, alerting, sa

typescriptpythonjava
View job →
A
17 days ago

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Affirm’s Identity team is mission-critical to the customer checkout experience. When a customer chooses Affirm, one of the first steps is an identity check, and our ability to make the right decision directly impacts conversion, revenue, fraud exposure, and regulatory compliance. Our team owns Identity for all markets outside North America, playing a key role in Affirm’s international expansion. We are responsible for KYC, user lifecycle management, and identity decisioning across multiple regulatory environments. This work is about shaping how Affirm adapts its core Identity platform to new market dynamics, customer expectations, and compliance requirements. We are a team of six engineers based across Spain and Poland. We are looking for a Senior Software Engineer who can turn ambiguous business and customer problems into reliable, well-designed technical solutions. In this role, you will help shape our quarterly technical direction, translate team goals into concrete projects, and identify cross-cutting risks, tradeoffs, and opportunities. You should be comfortable designing across system boundaries, driving collaboration with partner teams, and advocating for technical investments that improve long-term execution. We’re looking for someone who takes ownership beyond shipping code: raising code quality and review standards, making performance, availability, and scale tradeoffs explicit, and improving the operational health of the systems they own. You’ll help reduce toil, create useful playbooks, mentor engineers, support new hires and interns, and contribute to high-signal hiring. Most importantly, we’re excited to work with someone who leaves things better than they found them, communicates clearly across audiences, uses customer feedback to influence technical plans, and help

REMOTEpythonsqlmysql
View job →
A
Affirm
📍 Poland• Full-time• Remote• $384K – $576K/yr
17 days ago

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Affirm’s Identity team is mission-critical to the customer checkout experience. When a customer chooses Affirm, one of the first steps is an identity check, and our ability to make the right decision directly impacts conversion, revenue, fraud exposure, and regulatory compliance. Our team owns Identity for all markets outside North America, playing a key role in Affirm’s international expansion. We are responsible for KYC, user lifecycle management, and identity decisioning across multiple regulatory environments. This work is about shaping how Affirm adapts its core Identity platform to new market dynamics, customer expectations, and compliance requirements. We are a team of six engineers based across Spain and Poland. We are looking for a Senior Software Engineer who can turn ambiguous business and customer problems into reliable, well-designed technical solutions. In this role, you will help shape our quarterly technical direction, translate team goals into concrete projects, and identify cross-cutting risks, tradeoffs, and opportunities. You should be comfortable designing across system boundaries, driving collaboration with partner teams, and advocating for technical investments that improve long-term execution. We’re looking for someone who takes ownership beyond shipping code: raising code quality and review standards, making performance, availability, and scale tradeoffs explicit, and improving the operational health of the systems they own. You’ll help reduce toil, create useful playbooks, mentor engineers, support new hires and interns, and contribute to high-signal hiring. Most importantly, we’re excited to work with someone who leaves things better than they found them, communicates clearly across audiences, uses customer feedback to influence technical plans, and help

REMOTEpythonsqlmysql
View job →

About the Role & Team Every AI insight, every experiment, every cohort at Amplitude starts with a query. Our in-house OLAP engine, Nova , processes trillions of events in real time — turning raw behavioral data into fast, trustworthy answers that power decisions for thousands of product teams worldwide. We’re entering a world where AI agents don’t just assist product teams — they ship features, run experiments, and make prioritization calls autonomously. What makes that possible is agents’ ability to verify their work against real product data continuously. That makes Nova the critical infrastructure in the loop, and as non-stop agents become the main source of queries, the demand on Nova’s throughput, correctness, and operational rigor grows dramatically. We’re looking for a Staff Software Engineer who wants to go deep on both the engine internals and the infrastructure underneath it. You’ll work across the full stack of a modern OLAP system — query planning and execution, columnar storage and encoding, distributed compute, caching, and cloud infrastructure — while driving meaningful improvements to performance, cost-efficiency, and reliability at scale. You’ll influence technical direction through your work, your design reviews, and your mentorship of other engineers on a team of ~10. This role is ideal for someone who finds real satisfaction in making a complex distributed system faster, cheaper, and more reliable — and who wants to do that work on a system that directly powers the product experience for thousands of customers. What You’ll Do Build and evolve core query engine infrastructure Work across Nova's query execution engine and distributed compute layer: query planning, columnar storage formats, encoding and compression, caching, and cluster-level resource management. Design and implement new capabilities as Nova expands to support more warehouse-imported data types, such as metrics, profiles, and dimensions. Design for high-throughput automated quer

pythonjavaredis
View job →
AC
17 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Software Engineer (Senior) – Platform Engineering & AI-Powered Process Automation Location: Chennai, India | Work Model: On-site Architect the foundational systems powering the next generation of global software execution. In this role, you will build high-performance, reusable platform services and modern developer tooling that directly elevate developer velocity and scale enterprise software globally. About the Team The Platform Engineering, Foundations & Enablement team builds the bedrock of Appian’s technical ecosystem. By delivering resilient shared services, automated developer tooling, and core frameworks, this team enables engineering organizations to build scalable enterprise solutions and accelerates Appian’s growth in the AI-Powered Process Automation market. The Opportunity Top-tier engineers join Appian to tackle deep distributed systems challenges at massive enterprise scale. This role gives you direct ownership over the core platform architecture and internal developer tools that power critical workflows worldwide. If you want to eliminate engineering friction, set architecture standards, and shape high-concurrency systems using cutting-edge enterprise tech, this opportunity offers high visibility and immediate technical influence. What You’ll Do Architect and construct high-performance, resilient platform features, robust APIs , and shared frameworks to power mission-critical operations. Execute technical spikes to clear architectural runways, ensuring system stability, modularity, and seamless long-term scalability. Elevate engine

pythonjavaaws
View job →
B
Breeze
📍 Singapore• Full-time
17 days ago

At Breeze, we're building the AI-powered infrastructure layer for global commerce, making it radically simpler for businesses to sell, get paid, and operate across markets. We go far beyond traditional payment processing. Breeze combines global payments, AI, stablecoins, and a Merchant of Record-like model to take on the complexity businesses typically manage themselves, including compliance, risk, fraud, chargebacks, reconciliation, and customer support. Our goal is simple: let businesses focus on building and selling great products while Breeze handles the complexity behind getting paid. Backed by Sequoia Capital , Multicoin Capital , and The Chainsmokers , Breeze is a successful, rapidly growing, and exceptionally well-capitalized company. We have the runway to think long term while remaining early enough that every person joining today can have a meaningful impact on what we build. Overview We’re hiring a Staff Software Engineer to help build and scale the technology at the heart of Breeze’s payments platform. You’ll take on some of our most complex engineering challenges, building secure, reliable systems that move money while helping shape what we build and how we build it. You’ll work across the stack, own critical systems end to end, and partner closely with Product and Design to turn complex payments problems into simple, scalable solutions. This is a hands-on Staff role with real ownership and technical influence. You’ll help set architectural direction, raise the engineering bar, guide technical decisions across the team, and build the systems and foundations Breeze needs for its next stage of growth. If you’re a strong technical builder who wants to build, not just maintain, solve hard problems, and have a meaningful impact on both our product and engineering culture, we’d love to meet you. What You’ll Do: Build and scale the systems that power Breeze, with a focus on reliability, security, and performance across our payments platform Own critical system

typescriptreactnode.js
View job →

About the Role & Team Every AI insight, every experiment, every cohort at Amplitude starts with a query. Our in-house OLAP engine, Nova , processes trillions of events in real time — turning raw behavioral data into fast, trustworthy answers that power decisions for thousands of product teams worldwide. We're entering a world where AI agents don't just assist product teams — they ship features, run experiments, and make prioritization calls autonomously. What makes that possible is agents' ability to verify their work against real product data continuously. That makes Nova the critical infrastructure in the loop, and as non-stop agents become the main source of queries, the demand on Nova's throughput, correctness, and operational rigor grows dramatically. We're looking for a Senior Software Engineer who wants to go deep on the engine internals and the infrastructure underneath. You'll own significant components of a modern OLAP system — across query execution, columnar storage and encoding, distributed compute, caching, and cloud infrastructure — and drive meaningful improvements to performance, cost-efficiency, and reliability. You'll grow your technical influence through the quality of your code, your design contributions, and your collaboration with other engineers on a team of ~10. This role is ideal for someone who finds real satisfaction in making a complex distributed system faster, cheaper, and more reliable — and who wants to do that work on a system that directly powers the product experience for thousands of enterprise customers. What You'll Do Build and improve core query engine components Contribute across Nova's query execution engine and distributed compute layer: query planning, columnar storage formats, encoding and compression, caching, and cluster-level resource management. Implement new capabilities as Nova expands to support more warehouse-imported data types, such as metrics, profiles, and dimensions. Help ensure Nova's components support

pythonjavaredis
View job →
O
1mo ago

About the Team The Compute Strategy team works across research, engineering, product, finance, legal, and go-to-market teams to develop the partnerships, infrastructure capacity, and commercial models needed to advance AI infrastructure. About the Role As a member of the Compute Strategy team, you will develop commercial strategies for AI infrastructure partnerships and offerings. You’ll translate technical infrastructure opportunities into partnerships, transactions, and revenue. We’re looking for a commercially minded strategist who combines knowledge of semiconductors and AI infrastructure with strong financial judgment and the ability to execute complex partnerships. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Develop strategies for compute partnerships, vendor access, and infrastructure capacity. Structure and execute transactions with chipmakers, compute providers, and other infrastructure partners. Develop pricing frameworks and business cases for infrastructure-related partnerships. Evaluate partner technologies, strategic fit, commercial terms, and execution risks. Coordinate work across research, engineering, product, finance, legal, and go-to-market teams. Turn partnership learnings into repeatable operating models that can scale. You might thrive in this role if you: Have experience in strategy, corporate development, partnerships, or infrastructure transactions. Understand semiconductors and AI infrastructure. Can evaluate complex technical and commercial opportunities. Bring strong financial, analytical, and strategic judgment. Can influence and align technical and business stakeholders. Have negotiated or executed complex partnerships. Are comfortable operating in a fast-paced environment. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence ben

awsrestai
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Backend Engineer on the User Communities team, you’ll own the backend systems that power social engagement features (Announcements, Forums, Polls, Roles & Permissions systems) for millions of players and creators on Roblox. Your day-to-day will focus on building and scaling the infrastructure that makes both in-experience and on-platform communication and player engagement fast, safe, and reliable. That means driving architectural decisions, reasoning through trade-offs, influencing stakeholders, championing user-first safety standards, and defining developer-facing APIs so creators can build richer experiences. You'll work closely with frontend and backend engineers, product, design, and data scientists across teams, and you'll have real influence over our technical and product direction. If you're an experienced engineer who gets excited about large-scale distributed systems and wants to shape the way millions of people connect inside virtual worlds, we'd love to talk. You Will: Shape the backend engineering culture across the Communities team. Your influence extends beyond your immediate pod to improve how we build, review, and ship software across the entire org. Act as

javaawsgit
View job →
🔔

Get new inference technical lead jobs by email

Daily job updates · Unsubscribe anytime