Jobs in United States

Inference Technical Lead in United States

672 active opportunities · Updated October 2026

Explore current inference technical lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

S
📍 Portage, Michigan, United States· Hybrid
✓ High-confidence listingCompany trend +364.7%
Quick readStrong listing-quality and freshness signals

Work Flexibility: Hybrid *This is a hybrid role at our Portage office, with the expectation to be onsite 3 days/week As a Senior Marketing Associate , you will assist in developing, analyzing, and implementing marketing plans for Surgical Technologies . You'll also market the organization's products and services using customer marketing databases and create direct mail marketing plans, targeting specific market segments with specialized offers. You will collaborate with market research in developing response models and other database improvements and may conduct data mining analyses of customer data to develop marketing trends What you will do Understand key competitors and their relative strengths/weaknesses Understand customer groups, including why customers buy the product or service Advocate on the customer's behalf and influence product or portfolio changes based on customer feedback Support in the construction of the marketing plan with specific analysis and collaborate in the development of key strategic initiatives Responsible for achieving a budget target Work proactively with market research to collect customer insights for segmentation Responsible for product or portfolio's targeting and positioning Prioritize marketing initiatives appropriately given product or portfolio goal and metric Intimately understand the customers for the product or portfolio, including attitudes and behaviors Apply market data, use planning tools, and seek expert opinions when analyzing channel strategies Is the subject matter expert for applicable products/product lines and able to field technical questions. [note: some products are more tec

P
📍 New York City, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.7%
Quick readStrong listing-quality and freshness signals

About Pinecone: Pinecone is the trusted AI knowledge company. Its trusted AI knowledge platform—including its Database, Nexus, and Marketplace products—power accurate, fast, and cost-effective AI applications for more than 10,000 customers and 1M developers worldwide. Pinecone's mission is to make AI knowledgeable. For more information, visit pinecone.io . About the Team/Role: This is intended to be an entry level role for new grads/early career professionals. As an Associate Field Engineer, you'll work amongst our Customer Success and Solutions Engineering teams, building deep technical expertise and tooling to directly impact the core product and customer experience. You'll engage with technical stakeholders across the customer lifecycle, from helping prospects understand how Pinecone fits their architecture to guiding existing customers through implementation, optimization, and expansion. You'll collaborate closely with Sales, Product, and Engineering to scope solutions, troubleshoot complex issues, and surface customer insights that shape our product direction. Technical fluency, strong communication, intellectual curiosity, and an AI-first mindset are essential for this role. Responsibilities: Serve as a technical point of contact for customers across their journey from evaluation to production, ensuring best practices and accelerating time-to-value. Deeply understand customer architectures, use cases, and requirements to influence our roadmap and devise solutions that unblock and advance their goals. Build resources to monitor customer health, proactively identifying opportunities to address risk and drive expansion. Troubleshoot issues and deliver high-quality technical support, developing a strong command of Pinecone's product and the broader AI/ML ecosystem. Build processes and capture knowledge to scale and automate support, success, and outreach initiatives. Partner with Sales on technical discovery, deliver product demonstrations, and contribute to proof

JavaScriptPythonJavaAWS
S
📍 Chicago, Illinois, United States· Full-time
✓ High-confidence listingCompany trend -92.9%
Quick readStrong listing-quality and freshness signals

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our Solution Engineering organization is seeking an AI Specialist who can provide hands-on expertise and support while working with technical decision makers and data scientists to design and architect AI solutions built on the Snowflake AI Data Cloud. This is a strategic role that works closely with cross-functional teams, including product, engineering, and the broader field organization to ensure successful execution and customer adoption of Snowflake’s AI & ML solutions. IN THIS ROLE YOU WILL GET TO: Be the technical expert in the room that positions Snowflake’s AI and ML features and value to technical stakeholders at Snowflake’s customers across the Americas. Partner with Snowflake account team teams and customer champions to scope and drive POCs to success and technical wins that prove the value of Snowflake’s capabilities, including executive readouts and business value cases. Collaborate with Snowflake’s product and engineering teams to influence Snowflake’s AI and ML roadmaps based on customer feedback. Publish content that helps the team and company scale beyond your individual efforts, like blog posts, presentations at conferences, or technical collateral like notebooks and demos. Influence, tailor and maintain Sales Engineering AI and ML selling assets, inc

PythonAWSAzureGCP
A
📍 California, USA - Remote, United States· Remote
✓ Quality checkedCompany trend +90.9%

Job Requisition ID # 26WD97363 Position Overview We are seeking a Principal Software Engineer – Backend to join Autodesk’s Enterprise Data Management (EDM) organization within the COO-GET Engineering group. This is a senior individual contributor role operating at the Principal (P4) level , expected to drive technology direction for large, complex, and business-critical backend and distributed systems . This role is anchored in backend software engineering excellence : designing, building, and evolving scalable services, APIs, and event-driven systems that operate at enterprise scale. As a Principal Engineer, you will work with high autonomy and ambiguity , shape long-term architecture, and influence multiple teams and domains. Familiarity with data engineering concepts is valuable, but backend systems, service design, and distributed systems are the core competencies. You will function as a technical authority and force multiplier—guiding design decisions, setting standards, and ensuring Autodesk’s core data services are reliable, resilient, and evolvable over time. Responsibilities Provide principal-level technical

PythonAWSAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $216.7K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Privacy Software Engineer on the Privacy Infrastructure team, you will design and build the foundational platforms, services, and controls that enable Roblox to protect user data and meet global privacy obligations at scale. You will develop privacy-by-design solutions that support data governance, user privacy rights, data discovery, retention, auditing, and regulatory compliance across a rapidly growing ecosystem of products and services. This role sits at the intersection of distributed systems, data platforms, security, and privacy engineering. You will partner closely with engineers, product teams, security, legal, and policy stakeholders to build scalable privacy infrastructure that is deeply integrated into Roblox's development workflows. Your work will directly influence how we responsibly manage data for millions of users while enabling innovation across the platform. You Will Design and build scalable privacy infrastructure that enables Roblox to discover, govern, protect, and manage personal data at scale. Develop backend services, APIs, and data pipelines that embed privacy controls into engineering workflows and enable fulfillment of user privacy rights at global sc

PythonAWSGitAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $280.5K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the Role: AI models reshaping how our community creates, plays, and connects, all run on Compute Platform. As Senior Product Manager, Compute Platform , you'll set the strategy and roadmap for Roblox's next-generation AI infrastructure: the rapidly growing fleet of GPUs and AI accelerators spanning Roblox core and edge data centers, and public cloud that decides how fast we can train, serve, and scale every model on the platform. You'll own the products that turn raw GPU hosts into reliable, production-ready AI compute - driver and firmware management, fleet-wide health and performance, and the abstractions product teams across Roblox build on. You Will: Drive strategy and roadmap for Compute Platform spanning Managed Kubernetes (Roblox Kubernetes Service), Managed Compute Services and other critical distributed systems, and our fleet of GPU and CPU machines managed via unified Fleet APIs - all across on-prem and cloud. Drive the evolution of our Compute infrastructure to support Roblox’s most critical workloads - from AI to Storage to Data Analytics and more - each with their own distinct requirements. Build and scale our GPU infrastructure to support training and inferen

AWSAzureGCPKubernetes
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the team The Computer Use and New Interfaces team is focused on discovering and building the next generation of AI-native interfaces. We believe that the value of AI is increasingly constrained not by model capabilities, but by the ways people interact with those capabilities. Our mission is to create new interaction paradigms that unlock the full potential of AI and integrate it more deeply into people's lives and work. Our team does both near-term product development and longer-term product incubation that can influence experiences across ChatGPT, Codex, future OpenAI products, and emerging device platforms. We work in a highly collaborative, design-driven environment where engineering, product, and design operate as one team. We value rapid experimentation, prototyping, and iteration, creating the shortest possible path between an idea, a working system, and a product decision. About the role We're looking for exceptional engineers who are excited to invent entirely new ways for people to interact with AI. You'll work at the intersection of engineering, product, and design to explore, prototype, and build novel interface concepts that push beyond traditional software paradigms. This role requires comfort with ambiguity, strong product instincts, and a willingness to move fluidly between experimentation and production systems. You'll help shape both the capabilities and the user experiences that define how people engage with AI. In this role you will: Design, prototype, and build novel AI-native interfaces and interaction models. Develop foundational technologies and frameworks that enable generative UI and computer use experiences. Collaborate closely with designers, product thinkers, and engineers to rapidly explore and validate new concepts. Build end-to-end prototypes and production systems that can influence future OpenAI products. Contribute to platform technologies that can be adopted across multiple product surfaces. Help establish technical directio

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

We’re looking for a Software Engineer to architect and build backend systems that enforce data privacy and automate compliance at scale. You’ll work closely with product, infrastructure, security, and legal teams to embed privacy-by-design into our data and access layers. This is a hands-on, high-impact role for an experienced engineer who is passionate about protecting user data while enabling innovation. What You’ll Do Design, build, and operate backend services that enforce policy-driven data access, lifecycle controls, and privacy protections. Develop distributed authorization and identity-aware enforcement mechanisms integrated directly into data services and control planes. Implement auditability, policy hooks, and enforcement observability to ensure compliance is continuously verifiable. Partner with Security, Legal, and Compliance to convert privacy requirements into scalable technical designs and developer-friendly APIs. Harden data platforms and backend services through schema-level controls and data handling constraints by default. Collaborate with infrastructure teams to ensure consistent enforcement across systems while minimizing duplicated implementations. Contribute patterns, libraries, and education that elevate trustworthy data access patterns across the organization. You Might Thrive in This Role If You Have 5+ years of industry experience building and operating backend or infrastructure systems in production. Strong software engineering fundamentals , with fluency in at least one major programming language (e.g., Python, Go, Rust, C++, Java). Experience with distributed authorization, RBAC/ACL systems, encryption-based access, or policy engines. Familiarity with global privacy regulations and their architectural implications. Ability to influence and collaborate with teams across legal, compliance, product, and engineering. A bias toward practical, impactful solutions that balance privacy protections with product needs. Nice to Have Experience wi

PythonJavaAWSAzure
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our Solution Engineering organization is seeking an AI Specialist who can provide hands-on expertise and support while working with technical decision makers and data scientists to design and architect AI solutions built on the Snowflake AI Data Cloud. This is a strategic role that works closely with cross-functional teams, including product, engineering, and the broader field organization to ensure successful execution and customer adoption of Snowflake’s AI & ML solutions. IN THIS ROLE YOU WILL GET TO: Be the technical expert in the room that positions Snowflake’s AI and ML features and value to technical stakeholders at Snowflake’s customers across the Americas. Partner with Snowflake account team teams and customer champions to scope and drive POCs to success and technical wins that prove the value of Snowflake’s capabilities, including executive readouts and business value cases. Collaborate with Snowflake’s product and engineering teams to influence Snowflake’s AI and ML roadmaps based on customer feedback. Publish content that helps the team and company scale beyond your individual efforts, like blog posts, presentations at conferences, or technical collateral like notebooks and demos. Influence, tailor and maintain Sales Engineering AI and ML selling assets, inc

PythonAWSAzureGCP
L
📍 New York, New York, United States· Full-time
✓ High-confidence listing

$170K – $225K/yr

Quick readStrong listing-quality and freshness signals

Lithic is the modern card issuing and processing platform empowering ambitious financial companies to build the future of payments. Our infrastructure powers card programs for 100+ innovative clients, from fintechs reimagining credit and digital banking to platforms transforming disbursements and spend management. Companies like Mercury, Flex, and Novo rely on Lithic's developer-friendly APIs, direct network connections, and flawless reconciliation to launch and scale card programs in weeks, not years. We're building a future where access to better financial products materially improves people's lives, free from the constraints of 30-year-old mainframes and legacy processors. We're proud to be backed by world-class investors who share that vision, including Bessemer Venture Partners, Index Ventures, Spark Capital, Stripes, and Mastercard, along with many others. We're a team of 170+ across 26 states and 7 countries, headquartered in New York City. Lithic is hiring a Solutions Engineer to design and deliver technical product solutions that meet customer needs and highlight the value of our platform. In this role, you'll be the technical and strategic bridge between Lithic and our clients. You'll own complex client relationships from pre-sale through implementation, helping partners design integrations, navigate the Lithic platform, and unlock the full potential of card issuing infrastructure. This is a high-impact, highly visible role that requires equal parts technical credibility, client empathy, and cross-functional influence. You’ll be responsible to collaborate with internal teams to develop tailored solutions that clearly demonstrate the benefits of working with Lithic. If you’re passionate about technology, problem-solving, and creating exceptional customer experiences this role is for you. What You'll Do Client Engagement Serve as a trusted technical advisor to a portfolio of strategic clients, from early-stage fintechs to enterprise partne

GitAIGoRust
S
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.6%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About The Role The Security Team is responsible for securing all things Sentry: our customers, our code, and everything in between. We are a small but growing team with broad scope, high trust, and the autonomy to tackle hard security problems with creativity and an engineering mindset. We work at a company with a strong developer culture, building a product that millions of developers genuinely love and rely on. That context shapes everything about how we operate. As a Security Engineer on this team, you'll work across application and platform security domains. You'll contribute to the practices that keep Sentry secure as we grow: security reviews, threat modeling, vulnerability management, and embedding secure coding practices into an engineering organization that cares about doing things right. You'll partner closely with product and engineering teams to influence how features are designed and built from the start. You will work as a technical collaborator who helps make the secure path the obvious one. As Sentry expands our agentic product capabilities and development practices, you'll also find yourself at the frontier of a new set of security challenges. In this role, you will Support and help mature Sentry's security review program. From secure code review, to architecture review, and threat modeling. You'll help build the processes, tooling, and culture which make security a natural part of how we ship and operate. Contribute to mature vulnerability management practices. Intake, triage, prioritization, remediation tracking, and support of our bug bounty and responsible disclosure program. Advocate for secure-by-desig

TypeScriptPythonAWSAzure
R
📍 New York City, NY, United States· Full-time
✓ Quality checkedCompany trend -99.2%

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role The Credit Engineering team builds the systems that enable Ramp to manage credit lifecycles for businesses end-to-end, from initial underwriting through ongoing portfolio strategy optimization. As an engineer on Credit, you’ll own systems that influence $100+ billion in annual payment volume through millions of decisions across our product stack in support of the most ambitious FinTech portfolio in the United States. We are looking for a backend engineer to own the technical roadmap for underwriting, limit management, portfolio optimization, and agentic workflows. You will build robust systems, automate complex operational tasks, and maintain the high-frequency controls that power all of Ramp’s products. If you are obsessed with correctness, excited by ambiguity, and passionate about problems at the intersection of distributed systems and financial strategy, this role is for you. What You'll Do Architect scalable, stateful systems for automated limit management to optimize Ramp’s charge card portfolio. Envision and build the next gene

PythonRestAIGo
R
📍 New York City, NY, United States· Full-time
✓ High-confidence listingCompany trend -99.2%

From $120K/yr

Quick readStrong listing-quality and freshness signals

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. What You’ll Do Full stack development, building models to consume, transform, and expose data to stakeholders and production systems Drive a culture of experimental design, testing agenda, and best practices Contribute to the culture of Ramp’s data team by influencing processes, tools, and systems that will allow us to make better decisions in a scalable way Collaborate with Finance teams (e.g. GTM Finance, StratFin) to develop financial insights and influence business decisions Work closely with data engineering teams to capture, move, store, and transform raw data into highly actionable insights, and partner with business teams to turn those insights into action What You Need Minimum of 3 years of industry experience in Data Science / Software Engineering / Finance Strong AI proficiency as a lever to quickly adopt new skills and subject matter Track record of shipping high quality products and features at scale Ability to thrive in a fast-paced, constantly improving, start-up environment that focuses on solving problems with iterative technical so

SQLRestAIGo
SF
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $136K/yr

Quick readStrong listing-quality and freshness signals

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role As an ML Platform Engineer at Stitch Fix, you will play a key role in building and maintaining the critical infrastructure that powers machine learning and AI across our organization. You will design, develop, and support scalable, resilient services and frameworks for ML model training and deployment, feature engineering and serving, candidate generation, AI agent deployment and observability, and other core platform capabilities. In this role, you'll contribute to the day-to-day operations of the ML Platform team, ensuring the smooth functioning of existing systems while driving improvements. You’ll collaborate closely with full-stack data scientists, offering consultation and support to help them unlock the full potential of our platform. With significant autonomy, you’ll have the opportunity to shape the future of ML and AI at Stitch Fix. Your ideas and expertise will drive improvements, codify best practices, and influence how we approach machine learning and AI systems at scale. Responsibilities: Collaborate with cross-functional teams, including data scientists, engineers, and business partners, to solve complex distributed systems and business challenges at scale. Be part of a team with high visibility across the organization, driving impactful solutions that make a difference. Share your ideas and help guide the team’s investments toward high-value opportunities. Foster a culture of technical collaboration and contribute to the development of scalable, resilient systems. About You You bring

PythonRedisAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Personalization-Memory team, within OpenAI's broader Personal AGI organization, is focused on developing agents that can learn from prior interactions in order to become more helpful and efficient over time. We build general-purpose memory and personalization capabilities that transfer across ChatGPT and other agentic products, and we collaborate with applied engineering on the product surfaces that allow users to interact with memory. About the Role As a Research Engineer / Research Scientist on the Personalization-Memory team, you will research and develop improvements to memory usage and personalization in OpenAI's frontier models. Our team works on reinforcement learning, dataset creation, evaluations, and other post-training methods. We partner closely with research and product teams across the company to realize the vision of a truly personalized ChatGPT. We're looking for individuals who have a background in frontier model post-training, are able to iterate quickly, and who are passionate about product-driven research. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own and pursue a research agenda for improving memory use and personalization in frontier models. Build robust evaluations for tracking modeling improvements. Design, implement, test, and debug code across our research stack. Collaborate closely with the research and product teams to influence the shape of technical solutions in the product. You might thrive in this role if you: Are passionate about personalization and building personalized assistants. Have experience working with user signals and human data to turn feedback into reliable signals for training and evaluation. Have a deep understanding of frontier model post-training and machine learning applications. Value principled approaches and research craftsmanship. Are comfortable diving into a lar

AWSRestMachine LearningAI
🔔

Get new inference technical lead jobs in United States by email

Daily job updates · Unsubscribe anytime