Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Responsible for understanding the processes used to assess and mitigate risks related to the introduction of new package technologies so that you can improve efficiency and effectiveness by implementing automation, data analysis and machine learning/AI solutions. Set up data bases, develop data ingestion pipelines, and implement automated data analysis and reporting tools. Support problem solving and risk assessment for internal or customer quality issue by pulling and analyzing data. Contribute to the advancement of technology at Micron through mentoring, publishing technical papers (internal and external), and developing innovative solutions to challenging problems. Implement Automation, Data Analysis, and AI Solutions. Collaborate with Engineering teams to Map Package DDQA processes and data streams. Set up and optimize databases and develop solutions to improve efficiency and effectiveness. Understand the needs of internal customers and develop solutions. Support Problem Solving and Risk Assessment for Quality Issues. Pull relevant product information, manufacturing data, and reliability data based on given problem statements. Determine the appropriate dataset and treatment required to answer questions posed by problem solving teams. This could include producing data visualizations, machine learning models, statistical inferences, and web applications. Provide recommendations about root cause findings and product risk, based on data analysis. Collaboratively Communicate Findings and Best Practices. Share best practices with global teams to enable a cultur
Jobs in United States
Inference Technical Lead in United States
672 active opportunities · Updated October 2026
Showing
15 jobs
Explore current inference technical lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
From $278.5K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. ML Platform @ Roblox today supports hundreds of ML use cases and billions of inferences per day across Discovery, Safety, Engine, and much more. As an Infrastructure Engineer on the ML Platform team, you will design, scale, and maintain the foundational infrastructure powering our entire machine learning ecosystem. We are looking for accomplished engineers to spearhead the development of our next-generation ML tooling and platform capabilities. You will: Bootstrap and maintain Kubernetes and Cloud infrastructure for ML Platform components--Serving Layer, Metadata Store, Model Registry, and Pipeline Orchestrator. Set technical strategy and oversee development of high scale and reliable infrastructure systems. Propose and implement new platform tooling to improve time to production for MLEs and Data Scientists across the full ML lifecycle. Work on infrastructure projects such as GPU fleet management, hybrid-cloud orchestration, and writing custom Kubernetes controllers and resources. Stay abreast of industry trends in machine learning and infrastructure to ensure the adoption of leading-edge technologies and practices. Partner across organizations to build tooling, interfaces, and visualizati
At Freddie Mac, our mission of Making Home Possible is what motivates us, and it’s at the core of everything we do. Since our charter in 1970, we have made home possible for more than 90 million families across the country. Join an organization where your work contributes to a greater purpose. Position Overview: We need a highly innovative Technical Lead! How confident are you that you can build sophisticated analytic systems? If you believe you could contribute to the development of innovative principles and ideas in a matrixed environment, please keep reading as we are seeking an individual contributor who has experience with Java and Python and can lead and nurture an inspiring environment in our Virginia office. Our Impact: The Investments and Capital Markets (I&CM) division is looking for a capable technology lead for its trading and analytics development team. This could be you! To thrive in this division, you must have a comprehensive understanding of system implementation and design, experience working in capital markets, and be enthusiastic about leading development of new paradigms in software system architecture. Your Impact: As a Trading Analytics Development Tech Lead, you will develop and maintain software using Java and Python tech stack that adheres to software engineering best practices. You will influence technical decisions, mentor developers, resolve engineering blockers, and partner with engineering managers to help teams deliver secure, reliable, and maintainable solutions. You will provide hands-on directions for full-stack applications, APIs, microservices, and integration services while reinforcing engineering discipline across design, development, testing, deployment, observability, and production readiness. Partner closely with Product Owners, engineering managers, architecture, business stakeholders, and cross-functional tec
Work Flexibility: Remote As a Senior Lead, Data Engineering, you will serve as a technical leader who helps shape the future of enterprise data solutions. In this role, you will drive complex data initiatives, influence technical strategy, and partner with teams across the organization to build scalable, high-impact data products. This is an opportunity to solve challenging business problems while mentoring fellow engineers and elevating data engineering best practices. What You Will Do Lead the architecture, development, and modernization of scalable enterprise data platforms that support global procurement analytics and business transformation. Define and help execute a multi-year data engineering strategy focused on platform scalability, reliability, automation, technical debt reduction, and long-term maintainability. Design, build, and optimize Azure-based data solutions using technologies such as Databricks, Delta Lake, Azure Data Factory, Azure DevOps, CI/CD pipelines, and infrastructure automation. Integrate and harmonize data across multiple ERP systems by standardizing supplier, purchasing, and master data into common enterprise data models. Partner with procurement analysts, architects, engineers, and business stakeholders to translate complex business needs into reusable, scalable data products and engineering solutions. Establish engineering standards, conduct architecture reviews, improve documentation, and mentor engineers to raise the overall technical capability of the team. Identify and implement AI-enabled approaches that accelerate development, improve data quality, automate documentation, support testing, and enhance analyst productivity. Evaluate and recommend tools, frameworks, patterns, and platform investments that improve performance, reliability, security, governance, and operational ef
About the Role We are seeking a Cloud Infrastructure Engineer to help design and evolve the platforms that power OpenAI’s products. In this role, you will be a hands-on technical leader, driving the architecture, scalability, reliability, and security of critical infrastructure systems. You will help define how we build and operate infrastructure at the next order of magnitude, while influencing technical direction across teams. This role is both deeply technical and highly strategic, requiring strong ownership, sound judgment, and the ability to partner effectively across engineering, product, and research organizations. In this role, you will: Design and build scalable, reliable, and secure infrastructure platforms that power OpenAI products Evolve cloud infrastructure abstractions that enable rapid product development across teams Architect systems to support significant growth, performance, and operational complexity Improve server orchestration, networking, distributed systems reliability, and infrastructure security posture Influence technical direction and infrastructure strategy across multiple teams Partner closely with product, research, and engineering teams to align infrastructure with evolving needs Own operational excellence, including participation in on-call rotations, incident response, and production readiness Mentor engineers and raise the overall technical bar of the organization Contribute to a culture of high ownership, low ego, and thoughtful collaboration You might thrive in this role if you: 8+ years of experience building and operating large-scale infrastructure systems Deep expertise in Kubernetes and container orchestration at scale Strong experience designing cloud abstractions and platform infrastructure (AWS, GCP, Azure, or similar) Proven track record of leading complex technical initiatives across teams Experience operating highly reliable, secure, and scalable distributed systems Security engineering experience or security backgroun
From $154K/yr
We are Datadog’s in-house technical leaders. The Technical Account Management team drives Datadog’s continued global growth by ensuring our customers realize long-term value from our platform through successful adoption, expansion, and partnership. As a Manager 1 in Technical Account Management, you will lead and develop high-performing technical teams while influencing strategy, execution, and outcomes across customers, internal partners, and the broader organization. Manager 1 leaders at Datadog are people-first managers, trusted collaborators, and operational owners. You will coach and mentor individual contributors, drive execution against team and organizational goals, and serve as a strong voice for customer needs and technical excellence. At Datadog, we place value in our office culture; the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and coach a team of up to 6 Technical Account Managers, providing regular 1:1s, team meetings, and bi-annual performance feedback Own and track team KPIs , including scheduling, utilization, productivity, and delivery outcomes Partner closely with Sales, Customer Success, Presales, Product Management, Support, and Marketing to align post-sales strategy and execution Lead and participate in customer-facing engagements when appropriate, including escalations, strategic reviews, and key account discussions Drive account strategy discussions focused on product adoption, expansion, and services delivery Actively participate in recruiting , hiring, and onboarding efforts across your team and the broader organization Gather and synthesize customer feedback to influence product direction, process improvements, and internal initiatives Lead multiple OKR initiatives annually , coordinating and delegating efforts across your team Demonstrate thought leadership by identifyi
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role Is Different This is not a typical “Applied Scientist” or “ML Engineer” role. As a Member of Technical Staff, Applied ML, you will: Work directly with enterprise customers on problems that push LLMs to their limits. You’ll rapidly understand customer domains, design custom LLM solutions, and deliver production-ready models that solve high-value, real-world problems. Train and customize frontier models — not just use APIs. You’ll leverage Cohere’s full stack: CPT, post-training, retrieval + agent integrations, model evaluations, and SOTA modeling techniques. Influence the capabilities of Cohere’s foundation models. Techniques, datasets, evaluations, and insights you develop for customers will directly shape the next generation of Cohere’s frontier models. Operate with an early-startup level of ownership inside a frontier-model company. This role combines the breadth of an early-stage CTO with the infrastructure and scale of a deep-learning lab. Wear multiple hats, set a high technical bar, and define what Applied ML at Cohere becomes. Few roles in the industry combine application, research, customer-facing engineeri
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role Is Different This is not a typical “Applied Scientist” or “ML Engineer” role. As a Member of Technical Staff, Applied ML, you will: Work directly with enterprise customers on problems that push LLMs to their limits. You’ll rapidly understand customer domains, design custom LLM solutions, and deliver production-ready models that solve high-value, real-world problems. Train and customize frontier models — not just use APIs. You’ll leverage Cohere’s full stack: CPT, post-training, retrieval + agent integrations, model evaluations, and SOTA modeling techniques. Influence the capabilities of Cohere’s foundation models. Techniques, datasets, evaluations, and insights you develop for customers will directly shape the next generation of Cohere’s frontier models. Operate with an early-startup level of ownership inside a frontier-model company. This role combines the breadth of an early-stage CTO with the infrastructure and scale of a deep-learning lab. Wear multiple hats, set a high technical bar, and define what Applied ML at Cohere becomes. Few roles in the industry combine application, research, customer-facing engineeri
Become a part of our caring community We are seeking a highly experienced Lead Full Stack Engineer to drive engineering excellence across front-end, back-end, API, and data layers. This role is responsible for setting technical direction, ensuring consistent delivery across teams, and aligning engineering solutions with business priorities. The ideal candidate will be a strong technical leader who thrives in complex environments, drives modernization initiatives, and develops high-performing engineering talent. This role offers the opportunity to shape the future of our engineering landscape, influence enterprise-scale decisions, and lead transformation initiatives that directly impact customer experience and business growth. Location : Louisville, KY or Dallas, TX (Work At Home with occasional office visits) Key Responsibilities Engineering Leadership & Strategy Lead full stack engineering strategy and execution across UI, APIs, and data platforms Define architectural direction and ensure high standards for code quality, scalability, and maintainability Select frameworks, languages, and platforms across front-end, back-end, APIs, and data layers Approve architectural patterns (monolith vs. microservices, cloud strategy, and integrations) Identify opportunities to refactor, modernize, or retire legacy systems Delivery Execution & Prioritization Drive consistent delivery across multiple engineering teams through: Sprint planning and execution Dependency management Risk mitigation Removal of technical and organizational blockers Prioritize feature development, platf
About the team The OpenAI for Government team partners with federal, state, local, defense, national security, and international public-sector organizations to securely and responsibly adopt frontier AI, strengthen public services, and deliver meaningful mission impact. About the role OpenAI is seeking a strategic and deeply technical leader to serve as the Head of Government Technical Success. This leader will oversee the technical functions across the government customer lifecycle, spanning pre-sales engagement, prototype-to-production delivery, and post-sales adoption and value realization. You will define and operate a unified technical success strategy across federal civilian, defense and national security, state and local, international public sector, and industry partners. Your mission is to help customers identify the highest value applications of OpenAI’s technology, navigate the technical and organizational requirements of government environments, move those applications into production, and scale adoption in ways that deliver measurable mission impact. This role combines organizational leadership, technical judgment, executive customer engagement, operating rigor, and product influence. You will partner closely with Government Sales, Product, Engineering, Research, Security, Legal, Global Affairs and Policy, and other teams to ensure a seamless customer experience for governments in the United States and around the world. This role is based in Washington, DC. We offer relocation support to new employees. In this role, you will Set and continuously refine the strategy, operating model, and priorities for Government Technical Success, aligning the organization to OpenAI’s broader objectives and the distinct needs of government customers. Build and lead an organization of technical success personnel, including hiring, organizational design, manager development, career growth, and high standards for technical and customer-facing excellence. Create a seamless
Datadog's Forward Deployed Engineering function is in an active growth phase, and the FDE Lead will play a central role in shaping what comes next. Working in close partnership with the Head of Datadog for Startups and Forward Deployed Engineers, and the existing FDE team, you will help define and expand the FDE framework, build out the structures and processes that allow the team to operate at scale, and extend the program's reach well beyond any single customer segment. This role sits at the rare intersection of sales, execution, program design, and hands-on engineering leadership. You are part field technical leader, part program architect, and part cross-functional connector. You will help determine what the FDE motion looks like at Datadog, contribute to its playbook, and push the boundaries of what the team can deliver. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Evolve and scale the FDE operating model end to end: engagement intake, scoping, sprint delivery, handoff back to account teams, and the success metrics (time-to-value, adoption lift, ARR influence, NPS) and reporting infrastructure that support them. Build out a catalog of FDE offerings spanning observability quickstarts, custom integration development, LLM/AI observability accelerators, CI/CD pipeline instrumentation, and cost optimization deep dives, and contribute to their pricing and business models, including free-to-paid conversion plays, paid deployment packages, and post-deployment success motions. Capture product and feature gaps uncovered during deployments, translate them into structured prioritized briefs, and partner with PM and Engineering to strengthen the field-feedback channel that informs roadmap decisions on a regular cadence. Hire, onboard, an
From $196.8K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. WHY SAFETY? At Roblox, we strive to connect a billion people with optimism and civility, and the Safety organization's mission is to become the leader in civil immersive online communities. We systematically and proactively detect, remove, and prevent problematic content and behavior, and we make Roblox accounts secure and free from compromise. We also keep the platform compliant for changing regulations and growth markets. We cover a broad area of the tech spectrum, including machine learning, experimentation, automation, highly scalable distributed backend systems, detection workflows, and AI-powered text filters. Aligned and partnering with product teams, we use this tool belt to discover new opportunities, influence and shape the product roadmap and prioritization, build safety products, and measure the impact on our community of users and developers. In doing so, we keep Roblox safe, civil, and inclusive, and we foster positive relationships between people around the world. WHY CONTENT SUITABILITY? Join the Content Suitability team and play a pivotal role in shaping the future of content on Roblox. Our team is at the forefront of building tools and systems that enable creators to launc
$200K – $260K/yr
The Forward Deployed Engineer, Staff (FDE) is a high impact technical leader responsible for translating the immense power of Talkdesk's Agentic AI platform into transformative, production grade solutions for our strategic enterprise customers. You are a unique blend of a highly experienced software engineer, a technical project lead, and a hands-on builder. You own the technical execution for deployments, mitigate technical risks, and drive successful customer outcomes across assigned projects. This role is for the experienced builder who excels in high stakes, customer facing environments and is passionate about defining the future of Customer Experience Automation. Responsibilities Technical Project Leadership & Architecture: Act as the technical lead for complex deployments. Design and implement the architectural blueprint for AI agent solutions, manage cross-system dependencies, and ensure designs meet stringent enterprise standards for security and scale. Hands-On Engineering & Delivery: Write production-grade code and leverage Talkdesk and 3rd-party APIs/SDKs to design, build, test, and deploy AI agents. Drive the execution from prototype through to production deployment, ensuring technical quality. Technical Consultation & Alignment: Serve as a trusted technical expert for our AI solutions. Confidently address deep technical inquiries, mitigate technical risk, and build trust with customer Engineering Directors, and technical architects. Influence Product & Engineering Roadmap: Synthesize and codify deployment learnings into reusable solution patterns and tooling. Provide actionable feedback to core Product and Engineering teams to help inform future product direction. Technical Guidance: Mentor junior FDEs and technical specialists on best practices for complex AI architecture, production quality, and client-facing technical delivery. Who You Are We are looking for an autonomous, results-driven technical leader who thrives at the intersectio
NVIDIA is looking for a hands-on Solutions Architect Manager to lead a team of GPU, networking & software solution architects and engineers. Do you want to build and lead a group that designs, debugs, and deploys new AI hardware and software technologies into production in customer data centers? As part of the NVIDIA SA organization, you will drive people and technical leadership for end-to-end solutions deployments at some of NVIDIA's most strategic technology customers, while directly contributing to designs and deep-dive debugging and shaping our product roadmap with customer feedback. What you will be doing: Recruit & manage a team of solutions architects, system/network and software engineers focused on large-scale GPU and AI networking deployments. Set priorities, allocate resources, mentor, and ensure high-quality customer delivery across multiple concurrent projects - while remaining directly involved in key technical reviews, design decisions, and critical debug efforts. Provide deep subject-matter expertise in advanced GPU and network systems and serve as the senior technical point of contact for strategic customers. Personally lead and guide complex compute/network configuration and performance debugging, working side-by-side with your team to deliver performant, reliable clusters. Guide your team as they lead network / compute / software architecture discussions, and support server, network, and cluster bring-up, including on-site data center work where needed. Systematically collect and synthesize customer-specific requirements across your portfolio. Partner with GPU/Network Systems Engineering, Product Management, and Sales to influence roadmap priorities and packaging of reference designs and solutions. Demonstrate SME in advanced GPU & network systems and be a trusted technical advisor to NVIDIA's strategic customers. Bring customer-sp
About the Role As a Field CTO (Strategic Pursuits) , you will serve as a strategic bridge between our customers, go-to-market (GTM) teams, and product organization. You will partner closely with Sales, Technical Success, and Product to shape high-impact deals, guide customer architecture decisions, and influence our product roadmap based on real-world adoption and feedback. This is a highly cross-functional, externally facing leadership role for someone who combines deep technical expertise with strong business acumen and customer empathy. In this role, you will: Customer & Deal Strategy Partner with Sales, Product and Technical Success teams to support complex, high-value deals as a technical and strategic advisor. Translate customer business needs into scalable technical solutions and architectures. Engage with senior customer stakeholders (CTO/CIO/VP-level) to drive alignment on vision, roadmap, and adoption. Lead technical strategy discussions during key deal stages, including discovery, solution design, and executive presentations. Architecture & Implementation Guidance Guide customers on best practices for deploying and scaling AI-driven solutions in production. Provide architectural oversight across use cases such as LLM applications, integrations, data pipelines, and security. Act as a trusted advisor to ensure long-term success, not just short-term wins. Product & Feedback Loop Bring structured customer insights back to Product and Engineering teams to inform roadmap and prioritization. Identify gaps, opportunities, and emerging patterns from customer deployments. Influence product direction based on real-world usage, scalability needs, and enterprise requirements. GTM Strategy & Thought Leadership Help shape GTM strategies by identifying repeatable patterns across industries and customer segments. Develop scalable frameworks, reference architectures, and playbooks for broader field teams. Represent the company externally through customer en
Other cities to consider
More places hiring for this role
Get new inference technical lead jobs in United States by email
Daily job updates · Unsubscribe anytime