About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders About the Role We're hiring the first Account Managers at Modal. You'll report to the Regional Director of Account Management and be a founding member of the team. This function does not exist yet. There is no playbook, no territory map, no established motion. You'll own a book of business from day one and build the motion at the same time — from fast-moving AI startups to large enterprise teams running critical infrastructure on Modal. This is a commercial role with a revenue target. You'll be measured on retention and expansion across your accounts. While you won't be delivering the technical recommendations and implementation, the work is technical by nature. Our customers are engineers running GPU workloads, inference, and batch jobs in production, and you need to hold your own in those conversations. The profile we're hiring is a technical account manager. You've worked at companies that are deepl
Jobs in United States
Production Tech in United States
1,337 active opportunities · Updated October 2026
Showing
15 jobs
Explore current production tech jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are seeking experienced professionals with a strong background in Artificial Intelligence, Machine Learning, and Cloud Architecture to join our Services Delivery team to help create exciting new offerings and capabilities for our customers! In this strategic role, you will help customers expand their use of the Snowflake Data Cloud to bring AI/ML pipelines from ideation to full production. Leveraging Snowflake’s native features and extensive partner ecosystem, you will advise clients on best practices for scaling production-ready workloads. You will design tailored AI/ML solutions, coordinate closely with customer teams and Systems Integrators, and provide the technical leadership and oversight needed to ensure successful outcomes. AS A PRINCIPAL SOLUTIONS ARCHITECT AT SNOWFLAKE, YOU WILL: Be a technical expert on all aspects of Snowflake in relation to the AI/ML workload and provide customers with best practices given Snowflakes technology stack. Work with customers to understand their AI/ML use case, discover key requirements, and architect a Snowflake-centric solution to be delivered by Services Delivery. Understand how to build, deploy and AI and ML pipelines using Snowflake features and/or Snowflake ecosystem based on customer requirements. Work hands-on where neede
NVIDIA is seeking a Senior Technical Program Manager to join the CSP Engagements team, focused on deep technical engagement with hyperscale cloud service providers for NVIDIA’s next‑generation datacenter systems such as Vera Rubin NVL72. This role is intended for experienced systems and embedded software leaders—including software engineering managers, technical leads, or senior architects—who have led datacenter server and platform software programs and can operate as a trusted technical partner to hyperscale CSP engineering teams. As a member of the CSP Engagements team, you will act as the primary technical engagement leader between NVIDIA’s system software organizations and CSP platform, system software, and AI teams, ensuring alignment, readiness, and successful large‑scale deployment of NVIDIA‑based datacenter solutions. What you will be doing: Lead deep technical engagements with hyperscale CSPs as the primary NVIDIA point of contact for system software, firmware, and platform readiness for NVIDIA datacenter products. Partner directly with CSP system software, firmware, and infrastructure engineering leaders to align on software architecture, bring‑up plans, deployment readiness, and production requirements for NVIDIA‑based server and rack‑scale platforms. Represent CSP technical priorities internally, advocating for customer requirements and tradeoffs across NVIDIA’s system software, firmware, hardware, silicon, and product teams are aligned to customer needs, timelines, and constraints. Own the end‑to‑end CSP engagement lifecycle, from early technical alignment and pre‑production readiness through large‑scale deployment, escalation management, and sustained production support. Drive bi‑directional technical communication: translating CSP system‑level requirements into actionable focus areas for NVIDIA engineering teams, while clearly communicating N
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. We are the Data Foundation & AI team within Plaid’s Data organization. Our mission is to build the shared ML and AI infrastructure that powers intelligent capabilities across Plaid’s product suite. We develop the foundational systems, models, and data assets that transform Plaid’s unique financial network data into scalable, general-purpose representations that teams across the company can leverage. Our work spans the full ML lifecycle — from large-scale data curation and model pretraining to production serving, evaluation, and monitoring. As part of the team, you’ll work at the intersection of machine learning infrastructure, applied AI, and distributed systems, helping establish the core AI platform that enables innovation across Plaid. As a Staff Machine Learning Engineer, you will lead the technical strategy and development of Plaid’s foundation models, driving key decisions across pretraining objectives, model architecture, and fine-tuning approaches that power a wide range of downstream product applications. You will serve as the technical lead for the full machine learning lifecycle, overseeing everything from data curation and experimentation to production deployment, feature management,
From $243.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Infrastructure Compute Site Reliability Engineering mission is to own and manage the successful operation of our underlying cell infrastructure system, along with elements of service discovery, secrets management and related software layers. We’re looking for a skilled Senior Site Reliability Engineer with strong programming skills to help us build Roblox's private cloud, productionize our growing Kubernetes-based infrastructure, and institute reliability best practices across the Roblox Compute team. You will: Design and Develop systems & libraries that promote fault-tolerance and resilience, automate much of the management and lifecycle of our clusters, and ensure systems are observable. Promote and Institute reliability best practices across the Infra Compute group, drive common reliability initiatives. Provides collaborative technical reviews and operational guidance to strengthen system reliability. Build, Automate and Standardize process automation to create a "golden path" of tooling and platform support that powers the fundamental Roblox ecosystem. Create Tooling that provides production guardrails, by evaluating release candidate capacity with load testing tooling before de
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are seeking experienced professionals with a strong background in Artificial Intelligence, Machine Learning, and Cloud Architecture to join our Services Delivery team to help create exciting new offerings and capabilities for our customers! In this strategic role, you will help customers expand their use of the Snowflake Data Cloud to bring AI/ML pipelines from ideation to full production. Leveraging Snowflake’s native features and extensive partner ecosystem, you will advise clients on best practices for scaling production-ready workloads. You will design tailored AI/ML solutions, coordinate closely with customer teams and Systems Integrators, and provide the technical leadership and oversight needed to ensure successful outcomes. AS A SR. SOLUTIONS ARCHITECT AT SNOWFLAKE, YOU WILL: Be a technical expert on all aspects of Snowflake in relation to the AI/ML workload and provide customers with best practices given Snowflakes technology stack. Work with customers to understand their AI/ML use case, discover key requirements, and architect a Snowflake-centric solution to be delivered by Services Delivery. Understand how to build, deploy and AI and ML pipelines using Snowflake features and/or Snowflake ecosystem based on customer requirements. Work hands-on where needed usin
$1.2M – $1.3M/yr
About us EVERY™ is a leading VC-backed food tech ingredient company and market leader using precision fermentation to create animal proteins without the animal for the global food and beverage industry. EVERY™ is a team of passionate change-makers who are reimagining the factory farm model with a kinder, more sustainable alternative. Leveraging precision fermentation to produce hyper-functional and one-to-one replacement proteins from microorganisms, EVERY™ is on a mission to decouple the world’s proteins from the animals that make them. We are a passionate, determined (and fun!) team with a vital objective, and we're on the lookout for like-minded people to join our mission. For more information, visit www.every.com The Downstream Process Engineer II will be an integral member of our downstream process development team. You will use experimentation to optimize Every’s production process and then see the results of your changes in action at pilot and commercial-scale biomanufacturing sites. This is an excellent opportunity for someone with laboratory and tech transfer experience, who wants to make an impact at scale. What you'll Accomplish Optimize the Every downstream process via an iterative cycle. Improvements are developed in the laboratory, scaled up to an external pilot plant, and learnings are taken back to the lab for further troubleshooting and improvement. Perform various unit operations at the Every HQ such as TFF, depth filtration, chromatography, spray drying and other purification/separation processes. Develop and review tech transfer documentation to ensure successful scale up trials. Travel to external pilot plants to review the scale up of novel processes. Coordinate with third parties such as pilot scale equipment vendors to arrange internal and external trials. Partner with Every scientists and engineers to bring their bench ideas to pilot scale. Analyze results and report data to enable appropriate interpretation and
$100K – $500K/yr
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. We’re looking for a Staff Forward Deployed Engineer who’s excited to build with the engineers using the AI computers Tenstorrent makes. You will create continuity between customers, engineering, and AI inference service products. This is an engineering role first: you contribute production code, operate deployments, and you can explain a trade-off to customer leadership as clearly as to core engineering teams. This is a high-autonomy role with direct customer impact. This role is remote, based out of North America, with preference near one of our main hubs: Santa Clara, CA; Austin, TX; or Toronto, ON. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are You understand how accelerator compute, memory, and networking topology constrain AI workloads, and don't treat hardware as a black box. You're an early adopter of AI for your work from coding to building agentic workflows that multiply your impact. You work directly with customers to understand their challenges and provide effective solutions. You are comfortable debugging across the full inference stack: from failing requests, through the serving layer, down to OOMs or kernel dispatch if n
About the Team The Finance Platform & Technology team builds and scales the systems and data architecture that power OpenAI’s core financial operations. We enable business agility, compliance, and operational excellence across procure-to-pay, quote-to-cash, supply chain, financial planning, and asset management. We partner with Procurement, Accounting, Tax, Legal, Security, Data, and Engineering to modernize workflows through thoughtful platform design, reliable integrations, scalable automation, and trusted data. About the Role As a Business Systems Lead for Procure-to-Pay, you will be a hands-on engineer who designs, builds, and operates the integrations and first-party applications that power OpenAI’s procurement workflows. You will translate business needs into secure, scalable software, APIs, data flows, and automation across Oracle Fusion, Zip, and connected platforms. You will build the future of buying at OpenAI using OpenAI’s own technology, from guided intake and approval experiences to supplier onboarding, purchasing, receiving, invoicing, and downstream financial data flows. You will own the technical roadmap and support model for these capabilities, improving today’s platforms while deciding where to integrate, configure, or build as OpenAI scales. Your core strength will be software and integration engineering. You will personally write code, troubleshoot cross-system failures, and take solutions through testing, deployment, and production support. You will also make targeted functional configurations in procurement platforms and partner with functional specialists on deeper process and module design. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and operate integrations across Oracle Fusion, Zip, and connected systems using APIs, events, messaging, and batch interfaces where appropriate. Build first-party
$190.4K – $285.6K/yr
Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Lead the technical design and architecture of major platform initiatives, author design documents and build consensus across engineering teams. Define technical roadmaps for complex, multi-quarter projects that span multiple teams. Make critical architectural decisions for company documentation infrastructure, balancing scalability, reliability, and developer experience. Evaluate and set direction for integrating emerging technologies, including AI/LLM capabilities, into company documentation platforms and authoring tools. Establish and evolve engineering standards, best practices and technical guidelines for the team and broader organization. Partner with engineering teams across the company to understand documentation needs and design integrated solutions. Design, build and maintain scalable, reliable and performant services and systems. Contribute high-quality code across the full stack and navigate codebases with different languages and tools. Debug and resolve complex production issues and improve system reliability. Take ownership of system health and incident response. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Software Engineering, Engineering, or a related field, plus four (4) years of experience in Software Engineering. Must have four (4) years of experience in each of the following: - Working in a full stack environment with a foc
About the Team The Applied AI Engineering team partners closely with customers to help them move from experimentation to production with OpenAI’s technologies. We act as trusted technical advisors, working across customer strategy, architecture, deployment, and adoption to help organizations realize meaningful impact from frontier AI. The Startups segment serves fast-moving, high-growth companies that are often building new products, workflows, and businesses directly on top of AI. These customers move quickly, operate with high ambiguity, and expect practical, creative, and technically rigorous partnership. About the Role We are looking for an Applied AI Engineering Manager, Startups to lead and scale the Startups Applied AI Engineering motion. This team helps high-growth startups move quickly from experimentation to production, unlock meaningful usage, and build durable technical partnerships with OpenAI. This leader will operate in a high-velocity customer segment where founders, CTOs, and technical teams expect speed, judgment, and hands-on problem-solving. They will balance team leadership, technical depth, customer prioritization, and cross-functional influence across Sales, Product, Engineering, Research, and broader go-to-market teams. In this role, you will define how OpenAI supports startup customers at scale: identifying where deep technical engagement can unlock outsized impact, building repeatable deployment mechanisms, and ensuring the team can serve a broad and dynamic customer base without losing quality or strategic focus. In this role, you will: Craft and continuously refine the strategic vision and operating model for the Startups Applied AI Engineering team, aligning it with OpenAI’s broader company objectives and the evolving needs of high-growth startup customers. Lead, mentor, and grow a team of high-performing technical ICs supporting startup customers across AI-native, developer-led, and product-led companies. Help startups move from early e
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our FDE team and work directly with some of the world's largest organizations to turn their most ambitious ideas into production applications on Replit. As a Forward Deployed Engineer, you'll partner closely with customers to understand their technical and business needs, architect solutions, and build the integrations and applications required to deploy Replit successfully within complex enterprise environments. This is a deeply technical and hands-on role. You won’t just advise customers on what to build—you’ll build alongside them. You’ll take applications from initial idea and prototype through deployment and production, navigate complex enterprise environments, and solve the technical challenges that emerge when AI-powered software development meets real-world infrastructure, data, security, and organizational constraints. You will: Build with Customers: Embed with strategic enterprise customers to design and build high-impact applications and AI-powered workflows on Replit, taking projects from initial concept through production deployment. Architect Enterprise Solutions: Design secure, scalable architectures that connect Replit with customers’ existing systems, data, APIs, identity providers, and infrastructure. Integrate with the customer's stack: SSO, data warehouses, internal APIs, and SaaS systems. Own Technical Deployments: Serve as the technical owner for complex enterprise implementations, identifying blockers, debugging issues, and driving projects through to successful production adoption. Bridge Customers and Product: Develop a deep understanding of how enterprises use Replit and translate field insights, technical constraints, and recurring customer needs into actionable feedback
Work Flexibility: Onsite Stryker is seeking a Staff Advanced Manufacturing, Automation/Software Engineer to join our Advanced Operations team, supporting the Medical Division - Acute Care Business Unit. In this role, you will lead the design, development, and deployment of advanced automation and production test systems for new product introductions. This role is critical to ensuring reliable, scalable, and compliant manufacturing for patient support and patient environment products. You will serve as both a technical owner and a supplier-facing leader, architecting systems internally while guiding external partners to deliver high‑quality automation solutions. You will have the opportunity to work on the design transfer of new products from research through development and into production. This is a hybrid role based out of Portage, MI. The team works onsite 4-5 days per week to support collaboration and project needs. What you will do: Automation System Architecture & Development Lead the full lifecycle of industrial automation and test systems from requirements, architecture, and design through implementation, validation, and release. Define system-level requirements encompassing mechanical, electrical, software, controls, and safety considerations. Develop and integrate control software, embedded interfaces, test sequences, and operator interfaces (HMI/SCADA). Ensure robust performance, maintainability, reliability, and alignment with design intent and manufacturing needs. Troubleshoot complex processes, software, and equipment issues; optimize system performance and uptime. Supplier & Equipment Vendor Leadership Manage automation and equipment suppliers, including capability assessments, technical reviews, process monitoring, and on‑site visits. Create clear, comprehensive
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As the Senior Director of Solutions Architecture for the Americas at Cohere, you will own the US and Canada commercial Solutions Architecture function. You will lead the team that turns enterprise interest in agentic AI into deployed, production systems, and you will be accountable for the technical win in the most competitive AI market in the world. The United States and Canada are our largest commercial opportunity, and this seat owns how we win them. You will take an established, distributed team of strong technical people and raise what it can do — setting the bar and establishing the operating rhythm that lets a team of generalists run consistent, industry-fluent plays at enterprise scale. You will set direction for the function, sit on the Solution Architecture leadership team alongside the regional leaders for EMEA and Asia Pacific, and contribute to company-wide decisions with your peers across Sales, Product and Engineering. In this role, you will: Lead and Scale the Organization: Build, coach and develop a high-performing Solutions Architecture organization across the US and Canada, and grow the senior technical talent
$220K – $450K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role AI and machine learning are reshaping how developers debug, monitor, and ship software, and Sentry is uniquely positioned to lead that shift. We sit on a novel and massive dataset of real production errors, spans, and logs from tens of thousands of engineering organizations — the kind of signal that makes ML genuinely useful, whether it's a clustering model that groups related issues, a ranking system that surfaces the right alert at the right time, or an agent that proposes a fix. We're looking for an Engineering Manager to lead and grow our Machine Learning Engineering team. This team owns the full spectrum of ML at Sentry: classical techniques like clustering, ranking, anomaly detection, and embeddings that quietly power core product surfaces today, alongside the LLM-based and agentic systems shaping where the product is headed. You'll partner closely with product, design, and engineering leaders to decide where ML belongs in our products, what kind of ML actually fits the problem, and how we translate that work into experiences millions of developers rely on every day. In this role you will Set technical direction across the team's full ML surface area — from classical models for clustering, ranking, and anomaly detection to LLM-based and agentic systems — and make sharp calls about which approach fits each problem Define how the team evaluates and monitors ML systems in production, from offline metrics to online experimentation to model and agent observability Stay hands-on enough to review code and model designs, contribute to architecture discussions, and unblock engineers on complex ML problems Define
Other cities to consider
More places hiring for this role
Get new production tech jobs in United States by email
Daily job updates · Unsubscribe anytime