About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent, then turn that understanding into performance optimizations and models that project performance and capacity needs for future launches. About the Role In this role, you will model inference performance across application, model, and fleet layers with higher fidelity. You will build cost-to-serve estimates from microbenchmarks and create tools that help cross-functional teams reason about latency, capacity, utilization, and cost tradeoffs. In this role, you will Build and refine performance models that translate microbenchmark results into cost-to-serve estimates. Analyze inference workloads end to end across applications, models, and fleet infrastructure. Enhance tooling to identify bottlenecks across layers for latency and throughput. Partner with other teams to turn performance insights into concrete improvements and project how future changes affect inference. You might thrive in this role if you: Enjoy reasoning from first principles about distributed systems, model inference, and hardware efficiency. Are comfortable working across abstraction layers, from application behavior to kernels, accelerators, networking, and fleet scheduling. Have deep expertise with performance profiling, benchmarking, analysis, and optimization. Enjoy collaborating with engineering and research teams to improve real production systems. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve o
Jobs in United States
Production Tech in United States
1,337 active opportunities · Updated October 2026
Showing
15 jobs
Explore current production tech jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
$98K – $140K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role You'll own the quality bar for Notion AI products. You’ll work with product and engineering teams to build systems to define what “good” looks like, measure our progress, and drive changes to deliver reliable and high-quality AI experiences. Your work directly shapes how Notion's AI products behave for millions of users. This isn't a traditional software engineering role. It’s an art & science role . You won't spend your days writing code. Instead, you'll focus on understanding and shaping how our AI products behave through context engineering, designing evaluation systems, and analyzing data. This team sits in our AI engineering team, working directly with engineering, product, design, and data. This role is a unique blend of ops, strategy, and product thinking. Day to day, you'll live in production data, ship prompt fixes, run evals and, in effect, shape our quality strategy. As part of that you'll shape Notion's model strategy and work directly with frontier AI labs (OpenAI, Anthropic, Google) to evaluate and launch new models. We're looking for problem-seeking generalists interested in 0 → 1 : curious people wi
From $152K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role We’re looking for an AI Applications Engineer to help drive Notion’s business transformation efforts. In this role, you’ll be a strategic partner to our internal stakeholders (primarily GTM, Finance and People teams) and deliver and scale creative AI-driven solutions to multi-faceted problems with measurable business impact. You’ll also build reusable components, evaluation patterns, and operational guardrails that make AI delivery repeatable across teams. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You’ll Achieve Work with stakeholders to discover opportunities from ambiguous problem statements, translate them into scoped solutions, and drive iterative releases from idea to adoption Build and ship end‑to‑end AI solutions—from problem framing through data readiness, modeling, evaluation, and production rollout Establish evaluation and production-readiness patterns (met
From $299K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role We’re rolling out Go support at production scale at Notion, and we need an owner who can make it durable. You’ll lead the work to turn Go into a fully supported, well-operated platform: reliable and scalable service patterns, paved paths for our tooling stack, and the guardrails that make building in Go feel fast and safe. This role matters because our next wave of AI and agent-driven products will require backend services where Node won’t always be the right fit, and the platform decisions we make now will compound for years. While Go is the core focus, this role sits within Developer Experience and will regularly tackle other high-leverage engineering productivity challenges, developer experience ranging from AI-assisted development workflows and remote agent environments to CI performance, deployments, and reliability tooling. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days.
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. About Fern, a Postman Company Fern helps software companies build a world-class API experience. Our customers include industry leaders like Nvidia, Square, and Twilio, as well as fast-growing AI companies like ElevenLabs and OpenRouter. In the next year, the majority of API integrations will be implemented by AI agents. Agents don’t read marketing pages. They need strongly typed schemas, structured endpoints, deterministic contracts, and documentation that can be fed into a context window. Our team is small and talent-dense. We’re a team of builders from Google, Palantir, Amazon, and Uber, working together in New York City. About the Role As a Senior Software Engineer at Fern, you’ll build APIs, scale AI infrastructure, and design developer experiences that reach millions of people. Scale infrastructure to keep up with growth: Work on AI systems deployed on Vercel, AWS, and Turbopuffer that must be performant, reliable, and secure under real-world load. Stay close to customers: We work directly with API teams at companies like Nvidia, Twilio, and Square. You’ll debug real production edge cases, shape APIs ba
From $180K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Millions of people use Notion — and this number is increasing every day. That means millions of people trust us to deliver a fast, reliable, and secure experience, and we value this more than anything. We want to keep earning trust, while also continuing to amaze our users with the tools they can build in Notion. The AI Platform team is responsible for building the shared foundations that let Notion ship AI products quickly and operate them safely at scale. You’ll join a team of talented engineers focused on making speed and quality compatible: reliability and availability through provider changes, quality and correctness systems like evals and release gates, observability that makes failures explainable, and shared primitives for model integrations, context management, long-running actions, and cost/performance tradeoffs. Notion’s AI platform is vital to helping product teams move faster with production-grade guardrails as models, providers, and AI capabilities rapidly evolve. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days)
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity (Summer 2026 AI Internship - Applications Open Now) We're seeking an AI Engineer Intern to work alongside our AI team on large-scale AI and Agentic systems from data pipeline to production deployment. This role is scoped for someone with foundational experience who wants to deepen it: you'll own discrete pieces of real systems under the mentorship of senior engineers, not shadow work or isolated coursework-style projects. What You'll Do You’ll work directly with the AI team, taking responsibility for well-scoped pieces of real systems, with mentorship from senior engineers. Benchmarks & Evaluation Contribute to APIFlow-Bench , our open-source benchmark for real API-development work: design and review benchmark tasks and their mock API environments, extend the evaluation harness and task-generation pipeline in Python, and help maintain the public multi-model leaderboard with statistical confidence intervals. Help build a new action-level AI safety benchmark: instead of grading what a model says, it scores what an agent actually does inside a simulated enterprise API environment. You’ll work on scenari
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Cortex Apps team is building the future of AI for enterprise data. This role focuses on the backend infrastructure that powers our flagship products like Snowflake Intelligence , Cortex Agents and Search making agentic AI fast, reliable, scalable and secure at the enterprise level. You won’t just be using AI tools; you will be building the high-performance systems that orchestrate them. You’ll own and influence the architecture for agent execution environments, high-throughput context retrieval, or the ecosystem that allows our customers to iterate and launch agents in production. What you will do in this role: Architect Agentic Runtimes: Build and scale the orchestration engines that execute complex agentic workflows, ensuring low-latency tool execution and robust state management. Scale Context Engineering Infra: Design high-performance systems for RAG (Retrieval-Augmented Generation), including vector database integration, scalable and efficient search indexing, query processing, and result ranking, semantic caching, and automated metadata extraction. Build the "Evals Engine": Develop the automated infrastructure required to run massive-scale golden set simulations, error analysis pipelines, and "hillclimbing" experiments. Productionize AI Workflows: Collaborate with
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our team builds and operates Snowflake's Streaming Platform — the core infrastructure responsible for bringing data into Snowflake continuously and at scale. This includes Snowpipe Streaming and Datastream, services used by large set of enterprise customers to move mission-critical data in real time. Data is the fuel powering the new enterprise AI world, and we are the team responsible for getting it there reliably, continuously, and fast. Responsibilities Design, build, and maintain core components of Snowflake's streaming ingestion platform Improve service reliability, scalability, and latency under high-throughput production workloads Debug and resolve incidents in a distributed, multi-tenant cloud service Contribute to the design and evolution of streaming APIs, SDKs, and server-side protocols Write comprehensive tests including unit, integration, and chaos/fault-injection scenarios Collaborate with partner teams (storage, query, compute) on cross-cutting platform concerns Participate in code reviews, on-call rotations, and architecture discussions Required Qualifications 3–5 years of software engineering experience on large-scale distributed systems or cloud services Strong proficiency in Java or C++ Deep understanding of distributed systems concepts: consistency, faul
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lake using open formats like Apache Iceberg, delivering deep correlation and long-term analytics at dramatically lower cost. A dynamic Knowledge Graph and chat-based AI SRE provide rich context and guided workflows so teams can move from detection to root cause and resolution significantly faster. The Infrastructure team at Observe by Snowflake is responsible for building, scaling, and operating the development and production environments that power our observability platform. We are a small, highly collaborative team with a broad scope, focused on delivering reliable infrastructure while continuously improving the systems that support our engineers and customers. What You’ll Do Design, build, and operate scalable cloud infrastructure in AWS supporting a high-scale observability platform. Improve system reliability, performance, and operational visibility across development and production environments. Develop and maintain CI/CD pipelines and internal tooling to improve developer productivity and deployment safety. Identify and mitigate security risks, and help maintain intern
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause of production issue and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world’s leading data platforms. The Team On the AI Backend team at Observe by Snowflake you'll be at the forefront of how AI is reshaping the way engineering teams operate; building the platform that powers intelligent, automated workflows across observability and beyond. Our team is small, and moves fast, with real ownership over hard problems that span agentic APIs, real-time pipelines, and AI quality. You'll work alongside talented engineers across ML, product, and
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause of production issue and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world’s leading data platforms. In this role you will: Develop interactive, data-rich user interfaces using React, TypeScript, and Vega, with a focus on integrating LLM-driven features (e.g., natural language querying, generative UI, and AI-assisted data storytelling). Lead the end-to-end delivery of substantial product features, ensuring AI outputs are presented with high reliability and low latency. Work closely with PMs, UX designers, and AI/ML engineers to bridge t
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Product Data Science team is looking for a Full-stack Senior Data Scientist to come aboard and be part of Snowflake’s most critical initiatives. In this role, you will work closely with our Product and Engineering teams on everything from core operations, to innovative AI/ML tools, to our fastest-growing new and experimental features. You will work on long-running analytical initiatives yielding substantial product enhancements. This is a strategic, high-impact role that will help shape the future of Snowflake products and services. PLEASE NOTE: The position level is determined during the interview process and is influenced by various factors, including experience, educational attainment, skill level requirements, and overall interview performance. AS A SENIOR FULL STACK DATA SCIENTIST AT SNOWFLAKE YOU WILL: Be a proactive partner to Product Management and Engineering to shape feature roadmaps and help design metrics and analytical workflows to evaluate their success and effectiveness. Build efficient data models and high quality production pipelines and partner with Engineering for all necessary telemetry. Build scalable and extensible analytics/ML/causal inference frameworks to uncover feature usage patterns, subtle issues with the system, potential performance enhanc
Work Flexibility: Onsite Como Plastic Processing Operator III , serás responsable de operar, configurar y monitorear maquinaria de producción, asegurando el cumplimiento de los estándares de calidad, seguridad y productividad establecidos. ¿Qué harás? Opera equipos de procesamiento de plásticos, incluyendo máquinas automáticas y semiautomáticas, para cumplir con los objetivos de producción y calidad. Configura, inicia y monitorea equipos de manufactura siguiendo hojas de proceso, procedimientos establecidos e instrucciones operativas. Realiza cambios de herramientas, moldes y ajustes de configuración para garantizar las dimensiones y especificaciones requeridas del producto. Controla parámetros del proceso, como temperatura, presión y velocidad, y ejecuta ajustes rutinarios para mantener una producción consistente. Inspecciona piezas mediante controles visuales e instrumentos de medición para verificar el cumplimiento de estándares dimensionales y de calidad. Registra con precisión datos de producción, desperdicio, ajustes de proceso y actividades básicas de mantenimiento. Ejecuta mantenimiento preventivo básico y reparaciones menores, escalando incidencias complejas al equipo de mantenimiento correspondiente. Manipula y empaqueta componentes terminados, manteniendo el cumplimiento de los procedimientos de seguridad, calidad y uso de equipos de protección personal. ¿Qué necesitas? Requerido: Diploma de cuarto año completo o equivalente. Experiencia mínima de 2 años en operación de maquinaria de manufactura
Work Flexibility: Onsite Como Plastic Processing Operator III , serás responsable de operar, configurar y monitorear maquinaria de producción, asegurando el cumplimiento de los estándares de calidad, seguridad y productividad establecidos. ¿Qué harás? Opera equipos de procesamiento de plásticos, incluyendo máquinas automáticas y semiautomáticas, para cumplir con los objetivos de producción y calidad. Configura, inicia y monitorea equipos de manufactura siguiendo hojas de proceso, procedimientos establecidos e instrucciones operativas. Realiza cambios de herramientas, moldes y ajustes de configuración para garantizar las dimensiones y especificaciones requeridas del producto. Controla parámetros del proceso, como temperatura, presión y velocidad, y ejecuta ajustes rutinarios para mantener una producción consistente. Inspecciona piezas mediante controles visuales e instrumentos de medición para verificar el cumplimiento de estándares dimensionales y de calidad. Registra con precisión datos de producción, desperdicio, ajustes de proceso y actividades básicas de mantenimiento. Ejecuta mantenimiento preventivo básico y reparaciones menores, escalando incidencias complejas al equipo de mantenimiento correspondiente. Manipula y empaqueta componentes terminados, manteniendo el cumplimiento de los procedimientos de seguridad, calidad y uso de equipos de protección personal. ¿Qué necesitas? Requerido: Diploma de cuarto año completo o equivalente. Experiencia mínima de 2 años en operación de maquinaria de manufactura o procesamiento industrial. Conocimiento básico de interpretación de planos, diagramas o esquemas técnicos. Capacidad para utilizar herrami
Other cities to consider
More places hiring for this role
Get new production tech jobs in United States by email
Daily job updates · Unsubscribe anytime