About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hiring a Developer Productivity Engineer to help scale the engineering systems, safeguards, and developer workflows that enable our teams to move quickly without compromising reliability or performance. This role sits at the intersection of developer experience, CI/CD infrastructure, release engineering, production readiness, and inference systems reliability. You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack. About the Role We’re looking for an autonomous, high-ownership engineer who cares deeply about making other engineers faster, safer, and more confident. A major focus of this role will be improving the tooling and infrastructure around deploy gates for inference engine images. These systems help ensure that every image released to production and research is correct, numerically sound, free of regressions, and performant across key metrics like time-to-first-token (TTFT) and time-between-tokens (TBT). You’ll help harden the systems that catch issues before they reach production, reduce noise from flaky or infrastructure-related test failures, and improve automation around triage, ownership, debugging, and escalation when failures occur. You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack. This is not generic internal tools work. The systems you build directly impact OpenAI’s ability to support new model launches, safely ship inference optimizations to the world, onboard new infrastructure providers, and operate one of the largest and most p
Jobs in United States
Service Tech in United States
6,891 active opportunities · Updated October 2026
Showing
15 jobs
Explore current service tech jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Hardware Health and Observability team owns the end-to-end health lifecycle of OpenAI’s global compute fleet. Our mission is to maximize healthy, usable compute across accelerator vendors, generations, cloud providers, and regions through reliable health signals, automated remediation, and scalable operational tooling. We build the systems that observe, detect, remediate, and verify hardware issues across GPUs, CPUs, networking, and platform infrastructure, enabling frontier model training and inference workloads to run reliably at hyperscale. We are the last line of defense for the success of OAI’s production and research workloads. About the Role On the Hardware Health and Observability team, you’ll build critical infrastructure that keeps OpenAI’s largest compute clusters healthy and operational at scale. Even small numbers of unhealthy systems can impact large-scale training and inference workloads. This team focuses on minimizing downtime, improving fleet efficiency, and ensuring compute resources remain continuously available to researchers and product teams. Engineers on this team own problems end-to-end, from defining health signals and debugging failures to building automated remediation systems that operate across millions of GPUs globally. In this role, you will: Define and maintain health signals across GPUs, CPUs, networking, and platform infrastructure. Build and evolve health checks that detect, remediate, and verify failures at scale. Ensure critical health checks execute with minimal latency to maximize workload uptime. Investigate hardware failures and system-level issues across large-scale compute environments. Own node lifecycle workflows including drain, quarantine, repair, RMA, and return-to-service processes. Build automation and tooling that enables global cluster management with minimal manual intervention. Partner with workload, reliability, and provider teams to integrate health signals into training and inference system
About the Team ChatGPT is a rapidly evolving system: new capabilities ship continuously, product surfaces change quickly, and usage patterns shift week-to-week. Supporting that pace requires infrastructure that can handle real production constraints—high concurrency, unpredictable traffic patterns, complex dependency graphs, and frequent change. The ChatGPT Infrastructure team builds and operates the platforms that enable fast iteration without compromising performance or reliability. We design shared systems, data paths, rollout mechanisms, and reliability guardrails that teams rely on to ship changes to ChatGPT at scale. We focus on high-leverage infrastructure: primitives and “golden paths” that incorporate operational lessons as defaults, so engineers don’t need to rediscover failure modes, latency pitfalls, or integration issues each time they build something new. About the Role We’re hiring Senior and Staff Engineers to design and build infrastructure systems that underlie ChatGPT and multiply the effectiveness of teams building user experiences. This is not a support-only role. It’s a platform-building role: you’ll define interfaces, develop core abstractions, and create tooling to make safe, fast iteration the norm. Your work will reduce friction, prevent regressions, improve performance, and ensure systems scale gracefully as the product grows. Where You Can Have Impact You might work on one or more of the following areas (without being restricted to any single area): Platform foundations & frameworks: Core libraries, service frameworks, and shared components that standardize system building, integration, and evolution. Scalability & performance primitives: Patterns and infrastructure that reduce tail latency, improve throughput, and keep costs predictable as demand increases. Reliability guardrails: Mechanisms that prevent outages by design—rate limiting, load shedding, dependency isolation, backpressure, safe fallbacks, and robust regression contr
From $299K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role We’re rolling out Go support at production scale at Notion, and we need an owner who can make it durable. You’ll lead the work to turn Go into a fully supported, well-operated platform: reliable and scalable service patterns, paved paths for our tooling stack, and the guardrails that make building in Go feel fast and safe. This role matters because our next wave of AI and agent-driven products will require backend services where Node won’t always be the right fit, and the platform decisions we make now will compound for years. While Go is the core focus, this role sits within Developer Experience and will regularly tackle other high-leverage engineering productivity challenges, developer experience ranging from AI-assisted development workflows and remote agent environments to CI performance, deployments, and reliability tooling. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days.
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. About The Team We're looking for a self-motivated team member who craves a challenge, is obsessed with achieving key business objectives, and wants to work on one of the most loved developer products in the world. Product Marketing is responsible for developing crisp, highly differentiated, and compelling positioning and messaging for Postman and its services. We enable the revenue team to tell stories that educate customers about what is possible when they choose Postman to manage every API and service across their organization. As a Senior or Staff Product Marketing Manager, you will have the opportunity to create compelling content to help customers, especially platform engineers, and developers understand their use cases and value propositions, and build the right marketing programs to drive awareness, engagement, service adoption, and retention. This role sits in San Francisco, CA only. What You’ll Do Act as a product marketing lead for Postman features and services, and take ownership of the strategy to drive awareness and market penetration of key capabilities Work with engineering teams to distill key functional
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. About The Team We're looking for a self-motivated team member who craves a challenge, is obsessed with achieving key business objectives, and wants to work on one of the most loved developer products in the world. Product Marketing is responsible for developing crisp, highly differentiated, and compelling positioning and messaging for Postman and its services. We enable the revenue team to tell stories that educate customers about what is possible when they choose Postman to manage every API and service across their organization. As a Principal Product Marketing Manager, you will have the opportunity to define and drive the marketing and go-to-market strategy for an industry leading service. You will be creating compelling content to help customers, especially platform engineers, and developers understand their use cases and value propositions, and building the right marketing programs to drive awareness, engagement, service adoption, and retention. This role sits in San Francisco, CA USA only. What You’ll Do Act as a product marketing lead for Postman features and services, and take ownership of the strategy to drive awarene
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. PRINCIPAL PRACTICE MANAGER LEAD. STRATEGIZE. TRANSFORM. JOIN US AS A PROFESSIONAL SERVICES PRACTICE MANAGER! Are you ready to own your impact at the forefront of the AI and data revolution ? As a Principal Practice Manager , you’ll be a key driver in transforming the US Federal government through cutting-edge AI-cloud data solutions . You’ll collaborate with top-tier clients, strategize with elite sales teams, and win and lead groundbreaking professional service engagements —all while working with the best minds in the industry! This role is completely remote but preference will be given to those in the DC-metro area. WHY THIS ROLE? Be a Market Leader: Own and execute the strategic vision for Professional Services across all of the US Federal Government for Defense and Civilian. Drive Meaningful Change: Work with clients to unlock the power of AI and data and fuel digital transformations. Collaborate with the Best: Work cross-functionally with top industry experts to win transformative services opportunities and shape the future of cloud data solutions. Grow your Business: Work proactively across accounts to create and mature new opportunities, close deals and grow your business. Crush Your Goals: Meet and exceed bookings, revenue, and consumption targets while delivering e
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Staff Data Scientist, Finance About the Team The Finance Data Science team builds the forecasting and decision systems that power Snowflake's financial planning, operating cadence, and long-term strategy. Our work informs executive decision-making, product and go-to-market priorities, resource allocation, pricing, and cross-functional decisions across Finance, Product, Sales, and Data Science. We are expanding a driver-based revenue modeling platform that translates product and workload activity into trusted financial outcomes. The program began with one product category and will scale a common modeling and publishing framework across Snowflake's product categories. The models are highly visible, refreshed frequently, and designed for self-service scenario planning and business reviews. The Role We are hiring a Staff Data Scientist to lead the next phase of Snowflake's driver-based revenue modeling program. This role is not just about building models. It is about creating reliable, explainable, production-grade decision systems that connect upstream business and product levers to revenue outcomes. You will own high-impact, open-ended problems spanning driver identification, revenue decomposition, leading indicators, cohort and use-case modeling, scenario analysis, and multi
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a Senior Software Engineer across our Identity, Data Security and Trust organization. These teams builds the security backbone of Snowflake — enabling customers to confidently bring their most sensitive data and workloads to the Data Cloud, empowering every Snowflake engineer to deliver the most secure platform at scale, and giving customers the tools to monitor, achieve, and maintain a strong security, governance, privacy, and compliance posture. As the agentic era accelerates, these teams are at the forefront of building AI-native security capabilities that protect the world's data. AS A SENIOR SOFTWARE ENGINEER AT SNOWFLAKE, YOU WILL: Design and implement critical AI security capabilities — including controlled, audited agent workflows, MCP server and client security, agent identity frameworks, and admin guardrails — that form the foundation for secure agentic enterprise adoption Build and evolve foundational identity and access management systems, including authentication protocols, SSO integrations, multi-factor authentication, OAuth/OIDC, SAML, SCIM, and fine-grained RBAC that scales to millions of objects and users Design and develop core security infrastructure spanning secret management, key management, service identity, and end-to-end encryption acro
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. LEAD. STRATEGIZE. TRANSFORM. JOIN US AS A PROFESSIONAL SERVICES PRACTICE MANAGER! Are you ready to own your impact at the forefront of the data revolution ? As a Services Practice Manager , you’ll be a key driver in transforming businesses through cutting-edge cloud data solutions . You’ll collaborate with top-tier clients, strategize with elite sales teams, and lead groundbreaking professional service engagements —all while working with the best minds in the industry! This role is completely remote and can be based anywhere in the Eastern Timezone. WHY THIS ROLE? Be a Market Leader: Own and execute a strategic vision for Professional Services. Drive Meaningful Change: Work with clients to unlock the power of data and fuel digital transformations. Lead from the Front: Manage high-impact projects, ensure seamless delivery, and drive revenue growth. Collaborate with the Best: Work cross-functionally with top industry experts to shape the future of cloud data solutions. Crush Your Goals: Meet and exceed bookings, utilization, and consumption targets while delivering exceptional customer experiences . WHAT YOU’LL DO: Spearhead services sales execution and project delivery success for key accounts. Engage C-level executives to create and execute services roadmaps that drive real
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Identity & Access management (IAM) team’s charter is to enable our customers to confidently bring their most sensitive data and workloads to Snowflake. We provides the authentication and authorization capabilities for customers to secure their Snowflake accounts. We are heavily focused on critical AI adoption and security capabilities like Snowflake Intelligence access control, MCP server and clients, Agent identity, Admin guardrails for agents etc. Our feature set includes capabilities like user management, secret-less authentication for both human and service users, SSO integration with numerous IdPs, MFA, OAuth and OIDC support for 3P applications, and RBAC for granular access control. Our systems are critical to customer trust and maintaining Snowflake’s security, reliability and performance. The team culture is very collaborative with ample opportunities for growth and mentorship from Principal engineers. AS A SENIOR SOFTWARE ENGINEER - IDENTITY & ACCESS MANAGEMENT, YOU WILL: Design and implement critical AI security capabilities for controlled, audited, restricted agent workflows, both inbound and outbound. Design and implement features that provide critical identity and access management capabilities, including integration with the next generation identit
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake’s Release Engineering team builds and operates the systems that safely deliver infrastructure, platform, and product changes to production at global scale. We own the release platforms, rollout orchestration, and safety mechanisms that allow engineering teams across Snowflake to ship quickly while minimizing operational risk. Our mission is to make production deployments fast, safe, self-service, increasingly autonomous, and augmented by AI-driven intelligence and automation. This role sits at the intersection of developer productivity, distributed systems reliability, and large-scale multi-cloud infrastructure orchestration. At Snowflake, Release Engineering is a platform engineering function focused on building the systems, abstractions, and automation that make software delivery safe, scalable, and efficient across the company. In this role, you will Design and build continuous deployment and rollout infrastructure that safely ships changes across Snowflake’s large-scale, multi-cloud production environment. Build and evolve platform capabilities for progressive delivery, including staged rollouts, canarying, automated health checks, rollback controls, and guardrails that reduce blast radius during production change events. Improve engineering velocity by removi
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our team builds and operates Snowflake's Streaming Platform — the core infrastructure responsible for bringing data into Snowflake continuously and at scale. This includes Snowpipe Streaming and Datastream, services used by large set of enterprise customers to move mission-critical data in real time. Data is the fuel powering the new enterprise AI world, and we are the team responsible for getting it there reliably, continuously, and fast. Responsibilities Design, build, and maintain core components of Snowflake's streaming ingestion platform Improve service reliability, scalability, and latency under high-throughput production workloads Debug and resolve incidents in a distributed, multi-tenant cloud service Contribute to the design and evolution of streaming APIs, SDKs, and server-side protocols Write comprehensive tests including unit, integration, and chaos/fault-injection scenarios Collaborate with partner teams (storage, query, compute) on cross-cutting platform concerns Participate in code reviews, on-call rotations, and architecture discussions Required Qualifications 3–5 years of software engineering experience on large-scale distributed systems or cloud services Strong proficiency in Java or C++ Deep understanding of distributed systems concepts: consistency, faul
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer — Cortex Training The Snowflake ML Platform team's mission is to let customers run their most demanding ML/AI workloads inside Snowflake. Cortex Training is our LLM post-training platform: it turns scarce, expensive GPU capacity into a simple, composable service, so customers can adapt open-weight foundation models to their own business problems while we handle the hard distributed-systems parts, including scheduling, orchestration, multi-node training and inference, fault tolerance, and throughput. The platform already runs post-training at scale. Under the hood, it decouples GPU computation from the training loop and exposes it as primitive APIs that compose into everything from SFT to full RL workflows. You'll work alongside a team that ships fast & sweats reliability and the researchers behind DeepSpeed. We're looking for an engineer who thrives in the ML infrastructure layer and brings a solid understanding of LLMs and post-training to help us scale and grow it. YOU WILL: Design and build across the full stack — from the public training APIs and SDK through the control plane to the GPU data plane. Scale the distributed systems that make GPU compute serverless — multi-tenant scheduling, placement, and capacity-aware routing across regional G
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. LEAD. STRATEGIZE. TRANSFORM. JOIN US AS A PROFESSIONAL SERVICES PRACTICE MANAGER! Are you ready to own your impact at the forefront of the data revolution ? As a Services Practice Manager , you’ll be a key driver in transforming businesses through cutting-edge cloud data solutions . You’ll collaborate with top-tier clients, strategize with elite sales teams, and lead groundbreaking professional service engagements —all while working with the best minds in the industry! This role is completely remote and can be based anywhere in the Eastern Timezone. WHY THIS ROLE? Be a Market Leader: Own and execute a strategic vision for Professional Services. Drive Meaningful Change: Work with clients to unlock the power of data and fuel digital transformations. Lead from the Front: Manage high-impact projects, ensure seamless delivery, and drive revenue growth. Collaborate with the Best: Work cross-functionally with top industry experts to shape the future of cloud data solutions. Crush Your Goals: Meet and exceed bookings, utilization, and consumption targets while delivering exceptional customer experiences . WHAT YOU’LL DO: Spearhead services sales execution and project delivery success for key accounts. Engage C-level executives to create and execute services roadmaps that drive real
Other cities to consider
More places hiring for this role
Get new service tech jobs in United States by email
Daily job updates · Unsubscribe anytime