Jobs in United States

Platform Operations Specialist in United States

3,618 active opportunities · Updated October 2026

Explore current platform operations specialist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -89.3%

From $192K/yr

Quick readStrong listing-quality and freshness signals

We are the Experimentation team at Datadog, a startup operating inside one of the fastest-growing companies in software. The team brings together engineers who joined through Datadog’s May 2025 acquisition of Eppo with engineers from across Datadog, and together we are building feature flagging and experimentation into the platform. Both products were recently launched this year, and we are just getting started: we are migrating customers from Eppo onto Datadog while shipping an ambitious roadmap across both flagging and experimentation. Our statistics engineers hold PhDs and are pushing the boundaries of what an experimentation platform can do. Our full-stack engineers move fast, lean on the latest AI tooling, and prototype and ship new features weekly. We talk to customers directly to shape requirements, and our roadmap is driven by our own product vision and sharpened by real customer and sales input. We are looking for an Engineering Manager to lead the team building the Experimentation App. You will own one to two squads to start, partnering closely with our engineers, statisticians, product managers, and designers to turn a deep roadmap into a shipped product. To conform to US export control regulations, candidates should be eligible for any required authorizations from the US government. What You’ll Do: Lead, mentor, and grow a team of strong full-stack and frontend engineers. Help them grow to the next level and continuously provide them opportunities to develop. Own delivery for the Experimentation App from planning through production, balancing speed with quality in a fast-paced environment. Set technical direction in partnership with your team and our statistics engineers, engaging credibly on architecture and tradeoffs. Partner with Product, Design, and customers directly to shape requirements and turn customer and sales input into roadmap. Foster a culture of fast iteration and high engineering standards through code reviews, design reviews, and blamele

AIGoRustSpring
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the Team The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. Technical Program Managers at OpenAI play a key leadership role in scaling these efforts, partnering deeply with product, engineering, design, and go-to-market teams to bring ambitious ideas to life and ensure clarity and discipline in execution. About the Role We are hiring a Technical Program Manager to support OpenAI's critical AI deployments across strategic cloud partners. This role is designed for a candidate who can operate as an end-to-end owner across internal engineering teams and external partner organizations. This role will drive the technical strategy and execution required to bring OpenAI models and platform capabilities into partner environments responsibly and at scale. The work spans engineering deliverables, shared roadmaps, model launch pipelines, technical integration, launch readiness, and post-launch follow-through. You will work closely with senior leaders across OpenAI engineering, infrastructure, product, safety, security, legal, finance, and go-to-market, as well as technical counterparts at our partners. The job is to turn broad partnership commitments into concrete execution plans, align both sides on what must land, and build repeatable mechanisms for launching OpenAI capabilities on third-party platforms. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end execution for major cloud partner programs spanning model deployment, product integration, operational readiness, launch follow-through, and partner-platform adoption. Own integrated technical roadma

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the team OpenAI’s Forward Deployed Engineering team partners with healthcare organizations to deploy production AI systems across clinical, operational, and member-facing workflows. We work at the boundary of customer deployment and core platform development, using customer engagements to define repeatable architectures, evaluations, integrations, and operating standards for complex, regulated healthcare environments. About the role We are hiring a Forward Deployed Engineer (FDE) to own end-to-end deployments of our models within healthcare organizations, including payers, providers, health systems, and healthcare technology companies. You will lead technical discovery, architecture, implementation, evaluation, productionization, and handoff, translating complex customer workflows, data, infrastructure, and regulatory constraints into production AI systems. You will measure success through production adoption, measurable workflow impact, and evaluation loops that establish customer-specific benchmarks, acceptance criteria, and launch readiness. You’ll collaborate directly with customer technical and operational teams, alongside internal Business, Research, Product, Engineering, and Security partners, to deliver solutions and translate deployment learnings into product improvements. This role owns the technical solution; ownership of the commercial or executive relationship is not required. This role is based New York City. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own the technical solution end to end, from customer discovery and workflow scoping through architecture, hands-on implementation, evaluation, production deployment, adoption, and handoff. Partner credibly with customer engineers, operators, and domain experts to frame ambiguous problems, define scope, and translate payer, provider, or health-system workflows into technical requirements and measurab

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI’s Forward Deployed Engineering team partners with healthcare organizations to deploy production AI systems across clinical, operational, and member-facing workflows. We work at the boundary of customer deployment and core platform development, using customer engagements to define repeatable architectures, evaluations, integrations, and operating standards for complex, regulated healthcare environments. About the Role We are hiring a Forward Deployed Engineer (FDE) to own end-to-end deployments of our models within healthcare organizations, including payers, providers, health systems, and healthcare technology companies. You will lead technical discovery, architecture, implementation, evaluation, productionization, and handoff, translating complex customer workflows, data, infrastructure, and regulatory constraints into production AI systems. You will measure success through production adoption, measurable workflow impact, and evaluation loops that establish customer-specific benchmarks, acceptance criteria, and launch readiness. You’ll collaborate directly with customer technical and operational teams, alongside internal Business, Research, Product, Engineering, and Security partners, to deliver solutions and translate deployment learnings into product improvements. This role owns the technical solution; ownership of the commercial or executive relationship is not required. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role, you will: Own the technical solution end to end, from customer discovery and workflow scoping through architecture, hands-on implementation, evaluation, production deployment, adoption, and handoff. Partner credibly with customer engineers, operators, and domain experts to frame ambiguous problems, define scope, and translate payer, provider, or health-system workflows into technical requirements and mea

AWSRestAIGo
O
📍 Seattle, Washington, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

$293K – $325K/yr

Quick readStrong listing-quality and freshness signals

About the Team The Statsig team at OpenAI builds and operates the experimentation platform that powers product development, measurement, and decision-making across the company. We partner closely with product, engineering, and infrastructure teams to ensure experiments are trustworthy, statistically rigorous, and scalable to the needs of frontier AI products. Our mission is to help teams make better decisions through reliable experimentation. We care deeply about statistical correctness, pragmatic solutions, and building systems that researchers and engineers can trust at massive scale. The team operates at the intersection of experimentation methodology, data infrastructure, causal inference, and product analytics. We are looking for experienced experimentation experts who want to shape the future of experimentation in the AI era. About the Role We are hiring a Staff-level Data Scientist to help lead the evolution of OpenAI’s core experimentation platform. This role is focused on improving the statistical rigor, reliability, and practical usability of experimentation across the company. You’ll work on some of the hardest problems in online experimentation: sample ratio mismatch detection, variance reduction, bias mitigation, metric design, triggered analysis, heterogeneous treatment effects, sequential testing, and experimentation in complex ML systems. You’ll also help translate advanced statistical concepts into pragmatic systems and product experiences that teams can actually use. This is a highly technical individual contributor role with significant influence across methodology, platform architecture, and experimentation best practices. The ideal candidate combines deep statistical expertise with strong systems intuition and hands-on experience building or operating experimentation platforms at scale. In this role, you will: Drive the statistical direction and technical strategy for OpenAI’s experimentation platform Design and improve experimentation methodolo

PythonVueAWSRest
P
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity At Postman, we are revolutionizing the way developers build, trace, and automate API workflows with Postman Flows , a powerful visual programming tool designed to simplify the development and sharing of API-powered applications. With an intuitive drag-and-drop interface, Postman Flows enables teams to collaborate and showcase their APIs regardless of technical expertise. We are looking for a Software and Systems Engineer to help scale and maintain the Flows runtime system. This system runs mission-critical automations in the cloud, with a focus on low latency, high throughput, and high availability. You’ll play a vital role in developing, deploying, and operating our backend services and infrastructure in a Kubernetes-based cloud environment. We’re looking for an experienced engineer who is excited not only about hands-on building as we ship and iterate on a weekly basis to get our product ready for GA, but who can also serve as a role model and mentor to other engineers. This role involves making key technical decisions and improvements to the system, as well as effectively making impact through influence wit

Node.jsAWSAzureGCP
T
Golf. A golf role or an employer dedicated to golf.
📍 Tg, Dallas Office, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

The Engineering Manager, Memberships leads the team responsible for building and operating Topgolf’s Memberships systems, from guest-facing membership experiences through the backend services that power them. This role owns the people, process, and delivery of the Memberships engineering team, while staying technically credible across a full stack built on Go, Vue.js, and PostgreSQL. Requirements Lead, grow, and manage a team of full-stack engineers building Memberships systems, including hiring, performance management, career development, and mentorship Set clear goals and expectations for the team, run effective 1:1s, and build a culture of ownership, accountability, and continuous improvement Balance workload and staffing across Memberships initiatives, escalating resourcing gaps and continuity risks early Guide architecture and design decisions across Memberships systems, drawing on full-stack experience spanning Go, Vue.js, PostgreSQL, and API design Set engineering standards and best practices for code quality, testing, and release processes, and stay close to the codebase through code reviews and hands-on problem solving on critical issues Own delivery of the Memberships roadmap end to end, from technical planning through implementation, QA, release, and post-launch monitoring, ensuring systems are observable, testable, secure, and built to scale with guest demand Partner with product, design, QA, and platform engineering to translate guest needs and business priorities into a clear, prioritized Memberships roadmap Represent the Memberships team in cross-functional planning and architecture discussions, communicating progress, risks, and tradeoffs to engineering leadership and business stakeholders Critical Skills Strong architectural judgment and the ability to balance technical debt, delivery speed, and long-term maintainability <

PythonVuePostgreSQLAWS
H
📍 Texas, United States of America, United States
✓ High-confidence listingCompany trend -5.2%

$130.7K – $205.2K/yr

Quick readStrong listing-quality and freshness signals

Senior Infrastructure Architect — Enterprise Observability and Automation Description - Job Summary Senior individual contributor responsible for the architecture, implementation, and operational ownership of enterprise observability, monitoring, and automation platforms across HP's global IT environment. This role modernizes infrastructure visibility capabilities while ensuring operational stability, security, and compliance. Serves as a technical and operational bridge between infrastructure engineering, cybersecurity, SOX/compliance stakeholders, automation teams, and external technology partners — leading complex initiatives such as platform migrations, enterprise integrations, and governance enablement. Responsibilities Enterprise Observability and Monitoring Application owner and senior technical authority for enterprise monitoring and logging platforms (Datadog, Splunk), including platform governance, roadmap alignment, and operational oversight. Lead enterprise-scale monitoring platform migrations, including architecture design, agent strategy, data ingestion models, vendor coordination, and deployment across 5,000&#43; servers. Define standards for alerting, dashboards, observability data quality, and integration with ITSM platforms (ServiceNow). Design and manage multi-org Datadog architecture, including org structure, RBAC, SSO/SAML, secrets management, and cybersecurity compliance. Oversee SNMP-based monitoring of storage and network devices, including device profiling, syslog/event integration, and NetFlow collection. SOX Compliance and IT Governance SOX control owner for enterprise monitoring applications — approve monthly reviews, participate in internal/external audits (EY), and maintain ITGC/SOX compliance. Provide audit evidence, walkthrough docu

AWSAzureAnsible
C
📍 Tampa Florida United States, United States
✓ Quality checkedCompany trend +800%

We are seeking an experienced Senior Generative AI Developer to help drive the design, development, and integration of state-of-the-art Generative AI and agentic AI solutions across our enterprise Controls Technology platform. You will collaborate with cross-functional teams, contribute deep technical expertise in context engineering, retrieval systems, knowledge graphs, and multi-agent orchestration, and play a key role in delivering scalable, grounded AI solutions to enhance automation and operational efficiency. This role centers on architecting robust applications and agent systems on top of pre-trained and hosted foundation models — not on training or fine-tuning models. Key Responsibilities Collaborate with AI architects, leads, and stakeholders to design and implement generative and agentic AI solutions that address business challenges. Architect advanced context engineering strategies — context layering, chaining, compression, pruning/offloading, and memory management — to maximize reliability, provenance, and token efficiency in production. Design and implement advanced generative AI methods, including sophisticated prompt engineering and Retrieval-Augmented Generation (RAG) . Build and optimize RAG systems , including hybrid search, multi-vector retrieval, and re-ranking pipelines. Design and implement knowledge graphs and Graph RAG architectures to enable multi-hop reasoning, explainability, and traceable, grounded responses for high-value business domains. Architect agentic workflows and multi-agent systems using Google Agent Development Kit (ADK) and comparable frameworks (LangGraph, Microsoft Agent Framework, CrewAI), applying orchestration patterns such as supervisor/worker, hierarchical, and peer-to-peer. Design robust agent harnesses — governance, constraints, feedback loops, state/session management, and

PythonJavaSQLAWS
OS
📍 Aliso Viejo, United States· Full-time· Remote
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About OfficeSpace: OfficeSpace Software provides the leading AI operating system for the built world, that helps teams plan, connect, and perform in the workplace. As a performance-based, PE-backed company, we hire based on merit and a willingness to do what it takes to succeed long-term. You’re a great fit for the role if you’re entrepreneurial, passionate, motivated by building at light speed, and an Agentic AI early adopter. Our world-class teams operate in the US, Canada, and Costa Rica in a culture of trust, respect, growth, and impact. Role Summary: You win new logos. Full stop. As a Enterprise Account Executive, you own the end-to-end motion of bringing new customers into OfficeSpace. You run disciplined, value-led sales cycles with senior decision-makers and use AI-powered tools to prospect smarter, sell faster, and close with confidence. We give you a market-leading platform. You turn it into revenue. What You’ll Do: Own new-logo acquisition across a defined Enterprise territory—build pipeline, run deals, and close. Lead buyer-driven, value-based sales cycles with C-level, VP, and Director stakeholders. Orchestrate AI-assisted prospecting to identify high-intent accounts and create net-new opportunities. Translate workplace, facilities, and real estate challenges into clear, quantifiable business value. Deliver sharp, compelling product demos that connect OfficeSpace capabilities to real-world outcomes. Guide AI-generated proposals, pricing models, and forecasts—ensuring accuracy, clarity, and impact. Negotiate contracts and commercial terms with confidence and discipline. Maintain rigorous pipeline hygiene and forecasting using Salesforce and revenue intelligence tools. Partner with Marketing and Business Development to accelerate deal velocity and improve win rates. Share frontline market feedba

AgileAIGoRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team pAGI Infra team builds and operates the systems that make large-scale model training and evaluation reliable, efficient, and easy to run. Our work spans distributed training infrastructure, inference and grading platforms, compute scheduling, and research tooling. We partner closely with researchers and engineering teams to turn new research needs into dependable infrastructure, improve GPU efficiency, and shorten the path from an experiment to a validated model. About the Role We’re looking for an AI Systems Engineer to help scale the infrastructure behind our training and evaluation workflows. You’ll own projects from identifying bottlenecks and designing solutions through deployment and operation. The work combines distributed systems engineering, performance optimization, and close collaboration with researchers. You might build a shared grading service, improve resource allocation across workloads, or bring a new training stack into production — directly improving how quickly and reliably research moves forward. In this role, you will: Build and operate infrastructure for large-scale training and evaluation, improving reliability, throughput, and resource efficiency. Develop shared inference and grading platforms with automated capacity management, health monitoring, and visibility into performance. Improve compute scheduling and resource allocation to reduce idle GPU time and help workloads recover quickly from failures. Diagnose bottlenecks across training, inference, and orchestration, and work across teams to improve end-to-end performance. Build self-service tools, automated validation, and observability that help researchers launch experiments, diagnose issues, and compare results with less manual intervention. You might thrive in this role if you: Are excited about the potential of personal AGI and want to build the infrastructure that enables it. Have strong software engineering fundamentals and experience building or operating large-scal

AWSRestAIRust
F
📍 Bellevue, Washington, United States
✓ Quality checkedCompany trend -87.7%

About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The opportunity: We're looking for a Staff Product Manager to own and drive strategy across a product domain at Flexport. This is not a single-feature role. You will define the vision for how multiple product surfaces work together to serve clients and operators across global logistics, then execute against that vision with a high degree of autonomy. Flexport's platform powers every stage of global freight, from booking and pricing to visibility, customs, and delivery. The products you shape will be used daily by logistics managers, supply chain leaders, freight operators, and internal teams who keep goods moving around the world. You will work across multiple engineering teams and coordinate with other PMs to ensure your domain delivers cohesive, measurable outcomes. You will also raise the bar for the PM team through mentorship, better frameworks, and sharper thinking on the hardest product problems we face. This is a role for someone with deep product experience, strong analytical instincts, and the ability to operate at both the strategic and execution level in a complex, data-rich environment. The team: The Commerce team owns the end-to-end quality

SQLAISupply ChainLogistics
F
📍 Bellevue, Washington, United States
✓ Quality checkedCompany trend -87.7%

About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The opportunity: We're looking for a Staff Product Manager to own and drive strategy across a product domain at Flexport. This is not a single-feature role. You will define the vision for how multiple product surfaces work together to serve clients and operators across global logistics, then execute against that vision with a high degree of autonomy. Flexport's platform powers every stage of global freight, from booking and pricing to visibility, customs, and delivery. The products you shape will be used daily by logistics managers, supply chain leaders, freight operators, and internal teams who keep goods moving around the world. You will work across multiple engineering teams and coordinate with other PMs to ensure your domain delivers cohesive, measurable outcomes. You will also raise the bar for the PM team through mentorship, better frameworks, and sharper thinking on the hardest product problems we face. This is a role for someone with deep product experience, strong analytical instincts, and the ability to operate at both the strategic and execution level in a complex, data-rich environment. The team: The Visibility team owns the end-to-end mappi

SQLSupply ChainLogistics
H
📍 New York, NY, United States
✓ Quality checkedCompany trend +310%

Become a part of our caring community You have shipped AI products before. You understand the difference between a demo and a production system. You have strong opinions about evaluation frameworks because you have experienced the consequences of operating without them. You are at your best when you own architecture decisions while continuing to build and deliver critical code yourself. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to source documents, and route complex cases to human experts. The output of these systems supports healthcare decisions that impact real members. As a Lead AI Applied Engineer, you will provide technical leadership for AI-enabled products and platforms, define architectural direction, establish engineering standards, and personally design and build the most critical components of our systems. You will lead through both technical expertise and execution, helping the team deliver reliable, scalable, and auditable AI solutions in a highly regulated healthcare environment. Why Join Us Lead the architecture of production AI systems where LLMs are foundational to the product experience. Make key technical decisions regarding model selection, system boundaries, platform architecture, and build-versus-buy strategies. Own the highest-risk and highest-impact technical challenges involving reliability, explainability, and correctness. Influence engineering culture and establish standards that shape how the team builds and ships AI products. Work on systems operating at meaningful scale, processing millions of documents and supporting healthcare decisions across a large member population. Partner

JavaScriptTypeScriptPythonReact
P
📍 San Francisco, California, United States
✓ Quality checkedCompany trend -100%

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking a strategic and results-driven engineering leader who is passionate about cloud agnostic infrastructure, operational excellence, and enabling engineering teams to operate autonomously and build with confidence. As Head of Infrastructure, you'll lead a talented and geographically distributed team of engineers across the SF Bay Area, India, and Europe, fostering a culture of collaboration, ownership, and continuous improvement. You'll own the infrastructure that underpins one of the world's most widely used API platforms, an environment handling ~80,000 requests per second at the front door, and be responsible for its reliability, scalability, and evolution. In addition to infrastructure, you'll own the Site Reliability Engineering (SRE) function at Postman, setting the standards and practices that keep the platform reliable at scale. You'll work closely with engineering managers, product managers, and platform teams to drive the technical roadmap for our cloud agnostic infrastructure and reliability practices, ensuring we can support a large and rapidly growing engineering organization. If you're p

AWSAzureKubernetesProject Management
🔔

Get new platform operations specialist jobs in United States by email

Daily job updates · Unsubscribe anytime