Jobs in United States

Platform Delivery Specialist in United States

3,618 active opportunities · Updated October 2026

Explore current platform delivery specialist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.2%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE The largest, most demanding enterprises are starting to run on Baseten, and they arrive with a range of security, compliance, and procurement requirements. As a Senior Engineer on Baseten's enterprise engineering team, you'll build the capabilities that enable large organizations like Writer, HubSpot, and Notion to succeed on Baseten. Enterprise engineering authors the core building blocks, APIs, and user experiences powering the Baseten platform: identity and access management, billing, regional isolation, and self-hosted and single-tenant deployment options. This is deep product and systems work across the full stack, from designing authentication and authorization systems using standards like OAuth and OIDC to shipping the admin experiences enterprise IT teams use to manage their organization. EXAMPLE INITIATIVES Recent and upcoming work on the team: Fine-grained authorization for users, service accounts, and agentic workloads SSO and SCIM support, allowing customers to centralize and automate access to Baseten Expanding the billing platform to support evolving pricing models, advanced data exports, and controls to manage spend In-product management and enforcement of customer compliance requirements like data residency and HIPAA Securing network paths in and out of a customer's models with private connectivity and ingress and egress restrictions Allowing customers to run Baseten inside their own VPC, on-pr

KubernetesRestMachine LearningAI
G
📍 United States· Full-time
✓ High-confidence listingCompany trend -100%

From $215K/yr

Quick readStrong listing-quality and freshness signals

Location Details: USA Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This position may be a hybrid or fully remote position, as decided by your manager. If designated as hybrid, you’ll divide your time between working remotely from your home and an office location, so you should live within commuting distance. If designated as remote, you’ll be working remotely from your home and may occasionally visit a GoDaddy office to meet with your team for events or meetings. Your hiring manager can share more about this role’s hybrid or remote designation. This position is not eligible to be performed in Alaska, Mississippi, North Dakota, or the Virgin Islands. GoDaddy is not currently considering candidates for this role in California, Seattle, or NYC. Join our Team GoDaddy's Identity Platform powers authentication, authorization, and core identity infrastructure, which sits at the heart of how millions of small businesses securely access, manage, and grow their online presence, across our products. We’re looking for a Director of Engineering to lead this critical platform into its next chapter — driving company-wide adoption of OAuth/OpenID Connect (OIDC) and OpenFGA, strengthening the reliability and security of core identity services, and crafting customer experiences that are flawless, scalable, and trusted. In this role, you’ll set the technical vision and the execution rhythm for a 50+ person engineering organization. This organization spans the US and India. You will lead through multiple Senior Managers while staying close to the technology. This is a hands-on leadership role — we expect you to code, prototype, run technical experiments, and to lead by example in using AI-assisted development tools. What you’ll get to do Own and set the engin

PythonJavaReactAWS
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $345K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Creator Content Platform Team provides a secure, scalable, and extensible foundation for ingestion, processing, storing, managing, and serving user content. As a Principal Software Engineer (Backend, Distributed Systems), you will design and build backend services to help game creators reach the broadest audience possible. You will help us build & scale large distributed systems, data processing and analysis pipelines, an access control system, Open Cloud APIs all of which are central for all content created in Roblox. \ You Will: Solve on a variety of unique technical challenges Have the independence, opportunity and the end-to-end responsibility to design, build, test and deploy services within the Roblox ecosystem Guide the future technical direction of the team and have impact on engineering Be a technical bar-raiser for high code quality, architectural designs, and long-term approaches Mentor and develop fellow engineers on the team Design systems and services that are scalable and resilient Collaborate with passionate, Engineers, Product Managers and other Roblox team members, cross-functionally You have: Experience: You have 13+ years of experience working on backend, d

PythonJavaAWSGit
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -89.8%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Senior Software Engineer - Streaming Platform Client Data streams are mission-critical at Datadog, powering near real-time communication across the vast majority of our services. Our Streaming Platform group builds the core infrastructure and abstractions that ensure Datadog remains a trusted partner for engineers worldwide. See our blog post . The Streaming Platform Client team sits at the heart of this ecosystem. We own the Rust client library (producers and consumers) with language bindings for Java, Go, and Python. We focus on building intuitive APIs and abstractions that make a powerful distributed system easy to adopt and operate for the hundreds of internal users of our library. Our library runs on critical data paths that handle hundreds of millions of messages per second making performance, observability, and reliability paramount. We also develop and operate the service that bridges the clients fleet with the platform's control plane, handling complex balancing, scaling, and static stability challenges. We are seeking a Senior Software Engineer to help us evolve these features. You will collaborate directly with our users, tackle performance-critical code, and solve complex distributed systems challenges across the control plane, client libraries, and data plane. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Work within a distributed, high-impact team spanning Europe and the US, building critical technologies that power data pipelines for dozens of internal teams and hundreds of services. Architect and implement resilient interactions between our client libraries and the control plane. Optimize our high-throughput, low-level streaming library to push the boundaries of performance and efficiency. Champion the developer experience by providing

PythonJavaGitAI
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -89.8%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Senior Software Engineer - Streaming Platform Client Data streams are mission-critical at Datadog, powering near real-time communication across the vast majority of our services. Our Streaming Platform group builds the core infrastructure and abstractions that ensure Datadog remains a trusted partner for engineers worldwide. See our blog post . The Streaming Platform Client team sits at the heart of this ecosystem. We own the Rust client library (producers and consumers) with language bindings for Java, Go, and Python. We focus on building intuitive APIs and abstractions that make a powerful distributed system easy to adopt and operate for the hundreds of internal users of our library. Our library runs on critical data paths that handle hundreds of millions of messages per second making performance, observability, and reliability paramount. We also develop and operate the service that bridges the clients fleet with the platform's control plane, handling complex balancing, scaling, and static stability challenges. We are seeking a Senior Software Engineer to help us evolve these features. You will collaborate directly with our users, tackle performance-critical code, and solve complex distributed systems challenges across the control plane, client libraries, and data plane. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Work within a distributed, high-impact team spanning Europe and the US, building critical technologies that power data pipelines for dozens of internal teams and hundreds of services. Architect and implement resilient interactions between our client libraries and the control plane. Optimize our high-throughput, low-level streaming library to push the boundaries of performance and efficiency. Champion the developer experience by pro

PythonJavaGitAI
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -89.8%

From $280K/yr

Quick readStrong listing-quality and freshness signals

The Detection Platform organization is responsible for helping customers identify, understand, and act on issues across their environments through alerting, event intelligence, and autonomous detection capabilities. As Director, Detection Platform, you will lead a group of engineering managers and teams responsible for foundational alerting infrastructure, event management, monitor creation experiences, and AI-powered detection systems. This role sits at the center of Datadog’s efforts to evolve how customers detect, investigate, and respond to operational issues at massive scale. You will partner closely with Product Management, Applied Science, Design, and Engineering leaders to shape the future of detection and observability experiences for Datadog customers while leading a growing organization of engineers. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Lead a multi-team engineering organization responsible for alerting, event management, monitor creation experiences, and autonomous detection capabilities. Define and execute the technical and organizational strategy for the Detection Platform while aligning stakeholders across Engineering, Product, Design, and Applied Science. Drive innovation in AI-powered detection, anomaly identification, and signal generation that helps customers proactively identify and resolve issues. Scale highly available platform systems that process hundreds of millions of evaluations while maintaining reliability, performance, and operational excellence. Develop and mentor engineering managers and technical leaders, fostering a culture of execution, collaboration, and technical rigor. Champion customer-centric product thinking by balancing platform investments with intuitive user experiences and measurable customer

Machine LearningAIGoRust
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -86.4%

About the Team The Future of Computing Research team is an Applied Research team within the Consumer Devices group focused on developing new methods and models as we advance forward in our mission of building AGI that benefits all of humanity. As a Software Engineer on the Future of Computing Research team, you will work together with both the best ML researchers in the world and the greatest design talent of our generation to push the frontier of model capabilities. About the Role We are looking for a Software Engineer to join our team to build tools and services that enable AI research, evaluation, and data generation workflows. The best work in this role will start with an ambiguous design question and turn it into working research systems. You will work closely with researchers, designers, and engineers to build the evaluation systems, synthetic data generation pipelines, review tools, and supporting platform services. The goal is to make these workflows easier to create, run, and trust without requiring bespoke engineering support for each new design concept. You will help ensure that research artifacts have a clear lifecycle, runs are reproducible and observable, and results provide useful evidence for product and model-training decisions while the underlying systems remain reliable and reusable. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Build web applications, APIs, data models, and backend services for AI research workflows. Build tools to author and manage evaluation tasks, rubrics, graders, suites, and rollout configurations, including workflows for publishing, versioning, auditing, and sharing research artifacts. Automate evaluation runs and generate useful reports for design, research, and engineering teams. Support synthetic data generation workflows for multimodal and conversational research, including tools that comb

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -86.4%

About the Team The Platform Systems team at OpenAI operates at the intersection of cutting-edge AI and large-scale distributed systems. We build the engineering and research infrastructure required to train OpenAI’s flagship models on some of the world’s largest, custom-built supercomputers. Our team develops core model training software and works deep in the stack - spanning collective communication, compute efficiency, parallelism strategies, fault tolerance, failure detection, and observability. The systems we build are foundational to OpenAI’s research velocity, enabling reliable, efficient training at frontier scale. We collaborate closely with researchers across the organization, continuously incorporating learnings from across OpenAI into the evolution of our training platform. About the Role As a Software Engineer, Platform Systems, you will design and build distributed systems that provide visibility into large-scale training workloads and help operate them reliably at scale. You’ll work on failure detection, tracing, and observability systems that identify slow or faulty nodes, surface performance bottlenecks, and help engineers understand and optimize massive distributed training jobs. This infrastructure is critical to operating OpenAI’s training stack and is actively evolving to support new use cases and increasingly complex workloads. This role sits at the core of our training infrastructure, blending systems engineering, performance analysis, and large-scale debugging. In This Role, You Will Design and build distributed failure detection, tracing, and profiling systems for large-scale AI training jobs Develop tooling to identify slow, faulty, or misbehaving nodes and provide actionable visibility into system behavior Improve observability, reliability, and performance across OpenAI’s training platform Debug and resolve issues in complex, high-throughput distributed systems Collaborate with systems, infrastructure, and research teams to evolve platform

AWSRestAIRust
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As the Head of AI Platform Engineering at Postman, you will lead the alignment of AI development with our growing API platform. You will drive the AI roadmap with a focus on expanding AI-driven API collaboration and agentic capabilities across the platform. This role requires a leader who can identify market opportunities, coordinate cross-functional AI initiatives, and foster strong partnerships to amplify the Postman AI platform's impact What You’ll Do Lead the development and execution of Postman’s AI platform strategy, focused on API ecosystem growth and platform innovation. Drive the AI roadmap, concentrating on API integration, platform expansion, and AI-driven agent functionality. Identify and capitalize on market opportunities for AI-enhanced API collaboration and intelligent agent features. Collaborate closely with business units, product teams, engineering, and external partners to ensure alignment and successful AI initiatives deployment. Oversee implementation with core AI platforms (OpenAI, Anthropic, AWS, etc)), ensuring technical and strategic alignment with AI features and API lifecycle improvements.

AWSMachine LearningAIGo
C
📍 Tampa Florida United States, United States
✓ Quality checkedCompany trend +800%

The Engineering Lead Analyst – Test Automation Platform Engineering is a senior-level technical leadership role responsible for driving the architecture, implementation, management, operational support, and continuous enhancement of enterprise test management and automation platforms. In this role, you will lead efforts to modernize test automation capabilities across the global technology ecosystem. You will architect end-to-end integration workflows, embed automated quality gates into enterprise CI/CD pipelines, and administer as well as operationally support both vendor and internally developed enterprise platforms (e.g., Core Performance Engineering / Performance Center, ALM-Quality Center, Zephyr Enterprise, CSDP / Octane). Additionally, you will play a critical role in production operations—delivering tier-3 platform support to rapidly and safely troubleshoot, triage, and remediate performance and availability issues in complex, distributed production environments. The ideal candidate blends deep hands-on expertise in software testing frameworks, modern DevOps pipelines, containerized infrastructure, message-driven integration, and robust operational resilience practices with strong governance, compliance, and stakeholder leadership skills. Key Responsibilities 1. Platform Engineering & Operational Support Install, configure, upgrade, administer, and support enterprise test management and performance engineering toolsets (e.g., Core Performance Engineering / Performance Center, ALM-Quality Center, Zephyr Enterprise, Core Software Development Platform [CSDP] / Octane). Provide end-to-end operational support for

DockerKubernetesArtificial IntelligenceAI
C
📍 New York New York United States, United States
✓ Quality checkedCompany trend +800%

About the Role Discover your future at Citi Working at Citi is far more than just a job. A career with us means joining a team of more than 230,000 dedicated people from around the globe. At Citi, you'll have the opportunity to grow your career, give back to your community and make a real impact. Job Overview Citi's Integrated Digital Assets Platform (CIDAP) is at the vanguard of institutional blockchain adoption — and security is its foundation. As digital assets move from innovation to regulated infrastructure, the cryptographic integrity of every transaction, wallet, and key lifecycle operation becomes mission-critical. We are building the security layer that the world's most sophisticated financial institution can trust. We are seeking a Senior Security Engineer (VP) to join our New York-based Digital Assets Platform engineering team. This is a hands-on, Java-focused backend engineering role for a security-minded engineer who understands both the craft of secure software development and the cryptographic primitives that underpin digital asset custody, signing, and key management. You will sit inside the core engineering team — writing production code every day — while being the resident authority on cryptographic design patterns, HSM integration, MPC protocols, and security architecture. Your work will directly protect billions of dollars of digital asset infrastructure used by institutional clients worldwide. Key Responsibilities Design, develop, and maintain security-critical backend services in Java — including cryptographic libraries, key management APIs, signing wor

JavaArtificial IntelligenceAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

We're building the platform that lets long-running autonomous agents operate safely inside NVIDIA's enterprise. These are not assistants on a developer's laptop. They are fleets of agents deployed in the cloud, running continuously at scale on shared accelerated compute. They take on real work across enterprise systems, so people get far more done than they could before. This role defines the constructs that agents are built from: the blueprints they start from, the tools, skills, and plugins that power them against enterprise data, the runtime safety harness that keeps them in bounds, and the connections into credential management, sandbox, memory, and observability. The team designs and ships these building blocks so that agent developers across the company can stand up a new agent, wire it in, and run it for days or weeks. Security and safe execution come out of the box, not something each team has to get right on its own. Today an agent runs inside a single harness. Claude, Codex, and open-source agent harnesses each work differently underneath, with their own execution model, tool interface, and telemetry shape. The platform smooths over those differences, so a single skill, safety policy, or trace works the same no matter which harness is running. We want to enable agents that act on a person's behalf, governed and secure, continuously evaluated and self-improving. These agents coordinate and hand work off to each other, with identity and policy following every hop. They route and tune themselves across harnesses from live eval signals, and get better from their own production telemetry instead of waiting on a human to retrain them. Have you run agents on a harness like Claude or Codex and hit the walls that show up when they run for real, for days, against live systems — and wanted them to learn from it on their own? We're building the platform that solves those problems once, for every team. What you'll b

PythonAIFinanceHR
C
📍 Work At Home Texas, United States
✓ Quality checkedCompany trend +340.2%

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary Health100 is an AI‑native health technology platform that unifies pharmacies, providers, insurers, PBMs, and digital health solutions into a single, consumer‑focused ecosystem. Powered by Google Cloud AI, we’re reimagining personalized and connected health experiences. As a Senior Software Engineer for Health100, you will play a crucial role within a collaborative team — designing, developing, and maintaining backend services and APIs while ensuring releases are well-coordinated, fully prepared, and successfully deployed to production. The ideal candidate brings strong technical expertise in modern backend development, excellent problem-solving skills, and a proactive approach to production monitoring, issue triage, and cross-team coordination. This position is critical in maintaining high engineering standards, ensuring smooth release cycles, and driving operational excellence across the development lifecycle. *This role can be based anywhere in the US; hybrid or remote with preference for candidates to work out of our corporate headquarters in Woonsocket, RI. Responsibilities: Partner with technical leaders and the open-source community to contribute to technical designs, frameworks, roadmap definition, and requirements-gathering. Provide domain knowledge and engineering insight to guide early designs, ac

JavaAzureGCPAI
D
📍 Massachusetts, New York, United States· Full-time
✓ High-confidence listingCompany trend -89.8%

From $192K/yr

Quick readStrong listing-quality and freshness signals

About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—allowing for seamless collaboration and problem-solving among Dev, Ops and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Team: The Revenue Data Engineering Teams designs, builds and runs the data pipelines and helper systems to accurately and in a timely manner quantify our customers’ usage across all Datadog products. This team is at the leading edge of any new product we release. The Revenue Data Processing team builds and operates the data pipelines that does billing, and cost attribution for all Datadog products. We process terabytes of data daily to power revenue-critical systems and are at the center of every new product launch at Datadog. As a Senior Software Engineer, you will own meaningful parts of a large-scale, mission-critical processing platform — driving architectural improvements, building new billing capabilities, and maintaining the high reliability bar our downstream consumers depend on. You Will: Design and build high-throughput data pipelines for billing and cost attribution Drive platform improvements — latency reduction, Spark optimization, sharding, and cross-datacenter reliability Own root-cause investigations on billing accuracy issues in collaboration with Finance and Product teams Contribute to new billing features Work across Python and Scala, with technologies including Spark, Airflow, Trino, and Apache Iceberg Participate in on-call rotation and maintain a high reliability bar for production systems Contribute to engineering standards and help grow the technical culture of the team You Are: You have significant experience building and operating production data pipelines at scale using Spark and Airflow

PythonAIGoRust
🔔

Get new platform delivery specialist jobs in United States by email

Daily job updates · Unsubscribe anytime