Jobiba hiring network

Platform Engineer Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current platform engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

N
Nuro
📍 Mountain View• Full-time• From $258K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role The Eval Platform team owns the simulation and evaluation and validation platform that underpins autonomy development and driverless readiness validation. This platform is mission‑critical for both rapid iteration during development and rigorous validation used to assess safety, performance, and deployment readiness. We are looking for a seasoned engineering leader to lead the Eval Platform as a Director‑level role. This leader will define the technical vision, execution strategy, and organizational structure for a multi‑disciplinary org spanning simulation, evaluation infrastructure, and compute and storage platform. The evaluation platform brings together data, simulation, metrics, and analysis workflows into a coherent system that allows the teams to: Iterate quickly to improve the autonomy performance Measure autonomy performance and risk with credibility and rigor Make confident, data‑driven decisions about driverless

aigorust
View job →
C
Cloudflare
📍 Hybrid• Full-time• Hybrid
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations: Austin, Texas About the role Created in January 2025, our centralized Developer Go-to-Market (GTM) organization functions as a "speedboat" to maintain focus, control direction, and iterate quickly. Our mission is to build long-term relationships with developers by addressing their unique needs and challenges with specialized attention and resources. Our team operates across four core pillars: Marketing, Developer Relations, Gr

typescriptsqlaws
View job →

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe's Developer & End-user Experience Platform (DEEP) organization empowers all of Stripe's products with a shared product platform that helps with rapidly delivering high-quality, cross-product experiences across our UI and API surfaces. It focuses on providing a consistent and scalable developer experience that any developer (both internal and external) can leverage to accelerate a merchant's ability to create value using Stripe. Team matching — exact team matching for one of the subteams will begin during final stages. Please note we may also consider you for different orgs based on your experience, location, etc. More information on our team matching process can be found here . What you’ll do We're looking for Backend Engineers who want to make an impact on managing money at a global scale with a passion for building ergonomic APIs. You'll play a key role in extending our balance management platform and in building out a new funds accessibility platform that enterprises and SMBs use. Our team collaborates with many cross-functional teams—from Infrastructure to Product—at Stripe to deliver innovative solutions that address evolving user needs. Responsibilities Scope, design, build, and maintain APIs, services, and large-scale systems that reliably and efficiently handle billions of money movement requests Debug and solve critical production issues across services and multiple levels of the stack Mentor engineers to help them grow C

awsdockerkubernetes
View job →
O
1mo ago

About the Team The Platform Systems team at OpenAI operates at the intersection of cutting-edge AI and large-scale distributed systems. We build the engineering and research infrastructure required to train OpenAI’s flagship models on some of the world’s largest, custom-built supercomputers. Our team develops core model training software and works deep in the stack - spanning collective communication, compute efficiency, parallelism strategies, fault tolerance, failure detection, and observability. The systems we build are foundational to OpenAI’s research velocity, enabling reliable, efficient training at frontier scale. We collaborate closely with researchers across the organization, continuously incorporating learnings from across OpenAI into the evolution of our training platform. About the Role As a Software Engineer, Platform Systems, you will design and build distributed systems that provide visibility into large-scale training workloads and help operate them reliably at scale. You’ll work on failure detection, tracing, and observability systems that identify slow or faulty nodes, surface performance bottlenecks, and help engineers understand and optimize massive distributed training jobs. This infrastructure is critical to operating OpenAI’s training stack and is actively evolving to support new use cases and increasingly complex workloads. This role sits at the core of our training infrastructure, blending systems engineering, performance analysis, and large-scale debugging. In This Role, You Will Design and build distributed failure detection, tracing, and profiling systems for large-scale AI training jobs Develop tooling to identify slow, faulty, or misbehaving nodes and provide actionable visibility into system behavior Improve observability, reliability, and performance across OpenAI’s training platform Debug and resolve issues in complex, high-throughput distributed systems Collaborate with systems, infrastructure, and research teams to evolve platform

awsrestai
View job →

About the Team Revenue Platform sits at the intersection of customer experience, financial precision, and enterprise-grade reliability. We build both end-user experiences and the underlying platform capabilities that power invoicing, billing, payments, and revenue recognition across OpenAI. Our work spans high-leverage customer surfaces and deep, reusable platform primitives used by multiple teams. These are foundational systems that will support OpenAI’s growth for years to come, and we’re looking for engineers who care deeply about craftsmanship, correctness, and building platforms and experiences that scale gracefully and are a joy to build on. OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Achieving that mission requires more than great models, it demands a world-class commercial platform that enterprises can trust. The Revenue Platform team enables organizations to confidently adopt, scale, and manage their usage of OpenAI products by providing reliable, extensible financial infrastructure and shared services that power the broader ecosystem. We’re looking for a full-stack engineer to take end-to-end ownership of critical revenue workflows and the platform abstractions behind them. You’ll work across APIs, data models, shared services, and user interfaces to solve hard problems in scale, correctness, and usability. Partnering closely with Product, Finance, Sales, Operations, and other engineering teams, you’ll turn complex business requirements into durable platform capabilities, raising the bar on engineering quality and developer experience as our products and enterprise footprint grow. About the Role As a Full Stack Engineer on the Revenue Platform team, you will design, build, and operate platform services and user-facing interfaces that form the backbone of OpenAI’s commercial engine. You’ll collaborate with product, design, finance, and engineering partners to deliver intuitive customer experiences while also

typescriptpythonreact
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Platform Systems team at OpenAI operates at the intersection of cutting-edge AI and large-scale distributed systems. We build the engineering and research infrastructure required to train OpenAI’s flagship models on some of the world’s largest, custom-built supercomputers. Our team develops core model training software and works deep in the stack - spanning collective communication, compute efficiency, parallelism strategies, fault tolerance, failure detection, and observability. The systems we build are foundational to OpenAI’s research velocity, enabling reliable, efficient training at frontier scale. We collaborate closely with researchers across the organization, continuously incorporating learnings from across OpenAI into the evolution of our training platform. About the Role As a Software Engineer, Platform Systems, you will design and build distributed systems that provide visibility into large-scale training workloads and help operate them reliably at scale. You’ll work on failure detection, tracing, and observability systems that identify slow or faulty nodes, surface performance bottlenecks, and help engineers understand and optimize massive distributed training jobs. This infrastructure is critical to operating OpenAI’s training stack and is actively evolving to support new use cases and increasingly complex workloads. This role sits at the core of our training infrastructure, blending systems engineering, performance analysis, and large-scale debugging. In This Role, You Will Design and build distributed failure detection, tracing, and profiling systems for large-scale AI training jobs Develop tooling to identify slow, faulty, or misbehaving nodes and provide actionable visibility into system behavior Improve observability, reliability, and performance across OpenAI’s training platform Debug and resolve issues in complex, high-throughput distributed systems Collaborate with systems, infrastructure, and research teams to evolve platform

awsrestai
View job →

About the Team The Platform Analytics team builds the systems OpenAI researchers use to understand the quality and behavior of the models we train including what models are doing, why they behave in a particular way, and how that behavior changes across experiments. Neptune is a core part of this work. It ingests, stores, queries, and visualizes large volumes of metrics from pretraining, post-training, and reinforcement learning. Hundreds of researchers depend on these systems in their daily work to compare experiments, debug unexpected behavior, and decide what to try next. Our scope is broader than metrics. We also build platforms that help researchers analyze samples, traces, evaluation results, and other structured or unstructured data through dashboards, APIs, and increasingly agent-driven workflows. These systems need to remain fast, reliable, and understandable as the scale and complexity of research change quickly. We are not trying to become a consulting team that builds a separate solution for every research project. We work directly with researchers to understand recurring problems, then turn them into reusable infrastructure and platform capabilities that many teams can build on. About the Role We’re looking for a hands-on experienced software engineer who can take ownership of a critical system and drive it from problem definition through production adoption. This person should be able to own a platform such as CacheHouse end to end: define its technical direction, design its data model and storage architecture, integrate it with several research dashboards and workflows, guide one or two engineers, and ensure the system works reliably for its users. The right candidate should already bring the technical judgment, ownership, and execution expected at this level. The primary learning curve should be OpenAI’s stack and research problem space, not learning how to lead a complex engineering effort or deliver a production system. You will work directly with

awsrestai
View job →

A BOUT TIDE At Tide, we help SMEs save time and money in the running of their businesses by not only offering business accounts and related banking services, but also a comprehensive set of highly usable and connected administrative solutions, from invoicing to accounting. Tide is transforming the small business banking market and now supports over 2 million members globally across the UK, India, Germany and France. Using advanced technology, all solutions are designed with SMEs in mind. With quick onboarding, low fees and innovative features, we thrive on making data driven decisions to serve our mission: to help SMEs save time and money so they can get back to doing what they love. Tide facts: Tide is available for UK, Indian, German and French SMEs Over 2 million members across UK and India Over $300 million raised in funding Over 2,800 Tideans globally Recognised with Great Place to Work certification three years in a row, and among India’s Top 50 Best Workplaces in Banking, Financial Services, and Insurance in 2026 We have offices in Central London, with a member support and technology centre in Sofia, Bulgaria, technology centres in Serbia, Romania, Lithuania and Hyderabad and offices in Gurugram, New Delhi, Berlin, Paris and Luxembourg ABOUT THE ROLE Tide is hiring a Senior Staff Software Engineer to lead the architecture of our agentic platform. You will shape the shared capabilities that allow AI systems to operate safely, reliably, and at scale across Tide. That includes context, orchestration, tool execution, trust controls, observability, and evaluation. You will participate in key build versus buy decisions, integrate external components where they accelerate us, and ensure we own the parts that matter most for Tide’s trust, data advantage, and long-term platform leverage. WHAT YOU WILL DO Define the architecture for Tide’s agentic platform Drive the design of shared services such as context APIs, tool layers, policy controls, and auditability Partner w

pythonjavaangular
View job →

A BOUT TIDE At Tide, we help SMEs save time and money in the running of their businesses by not only offering business accounts and related banking services, but also a comprehensive set of highly usable and connected administrative solutions, from invoicing to accounting. Tide is transforming the small business banking market and now supports over 2 million members globally across the UK, India, Germany and France. Using advanced technology, all solutions are designed with SMEs in mind. With quick onboarding, low fees and innovative features, we thrive on making data driven decisions to serve our mission: to help SMEs save time and money so they can get back to doing what they love. Tide facts: Tide is available for UK, Indian, German and French SMEs Over 2 million members across UK and India Over $300 million raised in funding Over 2,800 Tideans globally Recognised with Great Place to Work certification three years in a row, and among India’s Top 50 Best Workplaces in Banking, Financial Services, and Insurance in 2026 We have offices in Central London, with a member support and technology centre in Sofia, Bulgaria, technology centres in Serbia, Romania, Lithuania and Hyderabad and offices in Gurugram, New Delhi, Berlin, Paris and Luxembourg ABOUT THE ROLE Tide is hiring a Senior Staff Software Engineer to lead the architecture of our agentic platform. You will shape the shared capabilities that allow AI systems to operate safely, reliably, and at scale across Tide. That includes context, orchestration, tool execution, trust controls, observability, and evaluation. You will participate in key build versus buy decisions, integrate external components where they accelerate us, and ensure we own the parts that matter most for Tide’s trust, data advantage, and long-term platform leverage. WHAT YOU WILL DO Define the architecture for Tide’s agentic platform Drive the design of shared services such as context APIs, tool layers, policy controls, and auditability Partner w

pythonjavaangular
View job →
DU
DoorDash USA
📍 San Francisco• Full-time• From $1.6M/yr
18 days ago

About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company at significant scale and continue to expand the workloads, capabilities, and consumer base we serve. Orchestrating and operating thousands of Spark cluster deployments is a complex distributed system problem which the team invests heavily in runtime optimization, systems architecture, multi-tenant scheduling, and end-user tooling. About the Role As a Software Engineer on Spark Platform, you will execute across the surfaces of our in-house Spark deployment that serves the entire company. The work spans Spark runtime upgrades and performance, multi-tenant scheduling and executor bin-packing on Kubernetes, cluster lifecycle automation, and the observability and incident automation that keep the platform sustainable. You will move between layers as the work demands — picking up the next high-leverage problem regardless of where it sits — and partner closely with the rest of the team and with platform consumers across the company. You must be located in San Francisco, Sunnyvale, Seattle, or New York City for this hybrid position. You will report into the Engineering Manager on our Spark Platform team. You're excited about this opportunity because you will… Build and operate an in-house Spark platform that runs at company-wide scale, spanning runtime, scheduler, reliability, and user-facing tooling. Drive multi-tenant scheduling, executor bin-packing, and cost-aware placement that let a small team serve dozens of consumer teams. Own pieces of cluster lifecycle automation — provisioning, upgrades, capacity changes, and node-failure handling — at a scale where these stop being manual events. Build the observability and incident automation that make the platform debuggable end-to-end and keep on-call sus

pythonjavasql
View job →
N
Nvidia
📍 Santa Clara, United States
1mo ago

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. We are looking for an excellent engineering manager to own and deliver an end to end manageability stack for Data Center Systems. We are seeking an experienced manager who is deeply technical, hands-on, and has a wide system view. You will manage a team of experts, design & build OpenBMC based manageability software stack for NVIDIA’s next generation Data Center Compute Systems. We want to grow our teams with the smartest people in the world. If you're creative and autonomous, we want to hear from you! What you’ll be doing: Own and deliver OpenBMC based manageability stack for next generation Data Center Compute Systems. Own firmware delivered to data centers in terms of quality, reliability and telemetry performance. Manage and lead a distributed team of software engineers to deliver firmware stack with high quality. Work with data center architects and cloud customers for correct requirements and scope implementation to ensure speed of light product development. Work closely with cross functional teams to ensure scalable manageability architecture for all data centers products Drive efficiency, reliability and optimization in firmware architecture from a data center view point. Work closely with customers and internal teams to resolve issues at Speed of Light. What we need to see: BS, MS, or PhD in EE/CS or related field o

pythongitai
View job →
R
Replit
📍 Foster City• Full-time
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Team Product Platform builds and owns the shared foundations the rest of Replit is built on, spanning the full stack so every other team can ship features safely and quickly: backend infrastructure, connectors, product primitives, and the frontend platform. Our work is high-leverage and horizontal: when our foundations are solid every other team moves faster, and the role gives you exposure across the whole of engineering. We are a small, collaborative team that values curiosity and clear thinking over pedigree, and we work in the open by bringing each other the problem rather than just the request. We care more about how you reason and build than the route you took to get here. About The Role As a Product Engineer , you can focus on frontend, backend, or full-stack work building the shared systems other teams depend on. The work is guided by a few simple questions: Are our shared systems fast, reliable, and cost-efficient as traffic grows? Are we making product development safe by default, consistent, and faster? Can a builder connect a third-party service once and have it work safely across every app they build? Are user-facing surfaces consistent and fast, with shared primitives teams can build on? Is our codebase easy to navigate, change, and extend, including for AI coding agents? What you’ll do Design reusable primitives and interfaces with clear contracts and documentation that other teams adopt Work directly with product teams to turn their friction into platform improvements Profile and instrument shared systems, then ship the improvements that move latency, cost, and reliability Harden systems against failure and abuse, and make safe defaults the path of least resistance Set technical direction in a

typescriptreactnode.js
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. Replit is a software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit helps more people build and ship software. About the Team Product Platform builds and owns the shared foundations the rest of Replit is built on: backend infrastructure, connectors, product primitives, and the frontend platform. When these foundations are solid, every other team moves faster. This role goes deep on the frontend platform: the architecture and platform layer behind every core product surface. The work is high-leverage and horizontal, and gives you exposure across the whole of engineering. We are a small, collaborative team that values curiosity and clear thinking over pedigree, and we care more about how you reason and build than the route you took to get here. If you like making other engineers faster, you will fit in here. About The Role As a Product Engineer focusing on the frontend platform , you will own the frontend architecture behind core product experiences: application frameworks, the API and data layer, testing infrastructure, and client performance. The goal is simple: product teams ship quickly and reliably on what you build. The team’s work is guided by a few simple questions: Is our core frontend architecture (frameworks, state, routing, SSR/CSR) sound, consistent, and easy to build on? Is our API and data layer reliable and ergonomic, with clear contracts, sensible error handling, and effective caching? Are user-facing surfaces fast and well-instrumented, with testing infrastructure that keeps them safe to change? Is the codebase easy to navigate, change, and extend, including for AI coding agents? You’ll partner closely with engineering, produc

typescriptreactnode.js
View job →
🔔

Get new platform engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More platform engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.

Countries hiring Platform Engineer

Country links use the same curated canonical inventory as Jobiba sitemaps.