About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives, spanning AI research specialists, silicon designers, software engineers and systems architects. Job Summary We are looking for an experienced Principal Engineer to join our System Management team and help lead the development of critical interfaces used by internal and external customers to manage system state. You will provide technical leadership within assigned areas of System Management, guide architecture and implementation choices, mentor engineers and translate broader technical direction into effective execution. This is a hands-on engineering role for someone who can lead complex technical work, improve reliability and operational readiness, and collaborate effectively across multiple engineering disciplines. The Team The System Management team sits within the Software Platform group and helps build Graphcore products into large-scale AI solutions for our customers. The team is responsible for developing the interfaces between hardware, AI software and frameworks, as well as providing interfaces for public and private cloud environments. This includes system management capabilities that abstract complex hardware administration and enable reliable deployment and operation at scale. As one of the first teams to work with new hardware and software, we regularly solve complex system-level problems
Jobs in United States
Lead Software Engineer Environment Platform Specialist in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead software engineer environment platform specialist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Hiring demand
58/100
steady · 483 related jobs
Hiring trend
-73.2%
Job postings compared with the previous 30 days
Remote options
13.3%
Share of matching jobs listed as remote
Typical salary
$210K – $210K/yr
Based on 179 salary observations
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Senior Principal Network Engineer to help design, deploy, and optimize next‑generation AI data center networks. AI training and inference workloads require extremely high bandwidth, deterministic low latency, and zero‑packet‑loss networking environments. In this role, you will partner closely with the Network Architecture Lead to design and scale high‑performance computing (HPC) network fabrics supporting GPU clusters. You will work across hardware, networking, and AI application layers to ensure Graphcore’s large‑scale AI infrastructure operates at peak performance. The ideal candidate brings deep experience operating hyperscale or HPC data center networks and has expertise in high‑speed Ethernet fabrics, RDMA technologies, advanced automation, and telemetry systems. The Team The Data Center Network Engineering team designs and operates the high‑performance network fabrics that power Graphcore’s AI compute platforms. The team collaborates closely with hardware engineering, AI researchers, and infrastructure teams to build scalable networking environments optimized for distributed training and infe
$153.1K – $293.8K/yr · Jobiba est.
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: This is a Product Engineering role specialized in monetization systems, and will lead that space at Replit. It’s a direct line to business impact. But it’s also critical to get right for our users. These are some of the most critical user journeys to get right. Getting them wrong creates the most frustrating experiences for users. So we’re looking for engineers who can build reliable and scalable billing systems and abstractions, while also translating that to an intuitive and friendly user experience. You will: Lead the architecture and implementation of monetization systems at Replit. Create seamless payment experiences for users for both product-led and sales-led motions. Build new abstractions and APIs for other engineers at Replit to monetize their new products. Iterate on pricing and packaging tactics to drive revenue growth. Examples include coupon codes and referral systems. Create monitoring and feedback systems so that we can proactively spot problems, fix them, and optimize performance. Required skills and experience: 6+ years of engineering experience, with strong skills working on the backend. Direct working experience in at least one of the following: Subscription platforms Usage-based billing SaaS Taxation Payment platforms Tokenization Self-directed and comfortable working autonomously in ambiguous environments. Excellent problem-solving skills with ability to debug complex billing issues and edge cases. Experience implementing customer-facing billing interfaces that simplify complex pricing structures. Tools + Tech Stack for this role: Python, TypeScript, React, Postgres, GraphQL, and Nodejs. Bonus Points : Experience working with Orb, Metronome, or Stripe usage based billing. Experienc
From $218K/yr
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Staff Software Engineer on the Core Automation team within the Platform group, you'll architect and build the Agentic AI systems that are transforming how Coinbase operates. This team is reimagining customer support and compliance processes for a fully AI-driven world, designing intelligent agents, orchestration frameworks, and measurement systems that deliver delightful customer experiences at scale. You'll own the technical direction for production AI systems, working across cross-functional teams to bring this vision to reality while building primitives that scale automation across the company. What you'll do: Architect and build Agentic AI systems that power Coinbase's compliance automation and other Operations, from intelligent agents through orchestration and guardrails Design foundational APIs and measurement frameworks that ensure AI agents are grounded, relevant, and reliably deliver customer delight with minimal hallucination Lead technical direction for distributed systems underpinning AI automation, defining architecture patterns and strategic roadmaps in partnership with engineering leadership Build reusable primitives and orchestration solutions that enable AI-powered automation to scale across multiple domains beyond the initial customer support and compliance focus Mentor engineers on AI system design techniques, coding standards, and production-
$153.1K – $293.8K/yr · Jobiba est.
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. Replit is building the world’s most accessible AI coding agent. Replit Agent can be used by anybody to bring their ideas to life. Whether it’s an app for yourself, the next great startup idea, or a tool to make you more productive at work, Replit Agent can help build it. Replit builds complete apps better than anybody thanks to our full suite of services that handle app integrations, storage, hosting, analytics, and more. We don’t just build apps in development, we handle the full lifecycle into production and beyond. About the role: Help power the development of Replit Agent as a technical leader for the Replit Cloud organization. You will report to the Vice President of Engineering. The Replit Cloud team builds Replit’s first party cloud infrastructure so users can build, scale, and succeed entirely on Replit. They manage databases, application storage, app publishing and hosting, development/production environment splitting, custom domains, and more. By having a set of first party services that integrate seamlessly, you will power one of Replit’s key product differentiators. You will: Help lead major projects, either by taking new products from 0->1 or doubling down on our first party primitives to keep winning users. Work closely with designers and product managers, to quickly iterate on Replit Cloud to continually grow and improve the product. Identify the hardest technical and/or quality problems holding us back, and then build solutions. Mentor and develop new senior engineers to help grow the team. Ship product and build infrastructure as a true full stack builder using: TypeScript, React, CSS, Postgres, Go, and Terraform. Examples of what you could do: Leverage our unique cloud infrastructure to build diffe
$153.1K – $293.8K/yr · Jobiba est.
About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu
$190.4K – $285.6K/yr
Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Lead the technical design and architecture of major platform initiatives, author design documents and build consensus across engineering teams. Define technical roadmaps for complex, multi-quarter projects that span multiple teams. Make critical architectural decisions for company documentation infrastructure, balancing scalability, reliability, and developer experience. Evaluate and set direction for integrating emerging technologies, including AI/LLM capabilities, into company documentation platforms and authoring tools. Establish and evolve engineering standards, best practices and technical guidelines for the team and broader organization. Partner with engineering teams across the company to understand documentation needs and design integrated solutions. Design, build and maintain scalable, reliable and performant services and systems. Contribute high-quality code across the full stack and navigate codebases with different languages and tools. Debug and resolve complex production issues and improve system reliability. Take ownership of system health and incident response. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Software Engineering, Engineering, or a related field, plus four (4) years of experience in Software Engineering. Must have four (4) years of experience in each of the following: - Working in a full stack environment with a foc
$170K – $250K/yr
At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You'll Play At Mindbody, Core Engineering builds and evolves the foundational systems that help our products run reliably at scale. In this staff role, you’ll bring clarity to complex technical problems, guide architecture, and strengthen how we design, deliver, and operate the backend services that power real-world experiences. Lead cross-team technical execution, aligning architecture and delivery across multiple squads and core domains Design and evolve microservices patterns that improve reliability, performance, and maintainability Drive cloud and deployment improvements across AWS, Mindbody’s cloud platform, and our containerized deployment environment Partner with engineering and product leaders to turn ambiguous problems into clear technical plans and milestones Establish and socialize standards for service design, APIs, and relational data modeling (SQL) Strengthen monitoring and operational visibility using New Relic and Kibana, turning insights into durable system improvements Mentor and unblock engineers through design reviews, pairing, and practical guidance Reduce technical risk and complexity while balancing
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Engineering Acceleration team delivers a world-class development experience to our 2000+ Roblox teammates who write, build, and ship the most incredible virtual experience platform in the world. As a Principal Software Engineer, you will set the technical direction for transforming development at Roblox from its solid roots to a fully scaled, globally mature, AI-powered development, code review and CI environment. If you are passionate about engineering productivity and you've ever wanted to define — not just build — the systems that fundamentally change how an entire engineering organization operates, this is your opportunity. You Will: Own, scale, and define the long-term technical vision for Roblox's source code search, storage, review, and CI infrastructure for the next 10 years of growth. Architect software that improves the lives of your coworkers — spanning local development tools, internal web apps and microservices, reusable CI workflows, productivity metrics toolchains, and integrations across many 1st and 3rd party tools. Lead and drive major company-wide efforts improving the productivity and user experience of our Source and CI tools, setting direction across teams a
$153.1K – $293.8K/yr · Jobiba est.
This is where your work makes a difference. At Baxter, we believe every person—regardless of who they are or where they are from—deserves a chance to live a healthy life. It was our founding belief in 1931 and continues to be our guiding principle. We are redefining healthcare delivery to make a greater impact today, tomorrow, and beyond. Our Baxter colleagues are united by our Mission to Save and Sustain Lives. Together, our community is driven by a culture of courage, trust, and collaboration. Every individual is empowered to take ownership and make a meaningful impact. We strive for efficient and effective operations, and we hold each other accountable for delivering exceptional results. Here, you will find more than just a job—you will find purpose and pride. This is where your work saves lives As a Principal Android Software Engineer, you will lead architecture and delivery of shared Android capabilities that power multiple clinical product applications worldwide. You will own significant portions of the Android system design—reusable services and libraries, hybrid WebView/React UI foundations, device connectivity patterns, and common clinical workflows—while mentoring engineers and raising the bar for quality in a regulated medical-device environment. An ideal candidate brings deep hands-on Android expertise, a clear technical vision for multi-product platform software, and the ability to drive complex cross-team decisions independently. Strong communication and collaboration with product teams, systems, hardware, and other platform partners are essential. What you'll be doing: Lead Platform Android Architecture and Delivery: Define and evolve shared Android platform architecture (Kotlin services, DI, WebView/JS bridge patterns, messaging/broker integr
From $244K/yr
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Job: Airbnb Payments team allows any two people in the world to frictionlessly exchange money with easy to use payments services. It is a core strategy to fulfill Airbnb's belongs anywhere mission. We are building a world-class payments platform that moves billions of dollars, in 191 countries, with 75 currencies, through a complex ecosystem of payments partners. We build and maintain our own in-house global payments platform because no solution exists with the global reach needed. As the platform grows we'll be adding new payment partners, global licenses, compliance and regulation controls, and building new payment experiences for our guests and hosts. Airbnb's business grows rapidly year over year, and so does Airbnb Payments' processing volume. Payments Compliance is an essential part of the Payments organization. Without it, Airbnb can be subject to fines and even forbidden from conducting business in certain markets. As a Senior Staff Engineer and Technical Lead for the Payments Compliance organization, you would own the technical vision and architectural direction across the full Compliance engineering landscape — spanning Policy Enforcement, Identity, Screening, Auditing, and Compliance Experience. You will ensure our compliance systems are cohesive, scalable, and aligned with broader company priorities as we navigate an increasingly complex global regulatory environment. The Role: We are looking for a seasoned technical leader who can operate at the intersection of deep systems expertise and organizational influence. As a Senior Staff Engineer, you will serve as the connective tissue across Complia
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking a highly experienced Principal / Director-level Full-Stack Software Development Engineer to join our Digital Caremark organization and lead the architecture, design, and delivery of next-generation digital solutions. This role spans AI-enabled applications, scalable digital platforms, and enterprise integrations that power critical healthcare and pharmacy experiences. This is a senior technical leadership role for a hands-on engineer who can operate across the full stack—from intuitive front-end applications to resilient backend services—while setting architectural direction, influencing engineering standards, and mentoring teams. The ideal candidate combines deep technical expertise, platform thinking, and strong collaboration skills to build secure, scalable, API-first solutions in a highly regulated environment. Key Responsibilities: Architecture & System Design • Define and drive architecture for large-scale distributed systems and digital platforms • Lead design reviews and set architecture standards and best practices • Champion API-first, microservices, and event-driven architecture patterns • Ensure systems meet scalability, reliability, security, and compliance requirements • Balance performance, cost, and speed in technical decision-making Fu
From $399.4K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Distinguished Machine Learning Engineer/Technical Director in the Safety organization at Roblox, you will drive the overall technical vision and execution for all machine learning initiatives focused on maintaining the safety and civility of our users. Our industry-leading safety features ensure Roblox remains a safe and inclusive environment for our community to express themselves creatively and share experiences without fear. The Safety org is the reason why Roblox is the safest place on the internet, protecting users You will provide technical leadership on AI/ML efforts for Trust and Safety as the Roblox platform scales to serve different age groups and geographic locations. You Will: Own the technical direction and implementation of machine learning solutions for safety-related systems Lead and mentor other engineers, fostering a culture of technical excellence and inclusivity Break down long-term product requirements into iterative deliverable stages, ensuring continuous improvement Craft and build large-scale machine learning models with billions of parameters, ensuring production-readiness Facilitate challenging technical decisions across multiple teams, demonstrating empathy a
C$100K – C$500K/yr
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a SoC Physical Design Verification Engineer to drive full-chip signoff and ensure manufacturable, high-quality silicon across advanced technology nodes. You’ll lead physical verification closure (DRC, LVS, ERC, etc.), debug issues using standard industry PV tools, and collaborate across RTL, PD, CAD, and packaging teams to achieve successful tapeouts. If you thrive in a fast-paced environment and enjoy solving complex challenges in cutting-edge silicon, we’d love to hear from you. This role is hybrid , based out of Santa Clara, CA or Austin, TX or Fort Collins, CO. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A seasoned engineer with a strong background in CPU/IP/SoC physical verification and tapeout closure. A hands-on problem solver who excels at debugging and driving signoff through complex verification flows. A collaborative team player who works effectively across RTL, PD, CAD, and foundry interfaces. A mentor and technical leader passionate about building efficient, manufacturable silicon. What We Need BS/MS in Electrical/Electronics Engineering (or related) with 7–14 years of hands-on CPU/IP/SoC physical verif
From $244K/yr
Datadog’s Cloud Networks team designs, builds, and maintains the production network infrastructure that powers everything built on top of our platform across AWS, GCP, Azure, and beyond. In this role, you’ll set technical direction for how we scale our multi-region, multi-cloud network footprint while keeping reliability and performance high. You’ll partner closely with internal teams and Cloud Service Providers to troubleshoot complex connectivity issues, integrate new networking capabilities, and improve the foundations our engineers and customers rely on. This is a high-impact opportunity to drive meaningful improvements in scale, resiliency, and cost efficiency. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Design, build, and operate cloud network infrastructure across AWS, GCP, Azure, and Neoclouds in a multi-region environment. Own connectivity between clouds, customers, and developers—ensuring scalable, secure, and reliable network paths. Set clear technical direction for expanding data centers and evolving the network while maintaining stability and performance. Improve cross-site and cross-region connectivity patterns to support Datadog’s growing platform needs. Lead deep investigations into latency, packet loss, and connectivity failures – from pcap and path analysis through to escalations with cloud providers that may originate from customer support Identify and deliver network-related efficiency and cost-saving opportunities that positively impact business health. Who You Are: You have deep networking expertise. You understand BGP, route policies, path selection, prefix advertisement, and what breaks in large-scale networking. You have substantial experience designing, building, and evolving large-scale Software-Defined Networks—inclu
Higher-paying openings
Jobs with higher listed pay
Staff Software Engineer - Fern
Postman · New York, California, United States
$3M – $3.7M/yr
Staff Software Engineer, Business Platform
Postman · San Francisco, California, United States
$2.9M – $3.6M/yr
Principal Software Engineer
Roblox · San Mateo, CA, United States
From $3.5M/yr
Principal Software Engineer, Game Safety
Roblox · San Mateo, CA, United States
From $3.5M/yr
Staff Software Engineer- Codegen
Postman · Austin, Texas, United States
$2.5M – $3.2M/yr
Sr. Staff Software Engineer, Merchants
Pinterest · San Francisco, CA, US
From $2.9M/yr
Related career options
Similar roles with stronger pay
Demand 35/100 · 7 jobs
$4.6M – $4.6M/yr
Salary →Demand 39/100 · 15 jobs
$1.8M – $1.8M/yr
Salary →Demand 48/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 45/100 · 5 jobs
$840K – $840K/yr
Salary →Demand 45/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 39/100 · 13 jobs
$345K – $345K/yr
Salary →Other cities to consider
More places hiring for this role
Get new lead software engineer environment platform specialist jobs in United States by email
Daily job updates · Unsubscribe anytime