At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Join our ML Feature Store team where we're building cutting-edge product capabilities that power complex feature transformations and low latency feature serving. We're revolutionizing machine learning feature management and serving capabilities as part of the Snowflake ML suite of products. In the era of GenAI and agents, our team delivers high-quality, fresh feature solutions that make a real difference for our customers. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Help define and own the roadmap for Snowflake Feature Store, working collaboratively with senior architects and ML team leadership Build and execute a vision for incorporating new advances in machine learning Ensure operational excellence of services and meet reliability, availability, and performance commitments Collaborate across ML partner teams to improve development velocity and capabilities Support team members in delivering high technical quality WE WOULD LOVE TO HEAR FROM YOU IF YOU HAVE: 10+ years of experience in designing and building data serving infrastructure and/or machine learning platforms. Strong track record working with machine learning systems and platforms. Strong understanding of computer science fundamentals. B.Sc . in Computer Science Fluency in Ja
Jobs in United States
Platform Engineering Lead in United States
3,618 active opportunities · Updated October 2026
Showing
15 jobs
Explore current platform engineering lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that deliver the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing some of the world's most transformative technologies. Our AI Engineering Campus in Austin plays an important role in building the future of AI computing. The Opportunity As Technical Services Director, you will lead the teams that operate and evolve Graphcore's engineering labs, high-performance computing (HPC) platforms, and data center environments globally. You will be accountable for reliable, secure, cost-effective infrastructure that supports demanding engineering, AI, silicon-development, and validation workloads. This role combines people leadership, infrastructure strategy, operational excellence, capacity and financial planning, procurement, and program delivery. You will partner with Engineering, Information Technology, Security, Finance, Facilities, Supply Chain, customers, and external suppliers. The position is based onsite in Austin and requires travel to company facilities, data centers, and supplier locations, including international travel. What You'll Do Lead, recruit, mentor, and develop the systems administration, lab operations, and technical services teams responsible for the facility supporting global Engineering and Research and Development. Own the reliability, efficiency, protection, safety, supportability, and continuous improvement of engineering labs, HPC systems, and infrastructure facilities. Establish service levels, operating standards, escalation paths, performance measures, monitoring, observability, automation, ticketing, and configuration-management practices. Translate engineering and customer requirements into infrastructure roadmaps, capacity p
NVIDIA is a leading artificial intelligence computing company, and we are paving the way with innovations in self-driving cars, machine learning, supercomputing, gaming, and visualization. We give automakers, tier-1 suppliers, automotive research institutions, and start-ups the power and flexibility to develop and deploy breakthrough artificial intelligence systems for self-driving vehicles. Our unified computing architecture enables training deep neural networks in the data center, and then seamlessly runs them on NVIDIA DRIVE Platforms inside the vehicle. The Hypervisor and RTOS Team within NVIDIA DRIVE Software plays a critical role in NVIDIA's expansion into the world of artificial intelligence and autonomous vehicles. Our job is to facilitate the sharing and separation of system resources while achieving real-time, safety, and security requirements. We develop Hypervisor and RTOS with a strong focus on automotive quality, safety and security needed for the real-time, highly available system level components of world-class Autonomous Vehicles. We are making extensive use of formal methods to automate our workflow and increase the quality of our SW. We are hiring now for the position of Senior System Software Engineer for Hypervisor and RTOS What you’ll be doing: Design and develop new features for RTOS and hypervisor software stack. Bring up and optimize RTOS and hypervisor stacks on new NVIDIA Tegra SoCs. Develop high-integrity software using best-in-class engineering, safety, and security practices. Debug complex system-level issues across hardware, firmware, RTOS, and virtualization layers. Lead team-wide technical initiatives by building alignment, coordinating execution, and driving them to completion. What we need to see: BS, MS in CS/CE/EE or a related engineering field or equivalent experience </
About the job Lead the team that proves our AI silicon performs reliably before it reaches customers. You will build and lead Graphcore's characterisation capability for next generation silicon and system platforms. Your work will help ensure our products perform consistently across real world conditions. You will define bring up and characterisation strategies, lead technical execution, and shape the lab infrastructure needed for success. You'll work across silicon, hardware, manufacturing, architecture and product teams to solve complex engineering challenges. This is a hands on leadership role with the opportunity to influence both product design and how Graphcore validates future AI systems. The team and culture This is a newly formed team within Manufacturing Operations. You'll have the opportunity to establish how the team works while building strong partnerships across engineering and operations. Day to day, you'll work closely with architecture, silicon, hardware, production test and product teams. Decisions are driven by data, technical evidence and close collaboration across disciplines. We value ownership and clear communication. You'll be trusted to lead technical direction, remove blockers and help teams make progress with confidence. What we're looking for Essential Proven track record of delivering complex technical projects as an individual contributor, manager, or project manager, with the ability to work independently and drive execution. Strong expertise in silicon digital device design, bring-up, characterisation, and silicon process technologies, with an understanding of their impact on transistor- and system-level performance. In-depth knowledge of high-performance processors, system-on-chip (SoC) architectures, and high-speed digital interfaces such as PCIe, Ethernet, and DDR. Experience with measurement automation, data analysis, and scripting/coding to develop automated test and analysis workflows, with familiarity of ATE systems and t
NVIDIA is seeking a strong technology leader to manage our Server Software Technical Program Management (TPM) team. This role is at the cross-section of execution and strategy, leading a team of Senior TPMs who drive the firmware and system software for NVIDIA's next-generation server platforms like DGX, MGX, and HGX. These platforms bring together the full power of NVIDIA GPUs, NVLink, InfiniBand networking, Grace CPUs, and our optimized AI/HPC software stack. This deep technical leadership role focused on the Software Development Processes that brings new server hardware to life. What you'll be doing: Lead a team of TPMs driving the technical software and firmware execution for NVIDIA's NPI (New Product Introduction) and sustaining engineering teams. Drive the end-to-end SDLC for low-level server components, including firmware (BMC, UEFI/BIOS), drivers, and system management software, ensuring alignment with hardware schedules. Collaborate closely with NVIDIA product management and hardware engineering teams to define release plans and program objectives. Build a strong connection and feedback loop between sustaining and NPI engineering teams to improve product quality and development velocity. Lead process improvement initiatives and help propagate SDLC standards across multiple engineering and TPM organizations. You will have the opportunity to interact with diverse technical groups, spanning all organizational levels. What we need to see: Bachelor of Science (or equivalent experience) or Master of Science degree in Computer Science, Electrical Engineering, or related field. 12+ overall years of experience developing and leading complex low-level or system software projects. and 7+ years of experience in a people management role. Deep understanding of system a
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Security Systems Engineer at Palantir, you are responsible for implementing, designing, and maintaining the physical security systems that ensure the protection of Palantir’s people, assets, intellectual property, and reputation. You lead research and evaluations for the latest technology, security infrastructure, policies, procedures, systems, and applications, looking for easy-to-use capabilities that reduce button clicks and ultimately give time back to the operator. You will be collaborating with technical teams to develop and maintain end-to-end project plans and ensure on-time delivery. Your technical expertise is second only to your integrity and genuine passion for security and technology. As a member of the Global Security and Investigations team, you work with security professionals, architects, lawyers, and software and network engineers to deliver a safe and secure environment for your fellow employees. Palantir is a global company, so you'll lead physical security projects all over the world. You'll work hard to break down barriers between teams by communicating your knowledge and goals effectively across time zones and other physical and virtual barriers. *Please note this role is NOT for cyber security or information security positions* Core Responsibilities Design and oversee the technical aspects of your projects from conception to full deployment. Systems naturally break and need maintenance. You'll work with vendors and end users to find the best solutions for a problem, in the shortest possible time. Leverage both system and software engineering skills in order to address the needs of all teams within Globa
Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Position Overview The Manager, Enterprise Analytics & Reporting leads a multidisciplinary team of analytics and reporting developers, business intelligence professionals, and data engineers responsible for developing, maintaining, and governing enterprise analytics, reporting, and data solutions across Tableau, Epic Clarity, Databricks, and other approved platforms. This role combines people leadership, technical expertise, and business partnership. The Manager works with commercial, clinical, operational, and technology stakeholders to translate business needs into scalable dashboards, reports, data products, and analytical solutions. The role is also responsible for establishing reporting standards, improving the usability and consistency of the reporting environment, managing demand and priorities, and ensuring the organization can confidently use data to support decision-making. The Manager intentionally grows the engineering capabilities of the team over time, helping reporting and business intelligence developers build the skills and practices needed to deliver increasingly scalable, reusable, automated, and engineering-oriented data and analytics solutions. This role requires consistent onsite presence in Madison, WI up to four days a week. Essential Duties Include, but are not limited to, the following: Lead, coach, and develop a multidisciplinary team of analytics and reporting developers, business intelligence professionals, and data engineers suppo
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Actuator Gear Design Engineer to lead the development of custom gears and gear stages for advanced robotic systems. You will own actuator development from early architecture and concept generation through prototype validation and system integration, partnering closely with mechanical, electrical, controls, firmware, and reliability teams. You will partner with external suppliers and internal manufacturing to create full gearbox assemblies. This role focuses on the design, integration, and validation of precision gearing, including broader knowledge around motor electromagnetics, transmission types, sensing, structural components, and thermal architectures. You will help drive actuator development across the full engineering lifecycle while establishing scalable design, test, and integration practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will: Lead the architecture, design, and integration of custom robotic actuator gearing. Define actuator requirements and system-level trade studies around torque density, bandwidth, efficiency, thermal performance, back drivability, inertia, reliability, manufacturability, and cost. Design precision electromechanical assemblies with strong attention to tolerances, alignment, load paths, thermal expansion, sealing, wear, and serviceability. Drive actuator integration into robotic systems, partnering closely with controls, firmware, electrical, and robotics software teams to optimize closed-lo
Become a part of our caring community The Principal Storage Engineer is a senior technical leader responsible for defining and advancing the enterprise storage architecture and long-term data infrastructure strategy. This role establishes standards, develops five-year technology roadmaps, and designs secure, resilient, scalable, and cost-effective storage platforms for business-critical, analytics, and artificial intelligence workloads. The engineer serves as the organization’s storage subject-matter expert and partners with infrastructure, cloud, security, data, application, architecture, finance, and vendor teams to translate business requirements into sustainable technology capabilities. The Principal Storage Engineer is a senior technical leader responsible for defining and advancing the enterprise storage architecture and long-term data infrastructure strategy. This role establishes standards, develops five-year technology roadmaps, and designs secure, resilient, scalable, and cost-effective storage platforms for business-critical, analytics, and artificial intelligence workloads. The engineer serves as the organization’s storage subject-matter expert and partners with infrastructure, cloud, security, data, application, architecture, finance, and vendor teams to translate business requirements into sustainable technology capabilities. Key Responsibilities Define the enterprise storage vision, reference architecture, engineering standards, and five-year roadmap across block, file, object, software-defined, and hybrid storage services. Lead architecture decisions for on-premises AI infrastructure, including high-throughput and low-latency storage fo
From $151K/yr
The Role MongoDB is seeking a Staff Product Manager to lead strategic initiatives across Identity and Security. In this role, you will serve as the senior product leader shaping the next generation of authentication, authorization, and access control across MongoDB Atlas, the core database engine, and emerging developer platforms. You will lead the strategy for non-human identity—spanning workload identity federation, zero-trust passwordless access, Model Context Protocol (MCP) integrations, and autonomous AI agentic identity. You will partner closely with engineering, UX, product marketing, and enterprise customers to build identity experiences that are enterprise-grade, seamless for developers, and secure by default. Responsibilities Set and champion a multi-year product strategy across interrelated identity areas, aligning dependencies and investments to MongoDB’s broader objectives Identify cross-cutting customer and market opportunities, build business cases for investment, and bring a clear point of view to senior stakeholders Lead workload identity federation and programmatic access initiatives across cloud and enterprise identity environments Define secure authorization approaches for AI agents and MCP clients, including delegation, identity lifecycle, policy, and administrative controls Drive developer-first access management, ensuring security controls enhance rather than obstruct developer velocity Partner with engineering, design, GTM, and cross-functional leadership to validate enterprise requirements, prioritize initiatives, and measure feature adoption Stay ahead of evolving cloud security paradigms, emerging standards (e.g., SPIFFE/SPIRE, O4AA), and AI access patterns to guide executive decision-making Represent MongoDB’s security and identity vision at industry events, customer advisory boards, and executive briefings Demonstrate creativity, an innovative spirit, and a bias towards action Requirements 8+ years (or equivalent experience) in Pro
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a Senior Mechanical Engineer to lead the design, integration, and sustaining engineering of mechanical subsystems in robotic platforms. You will work closely with experienced engineers and cross-functional partners to set functional requirements, develop and iterate on hardware to meet program expectations. This role is aimed at candidates with strong fundamentals in mechanical system design — including tolerance, alignment, load paths, wear, and failure modes — and robotics. You will contribute to real hardware programs moving from prototype through early production. You will lead both new subsystem development and ongoing improvements to existing systems based on testing, field performance, and manufacturing feedback. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will Lead the design and iteration of mechanical subsystems, including structures, mechanisms, and actuators. Create and maintain CAD models, assemblies, and drawings with appropriate tolerancing and documentation. Build and test prototypes, supporting debugging of mechanical issues such as fit, alignment, friction, and wear. Assist in developing test methods and executing validation to evaluate performance, durability, and failure modes. Work with cross-functional teams to integrate mechanical components with sensors, actuators, and control systems. Support transition of designs from prototype to manufacturable assemblies, incorporating DFM and DFA considerations. Collaborate with manufacturing partners
About ElevenLabs ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like Deutsche Telekom and Meta. Our investors are some of the world's most prominent, including Andreessen Horowitz, ICONIQ Growth and Sequoia. We've raised $781M in funding and our last valuation was $11B - multiples of 11, always. We have expanded from voice into three main platforms: ElevenAgents enables businesses to deliver seamless and intelligent customer experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale. ElevenCreative empowers creators and marketers to generate and edit speech, music, image, and video across 70+ languages. ElevenAPI gives developers access to our leading AI audio foundational models. Everything we do is the result of the creativity and commitment of our team - builders doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to work hard and create lasting positive impact, we want to hear from you. How we work High-velocity: Rapid experimentation, lean autonomous teams, and minimal bureaucracy. Impact not job titles: We don’t have job titles. Instead, it’s about the impact you have. No task is above or beneath you. AI first: We use AI to move faster with higher-quality results. We do this across the whole company—from engineering to growth to operations. Excellence everywhere: Everything we do should match the quality of our AI models. Global team: We prioritize your talent, not your location. What we offer Innovative culture: You’ll be part of a generational opportunity to define the trajectory of AI, surrounded by a team pushing the boundaries of what’s possible. Growth paths: Joining ElevenLabs means joining a
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking an accomplished Principal Cloud Storage Engineer to lead the design, engineering, and evolution of our private cloud storage platforms. This role will focus on large-scale storage architecture, data protection, cyber recovery, and resiliency technologies across complex enterprise environments. The ideal candidate will combine deep technical expertise in storage systems with strong leadership, architectural vision, and the ability to influence technical direction across the organization. Key Responsibilities Architect and engineer enterprise storage platforms that ensure data integrity, availability, security, and disaster recovery readiness Design and implement end-to-end storage solutions, including Software Defined Storage, SAN, NAS, and object storage across private cloud and data center environments Drive strategic technology decisions by evaluating emerging products, tools, and standards supporting storage, data protection, cloud, and compute platforms Lead infrastructure initiatives involving storage modernization, data protection, cyber recovery, data migration, and resilience engineering Develop and execute enterprise strategies for backup, recovery, cyber vaulting, and business continuity Create and maintain comprehensive documentation of storage architectures, configurations, policies, and operation
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Senior Principal Network Engineer to help design, deploy, and optimize next‑generation AI data center networks. AI training and inference workloads require extremely high bandwidth, deterministic low latency, and zero‑packet‑loss networking environments. In this role, you will partner closely with the Network Architecture Lead to design and scale high‑performance computing (HPC) network fabrics supporting GPU clusters. You will work across hardware, networking, and AI application layers to ensure Graphcore’s large‑scale AI infrastructure operates at peak performance. The ideal candidate brings deep experience operating hyperscale or HPC data center networks and has expertise in high‑speed Ethernet fabrics, RDMA technologies, advanced automation, and telemetry systems. The Team The Data Center Network Engineering team designs and operates the high‑performance network fabrics that power Graphcore’s AI compute platforms. The team collaborates closely with hardware engineering, AI researchers, and infrastructure teams to build scalable networking environments optimized for distributed training and infe
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The AI Engineer, People Technology is responsible for designing, developing, and deploying AI-powered solutions that transform work across the People Organization. This role combines software engineering, AI development, and HR domain expertise to build intelligent applications, automations, and agents that improve employee experiences, increase operational efficiency, and accelerate workforce transformation. The ideal candidate is a hands-on builder with demonstrated experience using AI Assisted Software Development Platforms to develop enterprise AI solutions. You must have successfully designed and deployed AI agents that collaborate across multiple platforms, systems, and business functions while operating within enterprise governance, security, and compliance standards. This role requires deep knowledge of HR technologies, including Workday, ServiceNow, and the Microsoft Copilot ecosystem, along with a passion for applying AI to solve complex business challenges. Responsibilities: AI Product & Solution Development: Experience with the end-to-end product lifecycle turning a vision into a roadmap while driving adoption and value delivery. Design, develop, test, and deploy AI-powered products, applications, and intelligent workflow solutions. Apply AI Assisted Development Tools to accelerate software development, solution building, and deployment activities. Support authorities in translating business needs into developed solutions and prototypes. Develop reusable frameworks, ser
Other cities to consider
More places hiring for this role
Get new platform engineering lead jobs in United States by email
Daily job updates · Unsubscribe anytime