The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure. Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration. Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks. Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads. Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification. What we need to see: BS in Computer Science, Information Sys
Jobs in United States
Platform Deployment Management Lead in United States
3,662 active opportunities · Updated October 2026
Showing
14 jobs
Explore current platform deployment management lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
NVIDIA's invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning - the next era of computing - with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as "the AI computing company." We're looking to grow our company and establish teams with the most thoughtful people in the world. We are looking for an excellent Senior Engineering Manager to lead a large firmware engineering organization delivering end-to-end manageability firmware for NVIDIA's next generation Data Center Compute Systems. This role owns HGX product line and OpenBMC-based management firmware and MCU firmware components in data center platforms, including architecture, execution, quality, reliability, telemetry, and customer readiness. We are seeking an experienced senior leader with strong technical depth, broad system perspective, and a proven ability to lead large teams through complex product cycles. This role is onsite in Santa Clara, CA, USA. If you're creative and autonomous, we want to hear from you! What you'll be doing: Lead a large firmware engineering organization delivering OpenBMC based firmware and MCU firmware for next-generation Data Center Compute Systems. Own HGX platform as a lead for Firmware and System software readiness working across the organization. Define and drive the long-term firmware roadmap, balancing architectural innovation with product execution and delivery milestones. Drive architecture strategy across BMC, MCU, platform software, manageability, health management, and data center firmware interfaces. <spa
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Senior Software Engineers independently drive complex technical work, shape the systems and technical decisions within their teams, and enable other engineers to deliver high-quality, scalable solutions. Vanta's product monitors the security posture of thousands of companies, pulling tens of millions of API calls of data per day, pushing information from hundreds of thousands of laptop agents, and running tests against that data continuously to identify potential security threats. Our infrastructure and tooling need to stay ahead of exponential growth in our customer base. As a Senior Software Engineer at Vanta, you'll drive complex projects across our technical stack, contribute to the technical direction of your team, and mentor other engineers. Your past experience will be leveraged to enable and accelerate Vanta's growth. Visit our Vanta Engineering Blog to learn more about what our team is working on! Tests are at the heart of how Vanta continuously monitors security and compliance for our customers. The Test Core team builds the runtime platform that powers these checks. We own how Tests are scheduled and executed, how their results are persisted and exposed, and the systems that keep this runtime reliable as Vanta grows. In this role, you'll work on some of the core systems behind Vanta's Tests platform. You'll tackle problems around the reliability, correctness, and performance of Test execution, evolve the systems and abstractions that allow the platform to scale, and make it easier for other engineering teams to build on the Tests runtime. Many of these problems span multiple systems and teams and require a deep u
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. We're looking for a highly skilled People Systems Administrator to drive the design, configuration, and optimization of Workday across Vanta, with deep functional ownership of the Payroll, Benefits, and/or Absence module(s). Reporting to the Sr. Manager, People Systems, you'll act as a trusted consultant and system expert, partnering with functional leaders across People, Payroll, Finance, IT, and Legal to identify opportunities, implement advanced solutions, and improve the employee experience. You'll shape the future of our Workday ecosystem by leading complex configurations, driving process improvement, and making sure the system evolves ahead of the business rather than behind it. What you’ll do as a People Systems Administrator, Workday at Vanta: Own the design and configuration of the Workday platform Lead the design and implementation of configurations across Workday modules (Core HCM, Payroll, Benefits, Absence, Time Tracking) including business processes and security groups along with condition rules and calculated fields. Serve as the primary technical expert for Workday enhancements, partnering with cross-functional teams to gather requirements and translate ambiguous business problems into scalable system solutions. Build the administration frameworks and standards the team runs on: security role design, tenant and environment management, change control, testing protocols, and documentation. Design reusable, scalable solutions rather than one-off fixes, so the same problem doesn't return in a different shape Own change management for system changes. Plan and deliver stakeholder communications, training, and enableme
As an Enterprise Customer Success Manager, you will proactively drive new product attachment and effective strong relationships across our largest and most strategic customers. You’ll advocate for the customer internally and focus on a positive customer experience. Interactions are rooted in relationship-management, first and foremost, while also advocating for growth opportunities. Enterprise Customer Success Managers follow a well-defined methodology that helps them identify the customer's unique needs and clearly convey the value of the Datadog product. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Act as a strategic partner to customers, orchestrating cross-functional internal teams and engaging executive, technical, and business stakeholders to understand customer goals and translate them into a clear, deliverable Datadog value narrative. Proactively build and maintain executive relationships to deliver clear, outcome-driven value stories that connect Datadog technical use cases to measurable business results. Lead QBRs and strategic reviews as a forum to demonstrate impact, align on priorities, and define next-step initiatives. Analyze adoption and usage trends to quantify value delivered, extract insights from large datasets, identify gaps, and drive financially grounded commercial recommendations and strategic opportunities. Position Datadog as a critical observability platform that enables reliability, efficiency, and informed decision-making. Own and project manage the on-boarding process for new customers Collaborate cross-functionally with AEs, SEs, TAM, Product, Support, Enablemen and other technical teams to ensure consistent value delivery and messaging. Who You Are: Customer-centric with 3+ years in a Customer Success
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. We are looking for a VP of Product Management to lead AI product strategy at Smartsheet. You will lead a product team responsible for defining our comprehensive AI vision and roadmap, ensuring we build a cohesive, coordinated approach to how AI transforms our product. Your team will establish the strategic direction for specialized AI agents, evaluation infrastructure, and AI-powered capabilities, from platform foundations through to customer-facing features. You will report to our Chief Product and Technology Officer located in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. You Will: Lead a high-performing product team focused on defining and executing Smartsheet's comprehensive Applied AI product strategy. Define the strategic roadmap for AI product initiatives — from conversational experiences, MCP, automations, specialized agents, and evaluation infrastructure through to AI-powered features in customer-facing products — ensuring a cohesive, coordinated vision. Own the business outcomes of the AI portfolio — define the success metrics (adoption, retention, expansion, and revenue impact), set the targets, and hold the team accountable to them. Instrument, measure, and iterate: build the data and evaluation foundation that tells us whether AI features are actually working, and use it to make hard prioritization calls — including what to sunset. Move fast in a fast-moving space — ship, learn, and adjust in short cycles rather than waiting for per
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Job Summary Snowflake is seeking an organized and detail-oriented Talent Management Program Manager to focus on the day-to-day execution and operational support of our global talent programs. This role is essential to the smooth running of our Performance Excellence cycle, ensuring that every phase of the process is coordinated effectively and on schedule. In this role, you will be the primary "executor" for our established talent milestones. You will handle the logistics for performance review cycles, the administrative coordination of talent and succession reviews, and the delivery of our employee listening surveys. Additionally, you will serve as the primary system administrator and program owner for our voice AI 360 platform, managing its scaling and cross-functional operational workflows. This is a perfect role for someone who loves continuous improvement, process, excels at task management, and enjoys building for scale and ensuring that large-scale, global programs run without a hitch. Responsibilities Drive Performance Excellence: Lead the day-to-day execution of the Performance Excellence cycle — tracking progress against key milestones, sending timely communications and reminders to employees, managers, and global People Team partners at each phase, and ensuring e
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. We believe Plaid has the power to be the next-gen Credit Bureau - supporting large scale adoption of cash flow into the credit underwriting process. The Credit Decisioning platform team is responsible for building best-in-class cashflow based insights products that enable lenders to make more holistic lending decisions and empower broader access to Credit products for prospective borrowers. We own the systems and tooling that form the platform to build and serve these insights at huge scale, partnering with our Data partners to release new products yearly. You will be defining the future architecture of Credit insights products and executing against an ambitious product roadmap. You will partner with our Product, Data Science, and Machine Learning team to iterate on and productionize new insights that enable our customers to make more holistic lending decisions. Responsibilities: Leading technical architecture and execution across credit insights products: everything from data fetching and online feature serving for API requests, to offline production pipelines and tooling for model training. Scaling and evolving the architecture through an expected ~100x increase in load from deterministic factors
From $81K/yr
We are Datadog's in-house product experts. The Technical Solutions team enables Datadog's worldwide growth by educating potential clients and ensuring that existing customers are happy and successful. Premier Support Engineers (PSEs) are primarily focused on assisting prospects and customers with any technical questions about Datadog. PSEs engage with Datadog’s Premier Customers via standard technical support channels, but are also involved with cadence calls, demos/presentations, conferences, and various side projects. You will work directly with Datadog’s Premier Customer base, and will be immersed in a fast-paced environment where you will be challenged, but will also immediately witness your contributions to Datadog and to our customers. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Respond to client requests (phone / chat / tickets) on our fast paced team while continuing to educate our clients on the use of the platform Develop relationships with our Premier Customers, working hand-in-hand to truly know their distinct environment Reproduce issues and dive into the 600+ integrations that Datadog works with Build out documentation and knowledge based articles for a variety of technology Drive product conversations based on needs and problems learned during client interactions Participate in routine health check meetings with Premier Customers Work from a Datadog office 3 - 5 days per week Who You Are: Experienced in multi-channel technical support at a SaaS company (2+ years of related experience) A tinkerer with some programming experience and a basic knowledge of Linux Self-motivated, detail-attentive, and have a desire for continuous learning A critical thinker who defaults to a client-centric approach A decision
From $90K/yr
As an Enterprise Customer Success Manager, you will proactively drive new product attachment and effective strong relationships across our largest and most strategic customers. You’ll advocate for the customer internally and focus on a positive customer experience. Interactions are rooted in relationship-management, first and foremost, while also advocating for growth opportunities. Enterprise Customer Success Managers follow a well-defined methodology that helps them identify the customer's unique needs and clearly convey the value of the Datadog product. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Act as a strategic partner to customers, orchestrating cross-functional internal teams and engaging executive, technical, and business stakeholders to understand customer goals and translate them into a clear, deliverable Datadog value narrative. Proactively build and maintain executive relationships to deliver clear, outcome-driven value stories that connect Datadog technical use cases to measurable business results. Lead QBRs and strategic reviews as a forum to demonstrate impact, align on priorities, and define next-step initiatives. Analyze adoption and usage trends to quantify value delivered, extract insights from large datasets, identify gaps, and drive financially grounded commercial recommendations and strategic opportunities. Position Datadog as a critical observability platform that enables reliability, efficiency, and informed decision-making. Own and project manage the on-boarding process for new customers Collaborate cross-functionally with AEs, SEs, TAM, Product, Support, Enablemen and other technical teams to ensure consistent value delivery and messaging. Who You Are: Customer-centric with 3+ years in a Customer Success
From $204K/yr
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: We connect Airbnb’s community with the right information, in the right place, at the right time. We tailor Messaging & Notifications so hosts on Airbnb can streamline their operations, and travelers get just the information they need to enjoy their stay worry-free. Additionally, we are building new connections within our community to help enrich the experience of hosting & traveling on Airbnb: easing the process of hosting, and adding meaning to our guest’s trips. The data team utilizes industry-leading tools, builds scalable data systems and applies cutting-edge ML models to provide insights and empower all products in the Communication and Connectivity (CnC) organization. The Difference You Will Make: At CnC, data is foundational to our organization’s success.This role will lead key initiatives to design and build large-scale, distributed data systems - both batch and real-time processing. The data will power machine learning models and unlock new product features. You’ll be at the center of cross-functional collaboration, bridging backend, frontend/client, and machine learning engineering teams. CnC is applying GenAI and large language models (LLMs) to power products that enhance the Airbnb experience in various surfaces including highly used ones like Messaging. We're building a robust ML platform to power our product ambitions. A Typical Day: Shape the team’s long-term vision and roadmap in close collaboration with cross-functional partners across Airbnb Build strong relationships with partner engineering teams, including backend, client, data science, analytics,
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our NYC or SF office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role We're looking for an Engineering Manager to lead a team of highly experienced engineers building the infrastructure that powers Modal's serverless GPU platform. This is a hands-on leadership role — expect to split your time between technical contribution and people management depending on what the team needs. You'll set direction, remove blockers, and build a strong engineering culture as your team tackles hard problems in distributed computing, large-scale data handling, and performance optimization. Who You Are You're an experienced engineering leader who stays close to the work and builds alongside your team when it counts. You earn trust through technical depth, not title. You communicate clearly, help strong engineers move fast without cutting corners, and stay calm and pragmatic under pressure. You care as much about how your team gets to an answer as the answ
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for Forward Deployed Engineers on our engineering team who want to work at the intersection of deep infrastructure work and direct customer impact. As an FDE, you'll partner with leading AI companies and foundation labs on cloud architecture, networking, storage, containerization, sandboxing, and more — helping them design and ship production infrastructure on Modal's platform. The FDE team today includes world-class software engineers, computational scientists, ML engineers, and former founders. We're looking for people with strong engineering fundamentals, deep curiosity across the infrastructure stack, and energy for working directly with customers on hard problems. You will: Work hands-on with companies like Suno, Lovable, Cognition, and Meta to architect and deploy massive-scale production workloads on Modal Lead technical discovery and architect
Other cities to consider
More places hiring for this role
Get new platform deployment management lead jobs in United States by email
Daily job updates · Unsubscribe anytime