NVIDIA is seeking a Senior Software Engineer to help us develop distributed storage services for AI/ML. In this role you will work closely with the broader NVIDIA team to design and build a reliable, scalable, and efficient storage-as-a-service tailored to AI applications that can be deployed anywhere and scale without limitations. This service supports the whole NVIDIA critical business from graphics drivers to autonomous vehicles to deep learning frameworks. To achieve this goal, we are looking for an engineer with a deep understanding of distributed systems, outstanding design skills, and a track record in building and delivering large-scale distributed services. What you will be doing: Leading the overall architecture and design of our distributed storage service optimized for AI/ML Develop and maintain distributed, robust and scalable Go programs deployed to state of the art open-source ecosystems, including Kubernetes. Develop and maintain user-space applications, containers, Go-bindings, and CLI tools. Building features for a distributed storage service to enhance availability and reliability for large-scale deployments Engaging and collaborating with NVIDIA Research, Computing, Product teams, cross-functional teams, and external customers to deliver Cloud services. Automating distributed storage service end-to-end, including deployment, management, and monitoring What we need to see: Bachelor’s of Science in Computer Science, or related field (or equivalent experience) with 8+ years of industry experience Strong background in developing distributed systems involving Golang, Kubernetes, and Cloud Service Provider integrations Strong track record of delivering distributed services in a variety of distributed computing environments Experience in i
Jobs in United States
Cloud Platform Engineer Salary Guide in United States
736 active opportunities · Updated October 2026
Showing
15 jobs
Explore current cloud platform engineer salary guide jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
NVIDIA is looking for an experienced software engineer with infrastructure experience to become a senior member of the Cloud Foundations Automation - Development Team. We build and manage the automation ecosystem supporting NVIDIA's GPU Cloud and NVIDIA SuperPod deployments. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most hard-working and dedicated people on the planet working for us. If you're creative and autonomous, we want to hear from you! What you'll be doing: Developing software to enable efficient network design, deployment and day 2 management. Building product focused software solutions, used by internal and external customers. Helping us as we transform our workflows and organization into a centrally orchestrated configuration management framework, operating at scale across geographies. Owning and driving integrations with various service APIs such as Cloud Service Providers, to automate creation of environments and auto populate data sources in turn. Building on open source software, designing and implementing data structures and UI interfaces to automate processes from equipment purchase to device config generation to deployment to operations. Streamlining deployment mechanisms and life cycle operations Developing modern service architectures around streaming data and event pipelines. Working with infrastructure domain experts on true, zero touch deployment solutions and utilizing best of breed high performance computing management solutions. Be a proactive problem solver, looking out for new opportunities to improve our services and customer experience. Communicate readily with your peers across the organization, b
About the Team OpenAI Finance ensures the organization is positioned for long-term success as we pursue our mission. The Order to Cash (OTC) team oversees the complete flow of commercial transactions from order intake and provisioning through billing, collections, and cash application — ensuring accuracy, compliance, and operational excellence in support of OpenAI’s mission to ensure artificial general intelligence benefits all of humanity. About the Role We are seeking a Senior Manager to build and scale Order to Cash across OpenAI’s API and ChatGPT businesses, cloud marketplaces, and strategic partnership channels. This role will own end-to-end Order Management and Billing operations across API and ChatGPT while leading OTC readiness and execution for AWS Marketplace, Google Cloud Marketplace, Oracle Cloud Marketplace, GovCloud, and future partner channels. The role will also own the end-to-end Order Management and Billing close, setting the close calendar, readiness standards, review and sign-off expectations, while leading the team responsible for execution. You will oversee the Order Management and Billing lifecycle across API, ChatGPT, marketplace, and partner transactions, spanning commercial readiness, order intake, provisioning, usage and transaction data, pricing validation, invoicing, credits, settlements, and product and partner reporting. Your work will ensure transactions are accurate, timely, complete, and supported by audit-ready controls. This is a leadership role that combines strategic ownership, cross-functional leadership, and hands-on operational execution. You will define the target operating model, lead first-of-kind launches, shape product and systems roadmaps, and oversee the resolution of complex contract modifications, non-standard pricing, usage disputes, reconciliation breaks, settlement variances, and customer- or partner-impacting escalations. You will collaborate with teams across Finance, GTM, Product, Engineering, Legal, Tax, Reven
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Technical Consultants are on the frontlines working to establish partnerships with Ramp customers, and act as a liaison between customers and our internal technical teams. You'll be partnering with Customer Activation & Account Management teams to identify, prioritize and build financial solutions for new and existing customers so that Ramp continues to scale with their needs. You will act as the technical and financial subject matter expert for our customers and design solutions that address the customer's needs through Ramp and partner integrations. You will partner with Post-Sales teams to drive Product, Design and Engineering teams' roadmaps to evolve Ramp to better serve our customers at scale and offer increasingly higher value. You will define the customer's financial & technical success criteria from onboarding through activation and expansion. You will represent Product, Design, and Engineering externally and will operate with deep conviction and understanding of Ramp's technical capabilities. The role requires techni
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Technical Consultants are on the frontlines working to establish partnerships with Ramp customers, and act as a liaison between customers and our internal technical teams. You'll be partnering with Customer Activation & Account Management teams to identify, prioritize and build financial solutions for new and existing customers so that Ramp continues to scale with their needs. You will act as the technical and financial subject matter expert for our customers and design solutions that address the customer's needs through Ramp and partner integrations. You will partner with Post-Sales teams to drive Product, Design and Engineering teams' roadmaps to evolve Ramp to better serve our customers at scale and offer increasingly higher value. You will define the customer's financial & technical success criteria from onboarding through activation and expansion. You will represent Product, Design, and Engineering externally and will operate with deep conviction and understanding of Ramp's technical capabilities. The role requires techni
From $195K/yr
The Team Datadog is an enterprise SaaS company featured in Forbes Cloud 100, Deloitte's Fast 500 Technology list, and Forrester's Intelligence Application and Service Monitoring Wave as a leader. Our team works with a best-of-breed product that solves real problems for our customers and partners. We are Datadog's in-house product experts and strategic architects of the partner ecosystem. The Partner Technical Solutions team drives Datadog's worldwide growth by shaping how our partners go to market, build services, and create lasting technical differentiation for their customers. We work at the intersection of product strategy, technical depth, and business outcomes - defining the frameworks, programs, and practices that determine how the partner ecosystem scales. Partner Technical Solutions is a growing global team that operates with high autonomy, collaborates constantly across functions, and sets the standard for what technical partnership looks like at Datadog. The Opportunity This role is built for someone who has already proven they can operate at the intersection of deep technical authority and executive influence - as comfortable presenting observability architecture to a federal agency CISO as they are rolling up their sleeves to get hands-on with a partner's most complex deployment challenges. Aligned to our Channels & Alliances business segment, the Principal PSA will report to the Head of North America Partner Solutions Architecture. They are accountable for the technical success of our most strategically important federal and public sector partners - a trusted advisor, a force multiplier, and the definitive Datadog expert in every room they enter. Specifically, you will Own the senior technical relationship with a portfolio of high-value federal and public sector partners, serving as a strategic advisor to partner CTO, CISO, and VP-level stakeholders; Lead joint technical strategy sessions and executive business reviews, ensuring alignm
About the Team The Codex team is responsible for building state-of-the-art AI systems that can write code, reason about software, and act as intelligent agents for developers and non-developers alike. Our mission is to push the frontier of code generation and agentic reasoning, and deploy these capabilities in real-world products such as ChatGPT and the API, as well as in next-generation tools specifically designed for agentic coding. We operate across research, engineering, product, and infrastructure—owning the full lifecycle of experimentation, deployment, and iteration on novel coding capabilities. About the Role As a Performance & Systems Engineer on the Codex team, you will be responsible for whole-system optimization across a complex, evolving stack. Codex spans LLM inference, cloud orchestration, agentic work management, and multiple product surfaces. Your job will be to identify and land high-leverage changes—across infrastructure, modeling, and product layers—that make Codex agents significantly faster and cheaper to serve. We’re looking for generalists who thrive in ambiguity and love chasing performance bottlenecks to ground. This is a high-ownership role where your work will directly improve the experience of millions of users. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Hunt down and address inefficiencies across the Codex system stack, from agent behavior to LLM inference to container orchestration, and beyond. Build tooling to measure, profile, and optimize system performance at scale. Collaborate with researchers and engineers to land high-ROI changes that improve latency and cost. You might thrive in this role if you: Have experience operating across both ML systems and cloud infrastructure. Enjoy diving into messy, ambiguous problems and emerging with clear wins. Think holistically about performance, balancing spee
NVIDIA is hiring an NCX Senior Engineer who is passionate about NVIDIA Cloud Partner (NCP) infrastructure operations to join our DSX team. This role involves working closely with strategic NVIDIA Cloud Partners to build and improve the operational capabilities essential for running large-scale NVIDIA accelerated infrastructure reliably in production. Your role involves guiding partners beyond the initial cluster deployment and validation phase into advanced Day 2 operations. These operations cover ongoing infrastructure health, observability, lifecycle management, quick remediation, performance validation, and operational readiness. You will engage directly with partner engineering and operations teams to develop consistent approaches that support NVIDIA workloads and the broader external customer environments of the partners. This is a highly technical, hands-on role at the intersection of NVIDIA accelerated computing, cloud infrastructure, distributed systems, and production operations. What you'll be doing: Lead NCP Day 2 operational readiness efforts. Collaborate directly with NVIDIA Cloud Partners to set up the systems, procedures, automation, and operational methods necessary to consistently manage NVIDIA accelerated infrastructure following initial deployment and activation. Build continuous infrastructure validation. Develop and implement methods to continuously validate GPU, CPU, storage, and network health. Do this across large-scale AI clusters to identify degraded infrastructure before it impacts critical training or inference workloads. Establish observability and operational telemetry. Help NCPs implement comprehensive telemetry, monitoring, alerting, dashboards, and operational signals across compute, GPU, InfiniBand/RoCE networking, storage, Kubernetes, and AI workloads. Devel
From $244K/yr
Role Summary: Datadog is seeking a Staff Software Engineer to help shape the future of our Bring Your Own Cloud (BYOC) Logs offering by unifying observability pipelines with log management software that customers deploy and manage in their own infrastructure. This role will focus on building and scaling systems that process, route, and store high-volume observability data within customer-managed infrastructure. You will operate as a hands-on technical leader, driving architecture, cross-team delivery, and product direction across a complex and evolving space. This is a high-impact opportunity to influence product strategy, mentor engineers, and solve deeply technical challenges at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Make customer-controlled deployments feel like a managed Datadog product: deployment, upgrades, configuration, observability, diagnostics, reliability, and secure operation across diverse customer cloud environments Build and scale high-throughput systems for log processing, routing, and transformation across distributed environments Lead cross-team initiatives, aligning engineers, product managers, and stakeholders to deliver complex, multi-team projects Design and implement software that runs reliably that customers deploy and operate within their own cloud infrastructure. Improve system performance, scalability, and cost efficiency through thoughtful trade-off analysis and capacity planning Contribute hands-on to critical code paths, debugging, and deployment challenges in customer environments Who You Are: You have significant experience building software that is installed, deployed, and operated in customer environments rather than only as a fully managed SaaS service. You have strong expertise in distributed systems,
Job Requisition ID # 26WD101160 Position Overview As an Implementation Consultant focused on our Forma Industry Cloud for Construction General Contractors and Trade Contractors, you will bring industry expertise, strong product knowledge, and engaging communication and presentation skills, to help customers put innovative solutions into practice. In this role, you will work with customers to evaluate current workflows, identify opportunities for transformation, and guide the adoption of digital solutions that support better business and project outcomes. You will partner with customers to define new workflows, configure technology, deliver training, and support lasting adoption. This role is designed for someone who can credibly speak to the full building lifecycle and connect Autodesk Forma Industry Cloud solutions to customer priorities across the plan, design, build, and operate journey. This role will primarily support North American accounts, with a particular focus on strategic General Contractors and Trade Contractors. You will help customers scale impactful solutions across their organizations, while also partnering closely with Autodesk product and sales teams to share feedback from the field and stay aligned with product roadmap direction. You will report to a manager on the Adoption Services team. Responsibilities Lead end-to-end Forma product implementation, training, and consulting engagements focused on Forma Build and Preconstruction Partner with customers to assess workflows across planning, design, construction, and operations, and recommend best practices using Forma Industr
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. About Modal Data: We’re growing our Data team and are looking for our first few key hires to build self-serve data tools and drive business strategy in the right direction. The mission of the Modal Data team is to make it easy to track company goals, make evidence-backed decisions, and prioritize the right work. We do this via: Self-serve AI analytics tools (Hex, Snowflake) Embedding with teams as a “data adviser”, providing strategic analysis and consulting What You'll Do: Contribute to building the most modern analytics stack in Data today to support AI-driven self-serve analysis, key metrics tracking, and external customer reporting Influence work on new products like LLM Inference Endpoints through product analytics tracking Identify millions of dollars of cost savings and optimization across our tools and financial operations Write data pipelines that power the operatio
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal considers high-quality documentation to be essential for developer experience, and we see docs becoming even more important as agents increasingly deploy and operate Modal Apps. We are looking for a content-minded engineer who will partner with our product teams to curate Modal’s technical documentation and maintain a high quality bar across multiple dimensions. Responsibilities: Thinking holistically about content architecture and how the docs should evolve as Modal introduces new products and features Innovating on novel documentation formats and delivery channels to optimize agent productivity, in collaboration with our Agent DX research team Developing content standards, style guides, and automated enforcement mechanisms to ensure consistent style and high quality Building and maintaining automated pipelines that will enforce the correctness of code examp
Are you a person who likes to work in a fast-paced organization? NVIDIA is the world leader in Visual Computing. We are passionate about four markets: Gaming, Automotive, Enterprise Graphics and HPC/Cloud Datacenters; in addition to our traditional OEM business. We are well positioned as the ‘AI Computing Company’, and our GPUs are the brains powering modern Deep Learning software frameworks, accelerated analytics, big data, modern data centers, smart cities, and driving autonomous vehicles. We have some of the most forward-thinking and talented people on the planet working for us. If you're forward-thinking, hardworking, driven and if working with extraordinary people across countries sounds interesting, this job is for you. We are now looking for a Human Resources Business Partner to provide HR support onsite in Santa Clara, CA for our Worldwide Field Organization in a dynamic and collaborative environment. This is a global organization, and we are looking for someone to be passionate about supporting and building strategies to enable NVIDIA to achieve success. You’ll partner with a cross-functional group of subject matter experts to design and execute strategies for how we staff, onboard, develop, motivate, retain and organize work. You will need excellent communication skills, critical thinking and planning ability, and the agility to function in a fast paced and innovative environment. What you'll be doing: This position will be an integral enabler of the mission of our Field organization. In this position you will work with the senior leaders and leadership teams within NVIDIA organizations to develop and execute the HR strategies that champion organizational and people effectiveness. You will think strategically as well as roll up your sleeves and dive deep into practical application. You must understand business priorities and translate them into an HR ag
About the Company Sigmoid enables business transformation using data and analytics, leveraging real-time insights to make accurate and fast business decisions, by building modern data architectures using cloud and open source. Some of the world’s largest data producers engage with Sigmoid to solve complex business problems. Sigmoid brings deep expertise in data engineering, predictive analytics, artificial intelligence, and DataOps. Sigmoid has been recognized as one of the fastest growing technology companies in North America, 2021, by Financial Times, Inc. 5000, and Deloitte Technology Fast 500. Job Description As an Account leader you will be responsible for ensuring customer success and growth in Fortune 1000 companies working with Sigmoid. A maverick self-starter, you understand brand building, how to sell innovation, drive deals forward and compress decision cycles. You will play a key role in driving our business to great heights, and drive our revenue growth in parallel. We're looking for a passionate farming growth hacker with a track record of proven success in Data Solutions selling in Fortune 1000. Prior experience in Analytics, Data Science & Big Data will be an added advantage. Job Responsibilities As an Account leader, you will have the opportunity to work on major business initiatives that contribute to Sigmoid’s growth and productivity objectives. In this role, you will have the responsibility of managing multiple account management strategy implementation assignments supporting the Account Management function and will work directly with the business, IT and strategy teams in catering to the end-to-end business needs. Essential Responsibilities Manage account management strategy implementation and validation. Effective communication and presentation ability. Ability to work as in a team as well as contributing as an individual. Lead and provide a road map for account. Able to establish priorities and coordinate work. Evaluate, scrutinize and str
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders About the Role We're hiring the first Account Managers at Modal. You'll report to the Regional Director of Account Management and be a founding member of the team. This function does not exist yet. There is no playbook, no territory map, no established motion. You'll own a book of business from day one and build the motion at the same time — from fast-moving AI startups to large enterprise teams running critical infrastructure on Modal. This is a commercial role with a revenue target. You'll be measured on retention and expansion across your accounts. While you won't be delivering the technical recommendations and implementation, the work is technical by nature. Our customers are engineers running GPU workloads, inference, and batch jobs in production, and you need to hold your own in those conversations. The profile we're hiring is a technical account manager. You've worked at companies that are deepl
Other cities to consider
More places hiring for this role
Get new cloud platform engineer salary guide jobs in United States by email
Daily job updates · Unsubscribe anytime