NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. Join NVIDIA's NIM team and be part of an exceptionally ambitious project in Santa Clara, CA! As a Senior Software Engineer, NIM Tools, you will have the remarkable opportunity to build a groundbreaking model customization and deployment lifecycle platform from inception. This isn't just another feature team—you will be defining the structure for a new product surface accessed by ISVs and CSPs internationally. Your work will empower customers to take models from selection through fine-tuning, evaluation, deployment, and compliance flawlessly. What you'll be doing: Compose and build the fine-tuning handoff pipeline, including LoRA adapter repackaging, re-quantization, and re-validation into NIM. Develop the evaluation harness, ensuring models meet our high standards. Implement the observability and attestation layer to produce auditable compliance artifacts. Work in close partnership with ISVs and CSPs to roll out NVIDIA NIMs on a large scale. Define and improve durable platform APIs, steering clear of one-off integrations. Ensure flawless completion of projects through strict attention to detail and proven methodologies. Wha
Jobs in United States
Deployment Lead in United States
636 active opportunities · Updated October 2026
Showing
15 jobs
Explore current deployment lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
NVIDIA is seeking a Senior Software Engineer to help us develop distributed storage services for AI/ML. In this role you will work closely with the broader NVIDIA team to design and build a reliable, scalable, and efficient storage-as-a-service tailored to AI applications that can be deployed anywhere and scale without limitations. This service supports the whole NVIDIA critical business from graphics drivers to autonomous vehicles to deep learning frameworks. To achieve this goal, we are looking for an engineer with a deep understanding of distributed systems, outstanding design skills, and a track record in building and delivering large-scale distributed services. What you will be doing: Leading the overall architecture and design of our distributed storage service optimized for AI/ML Develop and maintain distributed, robust and scalable Go programs deployed to state of the art open-source ecosystems, including Kubernetes. Develop and maintain user-space applications, containers, Go-bindings, and CLI tools. Building features for a distributed storage service to enhance availability and reliability for large-scale deployments Engaging and collaborating with NVIDIA Research, Computing, Product teams, cross-functional teams, and external customers to deliver Cloud services. Automating distributed storage service end-to-end, including deployment, management, and monitoring What we need to see: Bachelor’s of Science in Computer Science, or related field (or equivalent experience) with 8+ years of industry experience Strong background in developing distributed systems involving Golang, Kubernetes, and Cloud Service Provider integrations Strong track record of delivering distributed services in a variety of distributed computing environments Experience in i
$123.1K – $150.4K/yr
Software Product Security Engineer Description - This role supports the development and maintenance of secure software products under the guidance of senior engineers. The position focuses on learning software engineering and security best practices while contributing to the design, implementation, testing, and maintenance of desktop, web, and cloud-based applications and services. Key Responsibilities Assist in developing, testing, and maintaining software applications and security solutions. Participate in software development activities including coding, debugging, testing, and integration. Support the development and maintenance of Windows desktop applications and services. Assist in developing and maintaining web applications, APIs, and cloud-connected services. Troubleshoot software issues with guidance from senior team members. Write clean, maintainable, and well-documented code. Create and execute unit tests to verify software functionality and reliability. Participate in code reviews and learn software development best practices. Contribute to Agile ceremonies, sprint planning, and team activities. Learn and apply secure coding and software security principles. Support the deployment, monitoring, and maintenance of cloud-based applications and services. Collaborate with cross-functional teams to deliver end-to-end software solutions. Support product release and maintenance activities. Education & Experience Bachelor's or Master's Degree in Computer Science, Software Engineering, or a related discipline. 0-2 years of software development experience. <li
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Staff Software Engineer is responsible for the techno-functional impact analysis, design & code construction activities associated with development of software releases for the CVS Retail Pharmacy systems. Strong technical, functional, and interpersonal skills are key to perform this role successfully. The current suite of systems includes multiple applications including, but not limited, to the Tier 1 systems for Retail Pharmacy. The Staff Software Development Engineer will be involved throughout the entire development life cycle of a product and must be able to deliver an efficient solution, as well as identify and properly mitigate any risks to the product before deployment to production. ** This role is located in RI , Woonsocket** Required qualifications: 7+ years of software development experience in enterprise/web applications 5+ years of experience as full stack developer in Java-based technologies including microservices and spring boot framework /REST 5+ years of experience Cloud Technologies like Azure, GCP or other public cloud services and Knowledge of open-source packages especially those provided by Apache, Google, and Spring 2 + years of experienc
Become a part of our caring community Most AI engineering jobs are a thin wrapper around a model API. This role is different. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to the source document, and route ambiguous cases to human experts for review. Our users make decisions that impact real healthcare outcomes, so “good enough” is not good enough. Building AI systems that are accurate, reliable, auditable, and scalable is at the core of this role. As a Senior AI Applied Engineer, you will design, build, deploy, and operate production AI systems used at scale within one of the largest health insurers in the United States. You will own solutions end-to-end, from user experience and APIs to model orchestration, evaluation frameworks, infrastructure, and production operations. Why Join Us Build production AI systems where LLMs are in the critical path, not just demos or proofs of concept. Work on extraction, retrieval, agentic workflows, and human-review systems that process real healthcare data at scale. Own projects end-to-end across frontend, backend, AI orchestration, infrastructure, deployment, and operations. Solve challenging problems around accuracy, explainability, traceability, and reliability in regulated environments. Ship quickly in a small, high-impact team that embraces AI-assisted development and rigorous quality standards. Build systems that continuously improve through expert feedback, evaluations, and human-in-the-loop workflows. Key Responsibilities Design, develop, and deploy full-stack AI-powered application
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior AI Platform Engineer (DevOps) Who is Mastercard? Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payment choices, making transactions secure, simple, smart, and accessible. Our technology and innovation, partnerships, and networks combined to deliver a unique set of products and services that help people, businesses, and governments realize their greatest potential. Our decency quotient (DQ) drives our culture and everything we do inside and outside our company. We cultivate an environment where individuals can thrive, collaborate, and contribute to innovations that power the global economy. Overview: The AI Platform Engineering team is responsible for building, operating, and evolving Mastercard's enterprise AI platforms and capabilities. Our mission is to provide scalable, secure, and reliable AI infrastructure that enables teams across Mastercard to accelerate the development and deployment of AI-powered solutions. As a Senior AI Engineer, you will help design, implement, and operate the foundational platforms that support AI and machine learning workloads across the enterprise. You will work at the intersect
About the Team The Compute Strategy team works across research, engineering, product, finance, legal, and go-to-market teams to develop the partnerships, infrastructure capacity, and commercial models needed to advance AI infrastructure. About the Role As a member of the Compute Strategy team, you will develop commercial strategies for AI infrastructure partnerships and offerings. You’ll translate technical infrastructure opportunities into partnerships, transactions, and revenue. We’re looking for a commercially minded strategist who combines knowledge of semiconductors and AI infrastructure with strong financial judgment and the ability to execute complex partnerships. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Develop strategies for compute partnerships, vendor access, and infrastructure capacity. Structure and execute transactions with chipmakers, compute providers, and other infrastructure partners. Develop pricing frameworks and business cases for infrastructure-related partnerships. Evaluate partner technologies, strategic fit, commercial terms, and execution risks. Coordinate work across research, engineering, product, finance, legal, and go-to-market teams. Turn partnership learnings into repeatable operating models that can scale. You might thrive in this role if you: Have experience in strategy, corporate development, partnerships, or infrastructure transactions. Understand semiconductors and AI infrastructure. Can evaluate complex technical and commercial opportunities. Bring strong financial, analytical, and strategic judgment. Can influence and align technical and business stakeholders. Have negotiated or executed complex partnerships. Are comfortable operating in a fast-paced environment. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence ben
From $244K/yr
As a Forward Deployed Engineer on the Feature Flags team, you'll partner directly with customers to accelerate their feature flag implementations — from initial architecture consulting through prototype builds to full-scale migrations. This role is for someone who wants to write code with customers, not just advise them. You'll work hands-on inside customer codebases to unblock complex, high-stakes deployments, directly influencing deal velocity and customer success. Working closely with Sales, Solutions, and Engineering, you'll be the technical force that turns a signed contract into a live, adopted implementation. What You'll Do: Serve as the hands-on technical partner for strategic customers implementing Datadog Feature Flags, from pre-sales technical validation through post-sales delivery Consult on flag architecture and implementation approach for complex environments — multi-service, multi-platform, high-scale deployments Build prototype flag implementations directly in customer codebases to prove value and de-risk technical decisions early in the sales cycle Implement flags across diverse and advanced deployment modes (server-side, client-side, edge, mobile, streaming/real-time) tailored to each customer's stack Drive full flag migrations to completion — including legacy system cutover — efficiently and with minimal customer engineering burden Identify patterns across customer implementations and feed them back to Product and Engineering to improve the core product and reduce future implementation time Collaborate closely with Engineering on technical edge cases, product gaps, and implementation tooling Partner with Sales and Solutions to accelerate deal cycles by removing technical risk and uncertainty Who You Are: 5 years of professional software engineering experience, with hands-on coding ability across the stack you're deployed into Experience with feature flagging, experimentation, or config management systems (internal or vendor) Comfortable dropping i
About the Team The Future of Computing Research team is an applied research team within OpenAI’s Consumer Devices group. We study how AI systems perceive people and their surroundings, and we turn that research into capabilities for future products. Our work spans machine learning, sensing, and hardware, with a focus on building systems that work beyond controlled environments. About the Role We’re looking for a machine learning engineer to help shape how future AI systems understand the physical world and the people in it. The role focuses on multimodal perception and authentication, bringing together signals from cameras, microphones, and other sensors. You’ll work with specialized perception models and larger multimodal models, and partner with hardware, firmware, software, and product teams to bring new research into real-world systems. This role is based in San Francisco. We work in the office three days per week and offer relocation assistance. In this role, you will: Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals. Explore how specialized perception models and larger multimodal models can work together. Design data, training, and evaluation approaches that improve performance in real-world conditions. Study model behavior, robustness, and failure modes across sensing, data, and deployment environments. Integrate and validate new capabilities in real-time or resource-constrained systems. Work with hardware, firmware, software, and product teams to turn research into working systems. You might thrive in this role if you: Have a strong background in computer vision, audio or speech machine learning, multimodal learning, or sensing. Have experience developing specialized machine learning models, larger multimodal models, or both. Have brought research ideas into practical systems, prototypes, or products. Know how to design experiments, build evaluations, and investigate model behavior. Have wo
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Forward Deployed Engineers work directly with the largest and fastest-growing AI companies in the world, owning their technical outcomes on Baseten and taking on the hardest problems in serving and improving models at scale. The work spans the model lifecycle: inference, post-training, and the systems that tighten the loop between them. Act as each account's de facto CTO on Baseten, with final accountability for how their workloads are designed, run, and scaled. Take customer objectives from vague to shipped: frame the problem, define the spec and success criteria, build the PoC, and carry it through to production quickly, using the right tools for the problem. Design the evals and benchmarks that isolate where quality or performance falls short, then close the gap yourself, whether that means optimizing inference, improving the model through post-training, or reworking the eval itself. Be the first responder to mission-critical failures including triage, owning the fix directly or route to the owning team and stay accountable until it ships. Build internal systems so that each engagement is faster than the last. This includes tooling and automation for eval and deployment infrastructure, and the recipes and reference implementations that make the product more self-serve. Shape the product itself, channeling what your accounts need into the roadmap and shipping fixes and features into Baseten's codebase yourse
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Infrastructure Software Engineer at Baseten, you'll build and maintain components of our ML inference platform that powers production AI applications. You'll contribute to the core infrastructure, enabling developers to deploy, scale, and monitor ML models with high performance. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Develop infrastructure components for our ML inference platform using Python and Go Implement and maintain Kubernetes deployments for model serving Contribute to our inference orchestration layer for model deployments Build and enhance monitoring systems for model performance metrics Implement efficient resource management solutions for ML workloads Support infrastructure automation to improve ML deployment workflows Work closely with team members to implement technical solutions Help balance performance optimization with system reliability Participate in technical discussions around infrastructure improvements Learn and apply infrastructure best practices REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Go proficiency is a plus Working knowledge of Kubernetes and containeriza
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Cloud Platform Engineer, you'll envision and build robust systems and processes that ensure our infrastructure is scalable, reliable, and efficient. This can range from automating deployments and monitoring systems to optimizing performance and managing incidents. We all work closely with our users, learning from their past struggles in operationalizing ML, onboarding them onto our platform, and turning our learnings into ideas for improving Baseten. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Build and maintain scalable infrastructure to support the deployment and operation of machine learning models. Establish standards and best practices for reliability and performance across the infrastructure. Automate processes when relevant, particularly for managing CI/CD pipelines. Own products and projects end-to-end, functioning as both an engineer and a project manager, with a focus on user empathy, project specification, and end-to-end execution. Collaborate with cross-functional teams to understand project requirements and translate them into technical solutions. Mentor junior team members and contribute to knowledge sharing within the organization. Navigate ambiguity and exercise good judgment on tradeoffs and
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's engineers want to work in an AI-first way. What's missing isn't enthusiasm — it's the platform underneath it. Today everyone assembles their own agent config, context files, and MCP servers, so the good patterns stay trapped in individual setups instead of becoming defaults everyone inherits. You'll build that platform: the agent configurations tuned to our monorepo, the context and tooling layer that makes agents competent in our codebase, the evals that tell us which approaches actually work, and the rollout mechanics that get a new engineer productive with agents in week one. You are not here to mandate how engineers use AI — you're here to make the good path the easy path. Success looks like teams adopting what you build because it beats what they'd cobble together themselves, not because a policy requires it. Platform engineer, not AI evangelist. Ship infrastructure, measure it, kill what doesn't work, let adoption be the referee. The playbook for AI-first SDLC doesn't exist at any company yet. You'll write ours. WHAT YOU'LL BUILD Agent substrate — Repo-level context infrastructure that makes agents competent in our codebase ( CLAUDE.md/AGENTS.md conventions, architecture and domain context, and the tooling to keep it accurate as code moves). Internal MCP servers giving agents scoped access to CI, observability, incident tooling, deployment state, and docs. Shared skills, subagents, and hooks th
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
Other cities to consider
More places hiring for this role
Get new deployment lead jobs in United States by email
Daily job updates · Unsubscribe anytime