Jobs in United States

Senior Ai Compute Engineer in United States

1,941 active opportunities · Updated October 2026

Explore current senior ai compute engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $220K/yr

Quick readStrong listing-quality and freshness signals

The Applied AI team designs and builds algorithmically driven features in the Datadog app. We work across a range of applications, primarily focusing on analysis on streaming data such as anomaly detection, error outliers and faulty deployment analysis. As an Applied Scientist you will work on building models and algorithms for machine learning powered features within the Datadog platform. You will work closely with our engineering and product partners to explore, build, scale and deliver these features that we incubate within the Applied AI team. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Design solutions for our different use cases. You will research and benchmark relevant algorithms to find the best fit for our use-cases Leverage machine learning algorithms and statistical techniques to build new scalable product features Develop, deploy and monitor new and existing features to production Participate in our journal club by reading and presenting the latest academic research papers to the team Explore, analyze and tell the story behind high volumes of data flowing through Datadog systems Maintain and monitor the models, services and infrastructure owned by your team Participate in your team’s on-call rotation Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering, Machine Learning or related scientific field or equivalent experience You have experience working with high-scale systems and datasets including building models, applying machine learning to real business problems, and writing production data pipelines You can explain complex ideas and algorithms to non-technical audiences You care about code simplicity and performa

Machine LearningAIGoRust
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

As one of the technology industry's most desirable employers, NVIDIA has been redefining accelerated computing, computer graphics and leading the Artificial Intelligence revolution. NVIDIA's innovation is fueled by its great technology—and amazing people. We are seeking a Senior Silicon and System Product Lead to influence, innovate and take our next generation products to the market. As part of the Silicon Solutions Team, we architect and deliver groundbreaking system solutions that integrate all aspects of the system from silicon design, software design to operations and final deployment in multiple market segments that NVIDIA serves. This position offers an unique opportunity to collaborate with multiple organizations in the company and grow your career in a high impact role. We need a passionate, hard-working and creative individual to lead the products all the way from market analysis to delivering the features on the final product. What you'll be doing: Drive product performance and power targets, trade-off features/configurations and provide innovative solutions to complex silicon and system level problems. Evaluate new market segments and use cases; translate market requirements to engineering problem statements and metrics. Innovate Performance, power, yield and quality optimizations and features for the world’s fastest power-shipping products in the GPU and SoC market segments spanning gaming, automotive, datacenter and DL/AI. Develop methodologies and requirements for multi-functional teams to drive silicon and system product features to production. Incorporate productization feedback to improve the next generation. Lead the team for feature requirements and schedule from architecture to silicon phase of projects. Work alongside system architects, designers, marketing teams, chip and board designers, software/firmware engineers, HW/S

Artificial IntelligenceAI
A
📍 San Francisco, CA, United States
✓ Quality checkedCompany trend -100%

The Opportunity The world of design is changing rapidly, and the Pro Design team is leading that transformation. We are the Adobe organization behind Illustrator, InDesign, and emerging experiences that connect creativity, collaboration, and AI. Our teams are reimagining what professional design looks like for the next decade - building intelligent, connected tools that empower creators and teams to move faster without sacrificing craft. We are looking for a Senior Business Data Scientist who is creative, analytical, and unafraid to question the status quo and shape the decisions that move key business metrics at scale. Join us and build Adobe’s future products! What you'll Do Map the user funnel and build the metrics, cohorts, and dashboards that Product and Growth rely on to see how users move across free, trial, and paid tiers—and pinpoint where they drop off. Dig into the hard questions (what drives activation, which behaviors predict retention and expansion) and build propensity models for conversion, upgrade, churn, and expansion that feed real-time targeting and in-product nudges. Find and size growth bets and work with Product to ship them. Set north-star, driver, and guardrail metrics with your partners, and stand up multivariate experiments across onboarding, paywalls, in-product prompts, and pricing. What you need to succeed Minimum Requirements: Bachelor's degree in a quantitative field (Statistics, Mathematics, Computer Science, Economics, Engineering, or similar) or equivalent practical experience. 5+ years of experience in data science, product analytics, or a similar quantitative role. Proficiency in SQL and Python (or R) for data manipulation, analysis, and modeling. Hands-on experience designing and analyzing A/B tests and interpre

PythonSQLAI
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

NVIDIA is seeking a Senior System Architect: Heterogeneous EDA Systems to solve a complex challenge in accelerated computing: Failure Attribution at Scale. As EDA or equivalent experience workloads scale across thousands of heterogeneous nodes, a single failure can cause massive resource waste. We need an engineer to develop and build an automated framework. This framework will ingest telemetry from CPU and GPU clusters to identify the root cause of job failures in real-time. It will distinguish between hardware faults, infrastructure instability, and software defects. What you'll be doing: Architect Failure Attribution Frameworks: Build a scalable "flight recorder" for EDA jobs that captures high-fidelity state across the CPU, GPU, and Fabric at the moment of failure. Build automated diagnostics that correlate GPU XID errors, PCIe bus failures, and CUDA memory exceptions. Connect these errors with system-level events such as OOM kills or NUMA-related hangs. Distributed Logging & Tracing: Implement low-overhead tracing mechanisms (using tracing tools or custom agents) that provide access to job execution across multi-node Slurm or Kubernetes clusters. Root Cause Automation: Develop heuristics and models based on machine learning to classify failures as "Hardware Fault," "Software Bug," or "Environment Issue." This reduces the Mean Time to Identify (MTTI) for R&D teams. Resiliency Engineering: Work closely with hardware and infrastructure teams to define "signals of impending failure," enabling proactive job migration or check-pointing before a crash occurs. What we need to see: Distributed Systems Mastery: BS, MS, or PhD in Computer Science or Electrical Engineering (or equivalent experience) with 6+ years in systems programming. Experience building automated

PythonKubernetesLinuxMachine Learning
S
📍 Bellevue, Washington, United States· Full-time
✓ High-confidence listingCompany trend -92.9%
Quick readStrong listing-quality and freshness signals

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Join our ML Feature Store team where we're building cutting-edge product capabilities that power complex feature transformations and low latency feature serving. We're revolutionizing machine learning feature management and serving capabilities as part of the Snowflake ML suite of products. In the era of GenAI and agents, our team delivers high-quality, fresh feature solutions that make a real difference for our customers. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Help define and own the roadmap for Snowflake Feature Store, working collaboratively with senior architects and ML team leadership Build and execute a vision for incorporating new advances in machine learning Ensure operational excellence of services and meet reliability, availability, and performance commitments Collaborate across ML partner teams to improve development velocity and capabilities Support team members in delivering high technical quality WE WOULD LOVE TO HEAR FROM YOU IF YOU HAVE: 10+ years of experience in designing and building data serving infrastructure and/or machine learning platforms. Strong track record working with machine learning systems and platforms. Strong understanding of computer science fundamentals. B.Sc . in Computer Science Fluency in Ja

PythonJavaMachine LearningAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA is seeking a strong technology leader to manage our Server Software Technical Program Management (TPM) team. This role is at the cross-section of execution and strategy, leading a team of Senior TPMs who drive the firmware and system software for NVIDIA's next-generation server platforms like DGX, MGX, and HGX. These platforms bring together the full power of NVIDIA GPUs, NVLink, InfiniBand networking, Grace CPUs, and our optimized AI/HPC software stack. This deep technical leadership role focused on the Software Development Processes that brings new server hardware to life. What you'll be doing: Lead a team of TPMs driving the technical software and firmware execution for NVIDIA's NPI (New Product Introduction) and sustaining engineering teams. Drive the end-to-end SDLC for low-level server components, including firmware (BMC, UEFI/BIOS), drivers, and system management software, ensuring alignment with hardware schedules. Collaborate closely with NVIDIA product management and hardware engineering teams to define release plans and program objectives. Build a strong connection and feedback loop between sustaining and NPI engineering teams to improve product quality and development velocity. Lead process improvement initiatives and help propagate SDLC standards across multiple engineering and TPM organizations. You will have the opportunity to interact with diverse technical groups, spanning all organizational levels. What we need to see: Bachelor of Science (or equivalent experience) or Master of Science degree in Computer Science, Electrical Engineering, or related field. 12+ overall years of experience developing and leading complex low-level or system software projects. and 7+ years of experience in a people management role. Deep understanding of system a

Artificial IntelligenceAI
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

Own the end-to-end execution that takes NVIDIA DRIVE Autonomous Driving Software (NDAS) from product definition to production deployment across major OEM vehicle lines. We are seeking an accomplished engineering execution leader to lead adaptation, integration, validation, and production readiness of NVIDIA Autonomous driving solution (NDAS) for strategic automotive OEMs. You will work between NVIDIA product and engineering teams and the OEM's vehicle, software, systems, and validation groups. The role involves converting an agreed feature set into a deliverable vehicle program. This entails clarifying requirements, coordinating interfaces and responsibilities, setting the engineering plan, building consistent productization workflows, and managing issue resolution. You will influence teams such as systems engineering, autonomous-driving development, platform software, functional safety, cybersecurity, quality, validation, release, and customer engineering. What You'll Be Doing: Lead NDAS execution with major OEMs across the full program lifecycle — from feature definition and technical alignment through vehicle integration, validation, launch, and production ramp — owning the coordinated engineering plan per vehicle line, including requirements, architecture, adaptation scope, achievements, staffing, validation strategy, release criteria, and production-readiness gates. Translate OEM vehicle requirements and use cases into clear commitments for NVIDIA teams, while ensuring OEM partners understand product capabilities, constraints, assumptions, and the work required on their side; build scalable, repeatable workflows for adapting NDAS to specific OEM platforms, vehicle architectures, sensor configurations, compute platforms, networks, and development processes. Drive multi-functional implementation across autonomous-driving features, systems, platform software, vehicle integratio

L
📍 San Diego, California, Canada
✓ High-confidence listingCompany trend +500%
Quick readStrong listing-quality and freshness signals

Ignite your curiosity. Solve the unsolvable. At Leidos, we do more than write code—we decode the unknown. Our San Diego-based research and engineering team takes on some of the nation’s toughest defense challenges using advanced signal processing, ocean remote sensing, and high-performance computing. We’re seeking a Software Engineer / Computer Scientist who enjoys solving complex problems and pushing the limits of performance. In this role, you’ll work alongside a multidisciplinary team of scientists and engineers with expertise in hydrodynamics, physics, acoustics, and signal processing to build impactful software that turns massive, complex data sets into meaningful insight. If you are motivated by innovation, energized by collaboration, and excited to see your work support real-world missions, this could be the right opportunity for you. What You’ll Do Collaborate with scientists and engineers to design, develop, and optimize advanced algorithms for next-generation radar, optical, and infrared sensor systems. Build scalable, high-performance backend systems for scientific computing in distributed environments. Integrate, refactor, and improve scientific codebases to increase efficiency and scalability. Translate and optimize existing code for GPU/CUDA acceleration and parallel or distributed execution. Test, document, maintain, and enhance complex software in Linux/Unix environments. Contribute in a collaborative environment that values technical excellence, creativity, and continuous growth. Required Qualifications Bachelor’s degree in Computer Science, Applied Mathematics, Physics, or a related field with 4+ years of backend software development experience, or a Master’s degree with 2+ years of experience. Equivalent experience may be considered in place of a degree. U.S. citizenship and the ability to obtain a Top Secret clearance; active Top Secret clea

GitLinuxAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $280.5K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the Role: AI models reshaping how our community creates, plays, and connects, all run on Compute Platform. As Senior Product Manager, Compute Platform , you'll set the strategy and roadmap for Roblox's next-generation AI infrastructure: the rapidly growing fleet of GPUs and AI accelerators spanning Roblox core and edge data centers, and public cloud that decides how fast we can train, serve, and scale every model on the platform. You'll own the products that turn raw GPU hosts into reliable, production-ready AI compute - driver and firmware management, fleet-wide health and performance, and the abstractions product teams across Roblox build on. You Will: Drive strategy and roadmap for Compute Platform spanning Managed Kubernetes (Roblox Kubernetes Service), Managed Compute Services and other critical distributed systems, and our fleet of GPU and CPU machines managed via unified Fleet APIs - all across on-prem and cloud. Drive the evolution of our Compute infrastructure to support Roblox’s most critical workloads - from AI to Storage to Data Analytics and more - each with their own distinct requirements. Build and scale our GPU infrastructure to support training and inferen

AWSAzureGCPKubernetes
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%

$216K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Team OpenAI Finance is responsible for ensuring the organization is set up for success in pursuit of its mission. The Technical Accounting team plays a crucial role in helping OpenAI navigate complex, judgmental, and rapidly evolving accounting matters with rigor and clarity. We aim to bring both technical excellence and strong business partnership to some of the most novel accounting questions in the industry. About the Role As Senior Manager, Technical Accounting, Compute Infrastructure, you will lead the evaluation, documentation, and operationalization of complex accounting matters related to OpenAI's compute infrastructure, strategic investments, and other non-routine business activities.. This role sits at the intersection of U.S. GAAP technical accounting, infrastructure strategy, financial reporting, controls, and cross-functional execution. Key areas may include cloud compute arrangements, data center and colocation arrangements, lease accounting under ASC 842, power purchase agreements, strategic investments, consolidation evaluations under ASC 810, financial instruments, and other emerging or non-standard arrangements. This role is based in San Francisco, CA or remote. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead technical accounting analysis for complex, judgmental, and non-routine transactions under U.S. GAAP. Evaluate accounting implications for compute infrastructure arrangements, including cloud compute, data center, colocation, lease, PPA, infrastructure procurement, and related commercial arrangements. Partner with Controllership, Tax, Legal, FP&A, Procurement, Infrastructure, and other cross-functional teams to assess the accounting implications of new products, commercial arrangements, strategic transactions, and business initiatives. Prepare and review technical accounting memoranda, position papers, and other auditor-ready documentation.

AWSRestAIGo
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA is leading groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU -- our invention -- serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables groundbreaking creativity and discovery, and powers inventions that were once considered science fiction, including artificial intelligence to autonomous cars. We are the GPU Communications Libraries and Networking team at NVIDIA. We build communication libraries like NCCL, NVSHMEM, and UCX that are crucial for scaling Deep Learning and HPC. We're seeking a Senior Software Architect to help co-design next-gen data center platforms and scalable communications software. DL and HPC applications have a huge compute demands and already run at scales of up to tens of thousands of GPUs. GPUs are connected with high-speed interconnects (e.g. NVLink, PCIe) within a node and with high-speed networking (e.g. InfiniBand, Ethernet) across nodes. Efficient and fast communication between GPUs directly impacts end-to-end application performance. This impact continues to grow with the increasing scale of next generation systems. This is an outstanding opportunity to advance the state-of-the-art, break performance barriers, and deliver platforms the world has never seen before. Are you ready to build the new and innovative technologies that will help realize NVIDIA's vision? What you will be doing: Investigate opportunities to improve communication performance by identifying bottlenecks in today's systems. Design and implement new communication technologies to accelerate AI and HPC workloads. Explore innovative solutions in HW and SW for our next generation platforms as part of co-design efforts involving GPU, Networking, and SW architects. Build proofs-of-concept, conduct experiments,

LinuxArtificial IntelligenceAI
R
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At Render, we’re building the modern cloud platform for developers creating AI-native, full-stack, multi-service applications. Our mission is to eliminate the tradeoff between the power of hyperscalers and the simplicity of developer-friendly platforms—so teams can ship fast, scale reliably, and focus on their product, not infrastructure. Unlike complex hyperscalers or ephemeral edge/serverless solutions, Render offers a developer-first experience with persistent compute, dynamic autoscaling, built-in orchestration, and observability, allowing teams to launch, scale, and manage real-world applications without writing infrastructure code or managing servers. Whether you're building LLM-powered applications, scalable SaaS products, or async processing pipelines, Render empowers teams to move fast and scale confidently from MVP to millions of users. Our platform is trusted by over 6.5 million developers worldwide and continues to grow rapidly. In February 2026, we raised an additional $100M in Series C financing, bringing our total funding to $257M, to accelerate our vision of making cloud infrastructure both powerful and intuitive—designed for the speed of modern AI development. We’re a diverse and talented team that values craft, velocity, and user experience. If you’re excited to help shape the future of the intelligent cloud and empower developers everywhere, we’d love to hear from you. Applying to Render We're seeking candidates who possess high integrity, humility, and an insatiable drive to learn. Through reasoned discussions and continuous feedback, we strive to improve both individually and collectively. We foster an environment of mutual trust and respect, empowering effective debate to achieve the best outcomes for our customers and team. We especially encourage members of underrepresented groups in the tech community to apply and understand that not all successful candidates will meet each requirement listed. Our interview process is unique to each role, an

CI/CDRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Industrial Compute is building the infrastructure ecosystem that enables OpenAI to train and deploy increasingly capable AI systems at unprecedented scale. The organization operates across compute supply, demand, infrastructure, partnerships, and the physical and commercial systems required to make large-scale compute available. Industrial Compute Strategy & Operations serves as the connective operating layer across this ecosystem. The team works directly with senior leadership across Scaling, Finance, Partnerships, Research, and Infrastructure to translate ambiguous, high-impact challenges into clear strategies, scalable operating mechanisms, and decisive execution. This team is responsible for ensuring that OpenAI’s compute strategy evolves into durable competitive advantage by identifying systemic constraints, aligning stakeholders around critical decisions, and driving the operating mechanisms required to execute at scale. About the Role We are seeking a highly experienced Strategic Operations leader to help shape and operationalize OpenAI’s compute strategy across supply, demand, infrastructure, partnerships, and commercial strategy. This is a senior individual contributor role operating at the intersection of strategy, operations, infrastructure, and executive decision-making. You will work closely with compute leadership to identify the most consequential problems facing the organization, develop structured approaches to solving them, align stakeholders across the company, and drive initiatives from ambiguous concepts through execution. The role will span both strategic and operational work. You may develop long-range compute strategies and investment frameworks, evaluate build-versus-buy decisions, shape major commercial transactions, establish organizational planning mechanisms, or take ownership of a cross-functional initiative that does not have a clear organizational home. Success in this role requires exceptional judgment, analytical

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI Finance is responsible for ensuring the organization is set up for success in pursuit of its mission. The Technical Accounting team plays a crucial role in helping OpenAI navigate complex, judgmental, and rapidly evolving accounting matters with rigor and clarity. We aim to bring both technical excellence and strong business partnership to some of the most novel accounting questions in the industry. About the Role As Senior Manager, Technical Accounting, Compute Infrastructure, you will lead the evaluation, documentation, and operationalization of complex accounting matters related to OpenAI's compute infrastructure, strategic investments, and other non-routine business activities.. This role sits at the intersection of U.S. GAAP technical accounting, infrastructure strategy, financial reporting, controls, and cross-functional execution. Key areas may include cloud compute arrangements, data center and colocation arrangements, lease accounting under ASC 842, power purchase agreements, strategic investments, consolidation evaluations under ASC 810, financial instruments, and other emerging or non-standard arrangements. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead technical accounting analysis for complex, judgmental, and non-routine transactions under U.S. GAAP. Evaluate accounting implications for compute infrastructure arrangements, including cloud compute, data center, colocation, lease, PPA, infrastructure procurement, and related commercial arrangements. Partner with Controllership, Tax, Legal, FP&A, Procurement, Infrastructure, and other cross-functional teams to assess the accounting implications of new products, commercial arrangements, strategic transactions, and business initiatives. Prepare and review technical accounting memoranda, position papers, and other auditor-ready documentation. Translate

AWSRestAIGo
T
📍 Austin, Texas, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. At Tenstorrent, we are building the next generation of AI and RISC-V compute. This role supports our CVP of Operations by turning priorities into execution, driving cross-functional programs forward, and making sure important work does not stall between teams. It is a high-trust role for someone who can bring structure to ambiguity, keep leaders aligned, and move quickly without creating noise. This role is hybrid, based out of Santa Clara,CA; Austin,TX; Boston,MA; or Toronto,ON. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are High-agency operator who moves seamlessly between executive-level priorities and day-to-day execution, owning end-to-end outcomes. Proven at driving complex cross-functional work across operations, finance, legal, recruiting, and business partners. Deep understanding of the semiconductor business, including development timelines, product portfolios, manufacturing, sales and operations planning, and financial investments. Clear, direct, and organized communicator who brings structure, follow-through, and sound judgment to fast-moving, ambiguous environments. Trusted partner to senior leadership, anticipating needs, press

AWSAIGoRust
🔔

Get new senior ai compute engineer jobs in United States by email

Daily job updates · Unsubscribe anytime