Jobs in United States

Workload Porting And Performance Engineer in United States

382 active opportunities · Updated October 2026

Explore current workload porting and performance engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build cloud-native data and storage services for hybrid and multi-cloud infrastructure, including dataset discovery, ingestion, governance, checkpointing, observability, and low-latency access. Develop scalable cloud-native services and APIs that support exabyte-scale, high-performance GPU training and inference workflows. Work closely with product managers, internal AI teams, platform teams, and partner engineering teams to understand requirements and turn them into reliable production systems. Collaborate with SRE, operations, and support teams to improve service reliability, performance, observability, on-call readiness, and operational scale. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, and verification. What we need to see: BS in Computer Science, Information Systems, Computer Engineering, or equivalent experience, with 5+ years of software engineering experience. Strong foundation in algorithms, data structures, distributed systems, and practi

PythonJavaAWSAzure
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure. Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration. Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks. Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads. Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification. What we need to see: BS in Computer Science, Information Sys

PythonJavaKubernetesLinux
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -73.6%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Container runtimes were designed for general-purpose software workloads. AI inference is not a general-purpose workload. Running large models at production scale exposes cracks in every layer of the container stack: runtimes unaware of GPU memory constraints, images that take minutes to pull when a model needs to scale to thousands of replicas, and isolation mechanisms that weren't designed for the multi-tenant serving environments that production AI requires. The tools the industry has relied on for a decade weren't built for this, and patching around those limitations at higher layers only goes so far. Baseten owns the entire pipeline, from the moment a developer pushes a model to the moment a request gets a response. That vertical ownership means we can fix these problems at the root. The Runtime Fabrics team is doing exactly that: purpose-building the container runtime and storage layers for AI inference workloads, led by some of the world's top containerd maintainers. As Engineering Manager of the Runtime Fabrics team, you will lead this work, setting technical direction, growing a world-class team of systems engineers, and ensuring the team's output shapes not just Baseten's infrastructure but the open-source container ecosystem at large. If you've contributed to containerd, runc, or related OCI projects and are ready to lead a team solving some of the hardest problems in infrastructure today, we'd love

LinuxMachine LearningAIC++
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -83.5%

From $204K/yr

Quick readStrong listing-quality and freshness signals

The opportunity Datadog’s Infrastructure products help engineers understand and operate the systems their applications depend on. Our customers work in complex environments like Kubernetes and serverless, where infrastructure changes constantly, information is dense, and decisions about reliability, performance, and cost are closely connected. We’re looking for a Staff Product Designer to join Modern Compute, with an initial focus on Containers Autoscaling. Autoscaling helps engineering teams make better decisions about how their applications and infrastructure use resources. Designing these experiences requires making deeply technical systems understandable, helping customers act with confidence, and fitting into the tools and workflows they already use. The team is rethinking how workload and cluster autoscaling come together as a more coherent product experience. This includes how customers get started, understand recommendations, evaluate value, and safely apply changes across their environments. The work also connects to other parts of Datadog, including observability, Cloud Cost Management, permissions, and AI-assisted workflows. As a Staff Product Designer, you will help define that direction and lead the work from early problem framing through shipped product. You will partner closely with product and engineering, bring a high level of interaction and visual craft to complex workflows, and help raise the quality of design across Modern Compute. At Datadog, we place value in our office culture, the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to help our Datadogs find a work-life rhythm that works for them. What you’ll do Lead end-to-end product design for Modern Compute, initially focused on our Autoscaling product. Help define the product direction for an area that is still evolving, from early framing and exploration through detailed design and delivery. Design clear, trustwort

KubernetesGitAIGo
M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $151K/yr

Quick readStrong listing-quality and freshness signals

The Role MongoDB is seeking a Staff Product Manager to lead strategic initiatives across Identity and Security. In this role, you will serve as the senior product leader shaping the next generation of authentication, authorization, and access control across MongoDB Atlas, the core database engine, and emerging developer platforms. You will lead the strategy for non-human identity—spanning workload identity federation, zero-trust passwordless access, Model Context Protocol (MCP) integrations, and autonomous AI agentic identity. You will partner closely with engineering, UX, product marketing, and enterprise customers to build identity experiences that are enterprise-grade, seamless for developers, and secure by default. Responsibilities Set and champion a multi-year product strategy across interrelated identity areas, aligning dependencies and investments to MongoDB’s broader objectives Identify cross-cutting customer and market opportunities, build business cases for investment, and bring a clear point of view to senior stakeholders Lead workload identity federation and programmatic access initiatives across cloud and enterprise identity environments Define secure authorization approaches for AI agents and MCP clients, including delegation, identity lifecycle, policy, and administrative controls Drive developer-first access management, ensuring security controls enhance rather than obstruct developer velocity Partner with engineering, design, GTM, and cross-functional leadership to validate enterprise requirements, prioritize initiatives, and measure feature adoption Stay ahead of evolving cloud security paradigms, emerging standards (e.g., SPIFFE/SPIRE, O4AA), and AI access patterns to guide executive decision-making Represent MongoDB’s security and identity vision at industry events, customer advisory boards, and executive briefings Demonstrate creativity, an innovative spirit, and a bias towards action Requirements 8+ years (or equivalent experience) in Pro

MongoDBAWSAzureGCP
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Storage Infrastructure team builds and operates the storage foundation behind OpenAI’s most demanding workloads. We work directly with research to design storage systems for rapidly evolving experiments, while also powering production at scale. We own the platform end to end: backend systems, user-facing services and APIs, and the control planes that manage how data is placed, moved, and retained over time. Our stack spans cloud and in-house object stores across very different workload profiles, from GPU-attached systems to dedicated storage hardware. We also build the federation layer that unifies these backends behind a simple interface and routes each workload to the right storage solution. About the Role You will help build the storage platform that powers OpenAI’s research and production systems. This is a hands-on infrastructure role for engineers who want to work on deeply technical systems at scale and own them in production. You’ll work across object storage, cross-region data movement, lifecycle management, and the federation layer that provides a unified interface across multiple backends. Much of our stack runs on Kubernetes, and we primarily build services in Rust. In this role, you will: Build and operate storage services that underpin OpenAI’s research infrastructure Develop object storage systems across cloud and in-house environments Build systems for cross-region data movement, replication, and recovery Design lifecycle management capabilities that keep data durable, available, and cost-effective Evolve the federation layer that unifies multiple backend systems behind a simple interface Improve performance, reliability, and operational excellence across the platform Collaborate closely with researchers and infrastructure teams to support rapidly evolving workloads You might thrive in this role if you: Have experience building or operating distributed systems in production Have worked on storage infrastructure, object stores, dist

AWSKubernetesRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team At OpenAI, we’re building safe and beneficial artificial general intelligence. We deploy our models through ChatGPT, our APIs, and other cutting-edge products. Behind the scenes, making these systems fast, reliable, and cost-efficient requires world-class infrastructure. The Caching Infrastructure team is responsible for building a caching layer that powers many critical use cases at OpenAI. We aim to provide a high-availability, multi-tenant cache platform that scales automatically with workload, minimizes tail latency, and supports a diverse range of use cases. We’re looking for an experienced engineer to help design and scale this critical infrastructure. The ideal candidate has deep experience in distributed caching systems (e.g., Redis, Memcached), networking fundamentals, and Kubernetes-based service orchestration. In This Role, You Will: Design, build, and operate OpenAI’s multi-tenant caching platform used across inference, identity, quota, and product experiences. Define the long-term vision and roadmap for caching as a core infra capability, balancing performance, durability, and cost. Collaborate with other infra teams (e.g., networking, observability, databases) and product teams to ensure our caching platform meets their needs. You Might Thrive In This Role If You: Have 5+ years of experience building and scaling distributed systems, with a strong focus on caching, load balancing, or storage systems. Have deep expertise with Redis, Memcached, or similar solutions, including clustering, durability configurations, client-side connection patterns, and performance tuning. Have production experience with Kubernetes, service meshes (e.g., Envoy), and autoscaling systems. Think rigorously about latency, reliability, throughput, and cost in designing platform capabilities. Thrive in a fast-paced environment and enjoy balancing pragmatic engineering with long-term technical excellence. About OpenAI OpenAI is an AI research and deployment company d

RedisAWSKubernetesRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions tailored to the demands of advanced AI workloads. We work across the full stack—from silicon to system integration—partnering closely with internal teams and external vendors to define and deliver next-generation AI infrastructure. Our team focuses on defining scalable, high-performance system architectures and reference designs that balance performance, cost, and operational efficiency across rapidly evolving technologies. About the Role We are seeking a 3P Architect to define and drive rack- and cluster-level reference designs in collaboration with external partners. This role is responsible for translating workload requirements and system-level goals into concrete architectures, aligning partners on critical design attributes, and ensuring vendor roadmaps meet our infrastructure needs. You will work closely with performance modeling and internal architecture teams to evaluate tradeoffs, while owning the end-to-end definition and execution of third-party system designs. This includes identifying gaps in current technologies, driving vendor development, and shaping future infrastructure capabilities. This role requires strong system intuition, cross-functional leadership, and the ability to operate effectively across internal teams and external ecosystems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Define rack- and cluster-level reference architectures for AI infrastructure deployments. Translate workload requirements into clear system design specifications and partner deliverables. Collaborate with performance modeling teams to evaluate architectural tradeoffs and system behaviors. Align internal stakeholders and external partners on critical system attributes (performance, cost, power, reliability, scalability). Identify gaps in current technology offerings and dr

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team: Compute Infrastructure builds the platform that turns enormous amounts of compute into a reliable engine for frontier AI. We design, provision, schedule, operate, and optimize the systems that connect accelerators, CPUs, networks, storage, data centers, orchestration software, agent infrastructure, developer tools, and observability into one coherent experience for researchers and product teams. Our work spans the entire stack: capacity planning and cluster lifecycle, bare-metal automation, distributed systems, Kubernetes and scheduling, deep system optimization, high-performance networking, storage, fleet health, reliability, workload profiling, benchmarking, and the developer experience that lets teams use enormous compute systems with confidence. At this scale, small improvements to communication, scheduling, hardware efficiency, or debugging workflows can compound into meaningful research velocity. We are hiring across Compute Infrastructure rather than for a single narrow team, and we use this opening to match strong engineers to the problems where they can have the most leverage. About the Role We are looking for engineers who want to build the compute platform behind OpenAI's research and products. You may not be the strongest in low-level systems, high-performance computing, distributed infrastructure, reliability, CaaS, agent infrastructure, developer platforms, tooling, or the user experience around infrastructure. What matters is that you can reason carefully about complex systems, write durable software, and raise the quality and velocity of the people around you. Depending on your background and interests, you might work close to hardware, close to users, on CaaS and agent infrastructure, or on the control planes and data planes in between. You could help bring new supercomputing capacity online, optimize training workloads from profiler traces and benchmarks, improve NCCL and collective communication behavior, reason about GPUs, NICs, t

AWSKubernetesRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions optimized for advanced AI workloads. We collaborate across research, software, and external hardware partners to design and deploy next-generation AI systems at scale. Our team works closely with silicon vendors and system partners to evaluate emerging technologies, validate performance characteristics, and ensure that hardware capabilities translate effectively to real-world AI workloads. About the Role We are seeking a 3P Hardware Architecture Expert with deep expertise in GPU and accelerator architectures to engage directly with silicon vendors and guide hardware decisions for AI infrastructure. In this role, you will evaluate architectural tradeoffs across compute, memory, and interconnect systems, translating vendor specifications into real-world workload impact. You will play a critical role in early silicon evaluation, benchmarking, and performance validation, helping ensure that next-generation hardware meets the needs of our workloads. This role is highly hands-on and requires both deep technical understanding and the ability to engage at a high level with partners such as NVIDIA and AMD on architectural direction and design tradeoffs. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Engage deeply with silicon vendors (e.g NVIDIA & AMD) on GPU and accelerator architecture tradeoffs. Analyze and interpret performance, power, and efficiency characteristics of next-generation hardware. Translate vendor specifications into expected real-world performance for AI workloads. Evaluate architectural aspects including: compute throughput and utilization memory systems (HBM, cache hierarchies, bandwidth constraints) data types and precision tradeoffs (FP16, BF16, FP8, etc.) interconnect and scaling behavior. Run benchmarks and profiling to validate hardware performance a

AWSRestAIRust
S
📍 Bellevue, Washington, United States· Full-time
✓ Quality checkedCompany trend -91.7%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. ABOUT THE TEAM: We exist to improve Product leaders’ understanding and decision-making on how to profitably grow the business. Our North Star is to increase workload contribution margin dollars above the current forecast by serving as an objective, full-stack business partner. We do this by measuring key metrics, inferring actionable insights, catalyzing strategic decisions, and accelerating operational efficiency across Product, Engineering, and Sales. ABOUT THE ROLE: This is a high-visibility role that sits at the intersection of product strategy and financial analysis, giving you a unique vantage point to influence how we scale our business. We aren't looking for someone to just execute tasks and pull data; we need a partner who can navigate ambiguity, drive workstreams, and anticipate the next steps. This role is best for proactive problem-solvers with business intuition and a high level of curiosity. You will work on critical projects and collaborate with executives to turn complex data into decisive action. WHAT YOU WILL DO: Develop & Maintain Financial Models: Build and iterate on financial models that track both revenue and cost drivers, enabling accurate forecasting and real-time visibility into workload contribution margins. Drive Revenue & Margin Insights

SQLAIGoRust
T
📍 Springfield, IL 62704-6517, United States
✓ High-confidence listingCompany trend +89.4%

From $16/hr

Quick readStrong listing-quality and freshness signals

The Starting Hourly Rate / Salario por Hora Inicial is $16.00 USD per hour. The Pay Range / Rango salarial is $16.00 USD - $24.00 USD per hour. Effective, 09/20/26 to 01/02/2027, non-exempt team members will be eligible to receive a temporary additional $1 on top of an already existing $1/hr shift differential for eligible overnight or early morning hours worked. Please talk to HR for details. ALL ABOUT TARGET Working at Target means helping all families discover the joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here . ALL ABOUT SEASONAL JOBS Seasonal Specialty Sales Roles: A sales force of specialized consultants who provide tailored suggestions and solutions through active selling and compelling visual merchandising presentations that inspire guests and build the basket. Seasonal Service & Engagement: Advocates of guest experience who welcome, thank, and exceed guest service expectations by focusing on guest interaction and recovery. Seasonal General Merchandise & Food Sales: Experts of operations, process and efficiency who enable a consistent experience for our guests by ensuring product is set, in-stock, accurately priced and signed on the sales floor. At Target we believe our team members have meaningful experiences that help them build and develop skills for a career. These roles can provide you with the: Knowledge of guest service fundamentals and experience supporting a guest first culture across the store Experience in retail business fundamentals: department sales trends, inventory management, and process efficiency and improvement Experience executing daily/weekly workload to support business priorities and deliver on sales goals <p

T
📍 Saugus, MA 01906-3100, United States
✓ High-confidence listingCompany trend +89.4%

From $18/hr

Quick readStrong listing-quality and freshness signals

The Starting Hourly Rate / Salario por Hora Inicial is $18.00 USD per hour. The Pay Range / Rango salarial is $18.00 USD - $27.00 USD per hour. Effective, 09/20/26 to 01/02/2027, non-exempt team members will be eligible to receive a temporary additional $1 on top of an already existing $1/hr shift differential for eligible overnight or early morning hours worked. Please talk to HR for details. ALL ABOUT TARGET Working at Target means helping all families discover the joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here . ALL ABOUT SEASONAL JOBS Seasonal Specialty Sales Roles: A sales force of specialized consultants who provide tailored suggestions and solutions through active selling and compelling visual merchandising presentations that inspire guests and build the basket. Seasonal Service & Engagement: Advocates of guest experience who welcome, thank, and exceed guest service expectations by focusing on guest interaction and recovery. Seasonal General Merchandise & Food Sales: Experts of operations, process and efficiency who enable a consistent experience for our guests by ensuring product is set, in-stock, accurately priced and signed on the sales floor. At Target we believe our team members have meaningful experiences that help them build and develop skills for a career. These roles can provide you with the: Knowledge of guest service fundamentals and experience supporting a guest first culture across the store Experience in retail business fundamentals: department sales trends, inventory management, and process efficiency and improvement Experience executing daily/weekly workload to support business priorities and deliver on sales goals <p

T
📍 Santa Fe, NM 87507-2606, United States
✓ High-confidence listingCompany trend +89.4%

From $18/hr

Quick readStrong listing-quality and freshness signals

Starting Hourly Rate / Salario por Hora Inicial: $17.75 USD per hour Effective, 09/20/26 to 01/02/2027, non-exempt team members will be eligible to receive a temporary additional $1 on top of an already existing $1/hr shift differential for eligible overnight or early morning hours worked. Please talk to HR for details. ALL ABOUT TARGET Working at Target means helping all families discover the joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here . ALL ABOUT SEASONAL JOBS Seasonal Specialty Sales Roles: A sales force of specialized consultants who provide tailored suggestions and solutions through active selling and compelling visual merchandising presentations that inspire guests and build the basket. Seasonal Service & Engagement: Advocates of guest experience who welcome, thank, and exceed guest service expectations by focusing on guest interaction and recovery. Seasonal General Merchandise & Food Sales: Experts of operations, process and efficiency who enable a consistent experience for our guests by ensuring product is set, in-stock, accurately priced and signed on the sales floor. At Target we believe our team members have meaningful experiences that help them build and develop skills for a career. These roles can provide you with the: Knowledge of guest service fundamentals and experience supporting a guest first culture across the store Experience in retail business fundamentals: department sales trends, inventory management, and process efficiency and improvement Experience executing daily/weekly workload to support business priorities and deliver on sales goals WHAT WE ARE LOOKING FOR We might be a great match if: <

T
📍 Naples, FL 34109-2006, United States
✓ High-confidence listingCompany trend +89.4%

From $17/hr

Quick readStrong listing-quality and freshness signals

Starting Hourly Rate / Salario por Hora Inicial: $17.00 USD per hour Effective, 09/20/26 to 01/02/2027, non-exempt team members will be eligible to receive a temporary additional $1 on top of an already existing $1/hr shift differential for eligible overnight or early morning hours worked. Please talk to HR for details. ALL ABOUT TARGET Working at Target means helping all families discover the joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here . ALL ABOUT SEASONAL JOBS Seasonal Specialty Sales Roles: A sales force of specialized consultants who provide tailored suggestions and solutions through active selling and compelling visual merchandising presentations that inspire guests and build the basket. Seasonal Service & Engagement: Advocates of guest experience who welcome, thank, and exceed guest service expectations by focusing on guest interaction and recovery. Seasonal General Merchandise & Food Sales: Experts of operations, process and efficiency who enable a consistent experience for our guests by ensuring product is set, in-stock, accurately priced and signed on the sales floor. At Target we believe our team members have meaningful experiences that help them build and develop skills for a career. These roles can provide you with the: Knowledge of guest service fundamentals and experience supporting a guest first culture across the store Experience in retail business fundamentals: department sales trends, inventory management, and process efficiency and improvement Experience executing daily/weekly workload to support business priorities and deliver on sales goals WHAT WE ARE LOOKING FOR We might be a great match if: <ul

🔔

Get new workload porting and performance engineer jobs in United States by email

Daily job updates · Unsubscribe anytime