Jobs in United States

Senior Infrastructure Automation Engineer in United States

2,111 active opportunities · Updated October 2026

Explore current senior infrastructure automation engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

SF
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $200K/yr

Quick readStrong listing-quality and freshness signals

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role At Stitch Fix, we are at the forefront of innovation, creating cutting-edge solutions that blend fashion, technology, and data science. Our data science team combines machine learning with expert human judgment to generate innovative recommendations and insights that transform the way our clients discover what they love. We believe in a curiosity-driven data science culture where members are empowered to deliver impact through end-to-end model development. The diversity of the problems that we work on and the data-rich environment of our business make it possible, even essential, to bring the tools of multiple disciplines to bear on our hardest problems. We are looking for an experienced Styling Algorithms Team Manager to lead a group of talented machine learning engineers and data scientists. In this role, you will shape the future of fashion technology by driving the development and deployment of our styling algorithms, which empower our human stylists to delight clients by nailing their fit and style. This includes ML-, AI-, and product-driven feature curation and testing for our proprietary styling platform, as well as client-facing AI personalization experiences, such as Stitch Fix Vision, our virtual try-on. Responsibilities: Champion bold AI and ML interventions to improve our styling experiences, enabling our stylists to have a multiplicative impact on their client connection points. Likewise, actively shape the product roadmap for direct client-facing styling experiences, expand

PythonRestMachine LearningAI
S
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -94.1%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are looking for a Senior Solution Engineer who is accustomed to solving customer’s most complex problems and closing large deals. In this role you will work directly with the sales team and channel partners to understand the needs of our customers, strategize on how to navigate winning sales cycles, provide compelling value-based demonstrations, support enterprise Proof of Concepts, and ultimately close business. As a Snowflake Solution Engineer you must share our passion about reinventing the database space, thrive in a dynamic environment and have the flexibility and willingness to jump in and get things done. You are equally comfortable in both a business and technical context, interacting with executives and talking shop with technical audiences. This is a remote role that requires about 25% travel. Candidates can sit anywhere in the USA. IN THIS ROLE YOU WILL GET TO: Present Snowflake technology and vision to executives and technical contributors at prospects and customers Work hands-on with prospects and customers to demonstrate and communicate the value of Snowflake technology throughout the sales cycle, from demo to proof of concept to design and implementation Immerse yourself in the ever-evolving industry, maintaining a deep understanding of competitive and com

PythonSQLMachine LearningAI
I
📍 United States· Remote
✓ Quality checkedCompany trend -90.5%

We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Overview Instacarts Detection Engineering team sits at the core of our Security organization, building and operating the systems that identify, surface, and respond to threats across one of North America's largest grocery technology platforms. We own the full detection lifecycle, from telemetry collection and signal design to automated response, across a complex, cloud-native environment spanning endpoint, cloud, container, and SaaS. As a Senior Detection Engineer II, you'll be a technical anchor on the team: developing high-fidelity detection logic, hunting for novel attacker techniques, and raising the bar for how we think about coverage, quality, and scale. You'll work closely with Engineering, Red Team, Incident Response, Fraud, and Trust & Safety to ensure our detections reflect real-world adversary behavior; not just signatures. We operate with a detectio

PythonAWSAzureGCP
B
📍 Hazelwood, Macau S.a.r., United States
✓ Quality checkedCompany trend -20.1%

Senior Systems Engineer - Product Owner Company: The Boeing Company The Boeing Company is looking for a Senior Systems Engineer - Product Owner to join our Government Vehicle Health Management Systems (GVHMS) team in Hazelwood, MO . This dynamic team contributes to cutting-edge vehicle health management solutions that enhance mission readiness and operational efficiency across multiple high-profile aerospace platforms. The role offers a unique opportunity to lead efforts in lab testing, requirements development, and advanced Model-Based Systems Engineering (MBSE) practices. By driving system integration, testing, and verification, the Senior Systems Engineer will ensure seamless performance and adherence to stringent customer standards while collaborating with a diverse group of talented software developers, test engineers, and external partners. This position is ideal for professionals passionate about innovative aerospace technologies and solving complex engineering challenges in a fast-paced, mission-critical environment Position Responsibilities: Develop and maintain product roadmaps and system lifecycle plans Develop, document, and verify requirements Develop project schedules and track performance Develop and organize work items for the software team Be responsible for the integration of the components of the GVHMS system to successfully meet customer requirements Coordinate with the customer on a regular basis Document team plans and status, develop meeting agendas and presentations Analyze process compliance and actively work on improvements This position is expected to be 100% onsite. The sel

Recruitment
M
📍 Minnesota, United States of America, United States
✓ Quality checkedCompany trend -7.1%

We anticipate the application window for this opening will close on - 30 Sep 2026 Careers that change lives start here. Medtronic is a global leader in healthcare technology with a Mission to alleviate pain, restore health, and extend life. Our 95,000 employees work across more than 150 countries to put patients first — developing innovative medical technologies that improve the lives of 72+ million patients each year. Your unique talents will help shape the future of healthcare while building a career grounded in purpose, growth, and impact. A Day in the Life Our Automation Platforms team develops and scales reusable technologies that improve safety, quality, and productivity across Medtronic’s global manufacturing network. We're working onsite 4 days a week as part of our commitment to fostering a culture of professional growth and cross-functional collaboration as we work together to engineer the extraordinary. The person in this role will work from the Medtronic facility located in Fridley, Minnesota. This role will require 25% of travel to enhance collaboration and ensure successful completion of projects. The Sr Principal Robotics & AMR Automation Platforms Engineer will serve as the enterprise technical authority for autonomous mobile robotics, collaborative robotics, and emerging intelligent automation technologies, advancing reusable platforms from technology development through factory validation and global deployment. This role will define technical strategy, platform architectures, roadmaps, standards, and reusable capabilities while partnering with manufacturing sites, engineering teams, IT/OT, suppliers, system integrators, and Medtronic’s Surgical Robotics organization. The successful can

AIRecruitmentHR
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

We are now looking for a Senior SRAM Engineer within our Full Custom Memory (FCM) team! The FCM team designs specialized RAM implementations across NVIDIAs wide array of processing chips. Be it high speed, low power, multiport, we engage closely with processor architecture teams to build custom solutions across the entire NVIDIA silicon portfolio. Are you interested in designing circuits for the next generation of AI chips? Join a team of dedicated engineers developing the custom SRAM circuits that help power these chips. What you'll be doing: Design best-in-class SRAM circuits using state-of-the-art technology processes Optimize circuits for performance, area, and power Collaborate with mask designers to craft high quality and dense pitch-matched layout Verify functionality, electrical integrity, and robustness Improve/develop flows and methodologies to streamline design automation, data collection, and analysis to ensure working silicon What we need to see: BS/MS/PhD (or equivalent experience) in Electrical or Computer Engineering Minimum 12 years of circuit design experience Strong understanding of SRAM and memory design techniques and macro/block development Ways to stand out from the crowd: Self-motivation, attention to detail, clear data analysis and presentation skills Familiarity with industry tools such as Cadence Virtuoso for schematic and layout, SPICE simulators, waveform viewers Background with developing and using various flows and methodologies, in

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

We are developing advanced multi-rack, multi-tenant AI/ML datacenters with NVIDIA GB200, and upcoming GB300 GPUs. NVIDIA seeks a Senior Software Engineer for our CSP (Cloud Service Provider) Engagements team to focus on the cloud-native stack for datacenter products like GB200. In this role, You will define customer workflows, prototype stack enhancements, and debug the toughest Kubernetes + Slurm issues in multi-rack, multi-tenant AI datacenters. You'll tackle complex scheduling challenges across racks, tenants, and clouds as part of the CSP engagements team. What you’ll be doing: Perform deep-dive debugging of multi-rack, multi-tenant clusters: scheduler behavior, container runtime issues, device-plugin crashes, RDMA/IB fabric anomalies, etc. Gather customer requirements and prototype feature extensions for Kubernetes operators, Slurm plugins, and custom micro-services that expose new GPU capabilities. Drive joint architecture reviews and “whiteboard” sessions with CSP and internal platform teams; convert findings into RFCs and upstream pull requests. Create reproducible testbeds (Helm/Ansible/Terraform) that mirror customer environments; automate validation and benchmark suites. Deliver technical collateral-design docs, how-to guides, demo scripts-and present at customer on-sites, KubeCon, and SlurmUG. Collaborate with AE, FAE, and Solution Architect teams to deliver integrated customer solutions and technical documentation. What we need to see: Strong source-level expertise in Kubernetes internals (scheduler, CRI/CNI/CSI, operators) and Slurm (federation, power-save, plugins). Hands-on experience integrating next-gen GPUs (Blackwell/GB200/GB300) or comparable accelerators into containerized clusters. Proven track record debugging large-scale, cloud-native stacks across ne

PythonKubernetesArtificial IntelligenceAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA is seeking a Senior Firmware Engineer to join our CSP Engagements team, focusing on system software for Datacenter products such as GB200. This role combines deep technical expertise in embedded firmware development with customer-facing responsibilities to enable cloud service providers with next-generation computing platforms. You will work at the intersection of hardware and software, driving technical solutions from concept through deployment. What you will be doing: Design and develop firmware solutions for manageability and observability of data center servers. Actively participate in hardware bring-up activities, OOB firmware development, protocol stacks (Redfish, PLDM, MCTP, NSM) and hardware-software co-design for Cloud Service Provider deployments. Debug and troubleshoot NVIDIA GPU firmware issues, power management, performance, and thermal control problems for data center deployments, providing active support to CSPs. Partner directly with CSPs to deliver technical solutions, co-develop & co-debug features and optimizations, and provide support during new product introductions. Perform advanced system debugging, root cause analysis, and performance optimization for large-scale data center environments. Collaborate with AE, FAE, and Solution Architect teams to deliver integrated customer solutions and technical documentation. What we need to see: Deep expertise in data center server architectures, HPC systems, and hardware-software co-design. Deep expertise in embedded firmware, server management controllers, and hardware bring-up with proven track record of shipping production BMC solutions Strong knowledge of DMTF protocols (Redfish, IPMI, PLDM, MCTP, SPDM), telemetry frameworks, and out-of-band management architectures Expert-level skills in C/C&

Artificial IntelligenceAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA is the defining technology company of the Artificial Intelligence era. Our legacy of innovation is powered by extraordinary technology—and extraordinary people. Achieving what has never been done before demands vision, innovation, and the world’s best talent! We are seeking a deeply technical Product Manager to shape the enterprise software stack for next-generation AI factories, built on NVIDIA’s industry-leading GPUs and purpose-designed CPUs. This includes bare-metal orchestration, Kubernetes-native services, and full-stack AI platform capabilities to power the enterprise AI factory. At NVIDIA, you will be part of a diverse, collaborative environment where everyone is inspired to do their best work. Join us and see how you can make a lasting impact on the world. NVIDIA’s Enterprise Product Group is a small, dynamic, and highly motivated team behind DGX systems and DGX SuperPOD—the platforms that established NVIDIA as the gold standard in AI development and deployment. By delivering turnkey, integrated solutions, the team helped transform NVIDIA from a GPU hardware vendor into a full-stack AI and data-center company. We are looking for a proven product leader with a strong track record of delivering enterprise platforms and the expertise to drive continued adoption of these solutions across global enterprises! What you'll be doing: Understand NVIDIA’s strategic position and deliver an enterprise computing platform recognized as the best in the industry. Own the complete solution lifecycle—from solution definition through development and go-to-market—for NVIDIA Mission Control software stack to drive adoption of the NVIDIA technologies in Enterprise market. Leverage strong technical background to define product requirements and user stories for solutions by channeling customer needs, solution-architecture feedback, and cross-team input within NVIDIA. Establis

KubernetesArtificial IntelligenceAI
A
📍 San Francisco, CA, United States
✓ Quality checkedCompany trend -100%

The Opportunity The world of design is changing rapidly, and the Pro Design team is leading that transformation. We are the Adobe organization behind Illustrator, InDesign, and emerging experiences that connect creativity, collaboration, and AI. Our teams are reimagining what professional design looks like for the next decade - building intelligent, connected tools that empower creators and teams to move faster without sacrificing craft. We are looking for a Senior Business Data Scientist who is creative, analytical, and unafraid to question the status quo and shape the decisions that move key business metrics at scale. Join us and build Adobe’s future products! What you'll Do Map the user funnel and build the metrics, cohorts, and dashboards that Product and Growth rely on to see how users move across free, trial, and paid tiers—and pinpoint where they drop off. Dig into the hard questions (what drives activation, which behaviors predict retention and expansion) and build propensity models for conversion, upgrade, churn, and expansion that feed real-time targeting and in-product nudges. Find and size growth bets and work with Product to ship them. Set north-star, driver, and guardrail metrics with your partners, and stand up multivariate experiments across onboarding, paywalls, in-product prompts, and pricing. What you need to succeed Minimum Requirements: Bachelor's degree in a quantitative field (Statistics, Mathematics, Computer Science, Economics, Engineering, or similar) or equivalent practical experience. 5+ years of experience in data science, product analytics, or a similar quantitative role. Proficiency in SQL and Python (or R) for data manipulation, analysis, and modeling. Hands-on experience designing and analyzing A/B tests and interpre

PythonSQLAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA's invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning - the next era of computing - with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as &#34;the AI computing company.&#34; We're looking to grow our company and establish teams with the most thoughtful people in the world. We are looking for an excellent Senior Engineering Manager to lead a large firmware engineering organization delivering end-to-end manageability firmware for NVIDIA's next generation Data Center Compute Systems. This role owns HGX product line and OpenBMC-based management firmware and MCU firmware components in data center platforms, including architecture, execution, quality, reliability, telemetry, and customer readiness. We are seeking an experienced senior leader with strong technical depth, broad system perspective, and a proven ability to lead large teams through complex product cycles. This role is onsite in Santa Clara, CA, USA. If you're creative and autonomous, we want to hear from you! What you'll be doing: Lead a large firmware engineering organization delivering OpenBMC based firmware and MCU firmware for next-generation Data Center Compute Systems. Own HGX platform as a lead for Firmware and System software readiness working across the organization. Define and drive the long-term firmware roadmap, balancing architectural innovation with product execution and delivery milestones. Drive architecture strategy across BMC, MCU, platform software, manageability, health management, and data center firmware interfaces. <spa

PythonGitLinuxArtificial Intelligence
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA is leading groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU -- our invention -- serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables groundbreaking creativity and discovery, and powers inventions that were once considered science fiction, including artificial intelligence to autonomous cars. We are the GPU Communications Libraries and Networking team at NVIDIA. We build communication libraries like NCCL, NVSHMEM, and UCX that are crucial for scaling Deep Learning and HPC. We're seeking a Senior Software Architect to help co-design next-gen data center platforms and scalable communications software. DL and HPC applications have a huge compute demands and already run at scales of up to tens of thousands of GPUs. GPUs are connected with high-speed interconnects (e.g. NVLink, PCIe) within a node and with high-speed networking (e.g. InfiniBand, Ethernet) across nodes. Efficient and fast communication between GPUs directly impacts end-to-end application performance. This impact continues to grow with the increasing scale of next generation systems. This is an outstanding opportunity to advance the state-of-the-art, break performance barriers, and deliver platforms the world has never seen before. Are you ready to build the new and innovative technologies that will help realize NVIDIA's vision? What you will be doing: Investigate opportunities to improve communication performance by identifying bottlenecks in today's systems. Design and implement new communication technologies to accelerate AI and HPC workloads. Explore innovative solutions in HW and SW for our next generation platforms as part of co-design efforts involving GPU, Networking, and SW architects. Build proofs-of-concept, conduct experiments,

LinuxArtificial IntelligenceAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the team that created communication libraries like NCCL, NVSHMEM & technology like GPUDirect -- for scaling Deep Learning and HPC applications. Your customers will have diverse multi-GPU demands, ranging from training on scales up to 100K GPUs to inference down at microsecond latency. Communication performance between the GPUs has a direct impact on AI applications. Your work in AI toolkits will make all of those easier for the community. This is an outstanding opportunity for someone with an AI background to advance the state of the art in this space. Are you ready to contribute to the development of innovative technologies and help realize NVIDIA's vision? What you will be doing: Integrate new communication libraries features in AI frameworks: from PoC to performance analysis to production Perform deep analysis of AI workloads and frameworks to identify multi-GPU communication requirements and opportunities. Collaborate hands-on with teams working on the latest AI models. Improve AI compilers to hide communications or perform automatic fusion. Conduct in-depth AI workload performance characterization on multi-GPU clusters. Design fault-tolerant and elastic solutions for large-scale or dynamic AI workloads. Author

PythonArtificial IntelligenceAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. NVIDIA has a rapidly expanding ecosystem of data center platform designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each brings together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We are searching for a highly motivated engineer to lead performance benchmarking and optimization efforts for our data center products. You will be instrumental in ensuring our data center solutions deliver industry-leading performance for accelerated computing workloads. What you will be doing: Design and execute comprehensive performance benchmarking strategies for our data center platforms and products Characterize real-world AI training, inference, and HPC workloads at scale Define, track, and report key performance indicators (throughput, latency, efficiency, scaling) Build automation tools and frameworks for performance monitoring and analysis Identify and analyze performance bottlenecks across compute, memory, network and storage subsystems Work closely with architecture, hardware,

PythonDockerKubernetesLinux
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -13.7%

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. NVIDIA is looking for phenomenal people like you to help us accelerate the next wave of artificial intelligence. We are looking for a highly motivated senior software engineer for an exciting role in our communication libraries and network software team. The position will be part of a fast-paced crew that develops and maintains software for complex heterogeneous computing systems that power disruptive products in High Performance Computing and Deep Learning. What you will be doing: Design, implement and maintain highly-optimized communication runtimes for Deep Learning frameworks (e.g. NCCL for TensorFlow/Pytorch) and HPC programming interfaces (e.g. UCX for MPI/OpenSHMEM) on GPU clusters. Participating in and contributing to parallel programming interface specifications like MPI/OpenSHMEM. Design, implement and maintain system software that enables interactions among GPUs and interactions between GPUs and other system components. Creating proof-of-concepts to evaluate and motivate extensions in programming models, new designs in runtimes and new features in hardware. What we need to see: M.S./Ph.D. degree in CS/CE or equivalent experience. 5&#43; years of relevant experience. Excellent C/C&#43;&#43; programming and debugging skills. Strong experience with Linux. Expert understanding of computer syst

LinuxArtificial IntelligenceAI
🔔

Get new senior infrastructure automation engineer jobs in United States by email

Daily job updates · Unsubscribe anytime