Jobiba hiring network

Senior Infrastructure Automation Engineer Jobs

7,101 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior infrastructure automation engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

SA
Scale AI
📍 San Francisco• Full-time• From $216K/yr
20 days ago

The Public Sector software engineers (SWEs) create the core product building blocks forward-deployed teams use to develop agentic capabilities that function across multiple domains. SWEs responsibilities include building the systems required to ingest and process federal datasets to support real-time decision-making in contested environments. We develop novel agentic enabling capabilities that includes: Create multi-layered guardrails around agents Optimize data retrieval for agents Orchestrate fleets of asynchronous agents Automatically alerts users to deviations in data Illustrating how an agent reached a decision As a Senior Software Engineer, you will lead the development of a vertical feature or a horizontal capability to include defining requirements with stakeholders and implementation until it is accepted by the stakeholders. You will: Lead the design and implementation of scalable backend systems and distributed architectures for Federal customers. Manage the full lifecycle of feature development from requirement definition to deployment on classified networks. Direct the orchestration of asynchronous agent fleets to meet mission requirements. Lead customer engagements to translate mission needs into technical requirements. Own the communication with stakeholders to ensure implementation meets defined acceptance criteria. Conduct technical reviews and identify risks within machine learning infrastructure and model serving. Drive the platform roadmap by providing technical specifications for Federal product offerings. Ideally you will have: Full Stack Development: Proficiency in front-end, back-end development and infrastructure, including experience with modern web development frameworks, programming languages, and databases Cloud-Native Technologies: Familiarity with cloud platforms (e.g., AWS, Azure, GCP) and experience in developing and deploying applications in a cloud-native environment. Understanding of containerization (e.g., Docker) and contai

awsazuregcp
View job →
KH
K Health
📍 Tel Aviv• Full-time
20 days ago

About the role: We are seeking a Senior Backend Engineer with deep backend engineering expertise and proficiency in one or more major programming languages (e.g., Python, Java, Go, Rust, or Kotlin), along with a strong understanding of AI models and agents. As a core member of our AI Engineering team, you will collaborate with data scientists, ML engineers, and product managers to build scalable, production-ready infrastructure and APIs that power intelligent systems. What you'll be doing: As a Senior Backend Engineer in the AI Engineering team, you will: Build and maintain reliable, scalable backend services to support AI agent execution and orchestration. Develop AI agent systems for complex operational workflows using LangChain, LangGraph, LiteLLM, and Langfuse. Orchestrate a hybrid model stack that includes OpenAI and Google Gemini alongside self-hosted and fine-tuned LLMs like Gemma and Llama. Build and maintain integrations with clinical systems (FHIR, EMR). Drive observability and reliability using OpenTelemetry, Datadog, and Langfuse. Design APIs (GraphQL, REST), background workers, and event-driven systems that interface with AI inference engines and agent runtimes. Collaborate with Data Science, ML, and engineering teams to deploy AI features and improve the performance, scalability, and reliability of backend systems. Participate in code reviews, knowledge sharing, and mentoring to elevate the team’s technical capabilities. What we're looking for: 6+ years of backend engineering experience, with strong proficiency in more than one major programming language (such as Python, Java, Go, Rust, or Kotlin). Solid understanding of AI systems architecture and experience working in environments involving AI agents, LLMs, or inference pipelines. Proven experience in building and scaling backend APIs, microservices, and background jobs. Strong experience with relational and NoSQL databases (e.g., PostgreSQL, MySQL, MongoDB, Redis), including schema des

pythonjavanode.js
View job →
DU
20 days ago

About the Team The Consumer Engineering Team is responsible for helping consumers discover and order everything they love globally. Our work spans the entire consumer journey across homepage, search, store discovery, item exploration, checkout and post checkout. We aim to craft a hyper-personalized, delightful and frictionless experience for millions of our customers. About the Role As a Senior Staff Machine Learning Engineer on Core Cx, you will set the personalization (P13n) strategy for the entire consumer shopping journey and bring that strategy to life. You will use our robust data and machine learning infrastructure to implement new ML solutions to make the consumer search experience more relevant, seamless, and delightful across restaurant, grocery, retail and all business at DoorDash . You will modernize the recommendation system leveraging AI. You will demonstrate a strong command of production level machine learning, experience with solving end-user problems, and collaborate well with multi-disciplinary teams. You're excited about this opportunity because you will… Drive the engineering vision, strategy, and execution for an organization of 150+ Grow, build, and nurture impactful business-focused product engineering teams. Scale the team by developing leaders internally and attracting world-class talent Mentor and guide a fast-growing organization in setting the right architectural patterns, working with various vendors in the space, and making judicious investments in the right areas anticipating what the company needs a few years down the road. Partner with Business, Product, and other Engineering teams to transform DoorDash from local commerce to agentic commerce We're excited about you because you have… B.S. or M.S. in Computer Science or equivalent. 10+ years of industry experience developing machine learning models with business impact, and shipping ML solutions to production. Proficiency in using AI coding tools (e.g., Claude Code) in th

awsgitrest
View job →

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... Here at GoDaddy, the ML Engineering (MLE) team exists as the backbone of our machine learning infrastructure, enabling ML scientists and product teams across Domains to ship models to production reliably, efficiently, and at scale. This team owns the full lifecycle of ML systems — from CI/CD pipelines and model serving infrastructure to GPU workload orchestration and observability. Through disciplined engineering practices, thoughtful system design, and close collaboration with ML scientists, data engineers, and product teams, we deliver the platform that powers domain search, pricing, recommendations, and emerging AI experiences for millions of customers worldwide. We are currently looking for an experienced, highly motivated Senior Engineering Manager to lead our ML Engineering team based in India. This is an established team with existing engineers — we expect the candidate to ramp up quickly on our ML infrastructure stack, build strong relationships with the team, and partner with both India-based teams and US-based teams to drive execution and grow the team further. This individual will join us on our journey to build and scale ML infrastructure that serves real-time predictions at low latency, automates model deployment and promotion, and provides the observability and reliability guarantees that production ML systems demand. Become part of a team that bridges the gap between ML research and production engineering — shipping systems that directly impact GoDaddy's core revenue. What you'll get to do... Lead a team o

typescriptpythonaws
View job →
G
Godaddy
📍 India• Full-time
28 days ago

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Contribute to the development of GoDaddy’s eCommerce and SSO infrastructure and Kubernetes systems on AWS. On a day-to-day basis you will be working on the team who designs, writes, tests and deploys the infrastructure and application management software for GoDaddy’s eCommerce applications. Expect to learn every day. What you'll get to do... Work as a polyglot engineer, writing and maintaining Infrastructure as code with frameworks/ ecosystems such as Java, Unix CLI, and NodeJS Build and operate infrastructure workflows and deployment pipelines using Kubernetes, Argo Workflows, Argo CD, and GitOps practices Design, build, and own services and APIs in Java, running on Kubernetes-based platforms across AWS and distributed systems Develop and support application and infrastructure delivery pipelines, enabling reliable releases of eComm, Auth and Infrastructure services Collaborate closely with other GoDaddy departments to help advance security and technical standards, maintain regulatory compliances while operating eComm & Auth platforms Your experience should include... 5+ years of strong backend software engineering experience in Java Hands-on experience with Kubernetes, including Helm, Kustomize, or equivalent tools to deploy and manage backend services Experience building and operating high-volume, mission-critical production systems on AWS with continuous deployment (CD) practices Strong experience with infrastructure as code, supporting backend applications and services Experience with observability and l

javanodejssql
View job →
G
Godaddy
📍 United Kingdom• Full-time
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. About The Team The Commerce Site Reliability Engineering team is responsible for the reliability, scalability, and day-to-day operation of the platforms that power GoDaddy's Commerce ecosystem. We build and operate shared infrastructure, support critical production systems, and partner closely with engineering teams to ensure services remain secure, resilient, and highly available. As a Senior Site Reliability Engineer, you'll join a team that values ownership, operational excellence, and continuous improvement. Engineers are empowered to identify problems, drive meaningful change, and influence how reliability is delivered across the broader Commerce organisation. From improving operational maturity and reducing toil to modernising delivery platforms and strengthening incident response practices, this team plays a key role in enabling engineering teams to move quickly and safely. You'll work closely with engineers across infrastructure, cloud, security, networking, and application teams while helping shape the future of reliability engineering at GoDaddy. This role offers significant opportunity to broaden your impact, develop technical leadership skills, and grow toward Staff and Principal engineering positions over time. What you'll get to do... Lead reliability and operational improvement initiatives across GoDaddy's Commerce platform, helping engineering teams build and operate services safely and at scale. Own critical production systems, drive incident response and post-incident improvements, and continuously raise the bar

typescriptpythonaws
View job →

JLL empowers you to shape a brighter way . Our people at JLL are shaping the future of real estate for a better world by combining world class services, advisory and technology for our clients. We are committed to hiring the best, most talented people and empowering them to thrive, grow meaningful careers and to find a place where they belong. Whether you’ve got deep experience in commercial real estate, skilled trades or technology, or you’re looking to apply your relevant experience to a new industry, join our team as we help shape a brighter way forward. What this job involves: As a Senior Project Manager for our Healthcare team, you will serve as a dedicated Owner's Representative, guiding complex infrastructure projects from concept to completion. This position requires a unique blend of deep technical expertise in mechanical and electrical systems and sophisticated client-facing skills to champion our clients' interests. At JLL, we are collectively shaping a brighter way for our clients, and in this role, you will be their trusted advocate—ensuring healthcare facilities are delivered on schedule, within budget, and to the highest standards of quality and regulatory compliance. You will act as the central point of communication, translating intricate technical details into clear, strategic actions for all project stakeholders. What your day-to-day will look like: Serve as the owner’s primary advocate and liaison, fostering collaboration between the client, design teams, contractors, and all project stakeholders. Manage overall project performance, including scope, schedule, and budget, to ensure successful delivery against key milestones. Provide expert technical oversight for mechanical (HVAC, medical gas, plumbing) and electrical (power distribution, emergency power) infrastructure systems. Review design d

artificial intelligenceaiproject management
View job →
N
Nvidia
📍 Remote, United Kingdom• Remote
1mo ago

We are seeking a highly technical and strategic Developer Relations Manager to join our team, with a focus on engaging developer ecosystems across emerging technology domains. In this pivotal role, you will work directly with software solution providers, developers, and industry professionals to foster the adoption of NVIDIA’s advanced AI and computing platforms. The ideal candidate brings a blend of deep technical expertise and commercial go-to-market experience, combined with a passion for developer advocacy and a talent for communicating how NVIDIA technology can solve complex, real-world challenges. NVIDIA is seeking a senior Developer Relations leader to accelerate the integration of NVIDIA technologies across the Energy & Utilities ecosystem in EMEA. This role will focus primarily on the electric power and utilities industry , working with leading software vendors, technology partners, utilities, engineering companies, and energy infrastructure providers to help integrate NVIDIA accelerated computing, AI, simulation, digital twin, and edge technologies into industry software platforms and solutions. The ideal candidate combines strong energy industry domain expertise , technical depth, partner engagement experience, and the ability to identify and develop strategic opportunities with Independent Software Vendors (ISVs). Approximately 80% of the role will focus on utilities, electrical power systems, grid software and related ISVs , with approximately 20% supporting Oil & Gas and adjacent energy applications . What You'll

REMOTEai
View job →

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response. What you'll be doing: Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow) Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes Drive upstream contributions to OVN-Kubernetes and related open-source projects Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure Drive reliability through incident management, resource monitoring, and performance tuning<

pythonawsazure
View job →

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Mainframe Capacity & Performance Engineer to join our Enterprise Infrastructure organization. The Senior Mainframe Capacity & Performance Engineer will serve as a critical technical leader responsible for ensuring the performance, scalability, reliability, and efficiency of our enterprise mainframe environment supporting mission-critical healthcare, pharmacy, and retail applications. As a Senior Mainframe Capacity & Performance Engineer, you will play a key role in capacity planning, workload analysis, performance engineering, and infrastructure optimization across one of the nation's largest and most complex mainframe ecosystems. This position is responsible for proactively monitoring shared mainframe resources, evaluating system utilization trends, identifying performance risks, and providing actionable recommendations to improve overall system health and operational efficiency. The Senior Mainframe Capacity & Performance Engineer will partner closely with Application Development, Mainframe Systems Programming, Infrastructure Engineering, Architecture, Database Administration, Operations, and Business teams to analyze workload behavior, assess resource consumption, identify top consumers, and optimize application performance. This role requires deep expertise in z/OS performance analysis, capacity forecasting, workload managem

Job Details: Job Description: In this role, you'll build software capabilities to automate the build, test, and deployment of Intel's Process Design Kit (PDK). A PDK is a collection of artifacts representing Intel's semiconductor process, used by product designers to model, implement, and verify Intel's mobile, desktop, and server products before manufacturing. You'll be a member of the Design Technology Platform organization, working closely with teams in the United States, Bangalore, and Penang to build world-class DevOps infrastructure that shapes the future of deploying PDKs to silicon product development teams at Intel. Qualifications: Bachelor's or Master's degree in Computer Science, Computer Engineering, or another closely related field. 7-12 years of experience with a Bachelor's degree or Master's degree in software development and engineering. Excellent Python programming skills. Expert-level knowledge of pytest concepts. Expert-level experience with GitHub-based Jenkins CI/CD deployment, enablement, and debugging. Proven knowledge of Agile software engineering practices, IT environments, and DevOps principles and processes. Excellent debugging and problem-solving skills. Excellent written and verbal communication skills; ability to present complex issues with clarity to drive decisions. Must have hands-on experience with AI tools and demonstrated experience building AI agents/tools. Exposure to VLSI PDK domain is preferred. Strong team player with proven ability to collaborate effectively across cross-functional and geographically distributed teams. Job Type: Experienced Hire Shift: Shift 1

pythonaijenkins
View job →

NVIDIA's DGX Cloud (DGXC) powers AI for strategic research and product workloads. The company seeks a Senior Technical Program Manager (TPM) to lead complex, cross-functional programs powering NVIDIA’s next-generation AI software platforms. In this role, you will drive software initiatives across platform services, cloud infrastructure, and system integration. The focus is on enabling scalable, reliable, and supportable software for AI workloads. You will be responsible for managing high-impact engineering programs within a dynamic, fast-paced roadmap, aligning priorities across teams, and ensuring timely, high-quality delivery. This role requires strong technical competence, a proactive approach, and the ability to operate effectively across multiple levels of the organization. This is a software-first TPM role. The ideal candidate has extensive experience managing software initiatives. They also understand the full-stack environment, including infrastructure dependencies, system bring-up, integration readiness, and operational needs to support software across stack layers. What You'll Be Doing: Lead end-to-end execution of software platform initiatives, including planning, execution, delivery, and operationalization. Work together with software, infrastructure, product, and operations teams to ensure alignment on goals, deliverables, achievements, and schedules. Lead cross-functional initiatives encompassing cloud-native services, platform software, system integration, and release delivery. Help connect software roadmap execution to full-stack readiness, including dependencies across infrastructure, bring-up, validation, and downstream operational support. Identify cross-functional dependencies, mitigate risks, and drive resolution of complex technical and programmatic issues. Establish clear success metrics and reporting mechanis

kuberneteslinuxmachine learning
View job →
V
Vanta
📍 United States• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Senior Software Engineers independently drive complex technical work, shape the systems and technical decisions within their teams, and enable other engineers to deliver high-quality, scalable solutions. Vanta's product monitors the security posture of thousands of companies, pulling tens of millions of API calls of data per day, pushing information from hundreds of thousands of laptop agents, and running tests against that data continuously to identify potential security threats. Our infrastructure and tooling need to stay ahead of exponential growth in our customer base. As a Senior Software Engineer at Vanta, you'll drive complex projects across our technical stack, contribute to the technical direction of your team, and mentor other engineers. Your past experience will be leveraged to enable and accelerate Vanta's growth. Visit our Vanta Engineering Blog to learn more about what our team is working on! Tests are at the heart of how Vanta continuously monitors security and compliance for our customers. The Test Core team builds the runtime platform that powers these checks. We own how Tests are scheduled and executed, how their results are persisted and exposed, and the systems that keep this runtime reliable as Vanta grows. In this role, you'll work on some of the core systems behind Vanta's Tests platform. You'll tackle problems around the reliability, correctness, and performance of Test execution, evolve the systems and abstractions that allow the platform to scale, and make it easier for other engineering teams to build on the Tests runtime. Many of these problems span multiple systems and teams and require a deep u

REMOTEtypescriptmongodbgraphql
View job →
V
Vanta
📍 United States• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. You will build the layer that makes Vanta's data actually useful: a single interpretive layer that reads across every source of truth and makes them queryable, reasoned-over, and genuinely illuminating — for EPD, for GTM, and for anyone in the org trying to understand what's happening and why. The EPD Systems team is building the infrastructure Vanta's engineering, product, and design organization depends on to understand itself. We're constructing three interlocking layers: sources of truth at the foundation, an operating system that makes delivery legible, and the intelligence layer that reasons across all of it. This role owns the intelligence layer — the one that doesn't exist yet. This is a builder role. You will ship working things yourself — prototypes, internal tools, agent workflows. You will not hand specs to someone else and wait. Communication isn't a separate deliverable; it's how you learn what to build. What you’ll do as a Senior Product Builder at Vanta: Build the intelligence layer: a cross-source interpretive layer that reads across Vanta's sources of truth and makes them queryable and reasoned-over by EPD leadership, GTM, and beyond Define what to build: scope the problem yourself, make explicit tradeoffs about what to defer, and own the sequence of what gets built and when Ship working things yourself: prototypes, internal tools, agent workflows — using AI as part of how you work, not what you report on Understand the organization: go to the teams whose decisions this layer will serve — GTM, G&A, EPD — and come back knowing what they can't answer today Catch and address AI-specific quality problems — rel

REMOTEai
View job →

SkyLive, the flagship virtual chatroom and meeting-room platform developed by SkyVoice LLC, operates within the global digital communication and real-time interaction ecosystem. The platform enables users, creators, talent communities, corporate groups, and social communities to communicate through real-time voice, video, messaging, virtual rooms, and digital events. As SkyLive continues to expand its international user ecosystem and technology infrastructure, SkyVoice LLC is building a dedicated **Fundraising, Strategic Partnerships, Financial Management, Human Resources, and Risk Management function** to support sustainable business growth. We are seeking an experienced senior professional with approximately **5+ years of relevant experience** across fundraising, investor relations, financial management, business development, strategic partnerships, corporate finance, or risk management. The role will work closely with the Founding and Management teams and will be responsible for developing structured fundraising processes, maintaining strategic relationships, supporting financial planning, coordinating internal teams, and establishing appropriate business and operational risk controls. --- # Strategic Functional Areas The selected professional will work across the following areas: * Fundraising & Investor Relations * Strategic Partnership Development * Financial Planning & Management * Business Risk Management * Investor / Partner Communication * Corporate Documentation & Reporting * HR & Talent Coordination * Compliance & Operational Coordination * Business Development * Management Reporting --- # Key Responsibilities ## A. Fundraising & Investor Relationship Management Own and coordinate the company's structured fundraising and investor relationship activities. Build and maintain professional relationships with potential investors, strategic partners, family offices, venture capital firms, private equity firms, corporate investors, financial institu

accountingfinancerecruitment
View job →
🔔

Get new senior infrastructure automation engineer jobs by email

Daily job updates · Unsubscribe anytime