Jobiba hiring network

Senior Infrastructure Automation Engineer Jobs

7,101 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior infrastructure automation engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Product Manager, Product Platform Company Overview Okta is The World’s Identity Company. We free everyone to safely use any technology - anywhere, on any device, for any app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. Role Overview: Senior Product Manager, Product Platform This is a rare 0-to-1 opportunity to build a brand-new, high-visibility capability that touches every Okta customer: self-service visibility into subscription and license usage, delivered directly in the Admin Console. You’ll be responsible for building a modern, self-service experience that increases customer trust, unlocks new upsell and expansion motions, and sets the standard for how Okta communicates value to its largest accounts. You'll partner closely with Engineering, Field, Pricing & Packaging, and Product leadership to shape a capability with direct, measurable impact on customer retention and revenue growth. Key Responsibilities Orchestrate the long-term product vision and strategy for Okta's subscription usage and licensing visibility, ensuring the Admin Console roadmap scales with the breadth of Okta's customer base and product portfolio. Architect a self-service usage dashboard and reporting framework to ensure customers can seamlessly and securely track their

awsgitrest
View job →
O
Okta
📍 San Francisco• Full-time• From $1.9M/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Role Overview As the Senior Commission Manager, you will serve as a key operational leader driving execution, calculation integrity, and continuous process optimization across Okta’s end-to-end sales compensation lifecycle. Supporting a high-growth field organization of 2,300+ sellers, this role bridges hands-on commission administration—validations, payment submissions, dispute resolution, and operational modernization—with advanced reporting, enterprise analytics, and AI-driven process automation. With a sharp focus on execution, data and incentive design governance, process engineering, and emerging AI tools, you will eliminate manual bottlenecks, build scalable reporting infrastructure, and guarantee 100% payout accuracy within tight deadlines. Key Responsibilities Commission Execution & Calculation Integrity Oversee the monthly commission calculation and payout lifecycle for 2,300+ global sellers, actively streamlining and operationalizing processes to maintain near-zero error margins. Operationalize data reconciliations between upstream source systems (CRM and Workday) and Xactly Incent, optimizing workflows to safeguard data integrity, maintain accurate crediting rules, and ensure calculation accuracy. Process Engineering, Automation & AI Map core commission workflows to eliminate friction, slash monthly close processing times, and eliminate manual data manipulation. Identify and deploy automation tools, custom scripts, AI-driven effic

sqlawsrest
View job →
V
Vanta
📍 United States• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. You will build the operating system for EPD — the systems, agents, and practices that make engineering, product, and design unreasonably effective at delivery, with Jira as the substrate. The EPD Systems team is building the infrastructure Vanta's engineering, product, and design organization depends on to understand itself and operate well. We're constructing three interlocking layers: sources of truth at the foundation, a shared interpretive layer that reasons across all of them, and the operating system that makes the work itself legible. This role owns the operating system. This is not a traditional PMO or status-reporting role. You build — agents, automation, workflow design, and the practices that make teams want to use the system rather than route around it. What you’ll do as a Senior Systems Designer at Vanta: Design and own workflow and hierarchy across engineering, product, and design in Jira Make sure Jira reflects real practice, and real practice reflects what needs to show up in Jira — in both directions Build the agents and automation that run the system yourself, with AI as part of how you work Drive adoption: go to the teams whose practice doesn't match the substrate today and change that — through conversation, well-built artifacts, or clear instruction that works without you in the room Understand delivery breakdowns at the mechanism level and build fixes that address the root cause, not the symptom Build alongside teammates who are growing into more technical work — raise their ceiling, not just your own output How to be successful in this role: You build and ship, recently, with AI as part of how you work. "

REMOTEai
View job →
B
Baseten
📍 San Francisco• Full-time
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an OS / K8s Systems Engineer at Baseten, you’ll build the automation and systems that turn raw GPU hardware into production-ready compute. From provisioning to orchestration, you’ll own the software layer that makes our infrastructure reproducible, scalable, and reliable across data centers. This is a senior, hands-on role focused on building systems not operating them. You’ll work close to the metal designing OS images, building provisioning pipelines, and automating cluster bring-up from scratch. Your work will define how quickly we can turn new capacity into usable compute. EXAMPLE INITIATIVES Zero-to-cluster automation Build workflows that take new hardware from unprovisioned to fully operational cluster. Provisioning systems Design PXE-based or equivalent systems for imaging and lifecycle management. Reproducible infrastructure — Ensure clusters deploy consistently across data centers. RESPONSIBILITIES Own the end-to-end automation of cluster bring-up and lifecycle management. Build and maintain OS images, provisioning systems, and configuration pipelines. Deploy and operate cluster orchestration platforms (Kubernetes, Slurm, or similar). Design systems for reproducibility across sites and hardware generations. Automate upgrades, rollouts, and failure recovery. Optimize system performance, including GPU utilization and networking. Partner with hardware and network teams to validate and improve system b

pythonkuberneteslinux
View job →
H
1mo ago

We’re looking for an Senior Software Developer, Backend who can help us support the development organization to deliver value to customers in a reliable, efficient, and safe manner. You’ll be working in a focused team that owns one piece of the production application environment and the developer experience, you will execute on defined projects to achieve team-level goals. In line with Hootsuite's distributed workforce strategy, our flexible work arrangement allows for a hybrid model. This role is open to applicants within commutable distance to Luxembourg. WHAT YOU’LL DO: Write software - tools, libraries, automation, services Design and build our infrastructure platform Identify and implement new platform features Research and evaluate new technologies Refactor, rewrite or retire existing platform features Operate our developer experience and production application environments Diagnose and repair our distributed systems Perform maintenance, upgrades and migrations Control or eliminate repetitive tasks, alert noise, and business-as-usual work Enable development teams Provide executable interfaces to our infrastructure platform Provide tools and best practices to support the entire software development lifecycle Participate in a flexible on-call rotation Communicate by writing documentation, participating in meetings, and showing off your work at demos WHAT YOU’LL NEED: A degree in Computer Science or Engineering or equivalent experience working in a software engineering role An ability to write software and working knowledge of software engineering practice (Java programming language and strong working knowledge of object-oriented programming concepts) Proven experience creating stable, reliable, performing and maintainable code Familiarity with data modeling and schema design Knowledge of data structures and algorithms Open Communication: clearly conveys thoughts, both written and verbally, listening attentively and asking questions for clarific

javalinuxagile
View job →
D
1mo ago

As a Senior Platform Product Manager focused on AI SDLC Trusted Throughput, you will define and drive the product strategy for enabling safe, reliable software delivery at AI-native scale across Datadog’s Internal Developer Platform. As AI accelerates development velocity and system complexity, you will help evolve SDLC systems from human-supervised workflows to platforms with built-in safety, observability, and correctness guarantees. You will partner closely with engineering, security, and developer platform teams to improve deployment reliability, operational visibility, and governance while enabling both engineers and AI agents to move quickly with confidence. This role offers the opportunity to shape foundational developer infrastructure and influence how AI-powered software delivery operates across Datadog. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the product strategy, roadmap, and execution for AI-native SDLC throughput and reliability initiatives across Datadog’s Internal Developer Platform Define and drive platform outcomes aligned to DORA metrics, balancing deployment velocity with reliability, change failure reduction, and operational safety Partner with engineering, infrastructure, security, and developer experience teams to build automated validation, auditability, and risk-scoring capabilities into deployment workflows Deliver actionable SDLC observability and diagnostic capabilities that connect executive-level metrics to operational signals across the software delivery lifecycle Drive systems that monitor and validate AI-generated or AI-attributed changes to ensure correctness, compliance, and trustworthy automation Serve as a cross-functional product leader across SDLC Foundations, Security Engineering, and compl

aigorust
View job →
E
18 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a hands-on senior leader for our India SecOps team, you will shape and safeguard Everpure’s security posture at the intersection of detection engineering, threat hunting, attack surface management, and incident response. Positioned as a strategic cornerstone in Bangalore, you will empower an elite engineering team, optimize critical SecOps pipelines, and partner cross-functionally across global engineering and infrastructure groups. By driving execution excellence and high team morale, you ensure our enterprise platform and global telemetry remain resilient against evolving threats. WHAT YOU'LL DO Scale & Lead SecOps Operations: Architect, mentor, and grow the India SecOps team to foster an environment of high morale, technical excellence, and rapid execution across detection engineering and incident response. Proactively Manage & Remediate Attack Surface: Own end-to-end Attack Surface Management (ASM) across cloud environments, SaaS applications, endpoints, and secrets management to measurably minimize enterprise exposure and mitigate risk. Optimize Telemetry & Incident Response: Mature SIEM and SOAR automation pipelines to drastically reduce mean time to detect, contain, and respond (MTTD/MTTC/MTTR) while continuously elevating alert fidelity and signal confidence. Drive Cross-Functional Alignment & RCA Postmortems: Lead continuous validation through purple-teaming and incident postmortems alo

awsazureci/cd
View job →
A
Affirm
📍 Poland• Full-time• Remote• $192K – $288K/yr
18 days ago

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. We’re looking for a curious, driven professional to join our Revenue Analytics team. This builds and owns the data products, reporting infrastructure, semantic foundations, and analytical systems that power Affirm’s Revenue organization. As a Senior Analyst, Revenue Analytics, you’ll build scalable data products that power day-to-day decision-making - owning end-to-end work across data modeling, metric definitions, dashboards, automation, and enablement. You’ll also help strengthen our semantic layer and data governance, laying the foundation for reliable AI. The ideal candidate combines strong technical and analytical skills with the ability to turn ambiguous business questions into durable, well-tested data infrastructure. What you'll do Develop dbt data models, dashboards, metrics, and automation processes for the revenue field team and revenue analysts Build and maintain critical reporting data models that power external merchant reporting Build the semantic, metadata, and context layers that allow AI systems to accurately understand Revenue data, metrics, and business definitions Partner with Business Systems, engineering, and business stakeholders to translate requirements into durable, well-tested data products Contribute to the team’s best practices in version control, code review, documentation, and release hygiene (GitHub-based workflows) Develop processes, governance, and foundations to scale the impact of analytics within Revenue. What we look for 3+ years of work experience in an analytics engineering or business intelligence role Strong working knowledge of SQL, dbt, Python, data modeling, and data visualization Hands-on experience with BI tools (Sigma/Looker/Tableau), Databricks, and cloud data warehouses (Snowflake) Understanding of the data f

REMOTEpythonsqlgit
View job →
RS
Redwood Software
📍 Hyderabad• Full-time
18 days ago

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leading orchestration platform for the autonomous enterprise, driving business transformation at the lowest total cost of ownership. Redwood empowers organizations to intelligently automate and orchestrate mission-critical business and IT processes across complex ERP, hybrid cloud, data and emerging agentic AI systems. Through its SaaS-first automation fabric—with AI embedded across the automation lifecycle—Redwood accelerates the path to autonomous operations. Backed by 30 years of experience and trusted by more than 50% of the Fortune 50, Redwood helps organizations unlock human potential to focus on innovation, growth and what’s next. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for a Software Engineer, Platform & Integrations . You will own the end-to-end delivery of complex features, optimize backend performance, and collaborate on architectural decisions that impact more than 1,000 enterprise customers worldwide. This is an ideal role for an experienced engineer who thrives on technical autonomy, enjoys tackling complex cloud-native challenges, and wants to have a direct impact on product direction. Feature Ownership & Architecture: Design, build, and maintain scalable, secure, and highly observable backend services and microservices using Java and Spring Boot. Platform Evolution: Actively contribute to upgrading our core platform infrastructure, focusing on system resilience, performance tuning, and seamless component communication. Security & Compliance: Implement rigorous secure coding practices to safeguard data exchange and ensure compliance across our multi-tenant SaaS environment. Collaborative Execution: Work closely with Product, QA, and senior le

javasqlpostgresql
View job →

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a best-in-class family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from a diverse group of backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a senior validation lead engineer to lead at-scale rack validation efforts for next-generation AI hyperscale systems. This role focuses on post-silicon system validation across the full lifecycle, ensuring functional, electrical, and thermal performance meets product objectives. You will own end-to-end blade and rack validation including planning, development, execution, and debug while collaborating across firmware, systems, and hardware teams. The Team The Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Responsibilities and Duties Lead post-silicon validation of AI compute blades and racks including test planning, development, and automation. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness. Own execution against program achievements and report validation progress and risks. Triage test failures, collect debug data, and collaborate on root cause analysis. Track

pythonci/cdlinux
View job →

Staff -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debu

pythonlinuxai
View job →
G
18 days ago

Staff -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debu

pythonlinuxai
View job →
G
22 days ago

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team The Network Security team is responsible for securing GoDaddy's global hybrid infrastructure across data centres, cloud, edge, and remote-access environments. We partner across Security, Infrastructure, Cloud, and Engineering teams to build scalable, resilient, and secure solutions that support the business. As a Principal Network Security Engineer, you'll act as a senior technical leader, helping define network security strategy, influence architecture across teams, and drive security outcomes at enterprise scale through technical expertise, systems thinking, and cross-functional leadership. What you'll get to do... Define and drive network security architecture across hybrid environments, including data centres, cloud, edge, and remote-access technologies Design trust boundaries, segmentation strategies, secure connectivity patterns, and network controls that reduce risk and improve security posture Lead complex technical initiatives, migrations, and architectural decisions while balancing security, reliability, performance, and operational requirements Establish scalable approaches for policy governance, automation, monitoring, telemetry, and security control validation Partner across engineering organizations to drive large-scale initiatives, mentor engineers, and influence technical direction through architecture reviews and technical leadership Your experience should include... 10+ years of experience in Network Security Engineering, Network Architecture, or Security Engineering, including ownership of enterprise-scale secu

awsci/cdai
View job →
G
Godaddy
📍 United States• Full-time• From $154K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

pythonkubernetesai
View job →
P
Pendo
📍 Raleigh• Full-time• $105K – $120K/yr
1mo ago

The Team + The Role Pendo’s Business Systems team keeps the company running smoothly across our global office footprint. The team owns the systems, infrastructure, automations, and day-to-day support that help Pendonauts work securely, efficiently, and at scale. As Pendo grows, this team plays a critical role in finding smarter, more scalable ways to support the business. As a Senior IT Engineer, you will own key IT infrastructure initiatives, lead Okta administration, build automations that reduce manual work, and support the SaaS tools that power Pendo’s operations. This role goes beyond ticket resolution: you will work hands-on across systems, support, automation, A/V, SaaS administration, and infrastructure improvements while partnering closely with teams across the business. You will operate with a high degree of autonomy and help make Pendo’s IT function more reliable, secure, scalable, and forward-looking. This role is based in our Raleigh office. What this looks like day-to-day Okta administration: Own Okta administration as a core focus area, including SSO application implementation, onboarding and offboarding workflows, SCIM provisioning, and account lifecycle management. Build and maintain clean integrations that make access management seamless and scalable. Automation development: Use tools such as Workato, n8n, Okta Workflows, APIs, and AI tools to eliminate manual work across IT operations. Build automations for user account creation, workflow orchestration, tool integrations, and recurring support patterns. IT support: Resolve helpdesk tickets, troubleshoot network issues, and support A/V setup for meetings and events. You will own day-to-day support alongside larger initiatives, looking for repeatable patterns that can be automated or improved. User account management: Perform provisioning and deprovisioning as part of onboarding and offboarding workflows. Ensure access is accurate, timely, secure, and aligned to business needs. Infrastructure improv

🔔

Get new senior infrastructure automation engineer jobs by email

Daily job updates · Unsubscribe anytime