Jobiba hiring network

Senior Infrastructure Automation Engineer Jobs

7,101 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior infrastructure automation engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

NVIDIA is seeking a Senior Staff SRE to build and operate reliable, scalable compute platforms that support global engineering workloads. This role spans Kubernetes, KubeVirt, bare-metal infrastructure, automation, observability, and AI-enabled operations. Join a team that solves complex infrastructure challenges, builds durable automation, and improves the reliability and operational experience of critical compute services. What you’ll be doing: Build, operate, and improve large-scale Kubernetes, KubeVirt, Linux, container, and bare-metal compute platforms, with a focus on performance, capacity, reliability, and operational scale. Lead bare-metal provisioning and lifecycle management in data centers, including PXE boot, DHCP, DNS, OS provisioning, hardware validation, and fleet automation. Develop automation, self-service capabilities, and observability solutions using APIs, Python or Go, Infrastructure as Code, configuration management, metrics, logs, traces, and service-health data. Define and operate SLOs, SLIs, error budgets, alerting, and incident-response practices; lead complex incident investigations, corrective actions, and blameless postmortems. Partner with infrastructure, security, hardware, data-center, and application teams to deliver global platform initiatives, and participate in an on-call rotation. What we need to see: BS in Computer Science, Engineering, a related technical field, or equivalent experience, plus 10&#43; years operating production infrastructure or platform services. Strong expertise in Kubernetes administration, KubeVirt, Docker, containerization, microservices, Linux systems, and resolving distributed-system challenges. <l

pythondockerkubernetes
View job →

We at NVIDIA seek an Senior Developer Relations, Automated Synthetic Chemistry Science Lead who can coordinate science streams. This role exists at the crossroads of experimental science, optimization, analytical characterization, AI modeling, laboratory automation, and data infrastructure. It demands a practitioner with a background in research who can lead interdisciplinary science teams and accelerate experimentation without compromising scientific quality. What you'll be doing: Coordinate the operating plan across research priorities, technical execution, and program achievements. Convert research objectives into experimental priorities, parameter-space development, campaign planning, and success criteria. Lead research initiatives encompassing synthetic chemistry, catalysis, process chemistry, machine learning, automated systems, and data processing. Build standardized experimental traces capturing successful and unsuccessful outcomes, metadata, quality-control signals, analytical summaries. Guide AI systems for feasibility assessment, outcome prediction, optimization, scope exploration, and campaign orchestration. Integrate automated experimentation, analytical data streams, data curation, model retraining, and campaign decisions into a closed-loop operating model. Establish science-stream governance: build reviews, decision logs, risk and dependency tracking, quality thresholds, and paths for addressing blocking issues. Prepare recurring workstream readouts for program leadership, including scientific progress, critical decisions, cross-team dependencies, resource needs, and unresolved risks. What we need to see: PhD (or equivalent experience) with 10&#43; years of hands-on experience in experimental science, chemical engineering, material science, robotics, a

machine learningai
View job →
AC
Appian Corporation
📍 Toronto• Full-time• From C$75K/yr
18 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. About the Team Appian Customer Success is obsessed with delivering exceptional customer outcomes and high-velocity business impact. Our Public Sector team acts as a trusted strategic partner to government agencies and organizations, bringing their most vital ideas to life. By joining this high-performing team, you will strengthen and evolve your technical consulting skills while accelerating the adoption of our AI-Powered Process Automation platform across critical public infrastructure. As a Senior Consultant, you will step into a pivotal leadership role at the intersection of public sector strategy and modern technical innovation. This is your opportunity to move beyond isolated coding tasks and own the end-to-end delivery of highly visible, complex enterprise systems. You will lead small engineering teams, architect robust integrations, and serve as a trusted technical advisor to client stakeholders - enabling enterprise organizations to achieve true Enterprise-Grade Orchestration . What You’ll Do Lead Project Delivery: Oversee and drive the entire project lifecycle to define, design, and implement custom automation solutions using the Appian platform. Mentor & Develop Talent: Actively lead, coach, and mentor junior consultants through fast-paced software implementations, fostering a culture of technical excellence. Architect Integrations: Build seamless data pipelines and robust APIs to integrate multiple external enterprise systems using REST, SOAP, and JDBC connections. Drive Data Modeling: Design and launch sophisticated relational data

sqlawsrest
View job →
C
Coinbase
📍 - USA• Full-time• Remote• From $180.4K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Data Scientist on the CX Consumer Analytics team, you will serve as the foundational link between CX operations and top-line financial impact, owning the revenue calibration models, experimentation frameworks, and behavioral intelligence that connect every customer support interaction to Coinbase's asset accumulation flywheel. You will partner closely with CX Analytics Engineers, Program Managers, and Product teams to translate complex operational and behavioral data into defensible, executive-ready insights that drive measurable improvements in retention, product adoption, and automation quality. What you'll do: Own and evolve CX's Downstream Impact of Support (DSI) revenue calibration models, translating support interaction data into quantified revenue signals. Design and execute causal inference frameworks and experiments to measure the incremental impact of CX programs (Concierge, Proactive Outreach, automation interventions) on customer retention and product engagement. Build and maintain LLM-powered classification pipelines for CX contact taxonomy, customer friction detection, and issue attribution, partnering with Analytics Engineers to productionize models into CX's governed Source of Truth infrastructure. Partner with CX Program Managers and Product teams to define segmentation models and behavioral signals that enable personalized experiences an

REMOTEawsaigo
View job →
M
11 days ago

Enterprise Advanced is a distributed team across Europe and India that builds the software running MongoDB on any infrastructure, at global scale — from on-prem data centers to private cloud. You'll work primarily on Ops Manager and Automation, the systems that let customers deploy fault-tolerant, globally distributed MongoDB clusters in minutes. Our software manages some of the largest self-managed MongoDB deployments in the world, with production clusters running hundreds of shards and nodes under a single deployment. The main focus of this team is to adapt our software to manage MongoDB clusters which are deployed in data centers or private cloud platforms. You will work on the core functionality for all of our products, mainly on the Ops Manager , and Automation products. Our team's end users are some of the largest businesses in the world, deploying massive clusters and processing huge amounts of data. This role is based in our Gurgaon office, and can work in a hybrid fashion. This role will report to the Senior Engineering Manager also based in our Gurgaon office. What you’ll do Design, implement, test, and release features for Ops Manager Own end-to-end delivery of complex projects, from design through incremental shipping Troubleshoot and resolve issues surfaced in customer deployments running at scale Apply engineering judgment and MongoDB's core values across planning, design, and code review A great fit for this role will be You enjoy distributed-systems problems; consistency, fault tolerance, and scale are the daily reality, not edge cases People who like ambiguity and are comfortable defining their own approach with guidance, not step-by-step instruction You're flexible! You're willing to take on a wide variety of responsibilities, learning as you go You're a self-starter! You're comfortable organizing your own time, acting on feedback and prioritizing with guidance from senior members of your team Requirements 4+ years experience with a language

javascriptpythonjava
View job →
AC
18 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Are you looking to combine your passion for elite technology with a talent for strategic problem-solving and team leadership? Drive mission-critical modernization where impact matters most. Location: Mexico City, MX | Work Model: Hybrid About the Team Appian Customer Success is obsessed with delivering exceptional customer outcomes and high-velocity business impact. Our Public Sector team acts as a trusted strategic partner to government agencies and organizations, bringing their most vital ideas to life. By joining this high-performing team, you will strengthen and evolve your technical consulting skills while accelerating the adoption of our AI-Powered Process Automation platform across critical public infrastructure. The Opportunity As a Senior Technical Consultant, you will step into a pivotal leadership role at the intersection of public sector strategy and modern technical innovation. This is your opportunity to move beyond isolated coding tasks and own the end-to-end delivery of highly visible, complex enterprise systems. You will lead small engineering teams, architect robust integrations, and serve as a trusted technical advisor to client stakeholders—enabling public sector entities to achieve true Enterprise-Grade Orchestration . What You’ll Do Lead Project Delivery: Oversee and drive the entire software development lifecycle ( SDLC ) to define, design, and implement custom automation solutions using the Appian platform. Mentor & Develop Talent: Actively lead, coach, and mentor junior consultants through fast-paced software imp

sqlawsrest
View job →
G
18 days ago

Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug compl

pythonlinuxai
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering (SRE) team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide. As a Staff Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability. We are seeking Staff SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to proactively find and analyze reliability problems across our stack, then design and implement software and systems to create step-function improvements. You will design robust observability solutions, lead incident response, automate operational tasks, and continuously improve our infrastructure's reliability, all while mentoring and educating the broader engineering team to make reliability a core value at Replit. You Will: Architect and Implement Observability: Design, build, and lead the implementation of comprehensive monitoring, logging, and tracing solutions. Create dashboards and metrics that provide real-time visibility into system health and performance, enabling proactive issue detection. Define and Drive Reliability Standards: Work with product and engineering teams to define, implement, and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to monitor and report on these metrics, holding teams accountable and ensuring we maintain high reliability standards while balancing innovation speed. Lead Incident Management and Response: Act as a senior leader during high-impact incidents, guiding the team to rapid resolution

pythongcpdocker
View job →

As a Sr. Staff Technical Program Manager, you will partner with key Engineering, Product, Product Design, Marketing, and Analytics stakeholders to conduct data-driven experiments and deliver features for MongoDB. As a seasoned program leader, you will be responsible for one of our most mission-critical programs this fiscal year, and own executive level communication related to the program. We are looking to speak to candidates who are based in Dublin or Cork for our hybrid working model, or remote within Ireland. The right candidate for this role will be: Experienced with 15+ years of working in an engineering organization leading complex cross-functional technical programs Experienced with 10 years of Software development background, with Cloud storage and compute products Experience with Service Oriented Architecture and Cloud Infrastructure Able to leverage their knowledge and experience to influence technical discussions, summarizing outcomes and next steps Skilled at communicating across a diverse set of engineers and stakeholders Hyper-organized and capable of coordinating across multiple independent work streams and organizations A role model for effective execution practices, driven by an attuned sense of priority and urgency Able to leverage their experience in program delivery to influence improvements to our tools, operations, and architecture Trained in working with project tracking software (e.g. Jira, Rally, MS Project) Familiar with MongoDB or a comparable technology Interested in business automation work such as scripting in Python, Google Apps Script and Slack Position Expectations: Leverage technical acumen and analytical skills to drive engineering programs forward and maximize business value delivery Recognize patterns in a sea of information and take action accordingly Design, maintain, and improve the processes and tools that power program delivery Build strategic partnerships with Product and Engineering stakeholders Expand knowledge into new

pythonmongodbaws
View job →
O
Okta
📍 Toronto• Full-time• From C$118K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta is The World’s Identity Company. We free everyone to safely use any technology— anywhere, on any device or app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not looking for someone who checks every single box - we’re looking for lifelong learners and people who can make us better with their unique experiences. Join our team! We’re building a world where Identity belongs to you. Position Description: As a Senior Research Operations Program Manager at Okta, you will play a crucial role in building and enabling an impactful beta program. Via close collaboration across research, design, product, engineering teams, QA, and release teams, you will programmatically manage cross-functional teams’ research needs. You will proactively build and manage the operations for a beta program, collaborating with other cross functional functions to plan the necessary infrastructure for this programming to be successful, following compliance and governance best practices. In this role, you’ll get to: Lead research operations and panel management for beta programming Work collaboratively across multiple teams to empower Okta staff to carry out impactful research Help manage the general pipeline o

awsgitrest
View job →
O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Product Manager, Data Platform (India) Company Overview Okta is The World’s Identity Company. We free everyone to safely use any technology—anywhere, on any device, for any app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. Role Overview: Senior Product Manager, Data Platform As the Architect of Okta’s data-driven future, the Senior Product Manager for Data Platform orchestrates the evolution of our core data infrastructure, pipelines, and analytics capabilities. This role manages the critical tension between high-performance technical execution and the delivery of actionable business intelligence, translating high-volume identity event streams into strategic outcomes. You will serve as the essential bridge between our India-based engineering teams and a global suite of stakeholders, ensuring Okta's data remains a competitive differentiator and a cornerstone of trust for our customers. Key Responsibilities Orchestrate the long-term product vision and strategy for Okta’s data platform, ensuring infrastructure and analytics roadmaps align with the scale of the world’s identity traffic. Architect a robust API strategy and developer tooling framework to ensure data is seamlessly and securely consumed by services across the entire Okta ecosystem. Collaborate with India-base

awsgitrest
View job →
FB
18 days ago

Location : Barcelona, Berlin or Hamburg Join our innovative team as a Senior CRM Tech Manager, where you will drive the implementation, maintenance, and optimization of our CRM platform across Rider, Driver and B2B in 9 markets. As a key member of the CRM team, reporting directly to the Head of CRM, you’ll collaborate across marketing, operations and data functions to ideate and deliver high-quality, data-driven marketing automations and other solutions that enhance customer engagement and support our business growth. YOUR DAILY ADVENTURES WILL INCLUDE: Drive the technical design, implementation, optimization, and scalability of CRM systems. Collaborate with marketing, product and engineering teams, ensuring technical CRM solutions align with business goals. Develop and maintain our CRM infrastructure, including automating workflows, integrations, and personalization that drive efficiency. Ensure data quality, integrity, and compliance through audits and analysis. Troubleshoot CRM performance issues and optimize system functionality. Analyze CRM data to generate insights and recommendations. Manage technical vendors and maintain system documentation. Develop in-app messages and email templates using HTML, CSS, and JavaScript. Write SQL and Python scripts for segmentation and ETL processes. Our MarTech stack : Braze | Tableau | Mixpanel & more TO BE SUCCESSFUL IN THIS ROLE: Bachelor's degree in Information Technology, Computer Science, Business, or a related field. Minimum 3+ years of multi-country CRM experience, ideally in ride-hailing, mobility, or food delivery. Self-learner with hands-on experience configuring, customizing and optimizing CRM platforms, particularly Braze and for mobile applications. Proficiency in SQL, HTML, CSS, and JavaScript for CRM automation and customization. Strong analytical and problem-solving skills with a data-driven mindset. Experience with API integrations and understanding of data structures and microservices. Ability to c

javascriptpythonjava
View job →
FB
18 days ago

Location : Barcelona, Berlin or Hamburg Join our innovative team as a Senior CRM Tech Manager, where you will drive the implementation, maintenance, and optimization of our CRM platform across Rider, Driver and B2B in 9 markets. As a key member of the CRM team, reporting directly to the Head of CRM, you’ll collaborate across marketing, operations and data functions to ideate and deliver high-quality, data-driven marketing automations and other solutions that enhance customer engagement and support our business growth. YOUR DAILY ADVENTURES WILL INCLUDE: Drive the technical design, implementation, optimization, and scalability of CRM systems. Collaborate with marketing, product and engineering teams, ensuring technical CRM solutions align with business goals. Develop and maintain our CRM infrastructure, including automating workflows, integrations, and personalization that drive efficiency. Ensure data quality, integrity, and compliance through audits and analysis. Troubleshoot CRM performance issues and optimize system functionality. Analyze CRM data to generate insights and recommendations. Manage technical vendors and maintain system documentation. Develop in-app messages and email templates using HTML, CSS, and JavaScript. Write SQL and Python scripts for segmentation and ETL processes. Our MarTech stack : Braze | Tableau | Mixpanel & more TO BE SUCCESSFUL IN THIS ROLE: Bachelor's degree in Information Technology, Computer Science, Business, or a related field. Minimum 3+ years of multi-country CRM experience, ideally in ride-hailing, mobility, or food delivery. Self-learner with hands-on experience configuring, customizing and optimizing CRM platforms, particularly Braze and for mobile applications. Proficiency in SQL, HTML, CSS, and JavaScript for CRM automation and customization. Strong analytical and problem-solving skills with a data-driven mindset. Experience with API integrations and understanding of data structures and microservices. Ability to c

javascriptpythonjava
View job →
RS
18 days ago

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leading orchestration platform for the autonomous enterprise, driving business transformation at the lowest total cost of ownership. Redwood empowers organizations to intelligently automate and orchestrate mission-critical business and IT processes across complex ERP, hybrid cloud, data and emerging agentic AI systems. Through its SaaS-first automation fabric—with AI embedded across the automation lifecycle—Redwood accelerates the path to autonomous operations. Backed by 30 years of experience and trusted by more than 50% of the Fortune 50, Redwood helps organizations unlock human potential to focus on innovation, growth and what’s next. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for a Software Engineer, Platform & Integrations . Working closely with senior and lead engineers, you will design, develop, and maintain high-quality features that power enterprise data exchange for more than 1,000 customers worldwide. This is an incredible opportunity to deepen your expertise in cloud-native architectures, enterprise security, and modern DevOps practices in a fast-growing product environment. Feature Development & Design: Write clean, maintainable, and well-tested code using Java and Spring Boot to deliver scalable backend services and microservices. Platform Reliability: Contribute to enhancing the monitoring, logging, and observability of our core platform to ensure high availability and performance. Security & Compliance: Implement secure coding practices to safeguard data exchange and maintain compliance across our cloud infrastructure. Collaborative Execution: Work within an agile team, collaborating closely with QA, Product, and fellow engineers to deliver high-qual

javasqlpostgresql
View job →
G
18 days ago

Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug compl

pythonlinuxai
View job →
🔔

Get new senior infrastructure automation engineer jobs by email

Daily job updates · Unsubscribe anytime