Jobiba hiring network

Senior Infrastructure Engineer Jobs

7,101 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior infrastructure engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Mainframe Capacity & Performance Engineer to join our Enterprise Infrastructure organization. The Senior Mainframe Capacity & Performance Engineer will serve as a critical technical leader responsible for ensuring the performance, scalability, reliability, and efficiency of our enterprise mainframe environment supporting mission-critical healthcare, pharmacy, and retail applications. As a Senior Mainframe Capacity & Performance Engineer, you will play a key role in capacity planning, workload analysis, performance engineering, and infrastructure optimization across one of the nation's largest and most complex mainframe ecosystems. This position is responsible for proactively monitoring shared mainframe resources, evaluating system utilization trends, identifying performance risks, and providing actionable recommendations to improve overall system health and operational efficiency. The Senior Mainframe Capacity & Performance Engineer will partner closely with Application Development, Mainframe Systems Programming, Infrastructure Engineering, Architecture, Database Administration, Operations, and Business teams to analyze workload behavior, assess resource consumption, identify top consumers, and optimize application performance. This role requires deep expertise in z/OS performance analysis, capacity forecasting, workload managem

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As part of Micron's Technology Engineering & Innovation (TE&I) organization, you will have the opportunity to shape the future of our global infrastructure platforms while enabling business growth, operational resilience, and digital transformation at scale. The Opportunity Micron is seeking a transformational Senior Director of Technology Engineering & Innovation (TE&I) to lead the strategy, engineering, operations, and modernization of our global infrastructure ecosystem. This role is responsible for defining and executing Micron's vision across enterprise networks, cloud platforms, data center strategy, database services, infrastructure engineering, automation, observability, and global infrastructure operations. As a key member of the TE&I leadership team, you will drive innovation, operational excellence, and strategic transformation while building a high-performing organization focused on business outcomes and exceptional customer experiences. The successful candidate will be equally comfortable developing multi-year technology strategies, leading large-scale infrastructure transformations, driving operational performance, developing talent, and fostering a culture of collaboration, accountability, and continuous improvement. What You Will Do Lead Global Infrastructure Strategy & Transformation Define and execute Micron's global infrastructure vision and strategy. Develop multi-year roadmaps for: Enterprise Network Services <l

airecruitment
View job →
N
Notion
📍 Hyderabad• Full-time
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role The Hyderabad Infra team builds and maintains Notion's internal async task runner and configuration management platform. The async task runner plays a critical role in ensuring our millions of users have a fast, reliable, and secure experience. The configuration management platform enables safe, explicit configuration management for Notion's product and infrastructure engineers. As part of the Hyderabad Infra Team, you’ll have a unique opportunity to shape how Notion manages and scales its async task runner and configuration management platform, enabling innovation across the company. What You'll Achieve You will contribute to the evolution and maintenance of our async task runner to meet the needs of over 100 million global users and support the rapid growth of our product and business. With guidance from senior team members, you'll help ensure our systems remain reliable, efficient, and scalable. You'll evaluate and integrate new technologies to keep us ahead of emerging challenges. Your work will empower our engineering team to build features confidently while you grow your skills in distributed systems and infrastr

restaigo
View job →
DU
17 days ago

About the Team Our team exists to empower DoorDash Mobile engineers. We are the champions of three core pillars: Quality, Velocity, and Efficiency. Ultimately, our goal is to build the supportive infrastructure that allows our fellow engineers to build, ship, and operate apps at massive scale—with confidence and ease. We believe that a great developer experience leads to a great customer experience. About the Role We are looking for an Engineering Manager to lead our Mobile Developer Experience team (also known internally as the Mobile Foundations team). In this role, you’ll not only set the technical vision but also nurture the culture required to build world-class mobile applications. Your team will concentrate on one vertical—ZeroKit, which enables rapid mobile prototyping via agents—and two horizontal infrastructure pillars: the iOS and Android monorepos. You will partner with senior engineers and stakeholders across the company to design systems that make our platform faster, more reliable, and more efficient. You will be a partner to your customers—your fellow engineers—working side-by-side to understand their hurdles and solve their immediate challenges. You’ll also look to the future, anticipating needs so we can deliver solutions before they become blockers. Crucially, this role will support our growing international engineering footprint and unified technology stack, meaning you and your team will work across our DoorDash, Wolt, and Deliveroo brands. This is a unique opportunity to lead a team through a mix of exciting greenfield initiatives and the refinement of established, successful tools. You must be located in the following locations for this opportunity: San Francisco, CA; Sunnyvale, CA; Seattle, WA; Los Angeles, CA; New York, New York. You’re excited about this opportunity because you will… Develop and maintain foundational components to enable DoorDash Engineers to excel at mobile engineering. Lead the development and strategy for ZeroKit to enabl

awsgitrest
View job →
A
Asana
📍 Warsaw• Full-time• $249.6K – $275.8K/yr
1mo ago

Asana is looking for an experienced Audio Visual Engineer to join our IT Team and help manage AV and event operations across our worldwide offices. Based in our Warsaw office, you will primarily support our conference rooms throughout the EMEA region and oversee team/regional AV events. This role will be focused on keeping 100+ Zoom Rooms throughout Europe and APAC fully operational, while being available to assist in customer-facing events, Team All Hands, Regional syncs, and traveling for senior leadership AV needs. You ask questions and confidently admit when you don't have all the answers. You enjoy collaborating with others rather than working in isolation, but can work independently when necessary. While you are precise and detail-oriented, you can also think creatively and push boundaries when necessary. In addition to your proven success, you have a passion for technology and share our belief that hard work and fun are equally important. This role is based in our Warsaw office with an office-centric schedule. The current expectation is to be in the office 5 days a week for this position. Many Asana employees have the option of 1 or 2 work-from-home days a week, but these are critical times for testing, maintenance, and repairs. If you're interviewing for this role, your recruiter will share more about the in-office requirements. What you’ll achieve Collaborate with stakeholders to ensure specific needs for event technology. Perform daily checks of conference rooms and presentation spaces to ensure full functionality. Operate a wide range of A/V equipment ( Zoom, Crestron, QSYS, Logitech, Neat) Plan and support events ranging from 300+ person internal/external events to small team meetings. Edit videos using Adobe Creative Suite and assist in maintaining a living video library. Troubleshoot A/V and network-related issues affecting event setups and performance. Maintain detailed technical documentation, including runbooks and troubleshooting guides

gitrestai
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
17 days ago

Opportunity Overview: This is a unique opportunity to join a high-caliber software engineering team that is growing quickly. You will play a key role in building impactful healthcare technology on a modern technology stack, with a focus on our core data and AI platforms. Your work will focus on enhancing the platform's key features, while also balancing scalability, reusability, and performance. Role Overview: We're looking for a Staff Platform Engineer to serve as the technical backbone of our Engineering organization. You'll own the technical strategy, and delivery of our platform — spanning architecture, DevOps, SRE, security, Dev-ex. This is a hands-on staff level role: you'll set technical direction, drive cross-team alignment, and be the senior escalation point for platform challenges. What you’ll do: Drive platform reliability, scalability, security, and cost efficiency across all environments. Technical Leadership: Provide technical leadership for platform components, Influence the technical strategy and architecture of our cloud platform, from CI/CD pipelines to observability and incident response. Design and implement platform components and reusable integration patterns that minimize custom development efforts, reduce the time spent on repetitive tasks, and ensure that integrations scale across multiple healthcare systems Partner closely with Architecture, DevOps, SRE, and Security teams to deliver cohesive platform solutions Cross-Functional Collaboration: Work closely with product teams, and solutions architects to understand integration needs and ensure the platform meets current and future business requirements. Serve as a senior escalation point for infrastructure and platform incidents Establish frameworks for: AI governance and compliance. Observability of systems. Traceability of decisions and outputs. Ensure enterprise readiness with security, auditability, and reliability in production environments. Security & Compliance : Ensure all p

awsci/cdgit
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Infrastructure Engineering function sits within IT and is responsible for reliably building, deploying, and operating critical on prem and hybrid environments that power internal services and critical R&D environments. This is an early, high-leverage technical role focused on applying strong Site Reliability Engineering discipline to environments where uptime, safety, recoverability, and security are non-negotiable. This person helps replace bespoke, one-off infrastructure with standardized infrastructure-as-code building blocks that compound reliability and operational leverage as OpenAI scales. About the Role We are looking for an experienced Site Reliability Engineer working on security infrastructure to design, build, and operate reliable, secure, and scalable infrastructure that underpins identity, access, endpoint, and shared platform services across the company. In this role, you will be a senior technical owner for infrastructure and identity systems end to end, from architecture and implementation through policy enforcement, upgrades, recovery, and day-two operations. You will build durable, production-grade platforms that remove operational friction, enforce security by default, and enable teams to move faster with confidence. This role is well suited for a hands-on senior engineer who thrives in ambiguity, enjoys owning complex systems end to end, and raises the reliability and security bar by replacing fragile implementations with standardized, repeatable infrastructure. This role is based in our San Francisco HQ and requires in-office presence. In this role, you will: Design, build, and operate reliable infrastructure across on-prem, hybrid, shared, and product adjacent environments. Establish standardized infrastructure patterns that replace bespoke implementations with repeatable, auditable, secure-by-default systems. Own the lifecycle of critical infrastructure platforms, including provisioning, deployment, upgrades, patching,

awsazurerest
View job →
N
1mo ago

NVIDIA is seeking a Senior Technical Program Manager to join the CSP Engagements team, focused on deep technical engagement with hyperscale cloud service providers for NVIDIA’s next‑generation datacenter systems such as Vera Rubin NVL72. This role is intended for experienced systems and embedded software leaders—including software engineering managers, technical leads, or senior architects—who have led datacenter server and platform software programs and can operate as a trusted technical partner to hyperscale CSP engineering teams. As a member of the CSP Engagements team, you will act as the primary technical engagement leader between NVIDIA’s system software organizations and CSP platform, system software, and AI teams, ensuring alignment, readiness, and successful large‑scale deployment of NVIDIA‑based datacenter solutions. What you will be doing: Lead deep technical engagements with hyperscale CSPs as the primary NVIDIA point of contact for system software, firmware, and platform readiness for NVIDIA datacenter products. Partner directly with CSP system software, firmware, and infrastructure engineering leaders to align on software architecture, bring‑up plans, deployment readiness, and production requirements for NVIDIA‑based server and rack‑scale platforms. Represent CSP technical priorities internally, advocating for customer requirements and tradeoffs across NVIDIA’s system software, firmware, hardware, silicon, and product teams are aligned to customer needs, timelines, and constraints. Own the end‑to‑end CSP engagement lifecycle, from early technical alignment and pre‑production readiness through large‑scale deployment, escalation management, and sustained production support. Drive bi‑directional technical communication: translating CSP system‑level requirements into actionable focus areas for NVIDIA engineering teams, while clearly communicating N

linuxartificial intelligenceai
View job →
I
Instacart
📍 United States - Remote• Full-time• Remote• From $265K/yr
1mo ago

We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Overview Instacarts Data Infrastructure organization builds and operates the systems that power our company’s data ecosystem, including a modern open data lakehouse on Apache Iceberg, a multi-engine compute platform for stream and analytical workloads, and self-serve tooling that helps Product, Data Science, ML, Ads, Finance, and engineering teams move fast with data. We’re looking for a Staff Software Engineer, Data Infrastructure to join our Data Governance and Foundations Team. In this role, you’ll serve as a senior technical leader owning the architecture and delivery of our open lakehouse foundation, governance and access patterns, and multi-engine compute strategy—balancing today’s reliability with the next three to five years of scale, maturity, and cost efficiency. You’ll collaborate closely with engineering leadership and stakeholders across Data Science,

REMOTEpythonsqlaws
View job →
C
11 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a highly skilled Staff Data Engineer, Observability Engineering to join the Enterprise Observability Platform organization and help advance the next generation of observability, infrastructure, and security data capabilities. The Staff Data Engineer, Observability Engineering will play a critical role in designing, building, and operating scalable data pipelines and data products that power enterprise observability, operational intelligence, and security analytics across the organization. The Staff Data Engineer, Observability Engineering is a senior individual contributor responsible for developing and optimizing Databricks-based data engineering solutions that ingest, transform, govern, and deliver high-volume telemetry, infrastructure, application, and security data. This role combines deep hands-on technical execution with ownership of engineering excellence, operational reliability, performance optimization, and data platform best practices. Working closely with Observability Engineering, Security Engineering, Infrastructure Engineering, and Data Platform teams, the Staff Data Engineer, Observability Engineering will contribute to the evolution of the enterprise observability lakehouse by building resilient ingestion frameworks, establishing data quality standards, enhancing governance controls, and driving efficient, scalable data processing patterns. The id

REMOTEpythonsqlazure
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering (SRE) team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide. As a Staff Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability. We are seeking Staff SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to proactively find and analyze reliability problems across our stack, then design and implement software and systems to create step-function improvements. You will design robust observability solutions, lead incident response, automate operational tasks, and continuously improve our infrastructure's reliability, all while mentoring and educating the broader engineering team to make reliability a core value at Replit. You Will: Architect and Implement Observability: Design, build, and lead the implementation of comprehensive monitoring, logging, and tracing solutions. Create dashboards and metrics that provide real-time visibility into system health and performance, enabling proactive issue detection. Define and Drive Reliability Standards: Work with product and engineering teams to define, implement, and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to monitor and report on these metrics, holding teams accountable and ensuring we maintain high reliability standards while balancing innovation speed. Lead Incident Management and Response: Act as a senior leader during high-impact incidents, guiding the team to rapid resolution

pythongcpdocker
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Core Services organization builds and runs the mission-critical online services that product teams rely on in production. We own foundational distributed systems and platform capabilities that enable reliable execution, high-performance services, and large-scale file/data needs across our products. This team is distinct from developer infrastructure and data infrastructure—our focus is production service foundations and core runtime services. About the Role We’re hiring an Engineering Manager, Core Services to help lead teams responsible for highly reliable, high-scale distributed systems that sit on the critical path for OpenAI products. Your team will own foundational production systems that OpenAI’s product engineering teams build on. You’ll collaborate closely with product and infrastructure partners to ship reliable services quickly, and help scale systems and teams as OpenAI grows. You’ll partner closely with senior engineering leaders to scale the org, mature operations, and drive major platform initiatives. This role requires strong technical ability. You’ll be responsible for: Managing and growing a high-performing team of infrastructure engineers. Leading teams building and operating large, critical production platforms, including cluster reliability, scaling, and rollout safety. Building and operating mission-critical distributed systems with strong operational rigor (SLOs, incident response, capacity planning, reliability). Setting technical direction for platform foundations such as workflow/orchestration capabilities, large-scale file/blob/storage services, and core service foundations. Partnering with a broad set of stakeholders, including product engineering, adjacent infrastructure teams, and (where relevant) finance/cost partners. Coaching, mentoring, and developing engineers and emerging leaders. You might thrive in this role if you: Have significant experience leading teams that run mission-critical infrastructure in production

awsrestai
View job →

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Software Engineers at Palantir build software at scale to transform how organizations use data. Our Software Engineers are involved throughout the product lifecycle, from idea generation, design, prototyping, and production delivery. You will collaborate closely with technical and non-technical teammates to understand our customers' problems and build products that solve them. We encourage movement across teams to share context, skills, and experience, so you'll learn about many different technologies and aspects of each product. Engineers work autonomously and make decisions independently, within a community that will support and challenge you as you grow and develop, becoming a strong technical contributor and engineering leader. Our Product Development organization is made up of small teams of Software Engineers. Each team focuses on a specific aspect of a product. Our infrastructure teams are responsible for the lowest layers of our software stack, often focused on database technologies, distributed systems, large scale data systems, security, and application infrastructure. As a Software Engineer on infrastructure working on our Foundry platform, you'll contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks, enabling products deployed to the most important institutions in the public and private sector. You'll build the foundational capabilities that power our products used by research scientists, aerospace engineers, intelligence analysts, and economic forecasters, in countries around the world. We’re hiring engineers who are passionate about solving real-world problems and empowerin

PE
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Software Engineers at Palantir build software at scale to transform how organizations use data. Our Software Engineers are involved throughout the product lifecycle, from idea generation, design, prototyping, and production delivery. You will collaborate closely with technical and non-technical teammates to understand our customers' problems and build products that solve them. We encourage movement across teams to share context, skills, and experience, so you'll learn about many different technologies and aspects of each product. Engineers work autonomously and make decisions independently, within a community that will support and challenge you as you grow and develop, becoming a strong technical contributor and engineering leader. Our Product Development organization is made up of small teams of Software Engineers. Each team focuses on a specific aspect of a product. Our infrastructure teams are responsible for the lowest layers of our software stack, often focused on database technologies, distributed systems, large scale data systems, security, and application infrastructure. As a Software Engineer on infrastructure working on our Foundry platform, you'll contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks, enabling products deployed to the most important institutions in the public and private sector. You'll build the foundational capabilities that power our products used by research scientists, aerospace engineers, intelligence analysts, and economic forecasters, in countries around the world. We’re hiring engineers who are passionate about solving real-world problems and empowerin

R
Roblox
📍 San Mateo• Full-time• From $233.6K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Security Software Engineer for Infrastructure Security you will be a part of the Information Security organization and report to the Senior Manager of Infrastructure Security. You will help shape the future of Platform Security at Roblox. We work closely with Production IAM, Network Security, and Cloud Security at Roblox. You Will: Identify security gaps and threats in our cloud and on premise infrastructure, partnering with Governance Risk and Compliance teams to create standards and policies along the way. This will help Roblox meet regulatory and compliance requirements. Harden our infrastructure by introducing secure by default configurations, designs and guardrails for all developers at Roblox. Own and drive solutions that enable Roblox engineers to design, build, and use infrastructure securely at scale. Work closely with other InfoSec teams (AppSec, D&R, GRC, CorpSec, CloudSec, NetSec) and partner with engineering teams across Roblox, specifically the Infrastructure organization, to ensure the secure outcomes of security and product driven initiatives. You Have: 5+ years of experience writing code and/or relevant technical experience. Experience with

awskubernetesgit
View job →
🔔

Get new senior infrastructure engineer jobs by email

Daily job updates · Unsubscribe anytime