Jobiba hiring network

Infrastructure Team Manager Jobs

4,730 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current infrastructure team manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

D
Datadog
📍 New York• Full-time• From $99K/yr
1mo ago

Datadog's Finance team collaborates with teams across the organization, providing commercial, operational and analytical support to ensure that Datadog's business continues to scale rapidly and efficiently. The Financial Planning & Analysis (FP&A) team analyzes company financial data (revenue, customers, headcount, expenses, etc.) in order to support the business’ growth and success. As an analyst supporting the team, you will play a key role in delivering insights through the management of essential data infrastructure, including our financial planning tool, Pigment. Your role will be highly cross-functional, leveraging systems and data to unlock analytical capabilities for both FP&A and business leaders. Your role is critical in synthesizing information from across the organization to foster operational alignment and support informed strategic decisions. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the team’s forecasting and reporting software, Pigment, supporting data-driven insights through the development of dashboards and KPIs, both for standard FP&A reports and ad hoc projects Work cross-functionally with FP&A leaders to improve existing datasets and models Ensure data and system best practices in processes across the organization, including during planning and reporting cycles Represent FP&A in the data & analytics community, collaborating with analytics partners across the organization to democratize data and share insights Work on strategic projects and initiatives for senior management, assessing various business opportunities and proposing solutions Support Datadog’s data-based decision making and continued efficient growth Who You Are: 2+ years of professional experience in FP&A, Data Analy

pythonsqlrest
View job →
L
Lyft
📍 Nashville• Full-time
1mo ago

About Flexdrive At Flexdrive, we're at the forefront of revolutionizing transportation by building the operational backbone for autonomous vehicle (AV) fleets. As a leader in fleet management, we're leveraging our expertise to enter the AV space, forming strategic partnerships with cutting-edge technology providers. We're looking for dedicated team members to help us pioneer this new chapter, starting with our first AV depot in Nashville, Tennessee. The Opportunity Reporting to the Service Lead, Flexdrive is seeking a decisive and security-conscious Repair Coordinator to own the complete repair and parts lifecycle for our 24/7 AV operations — from the initial intake of a vehicle requiring service, through parts procurement and consumption, to the vehicle's final release back into the active fleet. This role is the central hub for service workflow, technician support, and inventory infrastructure: coordinating work assignments, comprehensive vehicle information, and correct parts availability to minimize vehicle downtime and uphold the high availability standards required for an autonomous fleet. The Repair Coordinator directly impacts the operational efficiency, inventory accuracy, and reliability of Flexdrive's AV deployment. This position offers unique exposure to cutting-edge AV technology and supply chain management in a highly regulated environment. The role requires someone who can balance operational efficiency with uncompromising security standards, provide 24/7 support coverage, and maintain perfect inventory accuracy. If you excel at detailed inventory management, shop-flow coordination, and thrive in mission-critical operations, we encourage you to apply. Work Schedule & Shift Availability Day Window Shift Pattern This role supports 24/7 AV depot operations and requires availability to work day window shifts, with hours scheduled between 6:30am - 9:30pm. Specific shift times and schedules will be determined based on operational needs and business requ

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team To advance Stripe’s important mission, we are building a world class Internal Audit (IA) team. Our purpose is to strengthen the organization’s ability to create, protect, and sustain value by providing the board and management with independent, risk-based, and objective assurance, advice, insight, and foresight. We are consumed with the goal of moving as fast as the business, being powered by technology, and increasing the maturity of Stripe’s controls where it matters the most. Our IA team is responsible for providing objective assurance on the design and operational effectiveness of Stripe’s internal controls and business processes, its compliance with laws and regulations, its risk management framework, and other governance processes. Currently Stripe IA supports various APAC licenced entities including Singapore, India, Australia, New Zealand, Thailand, Malaysia, Indonesia, and Japan. We’re looking for a candidate with deep finance, operations and regulatory compliance audit experience who will help us build and scale a global audit program and its relevance in the region. What you’ll do As a member of the APAC IA team you will support our audit landscape in one of the most dynamic sectors of FinTech. In this role, you will assist in the execution of a comprehensive, risk-based internal audit strategy that anticipates emerging risks and aligns with our management's vision and regulatory landscape. As a member of the global IA team,

S
1mo ago

Staff Engineer, Revenue and Financial Management Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe’s Revenue & Financial Management group is building new products that expand the scope of problems we tackle beyond payments into Revenue Management, Financial Operations and analytics. Right now this includes products like Billing, Invoicing, Revenue Recognition, Reconciliation and the underlying platform for batch and real time processing of large scale financial data to power these solutions. These solutions are going to be key pillars for Stripe’s growing SaaS business and a major revenue stream. What you’ll do We’re looking for a Staff Engineer that will help architect and design this system from ground up. You will need to set the technical direction across a variety of projects and initiatives while also mentoring and growing others on the team. Responsibilities Scope and lead large technical projects that are the foundational pillars for Financial Data Management Infrastructure Scrutinize and reason clearly about the technology and architecture choices we make in building these products. In many cases, you will be the decider of these decisions Directly contribute to core interface design and write code. Serve as a role model for how great software should be written for Stripe as a whole Arbitrate critical decisions correctly that fully consider software best practices, Stripe system realities, and numerous stakeholders’ preferences and concerns Advise Stripe’s leadership tea

restmicroservicesai
View job →

About the Team OpenAI’s GTM Partnerships team builds a strategic global partner ecosystem designed to accelerate customer success, support secure AI adoption, and drive growth in support of OpenAI’s mission. We collaborate closely across Sales, Technical Success, Product, Marketing, Operations, and leadership teams to turn strategic partnerships into durable commercial outcomes. Within that ecosystem, AWS is one of OpenAI’s most strategic partners. Our opportunity extends beyond the direct alliance relationship: we aim to enable AWS teams and the surrounding services-partner ecosystem to build, deliver, and scale OpenAI-powered solutions on AWS infrastructure. Success requires close collaboration across AWS Japan alliance and field teams, AWS-focused SI practices, cloud architects, field sellers, delivery leaders, OpenAI’s Japan Partner Director, and OpenAI GTM teams. It requires strong operating discipline, scalable co-sell execution, and the ability to translate a strategic alliance into measurable customer outcomes. About the Role We are hiring a Partner Director to shape and execute OpenAI’s AWS alliance partnership in Japan. This person will own the direct relationship with AWS Japan while mobilizing AWS Japan field teams and the broader services-partner ecosystem around high-value customer opportunities. This is a hands-on individual-contributor role operating through influence in Japan. You will work alongside OpenAI’s Japan Partner Director to develop Japan GTM strategies, drive joint account planning, coordinate AWS-related opportunities, and operationally manage selected cross-functional initiatives. You will build and own the Japan GTM strategy with AWS, translating the alliance into repeatable OpenAI and AWS solutions, field playbooks, governance structures, and co-sell motions that can be adopted across Japan’s customer segments and GTM teams. The right candidate combines alliance leadership with a builder’s mindset and the ability to move comfortably b

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team Our Cyber team builds AI systems and products that help trusted defenders understand and respond to cyber threats while improving the safety and reliability of frontier models in security-sensitive settings. The team works across product engineering, model training, evaluations, safeguards, and deployment to make advanced cyber capabilities useful to defenders and responsibly managed. We collaborate closely with Safety/Preparedness, Research, Security, Legal, Communications, GTM, and external partners across OpenAI’s broader cyber work. About the Role We’re looking for research and software engineers to join Codex Cyber. You’ll help define and ship security products, work with trusted defenders and customers, shape model training and access patterns, and build research and evaluation systems for assessing cyber capabilities, validating safeguards, and improving training data. This role is hands-on and cross-functional, connecting product launches, model development, safety work, and real-world security use cases. In this role, you will: Help define and execute the technical roadmap for Codex Cyber’s security products, including evaluations, safeguards, trusted-defender workflows, and deployment decisions. Work with trusted defenders, customers, and partner teams to understand cyber use cases, evaluate risk, and turn feedback into product and research priorities. Shape cyber-specific model training and access patterns, including data, evaluations, validation, and deployment criteria. Build and validate systems for measuring cyber capabilities, monitoring misuse risk, and proving safeguards work in practice. Collaborate with Safety/Preparedness, Research, Security, Legal, Communications, Go-to-Market, and external partners on company-wide cyber priorities. Translate frontier cyber research into launch-ready tools, operational playbooks, and durable infrastructure for Codex and security products. You might thrive in this role if you: Enjoy 0 -> 1 envi

javascripttypescriptpython
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Coding team is reimagining how software is built in the AI era. We build tools and workflows that help software engineers work faster, tackle more ambitious projects, and spend less time on repetitive tasks. AI has already transformed how code is written, but software engineering extends far beyond coding. Our mission is to apply AI across the entire software development lifecycle (SDLC) — from design and implementation to code review, testing, debugging, issue remediation, maintenance, documentation, and user support. The team is also responsible for developer-facing Codex experiences including the Codex IDE Extension and the terminal interface, which are used daily by developers ranging from individual open-source contributors to some of the world’s largest engineering organizations. The team also works closely with the open-source software community, building tools that help maintainers and contributors manage increasingly complex projects. We believe AI can make open-source development more sustainable by reducing the operational burden of reviewing contributions, triaging issues, maintaining quality, and supporting growing communities. By building the future of software development, we're helping advance OpenAI's mission of ensuring that the benefits of AI reach people around the world. About the Role We’re hiring a Full Stack Software Engineer to help invent the next generation of AI-powered software development workflows. “Full stack” in this role means much more than traditional frontend and backend development. You'll own complete product experiences, spanning user interfaces, workflow orchestration, agent and prompt design, backend systems, and cloud infrastructure. This is a highly product-oriented role. You'll work directly on the workflows developers use every day, identifying bottlenecks and rethinking how software gets built in a world where AI agents are active participants in the development process. The features you ship will inf

typescriptawsrest
View job →
O
OpenAI
📍 Seattle• Full-time• $293K – $325K/yr
1mo ago

About the Team The Statsig team at OpenAI builds and operates the experimentation platform that powers product development, measurement, and decision-making across the company. We partner closely with product, engineering, and infrastructure teams to ensure experiments are trustworthy, statistically rigorous, and scalable to the needs of frontier AI products. Our mission is to help teams make better decisions through reliable experimentation. We care deeply about statistical correctness, pragmatic solutions, and building systems that researchers and engineers can trust at massive scale. The team operates at the intersection of experimentation methodology, data infrastructure, causal inference, and product analytics. We are looking for experienced experimentation experts who want to shape the future of experimentation in the AI era. About the role: We're seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you're passionate about working with data and are eager to create solutions with significant impact, we'd love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI! In this role, you will: Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse. Develop canonical datasets to track key product metrics including user growth, engagement, and revenue. Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions. Implement ro

pythonjavavue
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Hardware Health and Observability team owns the end-to-end health lifecycle of OpenAI’s global compute fleet. Our mission is to maximize healthy, usable compute across accelerator vendors, generations, cloud providers, and regions through reliable health signals, automated remediation, and scalable operational tooling. We build the systems that observe, detect, remediate, and verify hardware issues across GPUs, CPUs, networking, and platform infrastructure, enabling frontier model training and inference workloads to run reliably at hyperscale. We are the last line of defense for the success of OAI’s production and research workloads. About the Role On the Hardware Health and Observability team, you’ll build critical infrastructure that keeps OpenAI’s largest compute clusters healthy and operational at scale. Even small numbers of unhealthy systems can impact large-scale training and inference workloads. This team focuses on minimizing downtime, improving fleet efficiency, and ensuring compute resources remain continuously available to researchers and product teams. Engineers on this team own problems end-to-end, from defining health signals and debugging failures to building automated remediation systems that operate across millions of GPUs globally. In this role, you will: Define and maintain health signals across GPUs, CPUs, networking, and platform infrastructure. Build and evolve health checks that detect, remediate, and verify failures at scale. Ensure critical health checks execute with minimal latency to maximize workload uptime. Investigate hardware failures and system-level issues across large-scale compute environments. Own node lifecycle workflows including drain, quarantine, repair, RMA, and return-to-service processes. Build automation and tooling that enables global cluster management with minimal manual intervention. Partner with workload, reliability, and provider teams to integrate health signals into training and inference system

pythonsqlaws
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI Finance ensures the organization is positioned for long-term success as we pursue our mission. The Strategic Sourcing & Procurement function plays a critical role in enabling OpenAI to deliver impact across research, product development, technology infrastructure, and services by helping the company scale responsibly, securely, and with strong commercial discipline. Our work sits at the intersection of innovation and execution. We partner closely with teams across OpenAI to translate rapidly evolving business needs into scalable, compliant, and economically sound external partnerships. As OpenAI continues to grow at pace, services sourcing is becoming increasingly strategic across the company. Every business unit relies on external service providers in different ways — to extend capacity, access specialized expertise, support operations, and accelerate execution. Done well, Procurement becomes a source of trust and momentum, helping OpenAI move faster with the right partners, stronger commercial outcomes, and the right level of protection. About the Role We are seeking an experienced Strategic Sourcing (GTM) Leader to lead strategic sourcing and commercial enablement for OpenAI’s Go-to-Market organization across B2B and B2C channels. You will manage substantial and rapidly growing spend while shaping sourcing strategies and scalable commercial pathways across Media, Creative, Production, Influencer, Agency, Sponsorships, Analytics, Communications, and Event suppliers in support of high-impact global initiatives. You’ll help evolve our GTM procurement function from reactive deal support into a speed-enabling, scalable commercial engine that delivers cost efficiency, launch readiness, and strong governance in a fast-moving environment. In this role, you will: Develop and execute sourcing strategies across GTM, Brand, Global Affairs, Events, Growth, and Partnership activities—spanning both B2B and B2C channels—that align with our mission and b

reactawsrest
View job →

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role As a Security Engineer on Detection & Response, you’ll help protect OpenAI’s most sensitive assets– including our intellectual property, customer data, and the infrastructure that supports them– by building and operating the systems we use to detect suspicious activity and respond effectively when it matters. You’ll work across endpoints, identity, cloud, hyperscale compute infrastructure, and datacenter-adjacent layers, partnering closely with security teams and infrastructure owners to define the telemetry and response requirements we need and building tooling and automation where it delivers the most leverage. In this role, you will: Build and evolve Detection & Response capabilities across OpenAI’s infrastructure, products, and research environments, with an emphasis on high-signal detection and reliable operational response. Engineer detection pipelines and tooling: develop rule lifecycle management, measurement/quality loops (coverage, precision, latency), tuning processes, and safe rollout patterns. Automate response and investigations by building workflows that reduce toil (triage, enrichment, containment, evidence capture) and improve time-to-understand/time-to-contain. Partner with other Security teams and system/infrastructure owners across the company to ensure new systems ship with the right telemetry, threat models, and response playbooks from day one. Define D&R requirements and drive visibility across endpoin

awsazuregcp
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The IT and Security organization builds the systems, data foundations, and automation that help OpenAI operate securely and reliably at scale. We support critical domains across identity, access, infrastructure security, enterprise systems, and internal productivity. As OpenAI grows, audit readiness and control assurance increasingly depend on reliable data: accurate system inventories, access populations, change records, configuration state, exception signals, and evidence generated directly from source systems. Our goal is to move beyond manual evidence collection and build scalable data products, automated validation, and continuous control monitoring that make security and IT controls measurable, repeatable, and defensible. About the Role We are looking for an IT Controls Data Engineer to build the data infrastructure that powers audit readiness, IT controls, evidence automation, and continuous control monitoring. In this role, you will design and maintain the pipelines, datasets, models, validation logic, dashboards, and evidence exports that make IT controls measurable, repeatable, and defensible. You will work across Security, IT, Infrastructure, Engineering, Finance Risk Management, and auditors to turn complex system behavior into reliable control data products. This is a technical builder role. The ideal candidate is strong in data engineering and analytics engineering, comfortable working with enterprise and security system data, and able to explain data lineage, source-system behavior, and control logic clearly to technical and audit stakeholders. You’ll be responsible for Building reliable data pipelines, models, and datasets for IT controls, including access, identity, configuration, change, ticketing, exception, and evidence data. Creating data quality, lineage, reconciliation, and completeness checks that make control data defensible for SOX and other audit use cases. Designing automated evidence generation workflows that produce compl

pythonsqlaws
View job →

About the Team The OpenAI Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role As a Software Engineer, Distributed Data Systems, you will design, build, and operate some of the largest distributed data systems in the world. You will be responsible for the end-to-end stack to deliver and consume top-quality data for robotics training at exabyte-scale. You’ll manage distributed data pipelines, collaborate closely with researchers to translate requirements into robust systems, and harden pipelines that serve as the backbone for OpenAI’s rapid iteration cycles. We’re looking for engineers who are detail-oriented, have strong experience with distributed systems, and excel at building reliable, large-scale systems in high-stakes environments. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and maintain data infrastructure such as exabyte-scale distributed data processing, data selection, automated labeling, and training data loaders. Ensure our data platform can scale by orders of magnitude while remaining reliable and efficient. Partner with researchers to deeply understand requirements and translate them into production-ready systems. Harden, optimize, and maintain critical data infrastructure systems that power multimodal training and evaluation. Deliver the best possible data for training robotics models. You might thrive in this role if you: Have strong experience with distributed systems and large-scale infrastructure with a strong interest in data. Are detail-oriented a

awsrestai
View job →
O
1mo ago

About the Team With Codex we’re building an AI software engineer. One that you can pair with, delegate to, or even ask to take on future tasks proactively. Our team is a fast-moving group within OpenAI, bringing together research, engineering, design, and product. We iteratively build the Codex agent harness and product to get the most out of the model, and we iteratively train the model to be great at complex software engineering tasks. The Codex team is responsible for building state-of-the-art AI systems that can write code, reason about software, and act as intelligent agents for developers and non-developers alike. We operate across research, engineering, product, and infrastructure; owning the full lifecycle of experimentation, deployment, and iteration on novel coding capabilities. Codex Enterprise builds the ecosystem, governance, and enterprise capabilities that help Codex spread across developers, teams, and organizations worldwide. The Enterprise Controls team owns the systems that allow companies to safely deploy Codex across their organization while protecting their most sensitive code, data, and internal knowledge. About the Role As Codex adoption grows inside large organizations, customers are increasingly trusting Codex with their most valuable assets: proprietary codebases, internal documentation, customer data, and sensitive workflows. This role will help build the enterprise control plane that makes Codex secure, governable, and trustworthy at scale. You will design and operate backend systems that give enterprise administrators visibility and control over how Codex is used across their organization. You will work across identity, access, encryption, policy enforcement, auditability, and admin controls. This may include systems that let customers manage encryption keys, control which Codex capabilities are enabled, enforce organizational policies, and understand how data flows through Codex. This role owns systems end-to-end: from architecture and

pythonjavaaws
View job →
O
1mo ago

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role As a Security Engineer on Detection & Response, you’ll help protect OpenAI’s most sensitive assets– including our intellectual property, customer data, and the infrastructure that supports them– by building and operating the systems we use to detect suspicious activity and respond effectively when it matters. You’ll work across endpoints, identity, cloud, hyperscale compute infrastructure, and datacenter-adjacent layers, partnering closely with security teams and infrastructure owners to define the telemetry and response requirements we need and building tooling and automation where it delivers the most leverage. In this role, you will: Build and evolve Detection & Response capabilities across OpenAI’s infrastructure, products, and research environments, with an emphasis on high-signal detection and reliable operational response. Engineer detection pipelines and tooling: develop rule lifecycle management, measurement/quality loops (coverage, precision, latency), tuning processes, and safe rollout patterns. Automate response and investigations by building workflows that reduce toil (triage, enrichment, containment, evidence capture) and improve time-to-understand/time-to-contain. Partner with other Security teams and system/infrastructure owners across the company to ensure new systems ship with the right telemetry, threat models, and response playbooks from day one. Define D&R requirements and drive visibility across endpoin

awsazuregcp
View job →
🔔

Get new infrastructure team manager jobs by email

Daily job updates · Unsubscribe anytime