Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com . Reddit has a flexible workforce! If you happen to live close to one of our physical office locations our doors are open for you to come into the office as often as you'd like. Don't live near one of our offices? No worries: You can apply to work remotely in any country in which we have a physical presence. Team Description Reddit is poised to rapidly innovate and grow like no other time in its history. We’re currently hiring across multiple teams, some of these teams include: Ads ML Serving Team The Ads ML Serving team is part of Reddit’s Ads ML Platform, which builds the infrastructure and tools that power machine learning across Ads. This team focuses on creating a highly reliable, scalable, and efficient ML serving stack. Their work includes evolving long-term serving architecture, integrating closely with the ads serving stack, optimizing CPU/GPU performance, and building model velocity tools like observability libraries and model quality gating. Attribution & Identity Team The Attribution & Identity team builds products that help advertisers understand and measure the impact of their campaigns. They focus on attribution systems, identity solutions, and advertiser experimentation tools that improve performance insights and usability. Their goal is to make Reddit’s advertising platform more effective, transparent, and data-driven. Ads Growth Team The Ads Growth team drives initiatives to expand Reddit’s advertiser base, with a focus on Small to Medium Businesses (SMBs). We build and scale the technical founda
Jobiba hiring network
Infrastructure Team Manager Jobs
4,730 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current infrastructure team manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. Plaid's Recruiting Operations team exists to build and sustain the infrastructure, programs, and data that enable our recruiting organization to hire effectively, efficiently, and equitably. We drive operational strategy, systems implementation, data integrity, and process consistency across the full hiring lifecycle. In 2026, we are focused on scaling our coordination function, fully operationalizing Ashby as our ATS, and launching programs that raise the bar for recruiting effectiveness across Plaid. We are a lean, high-impact team, and we are growing. As the Recruiting Operations Coordinator at Plaid, you will sit at the intersection of recruiting coordination and operational program support. You will own white-glove scheduling for high-touch hiring pipelines, serve as the on-site RC presence in our San Francisco office, and act as a core member of the Recruiting Operations team -- contributing to the systems, tools, and processes that keep the function running. This is an entry-level program management role that flexes between recruiting coordination and operational program support -- the balance will shift based on hiring volume and team needs. You will be one of the first pe
We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. ABOUT THE ROLE We are seeking a highly skilled and intellectually curious analyst to shape the future of financial data at Instacart, driving innovation at the intersection of systems, analytics, and business strategy. In this role, you will architect and deliver data models, pipelines, and reporting infrastructure that empower Finance and Accounting with trusted insights at scale. You’ll collaborate across technical and business teams to ensure financial data is accurate, reliable, and compliant, while also advancing large-scale systems initiatives, accelerating the month-end close, and optimizing cloud spend in our AWS environment. ABOUT THE TEAM The Financial Data Analytics team is the connective tissue between Finance, Accounting, and Engineering. Our mission is to deliver trusted financial data that is both a foundation for compliance and a cataly
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: As a member of the Global Technical Operations (TechOps), you will be a part of a team that focuses on operational reliability within a cloud-based infrastructure. You have hands-on cloud experience in architecting, building, deploying, managing databases, compute instances, and storage buckets. You have a passion for providing solutions through automation. You know that success is through collaboration and communication. What you'll be doing: Work in cross-functional teams to develop solutions and identify opportunities to bring efficiency and effectiveness. Research, evaluate, and incorporate new technologies/concepts into existing frameworks. Proactively identify areas to improve efficiency and effectiveness, recommend and implement solutions towards them. Develop and innovate operational practices, procedures for workflows, and documentation. Implement and contribute to IT security best practices. Automate tasks to ensure consistency and speed of deployment. Identify, analyze, and troubleshoot issues and work towards resolution. Explain technical solutions to bo
About the Team Compute Foundations builds the software that manages OpenAI’s GPU compute infrastructure across sites, data centers, and infrastructure providers, supporting model training and inference. Our systems turn large, heterogeneous fleets of machines into dependable compute for research and products. We build Kubernetes-based control planes, controllers, services, and APIs that coordinate the lifecycle of machines and clusters. We connect global infrastructure management with the realities of bare-metal systems, giving clients consistent interfaces across differences in hardware, topology, and provider behavior. About the Role You will build distributed systems that provision, configure, and manage compute throughout its lifecycle. Your work will connect global services and Kubernetes controllers with the systems that bring machines online, update them safely, and recover them when something goes wrong. This role combines software architecture with an understanding of how machines and data centers work. You might design a lifecycle API, improve controller performance under high concurrency and provider rate limits, or trace a provisioning failure from an API through reconciliation to network boot or host configuration. You will help these systems remain reliable as the fleet expands across sites and generations of GPU hardware. We value depth in relevant systems and the ability to connect layers. You do not need to arrive as an expert in every component of the stack. In this role, you will: Design, build, and operate Kubernetes-based controllers and distributed services that coordinate infrastructure across sites, isolate failures, and scale as GPU capacity grows. Define APIs and resource models that let clients request and track lifecycle operations through consistent interfaces across hardware platforms and providers. Build provisioning and configuration services that coordinate network boot, hardware management interfaces, and the deployment of firmware,
About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. Through a combination of strategic partnerships and self-built campuses, we are scaling the compute, storage, and networking platforms that power frontier AI training and inference. The Scaling Analytics team builds the data and software systems that help Industrial Compute understand, plan, and operate infrastructure at global scale. We work across capacity, hardware, storage, infrastructure software, and operational systems to connect fragmented sources of infrastructure data and make that information reliable and usable for engineering and planning. As OpenAI's infrastructure footprint grows, CPU and storage data increasingly spans internal platforms, vendor systems, APIs, databases, object storage, capacity management systems, and operational tooling. Building reliable connections across these environments is critical to understanding available capacity, utilization, fleet state, and infrastructure growth. About the Role We are seeking a Data Engineer to build the data systems and integrations that connect OpenAI's CPU, storage, and supporting infrastructure platforms. This role sits at the intersection of data engineering and backend software engineering. Rather than focusing primarily on traditional analytical pipelines, you will build the software and integrations required to collect, normalize, and make infrastructure data available across a heterogeneous set of systems. CPU and storage data may originate from internal infrastructure platforms, vendor APIs, databases, object storage, capacity systems, and operational services. You will determine how to reliably connect these systems and where those integrations should live—whether within an existing infrastructure service, an orchestration framework, a scheduled workload, or a purpose-built application. You will work closely with Infrastructure Engineering, Capacity Engineering, Storage,
About the Team OpenAI’s Industrial Compute organization is building the infrastructure required to support the next generation of frontier AI systems. Through a combination of strategic partnerships and self-built data center campuses, we are scaling the power, cooling, electrical, mechanical, and controls infrastructure needed to deliver compute at unprecedented scale. The Commissioning organization is responsible for ensuring this infrastructure is safely tested, validated, integrated, and transitioned into reliable operations. For our self-build campuses, the team operates through a hybrid delivery model: OpenAI provides commissioning leadership, discipline ownership, governance, and project integration, while commissioning partners provide field and test engineering capacity to support inspections, startup, testing, and turnover. About the Role We are seeking a Commissioning Project Lead to own the commissioning strategy and execution for a large-scale, self-build data center project. You will lead the overall commissioning program from early construction planning through startup, functional testing, integrated systems testing, and final turnover. You will establish the commissioning execution plan, integrate commissioning activities into the master project schedule, coordinate multidisciplinary readiness, and lead the vendor commissioning partners providing field and test engineering capacity. This role serves as the primary commissioning interface to project leadership, construction management, contractors, equipment vendors, operations, and commissioning partners. You will be responsible for creating clarity across organizations, identifying readiness and schedule risks early, and ensuring the facility progresses through testing and turnover against clearly defined acceptance criteria. The role will initially support planning and coordination in a hybrid capacity and transition to full-time onsite presence as construction, inspections, startup, testing, and t
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We’re looking for an IT Site Specialist to join our IT Team at Ramp based in our San Francisco office! This is a contract on-site role that blends hands-on, in-person site support with ownership of our core SaaS application stack, and we weigh both sides equally. On the site side, you’ll be the primary IT presence in SF — owning deskside support, onboarding/offboarding, endpoint and AV support, and office inventory. On the application side, you’ll administer the tools the whole company depends on — Okta, Google Workspace, Slack, JAMF, and a growing set of SaaS platforms — owning provisioning, access, portfolio, renewals, integrations, and automation. We’re looking for someone who takes pride in running a smooth on-the-ground operation while also thinking like a systems owner who scales IT through automation and sound identity practices. What You’ll Do Provide in-person support at our San Francisco office 5 days/week , acting as the site’s primary IT point of contact. Perform IT Support Specialist (L4) duties, encompassing all responsi
The team + the role Pendo's Applied AI team turns AI infrastructure into real business outcomes across Sales, Marketing, and Customer Engineering. AI here isn't a feature we bolt on — it's how we scale GTM capability across the company. We build AI that makes GTM teams measurably faster and more effective, and we measure success by whether those teams are actually using what we ship and getting real value from it. As an Applied AI Engineer, you'll own the full lifecycle of AI-powered solutions — from problem definition and prompt design through to production deployment and ongoing iteration. You'll work directly with GTM stakeholders to identify high-value problems, set realistic expectations about what AI can and can't do, and ship solutions that stick. The best person for this role has a strong engineering instinct, deep curiosity about how businesses operate, and the judgment to know when AI is the right tool — and when it isn't. This role is based in Raleigh, NC and follows Pendo's hybrid model: in-office 3 days per week. What this looks like day-to-day Own AI solutions end-to-end: problem definition, prompt design, production deployment, monitoring, and iteration — you ship, you watch, you improve. Build and manage GTM workflow automations that reduce manual work across Sales, Marketing, and Customer Engineering systems, including Slackbots and other integrations. Partner directly with GTM stakeholders to surface high-value problems, validate solutions, drive adoption, and set honest expectations about AI capabilities and limitations. Develop and maintain prompt management practices that make AI outputs reliable, auditable, and improvable over time — treat prompts as production code, not experiments. Work with the Data Platform and Systems teams to identify foundational tooling gaps and contribute clear, actionable requirements based on what you encounter in production. Share reusable AI workflow patterns and tool findings with the broader team — your impact sh
Who we are Stripe is a financial infrastructure platform for businesses. Millions of companies, from the world’s largest enterprises to the most ambitious startups, use Stripe to accept payments, grow their revenue, and pursue new business opportunities. Our mission is to increase the GDP of the internet. We have significant work ahead, which gives you the opportunity to do important work that expands access to the global economy. About the team The APAC Risk team helps Stripe launch and grow products responsibly. We work closely with global Product, Partnerships, and Risk teams to identify product risks early, define practical controls, and make risk decisions easier to execute at scale. What you’ll do As a Risk Strategist, you will help Stripe launch new Local Payment Methods (LPMs) and other products safely as part of our broader global expansion efforts. You will be part of a team that identifies, assesses, and manages risk across LPMs and other product launches. Your focus will be on building agent-based workflows that help scale LPM launches and related risk assessments across cross-functional teams. Responsibilities Identify risk scenarios associated with new products, product changes, and LPM initiatives. Translate those scenarios into clear requirements, controls, and actions for different cross-functional teams Improve and automate product-risk workflows through AI tools, data, and agent-based solutions. Define and build consistent, scalable processes with reduced manual load Help set the risk strategy and operating approach for product launches, including how teams assess risk, record decisions, and monitor outcomes Communicate complex risk issues clearly to technical and non-technical stakeholders, with practical recommendations that support responsible product growth Help define metrics that track risk exposure, control effectiveness, and operational efficiency Build strong working relationships across Risk and Product, and help align teams when trade-o
About the Team API Enterprise Controls is part of the API Infrastructure organization and owns the platform capabilities that help developers, startups, and enterprises adopt the OpenAI API securely and confidently. We build the systems underneath our APIs and developer platform across authentication and identity, service accounts and key management, secure networking, compliance, auditability, observability, and operational controls. Our users are developers and teams running critical applications on OpenAI, and we partner closely with Product, go-to-market, security, and infrastructure teams to turn their most important needs into reliable, intuitive platform capabilities. About the Role We are looking for an exceptional backend software engineer to help define and ship the enterprise capabilities our API Platform needs to scale.; this is a product-engineering role grounded in deep backend systems. You will work across databases, streaming systems, request routing, authentication, and developer-facing APIs while bringing strong product judgment, developer empathy, and attention to the small details that make a platform easier to understand, trust, and operate. You will lead large cross-functional initiatives, work closely with Product and go-to-market teams, engage directly with sophisticated users, and carry ambiguous needs from discovery through design, launch, and iteration. In this role, you will: Own backend product capabilities end to end across authentication and identity, service accounts and key controls, secure networking, compliance, observability, and operational workflows. Partner with Product, go-to-market, security, infrastructure teams, and sophisticated customers to identify needs, shape the roadmap, and lead large cross-functional projects from design through launch. Design developer-facing APIs, system behavior, configuration, error handling, safe defaults, auditing, and notifications with exceptional care for the details that define a great dev
About the Team OpenAI’s Finance and Revenue Operations organization builds the commercial infrastructure that enables the business to scale with speed, discipline, and financial integrity. Within Revenue Operations, Deal Desk partners closely with Sales, Partnerships, Product, Engineering, Legal, Technical Revenue, Finance, Billing, Order Management, and GTM Systems to turn complex commercial opportunities into executable, scalable transactions. About the Role We are hiring a Strategic Deals & Commercial Architecture Lead — Marketplaces & Partnerships to own the commercial architecture, execution, and governance layer for marketplace-enabled and partner-led transactions. This senior individual-contributor role is for someone who combines enterprise deal judgment, marketplace fluency, analytical rigor, and a builder mindset. You will take shaped opportunities from intake through approval and launch readiness, translating first-of-kind structures into clear economics, executable terms, quote-to-cash requirements, controls, and operating mechanisms. You will partner closely with teams that own business development and partner relationships; this role does not own partner sourcing, pipeline generation, or sales closing. You will support the commercial requirements for onboarding and integration without owning technical delivery. Success is measured both by the decisions and deals you enable and by the durable policies, systems, and controls you leave behind. This role is based in San Francisco, CA. We use a hybrid work model of three office days per week and offer relocation assistance. In This Role, You Will Lead the commercial architecture and execution of complex marketplace-enabled and partner-led enterprise transactions from intake through approval, contracting, launch readiness, and operational handoff. Structure private offers, pricing, fees, incentives, commitments, revenue share, credits, renewals, amendments, and other non-standard or multi-party terms
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Large, complex businesses on Stripe have historically had to create a new Stripe account for each global entity or business unit—making it difficult to unify reporting and manage payment operations across all accounts. Our customers want a way to manage these accounts under one umbrella so they can operate at scale more efficiently. We are responsible for building Organizations, which allows merchants to effectively and centrally manage their businesses across multiple Stripe accounts. As Stripe continues to grow its suite of product offerings beyond payments, Organizations will unlock a new phase of growth for Stripe. We will do this by enabling our platforms and businesses to model their complex businesses and distribute all of Stripe’s products to their end users. You can read more about what we are working on here (our Stripe Blog). What you’ll do As a leader on the User Auth Experience team, you will help shape the team’s vision and roadmap, guide and mentor your peers, and lead by example. Authentication, Login, and general Account Security experiences are deeply cross-cutting areas at Stripe with critical security, safety, and quality considerations. You will work with many partner teams: peer engineering organizations to build robust/reusable/standardized access controls, as well as non-engineer groups to shape the security, beauty, and usability of the end-to-end journey of Stripe account authentication. You’ll pla
About the Team OpenAI’s Finance and Revenue Operations organization builds the commercial infrastructure that enables the business to scale with speed, discipline, and financial integrity. Deal Desk operates as the commercial strategy and governance function at the intersection of Sales, Partnerships, Legal, Technical Revenue, Finance, Order Management, Billing Operations, Product, and GTM Systems. We architect complex enterprise transactions, turn ambiguity into executable decisions, and create the guardrails that let the business move quickly with operational and financial discipline. As OpenAI’s enterprise business grows in scale and complexity, Deal Desk defines how novel commercial motions become durable operating capabilities. We convert precedent-setting deal decisions into repeatable policy, controls, workflows, and systems requirements. About the Role We are hiring a Strategic Deals & Commercial Architecture Lead to lead the structuring and governance of OpenAI’s most complex enterprise transactions. This is a senior individual-contributor leadership role for someone with exceptional enterprise deal judgment, operational rigor, and a builder mindset. You will serve as the commercial architect for high-stakes opportunities, translating ambiguous requirements into coherent deal structures, approval strategies, and executable quote-to-cash plans. Your work will shape more than individual transactions. You will establish decision principles and precedent, clarify tradeoffs, and turn recurring patterns into scalable guidance, controls, systems requirements, and enablement. You will own the commercial decision and governance layer that helps strategic opportunities move decisively while managing downstream risk across contracting, billing, revenue recognition, provisioning, reporting, controls, auditability, and customer experience. The right candidate can move seamlessly between advising on a single high-value transaction and improving the operating model be
About the Team OpenAI, in close collaboration with our capital partners, is building the world’s most advanced AI infrastructure ecosystem. Our Industrial Compute organization develops and deploys large-scale AI campuses designed to support the next generation of frontier model training and inference workloads. The Hardware Operations team is responsible for ensuring the reliability, availability, and lifecycle health of OpenAI’s compute infrastructure. We partner closely with Data Center Operations, Fleet Health Engineering, Manufacturing, Network Infrastructure, Capacity Planning, and our infrastructure partners to maintain world-class operational performance across rapidly expanding AI environments. As we scale globally, we are building the operational frameworks, reliability standards, and sustaining engineering practices required to support thousands of GPUs and servers across multiple campuses. About the Role We are seeking a Datacenter Hardware Technician Lead to serve as the senior on-site technical authority for hardware reliability and fleet health at one of OpenAI’s flagship AI campuses. This role operates at the intersection of hardware operations, sustaining engineering, and fleet reliability. You will partner closely with Cloud Service Provider operations teams, OpenAI fleet-health engineers, hardware engineering teams, and OEM vendors to identify, diagnose, and resolve hardware issues affecting production systems. Beyond day-to-day operational support, you will drive root cause investigations, reliability improvement initiatives, lifecycle management programs, and operational readiness efforts. You will help establish hardware maintenance standards, operational procedures, and best practices that scale across future OpenAI infrastructure deployments. The ideal candidate combines deep hands-on datacenter hardware expertise with strong troubleshooting, failure analysis, and cross-functional leadership skills. Candidates must be able to sit onsite at our
Get new infrastructure team manager jobs by email
Daily job updates · Unsubscribe anytime