Jobiba hiring network

Infrastructure Team Manager Jobs

4,815 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current infrastructure team manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Engineering Manager, Home Infrastructure The Home Infrastructure team builds the mission-critical backend and data systems that power Roblox’s Homepage and Experience Details Page, two of the highest-traffic surfaces on Roblox. These surfaces reach the vast majority of Roblox’s daily active users and are core drivers of discovery, engagement, retention, and platform growth. We are a full-stack product infrastructure team responsible for content distribution across Roblox. Our systems support multiple modes of user interaction, including exploratory browsing, directed discovery, and personalized content recommendations across the many types of content that make up the Roblox ecosystem. This team sits at the intersection of large-scale distributed systems, machine learning-powered personalization, data infrastructure, and product experimentation. We partner closely with Machine Learning, Data Science, Product, Design, Frontend, Ads, Marketplace, Virtual Economy, and other teams across Roblox to build the platforms that help users find the most relevant and engaging content. As Engineering Manager for Home Infrastructure, you will lead a team of Backend and Data Engineers responsible for the e

awsgitmachine learning
View job →
D
Discord
📍 San Francisco Bay Area• Full-time• $180K – $220K/yr
1mo ago

Discord has a highly engaged community of millions of daily active users who use the platform for many different reasons, but there’s one thing that nearly everyone does: play video games. Discord plays a uniquely important role in the future of gaming, and we are focused on making it easier and more fun for people to hang out before, during, and after playing games. The Realtime Infrastructure team is responsible for building and maintaining some of Discord’s highest scale and most critical services. Those systems are at the core of our text chat infrastructure and facilitate the dispatching of every update to our users sessions. This role will have a significant impact on Discord’s overall reliability and performance. It will also help our product teams build new features on top of our infrastructure. This team is small but critical, and its work has a direct impact on Discord's success and ability to scale. This role reports to the Senior Engineering Manager of Realtime Infrastructure. What You'll Be Doing Build and operate large-scale, reliable and performant distributed systems. Collaborate with product teams to create new features. Ensure Discord “just works”. Write code but also manage our infrastructure. Work with a talented team of engineers who have built one of the largest communication platforms in the world. What you should have 2+ years of experience writing and designing backend systems. Experience solving complex distributed system problems. Experience operating and maintaining critical tier 0 services. Knowledge of monitoring and alerting best practices. Familiar with open source software, and not afraid to dig into the source code of a library to find the answer you’re looking for. Bonus Points Experience with Elixir or Rust. Experience working with systems deployed in a cloud environment (GCP, AWS, etc.) Knowledge of devops tools like Salt,Terraform or k8s. You have built or contributed to open source projects. You are a Discord power user and hav

awsgcprest
View job →
B
Baseten
📍 San Francisco• Full-time
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Cloud Platform Engineer, you'll envision and build robust systems and processes that ensure our infrastructure is scalable, reliable, and efficient. This can range from automating deployments and monitoring systems to optimizing performance and managing incidents. We all work closely with our users, learning from their past struggles in operationalizing ML, onboarding them onto our platform, and turning our learnings into ideas for improving Baseten. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Build and maintain scalable infrastructure to support the deployment and operation of machine learning models. Establish standards and best practices for reliability and performance across the infrastructure. Automate processes when relevant, particularly for managing CI/CD pipelines. Own products and projects end-to-end, functioning as both an engineer and a project manager, with a focus on user empathy, project specification, and end-to-end execution. Collaborate with cross-functional teams to understand project requirements and translate them into technical solutions. Mentor junior team members and contribute to knowledge sharing within the organization. Navigate ambiguity and exercise good judgment on tradeoffs and

kubernetesci/cdgit
View job →
O
29 days ago

Technical Program Manager – Applied Infrastructure About the Team The Applied team safely brings OpenAI’s technology to the world, powering products like ChatGPT, and the APIs for GPT and more. Behind these products is a complex and rapidly evolving infrastructure platform that enables scale, performance, and safety. The Applied Infrastructure TPM team partners across engineering to lead foundational programs that ensure OpenAI’s infrastructure can meet current and future demand. About the Role We’re looking for a seasoned Technical Program Manager to drive critical infrastructure programs across the Applied organization. This TPM will focus on cross-cutting initiatives such as general compute capacity planning, process transformation, cost and quota attribution and optimization, and coordination across infrastructure and product stakeholders. There will also be focus on evolving OpenAI’s infrastructure to support growth, scale and new products. This work is core to how OpenAI manages and grows its infrastructure footprint in a disciplined, scalable way. Location: San Francisco, CA (Hybrid – 3 days/week in-office) In this role, you will: Serve as the DRI for complex infrastructure programs spanning CPU planning, orchestration, and other resource management domains (e.g. networking, storage). Build and operationalize systems to capture demand signals, model future capacity needs, and align infrastructure planning across internal teams and partners external to the company. Partner closely with Infrastructure, Product and Finance teams to forecast infrastructure usage patterns and ensure supply/demand alignment. Lead cost attribution and quota enforcement programs to promote stability and ensure equitable access to resources across teams. Drive simplification and standardization of infrastructure tooling and processes across Applied and Infra organizations. Drive cross functional programs to evolve our infrastructure to support new growth and scale Work with external v

awsazurerest
View job →
S
Sentry
📍 San Francisco• Full-time• Remote
11 days ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry's Infrastructure Engineering team is what makes operating Sentry simple, safe, and seamless for every other engineering team in the company. They build the internal control platforms, configuration systems, traffic routing, and automation that let product engineers operate services safely at scale without needing deep infrastructure expertise themselves. As the Engineering Manager for Infrastructure Engineering, you'll lead a team of engineers building the tools that power Sentry's growth: internal admin and change management tools, configuration automation, and the routing layer that underlies Sentry's architecture. You'll be responsible for technical vision, team health, system reliability, and partnership with engineering teams across the company who depend on your team's tools every day. You'll work closely with leaders across Infrastructure, Platform, and Production Engineering to shape how Sentry scales its operational model as the company grows. In this role you will Lead a team of engineers building the internal control platforms that every engineering team at Sentry relies on to operate services safely. Drive the evolution of Infrastructure Engineering's platform, including configuration management, traffic routing and environment controls Own the team's technical direction, contributing to key decisions on API architecture, internal tooling design, and automation frameworks. Nurture and grow engineers at different levels, providing support through coaching, mentorship, and career development. Foster an inclusive, high-performing team culture focused on ownership, learning, and delivery. Partne

REMOTEpythonkubernetesai
View job →
S
Sentry
📍 San Francisco• Full-time• $220K – $450K/yr
1mo ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry's Infrastructure Engineering team is what makes operating Sentry simple, safe, and seamless for every other engineering team in the company. They build the internal control platforms, configuration systems, traffic routing, and automation that let product engineers operate services safely at scale without needing deep infrastructure expertise themselves. As the Engineering Manager for Infrastructure Engineering, you'll lead a team of engineers building the tools that power Sentry's growth: internal admin and change management tools, configuration automation, and the routing layer that underlies Sentry's architecture. You'll be responsible for technical vision, team health, system reliability, and partnership with engineering teams across the company who depend on your team's tools every day. You'll work closely with leaders across Infrastructure, Platform, and Production Engineering to shape how Sentry scales its operational model as the company grows. In this role you will Lead a team of engineers building the internal control platforms that every engineering team at Sentry relies on to operate services safely. Drive the evolution of Infrastructure Engineering's platform, including configuration management, traffic routing and environment controls Own the team's technical direction, contributing to key decisions on API architecture, internal tooling design, and automation frameworks. Nurture and grow engineers at different levels, providing support through coaching, mentorship, and career development. Foster an inclusive, high-performing team culture focused on ownership, learning, and delivery. Partne

pythonkubernetesai
View job →
SI
1mo ago

Within Infrastructure Management Team the candidate is acting as single point of contacts for all business requirements linked to the Cloud. He/she is responsible for understanding current and future customer demand for services. He is preparing the CAB (change advisory board) The candidate must have significant experience in Cloud / SAAS / PAAS / Infrastructure / system integration/migration. Address: No.110/1 M Krishnappa Layout Lalbagh Road Facility: SBS IS&T Global SA SA Qualification: Graduate Experience: 10 - 15 years Source: Sodexo India | Job Code: IJP561563

O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI is helping build the infrastructure that powers the next generation of artificial intelligence. Through Stargate, we are developing and operating large-scale AI compute campuses that require world-class execution across data center design, construction, commissioning, and operations. The Infrastructure Operations team is responsible for bringing AI infrastructure online and ensuring it operates reliably at scale. We partner closely with hardware, network, deployment, construction, and operations teams to deliver mission-critical environments capable of supporting frontier AI workloads. As our footprint expands, operational excellence becomes increasingly important to ensuring safe, reliable, and efficient campus operations. About the Role We are seeking a Facilities Operations Manager to support the commissioning, operational readiness, and long-term operation of next-generation AI data center campuses. This role sits at the intersection of construction, commissioning, hardware deployment, and facilities operations. You will be responsible for ensuring mission-critical infrastructure is prepared to support hardware deployment, transitioned successfully into production operations, and maintained to the highest standards of reliability and availability. You will lead day-to-day operational execution across electrical, mechanical, controls, and supporting infrastructure systems while partnering closely with commissioning teams, site operators, vendors, and engineering organizations. This role requires a strong blend of technical depth, operational leadership, and cross-functional execution. Key Responsibilities Lead day-to-day operations of mission-critical facility infrastructure across AI compute campuses. Own operational readiness activities supporting new campus deployments and infrastructure expansion. Partner with commissioning teams to transition facilities from construction and startup into steady-state operations. Develop, implement, and

awsrestai
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Trust is central to Stripe’s mission to grow the GDP of the internet and the Risk team advances that mission by building products and infrastructure that Stripe businesses, Stripe product teams, and Stripe’s network of regulators and financial partners rely on. From helping businesses onboard to Stripe to managing regulatory compliance at scale, Risk is at the center of what makes Stripe run. The Technical Program Manager team is responsible for delivering the cross-Stripe programs that help businesses grow on Stripe, enable products to launch safely at scale, and Stripe to maintain its position as the most trusted network in the financial ecosystem. What you’ll do You will lead a team of TPMs and PGMs to deliver on Stripe’s most complex, cross-cutting, and critical programs. As a manager, you will support your team in leading programs that advance Stripe and Risk’s mission. You’ll hire, coach, and grow a team of excellent TPMs that support Stripe’s growing needs and the individuals’ careers. You will also act as a TPM yourself on the most complex initiatives that uniquely need your hands-on leadership. Responsibilities Build, lead and manage high-performing teams, managers, and individuals, across multiple locations, including providing performance reviews, continual feedback, and career growth. Define and deliver on long-term vision and multi-year strategy for company-wide programs collaborating with product management and engineering lead

S
Stripe
📍 San Francisco Or Seattle• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The Manager of the Technical Solutions Engineering (TSE) team will grow and inspire the team responsible for the success of developers and our largest users integrating with Stripe's platform & products. What you’ll do At Stripe we consider the developer experience to be central to the overall experience of our customers. When we do our job well, developers all over the world are able to smoothly launch and grow their businesses on Stripe, whether they’re integrating payments for the first time or building & growing complex systems on our financial infrastructure. Stripe is beloved by developers for the simplicity of our APIs, the thoroughness of our documentation, and our focus on developer experience. The Technical Solutions Engineering team is the glue to make that possible. We make our users feel empowered when we show them about something Stripe could do that they didn't think was possible. Internally, we champion developer experience as central to the overall experience around Stripe. Responsibilities Define and deliver a comprehensive technical support experience for developers and our most complex users (via internal user-facing teams) integrating Stripe Develop both the long-term vision and strategy for the Technical Solutions Engineering team and manage day-to-day operations Develop relationships across the entire organization at Stripe to influence others in aiming for the best developer experience possible Work cross-

D
Datadog
📍 New York• Full-time• From $156K/yr
21 days ago

As a Product Manager – IaC Detection, you will define, build, and launch capabilities that proactively detect infrastructure issues in code (e.g. Terraform, Helm) before they can be deployed into production and escalate into production incidents. The Infrastructure Monitoring team has pioneered shift-left detection in the industry with Bits Infrastructure Operations , and we’re looking for a Product Manager to expand this capability to a broader set of use cases Customers (and thus developers) are increasingly standardizing on IaC tools to deploy and maintain ever-growing infrastructure in the cloud. At the same time, SREs and Infra teams struggle with an increasing number of production incidents. By shifting-left and identifying high-impact infra changes before they are deployed, we help reduce production incidents, reduce waste, and free up SRE time to focus on value-added tasks. You will own the roadmap to expand IaC detection to a broader set of use cases, including cost detection, blast radius impact, as well as configuration changes on infrastructure powering applications like nginx, postgres and more. You’ll partner closely with Engineering, Design, and customers to build and iterate on the roadmap, build product market fit, drive customer adoption (including internal usage), and focus on coverage and correctness of the AI system. This is an opportunity to lead an initiative at the intersection of AI, infrastructure operations, and autonomous observability. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the product roadmap for IaC Detection, enabling customers to proactively detect and catch high-impact infrastructure and configuration changes before they are deployed into production and escalate into incidents. Define the end-to

gitairust
View job →
C
Cloudflare
📍 Hybrid• Full-time• Hybrid
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Location: Singapore About the Team In this role, you will be focused on the build out and expansion of our global network. You'll work closely with Cloudflare’s Network Infrastructure planning team, Network Engineering team, Infrastructure automation team, Site Reliability Engineering (SRE) team, Project Managers and with various vendors and partners (including hardware vendors, logistics, datacenter and network providers, and ISPs) to p

pythonawslinux
View job →
D
Datadog
📍 United Kingdom; Paris, France• Full-time
1mo ago

Datadog is seeking a Director, Technical Account Management (TAM) for the EMEA region to join our high-growth organization and world-class Technical Post Sales (TPS) Team. You will own the strategy, growth, and execution of a team of 25–30 today, with a clear mandate to scale to 50–70+ over the next two to three years. You will be building organizational infrastructure, developing first-line leaders, and establishing scalable processes before they're needed. TAMs are technical experts who provide paid "white glove" technical services to our largest customers, supporting adoption and providing guidance across the comprehensive Datadog product suite. A TAM is held in high regard as an expert and trusted advisor for how IT Operations translates to business value. This is a senior leadership role where you will set the technical post-sales vision for all of EMEA, own executive relationships with key accounts, evangelize the TPS offering at scale, and grow services revenue across the region. At Datadog, we place value in our office culture—the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Scale the EMEA TAM organization from 25–30 to 50–70+, hiring first-line leaders, designing team structures, and building systems that sustain quality through rapid growth. Lead a team of managers and senior TAM ICs across EMEA, ensuring customer success, retention, and expansion at high standards of technical delivery. Own the commercial strategy for the TPS program, partnering with Sales to position, price, and sell TAM services as a core part of the Datadog value proposition. Define and execute regional OKRs aligned with global TPS goals, driving accountability through your leadership team each quarter. Lead executive-level recruiting and build a leadership pipeline from within, developing managers who can in turn scale thei

awsazuregcp
View job →

Position: Marketing Manager with Team Industry: Construction Chemicals & Road Project Chemicals Locations: Varanasi Mumbai Guwahati Gujarat Key Responsibilities: Business development and market expansion Dealer, distributor, and contractor network development Team handling and sales management Government and private project business development Client acquisition and relationship management Achieving sales targets and revenue growth Market survey and competitor analysis Eligibility: • Experience in Construction Chemicals, Road Project Chemicals, Building Materials, Infrastructure, or related industries • Proven experience in sales, marketing, and team management • Strong dealer, distributor, contractor, and project network preferred • Ability to lead and motivate a sales team • Willingness to travel and develop new markets Salary: Attractive Salary Package + Incentives (As per experience and industry standards) Apply Now: Experienced candidates with their existing sales and marketing teams are encouraged to apply. Limited vacancies available. Immediate selection and joining. Apply Now Limited vacancies available. Immediate selection and joining for suitable candidates. How to Apply? Note: Send your resumes at same number on what sap You may contact us between 9 AM to 8 PM Ph: 8777211016 HR: 9331205133 HR: 9231799122 Visit Our Office Ideal Career Zone 128/12A, Bidhan Sarani Shyambazar Metro Gate No. 1 Gandhi Market, Behind Sajjaa Dhaam Bed Sheet Showroom Kolkata 7 lakh 4 #UrgentHiring, #MarketingManager, #TeamManager, #Marketing Executive, #Customer Relationship Manager, #SeniorAccountant,#ConstructionChemicalsCompany, #RoadProjectChemicals, #kolkat, #India, #idealcareerzone, #kolkatajobs, #WestBengal, #Silliguri, #Bihar, #Jharkhand, #बैंगल र, #कर्न टक, #इंड य, #आइड यलकर यरज न, #क लक त ज ब्स, #वेस्टबंग ल, #स ल गुड़, #ब ह र, #झ रखंड, #হ ওড়, #সল্টল ক, #কলক ত, #ব্য ন্ড ল, #হুগল, #পশ্চ মবঙ্গ, #ह वड़, #स ल्टलेक, #Howrah, #SaltLake, #Bagnan, #midnapur, #Amta, #Santoshpur

S
Snowflake
📍 Bellevue• Full-time
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a talented Tech Lead Manager (TLM) to lead Snowflake’s System Under Test (SUT) team, responsible for evolving how Snowflake engineers test the Snowflake product locally and at scale in CI. The SUT empowers Snowflake engineers by delivering a reliable, low-latency and cost-efficient developer experience across a high-growth, high-demand surface area. As the TLM for SUT, you will lead a small and highly technical team at the intersection of CI, developer infrastructure, and product engineering. You will set direction, drive execution, and partner broadly across Engineering Systems and product teams to deliver a more reliable, faster, and more maintainable test platform for Snowflake’s engineers. In this role, you will: Lead, coach, and grow the SUT team while creating a high-energy, cohesive environment with strong planning, ownership, and career development. Own the roadmap and execution for SUT rollout across development environments, CI and AI workflows. Drive measurable improvements in startup reliability, latency, and cost, using clear SLOs, dashboards, and operational metrics to guide decisions and raise the bar on execution. Serve as the technical anchor for the SUT domain, shaping architecture and guiding the evolution from legacy systems to a composable

kubernetesaigo
View job →
🔔

Get new infrastructure team manager jobs by email

Daily job updates · Unsubscribe anytime