Jobiba hiring network

Capacity Planning Lead Jobs

600 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current capacity planning lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

G
1mo ago

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. Overview of the Role GitLab.com is being rebuilt as a horizontally scalable cluster. Two primitives make that possible: Organizations , the tenant boundary that makes a customer's data a single addressable, portable unit, and Cells , independent GitLab instances that host those Organizations and can be provisioned on demand across clouds and regions. The target is a 100x increase in headroom for GitLab.com, delivered without customers ever seeing a cell. Same domain, same URLs, same product — a substrate underneath that scales by adding capacity rather than by negotiating with a single database. The boundary we are building

gitrestai
View job →
G
Godaddy
📍 United States• Full-time• From $128K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team… GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the industry, powering the object, block, and file storage platforms that underpin hosting, applications, internal infrastructure, and next-generation AI/HPC workloads. If you're passionate about distributed systems, large-scale storage architecture, and solving complex reliability challenges, you'll work on infrastructure that few engineers ever experience. At GoDaddy, Ceph isn't a side project — it's a critical platform. Our environment spans 80+ production clusters, 20,000+ OSDs, and approximately 300 PB of raw storage capacity, supporting tens of billions of objects across multiple continents. The scale demands deep technical expertise in storage architecture, automation, observability, and performance engineering. As a Senior Site Reliability Engineer, you'll be a key technical owner of the platform, responsible for maintaining reliability, driving operational excellence, and influencing the future evolution of our storage ecosystem. You'll tackle challenging production problems, develop automation that operates at massive scale, contribute to architectural decisions, and collaborate with some of the industry's most experienced Ceph engineers. This is an opportunity to have direct impact on a storage platform that serves millions of customers worldwide. What You'll Get to Do… Own the reliability, performance, scalability, and capacity of large-scale production Ceph environments supporting object, block, and file storage wor

pythonkuberneteslinux
View job →
NR
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are seeking a highly analytical and strategic Director of Product Management to own product pricing, metering, and monetization at New Relic. This leader will ensure our pricing architecture is highly competitive, financially sound, and seamlessly aligns with customer value across our entire portfolio. What you'll do Pricing & Monetization Strategy: Design and battle-test list prices, discount curves, and price multiples for all existing and emerging technologies, with a specific focus on advanced AI capabilities and Compute Capacity Units (CCU). Financial Modeling & Analytics: Build comprehensive financial models encompassing COGS, revenue projections, and gross margin impact to ensure all pricing and packaging decisions drive sustainable business growth. Strategic Migration: Develop and execute the transition frameworks (the "carrots and sticks") required to effectively and smoothly migrate existing customers from legacy pricing. Meter Optimization: Own the ongoing effectiveness of the CCU meter, ensuring it remains compelling in the observability market and drives b

gitrestai
View job →
N
Nuro
📍 Mountain View• Full-time• $176.4K – $264.6K/yr
1mo ago

Who We Are Nuro is a self-driving technology company on a mission to make autonomy accessible to all. Founded in 2016, Nuro is building the world’s most scalable driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets to personally owned vehicles. With technology proven over years of self-driving deployments, Nuro gives the automakers and mobility platforms a clear path to AVs at commercial scale, empowering a safer, richer, and more connected future. About the Team The Talent Acquisition team is composed of Sourcing, Recruiting and Coordination, and serves as a strategic partner to all of Nuro’s functions, encompassing engineering and autonomy, as well as operations and business. We combine data-driven insights with deep market expertise to ensure Nuro continues to attract and hire exceptional talent. Our recruiters and sourcers are builders at heart, partnering closely with leadership and hiring teams to deliver a candidate experience that reflects Nuro’s mission and values. About the Role We’re seeking a Head of Talent Acquisition to lead and scale Nuro’s recruiting and sourcing organization through our next phase of growth and commercialization. This role will oversee both the recruiting teams and the dedicated sourcing team responsible for hiring across Nuro, including the technology, engineering, business, and operations functions. You’ll partner closely with senior executives and functional VPs to design capacity-driven hiring plans, leverage data to guide staffing decisions, and ensure our recruiting teams have the strategy, systems, and infrastructure needed to execute with excellence. This leader will also drive key executive searches, serving as a hands-on partner to Nuro’s leadership team to attract world-class talent in AI, ML, autonomy, and robotics. About the Work Lead and scale Nuro’s recruiting and sourcing org

S
Smartsheet
📍 Bellevue• Full-time• From $1.1M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Smartsheet is seeking a dynamic Manager of Sales Development to lead, coach, and motivate our team of Sales Development Representatives (SDRs). In this high-impact role, you will directly influence decision-making and drive growth in a fast-paced environment. This is a US-based remote position reporting directly to the Director of Sales Development, with a preference for candidates located in the Central, Mountain, or Pacific time zones. If you are an innovative leader who challenges the status quo, asks sharp questions, and elevates team performance, we want to hear from you. Here is what we are looking for: You Will: Hire, train and manage a team of Sales Development Representatives Develop strategies for career growth within the Smartsheet sales organization Report on activity metrics and forecast to the sales executives/directors Motivate individuals and team to exceed objectives through coaching and mentorship Actively use Salesforce.com and other SaaS tools to manage sales process, and set standards for performance metrics You Have: 2+ years sales experience in SaaS 1-3 years sales management experience Passion for coaching and mentoring emerging sales talent Experience giving both positive and constructive feedback Demonstrated ability to collaborate with a distributed sales team Capability to understand customer pain points and requirements, capacity to respond with value of Smartsheet products and services Current US Perks & Benefits: Employer subsidized medical/vision and dental coverage f

S
Smartsheet
📍 Dallas• Full-time• From $1.1M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Smartsheet is seeking a dynamic Manager of Sales Development to lead, coach, and motivate our team of Sales Development Representatives (SDRs). In this high-impact role, you will directly influence decision-making and drive growth in a fast-paced environment. This is a US-based remote position reporting directly to the Director of Sales Development, with a preference for candidates located in the Central, Mountain, or Pacific time zones. If you are an innovative leader who challenges the status quo, asks sharp questions, and elevates team performance, we want to hear from you. Here is what we are looking for: You Will: Hire, train and manage a team of Sales Development Representatives Develop strategies for career growth within the Smartsheet sales organization Report on activity metrics and forecast to the sales executives/directors Motivate individuals and team to exceed objectives through coaching and mentorship Actively use Salesforce.com and other SaaS tools to manage sales process, and set standards for performance metrics You Have: 2+ years sales experience in SaaS 1-3 years sales management experience Passion for coaching and mentoring emerging sales talent Experience giving both positive and constructive feedback Demonstrated ability to collaborate with a distributed sales team Capability to understand customer pain points and requirements, capacity to respond with value of Smartsheet products and services Current US Perks & Benefits: Employer subsidized medical/vision and dental coverage f

S
Sofi
📍 Cottonwood Heights• Full-time
1mo ago

Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we’re changing the way people think about and interact with personal finance. We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The role: This role sits within Enterprise Risk Management (ERM) as part of Independent Risk Management (IRM) and serves as a senior second-line-of-defense (2LOD) risk leader supporting SoFi’s international growth strategy. The Senior Director will lead risk oversight for new products, business initiatives, and international market expansion activities, with an emphasis on retail banking and consumer lending products. This leader will serve in a Business Unit Risk Officer (BURO)-style capacity, partnering closely with business, compliance, legal, and functional risk teams to establish scalable governance frameworks for international expansion and consumer financial products. The role is responsible for driving robust new product approval governance, ensuring alignment with SoFi’s risk appetite, regulatory expectations, and strategic objectives across multiple jurisdictions. The ideal candidate brings deep expertise in retail banking and consumer financial products, experience operating in global regulatory environments, and a demonstrated ability to build and mature international risk governance frameworks. What you’ll do: Serve as the primary IRM point of contact and escalation lead for international product, business, and expansion-related risk matters. Provide independent challenge and strategic risk gui

gitrestai
View job →
R
Roblox
📍 San Mateo• Full-time• From $345K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer on the Compute team, you will be the technical anchor for Roblox's GPU and AI accelerator capabilities. This is a battle-tested GPU expert role focused on the machine management layer and above: how GPU hosts are made production-ready, kept healthy, and turned into reliable compute for the workloads that depend on them. You will own the hard problems that show up only at scale, from driver and firmware management to GPU health, reliability, and performance across a rapidly growing fleet of accelerators spanning Roblox data centers and cloud environments. You will set the technical direction for GPU compute and up-level the entire organization's GPU expertise. You will: Serve as the GPU technical leader for the Compute team, partnering across Kubernetes, Machine Bootstrap, Networking, and Cloud to drive GPU strategy end to end. Own the GPU host lifecycle above raw fleet management: driver, firmware, and CUDA stack management, GPU health and telemetry, and remediation of GPU-specific failures (XID errors, ECC, thermal, NVLink and fabric faults). Architect how GPU capacity is exposed to compute platforms, including scheduling, isolation, and integration with Ku

awskubernetesgit
View job →
R
Roblox
📍 San Mateo• Full-time• From $345K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer leading Fleet Management, you will be the overall technical lead across three pods and the person who sets the technical direction for the fleet management layer of Roblox. This is a hands-on, deeply technical leadership role that owns all of Roblox's compute capacity end to end: from low-level provisioning and the data plane, up through the control planes that operate it, and all the way to the UI and internal-facing products that let teams self-serve capacity. Your org centralizes security, maintenance operations, and the uptime of every Roblox Kubernetes cluster, and governs the internal customer contracts that drive automation across the fleet spanning Roblox data centers and cloud providers. You will guide architecture, raise the engineering bar, and make sure compute capacity supply and demand stay in balance as the fleet grows. You will: Serve as the overall technical lead for three Fleet Management pods, setting and aligning the technical direction across low-level provisioning, the data plane, and the control plane and product surfaces above them. Architect the declarative, Kubernetes-style control planes that operate Roblox's compute fleet across o

sqlawskubernetes
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's Cache team is building a next-generation caching solution designed to deliver sub-millisecond average latency, horizontal scalability, and high efficiency—all at a drastically lower cost. Our ultimate vision is to shape a caching infrastructure capable of supporting 1 billion Daily Active Users while reducing costs by 90%. We are turning hours of onboarding and capacity expansion into seconds, freeing service owners entirely from managing cluster lifecycles. As a Senior Engineer on the Cache team (part of the Infra Storage org), you will innovate and operate large-scale, in-house distributed systems to solve Roblox's ever-growing caching challenges. You will report directly to the Engineering Manager for the Cache team. (Check out our recent engineering blog post here to learn more about the team's latest work!) You will: Lead the architectural transition to a next-generation, multitenant caching service built on ValKey, ensuring strict data, resource, and failure isolation for all tenants. Drive systemic optimizations to mitigate head-of-line blocking, manage hot keys, and maximize CPU and memory utilization across physical machine clusters. Design and build robust frameworks to a

redisawskubernetes
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Infrastructure Compute Site Reliability Engineering mission is to own and manage the successful operation of our underlying cell infrastructure system, along with elements of service discovery, secrets management and related software layers. We’re looking for a skilled Senior Site Reliability Engineer with strong programming skills to help us build Roblox's private cloud, productionize our growing Kubernetes-based infrastructure, and institute reliability best practices across the Roblox Compute team. You will: Design and Develop systems & libraries that promote fault-tolerance and resilience, automate much of the management and lifecycle of our clusters, and ensure systems are observable. Promote and Institute reliability best practices across the Infra Compute group, drive common reliability initiatives. Provides collaborative technical reviews and operational guidance to strengthen system reliability. Build, Automate and Standardize process automation to create a "golden path" of tooling and platform support that powers the fundamental Roblox ecosystem. Create Tooling that provides production guardrails, by evaluating release candidate capacity with load testing tooling before de

javaawskubernetes
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on the Orchestration pod within Engine Productivity, you'll design and run the platform that executes large-scale end-to-end and integration tests, running the real, shipping client on real hardware, across Roblox's data centers, cloud, and our own device labs, so our engineering teams can ship the engine, clients, Studio, and more with speed and confidence. Every Roblox engine, client, and Studio change, along with the experiences built on top of them, should ship with confidence, and the Orchestration team is the layer that makes that possible. We build large-scale distributed services that turn thousands of test suites into a reliable, push-button pipeline: fanning work out across fleets of machines and real devices, moving artifacts to where they're needed, managing single- and multi-client test state, and giving test owners and maintainers a system to validate their own runs. It looks a lot like building a specialized cloud platform, with capacity-aware scheduling, isolation and sandboxing, and smart retry and backoff, plus the classic distributed systems problems (fairness, efficiency, failure handling, and reliability) at Roblox scale. You Will: Design a

pythonawsgit
View job →
A
1mo ago

We're seeking a strategic, highly organized Program Manager to lead the operational backbone for Asana's incubation motions across new product lines. This is a unique opportunity to help Asana bring early-stage products to market by supporting the specialist sales, solutions, and post-sale teams that sit closest to our incubation customers. You will work across multiple products in the incubation stage in direct partnership with product GMs, sales and solutions leaders, post-sale partners, and R&D Ops leadership. Your role will be to create the systems that make these motions scalable and accountable: intake processes, rules of engagement, reporting and dashboards, capacity tracking, stage-gate reviews, field communications, and operating cadences. This role is ideal for someone who thrives in ambiguity, loves building operating systems from scratch, and knows how to bring structure to cross-functional work without slowing it down. You should be excited to work at the intersection of Product, R&D Operations, Revenue, Customer Success, and Finance to help Asana learn faster and scale the right motions at the right time. This role is based in our Chicago office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do and the teams with which you partner. If you're interviewing for this role, your recruiter will share more about the in-office requirements. What you’ll achieve Design, launch, and continuously improve the process to engage overlay sales teams across new product lines, including clear entry criteria, required context, routing rules, and escalation paths Define, document, and maintain rules of enga

reactrestai
View job →
A
Asana
📍 Warsaw• Full-time• $372K – $432K/yr
1mo ago

We're looking for a Senior Platform Reliability Engineer who brings strong software engineering skills and a deep understanding of system behavior under load and stress. This role is a good fit for someone who wants to own reliability as a first-class concern – building the foundational systems that protect Asana's platform, not just responding when things go wrong. You'll build core platform systems like load shedding, rate limiting, circuit breakers, and traffic controls that protect Asana under real-world load. This is deep, cross-cutting work that shapes stability and performance of our entire infrastructure – and you'll partner closely with other platform teams to make reliability something that's built in, not bolted on. Our tech stack includes: AWS, Kubernetes (EKS), CloudFront, Istio, Cilium, MySQL (RDS), OpenSearch, DynamoDB, Redis, Terraform, Datadog, TypeScript, Scala, Go, and Python. (Yeah, we know this sounds like buzzword bingo – but we want this post to actually show up in your searches.) Why this role? Reliability as a first-class feature : You won't be patching things up after the fact. You'll build the systems that make Asana resilient by design. Foundational work : Load shedding, traffic management, ingress/egress – these are the building blocks that protect everything else. You'll own them. Strong collaboration, reasonable hours : You'll work closely with infrastructure teams in Warsaw and Reykjavik, making deep collaboration practical without constant timezone gymnastics. Room to grow : This is a new team, and you'll help shape what Platform Reliability Engineering looks like at Asana – whether that means leading projects, mentoring others, or defining our technical direction. In this role, success means shipping systems that other teams rely on by default – because they make the platform safer, not because they're mandatory. We're especially interested in people who think like backend engineers but obsess over failure modes, capacity plan

typescriptpythonsql
View job →
A
Asana
📍 San Francisco• Full-time• $250K – $330K/yr
1mo ago

As the Head of Engineering Operations, you will be a key leadership partner to the CTO and the engineering leadership team, owning the operational backbone of our global R&D organization. In this role, you will serve as a critical bridge to other cross-functional groups, driving operational excellence, organizational effectiveness, and strategic initiatives. We are looking for a strategic leader who can translate company strategy into actionable engineering plans while building lightweight operating models that accelerate velocity and maintain high quality. You will empower our engineering organization to thrive and deliver exceptional impact at scale. This role is based in our San Francisco or Vancouver office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do and the teams with which you partner. If you're interviewing for this role, your recruiter will share more about the in-office requirements. What you’ll achieve Translate overarching company strategy into actionable engineering and operating plans in close collaboration with cross-functional partners. Design and build lightweight operating models, frameworks, and processes to optimize engineering execution and delivery, integrating evolving AI agentic engineering workflows. Hire, mentor, and lead a high-performing, global team of engineering operations professionals and Technical Program Managers to steer complex horizontal programs. Oversee financial and resource governance, including budget management, program spend, headcount strategy, and vendor/tooling allocations across public cloud and LLMs. Define and own the engineering metrics program (KPIs covering delivery, quality, reliability, efficiency, and capacity) to deliver data-driven insights to leadership. Serve as a trusted proxy and connective tissue for the CTO and R&D

aigorust
View job →
🔔

Get new capacity planning lead jobs by email

Daily job updates · Unsubscribe anytime