About the Team The ChatGPT Search Product Infrastructure team builds the foundational systems that power search experiences across ChatGPT. We develop the product infrastructure that connects models with search systems and other sources of real-time information, enabling ChatGPT to deliver timely, relevant, and trustworthy answers to users around the world. Our work sits at the intersection of product engineering, AI, and large-scale infrastructure. We build shared platforms and abstractions that enable product teams to independently develop, evaluate, and launch new search-powered experiences. These platforms provide the guardrails, testing capabilities, observability, and rollout controls needed to prevent reliability, scalability, quality, and latency regressions while supporting rapid product iteration. The team partners closely with: Post-Training on model launches, experimentation, and prompt optimization Search product verticals on new user experiences Inference on GPU efficiencies Indexing and Retrieval on the systems that identify and deliver relevant information Capacity/Fleet team to ensure optimal regionalized provisioning of GPUs and CPUs About the Role We are looking for an Engineering Manager to lead the team responsible for ChatGPT’s Search Product Infrastructure. You will set the technical and organizational direction for the systems that bring search capabilities into ChatGPT. You will guide architectural decisions across search orchestration, model and prompt integration, serving infrastructure, experimentation, observability, evaluation, and product integrations. You will balance immediate launch and product needs with the long-term reliability, scalability, latency, and maintainability of the platform. A central responsibility of this role is creating leverage for Search product verticals. You will lead the development of extensible platforms that allow those teams to independently build, test, and launch features without requiring ongoing invol
Jobs in United States
Capacity Strategy And Operations in United States
286 active opportunities · Updated October 2026
Showing
15 jobs
Explore current capacity strategy and operations jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
About the Team OpenAI's Industrial Compute organization is building and scaling the infrastructure required to support frontier AI. The Infrastructure Strategic Sourcing team connects technical and project requirements to supplier readiness, contracting, purchasing, equipment delivery, and portfolio-level risk visibility across owner-furnished contractor-installed equipment (OFCI), data center networking, rack systems and integration, fiber, cabling, optical interconnects, and related infrastructure. The team partners across Pre-Construction, Design, Construction, Electrical and Mechanical Engineering, Network Engineering, Hardware and Rack Delivery, Strategic Sourcing, Procurement, Legal, Finance, Accounts Payable, Logistics, and external suppliers. We build the operating mechanisms that keep sourcing decisions, purchase execution, long-lead equipment, network and fiber dependencies, rack readiness, and delivery commitments aligned to infrastructure schedules. About the Role We are seeking an Infrastructure Sourcing Operations Lead to own procurement operations across pre-construction, design, construction, and sourcing through purchase order issuance, while maintaining visibility through invoice resolution, production, logistics, delivery, installation, and readiness. The portfolio includes electrical and mechanical OFCI, networking equipment, rack systems and integration, fiber, cabling, optical interconnects, and other infrastructure required to bring capacity online. In this role, you will set priorities, make or escalate decisions that affect cost, supplier relationships, contractual position, and delivery schedules, and define the standards used by execution support for queue management, documentation, tracker maintenance, and recurring reporting. Success requires sound commercial and program judgment, operational rigor, systems thinking, and the ability to turn incomplete information across vendors, tools, and project teams into clear decisions, accountable
About the Team The Core Services organization builds and runs the mission-critical online services that product teams rely on in production. We own foundational distributed systems and platform capabilities that enable reliable execution, high-performance services, and large-scale file/data needs across our products. This team is distinct from developer infrastructure and data infrastructure—our focus is production service foundations and core runtime services. About the Role We’re hiring an Engineering Manager, Core Services to help lead teams responsible for highly reliable, high-scale distributed systems that sit on the critical path for OpenAI products. Your team will own foundational production systems that OpenAI’s product engineering teams build on. You’ll collaborate closely with product and infrastructure partners to ship reliable services quickly, and help scale systems and teams as OpenAI grows. You’ll partner closely with senior engineering leaders to scale the org, mature operations, and drive major platform initiatives. This role requires strong technical ability. You’ll be responsible for: Managing and growing a high-performing team of infrastructure engineers. Leading teams building and operating large, critical production platforms, including cluster reliability, scaling, and rollout safety. Building and operating mission-critical distributed systems with strong operational rigor (SLOs, incident response, capacity planning, reliability). Setting technical direction for platform foundations such as workflow/orchestration capabilities, large-scale file/blob/storage services, and core service foundations. Partnering with a broad set of stakeholders, including product engineering, adjacent infrastructure teams, and (where relevant) finance/cost partners. Coaching, mentoring, and developing engineers and emerging leaders. You might thrive in this role if you: Have significant experience leading teams that run mission-critical infrastructure in production
From $194K/yr
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Coinbase has built the world's leading compliant cryptocurrency platform serving over 73 million accounts in more than 100 countries. With multiple successful products, and our vocal advocacy for blockchain technology, we have played a major part in mainstream awareness and adoption of cryptocurrency. We are proud to offer an entire suite of products that are helping build the crypto economy and increase economic freedom around the world. There are a few things we look for across all hires we make at Coinbase, regardless of role or team. First, we look for signals that a candidate will thrive in a culture like ours, where we default to trust, embrace feedback, disrupt ourselves, and expect sustained high performance because we play as a championship team. Second, we expect all employees to commit to our mission-focused approach to our work. Finally, we seek people with the desire and capacity to build and share expertise in the frontier technologies of crypto and blockchain, in whatever way is most relevant to their role. JOB DUTIES Scale and grow the HR Engineering team by hiring, onboarding and training new analysts and engineers to support Workday and HR functional area Provide functional & technical leadership and mentoring to team members Supervise the day-to-day activities of our Workday instance Design user-friendly processes, guidelines, and documentation
Work Flexibility: Onsite Join Stryker in a role that combines continuous improvement leadership with project portfolio management to support operational excellence across the site. This position leads cross-functional initiatives, applies Lean and Six Sigma methodologies, and manages strategic projects that align with business and manufacturing objectives. The role offers the opportunity to work across functions, drive process improvement activities, and support the advancement of a high-performing operational environment. What You Will Do Lead cross-functional continuous improvement and project management initiatives from scope definition through implementation, ensuring projects are delivered on time and within budget. Facilitate value stream mapping, Kaizen events, Gemba walks, and problem-solving activities to identify waste, improve process flow, and support productivity targets. Manage the site project portfolio, track progress against commitments, and communicate performance through standardized project management processes. Analyze operational and manufacturing data, conduct time studies, labor modeling, capacity assessments, and line balancing activities to support process optimization decisions. Coach leaders and teams on Lean manufacturing principles, Daily Management, Leader Standard Work, Six Sigma methodologies, and continuous improvement tools while maintaining the site continuous improvement knowledge base and 12-month Kaizen roadmap. What You Will Need Required: Bachelor’s degree in Business Administration, Industrial Engineering, or a related field. Minimum 5 years of experience leading continuous improvement, Lean manufacturing, operational excellence, or project management initiatives in a manufacturing enviro
$116.2K – $182.4K/yr
Business Systems Manager Description - Position Summary Lead financial planning, manufacturing cost modeling, capital forecasting, and financial systems for MTD (MEMS Technology Development) and associated labs. Serve as a strategic partner to MTD labs, and executive leadership by translating financial and operational data into clear investment, resource, and long-range planning decisions. Core Responsibilities Financial Planning and Performance Lead annual budgeting, quarterly forecasts, Long-Term Planning, affordability reviews, capital prioritization, close activities, and executive financial reviews. Deliver variance, depreciation, headcount, scenario, and investment analyses with actionable recommendations. Provide timely reporting on financial results, operational trends, risks, and opportunities. Financial Systems and Data Governance Lead the vision, roadmap, and requirements for AMD (asset tracking), RaFT (forecasting tool), and related planning platforms. Govern cost objects, funding assignments, master data, reconciliations, and integrations across SAP, RaFT, AMD, eMagic (inventory finance system), and reporting tools. Improve data quality, usability, controls, and automation so systems remain trusted sources of record. Manufacturing Cost and Capital Leadership Lead manufacturing cost models and collaborate with Financial Analyst partners on ESC forecasts for wafer fabrication, development, and product platforms. Translate yield, capacity, and operating performance into cost, margin, funding, and investment insights. Help facilitate capital planning, depreciation and actuals reporting, purchase-order analysis, affordability assessments,
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As a Fab Engineer in the RDA & Metrology Team , you will collaborate with other equipment engineers, technicians, and application owners. Your responsibilities include overseeing the installation, modification, upgrade, and maintenance of manufacturing equipment, providing technical support to manufacturing equipment repair and process engineering organizations. You will define and write preventative maintenance schedules, maintain records on technical notices, upgrades, and safety issues, and study equipment performance to establish solutions that improve tool uptime. Responsibilities: Drive performance standards, critical measures, and accountability for operational decisions and results. Lead equipment issue resolution, troubleshooting efforts, and root cause investigations. Manage project priorities and recommend changes, continuation, or cancellation to meet fab objectives. Collaborate daily with Equipment Leads to align priorities, resolve issues, and optimize resource utilization. Plan for capacity and product mix requirements while enhancing equipment flexibility and performance. Maintain and improve equipment reliability through maintenance procedures, best-known methods (BKMs), program monitoring, and documentation updates. Partner with vendors and multi-functional teams to implement continuous improvement initiatives that enhance throughput, productivity, and equipment health. Leverage AI, analytics, and
Nurse Practitioner Company: The Boeing Company The Boeing Company’s Health Services organization is currently seeking a Nurse Practitioner to join their team in Everett, WA . The Nurse Practitioner will work in collaboration with others and partner with community health care providers, and suppliers such as Workers' Compensation Administrators to deliver services to employees, managers, and company organizations to determine how to increase the safety and well-being of Boeing employees. This position may require travel on a rare occasion to other Boeing sites in the Puget Sound area. Position Responsibilities: Provide medical services and / or medical consultation within the scope of an Advanced Registered Nurse Practitioner (ARNP) license in collaboration with company physicians and in support of company requirements Deliver services to employees, managers, and company organizations, and works with community health care providers, and suppliers such as workers' compensation administrators Provide clinical evaluation and urgent care assessment, treatment and/or referral for occupational and non-occupational medical conditions Determine fitness for duty, functional capacity, and medical restrictions for employees in relation to job demands, and provides assessment and consultation on accommodation and placement issues Perform health surveillance examinations and determines medical qualification for the occupational health examination program and international assignments Basic Qualifications (Required Skills/Experience): Must be licensed as an Advanced Registered Nurse Practitioner (ARNP) in the State of Washington or eligible
Upwork Inc.'s (Nasdaq: UPWK) family of companies connects businesses with global, AI-enabled talent across every contingent work type including freelance, fractional, and payrolled. This portfolio includes the Upwork Marketplace, which connects businesses with on-demand access to highly skilled talent across the globe, and Lifted, which provides a purpose-built solution for enterprise organizations to source, contract, manage, and pay talent across the full spectrum of contingent work. From Fortune 100 enterprises to entrepreneurs, businesses rely on Upwork Inc. to find and hire expert talent, leverage AI-powered work solutions, and drive business transformation. With access to professionals spanning more than 10,000 skills across AI & machine learning, software development, sales & marketing, customer support, finance & accounting, and more, the Upwork family of companies enables businesses of all sizes to scale, innovate, and transform their workforces for the age of AI and beyond. Since its founding, Upwork Inc. has facilitated more than $30 billion in total transactions and services as it fulfills its purpose to create opportunity in every era of work. Learn more about the Upwork Marketplace at Upwork.com and follow us on LinkedIn , Facebook , Instagram , TikTok , and X ; and learn more about Lifted at Go-Lifted and follow on LinkedIn . About the Role Upwork's COO and GM of Marketplace operates across one of the broadest organizational portfolios in the company - spanning Legal, Information Security, Marketing, Payments, Trust & Safety, Communications, Customer Support, Design, and Product. The Chief of Staff to the COO exists to make that model work: not as a coordinator, but as a true operating partner who extends the COO's capacity, sharpens decision-making, and holds together a complex, fast-moving organization. This role requires someone who has earned the confidence to walk into any room, speak with authority, and immediately earn cre
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Revenue Accounting Manager to own billing, contract review, and revenue recognition as Baseten scales. This is a hands-on, individual-contributor role for someone who wants full ownership of the revenue cycle at a company where deal structures are getting more complex: usage-based pricing, committed capacity, credits, multi-year enterprise contracts, and new business models coming online. You'll join Lauren, who built Baseten's quote-to-cash function from the ground up, to take on billing, contract review, and revenue recognition as deal volume and complexity grow. Having completed our first year-end audit, we're now focused on tightening contract review processes, close procedures, and reporting rigor to support the scale ahead. You'll partner closely with Lauren, FP&A, Sales, Legal, and Revenue Operations to make sure every deal is structured, billed, and recognized correctly from day one. Baseten is building the infrastructure layer for AI-native companies, and we're scaling quickly - in deal volume, contract complexity, and customer size. If you want real ownership over a growing function on a lean team, this role offers real scope. RESPONSIBILITIES Revenue Recognition and Technical Accounting Own revenue recognition under ASC 606 across all contract types, including usage-based/consumption arrangements, committed capacity deals, credits, multi-year contracts, and new business model
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Infrastructure Software Engineer at Baseten, you'll build and maintain components of our ML inference platform that powers production AI applications. You'll contribute to the core infrastructure, enabling developers to deploy, scale, and monitor ML models with high performance. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Develop infrastructure components for our ML inference platform using Python and Go Implement and maintain Kubernetes deployments for model serving Contribute to our inference orchestration layer for model deployments Build and enhance monitoring systems for model performance metrics Implement efficient resource management solutions for ML workloads Support infrastructure automation to improve ML deployment workflows Work closely with team members to implement technical solutions Help balance performance optimization with system reliability Participate in technical discussions around infrastructure improvements Learn and apply infrastructure best practices REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Go proficiency is a plus Working knowledge of Kubernetes and containeriza
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Cloud Platform Engineer, you'll envision and build robust systems and processes that ensure our infrastructure is scalable, reliable, and efficient. This can range from automating deployments and monitoring systems to optimizing performance and managing incidents. We all work closely with our users, learning from their past struggles in operationalizing ML, onboarding them onto our platform, and turning our learnings into ideas for improving Baseten. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Build and maintain scalable infrastructure to support the deployment and operation of machine learning models. Establish standards and best practices for reliability and performance across the infrastructure. Automate processes when relevant, particularly for managing CI/CD pipelines. Own products and projects end-to-end, functioning as both an engineer and a project manager, with a focus on user empathy, project specification, and end-to-end execution. Collaborate with cross-functional teams to understand project requirements and translate them into technical solutions. Mentor junior team members and contribute to knowledge sharing within the organization. Navigate ambiguity and exercise good judgment on tradeoffs and
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. ROLE We’re looking for a high-performing strategic finance professional to join our growing GTM Finance team. Our business grows with our customers' usage, which makes the finance function highly strategic at Baseten: growth, pricing, margin, and capacity decisions are business model decisions. You'll sit at the center of them, partnering directly with GTM leadership and reporting into a finance team with a seat at the table for the calls that shape the company's trajectory. This role is ideal for someone with 3 to 7 years of experience across strategic finance, investing, and/or investment banking who wants broad exposure to company-building inside a fast-scaling AI infrastructure company. Experience at a usage-based software company is a plus. RESPONSIBILITIES Own financial planning, forecasting, and budgeting processes for the GTM org Build and maintain financial models across revenue, S&M spend, headcount, and strategic bets Analyze the metrics that define a usage-based business – ARR, gross margin, consumption trends, retention, and GTM efficiency Partner with GTM leaders to set targets, evaluate growth initiatives, shape pricing, and design sales compensation Help prepare board materials, investor updates, and fundraising analyses Improve financial reporting, dashboards, and operational rigor so our infrastructure scales as fast as our revenue Work cross-functionally to turn ambiguous business questions into
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an OS / K8s Systems Engineer at Baseten, you’ll build the automation and systems that turn raw GPU hardware into production-ready compute. From provisioning to orchestration, you’ll own the software layer that makes our infrastructure reproducible, scalable, and reliable across data centers. This is a senior, hands-on role focused on building systems not operating them. You’ll work close to the metal designing OS images, building provisioning pipelines, and automating cluster bring-up from scratch. Your work will define how quickly we can turn new capacity into usable compute. EXAMPLE INITIATIVES Zero-to-cluster automation Build workflows that take new hardware from unprovisioned to fully operational cluster. Provisioning systems Design PXE-based or equivalent systems for imaging and lifecycle management. Reproducible infrastructure — Ensure clusters deploy consistently across data centers. RESPONSIBILITIES Own the end-to-end automation of cluster bring-up and lifecycle management. Build and maintain OS images, provisioning systems, and configuration pipelines. Deploy and operate cluster orchestration platforms (Kubernetes, Slurm, or similar). Design systems for reproducibility across sites and hardware generations. Automate upgrades, rollouts, and failure recovery. Optimize system performance, including GPU utilization and networking. Partner with hardware and network teams to validate and improve system b
Other cities to consider
More places hiring for this role
Get new capacity strategy and operations jobs in United States by email
Daily job updates · Unsubscribe anytime