About the Team Industrial Compute is building the infrastructure ecosystem that enables OpenAI to train and deploy increasingly capable AI systems at unprecedented scale. The organization operates across compute supply, demand, infrastructure, partnerships, and the physical and commercial systems required to make large-scale compute available. Industrial Compute Strategy & Operations serves as the connective operating layer across this ecosystem. The team works directly with senior leadership across Scaling, Finance, Partnerships, Research, and Infrastructure to translate ambiguous, high-impact challenges into clear strategies, scalable operating mechanisms, and decisive execution. This team is responsible for ensuring that OpenAI’s compute strategy evolves into durable competitive advantage by identifying systemic constraints, aligning stakeholders around critical decisions, and driving the operating mechanisms required to execute at scale. About the Role We are seeking a highly experienced Strategic Operations leader to help shape and operationalize OpenAI’s compute strategy across supply, demand, infrastructure, partnerships, and commercial strategy. This is a senior individual contributor role operating at the intersection of strategy, operations, infrastructure, and executive decision-making. You will work closely with compute leadership to identify the most consequential problems facing the organization, develop structured approaches to solving them, align stakeholders across the company, and drive initiatives from ambiguous concepts through execution. The role will span both strategic and operational work. You may develop long-range compute strategies and investment frameworks, evaluate build-versus-buy decisions, shape major commercial transactions, establish organizational planning mechanisms, or take ownership of a cross-functional initiative that does not have a clear organizational home. Success in this role requires exceptional judgment, analytical
Jobs in United States
Senior System And Manufacturing Codesign Architect in United States
2,019 active opportunities · Updated October 2026
Showing
15 jobs
Explore current senior system and manufacturing codesign architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the role We’re looking for an engineering manager to lead a team building software systems that detect and prevent harmful misuse of frontier AI models—before incidents occur. This is a builder’s role: you’ll lead engineers shipping production services, detection pipelines, and mitigation mechanisms that protect frontier model integrity and reduce high-severity misuse risk. While this work intersects with frontier model development, security and risk, we’re explicitly seeking someone with a software engineering foundation who is comfortable building reliable systems that can operate at billions of users scale. In this role you will: Lead a team of software engineers building detection + mitigation systems for frontier model misuse, with an emphasis on model IP protection / distillation detection and emerging risk surfaces from autonomous agents. Set the technical roadmap and execution strategy: prioritize, design, ship, iterate, measure impact. Build production systems: services, pipelines, tooling, instrumentation, and automation that scale with frontier model usage. Partner deeply with Research and Product to translate evolving model capabilities into concrete tests, signals, and mitigations that can be deployed at scale. Drive strong engineering fundamentals: architecture, reliability, monitoring, performance, and operational excellence. Hire and grow an exceptional team across backend, data systems, and applied ML engineering domains as needed. Anticipate what breaks at scale as agentic workflows become more capable. You might thrive in this role if you: Experience building systems in adversarial, fast-evolving environments Are comfortable with ambiguity and novelty Have experience adjacent to security (e.g., abuse prevention, fraud, integrity, platform defense, auth/identity, malware/spam, adversarial environments) Communicate clearly and build trust quickly with senior stakeholders—pragmatic, collaborative, and calm under scrutiny. Significant experience
About the Team The Strategic Finance team provides financial insights and guidance to support OpenAI’s long-term goals and strategies. We partner across the business to allocate and deploy our resources for the highest-impact outcomes.  Within Strategic Finance, the B2B team focuses on the financial performance of our products and GTM functions, ensuring tight alignment between financial objectives and company strategy. We partner with leaders across Product, GTM, Research, Partnerships, and Operations to: Drive operational planning, financial forecasting, and performance management. Provide decision-quality insights on product and financial performance to inform strategic resource allocation. Build the “0→1” financial foundations required to scale and accelerate growth. About the Role We are hiring a senior leader in B2B Strategic Finance to build and scale a new pillar within our finance organization. This is a highly visible role that reports into the Head of B2B Strategic Finance and supports some of our most critical executive stakeholders, including our COO, CFO, and CRO, among others on the B2B Leadership Team. This role is ideally based in our San Francisco HQ, but we are open to NYC and Seattle. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive B2B finance scale and rigor Build and scale core financial infrastructure across the B2B business, including forecasting methodology, variance management, and performance narratives that drive accountability and decision-making velocity. Lead consolidated planning across revenue, gross margin (including compute), and opex for annual budget, forecasts, regular business reviews, and long-range planning. Establish durable management reporting: KPI definitions, dashboards, month-end/quarter-end deliverables, and exec-ready readouts. Partner with Corporate FP&A, Accounting, and Finance Systems/Data to evolve processes and contro
About the Team The Infrastructure Engineering function sits within IT and is responsible for reliably building, deploying, and operating critical on prem and hybrid environments that power internal services and critical R&D environments. This is an early, high-leverage technical role focused on applying strong Site Reliability Engineering discipline to environments where uptime, safety, recoverability, and security are non-negotiable. This person helps replace bespoke, one-off infrastructure with standardized infrastructure-as-code building blocks that compound reliability and operational leverage as OpenAI scales. About the Role We are looking for an experienced Site Reliability Engineer working on security infrastructure to design, build, and operate reliable, secure, and scalable infrastructure that underpins identity, access, endpoint, and shared platform services across the company. In this role, you will be a senior technical owner for infrastructure and identity systems end to end, from architecture and implementation through policy enforcement, upgrades, recovery, and day-two operations. You will build durable, production-grade platforms that remove operational friction, enforce security by default, and enable teams to move faster with confidence. This role is well suited for a hands-on senior engineer who thrives in ambiguity, enjoys owning complex systems end to end, and raises the reliability and security bar by replacing fragile implementations with standardized, repeatable infrastructure. This role is based in our San Francisco HQ and requires in-office presence. In this role, you will: Design, build, and operate reliable infrastructure across on-prem, hybrid, shared, and product adjacent environments. Establish standardized infrastructure patterns that replace bespoke implementations with repeatable, auditable, secure-by-default systems. Own the lifecycle of critical infrastructure platforms, including provisioning, deployment, upgrades, patching,
Job Description Summary Topgolf is more than a venue business — we’re a technology-driven sports and entertainment platform. Toptracer, our tracking technology, powers ranges at hundreds of golf courses around the world, and our 100+ venues combine tech-enabled games, a strong food and beverage program, and large-scale event space, all in service of getting more people playing. We’re a large, multi-unit business backed by private equity, with significant growth ahead. Topgolf has invested in strong core financial platforms and infrastructure. The opportunity now is to unlock more of that investment — simplifying how work gets done, integrating our systems more deeply and building an accounting organization that is scalable, technology-enabled and increasingly focused on insight rather than transaction processing. We’re looking for an exceptional accounting leader to help make that happen. The Vice President, Accounting is Topgolf’s senior accounting executive, reporting to the Chief Financial Officer and leading a global organization of over 50 team members across financial reporting and technical accounting, accounts payable, accounts receivable, payroll, venue accounting and international operations, including venues in the UK and Toptracer. This is a role for someone who wants to design and run that organization, not simply operate the close — while continuing to deliver excellent accounting and controls. Enterprise Accounting Leadership Own the design and effectiveness of Topgolf’s accounting operating model — structure, process ownership, service delivery, systems, offshore resources and performance management. Maintain the integrity, accuracy and completeness of financial statements, and ensure compliance with U.S. GAAP, company policy, statutory requirements and internal controls. Oversee the monthly, quarterly and annual close and consolidation pro
About the Team OpenAI Finance helps ensure the organization is set up for success in pursuit of its mission of ensuring that artificial general intelligence benefits all of humanity. Within Finance, the Revenue Accounting, Technical team partners with Product, Strategic Partnerships, Business Development, Growth, Legal, Deal Desk, GTM, Strategic Finance, Revenue Accounting, and Finance Systems to enable scalable, operationally sound monetization. We advise on commercial and contract design, establish clear accounting positions under ASC 606, and translate those conclusions into repeatable processes, systems, controls, and reporting. About the Role We’re looking for a Manager of Revenue Accounting, Technical to drive the end-to-end technical revenue and operational workstream for OpenAI’s strategic commercial deals—from early structuring and deal-desk review through contracting, launch, and ongoing governance. Reporting to the Head of Revenue Accounting, Technical, you will serve as a key revenue advisor for complex and non-standard strategic deals, translating commercial objectives into revenue models that are economically sound, accounting-compliant, and operationally executable. You will partner with senior team members on the most novel or high-risk matters across strategic collaborations, cloud marketplaces, platform and distribution deals, licensing and revenue-share models, joint go-to-market initiatives, research collaborations, and other emerging commercial structures. You will engage from the earliest stages of a deal to shape economics, pricing, commercial terms, contract language, and product design; define the supporting data, billing, settlement, reporting, systems, controls, and revenue-recognition model; and drive cross-functional readiness through launch. After launch, you will help govern amendments, monitor whether the model is operating as intended, resolve execution issues, and evolve the approach as the partnership or business model changes. Thi
About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:
About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler’s high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You’ll Do (Role Expectations) Maintain h
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer on the Sharing team, you will lead large, multi-team initiatives with long-term technical vision and group-level impact. You will be expected to define and drive the platform-wide strategy powering how millions of users capture, share, and discover content on Roblox. In this role, you will set engineering standards, mentor senior engineers, and serve as a key architect of our long-term technical direction. You Will Drive Strategy & Execution: Own the outcome of complex, business-critical programs spanning several teams, often lasting years. Innovate at Scale: Develop and drive a multi-year technical vision for content creation and sharing, anticipating scale, technology, and business evolution. Elevate Reliability: Lead high-severity incident response across groups; drive durable systemic solutions that improve reliability and velocity. Architect Foundations: Regularly improve shared infrastructure and foundational systems, introducing frameworks that uplift development speed across the organization. Align Teams: Aligns multiple teams on shared technical direction, producing detailed design docs, phased roadmaps, and planning models that balance short an
From $183K/yr
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The Opportunity The Autonomous Freight Systems team is a brand new, AI-first engineering team in San Francisco, building Flexport's client-facing rates platform and self-serve freight booking experience from the ground up. We own two of the highest-leverage surfaces in the Client App: how clients see pricing and how they book freight without manual intervention. As a Senior Engineer, you will own significant technical domains end-to-end, not just features, but whole problem spaces. You will drive the engineering of systems that power rate visibility, pricing intelligence, and AI-assisted booking across ocean, air, and trucking. You don't wait to be told what to build next; you identify the highest-leverage problems, propose solutions, and see them through to production. You will be a technical force multiplier on the team: raising the bar on code quality, helping junior engineers grow, and shipping with a velocity and confidence that comes from deep ownership. This is a ground-floor opportunity to build the platform that moves Flexport from an assisted-sales model to a tech-run one for the long tail of our client base—turning the complexity of global trade into a s
From $192K/yr
Distributed Systems engineers at Datadog design, implement and run in production the foundational platforms powering our applications. Your data pipelines will ingest, store, analyze and query in real-time billions of events per second from companies all over the globe. The platforms are optimized for durability, high availability, low latency, internet-scale footprint and operability. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Build fault-tolerant, horizontally scalable solutions running in multi-tenant environments Write in Go, Java Rust or C++, amongst other languages Use Kafka, Redis, Cassandra, Elasticsearch and other open-source components Own meaningful parts of our service, have an impact, grow with the company Who You Are: 6+ years of experience You have a BS/MS/PhD in a scientific field or equivalent experience You have significant backend programming experience in one or more languages (Go, Java, Rust, C++) You have been exposed to working on problems (high durability / low latency /…) You can get down to the low-level when needed You care about simple designs and performance You want to work in a fast, high-growth startup environment that respects its engineers and customers You have demonstrated ability to use AI coding tools in day-to-day workflows and validate, critique, and refine AI-generated output. Bonus: you’re motivated to push the boundaries of how AI can improve software engineering best practices and contribute to building AI-enabled products. This job is available in various departments within our company; to conform to US export control regulations, some of these roles may require candidates to be eligible for any required authorizations from the US government. Datadog values peo
NVIDIA’s DGX Cloud organization is seeking a Senior Data Engineer to become part of its data team! We develop the reliable data foundation that supports fleet health, capacity, utilization, cost, reliability, and operational decision-making throughout DGX Cloud. Our platform supports engineering, operations, finance, and product teams managing and expanding large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We are looking for a practical engineer and technical lead to take charge of a key part of the Navigator data platform. We develop the systems that transform distributed infrastructure telemetry and operational data into dependable, managed data products that support fleet health, capacity, utilization, cost, and operational decisions. We are seeking a hands-on, platform-minded engineer to build and evolve the systems that turn distributed infrastructure telemetry and operational data into reliable, governed data products. You will work across ingestion, transformation, data quality, platform architecture, security, observability, and self-service consumption to help make Navigator and the DGXC data platform a dependable source of truth. We do expect strong engineering fundamentals, experience operating production systems, and the ability to learn new platforms and domains quickly. What you'll be doing: Own systems end to end. For example, work from ambiguous customer and operational needs through architecture, implementation, deployment, observability, incident response, and ongoing support. Construct data pipelines and products. Such as designing and maintain batch and streaming ingestion, transformation, reconciliation, and serving paths for fleet, capacity, utilization, cost, scheduling, and operational telemetry. Build shared libraries, workflow and DAG or equivalent experience abstractions to evolve the data platform. Develop deployment tooling, data
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Luau App Foundations team is responsible for the core infrastructure of the Roblox application. They bridge the gap between app and the performance-heavy Roblox game engine. This team is building the unified stack that powers mission-critical surfaces like Home, Avatar, and Social for millions of concurrent users. They are the ones who make it possible for "Web-style" development efficiency to exist within a high-performance C++ game engine. Why is this role exciting: Technical Pioneer: You will be writing libraries and modules using C++ inside a world class Roblox proprietary Game Engine. Systematic Impact: This is a "Force Multiplier" role. The frameworks and components you build will be used by dozens of other engineering teams to ship their features. Complex Problem Solving: You aren’t just building an app; you’re managing smooth data flow through a client that has to perform perfectly on a $100 Android phone and a $3,000 Gaming PC simultaneously. 0 to 1 Transitions: You will lead the charge in shaping some of the most crucial components of the App written in C++ inside the Game Engine. Key Challenges: Bridging Tech Stacks: The libraries you write sit between a modern UI (written in
From $345K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The mission of the search and discovery team is to connect a billion users with the best connections and content at the right time, covering an extremely diverse range of recommendations and search problems at high scale. This is fundamental to how we personalize the experience to drive engagement and retention for Roblox users. Our recommendation systems suggest all games that the users see and play with across all surfaces. As a Senior Engineering Manager for the Search & Discovery ML team, you will lead the architects of exploration for one of the world's largest immersive platforms. You will be responsible for the core algorithms that connect over 120 million daily active users with millions of 3D experiences, items, and social connections. This is a high-impact leadership role where your team’s models directly determine the growth, retention, and satisfaction of the Roblox community. You will move beyond traditional recommendation systems into the realm of multimodal, agentic, and generative discovery . You Will: Lead and grow a high-performing team of machine learning engineers, fostering a culture of ownership, collaboration, and engineering excellence. Establish a lo
Other cities to consider
More places hiring for this role
Get new senior system and manufacturing codesign architect jobs in United States by email
Daily job updates · Unsubscribe anytime