ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Finance Systems Lead to own and evolve the systems that support Baseten's finance operations, including ERP, planning tools, and the broader finance and adjacent systems stack, spanning HRIS/payroll, CPQ, and commissions, as we implement, integrate, and automate across them. As Baseten scales, so does the number of systems finance depends on and touches, from the ERP to planning to compensation and quote to cash tooling. This role exists to bring a dedicated owner to that stack, someone who can lead implementations (partnering with contractors or vendor CSMs for hands-on technical configuration where needed), keep systems integrated and clean as we grow, and find automation opportunities that save the team time. This is a highly cross-functional role. You will partner with Accounting, FP&A, People, Sales and Revenue Operations, Legal, and IT to understand how each team uses its systems, and design integrations and workflows that serve the business without creating more manual work. RESPONSIBILITIES Core Finance Systems Own finance systems strategy, configuration, and roadmap, partnering with contractors or vendor CSMs on hands-on technical configuration as needed Lead the implementation and ongoing administration of our finance systems and tools Own finance system integrations end to end, ensuring clean data flow between ERP, planning, billing, and other systems, partnering with contrac
Jobs in United States
Lead Systems Engineer in United States
2,434 active opportunities · Updated October 2026
Showing
15 jobs
Explore current lead systems engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will work on the systems software strategy and execution that brings new AI silicon from first power-on to a fully integrated system running production-representative models at expected functionality and performance. You will define how software exercises and validates compute, memory, interconnect, and I/O subsystems, then build the diagnostics, automation, and observability needed to find issues quickly. This role sits at the center of silicon, firmware, platform, systems, and workload teams. You will turn hardware specifications and performance targets into an end-to-end bringup plan, drive cross-functional debug, and establish the stress and regression infrastructure that makes each new platform reliable across operating environments. In this role, you will: Contribute to the end-to-end software bringup and validation strategy for new silicon and first-party systems. Define software-driven test coverage across compute, memory, interconnect, I/O, and their system-level interactions. Build diagnostics, test automation, telemetry, and regression infrastructure that accelerate first-silicon learning and issue isolation. Lead bringup from initial silicon arrival through board and system integration, docking, runtime enablement, and model execution. Design stress tests that characterize reliability, performance, and stability across workloads and operating conditions. Translate architecture specifications and performance models into measurable acceptance crit
From $192K/yr
Coordination Systems provides foundational distributed systems building blocks for internal Datadog platforms. Our services cover sharding, consensus, resource protection, configuration distribution, and much more. We are looking for a manager to lead the Coordination Systems - Storage team. This team provides essential configuration storage and distribution systems that are depended upon by almost every service and pod at Datadog. We power critical runtime configuration (e.g. feature flags), complex control planes (e.g. dynamic sharding configuration), and much more. Storage is one of four subteams within Coordination Systems. If successful, the candidate will have opportunities to lead other growing and impactful areas such as Resource Protection. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: (Describe role responsibilities here/max 6 bullets) Lead a core team of 5 engineers (distributed, with majority in NYC) Lead ceremonies, prioritize and delegate project Stay hands-on with the code, e.g. isolated features, small remediations, investigation follow ups Stay actively involved in operations, incidents, root cause analysis, etc. Constantly promote a culture of operational excellence, organizing gamedays, conducting operational reviews, staying proactive with reliability Who You Are: (Describe role qualifications here/max 6 bullets) Strong distributed systems skills, able to understand and account for a variety of failure modes, well-versed in end-to-end o11y, validation testing, simulation setup, etc. Worked on platform teams before, providing critical infrastructure to internal stakeholders Experienced in handling significant incidents, both as a responder and follow-up ow
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. As a Support Systems Lead at Replit, you'll build and own the systems that let support scale as fast as the product does. You'll own the tools our team works in every day, the technical setup behind self-service and AI-driven support, and the infrastructure that keeps all of it current as Replit ships. Replit is at the forefront of AI-driven software development, and how we support customers is constantly evolving. You'll shape how our support systems adapt to new products, new surfaces, and AI-assisted workflows, operating effectively in ambiguity and turning ad hoc fixes into infrastructure the whole team can rely on. You'll combine hands-on technical depth with systems thinking to keep builders moving, whether they get unblocked through self-service, an AI agent, or a person. This is an individual contributor role to start, with room to grow and build out a team as support scales. IN THIS ROLE YOU WILL: Configure and maintain Zendesk and the surrounding support stack, from the day-to-day workflow and automation setup to business rules and permissions, keeping it able to flex and scale as needs change. Build and maintain assignment logic, queues, tagging and taxonomy, and escalation paths, keeping them running cleanly as volume and workflows change. Set up and maintain support tooling, workflows, and access across internal agents, outsourced vendors, and regions, keeping the systems working for each group as the stack changes. Partner with Engineering on the technical requirements for self-service and in-product support surfaces, and build the entry points, help widgets, and routing behind them. Spot the repetitive steps in agent and admin workflows before they become bottlenecks, and build the automations, bulk acti
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. As a Support Systems Lead at Replit, you'll build and own the systems that let support scale as fast as the product does. You'll own the tools our team works in every day, the technical setup behind self-service and AI-driven support, and the infrastructure that keeps all of it current as Replit ships. Replit is at the forefront of AI-driven software development, and how we support customers is constantly evolving. You'll shape how our support systems adapt to new products, new surfaces, and AI-assisted workflows, operating effectively in ambiguity and turning ad hoc fixes into infrastructure the whole team can rely on. You'll combine hands-on technical depth with systems thinking to keep builders moving, whether they get unblocked through self-service, an AI agent, or a person. This is an individual contributor role to start, with room to grow and build out a team as support scales. IN THIS ROLE YOU WILL: Configure and maintain Zendesk and the surrounding support stack, from the day-to-day workflow and automation setup to business rules and permissions, keeping it able to flex and scale as needs change. Build and maintain assignment logic, queues, tagging and taxonomy, and escalation paths, keeping them running cleanly as volume and workflows change. Set up and maintain support tooling, workflows, and access across internal agents, outsourced vendors, and regions, keeping the systems working for each group as the stack changes. Partner with Engineering on the technical requirements for self-service and in-product support surfaces, and build the entry points, help widgets, and routing behind them. Spot the repetitive steps in agent and admin workflows before they become bottlenecks, and build the automations, bulk acti
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a rock star leader for our AI DevEx team. This team plays a crucial part in transforming how the Snowflake product is developed, evaluated and supported. The job is no less than building the ecosystem that powers our AI-pilled engineers and our agentic organizations and reinvents the SDLC for Snowflake. As the manager for this growing team, you will lead the group of engineers working to reinvent how code is written, reviewed and runs in production to usher scale and velocity for the AI era. With context as the new source code, you will make it possible for every engineer at Snowflake to deliver through intent and direction. In this role, you will: Lead, coach, and grow the ES AI DevEx team while creating a high-energy, cohesive environment with strong planning, ownership, and career development. Deliver the vision and roadmap for transforming the SDLC for all Snowflake engineers. Drive measurable improvements in code review latency, ensure Snowflake developers can use the latest harnesses and models, and create the golden paths that agentic codebases are built upon. Act as an agent of clarity in a space that is moving fast, making high quality decisions quickly by keeping up with the latest developments in the industry. Own operational excellence for the area
Senior Low Observables (LO) Mission Systems Integration Engineer Company: The Boeing Company The Senior LO Mission Systems Integration Engineer will lead design, integration, analysis, test, and production activities for low observable materials, structures, apertures, sensors, and radomes for mission systems. The role requires strong technical leadership, independent execution of complex tasks, advanced use of computational electromagnetic (CEM) solvers, and the ability to produce and defend technical results with minimal oversight. Key Responsibilities Lead design, modeling, and analysis efforts for LO materials, coatings, structures, and integrated mission systems to meet signature reduction and performance requirements. Independently apply CEM solvers (e.g., FEKO, HFSS, CST, WIPL‑D, xFDTD, or equivalent) to model antennas, apertures, radomes, and LO treatments; perform design optimization and trade studies. Drive integration of LO treatments with mechanical, thermal, and sensor/antenna system constraints; identify and mitigate producibility, testability, and sustainment risks. Plan, execute, and oversee laboratory and field tests (anechoic chamber measurements, antenna/radome characterization, RCS measurements); lead data collection, reduction, and validation activities. Develop and validate predictive models, reconcile simulations with measurements, perform uncertainty quantification, and provide actionable recommendations to engineering and program leadership. Produce high-quality technical documentation: test plans, test reports, technical memoranda, design reviews, and customer briefings. Mentor and provide technical oversight to junior engineers and technicians; ensure adherence to engineering processes, qual
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for Forward Deployed Engineers on our engineering team who want to work at the intersection of deep infrastructure work and direct customer impact. As an FDE, you'll partner with leading AI companies and foundation labs on cloud architecture, networking, storage, containerization, sandboxing, and more — helping them design and ship production infrastructure on Modal's platform. The FDE team today includes world-class software engineers, computational scientists, ML engineers, and former founders. We're looking for people with strong engineering fundamentals, deep curiosity across the infrastructure stack, and energy for working directly with customers on hard problems. You will: Work hands-on with companies like Suno, Lovable, Cognition, and Meta to architect and deploy massive-scale production workloads on Modal Lead technical discovery and architect
From $399.4K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Why Roblox Studio? Roblox Studio is the creation engine behind millions of immersive 3D experiences built by creators around the world. It serves everyone from first-time developers to large professional studios, combining a powerful real-time engine with a deeply expressive toolchain. As creation on Roblox grows more sophisticated, Studio must continue evolving as a world-class development environment for interactive 3D experiences. That requires deep investment in the technical systems that underpin the IDE experience from scripting, debugging, architecture, and core IDE capabilities including performance. We are looking for a leader who can raise the bar on these foundational systems and help shape the future of Studio as a high-quality, extensible, and performant creation environment. Familiarity with creator tools is important, and experience with AI-assisted creation is a plus. The Role We are seeking a Director of Engineering for Studio Systems to lead a critical area of Roblox Studio’s technical foundation. In this role, you will be accountable for the architecture, execution, and long-term evolution of core systems that power Studio’s IDE and developer workflows. This includes area
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking an experienced AI Systems Reliability Engineer to help define, build, and maintain the infrastructure and processes that ensure the reliability, scalability, and performance of Postman’s AI-powered API and agentic systems in production. This role focuses on monitoring, availability, incident response, and automation to support AI services and tools trusted by millions of developers globally. What You’ll Do Develop and manage reliability metrics (SLOs) for AI-driven API services and agentic AI platform features Implement comprehensive observability and monitoring systems for real-time performance and fault detection Design and drive automated failover, recovery, and incident response strategies for high-availability AI infrastructure Optimize resource utilization, particularly GPU/accelerator efficiency, ensuring cost-effective AI system operation Collaborate closely with engineering, platform, and product teams to align reliability efforts with broader organizational goals Lead efforts to build internal tooling and automation focused on AI system stability and operational excellence Drive continuo
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a talented Tech Lead Manager (TLM) to lead Snowflake’s System Under Test (SUT) team, responsible for evolving how Snowflake engineers test the Snowflake product locally and at scale in CI. The SUT empowers Snowflake engineers by delivering a reliable, low-latency and cost-efficient developer experience across a high-growth, high-demand surface area. As the TLM for SUT, you will lead a small and highly technical team at the intersection of CI, developer infrastructure, and product engineering. You will set direction, drive execution, and partner broadly across Engineering Systems and product teams to deliver a more reliable, faster, and more maintainable test platform for Snowflake’s engineers. In this role, you will: Lead, coach, and grow the SUT team while creating a high-energy, cohesive environment with strong planning, ownership, and career development. Own the roadmap and execution for SUT rollout across development environments, CI and AI workflows. Drive measurable improvements in startup reliability, latency, and cost, using clear SLOs, dashboards, and operational metrics to guide decisions and raise the bar on execution. Serve as the technical anchor for the SUT domain, shaping architecture and guiding the evolution from legacy systems to a composable
Leidos is seeking a Test and Integration Engineer to lead cutting-edge testing and validation efforts within the Undersea Systems Division (USD) . This role is a unique opportunity to drive innovation in testing underwater vehicle systems, maritime sensors, subsea telemetry, and ISR solutions that support critical defense and national security missions. You'll contribute to a multidisciplinary team focused on testing, integrating, and validating advanced maritime technologies, ensuring system reliability and performance from prototype design to full-scale system deployment for ongoing Navy missions. Leidos’ Undersea Systems Division is a recognized leader in C4ISR technologies, delivering innovative, mission-critical solutions across sensor networks, unmanned systems, and tactical platforms . We’re known for achieving “industry firsts” in the most challenging maritime domains. Join us and be part of a world-class team delivering unmatched solutions for today's most pressing maritime missions. Why Join Us? Make an Impact : Your work will directly support U.S. maritime dominance and national security. Lead Innovation : Be at the forefront of applying innovative technology and autonomy to real-world maritime systems. Work with Experts : Collaborate with a top-tier team of engineers, scientists, and technicians located across the U.S. Shape the Future : Influence both the strategic and tactical direction of next-generation subsea technologies. What You’ll Do Test and Validate: Write, develop and execute comprehensive test plans, procedures, and protocols to ensure system functionality, reliability, and compliance with requirements. Integrate systems and conduct hands-on testing: Write, develop, and execute integration plans to bring complex systems together. Perform field</b
About the Team The Consumer Devices team at OpenAI builds end-to-end hardware and software systems that bring AI into the physical world. We work at the intersection of custom silicon, embedded systems, operating systems, cloud services, mechanical engineering, electrical engineering, and product design to deliver reliable, production-ready devices at scale. Within Consumer Devices, Hardware Engineering eXperience, or HEX, is a new bootstrapped team building the environments, applications, compute, product-data systems, and workflows that let hardware engineers do their work without needing to troubleshoot the machinery underneath. HEX owns virtual engineering environments, HPC/GPU compute, storage, networking, licensing, MCAD/ECAD/CAE applications, PLM, product data, automation, validation, and support as one connected system. About the Role As a Staff PLM & Engineering Applications Engineer, you will be one of the first technical builders of HEX and the primary counterpart to the HEX lead. You will own the engineering-application and product-data side of the hardware engineering experience, with an initial focus on NX, Teamcenter, licensing, parts import, integrations, packaging, validation, and user workflows. This is not a traditional Teamcenter administration role and not a Corporate IT application-support role. You will take complex, fragile workflows and turn them into reliable engineering systems. This role is highly hands-on and systems-oriented. You will not inherit a mature environment and support queue. You will help build a fresh one, replacing manual setup guides, tribal knowledge, repeated support issues, and team handoffs with tested automation and reliable workflows. In This Role, You Will Own the technical architecture, deployment, configuration, integration, validation, and long-term operation of NX and Teamcenter. Build reliable workflows for parts import, product-data migration, metadata quality, BOMs, revisions, lifecycle states, and releas
From $272K/yr
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Opportunity: Datadog’s Senior Staff Engineers are technical leaders operating at the forefront of large-scale systems design, building the infrastructure that will support our next five years of growth and beyond. They do this in three major ways: As individual contributors, they bring world-class technical depth to build industry-leading systems in areas such as observability data platforms, distributed query engines, and real-time event streaming at global scale. As technical leaders, they apply broad architectural perspective and deep systems thinking to align design decisions across teams and domains. They work across complex, multi-team problem spaces to define long-term technical direction, drive large-scale initiatives forward, and ensure consistent execution. As engineering stewards, they play a key role in evolving our systems and engineering culture. They actively participate in Datadog’s senior technical community, bringing external insights and internal experience to elevate engineering standards and mentor the next generation of technical leaders. Examples of projects a Senior Staff Engineer may lead include designing and launching a new distributed data storage engine capable of handling hundreds of millions of records per second, building the real-time infrastructure behind a new observability product, or re-architecting a core service to support exponential growth in throughput and complexity. What You’ll Do: Be the technical owner of multiple critical systems or architecture areas, often spanning several t
The Applications Development Technology Senior Lead Analyst is a senior level position responsible for establishing and implementing new or revised application systems and programs in coordination with the Technology Team. The overall objective of this role is to lead applications systems analysis and programming activities. Responsibilities: Lead the architecture, design, development, and delivery of enterprise UI applications. Strong hands-on experience with Angular and Ext JS for developing scalable and responsive user interfaces. Strong experience with Java and Spring Boot for designing and developing backend services and APIs. Experience deploying, supporting, and troubleshooting applications in ECS/containerized environments . Define application architecture, technical standards, reusable components, and development best practices. Provide technical leadership to development teams, including design reviews, code reviews, performance optimization, and troubleshooting . Strong understanding of CI/CD pipelines , including automated build, testing, deployment, and release processes. Hands-on experience with GitHub , including source-code management, branching strategies, pull requests, code reviews, and integration with CI/CD pipelines. Strong knowledge of open-source technologies and frameworks , with hands-on implementation experience. Ability to evaluate and select appropriate open-source libraries, frameworks, and tools , considering security, licensing, maintainability, vulnerabilities, and enterprise standards. Drive application modernization, technical improvements, and adoption of engineering best practices. Work closely with architects, developers, infrastructure teams, product owners, and business stakeholders to deliver solutions successfully. Prov
Other cities to consider
More places hiring for this role
Get new lead systems engineer jobs in United States by email
Daily job updates · Unsubscribe anytime