About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking an experienced Principal Hardware Diagnostics Engineer to design and develop diagnostics software used to monitor hardware health and diagnose system-level issues across Graphcore’s AI infrastructure platforms. This role focuses on building diagnostics agents, tools, and analytics frameworks that enable engineers and automation systems to identify, isolate, and resolve hardware issues across blade-level servers and rack-scale clusters. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Systems Engineering and Platform Validation team ensures Graphcore’s AI compute platforms are reliable, diagnosable, and operationally robust at scale. The team co
Jobs in United States
Systems Architect in United States
4,866 active opportunities · Updated October 2026
Showing
15 jobs
Explore current systems architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products—including next-generation ads experiences—that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring research breakthroughs into real-world systems at global scale. About the Role We’re looking for an experienced Software Engineer to help build the core infrastructure behind OpenAI’s monetization and ads systems. In this foundational role, you’ll architect and implement distributed systems that power OpenAI’s monetization stack—focusing on reliability, performance, privacy, and large-scale operation. You’ll work across backend, systems, and platform layers to define and implement 0→1 infrastructure, partnering closely with Product, Design, and Research to shape the future of monetized AI experiences. Your work will enable both internal and external teams to build on safe, scalable, and robust monetization primitives. This role is exclusively based across our San Francisco & Seattles sites. We offer relocation assistance to new employees. In this role, you will: Design and build the foundational backend and infrastructure powering OpenAI’s monetization and ads systems Architect large-scale distributed systems that
About the Team The Cooperative AI team is scaling OpenAI with OpenAI. We are building an AI powered knowledge system that evolves and learns as our products, systems and customers evolve. We leverage our state of the art models, technologies, and products (some external, some still in the lab) to assist or completely automate robust operations supporting both internal and external customers. We support OpenAI customers and internal partners globally, powering systems from customer support to integrity to product insights. We are a self-contained multi-disciplinary team, who enjoy a lightning fast feedback loop with customers at scale, some of whom sit just a few pods away. We iterate fast, and engineer for reliable long-term impact. We're constantly looking for the similarities and patterns in different types of work, and focus on building simple primitives, to apply world class knowledge to many domains. The work of this team exemplifies use of OpenAI technologies. We build systems so everyone can see the leverage that is possible with well designed AI-based implementations. We do this by working through internal use cases focused on Customers (specifically knowledge systems, automation systems, and automated agent systems) to prove impact, then we scale. About the Role We’re looking for Software Engineers who're passionate about blending production-ready platform architecture with new tech and new paradigms. You’ll push the boundaries of OpenAI’s newest technologies to enable interactions and automations that are not only functional, but delightful. We value proactive, customer-centric engineers who can get the foundational details right (data models, architecture, security) in service of enabling great products. In this role, you will: Own the end-to-end development lifecycle for new platform capabilities and integrations with other systems Collaborate closely with engineers, data scientists, information systems architects, and internal customers to understand th
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity We're looking for a Senior Director/Director, GTM Tools & Technology to own the systems and tooling that power Postman's entire go-to-market organization. This is a strategic leadership role sitting at the intersection of business operations, systems architecture, and AI innovation. You'll be the internal product owner for our GTM tech stack — from Salesforce at its core, to every platform that plugs into it, across both pre-sale and post-sale. You'll define the vision and drive the roadmap for how we use technology and AI to make our GTM teams faster, smarter, and more effective. This role demands both strategic range and operational depth: you'll partner directly with GTM executives across Sales, Marketing, Customer Success, Professional Services, Revenue Operations, and Finance, while leading a team to deliver on a high bar. The standard we're hiring to is high. We want someone who can architect for where Postman is going, not just maintain what exists — and who will lead meaningfully on AI adoption as a first-class mandate. What You’ll Do GTM Systems Strategy & Roadmap Define and execute the roadmap for
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Forward Deployed Infrastructure Engineers (FDIEs) build, operate, and maintain the infrastructure that powers Palantir’s platforms and production deployments. As an FDIE intern, you’ll work alongside full-time FDIEs to deploy and operate Palantir software across real production environments, automate manual processes, and develop novel solutions to infrastructure challenges using tools like Foundry and Apollo. Every day looks different — you might be debugging a distributed systems issue, building automation to replace a manual runbook, or designing infrastructure improvements that scale across multiple deployments. You’ll be treated as a full member of the team, with real ownership over the work you take on. Core Responsibilities As an FDIE intern, your responsibilities look similar to those at a small startup, with the resources, stability, and mentorship of an established tech company. You’ll work in small teams with minimal supervision and own end-to-end execution of real infrastructure projects. Your day might span discussing systems architecture with fellow engineers, debugging a production issue, building automation to eliminate a manual process, or deploying new Palantir products across production environments. FDIE interns are treated just like full-time engineers, with significant freedom and ownership over their work. Specifically, you can expect to: Deploy and operate Palantir software across production environments, including monitoring, alerting, configuration management, and upgrades Debug, improve, and optimize Palantir’s services and infra
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE Baseten's Compute org is in hyper growth. As it scales, the systems and workflows that keep supply and demand balanced across our GPU fleet need to get more sophisticated, and this role exists to make sure they do. Compute sits at the center of how Baseten allocates, forecasts, and manages the capacity that powers every customer inference request. The team that supports this work, C3, runs on a mix of internal tooling, manual processes, and systems that haven't fully kept pace with the scale of the problem. This role exists to close that gap. You'll design, build, and ship AI-powered workflows that give the Compute and C3 teams real leverage, automating the manual, repetitive, and error-prone parts of the capacity lifecycle so the team can focus on judgment calls that actually need a human. We want someone who can walk in, audit what exists today, identify what's missing or broken, and start shipping fast. You know when to reach for an existing internal tool and when to build something custom in Claude Code. You think two to three steps ahead about how the thing you build today fits into the broader capacity systems architecture tomorrow. And you bring a point of view on our stack, on what we should be building, and on where AI can do something existing tooling simply can't. RESPONSIBILITIES Ship AI-powered workflows for Compute and C3 : build the agents and automations that give capacity analysts, ops leads, an
We're looking for a Senior Operations Director for Revenue & Services who will serve as a trusted advisor to the Chief Commercial Officer, Professional Services leadership and other senior leaders. This role will shape how Professional Services operates, scales, forecasts, and performs, while driving alignment across Operations, Revenue, Sales and Finance. This role will connect services to broader business strategy, with a focus on cross-functional alignment, profitable growth and corporate results. WHAT YOU’LL DO: Shape and evolve the operating model for Revenue and Services, includingProfessional Services to enable scalable, predictable and profitable growth: forecasting methodology, capacity planning, utilization and margin management, and delivery-to-revenue alignment Set the strategic roadmap for services operations and operating infrastructure, including systems architecture (resource management, CRM integration, financial systems, data and reporting) and where to invest next Act as the strategic partner to Professional Services leaders translating business strategy and operating plans into performance targets. Represent services performance directly to executive leadership and the board Partner with President, Chief Commercial Officer, Finance leaders to evaluate profitable and growth opportunities, operating leverage, organizational capabilities, technologies and external partnerships. Lead deal desk strategy for services and deal governance, including pricing frameworks, SOW structuring, and margin guardrails, in partnership with Sales and Legal Partner in quarterly and annual planning for services capacity and revenue targets, and defend the plan directly to Finance and executive leadership Establish and drive a consistent operating cadence and process framework, including KPIs, business reviews, planning, forecasting, and reporting across the full customer lifecycle, identifying where PS operations should own versus influence
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's GTM org is in hyper growth. As it grows and matures, the needs of the GTM stack get more sophisticated with scale — and this team exists to stay ahead of those needs. Our GTM Engineering team plays a critical role in building and maintaining the connective tissue across Baseten’s GTM tools, processes, and user experience for the field. The GTM tooling landscape is changing fast, and the teams that win are the ones that adapt and iterate the fastest. This role exists to make sure Baseten is one of them. You'll design, build, and ship AI-powered workflows that scale our GTM functions as a competitive advantage. We want someone who can walk in, audit what we have, identify what we're missing, and start shipping fast. You know when to reach for Clay and when to build something custom in Claude Code. You think two to three steps ahead about how the thing you build today fits into the broader systems architecture tomorrow. And you bring a point of view — on our stack, on what we should be building, and on where AI can do something low-code tooling simply can't. RESPONSIBILITIES Ship AI-powered workflows for the field — build the agents and automations that give reps and managers real leverage, off-loading the manual and repetitive work. Reach for AI where it does something low-code can't. Get insights in front of reps — turn Salesforce, warehouse, and usage data into the dashboards, scores, and alerts reps
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. The Grid Service and Platform Engineering team is looking for a highly motivated and collaborative Software Engineering Manager. This role involves leading the engineering of mission-critical, tier 0 service infrastructure, the foundational data platform that powers Smartsheet at scale. You will oversee services that handle millions requests per day, operate at 99.999% availability, and deliver low-latency, high-throughput performance for millions of customers worldwide. We are an agile team that operates iteratively, focused on building high-quality software and adhering to rigorous operational best practices across complex, cross-functional distributed systems. This full-time position reports to the Director, Engineering and can be located in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. You Will: Manage one or more related teams of 6–10+ software engineers, driving development of tier 0 grid services and platform infrastructure that millions of customers depend on daily. Own and uphold 99.999% service availability targets across critical platform services, embedding reliability engineering, incident management, and on-call rigor into team culture. Help architect and guide technical vision to evolve low-latency, high-throughput service platforms capable of sustaining millions requests per day with predictable, consistent performance under load. Guide and mentor engineers on distributed systems architecture, scalability patterns, and platform best pr
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Streamlit in Snowflake (SiS) is our flagship app development platform, enabling data engineers, analysts, and developers to build and deploy interactive, data-driven applications directly within Snowflake. As Principal Engineer for the Streamlit in Snowflake team, you will be the top technical voice shaping how millions of users build and deploy apps on Snowflake's data platform. This role requires a rare combination: deep hands-on engineering expertise in Python/container runtimes, a systems architect's instinct for platform design, and the organizational influence to drive cross-team programs. You will define how SiS evolves from its container-native SPCS runtime and embedding/iframe SDK, to its developer experience, performance at scale, and integration with Snowflake's broader AI and data ecosystem. AS THE PRINCIPAL ENGINEER FOR STREAMLIT IN SNOWFLAKE YOU WILL: Define and own the architectural vision for the SiS platform spanning the SPCS container runtime, the warehouse runtime, the embedding SDK, developer tooling, and the Snowsight integration layer. Drive the platform's evolution toward its next-generation capabilities building on publicly shipping features like chromeless viewer URLs and IdP integration, and shaping the architectural direction for areas still in de
$295K – $380K/yr
About the Team The OpenAI Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role As a Senior Software Engineer, ML Systems & Training Infrastructure, you will be a deeply hands-on engineering force multiplier for the robotics team. You will help keep the training framework and surrounding infrastructure healthy, review and improve code quickly, debug failures across ML systems and infrastructure, and unblock researchers and engineers when the path from idea to working training job gets rough. We’re looking for people who love writing, reading, reviewing, and fixing code; who can get productive quickly in unfamiliar systems; and who bring strong practical judgment without a lot of ego or process overhead. This role will be based in San Francisco, CA and be expected in office 5 days per week and offer relocation assistance to new employees. In this role, you will: Review, improve, and clean up code across training frameworks and adjacent infrastructure. Identify risky or low-quality changes before they land, and raise the code quality bar without slowing the team down. Debug issues across ML training systems, GPUs, clusters, networking, and related infrastructure. Help researchers and engineers unblock broken training jobs, flaky workflows, and brittle internal tooling. Improve the reliability, maintainability, and usability of the robotics team’s training framework. Move quickly on practical engineering problems that directly affect team velocity. You might thrive in this role if you: Have strong software engineering fundamentals and excellent code review judgment. Have experience with ML systems, training fr
From $124K/yr
Datadog’s Implementation Services team helps customers implement and deploy Datadog quickly. Our team of architects leads the discovery, design, build, and launch of the Datadog platform to help customers accelerate time to value and get the most out of their investment. As a member of our team, you will be responsible for developing the technical roadmap for a customer’s implementation, and leading them through it every step of the way. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Design and guide execution of Datadog implementations, including gathering requirements, building technical architecture, and supporting deployment and launch Guide customer onboarding through co-development working sessions and assist when they run into roadblocks Write automation and design architecture for deployment templates Manage the operation of implementation projects Collaborate with other teams, including Customer Success, Technical Account Management, and Sales to deliver an exceptional onboarding experience Who You Are: Experience designing and delivering technical solutions for DevOps Monitoring or architecture systems Hands-on experience with cloud platforms (AWS, Azure, GCP) and compute technologies (EC2, serverless, VMware, VMs, Docker, Kubernetes). Experience with Config Management, IAC and CI/CD tools, including Ansible, Puppet, Terraform, Chef, Jenkins, Circle CI, GitHub Actions, Azure Pipelines, etc… Experience programming/scripting with any of the following: Java, Python, Ruby, Go, Node.JS, PHP, and .NET etc Background in security operations, including SOC management, SIEM tooling (e.g. Splunk Professional Services), and designing or supporting zero-trust networking architectures Successful track record with 5+ years exper
From $177.2K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . We're looking for a Staff Software Engineer to lead the technical direction of the backend systems powering Pinterest's AI-driven product experiences — Pinterest Assistant, visual editing and content creation tools and future LLM-based products. You'll design and ship backend systems while architecting the broader platform strategy that enables these experiences to scale across surfaces and teams. This is a hands-on leadership role where you'll move between deep technical execution, system-level architecture and cross-team technical leadership. What you'll do: Define the backend and platform architecture for AI-driven product experiences — visual-chat, AI image generation and editing, and agentic or LLM-based products — partnering with Engineering, Product, ML and UX leaders to shape the technical vision and roadmap. Architect end-to-end systems
Work Flexibility: Onsite Stryker is seeking a Staff Advanced Manufacturing, Automation/Software Engineer to join our Advanced Operations team, supporting the Medical Division - Acute Care Business Unit. In this role, you will lead the design, development, and deployment of advanced automation and production test systems for new product introductions. This role is critical to ensuring reliable, scalable, and compliant manufacturing for patient support and patient environment products. You will serve as both a technical owner and a supplier-facing leader, architecting systems internally while guiding external partners to deliver high‑quality automation solutions. You will have the opportunity to work on the design transfer of new products from research through development and into production. This is a hybrid role based out of Portage, MI. The team works onsite 4-5 days per week to support collaboration and project needs. What you will do: Automation System Architecture & Development Lead the full lifecycle of industrial automation and test systems from requirements, architecture, and design through implementation, validation, and release. Define system-level requirements encompassing mechanical, electrical, software, controls, and safety considerations. Develop and integrate control software, embedded interfaces, test sequences, and operator interfaces (HMI/SCADA). Ensure robust performance, maintainability, reliability, and alignment with design intent and manufacturing needs. Troubleshoot complex processes, software, and equipment issues; optimize system performance and uptime. Supplier & Equipment Vendor Leadership Manage automation and equipment suppliers, including capability assessments, technical reviews, process monitoring, and on‑site visits. Create clear, comprehensive
From $244K/yr
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Opportunity: Datadog’s Staff Engineers are our technical leaders operating at the forefront of technology, building solutions that take us through at least our next five years of growth. They do this in three major ways: As individual contributors, they bring world class technical abilities to deliver industry leading systems in areas such as data visualization, virtual runtime profiling, and planet scale streaming. As technical leaders they bring experienced technical breadth and communication skills to tackling design and architectural problems spanning the organization, charting the right course, then leading delivery. In both roles they participate in the staff engineering community and help us learn from what the industry is doing and what we've built before, and so improve company wide standards around software and systems engineering. Some examples of projects a staff engineer may own include designing and building a new data storage engine handling hundreds of millions of records per second, being the lead engineer building a new product like synthetics or profiling, or rebuilding a critical service to handle the next two orders of magnitude of scale. What You'll Do: Be the technical owner of multiple pieces of critical architecture in your area of the business Own delivery of the systems you architect from beginning-to-end, doing what it takes to get things shipped and at full scale in production Dive deep into performance of systems; inventing new approaches that bring efficiency at scale Who You Are: You have a BS/MS/P
Other cities to consider
More places hiring for this role
Get new systems architect jobs in United States by email
Daily job updates · Unsubscribe anytime