Jobs in United States

Distributed Systems Engineer in United States

426 active opportunities · Updated October 2026

Explore current distributed systems engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As the Senior Director of Solutions Architecture for the Americas at Cohere, you will own the US and Canada commercial Solutions Architecture function. You will lead the team that turns enterprise interest in agentic AI into deployed, production systems, and you will be accountable for the technical win in the most competitive AI market in the world. The United States and Canada are our largest commercial opportunity, and this seat owns how we win them. You will take an established, distributed team of strong technical people and raise what it can do — setting the bar and establishing the operating rhythm that lets a team of generalists run consistent, industry-fluent plays at enterprise scale. You will set direction for the function, sit on the Solution Architecture leadership team alongside the regional leaders for EMEA and Asia Pacific, and contribute to company-wide decisions with your peers across Sales, Product and Engineering. In this role, you will: Lead and Scale the Organization: Build, coach and develop a high-performing Solutions Architecture organization across the US and Canada, and grow the senior technical talent

GitRestAIGo
Z
📍 Bellevue, Washington, United States· Full-time
✓ High-confidence listing

From $180K/yr

Quick readStrong listing-quality and freshness signals

Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Senior Staff Rust Developer to join our Platform Convergence Team. This is a hybrid role based in San Jose, CA reporting to the Sr. Director, Software Engineering. Join us to build a new platform from the ground up that can scale hundreds of millions of users with high reliability and low latency. You will design and implement distributed system and core infrastructure components while collaborating closely with various stakeholders. What you’ll do (Role Expectations) Design and build a low-latency, high-throughput data forwarding plane using Rust, leveraging its async/await model for efficient I/O and service-oriented infrastructure Develop distributed, scalable systems with a focus on concurrency, fault tolerance, and messaging Implement and maintain gRPC-based APIs and services to integrate forwarding plane capabilities with control and orchestration layers Optimize system

AWSKubernetesCI/CDGit
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Summary: You'll own ClickUp's social strategy and build the systems to execute it at scale. You deeply understand what wins on each platform: the formats, the hooks, the timing, the tone. But you're not just a strategist who hands off a plan. You build AI-powered workflows and automation to operationalize your strategy so it runs continuously, learns from data, and scales beyond what any team could do manually. You're a social-native operator who builds systems, not an engineer who dabbles in social. Responsibilities: Own platform-native social strategy across X, LinkedIn, TikTok, and emerging channels: define what ClickUp's voice, format, and engagement approach looks like on each, tailored to what works on that platform Develop and execute content and engagement strategies that drive measurable growth in reach, engagement, and audience quality Identify trends, conversations, and cultural moments worth engaging with, and move fast enough to capitalize on them Build AI-powered systems and automated workflows to execute social strategy at scale: monitoring, engagement, response, and content distribution Create feedback loops between social performance data and strategy; use signal to iterate what gets made and how it gets distributed Own proactive engagement: identify and engage relevant conversations, mentions, and opportunities using AI-powered monitoring and automated response workflows Develop automated systems that handle routine engagement while escalating high-value or brand-sensitive conversations to humans Own execution end-to-end: strategy through measurement, with clear accountability for out

AWSMachine LearningAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Future of Computing Research team is an applied research team in the Consumer Devices group focused on developing new methods and models to support our vision as we advance forward in our mission of building AGI that benefits all of humanity. About the Role As a Technical Lead on the Future of Computing Research team, you will work together with both the best ML researchers in the world and the greatest design talent of our generation to push the frontier of model capabilities. This role is based in San Francisco, CA. We follow a hybrid model with 3 days a week in the office and offer relocation assistance to new employees. In this role, you will: Evaluate and select silicon platforms (GPUs, NPUs, and specialized accelerators) for on-device and edge deployment of OpenAI models. Work closely with research teams to co-design model architectures that meet real-world deployment constraints such as latency, memory, power, and bandwidth. Analyze and model system performance, identifying tradeoffs between model design, memory hierarchy, compute throughput, and hardware capabilities. Partner with hardware vendors and internal infrastructure teams to bring up new accelerators and ensure efficient execution of transformer workloads. Build and lead a team of engineers responsible for implementing the low-level inference stack, including kernel development and runtime systems. Run through the necessary walls to take nascent research capabilities and turn them into capabilities we can build on top of. You might thrive in this role if you: Have experience evaluating or deploying workloads on GPUs, NPUs, or other specialized accelerators. Understand the performance characteristics of transformer models, including attention, KV-cache behavior, and memory bandwidth requirements. Have designed or optimized high-performance compute systems, such as inference engines, distributed runtimes, or hardware-aware ML pipelines. Have experience building or leading teams work

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and inform next-generation infrastructure design. About the Role We are seeking Performance Modeling Engineers to develop and apply modeling tools that evaluate AI system performance and inform architectural decisions. In this role, you will work closely with the Performance Modeling Lead and partner teams to analyze system behavior, run simulations or analytical models, and help quantify tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks and applying them to real-world questions that impact system design and vendor decisions. This role is well-suited for engineers with strong software or modeling backgrounds who are interested in developing deeper expertise in system architecture and AI infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Develop and maintain performance modeling tools and frameworks. Build models to evaluate system behavior across: compute, memory, and interconnect subsystems distributed system scaling and bottlenecks. Run simulations and analytical models to support architectural tradeoff analysis. Collaborate with performance modeling lead and system architects to answer forward-looking design questions. Analyze and interpret modeling outputs, translating results into actionable insights. Validate models against real system measurements and workload behavior. Contribute to improving modeling fidelity, usability, and scalability. Qualifications Strong software engineeri

AWSRestAIRust
C
📍 Hartford, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The Senior Software Engineer / Technical Lead (AI & Automation) will provide technical leadership for the design, development, modernization, and support of critical applications supporting Prior Authorization Operations (PAOps) under PBM line of business. This role will be responsible for building and maintaining scalable, cloud-native solutions that enable intelligent workflow automation, AI-driven decisioning, document processing, and business process optimization. The ideal candidate is a hands-on technical leader with strong software engineering and cloud architecture expertise, coupled with practical experience implementing Generative AI, Agentic AI, and Large Language Model (LLM) solutions in production environments. This individual will collaborate closely with Data Engineering, Data Science, Product, and Business teams to deliver highly available, secure, and scalable applications while driving innovation through AI-powered solutions. Key areas of focus include: Application architecture, development, and production support Cloud-native engineering and platform modernization Microservices and distributed systems AI/GenAI, Agentic AI, and LLM-based solutions Event-driven and streaming architectures Engineering best practices, mentoring, and technical leadership R

PythonSQLGCPDocker
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response. What you'll be doing: Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow) Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes Drive upstream contributions to OVN-Kubernetes and related open-source projects Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure Drive reliability through incident management, resource monitoring, and performance tuning<

PythonAWSAzureGCP
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

We are looking for a 100% hands-on Storage Services Software engineer to join the block storage group. You will be a member of a team that builds the next-generation block storage capabilities and architects a proprietary distributed file system solution from its inception. You will work closely with a variety of teams and architects including the networking team, and external customers. You will take part in defining the software architecture and implementation of the most advanced storage services! Services that will need to meet extreme performance and scalability demands! We have crafted a team of extraordinary people stretching around the globe, whose mission is to push the frontiers of what is possible today and define the platform of tomorrow. At NVIDIA, we work, think and learn as a team. We thrive in a deeply strong environment, and we're passionate about a culture that demands innovation and the highest standards. The rewards are sweet and include collaborating with some of the smartest people in the industry, an aggressive compensation plan that rewards top performers, and the opportunity to work on products that transform the way people work and play. What you’ll be doing: 100% hands-on coding role in C language, Kernel and Userspace Access advanced AI tools and a token budget for code development provided by NVIDIA, the world's AI factory leader. Research, design, implement and test, new and existing, distributed storage services and features of NVIDIA’s block and file storage solution, in both Host and DPU environments. Acquire understanding of the algorithms, the technicalities and the interaction with other components across NVIDIA’s block and file storage ecosystem. Analyze and solve challenging bugs and customer cases in la

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Applied AI team safely brings OpenAI's technology to the world. We released ChatGPT, Plugins, DALL·E, and the APIs for GPT-4, GPT-3, embeddings, and fine-tuning. We also operate inference infrastructure at scale. There's a lot more on the immediate horizon. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. We serve end-users directly through ChatGPT, and serve developers through our APIs, which power product features that were never before possible. About the Role The Engineering Acceleration team designs, builds and maintains the foundational systems that engineers use to build ChatGPT and the API. This is a fast-growing team and you will get a chance to own and define the strategy, vision, and plan for how to increase developer productivity. In this role, you will: Drive the design, development, and implementation of tools, systems, and processes that accelerate engineering velocity, reduce manual effort, and increase the quality of output. Use our latest AI tools to re-think how we can be the most productive team in the industry. Work closely with various teams within OpenAI to understand their workflows, challenges, and needs, and ensure the tools and systems built by the Engineering Acceleration team address these requirements. Bring new features and research capabilities to the world by partnering with product engineers to lay the necessary technical foundations. Guide and advise product engineering teams on best practices for ensuring observable, scalable systems. Like all other teams, we are responsible for the reliability of the systems we build. This includes an on-call rotation to respond to critical incidents as needed. You might thrive in this role if you: Have 5+ years of experience in engineering, including 3+ years of experience in infrastructure building tooling for developers. Have experi

PythonAWSKubernetesRest
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Snowhouse Foundation team builds our globally distributed data warehouse. We manage a vast array of petabyte scale data sets that are continuously ingested, processed and replicated from across all Snowflake environments and external data sources. Snowhouse powers all of Snowflake’s core business, engineering and data science needs and provides customers with full visibility into their account activities, usage, and resource consumption from all their global environments. The team is investing in multiple critical areas, including a pipeline authoring platform, high performance/high efficiency data export, ingestion and data layout. Our team is also responsible for a fundamental product for Snowflake’s customers: the Snowflake system database/application that provides customers with all usage insights they need to reason about their global Snowflake footprint as well as 1st party business logic such as ML powered functions and Budgeting applications. AS A PRINCIPAL SOFTWARE ENGINEER IN SNOWHOUSE FOUNDATION, YOU WILL: Design and implement innovative highly available distributed platforms and pipelines and enhance the overall Snowflake data infrastructure Lead and drive projects from idea formulation to design, implementation and successful productionization. Collaborate

AWSAzureGCPAI
L
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At Linear, we're building the product development system for teams and agents. AI is fundamentally changing how software gets built, and we’re shaping the tools this new era requires. Founded in 2019, Linear has become the platform of choice for more than 40,000 companies (including OpenAI, Coinbase, and Ramp) to plan, build, and ship their products. Today, our team is distributed across North America, Europe, and Australia, and we’re continuing to grow internationally. What unites us is relentless focus, fast execution, and a deep care for software craftsmanship. As a small team, we’re all generalists that work across the full stack (built in Typescript end-to-end). We’re looking for experienced engineers that thrive in an environment of autonomy and individual responsibility to help us build the future of product development. Location & work mode Linear is a remote-first company, with optional co-working offices in San Francisco, New York, and London. This role is open to candidates based in the US and Europe. You can work from anywhere within these regions. We value deep focus and async collaboration, with intentional moments to connect in person through team off-sites, optional co-working, and occasional travel. What you'll do Work closely with founders and design to implement new concepts and ideas Build AI-powered functionality into the core of Linear Update our realtime collaborative content editor used across all internal surfaces Build new user-facing features with beautiful and scalable UI components Obsessively improve application performance Refine our software development processes to keep the team operating at high velocity What we're looking for 5+ years of experience building customer-facing products at a high-quality software company Strong React and TypeScript fundamentals, with experience across the full stack (Browser technologies, Node, GraphQL, PostgreSQL) Track record of driving complex, end-to-end features (not just incremental improvemen

TypeScriptReactSQLPostgreSQL
L
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At Linear, we're building the product development system for teams and agents. AI is fundamentally changing how software gets built, and we’re shaping the tools this new era requires. Founded in 2019, Linear has become the platform of choice for more than 40,000 companies (including OpenAI, Coinbase, and Ramp) to plan, build, and ship their products. Today, our team is distributed across North America, Europe, and Australia, and we’re continuing to grow internationally. What unites us is relentless focus, fast execution, and a deep care for software craftsmanship. As a small team, we’re all generalists that work across the full stack (built in Typescript end-to-end). We’re looking for experienced engineers that thrive in an environment of autonomy and individual responsibility to help us build the future of product development. Location & work mode Linear is a remote-first company, with optional co-working offices in San Francisco, New York, and London. This role is open to candidates based in the US and Europe. You can work from anywhere within those regions. We value deep focus and async collaboration, with intentional moments to connect in person through team off-sites, optional co-working, and occasional travel. What you'll do Build new user-facing features with everything from database models to GraphQL resolvers and UI components Optimize our data synchronization stack by applying better serialization protocols Add real-time collaborative editing to our content editor Improve performance by profiling and tweaking virtualized list rendering Add analytics, monitoring, and alerts to our service so that we can better respond to operational incidents Open-source any non-trivial innovations that come out of our work on the product Redefine best-in-class software development processes so that we can build a purpose-built product. What we're looking for 5+ years of experience building customer-facing products at a high-quality software company Strong React and Typ

TypeScriptReactSQLPostgreSQL
L
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At Linear, we're building the product development system for teams and agents. AI is fundamentally changing how software gets built, and we’re shaping the tools this new era requires. Founded in 2019, Linear has become the platform of choice for more than 40,000 companies (including OpenAI, Coinbase, and Ramp) to plan, build, and ship their products. Today, our team is distributed across North America, Europe, and Australia, and we’re continuing to grow internationally. What unites us is relentless focus, fast execution, and a deep care for software craftsmanship. In this role, you will serve as the primary technical partner for our most strategic prospects and customers. You’ll play a dual role: driving complex technical engagements through end-to-end demos, migrations, and integrations, and accelerating ramp for new sales hires by codifying best practices. You will showcase our product’s value to prospects, uncover challenges and architect tailored solutions, drive successful sales evaluations, and champion customer needs internally. Location & work mode Linear is a remote-first company, with optional co-working offices in San Francisco, New York, and London. This role is open to candidates based in North America. You can work from anywhere within this region. We value deep focus and async collaboration, with intentional moments to connect in person through team off-sites, optional co-working, and occasional travel. What you’ll do Articulate Linear’s architecture, data model, and integration patterns to engineering and product teams; deliver tailored technical presentations and deep-dive demos for key stakeholders. Design, plan, and execute large-scale migrations, custom configurations, and API integrations; troubleshoot proactively to ensure seamless onboarding. Build and maintain realistic demo instances that mirror enterprise workflows, scale tests, and feature toggles—enabling sales reps to showcase Linear’s capabilities in context. Collaborate with Sales

JavaScriptPythonJavaSQL
L
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At Linear, we're building the product development system for teams and agents. AI is fundamentally changing how software gets built, and we’re shaping the tools this new era requires. Founded in 2019, Linear has become the platform of choice for more than 40,000 companies (including OpenAI, Coinbase, and Ramp) to plan, build, and ship their products. Today, our team is distributed across North America, Europe, and Australia, and we’re continuing to grow internationally. What unites us is relentless focus, fast execution, and a deep care for software craftsmanship. As a small team, we're all generalists that work across the full stack (built in TypeScript end-to-end). We’re looking for engineers that thrive in an environment of autonomy and individual responsibility to help us build the future of product development. Location & work mode Linear is a remote-first company, with optional co-working offices in San Francisco, New York, and London. This role is open to candidates based in North America. You can work from anywhere within this region. We value deep focus and async collaboration, with intentional moments to connect in person through team off-sites, optional co-working, and occasional travel. What you'll do Work closely with founders, product, and design to implement new concepts and ideas Build AI-powered functionality into the core of Linear Update our collaborative content editor used across all internal surfaces Build new user-facing features with beautiful and scalable UI components Obsessively improve application performance Refine our software development processes to keep the team operating at high velocity What we're looking for 2-5 years of experience building customer-facing products at a software company with a high engineering bar Strong React and TypeScript fundamentals, with experience across the full stack (Browser technologies, Node, GraphQL, PostgreSQL) Track record of driving complex, end-to-end features (not just incremental improvement

TypeScriptReactSQLPostgreSQL
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team We bring OpenAI's technology to the world through products like ChatGPT and the OpenAI API. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role OpenAI is looking for an experienced Performance Engineer to help us scale the performance, reliability, and efficiency of our systems. In this role, you'll apply deep technical expertise to optimize infrastructure and application-level performance across mission-critical products like ChatGPT and our developer API. You’ll work cross-functionally with teams building core services, training models, and developing real-time user experiences to push our latency, throughput, and cost-efficiency to the next level. We are looking for engineers who thrive in ambiguous environments, value deep systems understanding, and are motivated by delivering measurable impact. This is a highly technical, individual contributor role focused on root-cause analysis, profiling, instrumentation, and architecture-level performance improvements across our stack. In this role, you will: Analyze and optimize performance across application, middleware, runtime, and infrastructure layers—networking, storage, Python runtime, GPU utilization, and beyond. Develop tooling and metrics that provide deep observability into system performance. Collaborate closely with infra, platform, training, and product teams to identify key performance goals and drive systemic improvements. Influence architecture and design decisions to prioritize latency, throughput, and efficiency at scale. Lead investigations into high-impact performance regressions or scalability issues in production. Drive performance testing strategies and help define SLAs/SLOs around latency and throughput for critical systems. You might thrive in this role if you: Have 7+ years of experience in software engineering with a strong tr

PythonAWSRestAI
🔔

Get new distributed systems engineer jobs in United States by email

Daily job updates · Unsubscribe anytime