Jobiba hiring network

Remote Remote Software Engineer Distributed Systems Manager Manager Jobs

15 active opportunities · Updated for September 2026

Market range: $153K – $254.7K/yr

Fresh results

15 shown

Explore current remote remote software engineer distributed systems manager manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

S
Smartsheet
📍 -REMOTE, USA-Full-timeRemoteFrom $2.3M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. About the Role The AI Platform Engineering team is looking for a highly motivated and talented engineer who are passionate about continuous learning and excited to grow in a fast-paced, innovative environment. We are an agile team that operates iteratively, focused on building high-quality software and adhering to rigorous operational best practices across complex, cross-functional distributed systems. This full-time position reports to a Software Engineering Manager and can be located in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. What You'll Do Build the AI Platform Foundation : Lead the design and ownership of the core infrastructure that serves as the backbone for all Smartsheet AI experiences. Focus on building a robust, multi-tenant environment that reduces friction for internal teams, allowing them to deploy reliable and scalable AI features with ease. Standardize the AI Developer Path : Architect high-level abstractions and "Golden Path" APIs that democratize AI development across Smartsheet. By insulating product teams from infrastructure complexity, you will enable them to ship intelligent features with high velocity while guaranteeing safety and consistency at scale. Engineer AI Trust & Safety Systems : Establish the mission-critical monitoring and quality assurance layers that protect Smartsheet customers. By creating rigorous evaluation pipelines, you will ensure every AI-driven feature meets the high bar for safety, data privacy, a

REMOTEpythonvueaws
View job →
A
Anyscale
📍 RemoteFull-time$153K – $254.7K/yr · Jobiba est.
1mo ago

About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role: Ray aims to provide a universal API for building distributed applications (e.g. a machine learning pipeline of feature engineering, model training, and evaluation). Data is usually a core element connecting these different stages, and therefore plays a critical role in Ray’s usability, performance, and stability. We are looking for strong engineers to build, optimize, and scale Ray’s Datasets library and data processing capabilities in general. About the Ray Data team: The Ray Data team currently develops and maintains the Ray Datasets library, which is already powering critical production use cases (e.g. large scale data compaction at Amazon , and ML pipeline at Alibaba ). Ray Datasets is a Python library built on top of Apache Arrow and Ray Core (Ray’s C++ backend), and the Ray Data team interacts closely with Ray Core components including the scheduler and the memory & I/O subsystems. The Ray Data team also works closely with Ray’s ML libraries including Train, RLlib, and Serve. A snapshot of projects you will work on: - Performance of Ray Datasets at large scale (leveraging Arrow primitives, optimizing Ray object manager, etc.) - Integration with ML training and data sources - Stability an

pythonmachine learningai
View job →
S
Smartsheet
📍 -REMOTE, USA-Full-timeRemoteFrom $1.5M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. As a Software Engineer II at Smartsheet, you will help build the next generation of our platform while tackling meaningful distributed systems challenges alongside a talented engineering team. You will collaborate closely with product managers and fellow engineers through code reviews and architectural discussions, and you will have the opportunity to mentor junior engineers as you grow in your own career. This role is a strong fit for someone with a couple of years of hands-on development experience who is ready to take on more ownership, sharpen their skills in cloud infrastructure, and help shape how the team uses AI tools to work smarter and faster . You Will: Build scalable back-end services for the next generation of applications at Smartsheet (Kotlin, Java) Solve challenging distributed systems problems and work with modern cloud infrastructure (AWS, Kubernetes) Take part in code reviews and architectural discussions as you work with other software engineers and product managers Mentor junior engineers on code quality and other industry best practices Forge a strong partnership with product management and other key areas of the business Actively use AI tools to improve personal and team efficiency across coding, testing, design, and troubleshooting, exploring AI integration opportunities within team processes and features, and coaching others on effective AI use You Have: 2+ years software development experience building highly scalable, highly available applications 2+ years of programming experience with full st

REMOTEtypescriptjavaaws
View job →
S
Smartsheet
📍 -REMOTE, USA-Full-timeRemoteFrom $1.3M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. We are hiring a Software Engineer I to join our Engineering team. You will get your hands on full-stack services, dig into distributed systems problems, and work right at the intersection of traditional software engineering and AI-driven development. We are not just dabbling in AI here. We use it to write better code, catch bugs faster, and troubleshoot smarter, and we want engineers who want to get good at that too and help their teammates do the same. You will ship real features, sit in on architectural conversations that actually shape the product, and grow alongside a team that cares as much about doing good work as they do about doing meaningful work. You Will: Build scalable front-end and back-end services for the next generation of applications at Smartsheet (Kotlin, Java, Typescript, React) Solve challenging distributed systems problems and work with modern cloud infrastructure (AWS, Kubernetes) Take part in code reviews and architectural discussions as you work with other software engineers and product managers Forge a strong partnership with product management and other key areas of the business Enhance existing application code with new features and strike a balance when making technical decisions (build vs refactor vs simplify) Actively use AI tools to improve personal and team efficiency across coding, testing, design, and troubleshooting, exploring AI integration opportunities within team processes and features, and coaching others on effective AI use You Have: 1+ years software development experience build

REMOTEtypescriptpythonjava
View job →
A
1mo ago

About the Role & Team We’re looking for an Engineering Manager to lead the Data Infrastructure team within Statsig Experiment at Amplitude. You will lead a multidisciplinary team of software engineers, data engineers, and data scientists responsible for the systems that power experimentation at scale. The team owns three critical areas: Data ingestion: Collecting and importing experiment exposures, custom events, OpenTelemetry data, and real user monitoring data across SDKs, streaming systems, cloud storage, and customer data warehouses. Data computation: Building distributed computation systems that transform raw data into accurate, timely experiment results. Stats engine: Developing and productionizing the statistical methods that help customers make trustworthy decisions from their experiments. This is not a traditional data engineering management role. We are looking for a leader with a solid data science and statistical foundation who can connect advances in experimentation methodology with scalable production systems. You will help set our technical and scientific direction, translating new statistical methods and machine learning research into capabilities that customers can use reliably at scale. You’ll partner closely with data scientists, engineers, product managers, and customers to advance the state of experimentation. The ideal candidate is equally comfortable discussing causal inference and statistical power with data scientists, distributed computation architectures with engineers, and experimentation strategy with customers. What You’ll Do Lead and grow the team responsible for Statsig’s data ingestion, experiment computation, and stats engine. Define the technical and scientific strategy for advancing experimentation across both Statsig Cloud and warehouse-native deployments. Partner with data scientists and engineers to turn new statistical and causal inference methods into scalable, reliable product capabilities. Evolve our data and computatio

restmachine learningai
View job →
C
Coder
📍 United StatesFull-timeRemote
16 days ago

As an Engineering Manager on Coder’s Core Workspaces team, you’ll lead engineers building and evolving the systems behind our agentic development experience. You’ll help make agents more capable, reliable, and useful across real development environments. You’ll guide technical direction while growing the team and keeping execution sharp. You’ll work closely with Engineering, Product, and Design across the agent harness, integrations, and developer workflows. What you’ll do here Lead and grow a team within our Workspaces organization. Set technical direction across the agent harness, integrations, and workflows. Stay close to the code and contribute to architecture and implementation decisions. Evolve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Partner with Product and Design to turn agent capabilities into useful developer experiences. Improve reliability, performance, and operability across agentic systems. Coach engineers, raise the technical bar, and create clarity around priorities and tradeoffs. What we’re looking for Experience managing and growing software engineering teams. Strong hands-on engineering experience with React and TypeScript . Experience with Go . Hands-on experience building systems around LLMs and agentic workflows. Experience with model APIs, tool calling, context management, or agent loops. Strong distributed systems knowledge. Working knowledge of AWS . Strong technical judgment and comfort working through ambiguity. A track record of helping engineers grow while maintaining a high execution bar. Bonus tacos if you have Experience building coding agents, developer tools, or cloud development environments. Experience with MCP , agent tools, or multi-agent systems. Experience with remote execution, sandboxing, or isolated compute. Experience building abstractions across multiple model providers. Deep experience with AWS, Kube

REMOTEtypescriptreactaws
View job →
C
Coinbase
📍 - USAFull-timeRemoteFrom $218K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Staff Software Engineer (EAA) The Customer Experience (EAA-CX) sub-organization within Coinbase's Enterprise Applications and Architecture (EAA) group builds and operates the secure, Tier-1 platforms that power internal investigative and support workflows for approximately 7,000 specialists across CX, Compliance, Legal, Global Investigations, and Trust and Risk. As a Staff Software Engineer, you'll operate as a technical anchor across 8 globally distributed EAA-CX engineering squads, identifying high-leverage AI and developer-experience opportunities, architecting solutions on Coinbase's production AI infrastructure, and driving platform excellence that raises Tier-1 engineering quality across the portfolio. What you'll do: Own the EAA-CX AI maturity and developer-experience roadmap in partnership with engineering leadership, embedding with squads to surface opportunities, prototyping on Coinbase's existing AI platform, and driving adoption through working demos and feedback loops. Lead platform excellence workstreams that improve reliability, observability, change-failure rate, and operational readiness for Tier-1 systems, partnering with engineering managers to ensure practices persist on the teams. Define build-versus-buy strategy for developer-experience and AI tooling across EAA-CX, evaluating what to build internally, adopt from partner teams, procure from vend

REMOTEpythonsqlaws
View job →
C
Coinbase
📍 - CanadaFull-timeRemoteFrom C$191.1K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Software Engineer, Backend on the Growth Foundations team within the Consumer & Business group, you'll build and scale the core infrastructure that powers growth across Coinbase. This team owns the foundational systems behind powering efforts as: notifications at scale, incentives and referral platforms, user targeting and segmentation, lifecycle communications, and ML ranking that power more than 50% of Coinbase web and app surfaces. You'll design and ship backend systems that enable high-impact growth initiatives for Retail, Institutional, Base, and international teams while maintaining the reliability, scalability, and performance our customers depend on. What you'll do: Own the design and delivery of scalable distributed backend services that handle high QPS, low latency, and high reliability requirements for growth infrastructure Architect resilient service-oriented systems using modern cloud technologies, contributing across backend, data, and ML layers of the stack Partner with engineers, designers, product managers, and senior leadership to translate product and technical vision into quarterly roadmaps Lead investments in developer tooling, shared platform capabilities, and cross-team relationships that increase organizational velocity Build high-quality, well-tested production code for systems that are secure, reliable, and maintainable at sc

REMOTEmongodbawsdocker
View job →
A
Anyscale
📍 RemoteFull-timeRemote$153K – $254.7K/yr · Jobiba est.
22 days ago

About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We're commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we're building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role Anyscale is looking for a Software Engineer to join the ML Developer Experience (MLDevX) team. MLDevX owns the experience layer of the Anyscale platform: the interfaces through which users and coding agents discover, configure, run, observe, debug, and productionize AI workloads. Every user journey crosses this layer through the CLI, SDKs, APIs, UI, Workspaces, MCP, or the workflows and integrations built on top of them. Together, these form the user’s primary interface into Anyscale, turning distributed computing from a systems problem back into a coding problem. We build the common contracts, tools, control-plane services, and architecture that power these surfaces. You will work across the stack from developer tooling to cloud infrastructure and the Ray runtime. Manage long-running operations and make failures across jobs, tasks, actors, nodes, and GPUs easier to diagnose. The systems you build must scale with the platform, remain predictable through failures, and be intuitive for developers, programmable for applications, and operable by coding agents. This is a high impact individual-contributor role with end-to-end ownership. You will work directly with users and field teams to identify high

REMOTEmachine learningaigo
View job →
C
Coinbase
📍 - USAFull-timeRemoteFrom $253.9K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Staff Software Engineer on the Data Platform team within Platform , you'll define and lead the technical strategy for Coinbase's data infrastructure, spanning ingestion, transformation, warehousing, streaming, and serving systems. This is a foundational role at the intersection of distributed systems, data engineering, and AI-readiness, reporting to the Senior Director of Engineering. You'll set architectural direction, drive multi-quarter roadmaps, and transition the organization from managed-service dependency toward engineering-built, platform-grade infrastructure that powers everything from fraud detection to modern multi-agent AI architectures. What you'll do: Own the technical strategy and architecture for Data Platform, setting direction across data ingestion, transformation, warehousing, streaming, and serving systems while driving engineering-led cost reduction at the infrastructure layer. Architect data infrastructure to natively support AI and ML workloads, ensuring pipelines, data lake systems, and compute can power ML training, feature stores, real-time inference, and multi-agent AI architectures at scale. Drive the evolution to near-real-time data availability, enabling downstream teams across Coinbase to act on fresher data for fraud detection, financial reporting, and analytics. Build alignment and secure commitment from senior leadership

REMOTEawsaigo
View job →
T
Twilio
📍 - USFull-timeRemote$138.7K – $173.4K/yr
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Software Engineer, Platform Engineering (L3) About the job This position is a critical engineering role within Twilio Platform Engineering, requiring a hands-on engineer capable of developing, deploying, and managing highly available, massive-scale distributed systems. Our systems regularly process more than 12 billion emails during peak events like Black Friday, and our throughput requirements continue to scale rapidly. As an L3 engineer, you will build and operate resilient backend services at scale and contribute to the design and reliability of our dual-cloud infrastructure span across Amazon Web Services (AWS) and Microsoft Azure. You'll run Kubernetes beyond the boundaries of managed services, automate infrastructure with Terraform, and write production code to help keep distributed systems healthy under real production load while using modern AI-assisted tooling to move faster. Responsibilities In this role

REMOTEawsazuredocker
View job →
G
Godaddy
📍 CanadaFull-timeFrom C$107K/yr
1mo ago

Location Details: Canada, Remote At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) , and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Contribute to the development of GoDaddy’s eCommerce and SSO infrastructure and Kubernetes systems on AWS. On a day-to-day basis you will be working on the team who designs, writes, tests and deploys the infrastructure and application management software for GoDaddy’s eCommerce applications. Expect to learn every day. What you'll get to do... Work as a polyglot engineer, writing and maintaining Infrastructure as code with frameworks/ ecosystems such as Java, Unix CLI, and NodeJS Build and operate infrastructure workflows and deployment pipelines using Kubernetes, Argo Workflows, Argo CD, and GitOps practices Design, build, and own services and APIs in Java, running on Kubernetes-based platforms across AWS and distributed systems Develop and support application and infrastructure delivery pipelines, enabling reliable releases of eComm, Auth and Infrastructure services Collaborate closely with other GoDaddy departments to help advance security and technical standards, maintain regulatory compliances while operating eComm & Auth platforms Your experience should include... 5+ years of strong backend software engineering experience in Java Hands-on experience with Kubernetes, including Helm, Kustomize, or equivalent tools to deploy and manage backend services Experience building and operating high-volume, mission-critical production systems on AWS with continuous deployment (CD) practices Strong experience with infrastructure as code, supporting backend applications and services Experience with observability a

javanodejssql
View job →
S
Stripe
📍 NaFull-time$153K – $254.7K/yr · Jobiba est.
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Secrets Infrastructure team provides the cryptographic identity and secrets management foundation for Stripe. We build and operate the internal certificate authority that authenticates every person and service at Stripe, and the secrets platform that manages everything from financial partner credentials to infrastructure access keys. We build foundational security infrastructure at scale: our certificate authority issues mTLS client certificate identities for thousands of services and engineers, and our secrets platform and libraries protect access to critical financial systems and external partners across all of Stripe’s codebases, services, and platforms. The technical challenges include building systems with 99.99%+ availability, implementing TLS workload identity and attestation logic for new platforms, and designing secret management tools that are both secure and user-friendly. Our infrastructure must be both reliable and developer-friendly—we maintain libraries in Go, Java, Ruby, and Python. As a small team responsible for critical systems, engineers take on meaningful ownership. Through collaboration with teams across Stripe, you'll build and set direction for the authentication and secrets management underpin identity in distributed systems at scale. Secrets Infrastructure is a fully remote team, with a small presence in the Seattle and New York City offices. We pride ourselves on a friendly, technically rigorous, and supp

pythonjavareact
View job →
M
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azur

mongodbawsazure
View job →
M
Mongodb
📍 United StatesFull-timeFrom $127K/yr
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our NYC HQ, our smaller Austin, Palo Alto, or San Francisco offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design prim

mongodbawsazure
View job →
🔔

Get new remote remote software engineer distributed systems manager manager jobs by email

Daily job updates · Unsubscribe anytime