We are seeking an experienced IT/Lab Manager to lead the planning, deployment, and operations of our physical lab environment and IT systems. This role will focus on building and maintaining scalable, reliable, and secure environments to support engineering teams involved in research, quality assurance, validation, and related activities. It will also support internal collaborators. You will have an outstanding opportunity to drive innovation in a multidimensional, technology-focused company that is crafting the future of data-center and lab technologies. If you bring perfection and creative thinking while solving issues as they arise, and enjoy working with distributed teams – your place is with us! What You’ll Be Doing: Own day-to-day operations, planning, and roadmap for the engineering lab and IT infrastructure (servers, storage, networking, and related services). Lead and mentor an IT/Lab team, driving guidelines, standards, and a culture of ownership, partnership, and continuous improvement. Collaborate closely with R&D, QE, Verification, and other engineering teams to design, provision, and maintain environments that meet their performance, reliability, and security needs. Lead all aspects of running data center and lab operations, including rack layout, cabling, power and cooling, hardware lifecycle, and resource availability. Lead procurement and vendor management for hardware, software, and services, including evaluation, negotiation, and ongoing relationship management. Implement and maintain automation for system provisioning, configuration, and operations using tools such as shell/Perl/Ansible. Design and maintain monitoring, logging, and alerting for servers, network, and storage systems to ensure high availability and rapid incident response. Investigate and resolve sophisticated infrastructure issues across OS, networking, storage, virtualization, and appli
Jobiba hiring network
Distributed Systems Engineer Jobs
1,306 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About the Role We’re hiring the first Technical Program Manager to join a newly formed Technical Program Management function, reporting to the Head of TPM. This is a founding role, you’ll help design and build how strategic, cross-functional work gets delivered across a global organization. TPMs at Supabase own the connective tissue between teams and between functions, managing the dependencies of our largest programs, and author and maintain/improve the standard of using Linear. This role is a strong fit for someone who has a genuine curiosity about ambiguous, complex problems, is self-directed and has conviction in their craft. What You'll Be Responsible for Parter with Engineering, Product, Design (EPD) and other stakeholder to structure and drive complex, cross-team initiatives; clarifying scope and value, securing commitment across a distributed organization, setting timelines, and leading activities from design through launch. Design systems by building the tracking, reporting, and escalation mechanisms that surface risk and progress. Build process to ensure launch readiness is predictable, consistent and coordinated across all cross-functional participants. Ask the questions that clarify priority and scope, giving async teams the context to keep building and unblocking. Build trusted relationships with stakeholders and maintain transparency into program goals, commitments, risk and delivery quality, primarily through clear written communication. Identify bottlenecks before they become blockers, develop options to resolve them, and clearly communicate the trade-offs that each option carries. Map dependencies and relationships across projects and teams, coordinating delivery across groups that don’t share reporting lines. Keep Engineering, Product, and Design accountable for their own delivery outcomes while you track, surface risk, and unblock. Develop technical and domain fluency to be an active and credible member of the delivery team and triad. You Might Be
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. The Role Many enterprise deals Ema closes starts with a BDR earning five minutes with the right person. Our Business Development Representatives run a signal-led, account-based hunting motion against a defined enterprise ICP to book qualified meetings (SQLs) that convert into demos and paid POCs. The motion works: the team has a real playbook, real signal scoring in our tools, and real enterprise logos in the pipeline. What it needs now is a dedicated leader to run it with daily rigor. As Enterprise BDR Leader, you own the BDR function end to end. You will manage a distributed team of Business Development Representatives across the US and India, instill the metrics discipline and coaching cadence that turns activity into qualified pipeline, and rebuild onboarding and enablement so new reps ramp fast and hit quota on schedule. This is a newly created role — the BDR function has been running without a dedicated, full-time manager, and you are the person who changes that. You report into our GTM organization, and work closely with our Account Executives, Regional VPs, and Marketing/demand gen to keep the pipeline both full and qualified. What You Will Own 1. The team. Hire, ramp, and
Job Title Service Operations Manager Job Description Your role: Lead service operations for a global interventional cardiology and vascular device portfolio, owning field performance, delivery execution, and operational governance across a geographically distributed team of field service engineers, product support specialists, and service coordinators. You will set the operating rhythm for corrective, preventive, and installation service across North America while partnering with international service leaders to align standards globally. Own the service training and technical education function end to end, including training program management, instructor-led and digital course delivery, field certification programs, and training center operations across three continental sites. You will ensure engineers maintain verified competency on legacy, newly acquired, and next-generation product lines in a regulated medical device environment where patient safety depends on technician proficiency. Drive measurable improvement in service performance using indicators such as time to system restoration, first-visit resolution, installation quality, contract capture, and preventive maintenance completion. You will lead root-cause analysis on systemic delivery issues, design operational improvements with engineering and supply chain partners, and translate field data into changes that improve customer uptime and reduce repeat service events. Manage service quality and compliance activities including complaint escalation, field execution of corrective actions and compliance tracking, product lifecycle support for systems approaching end of service, and regulatory documentation across all served markets. You will work directly with quality, regulatory, and product engineering teams to ensure service processes satisfy medical device requirements while maintaining the speed and respo
We are hiring a Senior Technical Product Marketing Manager to lead positioning and messaging and to grow adoption of MongoDB Search and Vector Search as foundational components of our platform – the retrieval layer powering the next generation of grounded AI applications and agents. This is a high-impact role for a marketer who thinks like a builder. As developers architect increasingly sophisticated systems – RAG pipelines, agentic workflows, multi-modal search experiences – retrieval has moved from an implementation detail to a core design decision. You’ll join a high-performing, globally distributed team and partner closely with Marketing, Builder Relations, Product Management, Engineering, Partners, and Sales to develop, measure, and achieve cross-functional goals. The role requires technical depth in information retrieval — lexical and vector search, hybrid approaches, embeddings, re-ranking, agentic retrieval loops, and the tradeoffs that matter in production systems — paired with the product marketing instincts to turn that depth into crisp, differentiated messaging for distinct user and buyer personas. Hands-on experience building or shipping AI-enabled products is a strong advantage. Individuals with prior experience in technical sales, developer relations, or technical marketing are encouraged to apply. This person is a voracious consumer of AI research and pays close attention to shifting patterns in application architectures and development, including agentic systems. This individual is confident in communicating with technical practitioners and non-technical decision makers in one-to-few and one-to-many engagements for internal and external audiences. We are looking to speak to candidates who are based in the US for our hybrid working model. What You’ll Do Drive Strategy & Execution: Act as a strategic partner for high-impact initiatives that align with MongoDB’s long-term business goals in collaboration with Marketing, Developer Relations, Product
Hyliion is committed to creating innovative solutions that enable clean, flexible and affordable electricity production. The Company’s primary focus is to develop distributed power generators that can operate on various fuel sources to future-proof against an ever-changing energy economy. Job Purpose The Field Service Specialist provides installation support, commissioning, maintenance, troubleshooting, and on-site customer support for deployed KARNO Power Modules. This is the first dedicated field service role supporting KARNO and is a foundational position within Hyliion's field service organization, which is being built to support installations across the United States. Early deployments focus on defense and data center applications. The position begins with an intensive training period of approximately three to four months in Milford, OH, working alongside the research, development, and engineering teams as KARNO units are built and serviced during final testing, including training on heat engine fundamentals, PLCs, HMIs, and advanced control systems. The position then transitions to field installation, commissioning, and long-term on-site support at customer locations, initially across the West Coast and expanding to other regions as the installed base grows. As the service organization scales, this position helps define its processes and standards. AI at Hyliion At Hyliion, AI is core to how we work. We equip every team member with leading AI tools and count on you to use them — to move faster, solve harder problems, and help us realize the full potential of KARNO technology for the world. Duties and Responsibilities Provide on-site maintenance, troubleshooting, and break-fix support for deployed KARNO units, with remote support from the engineering team. Diagnose issues using remote monitoring systems and troubleshooting logs. Troubleshoot controls, instrumentation, and high-voltage electrical system issues. Super
NVIDIA is seeking a Senior Staff SRE to build and operate reliable, scalable compute platforms that support global engineering workloads. This role spans Kubernetes, KubeVirt, bare-metal infrastructure, automation, observability, and AI-enabled operations. Join a team that solves complex infrastructure challenges, builds durable automation, and improves the reliability and operational experience of critical compute services. What you’ll be doing: Build, operate, and improve large-scale Kubernetes, KubeVirt, Linux, container, and bare-metal compute platforms, with a focus on performance, capacity, reliability, and operational scale. Lead bare-metal provisioning and lifecycle management in data centers, including PXE boot, DHCP, DNS, OS provisioning, hardware validation, and fleet automation. Develop automation, self-service capabilities, and observability solutions using APIs, Python or Go, Infrastructure as Code, configuration management, metrics, logs, traces, and service-health data. Define and operate SLOs, SLIs, error budgets, alerting, and incident-response practices; lead complex incident investigations, corrective actions, and blameless postmortems. Partner with infrastructure, security, hardware, data-center, and application teams to deliver global platform initiatives, and participate in an on-call rotation. What we need to see: BS in Computer Science, Engineering, a related technical field, or equivalent experience, plus 10+ years operating production infrastructure or platform services. Strong expertise in Kubernetes administration, KubeVirt, Docker, containerization, microservices, Linux systems, and resolving distributed-system challenges. <l
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role? We're hiring a People Project Manager to join our HR PMO and help deliver an ambitious portfolio of HR programs, products, and initiatives as we scale globally. This is not a business-as-usual project management role . We're a hypergrowth company building at speed, and our HR team is building alongside it. We're looking for someone who gets genuine energy from bringing order to chaos, creating structure where there isn't any yet, and getting important things shipped. Hypergrowth, for real. Things move fast, priorities evolve, and there's always something meaningful to jump into. If you love pace, variety, and figuring things out as you go, you'll have a huge canvas here. A chance to build. Our HR PMO is taking shape within a hugely ambitious HR team. You'll help create the workflows, rhythms, and tools we need to build a world-class function at scale. Real ownership. Take projects from idea to delivery — creating structure, driving actions, managing dependencies, and keeping teams moving. Global from day one. Work across our distributed HR team and with stakeholders around the world. You'll thrive if you: Create o
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Customer Marketing Manager who can own the full customer evidence motion at Baseten: building the systems that capture customer stories, running the co-marketing programs that amplify them, and developing the channels and assets that get those stories in front of the right people. Our customers are ML engineers and AI teams deploying serious workloads — and the stories they tell about what they've built matter. We've earned trust with some of the most demanding technical teams in the industry, and this role exists to turn that trust into evidence. RESPONSIBILITIES Co-Marketing Execution Serve as the DRI for every customer co-marketing launch end to end — managing timelines, coordinating internal and external stakeholders, and driving the process from first outreach to final publication Own the single source of truth for what's in flight across all customer co-marketing activity Coordinate with design, social, and sales to ensure every asset is built, approved, and distributed correctly Customer Evidence & Asset Library Own the customer evidence library: written case studies, video stories, customer quote repository, logo library, and sales snippets ensuring all assets stay current and are tagged and accessible for sales and marketing use Run the monthly operating rhythm: new logo additions from closed-won opportunities, asset updates, and customer health checks Programs & Channels I
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? As a Member of Technical Staff on our Applied ML team, you will work directly with customers to quickly understand their greatest problems and design and implement solutions using Large Language Models. You’ll apply your problem-solving ability, creativity, and technical skills to close the last-mile gap in Enterprise AI adoption. You’ll be able to deliver products like early startup CTOs/CEOs do and disrupt some of the most important industries and institutions globally! As a Member of Technical Staff, you will: Plan and execute large-group projects that carry through from ideation to production. Bring cross-functional alignment across engineering, product and other disciplines. Mentor a distributed team of engineers in subject matter expertise. Identify opportunities and gaps in existing models and strategize what to work on. Work closely with product teams to develop solutions. Engage in collaborations with our partner organizations. Assist our legal teams with preparation of patents on developed IP. Join us at a pivotal moment, shape what we build and wear multiple hats! You may be a good fit if you have: Prio
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As the Senior Director of Solutions Architecture for the Americas at Cohere, you will own the US and Canada commercial Solutions Architecture function. You will lead the team that turns enterprise interest in agentic AI into deployed, production systems, and you will be accountable for the technical win in the most competitive AI market in the world. The United States and Canada are our largest commercial opportunity, and this seat owns how we win them. You will take an established, distributed team of strong technical people and raise what it can do — setting the bar and establishing the operating rhythm that lets a team of generalists run consistent, industry-fluent plays at enterprise scale. You will set direction for the function, sit on the Solution Architecture leadership team alongside the regional leaders for EMEA and Asia Pacific, and contribute to company-wide decisions with your peers across Sales, Product and Engineering. In this role, you will: Lead and Scale the Organization: Build, coach and develop a high-performing Solutions Architecture organization across the US and Canada, and grow the senior technical talent
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Senior Staff Rust Developer to join our Platform Convergence Team. This is a hybrid role based in San Jose, CA reporting to the Sr. Director, Software Engineering. Join us to build a new platform from the ground up that can scale hundreds of millions of users with high reliability and low latency. You will design and implement distributed system and core infrastructure components while collaborating closely with various stakeholders. What you’ll do (Role Expectations) Design and build a low-latency, high-throughput data forwarding plane using Rust, leveraging its async/await model for efficient I/O and service-oriented infrastructure Develop distributed, scalable systems with a focus on concurrency, fault tolerance, and messaging Implement and maintain gRPC-based APIs and services to integrate forwarding plane capabilities with control and orchestration layers Optimize system
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. We are looking for a Transformation Architect to join our Emerging markets team in a remote capacity in Mexico, reporting to the Senior Director, Sales Engineering within the Commercial Americas department. You will drive customer engagement and solution design, providing the technical leadership and expertise necessary for deployment design and optimization. What you’ll do (Role Expectations) Design and implement our Zero Trust architectures Gather requirements and understand specifications while evaluating existing systems Influence and mentor activities within the team and across the organization Who You Are (Success Profile) You thrive in ambiguity. You are comfortable building the path as you walk it and see ambiguity as raw material to build something meaningful. You act like an owner. You navigate seamlessly between high-level strategy and hands-on execution, fueled by a bias for action. You are a problem-sol
NVIDIA is leading company of AI computing. At NVIDIA, our employees are passionate about AI, HPC , VISUAL, GAMING. Our SA team is more focusing to bring NVIDIA new technology into difference industries. We help to design the architecture of AI computing platform, analysis the AI and HPC applications to deliver our value to customers, focusing on defining and solving computational challenges in LLM inference and training acceleration, as well as network communication and data transfer optimization. What You'll Be Doing: Contribute to the development of open-source inference frameworks such as SGLang and vLLM, including feature and operator development, performance optimization, and model support, in collaboration with the community. Develop and optimize KV cache offloading frameworks for LLM workloads, supporting multi-level cache offloading and reuse across CPU, SSD, and remote storage to improve inference efficiency. (Team project: FlexKV) Drive R&D on compute performance in distributed training, and explore methods and technologies for performance optimization. Study computational challenges in machine learning systems, identify common needs and bottlenecks, and build example code, acceleration libraries, or frameworks accordingly. What We Need to See: Over 5 years working experience in the technology industry, with master’s degree or above in computer science, mathematics, electrical engineering, automation, or related fields. Strong interest in accelerated computing, parallel computing, and heterogeneous computing, with the motivation to explore these areas in depth. Solid programming skills, with a good understanding of data structures and computer systems fundamentals. Strong learning agil
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Summary: You'll own ClickUp's social strategy and build the systems to execute it at scale. You deeply understand what wins on each platform: the formats, the hooks, the timing, the tone. But you're not just a strategist who hands off a plan. You build AI-powered workflows and automation to operationalize your strategy so it runs continuously, learns from data, and scales beyond what any team could do manually. You're a social-native operator who builds systems, not an engineer who dabbles in social. Responsibilities: Own platform-native social strategy across X, LinkedIn, TikTok, and emerging channels: define what ClickUp's voice, format, and engagement approach looks like on each, tailored to what works on that platform Develop and execute content and engagement strategies that drive measurable growth in reach, engagement, and audience quality Identify trends, conversations, and cultural moments worth engaging with, and move fast enough to capitalize on them Build AI-powered systems and automated workflows to execute social strategy at scale: monitoring, engagement, response, and content distribution Create feedback loops between social performance data and strategy; use signal to iterate what gets made and how it gets distributed Own proactive engagement: identify and engage relevant conversations, mentions, and opportunities using AI-powered monitoring and automated response workflows Develop automated systems that handle routine engagement while escalating high-value or brand-sensitive conversations to humans Own execution end-to-end: strategy through measurement, with clear accountability for out
Get new distributed systems engineer jobs by email
Daily job updates · Unsubscribe anytime