Jobiba hiring network

Senior System Software Safety Engineer Jobs

7,292 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior system software safety engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

L
Lyft
📍 Mexico City• Full-time
22 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Every day, millions of riders and drivers depend on Lyft to get where they're going. When something goes wrong along the way, they expect us to make it right — quickly, clearly, and without friction. How fast and how well we resolve those moments shapes whether people keep choosing Lyft. The Self-Serve Intelligence team, within the Safety & Customer Care org, is composed of engineers building the AI-powered systems that do exactly that: resolving rider and driver suboptimal experiences without agent involvement through AI Assist (e.g. AI Agents), automations, and self-serve workflows. Our goal is to make getting help feel effortless. We design and build backend services, APIs, and GenAI-powered products that combine robust engineering with applied AI to deliver reliable, scalable self-serve experiences. We are looking for a highly motivated, collaborative, team-focused and technically strong Software Engineer to join our Self-Serve Intelligence team. As a member of this team, you will build the services and AI-powered products that resolve customer issues autonomously. Every day, you'll partner with machine learning engineers, product, design, data science, and operations on high-impact projects — from shipping new AI Agent capabilities, to building the evaluation pipelines that keep their quality high, to improving the backend services underneath them. You'll bring strong engineering instincts, genuine curiosity about applied AI, and a willingness to work through ambiguity in a space that changes month to month. Responsibilities: Write well-crafted, well-tested, readable, and maintainable code Partner with senior engineers to design, build, and ship backend services and GenAI-powered products (e.g. AI Agents) that resolve rider and driver suboptimal experiences Independently lead tasks fro

pythonjavaredis
View job →
N
Nuro
📍 Mountain View• Full-time• From $258K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role As a Principal Software Engineer, you will help define and build the high-performance, highly reliable foundation of the Nuro Driver. This role spans Device Platform, Performance, and Onboard Systems, requiring deep technical leadership across sensor and compute integration, onboard runtime systems, distributed execution, and autonomy software performance. You will architect hardware-agnostic device interfaces, inter-device communication pipelines, runtime APIs, and onboard software platforms that enable Nuro’s autonomy stack to run safely and efficiently across current and future vehicle platforms. You will also lead system-wide performance and reliability efforts, including latency reduction, resource efficiency, observability, automated validation, and tooling for debugging complex onboard systems. This is a highly cross-functional and high-impact role. You will partner with autonomy, hardware, safety, validation, opera

linuxmachine learningai
View job →
M
Mindbody
📍 United States• Full-time
1mo ago

At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You'll Play: At Playlist, we're reimagining how technology can foster meaningful, real-world connections. As a Senior Software Engineer on the SmartDesk team, you'll be at the forefront of building an AI-powered front desk assistant that transforms how wellness businesses communicate and operate. Crafting end-to-end AI-powered experiences that seamlessly connect wellness businesses with their clients Designing and implementing robust web and messaging services that power conversational and workflow automation Integrating large language models (LLMs) and generative AI frameworks with a laser focus on safety, reliability, and user trust Collaborating closely with product, design, and applied AI teams to transform complex challenges into intuitive solutions Architecting scalable systems across frontend, backend, and integration layers Mentoring teammates through thoughtful code reviews and

javascripttypescriptpython
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE This role is for an individual contributor joining one of our five Everpure Resilience engineering teams within the Data & Digital Experience (DDX) organization. The product has been running for 3 years and operates like a startup within DDX. The teams are primarily composed of senior-level engineers who take full ownership of their workstreams - from early requirements definition (in close collaboration with Product Managers) to final delivery. Strong communication and collaboration skills are essential, as engineers work autonomously in a highly cross-functional environment. WHAT YOU'LL DO Architect & Deliver: Own the end-to-end design, development, and operation of high-throughput data processing services between edge devices and the Everpure cloud platform. Product Collaboration: Partner with Product Managers to translate complex requirements into scalable, resilient system architecture from concept to production. Performance & Scale: Drive continuous system improvements—experimenting with new tech to boost performance, security, and cost-efficiency for real-time data flows. Ship Fast & Safely: Utilize pragmatic testing, robust CI/CD, and automation to maintain high deployment velocity without sacrificing stability. System Integration: Lead the resolution of complex interoperability challenges between new and legacy components to ensure high availability across all regions. WHAT YOU BRING Distr

javaawsazure
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
15 days ago

Opportunity Overview: We’re looking for a senior-level automation engineer who will help raise the bar on release quality, environment reliability, and change safety across Cohere’s platform. You’ll partner closely with Product, Engineering, Platform, and SRE to build scalable automation, guardrails, and validation systems that reduce production risk while increasing delivery velocity. This is not a “test scripts only” role. You’ll shape automation strategy, embed quality into the SDLC, and help define how changes move safely from dev → staging → UAT → prod in a fast-moving healthcare platform. You’ll help define how quality scales as Cohere grows. This role has real influence over release safety, platform reliability, and how engineering teams ship software in a regulated, high-impact domain. You won’t just test features — you’ll shape how Cohere delivers them safely to production. What you’ll do: Own and evolve Cohere’s end-to-end test automation strategy across UI, API, config changes, and critical workflows Design and maintain scalable E2E automation frameworks for multi-tenant, payer-specific workflows Build automated validation for deployment guardrails, release readiness, and production change safety Partner with Platform/DevOps to integrate automation into CI/CD pipelines and deployment workflows Create automated coverage for high-risk paths (authorization flows, partner integrations, file pipelines, feature flags, config changes) Drive test reliability, flake reduction, and actionable failure signals Define and enforce quality gates for prod releases, blue/green and canary deployments, and config changes Collaborate with Product and Engineering to ensure business outcomes are testable, measurable, and observable Improve test data management and environment stability to enable reliable automation at scale Mentor engineers on testability, automation best practices, and quality-first development Partner with SRE and Security to ensure production readines

typescriptawsci/cd
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
15 days ago

Opportunity Overview: We’re looking for a senior-level automation engineer who will help raise the bar on release quality, environment reliability, and change safety across Cohere’s platform. You’ll partner closely with Product, Engineering, Platform, and SRE to build scalable automation, guardrails, and validation systems that reduce production risk while increasing delivery velocity. This is not a “test scripts only” role. You’ll shape automation strategy, embed quality into the SDLC, and help define how changes move safely from dev → staging → UAT → prod in a fast-moving healthcare platform. You’ll help define how quality scales as Cohere grows. This role has real influence over release safety, platform reliability, and how engineering teams ship software in a regulated, high-impact domain. You won’t just test features — you’ll shape how Cohere delivers them safely to production. What you’ll do: Own and evolve Cohere’s end-to-end test automation strategy across UI, API, config changes, and critical workflows Design and maintain scalable E2E automation frameworks for multi-tenant, payer-specific workflows Build automated validation for deployment guardrails, release readiness, and production change safety Partner with Platform/DevOps to integrate automation into CI/CD pipelines and deployment workflows Create automated coverage for high-risk paths (authorization flows, partner integrations, file pipelines, feature flags, config changes) Drive test reliability, flake reduction, and actionable failure signals Define and enforce quality gates for prod releases, blue/green and canary deployments, and config changes Collaborate with Product and Engineering to ensure business outcomes are testable, measurable, and observable Improve test data management and environment stability to enable reliable automation at scale Mentor engineers on testability, automation best practices, and quality-first development Partner with SRE and Security to ensure production readines

typescriptawsci/cd
View job →
V
Vanta
📍 United States• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's App Primitives team owns the shared building blocks that span data structures to UI: comments, notifications, event logs, feature flags, and task management. We own the general infrastructure; product teams own the logic built on top. Getting this layer right unlocks velocity for every team at Vanta. As the Engineering Manager, App Primitives at Vanta, you'll lead the team building the shared product primitives that every Vanta product team depends on, owning the infrastructure layer that makes collaboration, communication, and core workflows possible across the entire platform. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow the App Primitives team, owning hiring, team health, delivery, and the development of senior engineers and technical leads Own the strategy and roadmap for Vanta's shared product primitives: comments infrastructure, notifications platform, event log, feature flag system (Statsig), and task management Define and steward the interface model between App Primitives systems and product teams, ensuring product teams can build on top of shared infrastructure quickly and safely, without owning the underlying systems themselves Partner closely with product engineering leaders across Vanta to surface developer ne

REMOTErestairust
View job →
V
Vanta
📍 Toronto• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's App Primitives team owns the shared building blocks that span data structures to UI: comments, notifications, event logs, feature flags, and task management. We own the general infrastructure; product teams own the logic built on top. Getting this layer right unlocks velocity for every team at Vanta. As the Engineering Manager, App Primitives at Vanta, you'll lead the team building the shared product primitives that every Vanta product team depends on, owning the infrastructure layer that makes collaboration, communication, and core workflows possible across the entire platform. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow the App Primitives team, owning hiring, team health, delivery, and the development of senior engineers and technical leads Own the strategy and roadmap for Vanta's shared product primitives: comments infrastructure, notifications platform, event log, feature flag system (Statsig), and task management Define and steward the interface model between App Primitives systems and product teams, ensuring product teams can build on top of shared infrastructure quickly and safely, without owning the underlying systems themselves Partner closely with product engineering leaders across Vanta to surface developer ne

REMOTErestairust
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
15 days ago

Opportunity Overview: We’re looking for a senior-level automation engineer who will help raise the bar on release quality, environment reliability, and change safety across Cohere’s platform. You’ll partner closely with Product, Engineering, Platform, and SRE to build scalable automation, guardrails, and validation systems that reduce production risk while increasing delivery velocity. This is not a “test scripts only” role. You’ll shape automation strategy, embed quality into the SDLC, and help define how changes move safely from dev → staging → UAT → prod in a fast-moving healthcare platform. You’ll help define how quality scales as Cohere grows. This role has real influence over release safety, platform reliability, and how engineering teams ship software in a regulated, high-impact domain. You won’t just test features — you’ll shape how Cohere delivers them safely to production. What you’ll do: Own and evolve Cohere’s end-to-end test automation strategy across UI, API, config changes, and critical workflows Design and maintain scalable E2E automation frameworks for multi-tenant, payer-specific workflows Build automated validation for deployment guardrails, release readiness, and production change safety Partner with Platform/DevOps to integrate automation into CI/CD pipelines and deployment workflows Create automated coverage for high-risk paths (authorization flows, partner integrations, file pipelines, feature flags, config changes) Drive test reliability, flake reduction, and actionable failure signals Define and enforce quality gates for prod releases, blue/green and canary deployments, and config changes Collaborate with Product and Engineering to ensure business outcomes are testable, measurable, and observable Improve test data management and environment stability to enable reliable automation at scale Mentor engineers on testability, automation best practices, and quality-first development Partner with SRE and Security to ensure production readines

typescriptawsci/cd
View job →
OI
15 days ago

Job Overview: We are looking for a Senior GenAI Developer to design, build, and productionize agentic AI systems—LLM-powered agents that can plan, use tools, orchestrate workflows, and operate reliably under enterprise constraints. You will own key parts of the agent architecture (planning, tool use, memory, evaluation, safety/guardrails, and observability) and deliver end-to-end solutions across RAG, function/tool calling, multi-agent coordination, and scalable deployment. Key Responsibilities Design and implement agentic systems: single-agent and multi-agent architectures (planner/executor, supervisor-worker, routing, reflection, critique, task decomposition). Build robust tool-using agents: function calling, tool schemas, tool authorization, retries, rate limiting, and sandboxing. Implement RAG + memory patterns: retrieval strategies, hybrid search, context assembly, long-term memory, and grounding/citation behaviors. Develop workflow orchestration for agent execution (state machines/graphs), concurrency controls, and deterministic execution where possible. Productionize GenAI services: APIs, background jobs, streaming responses, caching, and cost/latency optimization. Establish agent evaluation: golden sets, simulation-based evals, LLM-as-judge with mitigations, task success metrics, regression testing. Build observability and safety: tracing, token/tool telemetry, anomaly detection, prompt injection defenses, data leakage prevention, policy enforcement. Collaborate with product, security, and platform teams to deliver enterprise-ready solutions and integrate with internal systems (data, identity, workflow). Mentor engineers, set coding standards, and contribute to architecture reviews and technical roadmaps. Required Qualifications 6+ years software engineering experience; 2+ years building LLM/GenAI systems in production. Strong programming skills in Python (required) and/or TypeScript/Node.js. H

typescriptpythonnode.js
View job →
G
15 days ago

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary The Workplace Coordinator is responsible for the day-to-day operation and maintenance of workplace infrastructure, ensuring that all building systems operate efficiently, safely, and reliably. This role oversees HVAC systems, including chillers and AHUs, Building Management Systems (BMS), electrical panels, and coordinates preventive and corrective maintenance activities to support uninterrupted business operations. Key Responsibilities Technical Operations · Monitor and maintain HVAC systems including Chillers, AHUs, FCUs, and ventilation systems. · Operate and monitor Building Management System (BMS) for alarms, trends, and equipment performance. · Inspect electrical LT panels, UPS systems, DG synchronization (if applicable), and power distribution systems. · Monitor critical utilities including temperature, humidity, pressure, and energy consumption. · Ensure uninterrupted operation of critical infrastructure and respond promptly to system failures. Preventive & Corrective Maintenance · Plan and execute preventive maintenance schedules for HVAC and electrical systems. · Coordinate breakdown maintenance with

aileanprocurement
View job →

About the Team OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future. The Operating Systems team is critical in this mission, turning sophisticated hardware and AI capabilities into a reliable, trusted platform. Security is central to that work: we define trust boundaries, integrate hardware-backed protections, isolate sensitive context, and establish the guardrails that let AI applications and agents act safely, privately, and under user control. About the Role We’re looking for a Software Security Architect to define the security architecture for OpenAI’s next-generation operating system. You’ll work alongside hardware security architects and partner with operating system, silicon, firmware, privacy, and product teams to protect users, their devices, and their data. This is a senior, hands-on role for someone who can connect operating system internals, hardware-backed security, and real-world product constraints. Your work will shape the platform’s trust model, protect sensitive information, and set the technical foundations for AI that is safe, private, and under user control. In this role, you will: Define the operating system’s security architecture, trust boundaries, privilege model, and protections for sensitive user data. Partner with hardware security architects to integrate roots of trust, secure elements, trusted execution environments, and processor security capabilities into the operating system. Desig

artificial intelligenceaic++
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
1mo ago

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Electrical Engineer to build and own the electrical backbone of our robotic actuator dynamometer and test infrastructure. You will design, integrate, and operate the load motor drives, power distribution, instrumentation wiring, DAQ interfaces, and safety systems that make high-performance robotic actuator testing repeatable, safe, and scalable. This role spans hands-on lab execution and system architecture: selecting and commissioning power electronics, designing robust test-cell electrical systems, bringing up sensors and DAQ, and partnering with mechanical and software engineers to turn robotic actuator hardware into trustworthy data. In this role, you will Own the electrical architecture of dynamometer and actuator test cells, from mains distribution and protection through load motor drives, braking, and auxiliary power. Specify, integrate, commission, and tune motor drives and load machines for robotic actuator torque, speed, efficiency, thermal, and durability testing. Design power distribution, grounding, shielding, cable routing, and connectorization for high-current, high-voltage, and low-level measurement systems. Integrate torque, position, speed, temperature, voltage, current, vibration, and other instrumentation from robotic actuators into DAQ and control systems. Develop electrical schematics, wiring diagrams, panel layouts, harness documentation, and test-cell interface definitions. Build, debug, and maintain test-cell electrical hardware, rapidly diagnosing noise, EMI, grounding, drive, sensor, and power-q

REMOTEartificial intelligenceai
View job →
DU
15 days ago

About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We are hiring a Firmware Validation & Integration Engineer for our autonomy software team. This is a critical role to build robust and scalable validation for our firmware and systems to ensure reliability at every level. In this role, you will work with our electrical, firmware, and autonomy engineers to build the infrastructure and test suites required to validate the system. This includes designing and implementing our Hardware-in-the-Loop (HIL) simulation environments and automation frameworks from the ground up. You will report to the Autonomy Platform Lead on our Autonomy Platform Team at DoorDash Labs. We expect this role to be hybrid with some time in-office and some time remote. You’re excited about this opportunity because you will… Play an integral role on a small and focused team. Design and build Hardware-in-the-Loop (HIL) systems to simulate vehicle dynamics and sensor data for comprehensive firmware and system-level validation. Develop automated test infrastructure and software tools to exercise multiple embedded platforms throughout our robot system. Interface many layers of our control system including vehicle controls, power management, and motion control to ensure seamless system integration. Implement low-level test sequences and validation algorithms to safely stress-test vehicle components such as batteries, drive-train, and thermal management devices. Collaborate with cross-functional teams to identify edge cases and hardware-software corner cases that impact vehicle safety and performance. We’re excited about you because… BS/MS degree in Computer Science, Robotics, Electrical Engineering, or related technical field. 5+ years of experience in validati

pythonawsgit
View job →

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. As the Site Reliability Engineering Manager for SLS, you will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll help grow and lead a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. We are looking to speak to candidates who are based in Dublin for our hybrid working model. Responsibilities Build and lead a team of 6-8 engineers, fostering a positive culture, handling career growth and performance conversations, and proactively removing blockers Define and drive a clear technical vision and comprehensive roadmap for our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs Contribute through hands-on technical work, such as leading architectural design reviews, reviewing PRs, and stepping in to guide the team through complex operational challenges Act as the primary liaison for the Storage Layer Services SRE team, collaborating closely with other engineering leaders to ensure platform alignment and manage stakeholder expectations You may be a good fit if you Have 10+ years of experience working on software and operating distributed systems, with 2+ years managing engineering teams Possess a customer-focused mindset, treating internal developers as your primary users Value efficiency in processes and operations, and have a track record of optimizing team workflows Pr

mongodbawsazure
View job →
🔔

Get new senior system software safety engineer jobs by email

Daily job updates · Unsubscribe anytime