Jobiba hiring network

Staff Software Reliability Engineer Data Platform Jobs

3,518 active opportunities · Updated for October 2026

Fresh results

13 shown

Explore current staff software reliability engineer data platform jobs. Use filters to narrow by work mode, employment type, experience and date posted.

E
16 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Pipeline Engineering team at Everpure™ as a Software Engineer to build, own, and operationally scale the microservices and Temporal workflows driving continuous integration for FlashArray, FlashBlade, and Hyperscale. In this developer-first role based in Bangalore, you will design durable production services and agentic AI tools that directly reduce time-to-signal, optimize compute utilization, and absorb operational toil across our global engineering workforce. WHAT YOU'LL DO Architect Deterministic CI Workflows: Design and deploy production microservices and durable Temporal workflows that make pipeline execution resilient, resumable, and fully debuggable for enterprise storage platforms. Optimize Compute & Testbed Utilization: Extend in-house scheduling logic and placement algorithms across bare-metal hardware and VM fleets to minimize queue times and maximize infrastructure efficiency. Build Operational AI Agents: Engineer RAG pipelines and agentic workflows over failure logs, vector stores, and test metadata to automate root-cause analysis, flake classification, and developer triage. Drive Platform Observability & Ownership: Establish key platform metrics—including time-to-signal and pass/flake rates—while operating your services end-to-end to ensure long-term stability and eliminate recurring failure modes. Collaborate Across Product Teams: Partner with platform and product engineering group

pythonsqlaws
View job →
E
Everpure
📍 Bengaluru• Full-time
16 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Pipeline Engineering team at Everpure™ as a Software Engineer to build, own, and operationally scale the microservices and Temporal workflows driving continuous integration for FlashArray, FlashBlade, and Hyperscale. In this developer-first role based in Bangalore, you will design durable production services and agentic AI tools that directly reduce time-to-signal, optimize compute utilization, and absorb operational toil across our global engineering workforce. WHAT YOU'LL DO Architect Deterministic CI Workflows: Design and deploy production microservices and durable Temporal workflows that make pipeline execution resilient, resumable, and fully debuggable for enterprise storage platforms. Optimize Compute & Testbed Utilization: Extend in-house scheduling logic and placement algorithms across bare-metal hardware and VM fleets to minimize queue times and maximize infrastructure efficiency. Build Operational AI Agents: Engineer RAG pipelines and agentic workflows over failure logs, vector stores, and test metadata to automate root-cause analysis, flake classification, and developer triage. Drive Platform Observability & Ownership: Establish key platform metrics—including time-to-signal and pass/flake rates—while operating your services end-to-end to ensure long-term stability and eliminate recurring failure modes. Collaborate Across Product Teams: Partner with platform and product engineering group

pythonsqlaws
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Overview We are looking for a strong software and platform engineer to join our Production Engineering team in Bangalore as an individual contributor in FA ProductionEng APJ. This role will help build and operate internal platforms that improve how we provision, observe, govern, and troubleshoot engineering infrastructure at scale. The fleet management use cases that give teams a single place to understand and operate the test infrastructure. If you enjoy building internal platforms that remove friction, improve visibility, and make engineering teams faster and more effective, this role is for you. Why This Role Is Unique This is not a typical application development role.You will work on internal platforms that directly shape how engineering teams consume and manage shared infrastructure. The role spans platform engineering, workflow automation, observability, API-driven services, and infrastructure lifecycle management. The right candidate will work on systems such as: Self-serviceability workflows and lease-based testbed governance. Developer Platform dashboards and APIs used for triage, visibility, and product trend observation. Testbed and workflow orchestration across fleet management domains. Impact This role is a high-leverage engineering investment. The work will improve how engineering teams provision testbeds, understand failures, operate shared infrastructure, and move faster with less friction. Better

awskuberneteslinux
View job →
V
Verse
📍 San Francisco• Full-time• $150K – $240K/yr
16 days ago

Location: San Francisco, CA (Remote/Hybrid Available) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role You will be a member of the technical staff developing product experiences for our Dispatch Intelligence users – customers who want and have battery energy storage systems for additional energy cost savings or faster interconnection times. In this role, you will serve in a “full stack” capacity designing, building, and maintaining frontend and backend components of our energy storage suite of applications. We use Typescript, React, Next.js, Tailwind CSS, Radix/ShadCN, Jest, Cypress, Playwright, Vitest, and Storybook with Echarts and D3/Observable for data visualization for our frontend, and Cloudflare Pages for hosting and content delivery. We rely on identity and auth platforms like Clerk for sign-in flows. Our backend is written in Go and Python with Postgres/AlloyDB and blob storage for data persistence. Key Responsibilities Foster a culture and mindset of well-designed systems, test-driven software, and proactive communication with a high degree of transparency, mutual resp

typescriptpythonreact
View job →
N
Nuro
📍 Mountain View• Full-time• From $160.4K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Work Lead system-level requirements definition, architecture design, integration strategy, and validation in one of the platform areas — autonomy sensors, AV compute, motion control, pose/localization, or end-to-end latency. Characterize and quantify system performance for that area: define nominal targets, minimum performance, and acceptance criteria across all ODDs, and defend those numbers with data rather than convention. Recommend operating constraints for current and future ODDs on different vehicle platforms — for example trajectory constraints for a given virtual driver, or sensor performance envelopes. Define and quantify test coverage for your area, including the realism and relevance of proposed tests across simulation, SIL/HIL, closed course, and on-road. Partner with software teams to define interfaces, performance metrics, failure modes, and validation strategies; validate platform modules against ground truth whe

pythonaic++
View job →
G
Gitlab
📍 United States• Full-time• Remote
1mo ago

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role Join us at a pivotal moment as we scale our revenue operations globally. As Chief of Staff to the CRO, you'll serve as a strategic partner and force multiplier to our revenue leadership team, driving operational excellence and orchestrating high-impact initiatives. As part of the Office of the CRO, you'll also manage our Communications lead and CRO Initiatives Program Manager, giving you direct ownership over how the CRO's key initiatives get built and run. This highly visible role offers a unique opportunity to shape strategy at the executive level while developing your path to senior leadership. Y

REMOTEgitrestai
View job →
U
1mo ago

Upwork Inc.'s (Nasdaq: UPWK) family of companies connects businesses with global, AI-enabled talent across every contingent work type including freelance, fractional, and payrolled. This portfolio includes the Upwork Marketplace, which connects businesses with on-demand access to highly skilled talent across the globe, and Lifted, which provides a purpose-built solution for enterprise organizations to source, contract, manage, and pay talent across the full spectrum of contingent work. From Fortune 100 enterprises to entrepreneurs, businesses rely on Upwork Inc. to find and hire expert talent, leverage AI-powered work solutions, and drive business transformation. With access to professionals spanning more than 10,000 skills across AI & machine learning, software development, sales & marketing, customer support, finance & accounting, and more, the Upwork family of companies enables businesses of all sizes to scale, innovate, and transform their workforces for the age of AI and beyond. Since its founding, Upwork Inc. has facilitated more than $30 billion in total transactions and services as it fulfills its purpose to create opportunity in every era of work. Learn more about the Upwork Marketplace at Upwork.com and follow us on LinkedIn , Facebook , Instagram , TikTok , and X ; and learn more about Lifted at Go-Lifted and follow on LinkedIn . About the Role Upwork's COO and GM of Marketplace operates across one of the broadest organizational portfolios in the company - spanning Legal, Information Security, Marketing, Payments, Trust & Safety, Communications, Customer Support, Design, and Product. The Chief of Staff to the COO exists to make that model work: not as a coordinator, but as a true operating partner who extends the COO's capacity, sharpens decision-making, and holds together a complex, fast-moving organization. This role requires someone who has earned the confidence to walk into any room, speak with authority, and immediately earn cre

reactrestmachine learning
View job →
PE
Private Employer
📍 Minneapolis• Full-time• Hybrid• $220K – $270K/yr
1mo ago

Perforce is a community of collaborative experts, problem solvers, and possibility seekers who believe work should be both challenging and fun. We are proud to inspire creativity, foster belonging, support collaboration, and encourage wellness. At Perforce, you’ll work with and learn from some of the best and brightest in business. Before you know it, you’ll be in the middle of a rewarding career at a company headed in one direction: upward. With a global footprint spanning more than 80 countries and including over 75% of the Fortune 100, Perforce Software, Inc. is trusted by the world’s leading brands to deliver solutions for the toughest challenges. The best run DevOps teams in the world choose Perforce. Position Summary: At Perforce, we build more than software. We build careers, communities, and impact. Our people are the driving force behind innovation, delighting customers, and global growth. Whether you're a builder or a coach, you’ll find purpose, connection, and opportunity here. We care deeply, move fast, and deliver excellence - together. Over half of the Global 500 rely on Perforce to move their business forward – Join us! We are seeking a strategic and execution-oriented Chief of Staff to partner directly with the CEO in scaling the company's business operations and global programs, while also acting as a high-impact driver of the CEO & Executive Team’s special initiatives and strategic agenda. This role sits at the intersection of strategy and execution, owning critical operating functions including Revenue Operations, Business Systems, Sales Operations, and Global AI Infrastructure while also leading a portfolio of special projects, related to strategic growth of the business. The ideal candidate is a versatile operator who moves fluidly between high-level strategic thinking and hands-on execution, serving as a force multiplier for the CEO and connective tissue across the leadership team.

airustdevops
View job →
M
1mo ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for strong backend engineers who love building a developer tools used by the largest AI companies in the world. You’ll be building for things at scale, but also for new AI workflows that change every day. Requirements: Experience building and shipping modern web applications end-to-end. We care more about what you’ve built than how many years you’ve been building. Comfort working across the stack: TypeScript on the frontend, Python services on the backend, and ClickHouse for data and analytics. Deep knowledge of observability tools and patterns used for large-scale workloads such as custom sandboxes, training and inference for large language (LLM) and diffusion models. Experience with at least one of: billing/payments systems, B2B SaaS tooling, or enterprise software, or LLM / diffusion models inference and training loads. Strong product instincts; yo

typescriptpythonai
View job →
N
Nuro
📍 Mountain View• Full-time• From $176.4K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role As a Senior/Staff Systems Engineer working on Autonomy Verification, you are responsible for verifying that the Nuro Driver is safe to deploy in our target ODD and complies with the rules of the road. This requires prior experience with the development or validation of autonomous systems, and a collaborative nature to work closely with a variety of teams across Nuro: Autonomy Software, Simulation, Product, and Operations. You will have end-to-end ownership from requirements definition, metrics design, and test strategy development. About the Work Create requirements for an autonomous system, ensuring safe operation within its ODD and compliance with the rules of the road. Design generalizable metrics and acceptance criteria to verify that the autonomous system satisfies these requirements, leveraging safety standards and methodologies. Propose and leverage diverse test strategies - synthetic and log simulation, on-ro

pythonaic++
View job →
N
Nuro
📍 Mountain View• Full-time• From $193.9K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role Our robotics team is growing and we are looking for a Software Engineer to join our Sensor Data and Calibration team. We are searching for an engineer with robotics and machine learning expertise to develop synthetic sensor simulation models and algorithms. The ideal candidate has hands-on experience in the research, development, and implementation of machine learning methods (e.g., NeRF or Gaussian splatting) for generating synthetic sensor data (photorealistic images, realistic lidar and/or radar, etc.). About the Work Research, develop, and implement state-of-the-art synthetic sensor simulation methods Analyze and characterize the realism and utility of synthetic sensor data Answer critical questions about sensor data and autonomy performance Collaborate with stakeholders across autonomy, infrastructure, and systems teams on map needs and requirements Role is scoped as a Senior/Staff IC with the flexibility to grow into

pythonmachine learningai
View job →
N
Nuro
📍 Mountain View• Full-time• From $183.8K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Team Our robotics team is growing and we are looking for an ML Software Engineer to join our Online Mapping team. We are searching for an engineer with robotics and machine learning expertise to work on challenging problems in the design and implementation of onboard mapping models and algorithms, data and label management, as well as training pipelines. About the Role Building robust ML and/or mapping systems that work with real data (cameras, LiDAR, etc.) in uncertain environments, and a strong desire to contribute to the future of robot navigation for logistics and transportation. About the Work Research, develop, and implement state-of-the-art online mapping models and algorithms. Analyze and characterize the performance of the online mapping system, identifying opportunities for architecture, data or evaluation improvements in a E2E ML system. Work cross functionally with other ML teams to integrate our models

pythonmachine learningai
View job →
N
Nuro
📍 Mountain View• Full-time• From $193.9K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role Controller is the critical link between Nuro Driver's upstream autonomy stack and physical vehicle platform — translating planned trajectories into safe, precise, real-world motions. As a Senior/Staff Controls Engineer, you'll design and deploy production controls software that powers L4 autonomous driving across an expanding operational design domain, from parking lots to highways, in all weather and road conditions. You'll work alongside world-class engineers across autonomy, hardware, and systems to solve problems that don't have textbook answers and see your work running on real vehicles globally. About the Work Develop robust, reliable and optimized production-level control software following safety standards (ISO 26262). Develop seamless interactions between planner and controller and advanced control strategies for expanded ODD including parking lot maneuvers through highway driving under various weather condi

machine learningaic++
View job →
🔔

Get new staff software reliability engineer data platform jobs by email

Daily job updates · Unsubscribe anytime