Jobiba hiring network

Software Reliability Engineer Jobs

6,326 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current software reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

L
Lyft
📍 Toronto• Full-time• From C$46/hr
25 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. With over half a billion rides and counting, Lyft is solving hard problems in a flourishing domain with a lot of data and creative solutions in Marketplace, Mapping, Fraud, Growth and beyond. We're actively building the next-generation Machine Learning (ML) platform for low-cost, ultra-immersive transportation to improve people’s lives using modern ML with peta-byte scale data. Our Machine Learning Engineers are excited to work on these challenging problems and redefine solutions to directly impact various aspects of Lyft's primary business. If you are a student with experience in machine learning workflows, passionate about solving challenging problems using data and working in a dynamic, creative, and collaborative environment, this opportunity is for you! Responsibilities: Contribute to the design, build, train and test of Machine Learning models Write production-level code to convert ML models into working pipelines Partner with Product Managers, Data Scientists, and fellow ML Engineers to frame Machine Learning problems within the business context Analyze experimental and observational data, communicate findings to support decisions Participate in code and spec reviews to ensure code quality and distribute knowledge Experience: Currently pursuing a Bachelor's, Master's, or PhD degree in Computer Science or a related technical field from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required). For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for the internship in Toronto Good understanding and knowledge of ML libraries like scikit-learn, Tensorflow, PyTorch, Keras, MXNet, et

pythonmachine learningai
View job →
C
Cohere
📍 San Francisco• Full-time
26 days ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? The Data Infrastructure team at Cohere is responsible for the storage and data movement layer underlying every model training run. We're building the unified storage layer that feeds our training workloads. It needs to serve petabytes of training data and model checkpoints fast enough to keep thousands of GPUs busy across several training clusters. In this role, you’d have an opportunity to build this system from the ground up. You’d be a key contributor, working on a problem few teams have had to solve at this scale. In this role, you will: Design, build, and operate the distributed storage system that feeds model training and evaluation. Run this system multiple on Kubernetes clusters at petabyte scale. Work with researchers and training-infra teams on how jobs actually read and write data, and turn that into throughput, latency, and durability requirements Work through the networking, I/O, and consistency problems of moving large datasets and checkpoints across regions and backends, with GPU idle time and time-to-insight as the measures of success You may be a good fit if you have: Strong storage fundamentals,

pythonkubernetesgit
View job →
L
Lyft
📍 Mexico City• Full-time
26 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Every day, millions of riders and drivers depend on Lyft to get where they're going. When something goes wrong along the way, they expect us to make it right — quickly, clearly, and without friction. How fast and how well we resolve those moments shapes whether people keep choosing Lyft. The Self-Serve Intelligence team, within the Safety & Customer Care org, is composed of engineers building the AI-powered systems that do exactly that: resolving rider and driver suboptimal experiences without agent involvement through AI Assist (e.g. AI Agents), automations, and self-serve workflows. Our goal is to make getting help feel effortless. We design and build backend services, APIs, and GenAI-powered products that combine robust engineering with applied AI to deliver reliable, scalable self-serve experiences. We are looking for a highly motivated, collaborative, team-focused and technically strong Software Engineer to join our Self-Serve Intelligence team. As a member of this team, you will build the services and AI-powered products that resolve customer issues autonomously. Every day, you'll partner with machine learning engineers, product, design, data science, and operations on high-impact projects — from shipping new AI Agent capabilities, to building the evaluation pipelines that keep their quality high, to improving the backend services underneath them. You'll bring strong engineering instincts, genuine curiosity about applied AI, and a willingness to work through ambiguity in a space that changes month to month. Responsibilities: Write well-crafted, well-tested, readable, and maintainable code Partner with senior engineers to design, build, and ship backend services and GenAI-powered products (e.g. AI Agents) that resolve rider and driver suboptimal experiences Independently lead tasks fro

pythonjavaredis
View job →
O
27 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. About the Team The Professional Services R&D team is a new, dynamic group at the forefront of innovation within Okta. Our mission is to design and build reusable, scalable assets and tools that empower our delivery teams and partners. By making customer engagements more efficient, streamlined, and cost-effective, we directly contribute to our customers' success and accelerate their time-to-value with Okta. This is a unique opportunity to join a strategic team from the ground up and shape the future of Okta's professional services. Position Summary As a Software Engineer on the R&D team, you will be a key player in developing the next generation of professional services assets. You will have the opportunity to work with a wide array of technologies, including Java, React, Node, Python, and .NET, to create solutions that will be leveraged by thousands of customers globally. We are looking for a creative and passionate engineer who thrives in a fast-paced, collaborative environment and is excited about building tools that have a massive impact. Responsibilities Design, develop, test, and deploy high-quality, reusable software assets and tools. Collaborate with technical architects, delivery consultants, and product managers to identify opportunities for innovation and define project requirements. Write clean, maintainable, and well-documented code across a variety of technology stacks. Create proofs-of-concept and prototypes to explore new ideas and te

pythonjavareact
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
27 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are looking for experienced backend software engineers to join our claims tech engineering team. Our vision is to tangibly reduce risk on the Lyft platform, and by extension reduce insurance cost. Our team is dedicated to centralizing the entire claims operation onto a unified risk platform. This consolidation of data, workflows, and communications aims to foster proactive measures, enhance efficiency, ensure consistency, and provide valuable insights. These efforts are designed to effectively reduce claims costs as Lyft's operations expand. Additionally, our team is responsible for maintaining robust relationships with our third-party insurance partners, guaranteeing timely, proactive, and precise sharing of claim data. Responsibilities: Write well-crafted, well-tested, readable, maintainable code Own feature from product spec to successful high quality development, deployment and maintenance Participate in code reviews to ensure code quality and distribute knowledge Respond to external questions and requests. Unblock, support and communicate with stakeholders to achieve results Experience: 3+ years of relevant professional experience Experience with object-oriented programming Experience in distributed systems Experience working with databases, relational or NoSQL Write clear, scalable and clear design documentation Design, build and improve a set of team owned components Benefits: Extended health and dental coverage options, along with life insurance and disability benefits Mental health benefits Family building benefits Child care and pet benefits Access to a Lyft funded Health Care Savings Account RRSP plan with company match to help save for your future In addition to provincial observed holidays, salaried team members are covered under Lyft's flexible paid time off policy. The policy allows

R
27 days ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Tokenization team's mission is to bring both public and private equity to be tradable 24/7 on decentralized exchanges (DEXs) across the Robinhood Chain. We work at the intersection of blockchain technology, financial market infrastructure, and modern distributed systems - with a focus on accessibility, security, and scalability. Equity markets have been closed nights and weekends for a century; we're building the infrastructure that changes that. As a Software Engineer , you'll build and own backend services that make tokenized equities work end to end - issuance, custody, and settlement on-chain, kept in sync with traditional brokerage and ledger systems. You'll ship product-facing features and platform capabilities in a domain where correctness is non-negotiable - the systems you build move real customer assets, continuously, around the clock. This role is based in our Toronto, ON office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-perfor

pythonjavaaws
View job →
A
27 days ago

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join The Quality Platform team is at the heart of Airbnb’s mission to deliver a seamless, high-quality experience for millions of hosts and guests. We don’t just find bugs — we build the systems that prevent them. Our team sits at the intersection of Quality Engineering, Infrastructure, and Applied AI. We are evolving how software quality is built by integrating LLMs, intelligent automation, and data-driven systems into the testing lifecycle. You will join a high-impact group of engineers focused on building AI-powered quality systems that scale across one of the world’s most complex codebases. The Difference You Will Make: As a Mobile Software Engineer, you will be a key contributor to the development of our Quality Platform across both native mobile stacks. You will help build the foundation for how quality is engineered across Airbnb's iOS and Android ecosystems, developing tools and frameworks that enable our mobile platform to scale while keeping developers productive and confident, regardless of which platform they build on. In this role, you will: Build AI-Driven Solutions: Contribute to AI-native agents that automate repetitive testing tasks and provide intelligent feedback to developers on both platforms.Deliver Scalable Infrastructure: Develop and maintain the high-scale platforms and testing environments used daily by the iOS and Android engineering organizations.Promote Engineering Craft: Implement best-in-class mobile patterns and modularity to improve testability and fault-tolerance across both native codebases.Contribute to Operational Excellence: Ensure our automated syste

ci/cdaikotlin
View job →
R
Replit
📍 Foster City• Full-time• Remote
28 days ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: As a New Grad Software Engineer, you'll join a team of exceptional builders working on products that are reshaping how the world creates software. You'll have the opportunity to work on everything from our AI-powered development platform to the distributed systems that enable real-time collaboration for millions of developers. This is a chance to define your career while defining the future of software development. You'll work on problems that matter, with the autonomy to drive solutions and the support to grow into a technical leader. What you will build: Product features that delight users and make it possible for anybody to create software AI coding agent that understands intent and generates production-ready applications Cloud infrastructure that provides instant, powerful development environments at global scale Platform features that enable one click deployments and scale to millions of users Required skills and experience: Recent graduate (2027) with a degree in Computer Science, Computer Engineering, or related field Strong programming skills in a modern language (JavaScript/TypeScript, Python, Go, Rust) Full-stack capabilities with experience in React, Node.js, and database technologies Growth orientation - eager to learn new technologies and take on increasing responsibility Collaborative spirit - you work well in cross-functional teams and value diverse perspectives What we value : Problem-solving mindset: Ability to approach complex operational challenges systematically and devise effective solutions Self-directed and autonomous: Capable of working independently while collaborating effectively with cross-functional teams Strong communication skills: Ability to explain complex technical conce

REMOTEjavascripttypescriptpython
View job →
R
Ramp
📍 San Fransisco• Full-time• Remote• From $10K/yr
28 days ago

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role As a Software Engineer, Forward Deployed, you’ll solve the hardest problems standing between Ramp and the world’s largest and most complex companies — shaping and shipping the product capabilities that unlock our growth upmarket. You'll be part of our Core FDE org, which is an agent-first, high-pace, customer-facing engineering team. On FDE, you will interact directly with customers and deliver solutions end to end — understanding pain points, shaping product decisions, and building agents that autonomously expand Ramp's capabilities. Check out our Engineering Blog and FDE post for more context on our work! What You’ll Do Deliver software end to end that meet the needs of our largest customers — understanding user pain points, scoping product specs, and building agents that autonomously implement solutions. Collaborate closely with Sales, Solutions, Customer Success, and Account Management to close deals, activate customers, and expand the value Ramp provides over time. Drive the core product engineering roadmap through our embedding

REMOTErestaigo
View job →
W
Writer
📍 San Francisco• Full-time• Remote
29 days ago

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We're looking for an exceptional software engineer to join our rapidly evolving team at WRITER. In this pivotal role, you'll be at the forefront of expanding human capacity by building the next generation of AI-powered solutions that transform how leading enterprises operate. You'll dive deep into developing a state-of-the-art platform that leverages cutting-edge generative AI technologies, from large language models to sophisticated agentic workflows, delivering seamless, scalable, and secure applications that redefine enterprise productivity. This is an unparalleled opportunity to make a tangible impact, shaping the future of AI and contributing to a product that’s changing how the world works. This role is hybrid, based out of our San Francisco, New York City, or Seattle hubs. You'll report to our senior director, engineering . 🦸🏻‍♀️ What you’ll do Design and deliver secure, scalable AI integration platforms that connect enterprise systems and power missio

REMOTEtypescriptpythonnode.js
View job →
G
Godaddy
📍 India• Full-time
29 days ago

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team We are looking for a Software Development Engineer to assist in building and strengthening our Identity UI team. GoDaddy is a global products seller to customers from all over the world. Our Identity UI serves login, account create as well as profile and preferences for all GoDaddy’s customers. The UI’s we build are among the first the users see on their customer journey, and helps the user keep their profile up to date. Does that sound exciting to you? Then you are made for the role. What you'll get to do... Develop and enhance user-facing features using React Optimize application performance, load times, and overall user experience while modernizing existing UIs Build and maintain unit, integration, and end-to-end test coverage, delivering through CI/CD pipelines Use AI-assisted development tools as part of your daily workflow while critically reviewing and validating generated code Collaborate within a global Agile team alongside product, design, and backend engineering partners Your experience should include... 3+ years of software engineering experience building large-scale applications or distributed systems Strong proficiency in JavaScript, HTML, and CSS, with an understanding of React.js and modern frontend development principles Experience working with RESTful APIs and a foundational understanding of accessibility (a11y) best practices Experience using version control systems (Git preferred), working in Agile teams, and testing and deploying code through CI/CD pipelines (e.g., GitHub Actions) Experience using AI-assisted

javascriptjavareact
View job →
T
Twilio
📍 India• Full-time• Remote
29 days ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . Recruiter: Aishwarya Gopalakrishna Hiring Manager: Vivek Kumar Tiwari Level: P3 See yourself at Twilio Join the team as Twilio’s next Senior Software Engineer (P3) About the job This position is needed to build new features and capabilities in the Twilio Segment team. As a Senior Software Engineer on this team, you’ll build and scale systems that process several hundred thousands of data points per second. Helping our customers unlock value of their data from external systems and activate it to various destinations via Segment capabilities. You will work on high-scale ingestion and data processing systems. We iterate quickly on these products and features and learn new things daily — all while writing quality code. We work closely with product and design and solve some of the toughest engineering problems to unlock new possibilities for our customers. If you get excited by building products with high customer impact — this is the place for you. Responsibilities In thi

REMOTEpythonjavadocker
View job →
S
Stripe
📍 Singapore• Full-time
1mo ago

Who we are About Stripe Stripe is a technology company focused on improving the conditions for economic growth and prosperity. We build programmable financial infrastructure, rethinking from first principles how financial services should work, to make it easier and cheaper for any business to start and scale. More than 10 million businesses build on Stripe, spanning the economic frontier—from solo founders to established enterprises—united by a practical focus on growth. The most ambitious companies in the world use Stripe as core infrastructure to grow faster. They process trillions of dollars a year on Stripe, equivalent to around 1.6% of global GDP. While economic growth makes everyone better off, open markets also enable greater variety. When any business can easily serve a global customer base, the quality and diversity of products in the world increase, and craft and creativity are unleashed into the smallest niches. Our own growth is wholly contingent on the success of the businesses building on Stripe. We therefore invest back into our technology at an unusual rate. We make upgrades to our products every single day to deliver compounding gains to our customers. We maintain some of the most reliable APIs on the internet. We build entirely new pieces of financial infrastructure to enable new ideas. And our significant advances in risk and fraud infrastructure over many years are making the internet economy safer and more accessible. Though people at Stripe don’t tend to take themselves seriously, Stripe is a fairly serious place: our customers are depending on us for their livelihoods. We admire ambition, intensity, curiosity, humility, and rigor. The most effective people become knowledgeable about many domains besides their own. Any company is an applied exercise in understanding some aspect of society or the market. In working with so many (especially the new and innovative ones), we think that Stripe is one of the very best places to learn about how

javascriptjavaai
View job →

We are hiring a Security Software Engineer to design and implement the hardware-backed security foundations used across OpenAI’s device ecosystem. A central focus of this role is hardening the boundary between our policy systems and the HSMs that protect sensitive cryptographic keys. This boundary determines which operations may be performed, what may be signed, which policies must be satisfied, and how changes to trusted software and policy are authorized. You will develop security-critical software and firmware within, or immediately adjacent to, an HSM trust boundary. Depending on your background, this may include HSM trusted applications, firmware services, cryptographic mechanisms, device drivers, PKCS#11 components, secure-provisioning protocols, or signing-policy enforcement systems. This is a hands-on software-engineering role. You will be expected to design systems, write and review production code, debug across hardware and software boundaries, and carry projects from initial requirements through deployment. It is not an HSM administration, PKI operations, compliance, or architecture-only position. In This Role, You Will Design and implement security-critical software and firmware for HSMs, secure elements, trusted execution environments, and hardware roots of trust. Build and harden the policy-to-HSM boundary responsible for authorizing certificate issuance and cryptographic signing operations. Develop HSM trusted applications, firmware components, host interfaces, device drivers, SDKs, or cryptographic service integrations. Implement or extend cryptographic interfaces such as PKCS#11, OpenSSL providers or engines, platform key-storage APIs, or comparable hardware-security interfaces. Build firmware and software that cryptographically enforces key generation, provisioning, usage, rotation, recovery, and destruction policies. Design and implement HSM-backed certificate authority, code-signing, key-management, and device-identity systems. Develop end-to-end

awsgitrest
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE The largest, most demanding enterprises run on Baseten, and they bring exacting requirements for how people, services, and agents access the platform. This is the founding role for our identity and authorization team within enterprise engineering. You'll own the identity and access layer of the Baseten platform: the authorization model, credential systems, and admin experiences that enterprise IT teams use to govern access for organizations like Harvey, HubSpot, and Notion. You'll design and build Baseten's fine-grained authorization system from the ground up to support the workflows customers depend on today while giving them cleaner, more precise ways to manage access as the platform grows. Authorization at Baseten requires low-latency permission checks at high request volume, consistent contracts and behaviors across the product suite, and strong security guarantees for mission-critical, highly regulated workloads. EXAMPLE INITIATIVES Recent and upcoming work in this area: Fine-grained authorization for users, service accounts, and agentic workloads: per-resource permissions at the organization, team, and workload scope to support both common workflows and complex enterprise access policies Programmatic authentication allowing high-compliance customers to connect service principles securely via short-lived, workload-based credentials Agent credentials that grant an agent exactly the access it needs for the gi

REMOTEpythonkubernetesmachine learning
View job →
🔔

Get new software reliability engineer jobs by email

Daily job updates · Unsubscribe anytime