At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You'll Play: Build and maintain performant backend systems and applications that drive real-world experiences Partner with Product, Design, and QA to bring features to life from ideation through deployment, always iterating with the end-user in mind. Champion engineering best practices—automated testing, peer reviews, observability, and elegant design Lead and influence architecture decisions that prioritize scalability and simplicity <span data-contrast=&quo
Jobs in United States
Software Reliability Engineer in United States
2,142 active opportunities · Updated October 2026
Showing
15 jobs
Explore current software reliability engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. About the Team Cloudflare handles traffic for almost 25% of the Internet. That’s a lot of data. On the Town Lake team, our mission is to make that data accessible and valuable for users across the company. We connect data from dozens of source systems and make it available so that any user in the company can answer any question in 5 minutes or less, using SQL or plain english. We’re building a modern, agentic-first data lakehouse platform ba
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Austin, US About the Role Cloudflare's People team supports 5,000+ employees globally. To scale, we are building an AI-driven operating layer on the Cloudflare Developer Platform to automate workflows, ensure data integrity, and streamline employee support. You will ship production systems for hiring, onboarding, and self-service, using AI to create leverage while designing rigorous guardrails for sensitive employee data. Lever
From $153K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As an Early Career Software Engineer at Roblox, your story begins with supportive mentorship and immediate, global impact. You’ll work alongside seasoned engineers and curious experts, designing, coding, and deploying real features that reach millions of people globally. By the end of your first year, you’ll have directly influenced the future of our platform and the experience of our global community with millions of daily active users. You Will: Join a community of curious, supportive engineers, actively engaging in architectural discussions and system design. Investigate and experiment with cutting-edge technologies, like machine learning frameworks and large language models (LLMs), to solve complex technical challenges and improve our engineering systems. Design, code, and test innovative features, navigating the full development lifecycle from initial design to production deployment. Partner closely with cross-functional teams, including Design, Product, Data, QA, and DevOps, to deliver cohesive products and features. Support the continuous evolution of our distributed systems, operating at our massive scale of 2 trillion analytics events a day. Engage in our team mat
From $200K/yr
We are looking for a talented engineer to lead evaluation of startup acquisition opportunities in the AI, cloud and security space. You will drive product evaluations, prepare and manage technical architecture discussions with target groups in Product and Engineering and provide roadmap suggestions for M&A and investments for Datadog. You will be a key partner to Datadog’s C-level leadership and highly visible at the most senior levels of Datadog. The role is reporting into the Senior Director of Product Strategy and falls within the Product organization. We are looking for an innovative and strategic thinker who is passionate about the latest tech being developed by startups in the cloud, AI and security space. The ideal candidate enjoys researching and evaluating new technologies, works effectively with cross-functional teams, and communicates opinions concisely to our leadership team. Broad understanding of relevant Cloud Technologies and deep understanding of the full coverage of Datadogs current offerings is necessary. The Corporate Development team is small and values authentic, strong-willed individuals who think creatively and proactively. This role leads technical due diligence from a product and architecture perspective across our acquisition pipeline. You'll scope and stand up proof-of-concept and sandbox environments to stress-test candidate products, then give an honest, unvarnished view of their quality and depth - the kind of assessment that holds up regardless of deal momentum. You'll assess technical architecture, flag the risks and open questions that matter most early, and turn that into a clear post-acquisition integration path. Working closely with engineering, you'll keep the evaluation focused on what's actually decision-relevant, then translate the findings into strategic recommendations for leadership and help carry the integration through by partnering with the right people on the other side. What You’l
From $244K/yr
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Opportunity: Datadog’s Staff Engineers are our technical leaders operating at the forefront of technology, building solutions that take us through at least our next five years of growth. They do this in three major ways: As individual contributors, they bring world class technical abilities to deliver industry leading systems in areas such as data visualization, virtual runtime profiling, and planet scale streaming. As technical leaders they bring experienced technical breadth and communication skills to tackling design and architectural problems spanning the organization, charting the right course, then leading delivery. In both roles they participate in the staff engineering community and help us learn from what the industry is doing and what we've built before, and so improve company wide standards around software and systems engineering. Some examples of projects a staff engineer may own include designing and building a new data storage engine handling hundreds of millions of records per second, being the lead engineer building a new product like synthetics or profiling, or rebuilding a critical service to handle the next two orders of magnitude of scale. What You'll Do: Be the technical owner of multiple pieces of critical architecture in your area of the business Own delivery of the systems you architect from beginning-to-end, doing what it takes to get things shipped and at full scale in production Dive deep into performance of systems; inventing new approaches that bring efficiency at scale Who You Are: You have a BS/MS/P
From $192K/yr
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. You will: Solve a scaling bottleneck in a critical service Deploy a new feature to production, progressively rolling it out with feature flags Investigate and fix a production issue from a service your team owns Design a way to scale up a service for more traffic With your team, plan the most important projects to work on next Who You Are: You have significant experience in one or more languages You value code simplicity and performance You can design architecture to solve problems at high scale You have a BS/MS/PhD in a scientific field or equivalent experience You want to work in a fast, high-growth startup environment that respects its engineers and customers You have demonstrated ability to use AI coding tools in day-to-day workflows and build, validate, and refine AI-generated output in products You can design AI Backend systems, with awareness of quality, cost, and latency tradeoffs 6+ years of experience Bonus points: You've worked at high scale with systems like Redis, Cassandra, Kafka You wrote your own data pipelines once or twice before You have a strong background in statistics You have significant experience with Go, C, or Python You’re excited about leveraging AI tools to enhance how you code, solve problems, and build – or eager to learn how You’re motivated to push the boundaries of how AI can improve software engineering best practices and contribute to building AI-enabled products Datadog values people from all walks of life. We understand not everyone will meet all the above qualificat
From $244K/yr
The Language Tools team enables ~1,500 Datadog developers to build, test, and package millions of lines of Go, Python, Java, Rust, and TypeScript in our backend monorepo. Our success is measured by their productivity and satisfaction. They use the tools that we develop and support several times a day, in both development and CI environments. We use the Bazel open source build system as a foundation. The team is growing rapidly, both with Datadog and as we absorb other repositories into the monorepo. As a senior software engineer on the team, you will own projects from start to finish, both greenfield and brownfield. You will gain first-hand understanding of what Datadog developers need, and inform our roadmap. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Invent build, test and packaging tools that are simpler and more reliable to use. Push performance and cost efficiency at scale, raising cache hit rates and cutting CI times and compute spend across millions of targets. Treat CI like SREs treat prod, making sure our pipelines are green and fast. Prepare, run, and finish complex migrations. Contribute back to the Bazel ecosystem, upstreaming fixes and shaping features we depend on. Who You Are: An expert in Bazel and/or one of the languages listed above. A well-rounded engineer. You must broadly understand the various types of software projects that are built, tested, and packaged with our tools. Both careful and fearless. The changes we make impact the velocity of hundreds of engineers. They are risky but necessary. User-focused. We help Datadog engineers to use the tools that we develop, and continuously improve their usability, so they don’t need our help the next time. Ideally, you have ex
From $234K/yr
The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in production, particularly those leveraging Large Language Models (LLMs) and generative AI. We provide robust, scalable observability for AI workloads, including drift detection and model evaluation, and behavior tracing, enabling customers to ship AI with confidence. As a Staff Engineer, you’ll lead the development of new features and foundational capabilities within Datadog’s LLM Observability product. You will shape product direction, drive experimentation, and apply your deep understanding of both AI systems and software engineering to solve open-ended problems in the fast-moving AI landscape. Your work will directly impact how our customers monitor, troubleshoot, and optimize LLM-based applications in production. Join us in building the foundational tools that make AI systems observable, understandable, and reliable in the real world. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Drive design and implementation of LLM observability features. Ideate, prototype, and scale new product features to provide insights and drive improvements for generative AI systems Work cross-functionally with other eng teams, product, UX, and applied science to iterate fast and find product-market fit Develop and extend tools for tracing, evaluating, and debugging LLMs Influence architecture decisions and mentor engineers to build resilient, high-performance systems Stay close to customer pain points and use those insights to guide product and engineering priorities Stay current with industry trends and advancements in machine learning and observability, driving innovation within the team Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or r
From $200K/yr
We are looking for a talented engineer to lead evaluation of startup acquisition opportunities in the AI, cloud and security space. You will drive product evaluations, prepare and manage technical architecture discussions with target groups in Product and Engineering and provide roadmap suggestions for M&A and investments for Datadog. You will be a key partner to Datadog’s C-level leadership and highly visible at the most senior levels of Datadog. The role is reporting into the Senior Director of Product Strategy and falls within the Product organization. We are looking for an innovative and strategic thinker who is passionate about the latest tech being developed by startups in the cloud, AI and security space. The ideal candidate enjoys researching and evaluating new technologies, works effectively with cross-functional teams, and communicates opinions concisely to our leadership team. Broad understanding of relevant Cloud Technologies and deep understanding of the full coverage of Datadogs current offerings is necessary. The Corporate Development team is small and values authentic, strong-willed individuals who think creatively and proactively. This role leads technical due diligence from a product and architecture perspective across our acquisition pipeline. You'll scope and stand up proof-of-concept and sandbox environments to stress-test candidate products, then give an honest, unvarnished view of their quality and depth - the kind of assessment that holds up regardless of deal momentum. You'll assess technical architecture, flag the risks and open questions that matter most early, and turn that into a clear post-acquisition integration path. Working closely with engineering, you'll keep the evaluation focused on what's actually decision-relevant, then translate the findings into strategic recommendations for leadership and help carry the integration through by partnering with the right people on the other side. What You’l
From $244K/yr
We're looking for a Staff Engineer to join the Logs organization at Datadog and help redefine how our customers ingest, query, and derive insights from logs data. In this role, you’ll work closely with Product Managers and customers to drive complex initiatives across ingestion pipelines, search infrastructure, and intelligent log management capabilities - all while pushing the boundaries of what’s possible with AI and distributed systems. You’ll have the opportunity to lead efforts that shape the future of log management. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Partner with Product Managers to define ambiguous product requirements and determine the most impactful solutions for customers Lead technical strategy and execution and design systems surrounding log query performance and ingestion at scale. Explore and prototype new capabilities and collaborate with peers on initiatives spanning AI-powered log management, security and business operations, advanced query capabilities, and external data sources query capabilities. Mentor engineers across levels and contribute to growing a high-performing, collaborative team culture Who You Are: You have deep experience architecting and scaling backend systems, with a strong focus on data-intensive or distributed infrastructure You excel in ambiguous environments, demonstrating a mix of drive, curiosity and pragmatic decision-making You’ve partnered effectively with Product Managers and customers to define product direction and ship impactful features You have expertise in debugging complex systems and optimizing performance across real-time data pipelines You have experience in using AI agents tools in your day-to-day engineering practices You lead by example and enjoy helping others grow through m
From $192K/yr
Distributed Systems engineers at Datadog design, implement and run in production the foundational platforms powering our applications. Your data pipelines will ingest, store, analyze and query in real-time billions of events per second from companies all over the globe. The platforms are optimized for durability, high availability, low latency, internet-scale footprint and operability. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Build fault-tolerant, horizontally scalable solutions running in multi-tenant environments Write in Go, Java Rust or C++, amongst other languages Use Kafka, Redis, Cassandra, Elasticsearch and other open-source components Own meaningful parts of our service, have an impact, grow with the company Who You Are: 6+ years of experience You have a BS/MS/PhD in a scientific field or equivalent experience You have significant backend programming experience in one or more languages (Go, Java, Rust, C++) You have been exposed to working on problems (high durability / low latency /…) You can get down to the low-level when needed You care about simple designs and performance You want to work in a fast, high-growth startup environment that respects its engineers and customers You have demonstrated ability to use AI coding tools in day-to-day workflows and validate, critique, and refine AI-generated output. Bonus: you’re motivated to push the boundaries of how AI can improve software engineering best practices and contribute to building AI-enabled products. This job is available in various departments within our company; to conform to US export control regulations, some of these roles may require candidates to be eligible for any required authorizations from the US government. Datadog values peo
From $151K/yr
Join the MongoDB Server Query Execution team, and help us build a world-class distributed open-source database. Our team plays a crucial role in the performance and efficiency of MongoDB's data processing. We are responsible for building and improving the core execution engine that powers all queries, taking a logical query plan produced by the optimizer and turning it into reality. This includes developing the physical operators for data retrieval and manipulation, improving the runtime for complex analytical and transactional workloads, and owning critical components such as our new execution engine. In addition to the core server, we support the query execution needs of other major products like Atlas Streams, Atlas Search and Vector Search, and mongosync, making our work vital to the entire MongoDB ecosystem. You will be joining a globally distributed team with a significant presence in both North America and Europe. While this role is based in the NAMER region, you will regularly collaborate closely with colleagues across different time zones. We support both office-based work in our North America hubs like New York, as well as remote work. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. To meet the ever-increasing data demands of modern applications, we are actively evolving our query system; this includes strategically re-architecting and improving key components of our query execution engine. We need your help to design and build the core of a distributed, flexible schema document database. This role can be based out of one of our North America offices, such as NYC or Palo Alto, or remotely across North America. Candidate Profile 10+ years of hands-on, professional experience in query engine development or database internals Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or
From $177.2K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business. We are seeking a Staff Data Engineer to lead the design, implementation, and evolution of our identity services and data governance platform. This role is critical to ensuring trusted, privacy-safe, and well-governed data across the organization. You will work at the intersection of da
From $177.2K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . We’re looking for a Staff Software Engineer to help build the next generation of Pinterest’s big data storage platform. You’ll work on some of the most exciting big data open source technologies — especially Apache Iceberg — at exabyte scale to power the data infrastructure that helps Pinners discover and do what they love. As a Staff Software Engineer, you’ll serve as a technical leader and hands-on contributor, designing and building highly scalable storage systems for Pinterest’s data lake. You’ll partner closely with teams across data, ML/AI, analytics, and infrastructure to evolve our storage and metadata management capabilities, enabling efficient, reliable, and governed access to data at massive scale. What you’ll do: Design, implement, and optimize Pinterest’s exabyte-scale data lake storage platform. Lead complex technical projects and
Other cities to consider
More places hiring for this role
Get new software reliability engineer jobs in United States by email
Daily job updates · Unsubscribe anytime