Jobs in United States

Software Reliability Engineer in United States

2,007 active opportunities · Updated October 2026

Explore current software reliability engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

This Engineering Manager will lead the Data Visualizations Explorations team within the Graphing organization, setting product direction, and coaching and developing team members. They will staff and drive projects that build end to end experiences for Datadog’s core users: observability engineers. This includes extending the capabilities of core widgets like Hostmap and Geomap Visualizations and finding innovative ways to leverage existing Datadog data sources. This role also involves close partnerships with other product teams to deeply understand customer needs and deliver compelling data experiences. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Work with Product Management and Design to plan and staff projects for the Data Visualization Explorations team Coach and develop engineers at various levels Ensure strong cross-team communication and design best practices around project management, code review, architecture patterns, and more. Proactively anticipate cross-team dependencies and blockers to goals. Identify opportunities to appropriately reuse or customize features across dashboards, notebooks, and product pages. Deeply understand the needs of our customers and other Datadog products we work with. Participate in customer conversations, review product briefs, read feature requests, and coach team members to adopt these practices as well. Define and maintain high standards for operations practices, including bug triage and remediation, incident response, and gathering and analyzing performance telemetry for our widgets. Who You Are: At least 2 years of people management experience in a software engineering or similar setting Strong TypeScript/JavaScript skills, including familiarity with front

JavaScriptTypeScriptJavaReact
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $99K/yr

Quick readStrong listing-quality and freshness signals

Datadog's Finance team collaborates with teams across the organization, providing commercial, operational and analytical support to ensure that Datadog's business continues to scale rapidly and efficiently. The Financial Planning & Analysis (FP&A) team analyzes company financial data (revenue, customers, headcount, expenses, etc.) in order to support the business’ growth and success. As an analyst supporting the team, you will play a key role in delivering insights through the management of essential data infrastructure, including our financial planning tool, Pigment. Your role will be highly cross-functional, leveraging systems and data to unlock analytical capabilities for both FP&A and business leaders. Your role is critical in synthesizing information from across the organization to foster operational alignment and support informed strategic decisions. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the team’s forecasting and reporting software, Pigment, supporting data-driven insights through the development of dashboards and KPIs, both for standard FP&A reports and ad hoc projects Work cross-functionally with FP&A leaders to improve existing datasets and models Ensure data and system best practices in processes across the organization, including during planning and reporting cycles Represent FP&A in the data & analytics community, collaborating with analytics partners across the organization to democratize data and share insights Work on strategic projects and initiatives for senior management, assessing various business opportunities and proposing solutions Support Datadog’s data-based decision making and continued efficient growth Who You Are: 2+ years of professional experience in FP&A, Data Analy

PythonSQLRestAI
D
📍 Massachusetts, California, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $185K/yr

Quick readStrong listing-quality and freshness signals

The Product Solutions Architecture (PSA) team acts as a technical multiplier across Datadog. PSAs are domain experts who partner with Field teams on complex customer use cases across pre- and post-sales engagements and scale their impact by producing reusable collateral, including reference architectures, technical guides, and enablement assets. By feeding real-world customer insights back to Datadog Product teams, PSAs help influence product roadmaps while accelerating adoption, usage, and long-term customer success. Datadog’s LLM Observability product enables organizations to monitor, troubleshoot, and optimize large-scale LLM-powered applications with confidence, while meeting requirements around data privacy, compliance, and cost management. As a Product Solutions Architect, you will partner closely with Datadog customers and the LLM Observability product team to design architectures, implement best practices, and drive adoption of LLM observability across customer environments. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Serve as the in-house subject matter expert for Datadog’s LLM Observability product Partner with Field teams to provide hands-on technical and architectural guidance to enterprise customers adopting LLM Observability Create high-impact technical collateral, including reference architectures, technical guides, cookbooks, and documentation to enable Field teams and the broader customer community Build proofs of concept and small-scale deployments to validate solutions and reproduce real-world customer environments Act as a trusted advisor to Product Management by delivering actionable feedback informed by real-world field experience Who You Are: You bring a strong software engineering foundation, with hands-on experience bu

JavaScriptTypeScriptPythonJava
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $195K/yr

Quick readStrong listing-quality and freshness signals

Here at Datadog, we think about offensive security a little bit differently. We embrace automation and AI to run adversary simulations continuously across a massive cloud-native environment, and we expect our offensive engineers to build the tooling that makes that possible. We're looking for a Senior Security Engineer who can execute sophisticated red team operations, write the code that scales them, and take an AI-first approach to offensive security engineering. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Plan and execute red team engagements end-to-end, simulating real-world threat actors across cloud infrastructure (AWS, GCP), Kubernetes, CI/CD pipelines, and corporate environments Build and maintain custom offensive tooling, automation frameworks, and engagement infrastructure, treating offensive operations as a software engineering problem Develop custom payloads and evasion capabilities tailored to Datadog's environment and modern defensive controls (EDR, SIEM, network monitoring) Improve the efficiency of offensive operations through thoughtful use of automation and AI, accelerating reconnaissance, vulnerability analysis, and reporting workflows Partner with the Detection & Response team on purple team exercises to validate detection logic, improve alert fidelity, and influence threat models Translate offensive findings into concrete improvements by working directly with defensive security and engineering teams to close gaps Who You Are: You have 5+ years of hands-on experience in offensive security (red teaming, penetration testing, or adversary simulation) with a track record of operating against mature, well-defended environments You write production-quality code (Python, Go, or similar), can build your own tools, and automate your w

PythonAWSAzureGCP
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $156K/yr

Quick readStrong listing-quality and freshness signals

The Team We are Datadog’s in-house product experts. The Technical Solutions team enables Datadog’s worldwide growth by educating potential partners and ensuring that our integration ecosystem is high-performing, secure, and valuable. Partner Technology Solutions Engineers (TSEs) are the technical bridge between Datadog and our third-party developer community. We act as consultants, helping partners build world-class monitoring solutions on the Integration Developer Platform (IDP) . The Opportunity Datadog is looking for a Partner Technology Solutions Engineer to join our fast-paced team. You will be the primary technical contact for our partners, guiding them through the entire integration lifecycle—from initial architectural design to final publication on the Datadog Marketplace. This is a unique role that combines deep technical troubleshooting with high-level consulting and platform advocacy. You will work directly with external developers and see your contributions immediately reflected in the Datadog ecosystem. You Will Act as the technical lead for partners, advising on OAuth flows, log pipelines, OpenTelemetry, and agent-based vs. API-based configurations Perform architectural assessments and deep-dive code reviews for partner integrations in the integrations-extras and marketplace repositories, ensuring they meet our Quality Rubric Solve complex technical challenges for partners via Zendesk, Slack, and dedicated technical consultations Identify friction points in our Integration Developer Platform (IDP) and partner with our internal Product and Engineering teams to build a better developer experience Maintain public-facing developer documentation and internal tracking systems ( JIRA ) to ensure transparency and scale You Are A technical expert with 3+ years of experience in a technical role (Support Engineering, Solutions Architecture, or Software Development) Proficient in at least one language (Python or Go preferred) An observability enthusiast who unders

PythonLinuxAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $89K/yr

Quick readStrong listing-quality and freshness signals

We’re looking for someone to join the Datadog Procurement team and help expand the Strategic Sourcing group. Make an impact by being a trusted analyst in several buying categories and continue to prove the value that Strategic Sourcing brings to the organization. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Leverage AI tools (e.g., Claude, ChatGPT) to enhance market research, accelerate data analysis, and improve efficiency in sourcing workflows Identify opportunities to incorporate automation and AI into procurement processes to drive scalability and smarter decision-making Support and execute sourcing strategies directed by sourcing managers with a specific focus in G&A categories (People Team, Finance, Legal, Real Estate) Develop strong relationships with stakeholders and assist teams to independently evaluate their best practices and adopt a value-driven mindset when purchasing on behalf of the organization Lead and execute Sourcing events (RFx) Negotiate pricing and other business terms with vendors on net-new purchases and renewals Manage the renewals list for designated categories and ensure accuracy of information and proactive communication Analyze spend data and identify potential cost savings opportunities Create pricing models based on vendor proposals to quantify various buying scenarios Build license forecast models and utilization analyses for enterprise software renewals to help uncover usage and savings opportunities Collaboratively align with adjacent functions like FP&A on utilization and forecast models Operate effectively in ambiguous or evolving problem spaces, helping bring structure, clarity, and forward momentum to loosely defined sourcing initiatives Monitor market trends that impact spend categories and

M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $151K/yr

Quick readStrong listing-quality and freshness signals

Join the MongoDB Server Query Optimization team, and help us build a world-class distributed open-source query optimizer. Our team plays a crucial role in the experience and performance of data processing. We are responsible for the MongoDB Query Language and the lifecycle of each query, through parsing, optimization and plan selection. We have a presence across the US and Europe including New York, Dublin, Seattle, Palo Alto, and Chicago. We support office-based and remote work and align projects with convenient work hours for each time zone. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. The team is endeavoring to systematically rewrite every major component of our optimization and execution systems. We need your help to design and build the heart of a distributed, flexible schema, document database. ​​This role can be based out of our US offices or remotely in the North America region. Candidate Profile 10+ years of experience in data management systems, distributed systems, or large-scale backend engineering Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or similar field, or equivalent practical experience, with strong competencies in data structures, algorithms, and software design/architecture Experience with large code bases written in C++ or another systems programming language. You'll need to trace down defects, estimate work complexity, and design evolution and integration strategies as we rewrite different components of the system A strong foundation in core database internals is essential. While direct experience in query optimization is a massive bonus, it is not a prerequisite. We are also excited to meet candidates with strong backgrounds in compilers, language transpilers, or distributed storage systems Position Expectations Innovate in the area of flexible schema d

MongoDBAWSAzureRest
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

As Engineering Manager for Threat Detection, you will lead a high-performing team that powers Datadog's detection program. Threat Detection is the organization responsible for keeping Datadog ahead of an evolving threat environment: closing coverage gaps faster, raising the bar on signal quality, and shipping detections that hold up under the scale and complexity of cloud-native infrastructure. Your team will combine direct detection expertise, platform engineering, and applied AI to ship detections at a pace and scale traditional rule-writing alone cannot match. Examples of what your team will work on include detection-authoring agents, the detection platform that powers every rule in production, coverage analysis, alert triage and response automation, and the evaluation infrastructure that holds these systems to a high bar of fidelity. Detection authorship is a shared responsibility across the organization, and your team will contribute both by building the systems that scale our authoring capacity and by writing detections directly when their domain expertise is the right tool. You will partner closely with our Security Incident & Response Team (SIRT), Cyber Threat Intelligence (CTI), AI Engineering teams, and Datadog's broader Security organization. This is a high-impact leadership role: you will grow a team of security and software engineers responsible for building and executing our detection and AI strategy. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog Security's shift to AI-accelerated detection and response. Drive development of high-fidelity detections as a shared responsibility across the organization, ensuring your team's systems and direct contributions raise the bar on coverage and

PythonCI/CDRestAI
P
📍 US; Remote, United States· Full-time· Remote
✓ High-confidence listingCompany trend -85.6%

From $189.3K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Conversion Visibility Modeling team enables a performant ads marketplace and helps prove value to advertisers by connecting Pinterest onsite activity with conversions that happen offsite (both digital and physical) in a privacy-preserving way. As a Machine Learning Engineering Manager on this team, you will lead a hybrid team of ML engineers and backend software engineers to build end-to-end identity and conversion visibility solutions across modeling, serving, and data infrastructure, so advertisers retain accurate, privacy-aware performance visibility as signals fragment and degrade. You will set the technical direction for high-impact ML systems that feed ranking, bidding, measurement, and reporting across Pinterest’s ads stack. What you’ll do: Attract, hire, develop, and lead a hybrid team of ML engineers and backend softwar

AWSGitRestMachine Learning
A
📍 Illinois, United States· Full-time
✓ High-confidence listingCompany trend -98.8%

From $111K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Our Supply Team is growing, and we want you to be part of it! As a Market Manager, you will play a crucial role in the retention, growth, and optimization of host entrepreneurs and rising hosts in your assigned territory. You will focus on building strong relationships with these hosts, helping them optimize their listings, improve pricing strategies, and enhance their overall success on Airbnb. The Difference You Will Make: As a Market Manager, you will be responsible for the growth of a book of business in your assigned territories by managing professional host accounts and supply acquisition of high quality inventory. You will build Airbnb’s strong market presence in your assigned region in collaborating with other teams as the local in-market expert. As someone who knows the Airbnb mission and values inside out, you will drive all phases of supply acquisition and market success from acquisition strategy development to strategic partner relationship management. You will help develop and iterate scalable, localized supply strategies in your territory aiming to secure our long-term success. A Typical Day: Build and manage partner relationships within assigned territory Directly manage accounts to meet and exceed quarterly/annual sales goals Analyze data and utilize data-driven recommendations to identify and action on strategic opportunities in your region to drive increase in sales Prospect and onboard new, high quality supply in your assigned geography Maintain a baseline understanding of the technical integration of various software partners so that you c

S
📍 United States· Full-time
✓ Quality checkedCompany trend -87.9%

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe Capital provides access to fast, flexible financing to small-and-medium businesses on Stripe to accelerate their growth, and we lent over $1B in 2024. Businesses use the funds for marketing, team growth, geographic expansion, working capital, new equipment purchases, and much more. Machine learning is core to Stripe Capital’s business—we use information about businesses from their activity within and outside of Stripe and our models to automatically underwrite uniquely tailored financing offers to their needs, which banks are often unable to do. We are doing so through models with an established performance history, data infrastructure that is Stripe scale, and a strong feedback loop that includes explainability, anomaly detection and a risk portfolio management layer. We're an end-to-end team going from ideas to models to shipping in production. What you’ll do As a machine learning engineer for Stripe Capital, you'll be responsible for designing, building, training, evaluating, deploying, and owning ML models in production with the goals of providing financing opportunities to as many users as possible while satisfying financial performance goals. You'll work closely with software engineers, data scientists, product managers, and risk managers to operate Stripe’s ML powered systems, features, and products. You'll also contribute to and influence ML architecture at Stripe and be a part of a larger ML community. Responsibilities Design

Machine LearningAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team We are a small and fast-moving partnerships team that shapes and executes OpenAI’s most important collaborations. Your mission is to build and grow partnerships across the cybersecurity ecosystem. Reporting to the Cybersecurity Partnerships Lead, you will own a portfolio of cybersecurity technology partners and help them validate, launch, and scale solutions powered by OpenAI models and platforms. About the Role You are an experienced partnerships operator who combines business development, partner management, technical curiosity, and strong execution. You can manage external relationships while driving detailed cross-functional work across product, engineering, technical success, sales, marketing, legal, security, and operations. You move with urgency, follow through consistently, and are comfortable managing a portfolio in a fast-changing market. In this role, you will: Source, close and manage a portfolio of cybersecurity technology partners. Identify high-value use cases across security operations, identity, cloud security, application security, threat intelligence, governance, risk, and compliance. Develop partner plans covering integration, launch, enablement, co-marketing, co-sell, and growth. Support partnership structuring and coordinate product, technical, commercial, legal, and security workstreams. Help partners move from concept and technical validation to production launch and scaled customer adoption. Run regular partner reviews, track commitments, resolve blockers, and identify expansion opportunities. Coordinate closely with product, engineering, technical success, sales, marketing, legal, security, and operations. Measure partner pipeline, launches, model adoption, consumption, customer outcomes, and partner health. You might thrive in this role if you have: 8+ years of experience in partnerships, business development, alliances, partner success, or ecosystem roles. Experience in cybersecurity, enterprise software, cloud platforms, o

AWSRestAIGo
O
📍 United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role Trusted Computing and Cryptography is a core security team at OpenAI focused on deploying high-performance cryptography at scale, secure key management, and trusted hardware enclaves—from boot measurements to GPU confidential computation. As a Hardware Platform Security Architect, you’ll own hardware platform security at OpenAI. In this role, you will: Co-Architect Secure Silicon: Collaborate with cross-functional silicon teams (Silicon Design, DV, FW) and silicon partners (silicon test facilities, foundries) to develop secure silicon that meets the end-to-end system requirements. Co-Architect Secure Hardware: Collaborate with hardware vendors and cross-functional teams (kernel, compiler, infra) to design secure hardware that meets performance and security needs. Co-Architect Secure Systems: Architect and deploy systems using TPM2, Secure Boot, Nitro Enclaves, Intel SGX, AMD-SEV, and other secure hardware technologies. Drive Innovation: Engage with internal and external partners to align hardware innovations with OpenAI’s trusted computing and cryptographic requirements. You might thrive in this role if you have: 10+ years of industry experience in hardware security or hardware–software co-design. Proven expertise in deploying secure hardware systems at scale and integrating secure hardware primitives. Strong coding skills in Rust and/or C/C++, with proficiency in Python. Proven ability to collaborate across teams, architect solutions,

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Governance, Risk, and Compliance team helps ensure security and privacy are grounded in how our products and systems actually operate. Assurance Operations partners with Security, Engineering, Infrastructure, Product, Privacy, and Legal to make controls provable, risk decisions explicit, and audit readiness a result of well-designed systems. About the Role We are hiring a technical, product-minded GRC builder who can own consequential audits while improving the control and evidence systems behind them. You will build a reusable common control framework, use Codex to automate assurance work, validate changing system scope, and turn repeated audit friction into measurable improvements. We are looking for someone who questions inherited assumptions, solves novel problems creatively, works closely with engineers, and makes the next audit easier by improving the underlying system. You’ll be responsible for: Lead external, internal, customer, and certification audit work from scoping through evidence review, fieldwork, remediation, and closeout. Build a common control framework linking risk, control intent, implementation, owner, system, environment, evidence, and applicable frameworks. Validate actual scope and ownership instead of assuming last year's controls, product boundaries, or evidence remain accurate. Use Codex to build and test evidence checks, control mappings, request triage, owner workflows, monitoring, and remediation reporting. Partner with engineers on cloud architecture, identity, logging, data flows, software changes, vulnerabilities, and control effectiveness. Design maintainable, permission-aware tools that preserve source provenance, human review, and evidence integrity. Reduce repeated requests and operational burden for control owners through measurable workflow improvements. Define roadmaps, decision rights, milestones, success metrics, and clear cross-functional escalations. We’re looking for someone with: Direct ownership

SQLAWSRestAI
O
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About OpenAI OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. We build models and products that help people learn, create, and solve problems—and we work to do so safely and responsibly. About the Team OpenAI’s products are talked about by people, not just press and pundits. More than 900 million people use our tools each week, learning from one other and passing along what works. Their stories shape our reputation and encourage others to try our products. Our community team builds direct relationships & channels with people who use our tech then amplifies their stories and use cases so peers can learn from them and channels their insights to our product and research teams. About the Role We are seeking an exceptional Technical Community Program Manager to help manage and scale a high-signal community of advanced ChatGPT Pro users. This is a hands-on, technically fluent role at the intersection of community-building, product enablement, user research, product education, and editorial storytelling. You do not need to be a full-time software engineer, but you should be comfortable understanding technical workflows, asking sharp technical questions, using the latest AI tools, and helping advanced users explain what they are building. You will report to the Head of Pro Subscriber Community and will be based in New York City. In this role, you will: Run and grow a high-signal community Help manage and deepen engagement with the ChatGPT Pro cohort of advanced users across disciplines Design and execute high-touch community programming, including product demos, office hours, show-and-tell sessions, peer-learning formats, and in-person events Build repeatable systems for onboarding, engagement, retention, and member communications Build deep relationships with exceptional users Identify, recruit, and onboard new individuals doing high-impact work with ChatGPT and Codex Conduct in-depth interviews and maintain ongoing r

AWSRestAIGo
🔔

Get new software reliability engineer jobs in United States by email

Daily job updates · Unsubscribe anytime