We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Mainframe Capacity & Performance Engineer to join our Enterprise Infrastructure organization. The Senior Mainframe Capacity & Performance Engineer will serve as a critical technical leader responsible for ensuring the performance, scalability, reliability, and efficiency of our enterprise mainframe environment supporting mission-critical healthcare, pharmacy, and retail applications. As a Senior Mainframe Capacity & Performance Engineer, you will play a key role in capacity planning, workload analysis, performance engineering, and infrastructure optimization across one of the nation's largest and most complex mainframe ecosystems. This position is responsible for proactively monitoring shared mainframe resources, evaluating system utilization trends, identifying performance risks, and providing actionable recommendations to improve overall system health and operational efficiency. The Senior Mainframe Capacity & Performance Engineer will partner closely with Application Development, Mainframe Systems Programming, Infrastructure Engineering, Architecture, Database Administration, Operations, and Business teams to analyze workload behavior, assess resource consumption, identify top consumers, and optimize application performance. This role requires deep expertise in z/OS performance analysis, capacity forecasting, workload managem
Jobiba hiring network
Senior Database Reliability Engineer Jobs
7,292 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current senior database reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Squarespace provides innovative solutions to empower our customers to focus on building their brand and growing their businesses on our platform. The Databases team manages all of the backend infrastructure that Squarespace runs on – MongoDB, CockroachDB, and Kafka clusters, to name a few examples. We are an accomplished, diverse group of people who develop the services that guarantee reliable and scalable infrastructure for both our cross-functional partners in product engineering, as well as our end users on the Squarespace platform. We believe that infrastructure excellence doesn't stop at just building for today; it needs to have a solid foundation of scalability, reliability, and a robust developer experience for the future. This is a hybrid role working from our Dublin office 3 days per week. You will report to the Databases Senior Engineering Manager. You’ll Get To… Nurture high-performing software engineers by guiding navigation when there is ambiguity. Distill the scope of the team and help hire a balanced group of engineers that will excel as a unit. Grow the career development of direct reports through regular 1:1s with direct, actionable feedback. Celebrate wins that motivate the team’s positive culture and robust dynamic. Evaluate consistently to improve team efficiency and effectiveness when required. Evolve a deep understanding of local systems to identify appropriate architectural decisions. Thread with Product, Design & Engineering to champion, define and execute an optimal roadmap. Bond across Engineering, Product, Design, Marketing, Data Science and Business Operations. Who We’re Looking For 3+ years of recent experience managing a Product Engineering team of four or more engineers. 7+ years of industry experience deploying apps across large codebases with many contributors. Ability to fluently translate, document and present technical concepts to non-technical stakeholders. Strong technical foundations to navigate the inherent tra
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Storage Platform team builds and operates the platform that powers database access across Robinhood. We own relational (Postgres/Aurora), key-value (DynamoDB), and caching systems, along with the SDKs, control plane automation, and data plane services that enable safe and reliable access at scale. Our mission is to standardize and strengthen how services connect to storage, improve reliability and performance, and reduce operational overhead through automation. We manage thousands of databases and hundreds of caching clusters supporting millions of users and critical brokerage workloads. Availability is our highest priority — our systems are designed to meet strict uptime targets, including no downtime during market hours. As a Senior Software Engineer , you will build and improve core infrastructure used by many engineering teams, with a focus on reliability, performance, and operational excellence. You’ll deliver key components of data plane and control plane systems (for example: connection pooling, query routing, automation workflows, and observability) and help evolve patterns for safe, consistent database access. You’ll work closely with peers to design pragmatic s
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Data Engineering team builds and maintains the foundational datasets that power decision-making across Robinhood. We design reliable, scalable data systems that support product analytics, growth strategy, financial reporting, experimentation, and machine learning. The team partners closely with Product, Engineering, Data Science, and Finance to ensure accurate, well-modeled data is available to teams across the company. Our work directly influences how Robinhood measures performance, improves customer experience, and scales its products. As a Senior Data Engineer, you will design, build, and evolve core datasets that track product performance and company-wide metrics. You will develop scalable data pipelines that ingest application events and database snapshots into our data lake, ensuring high data quality and reliability. You’ll collaborate with application engineers to improve data generation patterns and with analytics teams to design intuitive, well-documented data models. This is an opportunity to shape the technical foundation that supports data-informed decisions across the organization! This role is based in our Menlo Park, CA office, with in-person attend
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are You are an experienced Infrastructure Engineer Engineer who owns backend infrastructure end to end. You design multi-tenant, microservices-based systems that other engineering teams build on, and you make deliberate architectural tradeoffs around consistency, latency, scale, and cost. You are comfortable going deep — service mesh internals, database internals, distributed-systems failure modes — and equally comfortable defining the reliability and security contracts an enterprise AI platform depends on. Responsibilities Design, own, and evolve scalable microservices architectures on Kubernetes across GCP, Azure, and AWS, including multi-tenant isolation (namespaces, network policies, per-tenant resource quotas and RBAC). Build core platform and data-plane components in Golang and Python — data ingestion, knowledge-base indexing and vector/graph search, application connectivity, workflow automation, and ML operations — against explicit latency and throughput SLOs. Own service-to-service communication: gRPC/protobuf API contracts, service mesh (Istio/Linkerd), load balancing, retries, timeouts, and circuit breaking. Make and document architectural tradeoffs — partitioning
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is expanding the boundaries of the Data Cloud to support mission-critical transactional workloads. Our goal is to deliver OLTP capabilities with the performance, reliability, simplicity, and scale customers expect from Snowflake, while creating a seamless experience across transactional and analytical data. We are looking for a Senior Engineering Manager – OLTP to lead engineering teams and leaders building core transactional database technology and the cloud infrastructure required to operate it at scale. You will help define the architecture and roadmap, grow the organization, and drive technology from design through production. AS A SENIOR ENGINEERING MANAGER – OLTP AT SNOWFLAKE, YOU WILL: Set technical and execution strategy for key areas of Snowflake's OLTP platform, translating product goals into architecture, roadmaps, and team plans. Lead and grow multiple engineering teams, developing managers and senior technical leaders while fostering a culture of ownership, technical excellence, and execution. Drive adoption of AI and agentic development practices to improve engineering velocity, quality, and productivity across the software development lifecycle. Drive architecture and technical decisions in areas such as transactions, concurrency control, low-latenc
Position Overview We are looking for a Software Engineer II to build and deliver scalable software solutions across our products. You will work on modern web applications and cloud-based services using Node.js, React, TypeScript, AWS, PostgreSQL, MSSQL, and Docker, while contributing to AI-enabled features and integrations. You will collaborate closely with other engineers, product managers, and cross-functional teams to develop reliable, maintainable, and production-ready solutions. This role provides an opportunity to work with modern AI technologies including Python, AWS Bedrock, MCP, RAG, and agentic AI workflows while developing strong expertise in cloud-native software engineering. What You'll Do Develop and maintain scalable backend services and APIs using Node.js, TypeScript, and JavaScript. Build responsive and maintainable frontend applications using React. Design and implement integrations with AWS services and contribute to cloud-native application development. Develop and maintain applications using PostgreSQL and MSSQL, including writing efficient queries and working with database schemas. Build, test, and deploy applications using Docker and modern CI/CD practices. Contribute to AI-enabled product features using Python, AWS Bedrock, RAG, MCP, and AI integration patterns. Work with the team to integrate LLM capabilities, APIs, tools, and data sources into production applications. Write clean, maintainable, and well-tested code following established engineering practices. Participate in code reviews, technical discussions, debugging, and production issue resolution. Develop unit and integration tests and contribute to improving application quality and reliability. Monitor application performance and troubleshoot issues across development and production environments. Collaborate with senior engineers and architects to implement technical solutions aligned with product and engineering requirements. Stay current with emerging technologies, particularly in
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We’re hiring talented Software Engineers for the Snowflake Dynamic Tables team in Berlin, Germany. Join us to build the next generation data platform that enables customers to transform data with declarative SQL while maintaining control over cost, latency, and throughput. We are looking for strong engineers who are enthusiastic about building new cutting-edge technologies, who look forward to tackling complex database problems, and pick up and understand deep technical areas quickly. You will work alongside seasoned engineers and grow in your scope and influence. AS A SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Work with a talented and collaborative team of engineers and Product Managers in our globally distributed team to design and build Dynamic Tables capabilities Design, implement, support, and evolve new features and performance improvements Help shape technical and product direction with senior teammates Break ambiguous problems down, weigh the tradeoffs, and make technical and product decisions Analyze and solve performance, correctness, and fault-tolerance challenges at scale Dig into unfamiliar parts of a large system to root cause and solve problems Ensure operational readiness of what you build and help meet the commitments to our customers regarding reliability, a
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers and Architecture teams, your primary focus will be on ensuring production systems remain operational at all times, while continually setting and achieving long-term performance, reliability and scalability goals in a platform with an exponential growth plan for the coming years. With Okta’s increased dedication to ensuring customer availability expectations are exceeded in every way, you will play a key role as we evolve our system architecture to meet the demands of enormous growth and support the hundreds of millions of users who rely on us to provide uninterrupted access to business-critical enterprise and consumer applications. Skills Exceptional communication skills, including technical writing in the English language Systematic problem-solving approach, coupled with a strong sense of ownership and drive Understanding of microservices, cloud infrastructure (AWS, Azure), databases (SQL, No-SQL, Key/Value), containers (docker, kubernetes), web technologies (web sockets, http) and networking (SSL, routing, VPN) Live and breathe SLIs, SLOs, error budgets and SLAs Strong belief in automating everything and reducing toil for yourself and teammates Loves to work as a team, but is able to work effectively in a remote environment where tasks may be self-driven Knowledge
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge As a Senior Staff Software Engineer, you will serve as a technical leader for OneTrust’s AI Governance (AIG) platform, driving the design, scalability, and reliability of systems that enable enterprises to deploy and govern AI and LLM-powered applications responsibly. You will deeply understand how customers build, deploy, and operate AI systems, and translate those needs into secure, compliant, and observable platform capabilities. Your Mission Development Lead the design and development of Java/Python microservices and shared libraries integrating with AI platforms for OneTrust’s AI Governance product. Design, build, and test cloud-native applications deployed on Microsoft Azure using Core Java, REST, and the Spring ecosystem. Lead the architecture and development of reusable AIG reporting and dashboard capabilities that integrate governance data from SQL databases and analytical platforms with runtime observability signals. Design reusable semantic-layer and metric-abstraction capabilities, including dataset contracts, metric defini
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge As a Senior Staff Software Engineer, you will serve as a technical leader for OneTrust’s AI Governance (AIG) platform, driving the design, scalability, and reliability of systems that enable enterprises to deploy and govern AI and LLM-powered applications responsibly. You will deeply understand how customers build, deploy, and operate AI systems, and translate those needs into secure, compliant, and observable platform capabilities. Your Mission Development Lead the design and development of Java/Python microservices and shared libraries integrating with AI platforms for OneTrust’s AI Governance product. Design, build, and test cloud-native applications deployed on Microsoft Azure using Core Java, REST, and the Spring ecosystem. Lead the architecture and development of reusable AIG reporting and dashboard capabilities that integrate governance data from SQL databases and analytical platforms with runtime observability signals. Design reusable semantic-layer and metric-abstraction capabilities, including dataset contracts, metric defini
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: As a member of the Global Technical Operations (TechOps), you will be a part of a team that focuses on operational reliability within a cloud-based infrastructure. You have hands-on cloud experience in architecting, building, deploying, managing databases, compute instances, and storage buckets. You have a passion for providing solutions through automation. You know that success is through collaboration and communication. What you'll be doing: Work in cross-functional teams to develop solutions and identify opportunities to bring efficiency and effectiveness. Research, evaluate, and incorporate new technologies/concepts into existing frameworks. Proactively identify areas to improve efficiency and effectiveness, recommend and implement solutions towards them. Develop and innovate operational practices, procedures for workflows, and documentation. Implement and contribute to IT security best practices. Automate tasks to ensure consistency and speed of deployment. Identify, analyze, and troubleshoot issues and work towards resolution. Explain technical solutions to bo
Scale GP (Scale Generative AI Platform) is an enterprise-grade Generative AI platform providing APIs for knowledge retrieval, inference, evaluation, and more. We are seeking a strong Senior Full-Stack Engineer to help us build, scale, and refine our rapidly growing product. The ideal candidate is deeply grounded in software engineering best practices and experienced in developing and scaling modern web applications end-to-end. You will work across the stack—from React/TypeScript frontends to Python-based backends—while integrating with LLMs and machine learning systems. You will solve complex challenges in scalability, reliability, and product experience while owning significant product areas in a fast-paced environment. What You’ll Do Own major full-stack product areas , driving features from design through production deployment. Build modern frontend experiences using React and TypeScript, ensuring performance, usability, and responsiveness. Develop reliable backend services in Python, working with distributed systems, data pipelines, and ML/LLM components. Integrate with LLMs, vector databases, and AI infrastructure to power intelligent product experiences. Deliver experiments and new features quickly , maintaining high quality and tight feedback loops with customers. Collaborate across product, ML, and infrastructure teams to shape the direction of Scale GP. Adapt quickly —learning new technologies, frameworks, and tools as needed across the stack. Ideal Experience 5+ years of full-time engineering experience , post-graduation. Strong experience developing full-stack applications using React, TypeScript, and Python . Experience scaling or shipping products at high-growth startups . Familiarity with LLMs, vector databases, embeddings, or other modern AI tooling (tinkering or production experience welcome). Proficiency with SQL and modern API development. Experience with Kubernetes , containerization, and microservice architectures. Experience working with at leas
About the role: We are seeking a Senior Backend Engineer with deep backend engineering expertise and proficiency in one or more major programming languages (e.g., Python, Java, Go, Rust, or Kotlin), along with a strong understanding of AI models and agents. As a core member of our AI Engineering team, you will collaborate with data scientists, ML engineers, and product managers to build scalable, production-ready infrastructure and APIs that power intelligent systems. What you'll be doing: As a Senior Backend Engineer in the AI Engineering team, you will: Build and maintain reliable, scalable backend services to support AI agent execution and orchestration. Develop AI agent systems for complex operational workflows using LangChain, LangGraph, LiteLLM, and Langfuse. Orchestrate a hybrid model stack that includes OpenAI and Google Gemini alongside self-hosted and fine-tuned LLMs like Gemma and Llama. Build and maintain integrations with clinical systems (FHIR, EMR). Drive observability and reliability using OpenTelemetry, Datadog, and Langfuse. Design APIs (GraphQL, REST), background workers, and event-driven systems that interface with AI inference engines and agent runtimes. Collaborate with Data Science, ML, and engineering teams to deploy AI features and improve the performance, scalability, and reliability of backend systems. Participate in code reviews, knowledge sharing, and mentoring to elevate the team’s technical capabilities. What we're looking for: 6+ years of backend engineering experience, with strong proficiency in more than one major programming language (such as Python, Java, Go, Rust, or Kotlin). Solid understanding of AI systems architecture and experience working in environments involving AI agents, LLMs, or inference pipelines. Proven experience in building and scaling backend APIs, microservices, and background jobs. Strong experience with relational and NoSQL databases (e.g., PostgreSQL, MySQL, MongoDB, Redis), including schema des
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Help power the development of Replit Agent as an engineer in the Replit Cloud organization. The Replit Cloud team builds Replit’s first party cloud infrastructure so users can build, scale, and succeed entirely on Replit. They manage databases, application storage, app publishing and hosting, development/production environment splitting, custom domains, and more. By having a set of first party services that integrate seamlessly, you will power one of Replit’s key product differentiators. You will: Work closely with designers and product managers, to quickly iterate on Replit Cloud to continually grow and improve the product. Drive full-stack feature development from conception to deployment, taking ownership of key product initiatives. Contribute to architectural decisions that shape the future of our product. Ship product and build infrastructure as a true full stack builder using: TypeScript, React, CSS, Postgres, Go, and Terraform. Examples of what you could do: Leverage our unique cloud infrastructure to build differentiated full product experiences, helping non-technical or semi-technical users remove roadblocks to success. Leverage AI agents to proactively optimize or suggest app improvements on latency, reliability, SEO, and more. Be part of engineering leadership, steering teams towards the highest impact work and supporting initiatives across the company. Required skills and experience: Bachelor’s degree in Computer Science or related field, OR equivalent real-world experience in engineering roles. Comfortable building with our tech stack: TypeScript, React, Go Preferred Qualifications Experience building user facing platform as a service products. Experience with AI/agentic systems. Previous e
Get new senior database reliability engineer jobs by email
Daily job updates · Unsubscribe anytime