We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are a global team of innovators shaping the future of observability. Our intelligent platform gives customers real-time insight into complex systems so they can innovate faster and operate reliably in an AI-first world. If you’re excited by high-throughput distributed systems and want to contribute to one of the largest and fastest-growing observability platforms, we’d love to hear from you. Join a backend engineering team focused on building and operating JVM-based services that ingest, process, and serve massive volumes of telemetry data. You’ll work on high-scale, low-latency systems that power mission-critical observability features used by engineers worldwide. What you'll do Design, build, and operate JVM-based microservices (primarily Java and Kotlin) with a focus on performance, scalability, and reliability. Own services end-to-end: architecture, implementation, deployment, monitoring, on-call participation, and continuous improvement. Apply strong concurrency and performance practices: asynchronous programming, backpressure, efficient I/O, memory management, and GC tuning. Build and evolve event-driven systems; work with Kafka for streaming, partitioning, consumer groups, and schema evolution.Instrument services for deep observability (metrics, logs, traces), define SLIs/SLOs, and use e Experience with Kafka or similar streaming technologies (topic/partition strategy, consumer lag, idempotency, schema compatibility) strongly preferred. Proficiency w
Jobiba hiring network
Distributed Systems Engineer Data Platform Delivery Database Retrieval Jobs
1,301 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current distributed systems engineer data platform delivery database retrieval jobs. Use filters to narrow by work mode, employment type, experience and date posted.
MongoDB’s mission is to empower innovators to create, transform, and disrupt industries by unleashing the power of software and data. We enable organizations of all sizes to easily build, scale, and run modern applications by helping them modernize legacy workloads, embrace innovation, and unleash AI. Our industry-leading developer data platform, MongoDB Atlas, is the only globally distributed, multi-cloud database and is available in more than 115 regions across AWS, Google Cloud, and Microsoft Azure. Atlas allows customers to build and run applications anywhere—on premises, or across cloud providers. With offices worldwide and over 175,000 new developers signing up to use MongoDB every month, it’s no wonder that leading organizations, like Samsung and Toyota, trust MongoDB to build next-generation, AI-powered applications. MongoDB is seeking a Software Engineer 3 to join the Atlas Clusters Organization. The organization is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Clusters Security team creates a first-in-class cloud database security experience for our wide range of sophisticated customers. Our team develops and maintains systems for cluster networking, data encryption, database authentication, and more–enabling countless mission critical applications across the world. We are looking to speak to candidates who are based in New York for our hybrid working model. What you’ll do Build and design new features for MongoDB Atlas Contribute to and lead complex technical projects Work closely with product and design teams, considering the user’s perspective while building technical solutions
About the Team: Compute Infrastructure builds the platform that turns enormous amounts of compute into a reliable engine for frontier AI. We design, provision, schedule, operate, and optimize the systems that connect accelerators, CPUs, networks, storage, data centers, orchestration software, agent infrastructure, developer tools, and observability into one coherent experience for researchers and product teams. Our work spans the entire stack: capacity planning and cluster lifecycle, bare-metal automation, distributed systems, Kubernetes and scheduling, deep system optimization, high-performance networking, storage, fleet health, reliability, workload profiling, benchmarking, and the developer experience that lets teams use enormous compute systems with confidence. At this scale, small improvements to communication, scheduling, hardware efficiency, or debugging workflows can compound into meaningful research velocity. We are hiring across Compute Infrastructure rather than for a single narrow team, and we use this opening to match strong engineers to the problems where they can have the most leverage. About the Role We are looking for engineers who want to build the compute platform behind OpenAI's research and products. You may not be the strongest in low-level systems, high-performance computing, distributed infrastructure, reliability, CaaS, agent infrastructure, developer platforms, tooling, or the user experience around infrastructure. What matters is that you can reason carefully about complex systems, write durable software, and raise the quality and velocity of the people around you. Depending on your background and interests, you might work close to hardware, close to users, on CaaS and agent infrastructure, or on the control planes and data planes in between. You could help bring new supercomputing capacity online, optimize training workloads from profiler traces and benchmarks, improve NCCL and collective communication behavior, reason about GPUs, NICs, t
Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. We're looking for an engineer to help build the future of distributed Postgres. You'll work on Multigres, our open-source distributed database system that brings horizontal scaling to Postgres. In this role, you'll architect and implement critical distributed database infrastructure including sharding, consensus protocols, and materialization systems. You'll collaborate closely with our Postgres, networking, and infrastructure teams to push the boundaries of what's possible with distributed databases. What You’ll Be Responsible for: Design and implement query routing logic for sharded databases Build consensus and replication systems to support distribute durability Develop materialization pipelines for migrations and change data capture Contribute to connection pooling infrastructure and intelligent workload isolation Collaborate with the open-source community on Multigres development You Might Be a Good Fit If You have: This role requires deep technical expertise in distributed databases and systems. For detailed qualifications, see our contributor qualifications document . Key areas of expertise: Database sharding, relational algebra, and Postgres internals Consensus protocols (Raft, Paxos, FlexPaxos) and distributed transactions Stream processing, materialization, and change data capture Building robust, performant distributed systems with strong observability Low-latency infrastructure and network protocol optimization. What We Offer Fully Remote We hire globally. We believe you can do your best work from anywhere. There are no Supabase offices, but we provide a WeWork membership or co-working allowance you can use anywhere in the world. ESOP Every team member receives ESOP (equity ownership) in the comp
The Application Modernization Platform (AMP) team is tackling one of the industry's most critical challenges: leveraging Generative AI to transform rigid, legacy applications into modern, microservices-based architectures powered by MongoDB. We are building a comprehensive, SaaS-like platform, encompassing both the "brain" (multi-agent reasoning and orchestration) and the "hands" (the deployment platform and modernization toolset). This solution requires a robust platform foundation and infrastructure designed for a "build once, run anywhere" model, ensuring seamless operation regardless of a client's security or network constraints. A key challenge is balancing the need to tune our tools for each customer's unique tech stack and restrictive environments with making them easily extensible and scalable for common application modernization challenges. We seek an engineering leader for this high-visibility initiative. This role requires defining the high-level strategy and technical direction across all AMP engineering pillars, leading the execution of solving uniquely complex application modernization puzzles, and delivering an enterprise-grade product. The leader will minimize deployment friction, meet customer compliance requirements, and help shape the future of how global enterprises leverage GenAI. The ideal candidate is a hands-on technical leader who excels at leveraging GenAI capabilities, architecting complex distributed systems, and designing the orchestration agents necessary to reliably and fluidly run the entire software development lifecycle. This role will be based in North America's West Coast (PST), and offers a hybrid working model. The ideal candidate for this role will have 10+ years of software development and operations experience, with a focus on building platforms and distributable software infrastructure Deep experience in building data warehouses and core components for data processing systems Have experience in using GenAI in building comple
The Application Modernization Platform (AMP) team is tackling one of the industry's most critical challenges: leveraging Generative AI to transform rigid, legacy applications into modern, microservices-based architectures powered by MongoDB. We are building a comprehensive, SaaS-like platform, encompassing both the "brain" (multi-agent reasoning and orchestration) and the "hands" (the deployment platform and modernization toolset). This solution requires a robust platform foundation and infrastructure designed for a "build once, run anywhere" model, ensuring seamless operation regardless of a client's security or network constraints. A key challenge is balancing the need to tune our tools for each customer's unique tech stack and restrictive environments with making them easily extensible and scalable for common application modernization challenges. We seek an engineering leader for this high-visibility initiative. This role requires defining the high-level strategy and technical direction across all AMP engineering pillars, leading the execution of solving uniquely complex application modernization puzzles, and delivering an enterprise-grade product. The leader will minimize deployment friction, meet customer compliance requirements, and help shape the future of how global enterprises leverage GenAI. The ideal candidate is a hands-on technical leader who excels at leveraging GenAI capabilities, architecting complex distributed systems, and designing the orchestration agents necessary to reliably and fluidly run the entire software development lifecycle. This role will be based in North America's West Coast (PST), and offers a hybrid working model. The ideal candidate for this role will have 10+ years of software development and operations experience, with a focus on building platforms and distributable software infrastructure Deep experience in building data warehouses and core components for data processing systems Have experience in using GenAI in building comple
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. You Will: Data Architecture and Design: Designing and overseeing the architecture of scalable and reliable data platforms, including data pipelines, storage solutions, and processing systems Data Modelling and Management:Developing and implementing data models, ensuring data quality, and establishing data governance policies Data Pipeline Development: Building and optimising data pipelines for ingesting, processing, and transforming large datasets from various sources Performance Optimisation: Identifying and resolving performance bottlenecks in data pipelines and systems, ensuring efficient data retrieval and processing Technology Evaluation and Innovation: Staying abreast of emerging data technologies and exploring opportunities for innovation to improve the organisation’s data infrastructure Troubleshooting and Problem Solving: Diagnosing and resolving complex data-related issues, ensuring the stability and reliability of the data platform Data Security and Compliance: Implementing data security measures, ensuring compliance with data governance policies, and protecting sensitive data Perform other duties as assigned You Have: Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field. 10+ years of experience in data engineering or a similar role. Enterprise SaaS software solutions with high availability and scalability Solution handling large scale structured and unstructured data from varied data sources Experience in building and maintaining data platform systems such as distributed compute,
About Anyscale At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role The Customer Engineer will play a crucial role in the customers’ post-sale journey - helping them to onboard, adopt and grow on Anyscale, troubleshooting and resolving open customer tickets and driving consumption. Anyscale is an ever evolving platform and hence will require close co-ordination with our engineering teams to debug complex issues. This is an exciting role for those who are technically curious and passionate about ML/AI, LLM, vLLM and the role of AI in next generation applications. It’s an opportunity to make a significant impact in a collaborative, fast-paced environment while building a new segment in this space. In this role, you’ll be able to Resolve customer issues and help in their successful adoption of Anyscale platform Be a technical advisor, and internal champion for our key customers Own customer issues end-to-end, from troubleshooting, triaging, escalations and eventual resolution Participate in our follow-the-sun customer support model to ensure continuity in resolving high priority tickets Keep track of open customer bugs and feature requests to influence prioritization and provide timely customer updates upon resolution Contribute towards improvement of internal tools an
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world's largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Revenue and Financial Automation Sub-org within Revenue and Financial Automation: Billing. The Revenue and Financial Automation team at Stripe builds software tools that accelerate the economic and technological growth of global businesses by helping them operationalize their commercial relationships with customers. Our offerings include a billing platform, SaaS analytics, data services, and finance automation products that our customers creatively combine to support various revenue models. Team Matching: exact team matching for one of the subteams will begin during final stages. Please note we may also consider you for different orgs based on your experience, location, etc. More information on our team matching process can be found here. What you'll do We're looking for engineers who want to build the distributed systems, APIs, and backend services that power Stripe's revenue and billing platform. You'll focus on API design and performance, service reliability, and distributed systems challenges, while collaborating across the stack to ship complete solutions for millions of businesses. Responsibilities Scope, architect, and lead technical projects to build and scale distributed backend systems and APIs Design, build, and maintain reliable, high-performance APIs and backend services Own service reliability, including setting performance targets and driving improvements across the stack Debug production issues across distributed services
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Summary: As a Principal Software Engineer at Appian, you will be the primary technical strategist, responsible for shaping the architectural foundation of our platform. Your role involves anticipating future challenges and implementing innovative solutions today. You will drive complex cross-functional initiatives, ensuring Appian remains a leader in the low-code and automation industry. Responsibilities: Define the long-term architectural vision and governance for the Appian platform. Identify systemic technical risks and lead task forces to resolve architectural bottlenecks. Develop internal tools and SDKs to simplify infrastructure and enhance developer productivity. Conduct deep-dive troubleshooting for complex production issues. Promote AI-native engineering practices and integrate AI features into the platform. Lead the development of core platform libraries and high-risk prototypes. Participate in the Architectural Guild and review high-impact design documents. Mentor lead and senior engineers, and represent Appian in the tech community. Ensure operational resilience with self-healing and highly available systems. Required Qualifications: Bachelor’s or Master’s degree in Computer Science, Information Technology, or related field. 15+ years of software engineering experience, with significant experience in architecting large-scale distributed systems. Strong understanding of data structures, algorithms, and design patterns. Proven transformational leadership in technological migrations or strategies. Expertise in Java and Cloud-Native ecosyst
We're looking for a Senior Engineer with a strong background in computer science fundamentals, systems design, experience in the Java ecosystem, streaming systems, and data-intensive applications to join our engineering team. In this role, you will be instrumental in designing, building, and optimizing the underlying data structures, algorithms, and database interactions that power our generative AI platform, code generation and migration tools. This involves crafting sophisticated orchestration layers, robust integration points, and high-performance data systems that seamlessly connect and leverage advanced AI capabilities for code generation and building a sophisticated data migration suite using a modern technology stack, which includes Java, Spring Boot, Kafka, Debezium, and React.You will work on critical components that ensure the scalability, efficiency, and reliability of our services, collaborating closely with AI researchers, product management and other engineers to design and implement cutting-edge products that solve complex customer challenges.. We are looking to speak to candidates who are based in Sydney for our hybrid working model. The ideal candidate for this role will have 6+ years of engineering experience in backend systems, distributed systems, or core platform development. Proficiency in one or several of Java, Rust, C/C++, and/or Python, with a strong understanding of systems-level programming, memory management, and performance tuning. Extensive experience with streaming data platforms such as Apache Kafka and Change Data Capture (CDC) tools like Debezium Extensive experience with relational data modeling and hands-on experience with at least one SQL database (Postgres, MySQL, etc) Exposure to client-side technologies such as JavaScript and React is a plus Good understanding of algorithms, data structures and their time and space complexity Curiosity, a positive attitude, and a drive to continue learning Excellent verbal and wri
At Bolna, we’re building tools that change how businesses leverage voice AI. We’re looking for a Software Engineer to build reliable, scalable systems that power millions of production conversations across languages, industries, and telephony environments. This is a high-impact, high-ownership role where you’ll work on core platform problems across distributed systems, real-time communication, developer infrastructure, and customer-facing products. Our team includes IIT alumni with experience at Bain, Atlassian, Uber, Zomato, and LinkedIn, and is backed by leading investors. Responsibilities Build systems that operate at scale: Design and build backend services that support high-volume, real-time voice AI conversations with strong reliability, performance, and fault tolerance. Own features end to end: Take problems from product requirements and technical design through implementation, testing, deployment, monitoring, and iteration. Improve platform reliability: Build systems that are observable, resilient, and easy to debug. Identify bottlenecks, reduce failure rates, and improve system availability. Work on real-time infrastructure: Solve problems across telephony, streaming audio, webhooks, queues, scheduling, concurrency, and low-latency communication. Build for developers and customers: Improve APIs, SDKs, integrations, dashboards, and internal tools that make the Bolna platform easier to use and operate. Raise the engineering bar: Contribute to technical design reviews, code quality, testing standards, documentation, incident response, and engineering best practices. Required Skills Strong engineering fundamentals: Solid understanding of data structures, algorithms, databases, networking, operating systems, and distributed systems. Backend development experience: 2+ years of experience building and operating production backend systems using Python, Go, Java, Node.js, or a similar language. Production ownership: Experience shipping software to production and own
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are We are seeking an experienced DevOps Engineer to join our growing team and play a pivotal role in designing and building our platform and infrastructure as we continue to scale our product and user base. As a part of our team, you will be working in a dynamic, fast-paced environment to ensure the reliability, scalability, and performance of our systems, while focusing on service architecture and deployment, query optimization, distributed systems, data and machine learning infrastructure, and security and authentication. Most importantly, you are excited to be part of a mission-oriented, fast-paced, high-growth startup that can create a lasting impact. You will: Partner with product teams to architect, design, and build the foundational infrastructure for our products. Design, develop, and deploy highly available and scalable Multi-tenant SaaS solutions on any one of the public cloud networks like AWS, Azure and GCP. Leverage technologies such as Kubernetes, Helm, Terraform, and Istio to achieve infrastructure resilience. Drive the automation of infrastructure tasks, from provisioning to configuration management and deployment, utilizing tools like Terraform, Ansible, a
About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission is to engineer resilience from the ground up, enabling our product teams to innovate rapidly while ensuring our users have a stellar experience. We own the availability, latency, performance, and capacity of our platform, and we achieve our goals through a culture of data-driven decision-making, blameless learning, and relentless automation. As a Senior Site Reliability Engineer, you are a hands-on engineer who blends deep software development expertise with a passion for operational excellence. You will be responsible for designing, building, and running the resilient, scalable, and increasingly self-healing systems that power our products. You will apply sound engineering principles to solve our most complex reliability challenges, with a mandate to automate everything, eliminate toil, and write robust, maintainable code. You will be a force multiplier, mentoring other engineers and elevating the site reliability bar for the entire organization. This is a hybrid role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: System Architecture & Design: Design, build, and maintain scalable, highly available, and fault-tolerant distributed systems. Partner with development teams as a reliability consultant, reviewing designs and influencing architectural decisions to ensure new services are built with reliability, observability, and performance as core principles, not afterthoughts. Automation & Software Development: Write robust, performant, and maintainable code to automate operational tasks, and CI/CD pipelines. Build the internal tools, libraries, and frameworks that enable engineering teams to self-service their
We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. This is an opportunity to join a team that is responsible for all Observability systems that support metrics, metric visualization, logs, traces, and alerts for MongoDB. We are looking for engineers with high standards, and experience in setting direction and technical leadership for large engineering teams in designing and operating complex distributed systems, with strict SLO on security, durability, availability and performance. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Splunk, Flink, WarpStream/Kafka, Java, Golang Fluentbit. In addition to owning critical components of our observability infrastructure, as a Staff engineer on the team, you’ll also work closely with other SWE, Product and SRE teams to promote and implement best practices in instrumenting and monitoring their services. This is a highly collaborative role, and you will get to own some of the most relied upon internal infrastructure at Mongo. Our team champions a strong culture of inclusivity, diversity, and collaboration. If you want to be a deeply technical leader on a collaborative team that applies low-level systems expertise to build the foundational infrastructure of a popular database, join us! Let’s build a faster, more reliable, and exceptionally observable database system together. W
Get new distributed systems engineer data platform delivery database retrieval jobs by email
Daily job updates · Unsubscribe anytime