Jobiba hiring network

Distributed Systems Engineer Data Platform Delivery Database Retrieval Jobs

1,301 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current distributed systems engineer data platform delivery database retrieval jobs. Use filters to narrow by work mode, employment type, experience and date posted.

L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Our Infrastructure team is passionate about building software to solve problems at massive scale. We do this often, and when we believe our solution is worth sharing with the community, such as Envoy Proxy , we open source our ideas for the benefit of others. As an Observability team member, you are responsible for the operation and maintenance of our logging and metrics infrastructure. You ensure all teams at Lyft are aware of the operational health of their products by monitoring system availability and take a holistic view of our platform performance. You build software and platforms to automate infrastructure platform operations and management. By measuring and monitoring our operations you find opportunities to improve our systems in order to push our platform forward. You provide our partners with the support they need to help them build robust large scale distributed systems. We count on the reliability of our infrastructure to empower Lyft teams to provide our customers rich experiences that are highly available with rock solid performance to ensure our transportation platform continues to connect people and places. As we grow our team, we are seeking experienced Infrastructure Engineer to ensure that as our Infrastructure continues to scale, our platform continues to provide an essential and dependable service that transports millions of people every day. Specifically we are searching for someone who brings fresh perspectives, enjoys collaborating with cross-functional teams in order to continually improve our products and services for our customers. Responsibilities: Maintain, improve, and develop tooling and systems that enhance the reliability, scalability, and efficiency of our platform. Assist engineering teams in defining service-level objectives (SLOs) and provide the necessary toolin

pythonawskubernetes
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team API Multimodal builds the developer-facing products and infrastructure that bring OpenAI’s image, audio, and real-time model capabilities into the world. We are responsible for high-scale APIs for image generation, speech transcription, speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring frontier model capabilities to developers and use customer feedback to improve our models. About the Role As a software engineer on API Multimodal, you will build and operate the products and distributed systems behind OpenAI’s image, audio, and real-time APIs. You will work across model integration, API design, and production infrastructure to turn new research capabilities into reliable developer experiences. This hands-on role combines backend and systems depth with product judgment: you will own projects end to end, partner with Research, Inference, and Safety, and help make multimodal AI useful at scale. Model training experience is not required. In this role, you will: Design, build, and ship developer-facing APIs and backend services that serve frontier models. Architect low-latency streaming, request, session, and model integration systems that make complex multimodal interactions reliable and intuitive at scale. Work directly with Research to bring new model capabilities into production, shape the systems around them, and incorporate feedback from real-world developers and customers. Own the availability, latency, scalability, and cost efficiency of the services you build. Own projects from technical design and implementation through launch and ongoing iteration, while raising the team’s engineering standards. Your background might look something like: 7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles. A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed syste

typescriptpythonaws
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Core Services team is responsible for building and managing foundational services. It acts as the bridge between core infrastructure (e.g. compute, storage, networking) and product engineering teams, and enables product teams to move fast, build reliably, and scale efficiently. About the Role As a software engineer in the core services team, you will design and operate critical backend platforms such as caching systems, workflow orchestration, metadata stores, and file services. You’ll focus on building highly reliable, scalable, and performant systems that serve as the backbone of our products. We’re looking for people who are passionate about building infrastructure that empowers product teams, love working on distributed systems challenges, and enjoy creating well-designed APIs and abstractions that accelerate development. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and maintain shared infrastructure services such as caching layers, workflow orchestration (Temporal), metadata stores, and file storage services. Collaborate with product teams to provide scalable, reliable primitives that abstract the complexities of distributed systems. Improve performance, resilience, and scalability of core services that power customer-facing applications. You might thrive in this role if you: Have experience with distributed systems, caching infrastructure (e.g., Redis, Memcached), metadata storage (e.g., FoundationDB), or workflow orchestration (e.g., Temporal, Cadence). Have experience running containerized services in cloud environments and integrating them into automated build/test/release (CI/CD) workflows. Understand trade-offs in consistency models, replication strategies, and performance optimization in multi-region systems. Excel at communication and collaboration with cross-functional teams, and are obsesse

redisawsci/cd
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products—including next-generation ads experiences—that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring research breakthroughs into real-world systems at global scale. About the Role We’re looking for an experienced Software Engineer to help build the core infrastructure behind OpenAI’s monetization and ads systems. In this foundational role, you’ll architect and implement distributed systems that power OpenAI’s monetization stack—focusing on reliability, performance, privacy, and large-scale operation. You’ll work across backend, systems, and platform layers to define and implement 0→1 infrastructure, partnering closely with Product, Design, and Research to shape the future of monetized AI experiences. Your work will enable both internal and external teams to build on safe, scalable, and robust monetization primitives. This role is exclusively based across our San Francisco & Seattles sites. We offer relocation assistance to new employees. In this role, you will: Design and build the foundational backend and infrastructure powering OpenAI’s monetization and ads systems Architect large-scale distributed systems that

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for engineers to build the infrastructure that powers Codex agents in production. This role focuses on the systems that let models safely execute code, interact with tools, complete long-running tasks, and operate reliably and efficiently at scale. You’ll design and operate the infrastructure behind sandboxed execution, orchestration, stateful workflows, app-server and SDK boundaries, and model rollouts. You’ll work at the intersection of distributed systems, developer tooling, and AI, building primitives that make Codex faster, safer, more reliable, and easier for the rest of the organization to build on. What You’ll Do Design and build execution environments for AI agents, including sandboxing, isolation, and reproducibility. Develop systems for agent orchestration across multi-step, tool-using workflows. Build infrastructure for running, testing, and debugging code generated by models. Create state and memory systems that allow agents to persist context across long-running tasks. Optimize tokens, latency, reliability, and cost across Codex’s production fleet. Support model rollouts, capacity planning, and the core tradeoffs between quality, speed, and economics to manage a fleet of frontier agents at scale. Build shared platform capabilities that unblock product teams, partner teams, and open source Codex. Yo

awsci/cdrest
View job →

About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent, then turn that understanding into performance optimizations and models that project performance and capacity needs for future launches. About the Role In this role, you will model inference performance across application, model, and fleet layers with higher fidelity. You will build cost-to-serve estimates from microbenchmarks and create tools that help cross-functional teams reason about latency, capacity, utilization, and cost tradeoffs. In this role, you will Build and refine performance models that translate microbenchmark results into cost-to-serve estimates. Analyze inference workloads end to end across applications, models, and fleet infrastructure. Enhance tooling to identify bottlenecks across layers for latency and throughput. Partner with other teams to turn performance insights into concrete improvements and project how future changes affect inference. You might thrive in this role if you: Enjoy reasoning from first principles about distributed systems, model inference, and hardware efficiency. Are comfortable working across abstraction layers, from application behavior to kernels, accelerators, networking, and fleet scheduling. Have deep expertise with performance profiling, benchmarking, analysis, and optimization. Enjoy collaborating with engineering and research teams to improve real production systems. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve o

awsrestai
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake’s Release Engineering team builds and operates the systems that safely deliver infrastructure, platform, and product changes to production at global scale. We own the release platforms, rollout orchestration, and safety mechanisms that allow engineering teams across Snowflake to ship quickly while minimizing operational risk. Our mission is to make production deployments fast, safe, self-service, increasingly autonomous, and augmented by AI-driven intelligence and automation. This role sits at the intersection of developer productivity, distributed systems reliability, and large-scale multi-cloud infrastructure orchestration. At Snowflake, Release Engineering is a platform engineering function focused on building the systems, abstractions, and automation that make software delivery safe, scalable, and efficient across the company. In this role, you will Design and build continuous deployment and rollout infrastructure that safely ships changes across Snowflake’s large-scale, multi-cloud production environment. Build and evolve platform capabilities for progressive delivery, including staged rollouts, canarying, automated health checks, rollback controls, and guardrails that reduce blast radius during production change events. Improve engineering velocity by removi

pythonjavakubernetes
View job →
O
13 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Device Identity and Access Organization The Device Identity and Access organization is at the forefront of Okta’s Zero Trust vision. As a foundational pillar within Okta Research and Development (ORD), our mission is to transform the device itself into a secure, trusted, and effortless identity factor. We are the teams responsible for ensuring users can seamlessly interact with their work from any endpoint, anywhere in the world. Our organization is composed of engineers who thrive at the intersection of deep client-side platform engineering and massive-scale distributed systems. The work we do secures millions of enterprise endpoints globally, prevents modern identity attacks, and fundamentally changes how people work by making world-class security completely invisible to the end user. Opportunity We seek a highly impactful and influential Software Engineer to join our Device Authenticators engineering team. The ideal candidate will leverage their deep expertise in Apple (macOS and iOS) client software development to architect, build, and scale the critical software and services at the heart of our security and identity platform. This is a high-visibility, hands-on opportunity to define the architectural vision, pioneer new capabilities, and drive the technical strategy to solve complex, company-wide challenges and shape the future of Okta's device identity ecosystem. You will act as a key technical leader—a player-coach and force multiplier—by sett

machine learningartificial intelligenceai
View job →
W-
Wolt - English
📍 Helsinki• Full-time
19 days ago

About Wolt At Wolt, we create technology that brings joy, simplicity and earnings to the neighborhoods of the world. In 2014 we started with delivery of restaurant food. Now we’re building the delivery of (almost) everything and you’ll find us in over 500 cities in 30 countries around the world. In 2022 we joined forces with DoorDash and together we keep on dreaming big and expanding across the globe. Working at Wolt isn’t always easy, but it’s definitely exciting. Here you’ll learn more, build more, and ship more than in most other companies. You’ll be challenged a lot, but also have a lot of fun on the way. So, if you’re a self-starter with drive and entrepreneurial spirit, this could be the ride of your life. What you’ll do: Build and maintain high-throughput backend services using Go . Collaborate with product managers, designers, and frontend developers to ship features that support internal support agents across the globe. Design systems that are scalable , resilient , and easy to maintain. Lead and contribute to architectural discussions and technical decision-making. Write well-tested code and help the team maintain high code quality standards. Our humble expectations: 7+ years of professional software engineering experience, with a proven track record of building and scaling complex systems. 2+ years of production experience in Golang , with the ability to mentor others and drive best practices across the team. Strong hands-on experience with both SQL and NoSQL databases — especially Cassandra. Solid understanding of designing and operating low-latency, high-throughput distributed systems . Nice to have Background in Node.js or other backend languages. Familiarity with cloud infrastructure (AWS, GCP) and event-driven architectures. Previous on-call experience , with a pragmatic approach to reliability and incident management. What we value A product-oriented mindset — you think beyond the ticket, understand the “why” behind the work, and aim to create real

node.jssqlaws
View job →
OS
OfficeSpace Software
📍 Costa Rica• Full-time
19 days ago

About OfficeSpace: OfficeSpace Software provides the leading AI operating system for the built world, that helps teams plan, connect, and perform in the workplace. As a performance-based, PE-backed company, we hire based on merit and a willingness to do what it takes to succeed long-term. You’re a great fit for the role if you’re entrepreneurial, passionate, motivated by building at light speed, and an Agentic AI early adopter. Our world-class teams operate in the US, Canada, and Costa Rica in a culture of trust, respect, growth, and impact. About the role As a Lead Software Engineer at OfficeSpace, you'll shape the future of our workplace platform by building scalable products, leading technical execution, and raising the engineering bar across the team. You'll combine deep technical expertise with strong engineering leadership. From architecting distributed systems to mentoring teammates and designing AI-powered engineering workflows, you'll help us deliver enterprise software that is reliable, secure, and built to scale. This is a hands-on leadership role. We provide the platform. You drive the impact. What you'll do - Lead the architecture, design, and delivery of high-performance applications using Ruby on Rails, React, and modern cloud technologies. - Build scalable, maintainable software that supports enterprise customers across a rapidly growing platform. - Design AI-assisted engineering tools, workflows, and automation that improve developer productivity, code quality, and customer outcomes. - Own projects end-to-end—from technical discovery through production deployment and long-term ownership. - Shift quality left by embedding automated testing, code quality practices, and continuous validation early in the development lifecycle. - Drive predictable delivery by managing scope, balancing technical debt, protecting sprint commitments, and reducing reactive work. - Establish performance benchmarks and continuously optimize application spe

reactredisrest
View job →
P
19 days ago

A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Build software used to support global trading across time zones • Work closely with business users to establish and refine requirements • Participate in identifying new technologies to continuously improve software systems • Implement DevOps practices within the team (GitHub/GitLab, Jenkins) • Provide production support to diagnose and resolve elevated application software production incidents and implement proactive remediation measures WHAT’S REQUIRED • Bachelor’s degree in computer science or another technical/scientific field • Minimum 8 years object-oriented programming experience with C#/.NET • Significant experience working with / understanding databases - primarily MS SQL Server • Must have experience in writing automated tests, unit tests, Test-Driven Development • Knowledge of design/architecture patterns, distributed systems, microservices, observability and monitoring, and containerization (Docker) • Willingness to work as part of a distributed Dev Team - 3 time zones (USA, Poland, India) • Strong problem solving and analytical skills • Exceptional verbal and written communication skills • Commitment to the highest ethical standards WE TAKE CARE OF OUR PEOPLE We invest in our people, their careers, their health, and their well-being. When you work here, we provide: • Health care benefits • Maternity, Adoption & related leave policies • Generous paternity and family care leave policies • Employee Assistance Prog

sqldockergit
View job →

SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime. With the ability to build, scale and manage security across the cloud, hybrid and traditional environments in real-time, SonicWall provides relentless security against the most evasive cyberattacks across endless exposure points for increasingly remote, mobile and cloud-enabled users. With its own threat research center, SonicWall can quickly and economically provide purpose-built security solutions to enable any organization—enterprise, government agencies and SMBs—around the world. For more information, visit www.sonicwall.com or follow us on Twitter , LinkedIn , Facebook and Instagram . About the Role SonicWall is seeking a Software Dev Senior Engineer (Dataplane) with deep expertise in High Availability, Fast-Path Packet Processing, and Distributed Systems to architect and optimize the dataplane HA subsystem across SonicWall NGFW hardware and virtual firewalls — driving sub-millisecond session state synchronization, zero-packet-loss failover, and line-rate forwarding for IPsec, SSL-VPN, and TCP/UDP flows at 10G–100G+. Responsibilities Dataplane HA Architecture — Own fast-path HA engine design covering Active/Standby mirroring, Active/Active flow redistribution, and zero-drop VMAC/VIP failover. Stateful Session Mirroring — Build low-overhead dataplane sync for TCP/UDP 5-tuple tables, IPsec SA fast-path, SSL-VPN contexts, and L7 states across HA peers at line rate. Lock-Free Session Tables — Design NUMA-aware, lock-free (RCU/CAS/seqlock) session tables for multi-core pipelines sustaining tens of millions of concurrent sessions at sub-microsecond latency. High-Performance Sync Fabric — Engineer zero-copy ring buffer / DPDK mempool-based sync with prioritized queuing for IPsec/SSL-VPN state

redisawsazure
View job →
GW
Get Well Network
📍 Bengaluru• Full-time
20 days ago

Staff Software Engineer Bengaluru, Karnataka, India Opportunity Get Well is seeking a visionary and technically adept Staff Software Engineer to architect, design, develop, and optimize our cloud-native healthcare platform while driving the adoption of AI-First and AI-Augmented Software Engineering practices. This role is pivotal in shaping the future of software development at Get Well by combining deep technical expertise with modern AI-assisted engineering workflows. As we evolve toward an AI-First engineering organization, this leader will champion the use of Generative AI, AI development assistants, and Agentic AI to improve developer productivity, software quality, and engineering velocity. The ideal candidate brings deep expertise in software architecture, cloud-native application development, AI-enabled engineering, and distributed systems. This role provides technical leadership across multiple engineering teams, ensuring high standards for architecture, code quality, reliability, security, and AI adoption. This is a hands-on leadership role where strategic thinking meets deep engineering execution within a complex healthcare environment. This position reports to the Director, Product Development and requires close collaboration with software engineers, AI engineers, product managers, DevOps, QA, and compliance specialists. Key Responsibilities Technical Leadership Define and drive the architecture of scalable, distributed healthcare platforms. Champion AI-First Software Engineering practices across the development lifecycle. Lead the adoption of AI-Augmented Development , including spec-driven development, AI-assisted coding, code reviews, testing, and documentation. Establish engineering standards, Cursor/AI coding guidelines, reusable patterns, and governance for responsible AI usage. Provide hands-on leadership in architecture, coding, design reviews, debugging, and performance optimization. Mentor enginee

pythonjavareact
View job →
I
Instawork
📍 Bengaluru• Full-time
20 days ago

Instawork is on a mission to create meaningful economic opportunities for skilled hourly professionals in communities around the globe. Our AI-powered labor marketplace helps local businesses scale, and enables global technology companies to push the frontiers of robotics and AI. Backed by world-class investors like Benchmark, Spark Capital, Craft Ventures, Greylock, Y Combinator, and others, we’re looking for exceptional talent to reimagine the way the world works. Check out our Engineering blog here Who You Are: 5–8 years building and shipping high-traffic web or mobile apps in production. Strong coding and problem-solving skills. While experience with our stack (Python/Django, Celery, MySQL, Redis, ElasticSearch, React Native) is preferred, it's not mandatory. Proficiency in frontend technologies like React, Redux, and Responsive Design Systems. Experience in system design, distributed systems, and both relational and No-SQL databases. What You'll Do: End-to-end design, development, and deployment of high-scale web and mobile features. Prioritizing the highest ROI work—move quickly on the projects that matter most. Tackling large, complex software projects and delivering them on aggressive timelines. Maintaining peak team productivity: juggle multiple priorities in a fast-paced environment. Partnering with PMs and designers to choose the simplest, most maintainable solution. Raising the bar through thoughtful code reviews and clear technical feedback. Perks: Free snacks Health Insurance Personal Insurance Flexible Hours Maternity/Paternity Leave Broadband Reimbursement #LI-SS3 Our Values Empathy, Trust & Candor We put ourselves in the shoes of our colleagues and customers and don’t shy away from uncomfortable conversations, instead building trust through honest and direct feedback. Bias for Action We practice high-velocity decision-making, clear-eyed that we often operate with incomplete information. Growing quickly means it’s OK to be wrong, so l

pythonreactsql
View job →
O
Okta
📍 Toronto• Full-time• From C$136K/yr
20 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Datastores Engineer, Platform Infrastructure The Auth0 platform secures more than 100 million logins each day for customers all around the world - and we're growing fast! The Platform Infrastructure team enables Auth0 engineers to move faster by giving them tools to easily deploy and manage their services on AWS and Azure. This is a role with a huge impact. You will get to work with engineers throughout the organization and what you build will be a foundational piece of the infrastructure that allows Auth0 to scale for years to come. We are looking for Engineer who are passionate about distributed systems, availability, and delivering customer value to join our Platform Infrastructure Datastores team. Because we build and support the overall Auth0 platform, the ideal candidate is someone who is passionate about infrastructure, operations, databases and not intimidated by cross-organization coordination and collaboration. You will: Develop our large, distributed and highly-available infrastructure. Implement platform tools that allow feature teams to deploy and manage the datastores for their services. Research new technologies to accelerate new environment creation. Carry cross team initiatives from end to end: code reviews, design reviews, operational robustness, security hygiene, etc. Participate in the team's on-call rotation. You might be a good fit if you: Have 5-8 years of software development experience. Are proficient in or have a desire

sqlpostgresqlmongodb
View job →
🔔

Get new distributed systems engineer data platform delivery database retrieval jobs by email

Daily job updates · Unsubscribe anytime