Jobiba hiring network

Software Engineer Distributed Systems Salary India Jobs

6,428 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current software engineer distributed systems salary india jobs. Use filters to narrow by work mode, employment type, experience and date posted.

R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's Cache team is building a next-generation caching solution designed to deliver sub-millisecond average latency, horizontal scalability, and high efficiency—all at a drastically lower cost. Our ultimate vision is to shape a caching infrastructure capable of supporting 1 billion Daily Active Users while reducing costs by 90%. We are turning hours of onboarding and capacity expansion into seconds, freeing service owners entirely from managing cluster lifecycles. As a Senior Engineer on the Cache team (part of the Infra Storage org), you will innovate and operate large-scale, in-house distributed systems to solve Roblox's ever-growing caching challenges. You will report directly to the Engineering Manager for the Cache team. (Check out our recent engineering blog post here to learn more about the team's latest work!) You will: Lead the architectural transition to a next-generation, multitenant caching service built on ValKey, ensuring strict data, resource, and failure isolation for all tenants. Drive systemic optimizations to mitigate head-of-line blocking, manage hot keys, and maximize CPU and memory utilization across physical machine clusters. Design and build robust frameworks to a

redisawskubernetes
View job →
M
Mongodb
📍 New York City• Full-time• From $106K/yr
1mo ago

MongoDB’s Replication Team builds the infrastructure that enables high availability, fault tolerance, automatic failover, and tunable consistency. As an engineer on the team, you will design and implement distributed-systems features that protect data and keep applications available under demanding operating conditions. You will work primarily in C++ on core database code, partner with engineers across MongoDB, and help shape features that are central to major MongoDB releases. This is an opportunity to apply distributed-systems fundamentals to a widely used database while solving challenging problems in correctness, performance, and operability. We're looking to speak with candidates based in New York City for our in-office working model. What you will do Design and implement replication features based on the Raft consensus protocol Improve failover behavior, availability, correctness, and performance across the replication system Write production-quality C++ and the unit, integration, and system tests needed to demonstrate correctness Use JavaScript and Python where appropriate to extend test coverage and validate end-to-end behavior Diagnose test failures, investigate bugs, and drive issues through root-cause analysis and resolution Measure the performance impact of code changes and prevent or resolve regressions Collaborate with partner engineering teams and stakeholders on large, cross-functional initiatives Investigate distributed-systems issues raised by customers and Technical Support, communicate findings clearly, and help deliver durable fixes Participate in code reviews, design reviews, and technical discussions that improve the quality of the team’s work Interview candidates and mentor junior engineers and interns What you bring Required At least five years of experience programming, debugging, and performance-tuning distributed or highly concurrent software systems Strong systems fundamentals, including multithreaded programming, concurrency, debugging,

javascriptpythonjava
View job →

About the Team ChatGPT relies on a large and growing GPU fleet to serve inference workloads reliably and efficiently. Our team builds the software, tooling, and operational systems that help manage this fleet at scale. We work across production engineering, distributed systems, capacity management, and operational automation to improve reliability, reduce manual work, and make better use of available compute. About the Role We are looking for a software engineer with experience building or operating large-scale production systems. You will develop the systems that help manage the GPU fleet powering ChatGPT, including tooling for fleet health, capacity planning, operational automation, and incident response. You will work closely with infrastructure, research, and product engineering teams to improve reliability, developer productivity, and compute utilization. This role is a good fit for engineers who enjoy solving complex operational problems and building software that makes production infrastructure easier to run at scale. In This Role, You Will Build software and internal tools to manage large-scale GPU infrastructure supporting ChatGPT inference. Develop systems for capacity planning, fleet health monitoring, and resource utilization. Automate operational workflows, including incident detection, diagnosis, and response. Identify and address bottlenecks affecting fleet reliability, scalability, and performance. Partner with infrastructure, research, and product engineering teams to improve the compute platform. You Might Thrive in This Role If You Have experience operating large-scale production infrastructure, GPU clusters, or other compute-intensive distributed systems. Have a background in production engineering, site reliability engineering, infrastructure engineering, or platform engineering. Have built software that automates operational workflows and reduces manual work. Have worked with distributed infrastructure, cluster orchestration, or large-scale int

pythonawsrest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role On the Accelerators team, you will help OpenAI evaluate and bring up new compute platforms that can support large-scale AI training and inference. Your work will range from prototyping system software on new accelerators to enabling performance optimizations across our AI workloads. You’ll work across the stack, collaborating with both hardware and software aspects - working on kernels, sharding strategies, scaling across distributed systems, and performance modeling. You'll help adapt OpenAI's software stack to non-traditional hardware and drive efficiency improvements in core AI workloads. This is not a compiler-focused role, rather bridging ML algorithms with system performance - especially at scale. In this role, you will: Prototype and enable OpenAI's AI software stack on new, exploratory accelerator platforms. Optimize large-scale model performance (LLMs, recommender systems, distributed AI workloads) for diverse hardware environments. Develop kernels, sharding mechanisms, and system scaling strategies tailored to emerging accelerators. Collaborate on optimizations at the model code level (e.g. PyTorch) and below to enhance performance on non-traditional hardware. Perform system-level performance modeling, debug bottlenecks, and drive end-to-end optimization. Work with hardware teams and vendors to evaluate alternatives to existing platforms and adapt the software stack to their architectures. Contribute to runtime improvements, compute/communication over

awsrestai
View job →
O
1mo ago

About the Team: Compute Infrastructure builds the platform that turns enormous amounts of compute into a reliable engine for frontier AI. We design, provision, schedule, operate, and optimize the systems that connect accelerators, CPUs, networks, storage, data centers, orchestration software, agent infrastructure, developer tools, and observability into one coherent experience for researchers and product teams. Our work spans the entire stack: capacity planning and cluster lifecycle, bare-metal automation, distributed systems, Kubernetes and scheduling, deep system optimization, high-performance networking, storage, fleet health, reliability, workload profiling, benchmarking, and the developer experience that lets teams use enormous compute systems with confidence. At this scale, small improvements to communication, scheduling, hardware efficiency, or debugging workflows can compound into meaningful research velocity. We are hiring across Compute Infrastructure rather than for a single narrow team, and we use this opening to match strong engineers to the problems where they can have the most leverage. About the Role We are looking for engineers who want to build the compute platform behind OpenAI's research and products. You may not be the strongest in low-level systems, high-performance computing, distributed infrastructure, reliability, CaaS, agent infrastructure, developer platforms, tooling, or the user experience around infrastructure. What matters is that you can reason carefully about complex systems, write durable software, and raise the quality and velocity of the people around you. Depending on your background and interests, you might work close to hardware, close to users, on CaaS and agent infrastructure, or on the control planes and data planes in between. You could help bring new supercomputing capacity online, optimize training workloads from profiler traces and benchmarks, improve NCCL and collective communication behavior, reason about GPUs, NICs, t

awskubernetesrest
View job →

Role Description As a Software Engineer on the Metadata team, you’ll build and operate the large-scale distributed databases that every Dropbox service depends on. Metadata systems are mission-critical, in the live path for all user operations and must meet stringent requirements for latency, durability, and transactional consistency. You’ll design and evolve the core infrastructure that manages Dropbox’s databases at scale, enabling fast, reliable access to data for millions of users and hundreds of internal services. This work spans distributed systems, replication, caching, and transactional database systems. You’ll collaborate closely with engineers across Infrastructure and Product teams to ensure the metadata layer meets business needs and continues to scale with Dropbox’s growth. This is an opportunity to leverage your expertise in distributed systems and grow into broader technical leadership. Our Engineering Career Framework is viewable by anyone outside the company and describes what’s expected for our engineers at each of our career levels. Check out our blog post on this topic and more here . Responsibilities Design and maintain distributed database systems providing low-latency, strongly consistent data access Implement and optimize replication, consensus, and caching mechanisms to meet availability and performance goals Operate production systems, including participating in the on-call rotation, ensuring high availability and data durability Collaborate with infrastructure and product teams to assess current and future use cases and requirements, supporting the development of a mid- to long-term roadmap that reflects these needs Contribute to system design reviews, postmortems, and reliability improvements Write high-quality, efficient code in Go and Rust for performance-critical systems On-call work may be necessary occasionally to help address bugs, outages, or other operational issues, with the goal of maintaining a stable

REMOTEpythonjavaai
View job →
T
15 days ago

Join Truecaller – The place where innovation meets impact! Truecaller's mission is to build trust in communication by making it safer, smarter, and more efficient. Born in Sweden, trusted by the world, and here’s why we stand out: We are trusted by over 500 million active users every month across 190+ countries We identify over 15 billion calls daily, helping users avoid spam and scams We are powered by a team of 400+ employees from 45+ nationalities We always look for people who take initiative, own their work, and keep raising the bar. An entrepreneurial mindset matters here, especially when it turns bold ideas into real actions. We stay collaborative and focused, always searching for smarter paths forward. If you want to make an impact and grow with a team that inspires millions, you’ll fit right in. The role: As a Senior Software Engineer, Backend, you will leverage your deep technical expertise to ensure alignment, quality, and long-term maintainability across the organization. You will be building strong relationships with cross-functional stakeholders, you will drive project delivery, and ensure our backend infrastructure remains scalable and robust enough to support millions of users worldwide. What you will do Design and develop high-volume, low-latency services and cope with the challenge of working in a distributed environment. Operate mission-critical services at high availability. Collaborate across business units and drive solutions. Explore new solutions and technologies. Build a scalable and reliable system. What you bring in: 5+ years of experience with Scala, Java, or Go Experience of working with scalable, highly available, real-time distributed systems Experience of working with non-relational databases Good understanding of data structures and algorithms Mentoring/ Leadership skills Excellent communication skills in English. Ability to thrive in a dynamic, fast-paced environment. It would be great if you also have

javadockerkubernetes
View job →
AC
15 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Summary: As a Principal Software Engineer at Appian, you will be the primary technical strategist, responsible for shaping the architectural foundation of our platform. Your role involves anticipating future challenges and implementing innovative solutions today. You will drive complex cross-functional initiatives, ensuring Appian remains a leader in the low-code and automation industry. Responsibilities: Define the long-term architectural vision and governance for the Appian platform. Identify systemic technical risks and lead task forces to resolve architectural bottlenecks. Develop internal tools and SDKs to simplify infrastructure and enhance developer productivity. Conduct deep-dive troubleshooting for complex production issues. Promote AI-native engineering practices and integrate AI features into the platform. Lead the development of core platform libraries and high-risk prototypes. Participate in the Architectural Guild and review high-impact design documents. Mentor lead and senior engineers, and represent Appian in the tech community. Ensure operational resilience with self-healing and highly available systems. Required Qualifications: Bachelor’s or Master’s degree in Computer Science, Information Technology, or related field. 15+ years of software engineering experience, with significant experience in architecting large-scale distributed systems. Strong understanding of data structures, algorithms, and design patterns. Proven transformational leadership in technological migrations or strategies. Expertise in Java and Cloud-Native ecosyst

javasqlaws
View job →

Software Engineer Build technology where every nanosecond matters. At Graviton, software isn't just a tool that supports trading. It is the infrastructure behind every research breakthrough, every trading decision and every competitive advantage. As a Software Engineer , you'll work on systems where performance, reliability and precision matter at an extraordinary scale. You'll partner closely with software engineers and quantitative researchers to build technology that processes enormous volumes of market data, powers quantitative research and supports live trading. You'll take on problems that don't have obvious answers — from designing high-performance systems and distributed infrastructure to eliminating bottlenecks measured in microseconds and building tools that make our researchers and engineers faster. Your work will go into production, be measured against real-world performance and have a direct impact on how our trading systems operate. If you enjoy solving hard engineering problems, understanding systems at a deep level and pushing technology to its limits, you'll feel right at home. What You'll Work On You'll work across the engineering stack that powers our quantitative research and trading platforms. Depending on your team, your work may include: Designing and building high-performance, low-latency systems in modern C++. Building distributed systems that process and analyze massive volumes of market data . Designing systems where latency, throughput and reliability directly influence trading performance . Working on Linux systems, networking, concurrency and multithreaded applications. Profiling systems, identifying bottlenecks and optimizing performance at the hardware and software level. Building robust infrastructure that supports quantitative research and live trading. Debugging complex production systems and solving problems where correctness and reliability are critical. Designing internal platforms and developer tools that accelerate research an

linuxrestai
View job →
G
15 days ago

About Graphcore Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems. The Team An exciting opportunity to join a new team within the Software Operations group. The Build Engineering team is a new function within Software Infrastructure, which focuses on the overall process of building and integration of the Machine Learn ing S oftware S tack. You will work closely with the QA and development teams to get an understanding of how our ML SW stack is built, helping to ensure good build practices, and proving that the stack works together and is reproducible in secure, sandboxed environments. Responsibilities and Duties Developing our internal t

pythondockerlinux
View job →
CH
15 days ago

Opportunity Overview: We are seeking a Lead Software Engineer to join our Integrations team. In this role, you will be designing, developing, and scaling highly available healthcare integration systems supporting prior authorization workflows across providers, payers, and delegated entities. You'll direct a fast-paced, autonomous,agile team of software engineers in the design, development, and operational support of a growing enterprise integration platform. This is an opportunity to drive technical excellence at the intersection of healthcare interoperability and modern distributed systems. What you’ll do: Technical Leadership: Provide technical leadership across architecture, system design, platform scalability, reliability, and operational excellence. Platform Engineering: Design and build scalable, resilient, and high-performing systems that support critical business workflows and enterprise integrations. Integration Solutions: Lead the development and maintenance of secure integrations with internal and external platforms, partners, and third-party systems. Cloud & Automation: Drive cloud infrastructure, deployment automation, and software delivery practices that enable reliable and efficient releases. Distributed Systems: Design and support event-driven and distributed architectures that enable scalable and fault-tolerant processing. Operational Excellence: Establish monitoring, observability, and incident response practices to ensure system reliability, performance, and availability. Quality Engineering: Champion automated testing, quality assurance, and engineering best practices throughout the software development lifecycle. Production Support: Lead the resolution of complex production issues and drive continuous improvement in platform stability and operational efficiency. Cross-Functional Collaboration: Partner with product, operations, data, security, and business stakeholders to deliver solutions aligned with organizational goals. Agile Delive

javaawsdocker
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
15 days ago

Opportunity Overview: This is a unique opportunity to join a high-caliber software engineering team that is experiencing rapid growth. You’ll play a key role in building impactful healthcare technology on a modern technology stack, with a focus on our core data and AI platforms. Your work will focus on enhancing the platform's key features, while also balancing scalability, reusability, and performance. As a Staff Engineer on the Application Engineering team, you’ll serve as a senior technical leader - responsible for designing and delivering high-quality, scalable software systems that power Cohere Health’s core platform. You’ll act as a multiplier, elevating the technical bar for the team, mentoring engineers, and partnering with product, data, and clinical teams to deliver solutions that meet compliance, quality, and performance standards. This role is ideal for engineers who thrive on solving complex problems in healthcare, have deep expertise in building distributed systems, and want to influence architecture and engineering practices at scale. What you’ll do: Technical Leadership & Architecture Define and drive the architecture of large-scale, distributed application systems across the Cohere platform. Ensure solutions are secure, performant, maintainable, and compliant with NCQA, CMS, and payer requirements. Champion engineering best practices in CI/CD, testing, release management, and observability. Hands-On Engineering Write clean, maintainable, and well-tested code, primarily in modern frameworks (e.g., Python, TypeScript/React, Java/Kotlin). Lead the development of core features and APIs that directly impact providers, payers, and patients. Partner with DevOps and Data teams to ensure seamless integration, scalability, and operational readiness. Quality & Compliance Focus Embed automated testing, monitoring, and release safeguards into the development lifecycle. Proactively address compliance and audit-readiness requirements in application

typescriptpythonjava
View job →
O
OpenAI
📍 New York• Full-time• Remote
22 days ago

About the Team API Frontiers turns OpenAI’s frontier models into production APIs that developers can use to build reliable products and agents. We own the core path connecting models to developers through the Responses API, with a focus on safety, reliability, and speed. Working closely with Research, Safety, Codex, and other API teams, we bring new model capabilities into production and improve them through developer feedback. About the Role We are looking for a backend software engineer to build and operate the services behind the Responses API. You will shape API behavior, bring new capabilities from research into production, and make long-running agent workflows dependable and fast. The work combines distributed systems engineering with product judgment: designing useful developer interfaces, managing staged rollouts, and following production issues through to durable fixes. In this role, you will: Design, build, and operate APIs and backend services that bring frontier model capabilities to developers. Partner with Research, Safety, Codex, and API teams to define API behavior and deliver safe, staged launches. Build API capabilities for agent workflows, including task delegation, context sharing, and parallel execution. Strengthen long-running request reliability across timeouts, cancellation, streaming, and background execution. Improve request-processing performance and tail latency through profiling, efficient systems code, and persistent connections. Turn developer feedback and production failures into better observability, diagnostics, and lasting product improvements. Your background might look something like: 5+ years of experience building and operating backend services or developer-facing APIs in production. Strong software engineering fundamentals, with practical knowledge of distributed systems, concurrency, and asynchronous execution. Ability to diagnose production failures and performance bottlenecks using observability data and profiling. Product

REMOTEawsrestai
View job →
R
Replit
📍 Foster City• Full-time• Remote
24 days ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: As a New Grad Software Engineer, you'll join a team of exceptional builders working on products that are reshaping how the world creates software. You'll have the opportunity to work on everything from our AI-powered development platform to the distributed systems that enable real-time collaboration for millions of developers. This is a chance to define your career while defining the future of software development. You'll work on problems that matter, with the autonomy to drive solutions and the support to grow into a technical leader. What you will build: Product features that delight users and make it possible for anybody to create software AI coding agent that understands intent and generates production-ready applications Cloud infrastructure that provides instant, powerful development environments at global scale Platform features that enable one click deployments and scale to millions of users Required skills and experience: Recent graduate (2027) with a degree in Computer Science, Computer Engineering, or related field Strong programming skills in a modern language (JavaScript/TypeScript, Python, Go, Rust) Full-stack capabilities with experience in React, Node.js, and database technologies Growth orientation - eager to learn new technologies and take on increasing responsibility Collaborative spirit - you work well in cross-functional teams and value diverse perspectives What we value : Problem-solving mindset: Ability to approach complex operational challenges systematically and devise effective solutions Self-directed and autonomous: Capable of working independently while collaborating effectively with cross-functional teams Strong communication skills: Ability to explain complex technical conce

REMOTEjavascripttypescriptpython
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

The Application Modernization Platform (AMP) team is dedicated to solving one of the industry's biggest challenges: transforming rigid, legacy applications that suffer from poor scalability and high operating costs into modern, microservices-based architectures on MongoDB. To accelerate this transition, MongoDB is building a dedicated Platform and Infrastructure team to develop the Application Modernisation Platform and Infrastructure . As we help customers modernize their application and data ecosystems, we face the challenge of deploying complex tooling into highly restrictive client environments and architecting automated verification frameworks to ensure data equivalence. This new team will architect the platform foundation and infrastructure that enables a "build once, run anywhere" model, ensuring our AI-powered modernisation suite operates seamlessly regardless of a client's security or network constraints. We are looking for engineers to join this high-visibility initiative, where you will solve unique distributed systems puzzles and help shape the future of how global enterprises leverage data and AI. We are looking for an experienced Software Engineer who thrives on solving infrastructure constraints and building developer-centric modernisation platforms with a strong background in building software testkits/frameworks. The ideal candidate will be designing and building automated frameworks that validate functional equivalence, performance benchmarks, and data integrity. From leveraging LLMs for unit test generation to building contract testing frameworks, your work will be the safety net for the world’s largest enterprise migrations. This role will be based in our India office in Gurgaon and offers a hybrid working model. Position Expectations Contribute high-quality, well-tested code to the modernization and framework team and its surrounding services Collaborate effectively with Product Management, other engineers, and designers to build and deliver on

javasqlmysql
View job →
🔔

Get new software engineer distributed systems salary india jobs by email

Daily job updates · Unsubscribe anytime