Jobiba hiring network

Software Engineer Data Infrastructure Salary India Jobs

6,326 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current software engineer data infrastructure salary india jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Agent Post-Training team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what OpenAI's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role As a member of Agent Post-Training, Connectors, you will teach models how to interface with the top professional software using code. You will help train agents to use code, APIs, tools, and structured integrations to operate across applications like Slack, Google Workspace, GitHub, Notion, Linear, Salesforce, and other core systems of work. You will help enable models to take useful actions across a user’s digital context: finding information, updating systems, coordinating work, generating artifacts, and completing multi-step workflows through the tools teams already use. You will train models to be supercharged by the world’s most important productivity and enterprise software, turning connected tools into a powerful action surface for our agents. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure whether it worked, and ship improvements into products used by real people.

awsgitrest
View job →

About Highnote Founded in 2020 by a team of leaders from Braintree, PayPal, and Lending Club, Highnote is an embedded finance company that sets the standard in modern card platform management. As an all-in-one card issuer processor and program management platform, we provide digital-first organizations with the flexibility to seamlessly issue and process payment cards, embed virtual and physical card payments, and integrate ledger and wallet functionalities—empowering businesses to drive growth and profitability. We’ve raised $145M+ and have grown our team to 140+ employees. Headquartered in San Francisco, we’ve managed to build one of the most advanced payments teams in the industry, with team members in 25+ US states. Operating through our core values of customer obsession, executional excellence, intentional inclusion, we’re helping businesses grow for the future by creating the payment products demanded by tomorrow, with the ability to solve for use cases that don’t exist yet. We are fast-moving, hands-on, and strongly believe everyone deserves a seat at the table. We believe we’re unlocking incredible opportunities that can change the future of payments, as long as we have the right people to make it happen. Job Description The Data Platform Team is part of the Data Organization and is responsible for designing and building modern scalable, secure and reliable data platforms and data pipelines. With the data platforms, we will be able to serve the product/customer data needs, improve operational efficiency and increase revenue. What you’ll be doing Building secure and scalable data platforms that support our product/customer needs and power data driven decision making. Writing high quality and robustly tested code, mostly in Java and Python. Understanding business/data needs and growth to build the right solutions. Documenting your work to support our rapidly-scaling company. Interfacing with internal business stakeholders and technology partn

pythonjavasql
View job →
DC
19 days ago

Role Overview Build reliable software services that power products, platforms, and business decisions. As a Senior Software Developer, you’ll design and deliver scalable applications, backend services, and integrations that perform well in production and evolve with changing business needs. You’ll apply strong software engineering practices across APIs, data-intensive applications, cloud services, AI-enabled solutions, and deployment pipelines. You’ll help shape technical solutions, improve system reliability, and contribute to a high-quality engineering culture. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Design and develop scalable backend services and applications using Python or TypeScript. Lead the development of APIs, integrations, reusable software components, and AI-enabled features. Build reliable solutions for data ingestion, manipulation, service-to-service communication, and intelligent automation. Apply AI technologies and modern software engineering practices to improve product capabilities, developer productivity, and operational efficiency. Make sound technical decisions around architecture, performance, security, scalability, and maintainability. Deploy and operate applications using AWS services and CI/CD practices while improving testing, monitoring, documentation, and delivery standards. These are the essentials you’ll need to get an interview 5+ years of professional experience developing and delivering production software. Strong hands-on experience with Python; TypeScript or similar languages is also valuable. Proven experience building backend services, APIs, integrations, and service-oriented applications. Experience applying AI technologies, such as generative AI, machine learning services, intelligent automation, or AI-enabled application features. Strong understanding of software design principles, testing, debugging, performance optimization, and secure development. Experience working with cloud platfor

typescriptpythonaws
View job →
DC
19 days ago

Must be based in Vancouver The role We're hiring a dedicated data engineer to own the production data platform that our delivery, product, and engineering teams run on; designing integrated, governed data pipelines and delivering automated reporting, AI-assisted workflows, and predictive signals on top of them. You'll write production code, design systems, own CI/CD, and be accountable for the correctness of data that leaders make decisions on. What you'll do Design and operate our cloud data platform: ingestion, transformation, orchestration and serving. Integrate data from across the business (delivery tooling, CRM, product telemetry, finance, support and customer feedback systems) with shared identifiers, data contracts and lineage. Build automated and continuously refreshed reporting so teams manage by exception rather than chasing status. Connect approved AI agents to governed data with structured outputs, provenance, guardrails and human approval in the loop. Build feature pipelines and the MLOps controls behind predictive use cases: tests, versioning, promotion gates and drift monitoring. Own the engineering standards for data: testing, observability, environment promotion, PII classification and access control. What you'll bring Strong software engineering fundamentals: production-quality code, API and interface design, testing discipline, systems design. Real experience building and operating production data platforms on a cloud warehouse or lakehouse (Snowflake and AWS preferred) with dbt and a modern orchestrator. Practical AI tooling experience: something shipped, not prototyped. LLM-backed classification, extraction or structured-output pipelines; agent and tool-calling workflows; retrieval; evals. You can reaso

awsci/cdgit
View job →
P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Data Governance team makes sure Plaid handles consumer and customer data responsibly — and can prove it. Our mission is to enforce Plaid's privacy commitments and regulatory obligations in the systems themselves rather than in policy documents: we build the platform and controls that govern how data flows through Plaid — where it lives, who can use it, for what purpose, and for how long. That includes verifiable deletion of consumer data on request, enforcement of data-use restrictions so downstream systems can only use data in permitted ways, and the cataloging and classification that let Plaid know what data it holds and how sensitive it is. We operate at the scale of Plaid's entire data footprint, and correctness and auditability matter to us as much as throughput. As a Staff Software Engineer on Data Governance, you will set the technical direction for how Plaid enforces data governance at scale. You'll lead the design of distributed backend systems that reliably delete, restrict, and track data across dozens of services, making architectural decisions whose blast radius spans the whole company. You'll drive multi-quarter initiatives from ambiguous privacy and regulatory requirements through

javaawsrest
View job →
R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Who We Are: Shape the future of Roblox’s virtual economy. The Economy ML team is building the machine learning backbone that powers Roblox’s Marketplace, Developer Monetization, and Payments ecosystems. From intelligent pricing and personalized storefronts to dynamic layout optimization and avatar understanding, we’re reimagining how the Roblox economy drives user engagement, monetization, and creator success at scale. As a Principal Software Engineer (Data Systems) , you will architect, build and deploy high-scale, reliable real-time and batch data systems for personalization, search and recommendation across various product surfaces in Marketplace, Developer Monetization and Payments. You will be involved in key data projects from architecting event taxonomies and logging interfaces to real-time feature serving across multiple search and recommendation surfaces. What You’ll Do Act as data engineering lead for Economy ML, setting standards for batch vs streaming feature pipelines, table design, observability, and documentation used across the Economy group. Work as a hands-on contributor on our data systems to power content recommendation, search and personalization across Economy product

awsgitmachine learning
View job →
DC
19 days ago

Role Overview Build the software services that power products, platforms, and better business decisions. As a Software Engineer II, you’ll develop scalable backend applications, APIs, integrations, and AI-enabled features using Python and cloud technologies. You’ll contribute to solutions from design through production, helping improve reliability, performance, security, and developer productivity. This is an opportunity to solve meaningful engineering challenges, grow your technical ownership, and collaborate with experienced engineers across the development lifecycle. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Design and build scalable backend services, REST APIs, integrations, and reusable software components using Python and, where relevant TypeScript. Develop data ingestion, transformation, service-to-service communication, and automation capabilities that support reliable product experiences. Contribute to AI-enabled features and use AI development tools responsibly to improve coding, testing, research, documentation, and delivery. Apply sound engineering practices across architecture, performance, security, testing, debugging, and maintainability. Deploy and operate services using AWS and CI/CD workflows, contributing to monitoring, troubleshooting, documentation, and continuous improvement. Partner with engineers and cross-functional colleagues through design discussions, code reviews, technical problem-solving, and knowledge sharing. These are the essentials you’ll need to get an interview 3–5 years of professional experience building and delivering production software in an agile environment. Strong hands-on experience with Python and backend development, including APIs, integrations, or service-oriented applications. Experience working with cloud platforms, preferably AWS, and familiarity with deployment or CI/CD practices. Working knowledge of software design principles, testing, debugging, performance optimization, an

typescriptpythonreact
View job →
A
Anyscale
📍 Remote• Full-time
1mo ago

About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About Ray Data Team: Ray Data is Python-native data processing engine that is a one stop shop for all AI data processing needs. Ray Data provides performant, first-class integration with cutting edge AI frameworks using both multi-modal and structured data. The Ray Data team currently develops and maintains Ray Data . We are a team of engineers passionate about building a Data processing engine which is a one-stop shop for all of your ML/AI needs. We are looking for exceptional engineers to build, optimize, and scale Ray for modern and increasingly complex AI workloads. As part of this role, you will: Improve the performance of Ray Data and multi-modal batch inference use cases. Ensure efficient scaling across different stages of the Data pipeline in a heterogeneous environment. Building data loading solutions for production training workloads. Focus on stability and fault tolerance at high scale Working with customers and new age AI native companies in scaling their AI workloads. We'd love to hear from you if have: At least 3-4 years of relevant work experience Solid background in building scalable and fault-tolerant distributed systems Experience with data processing, database internals. Passionate about large

pythonmachine learningai
View job →
A
1mo ago

About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About Ray Data Team: Ray Data is Python-native data processing engine that is a one stop shop for all AI data processing needs. Ray Data provides performant, first-class integration with cutting edge AI frameworks using both multi-modal and structured data. The Ray Data team currently develops and maintains Ray Data . We are a team of engineers passionate about building a Data processing engine which is a one-stop shop for all of your ML/AI needs. We are looking for exceptional engineers to build, optimize, and scale Ray for modern and increasingly complex AI workloads. As part of this role, you will: Improve the performance of Ray Data and multi-modal batch inference use cases. Ensure efficient scaling across different stages of the Data pipeline in a heterogeneous environment. Building data loading solutions for production training workloads. Focus on stability and fault tolerance at high scale Working with customers and new age AI native companies in scaling their AI workloads. We'd love to hear from you if have: At least 3-4 years of relevant work experience Solid background in building scalable and fault-tolerant distributed systems Experience with data processing, database internals. Passionate about large

pythonmachine learningai
View job →
T
Twilio
📍 - US• Full-time• Remote• $138.7K – $173.4K/yr
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio's next Software Engineer on our Data & Analytics Platform Who we are & why we’re hiring Twilio powers real-time business communications and data solutions that help companies and developers worldwide build better applications and customer experiences. Although we're headquartered in San Francisco, we have presence throughout South America, Europe, Asia and Australia. We're on a journey to becoming a global company that actively opposes racism and all forms of oppression and bias. At Twilio, we support diversity, equity & inclusion wherever we do business. About the job We are looking for a talented and experienced Software Engineer to join our Data Platform team. In this role, you will play a crucial part in designing, building, and optimizing our platform to support a wide range of data-driven initiatives. You will work closely with cross-functional teams to understand business requirements, architect scalab

REMOTEpythonjavaaws
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Data Engineering team builds and maintains the foundational datasets that power decision-making across Robinhood. We design reliable, scalable data systems that support product analytics, growth strategy, financial reporting, experimentation, and machine learning. The team partners closely with Product, Engineering, Data Science, and Finance to ensure accurate, well-modeled data is available to teams across the company. Our work directly influences how Robinhood measures performance, improves customer experience, and scales its products. As a Senior Data Engineer, you will design, build, and evolve core datasets that track product performance and company-wide metrics. You will develop scalable data pipelines that ingest application events and database snapshots into our data lake, ensuring high data quality and reliability. You’ll collaborate with application engineers to improve data generation patterns and with analytics teams to design intuitive, well-documented data models. This is an opportunity to shape the technical foundation that supports data-informed decisions across the organization! This role is based in our Menlo Park, CA office, with in-person attend

pythonvuesql
View job →

About Bazaarvoice At Bazaarvoice, we create smart shopping experiences. Through our expansive global network, product-passionate community & enterprise technology, we connect thousands of brands and retailers with billions of consumers. Our solutions enable brands to connect with consumers and collect valuable user-generated content, at an unprecedented scale. This content achieves global reach by leveraging our extensive and ever-expanding retail, social & search syndication network. And we make it easy for brands & retailers to gain valuable business insights from real-time consumer feedback with intuitive tools and dashboards. The result is smarter shopping: loyal customers, increased sales, and improved products. The problem we are trying to solve : Brands and retailers struggle to make real connections with consumers. It's a challenge to deliver trustworthy and inspiring content in the moments that matter most during the discovery and purchase cycle. The result? Time and money spent on content that doesn't attract new consumers, convert them, or earn their long-term loyalty. Our brand promise : closing the gap between brands and consumers. Founded in 2005, Bazaarvoice is headquartered in Austin, Texas with offices in North America, Europe, Asia and Australia. It’s official: Bazaarvoice is a Great Place to Work in the US , Australia, India, Lithuania, France, Germany and the UK! About the Team We are building our next-generation Data Enrichment Platform and we need a heavy hitter to help us scale. We aren't just moving data; we are building the real-time engine that powers our business. As a core member of our engineering team, you will architect high-throughput pipelines, solve complex latency challenges, and define the standards for a robust "Lakehouse" architecture. If you obsess over JVM internals, distributed consistency, and Exactly-Once semantics, this is the role for you. How You'll Make an Impact: Architect at Scale: Lead the design of dis

javaawskubernetes
View job →
T
Twilio
📍 - US• Full-time• Remote• $171.1K – $213.9K/yr
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio's next Staff Software Engineer on our Data & Analytics Platform Who we are & why we’re hiring Twilio powers real-time business communications and data solutions that help companies and developers worldwide build better applications and customer experiences. Although we're headquartered in San Francisco, we have presence throughout South America, Europe, Asia and Australia. We're on a journey to becoming a global company that actively opposes racism and all forms of oppression and bias. At Twilio, we support diversity, equity & inclusion wherever we do business. About the job We are seeking an experienced Staff Engineer to join our Data Substrate team. In this role, you will be responsible for architecting scalable and reliable data solutions, collaborating closely with cross-functional partners driving technical innovation, and mentoring a team of talented engineers. The ideal candidate will have deep expertise in d

REMOTEpythonjavaaws
View job →
P
Pinterest
📍 San Francisco• Full-time• Remote• From $177.2K/yr
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business. We are seeking a Staff Data Engineer to lead the design, implementation, and evolution of our identity services and data governance platform. This role is critical to ensuring trusted, privacy-safe, and well-governed data across the organization. You will work at the intersection of da

REMOTEawsgitrest
View job →
🔔

Get new software engineer data infrastructure salary india jobs by email

Daily job updates · Unsubscribe anytime