About the Role Together AI runs one of the largest GPU fleets in the world. The Infra Agent Systems team builds the software systems that power and automate that infrastructure. We develop production AI agents that diagnose hardware failures, investigate incidents, correlate signals across the fleet, and automate operational workflows. Alongside these agents, we build the platform they run on, including knowledge graphs, retrieval systems, orchestration frameworks, and developer tooling. You’ll work across two areas: Infrastructure Agent Systems — Build production AI agents that help operate our GPU fleet by diagnosing failures, investigating incidents, gathering evidence from live systems, and assisting with remediation. These agents are used every day by our infrastructure and datacenter teams through APIs, CLI, dashboards, and Slack. Core Agent Platform — Build the platform that powers these agents, including knowledge graphs, search and retrieval, orchestration, evaluation, and the tooling that enables agents to reason, act, and continuously improve. We’re working on something that hasn’t really been done before: building knowledge graphs and self-improving AI agents that understand, operate, and continuously improve large-scale AI infrastructure. This is an opportunity to work at the intersection of AI agents, distributed systems, infrastructure, and automation , solving challenging engineering problems with real production impact. There’s an enormous amount to build, learn, and shape as we define the future of autonomous infrastructure. responsible for delivering the software but also for operating and supporting it in production. Why this Role You’ll work on two hard problems at the same time: making AI agents trustworthy enough to operate production infrastructure, and building the knowledge, retrieval, and distributed systems that make those agents effective. You’ll have the opportunity to build foundational systems from the ground up, work on infrastructur
Jobs in India
Senior Gpu Memory Architect in India
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current senior gpu memory architect jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
We are seeking a qualified Senior Software Tools Development Engineer to join our GPU SWQA team. The successful candidate will have strong experience applying AI technologies to automate test cases and a deep understanding of Windows operating systems. Extensive knowledge of GPU, CPU, SoC, x86, and ARM architectures is required, along with expertise in PC I/O architecture and common bus interfaces such as PCIe, USB, and SATA. Familiarity with specifications for general PC architecture components is a plus. What you’ll be doing: Design and implement automated tests incorporating AI technologies for NVIDIA's device driver software and SDKs on windows platforms. Build tools/utility/framework in Python, C# or equivalent which would help automate and optimize the testing workflows in GPU domain. Develop and carry out automated and manual tests, analyze results, identify and report defects. Rigorously drive test automation initiative. Build innovative ways to automate and expand our software testing. Expose defects and constraints; Isolate and debug the issue(s) and find the root cause; Contribute to the solution and drive to closure. Measure code coverage for the software under test, analyze and drive code coverage enhancements. Develop applications and tools that accelerate development and test workflows and write fast, effective, maintainable, reliable and well documented code. Generate and test compatibility across a range of products and interfaces and validate different key software applications across a test matrix designed to test both breadth and depth. Provide peer code reviews including feedback on performance, scalability and correctness. Report test coverage and Go/No-Go status for deliverables, escalate critical issues, and drive them to closure. Participate in root cause analysis and corrective actions to continuo
NVIDIA has continuously reinvented itself. Our invention of the GPU sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. Today, research in artificial intelligence is booming worldwide, which calls for highly scalable and massively parallel computation horsepower that NVIDIA GPUs excel. NVIDIA is a “learning machine” that constantly evolves by adapting to new opportunities that are hard to solve, that only we can address, and that matter to the world. This is our life’s work , to amplify human creativity and intelligence. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join our diverse team and see how you can make a lasting impact on the world! As a Formal Verification Engineer at NVIDIA, you will be responsible for formally verifying complex designs. NVIDIA has developed a strong functional formal verification methodology that not only enables hardware design and verification engineers to use lightweight FV tools and techniques successfully but also allows FV engineers to use advanced property proving techniques on complex and/or critical RTL logic. The job involves very close interaction with the design team, architecture team, with other validation teams, and with NVIDIA's internal FV R&D group that develops functional verification tools using formal verification technology. What you'll be doing: You will help decide on the best applications of formal verification techniques to various parts of the design. Review functional and micro-architectural specifications, define the scope for formal verification, and create high-quality formal verification testplans to sign-off on the corresponding design implementation. Build formal verification testbenches, code assertions and constraints, and apply abstra
NVIDIA's Deep Learning GPUs have ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company”. We are growing our company and the team with the smartest people in the world. We are looking for extraordinary Software Engineers to develop and productize NVIDIA's DRIVE OS software. As a member of NVIDIA's Solution Engineering team, you will adapt DRIVE OS solutions to various car platforms equipped with different sensors. We are looking to hire Senior System Software Engineer – AUTOSAR. Ideal candidate will have very strong programming skills, a good grasp of HW & SW Architectures, a solid exposure to AUTOSAR & related architecture, tools and frameworks. What you will be doing: Participate and provide inputs and recommendation into AUTOSAR Architecture evolution with design choices, tools and methodology Architectural explorations on both SW and HW fronts which include feasibility studies, quick prototyping, profiling, safety studies, data analysis and presentation of results Influence next-gen HW architectures and SW Architecture and design Drive complex technical issues to closure that may occur interacting with cross-teams What we need to see: BS/MS, or equivalent experience 5+ years of experience Strong programming skills in C/C++ and scripting skills in Perl, Python etc Good experience and com
About the Role REMOTE IN INDIA We're looking for a software engineer to build the Kubernetes-native control plane that provisions and runs our GPU inference fleet. You'll design a manifest-driven API where the inference team declares what they need, whether that's a cluster, a model deployment, or a capacity change, and our controllers handle the reconciliation, provider/runtime selection, and lifecycle management underneath, so the inference team never has to know or care which specific serving stack, scheduler, or hardware pool is doing the work. You'll also build the systems that keep the fleet efficient, not just running, including defragmentation and rebalancing logic that consolidates scattered workloads back into contiguous capacity, and scheduling/bin-packing improvements that push GPU utilization up without hurting latency. The core value we're after is decoupling the people building on top of the platform from the operational and runtime complexity underneath, while squeezing more usable capacity out of the same hardware. You'll build the controllers, reconciliation loops, and self-service surface (API/CLI, not tickets) that make that decoupling real, plus the event-driven health, remediation, and utilization systems that keep it running and efficient without a human in the loop. Strong candidates have hands-on experience with Kubernetes controller/CRD patterns, have built or operated a platform API that abstracts multiple backends behind one interface, understand GPU scheduling and capacity efficiency (fragmentation, bin-packing, right-sizing), and think about GPU infrastructure as software to be engineered. A product mindset - you've built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship. You build it, you own it. You are not only responsible for delivering the software but also for operating and supporting it in production. Responsibilities Build the provisioning state machine
Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... Here at GoDaddy, the ML Engineering (MLE) team exists as the backbone of our machine learning infrastructure, enabling ML scientists and product teams across Domains to ship models to production reliably, efficiently, and at scale. This team owns the full lifecycle of ML systems — from CI/CD pipelines and model serving infrastructure to GPU workload orchestration and observability. Through disciplined engineering practices, thoughtful system design, and close collaboration with ML scientists, data engineers, and product teams, we deliver the platform that powers domain search, pricing, recommendations, and emerging AI experiences for millions of customers worldwide. We are currently looking for an experienced, highly motivated Senior Engineering Manager to lead our ML Engineering team based in India. This is an established team with existing engineers — we expect the candidate to ramp up quickly on our ML infrastructure stack, build strong relationships with the team, and partner with both India-based teams and US-based teams to drive execution and grow the team further. This individual will join us on our journey to build and scale ML infrastructure that serves real-time predictions at low latency, automates model deployment and promotion, and provides the observability and reliability guarantees that production ML systems demand. Become part of a team that bridges the gap between ML research and production engineering — shipping systems that directly impact GoDaddy's core revenue. What you'll get to do... Lead a team o
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Pure Solutions team as a Senior MLOps Solutions Engineer to architect and build high-scale, enterprise-grade AI/ML solutions. You will be instrumental in integrating Pure Storage platforms with the evolving open-source MLOps ecosystem (Kubeflow, MLflow, Ray) to operationalize the complete machine learning lifecycle. This role requires a creative technologist with deep Python expertise to drive innovation and enable our customers and partners to achieve production AI success. WHAT YOU'LL DO Design and Automate MLOps Pipelines: Lead the development of end-to-end MLOps workflows using CI/CD tools (Git/Jenkins) and orchestration platforms (MLflow/Kubeflow), specifically integrating Pure Storage's FlashBlade, FlashArray, and Portworx as the high-performance data plane for data ingestion, training, and inference. Build High-Performance AI/ML Reference Architectures: Create validated, repeatable deployment models using Infrastructure as Code (e.g., Ansible, Terraform) for AI/ML environments spanning bare metal, virtual machines, and GPU-accelerated Kubernetes clusters, ensuring optimal performance for distributed training. Optimize and Operationalize GPU Inference: Architect and implement solutions for high-throughput, low-latency model serving, utilizing technologies like NVIDIA Triton Inference Server and advanced optimization techniques (quantization, model sharding like DeepSpeed/Megatron-LM, and dynamic bat
Senior Data Scientist Description - The Team We are an expanding team at HP that develops applications that make use of Generative AI and Large Language Models. We work with business units mainly from the Commercial Organization to develop and run solutions that help our sales teams and customers. Responsibilities Defines and implements AI solutions to create business value and innovation. Works with Data Science leaders to develop new innovative solutions to existing or new business challenges. Develops clear presentations for business stakeholders and managers. Manages relationships with business partners to evaluate and foster data driven innovation, provide domain-specific expertise in cross-organization projects/initiatives. Knowledge & Skills Proficiency in Python and PySpark Programming: Strong coding skills in Python for data manipulation, model development, and integration with Azure services. Experience with Azure Services: Knowledge of key Azure services like Azure Machine Learning, Azure AI Search, and Azure Functions for deploying RAG systems. Expertise in Databricks: Ability to design, develop, and optimize workflows in Azure Databricks for data processing and feature engineering. Understanding of NLP and GenAI concepts: Familiarity with Large Language Models, prompt engineering for LLMs, vector databases, and Retrieval Augmented Generation (RAG) systems. Experience in Applied Statistics and Algorithms: Use statistics, mathematics, algorithms, and programming to address business challenges. Familiarity with the deployment and scaling RAG systems in production, using Azure’s containerization options such as Docker, AKS (Azure Kubernetes Service), or Azure Funct
Senior Data Scientist Description - Job Summary • This role is responsible for enabling innovation and creativity by bringing cutting edge perspectives on adopting latest data mining and modelling techniques. The role understands current complex business problems and future business strategy to assess, build and deploy required data mining and modelling capabilities. The role is involved in driving standardization, productivity and cross team learning by establishing processes and SOPs for entire data model development lifecycle. The role drives excellence through continuous improvement in model accuracy and reliability. Responsibilities • Leads organization wide team or teams of other data science professionals in complex projects to mine data using modern tools and programming languages. • Defines models to uncover patterns and predictions creating business value and innovation. • Manages and creates relationships with business partners to evaluate and foster data driven innovation, provides domain-specific expertise in cross-organization projects/initiatives. • Ties insights into effective visualizations communicating business value and innovation potential. • Works with various stakeholders, including business leaders, engineers, product managers, and data analysts, to identify business problems and develop data-driven solutions. • Prepares and presents literature, presentations, invention disclosures for peer review & publication in industry data science domain initiatives and conferences. • Assures insights are communicated regularly and effectively, reviewing designs, models and data compliance. • Defines, communicates and drives data insights/innovation into the business. • Leverages recognized domain expertise, business acumen, and overall data systems leadership to influence decisions of executive business
Senior Field Technical Support Consultant Description - The Senior Field Technical Support Consultant is responsible for driving technical escalation management, service quality improvement, partner enablement, operational governance, and customer satisfaction across HP India operations. The role serves as a key technical advisor and escalation leader, collaborating with Field Operations, ATS, Care Center, Supply Chain, Quality, Engineering, and business stakeholders to resolve complex customer issues, improve service delivery performance, and drive continuous operational improvement. Key Responsibilities Service Quality & Operational Excellence Monitor and improve key service metrics including RR30, PPSNC, AIR, AFR, CE Scorecard, and repair quality indicators. Identify operational improvement opportunities through data analysis and trend reviews. Develop and implement corrective action plans to improve overall service performance. Support cost reduction, waste elimination, and serviceability improvement initiatives. Partner Enablement & Capability Development Design and conduct technical, process, and quality training programs for HP partners and subcontractors. Lead Train-the-Trainer (TTT) initiatives, field coaching, audits, and ride-along programs. Develop SOPs, knowledge articles, troubleshooting guides, checklists, and customer education materials. Drive partner readiness for new products, diagnostics, and service processes. Customer Experience & Governance Work closely with strategic customers and account teams to resolve business-critical issues. Improve customer experience through proactive communication, technical advisories, and escalation management. Ensure warranty compliance, service policy adherenc
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Specialist, Payroll Implementation & Integration-1 Position Summary The Senior Specialist, Payroll Implementation & Integration supports strategic payroll initiatives, including mergers and acquisitions (M&A), P&C Modernization programs, system implementations, and process improvement efforts. This role serves as a key execution resource responsible for coordinating testing activities, validating payroll and employee data, supporting implementation readiness, and assisting with solution deployments across payroll operations. Key Responsibilities M&A Integration Execution Support • Support payroll workstreams during mergers, acquisitions, divestitures, and organizational restructures. • Assist with payroll integration planning and execution activities. • Coordinate employee data collection, validation, and migration activities. • Prepare implementation documentation and integration trackers. • Support payroll readiness assessments and integration milestones. • Execute project tasks, risks, issues, and action items. • Assist in stabilization activities following acquisitions. P&C Modernization Implementation Support • Participate in payroll-related activities supporting Workday and P&C Modernization initiatives. • Support configuration validation
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Specialist, Payroll Implementation & Integration-2 Position Summary The Senior Specialist, Payroll Implementation & Integration supports strategic payroll initiatives, including mergers and acquisitions (M&A), P&C Modernization programs, system implementations, and process improvement efforts. This role serves as a key execution resource responsible for coordinating testing activities, validating payroll and employee data, supporting implementation readiness, and assisting with solution deployments across payroll operations. Key Responsibilities M&A Integration Execution Support • Support payroll workstreams during mergers, acquisitions, divestitures, and organizational restructures. • Assist with payroll integration planning and execution activities. • Coordinate employee data collection, validation, and migration activities. • Prepare implementation documentation and integration trackers. • Support payroll readiness assessments and integration milestones. • Execute project tasks, risks, issues, and action items. • Assist in stabilization activities following acquisitions. P&C Modernization Implementation Support • Participate in payroll-related activities supporting Workday and P&C Modernization initiatives. • Support configuration validation
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Technical Program Manager 1. Overview This role provides Tier 1 and Tier 2 support for Jira Cloud users across Mastercard programs, ensuring seamless onboarding, issue resolution, and vendor coordination. You will work closely with engineering, product, and vendor teams to maintain operational excellence and drive adoption of Jira Cloud and integrated tools like Aha! and X-ray. 2. Role • Serve as the primary contact for Jira Cloud support , handling Remedy tickets, onboarding queries, and user escalations. • Facilitate onboarding of programs to Jira Cloud, including license validation, test case migration, and user provisioning. • Conduct demos and office hours to guide users through Jira workflows and integrations. • Partner with engineering teams to troubleshoot configuration issues, permission models, and data sync errors between Jira and Rally/Aha • Support CRQ planning and execution for Jira schema updates, field validations, and migration logic enhancements. • Coordinate with vendors to resolve plugin issues, validate architecture diagrams, and ensure compliance with Mastercard’s security standards. • Monitor and clean up orphaned work items, duplicate epics, and initiatives across programs using Domo reports and Confluence documentation. • Maintain Confluence pages for s
Work Flexibility: Hybrid What you will do: Key Job Responsibilities · Develop and maintain automated test frameworks and test scripts. · Perform Functional, API, Integration, System, and Regression Testing. · Integrate automated tests into CI/CD pipelines. · Collaborate with Development, DevOps, and Product teams. · Analyze test failures, debug issues, and improve automation coverage. What You need: Experience · 4-6 years in Software Testing and Test Automation Educational Qualification · B.E./B.Tech/M.Tec in Computer Science, Information Technology, · Graduated from a reputed engineering institute such as VIT, NIT, IIIT, BITS Pilani, DTU, NSUT, Manipal, SRM, PSG, or an equivalent institution. Key Job Responsibilities · Develop and maintain automated test frameworks and test scripts. · Perform Functional, API, Integration, System, and Regression Testing.Integrate automated tests into CI/CD pipelines. · Collaborate with Development, DevOps, and Product teams.Analyze test failures, debug issues, and improve automation coverage. Required Technical Skills Programming Languages · Java, Python, C#, TypeScript Test Automation · Playwright, Selenium WebDriver, Appium, Postman, Cu
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Job Overview: We are looking for a Senior Engineer to join the FGA DevEx team and help evolve our end to end developer experience across both OSS and SaaS. This team owns the SDKs in Go, JavaScript, .NET, Python, Java and other languages, along with CLI workflows, IDE integrations, GitHub automation, developer documentation, and release strategy. All development is done in the open as open source, and we actively welcome and review community contributions. Our guiding principle is One developer experience, many deployment models. As a Senior Engineer, you will take ownership of significant portions of the SDK and tooling ecosystem, ensure high quality implementations across languages, and contribute to a consistent and reliable developer experience. Responsibilities: Maintain and enhance existing SDKs for FGA in Go, JavaScript, .NET, Python, and Java, leveraging our SDK generator framework. Customize and refine SDK templates and wrappers to ensure consistency across languages and support configuration overrides such as store ID, authorization model ID, headers, and parallelization limits. Implement and improve core SDK features including client credentials authentication flows, robust error mapping, retry logic with jitter, and rate limiting safeguards. Implement advanced capabilities such as BatchCheck, ListRelations, and non transactional write operations with appropriate parallelization and performance considerations. Contribute to the SDK generator tool
Other cities to consider
More places hiring for this role
Get new senior gpu memory architect jobs in India by email
Daily job updates · Unsubscribe anytime