About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Debug Validation Lead will drive post-silicon debug and validation activities for next-generation AI compute silicon and systems. The role is responsible for leading teams focused on identifying, reproducing, analysing and resolving complex silicon, firmware and system-level issues during bring-up, characterization and product readiness. This position combines deep technical debugging expertise with strong cross-functional collaboration across multiple engineering disciplines. The role will work closely with architecture, RTL, firmware, software and systems teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The role will work closely with architecture, RTL, firmware, software, systems and platform teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The Team The Post-Silicon Debug and Validation team sits within the Architecture and Validation organisation and is responsible for bring-up, debug and validation of Graphcore silicon and systems. The
Jobs in India
Lead Software Engineer 2c Inference Performance Optimization in Bengaluru
234 active opportunities · Updated October 2026
Showing
15 jobs
Explore current lead software engineer 2c inference performance optimization jobs in Bengaluru. Filter by work mode, employment type, experience, department, date posted and distance.
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the role We are looking for a Quality Architect/Lead with hands-on experience building quality systems and has deep expertise in building and scaling test automation frameworks.The role requires an individual who applies systems thinking to solving complex problems. They should be able to understand the product from various perspectives and be able to effectively create testing programs that validate not just functionality but performance, reliability and user experience.DevRev is building a next generation AI native product that requires us to build novel test systems for the Agent AI platform. The role is mult-faceted and is going to continuously evolve with time. What you'll do Test Case Design and Documentation Actively use AI and intelligent agents to accelerate test generation, test maintenance, and coverage expansion. Leverage LLMs to convert requirements, user stories, and production incidents into high-quality automated test cases. Reduce reliance on manual test case creation by introducing AI-assisted automation workflows, with human review and ownership. Apply AI to optimize test selection, prioritization, and execution based on risk, code changes, and historical failures. Use AI to assist in identi
JOB TITLE Site Reliability Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO You will play a highly critical operational role where you will apply a combination of software and systems engineering skills to develop and maintain a complex set of distributed, real-time systems that serve critical stakeholders in Point72’s Global Macro business. You will focus on optimizing the operations of existing systems and infrastructure in an efficient manner, through a strict adherence to automation and tooling Specifically, you will: Build out foundational technical components of an extensive SRE program across multiple complex systems, both new and existing • Collaborate with our development and quant teams to ensure that ongoing change is consistent with a pre-determined, measurable set of SLOs spanning multiple complex user interactions with our systems • Monitor system capacity and performance, identifying and addressing potential future bottlenecks and sources of instability before they become impactful to our stakeholders • Review and provide feedback on automation code developed by peers to maintain high standards of code quality and efficiency • Troubleshoot and resolve system issues, analyzing their impact on infrastructure and service operations • Participate in or lead design reviews with peers and stakeholders, evaluating and selecting the best technologies and automation strategies for our needs WHAT’S REQUIRED We are looking for highly motivated, proactive engineers
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role combines deep technical expertise with people leadership responsibilities, including team development, prioritisation, mentoring and delivery coordination across multiple projects and stakeholders. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug complex issues, optimize workloads and continuously imp
Staff -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debu
Senior -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to deb
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services that our customers can trust. We embrace an automation-first mindset and continuously invest in platform engineering, observability, and operational excellence to enable our engineering teams to move quickly and safely. This role is ideal for an experienced Site Reliability Engineer who enjoys solving complex technical challenges at scale, building automation, and improving the reliability of production systems. You will serve as a key contributor within the EPG SRE organization, partnering closely with software engineers, architects, and product teams to design, build, and operate world-class cloud services. What You'll Be Doing Reliability & Operations Design, build, and operate large-scale cloud infrastructure and production services. Participate in an on-call rotation supporting highly available customer-facing systems. Lead incident response efforts and drive post-incident reviews focused on systemic improvements. Define, measure, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. Partner with engineering teams to improve service availability, scalability, performance, and resilience. Continuously improve observability through metrics, logging, tracing, dashboards, and alerting. Eng
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As the Engineering Manager for the FlashArray (FA) Foundation Quality Engineering team, you will build and lead a high-performing team of Quality Engineers and Software Test Developers in Bangalore. You will combine people leadership, technical judgment, and execution discipline to drive the quality, test automation, and release readiness of foundational FlashArray capabilities. Working at the intersection of software, hardware, firmware, and cloud infrastructure, you will partner closely with Development, Product Management, Release, and Support teams. Your focus will be translating complex roadmap requirements into modern test strategies, establishing strong shift-left CI/CD signals, and ensuring our products consistently deliver industry-leading reliability to our customers. WHAT YOU’LL DO Build & Lead a High-Performing Team: Hire, mentor, and coach a newly forming team of quality and automation engineers in Bangalore, fostering a culture of technical excellence, continuous learning, and shared ownership. Drive Quality Strategy & Shift-Left Automation: Define and execute the FA Foundation quality roadmap—including risk-based coverage, automated CI/CD pipeline integration, fault injection, and release-readiness criteria across software, firmware, and hardware. Partner Cross-Functionally for Delivery: Collaborate early in the design cycle with Development, Architecture, Product, and Release teams to impro
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join The Integrations & Application Engineering team owns the integration and automation platforms at Airbnb. We enable teams across the company to self-serve in connecting their systems and automating solutions, and we integrate Airbnb’s enterprise systems end-to-end. The team also builds and manages custom in-house solutions and the application infrastructure that supports them. This is a hybrid IC and people-leadership role focused on Finance integrations and automations. As a Tech Lead Manager, you will lead and grow a team of engineers while operating hands-on at a senior IC level, setting technical direction, owning delivery, and writing and reviewing code yourself. We’re looking for someone with deep ERP product knowledge who builds integrations in general-purpose languages, takes ownership of systems and people, and is enthusiastic about applying AI, building tools and integrations with Claude Code that extend what the team can do. A Typical Day Lead the Finance and ERP integration practice: set direction, run initiatives in parallel, and own the systems and processes your team supports. Balance management with hands-on work: unblock the team, review designs and code, and build complex integrations, backend services, APIs and MCPs yourself. Partner with Finance, Procurement, and cross-functional teams to translate requirements into integrations, automations, and reporting. Coach and grow engineers and contractors through design reviews, architecture discussions, and career development. Own the smooth operation of month-end financial integrations and reporting, proactively surf
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Role We are seeking a seasoned Manager, Software Engineering with 12+ years of experience to lead our Database Engineering and Cloud Infrastructure team. In this role, you will lead a team of high-performing engineers responsible for architecting, scaling, and optimizing multi-cloud relational and in-memory database platforms. You will bridge technical execution, engineering leadership, and strategic infrastructure planning across AWS and Azure environments. Key Responsibilities Technical Leadership & Architecture Lead the architectural design and operations of enterprise-grade, multi-cloud relational databases across AWS (RDS PostgreSQL, MySQL, Aurora) and Azure (Database for PostgreSQL/MySQL, Azure SQL Managed Instance). Drive high-availability architecture strategies, including Multi-AZ deployments, auto-failover groups, read replica scaling, and cross-region disaster recovery (DR). Oversee zero-downtime operations, including major-version engine upgrades, schema migrations, and blue/green deployment strategies. In-Memory Infrastructure & Open-Source Strategy Manage scale operations for in-memory datastores (AWS ElastiCache, Azure Cache for Redis), focusing on cluster mode operations, eviction policies, and persistence tuning. Spearhead open-source caching initiatives and migration pathways from Redis to Valkey (e.g., AWS ElastiCache for Valkey) using zero-downtime tools like RedisShake to ensure open-source license compliance and optimize cloud spend. Aut
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights, so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand and tackle issues, analyze performance and get the most of their software and infrastructure. Service Levels is one of New Relic's most commercially critical products — it powers SLO compliance and reliability measurement for thousands of customers. We're looking for an experienced Engineering Manager to lead a senior, high-performing team building the next generation of service level management at scale. You'll lead a team that includes lead-level engineers with deep domain expertise, and your value will come from enabling their best work — not directing it. You'll own delivery, quality, and team health while partnering closely with product and design to ship features that directly impact New Relic's commercial momentum. What you'll do Lead a full-stack engineering team of 6-8 across backend (Java/Spring Boot) and frontend (React/TypeScript) Own end-to-end delivery — sprint planning, execution, quality, and release Set clear expectations, manage performance equitably, and develop engineers at every level Partner with Product Manager and XD to translate requirements into technically sound, deliverable plans Drive architectural discussions and hold the team to engineering excellence standards Identify
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. SHOULD YOU ACCEPT THIS CHALLENGE... In this role as an Engineering Manager, you will lead a team of engineers located in Bengaluru, India. You will focus on driving and shaping the direction of our observability software and enabling product engineers to deliver high-quality, reliable software to our customers. You will provide technical leadership and direction, mentor engineers on your team, and collaborate with product, engineering, and cross-functional stakeholders to deliver successful outcomes. The successful candidate must understand the dynamics of global R&D, possess deep knowledge of local culture, and have the ability to champion Pure values and leadership attributes. This role requires the ability to lead and influence multiple stakeholders across cross-functional teams and drive alignment across complex, distributed engineering environments. The team will help build and evolve observability capabilities that provide actionable insights into the health, performance, capacity, and reliability of Pure's products and infrastructure. WHAT YOU'LL NEED TO BRING TO THIS ROLE... 12+ years of combined experience as a software developer and manager 3+ years of technical management experience while staying hands-on 7+ years of hands-on software development experience Strong exposure to one or more of the following areas: distributed systems, systems programming, observability/telemetry, data platforms, or solving prob
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Team lead for a group of engineers with responsibility for Linux and VMWare initiator stack behavior as well as Fibre Channel and NIC drivers on Pure storage array. WHAT YOU'LL DO The Kernel and Driver development team is responsible for several areas Working for a team that deals primarily in various storage area network protocols such as Fiber Channel and Ethernet. On the initiator side, the team will be responsible for Linux initiator behavior attached to Flasharray. Focus is on NVME (ROCE, FC, TCP/IP) but also includes FC-SCSI (FCP) and iSCSI interfaces. This includes software development/fixes for Linux initiator stack, debugging initiator problems, and creating compatibility documents for Purity. For VMWare, focus is on debugging issues with VMWare as an initiator, and creating compatibility documentation. For FibreChannel and NIC Drivers, focus is responsibility for FC-SCSI driver stack for storage side and responsibility for NIC drivers on networking side. This includes responsibility for related code, utilities, enhancements supporting RAS, as well as debugging failures found internally and in the field. Maintain Linux kernels for internal testing Documenting supported configurations for customers Responsibility for evaluating Linux initiator behavior and optimizing for Pure Storage Flasharray. This includes correctness as well as optimizing for performances. Will contribute bug fixes and enhancemen
Other cities to consider
More places hiring for this role
Get new lead software engineer 2c inference performance optimization jobs in Bengaluru, India by email
Daily job updates · Unsubscribe anytime