Jobs in India

Lead Senior Staff Technical Program Manager 2c Site Reliability Engineering in India

1,055 active opportunities · Updated October 2026

Explore current lead senior staff technical program manager 2c site reliability engineering jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

34/100

watch · 38 related jobs

Hiring trend

-34.8%

Job postings compared with the previous 30 days

Remote options

2.6%

Share of matching jobs listed as remote

P
📍 India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Principal NetSuite Consultant | Non-Profit As a Principal NetSuite Consultant at Plative, you will be the senior functional authority for our nonprofit customer portfolio, including community foundations, food banks, and First Nations organizations. You will lead complex ERP transformations end to end, from executive discovery through go-live, and mentor consultants across the team. You will advise customer stakeholders on nonprofit finance and program processes, and partner with sales on pre-sales scoping and solution architecture. Key Responsibilities Lead executive-level discovery sessions and map fund accounting, grant management, and donor management requirements to recommended NetSuite configurations. Own the full project lifecycle for complex NetSuite implementations: solution design, configuration, and go-live readiness. Act as the senior subject matter expert for nonprofit finance: fund accounting, grant budgeting and reporting, restricted/unrestricted revenue, and multi-funder allocations. Guide program and grant workflows specific to community foundations, food banks, and First Nations/Indigenous organizations. Mentor senior consultants and analysts, and review their solution designs and configuration work across the portfolio. Partner with sales and account teams on pre-sales scoping, technical discovery, and SOW development. Identify project risks and technical constraints early, and work with delivery leadership to resolve them. Maintain executive-level customer relationships, and identify opportunities for platform optimization and expansion. Build reusable delivery assets, including blueprints and templates, for nonprofit engagements. Basic Qualifications 7+ years of experience implementing and configuring NetSuite ERP, with a focus on Nonprofit organizations, at least one full cycle implementation of NFP organization. Strong expertise in fund accounting, grant management, and restricted/unrestricted revenue recognition. Deep understanding of nonprof

RestAgileAIGo
E(
📍 India· Full-time
✓ Quality checkedCompany trend -93.7%

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Role Overview & Key Responsibilities This is a high-leverage leadership role that spans architecture, execution, and org-building, and will shape the direction of our AI / ML initiatives at Ema. We are seeking an AI / ML technical leader who can take a vision and build it. As a Principal ML Engineer at Ema, you will be a senior technical leader responsible for shaping the machine learning roadmap, architecting large-scale ML systems, driving innovation, and ensuring our mixture of expert models (LLM + SLM + Custom Model) is accurate and performant at scale. You will collaborate across teams (research, product, infra, data, etc.), mentor senior engineers, and influence strategy and execution at company-wide levels. Responsibilities Lead the technical direction of GenAI and agentic ML systems that power enterprise-grade AI agents — spanning reasoning, retrieval, tool use, and integrations across various SaaS products. Architect, design, and implement scalable production pipelines for model training, fine-tuning, retrieval (RAG), agent orchestration, and evaluation — ensuring robustness, latency efficiency, and continuous learning. Define and own the multi-year ML roadmap for GenA

PythonJavaMachine LearningAI
C
📍 India
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Managing Principal – Banking Transformation & Team Leadership (Chennai) Role Overview As a Managing Principal at Capco, you will be a senior leader responsible for driving client growth, leading complex transformation programmes, and shaping strategic direction across Tier 1 banking clients. You will combine deep domain expertise, delivery leadership, and commercial acumen to deliver measurable business outcomes. This role will play a critical leadership role in expanding Capco's footprint in India, working closely with APAC leadership to deliver large-scale banking transformation programmes and build sustainable, high-performing teams. Experience in workforce transformation and operating model optimisation would be advantageous. Key Responsibilities Client Leadership & Growth Build and manage C-level relationships across banking clients (WRB, Corporate, Operations, Technology). Support origination and closing large transformation deals (primarily in Operations & Technology). Define and lead account growth strategy, including pipeline development and GTM propositions. Act as a trusted advisor on banking transformation, operational excellence, and workforce strategy. Programme & Delivery Leadership Lead complex, multi-year transformation programmes across operations & technology. Drive design and implementation of target operating models, including process, technology, and workforce. Ensure delivery excellence through strong governance, risk management, and KPI tracking. Mobilise cross-functional teams (onshore/offshore) to deliver outcomes at scale. People & Practice Leadership Build and scale high-performing consulting teams in Chennai and broader India. Own capability development and workforce planning and support hiring aligned to growth ambitions. Mentor team and lead capability build across key areas (programme delivery, workforce management, operations). Commercial & Business Management Drive revenue growth, utilisation, and margin

SL
📍 Bengaluru, KARNATAKA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You

ReactSQLAWSGit
SL
📍 Noida, Uttar Pradesh, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You

ReactSQLAWSGit
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug compl

PythonLinuxAIC++
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role combines deep technical expertise with people leadership responsibilities, including team development, prioritisation, mentoring and delivery coordination across multiple projects and stakeholders. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug complex issues, optimize workloads and continuously imp

PythonLinuxAIC++
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Debug Validation Lead will drive post-silicon debug and validation activities for next-generation AI compute silicon and systems. The role is responsible for leading teams focused on identifying, reproducing, analysing and resolving complex silicon, firmware and system-level issues during bring-up, characterization and product readiness. This position combines deep technical debugging expertise with strong cross-functional collaboration across multiple engineering disciplines. The role will work closely with architecture, RTL, firmware, software and systems teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The role will work closely with architecture, RTL, firmware, software, systems and platform teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The Team The Post-Silicon Debug and Validation team sits within the Architecture and Validation organisation and is responsible for bring-up, debug and validation of Graphcore silicon and systems. The

PythonGitLinuxAI
CH
📍 Hyderabad, TELANGANA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Opportunity Overview: We are seeking an experienced and driven Sr. Technical Talent Acquisition Partner to join our Talent Acquisition team and spearhead hiring initiatives for our Healthcare Technology organization. The ideal candidate will have strong technical hiring expertise, a deep understanding of engineering roles (e.g., software development, data engineering, cybersecurity, cloud, etc.), and a passion for healthcare innovation. In this role, you will lead strategic hiring efforts, partner closely with business leaders to attract top-tier engineering talent aligned with the evolving needs of the healthcare tech landscape. You will play a key role in scaling high-performing technology teams that are building transformative healthcare solutions. What you’ll do: Recruitment: Lead and own the end-to-end recruitment lifecycle from sourcing, screening, and interviewing to offer negotiation and onboarding for critical IT roles across domains like software engineering, data science, cloud, DevOps, cybersecurity, and healthcare IT systems. Assess candidates for both technical and cultural fit; conduct initial technical interviews when appropriate. Drive diversity, equity, and inclusion (DEI) in hiring initiatives. Stakeholder Management: Serve as the primary point of contact for senior-level hiring, workforce planning, and executive recruitment. Partner with business, HRBPs, and leadership to forecast talent needs and develop proactive sourcing strategies. Collaborate with global TA and business teams on talent pipeline planning and role prioritization. Sourcing & Market Intelligence: Design and execute advanced sourcing strategies across platforms, leveraging multiple sourcing channels including job boards, LinkedIn, employee referrals, campus outreach, and recruitment agencies. Employer branding: Represent the company at tech events, career fairs, and employer branding campaigns to attract top talent. Conduct regular market mapping and talent intelligence repor

PythonJavaAWSAI
S
📍 Bengaluru, KARNATAKA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Job Title: Project Manager – Data & Analytics Location: Bangalore Role Overview We are looking for an experienced Project Manager to drive Agile delivery for large-scale data and analytics programs. The role requires strong expertise in Scrum, SAFe, and end-to-end delivery management, with a focus on project execution, stakeholder management, and delivering business outcomes. Key Responsibilities Drive end-to-end project/program delivery across data engineering, BI, and analytics initiatives. Act as Scrum Master and facilitate Agile ceremonies, including Sprint Planning, Daily Stand-ups, Reviews, and Retrospectives. Manage scope, timelines, risks, dependencies, and delivery quality. Lead stakeholder discussions and provide regular updates to senior leadership. Drive continuous improvement and proactively resolve project challenges. Prepare impactful PPTs, dashboards, and executive-level summaries. Work independently and ensure projects are delivered within timelines and quality standards. Required Skills & Experience 8–10+ years of progressive experience in Project/Program Management, with the majority of the career in these roles. Prior experience as a Scrum Master. Strong expertise in Agile/Scrum methodologies; SAFe experience is preferred. Excellent communication and stakeholder management skills, including with senior audiences. Strong ownership of project execution, timelines, and quality standards. Proactive problem-solving approach with the ability to take initiative and work independently. Strong presentation skills with the ability to create impactful PPTs and summarize content for senior-level discussions. Working knowledge of Databricks, Snowflake, AWS S3, and AWS Athena. Experience working in data engineering, analytics, or technology environments is preferred. Note: By submitting your application, you consent to being contacted by our Talent Acquisition team via phone call, email, SMS, WhatsApp, or other communication channels regarding your app

AWSAgileScrumAI
JT
📍 India
✓ Quality checkedCompany trend -100%

From ₹18L/yr

Dr Ashwani Kumar Vij have been retained to Select one Dean - Academics / Education by a Uttar Pradesh ( UP ) based Engineering, Technology and Management Institute in August, 2026 Person should have an Engineering / MBA background from IIT / IIM or Top 25 Premium University / Institute, along with PhD and possess experience in an academic, administrative and leadership role. He should have a minimum of 15 years of teaching, research and consultancy experience with at least 5/10 years at Dean / HOD in Higher Education Sector University / Institute / College. In the exceptional case, Top / Director / GM / Senior Management level Professionals from the Corporate Sector / Public sector undertaking can be considered. He should have published research papers at National and International levels. He should be able to lead and meet the Institute programme objectives. Preferred Age desired is between 45 to 55 Years. Those Drawing CTC of Less than Rs 18 LPA are less Likely to be considered. Higher Salary Package would be offered to the experienced Higher Education Sector Professional having IIT / IIM / Top 10 MBA / Engineering / Technology Education Background. Those who have applied earlier need not apply again. Excellent Salary Package would be offered to the right candidate. DR ASHWANI KUMAR VIJ : President : Sales Recruitment Delhi/NCR and Metro Based Candidates : Please Courier / Speed Post Hardcopy of your : Complete / Structured CV ( Preferably Brief Resume of 2-3 Pages ), with Hard Copy of Visiting Card, Passport size Photograph, Supporting Documents, 2 Professional References, Salary Drawn / Salary Slip and Salary Expected details and address to: Mrs Savita Vij - Manager And also Email Your : Complete / Structured CV ( Preferably Brief Resume of 2-3 Pages ), with Scanned copy of Visiting Card, Passport size Photograph, Supporting Documents, 2 Professional References, Salary Drawn / Salary Slip and Salary Expected details at :

Recruitment
PE
📍 India· Full-time
✓ Quality checked

About Team : The Internal Audit team at Paytm comprises seasoned professionals with diverse skill sets and experience across different verticals like process audits, technology audits and forensics. The team focuses on implementing the approved audit plan, ensuring delivery of qualitative audits and conducting internal / special reviews while leveraging technology & data analytics and gauging key risks across business processes. Manager – Internal Audit (6–8 Years Exp) Role Overview: As a Manager, you will be a strategic partner to the business. You won’t just be "checking boxes"; you’ll be designing the audit framework for complex ecosystems like UPI, Wallet, and Lending. You will lead teams and interface directly with Senior Leadership. Key Responsibilities ●Audit Strategy: Develop and execute an annual risk-based audit plan focusing on high-growth business units. ● Stakeholder Management: Act as the primary point of contact for Department Heads, explaining risk implications and negotiating remediation plans. ● Process Engineering: Evaluate the effectiveness of internal controls ($ICFR$) and suggest automation to move from manual testing to continuous monitoring. ● Team Leadership: Supervise 2–3 Assistant Managers/Associates, ensuring high-quality documentation and timely delivery. ● Special Investigations: Lead / perform ad-hoc investigations or fraud risk assessments as required. ● Execution: Conduct end-to-end audits of various functions (Operations, HR, Finance, or Tech) under the guidance of the Manager. ● Fieldwork: Perform walkthroughs, test of effectiveness ($ToE$), and document audit findings with clear "Cause and Effect" analysis. ● Cross-Functional Collaboration: Partner with cross-functional teams to gain deep insights into critical business processes, identify risks, and evaluate the adequacy of internal controls. ● Reporting: Prepare and present clear, concise audit reports to senior management and stakeh

SQLGitAIGo
F
📍 Bengaluru, KARNATAKA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Who are we? FalconX is a pioneering team of operators, investors, and builders committed to revolutionizing institutional access to the crypto markets. Operating at the intersection of traditional finance and cutting-edge technology, FalconX addresses the industry's foremost challenges: Navigating the digital asset market can be complex and fragmented, with limited products and services that support trading strategies, structures, and liquidity found in conventional financial markets. As a comprehensive solution for all digital asset strategies from start to scale, FalconX operates as the connective tissue empowering clients with seamless navigation through the ever- evolving cryptocurrency landscape. Role Overview As an Engineering Manager for the Credit team, you will lead the core engine driving our credit systems, including the margin notification engine and loan booking system. Your primary focus will be on maximizing data accuracy, system reliability, and architectural integrity. You will lead a high-performing group of senior engineers, stream-lining technical processes, preventing over-engineering, and maintaining a high standard of delivery alongside key stakeholders. Role Split & Focus Areas Technical Leadership & Architecture (40%): Drive long-term system maintenance, architectural refactoring, and code quality. Ensure technical debt is addressed without blocking impactful business features. Process & Stakeholder Management (20–30%): Partner with product and business stakeholders to prioritize deliverables. Make decisive trade-offs and cut unnecessary overhead to maintain lean execution. People Management & Mentorship (20–30%): Manage, mentor, and guide senior technical talent across performance cycles, career growth, and talent acquisition. Key Responsibilities Oversee and maintain the reliability, precision, and efficiency of core credit applications (Margin Notification Engine, Loan Booking Systems). Drive architectural roadmap decision

PythonAWSGitRest
P
📍 India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About Us Pearl is AI for professional services at global scale, combining advanced AI with verified human expertise to deliver help that is accurate, accountable, and fast. Since 2003, our network has connected millions of customers with licensed professionals across 196 countries, making real expertise available anytime, anywhere. Our Values Data driven: Start with truth, measure what matters. Courageous: Bias to action; run toward hard problems. Innovative: Seek novel, elegant solutions. Lean: Do more with less. Build, ship, learn fast. Humble: Strong opinions, lightly held. About the Role We’re looking for a Group Product Manager, Customer Acquisition to lead a portfolio of high-impact B2C growth products and develop a team of Product Managers. This leader will partner across Product, Engineering, Marketing, Design, and Analytics to identify meaningful customer and business opportunities, translate them into clear strategies, and deliver measurable improvements in growth, conversion, profitability, and customer lifetime value. This is a hands-on leadership role for someone who can operate at multiple altitudes: setting portfolio direction, coaching PMs, shaping senior-level decisions, and helping teams move quickly from insight to measurable learning. What You’ll Do Set a clear product strategy and portfolio of opportunities across multiple product teams. Lead and mentor 4 Product Managers (2 PMs and 2 SPMs), including hiring, coaching, performance management, career development, and raising the quality of product thinking. Coach PMs on customer discovery, strategic thinking, experimentation, execution, and business impact. Partner with functional leaders to set priorities, allocate resources, and resolve cross-team dependencies and trade-offs. Improve the quality and speed of product decision-making, helping teams move from customer insights to validated learning efficiently. Enable PMs to operate with autonomy while maintaining alignment and accountabilit

AIMarketingLeanHR
E
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Pure Solutions team as a Senior MLOps Solutions Engineer to architect and build high-scale, enterprise-grade AI/ML solutions. You will be instrumental in integrating Pure Storage platforms with the evolving open-source MLOps ecosystem (Kubeflow, MLflow, Ray) to operationalize the complete machine learning lifecycle. This role requires a creative technologist with deep Python expertise to drive innovation and enable our customers and partners to achieve production AI success. WHAT YOU'LL DO Design and Automate MLOps Pipelines: Lead the development of end-to-end MLOps workflows using CI/CD tools (Git/Jenkins) and orchestration platforms (MLflow/Kubeflow), specifically integrating Pure Storage's FlashBlade, FlashArray, and Portworx as the high-performance data plane for data ingestion, training, and inference. Build High-Performance AI/ML Reference Architectures: Create validated, repeatable deployment models using Infrastructure as Code (e.g., Ansible, Terraform) for AI/ML environments spanning bare metal, virtual machines, and GPU-accelerated Kubernetes clusters, ensuring optimal performance for distributed training. Optimize and Operationalize GPU Inference: Architect and implement solutions for high-throughput, low-latency model serving, utilizing technologies like NVIDIA Triton Inference Server and advanced optimization techniques (quantization, model sharding like DeepSpeed/Megatron-LM, and dynamic bat

PythonAWSKubernetesCI/CD
🔔

Get new lead senior staff technical program manager 2c site reliability engineering jobs in India by email

Daily job updates · Unsubscribe anytime