OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full stack automation fabric solutions for mission-critical business processes. With the first SaaS-based composable automation platform specifically built for ERP, we believe in the transformative power of automation. Our unparalleled solutions empower you to orchestrate, manage and monitor your workflows across any application, service or server — in the cloud or on premises — with confidence and control. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for a Software Engineer, Products & Platforms to join our Product engineering team to contribute to technical development across Redwood’s Workload Automation Platform, helping modernize and build reliable capabilities across our engineering teams. You will be instrumental in the design, development, and enhancement of our platform, building high-quality, scalable, and secure software that powers enterprise data exchange for more than 1,000 customers worldwide. As a Software Engineer, you will: Technical Delivery & Architecture: Solve complex engineering problems, follow established best practices, participate in design reviews, and collaborate with peers, while contributing to the development, security, compliance, and observability of our Java/Spring Boot microservices. Platform and Infrastructure Implementation: Develop a solid understanding of our core platform and infrastructure, contributing to its resilience, component communication, and scalability. Cross-Team Collaboration: Facilitate and drive cross-team collaboration with Product, QA, and other engineering groups to ensure end-to-end alignment and successful product delivery. AI Integration: Research and apply AI/ML co
Jobs in India
Workload Porting And Performance Engineer in India
84 active opportunities · Updated October 2026
Showing
15 jobs
Explore current workload porting and performance engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
Position Overview: A Junior Research Analyst within the Diligent Market Intelligence function supports subject matter expertise in specialized research areas such as compensation, governance, activism, voting, and risk. Their main responsibilities include assisting in the design and management of data collection methodologies, overseeing third-party service provider workflows, and ensuring high-quality, timely data delivery. They are also expected to execute manual data collection from global sources, maintain strong quality assurance practices, and collaborate closely with Product and Engineering teams to ensure operational efficiency and accuracy. Additionally, Junior Research Analysts take ownership of specific daily and weekly tasks, contribute to ongoing projects, and play a key role in training service providers while keeping process documentation clear and up to date. The role requires strong numeracy, attention to detail, independent work habits, and a proactive approach to continuously improving research products and processes. Key Responsibilities: Proactive communication and management of outsourcer workload. Conduct data collection tasks to uphold the integrity of our database and keep it current. Execute systematic checks to identify data inaccuracies and operational inefficiencies as per predefined protocols. Provide regular updates and reports to the Vertical lead on progress and challenges. Take ownership of specific daily and weekly tasks, as well as ongoing projects within the assigned research vertically. Required Experience/Skills: Entry level Analytical approach and problem-solving attitude Strong written and verbal communication skills Ability to manage deadlines Ability to adapt to difficult workload demands, e.g., time/resource constraints Proficiency in Microsoft Office, especially Excel About Us Diligent is the AI leader in governance, risk and compliance (GRC) SaaS solutions, helping more than 1 million users and 70
Identity has become a critical to an organizations security posture. This is true across their entire digital landscape include cloud, traditional on prem and emerging technologies like agentic ai. Britive is at the forefront of a modern approach to delivering identity security with the only modern privileged access management platform that provides unified Privileged Access Visibility, Dynamic Privilege Management and Secrets Governance across infrastructure, platforms & SaaS. Our patent-pending technology is deployed at companies of all sizes around the world, including Fortune 500 companies. We have repeatedly ranked among the hottest Cloud Security startups. Britive is founded by CyberSecurity industry veterans with a successful prior exit and is backed by top-tier VCs. About You Lead team(s) that operate in Kanban, Scrum or Scrumban model. Build positive and collaborative relationships globally with key stakeholders Monitor and oversee daily activities of the delivery teams, removing impediments and providing guidance as needed. Perform resource and workload allocations as required to achieve delivery timelines with quality. Reporting team productivity and output. Managing team dynamics and create an inclusive culture of innovation while working across product lines.. Your Impact Key Responsibilities: We are looking for a dynamic and experienced engineering leader with Agile exposure. You have prior experience delivering SaaS, PaaS, IaaS products / services. You come from a technical software development background and are able to effectively manage features and roadmap execution. You are an engineering leader but come from strong individual contributor background as a technical lead or architect. You are experienced at building highly available services in a cloud environment. You are proficient in programming (Java) , Microservices and cloud technologies. You have experience with production operations and good
Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida/ Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite
Title: Senior Site Reliability Engineer - I, Product Area Focus Location: Noida (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering teams. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedit
Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida / Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite
About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale. If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk. Responsibilities Design and build fleet automation systems that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention. Build AI Infrastructure Agents that automate deployment, root-cause failures, incident triage, and autonomous remediation. Develop Fleet Intelligence platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers. Build software that maximizes GPU availability, utilization, performance, and reliability across thousands of accelerators. Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads. Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations. Continuously improve deployment velocity, reliability, and operational efficiency through automation. Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure. Requirements 3+ years building distributed systems, infrastructure platforms, or large-scale backend software. Strong software engineering skills in Python, Go, or Rust . Experience building platforms, automation systems, or developer infrastructure. Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies. Strong systems thinking with the ability to understand problems across hardw
JOB TITLE Cloud Compute Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU'LL DO Design, build, and operate Kubernetes clusters on Amazon EKS, including cluster lifecycle management, networking, autoscaling, and workload scheduling. Manage and optimize EC2-based compute infrastructure, including instance selection, placement strategies, capacity planning, and utilization analysis. Operate and improve ECS-based services where applicable, ensuring consistency across our container runtime environments. Develop and maintain Infrastructure as Code (IaC) using Terraform to provision and manage compute resources at scale. Collaborate with development and platform teams to define compute patterns, containerization standards, and deployment best practices. Monitor compute environments for availability, performance, and cost, driving continuous optimization across the fleet. Contribute to architectural decisions around workload placement, multi-tenancy, OS image and container li
Opportunity Overview: We are seeking a Team leader, Intake to join our Intake Operations. In this role, as Team leader of Intake, you will work closely with the leadership team at Cohere and report to the Sr Director of India Operations. You will be responsible for providing supervision for the lead intake specialist and intake specialist team to meet or exceed operational objectives and metrics. You will leverage both your creative skills and communication skills to promote a high performing intake team. The Team leader of Intake will be highly organized in order to plan daily operational activities and provide oversight of the intake team. You will use your professionalism, personality, and communication skills to inspire the team to meet or exceed all performance standards. At a growing organization, this is a position that offers the ability to make a substantive mark on the company and its partners with exponential growth opportunities. What you’ll do: Gain a deep understanding of Cohere’s product and our health plan partners Provide daily operational direction to the intake staff. This includes interviewing new hires, training, coaching, mentoring, quality auditing, implementation and oversight of quality improvement plans identified based on trends and other process improvements Coordinates and provides day-to-day oversight of the intake staff Manage workload balancing needs of the intake team Assist in addressing case escalations and provider issues HR management to include performance evaluations, 1:1’s with the lead intake specialist and intake staff, payroll Other duties as assigned ISMS roles and responsibilities Good knowledge of Information practices. Assist the manager in all the information security activities implementation and maintenance process. Ensuring the team and imparted with Competence related to Information security Responsible for implementation of security policies and procedures and report any issues
Overview The Insights team delivers high-quality research, market intelligence, and data-driven analysis to support clients’ strategic and investment decisions. Working across industries-including healthcare and medtech-the team combines primary and secondary research with deep analytical capabilities to generate actionable insights. Team members collaborate closely with global stakeholders to translate complex information into clear, impactful outputs. The Project Manager – Insights is responsible for overseeing the end-to-end delivery of research projects, ensuring timelines, quality standards, and client expectations are consistently met. This role coordinates across internal teams and stakeholders to manage multiple concurrent projects, streamline workflows, and drive execution. The ideal candidate combines strong organizational and communication skills with an ability to translate complex project requirements into efficient delivery. This is a hybrid position out of our Mumbai OR Pune office What You’ll Do : Manage end-to-end delivery of research and insights projects, ensuring timelines and quality standards are met Coordinate across analysts, researchers, and stakeholders to execute project plans Define project scope, objectives, and deliverables in collaboration with internal teams and clients Track project progress, manage risks, and proactively address bottlenecks Oversee preparation and delivery of client-ready reports, presentations, and outputs Ensure consistency, accuracy, and quality of insights and deliverables Optimize workflows, processes, and tools to improve efficiency and scalability Support resource allocation and prioritize workload across multiple concurrent projects Maintain clear and consistent communication with stakeholders on timelines and outcomes What You Have : Bachelor’s degree from an accredited college/university (e.g., Business, Economics, Life Sciences, or related field) +4 years of experience in project management, consulting, r
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from conception to completion, managing timelines, and technical dependencies across teams. Mentor engineers across teams, fostering a culture of reliability, automation, and continuous learning. Collaborate with security and compliance partners to ensure infrastructure adheres to best practices and standards (e.g., IAM Federation, Workload Identity). Participate in t
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from conception to completion, managing timelines, and technical dependencies across teams. Mentor engineers across teams, fostering a culture of reliability, automation, and continuous learning. Collaborate with security and compliance partners to ensure infrastructure adheres to best practices and standards (e.g., IAM Federation, Workload Identity). Participate in t
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from conception to completion, managing timelines, and technical dependencies across teams. Mentor engineers across teams, fostering a culture of reliability, automation, and continuous learning. Collaborate with security and compliance partners to ensure infrastructure adheres to best practices and standards (e.g., IAM Federation, Workload Identity). Participate in t
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are looking for people who have a strong background in data science and cloud architecture to join our Professional Services team to help create exciting new offerings and capabilities for our customers! This team will be working with customers using Snowflake to expand their use of the Snowflake Data Cloud to bring data science pipelines from ideation to deployment, and beyond using Snowflake's features and its extensive partner ecosystem. The role will be more strategic, advising our clients on best practices and advice to implement Data Science workloads on Snowflake. You will be designing solutions based on requirements and coordinating with customer teams, and where needed Systems Integrators, while maintaining oversight and direction to ensure successful outcomes AS A TECHNICAL ARCHITECT - AI/ML AT SNOWFLAKE, YOU WILL : Be a technical expert on all aspects of Snowflake in relation to the AI/ML workload Provide customers with best practices and advise as it relates to Data Science workloads on Snowflake Build, deploy and ML pipelines using Snowflake features and/or Snowflake ecosystem partner tools based on customer requirements Work hands-on where needed using SQL, Python, to build POCs that demonstrate implementation techniques and best practices on Snowflake tech
- Proven experience deploying and managing Kubernetes clusters for AI/ML workloads. Experience of at scale deployments with Azure Kubernetes. Experience level - 5 Years or more Positions - 2 Proven experience deploying and managing Kubernetes clusters for AI/ML workloads. - Experience of at scale deployments with Azure Kubernetes Service, RedHat OpenShift, Microk8s and Helm Charts. - Expertise with infrastructure and resource management and virtualization tools such as VMWare/EXSi, KVM, Ansible, Redfish. - Strong understanding of Run:AI platform, including job scheduling, quota management, and GPU virtualization. - Knowledge of NVIDIA AI Enterprise components including, NIM, NeMO, TAO, Triton and Nucleus Servers - Familiarity with DGX systems, Jetson, and NVIDIA’s AI Factory components. - Proficiency in Python, C++, and optionally .NET/C# for enterprise integration.
Other cities to consider
More places hiring for this role
Get new workload porting and performance engineer jobs in India by email
Daily job updates · Unsubscribe anytime