Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput, and 99.999 availability. We're looking for a technical leader to help us to continue to scale the service with great people and reliable, cost-effective and efficient infrastructure, processes and tooling. As the Director of Site Reliability Engineering you will oversee the SRE organization focused on Okta platform, Databases, Edge networking, K8s platform, CI/CD, Observability, FinOps, and automation platform & tooling. Job Duties and Responsibilities: Build and lead a high-caliber India-based SRE organization supporting Okta’s production fleet. Partner with global engineering, product, and infrastructure leaders to deliver resilient, scalable, and secure services. Define and execute the India SRE strategy in alignment with global reliability goals. Lead post-incident reviews, drive root-cause analysis, and ensure long-term corrective actions. Participate in incident management, on-call rotations, and blameless RCAs. Implement automation and observability to reduce manual toil and improve operational efficiency. Drive adoption of modern infrastructure practices: infrastructure as code (Terraform), container orchestration (Kubernetes), and AI within Infrastructure org. H
Jobs in India
Production Planner in Bengaluru
113 active opportunities · Updated October 2026
Showing
15 jobs
Explore current production planner jobs in Bengaluru. Filter by work mode, employment type, experience, department, date posted and distance.
Senior Machine Learning Engineer Description - We are looking for a Senior MLOps Engineer to design, build, and operate the infrastructure that enables machine learning models and large language models to be deployed safely, reliably, and at scale. In this role, you will create the end-to-end capabilities required to move models from experimentation into production, expose them through secure and highly available endpoints, and enable users and applications to interact with AI-powered services. You will work across AWS and Databricks to establish robust CI/CD pipelines, model-serving infrastructure, observability, governance, rollback mechanisms, and operational standards. You will partner closely with data scientists, machine learning engineers, software engineers, security teams, and platform engineers. The ideal candidate combines strong cloud and DevOps engineering skills with a practical understanding of machine learning systems, LLM deployment patterns, and production reliability. Key Responsibilities MLOps Platform and Architecture Design and implement a scalable MLOps platform using AWS and Databricks. Define reference architectures and reusable deployment patterns for traditional machine learning models, deep learning models, and large language models. Build standardized workflows that move models from development and validation into staging and production. Develop self-service capabilities that allow data scientists and ML engineers to deploy models without manually managing infrastructure. Establish clear separation between development, testing, staging, and production environments. Design multi-region or multi-availability-zone architectures where required by business continuity and availability objectives. CI/CD and
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Job Description/ Responsibilities: Designing, developing and maintaining stable and reliable AI/ML Ops platforms / pipelines Minimum experience of 4-6 Years required in AI ML Ops Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time. Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with AWS Bedrock is preferable Resource Optimization: Manage GPU/CPU utilization to minimize cloud costs while maintaining low-latency inference for users Collaboration: Work closely with data scientists, data engineers, and software engineers to bridge the gap between model development and production. Version Control & Governance: Manage versioning for data, code, and models using tools like MLflow. Security & Compliance: Implementing data security measures, ensuring compliance with data governance
About Bolna Bolna is a YC-backed voice AI orchestration platform built for the Indian market—powering multilingual, vernacular voice agents across Hindi, Hinglish, Tamil, and 10+ languages at sub-500ms latency across collections, recruitment, sales, and e-commerce use cases. We are an orchestration layer, not a model company: our moat is outcome-labelled vernacular data, rigorous evaluation infrastructure, and a growing taxonomy of how Indian enterprise voice AI fails in production. Why This Role Exists Product decisions at Bolna increasingly hinge on rigorous, code-mixed-aware data analysis—and not just one kind. On one side, there is model and evaluation rigor: LLM benchmarking for post-call intelligence, ASR/WER evaluation, inter-rater reliability on human-labelled calls, and routing and latency economics. On the other, there is product and growth insight: understanding where self-serve users drop off in their journey, what patterns emerge across lakhs of monthly calls, and which use cases and configurations are actually working. Both currently sit with the Head of Product alongside strategy and roadmap ownership. We need a dedicated analyst to own the execution and recurring cadence across both-freeing product leadership to act on findings rather than produce them. What You’ll Do Model and Evaluation Analysis LLM and model benchmarking: Run structured comparisons across model providers such as Sarvam, DeepSeek, Gemini, and Claude variants for tasks including post-call extraction and LLM-as-judge scoring. Evaluate cost, accuracy, fill rate, and TTR, with particular attention to Hinglish and code-mixed content. Evaluation infrastructure: Build and maintain LLM-as-judge pipelines using tools such as DeepEval, design and track evaluation metrics, and run inter-rater reliability analysis such as Krippendorff’s alpha across human call reviewers. Golden dataset creation: Support the construction of golden datasets for ASR and transcript labelling, including flagging co
Database Migration Engineer JD We are looking for a hands-on Database Migration Engineer with experience in large-scale enterprise database migrations and Migration Factory environments. Key Responsibilities Execute migrations from Sybase ASE / IBM DB2 to PostgreSQL / Amazon Aurora PostgreSQL / SingleStore . Perform migration assessment, schema conversion, data validation, reconciliation, and performance testing . Execute database cutover, rollback, production go-live, and post-migration stabilization . Troubleshoot migration issues and independently execute migration runbooks across migration waves. Coordinate with application, infrastructure, cloud, and engineering teams . Perform database administration, performance tuning, backup/recovery, and optimization . Develop migration documentation, runbooks, status reports, and operational procedures. Ensure data integrity, availability, security, and compliance throughout the migration lifecycle. Must-Have Skills Sybase ASE + IBM DB2 + PostgreSQL Amazon Aurora PostgreSQL + AWS Hands-on database migration experience Database administration & performance tuning Shell / Python / PowerShell scripting and automation Strong troubleshooting and analytical skills Ability to independently manage end-to-end database migration activities Note: By submitting your application, you consent to being contacted by our Talent Acquisition team via phone call, email, SMS, WhatsApp, or other communication channels regarding your application and relevant career opportunities.
Position Overview: We are looking for a Senior Software Engineer to drive technical excellence, architect complex systems, and elevate our engineering team. You will own critical technical decisions, lead major initiatives from conception to delivery, and set the standard for engineering quality across our products. As a senior engineer, you will architect and lead the development of sophisticated AI-enabled features and infrastructure. This includes designing MCP server architectures, building advanced RAG systems, implementing agentic AI workflows, and establishing patterns that scale across our product portfolio. You will combine deep technical expertise in both traditional software engineering and AI/ML to deliver production-grade solutions. What You'll Do Lead the design and implementation of AI integration infrastructure (MCP servers, orchestration layers, API gateways) Build sophisticated AI features including advanced RAG systems, agentic workflows, and multi-step reasoning Establish AI engineering best practices, security patterns, and quality standards Lead technical initiatives from requirements through production deployment Make critical architectural decisions balancing performance, scalability, cost, and maintainability Design AI evaluation frameworks and implement quality benchmarks Debug and resolve complex production issues across traditional and AI systems Required Qualifications Tech Stack Core: Node.js, React, TypeScript, AWS, PostgreSQL, MSSQL, Docker AI & Integration: Python, MCP, AWS Bedrock, LangGraph/Semantic Kernel, Vector Databases, RAG Core Technical Skills 5+ years professional development with proven track record of delivering complex systems Strong Node.js and JavaScript/TypeScript expertise Advanced React and frontend architecture skills Extensive AWS architecture experience Expert PostgreSQL database design, optimization, and performance tuning Deep understanding of microservices, distributed systems, and
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Portworx team and play a pivotal role in delivering our highest-quality suite of products. In this role, you will contribute to building clean, robust code with a relentless focus on quality and customer success. You will be instrumental in developing a new SaaS platform designed to provide a secure, consistent, and best-in-class experience for customers purchasing and utilizing various Portworx offerings. As a core developer, you will take ownership of designing and implementing innovative features and products across the Portworx portfolio. WHAT YOU’LL DO Design & Develop SaaS Microservices: Build, scale, and integrate new features into Portworx products, ensuring seamless alignment with our distributed system architecture. Own the Lifecycle: Lead end-to-end engineering efforts, including design reviews, code reviews, unit/functional testing, documentation, and managing CI/CD deployment pipelines. Collaborate Globally: Partner closely with cross-functional peers and stakeholders (from Product Management to early-adopter customers) to take solutions from initial concept to production. Drive Iterative Quality: Take full ownership of feature development by proactively adapting to customer feedback and resolving issues identified during testing and deployment. Innovate & Experiment: Research and experiment with emerging technologies to push boundaries, optimize cloud infrastructure, and pione
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE We are looking for a highly skilled Senior Frontend Engineer to join the Portworx UI team, responsible for building intuitive, scalable user experiences for our Kubernetes-based platform. You will build production-quality web applications using React, collaborate closely with UX, Product, backend teams, and engineering leadership, and help shape frontend architecture and engineering practices. This role requires strong ownership, sound technical judgment, and the ability to solve complex problems independently while helping other engineers grow. WHAT YOU'LL DO We are primarily an in-office environment and therefore, you will be expected to work from the {{OFFICE_LOCATION}} office in compliance with Everpure's policies, unless you are on PTO, or work travel, or other approved leave. Design, develop, and maintain scalable, high-performance single-page applications using React and TypeScript. Own features end to end—from requirements and design through implementation, testing, release, and production support. Create reusable, accessible, responsive UI components using modern HTML, CSS, React patterns, and design systems. Collaborate with UX, Product Management, and backend teams to turn customer and product needs into polished, production-ready features. Contribute to frontend architecture, API integration, performance, maintainability, and engineering best practices. Write and maintain unit, integration, and end-to-
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Hardware Manufacturing Test Engineer at Everpure, you will spearhead manufacturing test enablement for our next-generation enterprise hardware platforms from New Product Introduction (NPI) through high-volume production ramp. Operating with high autonomy, you will design robust test strategies and collaborate cross-functionally with Firmware, Diagnostics, and global factory teams to deliver market-ready, high-yield hardware infrastructure. Your expertise will directly optimize manufacturing throughput, improve release quality, and accelerate the global deployment of the industry-defining Everpure Platform. WHAT YOU'LL DO Own NPI & Sustaining Test Readiness: Drive end-to-end manufacturing test strategy, test sequence definition, and coverage validation for multiple hardware infrastructure variants to ensure seamless factory execution and build readiness. Lead Debug & Yield Optimization: Independently lead technical triage and resolve complex system-level test fallout during critical build phases, analyzing failure paretos to isolate root causes across hardware, firmware, and test infrastructure. Scale Test Architecture & Frameworks: Enhance and standardize the Design for Manufacturing (DFM) hardware test architecture by deploying reusable automation frameworks and consolidating test execution into shared platforms to minimize factory cycle times. Drive Vendor & Cross-Functional Alignment: Colla
Senior Software Engineer (Backend) About Team Cloud Platform Engineering group designs, builds and manages platforms that allow Myntra’s tech product to be secured, reliable, deployed and run at scale. These horizontal platforms leverage cloud hosting platforms and ensure all Myntra’s hostings are agnostic to Cloud vendor. We also build a number of production automation like provisioning of infrastructure at scale and manage complex access management on servers. We have developed numerous in-house tools and platform for load/stress testing, zero trust, security and compliance, CI/CD, Observability at scale. And, we aggressively adopt from open sources and try to contribute back to the community. Tools and Platform Engineering This team builds and maintains centralized and high-scale platforms for Observability (centralized log collection, metric systems, monitoring systems), Security & Compliance (access management, secret management, database access, change management systems, Authentication and Authorization of services etc). The platforms developed by these teams are centralized tools used by all engineering teams for database access, and changes, infrastructure provisioning, on-call scheduling, onboarding new monitoring etc. The vault system and APIs, developed, deployed and managed by this team, is being used by all Myntra’s production services. This team consists of full-stack developers who are skilled in Python, Golang, ReactJs. Roles and Responsibilities Design, build and maintain central platform products to improve the security posture of Myntra Write maintainable, scalable, and efficient code. Design and architect technical solutions for the developer community at Myntra Work in a cross-functional team, collaborating with peers during the entire SDLC. Follow coding standards, code reviews, etc. Follow scrum sprint cycles and commitment to deadlines. Identify security gaps in or for software platforms and incorporate them into requirements
Partner Consultant (Photographer) Roles and Responsibilities Work with the creative producer on all styling and shoot requirements across catalogue , campaign, and social media shoots Program manage shoot calendar and relevant studios Collaborate with third party vendors to source styling props at the best possible rate Deliver output (still and video) following style guides and service level agreements Maintain and consistently update the casting database for models across multiple agencies & solve for fit challenges using knowledge of model pool Single threaded owner of all communication with production and onset crew to ensure productivity is efficient and MOQ is met. Highlight and escalate wherever required. Adapt to growing business needs and pro-actively accommodate ad-hoc requests Qualifications & Experience Graduate with Minimum 5-7 years of experience Good communication skills Stakeholder management
Key Responsibilities : Primary responsibilities :- Installation and configuration of MySQL instances on single or multiple ports. Hands-on experience of working with MysQL 5.7 and MySQL 8. Clear understanding of MysQL Replication process flows , threads , setting up multi node clusters and basic troubleshooting. Understanding of at least one of the backup and recovery methods for MySQL . Strong fundamentals of SQL and able to understand and tune complex SQL queries when needed. Strong fundamentals on the linux system side and monitoring tools like top , iostats , sar etc. At Least couple of years of production hands on experience on medium to big sized MySQL databases. Setting up and maintaining users and privileges management system and troubleshooting relevant access issues. Some exposure to external tools like Percona , ProxySQL , HAP etc. Understand the transaction flows and ACID compliance. Basic understanding of networking concepts . Performing on-call support and should be able to provide the first level support . Excellent verbal and written communication skills. Strong shell scripting skills . Good to have Python . Secondary responsibilities. :- Able to configure and setup NOSQL databases like Mongodb and Cassandra. Ability to learn new technologies along with a team and a positive outlook to understand problems from the business point of view. Qualifications: Proficiency in database management systems such as , MySQL or NoSQL databases. SQL programming and database design skills. Knowledge of database performance tuning and optimization techniques. Familiarity with database security best practices. Scripting and automation skills (Good to have- Python). Good problem-solving and analytical skills. Excellent communication and teamwork skills.
Key Responsibilities : Primary responsibilities :- Installation and configuration of MySQL instances on single or multiple ports. Hands-on experience of working with MysQL 5.7 and MySQL 8. Clear understanding of MysQL Replication process flows , threads , setting up multi node clusters and basic troubleshooting. Understanding of at least one of the backup and recovery methods for MySQL . Strong fundamentals of SQL and able to understand and tune complex SQL queries when needed. Strong fundamentals on the linux system side and monitoring tools like top , iostats , sar etc. At Least couple of years of production hands on experience on medium to big sized MySQL databases. Setting up and maintaining users and privileges management system and troubleshooting relevant access issues. Some exposure to external tools like Percona , ProxySQL , HAP etc. Understand the transaction flows and ACID compliance. Basic understanding of networking concepts . Performing on-call support and should be able to provide the first level support . Excellent verbal and written communication skills. Strong shell scripting skills . Good to have Python . Secondary responsibilities. :- Able to configure and setup NOSQL databases like Mongodb and Cassandra. Ability to learn new technologies along with a team and a positive outlook to understand problems from the business point of view. Qualifications: Bachelor's degree in Computer Science, Information Technology, or a related field (or equivalent experience). Proficiency in database management systems such as , MySQL or NoSQL databases. SQL programming and database design skills. Knowledge of database performance tuning and optimization techniques. Familiarity with database security best practices. Scripting and automation skills (Good to have- Python). Good problem-solving and analytical skills. Excellent communication and teamwork ski
Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You
AI/ML – Investment Services A Career with Point72's AI/ML – Investment Services Team The AI/ML – Investment Services team at Point72 spearheads the development of cutting-edge AI solutions that seek to transform our business processes and enhance enterprise intelligence. The team aims to bridge the gap between business challenges and technological innovation, collaborating with stakeholders across the firm and leveraging expertise in generative AI, data engineering, and machine learning. WHAT YOU'LL DO Build and scale core backend services and platforms that power generative AI applications and data infrastructure used across the firm’s investment workflows Design and implement high-throughput, low-latency data pipelines to ingest, normalize, and serve both structured and unstructured data Develop robust APIs and microservices to support model inference, feature serving, and downstream applications Integrate generative AI tools and model-serving workflows into production, including embedding stores, retrieval components, and fine-tuning pipelines Optimize system performance, cost, and reliability through profiling, capacity planning, and architectural improvements Implement automated testing, continuous delivery pipelines, monitoring, and incident response practices to maintain production health Partner with data scientists, AI engineers, product owners, and operations to translate models and prototypes into scalable, production-grade solutions Mentor engineers, lead code reviews, and establish engineering best practices for maintainability, security, and observability Own end-to-end delivery, operational runbooks, and metrics-driven measurement of feature impact and system reliability WHAT'S REQUIRED Bachelor’s degree in computer science, software engineering, or a related technical field Minimum 5+ years of professional experience building backend systems and production services Demonstrated experience designing and operating large-scale data engineering pipelines
Other cities to consider
More places hiring for this role
Get new production planner jobs in Bengaluru, India by email
Daily job updates · Unsubscribe anytime