For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Job Description/ Responsibilities: Designing, developing and maintaining stable and reliable AI/ML Ops platforms / pipelines Minimum experience of 4-6 Years required in AI ML Ops Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time. Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with AWS Bedrock is preferable Resource Optimization: Manage GPU/CPU utilization to minimize cloud costs while maintaining low-latency inference for users Collaboration: Work closely with data scientists, data engineers, and software engineers to bridge the gap between model development and production. Version Control & Governance: Manage versioning for data, code, and models using tools like MLflow. Security & Compliance: Implementing data security measures, ensuring compliance with data governance
Jobs in India
Terraform in Bengaluru
40 active opportunities · Updated October 2026
Showing
15 jobs
Explore current terraform jobs in Bengaluru. Filter by work mode, employment type, experience, department, date posted and distance.
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Role We are seeking a seasoned Manager, Software Engineering with 12+ years of experience to lead our Database Engineering and Cloud Infrastructure team. In this role, you will lead a team of high-performing engineers responsible for architecting, scaling, and optimizing multi-cloud relational and in-memory database platforms. You will bridge technical execution, engineering leadership, and strategic infrastructure planning across AWS and Azure environments. Key Responsibilities Technical Leadership & Architecture Lead the architectural design and operations of enterprise-grade, multi-cloud relational databases across AWS (RDS PostgreSQL, MySQL, Aurora) and Azure (Database for PostgreSQL/MySQL, Azure SQL Managed Instance). Drive high-availability architecture strategies, including Multi-AZ deployments, auto-failover groups, read replica scaling, and cross-region disaster recovery (DR). Oversee zero-downtime operations, including major-version engine upgrades, schema migrations, and blue/green deployment strategies. In-Memory Infrastructure & Open-Source Strategy Manage scale operations for in-memory datastores (AWS ElastiCache, Azure Cache for Redis), focusing on cluster mode operations, eviction policies, and persistence tuning. Spearhead open-source caching initiatives and migration pathways from Redis to Valkey (e.g., AWS ElastiCache for Valkey) using zero-downtime tools like RedisShake to ensure open-source license compliance and optimize cloud spend. Aut
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Identity is the foundation of trust on the internet, and at Okta, we are building a world where anyone can safely use any technology. We are looking for a driven, curious, and empathetic product-minded engineering leader to expand Okta’s market footprint by growing our global integration ecosystem. As the Director of Engineering for the Okta Integration Ecosystem in Bangalore, you will lead, scale, and inspire a talented organization of engineering managers and engineers dedicated to evolving Okta-built integrations across our entire product portfolio. This is a high-impact, high-visibility role. If you are passionate about customer security, obsessed with developer experience, and love engaging with a vibrant technology community, we want to hear from you. What You’ll Do Lead and Scale the Team: Mentor, inspire, and grow a world-class engineering organization in Bangalore, directly managing engineering managers and senior individual contributors while fostering a high-performance culture. Drive Integration Strategy: Own the technical roadmap and execution for Okta’s integrations across a massive gamut of applications—spanning traditional Enterprise systems, SaaS, On-prem infrastructure, Active Directory (AD), Azure, and next-generation AI agents. Advance Modern Infra & Governance: Drive the development, rollout, and governance of the official Okta MCP (Model Context Protocol) Server. Streamline and enable partner integrations through robust Terraform P
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Technical Services Engineer for Portworx® by Everpure™, you will empower innovators and enterprise customers to seamlessly deploy and scale multi-cloud data services for Kubernetes and containerized applications. In this high-impact role, you will redefine the cloud-native storage experience by bridging deep technical expertise with direct customer engagement. You will serve as a technical expert across all deployment phases, partnering closely with customer engineers, account teams, and core R&D to solve complex cloud-native architecture challenges. WHAT YOU'LL DO Drive Enterprise Customer Success: Deliver end-to-end technical expertise across all stages of Portworx® deployment and production environments, ensuring rapid issue resolution, minimal downtime, and maximum platform reliability for key enterprise accounts. Perform Complex Problem Diagnosis & Triage: Analyze application workflows, system logs, and containerized workloads across public/private cloud stacks to isolate, reproduce, and resolve multi-layer technical issues alongside engineering teams. Enable Pre-Sales and Onboarding: Partner with field systems engineers to guide technical installation, integration, and proof-of-concept (POC) execution for prospective and onboarding enterprise clients. Scale Organizational Knowledge: Author and maintain high-impact knowledge base articles, technical documentation, and FAQ guides to build self-ser
About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale. If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk. Responsibilities Design and build fleet automation systems that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention. Build AI Infrastructure Agents that automate deployment, root-cause failures, incident triage, and autonomous remediation. Develop Fleet Intelligence platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers. Build software that maximizes GPU availability, utilization, performance, and reliability across thousands of accelerators. Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads. Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations. Continuously improve deployment velocity, reliability, and operational efficiency through automation. Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure. Requirements 3+ years building distributed systems, infrastructure platforms, or large-scale backend software. Strong software engineering skills in Python, Go, or Rust . Experience building platforms, automation systems, or developer infrastructure. Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies. Strong systems thinking with the ability to understand problems across hardw
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta’s TDI Network Engineering team is responsible for the global corporate network, building and supporting a high-performing, reliable network at scale. As a member of this team, you will have a direct impact on network design, deployment, and reliability, enabling our employees to work effectively from any location globally. Your role ensures the overall security and integrity of our corporate network by leveraging network security best practices, innovative products, and rigorous security validation. Reporting to the Network Engineering Manager, this operations-focused role is distinct from core Network Engineering and Network Security, centering primarily on operational execution—including responding to alerts, maintaining service availability, and ensuring system health across our global enterprise network. You will drive the strategic reduction of systemic toil and technical debt across multiple teams, applying a systems-level perspective and leveraging deep expertise in Distributed Systems, Networking fundamentals, Infrastructure as Code, and observability to architect scalable platforms and lead technical efforts to ensure an "Always Secure. Always On." environment. You will own multi-quarter objectives and establish long-term strategies for network reliability. What you'll be doing : Design and Own the resilience, health and availability of our entire global corporate network domain, managing operational responsibilities such as responding to aler
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput, and 99.999 availability. We're looking for a technical leader to help us to continue to scale the service with great people and reliable, cost-effective and efficient infrastructure, processes and tooling. As the Director of Site Reliability Engineering you will oversee the SRE organization focused on Okta platform, Databases, Edge networking, K8s platform, CI/CD, Observability, FinOps, and automation platform & tooling. Job Duties and Responsibilities: Build and lead a high-caliber India-based SRE organization supporting Okta’s production fleet. Partner with global engineering, product, and infrastructure leaders to deliver resilient, scalable, and secure services. Define and execute the India SRE strategy in alignment with global reliability goals. Lead post-incident reviews, drive root-cause analysis, and ensure long-term corrective actions. Participate in incident management, on-call rotations, and blameless RCAs. Implement automation and observability to reduce manual toil and improve operational efficiency. Drive adoption of modern infrastructure practices: infrastructure as code (Terraform), container orchestration (Kubernetes), and AI within Infrastructure org. H
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. About the Role At Okta, identity is at the heart of everything we build. As an Associate Solutions Engineer for Okta, you are a builder, owner, and collaborator who serves as a key technical and business advisor to developer personas across the Australian commercial customers/businesses and the APJ timezone (mostly ANZ). Anchored by Okta’s core values— Love our customers , Build and own it , Always secure. Always on. , and Drive what’s next —you will craft, position, and demonstrate how Auth0 solves critical security challenges throughout the sales cycle. Key Responsibilities Love Our Customers: Partner closely with sales teams as the go-to identity authority, aligning Auth0 solutions directly with developer pain points and business goals through tailored, value-driven product demonstrations. Build and Own It: Architect and execute hands-on Proof of Concepts (POCs) for complex customer use cases, taking full accountability for technical validation alongside cross-functional engineering teams. Always Secure. Always On.: Confidently address in-depth technical, application security, and identity infrastructure inquiries from developers, architects, and CTOs to ensure reliable, high-trust deployments. Drive What’s Next: Serve as the strategic voice of the customer, bringing critical market insights and developer feedback back to Product Management to continuously evolve our platform. Innovate & Share Knowledge: Champion scalable best practices, build reusab
Senior Machine Learning Engineer Description - We are looking for a Senior MLOps Engineer to design, build, and operate the infrastructure that enables machine learning models and large language models to be deployed safely, reliably, and at scale. In this role, you will create the end-to-end capabilities required to move models from experimentation into production, expose them through secure and highly available endpoints, and enable users and applications to interact with AI-powered services. You will work across AWS and Databricks to establish robust CI/CD pipelines, model-serving infrastructure, observability, governance, rollback mechanisms, and operational standards. You will partner closely with data scientists, machine learning engineers, software engineers, security teams, and platform engineers. The ideal candidate combines strong cloud and DevOps engineering skills with a practical understanding of machine learning systems, LLM deployment patterns, and production reliability. Key Responsibilities MLOps Platform and Architecture Design and implement a scalable MLOps platform using AWS and Databricks. Define reference architectures and reusable deployment patterns for traditional machine learning models, deep learning models, and large language models. Build standardized workflows that move models from development and validation into staging and production. Develop self-service capabilities that allow data scientists and ML engineers to deploy models without manually managing infrastructure. Establish clear separation between development, testing, staging, and production environments. Design multi-region or multi-availability-zone architectures where required by business continuity and availability objectives. CI/CD and
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Staff Backend Engineer We’re redefining Privileged Access Management (PAM) from the ground up, purpose-built for Cloud, SaaS, Databases, Containers, and any virtualized environment. Our mission is to simplify and secure workforce access with seamless, secure-by-default workflows that adapt dynamically to modern infrastructure. We eliminate standing privileges, enforce least privilege, and embed Zero Trust principles into every access workflow by default. About the Role We are seeking a Staff Backend Engineer to serve as the core technical anchor and senior Individual Contributor (IC) for our newly established engineering pod in India. At the P4 level, your primary sphere of influence will be at the team level —taking ownership of complex, ambiguous problems and defining how to solve them cleanly, securely, and efficiently. In this role, you will lead by example through hands-on architecture, high-velocity coding, and end-to-end execution. You will drive the implementation of secure database and network device connectors (routers, switches, firewalls) on top of our core Zero Standing Privileges (ZSP) platform. You will work closely with our local Technical Team Lead to elevate the pod’s engineering craft, acting as a technical multiplier for mid-level developers while ensuring tight architectural alignment with our global team. What You’ll Be Doing Execution & Technical Impact End-to-End Ownership: Consistently design, code, debug, test, moni
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. As a Staff Site Reliability Engineer (SRE) at GitLab, you’ll help keep all user-facing services and production systems reliable, scalable, and efficient. Our SREs combine a pragmatic operations mindset with strong software engineering practices to drive automation, reduce toil, and improve resilience across our platform. In the Environment Automation specialization, your focus is on operating and automating hundreds of GitLab environments—from initial provisioning to day-to-day maintenance tasks. Unlike other SRE roles, this position centers on automating the lifecycle of many tenant environments, ensuring they remain secur
Couchbase, the operational data platform for AI, empowers businesses to succeed by bringing data to life in new ways. Major market-leading companies rely on Couchbase for mission critical operational, analytical, mobile and AI workloads. Built to replace legacy infrastructure and fragmented data services, Couchbase empowers enterprises with a unified platform architected for performance, flexibility and global scale. With Couchbase, organizations bring their data to life, launching game‑changing customer experiences, exploring the limitless potential of AI, and seamlessly extending applications from the cloud to the edge and beyond. Couchbase’s AI‑ready technology and enterprise partnership model eliminate complexity and reduce total cost of ownership, enabling teams to stay agile, innovative and secure. Couchbase believes data should never slow you down, but act as the foundation for your next breakthrough. Discover why Couchbase is trusted to help the world’s biggest players scale, move fast and stay resilient, no matter what’s next on their roadmap. Visit couchbase.com and follow us on LinkedIn and X. Want to be part of our story? Apply today! About Couchbase Couchbase is the Operational Data Platform for AI. Our customers don't run AI in a lab — they run it in production, where agents remember, reason, and act on live operational data. With the Couchbase AI Data Plane, we give those agents one governed layer for memory, context, tool access, and MCP, deployed anywhere from public cloud to Kubernetes to the edge to air-gapped environments. Amadeus, Cisco, Comcast, FICO, PepsiCo, United, Verizon, and Wells Fargo trust us with their data. That means AI security isn't a side topic here. It's adjacent to the product, and it's the environment we operate in every day. The Role We're hiring a Sr. Security Operations Engineer to join Couchbase's global Information Security team as our second dedicated SecOps engineer. We think the interesting work in security
Want to be a bswifter? At bswift we’ve been transforming benefits administration since 1996, making it simpler, smarter, and more human. Our state-of-the-art, cloud-based technology and services empower employees to understand, manage, and love their benefits. From downtown Chicago, and remotely across the country, we serve thousands of companies and millions of people nationwide, reducing administrative burdens and freeing HR teams to focus on creating thriving, people-first workplaces. We’re looking for motivated and goal-driven individuals who share our passion for delivering excellence and creating solutions that make a difference. The reward is a fun, flexible and creative environment with ample opportunity for professional and personal growth. If you love the bswift values of pursue excellence, embrace accountability, deliver superior service, and be a great place to work, we want to hear from you! About bswift bswift is a leading benefits administration technology company, delivering modern, cloud-based solutions to employers and health plans across the United States. Our platform supports millions of employees and processes some of the most complex benefits transactions in the industry. As bswift continues to scale and modernize its data platform, we are building a world-class database engineering team — and our India team is a critical part of that vision. Position Summary We are seeking an experienced Database Administrator to join bswift’s growing India team. This is a high-impact, hands-on role that owns the health, performance, and reliability of our SQL Server and AWS database environments during the off-shift window — a critical coverage period for our production systems. You will join bswift’s India engineering team and partner closely with both local engineering colleagues and the US-based database architecture and engineering team. This is not just a maintenance role — you will actively contribute to platform modernizati
SmartBear delivers application integrity for modern tech stacks, ensuring continuous, measurable assurance that software just works as intended with governance to operate at AI speed and scale. SmartBear offers deep test automation, API lifecycle management, and observability capabilities. With integrations across the SDLC, it sets a new quality standard for application delivery teams. SmartBear is trusted by developers, testers, and software engineers across 32,000 organizations, including 75% of the largest financial institutions and industry leaders such as Adobe, JetBlue, and Microsoft. SmartBear’s open source tools are downloaded more than 100 million times a month and have earned over 30,000 GitHub stars from the developer community. With its best-loved brands, including Swagger, TestComplete, Reflect, QMetry, Zephyr, and more, SmartBear meets customers where they are to make our technology-driven world a better place. Learn more at www.smartbear.com , or follow us on LinkedIn , X , and Reddit . At SmartBear, you will be part of a dynamic team solving one of the most critical challenges facing modern businesses: ensuring the integrity of software in an AI-driven world. Whether you are working directly with customers, driving go to market strategies, supporting operations, building products, or enabling teams, your contributions help shape the future of software quality for organizations worldwide. Join us in our mission. Senior Software Engineer – Java Solve challenging business problems and build highly scalable applications. Design, document and implement solutions in Java 8 and 17. Build and deploy services based on the AWS platform. Work with a high-performance team, working with continuous delivery and heavy test automation Product intro Zephyr provides test management capabilities to Atlassian products, working as a plugin/extension for Jira and Confluence, for managing both manual and automated testing processes. Using AI, users can transform thei
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE We are seeking a highly technical Lead Release Engineer with a strong engineering foundation to lead the end-to-end release lifecycle of our storage products. You will serve as the bridge between Development, QA, and Product Management — ensuring that complex storage stacks are delivered with high quality and predictable cadences. Unlike traditional project-based release management, this role demands deep hands-on expertise across CI/CD orchestration, codeline management, system-level triaging, fleet operations, and the engineering rigor required for data-critical products. You will own the health of our release pipelines, lead triage war rooms, drive automation initiatives, and participate in on-call rotations to keep CI and test-orchestration infrastructure running reliably. You will also build developer-facing tooling, manage HW test fleet operations, and maintain high-quality integration workflows across our code lines. WHAT YOU'LL DO Release Orchestration: Own the end-to-end release process for storage software and firmware, from development to GA (General Availability). CI/CD Leadership: Design and build optimized pipelines and tools to scale code management and merge operations. Work closely with systems such as Jenkins, test frameworks, Premerge, Orchestrator, and related developer productivity tooling to keep the codeline healthy and actionable. Technical Triaging: Act as the primary technical poin
Other cities to consider
More places hiring for this role
Get new terraform jobs in Bengaluru, India by email
Daily job updates · Unsubscribe anytime