About the team The AI Deployment Engineering team ensures the safe and effective deployment of Generative AI applications for developers and enterprises. We serve as trusted technical advisors, helping customers and partners move from early experimentation to production-scale AI systems. As a Partner AI Deployment Engineer focused on AWS, you will operate at the center of one of our most strategic partnerships, driving joint customer success and enabling AWS and partner ecosystems to scale adoption of OpenAI-powered solutions. About the role We are looking for a highly experienced technical leader to serve as the primary technical counterpart to AWS field leadership (Solutions Architects, Specialists, and Partner teams). This role goes beyond individual deal support—you will shape strategy, define engagement models, and build repeatable systems that scale across AWS globally. You will work across pre- and post-sales, guiding complex enterprise customers from ideation to production while enabling AWS and partners to independently drive deployments. You will combine deep technical expertise, strong judgment, and ecosystem leadership to maximize impact across a portfolio of high-priority opportunities. This role is based in Bangalore . In this role, you will: Strategic AWS Engagement & Influence Serve as the senior technical counterpart to AWS field leadership, building trust and credibility across regions and teams. Influence joint account strategy and technical direction for high-priority opportunities. Shape how OpenAI engages with AWS by defining engagement models, prioritization frameworks, and best practices. Proactively identify and drive net-new opportunities and high-impact use cases across the AWS ecosystem. Complex Deal Leadership & Execution Lead technical strategy for large, ambiguous, and high-stakes enterprise engagements. Guide customers from early ideation through architecture design, prototyping, and production deployment. Act as a technical d
Jobs in India
Ai Deployment Engineer in India
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current ai deployment engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. L9 - Senior Business Solutions Engineer, Legal Tech Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The Legal Technology team, within the BizTech organization, leads the mission to deliver innovative technology, empowering our legal function to utilize technology productively, driving connection, and scale support. This role sits within BizTech and partners directly with Airbnb's global Legal organization to accelerate the deployment of AI-powered tools, implementation of legal systems & integrations, and workflow automation across the CLO org. The Difference You Will Make: We're looking for a world-class Senior Business Systems Engineer to help redefine how Legal operates. You'll deliver fast, practical solutions using internal tools and agentic AI — owning system and tool changes end-to-end, from design through support. You'll spot what's slowing teams down, champion best practices, and safeguard data integrity across our platforms — all while staying ahead of how AI is reshaping the way we work. This isn't conventional IT or Legal Ops. You'll be embedded directly within our Legal teams, solving real challenges in real time. As a key driver of our Legal Tech strategy, you'll lead the deployment of AI agents, legal application implementations, and automation that
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Required Skills: 9+ years in Python and its libraries (e.g., Pandas, Boto3) for data manipulation, ETL processes, and developing serverless functions to interact with foundational AI services on platforms like AWS Bedrock . 6+ years with modern front-end frameworks (e.g., React, Next.js, TypeScript), and an ability to collaborate effectively with Product & Design. Deep hands-on experience with AWS solutioning , including designing and deploying production applications using core services such as AWS EC2, AWS Lambda, Amazon S3, Amazon RDS, and API Gateway . Familiarity with Infrastructure as Code (IaC) tools like CloudFormation or Terraform is essential. Proven experience building and deploying production-grade AI applications and services. Experience with agentic frameworks like LangChain, LangGraph, and LangSmith is a strong plus. Strong background in building distributed systems, microservices, and resilient APIs (REST/GraphQL) on cloud platforms (AWS/GCP/Azure). Demonstrated ability to influence technical strategy and lead architecture decisions across global teams. Excellent communication skills with experience working effectively across time zones and cultures. Working experience with foundational AI services, with a desire to use a unified platform like AWS Bedrock for Generative AI development and deployment Education and Certifications A Bachelor’s degree in Computer Science, Information Systems, or equivalent years of industry experience AWS cl
About the Role At FourKites we have the opportunity to tackle complex challenges with real-world impacts. Whether it’s medical supplies from Cardinal Health or groceries for Walmart, the FourKites platform helps customers operate global supply chains that are efficient, agile and sustainable. Join a team of curious problem solvers that celebrates differences, leads with empathy and values inclusivity. As a Senior Customer Engineer, you own the technical customer relationship end to end. You run discovery independently, design integration and agentic workflow architectures for complex enterprise problems, and deploy solutions live with customers — often before they know exactly how to articulate what they need. You write optimized, production-grade code at speed, grounded in strong data structures and algorithms fundamentals, because compressing the time from customer problem to working solution is how FDE delivers its value. You bring genuine innovation to hard problems — your solutions are technically sound, elegant, and often non-obvious. You are the primary technical contact for 2–4 major enterprise accounts, you mentor FDEs on the team, and you are building the skills that will take you into Staff-level technical leadership. What You'll Do Own complex integration and AI agent implementations end-to-end — from technical discovery through go-live and post-launch enhancement — as the primary technical contact for 2–4 enterprise accounts Design integration architectures with explicit attention to error handling, retry logic, observability, failure recovery, and multi-system authentication Design and deploy agentic AI workflows that orchestrate supply chain operations — from requirements through production, including regression testing and validation before each customer deployment Deploy AI agent workflows live with customers present — configuring and troubleshooting in the room during customer calls, not gathering requirements to build later Run customer discovery
We are looking for a Full Stack Developer to join our growing engineering team. In this role, you will design, build, and operate scalable software platforms that support analytics and AI solutions — owning the full journey from intuitive user interfaces to robust cloud-native backends. What This Involves: Front-End Development Develop and maintain user interfaces for web applications using React and Next.js. Translate wireframes and design mockups into functional, accessible UI components. Identify and resolve front-end performance bottlenecks to ensure smooth user experiences. Write and maintain unit and integration tests for front-end components (e.g. Jest, React Testing Library). Back-End Development Develop and maintain high-quality back-end services and APIs using Python. Deploy, operate, and monitor applications in cloud environments (AWS, Azure, or GCP). Manage containerized applications using Docker and Kubernetes. Contribute to the design and evolution of scalable, cloud-native software architectures. Contribute to and maintain CI/CD pipelines for web and back-end applications. Collaboration and Quality Work closely with data scientists, engineers and project managers to deliver integrated end-to-end solutions. Support the development and deployment of AI and analytics solutions. Write clean, well-documented, and maintainable code across the full stack. Participate in technical discussions, code reviews, and continuous improvement initiatives. Adhere to internal and client-mandated data protection and compliance policies, ensuring all handling, storage, and sharing of data meets required security and privacy standards. Requirements: Bachelor’s degree in Computer Science or related fields. 5+ years of software development experience, with meaningful time on both front-end and back-end systems. Experience designing systems in cloud-native or distributed environments is a plus. Excellent communication and collaboration skills — comfortable working acros
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Senior Engineering Manager, Continuous Deployment at GitLab, you'll build and manage a globally distributed team focused on Continuous Deployment capabilities within GitLab's artificial intelligence-powered DevSecOps platform. You'll hire engineers, shape how the team works, and guide delivery of a reliable product experience for customers. Your team will build a Continuous Deployment engine that goes beyond script execution. It will reconcile live state, coordinate durable workflows, and support artificial intelligence-native governance. You'll align technical direction with customer needs and busin
EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute efficiency and density compared to today’s best-in-class solutions. The high-performance architecture is coupled with seamless software integration and will enable the immense potential of AI to be accessible in power, energy, and space constrained applications. EnCharge AI launched in 2022 and is led by veteran technologists with backgrounds in semiconductor design and AI systems. About the Role EnCharge AI is seeking a highly skilled and experienced AI Compiler Engineer to spearhead the efforts in developing and optimizing graph compilers tailored to cutting-edge AI and ML workloads. You will collaborate with hardware architects, and AI researchers to enhance performance, optimize computation graphs, and enable efficient model deployment on EnCharge’s Inference Accelerators. Responsibilities Architect, design, and implement optimizations for AI model execution on graph compilers to improve performance, reduce latency, and maximize hardware utilization. Work closely with ML researchers, hardware engineers, and software developers to design and deploy AI models, understanding and addressing hardware-specific challenges. Work on performance optimizations for neural network models, such as layer fusion, operator fusion, and graph-level transformations. Develop compiler optimizations and passes that convert high-level AI models (e.g., from TensorFlow, PyTorch) into intermediate representations (IR). Implement parsing, semantic analysis, and IR generation for deep learning frameworks. Research and integrate the latest advancements in compiler design, ML model optimizations, and hardware acceleration into graph compilers. Provide leadership, mentorship, and technical guidance to a team of engineers focused on graph compiler optimizations. Qual
About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale. If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk. Responsibilities Design and build fleet automation systems that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention. Build AI Infrastructure Agents that automate deployment, root-cause failures, incident triage, and autonomous remediation. Develop Fleet Intelligence platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers. Build software that maximizes GPU availability, utilization, performance, and reliability across thousands of accelerators. Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads. Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations. Continuously improve deployment velocity, reliability, and operational efficiency through automation. Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure. Requirements 3+ years building distributed systems, infrastructure platforms, or large-scale backend software. Strong software engineering skills in Python, Go, or Rust . Experience building platforms, automation systems, or developer infrastructure. Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies. Strong systems thinking with the ability to understand problems across hardw
Role: Senior AI Engineer Location: Hyderabad, India (Hybrid) Department: Product Development About the Role GHX is building a cutting-edge LLM-powered document understanding platform focused on classification, structured data extraction, and intelligent orchestration at scale. This is a high-impact AI engineering role where you will own the full lifecycle—from problem framing to production deployment . Initially, you will focus on prompt engineering and evaluation systems , building the quality foundation for AI performance. Over time, the role expands into agent orchestration, system architecture, and migration of rule-based systems to LLM-driven pipelines . A strong foundation in software engineering (5+ years) is essential. This role demands engineering rigor across both traditional system design and AI system behavior . Core Responsibilities 1. Prompt Engineering Design prompts for diverse document classification and extraction tasks Treat prompts as formal specifications (precise, structured, and edge-case-aware) Develop few-shot, chain-of-thought, and structured output templates Manage prompt lifecycle: versioning, testing, and rollback 2. LLM Output Evaluation Create and maintain ground truth datasets Build automated evaluation pipelines (precision, recall, field-level accuracy) Identify and resolve conceptually incorrect outputs despite surface correctness 3. AI Agent Orchestration Design multi-agent workflows for document processing Implement tool-use patterns and integrate MCP servers Optimize orchestration for scale and efficiency 4. Software Engineering Develop production-grade APIs and backend services Apply Clean Architecture / DDD principles Write maintainable, testable Python code Contribute to CI/CD, deployment, and observability systems 5. Stakeholder Collaboration Act as a bridge between business stakeholders and AI systems Translate product requirements into technical architectures Communicate system behavior, limitations, and quality
₹2K – ₹2K/yr
Opportunity Overview: We are seeking a Senior Software Engineer - AI to join our Engineering team. In this role, you will build the intelligent agents and applications serving health insurance plans covering over 15 million people. You'll partner closely with product, data science, and clinical teams to build the foundation that transforms how clinical intelligence is delivered at scale. This is an opportunity to make a direct impact on healthcare outcomes while working with modern technologies in a fast-paced, collaborative environment. What you’ll do: Agent Development : Participate in the development, evaluation, and deployment of Cohere’s AI-powered agents and applications. Data & Retrieval Architecture : Design data and retrieval architectures that give AI agents the right context across diverse healthcare sources. Hands-On Engineering : Design, review, and own high-quality agentic code — leading releases, production deployments and on-call support. Cross-Functional Collaboration : Work closely with ML/DS teams, product teams, and architects to understand requirements and ensure systems meet business needs. Security & Compliance : Ensure all agentic components comply with healthcare security and privacy regulations (e.g., HIPAA) and adhere to industry best practices for security. Agile Development : Contribute to sprint planning, execution, and retrospectives, driving efficiency and velocity within the engineering team and collaborating closely with stakeholders to align with product goals and business priorities. ISMS roles and responsibilities: Good knowledge of Information practices. Assist the manager in all the information security activities implementation and maintenance process. Ensuring the team and imparted with Competence related to Information security Responsible for implementation of security policies and procedures and report any issues to the Information Security Manager. Required Qualifications: Must-haves Bachelor’s degree
NK Securities Research is a leading financial firm that leverages cutting-edge technology and sophisticated algorithms to trade the financial markets. Founded in 2011, we have gained invaluable experience in the field of High-Frequency Trading (HFT) across different asset classes. Role Overview We’re looking for engineers who can take AI work beyond experiments and make it hold up in production. You’ll work closely with quant researchers and infra engineers to build AI systems that actually get used improving research speed and internal tooling without slowing down the core stack. We value engineers who think about trade-offs, test what they build, and care about how things run in production. What You’ll Build Production AI Ship models that meet defined latency and reliability expectation Add monitoring, rollback, and guardrails before anything goes live Optimise inference across CPU/GPU environments when it matters Integration into Real Systems Plug AI into data-heavy workflows without hurting performance Work within existing low-latency architecture instead of fighting it Profile and remove bottlenecks rather than guessing AI for Engineers & Researchers Build tools that genuinely speed up research and development Improve code understanding, review workflows, and internal knowledge retrieval Keep systems auditable and predictable LLM & Retrieval Systems Implement structured RAG and embedding pipelines with validation in place Create safe integration layers between models and internal systems Performance & Standards Track latency, drift, and stability — not just accuracy Build observability into everything you ship Help raise the bar for how AI is engineered here What We’re Looking For Strong Python fundamentals Clear thinking around system design and performance trade-offs Experience deploying AI systems in production (1–5 years is typical) Familiarity with transformers, embeddings, or LLM deployment Nice to have: Exposure to C++ / Rust / Go E
Title: AI Application Packaging Engineer Location: Andheri (East), Mumbai Reports to: Service Director ABOUT BLENHEIM CHALCOT: As part of the Blenheim Chalcot portfolio, we benefit from the expertise, infrastructure, and scale of the leading global venture builder. With over 25 years of experience creating and growing SaaS businesses powered by Generative AI, Blenheim Chalcot has built 60+ ventures across sectors such as financial services, education, health, and marketing. Our global ecosystem—including Scale Space in London, the Rajasthan Royals in Mumbai, and a go-to-market base in Austin—enables us to access world-class talent, tools, and support to accelerate our growth and build a market-leading business. THE ROLE: We’re looking for an AI-first ITS Engineer with strong Application Packaging expertise to join our growing team at Agilisys. In this role, you’ll work across Application Packaging, Endpoint Management, Cloud Deployment, and Automation, helping modernise IT operations through AI-driven and scalable solutions. This is an opportunity to work in a fast-paced environment where automation, innovation, and operational efficiency are at the core of what we do. ABOUT YOU: Must-Have's 5+ years of hands-on experience with SCCM and/or Microsoft Intune for application packaging, deployment, and endpoint management. Strong experience with PowerShell scripting and a good understanding of Windows Registry, File System internals, and deployment workflows. Experience managing application packaging, deployment, patching, and remediation processes in enterprise environments. Exposure to vulnerability management and patch lifecycle activities using tools such as WSUS, Active Directory, or endpoint security solutions. An AI-first mindset with interest or experience leveraging Generative AI tools to improve efficiency across packaging, testing, deployment, troubleshooting, and operational tasks. Hands-on experience using automation or AI-assisted tools
SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime. With the ability to build, scale and manage security across the cloud, hybrid and traditional environments in real-time, SonicWall provides relentless security against the most evasive cyberattacks across endless exposure points for increasingly remote, mobile and cloud-enabled users. With its own threat research center, SonicWall can quickly and economically provide purpose-built security solutions to enable any organization—enterprise, government agencies and SMBs—around the world. For more information, visit www.sonicwall.com or follow us on Twitter , LinkedIn , Facebook and Instagram . Position Summary We are hiring a Software Dev Engineer to design, build, test, and deploy AI-powered applications. You will work across the full application lifecycle — from architecture and implementation through automated testing, CI/CD deployment, and production monitoring — building features that put large language models and agentic tooling to work inside SonicWall's internal systems. This is a hands-on engineering role for someone with 6–8 years of professional software development experience who is comfortable owning services end to end: writing production-quality code, integrating LLM and agentic APIs, standing up reliable data and retrieval pipelines, and shipping through a disciplined test-and-release process. You will collaborate closely with senior engineers, product stakeholders, and platform teams to turn requirements into dependable, well-tested applications. Key Responsibilities Develop AI applications: Design and build features and services that use LLM and agentic capabilities — prompting, tool/function calling, retrieval-augmented generation (RAG
We are looking for a highly motivated AI/ML Software Engineer to join the Enterprise Agentic AI Platform team within IT. You will work closely with Business Analysts, and Engineering teams to design, develop, and deploy enterprise AI solutions that improve productivity and automate business workflows across Engineering, Operations, and Manufacturing. What you'll be doing: Design, develop, and deploy Agentic AI applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and AI orchestration frameworks. Build scalable AI services and reusable components integrated with enterprise applications such as PLM, SAP, and other business systems. Collaborate with business and IT teams to translate business requirements into AI-driven solutions. Develop secure, scalable APIs and enterprise integrations to enable intelligent workflows and automation. Improve AI solution quality, performance, and reliability through prompt engineering, evaluation, and continuous optimization. Partner with cross-functional teams throughout the Software Development Lifecycle (SDLC), from solution design through deployment and production support. What we need to see: Bachelor's or Master's degree in Computer Science, Information Technology, AI/ML, or a related field. 6+ years of software engineering experience with strong proficiency in Python and backend application development. Hands-on experience with Generative AI, LLMs, RAG, AI agents, REST APIs, and cloud-native application development. Experience integrating enterprise applications and building scalable, production-ready software solutions. Strong analytical, problem-solving, communicatio
Remote Web Developer (Equity Partner 10%) Role Type: Founding Web Developer (Equity-Based) Location: Remote Experience: Freshers Welcome (Strong builders preferred) Compensation: 10% Equity (Vesting-Based) Cash Compensation: None initially About the Startup We are building two high-impact digital platforms: 1️ Solar Fintech Platform A fractional ownership platform where: • Users buy micro-shares (e.g., 100 units) of solar panels. • Once fully funded, solar panels are deployed to corporates. • Energy generation produces returns distributed proportionally to shareholders. Core Components: • Investor dashboard • Payment integration • Solar performance analytics • Admin & corporate PPA tracking 2️ Carpooling Platform (Similar to Mitfahrgelegenheit) A ride-sharing website connecting drivers and passengers traveling between cities. Core Components: • Driver ride listing system • Seat booking & payment • Route matching • Ratings & verification • Real-time notifications Role Overview As a Founding Web Developer, you will architect, build, and deploy both platforms from scratch. You are not just a coder you are a technical co-builder. Key Responsibilities Product Development • Develop full-stack web applications (frontend + backend) • Build scalable database architecture • Implement secure payment gateway integrations • Develop admin dashboards • Deploy MVP and iterate rapidly Solar Fintech Specific • Micro-share allocation logic • Wallet & transaction history system • Solar data API integration • ROI calculation engine • Investor performance dashboard Carpooling Platform Specific • Ride-matching algorithm • Location-based search • Booking & seat allocation system • Driver/passenger profile system Infrastructure • Cloud deployment (AWS / GCP / Azure) • CI/CD pipelines • Basic cybersecurity implementation • API architecture Skills Required • React / Next.js / Vue (Frontend) • Node.js / Django / Laravel (Backend) • SQL / NoSQL databases • Payment gateway integration exper
Other cities to consider
More places hiring for this role
Get new ai deployment engineer jobs in India by email
Daily job updates · Unsubscribe anytime