Jobs in United States

Ai Infrastructure System Engineer Bangalore in United States

5,418 active opportunities · Updated October 2026

Explore current ai infrastructure system engineer bangalore jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team Industrial Compute is building the world's most advanced AI infrastructure. Working alongside our capital partners, engineering teams, and construction organizations, we design and deliver large-scale, mission-critical compute campuses that power the next generation of AI. Our Design organization brings together engineering, construction, and digital design to ensure facilities are coordinated, constructible, and optimized from the earliest planning phases through deployment. About the Role We are seeking a BIM Designer & Coordinator to support the planning and design of large-scale industrial and mission-critical facilities. In this role, you will develop, coordinate, and maintain multidisciplinary BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines, enabling early design validation, constructability reviews, equipment planning, and cross-functional coordination. You will partner closely with internal engineering teams, external design consultants, contractors, and project stakeholders to produce coordinated BIM deliverables that improve design quality, reduce project risk, and support efficient execution across Industrial Compute's global infrastructure portfolio. Key Responsibilities: Develop and maintain conceptual and schematic BIM models for large-scale industrial and MEP-intensive facilities. Coordinate BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines. Create and maintain federated models used for design reviews, spatial coordination, constructability analysis, and clash detection. Model major building systems including equipment layouts, utility corridors, electrical rooms, mechanical rooms, structural framing, site infrastructure, and architectural constraints. Coordinate equipment clearances, maintenance access, routing zones, shafts, risers, utility entrances, and major MEP pathways. Translate engineering sketches, basis-of-design documents, equipment lists

PythonAWSGitRest
JI
📍 New York, NY, United States
✓ Quality checked

JLL empowers you to shape a brighter way . Our people at JLL are shaping the future of real estate for a better world by combining world class services, advisory and technology for our clients. We are committed to hiring the best, most talented people and empowering them to thrive, grow meaningful careers and to find a place where they belong. Whether you’ve got deep experience in commercial real estate, skilled trades or technology, or you’re looking to apply your relevant experience to a new industry, join our team as we help shape a brighter way forward. What this job involves: As a Senior Project Manager for our Healthcare team, you will serve as a dedicated Owner's Representative, guiding complex infrastructure projects from concept to completion. This position requires a unique blend of deep technical expertise in mechanical and electrical systems and sophisticated client-facing skills to champion our clients' interests. At JLL, we are collectively shaping a brighter way for our clients, and in this role, you will be their trusted advocate—ensuring healthcare facilities are delivered on schedule, within budget, and to the highest standards of quality and regulatory compliance. You will act as the central point of communication, translating intricate technical details into clear, strategic actions for all project stakeholders. What your day-to-day will look like: Serve as the owner’s primary advocate and liaison, fostering collaboration between the client, design teams, contractors, and all project stakeholders. Manage overall project performance, including scope, schedule, and budget, to ensure successful delivery against key milestones. Provide expert technical oversight for mechanical (HVAC, medical gas, plumbing) and electrical (power distribution, emergency power) infrastructure systems. Review design d

Artificial IntelligenceAIProject ManagementPmp
O
📍 United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. The Industrial Compute team is central to this mission, setting the core infra strategy and implementing this vision. From site selection to the buildout process, this team sits at the intersection of commercial, technical, strategy, and operations, interacting with teams and executives inside and outside of OpenAI. About the Role Responsible for validating that proposed sites are buildable, compliant, and cost-effective. You will lead diligence across civil, geotechnical, environmental, and entitlement dimensions, identifying risks and driving mitigation strategies. Key Responsibilities Lead all technical diligence: geotech, soils, title/ALTA surveys, mineral rights, and access. Oversee permitting/entitlement path and schedule governance with agencies. Evaluate generator air permits, wetlands, floodplain, and stormwater constraints. Manage consultants performing feasibility studies and environmental assessments. Deliver go/no-go recommendations with risk and mitigation options. Build diligence templates and playbooks to scale future site reviews. Qualifications 8+ years in land development, civil/environmental engineering, or data center diligence. Knowledge of permitting, entitlements, and AHJ engagement. Strong project management and technical review skills. Experience managing consultants and interpreting complex studies. Regularly communicate site readiness updates, risks, and milestones to executive stakeholders Establish and track key performance indicators to assess the effectiveness of the site selection program and the contributions of external vendors and partners. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely depl

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. The Infrastructure team is central to this mission, setting the core strategy and implementing the vision. From site selection to deployment to operations, this team sits at the intersection of commercial, technical, and operational domains, interacting with experts and executives inside and outside of OpenAI. We design and operate mission-critical facilities that support cutting-edge AI workloads at scale. About the Role We are seeking a Facilities Operations Lead to support the commissioning, deployment, and long-term operation of our next-generation AI data centers. This role bridges the interface between data center construction and hardware landing, ensuring seamless integration of mission-critical infrastructure with hardware deployment timelines. You will define and execute commissioning plans, support infrastructure bring-up, and take ownership of operations and maintenance for cutting-edge, large-scale, AI data centers. You will collaborate closely with design, construction, and hardware teams to define repeatable processes for new data center builds and lead hands-on operations to uphold the performance and reliability of our deployed infrastructure. Key Responsibilities Define and execute sequences of operations, commissioning steps, and bring-up processes for mission-critical data center facilities. Interface with the design and hardware teams to define deployment procedures tailored to each data center and hardware configuration. Oversee installation, commissioning, and operational readiness of large-scale data center campuses. Manage monitoring, maintenance, and quality control of the data center infrastructure, including high-performance liquid cooling systems. Develop on-site operations staffing strategy. Develop and enforce procedures for planed and unplanned downtime and SLAs for critical

AWSRestAIRust
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -93.7%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is transforming how the world uses data and AI — and the networking and traffic infrastructure that powers these experiences is mission-critical. As a Product Manager focused on Traffic & Networking, you will define how Snowflake delivers secure, reliable, and high-performance connectivity at global scale, including the networking foundations required to support AI-driven products and workloads. You will own the product vision and roadmap for internal traffic management, service-to-service networking, customer connectivity, and performance optimization across multi-cloud environments. A core part of this role is defining and evolving Snowflake’s network strategy to support AI products , including latency-sensitive inference, large-scale model training pipelines, vector search, streaming ingestion, and cross-region data movement. This is a high-impact role at the intersection of distributed systems, cloud networking, and AI infrastructure. AS A PRODUCT MANAGER AT SNOWFLAKE YOU WILL: Define the networking strategy required to support Snowflake’s AI products , including low-latency inference paths, high-throughput data pipelines, GPU-adjacent services, and elastic scaling for AI workloads. Partner with AI platform, compute, and storage teams to ensure networking

AWSAzureGCPAI
R
📍 Foster City, California, United States· Full-time
✓ Quality checkedCompany trend -89%

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. We are looking for a Security Operations Lead (SOC Lead) to build, mature, and operate our 24/7 detection and response capabilities across a modern cloud-native and AI-driven environment. This role leads the global SOC function—monitoring, SIEM ownership, detection engineering, alert triage, and operational readiness—while also evaluating and integrating emerging AI-based SOC products and autonomous response platforms . You will oversee monitoring across multi-cloud environments (GCP primary, AWS/Azure secondary), Kubernetes, SaaS services, endpoints, developer tools, and AI workloads . You’ll collaborate closely with Cloud Security, Compliance/GRC, SRE, Platform Engineering, IT/Endpoint teams, and AI Infrastructure to ensure our detection strategy scales and stays ahead of evolving threats. This is a hands-on leadership role perfect for someone who wants to shape the SOC of the future while solving complex challenges in a high-scale AI setting. What You’ll Do SOC Leadership & 24/7 Monitoring Lead, mentor, and scale a global SOC team responsible for 24/7 monitoring, alert intake, triage, correlation, and escalation. Build operational rigor: processes, runbooks, SLAs, metrics, and quality standards for high-scale environments. Cover monitoring across: Cloud infrastructure (GCP, AWS, Azure) Kubernetes/GKE/EKS/AKS clusters SaaS platforms (Google Workspace, GitHub, Slack, Okta, etc.) Endpoints (macOS, Linux, Windows) including EDR/XDR telemetry Developer platforms + CI/CD pipelines AI/ML systems and model-serving workflows AI-Based SOC Integration & Innovation Evaluate, adopt, and integrate AI-native SOC technologies for triaging, detection, and correlation Identify opportunities to automate triage, investigations,

PythonAWSAzureGCP
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $280.5K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the Role: AI models reshaping how our community creates, plays, and connects, all run on Compute Platform. As Senior Product Manager, Compute Platform , you'll set the strategy and roadmap for Roblox's next-generation AI infrastructure: the rapidly growing fleet of GPUs and AI accelerators spanning Roblox core and edge data centers, and public cloud that decides how fast we can train, serve, and scale every model on the platform. You'll own the products that turn raw GPU hosts into reliable, production-ready AI compute - driver and firmware management, fleet-wide health and performance, and the abstractions product teams across Roblox build on. You Will: Drive strategy and roadmap for Compute Platform spanning Managed Kubernetes (Roblox Kubernetes Service), Managed Compute Services and other critical distributed systems, and our fleet of GPU and CPU machines managed via unified Fleet APIs - all across on-prem and cloud. Drive the evolution of our Compute infrastructure to support Roblox’s most critical workloads - from AI to Storage to Data Analytics and more - each with their own distinct requirements. Build and scale our GPU infrastructure to support training and inferen

AWSAzureGCPKubernetes
NR
📍 Georgia, Washington, United States· Full-time
✓ High-confidence listingCompany trend -75%

From $12.6K/yr

Quick readStrong listing-quality and freshness signals

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are seeking a dynamic, analytical, and execution-oriented Treasury Manager to own and elevate our treasury operations and contribute to the development and execution of Treasury strategy. This isn't a role for someone who just wants to maintain the status quo—we need a builder. You will be responsible for managing day-to-day cash operations, liquidity, and debt management, while actively designing, setting up, and modernizing our banking & treasury management systems infrastructure. This role is operational and strategic at the same time. You will partner closely with teams across the entire organization to help streamline and scale treasury-related processes. If you thrive in a fast-paced, dynamic and challenging environment, we want to talk to you. What you'll do Treasury Operations & Liquidity Management Cash & Liquidity: Manage daily cash management & treasury operations, ensuring optimal liquidity levels across all entities while positioning the company for growth. Credit Facility Management: Manage credit facility activities and support the lender-related processes. Forecasting: Build, maintain, and continuously improve robust short-term cash flow forecasts. Implement improvements to help automate workflows and upstream processes to make short-term forecasting more efficient and accurate. Support long-term liquidity planning exercises. KYC & Compliance : Support treasury-related KYC and compliance requiremen

SQLGitRestAI
A
📍 New York, New York, United States· Full-time
✓ High-confidence listing

$960K – $1.2M/yr

Quick readStrong listing-quality and freshness signals

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. The IT team helps keep Affirm’s people productive, connected, and secure by supporting the technology employees rely on every day. We manage and support end-user devices, enterprise applications, office technology, and IT infrastructure while maintaining a strong focus on security, compliance, and a seamless employee experience. Our team works across a broad range of technologies and partners closely with employees throughout the company to troubleshoot issues, improve processes, and build reliable, scalable solutions as Affirm grows. We’re seeking a flexible IT Support Administrator to help manage and support Affirm’s New York City office and technology infrastructure. Our ideal candidate excels in troubleshooting and problem-solving, with respect for compliance, and documentation. As a Support Administrator of a rapidly growing company, you will be the backbone of IT, with efficiency, security, and customer satisfaction being your primary objectives. What You'll Do: Locally support and troubleshoot users within our NYC office Remotely support and troubleshoot our end user’s problems Educate on and enforce our IT compliance and policy obligations Maintain our Zoom Rooms and AV hardware within conference rooms Monitor for security threats and malware, staying ahead of the curve to protect our users Occasionally maintain the in-office, local network infrastructure Maintain enterprise systems and company IT assets such as servers and desktops What We Look For: Experience with macOS administration and support Ability to improvise and find immediate solutions Experience managing formal ticketing systems Experience administering SaaS technologies Flexibility to work with all types of technology macOS, Windows, Printers, Network devices, etc. Experience with machine and account prov

RestAIGoExcel
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -84.7%
Quick readStrong listing-quality and freshness signals

About the Team The ChatGPT Search Product Infrastructure team builds the foundational systems that power search experiences across ChatGPT. We develop the product infrastructure that connects models with search systems and other sources of real-time information, enabling ChatGPT to deliver timely, relevant, and trustworthy answers to users around the world. Our work sits at the intersection of product engineering, AI, and large-scale infrastructure. We build shared platforms and abstractions that enable product teams to independently develop, evaluate, and launch new search-powered experiences. These platforms provide the guardrails, testing capabilities, observability, and rollout controls needed to prevent reliability, scalability, quality, and latency regressions while supporting rapid product iteration. The team partners closely with: Post-Training on model launches, experimentation, and prompt optimization Search product verticals on new user experiences Inference on GPU efficiencies Indexing and Retrieval on the systems that identify and deliver relevant information Capacity/Fleet team to ensure optimal regionalized provisioning of GPUs and CPUs About the Role We are looking for an Engineering Manager to lead the team responsible for ChatGPT’s Search Product Infrastructure. You will set the technical and organizational direction for the systems that bring search capabilities into ChatGPT. You will guide architectural decisions across search orchestration, model and prompt integration, serving infrastructure, experimentation, observability, evaluation, and product integrations. You will balance immediate launch and product needs with the long-term reliability, scalability, latency, and maintainability of the platform. A central responsibility of this role is creating leverage for Search product verticals. You will lead the development of extensible platforms that allow those teams to independently build, test, and launch features without requiring ongoing invol

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

$293K – $405K/yr

Quick readStrong listing-quality and freshness signals

About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity. Our work involves: Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping misuse safeguards, alignment tools, and security measures on track to adequately address extreme threats that might arise in the future. Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework , and partnering with other staff to achieve these targets. This is urgent, fast-paced work that has far-reaching implications for the company and for society. About the role As AI agents become more capable at software engineering, and automate more of our internal work, they could become a dangerous cyber threat. People in this role will help OpenAI prepare for security threats from advanced AI agent insiders. In this role, you will: Identify paths by which capable future internal AI agents could compromise OpenAI. Design security controls - focusing on measures with long lead times that benefit from advanced preparation. Stress-test defenses with AI agent evaluations and penetration tests You might thrive in this role if you: Are deeply technical across security and modern infrastructure, and are comfortable digging into the details of operating systems, cloud, containers, CI/CD, or distributed systems. Have strong software engineering skills and enjoy building prototypes yourself. Are interested in engaging with stakeholders and can do so effectively. Bonus: have experience securing cloud infrastructure, and are deeply familiar with core components of the AI stack. Compensation Range: $293K - $405K USD About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy

AWSCI/CDRestAI
R
📍 New York, NY, United States· Full-time· Remote
✓ High-confidence listingCompany trend -99.2%

From $10K/yr

Quick readStrong listing-quality and freshness signals

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Ramp is in a critical phase of growth. We grew immensely last year and are building out a talented business systems team to ensure we maintain this trajectory for years to come. You’ll work directly with our Sales, Account Management, Partnerships, and Product teams to execute mission-critical business systems projects across the organization. This is a key role where you will be uniquely positioned to impact the full picture of Ramp’s growth efforts through systems development. What You’ll Do Work alongside Sales Operations to administer key go-to-market business systems, including Salesforce, Outreach, Qualified, Zendesk, Hubspot, Looker, Gong.io Build and deploy automation (flows), validations, and applications in Salesforce Implement new systems and integrations as needed Analyze key business requirements and systems capabilities to write specifications for systems build and run end to end implementation Create key reports and dashboards to track systems performance and data accuracy Write and maintain clear documentation on syste

RestAIGoSalesforce
V
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -92.7%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. You will own the data Vanta's EPD organization actually runs on — bringing new signal sources online, standardizing them at the point of ingestion, and building the systems that make the underlying data trustworthy rather than merely stored. The EPD Systems team is building the infrastructure Vanta's engineering, product, and design organization depends on to understand itself. Not by asking teams to be more diligent but by going to the source, processing it, standardizing it, and pushing value back out so each source of truth earns its own adoption. What you’ll do as a Operations Manager, Signal Systems at Vanta: Bring new signal sources online end to end — from discovery and scoping through ingestion, standardization, and live operation Identify the specific failure modes in each information source and build systems that mitigate them at the point of ingestion Go to the teams that produce and consume a signal, understand what the information actually means and what they need back from it, and build accordingly Replace human-diligence dependencies with engineering solutions: derive fields, pull from source systems, validate at write time Build alongside teammates who are growing into building — raise their technical ceiling, not just your own output Partner with the inference and systems layers to ensure what you produce is queryable, trustworthy, and ready to build on How to be successful in this role: You look at a data source and see its failure modes before its contents: where it lies, where it goes stale, where it's duplicated, where the schema won't hold at 10x Your first move on an adherence problem is an engineering an

G
📍 Austin, Texas, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that deliver the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing some of the world's most transformative technologies. Our AI Engineering Campus in Austin plays an important role in building the future of AI computing. The Opportunity As Technical Services Director, you will lead the teams that operate and evolve Graphcore's engineering labs, high-performance computing (HPC) platforms, and data center environments globally. You will be accountable for reliable, secure, cost-effective infrastructure that supports demanding engineering, AI, silicon-development, and validation workloads. This role combines people leadership, infrastructure strategy, operational excellence, capacity and financial planning, procurement, and program delivery. You will partner with Engineering, Information Technology, Security, Finance, Facilities, Supply Chain, customers, and external suppliers. The position is based onsite in Austin and requires travel to company facilities, data centers, and supplier locations, including international travel. What You'll Do Lead, recruit, mentor, and develop the systems administration, lab operations, and technical services teams responsible for the facility supporting global Engineering and Research and Development. Own the reliability, efficiency, protection, safety, supportability, and continuous improvement of engineering labs, HPC systems, and infrastructure facilities. Establish service levels, operating standards, escalation paths, performance measures, monitoring, observability, automation, ticketing, and configuration-management practices. Translate engineering and customer requirements into infrastructure roadmaps, capacity p

LinuxAIGoExcel
R
📍 Foster City, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -89%

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. As a Support Systems Lead at Replit, you'll build and own the systems that let support scale as fast as the product does. You'll own the tools our team works in every day, the technical setup behind self-service and AI-driven support, and the infrastructure that keeps all of it current as Replit ships. Replit is at the forefront of AI-driven software development, and how we support customers is constantly evolving. You'll shape how our support systems adapt to new products, new surfaces, and AI-assisted workflows, operating effectively in ambiguity and turning ad hoc fixes into infrastructure the whole team can rely on. You'll combine hands-on technical depth with systems thinking to keep builders moving, whether they get unblocked through self-service, an AI agent, or a person. This is an individual contributor role to start, with room to grow and build out a team as support scales. IN THIS ROLE YOU WILL: Configure and maintain Zendesk and the surrounding support stack, from the day-to-day workflow and automation setup to business rules and permissions, keeping it able to flex and scale as needs change. Build and maintain assignment logic, queues, tagging and taxonomy, and escalation paths, keeping them running cleanly as volume and workflows change. Set up and maintain support tooling, workflows, and access across internal agents, outsourced vendors, and regions, keeping the systems working for each group as the stack changes. Partner with Engineering on the technical requirements for self-service and in-product support surfaces, and build the entry points, help widgets, and routing behind them. Spot the repetitive steps in agent and admin workflows before they become bottlenecks, and build the automations, bulk acti

PythonAISEMLean
🔔

Get new ai infrastructure system engineer bangalore jobs in United States by email

Daily job updates · Unsubscribe anytime