Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The Payments organization owns some of Stripe's most critical payment flows and a platform that processes hundreds of billions of dollars in payments a year. Our team is responsible for translating complex partner specifications related to network costs (interchange and scheme fees) into simplified logic for internal and external consumption. This also drives decisions and product recommendations to manage the underlying network costs paid by Stripe and our users. The team partners closely with the engineering, product, finance, and partnership groups to manage and understand Stripe's network costs. Our work is core to Stripe's business, as Technical Operations roles in Payments are a dynamic and key component of Stripe's success. We sit at the intersection of product and platform engineers and financial partners, connecting them to ensure that everyone thrives and nothing is lost in translation. What you'll do We're looking to add payment enthusiasts who enjoy interpreting complex cost structures and ever-changing payment network systems to optimize on behalf of Stripe and our users. You'll be instrumental in building Stripe's approach to managing our global network cost base. Responsibilities • Collaborate across the company, including engineering, accounting, financial partnerships, and product teams, to analyze billions of dollars moving through the Stripe platform • Translate network specifications into implementation instructions to cr
Jobiba hiring network
Infrastructure Team Manager Jobs
4,729 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current infrastructure team manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe is seeking a highly experienced and strategic Chief Compliance Officer (CCO) to lead our regulatory compliance, financial crimes and enterprise risk officers and program across Asia-Pacific. This role reports to the Global Head of Compliance Officers & MLROs and sits within the Financial Crime, Compliance & Risk Oversight (FinCRO) organization. You will also serve as the formal CCO for Stripe's Singapore-regulated entity. You will be the principal decision-maker and strategic lead on regulatory compliance, financial crime and enterprise risk management for APAC, serving a critical role in ensuring Stripe's continued safe growth in the region. You will be the public face of the compliance team across APAC and a primary point of contact with regulators, financial partners, and internal stakeholders. The ideal candidate will have deep subject matter expertise in regulatory compliance, financial crimes and ERM across APAC markets, and a proven track record building and evolving tech-forward compliance programs in a regulated financial services or payments environment. Key Responsibilities: Strategic Leadership: Develop and implement a robust and scalable compliance strategy for APAC, aligning with global FinCRO objectives and business goals. Drive a culture of compliance and sound regulatory risk management across the region as Stripe expands its product offerings in existing markets and enters new ones. AML/Fi
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team To further this important mission, we are building the foundation for a long-term talent bench at Stripe. We believe every campus hire can build products with meaningful impact at Stripe, and provide a far-reaching impact for anyone trying to grow their business online. We are looking for a University Recruiter to join our team to attract and hire the very best junior talent, while providing each candidate with an exceptional experience. This is a 12 month fixed term contract. What you’ll do We’re a small but mighty team, and are committed to big results. You will help shape the future of our early-stage programs and be responsible for all parts of the campus recruiting life cycle, including, but not limited to, organizing and leading on campus events, driving a strong and diverse candidate pipeline, branding, building relationships with internal stakeholders, and managing our global internship program. You must be a builder who thrives in a learning environment - unafraid to try new things, embrace new ideas, and welcome suggestions for how we can iterate on our processes. Responsibilities Partner closely with Stripe’s engineering organization to build the foundation of our early-stage university program Manage the full-cycle recruiting process for new graduate and intern candidates across multiple roles and offices Develop deep relationships with university groups, including faculty and student organizations, to design school-specific stra
A Career with Point72’s Technology Team As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. What you’ll do Optimize cloud financial operations to maximize value from cloud investments, including rapidly growing artificial intelligence (AI) and machine learning workloads Provide actionable insights on cloud spend, SaaS license optimization, and emerging AI cost drivers, including model inference and usage-based consumption Implement tooling, tagging standards, and processes that improve cost visibility and optimization across cloud, SaaS, and AI workloads Monitor large language model API consumption and GPU-intensive infrastructure to identify cost trends, anomalies, and optimization opportunities Build financial models to forecast cloud, SaaS, and AI expenditures for budgeting cycles, commitment decisions, and vendor negotiations Design cost allocation, tagging, showback, and chargeback models that attribute spend to the teams, applications, and use cases driving it Educate engineering and business owners on cloud financial management practices th
NVIDIA’s DGX Cloud organization is seeking a Senior Data Engineer to become part of its data team! We develop the reliable data foundation that supports fleet health, capacity, utilization, cost, reliability, and operational decision-making throughout DGX Cloud. Our platform supports engineering, operations, finance, and product teams managing and expanding large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We are looking for a practical engineer and technical lead to take charge of a key part of the Navigator data platform. We develop the systems that transform distributed infrastructure telemetry and operational data into dependable, managed data products that support fleet health, capacity, utilization, cost, and operational decisions. We are seeking a hands-on, platform-minded engineer to build and evolve the systems that turn distributed infrastructure telemetry and operational data into reliable, governed data products. You will work across ingestion, transformation, data quality, platform architecture, security, observability, and self-service consumption to help make Navigator and the DGXC data platform a dependable source of truth. We do expect strong engineering fundamentals, experience operating production systems, and the ability to learn new platforms and domains quickly. What you'll be doing: Own systems end to end. For example, work from ambiguous customer and operational needs through architecture, implementation, deployment, observability, incident response, and ongoing support. Construct data pipelines and products. Such as designing and maintain batch and streaming ingestion, transformation, reconciliation, and serving paths for fleet, capacity, utilization, cost, scheduling, and operational telemetry. Build shared libraries, workflow and DAG or equivalent experience abstractions to evolve the data platform. Develop deployment tooling, data
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Startup and SMB Account Executive team is a highly consultative sales team that is responsible for the acquisition and growth of Stripe’s largest and most promising Startup and SMB customers. As an Account Executive, you’ll identify new opportunities for prospects to get the most out of Stripe by quickly understanding our clients’ needs, execute scaled sales strategies to acquire new customers and drive product adoption. What you’ll do Responsibilities Identify high-potential prospective users from inbound leads and outbound prospecting Own the full sales cycle from lead to close for SMBs Develop and implement full sales cycle strategies for driving new logo and product adoption Solve complex client needs and work across product, sales, risk, and operations teams to improve our platform Generate your own leads through cold calling, blitzing, research, networking, and driving your territory Make every potential Stripe user happy with every interaction, regardless of deal size Identify and understand users’ pain points to propose Stripe solutions Set up users for success by coordinating internally with other Stripe teams Who you are We’re looking for a well-rounded strategic seller who can build strong relationships with Stripe’s clients and manage high-velocity deal cycles. If you’re smart, persistent, and a great teammate, we want to hear from you! Minimum requirements 2+ years of closing sales experience with a track record of top perfo
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Link is a digital wallet designed for fast and secure online payments. It allows consumers to save and use their preferred payment methods across the Link network, helping them check out quickly and securely wherever Link is accepted. The Link Fraud and Auth team works to make Link the most trusted and highest-performing way to pay. We protect consumers and merchants from fraud, abuse, and financial loss while maximizing authorization rates for good users. Our work spans consumer-facing experiences, payment infrastructure, and ML powered risk systems. We manage fraud and financial risk across a growing range of novel Link features, including Link’s agentic wallet, stored balance, and LPMs. The team also owns Instant Bank Payments, a proprietary payment method built on ACH rails, offering merchants immediate confirmation while protecting them from bank-initiated returns. IBP is the heart of Link’s revenue engine, giving LFA engineers the opportunity to shape and scale one of Link’s most important products. What you’ll do As a machine learning engineer on Link Fraud and Auth, you’ll build and operate models and risk decisioning systems that protect Link while helping more legitimate payments succeed. You’ll work across the full machine learning lifecycle, from analyzing fraud patterns and identifying opportunities to building, deploying, monitoring, and improving models in production. You’ll use data to form hypotheses, make practical modeling
Principal Targeting Analyst Target development for cyber operations, red-team engagements, and applied R&D Position Overview: Position: Principal Targeting Analyst Job Type: Full-time Location: Ft. Meade, MD (Hybrid) Clearance Requirements: US Citizen &TS/SCI Experience: 8+ years What You'll Do SIXGEN is seeking a Principal Targeting Analyst to build and lead an organic targeting and collection capability that strengthens our offensive cyber, red-team, and adversary-emulation missions. This is a hands-on technical leadership role for an analyst-engineer hybrid who enjoys solving hard collection problems, developing targets from a cold start, and turning operational experience into automated, mission-ready capability. As a Principal Targeting Analyst, you'll pair deep OSINT and managed-attribution tradecraft with professional foreign-language proficiency, applied AI engineering, and real software-development ability. You'll partner with operators, engineers, and program teams to deliver timely targeting and access intelligence, mature SIXGEN's collection methodologies and tooling, and mentor analysts as the capability grows. Whether you're developing a target for a critical customer mission, automating collection at scale, or shaping the technical direction of a new capability, your operational credibility and technical judgment will help drive mission success across the organization. Key Responsibilities Target Development and Key Operations Conduct end-to-end target development against threat actors, and adversary infrastructure — from initial requirement through finished, decision-ready intelligence. Direct collection across the surface, deep, and dark web, closed forums, encrypted messaging platforms, and regional ecosystems beyond the coverage of commercial data sources. Perform authorized threat-actor engagement and elicitation; track initial-access brokers, exploit sellers, and access markets relevant to mission objectives. Conduct
About Hexnode Hexnode, the Enterprise software division of Mitsogo Inc., was founded with a mission to simplify the way people work. Operating in over 100 countries, Hexnode UEM empowers organizations in diverse sectors. Fueling the transformation to a seamless ecosystem of connected tools, Hexnode is revolutionizing the enterprise software and cybersecurity landscape. Role Overview We are seeking a AWS Operations Specialist to manage and maintain our cloud infrastructure and device ecosystems. This is a highly operational, execution-focused role—not an architecture position. The ideal candidate has 2 to 4 years of experience executing infrastructure as code, monitoring environments, and following documented playbooks to keep our systems secure and resilient. Because this role handles secure environments, candidates must be US Citizens and capable of passing a comprehensive federal background check. Key Responsibilities Infrastructure Execution: Run, maintain, and execute existing Terraform and Ansible scripts to deploy and update infrastructure. GovCloud Monitoring: Actively monitor our AWS GovCloud dashboards, keeping a close eye on system health, performance metrics, and security baselines. Mobile Device Management: Manage Android Enterprise kiosk configurations, ensuring secure deployments and smooth device operations. Incident Response & Triage: Respond swiftly to operational alerts by strictly following our documented team playbooks. Escalation: Identify anomalies or issues that fall outside established, documented procedures and escalate them accurately to the engineering team. Required Qualifications & Profile Citizenship: Must be a US Citizen (required for GovCloud infrastructure management). Background: Must be able to successfully clear a rigorous federal background investigation. Experience: 2 to 4 years of hands-on experience in a technical operations, DevOps, or SysAdmin role. Technical Familiarity: Comfort executing/running Terraform and
Role Purpose: At Jumio, you will work for one of the market leaders in the global identity verification space that is helping to make the digital world a safer place for everyone. As a Software Development Engineer in the MLOpsTeam, you will develop the blueprint for highly scalable and performant ML model serving. Role Value: As a Software Engineer (SDE III), you will drive the continuous improvement of the infrastructure and applications to manage the lifecycle of ML assets (data, models) to better developer experience and strengthen governance capabilities. Secondly, you will design and implement robust ML infrastructure for model deployment, serving, and optimization. You will work on efficient CI/CD pipelines for ML models and leverage advanced compilers or hardware optimization to maximize inference performance while optimizing costs. We welcome you to challenge us to impact our software development processes and tools. Example Responsibilities: Upgrade ML assets (models, data) management systems for better developer experience and robust governance capabilities Build and optimize model serving infrastructure with a focus on inference latency and cost optimization Architect efficient inference pipelines that balance latency, throughput, and cost across various acceleration options Implement cost-efficient, enterprise-scale solutions Collaborate in a cross-functional, distributed team for continuous system improvement Work with MLEs, QA Engineers, and DevOps Engineers Evaluate and implement new technologies and tools Contribute to architectural decisions for distributed ML systems Experience and Qualifications : 5+ years of experience in software engineering with Python Experience with model lifecycle management (MLFlow, Weights & Biases or equivalent) Experience with data management ecosystem (quality, transformation, catalog) Experience with ML frameworks, particularly PyTorch Experience optimizing ML models with hardwar
Scale’s rapidly growing Global Public Sector team is focused on using AI to address critical challenges facing the public sector around the world. Our core work consists of: Creating custom AI applications that will impact millions of citizens Generating high-quality training data for custom LLMs Upskilling and advisory services to spread the impact of AI As a Full Stack Software Engineer (Forward Deployed), you’ll collaborate directly with public sector counterparts to quickly build full-stack, AI applications, to solve their most pressing challenges and achieve meaningful impact for citizens. At Scale, we’re not just building AI solutions—we’re enabling the public sector to transform their operations and better serve citizens through cutting-edge technology. If you’re ready to shape the future of AI in the public sector and be a founding member of our team, we’d love to hear from you. You will: Partner with public sector clients to scope, collect feedback and implement solutions for complex problems, including spending up to two weeks per month in client offices for feedback and delivery. Architect production-grade applications that integrate AI models with full-stack frameworks, managing everything from interactive UIs to backend APIs and systems. Deploy and manage infrastructure within cloud environments, ensuring the highest levels of system integrity, security, scalability, and long-term reliability. Contribute to core platform features designed to be reused across diverse international client use cases. Partner with design, product, and data teams to build robust applications aligned with the broader technical architecture. Ideally you’d have: Bachelor’s degree in Computer Science or a related quantitative field 5+ years of post-graduation, full-stack engineering experience with demonstrated proficiency in React (required), TypeScript, Next.js, Python, Node.js, PostgreSQL or MongoDB plus hands-on experience with Docker, Kubernetes, and Azure
About the Role: We are looking for a Senior DevOps Engineer to join our DevOps team at K Health. You will own and evolve the infrastructure underpinning a healthcare AI platform serving patients and enterprise health system partners. This is a high-ownership role: you will architect and operate cloud environments across K Health and its enterprise partners, lead complex infrastructure migrations, drive disaster recovery programs, and help build the next generation of AI-powered operations tooling. You will also mentor junior engineers and collaborate closely with product and engineering teams across the company. This is a hybrid role based in New York City (4 days/week in office) and includes participation in a daytime on-call rotation. What you will do: Own the design, implementation, and evolution of our GKE-based Kubernetes infrastructure across K Health and enterprise partner environments. Build and maintain our Terraform modular infrastructure library, including reusable modules with automated testing, across GCP, Cloudflare, and AWS. Architect, build, and maintain GitLab CI/CD shared pipeline templates used by all engineering teams (build, test, security scanning, deployment). Own and maintain self-hosted infrastructure software running in-cluster, including GitLab, ArgoCD, Langfuse, DependencyTrack, NGINX Ingress, and others. Implement and support security and compliance controls across infrastructure and the software supply chain - secrets management, pipeline secret detection, container scanning, SOC2 and HIPAA. Drive disaster recovery readiness: design failover scenarios, author runbooks, and lead periodic DR tests. Lead development of AI-powered operations tooling and agentic infrastructure. Monitor, troubleshoot, and improve production system reliability; respond to incidents during on-call shifts. Mentor junior DevOps engineers and establish team-wide engineering standards. What we are looking for: 5+ years of experience in DevOps, platform engineering,
We are expanding our agentic AI capability and are looking for an AI Engineer to join the team. You will work alongside senior engineers to build and maintain AI systems — contributing to agentic pipelines, retrieval infrastructure, and the integrations that tie these systems together. This is a hands-on implementation role with real ownership of components. You will grow your skills in a fast-moving AI practice, working on production systems that directly affect client outcomes. What This Involves: Build and maintain agentic pipelines and workflows under the guidance of senior engineers: tool use, orchestration, and multi-step reasoning. Implement and tune RAG pipelines — including embedding, chunking strategies, vector retrieval, and retrieval evaluation. Contribute to memory and context layer components: integrating vector databases, supporting knowledge graph pipelines, and helping maintain state management across agentic systems. Write clean, well-tested Python code and participate in code reviews. Debug and improve existing AI systems based on evaluation results and production feedback. Collaborate with data engineers and domain experts to integrate AI components with upstream data sources and downstream applications. Document implementations clearly and contribute to shared internal tooling. Requirements: 2–4 years of software or ML engineering experience, with at least 1 year working with LLMs or AI systems in a professional setting. Working knowledge of LLM APIs (OpenAI, Anthropic, or similar) and at least one agentic or RAG framework (LangChain, LlamaIndex, or equivalent). Solid Python skills and comfort with software engineering basics: version control, testing, REST APIs. Familiarity with vector databases or embedding-based search. Curiosity about agentic AI — you follow developments in the space and are eager to apply new techniques. Excellent communication and collaboration skills — comfortable working across cross-functional and client-facing te
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role Modal's LLM inference platform delivers frontier performance for open-source models with best-in-class elasticity and developer experience, made in part possible by our custom runtime with GPU memory snapshots and multi-cloud substrate . We're looking for a leader to own the direction and execution of this platform to continue to establish us as the clear market leader, working closely with customers like Cognition, Doordash, Ramp, and many more. You'll be leading a group of highly talented engineers working on our market-leading LLM inference offering, spanning the serving stack, routing infrastructure, internal agentic optimization platform, and the user-facing product surface area. This is a hands-on leadership role — expect to split your time between technical contribution, product shaping and people management depending on what the team needs. You'll set direct
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role As a Software Engineer, Forward Deployed, you’ll solve the hardest problems standing between Ramp and the world’s largest and most complex companies — shaping and shipping the product capabilities that unlock our growth upmarket. You'll be part of our Core FDE org, which is an agent-first, high-pace, customer-facing engineering team. On FDE, you will interact directly with customers and deliver solutions end to end — understanding pain points, shaping product decisions, and building agents that autonomously expand Ramp's capabilities. Check out our Engineering Blog and FDE post for more context on our work! What You’ll Do Deliver software end to end that meet the needs of our largest customers — understanding user pain points, scoping product specs, and building agents that autonomously implement solutions. Collaborate closely with Sales, Solutions, Customer Success, and Account Management to close deals, activate customers, and expand the value Ramp provides over time. Drive the core product engineering roadmap through our embedding
Get new infrastructure team manager jobs by email
Daily job updates · Unsubscribe anytime