About the Team The Core Network Engineering team owns the end-to-end networking stack that connects OpenAI’s compute infrastructure — spanning global WAN/edge connectivity, data-center networking, and high-performance host/xPU networking used for large-scale training and inference workloads. This team is responsible for ensuring networking is never the bottleneck to model training efficiency, cluster reliability, or fleet expansion. They design and operate the systems that provide predictable, high-throughput, low-latency connectivity across some of the world’s most advanced AI infrastructure. About the Role We’re looking for engineers to help build and operate the networking foundation behind OpenAI’s frontier AI systems. Depending on your background and area of focus, you may work across host networking, datacenter fabrics, or global WAN infrastructure. The problems span low-level systems software, distributed infrastructure, protocol readiness, observability, performance engineering, automation, and large-scale network operations. You’ll work on systems where microseconds of latency, tail performance, and network reliability directly impact model training efficiency and production serving performance. This role is ideal for engineers who enjoy operating close to the hardware/software boundary and solving performance-critical infrastructure problems at massive scale. In this role, you will: Design, build, and operate networking systems that support large-scale AI training and inference infrastructure Improve performance, reliability, and scalability across host networking, datacenter fabrics, and WAN systems Develop automation for provisioning, configuration management, validation, upgrades, and lifecycle management of networking infrastructure Build tooling and observability systems for network health, performance analysis, debugging, and automated remediation Optimize network performance across technologies such as RDMA, RoCE, InfiniBand, Ethernet, and high-perf
Jobiba hiring network
Software Engineer Infrastructure Jobs
6,326 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current software engineer infrastructure jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About the Team The Scaling team is responsible for the architectural and engineering backbone of OpenAI’s infrastructure. We design and deliver advanced systems that support the deployment and operation of cutting-edge AI models. Our work spans system software, networking, platform architecture, fleet-level monitoring, and performance optimization. About the Role We’re hiring an SW Engineer to enable production workloads and end-to-end testing on new platforms. This role will include creating new test harnesses and platform stress benchmarks, porting existing inference and training workloads to new, sometimes early-access, systems/hardware, analyzing performance and bottlenecks, and characterizing the end-to-end behavior of new systems (compute, comms, storage, control plane, and failure modes). Key Responsibilities Port and validate key inference and training workloads on new platforms/SKUs as they arrive; drive correctness, performance, and stability to an internal readiness bar. Build a suite of benchmarks and stress tests that capture real E2E behavior of our workloads by exercising all aspects of a system, including CPU, GPU, memory subsystem, frontend, scale-up, and scale-out networking (including WAN traffic, NVlink and RDMA collectives), storage, thermals, and any other relevant parts. Deep-dive performance on distributed training/inference: Collective performance and tuning (across NCCL/RCCL and internal libraries) Overlap of compute/communication, kernel-level bottlenecks, memory bandwidth and scheduling effects Create repeatable test harnesses that run in CI / lab environments and produce actionable outputs (pass/fail, performance score, regression detection). Partner with systems + fleet bring-up engineers to ensure the platform is not only stable and performant, but also operationally usable and scalable (containerization, K8s integration, telemetry hooks, failure triage loops). Work cross-functionally with vendors and internal stakeholders by producing
About the Role The Engineering Acceleration Delivery / Continuous Deployment team builds and operates the systems that safely ship OpenAI’s infrastructure and product code to production. We own the deployment platform, release pipelines, and rollout safety mechanisms that allow engineers across OpenAI to deploy changes rapidly while minimizing operational risk. Our mission is to make production deployments fast, safe, and increasingly autonomous. This role sits at the intersection of developer productivity, distributed systems reliability, and large-scale infrastructure orchestration. In This Role, You Will Design and build continuous deployment infrastructure that safely rolls out changes across dozens of Kubernetes clusters and global regions. Develop systems for progressive delivery, including canary releases, staged rollouts, and automated rollback. Improve engineering velocity by reducing friction in the release pipeline and automating manual operational workflows. Work with product and infrastructure teams to ensure their services are deployable, observable, and resilient at scale. Implement and evolve deployment methodologies such as GitOps, infrastructure-as-code, and progressive delivery patterns. Build systems that automatically evaluate deployment health using metrics, logs, traces, and alerts to detect regressions and trigger safe rollbacks. Build systems that support agent-assisted or autonomous deployment workflows using modern AI tooling. Technologies commonly used in this environment include: Kubernetes for large-scale container orchestration and runtime infrastructure Python and FastAPI for internal services Terraform for infrastructure as code GitOps-based deployment workflows (e.g., ArgoCD, Flux, or similar systems) Buildkite for CI orchestration You may be a strong fit if you: Have worked with Kubernetes-based deployment systems at scale Have experience building or operating continuous deployment platforms Are familiar with GitOps tooling such as
Lithic is the modern card issuing and processing platform empowering ambitious financial companies to build the future of payments. Our infrastructure powers card programs for 100+ innovative clients, from fintechs reimagining credit and digital banking to platforms transforming disbursements and spend management. Companies like Mercury, Flex, and Novo rely on Lithic's developer-friendly APIs, direct network connections, and flawless reconciliation to launch and scale card programs in weeks, not years. We're building a future where access to better financial products materially improves people's lives, free from the constraints of 30-year-old mainframes and legacy processors. We're proud to be backed by world-class investors who share that vision, including Bessemer Venture Partners, Index Ventures, Spark Capital, Stripes, and Mastercard, along with many others. We're a team of 170+ across 26 states and 7 countries, headquartered in New York City. We are hiring for our Treasury team Software Engineers at various levels (II and Senior) who are curious and willing to dive deep and understand our technology and domain in order to solve interesting and hard problems.The Treasury team maintains and builds the backend services that manage the flow of funds between Lithic and third parties. This includes our ledger, ACH and wire infrastructure, and associated reconciliation. The systems we maintain have high standards of reliability and correctness. You will become an expert in the card payments space. The Treasury team primarily uses Python for their tech stack. What You'll Do: Ensure high reliability and correctness for Lithic’s ledger and orchestrated funds flows Develop new features to better serve Lithic customers Ensure that the team is delivering reliable, secure, and scalable code with minimal tech debt Own initiatives from planning to launch, keeping stakeholders informed and aligned along the way Lead efforts to improve systems and processes wit
At Breeze, we're building the AI-powered infrastructure layer for global commerce, making it radically simpler for businesses to sell, get paid, and operate across markets. We go far beyond traditional payment processing. Breeze combines global payments, AI, stablecoins, and a Merchant of Record-like model to take on the complexity businesses typically manage themselves, including compliance, risk, fraud, chargebacks, reconciliation, and customer support. Our goal is simple: let businesses focus on building and selling great products while Breeze handles the complexity behind getting paid. Backed by Sequoia Capital , Multicoin Capital , and The Chainsmokers , Breeze is a successful, rapidly growing, and exceptionally well-capitalized company. We have the runway to think long term while remaining early enough that every person joining today can have a meaningful impact on what we build. Overview Want to build products that sit at the heart of how money moves? We’re looking for a Senior Software Engineer to help build and scale the technology powering Breeze. This is a hands-on role with real ownership. You’ll tackle complex engineering challenges, ship products customers rely on every day, and help shape technical decisions as we scale. You’ll work across the stack, partner closely with Product and Design, and have plenty of white space to improve how we build. We’re looking for an engineer who combines strong technical judgment with a product mindset. Someone who moves fast, cares deeply about quality, and gets excited about turning complex payments problems into simple, reliable experiences. If you want to build, not just maintain, and have a meaningful impact on both our product and engineering culture, we’d love to meet you. What You’ll Do Build and scale the core payment systems that power Breeze, with a focus on reliability, security, and performance Ship high-quality customer experiences using React.js and modern serverless infrastructure Own meaningful featur
At Breeze, we're building the AI-powered infrastructure layer for global commerce, making it radically simpler for businesses to sell, get paid, and operate across markets. We go far beyond traditional payment processing. Breeze combines global payments, AI, stablecoins, and a Merchant of Record-like model to take on the complexity businesses typically manage themselves, including compliance, risk, fraud, chargebacks, reconciliation, and customer support. Our goal is simple: let businesses focus on building and selling great products while Breeze handles the complexity behind getting paid. Backed by Sequoia Capital , Multicoin Capital , and The Chainsmokers , Breeze is a successful, rapidly growing, and exceptionally well-capitalized company. We have the runway to think long term while remaining early enough that every person joining today can have a meaningful impact on what we build. Overview We’re hiring a Staff Software Engineer to help build and scale the technology at the heart of Breeze’s payments platform. You’ll take on some of our most complex engineering challenges, building secure, reliable systems that move money while helping shape what we build and how we build it. You’ll work across the stack, own critical systems end to end, and partner closely with Product and Design to turn complex payments problems into simple, scalable solutions. This is a hands-on Staff role with real ownership and technical influence. You’ll help set architectural direction, raise the engineering bar, guide technical decisions across the team, and build the systems and foundations Breeze needs for its next stage of growth. If you’re a strong technical builder who wants to build, not just maintain, solve hard problems, and have a meaningful impact on both our product and engineering culture, we’d love to meet you. What You’ll Do: Build and scale the systems that power Breeze, with a focus on reliability, security, and performance across our payments platform Own critical system
Software Engineer - III About the Role GHX is building a next-generation Intelligent Process Automation (IPA) platform powered by LLMs and AI-native document understanding. We extract structured data from complex healthcare procurement documents — Purchase Orders, invoices, contracts — at scale across cloud environments. As a member of the IPA engineering team, you will bridge strong software engineering with applied AI. You will design and ship Python services, integrate LLM APIs, build AI agents, create skills and validate the output of LLMs to build document extraction pipelines, and own evaluation infrastructure that ensures production quality. This is not a research role — it is a hands-on engineering role where AI fluency amplifies solid software craft. Core Responsibilities Python Development & Automation Build and maintain Python-based automation services and IPA workflows Develop platform-agnostic solutions supporting future migration across automation tooling Build reusable libraries, frameworks, and components for automation projects Integrate automation solutions with REST APIs and enterprise applications Deploy and monitor automation bots on AWS (Lambda, ECS, SQS, S3) AI & LLM Integration Integrate LLM APIs (like Anthropic Claude, OpenAI, Azure AI) into production pipelines Design classification and extraction prompts for diverse document types (POs, invoices, contracts) Write prompts that function as formal specifications — unambiguous, edge-case-aware Build and iterate few-shot, chain-of-thought, and structured output templates Own prompt library versioning, rollback strategy, and prompt lifecycle management AI & LLM Integration Integrate LLM APIs (like Anthropic Claude, OpenAI, Azure AI) into production pipelines Design classification and extraction prompts f
Software Engineer Build technology where every nanosecond matters. At Graviton, software isn't just a tool that supports trading. It is the infrastructure behind every research breakthrough, every trading decision and every competitive advantage. As a Software Engineer , you'll work on systems where performance, reliability and precision matter at an extraordinary scale. You'll partner closely with software engineers and quantitative researchers to build technology that processes enormous volumes of market data, powers quantitative research and supports live trading. You'll take on problems that don't have obvious answers — from designing high-performance systems and distributed infrastructure to eliminating bottlenecks measured in microseconds and building tools that make our researchers and engineers faster. Your work will go into production, be measured against real-world performance and have a direct impact on how our trading systems operate. If you enjoy solving hard engineering problems, understanding systems at a deep level and pushing technology to its limits, you'll feel right at home. What You'll Work On You'll work across the engineering stack that powers our quantitative research and trading platforms. Depending on your team, your work may include: Designing and building high-performance, low-latency systems in modern C++. Building distributed systems that process and analyze massive volumes of market data . Designing systems where latency, throughput and reliability directly influence trading performance . Working on Linux systems, networking, concurrency and multithreaded applications. Profiling systems, identifying bottlenecks and optimizing performance at the hardware and software level. Building robust infrastructure that supports quantitative research and live trading. Debugging complex production systems and solving problems where correctness and reliability are critical. Designing internal platforms and developer tools that accelerate research an
Description: Graviton Research Capital LLP, Gurgaon is looking to hire Software Engineers for our Core Technology team which has some of the best programmers in India working on cutting edge technologies to build a super fast and robust trading infrastructure handling millions of dollars worth of trading transactions every day. As a Senior Software Engineer with Graviton your responsibilities will include: Designing and implementing a high-frequency automated trading system, that trades on multiple exchanges Building live reporting and administration tools for the trading system Performance optimization and improving the overall latency of systems, through algorithm research and using cutting edge tools and techniques End-to-end ownership of modules, including designing, development, deployment and support Growing the team through involvement in the regular hiring process and occasional campus recruitments Requirements : The ideal requirements for our candidates are: A degree in Computer Science 3-5 yrs Experience with C/C++ and object-oriented programming Experience in HFT industry Expertise in algorithms and data structures Excellent problem solving skills Strong communication skills A working knowledge of Linux systems Any of the following is a plus: A good understanding of TCP/IP and Ethernet Knowledge of any other programming language e.g. Java, Scala, Python, bash, Lisp, etc. Familiarity with parallel programming models and parallel algorithms Experience with big data environments e.g. Hadoop, Spark etc. Benefits: Our open and collaborative work culture gives you the freedom to innovate and experiment. Our cubicle free offices, non-hierarchical work culture and insistence to hire the very best creates a melting pot for great ideas and technological innovations. Everyone on the team is approachable, there is nothing better than working with friends! Our perks have you covered. Competitive compensation Annual international team outing Fully covered commuti
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Build software used to support global trading across time zones • Work closely with business users to establish and refine requirements • Participate in identifying new technologies to continuously improve software systems • Implement DevOps practices within the team (GitHub/GitLab, Jenkins) • Provide production support to diagnose and resolve elevated application software production incidents and implement proactive remediation measures WHAT’S REQUIRED • Bachelor’s degree in computer science or another technical/scientific field • Minimum 8 years object-oriented programming experience with C#/.NET • Significant experience working with / understanding databases - primarily MS SQL Server • Must have experience in writing automated tests, unit tests, Test-Driven Development • Knowledge of design/architecture patterns, distributed systems, microservices, observability and monitoring, and containerization (Docker) • Willingness to work as part of a distributed Dev Team - 3 time zones (USA, Poland, India) • Strong problem solving and analytical skills • Exceptional verbal and written communication skills • Commitment to the highest ethical standards WE TAKE CARE OF OUR PEOPLE We invest in our people, their careers, their health, and their well-being. When you work here, we provide: • Health care benefits • Maternity, Adoption & related leave policies • Generous paternity and family care leave policies • Employee Assistance Prog
A Career with point72’s technology Team As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you’ll do As a software engineer on our Risk Technology team, you will use your programming expertise to build a scalable reporting framework capable of processing terabytes of data for complex risk calculations within Point72 Risk Department. Build analytical framework to enable flexible large-scale compute on big datasets Develop quantitative understanding of risk to anticipate the requirements of our team Innovate and improve our capabilities to deliver high quality products to our business partners Solve complex and challenging problems through the application of creative and unique solutions using the latest technology What’s REQUIRED Minimum of 5 years’ programming experience in Python and/or one of functional languages Bachelor's in computer science, math, physics or related technical field Programming experience in a quantitative environment with statistic models or complex aggregations Awesome team player with excellent written and verbal communication skills with peers and clients Problem solving skills with deep understanding of core computer science algorithms and data structures Commitment to the highest ethical standards We take care of our people We invest in our people, their careers, their health, and their well-being. When you work here, we provide: Health care benefits Maternity, Adoption & related leave policies Generous paternity and family care leave policies Employee Assistance
A Career with Point72’s Technology Team As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. What you’ll do Build software applications and deliver software enhancements and projects supporting fund accounting and trade processing technology. Work closely with business stakeholders to develop software solutions using test-driven and agile software development methodologies. Be responsible for system upgrades and features supporting resiliency and capacity improvements, automation and controls, and integration with internal and external vendors and services. Driving architecture of core platforms and accelerating modernizing leveraging AI tools. Work with DevOps teams to manage and resolve operational issues and leverage CI/CD platforms while following DevOps practices within the team and projects. Continuously improve the platforms using the latest technologies and software development ideas. Participate in initiatives to transition select applications to cloud platforms, enhancing stability, scalability and performance of the existing platform. Develop close wo
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our technology group is constantly improving our company’s technology infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. The Back Office Technology Team supports trade processing, position keeping, clearing/settlement, fund accounting, trade reconciliation and prime broker integrations. The team partners with Middle and Back Office business users to customize and implement solutions supporting trade processing and new business developments. WHAT YOU’LL DO We are looking for an experienced professional to work as part of the Back Office Technology team. In addition to tactical development, you will be responsible for delivering and creating programs to modernize and scale the platform through technology upgrades, cloud technology adoption, and re-architecting business processes. You will work in the Back Office Technology team with world class engineers, developing high-capacity integrations and development capabilities with in-house and vendor build applications, implementing new financial products, managing internal and prime broker data, and developing the data warehouse of the future to support growing business. This position assumes close interaction with business and opportunity to build a high-demanding domain knowledge in post-trade flow. Build software applications and deliver software enhancements and projects supporting fund accounting and trade processing technology. Work closely with business stakeholders to develop software solutions using test-driven and agile software development methodologies. Be responsible for system upgrades and features supporting resiliency and capacity improvements, automation and controls, and integration with internal and external ven
You’ll shape the future of a business‑critical platform as the technical lead across both product engineering and cloud infrastructure. You’ll modernize a mature .NET application running on AWS today, while steering its evolution toward a cloud‑native, React/Node.js, AI‑enabled architecture. If you enjoy owning architecture end‑to‑end, from backend and frontend through CI/CD, DevOps, and AWS infrastructure, this role gives you real influence at Staff Engineer level and the opportunity to set engineering standards that others follow. You’ll spend your time leading complex .NET and React features, designing scalable AWS infrastructure with Infrastructure as Code, and building automation that makes releases fast, safe, and repeatable. You’ll work on performance, reliability, and modernization in equal measure—fixing what’s slowing the platform down today and designing what it will look like in the next generation. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead the architecture and development of enterprise .NET services and APIs that power a business‑critical platform. Design and operate AWS infrastructure (using AWS CDK in TypeScript) to support secure, scalable, multi‑environment deployments. Build and optimize CI/CD pipelines (AWS CodePipeline, CodeBuild, Windows build agents) to make shipping .NET and React changes fast and reliable. Drive modernization initiatives across the stack, including clean architecture, refactoring legacy components, and reducing technical debt. Design and tune PostgreSQL and MSSQL database solutions for performance, scalability, and reliability. Mentor engineers and influence engineering practices across teams, raising the bar on cloud, DevOps, and software design. These are the essentials you’ll need to get an interview Significant experience (typically 8+ years) delivering and operating scalable enterprise software, owning both application code and cloud infrastructure. Deep hands‑on expertise with C
ROLE: SOFTWARE ENGINEER TEAM: MEDIA LOCATION: TORONTO (HYBRID) COMPANY OVERVIEW Salt is a North American marketing agency that creates connected experiences through creative, digital & media innovation. Our mission is to “Earn The World’s Attention” and Salt is built to find, develop, execute, and amplify the ideas that are worthy of our clients’ audiences. We’ve structured our agency to do what’s right for our clients – to connect different perspectives, to work across mediums, and to focus on delivering meaningful, effective results. We’re committed to living up to the values in our name. “Salt of The Earth” means we value collaborative, humble, hard-working people here. We’re looking for people who are as smart as they are kind because we believe the right talent and the right culture help us do what’s right for our clients. ROLE OVERVIEW Join us as a Software Engineer building the next generation of AI products. This hands-on engineering role combines Platform Engineering, Security, and Software Development to create the cloud platform that powers our applications at scale. You'll spend approximately 50% of your time building secure, automated, and reliable infrastructure, and the other 50% developing innovative product features - working with the latest cloud, AI, and developer technologies to deliver solutions used across the organization. You will build, secure, and operate modern cloud-based applications and AI platforms, developing new product features while helping ensure our infrastructure is secure, scalable, and reliable. Working closely with product and engineering, you will develop new product features while managing cloud infrastructure, CI/CD pipelines, application security, and production operations across our AI platforms. CORE RESPONSIBILITIES Infrastructure & Data Platform Engineering Maintain efficient cloud Infrastructure, and development environments. Develop i
Get new software engineer infrastructure jobs by email
Daily job updates · Unsubscribe anytime