ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for a Head of Environment to build the physical expression of Baseten as we grow from roughly 300 employees today to more than 2,000 over the next couple of years. You’ll own the environments, spaces, and experiences that define what it feels like to work at Baseten. This is equal parts operations, hospitality, design, and strategy. You’ll partner closely with our founders and leadership team to create an environment that scales with the company while remaining unmistakably Baseten. You’ll inherit a strong Workplace Experience team and continue to evolve the function into one of the defining strengths of the company. We’re looking for someone who has built and scaled world-class workplace functions before; someone who has seen what’s ahead and can help us get there faster. RESPONSIBILITIES Own Baseten’s global workplace and real estate strategy, developing the long-term roadmap for how our physical footprint evolves as we scale from hundreds to thousands of employees. Lead real estate planning, site selection, expansions, and significant capital investments across our headquarters, growing network of smaller offices, and future international locations. Build and operate exceptional workplaces. Own end-to-end workplace operations, office buildouts, relocations, and launches, ensuring every office runs smoothly with an uncompromising bar for quality, hospitality, and attention to detail. Define the
Jobiba hiring network
Lead Network Reliability Engineer Jobs
6,876 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current lead network reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building its own GPU infrastructure for large-scale inference. As we move into large scale, high-density NVIDIA systems, the hardest failures are intermittent, cross-layer, and difficult to prove: RoCE congestion, InfiniBand stalls, ECN/DCQCN mis-tuning, bad optics, RNIC issues, host kernel stalls, GPU driver problems, and workload symptoms that look like network problems, but are not. We are hiring a Lead Software Engineer to build a first-class observability and root-cause analysis system for GPU fabrics. This is a hard distributed systems problem, not a dashboarding problem. The system will collect high-volume signals from switches, hosts, active probes, and inference services; reduce and correlate them in real time; understand topology and service ownership; and produce actionable diagnosis while an incident is still unfolding. This role sits at the boundary between networking and inference software. RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request routing, and workload backpressure can all create fabric symptoms or hide real fabric failures. The goal is to tell an operator, quickly and with evidence, whether an incident is caused by the fabric, host, NIC, GPU, RDMA path, scheduler, or serving layer — and what to do next. EXAMPLE INITIATIVES Real-time telemetry engine — Build the ingestion, reduction, storage, and query path for high-cardinality fab
Coder builds enterprise software that keeps developers in flow. As a Senior Customer Support Engineer, you will help customers solve complex technical issues, get unstuck faster, and build with confidence. You will work closely with Product, Customer Success, and Enterprise Account Managers to resolve urgent issues, improve support workflows, and turn customer feedback into better product experiences. This is a highly visible role on a growing team, with direct impact on how customers experience Coder. What you’ll do here Reproduce and debug customer issues using existing tools, test environments, or custom reproduction setups Solve incoming technical support requests in a timely manner, including urgent, high-severity cases Communicate clearly and tactfully with customers throughout the support lifecycle Gather context, provide diagnostic steps, share resolution guidance, and follow up after issues are resolved Identify and communicate product usage trends, bugs, and feature requests Collaborate with Enterprise Account Managers to schedule, coordinate, and lead customer debugging calls Document customer activity in accordance with internal and external security standards Guide and mentor new team members on support processes and procedures Contribute to product documentation, customer knowledge base articles, and best practice guides Improve support processes, tools, and workflows in collaboration with the broader team Participate in a periodic on-call rotation for production-down issues What we’re looking for 5+ years of customer support engineering experience, or a comparable customer-facing technical role, preferably with critical software products 5+ years of experience supporting enterprise customers 2+ years of experience diagnosing production network connectivity and performance issues Strong problem-solving, analytical, and troubleshooting skills Clear, thoughtful communication skills, both verbal and written Experience with Terraform Experience with major
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As the Director of Communications at Vanta you’ll be part of a growing Communications & Content team responsible for telling Vanta’s company story across external and internal audiences,including media, influencers and 3rd party validators, customers, prospects and partners. The Director of Communications will lead and drive our integrated communications program, increasing awareness of Vanta and our market leadership across external and internal audiences. This role is a critical member of our team focused on amplifying Vanta’s brand and market vision, driving both enterprise and start-up credibility and illustrating the value we deliver for global customers across traditional and new media audiences. The ideal candidate is a hands-on communications pro with a passion for crafting compelling corporate narratives, developing close relationships with key industry influencers, and working across teams to drive results. We're looking for a critical thinker with strong writing and communication skills, deep AI/B2B/SaaS/Security product knowledge, a wide network of relationships across traditional and new media, excitement to experiment with AI and a passion for pioneering the future of trust. What you’ll do as a Director of Communications at Vanta: Drive the execution of our external commu
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As the Senior Director of Solutions Architecture for Asia Pacific at Cohere, you will own Solutions Architecture across the region. Our customers here are banks, telcos, industrials and governments who cannot send their data to someone else’s cloud, and who are past experimenting and into production. You will lead the team that makes those deployments real, and you will be accountable for the technical win across the region. This is a build. Today a Solutions Architecture team covers Korea, Japan and Southeast Asia. You will own that team, grow it, and build the in-region depth this market needs. You will be expected to open doors on your own credibility and network from your first weeks, set direction for the function, sit on the Solution Architecture leadership team alongside the regional leaders for the Americas and EMEA, and contribute to company-wide decisions with your peers across Sales, Product and Engineering. In this role, you will: Build and Lead the Team: Hire, coach and develop the Asia Pacific Solutions Architecture organization, and establish in-region depth rather than relying on support flown in from other regio
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As a Manager of Security Engineering, your key responsibilities include: Serve as trusted advisor to team’s leadership and partner teams by clearly articulating business risks associated with security issues Execute the long-term vision for the Security team in alignment with Cohere’s product and business goals. Collaborate closely with leadership to prioritize high-impact initiatives and strategic customer engagements. Vulnerability Management: Develop and implement enterprise-wide vulnerability management processes and tooling, including identification, prioritization, remediation tracking, and reporting, including customer artifacts Static Application Security Testing (SAST): Establish SAST programs, integrate tools into CI/CD pipelines, and analyze results to identify and remediate security flaws in source code Dynamic Application Security Testing (DAST): Implement DAST methodologies, configure scanning tools, and conduct regular assessments of running applications Penetration Testing: Lead and oversee internal and external penetration testing engagements, including web application, API, network and agentic AI platform inclu
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The Opportunity: Flexport is insourcing operations in Japan — one of the world's most quality-driven and relationship-led freight markets. To stand up our in-house Tokyo operations team, we are seeking a Senior Operations Developer: a deeply experienced freight forwarding operator who will help us build the local team, establish the right vendor relationships, and accelerate the velocity of our insourcing initiative. Japan is a market where seniority and trust take precedence in business dealings, and where local relationships with carriers, CFS providers, and drayage partners are difficult and slow to build from the outside. As Senior Operations Developer, you will use your existing network and operational expertise to dramatically shorten that ramp time — laying the foundation for a Tokyo operations function that supports a projected ~150 FCL, ~30 LCL, and ~44 Air shipments per month by year-end 2026. You will: Establish and strengthen Flexport's relationships and operating arrangements with key local vendors in Japan, including CFS, drayage providers, ocean carriers, and customs partners. Lead the sourcing, screening, and hiring of the local Tokyo operations tea
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The opportunity: As Flexport embarks on our goals for 2026 and beyond, we will be rapidly expanding our presence in new markets. As Country Manager, you will build and lead our local teams, drive our market presence, oversee end-to-end operations, and cultivate partnerships with local partners to ensure seamless service delivery and growth. You will establish and infuse Flexport’s culture and values in daily operations, exhibiting grit and a bias to action. This role is pivotal to our success in expanding across LATAM and delivering exceptional value to our customers. You will: Lead and Manage Operations: Oversee all operational aspects of our business in the country, ensuring efficiency, compliance, and alignment with global standards. Develop and Execute Strategy: Build and implement local business strategies that align with regional and global objectives, driving growth and profitability. Foster Partnerships: Identify, negotiate, and maintain partnerships with local partners to optimize service offerings and expand our network. Manage and Grow the Team: Recruit, develop, and lead a high-performing local team, fostering a culture of excellence and accountability.
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The opportunity: As Flexport embarks on our goals for 2026 and beyond, we will be rapidly expanding our presence in new markets. As General Manager, you will build and lead our local teams, drive our market presence, oversee end-to-end operations, and cultivate partnerships with local partners to ensure seamless service delivery and growth. You will establish and infuse Flexport’s culture and values in daily operations, exhibiting grit and a bias to action. This role is pivotal to our success in expanding across Asia and delivering exceptional value to our customers. You will: Lead and Manage Operations : Oversee all operational aspects of our business in the country, ensuring efficiency, compliance, and alignment with global standards. Develop and Execute Strategy : Build and implement local business strategies that align with regional and global objectives, driving growth and profitability. Foster Partnerships : Identify, negotiate, and maintain partnerships with local partners to optimize service offerings and expand our network. Manage and Grow the Team : Recruit, develop, and lead a high-performing local team, fostering a culture of excellence and accountabili
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Engine Networking Team pulls the players together by ensuring the communication of the game state to all. As a Principal Engineer on this team you will help the players experience the game as a nearly synchronous world. The networking and asset loading team plays a key role in a smooth experience for the players. You will work in all areas of the game platform in your quest for real-time communication of every part of Roblox. You Will: Lead engineers with 8+ years of industry experience Be experienced with one of these area: asset loading, rendering, and networking coming from a Game Engine/Studio. Be an amazing systems-level C++ programmer and be fascinated by the actual work the CPU does when you use smart pointers, templates, virtual functions, and blocks of memory, both structured and raw Have a keen to each millisecond of the network exchanges: You know where the time goes and how to reduce the waste Understand what happens on the operating system level when certain code is completed You Have: Worked on the guts of a multi-player game engine, solving problems related to scale, performance, latency, and throughput in client/server environments. Worked on a very large multithreaded d
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Manager, Data Center Operations, you'll help us scale our Core Data Center and hardware infrastructure at a time of incredible growth for our business. At Roblox, you'll have boundless opportunities to shape the future of the Imagination Platform™ and demonstrate your passion for delivering thoughtful solutions in front of a global audience. If you know what it takes to build and operate hardware infrastructure that can sustain millions of concurrent players year-round and you take play as seriously as we do, you'll fit right into our highly experienced and ever-expanding engineering team. You will report to the Senior Manager of Data Center Operations. This will be a position based in Goodyear, AZ. You will: Develop and maintain the Core Data Center and hardware infrastructure to meet the large-scale and real-time requirements of our Imagination Platform™ to ensure our community has an awesome experience anywhere in the world. This includes all aspects of the server, network infrastructure, power, and environmental monitoring. Lead a growing team of data center engineers focusing on rack deployments, hardware troubleshooting and break-fix, and decommissioning. Identify and solve criti
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Threat Investigator on the Safety Investigations team, you will help lead Roblox’s deep-dive investigative capability for the most severe, complex, and externally significant safety matters. You will conduct actor-centric and network-centric investigations, proactively identify sophisticated threat actors and abuse patterns, and develop high-quality investigative work product that supports internal decision-making and, where appropriate, external law-enforcement engagement. The ideal candidate will have an exceptional investigations background, strong analytical tradecraft, and a demonstrated ability to connect fragmented internal and external signals into clear, defensible investigative findings. They will be an innovative self-starter, a collaborative partner across functions, and someone motivated by Roblox’s mission of connecting a billion people with optimism and civility. Through your work, you will become an expert in Roblox’s internal and external safety practices, advance our investigative capabilities and playbooks, and work across a variety of systems and data sources to surface, analyze, and ultimately disrupt high-risk actors, networks, and behaviors. Please note: T
Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Staff Offensive Security Engineer. About the job The Staff Engineer acts as a Technical Lead. You don't just find bugs; you design complex attack chains that demonstrate systemic risk. You spend as much time writing custom code and researching new bypasses as you do executing tests. Responsibilities In this role, you’ll: Full-Stack Penetration Testing: Perform manual and automated testing of web applications, APIs, and mobile apps (iOS/Android). Internal/External Network Audits: Conduct network and cloud level assessments with various tooling Vulnerability Validation: Triage and validate reports from automated scanners or bug bounty hunters to eliminate false positives and escalate true positives AI/LLM Probing: Perform initial prompt injection and jailbreak tests on AI prototypes, services, and applications using established checklists (OWASP Top 10 for LLMs). Technical Reporting: Draft high-quality reports that detail
Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Product Manager, Super Network About the job This position is needed to play a pivotal role in shaping, defining and executing our Twilio’s Super Network Data & Analytics product vision and strategy. You will collaborate closely with cross-functional teams, including engineering, design, and GTM to bring to life groundbreaking solutions that will shape the future of the global connectivity domain. You will lead the ideation, development, and launch of high-impact products, while also mentoring and guiding a team of product managers Responsibilities In this role, you’ll: Develop, communicate, generate excitement for and obtain buy-in of the product vision, strategy, and roadmap. Partner with cross functional peers to drive product planning, design, execution, and delivery of product investments against defined success metrics. Partner with Industry Relations & Strategic Growth teams on business develop
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Staff Backend Engineer We’re redefining Privileged Access Management (PAM) from the ground up, purpose-built for Cloud, SaaS, Databases, Containers, and any virtualized environment. Our mission is to simplify and secure workforce access with seamless, secure-by-default workflows that adapt dynamically to modern infrastructure. We eliminate standing privileges, enforce least privilege, and embed Zero Trust principles into every access workflow by default. About the Role We are seeking a Staff Backend Engineer to serve as the core technical anchor and senior Individual Contributor (IC) for our newly established engineering pod in India. At the P4 level, your primary sphere of influence will be at the team level —taking ownership of complex, ambiguous problems and defining how to solve them cleanly, securely, and efficiently. In this role, you will lead by example through hands-on architecture, high-velocity coding, and end-to-end execution. You will drive the implementation of secure database and network device connectors (routers, switches, firewalls) on top of our core Zero Standing Privileges (ZSP) platform. You will work closely with our local Technical Team Lead to elevate the pod’s engineering craft, acting as a technical multiplier for mid-level developers while ensuring tight architectural alignment with our global team. What You’ll Be Doing Execution & Technical Impact End-to-End Ownership: Consistently design, code, debug, test, moni
Get new lead network reliability engineer jobs by email
Daily job updates · Unsubscribe anytime