Jobiba hiring network

Network Reliability Engineer Jobs

1,954 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current network reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

About the Team: OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. Our Stargate program develops and deploys massive, state-of-the-art data center campuses in partnership with industry leaders today—and through future OpenAI infrastructure projects tomorrow. We design for scale, speed, and reliability, and we need experienced technicians who can translate network blueprints into physical reality. About the Role: We are seeking a Senior Data Center Networking Technician who thrives in fast-moving build environments and is eager to roll up their sleeves during active datacenter deployments. Your first assignment will focus on the physical bring-up of network infrastructure at a large partner-operated campus, collaborating with partner teams and their delivery vendors to achieve agreed performance and reliability targets. As that campus reaches steady state, you will transition to lead network deployment for future OpenAI data center projects, defining standards and guiding implementation across multiple locations. Candidates must be able to sit onsite in Abilene, Texas 5 days per week Key Responsibilities Serve as OpenAI’s technical lead technician during the current campus build, partnering with internal engineers and external contractors on design reviews, installation plans, and acceptance criteria. Spend significant time on the data-center floor performing inspections, assisting with cable routing/termination when needed, conducting fiber testing (OTDR, power levels, continuity), and resolving installation challenges in real time. Troubleshoot and optimize cabling routes, patching, and equipment turn-up to ensure clean, reliable handoff to network operations. Contribute to design discussions and peer reviews for structured cabling and physical network layouts, providing practical field feedback to engineering teams. Develop repeatable engineering standards, as-built do

pythonawslinux
View job →

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect to join its NVIDIA Infrastructure Specialist Team. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer! What you'll be doing: Primary responsibilities will include building AI/HPC infrastructure for new and existing customers. Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, real-time monitoring, logging, and alerting. Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation, and refinement. Maintain services once they are live by measuring and monitoring availability, latency, and overall system health. Provide feedback to internal teams such as opening bugs, documenting workarounds, and suggesting improvements. What we need to see: BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields. At least 5+ years of professional experience in networking fundamentals, Ethernet or InfiniBand World. Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS, etc. Possess solid working knowl

pythonlinuxai
View job →
P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. We lead complex technical programs that help Plaid scale its engineering platform. We partner across Engineering, Infrastructure, Data, Security, ML, Legal, and Product to deliver company-wide technical initiatives that improve reliability, scalability, and developer productivity. You'll lead strategic technical programs from planning through execution. You'll partner with engineering leaders to align stakeholders, manage dependencies, drive decisions, and ensure successful delivery of complex initiatives. You'll work across a variety of technical domains, adapting quickly to new challenges and helping teams execute effectively. As a Technical Program Manager, you will lead high-impact, cross-functional initiatives. As a generalist, you may work on a variety of programs. An example is one that strengthens Plaid's data and machine learning platforms. You will partner with engineering, product, data, legal, privacy, and business stakeholders to drive complex technical programs from planning through execution. Your work will help improve data governance, modernize machine learning infrastructure, and accelerate the adoption of trusted, high-quality datasets that power analytics, artificial intelligence

awsmachine learningai
View job →
JT
Jobiba Technologies
📍 India• Full-time• From $35K/yr
1mo ago

MAINTENANCE HEAD Location: Kashipur, Uttrakhand Employment Type: Full Time CTC: 35,000 per month Minimum Experience: 5 years Hiring Timeline: Within 30 days About the Role We are looking for a hands-on Maintenance Head to oversee maintenance across a multi-location restaurant and food-service outlet network. The role will focus on preventive maintenance, equipment reliability and minimising operational disruption caused by breakdowns. Key Responsibilities • Manage HVAC and AC systems across multiple outlets. • Track servicing, repairs and issue resolution for AC/HVAC equipment. • Maintain refrigeration equipment, cold rooms, freezers and display units. • Handle basic plumbing and fitting requirements. • Carry out basic electrical and motor troubleshooting. • Establish and maintain preventive maintenance schedules across all outlets. • Minimise breakdowns through proactive maintenance rather than reactive firefighting. • Coordinate equipment refurbishment and cleaning cycles. Candidate Profile • Minimum 5 years of relevant maintenance experience. • ITI preferred/required, or equivalent hands-on experience in kitchen maintenance. • Strong hands-on experience in multi-site maintenance. • Knowledge of HVAC, refrigeration, plumbing and basic electrical systems. • Ability to troubleshoot equipment issues independently. • Experience in restaurant, hospitality, food-service, retail or similar multi-site operations preferred. • Strong preventive-maintenance orientation. • Good vendor-management and coordination skills. • Willingness to travel between outlets as required. Benefits: As per company policy. Skills: Core Technical & Facilities Engineering Skills • HVAC & AC Maintenance • Commercial Refrigeration Systems • Cold Room & Freezer Repair • Preventative Maintenance Scheduling • Breakdown Troubleshooting • Electrical Systems Vetting • Plumbing & Fittings • Motor Repair & Troubleshooting • Equipment Refurbishment Multi-Site Operations & Vendor Coordination • Mult

IE
10 days ago

About Inspira Education Inspira Education Group is one of the fastest-growing edtech startups in the US. We started with a simple mission to democratize access to high-quality coaching so that every student in the world has an equal opportunity to access the best opportunities. As the world’s leading network of top admissions coaches in medical, legal, business, and college studies, we’re building software and services in one place—disrupting long-entrenched application processes with products and experiences that strive to provide an equal platform for candidates from diverse backgrounds worldwide. As one of the fastest-growing edtech firms in the world, we are backed by some of the leading venture capital firms and investors in the world, including Zeev Ventures, Quiet Capital, Craft Ventures and Jeff Fluhr (Founder of Stubhub). About the role We’re looking for a strong full-stack engineer who can own the complete product development process: understand a business problem, define the solution, design the user experience, build the software, and improve it after launch. You’ll work closely with leadership and business teams, combining hands-on engineering with product management and design responsibilities. You should be highly effective with AI coding tools and have the technical depth to independently review, debug, secure, and maintain everything you ship. This is an in-person role requiring 5 day/week in our NYC office. What you’ll own Translate business needs and user feedback into product requirements, user flows, prototypes, and prioritized development plans. Design and build polished applications across the front end, back end, database, and integrations. Make architecture decisions and scope releases that balance speed, reliability, and future maintainability. Use AI tools throughout development to accelerate implementation, testing, debugging, and documentation. Own deployment, production monitoring, incident resolution, and ongoing improvemen

javascripttypescriptpython
View job →
G
GHX
📍 Hyderabad• Full-time
16 days ago

General Summary: The Services Portfolio Owner will have a technical focus and work with product managers, internal stakeholders, 3rd-party partners, and customers to gain a thorough understanding of their product and market needs, translating them into feature stories and tasks for development. You will be responsible for turning the product vision into an actionable backlog and advocating the customers’ needs to the development team. The Services Portfolio Owner is responsible for the development backlog, grooming, and acceptance. This individual will execute converting the product vision and roadmap into workable product requirements for engineering to meet product roadmap expectations defined by the Portfolio Manager. This role will report to the Manager, Services Portfolio and will regularly engage with cross-functional leaders. The ideal candidate must be well-versed in agile scrum and comfortable working in all aspects of a multifaceted delivery network. You will demonstrate excellent problem-solving skills, be self-motivated and detail-oriented, and be a good team player. Other duties include prioritization of key backlog items that drive growth and success, developing supporting documentation, and keeping customers and stakeholders informed of the status of the product. ROLES & RESPONSIBILITIES Owns and prioritizes the sprint backlog for assigned platform and technical domains, balancing technical debt, reliability, and roadmap commitments Translates platform and technical requirements into well-defined epics, user stories, technical specifications, and acceptance criteria Converts product vision into actionable backlog items that deliver platform enablement, reliability, scalability and technical debt reduction Participates in sprint planning, grooming, retrospectives and reviews for assigned technical and platform teams Works with Portfolio Management, Engineering leads, architects and security teams to support platform decision

agilescrumai
View job →

Cloud Infrastructure Administrator (Mid-Level, Senior or Lead) **Sign on Bonus Potential** Company: The Boeing Company The Boeing Company’s Specialized United States Infrastructure Operations organization is currently seeking a Cloud Infrastructure Administrator (Mid-Level, Senior or Lead) to join the team in Berkeley, MO; Seattle, WA; or Daytona Beach, FL . The Infrastructure team is seeking an experienced cloud infrastructure professional to help design, build, and sustain the foundational cloud environment supporting critical program needs. In this role, the selected candidate will help establish and operate secure, scalable, and resilient cloud infrastructure environments in Microsoft Azure to enable enterprise applications, software toolchains, and digital engineering workloads. As both an individual contributor and technical leader, this position will work across network, computer, storage, identity, security, and automation domains to deliver repeatable cloud infrastructure patterns and operational excellence. This role is focused on infrastructure operations, sustainment, automation, and reliability, rather than application software development. Position Responsibilities: Design, implement, and maintain Microsoft Azure-based infrastructure solutions including networking, compute, storage, identity integration, and supporting services Develop and maintain Infrastructure as Code (IaC) and configuration automation solutions using Terraform, Ansible, PowerShell, and Bash Implement cloud policies to enforce security, ensure regulatory compliance, and manage user access Build repeatable landing zones and cloud infrastructure patterns that support mul

azureterraformansible
View job →
N
1mo ago

We are seeking an experienced IT/Lab Manager to lead the planning, deployment, and operations of our physical lab environment and IT systems. This role will focus on building and maintaining scalable, reliable, and secure environments to support engineering teams involved in research, quality assurance, validation, and related activities. It will also support internal collaborators. You will have an outstanding opportunity to drive innovation in a multidimensional, technology-focused company that is crafting the future of data-center and lab technologies. If you bring perfection and creative thinking while solving issues as they arise, and enjoy working with distributed teams – your place is with us! What You’ll Be Doing: Own day-to-day operations, planning, and roadmap for the engineering lab and IT infrastructure (servers, storage, networking, and related services). Lead and mentor an IT/Lab team, driving guidelines, standards, and a culture of ownership, partnership, and continuous improvement. Collaborate closely with R&D, QE, Verification, and other engineering teams to design, provision, and maintain environments that meet their performance, reliability, and security needs. Lead all aspects of running data center and lab operations, including rack layout, cabling, power and cooling, hardware lifecycle, and resource availability. Lead procurement and vendor management for hardware, software, and services, including evaluation, negotiation, and ongoing relationship management. Implement and maintain automation for system provisioning, configuration, and operations using tools such as shell/Perl/Ansible. Design and maintain monitoring, logging, and alerting for servers, network, and storage systems to ensure high availability and rapid incident response. Investigate and resolve sophisticated infrastructure issues across OS, networking, storage, virtualization, and appli

kuberneteslinuxansible
View job →
P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s Product team builds the network that powers open finance. We unlock financial freedom by enabling innovation, reliability, and customer success. Our PMs are curious, fast-moving, and customer-obsessed, making intuitive, reliable experiences that help people thrive financially. As a Product Manager at Plaid, you’ll define and drive products that help millions connect, move, and manage their money. You’ll partner with engineering, design, and go-to-market teams to translate customer needs into impactful, high-quality products. This role suits someone early in their PM career who loves to learn, collaborate, and build meaningful solutions that improve financial lives. Responsibilities Own a product area: define problems, write clear requirements, and drive execution with cross-functional teams. Shape the roadmap using data, feedback, and market signals. Collaborate with Engineering and Design to deliver intuitive, high-quality experiences. Use metrics and user insights to measure and improve outcomes. Communicate clearly, align stakeholders, and guide decisions. Act with a founder mindset–simplify, move fast, and embrace feedback. Learn continuously and help raise the bar for Plaid’s products a

P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s Product team builds the network that powers open finance. We unlock financial freedom by enabling innovation, reliability, and customer success. Our PMs are curious, fast-moving, and customer-obsessed, making intuitive, reliable experiences that help people thrive financially. As a Product Manager at Plaid, you’ll define and drive products that help millions connect, move, and manage their money. You’ll partner with engineering, design, and go-to-market teams to translate customer needs into impactful, high-quality products. This role suits someone early in their PM career who loves to learn, collaborate, and build meaningful solutions that improve financial lives. Responsibilities Own a product area: define problems, write clear requirements, and drive execution with cross-functional teams. Shape the roadmap using data, feedback, and market signals. Collaborate with Engineering and Design to deliver intuitive, high-quality experiences. Use metrics and user insights to measure and improve outcomes. Communicate clearly, align stakeholders, and guide decisions. Act with a founder mindset–simplify, move fast, and embrace feedback. Learn continuously and help raise the bar for Plaid’s products a

P
Peloton
📍 New York• Full-time• From $244K/yr
1mo ago

ABOUT THE ROLE The Director of Supply Chain Strategy & Excellence is a high-impact leadership role responsible for architecting and executing a comprehensive end-to-end supply chain strategy. This position focuses on driving operational excellence, long-term scalability, and rigorous cost management across the enterprise. The successful candidate will use advanced SQL and Python to build rigorous data models, surface operational insights to craft a strategic vision, identify opportunities to deploy advanced AI-driven technologies to automate decision-making and improve predictive accuracy, and optimize the global network to enhance service levels and deliver significant bottom-line value. YOUR DAILY IMPACT AT PELOTON Develop and implement a multi-year supply chain roadmap aligned with corporate growth and fiscal objectives Network Design: Lead the design and optimization of the global distribution and logistics network to ensure maximum efficiency and customer satisfaction Serve as a key advisor to executive leadership on supply chain trends, risks, and investment opportunities Drive cross-functional alignment between Finance, Operations, Sales, and Marketing to ensure integrated business planning Targeted Savings: Establish and manage robust cost-reduction programs targeting delivery, service & repair, refurbishment, and transportation Cost of Quality (CoQ): Drive "Quality-by-Design" initiatives, incorporating member and engineering feedback to optimize product reliability and reduce total CoQ Oversee the supply chain budget, ensuring financial targets are met through rigorous expense management and productivity gains Implement total landed cost models to drive data-driven decision-making in sourcing and logistics Partner with analytics engineering to define and build dbt data models that translate exploratory supply chain analyses into reliable, reusable data infrastructure AI Deployment: Incorporate AI tools into analytical workflows and build AI-assisted

pythonsqlredis
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI is helping build the infrastructure that powers the next generation of artificial intelligence. Through Stargate, we are developing and operating large-scale AI compute campuses that require world-class execution across data center design, construction, commissioning, and operations. The Infrastructure Operations team is responsible for bringing AI infrastructure online and ensuring it operates reliably at scale. We partner closely with hardware, network, deployment, construction, and operations teams to deliver mission-critical environments capable of supporting frontier AI workloads. As our footprint expands, operational excellence becomes increasingly important to ensuring safe, reliable, and efficient campus operations. About the Role We are seeking a Facilities Operations Manager to support the commissioning, operational readiness, and long-term operation of next-generation AI data center campuses. This role sits at the intersection of construction, commissioning, hardware deployment, and facilities operations. You will be responsible for ensuring mission-critical infrastructure is prepared to support hardware deployment, transitioned successfully into production operations, and maintained to the highest standards of reliability and availability. You will lead day-to-day operational execution across electrical, mechanical, controls, and supporting infrastructure systems while partnering closely with commissioning teams, site operators, vendors, and engineering organizations. This role requires a strong blend of technical depth, operational leadership, and cross-functional execution. Key Responsibilities Lead day-to-day operations of mission-critical facility infrastructure across AI compute campuses. Own operational readiness activities supporting new campus deployments and infrastructure expansion. Partner with commissioning teams to transition facilities from construction and startup into steady-state operations. Develop, implement, and

awsrestai
View job →
S
1mo ago

Supabase is the open-source Postgres development platform that 7M+ developers and thousands of enterprises depend on every day. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the role We’re hiring a Product Manager to own the platform primitives every Supabase product runs on, including internal features like compute , disks , networking, and the API gateway, and external features like read replicas , custom domains , PrivateLink , and Bring Your Own Cloud. What you'll be responsible for: Talk to customers across the full spectrum. Indie developers running a single nano project, fast-growing startups whose costs are dominated by compute and disk, enterprises walking through a network-architecture review, and partners building on top of Supabase. Find the real blockers and bring them back to the roadmap. Own the problem statement and requirements behind every platform bet. Capture the customer evidence behind each decision, name the cost, capacity, and reliability constraints, and give the team a target it can hit. Decide what gets built, what gets deferred, and what gets cut. Every quarter you're choosing between enterprise unlocks blocking deals, reliability and cost wins for the long tail of projects, and net-new capabilities that change what Supabase can run. Set the priorities and defend them. Define how each launch is measured before it ships. Set the metric, agree on the threshold, and track it after launch. Know whether a feature moved enterprise deal velocity, project economics, or platform reliability. Use that to sharpen the next call. Keep engineering, design, and leadership aligned. The platform touches every other Supabase product, every region, and every customer tier. Write the roadmap, surface dependencies before they become blockers, and keep decisions moving. You might be a good fit if you: Have 7+ years of produ

aigorust
View job →
DU
16 days ago

About the Team DoorDash Labs, established in 2018, serves as the innovation hub for DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. The team's mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We’re ruthlessly focused on business impact. We are a highly senior team composed of former pioneers from a variety of different robotics industries. As of 2025, DoorDash has completed 10B lifetime deliveries. We’re focused on how to do the next 10B even better. About the Role We are seeking a highly motivated Senior Reliability & Test Engineer to join our team. This individual will play a key role in the development and validation of our unmanned platforms at the system and component levels. You will partner closely with EE, ME, and Autonomy teams to translate mission needs into robust, reliable hardware. The ideal candidate thrives in a fast-moving, cross-functional environment where reliability and test rigor determine program success. You will be hands-on in developing test methods and equipment to uncover failures before they happen in the field. You will partner closely with EE, ME, and Autonomy teams to translate mission needs into robust, reliable hardware. The ideal candidate thrives in a fast-moving, cross-functional environment where reliability and test rigor determine program success. You’re excited about this opportunity because you will… Architect and implement rigorous validation strategies, utilizing Python scripts for automation while leveraging CAD and shop tools to engineer bespoke test fixtures and hardware rigs. Oversee experimental execution across internal facilities and external laboratories, maintaining technical mastery over vibration tables, environmental chambers, DAQ systems, and ingress protection testing. Translate high-level vehicle reliability requirements into granula

pythonawsgit
View job →
DU
16 days ago

About the Team DoorDash Labs, established in 2018, serves as the innovation hub for DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. The team's mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We’re ruthlessly focused on business impact. We are a highly senior team composed of former pioneers from a variety of different robotics industries. As of 2025, DoorDash has completed 10B lifetime deliveries. We’re focused on how to do the next 10B even better. About the Role We are seeking a highly motivated Senior/Staff Test Engineer to join our team. This individual will play a key role in the development and validation of our unmanned platforms at the system and component levels. The ideal candidate has a strong background in test development, test execution, and root cause analysis with a proven track record of collaboratively managing risk throughout a fast paced development process. You’re excited about this opportunity because you will… Run and monitor tests within our facility as well as at outside test labs. Collaborate with a tight knit team to identify and understand test failures. Be hands-on in developing test methods and equipment to uncover failures before they happen in the field. Find clarity through root cause analysis of lab and field failures and suggest design changes to prevent them. Use your creativity to create novel and scaled tests for autonomous systems. We’re excited about you because you have… A bachelors or advanced degree in a relevant engineering discipline. Mastery of test equipment such as environmental chambers, vibration tables, water testers, DAQs, etc.. Ability to bring order to complex test and development programs via clear technical communication and documentation. Experience designing and building testers and equipment. Ability to write Python scripts to automate

pythonawsgit
View job →
🔔

Get new network reliability engineer jobs by email

Daily job updates · Unsubscribe anytime