About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations: Austin, TX What You’ll Do The Argo team was formed to own a very important aspect of Cloudflare's systems: enable more reliable network connectivity for Cloudflare’s products than the Internet itself provides. Almost all products in Cloudflare’s portfolio are or will be powered by Argo technology, including CDN, Spectrum, Magic Transit, Stream, Workers, Workers AI, R2, WARP, and more. As a member of the Argo team, you’ll
Jobiba hiring network
Network Reliability Engineer Jobs
1,954 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current network reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Austin, US Responsibilities The Argo team was formed to own a very important aspect of Cloudflare's systems: enable more reliable network connectivity for Cloudflare’s products than the Internet itself provides. Almost all products in Cloudflare’s portfolio are or will be powered by Argo technology, including CDN, Spectrum, Magic Transit, Stream, Workers, Workers AI, R2, WARP, and more. As a member of the Argo team, you’ll be a
About the Team OpenAI’s Infrastructure Operations team is responsible for the availability, reliability, and operational excellence of one of the world’s largest AI infrastructure networks. The team owns day-to-day operations of production AI networks across Industrial Compute's data centers, working with colocation providers, deployment teams, and hardware vendors to deliver highly available GPU infrastructure for AI training and inference workloads. About the Role We are seeking an Infrastructure Operations Engineer to operate and improve the large-scale Ethernet fabrics that support GPU clusters, storage systems, and management infrastructure. This role combines hands-on production operations with automation, observability, and incident response across a global AI network. The ideal candidate has experience operating high-availability data center, cloud, AI, or HPC networks and can move comfortably from physical-layer troubleshooting to routing and fabric behavior, change execution, and root-cause analysis. You will partner closely with network architecture, systems engineering, GPU engineering, storage engineering, security, deployment, site operations, service providers, colocation partners, and hardware vendors to raise reliability and reduce operational toil. Key Responsibilities Own the operational health, availability, and reliability of production AI network infrastructure across Industrial Compute's data centers. Monitor, troubleshoot, and resolve network incidents while meeting service-level objectives (SLOs), reducing Mean Time to Detect (MTTD), and minimizing Mean Time to Recovery (MTTR). Operate and maintain large-scale Ethernet fabrics supporting GPU compute, storage, and management networks. Execute production network changes, maintenance windows, and capacity expansions with minimal customer impact. Manage the hardware lifecycle, including switch and optics replacements, RMA coordination, software upgrades, and preventive maintenance. Support new A
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an IT Network Engineer at Micron, you will be responsible for the operational health (Security, Availability, Performance, Interoperability and Reliability) of Micron’s data communications systems. You will support the acquisition, installation, maintenance, and administration of Micron’s data and voice communications systems. Network Engineers specialize in LAN/WAN, Wireless, and Telecommunications hardware, software and cabling. You will regularly work with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. Responsibility Collaborate with various teams and functional areas in developing and improving business processes with clear and agreed-upon business rules Develop process improvement strategies across the network Prioritize and manage multiple projects based on Department and corporate objectives Lead and collaborate with various engineering groups to optimize current processes and align the network to improve manufacturability, yield, quality and cost performance of products Work with process integration team to implement and integrate new/improved process workflow Maintain documentation for new and improved processes Communicate project strategy, status and actions through meetings, presentations and reports to
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an IT Network Engineer at Micron, you will be responsible for the operational health (Security, Availability, Performance, Interoperability and Reliability) of Micron’s data communications systems. You will support the acquisition, installation, maintenance, and administration of Micron’s data and voice communications systems. Network Engineers specialize in LAN/WAN, Wireless, and Telecommunications hardware, software and cabling. You will regularly work with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. Responsibility Collaborate with various teams and functional areas in developing and improving business processes with clear and agreed-upon business rules Develop process improvement strategies across the network Prioritize and manage multiple projects based on Department and corporate objectives Lead and collaborate with various engineering groups to optimize current processes and align the network to improve manufacturability, yield, quality and cost performance of products Work with process integration team to implement and integrate new/improved process workflow Maintain documentation for new and improved processes Communicate project strategy, status and actions through meetings, presentations and reports to
About the Role Join Peloton’s Global Network Services team as a Senior Engineer, sitting at the unique intersection of high-scale enterprise networking and high-stakes media production. In this role, you will architect and support the global infrastructure that powers our corporate offices, warehouses, and flagship New York Broadcast Studio. Acting as the key bridge between Global Cloud Engineering and Studio Operations, you will deploy highly resilient, scalable architectures that deliver live streaming content seamlessly to millions of members worldwide. Your Daily Impact Global Architecture & Deployment: Design, optimize, and secure Peloton's global network footprint, integrating Cisco, Meraki, Aruba, and Palo Alto Networks across hybrid on-premises and AWS environments. Studio & Broadcast Reliability: Lead deep-dive traffic analysis and troubleshooting for our NY studio, ensuring 24/7 uptime for live broadcasts, real-time video streaming, and OTT media delivery. Operational Leadership & Lifecycle Management: Manage the end-to-end network project lifecycle—from traffic shaping and SD-WAN optimization to establishing SOPs, security policies (with InfoSec), and handling Tier-3 disaster recovery. You Bring To Peloton Broad & Deep Network Expertise: 8+ years in Network Engineering, including 6+ years in complex SaaS environments and 4+ years designing public cloud networking (specifically AWS). Advanced Protocol & Security Mastery: Expert knowledge of L2/L3 protocols (BGP, OSPF, EIGRP), security protocols (IPsec, 802.1x, RADIUS), SD-WAN, and SDN/SDDC full-stack solutions. Media & Streaming Specialization: 3+ years optimizing IP networks specifically for live video delivery, utilizing multicast/unicast technologies and streaming protocols like HLS, RTMP, and SRT. Education & Elite Certifications: A Bachelor’s degree in Engineering or Computer Science, backed by active CCIE or JNCIE certifications (required). Collaborative Mindset: A curious
About the Team OpenAI’s Network Engineering team within IT and Security advances the mission of deploying artificial general intelligence (AGI) for the benefit of all by delivering secure, scalable, and resilient network services. We build and operate the connectivity that supports OpenAI’s offices, labs, campuses, cloud environments, people, and devices. By combining strong network fundamentals with security, reliability, automation, and user-centered design, we enable impactful AI research, corporate operations, and product innovation. About the Role As a Network Engineer at OpenAI, you will design, operate, and continuously improve the global networks that connect our offices, labs, campuses, PoPs, cloud environments, people, and devices. The role spans strategic platform engineering and responsive production operations: you will shape architecture, standards, roadmaps, lifecycle plans, and automation while supporting incidents, escalations, and time-sensitive delivery. Operational signals will inform what we stabilize, simplify, standardize, or automate next. We work backward from user needs, investigate root causes, own outcomes end-to-end, and move quickly without compromising security. We are looking for a versatile engineer who can make pragmatic reliability and security tradeoffs, communicate clearly, and turn recurring operational work into durable platforms, tooling, and standards. You will partner across IT, Security, AppEng, Research, Applied, workplace teams, carriers, and vendors. In this role, you will: Design, implement, and operate secure, scalable enterprise networks across offices, labs, campuses, PoPs, cloud connectivity, and hybrid environments. Set strategic direction for network services through architecture, standards, roadmaps, lifecycle planning, capacity strategy, and measurable reliability outcomes. Own production operations, including on-call, incident response, escalations, and time-sensitive delivery, while protecting user experience,
About the team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the role: We are seeking an experienced Optical Network Engineer to lead Laser related work within our optical interconnect efforts for large-scale compute systems. The role also requires broad, hands-on optical validation experience across IM/DD-based interconnects, working from lab characterization through production readiness and scaled deployment. In this role you will: Drive laser-focused requirements and technical direction within the broader optical interconnect roadmap. Lead evaluation and validation of optical components and subsystems, including laser-based elements, in lab and production-representative environments. Support end-to-end optical testing for IM/DD interconnects (e.g., module/system bring-up, characterization, debug, and readiness for scale). Work with external partners to align on development milestones, performance targets, and quality expectations. Own technical issue triage and resolution across performance, reliability, and manufacturability topics. Collaborate across internal teams to support integration, rollout, and operational success at scale. You might thrive in this role if you have: Strong experience in laser-focused optical engineering (development, validation, manufacturing readiness, or field support). Broad hands-on background with IM/DD optical technologies and optical test/debug workflows. Experience working with external suppliers/manufacturing partners and production-oriented execution. Demonstrated ability to debug complex t
About Us: Paytm is India's leading mobile payments and financial services distribution company. Pioneer of the mobile QR payments revolution in India, Paytm builds technologies that help small businesses with payments and commerce. Paytm’s mission is to serve half a billion Indians and bring them to the mainstream economy with the help of technology. About Role: We are seeking an experienced L2 Network & Security Engineer to join our team. As an integral part of our network operations, you will play a crucial role in maintaining and securing our infrastructure. If you have a passion for networking, security, and troubleshooting, we’d love to hear from you! Job Location: Noida Responsibilities: Install and Support: Deploy and maintain LANs, WANs, and network segments. Support Coordination: Collaborate with equipment vendors for troubleshooting and configuration standardization. Firewall Management: Handle firewall configurations to ensure security and optimal performance. Switch Expertise: Maintain a strong understanding of LAN switching technologies, including VLANs, RSTP, ACLs, and Virtual Chassis. Capacity Planning: Recommend network capacity planning strategies. Proactive Maintenance: Plan and execute proactive maintenance for disaster recovery. Network Monitoring: Monitor networks for security, reliability, and availability. Key Skills Required: He/She/They should have 3-6 Yrs. of overall IT experience. Firewalls: Proficient in configuring, managing, and troubleshooting FortiGate and Palo-Alto firewall. Network Switches: Skilled in handling Cisco, Juniper, and Huawei switches. Wireless: Familiarity with Aruba, Cisco controllers and access points. Cloud Expertise: Experience with AWS and Zscaler Private Access. Network Protocols: Sound knowledge of various network protocols and ports. Security Technologies: Familiarity with IPSec, SSL VPN, IDS, and IPS. Additional Skills: Routing and Switching: Prior experience in routing and switching. Communication: S
Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team The Network Security team is responsible for securing GoDaddy's global hybrid infrastructure across data centres, cloud, edge, and remote-access environments. We partner across Security, Infrastructure, Cloud, and Engineering teams to build scalable, resilient, and secure solutions that support the business. As a Principal Network Security Engineer, you'll act as a senior technical leader, helping define network security strategy, influence architecture across teams, and drive security outcomes at enterprise scale through technical expertise, systems thinking, and cross-functional leadership. What you'll get to do... Define and drive network security architecture across hybrid environments, including data centres, cloud, edge, and remote-access technologies Design trust boundaries, segmentation strategies, secure connectivity patterns, and network controls that reduce risk and improve security posture Lead complex technical initiatives, migrations, and architectural decisions while balancing security, reliability, performance, and operational requirements Establish scalable approaches for policy governance, automation, monitoring, telemetry, and security control validation Partner across engineering organizations to drive large-scale initiatives, mentor engineers, and influence technical direction through architecture reviews and technical leadership Your experience should include... 10+ years of experience in Network Security Engineering, Network Architecture, or Security Engineering, including ownership of enterprise-scale secu
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Software Engineer on the Blockchain Networks team within the Platform group, you'll build the critical infrastructure that integrates blockchain protocols with Coinbase's internal services for new assets, stablecoins, and staking. This team translates complex blockchain operations and state machines into simple, reliable APIs that product teams across Coinbase depend on. You'll own multi-quarter technical initiatives and ship platform primitives that directly improve the latency, reliability, and cost of our crypto stack. What you'll do: Own the design and delivery of blockchain network infrastructure that abstracts blockchain complexity into reliable, platform APIs Lead multi-quarter initiatives including new chain integrations, re-architecture efforts, and data migrations that improve system latency, reliability, and cost Define and maintain APIs, SLOs, and observability for the systems you build, including on-call ownership Partner with product and platform teams to establish data contracts and ship SDKs and platform primitives that drive adoption across Coinbase Build deep technical context across multiple top blockchain protocols (e.g., Bitcoin, Ethereum) and apply that knowledge to simplify cross-chain operations Required Skills and Experience: 5+ years of software engineering experience, with demonstrated ownership of production services in a servi
Join the engineering teams that bring OpenAI’s ideas safely to the world!! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role As OpenAI continues to grow, we are looking for experienced, problem-solving engineers to ensure our systems scale. Our success depends on our ability to quickly iterate on products while also ensuring that they are performant and reliable. You will work in a deeply iterative, collaborative, fast-paced environment to bring our technology to millions of users around the world, and ensure it’s delivered with safety and reliability in mind. Successful candidates will play a crucial role in ensuring the reliability, scalability, and performance of our systems as we continue to expand. As a reliability expert, you will be at the forefront of maintaining and enhancing the stability, scalability, and performance of our rapidly evolving infrastructure. You will work closely with cross-functional teams, including software engineers, product managers, and data scientists, to build and maintain resilient systems that can handle our growing user base and workload. In this role, you will: Design and implement solutions to ensure the scalability of our infrastructure to meet rapidly increasing demands. Build and maintain the load, chaos and synthetic testing software leveraged by development teams to make the systems they design and operate more reliable. Build and maintain automation tools to streamline repetitive tasks and improve system reliability. Build and maintain the platform for CPU/storage, GPU, and network lifecycle management to drive efficiency, accountability and support dynamic optimization of our resources. Implement fault-tolerant and resilient
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Sr. Staff Network Engineer to join our team. This is a fully remote capacity in the Netherlands role, reporting to the Manager, Network Engineering in the Cloud Ops - Network Engineering department. This position involves a mix of infrastructure, project management, and network engineering responsibilities. You will join our global team as an integral member, taking ownership of the deployment, monitoring, and ongoing operation of our worldwide production infrastructure across diverse data center locations. We are seeking a high-trust collaborator with a background in large-scale enterprise or telecom networks who approaches challenges with a growth mindset and a commitment to achieving engineering excellence. This role requires active participation in our technical on-call rotations, which include support during weekends and holidays to ensure continuous service reliability. What you’ll d
Staff Software Engineer Bengaluru, Karnataka, India Opportunity Get Well is seeking a visionary and technically adept Staff Software Engineer to architect, design, develop, and optimize our cloud-native healthcare platform while driving the adoption of AI-First and AI-Augmented Software Engineering practices. This role is pivotal in shaping the future of software development at Get Well by combining deep technical expertise with modern AI-assisted engineering workflows. As we evolve toward an AI-First engineering organization, this leader will champion the use of Generative AI, AI development assistants, and Agentic AI to improve developer productivity, software quality, and engineering velocity. The ideal candidate brings deep expertise in software architecture, cloud-native application development, AI-enabled engineering, and distributed systems. This role provides technical leadership across multiple engineering teams, ensuring high standards for architecture, code quality, reliability, security, and AI adoption. This is a hands-on leadership role where strategic thinking meets deep engineering execution within a complex healthcare environment. This position reports to the Director, Product Development and requires close collaboration with software engineers, AI engineers, product managers, DevOps, QA, and compliance specialists. Key Responsibilities Technical Leadership Define and drive the architecture of scalable, distributed healthcare platforms. Champion AI-First Software Engineering practices across the development lifecycle. Lead the adoption of AI-Augmented Development , including spec-driven development, AI-assisted coding, code reviews, testing, and documentation. Establish engineering standards, Cursor/AI coding guidelines, reusable patterns, and governance for responsible AI usage. Provide hands-on leadership in architecture, coding, design reviews, debugging, and performance optimization. Mentor enginee
Job Details: Job Description: As one of the world's largest semiconductor manufacturers, Intel is committed to advancing every aspect of semiconductor technology, from process development and manufacturing to advanced packaging and reliability characterization. Employees within Intel Foundry are part of a global network spanning technology development, manufacturing, assembly, test, and quality organizations across both front-end silicon and advanced packaging facilities. You will join the Foundry Lab Network (FLN), a key organization within Foundry Quality, Reliability, and Labs (FQRL), located at Intel's growing Chandler, Arizona site. This site serves as the technology development hub for Intel's most advanced packaging technologies, including Hybrid Bond Interconnect (HBI), Embedded Multi-die Interconnect Bridge (EMIB-T), glass substrates, Foveros, and future advanced packaging innovations. FLN is actively expanding its laboratory capabilities and capacity to support Intel Foundry's roadmap, enabling faster technology qualification and accelerating development cycles through Quick Turn Monitoring (QTM) solutions that significantly reduce time-to-data. As the Stress Test and Reliability (STaR) Back-End Laboratory Manager, you will play a critical leadership role within FLN. You will partner closely with Packaging Technology Development, Quality and Reliability, and Failure Analysis teams to establish and enhance laboratory capabilities that support qualification and certification of next-generation packaging technologies and products. This position offers a unique combination of people leadership, technical problem solving, and organizational strategy. You will lead a team of 7-10 talented engineers, spearhead cross-functional efforts to resolve complex reliability and technology challenges, drive strong quality and operati
Get new network reliability engineer jobs by email
Daily job updates · Unsubscribe anytime