Jobiba hiring network

Network Reliability Engineer Jobs

1,954 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current network reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
1mo ago

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but operational in how we execute, and we support every product and research effort at OpenAI. Our tenets include prioritizing for impact, enabling researchers and developers, preparing for future transformative technologies, and fostering a strong, collaborative security culture. About the Role OpenAI is seeking a Principal Software Engineer to join the Infrastructure Security (InfraSec) team. InfraSec safeguards the core of OpenAI’s research and production environments: GPU supercomputing clusters, multi-cloud infrastructure, datacenters, networking, storage, and the critical services that power our frontier AI models. Our charter spans everything from bare-metal hardware and firmware to Kubernetes clusters, service meshes, and the data pathways that carry highly sensitive model weights and user data. As a Principal Software Engineer, you will set technical direction and drive execution of critical foundational services, such as authentication systems, egress/ingress proxies, access brokers, and key management platforms, that demand high standards of reliability, scalability, and software craftsmanship. These systems form the security backbone of OpenAI’s customer and supercomputing environment and must remain robust under intense scale and adversarial pressure. In this role, you will: Own the architecture and roadmap for one or more core security services (e.g., authN/Z, policy enforcement, secure proxies, key management), taking them from design to rollout to long-term operation. Design and implement planet-scale security systems that provide strong guarantees across hardware, operating systems, Kubernetes, networks, and CI/CD: balancing security, reliability, latency, and developer ergonomics. Lead cross-functional launches

awsazuregcp
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are building and maintaining a highly scalable asynchronous platform that empowers our organization to handle critical business cases. As a software engineering team, our mission is to create robust and innovative solutions that drive the success of our business and deliver unparalleled value to our customers. We adopt Infrastructure as Code practice to automate the provisioning and configuration of our resources, which helps reduce manual configuration and improve consistency. Our team culture is built on collaboration, open communication, and a supportive environment where each member's ideas are valued and contributions are recognized. We believe in the importance of fostering a positive workplace culture that inspires innovation and creativity. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and production readiness reviews, platform management and capacity planning ceremonies with cross-functional teams Document Infrastructure operations process and insights, identify repeatable actions and ruthlessly automate repetitive tasks Participate in our teams on-call rotations, respond to incidents and support other teams mitigate customer impacting events Experience: 5+ years experience working on teams responsible for software development, automation and systems engineering Experience building large-scale infrastructure, distributed systems or networks. Knowledge with SQS,

pythonawsazure
View job →
O
1mo ago

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but operational in how we execute, and we support every product and research effort at OpenAI. Our tenets include prioritizing for impact, enabling researchers and developers, preparing for future transformative technologies, and fostering a strong, collaborative security culture. About the Role OpenAI is seeking a Security Software Engineer to join the Infrastructure Security (InfraSec) team. InfraSec safeguards the core of OpenAI’s research and production environments—GPU supercomputing clusters, multi-cloud infrastructure, datacenters, networking, storage, and the critical services that power our frontier AI models. Our charter spans everything from bare-metal hardware and firmware to Kubernetes clusters, service meshes, and the data pathways that carry highly sensitive model weights and user data. As a Security Software Engineer, you will design and build critical foundational services, such as authentication systems, egress/ingress proxies, access brokers, and key management platforms, that demand high standards of reliability, scalability, and software craftsmanship. These systems form the security backbone of OpenAI’s supercomputing environment and must remain robust under intense scale and adversarial pressure. In this role, you will: Architect and implement production-grade security services (e.g., auth services, access brokers, secure proxies, key-management infrastructure) that provide strong guarantees across hardware, operating systems, Kubernetes, networks, and CI/CD. Partner with infrastructure and research engineers to embed security into high-performance compute clusters, enabling rapid model training and deployment without compromising protection. Develop automation and detection tooling to continuously identif

pythonawsazure
View job →
E(
Ema (Enterprise Machine Assistant)
📍 San Francisco Bay Area• Full-time• $135K – $225K/yr
1mo ago

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are We are seeking an experienced DevOps Engineer to join our growing team and play a pivotal role in designing and building our platform and infrastructure as we continue to scale our product and user base. As a part of our team, you will be working in a dynamic, fast-paced environment to ensure the reliability, scalability, and performance of our systems, while focusing on service architecture and deployment, query optimization, distributed systems, data and machine learning infrastructure, and security and authentication. Most importantly, you are excited to be part of a mission-oriented, fast-paced, high-growth startup that can create a lasting impact. You will: Partner with product teams to architect, design, and build the foundational infrastructure for our products. Design, develop, and deploy highly available and scalable Multi-tenant SaaS solutions on any one of the public cloud networks like AWS, Azure and GCP. Leverage technologies such as Kubernetes, Helm, Terraform, and Istio to achieve infrastructure resilience. Drive the automation of infrastructure tasks, from provisioning to configuration management and deployment, utilizing tools like Terraform, Ansible, a

awsazuregcp
View job →

Our Purpose Mastercard powers economies and empowers people in 200&#43; countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Manager, Software Engineering Overview We are seeking a hands-on Manager, Software Engineering to lead the development of scalable data platforms, backend systems, and data pipelines supporting Mastercard's Portfolio Intelligence products. This role will initially focus on leading a team of Data Engineering contractors while providing strong technical leadership across architecture, delivery, and engineering excellence. The ideal candidate is a recent practitioner who has built and operated modern data platforms and backend services, can confidently review architecture proposals and code, and leverages AI-powered engineering tools to improve productivity, quality, and delivery speed. As the team evolves, this leader will play a key role in building and developing a high-performing organization of full-time engineers. What You'll Do Technical Leadership Provide technical leadership for backend services, data pipelines, and platform capabilities that support analytics, reporting, and AI-driven products. Review architecture and design documents to ensure solutions are scalable, maintainable, secure, and aligned with long-term platform strategy. Conduct and oversee code reviews, promoting engineering best practices, reliability, security, and operational excellence. <br

pythonjavaaws
View job →
F
Fin
📍 Ireland• Full-time
1mo ago

Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences. Our AI Agent Fin is the highest-performing AI Customer Agent on the market today, enabling businesses to deliver impeccable, always-on customer support across the customer journey – from service, to sales, to ecommerce. Powered by our own AI models, Fin resolves complex customer issues end-to-end across every channel, with minimal set-up and integration. Fin can also be combined with our natively integrated Intercom help desk for one single system that is designed to meet the needs of modern day support teams. Founded in 2011, Fin became one of the fastest growing companies and remains one of the largest private software companies in the world with nearly 30,000 global businesses using our products to transform their customer support. Driven by our core values, we push boundaries, build with speed and intensity, and relentlessly deliver incredible value to our customers. What's the opportunity? Fin's Machine Learning team is responsible for defining new ML features, researching appropriate algorithms and technologies, and rapidly getting first prototypes in our customers’ hands. We are an extremely product focussed team. We work in partnership with Product and Design functions of teams we support. Our team's dedicated ML product engineers enable us to move to production fast, often shipping to beta in weeks after a successful offline test. We are very passionate about applying machine learning technology, and have productized everything from classic supervised models, to cutting-edge unsupervised clustering algorithms, to novel applications of transformer neural networks. We test and measure the real customer impact of each model we deploy. What will I be doing? Play an active role in hiring, mentoring and career development of other engineers Raise the bar for technical standards, performance, reliability, and operational excellence Identify areas

sqlrestmachine learning
View job →
F
Fin
📍 England• Full-time
1mo ago

Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences. Our AI Agent Fin is the highest-performing AI Customer Agent on the market today, enabling businesses to deliver impeccable, always-on customer support across the customer journey – from service, to sales, to ecommerce. Powered by our own AI models, Fin resolves complex customer issues end-to-end across every channel, with minimal set-up and integration. Fin can also be combined with our natively integrated Intercom help desk for one single system that is designed to meet the needs of modern day support teams. Founded in 2011, Fin became one of the fastest growing companies and remains one of the largest private software companies in the world with nearly 30,000 global businesses using our products to transform their customer support. Driven by our core values, we push boundaries, build with speed and intensity, and relentlessly deliver incredible value to our customers. What's the opportunity? Fin's Machine Learning team is responsible for defining new ML features, researching appropriate algorithms and technologies, and rapidly getting first prototypes in our customers’ hands. We are an extremely product focussed team. We work in partnership with Product and Design functions of teams we support. Our team's dedicated ML product engineers enable us to move to production fast, often shipping to beta in weeks after a successful offline test. We are very passionate about applying machine learning technology, and have productized everything from classic supervised models, to cutting-edge unsupervised clustering algorithms, to novel applications of transformer neural networks. We test and measure the real customer impact of each model we deploy. What will I be doing? Play an active role in hiring, mentoring and career development of other engineers Raise the bar for technical standards, performance, reliability, and operational excellence Identify areas

sqlrestmachine learning
View job →
AC
16 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Senior Consultant Location: Chennai, India | Work Model: In Office Appian Customer Success is obsessed with delivering exceptional customer outcomes and driving mission-critical business impact. Grounded in our core values of Excellence and Intensity, we act as elite technical advisors to our commercial clients. By joining this high-performance team, you will champion a culture of candid communication and excellence while accelerating global adoption of our AI-Powered Process Automation platform. What You’ll Do Implement Scalable Platform Architecture: Build and maintain standards and tools to ensure high performance and resilience with a focus on reusability, schema design, data modeling , query performance, and responsive UX/UI best practices. Engineered Integrations: Perform hands-on design, code implementation, and unit testing using SQL views , stored procedures , and high-concurrency REST/SOAP web services/APIs . System-Level Troubleshooting: Rapidly diagnose performance bottlenecks, memory usage, CPU constraints, and network latency spikes across the web application stack. Performance Profiling & Analysis: Analyze database execution plans, system logs, and Locust load test outputs to isolate and resolve root-cause performance bottlenecks. Technical Quality & Enablement: Ensure continuous integration and delivery alignment by applying strong DevSecOps practices, robust automated testing, and clean coding standards. Required Qualifications Bachelor’s Degree in Computer Science, Engineering, Information Systems, or Mathematics. 8+ years of pro

javascriptpythonjava
View job →
AG
1mo ago

This role is responsible to manage the continuous availability, reliability, and functionality of DCS and PLC systems. This includes scheduling and performing regular system backups, maintenance, and troubleshooting, as well as coordinating with OEMs for system upgrades. The role also involves managing hardware and software resources, network integrity, and cybersecurity measures to prevent data loss and system vulnerabilities. Source: Adani Group | Job ID: 44787

PE
Private Employer
📍 Bangalore, Karnataka• Full-time
1mo ago

About the Team ValMo is Meesho's logistics and supply chain organization, responsible for building India's most cost-efficient, reliable, and scalable fulfillment network. From first-mile pickup to last-mile delivery and reverse logistics, ValMo powers millions of shipments every month for small businesses across the country. As we scale, the complexity of our network increases exponentially. ValMo operates at the intersection of technology, operations, and partner ecosystems — where execution precision and systems thinking are critical to prevent cascading (domino) failures across the network. About the Role The Control Tower is where the stability, reliability, and cost-efficiency of ValMo's pan-India network is owned end-to-end. This is a deliberately generalist role in Fulfilment & Experience - your mandate is defined by an outcome, not a fixed function. Put simply: you are the person we point at the highest-leverage problem threatening the network, and you figure it out from first principles. The specific problem changes with what the network needs. In one quarter you might be redesigning the topology of a region; in the next, rebuilding the capacity-planning engine, running a structural cost-out program, firefighting an RTO or CX spike, or standing up readiness for a peak event - including problems that don't yet have an owner. What stays constant is the accountability: keep the network stable, predictable, and cost-efficient no matter what breaks or scales.

AG
Adani Group
📍 Mumbai• Full-time
1mo ago

Lead - SCADA is responsible for the operational integrity, maintenance, and security of the SCADA systems within assigned plant segments. This role ensures the smooth functioning of control systems essential for monitoring and managing automated processes, contributing to operational efficiency, safety, and productivity. By providing technical expertise in SCADA programming, network management, and system troubleshooting, the Associate enhances the reliability and security of plant operations. Additionally, they play a key role in supporting business development through technical advisement, compliance management, and assisting in the evaluation of new SCADA technologies, ultimately contributing to both the operational and strategic goals of the organization. Source: Adani Group | Job ID: 51756

P
Plaid
📍 San Francisco• Full-time• Remote
12 days ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. About the Team Our Network Enablement and Access team works to unlock the potential of Plaid's network by broadening and deepening our connections with data partners. We build the capabilities that help data providers participate in the network, strengthen the quality and reliability of those connections, and enable great products and experiences for Plaid's customers. Within Network Enablement and Access, the Data Supply Traffic and Health team owns how Plaid's requests flow to the data providers we depend on. We manage the load placed on each provider, the constraints that shape our access, and the fair allocation of capacity across Plaid's products and new initiatives. As paid access expands across the network, we also work to keep that traffic reliable, efficient, and cost-effective. As the Product Manager for Data Supply Health and Traffic, you will establish and lead a new product area at the foundation of every Plaid product. You will define how Plaid allocates constrained provider capacity, scales traffic across products, and manages the economics of paid data access. You will also optimize for data freshness, balancing timeliness with provider capacity and cost so Plaid's

P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s Account Verification team builds the foundation of trust for open finance. We help fintechs and financial institutions connect and verify bank accounts securely so that money can move safely and instantly. Account Verification is the entry point for all of Plaid’s payment-focused consumer experiences and is a critical building block for our customers. The team obsesses over creating seamless verification journeys that balance speed, reliability, and security, enabling consumers to confidently connect to the financial ecosystem. As a PM for Account Verification, you’ll own one of Plaid’s most critical and high-impact product areas. You’ll lead the evolution of our verification platform across Auth, Balance, and Identity Match, defining how millions of people and businesses connect their financial accounts every day. We are looking for a high-ownership builder who thrives in ambiguity, loves building with customers, and is excited to define what’s next for one of Plaid’s most established and strategically important product lines. You’ll set vision and strategy, drive execution across a cross-functional team, and shape how Plaid competes in an increasingly complex and, eventually, AI-driven ver

C
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Sweden About the Role You aren't just selling a product; you’re selling the security, performance, and reliability that major enterprises require to stay competitive. You will be a foundational part of our growth story in the region, bridging the gap between complex technical challenges and the business value that keeps our customers ahead of the curve. Why You’ll Love This Role This is a builder’s role. We’re looking for a hun

awsgitai
View job →

Our Purpose Mastercard powers economies and empowers people in 200&#43; countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Technical Program Manager Overview As a Senior Technical Program Manager at Mastercard, you’ll bring your expertise with conceptualizing, coordinating and driving technology projects, and coaching Agile practices in high-performing, self-organizing development teams. We’re building a global B2B, microservices-based platform to help businesses of all sizes streamline how they manage payments when buying or selling products & services. As a global business, the projects you lead for Mastercard will deliver software operating at massive scale requiring a focus on performance, security, and reliability. This role will support our Network Solutions team within Payment Networks, assisting the development team with building out software solutions for Mastercard. Role: • Dive as deep as you want into the tech stack, the integration patterns, the organizational capabilities, and the company wide assets that can be leveraged to provide technical solutions to customer problems. • Contribute to the strategies, design choices, and even the cloud infrastructure necessary to build comprehensive and achievable execution plans to deliver high-profile new features and capabilities for our customers. • Drive the execution of an initiative that may span multiple teams and integrations, reporting meani

airecruitment
View job →
🔔

Get new network reliability engineer jobs by email

Daily job updates · Unsubscribe anytime