At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is expanding the boundaries of the Data Cloud to support mission-critical transactional workloads. Our goal is to deliver OLTP capabilities with the performance, reliability, simplicity, and scale customers expect from Snowflake, while creating a seamless experience across transactional and analytical data. We are looking for a Senior Engineering Manager – OLTP to lead engineering teams and leaders building core transactional database technology and the cloud infrastructure required to operate it at scale. You will help define the architecture and roadmap, grow the organization, and drive technology from design through production. AS A SENIOR ENGINEERING MANAGER – OLTP AT SNOWFLAKE, YOU WILL: Set technical and execution strategy for key areas of Snowflake's OLTP platform, translating product goals into architecture, roadmaps, and team plans. Lead and grow multiple engineering teams, developing managers and senior technical leaders while fostering a culture of ownership, technical excellence, and execution. Drive adoption of AI and agentic development practices to improve engineering velocity, quality, and productivity across the software development lifecycle. Drive architecture and technical decisions in areas such as transactions, concurrency control, low-latenc
Jobiba hiring network
Senior Infrastructure Architect Jobs
7,292 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current senior infrastructure architect jobs. Use filters to narrow by work mode, employment type, experience and date posted.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer, Snowflake Natsec Running Snowflake in public sectors in different countries and regions, even in different industry verticals, requires us to build a compliant, secure, and auditable infrastructure. Many key design decisions are deeply rooted in the Snowflake product architecture. As a Senior Software Engineer, you will be responsible for leading several key areas and collaborating with various engineering groups in addition to the Public Sector team. To be successful in the area, you will need to have (and continue to build) a broad and in-depth knowledge base on cloud infrastructure, privacy, and governance, compliance controls, data security and data residency in various aspects of Snowflake. AS A SENIOR SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business needs at large scale by applying your software engineering and analytical problem solving skills. Design, implement and maintain scalable distributed systems for our cloud automation platform that include cloud control plane, Kubernetes container platform and traffic and networking. Work directly with customers to quickly understand their critical problems and design and implement solutions Deploy and maintain availability of cloud compute servers and Kubernetes cluster that power the
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer, Snowflake Natsec Running Snowflake in public sectors in different countries and regions, even in different industry verticals, requires us to build a compliant, secure, and auditable infrastructure. Many key design decisions are deeply rooted in the Snowflake product architecture. As a Senior Software Engineer, you will be responsible for leading several key areas and collaborating with various engineering groups in addition to the Public Sector team. To be successful in the area, you will need to have (and continue to build) a broad and in-depth knowledge base on cloud infrastructure, privacy, and governance, compliance controls, data security and data residency in various aspects of Snowflake. AS A SENIOR SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business needs at large scale by applying your software engineering and analytical problem solving skills. Design, implement and maintain scalable distributed systems for our cloud automation platform that include cloud control plane, Kubernetes container platform and traffic and networking. Work directly with customers to quickly understand their critical problems and design and implement solutions Deploy and maintain availability of cloud compute servers and Kubernetes cluster that power the
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the Role We are hiring a Senior Software Engineer for our Data Platform team - the team that owns the core streaming execution and database internals powering Snowflake's platform. You'll work deep in the stack: stream processing runtimes, query execution engines, storage layers, and the distributed systems primitive that everything else is built on. AS A SOFTWARE ENGINEER, DATA PLATFORM AT SNOWFLAKE, YOU WILL: Design and implement high-throughput stream processing systems with strong correctness guarantees - exactly-once semantics, out-of-order event handling, low-latency execution Work on database internals - query execution, operator design, storage engine components, or transaction processing at cloud scale Own features and systems end-to-end: design, implementation, testing, deployment, and production observability Drive architectural decisions that shape the long-term evolution of Snowflake's streaming and core platform layers Mentor engineers through design reviews, code reviews, and technical guidance Partner across product, infrastructure, and research teams to deliver high-impact, foundational capabilities OUR IDEAL CANDIDATE WILL HAVE: 5+ years of software engineering experience in distributed systems, query engines, streaming infrastructure, or database in
Shape the Future with Dun & Bradstreet At Dun & Bradstreet, we believe data has the power to create a better tomorrow. As a global leader in business decisioning data and analytics, we help companies worldwide grow, manage risk, and innovate. For over 180 years, businesses have trusted us to turn uncertainty into opportunity. We’re a diverse, global team that values creativity, collaboration, and bold ideas. Are you ready to make an impact and help shape what’s next? Join us! Explore opportunities at dnb.com/careers. Sr. Data Cloud Engineer(s)(multiple positions) – Duties are using relational database systems, CI/CD pipeline, Kubernetes for application deployments, terraform IAC, Observability using Splunk, server-side GitHub REST API, & Snowflake, Spark, & Python to design, automate, implement & maintain data architecture patterns in AWS & GCP for data collection infrastructure & pipelines for multi-region cloud-based big data IaaS & PaaS solutions including performing requirements & document use technology analysis; coordinating data integration & ETL processes; supporting data security, integrity & cost controls; automating data processes utilizing Cloud Composer & develop SQL scripts for data modeling; performing end-to-end pipeline load performance testing; supporting technology upgrades through PostgreSQL & Cloud SQL cloud migrations; performing SQL & NoSQL database tuning & maintenance; providing guidance on data management & SQL optimization; & developing data warehousing management policies. Requires Bach degree in Comp Science, Comp Engineering, Information Technology (IT) or related field & 5 yrs exp in job duties as stated. Position is with Dun & Bradstreet in Jacksonville, FL. Inquire and send resume through Dun & Bradstreet’s job board at https://jobs.lever.co/dnb. Position is under Sr. Data Cloud Engineer for Jacksonville, FL. #LI-DNI
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Transformation Architect to join our team. This is a fully remote role, reporting to the Solution Consulting Director in the Solutions Consulting department. Transformation architects work at the cutting-edge of Zscaler’s platform to create a multi-year cybersecurity vision together with our largest customers. This role is strongly customer-oriented, requiring engagement at mid-senior to practitioner levels within organizations while rewarding individuals who relish in-depth analysis of both business imperatives and technology infrastructure. Additionally, this role encourages lateral and creative thinking to guide customers drawn across all verticals towards a more secure, simplified future while constantly evangelizing the benefits of a true zero trust approach along the way. What you’ll do (Role Expectations) Drive Architecture Value by demonstrating the strategic value and ROI of platform of
Forward Deployed Senior Software Engineer (Migration Tooling – RunMyJobs) OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full-stack automation fabric solutions for mission-critical business processes. Our flagship SaaS platform, RunMyJobs (RMJ) , is the first composable automation platform specifically built for ERP environments. We enable organizations to orchestrate, manage, and monitor workflows across applications, services, and infrastructure — in the cloud or on premises. Our global team of automation experts, engineers, and customer success professionals work together to deliver seamless automation transformations. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are seeking a Forward Deployed Software Engineer focused on migration tooling and customer onboarding to RunMyJobs (RMJ) . This role is part of an exciting new Forward Deployed Engineering team within the Global Professional Services team, at the intersection of Product Engineering and Go To Market teams. In this role, you will operate at the intersection of engineering and delivery. Your primary focus will be designing, building, and enhancing migration frameworks, tooling, and automation accelerators that enable customers to smoothly transition from legacy schedulers and automation platforms into RMJ. You will work closely with: Migration Architects to design scalable and reusable migration patterns Professional Services to enable efficient customer onboarding Engineering & Product to improve platform capabilities based on field learnings Customers (occasionally) to validate requirements, troubleshoot edge cases, and ensure successful implementations This is a forward-deployed engineering role — highly technical, impact-driv
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Team The NR Lens team builds New Relic's Federated Data Platform — a distributed SQL query engine that lets customers connect external data sources (Snowflake, PostgreSQL, Google Sheets, AWS CloudWatch) and query them directly from New Relic. You'll work on the distributed query execution layer, connector architecture, and API gateway that powers cross-source JOINs and unified analytics across customer data stores. What You'll Do Design, build, and maintain cloud-native Java microservices in the NR Lens query path: SQL Gateway, Query Gateway, and data source connector plugins Develop and harden JDBC connector integrations — including connection lifecycle management, credential handling, query pushdown optimization, and security validation Improve query reliability and performance across a multi-tenant distributed SQL deployment serving 50+ customers Build and operate services on AWS (EKS, IAM/STS, S3) with infrastructure-as-code practices Implement
Reolink , a leader in intelligent visual technology for homes and businesses, was founded in 2009 by a group of engineers with a strong commitment to and passion for smarter security solutions. Our products are now trusted by millions of users across more than 110 countries and regions worldwide. Building on this trust, we continue expanding our presence and bringing our innovations to more markets around the globe. Reolink remains committed to delivering advanced, reliable, and user‑centric solutions that empower people to protect what matters most. 5 Work Days Per Week Office at Tai Seng Exchange Tower B Near Tai Seng MRT, Singapore Insurance Coverage Entitled to Yearly Bonus & Performance Bonus Responsibilities (Site Reliability Engineer - Senior / Lead ) High Availability and Stability Maintenance of Application Systems: Includes daily monitoring, alert response, emergency handling, on-call duties, regular system health checks, and performance optimization. Compliance and Secure Access Construction for Application Systems: Ensure operational design, processes, and data management comply with relevant privacy and data protection laws. Ensure compliance with full auditing and regulatory checks and provide auditing materials as required. Change and Release Management: Best practices for application system changes, including change control, version management, and rollback strategies, while ensuring operational duties during release windows. Automation and Infrastructure Optimization: Drive operational automation by designing and implementing automated tools and processes, ensuring resource allocation is optimized and supporting business scalability. Other Operational Practices and Work Arrangements: Providefeedback and suggestions for business architecture design and continuously produce operational technical documentation. Qualifications Bachelor's Degree or above; a degree in computer science or a related field is preferred. Experiences as Senior SRE or
The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our NYC HQ, our smaller Austin, Palo Alto, or San Francisco offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design prim
The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azur
About the Team The OpenAI Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role As a Senior Software Engineer, ML Systems & Training Infrastructure, you will be a deeply hands-on engineering force multiplier for the robotics team. You will help keep the training framework and surrounding infrastructure healthy, review and improve code quickly, debug failures across ML systems and infrastructure, and unblock researchers and engineers when the path from idea to working training job gets rough. We’re looking for people who love writing, reading, reviewing, and fixing code; who can get productive quickly in unfamiliar systems; and who bring strong practical judgment without a lot of ego or process overhead. This role will be based in San Francisco, CA and be expected in office 5 days per week and offer relocation assistance to new employees. In this role, you will: Review, improve, and clean up code across training frameworks and adjacent infrastructure. Identify risky or low-quality changes before they land, and raise the code quality bar without slowing the team down. Debug issues across ML training systems, GPUs, clusters, networking, and related infrastructure. Help researchers and engineers unblock broken training jobs, flaky workflows, and brittle internal tooling. Improve the reliability, maintainability, and usability of the robotics team’s training framework. Move quickly on practical engineering problems that directly affect team velocity. You might thrive in this role if you: Have strong software engineering fundamentals and excellent code review judgment. Have experience with ML systems, training fr
About Backblaze Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our partners, we're helping customers break free from the restrictive, overpriced legacy solutions that hold them back and blaze forward with the full power of the open cloud in their hands. Founded in 2007, we scaled the business with less than $3 million in outside funding until 2021, when we did a traditional IPO on the Nasdaq stock exchange. Today, Backblaze generates over $125M in revenue and is the leading specialized storage cloud, managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals. Role Overview Backblaze is seeking a Director, Enterprise Systems & Architecture to own the technical vision, architecture, and operations of the company’s enterprise systems and emerging AI use cases. This is a senior leadership role with organization-wide scope — spanning commercial and financial systems, GTM technology, data architecture, and the deployment and governance of AI agents across the business. This leader will serve as the company’s architectural authority for enterprise platforms including Salesforce, NetSuite, Zendesk, Snowflake, and their surrounding integration ecosystem. Equally important, they will support the governed infrastructure layer that enables AI agents, automated workflows, and headless capabilities to operate safely at a public company. They will partner closely with the CRO, CISO, and CFO organizations to enable velocity while ensuring systems are compliant, auditable, and built to scale. The ideal candidate combines deep enterprise architecture expertise with hands-on platform engineering instincts and practical experience deploying AI agents in production environments. This is a build-and-operate role, no
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Staff Software Engineer, Core AI Infrastructure You'll join a high-performing team of engineers driving AI transformation at Coinbase as a Senior Software Engineer on the IT Operations team within ESTO. This team builds custom products through full-stack engineering and scales the infrastructure powering Coinbase's AI products, with direct exposure to senior leadership in a fast-paced, incubator-style environment. You'll own the reliability and automation of critical AI infrastructure, ensuring our systems are resilient, observable, and secure at scale. What you'll do: Own end-to-end delivery of AI products by building production-grade distributed systems, including serving infrastructure, data pipelines, and deployment orchestration across the full stack throughout the SDLC. Drive platform adoption by designing clean APIs, abstractions, and developer-facing tooling that enable product teams to integrate AI capabilities without bespoke infrastructure requests. Partner with engineering and product leadership across Platform and other product groups to align infrastructure requirements, resolve cross-team technical dependencies, and define shared platform contracts. Shape engineering standards and technical culture by establishing architectural patterns, mentoring engineers, and raising the bar on code quality, observability, and operational excellence. Build
NVIDIA DGX Cloud is an AI Factory designed to power the next generation of AI and industrial-scale breakthroughs. As a Principal Engineer for Security Architecture, within our Security Engineering organization, you will own a core security domain of the AI factory: the architecture, the paved road that delivers it, and much of the code underneath. You will hold the security design bar across DGX Cloud from inside the teams doing the building, and this is a founding seat on a new team. Security Engineering is a new organization at DGX Cloud, accountable for the security outcome of the platform, and Security Architecture is the function inside it that holds the design bar. Security here is fleet horizontal and stack vertical, so your work will cross every DGX Cloud engineering organization: you will embed with the teams building GPU clusters, control planes, and services, join their designs as a participant rather than an approver, and leave behind systems in which an entire class of risk is no longer possible. There is no architecture review board here and no approval queue. You are a senior IC with deep security domain knowledge, and the security bar holds because you helped set it and then helped ship it. What You Will Be Doing: Own a Security Domain End to End: Take architectural ownership of a core domain of DGX Cloud security, from the design through the system running in production. That could be tenant and GPU workload isolation, workload identity, infrastructure and network, supply-chain provenance, hardened baselines and patching, or deploy-time policy and admission control. Embed with the Teams Building It: Join the design early, write the code, and help land it. The posture is not "you did this wrong." It is "here are the considerations we need to meet, I will help, let's go to work." Build Paved Roads, Not
Get new senior infrastructure architect jobs by email
Daily job updates · Unsubscribe anytime