Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies — from the world's largest enterprises to the most ambitious startups — use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the Organization The Core Infrastructure organization operates the foundational systems that power Stripe globally — including databases (MongoDB, PostgreSQL), high availability and disaster recovery (HADR), AWS cloud infrastructure, Linux servers, container orchestration, mesh networking, service discovery, and network edge infrastructure. Within Core Infra, the Regional Enablement Platform (REP) team helps Stripe launch and operate new regions without learning about broken dependencies from users. REP builds the regionalization, validation, deploy-safety, and operator tooling needed to answer practical launch-readiness questions: can critical payment paths run from the new region, which services still depend on a remote control plane, what breaks under packet loss or failover, and what must be fixed before deploys, launches, traffic shifts, or failovers proceed. The team uses traffic replay, synthetics, failover drills, dependency analysis, CI/CD gates, and incident data to turn those findings into platform fixes, service-owner asks, and reusable readiness checks across networking, HADR, and service teams. This role is based in Bangalore and serves as a senior technical anchor for Core Infrastructure in India, with direct cross-region influence across AMER, EU, and APAC. What you'll do As a Staff Engineer on REP, you will play a key leadership role in enabling Stripe's infrastructure to power all of our products, globally and at scale. You will
Jobiba hiring network
Senior Staff Engineer Advanced Operations Jobs
7,292 active opportunities · Updated for October 2026
Fresh results
14 shown
Explore current senior staff engineer advanced operations jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Premium Merchant Experiences team helps businesses using Stripe succeed by providing high-quality, contextual, and timely premium support options. Our Dublin-based team is a startup within Stripe, building a greenfield product – offering both automated, AI-driven support and improved access to technical experts across Stripe. In addition to building directly for Stripe's merchants, we are responsible for building robust internal tooling that empowers Stripe employees in support roles to deliver the best possible experience for our customers. What you’ll do As an early hire on the Premium Merchant Experiences team, your role more closely resembles that of a founding engineer at a startup within Stripe. As a Staff Engineer, you will have widespread impact by setting the technical strategy of the team, mentoring and growing senior engineers, and building a culture of technical excellence. As a technical leader within the team, you will both drive fast-paced, sustained product delivery while also upholding high standards for building reliable systems. Responsibilities: Design, build, and operate production systems, with an emphasis on extensibility, reliability, and usability Write and review code regularly, setting the pace and elevating standards for the team Build high-quality end-to-end product experiences, demonstrating deep care for users' needs and serving as a steward of crafting great experiences Collaborate with
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role Is Different This is not a typical “Applied Scientist” or “ML Engineer” role. As a Member of Technical Staff, Applied ML, you will: Work directly with enterprise customers on problems that push LLMs to their limits. You’ll rapidly understand customer domains, design custom LLM solutions, and deliver production-ready models that solve high-value, real-world problems. Train and customize frontier models — not just use APIs. You’ll leverage Cohere’s full stack: CPT, post-training, retrieval + agent integrations, model evaluations, and SOTA modeling techniques. Influence the capabilities of Cohere’s foundation models. Techniques, datasets, evaluations, and insights you develop for customers will directly shape the next generation of Cohere’s frontier models. Operate with an early-startup level of ownership inside a frontier-model company. This role combines the breadth of an early-stage CTO with the infrastructure and scale of a deep-learning lab. Wear multiple hats, set a high technical bar, and define what Applied ML at Cohere becomes. Few roles in the industry combine application, research, customer-facing engineeri
This position is based in Vancouver, BC , within Diligent’s Technical Center of Excellence. We are currently hiring candidates who are based in or able to work from Vancouver . Software Engineer — Platform AI Service Levels: Software Engineer II Senior Software Engineer Staff Software Engineer Location: Vancouver Position Overview As a Software Engineer on Diligent's Platform AI team, you'll help design, build, and operate the core services that power AI-driven capabilities across Diligent's global product suite. You'll build secure, scalable, serverless services on AWS that translate AI research and models into commercial-quality, production-ready solutions — enabling customers to derive insights from their governance data. You'll work closely with AI researchers, product managers, and other engineering teams, owning your services end-to-end: architecture, implementation, deployment, and monitoring. The team operates with a strong AI-augmented engineering culture — using AI tools to accelerate coding, testing, debugging, and delivery — while applying sound judgment about when and how to apply them. Key Responsibilities Design and implement secure, scalable, fault-tolerant, high-performing solutions using AWS serverless technology — event-driven, highly observable, and built with infrastructure as code. Collaborate with AI researchers/engineers to translate AI and LLM capabilities into robust, production-grade services, and help other teams integrate them. Build and maintain the pipelines needed to deploy, monitor, and manage AI services at scale — observable, resilient, and cost-effective. Use AI-powered development tools (code assistants, test generation, architecture exploration) responsibly to accelerate delivery and improve quality, always validating outputs. Participate in architecture discussions and design reviews, and contribute to product design by understanding customer problems — especially where AI can offer a breakthrough solution. Work in
About the Team DoorDash Labs, established in 2018, serves as the innovation hub for DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. The team's mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We’re ruthlessly focused on business impact. We are a highly senior team composed of former pioneers from a variety of different robotics industries. As of 2025, DoorDash has completed 10B lifetime deliveries. We’re focused on how to do the next 10B even better. About the Role We’re looking for a Senior/Staff electrical and firmware engineer who designs the board and writes the firmware for connected consumer and enterprise devices, including tablets, POS systems, peripherals, and emerging robotics applications. In this role, you will take connected devices from ambiguous user or business needs through system architecture, rapid prototyping, board design, firmware implementation, bring-up, field deployment, and production transition. The ideal candidate has a passion for building and shipping reliable hardware at scale and thrives in an environment that demands technical depth and high-quality execution across multiple concurrent programs. This is a build-heavy role on a small, senior team. You’re excited about this opportunity because you will… Own connected devices end to end, including: Own connected devices from an ambiguous product need through field deployment. Work across the hardware and firmware boundary instead of handing work between specialists. Explore uncertain product ideas through rapid prototypes and experiments. Make consequential tradeoffs on a small, senior team with direct exposure to users and field behavior. Key Responsibilities Translate user and business needs into an integrated electrical and firmware architecture. Design compute-based, mixed-signal boards incorporating processors
Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we’re changing the way people think about and interact with personal finance. We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The role: SoFi's Associate AI Engineer, Finance Transformation is a hands-on builder within SoFi's Finance organization, focused on building agentic AI workflows that transform how Finance works from close and reconciliations to forecasting and reporting. Finance has one of the largest AI opportunity surfaces at SoFi: over a hundred identified use cases, an active champions network, and executive sponsorship. In this role you will build multi-step AI workflows on approved enterprise AI platforms, stand up the telemetry that measures AI usage, cost, and ROI across Finance, and help make AI outputs trustworthy enough for Finance decision-making in a controlled environment where outputs must be explainable, auditable, and reconciled to the number. You will work directly with the AI Transformation Manager for Finance, who owns use-case strategy and stakeholder engagement, and in close partnership with SoFi's AI SDLC and platform teams, who support the path from prototype to production. This is a build-focused role with an unusual growth surface: SoFi's AI Engineering ladder (through Staff and Senior Staff) is the visible progression path. What you’ll do: Build agentic AI workflows: Develop multi-step AI workflows such as planning, tool use, retrieval, structured orchestration on approved enterprise AI pla
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Come lead the Identity & Access Management Platform team (IAM Platform) — the foundational layer that decides who can see and do what across all of ClickUp. This team owns the authorization engine, the permission and role model, authentication (SSO, MFA, SCIM, OIDC), sharing primitives, audit logs, and the core data model and APIs that define how customers' work is organized and nested across the product — the foundation every other part of ClickUp is built on. It is backend-heavy, high-blast-radius work. As we move upmarket, this is some of the most important and security-sensitive work we do — enterprises decide whether they can trust us based on how precise, predictable, auditable, and manageable our access controls are. We're looking for a Senior Engineering Manager to lead and grow this team. Your mandate is twofold: deliver an enterprise-grade access management platform — extensible, configuration-driven, correct and auditable by design — and build the team that will carry it , hiring and developing engineers as we invest heavily in enterprise. You'll raise the access, identity, and admin capabilities enterprises depend on to an enterprise-grade standard and keep them there, partnering with a Staff engineer on technical direction while you own delivery, people, and priorities. Just as important: you'll run this team the way ClickUp runs — AI-native . We structure work so AI agents can read it, route it, roll it up, and report on it, so a small team amplified by agents operates at a different scale. There are no status-reporting meetings; agents keep status, progress, and health current from r
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role We've been doing a lot of thinking as a team, and what we've landed on is that this person needs to be someone who takes ownership. Real ownership. Not just "I'll close the ticket", but "I'm going to make sure this never happens again" type of ownership. That's the kind of person we're looking for, and frankly, that's the kind of person who's going to thrive here at Palantir. As an IT Support Engineer, you take real pride of our internal ecosystem, from individual workstations to conference rooms to the systems employees rely on every day. You're the person people turn to when something isn't working, whether they're at their desk, travelling, or preparing for an important meeting. You're thorough when it comes to troubleshooting. You don't just close the ticket, you make sure the underlying issue is actually resolved. When something keeps coming back, you dig into why and strive to look for a permanent fix. You're proactive by nature and complacency isn't something you settle for, and it’s not something we settle for either. You're constantly looking for ways to eliminate friction, automate the mundane, and elevate the bar for TechOps. People feel comfortable coming to you with problems because you're approachable and follow through. You're familiar with the needs of executives and senior staff, and you check in proactively when you know something critical is on the horizon like a board meeting, an earnings call, an internal conference or a big reveal rather than taking a back seat. What sets you apart is how much you care about the growth of the people around you. When you work through a tough ticket with a colleague, you make it a teaching moment. When you
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role We've been doing a lot of thinking as a team, and what we've landed on is that this person needs to be someone who takes ownership. Real ownership. Not just "I'll close the ticket", but "I'm going to make sure this never happens again" type of ownership. That's the kind of person we're looking for, and frankly, that's the kind of person who's going to thrive here at Palantir. As an IT Support Engineer, you take real pride of our internal ecosystem, from individual workstations to conference rooms to the systems employees rely on every day. You're the person people turn to when something isn't working, whether they're at their desk, traveling, or preparing for an important meeting. You're thorough when it comes to troubleshooting. You don't just close the ticket, you make sure the underlying issue is actually resolved. When something keeps coming back, you dig into why and strive to look for a permanent fix. You're proactive by nature and complacency isn't something you settle for, and it’s not something we settle for either. You're constantly looking for ways to eliminate friction, automate the mundane, and elevate the bar for TechOps. People feel comfortable coming to you with problems because you're approachable and follow through. You're familiar with the needs of executives and senior staff, and you check in proactively when you know something critical is on the horizon like a board meeting, an earnings call, an internal conference or a big reveal rather than taking a back seat. What sets you apart is how much you care about the growth of the people around you. When you work through a tough ticket with a colleague, you make it a teaching moment. Whe
Coder is looking for an experienced Senior Technical Program Manager to lead and shape our most complex, high-stakes engineering programs. You'll operate with a high degree of autonomy, setting the standard for program execution across the organization and acting as a trusted partner to senior engineering and product leadership. The role sits at the intersection of strategy and execution: you'll own the definition, cross-functional planning, and delivery of initiatives that span multiple teams and systems, and give leadership the visibility they need to steer. What you'll do here Own engineering programs end to end, from definition through delivery and iteration, across teams that don't report to one another Own scope, milestones, timelines, dependencies, and risk, driving trade-offs to a written decision and holding accountability for delivery Work hands-on with engineers and product managers, close enough to the work to influence architectural and interface decisions Serve as the central coordination point across engineering, product, design, security, and go-to-market Run the operating cadence and deliver crisp updates on program health, risks, and trade-offs to technical, executive, and customer stakeholders Improve engineering productivity through better tooling and light-weight process, including AI-native tooling and our own product What we're looking for 7+ years in technical program management, including complex multi-team programs where no single team owned the whole outcome Deep technical intuition—you can read a design doc, ask the question that exposes the unstated assumption, and argue sequencing with a staff engineer A track record of delivering at scale, including ownership of a dated external commitment such as a GA, a launch, or a regulated enterprise deadline Strong organizational influence; you set the agenda with minimal oversight and move teams that don't report to you, without being slowed down by ambiguity or competing priorities Excellent ve
Staff/Senior Photolithography Process Development Engineer — Fab 10N/X, Singapore. Apply via Workday.
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Deployments team designs and maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering teams. This infrastructure is primarily composed of Argo Workflows and ArgoCD. The team also provides tooling that enables clear system ownership and facilitates self-service onboarding for development teams. We are looking to speak to candidates who can work East Coast hours. The ideal candidate should Have 6+ years of experience in software development and operating distributed systems Proficiency in Python, Go, or a similar language Proven experience building and operating large-scale continuous integration and continuous deployment (CI/CD) pipelines Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual process (“allergic to ops work”). We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Expectations Contribute to developing a world-class continuous deployment experience, enabling the rapid and reliable shipment of MongoDB products This includes, but is not limited to, contributing to open-source projects, or engineering software-based
MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. This role can be based out of our Boston, New York City, Raleigh, Miami, Pittsburgh or remotely in the United States while physically based in an Eastern or Central time zone location. The ideal candidate should Have 6+ years of experience working on software development and operating distributed systems Proficiency in Python, Go, or a similar language Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs. Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Responsibilities Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure g
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Deployments team designs and maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering teams. This infrastructure is primarily composed of Argo Workflows and ArgoCD. The team also provides tooling that enables clear system ownership and facilitates self-service onboarding for development teams. We are looking to speak to candidates who can work East Coast hours. The ideal candidate should Have 6+ years of experience in software development and operating distributed systems Proficiency in Python, Go, or a similar language Proven experience building and operating large-scale continuous integration and continuous deployment (CI/CD) pipelines Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual process (“allergic to ops work”). We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Expectations Contribute to developing a world-class continuous deployment experience, enabling the rapid and reliable shipment of MongoDB products This includes, but is not limited to, contributing to open-source projects, or engineering software-based
Get new senior staff engineer advanced operations jobs by email
Daily job updates · Unsubscribe anytime