Jobiba hiring network

Lead Principal Systems Engineer Devtools Manager Jobs

15 active opportunities · Updated for September 2026

Fresh results

15 shown

Explore current lead principal systems engineer devtools manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

R
Roblox
📍 San MateoFull-timeFrom $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Who We Are: Shape the future of Roblox’s virtual economy. The Economy ML team is building the machine learning backbone that powers Roblox’s Marketplace, Developer Monetization, and Payments ecosystems. From intelligent pricing and personalized storefronts to dynamic layout optimization and avatar understanding, we’re reimagining how the Roblox economy drives user engagement, monetization, and creator success at scale. As a Principal Software Engineer (Data Systems) , you will architect, build and deploy high-scale, reliable real-time and batch data systems for personalization, search and recommendation across various product surfaces in Marketplace, Developer Monetization and Payments. You will be involved in key data projects from architecting event taxonomies and logging interfaces to real-time feature serving across multiple search and recommendation surfaces. What You’ll Do Act as data engineering lead for Economy ML, setting standards for batch vs streaming feature pipelines, table design, observability, and documentation used across the Economy group. Work as a hands-on contributor on our data systems to power content recommendation, search and personalization across Economy product

awsgitmachine learning
View job →
R
Roblox
📍 San MateoFull-timeFrom $293.8K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Engine Networking Team pulls the players together by ensuring the communication of the game state to all. As a Principal Engineer on this team you will help the players experience the game as a nearly synchronous world. The networking and asset loading team plays a key role in a smooth experience for the players. You will work in all areas of the game platform in your quest for real-time communication of every part of Roblox. You Will: Lead engineers with 8+ years of industry experience Be experienced with one of these area: asset loading, rendering, and networking coming from a Game Engine/Studio. Be an amazing systems-level C++ programmer and be fascinated by the actual work the CPU does when you use smart pointers, templates, virtual functions, and blocks of memory, both structured and raw Have a keen to each millisecond of the network exchanges: You know where the time goes and how to reduce the waste Understand what happens on the operating system level when certain code is completed You Have: Worked on the guts of a multi-player game engine, solving problems related to scale, performance, latency, and throughput in client/server environments. Worked on a very large multithreaded d

awsgitai
View job →
P
Pinterest
📍 United StatesFull-timeFrom $285.5K/yr
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . As a principal engineer on the Online Systems team, you’ll join a team that powers Pinterest’s most business-critical online systems at massive scale, driving the reliability, efficiency, and evolution behind every core Pinner and Advertiser experience. You'll lead major efforts like multi-region deployment and Kubernetes migration, set the standard for operational excellence, and define the long-term vision for our online serving infrastructure, supporting machine learning and product innovation across the company. This is an opportunity for high-impact technical leadership, broad visibility, and cross-functional influence at the heart of Pinterest’s platform. What you’ll do: Improve reliability, scalability and infra efficiency for Pinterest’s critical online systems across storage and caching, online service and realtime analytics syste

pythonjavaaws
View job →
M
Midjourney
📍 San FranciscoFull-time
1mo ago

What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.

pythonaigo
View job →
G
Godaddy
📍 United StatesFull-timeFrom $154K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

pythonkubernetesai
View job →
O
OpenAI
📍 San FranciscoFull-time
1mo ago

About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu

pythonawsrest
View job →

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Role Overview & Key Responsibilities This is a high-leverage leadership role that spans architecture, execution, and org-building, and will shape the direction of our AI / ML initiatives at Ema. We are seeking an AI / ML technical leader who can take a vision and build it. As a Principal ML Engineer at Ema, you will be a senior technical leader responsible for shaping the machine learning roadmap, architecting large-scale ML systems, driving innovation, and ensuring our mixture of expert models (LLM + SLM + Custom Model) is accurate and performant at scale. You will collaborate across teams (research, product, infra, data, etc.), mentor senior engineers, and influence strategy and execution at company-wide levels. Responsibilities Lead the technical direction of GenAI and agentic ML systems that power enterprise-grade AI agents — spanning reasoning, retrieval, tool use, and integrations across various SaaS products. Architect, design, and implement scalable production pipelines for model training, fine-tuning, retrieval (RAG), agent orchestration, and evaluation — ensuring robustness, latency efficiency, and continuous learning. Define and own the multi-year ML roadmap for GenA

pythonjavamachine learning
View job →
C
Clickup
📍 CanadaFull-time
1mo ago

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview We’re seeking a Principal Frontend Engineer to drive the long-term technical vision for frontend engineering at ClickUp. In this role, you will lead architectural strategy across major product surfaces, solve company-wide frontend challenges, and influence how teams build and ship at scale. You will partner deeply with engineering, product, design, backend, and integrations leadership to ensure ClickUp’s frontend systems remain fast, reliable, and adaptable as the business grows. This is a hands-on technical leadership role for someone who can operate at both strategic and implementation levels, especially in a high-growth, fast-paced environment. What You’ll Do Define and drive frontend architecture strategy across product areas and platform investments Lead the design of scalable, performant systems in Angular 2+ and React that support rapid product development Solve complex cross-team technical challenges related to performance, state management, scalability, and developer productivity Partner with backend and integrations teams to shape end-to-end architecture for major features and systems Raise the bar for quality, testing, maintainability, and engineering rigor across the frontend organization Provide technical leadership on the highest-impact initiatives and unblock teams working through ambiguous or difficult problems Guide teams toward pragmatic decisions that support both speed and long-term product quality Mentor senior and staff engineers, and help shape frontend engineering culture and standards Influence roadmap and organizational decisions through strong technical judgment

typescriptreactangular
View job →
C
Clickup
📍 United StatesFull-time
1mo ago

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview We’re seeking a Principal Frontend Engineer to drive the long-term technical vision for frontend engineering at ClickUp. In this role, you will lead architectural strategy across major product surfaces, solve company-wide frontend challenges, and influence how teams build and ship at scale. You will partner deeply with engineering, product, design, backend, and integrations leadership to ensure ClickUp’s frontend systems remain fast, reliable, and adaptable as the business grows. This is a hands-on technical leadership role for someone who can operate at both strategic and implementation levels, especially in a high-growth, fast-paced environment. What You’ll Do Define and drive frontend architecture strategy across product areas and platform investments Lead the design of scalable, performant systems in Angular 2+ and React that support rapid product development Solve complex cross-team technical challenges related to performance, state management, scalability, and developer productivity Partner with backend and integrations teams to shape end-to-end architecture for major features and systems Raise the bar for quality, testing, maintainability, and engineering rigor across the frontend organization Provide technical leadership on the highest-impact initiatives and unblock teams working through ambiguous or difficult problems Guide teams toward pragmatic decisions that support both speed and long-term product quality Mentor senior and staff engineers, and help shape frontend engineering culture and standards Influence roadmap and organizational decisions through strong technical judgment

typescriptreactangular
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role: This is a Principal Product Engineering role focused on Money Infrastructure at Replit. You’ll work on the financial backbone that powers how Replit earns money, how builders earn money, and how Agents transact. This role sits at the intersection of engineering, product, and the business. The systems you build directly impact revenue, trust, and some of the most critical user journeys on the platform. Getting them right enables growth, experimentation, and global scale. Getting them wrong creates broken payments, confusing pricing, and lost trust. We’re looking for engineers who can design and scale reliable financial systems while translating complex monetary logic into intuitive, user-friendly experiences for both Replit customers and builders on the platform. We love folks who have a passion for monetizing innovation and being a part of the greater pricing story. You will: Lead the design, architecture, and implementation of Replit’s core money infrastructure, spanning pricing, billing, payments, and monetization. Own and scale the global order-to-cash foundation supporting credit-based subscriptions, usage-based billing, marketplaces, in-app payments, and commerce for Agents. Enable rapid pricing and packaging experimentation across the company by building flexible abstractions and APIs for new SKUs, plans, and monetization models. Build high-converting, localized payment experiences across geographies — thinking globally while enabling users to pay locally. Power builder monetization by creating payment rails for apps, Agents, subscriptions, and new monetization primitives. Partner closely on data specifications with finance, accounting, and data teams to produce accurate, auditable, and reliable f

typescriptreactnode.js
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Cortex CoWork team is defining the future of AI for enterprise data. Our mission is to transform how the world’s largest enterprises interact with their data through flagship products like Snowflake (CoWork) Intelligence . As a Principal AI Engineer , you will be a technical North Star for our AI initiatives. You won't just execute on a roadmap; you will help define it. You will tackle the most complex, "frontier" problems in agentic reasoning, NL-to-SQL, and enterprise-scale RAG, ensuring our AI products are not only innovative but fundamentally reliable and scalable for the Fortune 500. What you will do in this role: Technical Strategy & Architecture: Define the long-term technical vision for Snowflake Intelligence. Lead the architectural design of multi-agent systems, complex tool-use frameworks, and self-correcting NL-to-SQL engines. Drive Industry-Leading Reliability: Move beyond simple evals to build world-class, automated "hill-climbing" infrastructure. You will establish the methodology for how Snowflake measures and guarantees LLM performance across diverse customer schemas. Cross-Functional Influence: Partner with Product and Engineering leadership to align AI capabilities with business goals. You will bridge the gap between Research (modeling) and Product

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. We are looking for a skilled and motivated Principal Software Engineer who is passionate about continuous learning and eager to grow along with us in a fast-paced, innovative environment. You will work remotely from Bulgaria and will be reporting to an engineering leader located in Bulgaria. You Will: Lead the design and implementation of Smartsheet's next-generation architecture, ensuring scalability, security, and performance for millions of global users. Define and drive architecture strategy, making key technical decisions that shape the future of the platform. Review and guide technical project designs, providing feedback during design review presentations to ensure system resilience and scalability. Take ownership of cross-functional technical initiatives, aligning teams around common architectural goals while driving large-scale projects to completion. Foster strong technical leadership, mentoring senior engineers and influencing best practices across multiple engineering teams. Lead deployment reviews for high-impact projects, ensuring they meet scalability, performance, and security requirements. Collaborate closely with product management and other business stakeholders to balance market needs with technical constraints, driving innovation while maintaining technical rigor. Advocate for quality and operational excellence, ensuring systems are monitored, tested, and maintained to meet the highest reliability standards. Perform other duties as assigned. You Have: Proven experience in system architecture and the d

REMOTEpythonjavasql
View job →
R
Roblox
📍 San MateoFull-timeFrom $385.1K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Security Software Engineer in the Enterprise Security team, you will advance Roblox's Enterprise Security strategy by building the systems and integrations that protect Roblox's corporate infrastructure. Where traditional security engineers evaluate and deploy vendor solutions, you will design and build production-grade security software - Identity and Access governance, policy enforcement engines, and security integrations that scale with Roblox's business. You'll partner with security professionals across InfoSec and work cross-functionally with Corporate Engineering, DevOps, and Product teams to drive security initiatives. You will: Build identity and access systems : Design, implement, and own integrations across Roblox's IAM ecosystem, including SSO federation, SCIM provisioning/deprovisioning pipelines, OAuth 2.0 authorization servers, and token lifecycle management. Lead security automation : Develop production-quality tools and services that enforce security policies at scale, replace manual workflows, and surface actionable signals Drive secure-by-design implementations : Partner with Corporate Engineering, DevOps, and product teams to embed security controls in

pythonawsgit
View job →
R
Roblox
📍 San MateoFull-timeFrom $326.1K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Security Software Engineer on the Production IAM team, you will set the technical direction for how identity and access work across Roblox's production infrastructure, from the mTLS-based identity that services use to authenticate to one another, to the privileged access controls that govern how engineers reach production. The team is accountable for Roblox's machine and workload identity platform, its centralized authorization engine, its production access management platform, production PKI and certificate lifecycle, and just-in-time privileged access for engineers. As an individual contributor in Production IAM, you will define multi-year strategy, drive alignment across Roblox Platform, mentor senior and staff engineers, and personally build the hardest parts of these systems. As AI agents become first-class actors in production, you will also help pioneer how they get identity, prove who they are, and receive safely-scoped access. You will Lead the architecture for production identity and access. Define and evolve the end-to-end design for machine, workload, human, and AI-agent identity across our hybrid on-prem and cloud fleet, making secure access invisible when

pythonjavaaws
View job →
R
Roblox
📍 San MateoFull-timeFrom $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Why Content Safety? As a Principal Machine Learning Engineer for Content Safety, you will define the future of proactive moderation, driving immense social impact through cutting-edge, innovative ML solutions, focused on critical and ambiguous safety challenges. You will set the 3-5 year technical strategy and architectural blueprint for how Roblox uses machine learning for content moderation. You will own the architectural and execution roadmap of massive-scale ML systems that mitigate violative UGC content before it impacts our community. You will feel a deep sense of responsibility in proactively protecting our community thoughtfully and fairly, while balancing user freedom with platform civility. Your efforts will ensure Roblox remains one of the safest places on the internet for our broad community of over 100 million daily active users. You will: Define and Own the Technical Vision: Define and lead the multi-year technical vision, architectural strategy, and execution for machine learning solutions in Content Safety, ensuring these systems proactively and effectively detect and mitigate violative content at massive scale. Strategic Stakeholder Partnership: Collaborate with execu

awsgitmachine learning
View job →
🔔

Get new lead principal systems engineer devtools manager jobs by email

Daily job updates · Unsubscribe anytime