What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.
Jobs in United States
Lead Lead Software Architect Java Backend Manager Manager Manager Manager in United States
1,193 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead lead software architect java backend manager manager manager manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
From $200K/yr
We are looking for a talented engineer to lead evaluation of startup acquisition opportunities in the AI, cloud and security space. You will drive product evaluations, prepare and manage technical architecture discussions with target groups in Product and Engineering and provide roadmap suggestions for M&A and investments for Datadog. You will be a key partner to Datadog’s C-level leadership and highly visible at the most senior levels of Datadog. The role is reporting into the Senior Director of Product Strategy and falls within the Product organization. We are looking for an innovative and strategic thinker who is passionate about the latest tech being developed by startups in the cloud, AI and security space. The ideal candidate enjoys researching and evaluating new technologies, works effectively with cross-functional teams, and communicates opinions concisely to our leadership team. Broad understanding of relevant Cloud Technologies and deep understanding of the full coverage of Datadogs current offerings is necessary. The Corporate Development team is small and values authentic, strong-willed individuals who think creatively and proactively. This role leads technical due diligence from a product and architecture perspective across our acquisition pipeline. You'll scope and stand up proof-of-concept and sandbox environments to stress-test candidate products, then give an honest, unvarnished view of their quality and depth - the kind of assessment that holds up regardless of deal momentum. You'll assess technical architecture, flag the risks and open questions that matter most early, and turn that into a clear post-acquisition integration path. Working closely with engineering, you'll keep the evaluation focused on what's actually decision-relevant, then translate the findings into strategic recommendations for leadership and help carry the integration through by partnering with the right people on the other side. What You’l
From $200K/yr
We are looking for a talented engineer to lead evaluation of startup acquisition opportunities in the AI, cloud and security space. You will drive product evaluations, prepare and manage technical architecture discussions with target groups in Product and Engineering and provide roadmap suggestions for M&A and investments for Datadog. You will be a key partner to Datadog’s C-level leadership and highly visible at the most senior levels of Datadog. The role is reporting into the Senior Director of Product Strategy and falls within the Product organization. We are looking for an innovative and strategic thinker who is passionate about the latest tech being developed by startups in the cloud, AI and security space. The ideal candidate enjoys researching and evaluating new technologies, works effectively with cross-functional teams, and communicates opinions concisely to our leadership team. Broad understanding of relevant Cloud Technologies and deep understanding of the full coverage of Datadogs current offerings is necessary. The Corporate Development team is small and values authentic, strong-willed individuals who think creatively and proactively. This role leads technical due diligence from a product and architecture perspective across our acquisition pipeline. You'll scope and stand up proof-of-concept and sandbox environments to stress-test candidate products, then give an honest, unvarnished view of their quality and depth - the kind of assessment that holds up regardless of deal momentum. You'll assess technical architecture, flag the risks and open questions that matter most early, and turn that into a clear post-acquisition integration path. Working closely with engineering, you'll keep the evaluation focused on what's actually decision-relevant, then translate the findings into strategic recommendations for leadership and help carry the integration through by partnering with the right people on the other side. What You’l
As an Engineering Manager on Coder’s Core Workspaces team, you’ll lead engineers building and evolving the systems behind our agentic development experience. You’ll help make agents more capable, reliable, and useful across real development environments. You’ll guide technical direction while growing the team and keeping execution sharp. You’ll work closely with Engineering, Product, and Design across the agent harness, integrations, and developer workflows. What you’ll do here Lead and grow a team within our Workspaces organization. Set technical direction across the agent harness, integrations, and workflows. Stay close to the code and contribute to architecture and implementation decisions. Evolve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Partner with Product and Design to turn agent capabilities into useful developer experiences. Improve reliability, performance, and operability across agentic systems. Coach engineers, raise the technical bar, and create clarity around priorities and tradeoffs. What we’re looking for Experience managing and growing software engineering teams. Strong hands-on engineering experience with React and TypeScript . Experience with Go . Hands-on experience building systems around LLMs and agentic workflows. Experience with model APIs, tool calling, context management, or agent loops. Strong distributed systems knowledge. Working knowledge of AWS . Strong technical judgment and comfort working through ambiguity. A track record of helping engineers grow while maintaining a high execution bar. Bonus tacos if you have Experience building coding agents, developer tools, or cloud development environments. Experience with MCP , agent tools, or multi-agent systems. Experience with remote execution, sandboxing, or isolated compute. Experience building abstractions across multiple model providers. Deep experience with AWS, Kube
$190.4K – $285.6K/yr
Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Lead the technical design and architecture of major platform initiatives, author design documents and build consensus across engineering teams. Define technical roadmaps for complex, multi-quarter projects that span multiple teams. Make critical architectural decisions for company documentation infrastructure, balancing scalability, reliability, and developer experience. Evaluate and set direction for integrating emerging technologies, including AI/LLM capabilities, into company documentation platforms and authoring tools. Establish and evolve engineering standards, best practices and technical guidelines for the team and broader organization. Partner with engineering teams across the company to understand documentation needs and design integrated solutions. Design, build and maintain scalable, reliable and performant services and systems. Contribute high-quality code across the full stack and navigate codebases with different languages and tools. Debug and resolve complex production issues and improve system reliability. Take ownership of system health and incident response. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Software Engineering, Engineering, or a related field, plus four (4) years of experience in Software Engineering. Must have four (4) years of experience in each of the following: - Working in a full stack environment with a foc
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As a Staff Software Engineer on the Automation - Foundations team, you will lead the re-platforming of the data layer underneath Vanta’s entire compliance product. This is a migration that spans multiple teams, has to preserve every public API contract along the way, and cannot lose a single customer’s evidence while it happens. Foundations is Vanta’s data platform team. We ingest security and compliance data from across our customer’s environments, currently tens of thousands of resources per second, with single-customer bursts running into the millions. We store the data, catalog it, make it queryable, and turn it into evidence that has to survive a real SOC 2 or FedRAMP audit. We are in the middle of moving our platform from a Mongo-centric architecture to a schema-aware, Postgres-backed architecture, on a Kafka and S3 pipeline that decouples data fetching from processing. Both pipelines run in parallel today, the hard problems here are correctness under migration, eventual consistency, and multi-tenancy, in a domain where “mostly right” is not an acceptable failure mode. Visit our Vanta Engineering Blog to learn more about what our team is working on. What you'll do as a Staff Software Engineer at Vanta: Lead the migration of Vanta’s resource data model from a Mongo-centric solution to a schema-aware Postgres-backed solution and running both generations in parallel without breaking a customer integration. Drive solutions across teams that you do not own but are dependent on the platform built by your team. Design for correctness under eventual consistency with idempotent session handling, conditional writes that survive out
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: This is a Product Engineering role specialized in monetization systems, and will lead that space at Replit. It’s a direct line to business impact. But it’s also critical to get right for our users. These are some of the most critical user journeys to get right. Getting them wrong creates the most frustrating experiences for users. So we’re looking for engineers who can build reliable and scalable billing systems and abstractions, while also translating that to an intuitive and friendly user experience. You will: Lead the architecture and implementation of monetization systems at Replit. Create seamless payment experiences for users for both product-led and sales-led motions. Build new abstractions and APIs for other engineers at Replit to monetize their new products. Iterate on pricing and packaging tactics to drive revenue growth. Examples include coupon codes and referral systems. Create monitoring and feedback systems so that we can proactively spot problems, fix them, and optimize performance. Required skills and experience: 6+ years of engineering experience, with strong skills working on the backend. Direct working experience in at least one of the following: Subscription platforms Usage-based billing SaaS Taxation Payment platforms Tokenization Self-directed and comfortable working autonomously in ambiguous environments. Excellent problem-solving skills with ability to debug complex billing issues and edge cases. Experience implementing customer-facing billing interfaces that simplify complex pricing structures. Tools + Tech Stack for this role: Python, TypeScript, React, Postgres, GraphQL, and Nodejs. Bonus Points : Experience working with Orb, Metronome, or Stripe usage based billing. Experienc
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: As a Staff Product Engineer at Replit, you’ll work closely with other product and platform engineers, designers, sales representative, and product managers to build features that help users collaborate with their team to go from idea to software fast. You’ll be at the forefront of shaping and experimenting on what our tens of millions of users love. You will: Help lead major projects and take new products from 0->1 Identify the hardest technical and/or quality problems holding us back, and then build solutions Chart high level technical direction and follow up to make sure those projects come together to deliver on results Mentor and develop new senior engineers to help grow the team Ship new features and build infrastructure using: TypeScript, React, CSS, GraphQL, Node.js, and Postgres Required skills and experience: A minimum of 7 years of professional software development experience Experience in a technical leadership role, working cross functionally Working experience building full stack applications with TypeScript Working experience building directly for users Bonus Points : You’re excited about the future of programming and have experience working with IDEs, terminals, or other common developer tools You’ve had previous experience working at a startup in a cross-functional engineering role This is a full-time role that can be held from our Foster City, CA office. The hybrid role has an in-office requirement of Monday, Wednesday, and Friday. Full-Time Employee Benefits Include: 💰 Competitive Salary & Equity 💹 401(k) Program with a 4% match ( US Only ) ⚕️ Health, Dental, Vision and Life Insurance 🩼 Short Term and Long Term Disability 🚼 Paid Parental, Medical, Caregiver Leave 🏝 Flexible
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Team: Plaid is becoming an AI-first company, and Intelligent Tooling builds internal platforms and tools to lead the transformation. Our biggest opportunity isn't just better tools for engineers, it's extending AI-native internal tooling to the rest of Plaid. Tools built for engineers assume things non-engineers don't have: local toolchains, monorepos, engineer credentials, PR-based workflows. That mismatch means Ops, Support, and other teams can't easily inherit what we build for engineering. They need their own path and we're building that path. Role: As a Senior Software Engineer on Intelligent Tooling, you will build and operate internal systems that empower non engineering teams to automate their workflows with AI. You will own the product and platform layer for internal tools, including the constraints and infrastructure that keep those tools safe and maintainable. There's no existing playbook for this at Plaid. You will define what the right non-eng AI surface looks like, ship its first durable versions, and partner closely with internal users to make sure it solves real problems. You will act as the engineering point of contact embedded with non-engineering teams, running discovery and trans
From $234K/yr
The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in production, particularly those leveraging Large Language Models (LLMs) and generative AI. We provide robust, scalable observability for AI workloads, including drift detection and model evaluation, and behavior tracing, enabling customers to ship AI with confidence. As a Staff Engineer, you’ll lead the development of new features and foundational capabilities within Datadog’s LLM Observability product. You will shape product direction, drive experimentation, and apply your deep understanding of both AI systems and software engineering to solve open-ended problems in the fast-moving AI landscape. Your work will directly impact how our customers monitor, troubleshoot, and optimize LLM-based applications in production. Join us in building the foundational tools that make AI systems observable, understandable, and reliable in the real world. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Drive design and implementation of LLM observability features. Ideate, prototype, and scale new product features to provide insights and drive improvements for generative AI systems Work cross-functionally with other eng teams, product, UX, and applied science to iterate fast and find product-market fit Develop and extend tools for tracing, evaluating, and debugging LLMs Influence architecture decisions and mentor engineers to build resilient, high-performance systems Stay close to customer pain points and use those insights to guide product and engineering priorities Stay current with industry trends and advancements in machine learning and observability, driving innovation within the team Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or r
From $244K/yr
We're looking for a Staff Engineer to join the Logs organization at Datadog and help redefine how our customers ingest, query, and derive insights from logs data. In this role, you’ll work closely with Product Managers and customers to drive complex initiatives across ingestion pipelines, search infrastructure, and intelligent log management capabilities - all while pushing the boundaries of what’s possible with AI and distributed systems. You’ll have the opportunity to lead efforts that shape the future of log management. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Partner with Product Managers to define ambiguous product requirements and determine the most impactful solutions for customers Lead technical strategy and execution and design systems surrounding log query performance and ingestion at scale. Explore and prototype new capabilities and collaborate with peers on initiatives spanning AI-powered log management, security and business operations, advanced query capabilities, and external data sources query capabilities. Mentor engineers across levels and contribute to growing a high-performing, collaborative team culture Who You Are: You have deep experience architecting and scaling backend systems, with a strong focus on data-intensive or distributed infrastructure You excel in ambiguous environments, demonstrating a mix of drive, curiosity and pragmatic decision-making You’ve partnered effectively with Product Managers and customers to define product direction and ship impactful features You have expertise in debugging complex systems and optimizing performance across real-time data pipelines You have experience in using AI agents tools in your day-to-day engineering practices You lead by example and enjoy helping others grow through m
About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Mission: Building the Data Foundations for AI We are the Snowflake Interoperable Foundations organization - the foundational layer that powers Snowflake’s AI, Analytics and Data Engineering capabilities. We lead innovations across open table formats such as Apache Iceberg, helping customers build peta-byte scale multi-cloud data lakes on Snowflake. We deliver core Metadata capabilities that power Snowflake’s industry-leading performance, AI, governance and platform features. We are embarking on a 0->1 redesign of our core systems across Interoperable Foundations. While we already manage exabyte-scale data supporting Snowflake’s AI capabilities, the next frontier is providing the foundational data layer that accelerates agentic innovation in an open, multi-format data world, You will be setting the technical vision across our investments in metadata platforms, Apache Iceberg and AI-ready storage. Your Impact: From Redesign to Reality 0->1 Architectural Leadership: Lead the ground-up redesign of our core Metadata systems, influencing the transaction frameworks that power query, DML, and AI-driven data interactions in addition to extending our lead on platform capabilities such as Zero Copy Cloning and Cross-Region / Cross-Cloud Replication. Iceberg Innovation: Drive
👋 Welcome to Glide! At Glide we’re reimagining the banking experience for the modern world . Our embedded fintech platform empowers legacy financial institutions, like community banks and credit unions, to pioneer novel digital experiences for their customers. You’ll be joining an all-star team with engineering, product, and growth experience from Stripe, Google, and Amazon. We’re looking for a talented Fullstack Software Engineer to help us build our initial product. We’re bringing a new perspective to the decades-old financial world , and we’re hoping you can help us do that! Your Responsibilities Lead and mentor a team of engineers, fostering growth, collaboration, and technical excellence. Partner closely with product and design to define requirements, prioritize work, and translate business needs into scalable technical solutions. Oversee development across frontend and backend, ensuring well-tested, secure, and performant code. Establish and maintain best practices in engineering, including CI/CD, test automation, and code review. Provide long-term technical vision for the evolution of Glide’s platform and infrastructure. Help recruit and build a diverse, world-class engineering team. Need-to-Haves Proven experience leading software engineering teams, including mentoring and performance management. Strong technical foundation in fullstack development (JavaScript/TypeScript, React, Node.js, Next.js). Experience designing and maintaining scalable, secure architectures. Deep knowledge of modern frontend technologies (responsive HTML/CSS, state/data fetching libraries like React Query/TanStack or tRPC). Familiarity with cloud infrastructure and DevOps practices (AWS, Docker, Git, CI/CD). Excellent understanding of software engineering best practices: architecture, testing, and security. Strong communication and collaboration skills, with the ability to partner effectively across functions. Nice-to-Haves 7+ years of professional software engineering experience, wi
From $385.1K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Security Software Engineer in the Enterprise Security team, you will advance Roblox's Enterprise Security strategy by building the systems and integrations that protect Roblox's corporate infrastructure. Where traditional security engineers evaluate and deploy vendor solutions, you will design and build production-grade security software - Identity and Access governance, policy enforcement engines, and security integrations that scale with Roblox's business. You'll partner with security professionals across InfoSec and work cross-functionally with Corporate Engineering, DevOps, and Product teams to drive security initiatives. You will: Build identity and access systems : Design, implement, and own integrations across Roblox's IAM ecosystem, including SSO federation, SCIM provisioning/deprovisioning pipelines, OAuth 2.0 authorization servers, and token lifecycle management. Lead security automation : Develop production-quality tools and services that enforce security policies at scale, replace manual workflows, and surface actionable signals Drive secure-by-design implementations : Partner with Corporate Engineering, DevOps, and product teams to embed security controls in
Other cities to consider
More places hiring for this role
Get new lead lead software architect java backend manager manager manager manager jobs in United States by email
Daily job updates · Unsubscribe anytime