We’re looking for a Senior Staff Software Engineer with deep experience in GenAI/ML to join Datadog’s Application Performance Monitoring (APM) team. APM is a product which provides deep visibility into applications, enabling users to identify performance bottlenecks, troubleshoot issues, and optimize services. With distributed tracing, profiling, out-of-the-box dashboards, and seamless correlation with other telemetry data, Datadog APM provides some of the deepest and most structured visibility into the health and performance of applications. This context sets us up for an opportunity to be the world leaders in agentic investigations and incident troubleshooting. You’ll act as a technical leader within the APM group, focused on agentic workflows. You’ll lead efforts to design, train, evaluate, and deploy GenAI/ML models at scale. We’re looking for a product-minded ML engineer with strong technical expertise, excellent communication skills, and a track record of driving impactful initiatives end to end. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Serve as the technical owner for GenAI initiatives within APM, leading design, development, and deployment of ML/AI-powered features across multiple teams. Guide long-term strategy and technical direction for GenAI workflows across APM and related products. Build and benchmark GenAI/ML models using state-of-the-art techniques. Contribute to Datadog’s broader senior engineering community through thought leadership and collaboration on company-wide initiatives. Collaborate with cross-functional teams to build automated investigation and triaging tools. Influence product direction by bringing a strong product mindset to your work, always advocating for the end user. Guide teams through ambiguity, sc
Jobiba hiring network
Lead Senior Engineering Manager 2c Platform Reliability Jobs
6,876 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current lead senior engineering manager 2c platform reliability jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s
Level Up Your Career with Zynga! At Zynga, we bring people together through the power of play. As a global leader in interactive entertainment and a proud label of Take-Two Interactive, our games have been downloaded over 6 billion times—connecting players in 175+ countries through fun, strategy, and a little friendly competition. From thrilling casino spins to epic strategy battles, mind-bending puzzles, and social word challenges, our diverse game portfolio has something for everyone. Fan-favorites and latest hits include FarmVille™, Words With Friends™, Zynga Poker™, Game of Thrones Slots Casino™, Wizard of Oz Slots™, Hit it Rich! Slots™, Wonka Slots™, Top Eleven™, Toon Blast™, Empires & Puzzles™, Merge Dragons!™, CSR Racing™, Harry Potter: Puzzles & Spells™, Match Factory™, and Color Block Jam™—plus many more! Founded in 2007 and headquartered in California, our teams span North America, Europe, and Asia, working together to craft unforgettable gaming experiences. Whether you're spinning, strategizing, matching, or competing, Zynga is where fun meets innovation—and where you can take your career to the next level. Join us and be part of the play! Position Overview CSR2 is one of the most successful mobile racing games in the world, and we’re evolving the experience for millions of players. We’re looking for a Lead/Senior Game Designer who can shape impactful systems, elevate the player experience, and strengthen the overall design function across pods. This role combines strong systems design craft with leadership, collaboration, and an ability to drive clarity and alignment. What You'll Do : Lead the design of major systems and features that deepen CSR2’s progression, engagement, and long-term meta Drive end-to-end feature delivery—including concepts, specs, flows, balancing, prototyping, iteration, and final launch Partner closely with Product, Engineering, UX, Art, and LiveOps to align goals, define KPIs, and ensure high-quality execution
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Snowflake is seeking an entrepreneurial and visionary engineering leader to build and scale our new AI Solutions team. As the Director of Engineering for AI Enterprise , you will be at the forefront of the generative AI revolution, leading a world-class engineering organization that designs and implements cutting-edge solutions on the AI Data Cloud. This is a critical, high-impact leadership role where you will partner with the world's largest companies to solve their most complex challenges and unlock transformative business value using Snowflake's powerful Cortex AI Platform. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Recruit, mentor, and scale a high-performing, global engineering organization. Foster a culture of innovation, ownership, and engineering excellence. Define the long-term technical vision and organizational structure for the AI Solutions team, ensuring alignment with Snowflake's product evolution and business goals. Lead the technical design and development of advanced AI and machine learning solutions using Snowflake Cortex and Snowflake ML. Own the end-to-end implementation of the AI Solutions product line, from initial concept to enterprise-grade production deployments. Act as the senior engineering authority in cu
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Vice President, Software Engineering Role Overview Mastercard is seeking a Vice President of Engineering to lead the Developer Workbench, a strategic platform designed to deliver a unified, AI-enabled, end-to-end software engineering experience across the enterprise. The Developer Workbench will bring together core engineering products, developer tools, AI-assisted coding capabilities, development environments, testing workflows, deployment pipelines, cloud development experiences, and developer insights into one cohesive platform. This leader will be accountable for transforming the Workbench from a set of disconnected tools and services into a productized developer experience that improves productivity, accelerates onboarding, increases adoption of modern engineering capabilities, and strengthens governance across the software development lifecycle. This is a senior engineering leadership role for a builder, integrator, and enterprise change leader who can operate across product, engineering, architecture, security, finance, learning, and senior technology leadership. ________________________________________ Key Responsibilities Lead the Developer Workbench Engineering Strategy • Define and execute the engineering strategy for the Developer Workbench. • Establish the technical architectu
Datadog (NASDAQ: DDOG) is looking for a Product Strategy and Corporate Development Lead to drive Product vision, M&A, and venture investments for Datadog. You will work directly with Datadog's founders, partnering with Datadog's global engineering and product leaders to identify and execute transactions that expand our platform into new markets. This is not a traditional Corp Dev seat. You will operate on a small, high-autonomy team where technical depth matters as much as deal execution. You can expect to focus on the EMEA landscape, which is one of the most dynamic ecosystems in enterprise software right now, and you will be at the center of it. At the same time, you'll be working with our Product and Engineering leaders based in Europe, who will depend on you to be the fabric between our NYC and Paris HQs. Together, you'll be expected to go deep and autonomously explore new areas of expansion for us. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: (Describe role responsibilities here) Be close to our Product & Engineering leaders to discover new areas of interest and acceleration before an M&A target is even identified, and stay ahead of market trends in AI/ML, observability, security, and cloud infrastructure to inform Datadog's growth strategy Identify and evaluate M&A targets and venture investment opportunities across European markets, with a focus on AI-native and infrastructure companies Build deep relationships with founders, VCs, and accelerators across the Paris, London, and Iberian ecosystems to generate proprietary deal flow Lead end-to-end due diligence: technical product assessments, financial modeling, valuation analysis, and integration planning Partner with senior engineering and product leaders to assess techni
Squarespace is looking for a Staff Software Engineer for the Domains group to partner with our product and engineering teams. Our group's products offer customers the tools to establish the foundation of their digital presence, namely their domain and corresponding email. The teams are currently focused on building new features to allow customers to do more with their domains to further our goal of being one of the top domain providers in the business. In this role, you will lead technical strategy, architecture, and implementation while mentoring engineers, collaborating across teams, and contributing code to complex projects that align technical investments with business priorities. This is a hybrid role working from our Aveiro office 3 days per week. You will report to the Senior Engineering Director based in Portugal. You'll Get To… Be the architect for a major area of our business Influence: Drive best practices for engineering excellence and technical leadership, influence design reviews, question assumptions, point out pitfalls and foster shared understanding Guide: Provide guidance to the engineering team to help architect solutions at scale and become the go-to whiteboard partner for the team. Mentor: Multiply your effectiveness through influence by growing our experienced team of Software Engineers at several levels. Plan: Tie technical investments to business priorities, lead architecture discussions and design reviews, and contribute to product roadmap planning Collaborate: Work with engineering teams across the company Write: Create medium to long-term technical strategies for our online store product and platform. Clearly and concisely write that strategy and collaborate with engineers across the organization to implement the strategy Code: Wants to contribute directly to projects that require more senior technical expertise or to quickly address ambiguous engineering challenges across the organization Who We're Looking For Have 8+ years of experience
Work Flexibility: Remote As a Senior Lead, Data Engineering, you will serve as a technical leader who helps shape the future of enterprise data solutions. In this role, you will drive complex data initiatives, influence technical strategy, and partner with teams across the organization to build scalable, high-impact data products. This is an opportunity to solve challenging business problems while mentoring fellow engineers and elevating data engineering best practices. What You Will Do Lead the architecture, development, and modernization of scalable enterprise data platforms that support global procurement analytics and business transformation. Define and help execute a multi-year data engineering strategy focused on platform scalability, reliability, automation, technical debt reduction, and long-term maintainability. Design, build, and optimize Azure-based data solutions using technologies such as Databricks, Delta Lake, Azure Data Factory, Azure DevOps, CI/CD pipelines, and infrastructure automation. Integrate and harmonize data across multiple ERP systems by standardizing supplier, purchasing, and master data into common enterprise data models. Partner with procurement analysts, architects, engineers, and business stakeholders to translate complex business needs into reusable, scalable data products and engineering solutions. Establish engineering standards, conduct architecture reviews, improve documentation, and mentor engineers to raise the overall technical capability of the team. Identify and implement AI-enabled approaches that accelerate development, improve data quality, automate documentation, support testing, and enhance analyst productivity. Evaluate and recommend tools, frameworks, patterns, and platform investments that improve performance, reliability, security, governance, and operational ef
About The Role & Team Amplitude is the leading AI analytics platform, and our ability to deliver measurable customer outcomes quickly is a key part of how we keep that lead. The Customer FDE (Forward Deployed Engineering) team sits at the intersection of engineering, product, and customer success — owning the technical delivery that takes validated products from co-development and implements them across enterprise customers. As a Customer Forward Deployed Engineer, you will own end-to-end technical delivery for enterprise customer implementations, from sales engagement through post-deployment validation. You'll work directly in Amplitude's product codebase, submitting PRs, shipping customer-specific solutions, and building reusable patterns that make every successive engagement faster. This is not a traditional support or solutions role. Customer FDEs are engineers first who partner closely with Sales, Customer Success, Product, and Engineering to turn customer needs into working software. The right person thrives in ambiguity, learns new domains quickly, and cares as much about the customer's outcome as about technical elegance. As a Customer Forward Deployed Engineer, you will: Own enterprise implementations end-to-end — engage during sales to map customer needs to capabilities, scope implementation plans, and deliver through post-deployment validation with measurable customer outcomes Work directly in the product codebase — submit PRs, test, and ship customer-specific solutions without requiring constant oversight; over time, the gap between "what we built" and "what the customer needs" keeps shrinking because you're the one closing it Ship early, ship often, and gather feedback — treat every customer interaction as an opportunity to validate your approach and course-correct quickly Build reusable patterns and playbooks — when you solve something once, make sure the next person doesn't have to solve it again; identify and document implementation patterns that
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About The Role & Team Amplitude is the leading AI analytics platform, and our ability to deliver measurable customer outcomes quickly is a key part of how we keep that lead. The Customer FDE (Forward Deployed Engineering) team sits at the intersection of engineering, product, and customer success — owning the technical delivery that takes validated products from co-development and implements them across enterprise customers. As a Customer Forward Deployed Engineer, you will own end-to-end technical delivery for enterprise customer implementations, from sales engagement through post-deployment validation. You'll work directly in Amplitude's product codebase, submitting PRs, shipping customer-specific solutions, and building reusable patterns that make every successive engagement faster. This is not a traditional support or solutions role. Customer FDEs a
Citi, the leading global bank, has approximately 200 million customer accounts and does business in more than 160 countries and jurisdictions. Citi provides consumers, corporations, governments, and institutions with a broad range of financial products and services, including consumer banking and credit, corporate and investment banking, securities brokerage, transaction services, and wealth management. As a bank with a brain and a soul, Citi creates economic value that is systemically responsible and in our clients’ best interests. As a financial institution that touches every region of the world and every sector that shapes your daily life, our Enterprise Operations & Technology teams are charged with a mission that rivals any large tech company. Our technology solutions are the foundations of everything we do from keeping the bank safe, managing global resources, and providing the technical tools our workers need to be successful to designing our digital architecture and ensuring our platforms provide a first-class customer experience. We reimagine client and partner experiences to deliver excellence through secure, reliable, and efficient services. Our commitment to diversity includes a workforce that represents the clients we serve from all walks of life, backgrounds, and origins. We foster an environment where the best people want to work. We value and demand respect for others, promote individuals based on merit, and ensure opportunities for personal development are widely available to all. Ideal candidates are innovators with well-rounded backgrounds who bring their authentic selves to work and complement our culture of delivering results with pride. If you are a problem solver who seeks passion in your work, come join us. We’ll enable growth and progress together. Position Overview: The Senior Platform Engineering Lead is a pivotal senior-level engineering position responsible for driving the
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Summary: As a Principal Software Engineer at Appian, you will be the primary technical strategist, responsible for shaping the architectural foundation of our platform. Your role involves anticipating future challenges and implementing innovative solutions today. You will drive complex cross-functional initiatives, ensuring Appian remains a leader in the low-code and automation industry. Responsibilities: Define the long-term architectural vision and governance for the Appian platform. Identify systemic technical risks and lead task forces to resolve architectural bottlenecks. Develop internal tools and SDKs to simplify infrastructure and enhance developer productivity. Conduct deep-dive troubleshooting for complex production issues. Promote AI-native engineering practices and integrate AI features into the platform. Lead the development of core platform libraries and high-risk prototypes. Participate in the Architectural Guild and review high-impact design documents. Mentor lead and senior engineers, and represent Appian in the tech community. Ensure operational resilience with self-healing and highly available systems. Required Qualifications: Bachelor’s or Master’s degree in Computer Science, Information Technology, or related field. 15+ years of software engineering experience, with significant experience in architecting large-scale distributed systems. Strong understanding of data structures, algorithms, and design patterns. Proven transformational leadership in technological migrations or strategies. Expertise in Java and Cloud-Native ecosyst
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Location- Chennai Team: Engineering Enablement Group As a Senior Software Engineer in our Engineering Enablement Group, you will lead the re-design and evolution of our Mobile Branding framework — the system that enables customers to create custom-branded versions of the Appian mobile application for both iOS and Android. You will drive the architectural modernization of the end-to-end branding pipeline, from the customer-facing Forum application and provisioning tools to the backend build service running on Mac EC2 runners in AWS. By leveraging modern microservices, CI/CD automation, and cloud-native infrastructure, you will transform the current system into a more reliable, scalable, and maintainable platform that reduces manual intervention and accelerates customer delivery. We are looking for a technical leader who can bridge the gap between complex Ruby/Bash-based tooling, Appian process models, and AWS infrastructure to deliver a seamless mobile branding experience. Primary Qualifications: 6-9 Strong working experience with Android and iOS frameworks and mobile application development workflows. Familiarity with mobile build systems (Fastlane, Xcode, Gradle) and code-signing workflows. Experience with proficiency in Python, with experience in Ruby, Bash, or Go being a plus. Advanced experience with AWS infrastructure (S3, Lambda, EC2) and CI/CD pipeline design. Strong end-to-end knowledge of pipeline creation, deployment automation, and infrastructure-as-code (Terraform). Familiarity with monitoring, observability, and performanc
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Senior Director, Generative AI About the Role Roblox Build is our generative creation product, the platform where creators design, build, and publish 3D experiences. We are looking for a Senior Director of Generative AI to lead the Applied AI organization inside Build, responsible for turning state-of-the-art foundation models into high-quality, reliable creation systems at Roblox scale. This leader will own the full applied AI stack: model strategy and routing, model adaptation and fine-tuning, code generation (CodeGen), 3D layout generation (LayoutGen), and the evaluation science and infrastructure that tells us what actually works. You Will Own model strategy and routing for Build. Design and build an intelligent model layer that selects the right model for each creation task based on quality, capability, latency, cost, and safety, leveraging both frontier models and Roblox-adapted open-source models. Lead model adaptation across the Applied AI org, including fine-tuning, distillation, synthetic data generation, human feedback pipelines, and preference optimization for Roblox-specific creation tasks such as Luau code generation and 3D scene understanding. Drive CodeGen capabilities
NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology – and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are seeking a Senior Site Reliability Engineer – Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale. What you will be doing: Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security. Capture requirements from partner teams, architect storage solutions, and drive end‑to‑end implementation for new and existing services. Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure. Participate in on‑call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions. Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuous
Get new lead senior engineering manager 2c platform reliability jobs by email
Daily job updates · Unsubscribe anytime