This is where your work makes a difference. At Baxter, we believe every person—regardless of who they are or where they are from—deserves a chance to live a healthy life. It was our founding belief in 1931 and continues to be our guiding principle. We are redefining healthcare delivery to make a greater impact today, tomorrow, and beyond. Our Baxter colleagues are united by our Mission to Save and Sustain Lives. Together, our community is driven by a culture of courage, trust, and collaboration. Every individual is empowered to take ownership and make a meaningful impact. We strive for efficient and effective operations, and we hold each other accountable for delivering exceptional results. Here, you will find more than just a job—you will find purpose and pride. Your Role at Baxter As a Principal DevOps Engineer, you will provide technical leadership in the design, implementation, deployment, and support of cloud and platform engineering solutions that enable product development teams. You will partner across engineering functions to build scalable, reliable, and secure infrastructure while helping drive operational excellence through automation, observability, and continuous improvement. What You'll Do: Design, implement, and support cloud infrastructure solutions that enable the development and operation of Baxter products and applications. Manage infrastructure, platform, and deployment processes for product teams to ensure reliability, scalability, and performance. Deploy, manage, and troubleshoot containerized applications using Kubernetes and cloud-native technologies. Develop and maintain Infrastructure as Code (IaC) solutions using Terraform to automate infrastructure provisioning and management. Implement and support observability, monitoring, and alert
Jobs in United States
Devops Engineer Observability Manager Specialist in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current devops engineer observability manager specialist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Hiring demand
35/100
watch · 21 related jobs
Hiring trend
-38.5%
Job postings compared with the previous 30 days
Remote options
14.3%
Share of matching jobs listed as remote
Typical salary
$128K – $128K/yr
Based on 7 salary observations
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Observability team builds the infrastructure that empowers engineers to understand, operate, and improve the Roblox platform and ecosystem. Our team owns the end-to-end observability stack across telemetry, distributed tracing, logging, profiling, storage systems, and developer-facing visualization tools. We are looking for an Engineering Manager to lead the next generation of AI-powered observability platforms. In this role, you will help build intelligent systems that leverage AI to revolutionize CI/CD, testing, and DevOps workflows — enabling engineers to move faster, improve reliability, and operate large-scale distributed systems with greater efficiency and confidence. This is a highly impactful leadership role at the center of Roblox infrastructure. Your work will directly improve developer productivity, platform reliability, and operational excellence across the company. You will partner closely with infrastructure, product engineering, and AI platform teams to shape the future of developer tooling and autonomous operations at scale. You Have 3+ years of engineering management experience with a proven track record of hiring, mentoring, and growing high-performing teams. Strong ex
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Senior Manager, Platform Engineering / DevOps Who Are You You are an experienced Senior Manager / emerging Staff-level leader in DevOps and Platform Engineering with strong technical depth and demonstrated leadership in delivering enterprise-scale cloud platforms. You bring a balanced mix of hands-on engineering expertise, team leadership, and execution rigor. You excel in driving outcomes in complex, multi-stakeholder environments, guiding teams to deliver secure, scalable, and high-quality platform solutions. You are comfortable leading engineers, managing stakeholders, and owning delivery across multiple workstreams. You demonstrate: A strong ownership mindset with accountability for delivery and outcomes Ability to translate business needs into actionable engineering roadmaps Solid expertise in cloud-native platforms, DevOps practices, and SRE principles Capability to lead teams and influence without requiring extensive tenure Role Responsibilities Development & Enforcement Own and execute the H100 platform engineering roadmap, aligned to enterprise priorities and program milestones Drive delivery of GCP-based platform capabilities (GKE, networking, IAM, CI/CD, observability) Establish and enforce engineering standards, best practices, and ADR compliance <li
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Senior Manager, Platform Engineering / DevOps Who Are You You are an experienced Senior Manager / emerging Staff-level leader in DevOps and Platform Engineering with strong technical depth and demonstrated leadership in delivering enterprise-scale cloud platforms. You bring a balanced mix of hands-on engineering expertise, team leadership, and execution rigor. You excel in driving outcomes in complex, multi-stakeholder environments, guiding teams to deliver secure, scalable, and high-quality platform solutions. You are comfortable leading engineers, managing stakeholders, and owning delivery across multiple workstreams. You demonstrate: A strong ownership mindset with accountability for delivery and outcomes Ability to translate business needs into actionable engineering roadmaps Solid expertise in cloud-native platforms, DevOps practices, and SRE principles Capability to lead teams and influence without requiring extensive tenure Role Responsibilities Development & Enforcement Own and execute the H100 platform engineering roadmap, aligned to enterprise priorities and program milestones Drive delivery of GCP-based platform capabilities (GKE, networking, IAM, CI/CD, observability) Establish and enforce engineering standards, best practices, and ADR compliance</li
From $151K/yr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are seeking an experienced and vision-driven Lead Enterprise Systems Engineer to join our engineering team. In this role, you will bridge the gap between business objectives, solution architecture, and hands-on execution. The ideal candidate remains actively involved in coding (roughly 70–80% of the time) while serving as the primary technical point of contact for project stakeholders. What you'll do Technical Vision & Solution Architecture Lead the architectural design, development, and deployment of resilient, scalable solutions in our Salesforce Platform for both Sales & CPQ. Translate business and product requirements into clear, technical roadmaps and system specifications. Establish engineering best practices, design patterns, coding standards, and testing strategies. Hands-On Execution & Quality Assurance Write clean, maintainable, and highly efficient APEX code alongside the Salesforce development team. Conduct thorough code reviews to ensure quality, security, and performance. Manage technical debt, proactively balancing speed of delivery with long-term system health. Team Leadership & Mentorship Provide technical guidance, direct support, and actionable feedback to Salesforce engineers. Mentor team members to foster technical growth and career advancement. Lead agile ceremonies (sprint planning, daily stand-ups, technical grooming, post-mortems). Cross-Functional Collaboration Partner closely with Technical Managers, Enterpri
We're on a mission to build the best platform in the world to defend the enterprise from code-to-cloud-to-runtime. Used by thousands of companies globally, Datadog security products uniquely leverage Datadog’s unified security and observability platform so Security, DevOps and SRE can collaborate rapidly and seamlessly to deliver better detection, prioritization and remediation. Our product and engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. In this competitive market, the Group Product Manager for Code Security will play a mission-critical role in providing product and strategy leadership to grow Datadog’s market share through differentiation, innovation and compelling customer value. This leader will lead a talented and growing team of product managers and work with world class engineers to build and grow multiple Code Security products that play an essential role for our customers’ code security programs, and growing Datadog into a security industry leader. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Run and grow multiple Code Security products to meet revenue and business targets with the goal of building a multi-hundred million dollar annual business. Lead and own product strategy and roadmap for accountable security products, fully aligned to revenue and business goals and with compelling differentiation and customer value. Ensure predictable roadmap execution across direct and partner teams to achieve product and business outcomes required to meet the revenue and business goals. Analyze and develop pricing and packaging strategies to maximize revenue through attaching deep understanding of market dynamics and other strategic leverage points. Drive GTM strategy with GTM partner teams
From $80K/yr
Datadog Sales Engineers help qualify and close opportunities with customers and partners by providing technical expertise through sales presentations, product demonstrations, and supporting technical evaluations (POCs). Sales Engineers have a voice with the product team to prioritize features based on input from customers, competitors, and partners. Additionally, you will work with various cross-functional teams to resolve customer concerns, escalate issues, advocate for customer needs, and serve as an ambassador for our brand. If you want to join a friendly, passionate team with limitless potential, we'd love to meet you! At Datadog, we place value in our office culture, the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Partner with the Customer Success team to articulate Datadog's value proposition, vision, and technical strategy throughout the sales cycle. Continually learn new technologies to build competitive knowledge, technical skill, and credibility Deliver value-driven product and technical presentations to Datadog customers Have a direct line of communication with the product team to collaborate on feature requests Help clients onboard the product and assist when they run into roadblocks Think creatively to solve a wide variety of technical challenges throughout the sales cycle Who You Are: Has 1 year or more of experience in a technical sales and/or post-sales role Has working knowledge of cloud infrastructure, modern application architectures, observability, and/or DevOps technologies Enjoys collaborating and teaming up with others Demonstrates strong written and oral communication skills Comfortable and confident in delivering technical presentations & demos Can multi-task, manage time well, and be responsive to others Curious, adaptable, and eager to continually learn and expand technica
At Render, we’re building the modern cloud platform for developers creating AI-native, full-stack, multi-service applications. Our mission is to eliminate the tradeoff between the power of hyperscalers and the simplicity of developer-friendly platforms—so teams can ship fast, scale reliably, and focus on their product, not infrastructure. Unlike complex hyperscalers or ephemeral edge/serverless solutions, Render offers a developer-first experience with persistent compute, dynamic autoscaling, built-in orchestration, and observability, allowing teams to launch, scale, and manage real-world applications without writing infrastructure code or managing servers. Whether you're building LLM-powered applications, scalable SaaS products, or async processing pipelines, Render empowers teams to move fast and scale confidently from MVP to millions of users. Our platform is trusted by over 6.5 million developers worldwide and continues to grow rapidly. In February 2026, we raised an additional $100M in Series C financing, bringing our total funding to $257M, to accelerate our vision of making cloud infrastructure both powerful and intuitive—designed for the speed of modern AI development. We’re a diverse and talented team that values craft, velocity, and user experience. If you’re excited to help shape the future of the intelligent cloud and empower developers everywhere, we’d love to hear from you. Applying to Render We're seeking candidates who possess high integrity, humility, and an insatiable drive to learn. Through reasoned discussions and continuous feedback, we strive to improve both individually and collectively. We foster an environment of mutual trust and respect, empowering effective debate to achieve the best outcomes for our customers and team. We especially encourage members of underrepresented groups in the tech community to apply and understand that not all successful candidates will meet each requirement listed. Our interview process is unique to each role, an
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. REQUIREMENTS: Product Positioning & Technical Narrative Own the positioning and messaging for Baseten’s dedicated inference and platform capabilities, including autoscaling, routing, failover, release safety, observability, and cost/performance capabilities. Translate infrastructure-heavy product work into buyer narratives for ML engineering, platform engineering, security, compliance, procurement, and executive audiences. Partner with Product and Engineering to understand the technical architecture, customer value, roadmap tradeoffs, and proof points behind each capability. Define when a capability should be positioned as a platform differentiator, a dedicated inference requirement, a reliability story, a compliance story, or sales enablement. Build messaging that is technically credible without being overly implementation-focused or generic. Launch Strategy & GTM Execution Build and execute launch plans for major dedicated inference and serving platform capabilities, from early internal enablement through external announcement. Decide what deserves a full launch versus what should ship through docs, sales enablement, customer-specific materials, or targeted enterprise outreach. Create launch assets including messaging briefs, landing pages, blog posts, sales decks, one-pagers, FAQs, demo storylines, competitive talk tracks, and customer-facing proof points. Sequence launches and supporting assets based on custo
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. As a Senior Solutions Engineer for Observe, you will play a critical, highly-visible, role within our organization, serving as the primary technical resource for our Sales team. You will be responsible for driving the technical closure of sales opportunities by demonstrating the value of Observe to prospective customers. This role requires a blend of deep technical knowledge, strong presentation skills, and a customer-focused approach. KEY RESPONSIBILITIES: Technical Discovery and Presentation: Conduct in-depth technical discovery sessions with prospects to understand their current environment, challenges, and specific observability requirements. Tailor and deliver compelling product demonstrations and technical presentations that showcase how our solution addresses their needs. Proof of Concepts (POCs): Design, scope, and manage te
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking an experienced AI Systems Reliability Engineer to help define, build, and maintain the infrastructure and processes that ensure the reliability, scalability, and performance of Postman’s AI-powered API and agentic systems in production. This role focuses on monitoring, availability, incident response, and automation to support AI services and tools trusted by millions of developers globally. What You’ll Do Develop and manage reliability metrics (SLOs) for AI-driven API services and agentic AI platform features Implement comprehensive observability and monitoring systems for real-time performance and fault detection Design and drive automated failover, recovery, and incident response strategies for high-availability AI infrastructure Optimize resource utilization, particularly GPU/accelerator efficiency, ensuring cost-effective AI system operation Collaborate closely with engineering, platform, and product teams to align reliability efforts with broader organizational goals Lead efforts to build internal tooling and automation focused on AI system stability and operational excellence Drive continuo
From $98K/yr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity At New Relic, we provide our customers with real-time insights, so they can innovate faster. Our software provides deep observability across the stack, enabling software teams to solve their customer’s problems, accelerate digital transformation, and make DevOps work. You will be at the heart of the teams supporting New Relic’s infrastructure and will work on a team that provides global service mesh and load balancing solutions. We provide these services on-premises, as well as using our multi-cloud infrastructure. We support each other to do our best work through positive communication and continuous improvement. What You’ll Do As a key member of our Infrastructure team, you will design and operate a scalable, resilient ingress data plane that directly impacts the value we provide to our customers. By ensuring the stability and performance of our global service mesh and load balancing solutions, you drive the foundational reliability that the entire New Relic organization depends on to deliver real-time insights. You will leverage advanced automation and infrastructure-as-code to accelerate development speed, allowing our engineering teams to ship safe, incremental changes across a massive fleet with confidence. Your work in evolving our DNS and CDN infrastructure is not just about maintenance; it is about creating a seamless, high-performance environment that enables innovation at scale. Through deep collaboration with Product, Design, and partner platform t
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. We’re looking for an Implementation Engineer to help enterprise customers successfully deploy, configure, and operationalize Observe. This is a hands-on, post-sales technical role focused on delivering strong first outcomes, accelerating time-to-value, and establishing a solid foundation for long-term customer success. Implementation Engineers are deeply technical, customer-facing practitioners who work closely with customer platform, SRE, DevOps, and application teams during onboarding and early adoption. In this role, you’ll translate existing observability architectures (including OpenTelemetry-based pipelines, Splunk, ELK, and other monitoring solutions) into scalable, production-ready implementations on Observe—using best practices while balancing speed, quality, and customer enablement. Implementation Engineers focus on initia
About the Team API Enterprise Controls is part of the API Infrastructure organization and owns the platform capabilities that help developers, startups, and enterprises adopt the OpenAI API securely and confidently. We build the systems underneath our APIs and developer platform across authentication and identity, service accounts and key management, secure networking, compliance, auditability, observability, and operational controls. Our users are developers and teams running critical applications on OpenAI, and we partner closely with Product, go-to-market, security, and infrastructure teams to turn their most important needs into reliable, intuitive platform capabilities. About the Role We are looking for an exceptional backend software engineer to help define and ship the enterprise capabilities our API Platform needs to scale.; this is a product-engineering role grounded in deep backend systems. You will work across databases, streaming systems, request routing, authentication, and developer-facing APIs while bringing strong product judgment, developer empathy, and attention to the small details that make a platform easier to understand, trust, and operate. You will lead large cross-functional initiatives, work closely with Product and go-to-market teams, engage directly with sophisticated users, and carry ambiguous needs from discovery through design, launch, and iteration. In this role, you will: Own backend product capabilities end to end across authentication and identity, service accounts and key controls, secure networking, compliance, observability, and operational workflows. Partner with Product, go-to-market, security, infrastructure teams, and sophisticated customers to identify needs, shape the roadmap, and lead large cross-functional projects from design through launch. Design developer-facing APIs, system behavior, configuration, error handling, safe defaults, auditing, and notifications with exceptional care for the details that define a great dev
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lake using open formats like Apache Iceberg, delivering deep correlation and long-term analytics at dramatically lower cost. A dynamic Knowledge Graph and chat-based AI SRE provide rich context and guided workflows so teams can move from detection to root cause and resolution significantly faster. The Infrastructure team at Observe by Snowflake is responsible for building, scaling, and operating the development and production environments that power our observability platform. We are a small, highly collaborative team with a broad scope, focused on delivering reliable infrastructure while continuously improving the systems that support our engineers and customers. What You’ll Do Design, build, and operate scalable cloud infrastructure in AWS supporting a high-scale observability platform. Improve system reliability, performance, and operational visibility across development and production environments. Develop and maintain CI/CD pipelines and internal tooling to improve developer productivity and deployment safety. Identify and mitigate security risks, and help maintain intern
Higher-paying openings
Jobs with higher listed pay
Related career options
Similar roles with stronger pay
Demand 33/100 · 7 jobs
$4.6M – $4.6M/yr
Salary →Demand 33/100 · 15 jobs
$1.8M – $1.8M/yr
Salary →Demand 46/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 5 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 44/100 · 18 jobs
$345K – $345K/yr
Salary →Other cities to consider
More places hiring for this role
Get new devops engineer observability manager specialist jobs in United States by email
Daily job updates · Unsubscribe anytime