Jobs in Canada

Devops Engineer Observability in San Francisco

9 active opportunities · Updated October 2026

Explore current devops engineer observability jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

46/100

steady · 6 related jobs

Remote options

0%

Share of matching jobs listed as remote

TI
📍 San Francisco, Canada· Full-time
✓ High-confidence listingDemand 46/100

C$113.4K – C$162K/yr

Quick readStrong listing-quality and freshness signals

We believe communication belongs to everyone. We exist to democratize phone service. TextNow is evolving the way the world connects and that's because we're made up of people with curious minds who bring an optimistic, yet critical lens into the work we do. We're the largest provider of free phone service in the nation. And we're just getting started. Join us in our mission to break down barriers to communication and free the flow of conversation for people everywhere. TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between! This role is about impact at scale. You’ll shape how TextNow builds and operates its systems in an AI-first environment where intelligent tooling is embedded into everyday engineering practice. Using AI is not optional, it’s expected. From design and architecture to implementation, testing, debugging, documentation, and operational analysis, you will actively leverage AI tools to increase velocity, improve code quality, and make better technical decisions. We provide a robust suite of AI-powered development tools and workflows to support you, and we expect you to continuously evolve how you use them to raise the bar for efficiency, clarity, and product excellence across the organization. What You'll Do Ensure System Reliability: Design, build, and maintain scalable, resilient, and highly available systems to support TextNow’s infrastructure and services. Automation & Infrastructure as Code: Develop and maintain automation using Terraform, Ansible, and other tools to enable efficient deployment, scaling, and operations of cloud-based systems (AWS preferred). Incident Response & On-Call Support: Participate in an on-call rotation, troubleshoot issues, and drive incident resolution to minimize downtime and improve syste

AWSCI/CDGitAI
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$90K – $115K/yr

Quick readStrong listing-quality and freshness signals

Sigma is growing rapidly, and our Technical Support Engineering team is scaling alongside it to meet the needs of an expanding global user base. As a Technical Support Engineer at Sigma, you will be part of an award-winning team recognized with the 2024 Stevie Gold Award for Customer Service, helping customers solve technical, business, and data challenges using the Sigma platform. You'll work closely with Product, Engineering, and Go-to-Market teams to diagnose complex issues, drive solutions, and contribute to the continuous improvement of our product and support operations. Minimum Education Requirement This position requires a U.S. Bachelor's degree (or foreign equivalent) in Computer Science, Software Engineering, Information Systems, Data Science, or a closely related technical field. This requirement is a minimum and cannot be substituted by work experience alone. What You Will Be Doing You will work with Sigma's customers and the pre-sales team to assist with the diagnosis and resolution of complex technical issues. Working closely with the development team, you will develop best practices and tools for diagnosing issues and optimizing the service for performance. Collaborate with cross-functional groups — backend, frontend, DevOps, design, product, and the go-to-market teams to create a first-class experience for users of our product. Participate in quarterly projects and perform periodic on-call duties to improve automation and processes. Qualifications We Are Looking For 2+ years of experience in a customer-facing technical role (Technical Support, Solutions Engineering, or Software Engineering) at a cloud or SaaS provider. SQL proficiency — strong grasp of JOINs, Partitions, Window Functions, Aggregations, CTEs, and sub-queries. SQL query performance troubleshooting and query plan analysis. Proficient in data modeling concepts. Ability to chart data into logical visualizations. A proven track record of building trust with customers and brin

PythonSQLAWSGCP
A
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$144K – $216K/yr

Quick readStrong listing-quality and freshness signals

Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. Amplitude is looking for a Senior Salesforce Platform Engineer to support the scaling and ongoing evolution of our go-to-market (GTM) systems. You'll own end-to-end visibility into our Salesforce platform, from day-to-day configuration and support to long-term architecture and improvement. As part of our GTM Engineering team, you’ll partner with Sales, GTM Operations, Finance, Product & Engineering, and GTM leadership to deliver well-governed solutions across Salesforce and connected platforms like CPQ, NetSuite, Amplitude, and Snowflake. What You'll Do Hands-on platform administration, configuration, troubleshooting, maintenance, testing, upgrades, and user support Maintain and optimize CPQ, pricing, quoting, and opportunity workflows to support the quote-to-cash process Serve as the escalation line of support for Salesforce, handling escalated i

ReactCI/CDGitAI
A
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$165K – $247K/yr

Quick readStrong listing-quality and freshness signals

Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Role Amplitude's Cloud Platform team builds the systems that every Amplitude engineer relies on every day to ship code — and we're rebuilding them for the AI era. As a Senior Platform Engineer, you'll own medium-to-high-complexity platform projects end-to-end and help shape a platform where AI agents are first-class users alongside humans: kicking off deploys, opening pull requests against infrastructure, and triaging incidents, so a single engineer can get the throughput of a team. You'll partner with Staff engineers and product teams to make Kubernetes effortless across the engineering org, building self-service automation and scalable AWS infrastructure that lets product teams ship faster, safer, and with less cognitive load. If you're excited about building the systems that other engineers will rely on every day, this role is for you. Key Resp

PythonAWSGCPKubernetes
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $288K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. Role Overview As a Senior Staff Frontier Agents Engineer on our Enterprise team, you'll be the technical bridge between Scale AI's cutting-edge AI capabilities and our most strategic customers. You'll work with enterprise clients to understand their unique challenges, architect custom AI solutions, and ensure successful deployment and adoption of AI systems in production environments. This is a hands-on technical role that combines deep engineering expertise with customer-facing problem solving. You'll work directly with customer engineering teams to integrate AI into their critical workflows. Key Responsibilities Customer Integration & Deployment Partner directly with enterprise customers to understand their technical infrastructure, data pipelines, and business requirements Design and implement custom integrations between Scale AI's platform and customer data environments (cloud platforms, data warehouses, internal APIs) Build robust data connectors and ETL pipelines to ingest, process, and prepare customer data for AI workflows Deploy and configure AI models and agents within customer security and compliance boundaries AI Agent Development Develop production-grade AI agents tailored to customer use cases across domains like customer support, data analysis, content generation, and workflow automation Architect multi-agent systems that orchestrate between different models, tools, and data sources Implement evaluation frameworks to measure agent performance and iterate toward business objectives Design human-in-the-loop workflows and feedback mechanisms for continuous agent improvement Prompt Engineering & Optimization Create sophisticate

PythonAWSAzureGCP
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. Role Overview As a Forward Deployed AI Engineering Manager on our Enterprise team, you'll be the technical bridge between Scale AI's cutting-edge AI capabilities and our most strategic customers. You'll work with enterprise clients to understand their unique challenges, lead a team that architects specific AI solutions, and ensure successful deployment and adoption of AI systems in production environments. This is a Management role that combines deep engineering and AI expertise, leading a team, and working on customer-facing problems. You'll work directly with customer engineering teams to integrate AI into their critical workflows. Key Responsibilities Customer Integration & Deployment Partner directly with enterprise customers to understand their technical infrastructure, data pipelines, and business requirements Design and implement custom integrations between Scale AI's platform and customer data environments (cloud platforms, data warehouses, internal APIs) Build robust data connectors and ETL pipelines to ingest, process, and prepare customer data for AI workflows Deploy and configure AI models and agents within customer security and compliance boundaries AI Agent Development Develop production-grade AI agents tailored to customer use cases across domains like customer support, data analysis, content generation, and workflow automation Architect multi-agent systems that orchestrate between different models, tools, and data sources Implement evaluation frameworks to measure agent performance and iterate toward business objectives Design human-in-the-loop workflows and feedback mechanisms for continuous agent improvement Prompt Engineeri

PythonAWSAzureGCP
TI
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$136.3K – $273.9K/yr

Quick readStrong listing-quality and freshness signals

TextNow is on a mission to make communications affordable and accessible for everyone. As a full MVNO operating our own mobile core network over LTE and 5G NSA, we have the unique advantage of controlling our network infrastructure end-to-end. We operate the HSS, PGW, and other critical network functions, giving us the flexibility to innovate and deliver exceptional service to millions of users. About the Role Join us in our mission to break down barriers to communication and free the flow of conversation for people everywhere. T extNow is looking for a new SecOps team member to secure, monitor , and enable automated response within our infrastructure. What You’ll Do Ensure Secure & Reliable Systems: Design, implement, and maintain security-focused infrastructure to protect TextNow’s services while ensuring reliability and scalability. Security Automation & Infrastructure as Code: Develop and enforce best practices using Terraform, Ansible, Crowdstrike , and AWS security tools , ensuring secure configurations, automated compliance checks, and infrastructure as code. Threat Detection & Incident Response: Participate in an on-call rotation to respond to security incidents, investigate vulnerabilities, and implement proactive measures to prevent future threats. Work closely with engineering teams to remediate security risks. Monitoring & Logging for Security: Improve observability by implementing security monitoring solutions, logging best practices, and alerting mechanisms to detect anomalies and suspicious activity. Access Control & Identity Management: Manage IAM roles, permissions, and policies to ensure least privilege access and enforce security controls across cloud and internal systems. Collaboration & Security Advocacy: Wo

AWSCI/CDGitAI
SL
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $230K/yr

Quick readStrong listing-quality and freshness signals

Location - Hybrid This is a hybrid role based out of our San Francisco, Corporate Headquarter office, 3 days in the office, 2 days remote OR one of our Hub locations, Boston, MA or Raleigh, NC, which means you may be expected to work from a designated co-working space from time to time, and will otherwise work remotely from home, until such time as a dedicated office is established. About Us Sauce Labs is the world’s largest full-lifecycle, test automation platform, and the company behind Selenium. Trusted by 80% of the world’s top ten largest financial institutions and over 300,000 enterprise users, Sauce Labs provides the only AI platform capable of turning business intent into autonomous testing and quality assurance. With a proprietary dataset of 8.7 billion test runs, Sauce Labs empowers the Fortune 2000 to bridge the gap between AI-driven code generation and enterprise-grade software quality. Learn more at http://saucelabs.com . Release Assurance at the Speed of AI | Meet the new Sauce Labs The Role: As the Head of Product Marketing (Director Level), you will be a strategic leader responsible for defining and executing the comprehensive product vision for Sauce Labs. You will own the overall product marketing strategy, guiding a high-performing team to drive leadership, accelerate product adoption, and significantly impact revenue growth. This role requires a blend of strategic thinking, hands-on leadership, and a deep understanding of the continuous testing and software development lifecycle market. You will serve as a key bridge between product development, sales, and broader marketing functions, ensuring our market narrative is compelling, differentiated, and aligned with business objectives. Responsibilities: Strategic Leadership & Vision: Define and articulate the overarching product marketing strategy that aligns with company goals and market opportunities. Drive the narrative, positioning, and messagi

AIGoRustDevOps
G
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$140K – $200K/yr

Quick readStrong listing-quality and freshness signals

About Glean: Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry’s most advanced enterprise search has evolved into a full-scale Work AI ecosystem, powering intelligent Search, an AI Assistant, and scalable AI agents on one secure, open platform. With over 100 enterprise SaaS connectors, flexible LLM choice, and robust APIs, Glean gives organizations the infrastructure to govern, scale, and customize AI across their entire business - without vendor lock-in or costly implementation cycles. At its core, Glean is redefining how enterprises find, use, and act on knowledge. Its Enterprise Graph and Personal Knowledge Graph map the relationships between people, content, and activity, delivering deeply personalized, context-aware responses for every employee. This foundation powers Glean’s agentic capabilities - AI agents that automate real work across teams by accessing the industry’s broadest range of data: enterprise and world, structured and unstructured, historical and real-time. The result: measurable business impact through faster onboarding, hours of productivity gained each week, and smarter, safer decisions at every level. Recognized by Fast Company as one of the World’s Most Innovative Companies (Top 10, 2025), by CNBC’s Disruptor 50, Bloomberg’s AI Startups to Watch (2026), Forbes AI 50, and Gartner’s Tech Innovators in Agentic AI, Glean continues to accelerate its global impact. With customers across 50+ industries and 1,000+ employees in more than 25 countries, we’re helping the world’s largest organizations make every employee AI-fluent, and turning the superintelligent enterprise from concept into reality. If you’re excited to shape how the world works, you’ll help build systems used daily across Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, and many more - deeply embedded where people get things done. You’ll ship agentic capabilities on an open, extensible stack, with the craf

AWSAzureGCPGit
🔔

Get new devops engineer observability jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime