Jobiba hiring network

Engineering Excellence Engineer 2 Jobs

8,135 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current engineering excellence engineer 2 jobs. Use filters to narrow by work mode, employment type, experience and date posted.

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. About the Role As a Site Reliability Engineer at Ema, you will own the stability, availability, and operational health of our agentic AI platform across customer environments. You'll work closely with Engineering and DevOps to provision infrastructure, drive deployment excellence, and keep production running at the quality bar our enterprise customers expect — 99.9%+ uptime, proactive incident response, and continuous improvement. What You'll Do Infrastructure & Deployment Design and provision cloud infrastructure (GCP, Azure, AWS) tailored to customer environments, with security, scalability, and compliance built in Execute on-call SaaS deployments with minimal downtime; automate and optimize deployment workflows end-to-end Production Stability & Observability Monitor logs, alerts, and metrics to maintain SLA commitments and catch issues before they escalate Diagnose and resolve production incidents with speed and rigor; drive root cause analysis and permanent fixes Collaborate with DevOps to enhance monitoring dashboards and alerting frameworks; deliver clear system health reporting to internal and customer stakeholders Documentation & Knowledge Management Maintain de

awsazuregcp
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We’re hiring a talented Senior Software Engineer to help us build a world-class security experience for Snowflake engineers. As an engineer in the Product Security team, you will play a central role in delivering the next generation of tools used by our engineers to develop and secure our flagship product. Together with industry-wide experts in distributed systems, cross-cloud development and security excellence, you will evolve our security infrastructure and tooling to be elastic, large-scale, and highly performant with simplicity at its core. As a software engineer in the Product Security Team, you will help Snowflake realize its mission and vision of becoming the premier Data Cloud by accelerating the delivery of our high-quality software to production. You will be a key stakeholder in driving clarity on our strategy and partner with product managers to chart the quarterly and long-term roadmaps for the team. You will ensure that the team is executing towards serving the current needs of its customers, while ensuring that we stay ahead of technological trends and future demands, including AI security and security for AI-driven systems. Our ideal Senior Software Engineer will have: Strong passion for making developers highly productive. 6+ years of industry experience de

pythonjavaaws
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse using open formats like Apache Iceberg — at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world's leading data platforms. We are hiring a Senior Software Engineer for the Observe Data Management team. This team owns the core pipelines that ingest and process over 1 petabyte of telemetry data per day — the foundational infrastructure powering Observe's entire observability stack. You'll be working at the intersection of massive scale, open-source innovation, and real-world reliability challenges for enterprise customers around the globe. AS A SENIOR SOFTWARE ENGINEER - OBSERVE

awsazureai
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause of production issue and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world’s leading data platforms. In this role you will: Develop interactive, data-rich user interfaces using React, TypeScript, and Vega, with a focus on integrating LLM-driven features (e.g., natural language querying, generative UI, and AI-assisted data storytelling). Lead the end-to-end delivery of substantial product features, ensuring AI outputs are presented with high reliability and low latency. Work closely with PMs, UX designers, and AI/ML engineers to bridge t

javascripttypescriptjava
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Product Security team ensures that Snowflake products are built and shipped with the highest level of security. Our team drives the security posture of Snowflake products and is responsible for embedding security into every stage of the product lifecycle, from design through deployment and beyond. We design and build frameworks, systems and services that keep Snowflake secure. As a Principal Software Engineer II on the Product Security team, you will be the senior technical authority for Product Security and play a critical leadership role in shaping and advancing Snowflake’s security. This is a unique opportunity to define and influence our long-term security strategy and have a direct impact on the security of the Snowflake platform and the trust of our customers. You will operate across organizational boundaries, guiding major security initiatives, influencing architectural decisions at the highest levels, setting the technical direction for the organization, and ensuring consistent security excellence across all product teams while working closely with business leaders to advance Snowflake’s business. The role requires deep expertise in security, software engineering, distributed systems, software infrastructure, AI/ML, applied cryptography, threat modeling and clou

pythonjavaai
View job →
T-
Tubi - Canada
📍 Toronto• Full-time• From C$1.2M/yr
18 days ago

About the Role: Tubi is seeking a highly skilled and experienced Senior QA Automation Engineer to lead quality assurance initiatives for our cutting-edge streaming and AI-driven product features. This pivotal role involves ensuring exceptional end-to-end user experiences, robust streaming playback, and the accuracy and integrity of our AI/ML features across web, mobile, and OTT platforms. We're looking for a candidate with a strong background in streaming QA and deep technical knowledge of media workflows. You'll be instrumental in collaborating with engineering, product, and data science teams to define comprehensive QA strategies that guarantee both functional excellence and data-level quality. This is a hybrid role based out of our Toronto office. You must be willing to travel to our Toronto office three days/week. What You'll Do: Design and lead test strategies for streaming workflows, playback systems, and AI-powered features. Test across platforms (web, mobile, and connected TV) to ensure functional parity and playback stability. Validate streaming performance—including ABR logic, encoding pipelines, and DRM integrations—under diverse real-world conditions. Debug with precision using tools like Charles Proxy, Chrome DevTools, ADB, and Xcode. Collaborate with data and ML teams to validate AI model updates, recommendations, and personalization accuracy. Leverage AI-assisted QA tools to enhance regression coverage, UI validation, and anomaly detection. Contribute to automation and CI/CD frameworks, driving faster, more reliable releases. Help drive a shift-left testing approach by engaging early in the software development lifecycle, partnering with product managers, engineers, and data scientists to identify quality risks, define test strategies, and ensure testability during requirements and design phases. Oversee QA deliverables for multiple concurrent releases and ensure seamless sign-off for production launches. Monitor live environments for playback or reco

javascripttypescriptpython
View job →
T-
Tubi - Canada
📍 Toronto• Full-time• From C$1.4M/yr
18 days ago

About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission is to engineer resilience from the ground up, enabling our product teams to innovate rapidly while ensuring our users have a stellar experience. We own the availability, latency, performance, and capacity of our platform, and we achieve our goals through a culture of data-driven decision-making, blameless learning, and relentless automation. As a Senior Site Reliability Engineer, you are a hands-on engineer who blends deep software development expertise with a passion for operational excellence. You will be responsible for designing, building, and running the resilient, scalable, and increasingly self-healing systems that power our products. You will apply sound engineering principles to solve our most complex reliability challenges, with a mandate to automate everything, eliminate toil, and write robust, maintainable code. You will be a force multiplier, mentoring other engineers and elevating the site reliability bar for the entire organization. This is a hybrid role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: System Architecture & Design: Design, build, and maintain scalable, highly available, and fault-tolerant distributed systems. Partner with development teams as a reliability consultant, reviewing designs and influencing architectural decisions to ensure new services are built with reliability, observability, and performance as core principles, not afterthoughts. Automation & Software Development: Write robust, performant, and maintainable code to automate operational tasks, and CI/CD pipelines. Build the internal tools, libraries, and frameworks that enable engineering teams to self-service their

typescriptpythonaws
View job →
SL
18 days ago

Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida/ Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite

pythonjavareact
View job →
SL
18 days ago

Title: Senior Site Reliability Engineer - I, Product Area Focus Location: Noida (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering teams. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedit

pythonjavareact
View job →
SL
18 days ago

Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida / Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite

pythonjavareact
View job →
A
Affirm
📍 Spain• Full-time• Remote
18 days ago

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Affirm is seeking a Staff Full Stack Software Engineer to join the Acquisition & Onboarding team within the Consumer org. This is a high-impact leadership role for an engineer who can set technical direction, drive architectural decisions, and elevate the quality and velocity of full stack development across mobile, web, and backend systems. The team plays a critical role in shaping the first experience customers have with Affirm—building trust, clarity, and value from the very first interaction. As a Staff Engineer, you will be responsible for defining long-term technical strategy, mentoring senior engineers, and acting as a force multiplier through your technical depth, operational excellence, and ability to navigate ambiguity. You'll work at the intersection of product, design, and engineering to build polished, performant, and accessible user experiences that directly impact conversion, retention, and business growth. What You'll Do You will be responsible for setting technical strategy for your team on a year-long time scale, and help your team tie it together with critical, business-impacting projects. You will collaborate across teams in the product development lifecycle by collaborating with product management, design & analytics to ensure technical sustainability, risks and trade-offs are well understood and managed. You will act as a force-multiplier for your team through your definition and advocacy of technical solutions and operational processes. You take ownership of your team’s operations and availability by ensuring you have the right monitoring, triage rotations, playbooks, policies, testing and alerting in place to support “keep the lights on” & on-call efforts. You will foster a culture of quality and ownership on your team by setting code review a

REMOTEpythonreactvue
View job →
A
Affirm
📍 Poland• Full-time• Remote
18 days ago

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Affirm is seeking a Staff Full Stack Software Engineer to join the Acquisition & Onboarding team within the Consumer org. This is a high-impact leadership role for an engineer who can set technical direction, drive architectural decisions, and elevate the quality and velocity of full stack development across mobile, web, and backend systems. The team plays a critical role in shaping the first experience customers have with Affirm—building trust, clarity, and value from the very first interaction. As a Staff Engineer, you will be responsible for defining long-term technical strategy, mentoring senior engineers, and acting as a force multiplier through your technical depth, operational excellence, and ability to navigate ambiguity. You'll work at the intersection of product, design, and engineering to build polished, performant, and accessible user experiences that directly impact conversion, retention, and business growth. What You'll Do You will be responsible for setting technical strategy for your team on a year-long time scale, and help your team tie it together with critical, business-impacting projects. You will collaborate across teams in the product development lifecycle by collaborating with product management, design & analytics to ensure technical sustainability, risks and trade-offs are well understood and managed. You will act as a force-multiplier for your team through your definition and advocacy of technical solutions and operational processes. You take ownership of your team’s operations and availability by ensuring you have the right monitoring, triage rotations, playbooks, policies, testing and alerting in place to support “keep the lights on” & on-call efforts. You will foster a culture of quality and ownership on your team by setting code review and de

REMOTEpythonreactvue
View job →
GR
18 days ago

Role: VPN Engineer Location: Gurgaon Graviton is a privately funded quantitative trading firm striving for excellence in financial markets' research. We are seeking a Network Engineer for our team in Gurgaon. Graviton trades across a multitude of asset classes and trading venues using a gamut of concepts and techniques ranging from time series analysis, filtering, classification, stochastic models, pattern recognition to statistical inference analysing terabytes of data to come up with ideas to identify pricing anomalies in financial markets. Responsibilities Manage and support corporate network infrastructure across multiple locations, including routers, firewalls, switches, wireless access points, VPN gateways, Internet links, and LAN/WAN connectivity. Configure and troubleshoot VPN technologies such as IPsec, SSL VPN, site-to-site VPN, remote-access VPN, WireGuard, OpenVPN, and FortiClient/FortiGate VPN, including issues related to authentication, tunnels, routing, DNS, packet loss, performance, split tunnelling, and firewall policies. Manage secure connectivity between offices, data centres, cloud environments, and remote users. Configure and maintain office LAN infrastructure, including VLANs, trunk/access ports, inter-VLAN routing, DHCP, DNS, NAT, ACLs, static routing, BGP, and OSPF where required. Manage multiple ISP connections, including primary and backup Internet links, automatic failover, and monitoring of utilization, latency, jitter, packet loss, and link availability. Coordinate with ISPs and telecom providers for new circuits, link failures, bandwidth upgrades, routing issues, packet-loss investigations, and service escalations. Manage firewall policies, NAT rules, VPN policies, network objects, and routing, while regularly reviewing and removing unnecessary access. Implement network segmentation across user, server, management, guest, and other business networks, while maintaining secure administrative access to network equip

pythonlinuxrest
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
18 days ago

Opportunity Overview: We are seeking a Senior Data Engineer to contribute to the design and delivery of our cloud-native healthcare data platform. You will implement scalable data solutions built on AWS, Apache Iceberg, Lake Formation, Glue Catalog, Athena, dbt, and modern orchestration frameworks. This role combines strong hands-on engineering with collaboration across platform, analytics, and business teams. What You'll Do Data Engineering Delivery Deliver complex data engineering projects in collaboration with cross-functional teams Drive technical execution from design through production deployment Implement scalable data patterns and reusable frameworks Design and implement batch and near-real-time pipelines Build reusable ingestion, transformation, validation, and publishing frameworks Support modernization of legacy workloads Contribute to Apache Iceberg implementation and optimization Apply standards for schema evolution, partitioning, compaction, and metadata management Ensure efficient storage and query performance Implement data quality frameworks and validation layers Support observability and monitoring practices Contribute to operational excellence and reliability improvements Participate in architecture and design discussions Conduct and participate in code reviews Mentor junior engineers and share best practices ISMS roles and responsibilities Good knowledge of Information security Oversee specific business processes within the ISMS. Responsible to manage the ISMS documentation, conduct risk assessments, and implement risk treatment plans. Risk Owners are responsible for identifying, assessing, and managing risks within their areas of responsibility. They are also responsible for implementing risk treatment plans. Conduct the BCP and other test related to information security continuity along with CISO Responsible for monitoring and reporting on the performance of the ISMS. Responsible for implementation of security policies and procedures and report

pythonsqlaws
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
18 days ago

Opportunity Overview: We are seeking a Lead Data Engineer to drive the design and delivery of our cloud-native healthcare data platform. You will lead the implementation of scalable data solutions built on AWS, Apache Iceberg, Lake Formation, Glue Catalog, Athena, dbt, and modern orchestration frameworks. This role combines deep hands-on engineering with technical leadership and collaboration across platform, analytics, and business teams. What you’ll do: Lead Data Engineering Initiatives Lead delivery of complex data engineering projects across multiple teams Drive technical execution from design through production deployment Establish scalable implementation patterns Build and Optimize Data Platforms Design and implement batch and near-real-time pipelines Build reusable ingestion, transformation, validation, and publishing frameworks Support modernization of legacy workloads Lakehouse Engineering Lead Apache Iceberg implementation and optimization Define standards for schema evolution, partitioning, compaction, and metadata management Ensure efficient storage and query performance Data Quality and Reliability Implement data quality frameworks Drive observability and monitoring practices Improve operational excellence and reliability Technical Leadership Review architecture and design proposals Conduct code reviews and engineering reviews Mentor engineers and establish best practices ISMS roles and responsibilities: Good knowledge of Information practices. Assist the manager in all the information security activities implementation and maintenance process. Ensuring the team and imparted with Competence related to Information security Responsible for implementation of security policies and procedures and report any issues to the Information Security Manager. Required Qualifications: 8–12 years of Data Engineering experience. Experience leading enterprise-scale data initia

pythonsqlaws
View job →
🔔

Get new engineering excellence engineer 2 jobs by email

Daily job updates · Unsubscribe anytime