Jobiba hiring network

Detection And Mitigation Engineer Jobs

2,196 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current detection and mitigation engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. Today, we are tapping into the unlimited potential of AI to define the next era of computing. As an NVIDIAN, you will address challenges spanning architecture, silicon, firmware, software, and production — and excellent judgment matters as much as technical depth! We are the Silicon Power Team within the Silicon Co-Design Group. We architect and deliver groundbreaking solutions for productizing NVIDIA's chips across consumer, professional, server, embedded, mobile, and automotive markets. Silicon characterization, correlation to arch and design expectations, product spec finalization, and productization techniques and infrastructure are our day-to-day work — always on the bleeding edge of the industry. Small decisions here have outsized impact on performance, efficiency, reliability, bring-up speed, and ultimately what the product delivers in the field. We are hiring a Senior Silicon Power Engineer to own power-feature productization on a flagship silicon program. This is not a coordination role, and it is not a compliance role — it is the seat where power features either work at scale or become the reason a program slips. The two highest-leverage problems in this seat: Close the hardest multi-functional power failures before they gate a program. Take ambiguous, cross-boundary issues across architecture, firmware, validation, and platform to root-cause closure — with productized fixes and reusable methodology the next program can inherit. Build AI-enabled characterization as a real capability, not a demo. Every bring-up generates terabytes of characterization, shmoo, and telemetry data. Deploy AI workflows for data analysis, metric extraction, trend detection, and cross-bring-up correlation — with the guardrails and validation discipline to make them trustworthy enough to gate production decisions! <

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The Realtime Defect Analysis (RDA) Yield Technology team supports advanced DRAM research and development by identifying, analyzing, and reducing manufacturing defects that impact yield and product performance. The team partners closely with engineers across R&D and High Volume Manufacturing to improve process stability, accelerate learning cycles, and drive continuous improvement through data-driven decision making. As an RDA Yield Technology Intern, you will gain hands-on experience with innovative semiconductor processing, defect inspection systems, and advanced analytical techniques. You will contribute to projects focused on experimentation, defect detection, data analysis, and process optimization while collaborating with multi-functional engineering teams. This role is ideal for individuals who are passionate about problem solving, root-cause investigation, and leverausingology to improve manufacturing performance. Responsibilities Analyze experimental process flows and defect inspection data to identify anomalies, investigate root causes, and support corrective actions. Partner with R&D and manufacturing teams to improve yield performance through defect detection, process monitoring, and continuous improvement initiatives. Leverage Artificial Intelligence (AI) and data analytics tools to accelerate defect detection, identify yield-impacting trends, automate routine analysis workflows, and generate actionable insights. Use inline defect signals, statistical analys

pythonsqlartificial intelligence
View job →

NVIDIA has transformed computer graphics, PC gaming, and accelerated computing for more than 25 years through exceptional technology and the people who build it. In semiconductor manufacturing, our role is to enable the ecosystem, not compete within it. We partner with fabs, equipment manufacturers, and software providers to make inspection, metrology, and manufacturing intelligence dramatically faster on the NVIDIA platform. Our team builds the software that makes this possible: models, adaptation and evaluation workflows, and deployable inference capabilities that partners integrate into their own tools. We work in environments where labeled data is limited and proprietary, distributions shift across tools and fabs, production budgets are tight, and software must operate inside air-gapped facilities. We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa Clara. This is a hands-on architect role: you will define the approach, build it, evaluate it, and demonstrate the results. You will work across computer vision, time-series modeling, multimodal AI, anomaly detection, model adaptation, evaluation, and production inference. Success means technology that a fab or equipment vendor can integrate, operate, and trust—not only a successful internal demonstration. What you’ll be doing: Define and prototype AI system architectures spanning optical and e-beam inspection, wafer and mask inspection, metrology, defect review, equipment signals, and process data. Advance world foundation model capabilities for semiconductor manufacturing, including vision, time-series and multimodal representation learning, model adaptation, domain transfer, and data-scarce defect understanding. Develop workflows for defect detection, classification, localization, segmentation, nuisance filtering, ADC, AD

pythonmachine learningai
View job →

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. What you’ll be doing: Use and develop AI-powered tools to make software testing smarter, faster, and more effective! Improve test case generation, defect detection, flaky test analysis, regression testing, and test coverage optimization. Work with product, engineering, and cross-functional teams to review requirements and define test strategies. Build test plans, design and execute test cases, and report quality status, risks, bugs, and results. Perform functional, performance, fault-injection, reliability, and regression testing for cloud-native systems. Automate test cases and contribute to scalable test frameworks. Manage the bug lifecycle, reproduce customer issues, and verify fixes. What we need to see: MS or PhD in Computer Science, Engineering, or a related field. 5&#43; years of QA, test automation, or software testing experience. Hands-on experience using AI tools to improve QA workflows. Strong QA fundamentals, test strategy, test planning, and failure analysis skills. Proficiency with Unix/Linux and shell or Python programming. Exp

pythonkuberneteslinux
View job →
P
Plaid
📍 San Francisco• Full-time
1mo ago

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Our Fraud team's mission is to help companies detect and prevent fraud using Plaid's financial network data. We believe that transaction patterns, device signals, identity linkages, and behavioral data are dramatically underleveraged tools in fraud prevention. Our products — including Protect and Signal — operate at network scale and depend on real-world investigation and research to stay ahead of adaptive adversaries. As a Senior Fraud Researcher, you will sit at the intersection of live fraud investigation, applied data science, and product innovation. You will lead complex investigations, translate findings into detection improvements, and collaborate tightly with Data Science, ML, and Product teams to shape the next generation of Plaid's fraud capabilities. This is not a purely operational role — your research directly drives features, model inputs, and product design. Responsibilities: Live Fraud Investigation & Reconstruction Lead investigations into complex fraud cases across identities, accounts, devices, and transaction surfaces Provide support to day-to-day fraud operations including SEVs and alert triage Reconstruct attacker sequences and hypothesize actor intent and tooling Distill p

pythonsqlaws
View job →
G
Godaddy
📍 Bulgaria• Full-time
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Our Global Sustaining Engineering team sits at the intersection of software engineering and infrastructure, ensuring the services our customers depend on are fast, resilient, and always available. As a Senior Site Reliability Engineer, you'll take direct ownership of production services — from initial design through day-to-day operation — while partnering with product, engineering, and security teams to build and maintain business-critical systems. In this role, you will deepen your technical expertise and grow your leadership presence by mentoring the next generation of SREs. You will also gain hands-on experience with intelligent tooling in real-world workflows. What you'll get to do... Design, implement, and operate scalable, highly available production services while diagnosing and resolving complex infrastructure, network, and application issues Build and maintain alerting pipelines, dashboards, and SLO-driven monitoring strategies using Icinga, Prometheus, and Grafana Lead incident response end-to-end — performing root-cause analysis, authoring blameless post-mortems, and driving corrective actions to closure Develop and extend Infrastructure as Code coverage and build internal tooling that eliminates manual, repetitive operational work Mentor SRE I and SRE II engineers through code reviews, debugging sessions, and knowledge-sharing talks Apply LLM-driven log analysis, anomaly detection, and generative AI tools to accelerate incident response and runbook creation — validating all outputs before use Your experien

pythondockerkubernetes
View job →
P
Pendo
📍 Raleigh• Full-time• From $150K/yr
1mo ago

Sr. Product Manager, Analytics The team + the role Pendo is the AI-powered analytics and adoption platform for builders. The Product team builds the tools that help software teams understand how their products are used, where adoption breaks down, and what drives business outcomes. AI features are increasingly central to the analytics roadmap — surfacing intelligent insights, automating pattern detection, and embedding AI-assisted workflows directly into the product experience. This team is embedded in its own customer base, runs continuous discovery, and treats customer outcomes as the primary measure of whether the work is good. This role owns an analytics product area end-to-end: strategy, roadmap, discovery, and delivery across a globally distributed cross-functional team spanning Raleigh, New York, and Pendo's India-based engineering and product teams. Customer outcomes are the starting point for every significant decision. Discovery is continuous, not episodic. Operating across time zones is a practical requirement — the person in this role adapts their working hours to maintain meaningful overlap with India-based teams and builds the async artifacts that let those teams work independently between overlap windows. This role is based in Raleigh, NC (primary), with New York, NY and East Coast remote candidates also considered. Pendo's hybrid model requires in-office 3 days per week for local candidates. What this looks like day-to-day Own the analytics product roadmap — define priorities, defend tradeoffs with data and reasoning, and keep the long-term direction aligned with enterprise customer needs and the broader Pendo platform. Drive AI-powered capabilities within the analytics area: define what the AI does, how outputs are surfaced, how confidence is communicated, and where human judgment stays in the loop — with the customer problem as the starting point, not the model capability. Run continuous discovery with enterprise customers and India-based product a

R
Roblox
📍 San Mateo• Full-time• From $196.8K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. WHY SAFETY? At Roblox, we strive to connect a billion people with optimism and civility, and the Safety organization's mission is to become the leader in civil immersive online communities. We systematically and proactively detect, remove, and prevent problematic content and behavior, and we make Roblox accounts secure and free from compromise. We also keep the platform compliant for changing regulations and growth markets. We cover a broad area of the tech spectrum, including machine learning, experimentation, automation, highly scalable distributed backend systems, detection workflows, and AI-powered text filters. Aligned and partnering with product teams, we use this tool belt to discover new opportunities, influence and shape the product roadmap and prioritization, build safety products, and measure the impact on our community of users and developers. In doing so, we keep Roblox safe, civil, and inclusive, and we foster positive relationships between people around the world. WHY CONTENT SUITABILITY? Join the Content Suitability team and play a pivotal role in shaping the future of content on Roblox. Our team is at the forefront of building tools and systems that enable creators to launc

javaawsgit
View job →
T
Twilio
📍 - US• Full-time• Remote• $155.5K – $194.4K/yr
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Machine Learning Engineer. About the job This position is needed to drive innovation and the development of cutting-edge products that serve developers, builders, and operators within Twilio’s Data & Observability Substrate organization. This is a hands-on, builder-focused engineering role that bridges Product, Design, and Engineering to develop, evaluate, and maintain scalable, low-latency, ML-based systems for real-time applications. You will lead rapid research-to-production cycles that translate business ideas into solutions for complex problems—such as streaming anomaly detection, recommendation systems, predictive modeling, and agentic AI frameworks—with the goal of delivering personalized customer experiences. You will collaborate closely with a cross-functional team of engineers, architects, product managers, UI/UX designers, and ML/data science partners to deliver robust, reliable solutions that power c

REMOTEpythonjavasql
View job →
D
1mo ago

We're on a mission to build the best platform in the world to defend the enterprise from code-to-cloud-to-runtime. Used by thousands of companies globally, Datadog security products uniquely leverage Datadog’s unified security and observability platform so Security, DevOps and SRE can collaborate rapidly and seamlessly to deliver better detection, prioritization and remediation. Our product and engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. In this competitive market, the Group Product Manager for Code Security will play a mission-critical role in providing product and strategy leadership to grow Datadog’s market share through differentiation, innovation and compelling customer value. This leader will lead a talented and growing team of product managers and work with world class engineers to build and grow multiple Code Security products that play an essential role for our customers’ code security programs, and growing Datadog into a security industry leader. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Run and grow multiple Code Security products to meet revenue and business targets with the goal of building a multi-hundred million dollar annual business. Lead and own product strategy and roadmap for accountable security products, fully aligned to revenue and business goals and with compelling differentiation and customer value. Ensure predictable roadmap execution across direct and partner teams to achieve product and business outcomes required to meet the revenue and business goals. Analyze and develop pricing and packaging strategies to maximize revenue through attaching deep understanding of market dynamics and other strategic leverage points. Drive GTM strategy with GTM partner teams

aigorust
View job →
D
Datadog
📍 Tel Aviv• Full-time
1mo ago

The eBPF APM team is building a zero-instrumentation observability solution that automatically discovers services on every host, supports both plaintext and TLS-encrypted traffic, classifies Layer 7 protocols, decodes service-level traffic, and reports RED (requests, errors, duration) metrics. Leveraging deep expertise in eBPF, the team operates across a wide range of Linux kernel versions, distributions, and complex customer environments. In addition to low-level networking, the team solves challenges related to protocol versioning, TLS detection across diverse languages and runtimes, and resilient performance in production systems We’re looking for a senior engineer with strong systems-level thinking and a good understanding of Linux. You should be comfortable working close to the kernel, ideally with experience in eBPF, or with a strong desire to dive into it. Proficiency in C/C++/ Go is essential, and familiarity with networking protocols, TLS internals, or distributed tracing is a strong advantage. You’ll join a high-impact team tackling ambitious technical challenges—like decoding traffic across multiple protocols, and ensuring high-fidelity metrics in complex, real-world environments. You’ll be expected to lead design and implementation efforts, contribute to roadmap planning, and collaborate across teams to ensure our solution remains robust, scalable, and frictionless for our users. This role is a great fit for engineers who thrive on low-level, performance-sensitive problems, and want to shape the future of observability through cutting-edge kernel technology. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Design and build core components of our zero-instrumentation APM product using eBPF and Go Develop systems to aut

linuxrestai
View job →

We’re looking for an Engineering Manager to lead our Sensitive Data Scanner (SDS) Telemetry team. The SDS group’s mission is to be the world’s easiest-to-use tool to discover, classify, manage, and report sensitive data risks across cloud, on-premise, and code environments. This team builds and scales the detection capabilities that scan all telemetry data flowing into Datadog — logs, APM spans, and RUM events — operating in streaming, at processing time, and at very large scale. You’ll lead a small, close-knit team based in Paris, with the opportunity to shape how the team grows as SDS Telemetry’s scope expands. It’s a chance to combine hands-on technical leadership with direct customer and product impact in the security and observability space. At Datadog, we place value in our office culture — the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team of engineers building real-time sensitive data detection across Datadog’s Logs, APM, and RUM telemetry pipelines Partner closely with the Logs, APM, and RUM teams, plus Datadog’s Trust & Safety team, to align on roadmap and integration priorities Shape product direction by working closely with Product, grounding decisions in customer needs and business impact Stay hands-on: contribute to design decisions and participate in the team’s on-call rotation Recruit, mentor, and develop engineers as the team grows beyond its initial size Help build a strong engineering culture as part of Datadog’s broader Sensitive Data Scanner group Who You Are: You have experience building and shipping revenue-generating products, with strong product acumen and a customer-first mindset You have hands-on experience with Go and/or Java, and a track record building distributed, streaming systems at scale You have experience managing engineers — or are

javaaigo
View job →

Datadog’s Product Analytics suite spans Product Analytics, Session Replay, Feature Flags, and Experimentation. Together they give product teams a complete, quantitative and qualitative picture of how users experience their applications, plus the tools to release, measure, and improve those experiences with confidence. Our team of Applied Scientists makes this space smarter and more autonomous, researching, prototyping, and industrializing AI/ML capabilities that make the suite proactive and agentic by default: AI-driven event understanding and instrumentation, conversational analytics, automated insight reporting, and UI/UX issue detection from session replays. The ambition is to let product teams act with autonomy, moving from question to insight to safely released change without depending on engineering or analysts. As the Engineering Manager for this team, you’ll lead a team of Applied Scientists through an early-stage, high-impact opportunity: defining the team’s technical vision, growing its footprint, and shaping how AI capabilities get built and delivered across the suite. This is greenfield work, both in the R&D itself and in the team you’ll grow around it. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team of Applied Scientists researching, prototyping, and industrializing AI/ML capabilities across the suite Define the technical vision and roadmap for the suite’s agentic and proactive AI/ML capabilities: event understanding, conversational analytics, insight reporting, and session-replay issue detection Stay hands-on as a technical contributor and reviewer, helping the team move from research prototypes to production-grade capabilities Shape how AI capabilities get built and delivered across the suite, partnering closely wi

aigorust
View job →

As a Research Scientist on our team, you will partner with Research Engineers, working on fundamental research problems and collaborating with Datadog's product and engineering teams to translate research advances into products. Building on our track record of AI-powered solutions (e.g., Bits AI , Bits Evolve , and our time series foundation model ), Datadog AI Research tackles high-risk, high-reward problems grounded in real-world challenges in cloud observability and security. We are focused on two research areas: World Models for Observability -- Training multimodal foundation models that learn the joint dynamics of distributed systems across metrics, traces, logs, topology, and events. These models power advanced forecasting, anomaly detection, root cause analysis, counterfactual simulation ("what if?"), and provide a learned planning backbone for our autonomous agents. Trained Agents for Observability -- Post-training models to operate autonomously across Datadog's domain. SRE incident response is our first target, with a clear path to code repair, security response, and infrastructure optimization. We build the simulation environments, RL training loops, and evaluation infrastructure needed to train agents that match or surpass frontier models at a fraction of the cost. What You'll Do: Conduct research in generative AI and machine learning, building specialized foundation models and trained agents for observability Train multimodal models on large-scale, diverse telemetry data (metrics, logs, traces, topology, events) using distributed training infrastructure Design and build simulated environments and RL training loops for on-policy agent training and evaluation Collaborate with cross-functional teams (Product, Engineering) to integrate capabilities like multimodal world modeling and autonomous agents into Datadog's products Stay at the forefront of foundation models, world models, and RL-based agent research Contribute to r

gitmachine learningai
View job →
🔔

Get new detection and mitigation engineer jobs by email

Daily job updates · Unsubscribe anytime