Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. We are looking for a Security Operations Lead (SOC Lead) to build, mature, and operate our 24/7 detection and response capabilities across a modern cloud-native and AI-driven environment. This role leads the global SOC function—monitoring, SIEM ownership, detection engineering, alert triage, and operational readiness—while also evaluating and integrating emerging AI-based SOC products and autonomous response platforms . You will oversee monitoring across multi-cloud environments (GCP primary, AWS/Azure secondary), Kubernetes, SaaS services, endpoints, developer tools, and AI workloads . You’ll collaborate closely with Cloud Security, Compliance/GRC, SRE, Platform Engineering, IT/Endpoint teams, and AI Infrastructure to ensure our detection strategy scales and stays ahead of evolving threats. This is a hands-on leadership role perfect for someone who wants to shape the SOC of the future while solving complex challenges in a high-scale AI setting. What You’ll Do SOC Leadership & 24/7 Monitoring Lead, mentor, and scale a global SOC team responsible for 24/7 monitoring, alert intake, triage, correlation, and escalation. Build operational rigor: processes, runbooks, SLAs, metrics, and quality standards for high-scale environments. Cover monitoring across: Cloud infrastructure (GCP, AWS, Azure) Kubernetes/GKE/EKS/AKS clusters SaaS platforms (Google Workspace, GitHub, Slack, Okta, etc.) Endpoints (macOS, Linux, Windows) including EDR/XDR telemetry Developer platforms + CI/CD pipelines AI/ML systems and model-serving workflows AI-Based SOC Integration & Innovation Evaluate, adopt, and integrate AI-native SOC technologies for triaging, detection, and correlation Identify opportunities to automate triage, investigations,
Jobs in United States
It Systems Engineer in United States
3,028 active opportunities · Updated October 2026
Showing
15 jobs
Explore current it systems engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. You will build the layer that makes Vanta's data actually useful: a single interpretive layer that reads across every source of truth and makes them queryable, reasoned-over, and genuinely illuminating — for EPD, for GTM, and for anyone in the org trying to understand what's happening and why. The EPD Systems team is building the infrastructure Vanta's engineering, product, and design organization depends on to understand itself. We're constructing three interlocking layers: sources of truth at the foundation, an operating system that makes delivery legible, and the intelligence layer that reasons across all of it. This role owns the intelligence layer — the one that doesn't exist yet. This is a builder role. You will ship working things yourself — prototypes, internal tools, agent workflows. You will not hand specs to someone else and wait. Communication isn't a separate deliverable; it's how you learn what to build. What you’ll do as a Senior Product Builder at Vanta: Build the intelligence layer: a cross-source interpretive layer that reads across Vanta's sources of truth and makes them queryable and reasoned-over by EPD leadership, GTM, and beyond Define what to build: scope the problem yourself, make explicit tradeoffs about what to defer, and own the sequence of what gets built and when Ship working things yourself: prototypes, internal tools, agent workflows — using AI as part of how you work, not what you report on Understand the organization: go to the teams whose decisions this layer will serve — GTM, G&A, EPD — and come back knowing what they can't answer today Catch and address AI-specific quality problems — rel
From $204K/yr
The opportunity Datadog’s Infrastructure products help engineers understand and operate the systems their applications depend on. Our customers work in complex environments like Kubernetes and serverless, where infrastructure changes constantly, information is dense, and decisions about reliability, performance, and cost are closely connected. We’re looking for a Staff Product Designer to join Modern Compute, with an initial focus on Containers Autoscaling. Autoscaling helps engineering teams make better decisions about how their applications and infrastructure use resources. Designing these experiences requires making deeply technical systems understandable, helping customers act with confidence, and fitting into the tools and workflows they already use. The team is rethinking how workload and cluster autoscaling come together as a more coherent product experience. This includes how customers get started, understand recommendations, evaluate value, and safely apply changes across their environments. The work also connects to other parts of Datadog, including observability, Cloud Cost Management, permissions, and AI-assisted workflows. As a Staff Product Designer, you will help define that direction and lead the work from early problem framing through shipped product. You will partner closely with product and engineering, bring a high level of interaction and visual craft to complex workflows, and help raise the quality of design across Modern Compute. At Datadog, we place value in our office culture, the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to help our Datadogs find a work-life rhythm that works for them. What you’ll do Lead end-to-end product design for Modern Compute, initially focused on our Autoscaling product. Help define the product direction for an area that is still evolving, from early framing and exploration through detailed design and delivery. Design clear, trustwort
From $244K/yr
We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—allowing for seamless collaboration and problem-solving among Dev, Ops and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Team: As organizations rapidly invest in AI applications and build out AI labs, telemetry volumes are growing exponentially and costs are becoming unpredictable. From LLM interactions to agentic workflows, AI systems generate unpredictable streams of logs, driving up costs and making it harder to maintain efficient observability. These challenges are critical for organizations in regulated industries with strict data residency requirements, where data must remain within controlled environments. Datadog’s Bring Your Own Cloud (BYOC) team is reimagining what observability and security look like at petabyte scale in the AI era. The Opportunity: The Group Product Manager - Bring Your Own Cloud (BYOC) role is responsible for defining and bringing to market the next generation of telemetry analytics and insights capabilities in an AI-first environment. This role is highly technical and creative in nature as you will envision novel ways to enable customers to cost-effectively explore, analyze and report over petabytes of data through a welcoming and easy-to-use interface. You will partner with various teams to take advantage of BitsAI capabilities and surface critical insights on volume usage and retention for popular use cases. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team
NVIDIA is seeking a Senior System Architect: Heterogeneous EDA Systems to solve a complex challenge in accelerated computing: Failure Attribution at Scale. As EDA or equivalent experience workloads scale across thousands of heterogeneous nodes, a single failure can cause massive resource waste. We need an engineer to develop and build an automated framework. This framework will ingest telemetry from CPU and GPU clusters to identify the root cause of job failures in real-time. It will distinguish between hardware faults, infrastructure instability, and software defects. What you'll be doing: Architect Failure Attribution Frameworks: Build a scalable "flight recorder" for EDA jobs that captures high-fidelity state across the CPU, GPU, and Fabric at the moment of failure. Build automated diagnostics that correlate GPU XID errors, PCIe bus failures, and CUDA memory exceptions. Connect these errors with system-level events such as OOM kills or NUMA-related hangs. Distributed Logging & Tracing: Implement low-overhead tracing mechanisms (using tracing tools or custom agents) that provide access to job execution across multi-node Slurm or Kubernetes clusters. Root Cause Automation: Develop heuristics and models based on machine learning to classify failures as "Hardware Fault," "Software Bug," or "Environment Issue." This reduces the Mean Time to Identify (MTTI) for R&D teams. Resiliency Engineering: Work closely with hardware and infrastructure teams to define "signals of impending failure," enabling proactive job migration or check-pointing before a crash occurs. What we need to see: Distributed Systems Mastery: BS, MS, or PhD in Computer Science or Electrical Engineering (or equivalent experience) with 6+ years in systems programming. Experience building automated
About the Team The Support team is central to ensuring that our customers' experience with our products is nothing short of exceptional. We resolve complex issues, provide technical guidance, and support customers in maximizing value and adoption from deploying our products. We work closely with Sales, Technical Success, Product, Engineering and others to deliver the best possible experience to our customers at scale. OpenAI's customers represent a range of diverse backgrounds and maturity, from individual customers to early-stage startups and established global enterprises. Given OpenAI’s breakneck shipping cadence and growth – and the expectation that it will only accelerate – our ability to architect automation systems and agentic workflows for scale is central to our ability to maintain exceptional support quality in the face of AGI. About the Role We are seeking a Support Operations Lead who combines operational leadership, systems thinking, vendor management, and hands-on execution. You’ll own service health, automation programs, partner and vendor management. In addition to delivering high-quality service, you’ll identify opportunities to reduce manual work, experiment with tools and help operationalize AI across support at scale. This is not a traditional support operations lead role. We’re looking for someone to help us define the future of support, who thrive at the intersection of team/project management, systems building, data science/engineering, and with deep craft experience in the support operations space. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead and evolve organizational design for our frontline operations across partner and vendor management, coaching teams to expand automation and deliver measurable capacity gains. Lead multi-site partner management, including commercial ownership, capacity planning and workfor
About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role The Safety Measurement Product Manager owns OpenAI's approach to measuring harm and safeguard efficacy in production, including driving the strategy for our suite of safety measurement platforms and products used across the company. You will partner closely with our safety research and engineering teams to determine what we measure, where we measure it, and how we measure it, feeding those insights directly into critical leadership decisions and back into our safety work. You will also represent the company's topline safety metric as well as prioritize incoming requests from partner teams to expand our safety measurement platform to more use cases. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with data science, research, engineering, policy teams, and other stakeholders to craft a vision for understanding safety outcomes and prevalence on our platforms. Define strategic priorities and product roadmaps focused on improving safety measurement approaches will scaling our measurement platform to more use cases, products, and cross-functional team needs. Establish repeatable processes to integrate cutting-edge AI safety research into OpenAI’s safety m
About the Team OpenAI is building AI systems that can help professionals perform complex, high-value work with greater speed, rigor, and creativity. Investment banking is one of the most demanding environments for knowledge work: bankers must synthesize fragmented information, exercise judgment under pressure, and produce precise, defensible models, analyses, and client materials. Our team works across Research, Product, Engineering, and Go-to-Market to make OpenAI's models genuinely useful for these workflows. We translate real professional work into product requirements, evaluations, training signals, and repeatable customer solutions. We care not only whether a model can generate an answer, but whether it can deliver accurate, defensible work that experienced bankers can trust and use. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. About the Role We are looking for a Subject Matter Expert in Investment Banking to help define what excellent AI-assisted banking work looks like and turn that standard into better models and products. You will bring deep, current knowledge of how investment banking work is actually performed, including company and industry research, financial analysis and modeling, valuation, diligence, transaction execution, and the creation and review of client materials. You will use that expertise to design realistic tasks and evaluations, create and assess high-quality reference work, diagnose model failures, and help our technical teams improve model behavior and product experiences. This is a hands-on individual-contributor role for someone who enjoys both doing the work and explaining what makes it good. You should be comfortable moving between an Excel model, a presentation, a source document, an evaluation rubric, a product prototype, and a conversation with researchers or customers. You will help us distinguish outputs that merely look pl
About the Team OpenAI builds powerful AI systems like ChatGPT, the OpenAI API, and enterprise products that serve millions of users across the globe. As we scale, securing our infrastructure, protecting sensitive data, and meeting global compliance standards are essential to our success and societal impact. Security at OpenAI is a cross-cutting function that spans infrastructure, applied engineering, legal, policy, and product. Technical Program Managers (TPMs) play a critical leadership role in aligning teams and delivering execution at scale and this role will be foundational in shaping how we secure OpenAI’s systems, users, and commitments. About the Role We’re seeking a Senior Technical Program Manager to drive cross-functional security, privacy, IT and compliance initiatives at the intersection of infrastructure, product, and policy. You will execute complex programs that reduce internal data access, prevent misuse, and ship security capabilities. You will focus your efforts on the most critical initiatives within Security, crossing the spectrum of insider threat, information security, physical security, and information technology challenges. This role is deeply technical and execution-focused. It requires a structured operator who thrives in ambiguity, partners effectively across boundaries, and applies principled judgment to scale trust, governance, and security across OpenAI’s systems and products. In this role, you will: Drive execution of critical security and compliance programs such as vulnerability management, merger and acquisition security and integration, infrastructure hardening, and datacenter security management. You will need to deeply collaborate on technical architecture and resolve technical problems in partnership with engineering. Partner with IT, Infrastructure, Application, Legal, Privacy, and Security teams to build scalable programs, and deliver critical security outcomes across multiple disciplines, including insider threat, information
From $156K/yr
As organizations rapidly adopt AI applications and agentic systems, security teams need visibility and control over how these technologies are being used. Datadog's AI & Data Security product helps customers discover, secure, and govern AI usage across their environments ensuring sensitive data is properly managed from model training through production. As a Product Manager II for AI & Data Security, you will own capabilities for AI discovery, posture management, and data security; giving customers complete visibility into their AI applications, agents, and enabling ecosystem, with prioritized actions to operate AI systems securely. You'll partner with engineering, design, security research, and GTM teams to define and ship platform capabilities that help organizations adopt AI at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the roadmap for AI & Data Security capabilities, including AI and data discovery & posture management. Define how security teams can assess and manage the security posture of AI-enabled systems, including configuration risks, sensitive data exposure, and policy violations. Work closely with engineers and designers to deliver new product capabilities end-to-end, from early concept through launch and iteration. Partner with Datadog security researchers to identify emerging risks in AI systems and translate them into actionable product features. Engage with customers to understand how they are adopting AI and validate solutions that help them operate these systems securely. Collaborate with go-to-market teams to enable adoption and communicate the value of AI security capabilities to customers. Who You Are: You have 3+ years of product management experience building technical products, ideally in security,
At Freddie Mac, our mission of Making Home Possible is what motivates us, and it’s at the core of everything we do. Since our charter in 1970, we have made home possible for more than 90 million families across the country. Join an organization where your work contributes to a greater purpose. Position Overview: We need a highly innovative Technical Lead! How confident are you that you can build sophisticated analytic systems? If you believe you could contribute to the development of innovative principles and ideas in a matrixed environment, please keep reading as we are seeking an individual contributor who has experience with Java and Python and can lead and nurture an inspiring environment in our Virginia office. Our Impact: The Investments and Capital Markets (I&CM) division is looking for a capable technology lead for its trading and analytics development team. This could be you! To thrive in this division, you must have a comprehensive understanding of system implementation and design, experience working in capital markets, and be enthusiastic about leading development of new paradigms in software system architecture. Your Impact: As a Trading Analytics Development Tech Lead, you will develop and maintain software using Java and Python tech stack that adheres to software engineering best practices. You will influence technical decisions, mentor developers, resolve engineering blockers, and partner with engineering managers to help teams deliver secure, reliable, and maintainable solutions. You will provide hands-on directions for full-stack applications, APIs, microservices, and integration services while reinforcing engineering discipline across design, development, testing, deployment, observability, and production readiness. Partner closely with Product Owners, engineering managers, architecture, business stakeholders, and cross-functional tec
About the job Lead the team that proves our AI silicon performs reliably before it reaches customers. You will build and lead Graphcore's characterisation capability for next generation silicon and system platforms. Your work will help ensure our products perform consistently across real world conditions. You will define bring up and characterisation strategies, lead technical execution, and shape the lab infrastructure needed for success. You'll work across silicon, hardware, manufacturing, architecture and product teams to solve complex engineering challenges. This is a hands on leadership role with the opportunity to influence both product design and how Graphcore validates future AI systems. The team and culture This is a newly formed team within Manufacturing Operations. You'll have the opportunity to establish how the team works while building strong partnerships across engineering and operations. Day to day, you'll work closely with architecture, silicon, hardware, production test and product teams. Decisions are driven by data, technical evidence and close collaboration across disciplines. We value ownership and clear communication. You'll be trusted to lead technical direction, remove blockers and help teams make progress with confidence. What we're looking for Essential Proven track record of delivering complex technical projects as an individual contributor, manager, or project manager, with the ability to work independently and drive execution. Strong expertise in silicon digital device design, bring-up, characterisation, and silicon process technologies, with an understanding of their impact on transistor- and system-level performance. In-depth knowledge of high-performance processors, system-on-chip (SoC) architectures, and high-speed digital interfaces such as PCIe, Ethernet, and DDR. Experience with measurement automation, data analysis, and scripting/coding to develop automated test and analysis workflows, with familiarity of ATE systems and t
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an Intern Product Yield Enhancement (PYE) Engineer at Micron Technology Inc., you will assist in taking next generation DRAM devices from design to mass production! PYE Engineers are crucial partners between Research and Development, Manufacturing, Product Engineering and Global Quality teams. Our goal is to offer hands-on and substantial engineering projects that allow interns to gain valuable experience with memory systems and the product life cycle while adding impact to Micron’s business. Our intern programs help prepare engineers for future roles and offer an extraordinary experience that supplements your academic experience! The PYE Engineering role is critical to Micron as it serves the company's charter to be the most cost-effective producer of DRAM memory while delivering the quality and reliability that customers expect from the Micron brand. This position is for a 3-month internship. Responsibilities: Your job duties will include detailed DRAM electrical and physical failure analysis. Identify defects introduced by the semiconductor fabrication process. You will data mine using statistical analysis, AI assisted engineering tools, and automation workflows. Find the root cause of failure mechanisms, communicating results with Management, Product, Process and Test Engineering teams. Complete 3-month internship. Minimum Qualifications: Must be pursuing a minimum of a bachelor’s degree in Electrical Engineering, Computer Engineering,
From $193K/yr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity Every company shipping AI right now has the same problem: they can't see what their AI is actually doing in production. New Relic is building the answer, AI Observability, and it's one of the company's top strategic initiatives. We're running it like a startup inside the company: small teams, fast iteration, direct executive sponsorship, and a mandate to ship. As the Principal Product Manager and business owner of this initiative, you'll take it from vision through GA, then keep pushing past GA, where most of the real work actually happens, working directly with some of the largest AI deployments in the world. What you'll do Own the AI Observability roadmap and the business outcomes leadership tracks (adoption, retention, expansion), plus the parts of the job that aren't purely product: pricing conversations, exec reporting, and cross-functional escalations. Drive multiple engineering and design teams in parallel toward a coherent product vision, without formal authority over them. Partner directly with strategic enterprise customers, including their platform and infra teams deciding how to operationalize AI safely, and turn what breaks at their scale into product requirements fast enough to matter. Build working prototypes yourself using agentic AI workflows, so ideas get pressure-tested before engineering time is committed to them. Represent AI Observability in executive strategy and prioritization decisions, and make the tradeoff calls when prioritie
From $1.3M/yr
The Opportunity MongoDB’s partner ecosystem — systems integrators, ISVs, and the major cloud providers — is a strategic growth engine for the business. We’re looking for a Senior Manager, Sales Plays and Offerings to design the joint go-to-market motions that turn partner relationships into pipeline and revenue. This is a highly cross-functional, strategic role: you’ll build the sales plays and joint offerings themselves, partner with Enablement to get the field and our partners ready to sell them, instrument how they perform, and work with the Programs lead to make sure incentives reward the behavior we want to see. You will report directly to the VP of Partner Strategic Operations and act as a connective layer between Partnerships, Sales, Enablement, and Programs — translating ecosystem strategy into repeatable, measurable, field-ready motions. We are looking to speak to candidates who are based anywhere in the US for our hybrid working model. What You'll Do Build sales plays and joint offerings Design and package partner sales plays and joint solution offerings with priority ISVs, SIs, and cloud partners — defining the joint value proposition, target segment, competitive positioning, and playbook for how field and partner sellers execute it Partner with Product Marketing, Solutions Engineering, and partner counterparts to validate technical integration stories and translate them into a compelling, sellable narrative Prioritize which plays to build and scale based on market opportunity, partner readiness, and alignment to MongoDB’s strategic pillars (e.g., AI, migrations, industry verticals) Own the lifecycle of each play from concept through launch, iteration, and eventual retirement or refresh Drive partner and field enablement Partner closely with the Enablement team to translate each sales play into field- and partner-facing assets: pitch decks, battlecards, demo scripts, certification conte
Other cities to consider
More places hiring for this role
Get new it systems engineer jobs in United States by email
Daily job updates · Unsubscribe anytime