Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Job Summary We are seeking a motivated engineer to join the DRAM Systems Engineering team, focusing on the development, evaluation, and optimization of next-generation memory systems for AI accelerators. This role emphasizes research and development across hardware architecture, operating systems, and performance analysis to support Agentic AI inference workloads. If you are ambitious and eager to make an impact in the exciting world of AI and memory systems, this is the perfect opportunity for you! Responsibilities Characterize AI inference workloads and examine memory behavior Build and evaluate tiered memory hierarchies for AI accelerators Study KV cache lifecycles, MoE models, and data placement strategies Compare and optimize explicit versus hardware-assisted data movement Develop, test, debug, and detail system-level and OS components Prototype and evaluate agentic AI systems by building agents and multi-agent workflows using modern frameworks and orchestration patterns (planning, tool use, memory, and context management). Apply these technologies both as workloads under study and as accelerators for internal engineering workflows <h2 style="color:!importan
Jobiba hiring network
System Engineer Jobs
10,000 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current system engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Leidos has an exciting opportunity for a Sr. Software Engineer in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary As a Software Engineer on this program, you will have the opportunity to build strong systems, software, and cloud environments while providing operations and maintenance for critical systems. This role will provide technical expertise in the design, development, implementation and testing of customer tools and applications. Based in a DevOps framework, this role participates in and/or directs major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Participates in and/or directs software programming initiatives using Java, JavaScript, Python, SpringBoot, and Hibernate. Develops software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within Commercial Cloud Solutions leveraging infrastructure platform services Coordinate closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases Support th
We’re here for one reason and one reason only – to cure cancer. Every moment is dedicated to developing treatments and every action moves us one step closer to our goal. We’ve made incredible scientific breakthroughs and our pioneering personalized CAR T-cell therapies have changed the paradigm. But we're not finished yet. Join Kite, as we make even bigger advances in cancer therapies, and help shape where our business and medical science goes next. We believe every employee deserves a great leader. People Leaders are the cornerstone to the employee experience at Gilead and Kite. As a people leader now or in the future, you are the key driver in evolving our culture and creating an environment where every employee feels included, developed and empowered to fulfil their aspirations. Join Kite and help create more tomorrows. Job Description Position Summary The Senior IT Quality Engineering Specialist serves as the technical lead for the Kite Laboratory Information Management System (KLIMS) within North America West Coast (El Segundo). This role is responsible for the operational stability, compliance, maintenance, enhancement, and technical governance of the LabVantage platform and associated integrations. The position partners closely with Quality Control, Manufacturing, Validation, Infrastructure, and Kite Business stakeholders to ensure reliable and compliant laboratory operations. Key Responsibilities System Administration & Operational Support • Provide day-to-day administration and technical support for KLIMS and related
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate
NVIDIA's invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning - the next era of computing - with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as "the AI computing company." We're looking to grow our company and establish teams with the most thoughtful people in the world. We are looking for an excellent Senior Engineering Manager to lead a large firmware engineering organization delivering end-to-end manageability firmware for NVIDIA's next generation Data Center Compute Systems. This role owns HGX product line and OpenBMC-based management firmware and MCU firmware components in data center platforms, including architecture, execution, quality, reliability, telemetry, and customer readiness. We are seeking an experienced senior leader with strong technical depth, broad system perspective, and a proven ability to lead large teams through complex product cycles. This role is onsite in Santa Clara, CA, USA. If you're creative and autonomous, we want to hear from you! What you'll be doing: Lead a large firmware engineering organization delivering OpenBMC based firmware and MCU firmware for next-generation Data Center Compute Systems. Own HGX platform as a lead for Firmware and System software readiness working across the organization. Define and drive the long-term firmware roadmap, balancing architectural innovation with product execution and delivery milestones. Drive architecture strategy across BMC, MCU, platform software, manageability, health management, and data center firmware interfaces. <spa
Applied AI is where Datadog's ambitious AI bets get built and shipped ( Bits Chat , updog ). We sit at the intersection of research and product: turning promising capabilities from Datadog AI Research lab and the research community into production systems that reach real customers. The team builds the foundations for agentic systems capable of operating at scale in complex production environments. Current bets span agents that run autonomously at scale, context and memory layers that make those agents more intelligent over time, and tools that help customers build and validate AI-native services in production. The mandate is to move fast from idea to customer impact, and when a product finds its footing, to set it up for growth. As an Engineering Manager I in Applied AI, you will lead a team of engineers and applied scientists working on one of these challenges. You will define technical direction, run short feedback loops, make deliberate decisions about what to pursue or stop, and work closely with product managers, research teams, and cross-functional partners to ship AI capabilities that matter. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do Lead and develop a team of engineers and applied scientists focused on building the foundations for agents operating at scale Work closely with product managers, research teams, and cross-functional partners to shape the team's bets from initial framing through to broader adoption, with a clear definition of success criteria at each stage Own end-to-end delivery of high-quality AI systems, from early research exploration to production-grade reliability, with high standards for operational excellence, system reliability, and technical quality Navigate the unique challenges of shipping AI-powered products: balancing quali
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking an experienced AI Systems Reliability Engineer to help define, build, and maintain the infrastructure and processes that ensure the reliability, scalability, and performance of Postman’s AI-powered API and agentic systems in production. This role focuses on monitoring, availability, incident response, and automation to support AI services and tools trusted by millions of developers globally. What You’ll Do Develop and manage reliability metrics (SLOs) for AI-driven API services and agentic AI platform features Implement comprehensive observability and monitoring systems for real-time performance and fault detection Design and drive automated failover, recovery, and incident response strategies for high-availability AI infrastructure Optimize resource utilization, particularly GPU/accelerator efficiency, ensuring cost-effective AI system operation Collaborate closely with engineering, platform, and product teams to align reliability efforts with broader organizational goals Lead efforts to build internal tooling and automation focused on AI system stability and operational excellence Drive continuo
About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight : Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review ). About the Role This role focuses on oversight and system-level mitigations that enable increasingly capable agents to operate safely and autonomously in real environments. We prioritize building oversight systems that are used in practice today, both internally and externally (see our recent work on action monitoring for codex and former code review ). We also study longer-term questions about how increasingly capable agentis systems can be supervised, constrained, and corrected. We’re looking for a safety&security minded researcher or engineer who can reason rigorously about security boundaries and agent behavior, then build and test practical mitigations. A background in AI control or security is welcome but not required. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and evaluate system-level controls for agent actions like agent-based review. Plan how they fit in a broader syste
Experienced Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer Company: The Boeing Company Boeing Defense, Space and Security (BDS) is looking for an Experienced Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer (Level 3) to join our team in Huntington Beach or El Segundo, California. This position will be joining a team of engineers, analysts and staff within the Space and Missile Systems (SMS) and Boeing Technology and Innovation (BTI) Mission Systems organizations in the development, assembly, integration, and test of a constellation of satellites. The SMS and BTI organizations develop and capture technology for designing disruptive Mission Systems solutions. Focused on visible and infrared spectrum Electro-Optics Infrared (EO/IR) sensors operating in Space and Air domains, our growing team is leaping ahead of our competition with an exceptional mix of mission architectures, sensor designs, and algorithms for advanced image and data processing. Position Responsibilities: Develops and validates requirements for various communication, sensor, electronic warfare and other electromagnetic systems and components Develops and validates electromagnetic requirements for electrical\electronic systems, mechanical systems, interconnects and structures Develops architectures to integrate systems and components into higher level systems and platforms Performs trade studies, modeling, simulation and other forms of analysis to predict component, interconnects and system performance and to optimize design around established requirements Defines and conducts tests to validate performance of designs to requirements Manages appropriate aspe
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. As a Datacenter Liquid Cooling Architect, you will define, design, and architect next-generation liquid cooling infrastructure for Tenstorrent’s large-scale AI training and inference clusters. You will partner with systems engineering, mechanical engineering, software, and cross-functional design teams to develop chassis-, rack-, and cluster-scale cooling solutions, including CDU integration, telemetry and control, leak detection, and resilient operating strategies. This role will help shape reliable AI datacenter architectures and deployments for both internal and external customers. This role is on-site, based out of Toronto, Canada, Austin, Texas or Santa Clara, California. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A datacenter and system thermal design professional with 10+ years of experience architecting cooling infrastructure for complex computing environments. An experienced liquid cooling architect who can design chassis- and rack-scale solutions for large AI training and inference clusters. A systems thinker who understands how mechanical, electrical, software, facility, and systems engineering decisions come toge
Gong harnesses the power of AI to transform how revenue teams win. The Gong Revenue AI Operating System unifies data, insights, and workflows into a single, trusted system that observes, guides, and acts alongside the world’s most successful revenue teams. Powered by the Gong Revenue Graph, AI-powered intelligence, specialized agents, and trusted applications, Gong helps more than 5,000 companies around the world deeply understand their teams and customers, automate critical sales workflows, and close more deals with less effort. For more information, visit www.gong.io. At Gong, you will join a company built on innovative products, ambitious goals, and passionate people. We are shaping the future of revenue intelligence and we want people who are excited to build what comes next. You will work with a team that dreams big, moves fast, and cares deeply about the craft and about each other. Here, transparency and trust are core to how we operate, and every person has the opportunity to make a visible impact. If you want to grow, stretch, and do work that truly matters, Gong is the place to do the best work of your career. Gong is looking for an experienced Director of Quality Engineering to join our R&D team and define how quality is managed, measured, and improved across Gong. In this role, you will strengthen Gong’s engineering-owned quality model by creating the methodology, tooling, standards, and operating rhythm that help engineering teams own quality effectively. You will work closely with engineering leaders to make quality risks visible, improve release confidence, and give teams the practical capabilities they need to ship quickly and safely at scale. As Director of Quality Engineering at Gong, you will: Define and manage Gong’s quality operating model across R&D, including ownership expectations, release-readiness standards, metrics, and review routines. Build the Quality Engineering methodology and tooling that help engineering teams improve automat
MeltPlan | Planning Engine for the Built Environment MeltPlan is building the “planning engine” for the $14 Tn construction industry, an AI system designed specifically to optimize decisions before construction begins. While design software optimizes use and aesthetics and construction software optimizes execution and control, MeltPlan is building the missing layer - software that optimizes decisions and tradeoffs upstream, before scope is locked, procurement begins, and change orders become inevitable. MeltPlan’s long-term goal is to help teams make construction “boring” by making planning more intense: surfacing constraints and tradeoffs early, aligning stakeholders before plans are frozen, and reducing the need for late-stage redlines, rework, and change orders. MeltPlan is founded by operators who have built at scale. Kanav previously co-founded Innovaccer, a $3Bn healthtech company focused on making US healthcare more affordable and accessible. He’s now applying that systems-level thinking to construction.He’s joined by Tanmaya Kala, former Project Executive at DPR Construction, who led large commercial, healthcare, and life sciences projects. We combine deep tech scale with real construction execution. What This Role Really Is We are looking for an AI Research Scientist – Computer Vision to enhance and manage the PlanGraph model, which transforms 2D drawings into structured graphical representations of building elements. The role involves solving downstream business use cases such as quantity takeoff, code compliance, value engineering, and constructability analysis.We are specifically looking for hands-on researchers with experience in solving real-world Computer Vision problems and building custom vision models, VLMs, or VLLMs for production-grade applications. What You’ll Do Build and optimize custom Computer Vision models, VLMs, and VLLMs for construction intelligence workflows. Solve downstream business use cases including quantity takeoff, code complianc
Opportunity Overview: We’re looking for a Manager, Platform Engineering that can lead and grow a high-performing engineering team focused on Developer Experience, DevOps, SRE, and Quality. You will own the systems and processes that enable teams to build, test, release, and operate software with high velocity and reliability, driving engineering efficiency and operational excellence across the organization. What you’ll do: Lead a fast-paced, autonomous team of engineers focused on platform engineering, developer experience, DevOps, SRE, and quality engineering Own and drive the internal developer platform strategy and roadmap, improving how engineering teams build, test, deploy, and operate services Create transparency into engineering efficiency and system health through meaningful metrics across delivery, reliability, and quality Enable teams to move faster by improving CI CD pipelines, environments, tooling, and overall developer workflows Provide technical leadership across platform, infrastructure, and reliability, helping teams build scalable and resilient systems Ensure strong engineering practices across release processes, testing, quality, reliability, and security Define and enforce release guardrails, validation standards, and rollback mechanisms to improve production safety Improve environment stability and consistency across development, QA, and pre production environments Drive test strategy and automation maturity to improve overall product quality and confidence in releases Define and implement observability, monitoring, and alerting standards across systems Improve incident detection, response, and RCA practices, ensuring learnings translate into platform and system improvements Drive cloud infrastructure best practices across AWS, containers, and infrastructure as code Foster a culture of ownership, reliability, and continuous improvement within the team Provide innovative solutions for attracting, developing, and retaining top engineering talent I
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are hiring a Business Systems Lead to own the sales technology behind Lyft Ads. As the senior technical owner of the Lyft Ads CRM and commercial systems, you will lead solution design across Salesforce and the connected ad-sales stack and own the architecture for how ad deals move from opportunity through order management and billing. You will work alongside a team of Salesforce engineers and partner closely with Ads Sales, Sales Operations, Ad Operations, Finance, Legal, and the Media engineering teams. This is a hands-on leadership role for someone who knows ad-sales operations and can build the systems that run them. Responsibilities: Own the Salesforce architecture for the ad-sales lifecycle, from lead sourcing and opportunity creation through Order Management System (OMS) integration, contract execution, delivery reconciliation, and billing. Identify problems and opportunities across the Ads sales technology stack and shape them into scoped, funded initiatives. Lead requirements with Ads Sales, Sales Operations, Ad Operations, Finance, and Legal partners, and translate them into functional and technical designs. Own the integrations that connect Salesforce to the ad-sales ecosystem, including order management and contract systems, the ad server, and billing, through middleware such as Workato or Boomi. Drive AI best practices across the Sales Technology team, modeling effective usage and ensuring all AI-generated output meets Lyft's standards for quality, security, and performance. Guide the evolution of the Ads CRM platform, including the media-specific data model (agency, advertiser, and brand hierarchies, insertion orders, and flights) and own the Ad Sales Technology roadmap . Lead contract lifecycle management for Ads, including intake automation and Salesforce integration. Remain h
Get new system engineer jobs by email
Daily job updates · Unsubscribe anytime