Jobiba hiring network

Controls Engineer Jobs

2,131 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current controls engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

PE
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role We are a software engineering team with expertise in enabling ML models in production. We deploy AI models to run in variety of environments: air-gapped government networks, forward-deployed defense environments, edge nodes, and enterprises with strict data sovereignty requirements. Our customers rely on us for frontier AI capabilities running on hardware they control, often with constrained GPU resources and limited direct access. Rising to that challenge and meeting those expectations is what Palantir's excels at. We treat models like any other software: continuously tested, continually delivered, packaged for reproducible deployment, and built for long-term maintainability. You will own services end-to-end, and work across the full stack, from inference engines, GPU scheduling to deployment pipelines, observability, and integration with Palantir's platform. The goal is to deliver new models and capabilities quickly and continuously. Join us if you want to solve problems at the intersection of infrastructure and machine learning that directly enable critical customers.

machine learningaigo
View job →
M
Mongodb
📍 Dublin• Full-time
1mo ago

We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. The Collections team is a newly formed team within MongoDB's Observability & Adoption Org focused on making telemetry onboarding and collection significantly easier across MongoDB. We own key parts of the observability collection stack, including onboarding experience, telemetry collection agents across the data and control planes, and ingestion services for metrics, logs, and traces that support both internal and customer observability in MongoDB, driving insights, recommendations, and alerting. Our mission is to reduce friction for teams implementing and iterating on Observability while partnering closely with development teams to instrument their services using shared best practices, helping define the conventions our telemetry should follow, and building collection and ingestion systems that are stable, performant, secure, well-documented, and self-service. We also work closely with the Data Pipeline and Storage & Query teams to help ensure MongoDB has a stable and performant observability stack end to end. This is an opportunity to join a team shaping how observability works across MongoDB and to have outsized impact on both the developer experience and the reliability of the platform underneath it. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Grafana, Sp

javamongodbaws
View job →
M
Mongodb
📍 Ireland• Full-time
1mo ago

We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. The Collections team is a newly formed team within MongoDB's Observability & Adoption Org focused on making telemetry onboarding and collection significantly easier across MongoDB. We own key parts of the observability collection stack, including onboarding experience, telemetry collection agents across the data and control planes, and ingestion services for metrics, logs, and traces that support both internal and customer observability in MongoDB, driving insights, recommendations, and alerting. Our mission is to reduce friction for teams implementing and iterating on Observability while partnering closely with development teams to instrument their services using shared best practices, helping define the conventions our telemetry should follow, and building collection and ingestion systems that are stable, performant, secure, well-documented, and self-service. We also work closely with the Data Pipeline and Storage & Query teams to help ensure MongoDB has a stable and performant observability stack end to end. This is an opportunity to join a team shaping how observability works across MongoDB and to have outsized impact on both the developer experience and the reliability of the platform underneath it. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Grafana, Sp

javamongodbaws
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are building and maintaining a highly scalable asynchronous platform that empowers our organization to handle critical business cases. As a software engineering team, our mission is to create robust and innovative solutions that drive the success of our business and deliver unparalleled value to our customers. We adopt Infrastructure as Code practice to automate the provisioning and configuration of our resources, which helps reduce manual configuration and improve consistency. Our team culture is built on collaboration, open communication, and a supportive environment where each member's ideas are valued and contributions are recognized. We believe in the importance of fostering a positive workplace culture that inspires innovation and creativity. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and production readiness reviews, platform management and capacity planning ceremonies with cross-functional teams Document Infrastructure operations process and insights, identify repeatable actions and ruthlessly automate repetitive tasks Participate in our teams on-call rotations, respond to incidents and support other teams mitigate customer impacting events Experience: 5+ years experience working on teams responsible for software development, automation and systems engineering Experience building large-scale infrastructure, distributed systems or networks. Knowledge with SQS,

pythonawsazure
View job →
S
Stripe
📍 US - Remote• Full-time• Remote
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The Abuse Research team is dedicated to proactively hunting for emerging abuse vectors and studying complex attacker behaviors. Rather than just reacting to alerts, the team maps complete abuse paths across Stripe products and external systems to validate novel findings and explain the underlying product conditions that enable fraud. By building continuous abuse tests with agentic testing and related systems, they translate their deep research into actionable threat advisories, strategic control recommendations, and regression scenarios that fortify Stripe’s defenses. What you’ll do You will lead the Abuse Research Group (ARG)—a team of threat intelligence analysts, fraud researchers, and detection engineers focused on proactively identifying and mitigating threats to Stripe and our merchants. You will set the research agenda, guiding work across threat actor tracking, hands-on fraud investigations, merchant ecosystem defense, and the detection pipelines that turn findings into production enforcement. You will scale the team through hiring, coaching, and talent development, while serving as a technical advisor on complex investigations. You will also own ARG's relationships with key partners—including Risk, Trust & Safety, Fraud Platform, and Security Engineering. Lead, develop, and retain a team of threat intelligence analysts, fraud researchers, and detection engineers who thrive at the intersection of adversarial research an

REMOTEreactairust
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI’s Application Engineering team builds the internal products and platforms that help OpenAI operate securely and at scale. We engineer, own, and evolve OpenAI’s core productivity ecosystem, creating secure applications, integrations, automation, and reusable tooling where off-the-shelf software is not enough. Our work spans employee-facing experiences and the services, APIs, control planes, and governance that make them reliable, permission-aware, and scalable. We also act as a customer zero for OpenAI’s technology, building the enterprise foundations that let employees and agents safely access the context, tools, and actions they need. We partner closely with IT, Security, product teams, and platform providers to turn company-wide problems into durable systems, learn from real internal workflows, and help shape the products we deploy. We create paved paths that let teams move quickly without compromising security or operational quality. About the Role As a Staff Software Engineer on Agent Productivity, you will shape the foundation that enables teams to build agents with secure access to the context and capabilities they need. Slack will be the first and deepest implementation surface—and where you spend most of your time—owning its application architecture, integrations, APIs, governance, and administration automation while building patterns that extend to internal systems, identity platforms, and other enterprise applications. This is a hands-on engineering role with broad organizational impact as agents support more employee workflows. You will define platform architecture, build reusable foundations, and establish secure patterns for identity, permissions, connectivity, and operations that make agents easier to develop, deploy, and manage. In this role, you will: Own the technical strategy and architecture that enable teams to build, connect, and deploy agents quickly and safely, using Slack as the primary implementation surface. Design and

awsgitrest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Storage Infrastructure team builds and operates the storage foundation behind OpenAI’s most demanding workloads. We work directly with research to design storage systems for rapidly evolving experiments, while also powering production at scale. We own the platform end to end: backend systems, user-facing services and APIs, and the control planes that manage how data is placed, moved, and retained over time. Our stack spans cloud and in-house object stores across very different workload profiles, from GPU-attached systems to dedicated storage hardware. We also build the federation layer that unifies these backends behind a simple interface and routes each workload to the right storage solution. About the Role You will help build the storage platform that powers OpenAI’s research and production systems. This is a hands-on infrastructure role for engineers who want to work on deeply technical systems at scale and own them in production. You’ll work across object storage, cross-region data movement, lifecycle management, and the federation layer that provides a unified interface across multiple backends. Much of our stack runs on Kubernetes, and we primarily build services in Rust. In this role, you will: Build and operate storage services that underpin OpenAI’s research infrastructure Develop object storage systems across cloud and in-house environments Build systems for cross-region data movement, replication, and recovery Design lifecycle management capabilities that keep data durable, available, and cost-effective Evolve the federation layer that unifies multiple backend systems behind a simple interface Improve performance, reliability, and operational excellence across the platform Collaborate closely with researchers and infrastructure teams to support rapidly evolving workloads You might thrive in this role if you: Have experience building or operating distributed systems in production Have worked on storage infrastructure, object stores, dist

awskubernetesrest
View job →
O
1mo ago

About the team Online Data builds and operates Habitat, the single product surface of Online Data and the system of record for OpenAI’s online user data. As OpenAI’s scale and product requirements evolve, Habitat is becoming a full-stack, one-size-fits-most database platform with end-to-end ownership of: Provisioning and developer experience APIs and guardrails Scaling, performance, and reliability Data movement, caching, routing, and placement Privacy enforcement and access control Change Data Capture (CDC) as a first-class primitive The foundation for future storage backends You’ll work on the core online database platform behind OpenAI’s products, building and operating Habitat services that handle high-QPS, latency-sensitive workloads across regions. You’ll partner closely with internal platform and product teams to ship safe, reliable systems, then push them to be faster and more cost-efficient through better caching, routing, observability, and operational tooling. This is a critical role for engineers who like owning hard distributed-systems problems end to end and sweating the details from p99 latency to production operations at massive scale. In this role, you will Design and build core abstractions spanning storage, caching, routing, CDC, and privacy enforcement Own a major surface area end to end, from product and API design to operational excellence Improve latency, correctness, and cost efficiency for real production workloads at massive scale Build strong instrumentation, debugging workflows, and developer-first tooling Collaborate closely with internal product and infrastructure teams to understand requirements and ship pragmatic solutions Participate in an on-call rotation and raise the bar on reliability while aggressively improving performance and usability You might thrive in this role if you have A strong track record building and operating high-scale backend or data-intensive distributed systems in production Excellent systems judgment and the a

pythonawsrest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Scaling team is responsible for the architectural and engineering backbone of OpenAI’s infrastructure. We design and deliver advanced systems that support the deployment and operation of cutting-edge AI models. Our work spans system software, networking, platform architecture, fleet-level monitoring, and performance optimization. About the Role We’re hiring an SW Engineer to enable production workloads and end-to-end testing on new platforms. This role will include creating new test harnesses and platform stress benchmarks, porting existing inference and training workloads to new, sometimes early-access, systems/hardware, analyzing performance and bottlenecks, and characterizing the end-to-end behavior of new systems (compute, comms, storage, control plane, and failure modes). Key Responsibilities Port and validate key inference and training workloads on new platforms/SKUs as they arrive; drive correctness, performance, and stability to an internal readiness bar. Build a suite of benchmarks and stress tests that capture real E2E behavior of our workloads by exercising all aspects of a system, including CPU, GPU, memory subsystem, frontend, scale-up, and scale-out networking (including WAN traffic, NVlink and RDMA collectives), storage, thermals, and any other relevant parts. Deep-dive performance on distributed training/inference: Collective performance and tuning (across NCCL/RCCL and internal libraries) Overlap of compute/communication, kernel-level bottlenecks, memory bandwidth and scheduling effects Create repeatable test harnesses that run in CI / lab environments and produce actionable outputs (pass/fail, performance score, regression detection). Partner with systems + fleet bring-up engineers to ensure the platform is not only stable and performant, but also operationally usable and scalable (containerization, K8s integration, telemetry hooks, failure triage loops). Work cross-functionally with vendors and internal stakeholders by producing

pythonawskubernetes
View job →

The Product Manufacturing Engineer will lead industrialization activities of new golf ball designs, partnering with our manufacturing sites in Taiwan, South Korea, and other Asia-Pacific countries from prototype to commercialization. This role troubleshoots production issues, drives cost/quality/speed/yield improvements, and reports test data and summarized results to management and key stakeholders. Essential Functions and Key Responsibilities: Maintain regular communication (e-mail, video conference, etc.) with global vendors to align specs, timelines, and project direction. Coordinate tooling, prototypes, and first article builds through evaluation and testing. Support manufacturing execution to ensure products meet performance, conformance, quality, and cost targets through statistical process control data and charts. Own manufacturing readiness timelines for new product launches, driving tooling, first-article, and qualification milestones to meet product launch dates. Generate test data reports and summaries for Engineering Management supporting supplier and R&D development activities. Work cross-functionally to develop, manage, and distribute accurate and controlled documents / specifications for new product designs and manufacturing processes. Process Engineering Changes and documentation updates through controlled channels. Evaluate and contribute to the design of test equipment, set-up/test procedures, and QC/QA inspection processes. Support Continuous Improvement efforts across production processes, qualifications, test methods, and test equipment. Drive supply chain Continuous Improvement projects designed to reduce cost and defective rates. Ensure component measurement

Excelsupply chain
View job →

Manufacturing Engineer (Associate or Experienced) Company: The Boeing Company Boeing Defense, Space & Security (BDS) Laser and Electro-Optical Systems (LEOS) site is seeking an Experienced Manufacturing Engineer for our advanced development and production manufacturing team located in Albuquerque, NM. The Boeing Laser and Electro-Optical systems (LEOS) group works to develop and implement next-generation technologies and products serving a variety of commercial and military customers. Our team focuses on rapid prototyping and precision production of advanced laser and electro-optical systems that satisfy challenging product performance requirements built in accordance with aerospace (AS9100) quality standards. Position Responsibilities: Design, develop and optimize manufacturing processes and tooling approaches for complex electro-optical aerospace parts and assemblies Collaborate with design engineers to ensure manufacturability and cost-effective production Review engineering drawings and provide comments to responsible engineer Participate in Integrated Product Teams (IPTs) to integrate technical solutions across multiple disciplines Participate in supplier selection and evaluation to ensure adherence to quality and delivery standards Write content for work instructions, travelers, and standard shop procedures Utilize electronic work instruction, material control and non-conformance management systems as needed to support production Ensure compliance with aerospace industry quality standards (e.g., AS9100, J-STD) Ensure the manufacturing work instructions and special processes follow safety and environmental regulations Assist in shop layout and

supply chainrecruitment
View job →

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full stack automation fabric solutions for mission-critical business processes. With the first SaaS-based composable automation platform specifically built for ERP, we believe in the transformative power of automation. Our unparalleled solutions empower you to orchestrate, manage and monitor your workflows across any application, service or server — in the cloud or on premises — with confidence and control. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for a Principal Engineer, Products & Platforms to join our Product engineering team to provide technical leadership across Redwood’s Workload Automation Platform, defining architecture, driving modernization, and influencing engineering strategy across multiple teams. You will be instrumental in the design, development, and enhancement of our platform, building high-quality, scalable, and secure software that powers enterprise data exchange for more than 1,000 customers worldwide. As a Principal Engineer, you will: Technical Leadership & Architecture: Define the path forward for complex engineering problems, establish best practices, lead design review and technology decisions, and mentor the team to excel, while driving the architecture, security, compliance, and observability of our Java/Spring Boot microservices. Platform and Infrastructure Ownership: Drive the deep understanding, architecture, and evolution of our core platform and infrastructure, focusing on resilience, communication between components, and scalability. Cross-Team Collaboration: Facilitate and drive cross-team collaboration with Product, QA, and other engineering groups to ensure end-to-end alignment and successful product deliv

javaawskubernetes
View job →
A
10 days ago

Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Meet Lingo, a new biosensing technology that provides users a window into their body. Lingo tracks key biomarkers such as glucose, ketones, and lactate to help people make better decisions about their health and nutrition. Biowearable technology will digitize, decentralize and democratize healthcare, enabling consumers to take control of their own health. At Abbott, you can do work that matters, grow, and learn, care for yourself and family, be your true self and live a full life. You’ll also have access to: Career development with an international company where you can grow the career you dream of. Employees can qualify for free medical coverage in our Health Investment Plan (HIP) PPO medical plan in the next calendar year An excellent retirement savings plan with high employer contribution Tuition reimbursement, the Freedom 2 Save student debt program and FreeU education benefit - an affordable and convenient path to getting a bachelor’s degree. A company recognized as a great place to work in dozens of countries around the world and named one of the most admired companies in the world by Fortune. A company that is recognized as one of the best big companies to work for as well as a best place to work fo

reactgraphqlai
View job →

Datadog's Software Engineers with Systems depth leverage their experience with systems and tooling to build software that ensures Datadog remains reliable, performant, and secure. For this track, their Software Engineering experience may resemble the Distributed Systems track, but is typically applied in combination with their systems experience to build and run internal platforms and tools that our products are built on. These people typically have deep experience building and managing large cloud infrastructure deployments, or leading reliability efforts for orgs similar to ours, or building release machinery to allow hundreds or thousands of devs to do their jobs without stepping on each others' toes. The systems and tooling where they may have experience depth may include (but not limited to): bazel, build tooling, cassandra, CDN, chef, configuration management, container orchestration, consul, docker, elasticsearch envoy, haproxy, kafka, kubernetes, load balancing, network architecture, postgres, redis, release management, RPC frameworks, service discovery, spinnaker, terraform, zookeeper. Bonus: You’re excited about leveraging AI tools to enhance how you code, solve problems, and build – or eager to learn how This job is available in various departments within our company; to conform to US export control regulations, some of these roles may require candidates to be eligible for any required authorizations from the US government. #LI-KM5 Datadog offers a competitive salary and equity package, and may include variable compensation. Actual compensation is based on factors such as the candidate's skills, qualifications, and experience. In addition, Datadog offers a wide range of best in class, comprehensive and inclusive employee benefits for this role including healthcare, dental, parental planning, and mental health benefits, a 401(k) plan and match, paid time off, fitness reimbursements, and a discounted employee stock purchase plan. Th

postgresqlredisdocker
View job →

About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors, software, and data center systems that provide the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing transformative technologies. Our U.S. engineering teams contribute to the hardware and software platforms that support the next generation of AI systems. The Opportunity We are looking for a recent graduate or early-career engineer to join the BMC Engineering team as a Graduate Firmware Engineer. You will develop low-level and embedded firmware that supports the operation, control, monitoring, and validation of advanced compute systems. You will work with experienced firmware, hardware, systems, and software engineers throughout the development lifecycle. The role combines hands-on implementation with automated testing, lab-based debugging, hardware bring-up, and analysis of interactions between firmware and the underlying platform. Start: September, 2027 Location: Austin, Texas, USA What You Will Do Design, implement, test, and maintain system and embedded firmware in C, C++, or Python. Take ownership of defined firmware features and deliver them from requirements and design through implementation, validation, and documentation. Develop and debug firmware in a Linux-based engineering environment using appropriate diagnostic tools and techniques. Create automated tests and scripts that improve firmware validation, test coverage, and engineering efficiency. Contribute to continuous integration and delivery workflows for firmware development and testing. Plan and conduct engineering experiments, analyze test data, and communicate findings clearly. Support lab setup, system configuration, hardware bring-up, and firmware validation on development platforms. Investigate firmware behavior and hardware-software

pythonlinuxartificial intelligence
View job →
🔔

Get new controls engineer jobs by email

Daily job updates · Unsubscribe anytime