JOB TITLE Observability Engineer A CAREER WITH POINT72'S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU’LL DO Design observability capabilities that give engineering teams clear insight into application health, platform performance, and issues affecting users Build scalable collection pipelines for metrics, logs, and traces across cloud-based and on-premises environments Develop actionable alerting standards that reduce noise, shorten incident response, and highlight the most important signals Partner with application and infrastructure teams to define service health indicators and improve operational readiness before production launches Automate monitoring configuration, dashboard deployment, and reliability checks to support consistent observability across the technology environment Analyze production incidents to identify telemetry gaps and improve detection, diagnosis, and recovery Create dashboards and reporting views that help teams understand trends, capacity risks, and reliability outcomes Establish practical observability
Jobiba hiring network
Deployment Strategist Lead Jobs
1,782 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current deployment strategist lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.
JOB TITLE IT Operations Engineer, Application Support A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Provide technical support for software applications and investigate, diagnose, and resolve application issues • Automate start-of-day and end-of-day checks for key applications. • Log and track incidents across applications in the production environment. • Implement monitoring and automation initiatives and develop custom solutions using Python, Shell, and/or Powershell scripts. • Create tactical support tools and scripts to improve the incident investigation process and enable • transparency into potential business impacts. • Prioritize and categorize incidents based on severity and impact. • Collaborate with the development team to improve applications based on user feedback. • Create and maintain documentation for responding to common errors and application incidents. • Assist with software applications deployment and configuration . • Provide training and assistance to users to ensure effective use of applications and systems. • Develop knowledge base resources to empower users to independently resolve common problems. WHAT’S REQUIRED • Bachelor's degree in computer science, information technology, or a related field. • Experience supporting middle- and back-office applications created in .Net, Java, C# etc. Ability to debug apps using of code, logs, alerts etc. • Literacy in complex SQL procedures/queries. • Ability to diagnose and troubleshoot technical issues.
Role Overview You’re a seasoned Site Reliability Engineer who loves owning complex infrastructure, making things run faster, safer, and with less manual effort. In this Staff‑level role, you’ll design and operate VMware‑based private cloud platforms that power mission‑critical SaaS products used by customers around the world. You’ll work across Linux, Windows Server, networking, storage, and automation frameworks to increase reliability, reduce toil, and modernize a global datacenter environment. You’ll have the scope to set technical direction, build automation at scale, and mentor engineers while staying hands‑on with VMware vSphere, F5/AVI load balancers, and hybrid Active Directory. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead the architecture, deployment, and ongoing optimization of VMware vSphere–based private cloud infrastructure across multiple global datacenters. Design and build automation using PowerShell/PowerCLI, Ansible, Python, and CI/CD tools to streamline provisioning, configuration, and compliance. Administer, harden, and troubleshoot Linux (RHEL/CentOS/Ubuntu) and Windows Server environments that host enterprise and SaaS workloads. Integrate and manage Active Directory for authentication, access control, and service accounts across hybrid on‑prem and cloud environments. Partner with network and security teams to manage firewalls, VPNs, storage, and load balancers (F5 BIG‑IP, AVI/NSX Advanced Load Balancer) for highly available services. Document architectures and runbooks, participate in on‑call and change management, and mentor engineers while influencing long‑term reliability and automation strategy. These are the essentials you’ll need to get an interview 10+ years of experience in systems or infrastructure engineering, including operating large‑scale enterprise or SaaS datacenter environments. Deep hands‑on expertise with VMware vSphere (ESXi, vCenter, DRS, HA, vMotion, distributed switches) in production
Role Overview Build the software services that power products, platforms, and better business decisions. As a Software Engineer II, you’ll develop scalable backend applications, APIs, integrations, and AI-enabled features using Python and cloud technologies. You’ll contribute to solutions from design through production, helping improve reliability, performance, security, and developer productivity. This is an opportunity to solve meaningful engineering challenges, grow your technical ownership, and collaborate with experienced engineers across the development lifecycle. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Design and build scalable backend services, REST APIs, integrations, and reusable software components using Python and, where relevant TypeScript. Develop data ingestion, transformation, service-to-service communication, and automation capabilities that support reliable product experiences. Contribute to AI-enabled features and use AI development tools responsibly to improve coding, testing, research, documentation, and delivery. Apply sound engineering practices across architecture, performance, security, testing, debugging, and maintainability. Deploy and operate services using AWS and CI/CD workflows, contributing to monitoring, troubleshooting, documentation, and continuous improvement. Partner with engineers and cross-functional colleagues through design discussions, code reviews, technical problem-solving, and knowledge sharing. These are the essentials you’ll need to get an interview 3–5 years of professional experience building and delivering production software in an agile environment. Strong hands-on experience with Python and backend development, including APIs, integrations, or service-oriented applications. Experience working with cloud platforms, preferably AWS, and familiarity with deployment or CI/CD practices. Working knowledge of software design principles, testing, debugging, performance optimization, an
This position is based in Vancouver, BC , within Diligent’s Technical Center of Excellence. We are currently hiring candidates who are based in or able to work from Vancouver . Software Engineer — Platform AI Service Levels: Software Engineer II Senior Software Engineer Staff Software Engineer Location: Vancouver Position Overview As a Software Engineer on Diligent's Platform AI team, you'll help design, build, and operate the core services that power AI-driven capabilities across Diligent's global product suite. You'll build secure, scalable, serverless services on AWS that translate AI research and models into commercial-quality, production-ready solutions — enabling customers to derive insights from their governance data. You'll work closely with AI researchers, product managers, and other engineering teams, owning your services end-to-end: architecture, implementation, deployment, and monitoring. The team operates with a strong AI-augmented engineering culture — using AI tools to accelerate coding, testing, debugging, and delivery — while applying sound judgment about when and how to apply them. Key Responsibilities Design and implement secure, scalable, fault-tolerant, high-performing solutions using AWS serverless technology — event-driven, highly observable, and built with infrastructure as code. Collaborate with AI researchers/engineers to translate AI and LLM capabilities into robust, production-grade services, and help other teams integrate them. Build and maintain the pipelines needed to deploy, monitor, and manage AI services at scale — observable, resilient, and cost-effective. Use AI-powered development tools (code assistants, test generation, architecture exploration) responsibly to accelerate delivery and improve quality, always validating outputs. Participate in architecture discussions and design reviews, and contribute to product design by understanding customer problems — especially where AI can offer a breakthrough solution. Work in
Here’s a summary of the role: Build cloud software that matters, grow your technical depth, and use modern AI tooling to do your best work. This is a hands-on engineering role for someone who enjoys solving product problems, writing clean code, and helping services run reliably at scale. You’ll work on secure, scalable microservices and APIs using TypeScript, AWS , and modern engineering practices. You’ll be part of a collaborative product engineering team where you can own features, contribute to design discussions, support production systems, and keep growing across backend, cloud, and AI-assisted development workflows. Here’s a breakdown of what you’ll do, not all of it, just the important stuff: Design, build, test, and improve backend services and APIs using Node.js, TypeScript, and AWS . Take ownership of well-defined features from planning through release, including code quality, deployment, and production support . Work closely with product managers, designers, and other engineers to turn requirements into practical, reliable solutions. Contribute to technical design conversations, code reviews, and engineering standards that keep the team moving well. Use AI tools to speed up research, coding, debugging, testing, and documentation, while checking outputs carefully and applying sound judgment. Help keep systems secure, observable, and maintainable by improving monitoring, reliability, and day-to-day development practices. These are the essentials you’ll need to get an interview: 3 to 5 years of professional software engineering experience building production applications in an agile environment. Strong backend development skills with Node.js and TypeScript, including experience building APIs or microservices. Experience with React or Angular in a product engineering environment. Hands-on experience with
You love turning real business problems into working AI solutions — and you’re not afraid to roll up your sleeves to ship them. In this role, you’ll lead Diligent’s internal AI Solutions function , a small, high-impact team that is embedding AI and GenA I into the systems thousands of colleagues use every day across Marketing, Sales, Customer Success, Finance, HR, Legal, and Product & Engineering. You’ll set the AI vision and roadmap for internal tools, architect solutions, and still build hands-on — from prototypes and reference implementations through to production-grade integrations . You’ll own how AI shows up inside ERP, CRM, BI and core IT platforms, and you’ll be accountable for making those solutions reliable, secure, compliant, and measurably valuable for the business. If you enjoy being a “player-coach” who can move seamlessly between executive conversations and deep technical reviews, this role gives you the scope, visibility, and impact to shape how a global SaaS leader uses AI to run smarter and faster. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead, coach, and grow a high-performing team of AI Solutions Architects and Engineers, setting clear goals and building the capabilities the business needs as AI demand scales. Define and own the strategy, vision, and roadmap for internal AI solutions, translating business priorities into a focused portfolio of AI and platform initiatives. Design and deliver end-to-end AI/GenAI solutions — from ideation and prototyping through production deployment, monitoring, and continuous improvement. Embed AI capabilities (such as RAG, copilots, agents, summarization, and classification) into core business applications including ERP, CRM, BI, and other enterprise systems in a robust, maintainable way.  
Overview Scale’s Finance Systems and Automation team is looking for a builder-oriented team member to help design and develop integrations, automations, and AI agents that streamline workflows across Finance, Accounting, People, and Recruiting. In this role, you will work closely with stakeholders across Finance, Accounting, People Operations, and Recruiting to understand their workflows and build integrations, automations, and agentic workflows that reduce manual effort and accelerate execution across these teams. You will leverage our internal data infrastructure, system integration tooling, and emerging AI platforms to architect scalable solutions — from traditional system integrations to intelligent agent-driven workflows — that serve as the foundation for long-term operational efficiency. We’re looking for someone who thrives on connecting systems, automating repetitive processes, and pushing toward more autonomous, AI-assisted operations. You should be comfortable navigating ambiguity, designing solutions that scale, and rigorously validating outcomes end to end. What You’ll Do Design and build agent-driven workflows and automation systems across People Operations, Recruiting, Finance, and Accounting Identify opportunities to replace manual or rules-based processes with agentic workflows Partner with stakeholders to translate business processes into scalable, automated solutions Lead the implementation of end-to-end workflows, from requirements through deployment and validation Automate candidate-to-employee transitions (e.g., Greenhouse → HRIS → provisioning systems) Build workflows to manage employee lifecycle events such as onboarding, transfers, and offboarding Automate approval flows and data synchronization across People, Finance, and Recruiting systems Support accounting and finance workflows through scalable integrations and automation Design and implement the underlying integrations and data flows that enable reliable automation and agent behavior Est
About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with the world's leading enterprises and government organizations to accelerate their AI transformation through frontier AI systems that solve real business problems. Every day, we work with organizations across finance, healthcare, manufacturing, media and telecommunications to build production AI agents that automate complex workflows, help humans, reason over enterprise knowledge, and operate safely at scale. The Opportunity Applied AI is moving faster than ever. New foundation models, reasoning techniques, agent architectures, and research papers emerge every week. Yet building AI systems that reliably solve real-world problems remains one of the hardest engineering challenges. As a Frontier Agent Engineer (Applied AI) , you'll bridge the gap between cutting-edge AI research and production deployment. You'll work directly with enterprise customers to design, evaluate, and deploy intelligent systems that combine frontier models with structured knowledge, retrieval, traditional machine learning, and enterprise software. Unlike traditional ML roles that focus on a single model or product, you'll work across a diverse portfolio of AI challenges spanning multiple industries and use cases. You may build a multi-agent research system and then participate in designing a customer intelligence platform, a healthcare copilot, or an autonomous workflow for a Fortune 100 company. If you enjoy reading new AI papers, experimenting with the latest models, and shipping production systems that create measurable business impact, you'll fit right in. What You'll Build Frontier AI Systems Design and deploy production AI agents that leverage the latest advances in large language models, reasoning, retrieval, memory, and tool use. Architect intelligent systems that combine LLMs, traditional machine learning, structured knowledge, enterprise data
Scale GP is Scale's enterprise Generative AI platform—APIs and infrastructure for knowledge retrieval, inference, evaluation, and intelligent automation. We power mission-critical workflows for leading enterprises, helping teams turn complex data and models into reliable, production-ready AI systems. We're building a new AI Enablement team to create the next generation of agent-powered tools that ground AI in real operational workflows. Our goal: help internal teams demystify their own workflows, then deploy agentic systems that reason over data, take action, and deliver measurable outcomes. We don't build in a vacuum. You'll use our own platform to solve real business problems internally—then selectively commercialize that same stack for customers. What we run on is what we sell. This is a 0→1 team. We're looking for a sharp, product-minded engineer who thrives in ambiguity, moves fast, and loves building systems from scratch alongside customers and cross-functional partners. You'll work closely with product, forward-deployed engineers, data scientists, and applied AI teams to turn real-world problems into scalable production solutions. If you like shipping fast, owning outcomes, and working across the stack—from polished frontends to distributed backends to LLM integrations—this role is for you. What You’ll Do Own full-stack features and projects end-to-end — from design through production deployment — within a larger product area Sample surfaces - Accounting Agents, Finance Copilots, GTM Agents, Agentic Experimentation Platforms Develop reliable backend services in Typescript/Python, work with distributed systems, data pipelines, and AI/ML infrastructure Integrate LLMs, vector databases, and agentic frameworks to power intelligent workflows Ship quickly through tight experimentation loops while maintaining high quality and reliability Adapt across the stack and learn new tools as needed to solve real problems end-to-end Ideal Experience 3+ years of full-tim
About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with the world's leading enterprises and government organizations to accelerate their AI transformation through frontier AI systems that solve real business problems. Every day, we work with organizations across finance, healthcare, manufacturing, media and telecommunications to build production AI agents that automate complex workflows, help humans, reason over enterprise knowledge, and operate safely at scale. The Opportunity Applied AI is moving faster than ever. New foundation models, reasoning techniques, agent architectures, and research papers emerge every week. Yet building AI systems that reliably solve real-world problems remains one of the hardest engineering challenges. As a Senior Frontier Agent Engineer (Applied AI) , you'll bridge the gap between cutting-edge AI research and production deployment. You'll work directly with enterprise customers to design, evaluate, and deploy intelligent systems that combine frontier models with structured knowledge, retrieval, traditional machine learning, and enterprise software. Unlike traditional ML roles that focus on a single model or product, you'll work across a diverse portfolio of AI challenges spanning multiple industries and use cases. You may build a multi-agent research system and then participate in designing a customer intelligence platform, a healthcare copilot, or an autonomous workflow for a Fortune 100 company. If you enjoy reading new AI papers, experimenting with the latest models, and shipping production systems that create measurable business impact, you'll fit right in. What You'll Build Frontier AI Systems Design and deploy production AI agents that leverage the latest advances in large language models, reasoning, retrieval, memory, and tool use. Architect intelligent systems that combine LLMs, traditional machine learning, structured knowledge, enterpri
OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leading orchestration platform for the autonomous enterprise, driving business transformation at the lowest total cost of ownership. Redwood empowers organizations to intelligently automate and orchestrate mission-critical business and IT processes across complex ERP, hybrid cloud, data and emerging agentic AI systems. Through its SaaS-first automation fabric—with AI embedded across the automation lifecycle—Redwood accelerates the path to autonomous operations. Backed by 30 years of experience and trusted by more than 50% of the Fortune 50, Redwood helps organizations unlock human potential to focus on innovation, growth and what’s next. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are seeking a highly skilled and passionate Full Stack Software Developer with a strong focus on Java to join our growing engineering team. In this role, you will be instrumental in designing, developing, and maintaining robust and scalable full-stack applications that power our automation and SaaS platforms. You will work across the entire software development lifecycle, from concept to deployment, collaborating closely with product managers, designers, and other engineers to deliver high-quality, impactful solutions. Design, develop, and implement highly performant and scalable full-stack applications using Java, Javascript, and related technologies. Build and maintain robust back-end services, APIs, and microservices. Develop responsive and intuitive front-end user interfaces. Collaborate with product management to understand requirements and translate them into technical specifications. Participate in all phases of the software development lifecycle, including planning, design, codin
OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leading orchestration platform for the autonomous enterprise, driving business transformation at the lowest total cost of ownership. Redwood empowers organizations to intelligently automate and orchestrate mission-critical business and IT processes across complex ERP, hybrid cloud, data and emerging agentic AI systems. Through its SaaS-first automation fabric—with AI embedded across the automation lifecycle—Redwood accelerates the path to autonomous operations. Backed by 30 years of experience and trusted by more than 50% of the Fortune 50, Redwood helps organizations unlock human potential to focus on innovation, growth and what’s next. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT As a Senior Full Stack Software Developer, you will be responsible for leading the design, development, and delivery of scalable full-stack applications, shaping system architecture, and driving engineering excellence across Redwood’s automation and SaaS platforms. Design, develop, and implement scalable, secure, and high-performance full-stack applications using Java, JavaScript, and related technologies Architect and build backend services, APIs, and microservices with a focus on scalability, reliability, and maintainability Develop responsive, accessible, and high-quality front-end user experiences Partner with product managers and stakeholders to define technical strategy and translate business requirements into system designs Own and contribute across the full software development lifecycle, from architecture and design to deployment and optimization Establish and promote best practices in coding, testing, observability, performance optimization, and AI usage Lead architectural discussions a
Graphcore Senior Principal AI SoC Validation (Bring-up lead) Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. We are developing the next generation of AI compute, a large-scale system-on-chip (SoC) designed to power future high-performance AI systems. As the SoC Validation Lead, you will be responsible for enabling pre-production software to run reliably on new silicon quickly and efficiently, before showing that the silicon meets the highest standards of quality, reliability and functionality, ready for production deployment. You will lead a team delivering post-silicon validation across the full AI SoC, working across silicon, firmware, and platform levels. The role requires a deep technical understanding, strong hands-on debug experience, and the ability to collaborate effectively with hardware, software, and systems engineering teams. Key responsibilities Define and lead post-silicon validation strategy Develop and refine the overall post-silicon validation approach for our AI SoCs, ensuring reliable and timely delivery of validated silicon, architectural correctness, feature robustness, and at-scale system reliability. Drive cross-domain debug and issue resolution Lead investigation and resolution of complex issues spanning silicon, firmware, operating systems, and platform interactions. Ensure that fixes are effective and sustainable. Promote collaboration and shared understanding Work closely with
About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence . Job Summary Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems The Team The Software Infrastructure team provides critical platforms and services for software development teams across the business. Our responsibilities include managing the CI platform and services, build engineering, component integration, and packaging and release systems. We operate in squads, fostering a culture of service ownership and empowerment for our engineers. We focus on long-term engineering solutions and strive to eliminate toil wherever possible. Responsibilities and Duties Develop, own, and maintain tools and services to support AI research and engineering teams Deploy and maintain services with Kubernetes and Docker Manage our Cloud Infrastructure using tools such as Terraform Candidate Profile Essential:
Get new deployment strategist lead jobs by email
Daily job updates · Unsubscribe anytime