Overview: Qsight is a high-growth division of Guidepoint focused on building data intelligence solutions for the healthcare sector. Qsight leverages proprietary datasets and rigorous analysis of alternative data sources to generate actionable insights for top-tier institutional investors, medical device manufacturers, and pharmaceutical companies. The Qsight team develops market intelligence products designed to be highly relevant, accurate, and scalable – delivering superior insights to a diverse, global client base. We are seeking an experienced, motivated Tehnical Operations Engineer to join our growing team. This is a multiple-hats role focused on SaaS/platform operations and tier-2 support for client-facing systems. You will own the administration and reliability of key tools, troubleshoot and resolve escalations with clear documentation, and build lightweight automation and reporting to reduce manual work as we scale. You will partner closely with Customer Success, Product, and Engineering to proactively monitor, support, and improve critical systems. Through practical, creative problem-solving, you will strengthen reliability, accelerate time to resolution, and increase operational visibility. Day to day, you will triage and resolve client technical questions, manage vendor license administration and renewals, and produce reporting that informs operational decisions. This role is a launchpad toward an SRE/Platform Engineering track as you grow into deeper automation, reliability engineering, and systems design work. This is a hybrid position based out of our Toronto office. What You’ll Do: Platform Support Own routine ops and configuration changes for critical SaaS platforms – Including Tableau, Freshdesk, Datadog, and our own client facing and internal portals Configure and maintain Freshdesk portals, routing, SLAs, permissions, integrations, etc. based on business requirements. Automate manual operations with Python, PowerAutomate, and shell scri
Jobiba hiring network
Technical Operations Engineer Jobs
15 active opportunities · Updated for September 2026
Fresh results
15 shown
Explore current technical operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Payments organization owns some of Stripe’s most critical payment flows and a platform that processes hundreds of billions of dollars in payments a year. Our team is responsible for translating complex partner specifications related to network costs (interchange and scheme fees) into simplified logic for internal and external consumption. This also drives decisions and product recommendations to manage the underlying network costs paid by Stripe and our users. The team partners closely with the engineering, product, finance, and partnership groups to manage and understand Stripe’s network costs. Our work is core to Stripe’s business, as Technical Operations roles in Payments are a dynamic and key component of Stripe's success. We sit at the intersection of product/platform engineers and financial partners, connecting them to ensure that everyone thrives and nothing is lost in translation. What you’ll do We are looking to add payment enthusiasts who enjoy interpreting complex cost structures and ever changing payment network systems in order to optimize on behalf of Stripe and our users. You will be instrumental in building Stripe’s approach to managing our global network cost base.</p
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Staff Hardware Engineer to provide advanced operational, diagnostic, and engineering support for Graphcore’s Arm-based hardware platforms across lab and data center environments. This role focuses on supporting hardware bring-up, validation, and troubleshooting of complex AI compute platforms, including server blades, racks, and rack-scale infrastructure. The successful candidate will collaborate closely with engineering, platform, and data center teams to ensure the reliability and performance of next-generation AI systems. The Team The Systems Engineering and Hardware Engineering teams are responsible for enabling the bring-up, validation, and operational reliability of Graphcore’s AI infrastructure platforms. The team works closely with server engineering, firmware teams, platform architects, and data center operations to support the development, testing, and deployment of next-generation AI compute systems. This collaborative environment enables rapid problem-solving and continuous improvement of Graphcore’s hardware platforms from early development through production deployment.
Integration Reliability Engineer, Technical Operations About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team APAC Payins TechOps is a newly formed team based in Singapore. Our charter is to improve the health and resilience of our payment acquiring volume, making it easier for Stripe to build and operate the integrations that support hundreds of billions of dollars in payments annually. Our team partners closely with payments and platform engineering teams to assess systems or processes that create high operational workloads, then develops durable solutions via code, tooling, data, and process improvements. We are responsible for the financial data quality of key systems at Stripe, ensuring that data is quarantined without impacting downstream systems while implementing the right changes upstream to prevent recurring issues. The team's work has a direct impact on Stripe's ability to expand into new markets and offer more sophisticated payment features to merchants. What you’ll do Responsibilities Scope and lead technical initiatives end-to-end: identify problems worth solving, propose the right solution approach, and deliver on that solution — not just execute on a pre-defined plan. Investigate problems in systems by tracing problems through Stripe’s stack. You’ll examine code, write SQL queries, read logs, and inspect data pipelines to understand system behavior, then make changes to address. Examine updates being made by Stripe’s financial partners to understand impact, and make the changes within Strip
Position Overview As SingleStore’s IT Operations Engineer, you will help shape the IT toolset used by our end users. This is an active, hands-on position responsible for the planning, design, development, and Tier 1 support of several key technical areas at the SingleStore IT team, including end-user support, client engineering, executive support, and infrastructure application support. This is an incredible opportunity for someone to build upon their technical strengths and be a part of IT at SingleStore team . Roles and Responsibilities: Administering a wide variety of SaaS applications. Some main applications that need to be supported are OKTA (+ Workflows), Google Workspace, Slack, and Atlassian tools (JIRA + Confluence), MDM administration. Keep up to date with new features and new releases in these applications to identify opportunities for better automation or features that could be useful for our environment. Seize opportunities across the IT Operations team to eliminate manual work through tooling, integrations, and automation of IT workflows. Respond to tickets and execute new hire onboarding and user separation processes. Support members of the team with troubleshooting and resolution of complex issues. Design, architect, implement and maintain systems and solutions for various IT-related topics, including but not limited to staff computer hardware, operating systems, software applications, networking, videoconferencing, and printers. Partner and collaborate with all business units to help them evaluate hardware and software solutions. Able to communicate effectively and concisely with the entire company. Analyze existing processes, suggest and make improvements, and implement business processes where none exists. A desire to learn and expand your horizons; take on new challenges as the business scales Required Skills and Experience: Minimum 2 years of relevant experience Prior experience in implementing and administering Google Workspac
JOB TITLE IT Operations Engineer, Application Support A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Provide technical support for software applications and investigate, diagnose, and resolve application issues • Automate start-of-day and end-of-day checks for key applications. • Log and track incidents across applications in the production environment. • Implement monitoring and automation initiatives and develop custom solutions using Python, Shell, and/or Powershell scripts. • Create tactical support tools and scripts to improve the incident investigation process and enable • transparency into potential business impacts. • Prioritize and categorize incidents based on severity and impact. • Collaborate with the development team to improve applications based on user feedback. • Create and maintain documentation for responding to common errors and application incidents. • Assist with software applications deployment and configuration . • Provide training and assistance to users to ensure effective use of applications and systems. • Develop knowledge base resources to empower users to independently resolve common problems. WHAT’S REQUIRED • Bachelor's degree in computer science, information technology, or a related field. • Experience supporting middle- and back-office applications created in .Net, Java, C# etc. Ability to debug apps using of code, logs, alerts etc. • Literacy in complex SQL procedures/queries. • Ability to diagnose and troubleshoot technical issues.
JOB TITLE IT Operations Engineer, EQUITY TRADING technology A Career with point72’s technology TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source and AI solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you’ll do Provide operational and technical support for the firm’s trading platforms to ensure optimal performance Coordinate and execute software upgrades and releases across trading platforms Manage all production and UAT trading platforms Support trading systems during incidents, including both remediating the issue and ensuring ongoing communication with end-users and stakeholders Liaise with brokers, service providers, and other internal technology groups and stakeholders Design and implement tools and reports to enhance department efficiency Assist platform users during onboarding processes What’s REQUIRED 7+ years of application support experience within the financial services industry Experience working with Linux and other languages Ability to work effectively within a global team, adapting to varying time zones and flexible shift schedules Strong understanding of order management workflows Commitment to the highest ethical standards About point72 Point72 is a leading global alternative investment firm led by Steven A. Cohen. Building on more than 30 years of investing experience, Point72 seeks to deliver superior returns for its investors through fundamental and systematic investing strategies across asset classes and geographies. We aim to attract and retain the industry’s brightest talent by cultivating an investor-led culture and co
NVIDIA pioneers computer graphics, gaming, AI, and accelerated computing. We are looking for a Technical Platform Operations Lead to join our team and play an important role in scaling Sales AI applications and platforms. This position offers the opportunity to shape how these solutions operate after launch and help ensure they remain reliable, secure, well governed, widely adopted, and continuously improved. You will collaborate with Sales, Product, Engineering, Data, Security, and IT teams to strengthen platform health, improve the user experience, and increase business impact. What you’ll be doing: Lead end-to-end post-launch operations for Sales AI applications, including availability, performance, support readiness, releases, upgrades, and lifecycle planning. Develop effective processes for incident response, problem management, changes, and issue resolution. Coordinate timely recovery and lasting improvements. Analyze service-level indicators and objectives, adoption metrics, dashboards, alerts, and user feedback to identify risks, performance degradation, and usage gaps. Collaborate with partner teams to translate operational signals and user needs into prioritized improvements and roadmap inputs. Improve adoption and business value through usage analytics, enablement, feedback loops, and user experience enhancements. Establish governance practices for security, access controls, compliance, documentation, and platform support. Develop automation, observability, and self-service capabilities that simplify operations and reduce repetitive work and recurring incidents. Prepare new AI capabilities and releases for production with runbooks, monitoring, rollback plans, support models, and partner enablement. What we need to see: 8+ years of experience in technical operations, pl
Cloud Operations Engineers are responsible for building internal tools and process automation. Day-to-day duties are creating and monitoring systems alert dashboards, reviewing critical event and system logs, accessing customer instances that underpin their production databases, and performing server administration duties including performance troubleshooting. Applicants must be critical thinkers who are quick to detect, resolve, or escalate issues that are sometimes broad in scope and difficult to trace. We are looking for a Lead with strong technical leadership experience as well as technical depth who is looking to collaborate closely with Cloud Operations Engineering Management in building and maintaining a high-performing team that delivers high quality outcomes while fostering psychological safety and professional growth. We are looking to speak to candidates who are based in Dublin for our hybrid working model. Core responsibilities Team leadership: partner with and assist COE Management with the tasks of providing ongoing technical feedback to engineers, support their growth and creating an inclusive team environment Execution and delivery: play a key role in guiding team members through project deliverables ensuring high quality outcomes while also assisting in meeting or resetting timelines when required Time management: between assisting team members with day to day tasks ranging from incident to project management Cross-functional collaboration: work closely with Product, Technical Services and R&D to surface team’s pain points and drive alignment with the goal of providing an excellent user experience to the end customer Coordinate with Lead counterparts within Cloud Operations as well as Technical Services to ensure our uptime guarantees to the MongoDB Atlas customer base Assist and collaborate with the team on scoping, designing, deploying and ongoing maintenance of systems that focus on reducing mean time to resolve customer incidents Detec
About Paytm Group: Paytm is India's leading mobile payments and financial services distribution company. Pioneer of the mobile QR payments revolution in India, Paytm builds technologies that help small businesses with payments and commerce. Paytm's mission is to serve half a billion Indians and bring them to the mainstream economy with the help of technology. OVERALL ROLE: The Technical Manager will (lead a team to) manage the day-to-day M&E Operations and activities for the assigned property/facility, and be the on-site key point of contact for key location stakeholders. The role will assume overall responsibility for site budgets, accounting and finance, maintenance and operations, contract services, purchasing of material, equipment & supplies, occupancy services and helpdesk. MAJOR RESPONSIBILITIES ▪ Operations Management: ● Establish Engineering & Operational procedures and roll out the same for site staff ● Establish contacts with local authorities on the facility related issues and maintain the relationship. Responsible for all legal & authorities related compliances pertaining to facility & engineering systems ● Plan and manage the budgets for Engineering & Operational contracts ● Carry out Technical Audits for all installations at periodical intervals ● Review the maintenance/service practices of M&E Contractors to deliver quality work practices in line with the manufacturer recommendations ● Plan & take responsibility for smooth operations of all Mechanical, Electrical, Plumbing installations and Civil works pertaining to the facility ● Responsible for planning a critical spares list for all installations as per manufacturer's recommendations and inventory ● Responsible for development of all maintenance related schedules and shutdowns in consultation with Clients / OEMs ● Periodically inspect the logbooks, checklists and PPM schedules for a better management of Engineering systems ● Work towards the 'ZERO' down time and set
CAPCO POLAND *We are looking for Poland based candidate. Capco is a fully independent, global management and technology consultancy. For 25 years we have combined innovative thinking with deep industry knowledge to deliver business consulting, digital transformation and technology services to Finance and Energy markets. Our collaborative and efficient approach helps clients reduce costs and manage risk and regulatory change while increasing revenues. We are thinkers, innovators, and disruptors. We are small enough to care but large enough to matter. We are seeking a highly skilled Security Operations Engineer to support the expansion of a strategic security program focused on onboarding critical applications into enhanced monitoring capabilities.In this role, you will play a key part in building and optimizing SIEM detection capabilities, supporting threat verification, and enabling regulatory alignment with DORA (Digital Operational Resilience Act) requirements by the end of 2026. You will work at the intersection of SIEM engineering, threat modelling, and security operations , contributing directly to improving detection accuracy and strengthening overall security posture. Key Responsibilities: Detection Engineering: Design, build, and optimize SIEM detection rules (with a focus on Microsoft Sentinel) Testing & Automation: Develop and execute test cases for detection logic; automate validation processes using scripting Application Onboarding: Support onboarding of critical applications into the security monitoring ecosystem Requirements Gathering: Collaborate with application teams to define logging requirements and detection use cases Workshop Facilitation: Lead and moderate workshops with stakeholders to align on threat scenarios and security capabilities Technical Documentation: Produce clear and comprehensive documentation covering detection logic, threat models, and validation results Collaboration: Work closely with SOC, enginee
About the Team OpenAI, in close collaboration with our capital partners, is building the world’s most advanced AI infrastructure ecosystem. Our Industrial Compute organization develops and deploys large-scale AI campuses designed to support the next generation of frontier model training and inference workloads. The Hardware Operations team is responsible for ensuring the reliability, availability, and lifecycle health of OpenAI’s compute infrastructure. We partner closely with Data Center Operations, Fleet Health Engineering, Manufacturing, Network Infrastructure, Capacity Planning, and our infrastructure partners to maintain world-class operational performance across rapidly expanding AI environments. As we scale globally, we are building the operational frameworks, reliability standards, and sustaining engineering practices required to support thousands of GPUs and servers across multiple campuses. About the Role We are seeking a Datacenter Hardware Technician Lead to serve as the senior on-site technical authority for hardware reliability and fleet health at one of OpenAI’s flagship AI campuses. This role operates at the intersection of hardware operations, sustaining engineering, and fleet reliability. You will partner closely with Cloud Service Provider operations teams, OpenAI fleet-health engineers, hardware engineering teams, and OEM vendors to identify, diagnose, and resolve hardware issues affecting production systems. Beyond day-to-day operational support, you will drive root cause investigations, reliability improvement initiatives, lifecycle management programs, and operational readiness efforts. You will help establish hardware maintenance standards, operational procedures, and best practices that scale across future OpenAI infrastructure deployments. The ideal candidate combines deep hands-on datacenter hardware expertise with strong troubleshooting, failure analysis, and cross-functional leadership skills. Candidates must be able to sit onsite at our
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Staff Security Operations Engineer , you will own and continuously mature capabilities across Attack Surface Management, Vulnerability Management, Zero Trust, Secrets and Credential Security, Detection Engineering, and Security Automation . This is a hands-on role requiring strong security engineering expertise combined with the ability to build teams, establish operating processes, measure outcomes, and drive remediation across Engineering, Cloud, Infrastructure, and Product organizations. You will partner closely with GISO leadership and global security teams to translate security strategy into measurable execution and risk reduction. WHAT YOU’LL DO Lead and mature enterprise Attack Surface and Vulnerability Management capabilities across cloud, infrastructure, endpoints, applications, and internet-facing environments. Drive risk-based vulnerability prioritization using asset criticality, exposure, exploitability, known exploitation, threat intelligence, and compensating controls. Establish operating processes, remediation SLAs, KPIs/KRIs, dashboards, and governance to measure and drive security risk reduction. Identify systemic security gaps and develop scalable technical and operational solutions. Provide technical leadership across Zscaler/Zero Trust, secrets and credential security, SIEM/detection engineering, EDR, cloud security, and security automation. Drive automation and integrations using APIs, sc
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are looking for an IT Support / Operations Engineer to join Baseten as we continue to scale our IT team. In this role, you will play a critical part in bringing our technical support entirely in-house to provide a seamless, high-touch experience for all Baseten employees. As we continue to scale, you will be the primary point of contact for day-to-day technical issues, allowing you to have a direct impact on our team's productivity and overall office environment. This position is ideal for a hands-on problem solver who enjoys a mix of hardware and software troubleshooting, user lifecycle management, and maintaining the physical IT infrastructure of a modern office. While you will focus heavily on elevating our internal support standards, you will also assist with systems administration and workflow automation as our company evolves. This is a hybrid role based out of our San Francisco or New York office, following our standard policy of three days per week in-person to ensure our physical office and AV systems remain high-performing and reliable. RESPONSIBILITIES Serve as the escalation point for day-to-day technical support, diagnosing and resolving hardware and software issues across our Mac and Windows fleet Manage user lifecycle administration including provisioning, deprovisioning, and access management across all systems and services Own the IT onboarding experience for new employees — from laptop set
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: BizTech fosters culture and connection at Airbnb by providing reliable corporate tools, innovative products, and technical support for all teams. We drive technical breakthroughs and strategies that redefine what it means to belong anywhere, delivering greater value for the business and our people. The Global Operations team at BizTech manages production services across Airbnb’s corporate environment, delivering reliable operations through Observability, Incident Management, Core Operations, and AI-enabled automation. We partner across BizTech to scale service quality, efficiency, and resilience. The Difference You Will Make: As an Operations Engineer, you'll apply AI at the forefront of BizTech's operational health: using LLM-powered triage and intelligent automation to resolve tickets, speed up incident response, and build self-healing observability that catches problems before they escalate. AI fluency is core to this role, not an add-on. You'll prototype agentic workflows, embed AI into runbooks and diagnostics, and continuously look for repetitive work automation can take over. Success looks like a shrinking backlog of recurring ticket categories through AI-assisted automation, faster MTTR powered by intelligent alerting and root-cause suggestions, and dashboards/reporting enhanced with AI-driven insights that give stakeholders clear, trustworthy visibility into service health and data quality. A Typical Day: Manage the ticket queue prioritizing and resolving requests while identifying recurring categories to automate or deflect Participate in a rotating on-call and incid
Get new technical operations engineer jobs by email
Daily job updates · Unsubscribe anytime