Jobiba hiring network

Trading Flash Usdt On P2p And Other S Blockchain Jobs

286 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current trading flash usdt on p2p and other s blockchain jobs. Use filters to narrow by work mode, employment type, experience and date posted.

N
11 days ago

NVIDIA is seeking a Senior System Architect: Heterogeneous EDA Systems to solve a complex challenge in accelerated computing: Failure Attribution at Scale. As EDA or equivalent experience workloads scale across thousands of heterogeneous nodes, a single failure can cause massive resource waste. We need an engineer to develop and build an automated framework. This framework will ingest telemetry from CPU and GPU clusters to identify the root cause of job failures in real-time. It will distinguish between hardware faults, infrastructure instability, and software defects. What you'll be doing: Architect Failure Attribution Frameworks: Build a scalable "flight recorder" for EDA jobs that captures high-fidelity state across the CPU, GPU, and Fabric at the moment of failure. Build automated diagnostics that correlate GPU XID errors, PCIe bus failures, and CUDA memory exceptions. Connect these errors with system-level events such as OOM kills or NUMA-related hangs. Distributed Logging & Tracing: Implement low-overhead tracing mechanisms (using tracing tools or custom agents) that provide access to job execution across multi-node Slurm or Kubernetes clusters. Root Cause Automation: Develop heuristics and models based on machine learning to classify failures as "Hardware Fault," "Software Bug," or "Environment Issue." This reduces the Mean Time to Identify (MTTI) for R&D teams. Resiliency Engineering: Work closely with hardware and infrastructure teams to define "signals of impending failure," enabling proactive job migration or check-pointing before a crash occurs. What we need to see: Distributed Systems Mastery: BS, MS, or PhD in Computer Science or Electrical Engineering (or equivalent experience) with 6+ years in systems programming. Experience building automated

pythonkuberneteslinux
View job →

About the Team OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future. Our team works across custom silicon, embedded systems, operating systems, and cloud services to build reliable consumer devices and the platforms behind them. We connect kernel development with the broader software stack to deliver complete product capabilities. About the Role As an Operating Systems Engineer focused on the Linux kernel, you will design, develop, and maintain the kernel capabilities that underpin OpenAI’s consumer devices. You’ll bring deep expertise in one or more Linux kernel subsystems and carry solutions through the higher-level software stack. Your ownership will extend into the userspace services, libraries, tools, and interfaces needed to deliver complete product features. You’ll shape the boundaries between kernel and userspace, make design decisions across the stack, and see your work through development, integration, and production. In this role, you will: Build kernel capabilities: Design, implement, and maintain Linux kernel subsystem changes that support device capabilities and product requirements. Own features across the stack: Choose appropriate kernel and userspace boundaries, and build the interfaces and supporting components needed to deliver reliable features in shipped products. Debug complex system behavior: Use tracing, profiling, instrumentation, and diagnostic tools to resolve correctness, concurrency, p

linuxartificial intelligenceai
View job →
OI
15 days ago

Job Overview: We are looking for a Senior GenAI Developer to design, build, and productionize agentic AI systems—LLM-powered agents that can plan, use tools, orchestrate workflows, and operate reliably under enterprise constraints. You will own key parts of the agent architecture (planning, tool use, memory, evaluation, safety/guardrails, and observability) and deliver end-to-end solutions across RAG, function/tool calling, multi-agent coordination, and scalable deployment. Key Responsibilities Design and implement agentic systems: single-agent and multi-agent architectures (planner/executor, supervisor-worker, routing, reflection, critique, task decomposition). Build robust tool-using agents: function calling, tool schemas, tool authorization, retries, rate limiting, and sandboxing. Implement RAG + memory patterns: retrieval strategies, hybrid search, context assembly, long-term memory, and grounding/citation behaviors. Develop workflow orchestration for agent execution (state machines/graphs), concurrency controls, and deterministic execution where possible. Productionize GenAI services: APIs, background jobs, streaming responses, caching, and cost/latency optimization. Establish agent evaluation: golden sets, simulation-based evals, LLM-as-judge with mitigations, task success metrics, regression testing. Build observability and safety: tracing, token/tool telemetry, anomaly detection, prompt injection defenses, data leakage prevention, policy enforcement. Collaborate with product, security, and platform teams to deliver enterprise-ready solutions and integrate with internal systems (data, identity, workflow). Mentor engineers, set coding standards, and contribute to architecture reviews and technical roadmaps. Required Qualifications 6+ years software engineering experience; 2+ years building LLM/GenAI systems in production. Strong programming skills in Python (required) and/or TypeScript/Node.js. H

typescriptpythonnode.js
View job →
W
Wellhub
📍 Brazil• Full-time• Remote
15 days ago

Your wellbeing, our mission. Join a company shaping a healthier world. GET TO KNOW US At Wellhub we're revolutionizing workplace wellness. Our platform connects employees worldwide to the best partners for fitness, mindfulness, therapy, nutrition, and sleep—all in one simple subscription. Headquartered in NYC with team members in Europe, North America and South America, we’re on a mission to make every company a wellness company. We believe work should be fulfilling, inspiring, and balanced. Here, you’ll find a team that values wellbeing, collaboration, and different perspectives, where passion and creativity push boundaries to create real impact. Your contributions will help shape a healthier, more balanced world for you and millions of people globally. Join us in redefining the future of wellbeing! THE OPPORTUNITY We are hiring a Staff Platform Engineer with a dedicated focus on our Observability ecosystem for our Platform area in Brazil ! This is a Remote – Brazil position, meaning you can work from anywhere within the country. Please note that this role is only open to candidates in Brazil. In an environment of rapid growth and high-scale distributed architecture, your mission is to transform Observability from a passive toolset into a strategic asset using open source standards. You will act as an architect of efficiency and reliability , building a global platform that empowers engineering teams to "own what they build" with confidence. We are moving beyond basic monitoring to build a comprehensive "Observability as a Service" ecosystem. You will be responsible for evolving a self-service platform that balances performance with cost-effectiveness, solving complex challenges related to high-cardinality metrics, log retention strategies, and distributed tracing. We strive to eliminate friction. You will design the "Golden Paths" that allow developers to instrument their code instantly and gain high-fidelity signals without operatio

REMOTEpythonawsazure
View job →
SL
Sumo Logic
📍 Bengaluru• Full-time
15 days ago

Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You

reactsqlaws
View job →
SL
15 days ago

Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You

reactsqlaws
View job →
C-
CLEAR - Corporate
📍 New York• Full-time• $225K – $300K/yr
15 days ago

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. Today, CLEAR is well-known as a leader in digital and biometric identification, reducing friction for our members wherever an ID check is needed. We’re looking for a Senior Software Engineer to establish our Observability framework and foundations. You will join us to accelerate building and scaling our innovative systems that support our growing identity platform. You will drive on Observability best practices to find and fix gaps in our observability and our overall systems. You will also lead practices such as load testing, capacity planning, game days, chaos testing, and incident post-mortems. What You Will Do: Embed within the Engineering pillar to deeply understand the product and implement observability across all key flows Facilitate and build load testing cases, ensuring we understand the limits and scaling factors of our services and systems Contribute to observability and support the design of new services and systems, ensuring highly reliable and scalable concepts are implemented Build and lead practices such as game days, chaos engineering, and failure analysis Build long-term capacity plans, with an eye toward reliability and cost-efficiency Who You Are: 6+ experience writing production-grade software in a modern language, such as Java and Python. Strong knowledge of distributed systems concepts (think CAP theorem), microservices architecture, and distributed tracing . Experience with modern observability systems such as Datadog. Experience with performance debugging tools and patterns. You should be able to read a f

pythonjavagit
View job →
DU
15 days ago

About the Team DoorDash Labs is an independent innovation hub within DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. Our mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We are a highly senior team composed of former pioneers from a variety of different robotics industries. If you have a passion for applying AI and robotics solutions in a service used by millions of people, then we want to talk to you! Please Note: This role is based in San Francisco, CA, and requires being in-person 5 days a week. About the Role We are seeking a Senior RF Engineer to join our multidisciplinary team developing mission-critical RF/electrical systems for unmanned platforms.This role offers a rare opportunity to gain deep insight into RF design and validation, contributing across the full product lifecycle, from early concept and prototyping through production. In this role, you will apply deep expertise in electromagnetic theory and wireless system design to ensure autonomous platforms operate with exceptional reliability and precision. You will design, optimize and validate RF / mixed-signal PCBAs with a strong focus on interference mitigation, troubleshoot system-level desense, and leverage a thorough understanding of wireless technologies including LTE, 5G NR, Wi-Fi, GPS, and UWB. You’re excited about this opportunity because you will… Lead the development of innovative solutions that span across RF, Mixed signal and mechanical domain. Design RF hardware as well as wireless testbeds to support system validation, and lead the execution and documentation of system-level design and verification Provide expertise, mentorship to partner teams regarding EMI mitigation strategies, and cross-functional debug processes Characterize EMI caused by internal & external aggressors by tracing root causes using EMC probes

awsgitrest
View job →

SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime. With the ability to build, scale and manage security across the cloud, hybrid and traditional environments in real-time, SonicWall provides relentless security against the most evasive cyberattacks across endless exposure points for increasingly remote, mobile and cloud-enabled users. With its own threat research center, SonicWall can quickly and economically provide purpose-built security solutions to enable any organization—enterprise, government agencies and SMBs—around the world. For more information, visit www.sonicwall.com or follow us on Twitter , LinkedIn , Facebook and Instagram . As a Software Dev Senior Engineer , you will own the reliability, scalability, and operational excellence of our Cloud-based services. You will define and enforce reliability standards, drive the adoption of SRE practices across engineering teams, and build the systems and tooling that keep our production infrastructure healthy. We follow a DevOps model: Development and Operations teams are integrated, and the SRE function acts as the reliability layer — setting Service Level Objectives, managing error budgets, and continuously reducing toil through engineering. Key Responsibilities: Define, publish, and continuously refine Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs ) for all critical services, partnering with product and engineering leadership. Own the error budget framework: track consumption, enforce error budget policies, and drive reliability investments when budgets are at risk. Lead the design and implementation of comprehensive observability platforms — metrics, structured logging, and distributed tracing — to ensure full visibility into pro

pythonsqlpostgresql
View job →

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services that our customers can trust. We embrace an automation-first mindset and continuously invest in platform engineering, observability, and operational excellence to enable our engineering teams to move quickly and safely. This role is ideal for an experienced Site Reliability Engineer who enjoys solving complex technical challenges at scale, building automation, and improving the reliability of production systems. You will serve as a key contributor within the EPG SRE organization, partnering closely with software engineers, architects, and product teams to design, build, and operate world-class cloud services. What You'll Be Doing Reliability & Operations Design, build, and operate large-scale cloud infrastructure and production services. Participate in an on-call rotation supporting highly available customer-facing systems. Lead incident response efforts and drive post-incident reviews focused on systemic improvements. Define, measure, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. Partner with engineering teams to improve service availability, scalability, performance, and resilience. Continuously improve observability through metrics, logging, tracing, dashboards, and alerting. Eng

pythonsqlpostgresql
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team We’re hiring Software Engineers to join our broader Infrastructure organization, which supports multiple high-impact teams. Depending on your interests and experience, you could work on one of several focus areas—including Core Distributed Systems, Reliability Engineering, Observability, Developer Productivity or Cloud Infrastructure. About the Role All teams are deeply collaborative, work on mission-critical services, and are responsible for building distributed, scalable infrastructure to bring OpenAI’s technology to the world through products like ChatGPT and the OpenAI API. You’ll work closely with stakeholders to understand infrastructure, data and compute needs, setting the technical strategy that supports cutting-edge research and product development. This is a critical role for someone who is passionate about solving complex engineering problems at scale, ensuring their performance, scalability and reliability Team Focus Areas Distributed Systems: Owning and building important, highly scalable, available, performant, and reliable distributed systems (and their building blocks) to power the entire stack at OpenAI Systems Engineering: Work across layers of the stack—debugging system bottlenecks, evolving core infrastructure, and solving novel problems in performance and scalability. Reliability Engineering: Build scalable, fault-tolerant systems and lead efforts around service health, incident response, and resilience. Observability: Design and maintain observability tooling (metrics, logs, tracing) to give teams visibility into production systems at scale. Developer Productivity: Create tools, environments, and workflows that help engineers ship high-quality software faster and more safely. Cloud Infrastructure: Own the cloud-native infrastructure (compute, networking, storage) that underpins all services and research workloads. Databases: Building high performance, distributed database systems that power all of OpenAI's product stack. In this

pythonawskubernetes
View job →

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response. What you'll be doing: Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow) Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes Drive upstream contributions to OVN-Kubernetes and related open-source projects Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure Drive reliability through incident management, resource monitoring, and performance tuning<

pythonawsazure
View job →
A
1mo ago

About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role: As a Site Reliability Engineer, you will play a crucial role in ensuring the smooth operation of all user-facing services and other Anyscale production systems. Anyscale values diversity and inclusion, and we encourage applications from individuals of all backgrounds. This includes processes for provisioning, negotiating prices, managing costs, seeing opportunities for teams to reduce wastage by finding applications across the company. You will apply sound engineering principles, operational discipline, and mature automation to our environments and the Anyscale codebase as we scale. As part of this role, you will: Develop a unified perspective on how cloud components are utilized across the company, taking into account diverse needs and requirements. Ensure that deployment methodologies align with the company's reliability goals. Build systems that promote understanding of production environments, facilitating quick identification of issues through robust observability infrastructure for metrics, logging, and tracing. Create monitoring and alerting systems at different levels, enabling teams to easily contribute and enhance the overall monitoring capabilities. Establish testing infrastructure to s

machine learningai
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer - External Observability Platform Location: Bellevue, WA (Hybrid: 3 days/week in-office) Team: Infrastructure & Observability Platform Engineering About the Role Snowflake’s Data Cloud processes exabytes of data across multi-cloud global environments every day. Delivering seamless reliability and real-time visibility to thousands of global enterprise customers requires an Observability Platform built on hyper-scalable backend distributed systems. We are seeking a Senior Software Engineer to own key components of our AI native External Observability Platform . In this role, you will contribute to the technical road map for customer-facing telemetry, system metrics, audit logs, distributed tracing, and actionable operational insights. You will build high-throughput, low-latency infrastructure capable of ingesting, processing, and serving petabytes of telemetry data with strict SLA guarantees. You will join a team of world-class engineers in our Bellevue, WA office. To be successful, you must be deeply technical, capable of leading complex technical projects, and skilled at collaborating with the brightest technical minds in the industry. Key Responsibilities Develop and Scale Distributed Infrastructure: Design and implement key components of Snowf

javavueaws
View job →

The Senior Enterprise Engineer opportunity Statsig is building an Enterprise Engineering team for engineers who want unusually direct ownership of high-stakes customer outcomes. Like a Forward Deployed Engineering team, Enterprise Engineering works extremely close to customers and tackles the technical problems that matter most to them. The difference is the scope of ownership. Forward Deployed Engineering is often concentrated around implementation or onboarding; Enterprise Engineering is responsible for a customer’s technical success throughout the entire customer journey. You’ll stay close as needs evolve, answering questions, investigating across the stack, identifying root causes, shipping fixes and features, and remaining accountable through resolution. This is a software engineering role with a tight customer feedback loop. It is ideal for a pragmatic generalist, especially someone with backend or infrastructure depth who enjoys ambiguous problems, production systems, fast decisions, and immediate impact. What you’ll do Serve on the engineering front line in Unthread, our Slack aggregation and workflow layer, for key customer accounts. Answer technical questions by reading the code, tracing behavior, inspecting production signals, and building a clear explanation, not by forwarding the thread. Diagnose and fix bugs, then validate the outcome with the customer. Design and implement product or platform features when doing so is the fastest, highest-quality way to solve the customer’s problem. Act as the engineering point of contact for urgent issues in key customer accounts’ Slack channels, coordinating the response while retaining technical ownership. Represent engineering in customer office hours and, for selected accounts, recurring weekly meetings. Partner closely with customer-facing and product teams while maintaining crisp ownership, communication, and escalation paths. Identify recurring patterns and improve tooling, documentation, observability, APIs,

gitrestai
View job →
🔔

Get new trading flash usdt on p2p and other s blockchain jobs by email

Daily job updates · Unsubscribe anytime