About the Team The Integrated Marketing team sits at the center of product, brand, creative, media, research, and GTM work. We partner closely with PMM, Product, Developer Relations, Design, Comms, and agency teams to shape launches and campaigns that are clear, distinctive, and grounded in real audience insight. This is a growing function, so the team needs people who can both raise the quality of the work and build the operating model around it: bringing strong judgment, high agency, and trusted partnership to complex, visible moments. About the Role OpenAI business launches move quickly and often involve new products, model capabilities, or research that can change how businesses operate. We’re looking for an integrated marketing manager to help shape those launches from early strategy through execution. You’ll develop a point of view on the audience, positioning, creative direction, and channels, and help turn complex product and research advances into campaigns that connect with customers. You’ll partner across Product, Product Marketing, Research, Brand, Creative, Communications, Growth, and Sales, along with external agencies. The right person understands technology, has strong strategic and creative instincts, and knows how to bring together the right people and ideas to produce work that is clear, credible, and effective. This is a hybrid role based in San Francisco, with three days a week in the office. In this role, you will: Develop integrated marketing strategies for business product launches, new capabilities, model releases, and research moments. Partner with Product, Product Marketing, and Research to understand what’s changing, identify the right audiences, and shape positioning and launch narratives. Lead campaigns that bring launches to life across creative, communications, customer marketing, developer marketing, growth, sales, and paid, owned, and earned channels. Identify customer stories, product demonstrations, and creative concepts that make
Jobiba hiring network
Model Behavior Engineer Jobs
4,916 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current model behavior engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About the Team ChatGPT relies on a large and growing GPU fleet to serve inference workloads reliably and efficiently. We develop the systems and tools that make it possible to introduce new models, manage production deployments, respond to operational issues, and use infrastructure effectively at scale. Our work spans distributed systems, platform engineering, infrastructure automation, and developer experience. We partner closely with research, infrastructure, and product teams to make model deployment more reliable, more efficient, and easier to manage. About the Role We are looking for a software engineer with experience building or operating large-scale production systems. You will design and develop systems that support the model lifecycle in production, including deployment orchestration, configuration management, operational automation, reliability, and capacity management. You will help transform complex operational processes into scalable platform capabilities that enable teams across OpenAI to deploy and manage models with greater confidence and less manual effort. This role is a good fit for engineers who enjoy solving complex operational problems and building software that makes production infrastructure easier to run at scale. In This Role, You Will Build and evolve the platform used to deploy, configure, and manage models across ChatGPT. Develop systems for deployment orchestration, model rollouts, operational visibility, and production readiness. Create abstractions and tooling that simplify complex infrastructure and improve the developer experience. Automate operational workflows, including incident detection, diagnosis, mitigation, and recovery. Improve the reliability, scalability, and efficiency of model deployments and the infrastructure that supports them. Build systems that support capacity planning, resource allocation, and infrastructure utilization. Partner with research, infrastructure, and product engineering teams to identify common chal
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: We are seeking an experienced Product Marketing Manager with a strong background in engaging developer audiences and delivering impactful go-to-market programs for native AI and enterprise companies. This role requires someone who is both technically savvy and strategic, with a proven track record of crafting compelling product narratives and building marketing assets that resonate with technical decision-makers. This role is specifically focused on our Model API product offering at Baseten. If you’re passionate about AI infrastructure, developer engagement, and simplifying complex technologies for real-world adoption, we want to hear from you. RESPONSIBILITIES: Positioning & Messaging: Develop clear and differentiated messaging that articulates the value of Baseten’s inference platform to developers and enterprise customers. Narrative Development: Shape how the market thinks about closed-to-open weights models and what matters most when building inference. Go-to-Market Strategy: Own the launch process for new features and products, collaborating closely with product, engineering, sales, and growth teams. Content Development: Create high-quality marketing assets, including white papers, technical blogs, demos, and customer case studies. Sales Enablement: Build resources and programs that empower our sales teams to effectively communicate Baseten’s capabilities and benefits. Market Insights: Understand the
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of committed researchers and engineers. The mission of the team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will work on advancing core audio model serving metrics, including latency, throughput, and quality by diving deep into our systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You’ll collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. You may
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Evaluation is critical to making progress in scaling intelligence. As models continue to become superhuman in many real-world use cases, we must continue to develop new evaluation techniques that accurately reflect what models are already capable of, as well as set the agenda for what future models should be capable of. In this role, you are responsible for creating these next-generation evaluation methods and infrastructure to measure LLM progress. As a Senior Research Scientist, Model Evaluation, you will: Create ambitious new evaluation benchmarks that push the limits of what our models can accomplish. Work on highly cross-functional teams to translate model feedback into trustworthy, repeatable evaluations. Conduct research to advance the state-of-the-art in LLM evaluation methods, including training LLM judges; refining LLM-based data synthesis pipelines; and improving evaluation efficiency. Build scalable and reusable tools for digging into model performance. You may be a good fit if: You enjoy rapidly building prototypes that demonstrate the boundaries of what LLMs are capable of, and you have developed res
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . As a Senior Product Manager for Signal Lifecycle within Trust & Safety, you'll own the product strategy for the ML platform that powers how Pinterest trains, evaluates, deploys, and measures content safety models at scale. You'll lead the development of ML Signal Management — making ML signals first-class entities with unified metadata and identity across systems. Partnering deeply with ML engineering, data science, content safety, and enforcement systems, you'll drive a platform whose scope is expanding from T&S into content quality, ads safety, and beyond. What you'll do: Own and drive the Signal Lifecycle product roadmap, including ML Flywheel infrastructure, auto-deployment, model onboarding, golden dataset management, and signal performance measurement Define and ship ML Signal Management — a unified backbone that elevates ML
About the Team Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never been able to before. We focus on performant and efficient model inference, as well as accelerating research progression via model inference. About the Role We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment. In this role, you will: Work alongside machine learning researchers, engineers, and product managers to bring our latest technologies into production. Work alongside researchers to enable advanced research through awesome engineering. Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack. Build tools to give us visibility into our bottlenecks and sources of instability and then design and implement solutions to address the highest priority issues. Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware. You might thrive in this role if you: Have an understanding of modern ML architectures and an intuition for how to optimize their performance, particularly for inference. Own problems end-to-end, and are willing to pick up whatever knowledge you're missing to get the job done. Have at least 5 years of professional software engineering experience. Have or can quickly gain familiarity with PyTorch, NVidia GPUs and the software stacks that optimize them (e.g. NCCL, CUDA), as well as HPC technologies such as InfiniBand, MPI, NVLink, etc. Have experience architecting, building, observing, and debugging production distributed systems. Bonus point if worked on performance-critical distributed systems. Have need
About the Team We’re hiring software engineers to make OpenAI’s Model Performance teams more productive. These teams work on the systems, tooling, and infrastructure that help improve model performance across OpenAI’s training and inference workloads at frontier scale. About the Role We’re looking for an autonomous, high-ownership developer productivity engineer who cares deeply about helping other engineers move faster, safer, and with more confidence. This role will sit within OpenAI’s Model Performance organization, contributing to developer infrastructure, CI systems, testing workflows, tooling, and broader performance infrastructure efforts. There is also a strong opportunity to contribute to the Triton project and help improve the systems that support performance-critical engineering work across OpenAI. In this role you will: Improve development workflows for engineers working on model performance infrastructure Design and improve CI/CD, release, validation, and testing pipelines Build and maintain tools that improve reliability, iteration speed, and engineering confidence Partner closely with engineers to identify friction in testing, debugging, deployment, and development workflows Contribute to infrastructure efforts that support performance-critical training and inference systems Help improve developer experience across Python-heavy codebases and performance-oriented infrastructure Work in a high-context, ambiguous environment where ownership and good judgment matter You might thrive in this role if: You are motivated by enabling the people around you and helping engineers do their best work You have strong experience with CI/CD, developer infrastructure, testing systems, tooling, or build/release workflows You are highly collaborative, empathetic, and comfortable partnering deeply with technical teams You are strong in Python and enjoy building reliable, scalable developer tools and infrastructure You have experience improving large-scale engineering work
Our team is seeking to extend the internship of our current AI Research Intern for the TAO (Train, Adapt, Optimize) Multi-Modal Model Development project, recognizing their exceptional performance and strong alignment with the team’s research goals. Their innovative ideas and technical contributions have significantly enhanced our work. Given the rapidly evolving field of multi-modal AI, encompassing vision-language modeling, universal segmentation, and large-scale model training, extending this internship will provide further growth opportunities for the intern while strengthening our team’s capacity to develop scalable, high-impact AI solutions. Embark on an exciting journey with NVIDIA, a global leader in AI and accelerated computing. As an AI Research Intern focusing on multi-modal AI and vision-language model development within the TAO framework in Hanoi/HCM City, Vietnam, you will be at the forefront of advancing cutting-edge machine learning research. You’ll collaborate with a talented team of engineers and researchers dedicated to developing state-of-the-art deep learning models for tasks such as image segmentation, cross-modal understanding, and universal representation learning. This internship offers a unique opportunity to contribute to next-generation AI systems with real-world impact across industries—from autonomous vehicles to intelligent content understanding. What you'll be doing: Develop and fine-tune multi-modal AI models using NVIDIA’s TAO Toolkit and deep learning frameworks. Contribute to the design and implementation of vision-language models (VLMs) and universal segmentation systems. Conduct experiments and benchmarking to evaluate model accuracy, robustness, and scalability. Collaborate with cross-functional teams to integrate your research into production-le
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As an Applied Scientist specializing in Small Language Models and AI Training, you will lead research and development efforts focused on building efficient, high-performance language models tailored for practical applications. You will work closely with research, engineering, and product teams to advance model training techniques, optimize architectures, and scale AI solutions. Your work will directly contribute to AI systems that are safe, interpretable, and impactful across diverse usage scenarios. What You’ll Do Lead research and development of novel training methodologies and architectures for small and efficient language models. Design, implement, and evaluate model training experiments to improve performance, robustness, and generalization of language models. Collaborate closely with research scientists and engineers on scalable training pipelines and model deployment strategies. Develop techniques for model compression, fine-tuning, and domain adaptation to optimize models for real-world applications. Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluat
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is building next-generation processors and systems, bringing together world-class expertise across silicon, systems, and software. We are looking for a CPU Performance Modeling Architect to help evaluate, shape, and optimize the performance of future CPU architectures. In this role, you’ll use performance modeling, workload analysis, and deep understanding of CPU architecture to answer complex questions about how a processor should be designed. You’ll work closely with CPU architects, RTL designers, software and compiler teams, and system engineers to identify performance opportunities, evaluate architectural tradeoffs, and turn modeling insights into actionable design decisions. This role is hybrid, based out of Santa Clara, CA or Austin, TX. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A CPU architect, performance architect, or performance modeling engineer with experience influencing CPU architecture or microarchitecture decisions. You have a strong understanding of modern processor architecture and enjoy digging into why a CPU performs the way it does. You are comfortable combining hardware architecture, software, data, and
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO As Database Support Engineer, you’ll support various critical database platforms across Development, QA, UAT, and Production environments. The role partners closely with application teams, application support, and database engineers and operates within a Follow‑the‑Sun model to ensure availability, performance, and reliability of database services. Key responsibilities include: • Provide operational support for enterprise database platforms in both on-prem private cloud and public cloud • Monitor database health, capacity, performance, and availability, and respond to alerts, diagnose issues, and perform timely remediation • Perform routine maintenance activities (patching, upgrades, housekeeping etc) • Troubleshoot database‑related incidents and collaborate on root cause analysis • Work closely with application owners, application support teams, and DB Engineers • Provide guidance on database best practices and operational standards • Participate in cross‑team problem resolution and continuous improvement initiatives • Contribute to design, implementation and testing of automation and self service capabilities of DB platforms • Drive continuous improvement, identifying opportunities to reduce toil and increase platform efficiency. • Participate in a Follow‑the‑Sun operating model, including shift‑based coverage and handoffs WHAT’S REQUIRED • Bachelor’s degr
CAPCO POLAND We offer a flexible collaboration model based on a B2B contract. At Capco Poland, we're not just another consultancy – we're the spark behind digital transformation in the financial world. As a global leader in technology and management consulting, we help clients tackle complex challenges across banking, payments, capital markets, wealth, and asset management. Engagement Overview As a GenAI Developer, you will provide services related to the design, development, and deployment of scalable AI-powered applications using Large Language Models. Collaborating with cross-functional Agile teams, you will deliver production-ready solutions integrated into enterprise environments, helping clients unlock value from Generative AI. This engagement is well suited to professionals passionate about GenAI who enjoy combining strong backend and cloud expertise with modern AI capabilities. What You’ll Do Design, build, and deploy AI applications leveraging LLMs Develop scalable solutions using GCP services (Vertex AI, BigQuery, Cloud Run / Functions) Integrate LLM APIs (e.g. OpenAI, Vertex AI) into enterprise systems Design and implement RAG architectures Apply prompt engineering techniques to optimize model performance Build and maintain REST APIs and microservices Collaborate with cross-functional teams including data, backend, and business stakeholders Deliver high-quality solutions in agile, client-facing environments What We’re Looking For 3–6 years of experience in software development Strong hands-on experience with Python Experience with Google Cloud Platform (Vertex AI, BigQuery, Cloud Run / Functions) Practical experience working with LLM APIs (OpenAI, Vertex AI, etc.) Understanding of prompt engineering and RAG architectures Experience building REST APIs and microservices Strong communication skills and ability to work in a consulting environment Nice to Have Experience with LangChain or LlamaIndex Knowledge of embeddings and vector searc
Roles & Responsibilities: Build a Creator-Led Affiliate Model • Identify and shortlist regional content creators, micro-influencers, YouTubers and Instagram creators to participate in an affiliate program where we pay the creator on every new user acquisition. • Research creators through the internet, social media platforms, YouTube, Instagram and local content pages. • Build creator lists by language, city, category, audience type and performance potential. • Reach out to creators through email, WhatsApp, Instagram DM and calls. • Coordinate with creators for briefs, timelines, content posting, tracking links and performance reports. • Maintain and update creator data in Excel or Google Sheets, including status updates and campaign reports. • Track creator-wise daily performance across views, clicks, installs, payments and cost per acquired user. • Recommend creators based on business impact, not just follower count. • Share weekly updates on creator outreach, onboarding, content posted and results delivered. What We Are Looking For: • Fluency in the assigned language. • Good understanding of regional content, creators and local culture. • An existing creator network is useful, but not mandatory. • Ability to research creators independently on the internet. • Comfort with email, WhatsApp, Excel, Google Sheets and basic reporting. • Experience is not mandatory. Freshers can also apply. • A quick learner who is self-driven and able to deliver monthly targets.
Get new model behavior engineer jobs by email
Daily job updates · Unsubscribe anytime