Weights & Biases

Product newsletter: Updates and new features for June 2026


AI agent observability

Evaluations

Evaluations

Agents

Agents

Inference

Inference

Safety and scoring

Safety and scoring

Build models

Computer Vision

Computer Vision

NLP

NLP

MLOps

MLOps

LLM

LLM

About Weights & Biases

Integrations

Integrations

W&B releases

W&B releases

Events

Events

Podcast

Podcast

W&B Community Posts Agents

Track, guard, and remediate AI usage with W&B Weave and CrowdStrike Falcon AIDR

Learn how Weave and CrowdStrike Falcon AI Detection and Response can help ensure internal AI systems are handling sensitive data appropriately Read more

Tabular ML didn't die. It was just waiting for agents

Learn how our tabular data agent took #1 on the MLE-bench tabular slice with two golds, three above-median, four-for-four valid submissions.
Read more

Introducing CoreWeave ARIA: AI Research and Iteration Agent

Your experiments are already tracked. Now let the agent read them, analyze them, and help you turn every experiment into continuous improvement.
Read more

We'll be at the AI Engineer World's Fair - Come Say Hi

Join Weights & Biases by CoreWeave at AI Engineer World’s Fair in SF. Visit booth UG 24 for demos, swag, speaker sessions, and new Weave agent tracing features. Read more

Building a production-grade MLOps pipeline for VLA models

This report covers the development of a compact vision-language-action model built around Gemma 3 270M.
Read more

Getting started with embodied chain-of-thought (ECoT)

Learn how to get started with Embodied Chain-of-Thought, a robotics reasoning approach that helps vision-language-action models plan, explain, and execute grounded actions. Read more

Governance workflows for AI agents: an evidence-backed review gate on W&B Weave

A technical walkthrough of an open-source AI governance toolkit that turns scattered evals, red-team runs, and approvals into one reproducible review gate, with every finding traced in W&B Weave. Read more

Tutorial: Benchmarking Claude Opus 4.8 honesty with BeHonest and W&B Weave

How honest is Claude Opus 4.8 compared to 4.7? We ran both models through the BeHonest benchmark and logged every result into W&B Weave. Read more

New in W&B Weave: Observability and continuous improvement for production agents

We're shipping a new set of capabilities that make Weave the best tool to raise and maintain agent quality in production. Here's what you need to know. Read more

Product newsletter: Updates and new features for May 2026

From a brand new CoreWeave Sandboxes offering to updates in Models and to our iOS app, here's what we released in May. Read more

Building a text-to-SQL agent

In the second edition of our AI Builders series, learn how to build a text-to-SQL agent Read more

Product newsletter: Updates and new features for April 2026

From universally available automations to brand new Weave dashboards, here are the big features we released in April Read more

Training a Gemma-3 powered VLA model

This report covers the development of a compact vision-language-action model built around Gemma 3 270M. Read more

From sequence to structure: Tracking protein fold prediction with W&B

Learn how to track protein fold prediction using ESMFold and Weights & Biases. Explore a complete biotech pipeline, from 3D structure visualization and pLDDT tracking to Bayesian sweeps and model versioning. Read more

Building an AI agent for interior design

In the inaugural edition of our AI Builders series, learn how to build an AI agent for home decor Read more

Tutorial: Building a production-ready fraud triage Copilot

Build, evaluate, guard, and monitor an LLM fraud analyst end-to-end with W&B Weave. Read more