All guides

Lunary vs W&B Weave

Compare Lunary and W&B Weave for agent conversations, feedback APIs, prompt versions and evaluation. See which fits a customer-facing AI product.

$20 per user / month. Payment required to activate your workspace.

Lunary conversation replay
Conversation replay in Lunary

Lunary is our pick for application teams that want to connect what a user saw, how they reacted, and what the agent actually did.

Lunary fit

Conversation replay with application feedback.

Weave fit

Agent tracing and versioned evaluation artifacts in W&B.

Check SDK parity

Confirm feedback and object operations in your language.

At a glance

Lunary vs W&B Weave: at a glance
WorkflowLunaryW&B Weave
Review interactionsChat replay with linked model and agent runsAgent sessions and turns
Capture feedback in codeJavaScript and Python APIsPython API; TypeScript unsupported in current guide
Version promptsTemplates with drafts and team reviewVersioned objects, tags and aliases
Compare qualityEvaluate models and prompt configurationsDataset evaluations and scoring functions

See your workflow in Lunary.

Explore conversations, traces and prompt versions with our team.

See Lunary in action

Review the customer experience

In Lunary, messages can link to LLM calls and agent runs, with custom events and feedback alongside the interaction. Weave also has an Agents view for sessions and turns, plus a separate Ops and Calls tracing workflow.

Check the code you need to ship

Lunary provides JavaScript and Python feedback examples for frontend or backend capture. Weave’s feedback guide documents its UI and Python SDK, with TypeScript support marked unavailable. That distinction matters if feedback is part of your TypeScript application.

Make prompt changes reviewable

Lunary combines versioned templates with draft sharing and a playground. Weave versions prompts, datasets and model configurations as objects. Trial the edit, review and release process your team will repeat each week.

Questions & answers

Is Weave only for ML experiments?

No. Weave provides agent observability, LLM tracing, evaluations and feedback.

Which should a TypeScript team try first?

We recommend Lunary if collecting in-app feedback is essential. Validate the exact SDK operations your application needs on both platforms.

What would a useful Lunary pilot prove?

That your team can locate a poor reply, inspect its trace, review a prompt revision and assess the change on representative examples.

By Lunary · Documentation reviewed Sep 19, 2026

Sources & methodology (9)

Based on vendor documentation, not an independent benchmark or hands-on rating. Features and plans can change.

Bring your next AI issue to Lunary.

Follow the conversation. Find the cause. Test the fix.

$20 per user / month. Payment required to activate your workspace.

Compare plans and pricing