Repository navigation
docs: add a showcase of six real multi-agent runs - #517
Conversation
Adds a Showcase section between the built-in agents and the third-party presets. The agent sections above it carry benchmark charts; this one shows what the agents produce together. Four runs laid out two by two. Each cell holds the task graph as Raven ran it and a contact sheet of the deck that came out, so the same orchestration shape is visible four times: three research legs in parallel, a node that reduces them to one argument, one design node. Images are GitHub user attachments, matching the existing README images, so no assets land in the repository. Co-authored-by: Claude (claude-opus-5) <noreply@anthropic.com>
The caption line under each cell read as a string of parameters rather than a sentence, so each cell is now just its title and its two images. Adds a second table with two runs that do not end in a deck: the six-framework comparison board, and the parameter sweep that Raven-Code wrote, Raven-Oncall ran and Raven-Design plotted. Both have an English task graph, so the two tables stay consistent. Co-authored-by: Claude (claude-opus-5) <noreply@anthropic.com>
The comparison board was rendered from its own source in English, so the Showcase no longer mixes languages. English wraps wider than Chinese, so the board comes out taller than before, at 1.50 to 1. The sweep chart is padded to the same aspect on a white field matching its own canvas, so the two cells in that table end at the same height. The chart itself is unchanged. Co-authored-by: Claude (claude-opus-5) <noreply@anthropic.com>
gloryfromca
left a comment
There was a problem hiding this comment.
No blockers; this can merge as far as I am concerned.
I reviewed the full github/main...HEAD diff and its three-commit history, the surrounding README structure, AGENTS.md/CLAUDE.md and the context map. I also checked backward compatibility, test integrity, and architecture impact: this is an additive README-only change, touches no callers or runtime boundaries, and changes or weakens no tests.
All twelve new attachments load anonymously and visually match their titles and alt text. GitHub's Markdown renderer preserves both tables and all twelve images. git diff --check, the large-file gate, source-language gate, and commitlint all pass. The documented make wrapper could not run because make is not installed here, so I ran its exact uv-based gate commands directly. The complete GitHub suite, including docs build, repository-file checks, and unit tests, is green.
|
Not a blocker -- the section is sound and every image in it is real. One thing it leaves What I measured rather than read. All twelve asset URLs resolve to actual images -- 200, The parity. README.md and README.zh-CN.md have matched section for section -- thirteen and the diff touches one file. So a reader who arrives at the Chinese README -- which the English I am filing it as a note rather than a block because nothing is wrong with what landed and nothing The cheapest close is the same block with the six titles and the two lines of prose translated; For the record, what else I checked: the whole diff read line by line (49 added lines, one file, One thing I could not check and am not claiming either way: whether each task-graph image shows the |
## Summary Swaps the sweep chart in the Showcase for an annotated version. The chart that shipped with #517 was the run's raw matplotlib output: four flat lines and four noisy ones, with the finding left for the reader to infer. The new one states it in its own panel titles, marks the best cell, labels the recall values on the right edge, and carries a footnote saying the source is 18 rows with duplicate cells averaged. That footnote also explains the two points that are means of two runs, which the earlier chart showed without comment. The image is padded to the same aspect as the comparison board beside it, on the chart's own white canvas, so the two cells in that table still end at the same height. Hosted as a GitHub user attachment, like the rest of the section. ## Type - [ ] Fix - [ ] Feature - [x] Docs - [ ] CI / tooling - [ ] Refactor - [ ] Other ## Verification - `make check-large-files` and `make check-source-language` both pass; the change is one line of README markup. - Confirmed the new attachment returns HTTP 200 to an anonymous request and matches its source file by checksum. - Rendered the section locally at GitHub's 860px README column width to check that the two cells in the second table still align. - [ ] Relevant tests pass locally - [ ] Relevant lint / type checks pass locally - [x] User-facing docs or screenshots are updated when needed No tests or lint were run: this change touches README markup only. ## Risk - [ ] Security impact considered - [ ] Backward compatibility considered - [x] Rollback path is clear for risky changes README-only, with no runtime behaviour change. Rollback is reverting the commit; the previous image is still a live attachment, so the old chart can come back by restoring the one line. ## Related Issues Follows #517. Co-authored-by: Claude (claude-opus-5) <noreply@anthropic.com>
## Summary Adds a second row to the Showcase's non-deck table, bringing the design engine's two output modes into the section: a poster campaign and an interactive page. **Light-pollution poster campaign.** The first case whose graph takes an edge from the code side. Three Raven-Research nodes and a plate-generation node run in parallel; a Raven-Code node computes the sky-brightness panel from the research findings; a Raven-Design node composes the key visual and the street, data-panel, social and banner cuts off one master plate, with every headline and statistic set as live type rather than baked into the generated pixels. The artifact image is the key visual over the four derived formats. **Interactive GPS explainer.** Two Raven-Code nodes write a trilateration solver and an independent oracle in parallel, a Raven-Oncall node cross-checks them, and a Raven-Design node builds the page on a generated orbital plate. The artifact image is the live page above the fix at two, three and four satellites, which is what shows the geometry is computed rather than drawn. Both images are hosted as GitHub user-attachments and referenced by URL. Nothing is committed to the repository, per the repository-assets rule. ## Type - [ ] Fix - [ ] Feature - [x] Docs - [ ] CI / tooling - [ ] Refactor - [ ] Other ## Verification - `make check-large-files` -> exit 0 - `make check-source-language` -> exit 0 - All four attachment URLs fetched anonymously with curl: HTTP 200, and each one's md5 compared against its local source file to confirm the right image sits behind the right URL. - [ ] Relevant tests pass locally - [x] Relevant lint / type checks pass locally - [x] User-facing docs or screenshots are updated when needed ## Risk README-only change. The two new cells add four remote images to the page; if an attachment URL ever stops resolving the cell degrades to alt text. Rollback is reverting the commit. - [ ] Security impact considered - [x] Backward compatibility considered - [x] Rollback path is clear for risky changes ## Related Issues Follows #517 and #520. --------- Co-authored-by: Claude (claude-opus-5) <noreply@anthropic.com>
Summary
Adds a Showcase section to the README, between the built-in agents and the
third-party agent presets. The agent sections above it each carry benchmark
charts; this one shows what the agents produce together on real work.
Six runs, laid out as two tables of cells. Each cell holds two screenshots:
the task graph exactly as Raven ran it, and what the run produced. The first
table is the four runs that end in a slide deck, and all four graphs show the
same orchestration shape, three research legs in parallel, a node that reduces
them to a single argument, and one design node, so they read as one family
rather than four unrelated demos. The second table is two runs that end in
something else: a comparison board across six orchestration frameworks, and a
parameter sweep that Raven-Code wrote, Raven-Oncall ran and Raven-Design
plotted.
Everything on display is an output of the run itself, not an illustration.
The comparison board was re-rendered from its own source in English so the
section does not mix languages; the one remaining Chinese artifact is the
Song-dynasty deck, which was asked for in Chinese.
Images are hosted as GitHub user attachments, matching how the existing README
images are served, so no assets land in the repository.
Two things left out on purpose, each worth its own change: per-deck download
links, because both larger decks exceed the 25 MiB attachment limit and need
release assets; and a run whose task graph is not yet captured in English.
Type
Verification
make check-large-filesandmake check-source-languageboth pass; thechange adds no files, only README markup.
Rebased onto f10ac9f and re-ran both gates afterwards.
Rendered the section locally at GitHub's 860px README column width to check
cell alignment, image scaling and the two-column layout.
Confirmed all twelve attachments return HTTP 200 to an anonymous request,
and that each one matches its source file by checksum.
Paired every graph with its own artifact by content rather than by file
name, because the source screenshots use two different numbering schemes.
Relevant tests pass locally
Relevant lint / type checks pass locally
User-facing docs or screenshots are updated when needed
No tests or lint were run: this change touches README markup only.
Risk
README-only, with no runtime behaviour change. Rollback is reverting the
commits; the twelve images are user attachments, so nothing is left behind in
the repository.
Related Issues
N/A