Skip to content

evidence(bakeoffs): Verdana, Tahoma, Trebuchet MS vs the open corpus - #25

Merged
caio-pizzol merged 1 commit into
mainfrom
caio-pizzol/bakeoff-sans
Jun 5, 2026
Merged

evidence(bakeoffs): Verdana, Tahoma, Trebuchet MS vs the open corpus#25
caio-pizzol merged 1 commit into
mainfrom
caio-pizzol/bakeoff-sans

Conversation

@caio-pizzol

@caio-pizzol caio-pizzol commented Jun 5, 2026

Copy link
Copy Markdown
Contributor

Runs the bakeoff (#24) for three high-value unresolved Word/UI sans fonts against the full Google discovery corpus, to test whether the broad corpus changes our current "no good substitute" answer.

It does not - the answer holds. Every top candidate for all three is visual-tier: a low-ish mean but a max of 32-85%, so at least one glyph always reflows. No direct/likely candidate exists, matching the existing visual_only verdicts. Nothing is promoted; records.json is untouched.

What the broad corpus surfaced, vs the registry's CURRENT committed top_candidate:

Target Bakeoff closest by mean Registry's current top_candidate
Tahoma Viga (0.58% mean / 32% max) Viga - same font, confirmed
Trebuchet MS Ruda (1.06% / 59%) Fira Sans - Ruda is a new, closer-by-mean lead
Verdana LINE Seed JP (1.79% / 85%) Tsukimi Rounded - LINE Seed JP is a new, closer-by-mean lead

So the bakeoff confirms Tahoma's existing pick (Viga) and surfaces two new leads (Ruda, LINE Seed JP) the registry doesn't currently name. These are leads for a reviewer, not promotions.

Note on the numbers: the bakeoff uses docfonts' own prose-weighted mean + strict per-glyph LATIN_CORE max, which differs from the apryse summary the committed top_candidates were measured with (e.g. the apryse Viga figure is 1.53% / 7.15%). So compare which font is closest, not the raw deltas across methods - and the huge max deltas are exactly why none of these is a metric candidate.

Data-only PR: three committed BakeoffResult JSONs (each with targetFace), validated by the existing bakeoff tests (face-scoped, ranked, public-safe, no verdict field).

Verified locally: bun test -> 114 pass / 0 fail; tsc clean; biome clean.

Three high-value unresolved Word/UI sans fonts run against the full discovery
corpus. The answer holds: no metric clone exists - every top candidate is
visual-tier (best mean 0.6-1.8%, but max 32-85%, so one glyph always reflows).

What the broad corpus DID surface: closer-by-mean visual leads than the
registry's current top_candidates - e.g. Tahoma -> Viga (0.58% mean) vs the
prior Open Sans (4.85%), Trebuchet MS -> Ruda (1.06%). Useful leads for a
reviewer, but still not metric-safe; nothing is promoted - records.json is
untouched.
@caio-pizzol
caio-pizzol merged commit a945049 into main Jun 5, 2026
1 check passed
@caio-pizzol
caio-pizzol deleted the caio-pizzol/bakeoff-sans branch June 5, 2026 12:33
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants