Local Model Bench
  • Home
  • Cases
  • Methodology
  • About
Sign in Subscribe

benchmark note

A collection of 2 posts
Bigger was not always better in the paperwork benchmark
benchmark note

Bigger was not always better in the paperwork benchmark

In the current nine-case paperwork suite, two larger-sounding local model choices produced fewer finished cases than leaner alternatives. That does not prove small models are better. It shows that parameter count and model labels a
27 May 2026 3 min read
Gemini 3.5 Flash can draw SVG. The runner was the problem.
benchmark note

Gemini 3.5 Flash can draw SVG. The runner was the problem.

A first Gemini 3.5 Flash SVG run looked like a model failure. It was not. The OpenAI-compatible API path let hidden thinking consume the output budget, so the model returned planning fragments instead of a complete SVG. The native Gemini API with thinkingBudget: 0 produced a valid City Plan SVG and
21 May 2026 3 min read
Page 1 of 1
Local Model Bench © 2026
  • Sign up
Powered by Ghost