Ghost Palette

local/sefi-image-5b-base — ImageBench V1

ImageBench V1 —
192 evaluations across 6 categories
Benchmark V1 verdicts are produced by VLM judges and can contain mistakes. Treat PASS/FAIL labels as machine-assisted assessments, and inspect the images yourself. Learn more about the methodology.

Generation Details

Source-backed model context, size, cost, and request settings for this run.

MakerSefi
FamilySefi Image
Model size5B5B params
Costself-hostedfree
Latency131s median
Run targetlocal/sefi-image-5b-base
Effective requestmodel: sefi-image-5b-base · size: 1024x1024 · seed: 42
SourcesModel cardWeights & checkpoint
Text47%Spatial60%Human59%Pro Studio60%Graphical76%Truth51%Preference40%Latency0%

All 192 generations

Every challenge, grouped by category and subcategory. Solid-bordered tiles with a check passed the VLM judge; dashed tiles with a cross failed. Images are placeholders until generated.

Text Rendering

47%
Typography Style67%
Writing accuracy50%

Spatial Reasoning

60%
Attributes Binding33%
Compositionality44%
Counting56%
Negation44%
Relative Position50%
Scale & Proportions44%

Human realism

59%
Faces & Expressions50%
Full Body42%
Hands42%
Multi-Subject50%

Professional Studio

60%
Camera & Lighting58%
Color Precision42%
Photorealism67%

Graphical design

76%
Data Visualisation67%
Layout & Design67%
Style Diversity58%

Truthfulness

51%
Photorealism33%
Physics & Reflections42%
World Knowledge58%