RSS Amplifier

Machine Learning · Aug 26, 2026

A dataset with 52 Text to image model evaluation [P]

0
Sign in to vote or save

This site does not allow itself to be embedded. You can still read it on the original site — the toolbar below keeps your place in the directory.

I created a simple text to image benchmark. I curated 192 prompts that are difficult for T2I models in various ways: text rendering, spatial reasoning, human realism, negations, etc... I then asked a VLM to judge every output against a pre-specified binary question with the ground truth baked in. I'm publishing all the results including the images. (Most public T2I leaderboards don't publish the…

Read on reddit.com

Comments

Nothing yet. Say the first thing.

    Sign in to join the conversation.