Benchmark
An independent benchmark for image and video generation models. Public tasks, blind judging, and a match journal anyone can re-run. Genfeed competes; it does not score itself.
AnnouncedRead from the public bench repository. Nothing on this page is maintained by hand.
6Tasks
6Contestants
0Matches
Season One — Image
Elo from recorded pairwise matches only. A voided match is counted and shown but never moves a rating.Standings.
No matches yet
The ladder is empty on purpose.
Season One — Image is announced with 6 contestants entered and no match recorded. Ranking models that have never met would be inventing a result, so there is nothing here until the first verdict lands in the journal.The suite
Chosen for where models differ in work people get paid for, not where they demo well. Every prompt is public — there is no hidden set.What we ask models to make.
Brand kit fidelity
imageCharacter consistency
imageProduct consistency
imagePrompt adherence
imageLegible text
imageUnbranded craft
imageSeason Two
Published now so the tasks are fixed in public before anyone knows which model they will favour. The harness refuses to draw a draft into a match.Video, not yet running.
Camera instruction
videoSeason Two draftFirst-frame fidelity
videoSeason Two draftMotion coherence
videoSeason Two draftGenfeed is entered as a contestant, not as the scoreboard.
One route runs through Genfeed's brief compiler; the rest are raw provider routes, including the same model the compiled route wraps. If the compile step makes a task worse, the ladder says so.Genfeed compile → Seedream 5 ProGenfeedSeedream 5 ProNano Banana 2FLUX.2 ProGPT Image 2Recraft V4 Pro