Can You Tell Which AI Model Made This? — elvex

You could be saving 90% of your token spend by using open-weight models.

GLM Mistral Kimi DeepSeek

But is it worth it? You tell us.

Pick a task, choose your models, and judge the outputs blind. Then see what each one actually cost.

Step 1: Pick a task

Each one was run identically across every model.

Step 2: Choose models to compare

Pick at least 2, up to 4.

Select at least 2 models to continue

Step 3: Read the outputs, then rank each tab

Use the dropdown on each tab to give it a rank. 1 is best.

Nothing here yet Pick a task and choose your models above.

Step 4: See everyone else's preferences & costs

Get the full report
Which model is best at which job
Reporting automation, data analysis, drafting, documentation, measuring ROI, integrations. These are rated per job based on real practitioner evidence.
Intelligence output per dollar spent, ranked
One model delivers 29x more output per dollar than the most expensive option.
Honest profiles, cautions included
What each model is best at and where each one falls short, with evidence.
Built on recent independent research
82 sources screened with 168 traceable claims between May - August 2026.
Everything above, in one PDF. Sent straight to your inbox.

Check your inbox.

Your report is heading to now.