Public benchmark snapshot: 2026-07-19
Best AI For Emails: research draft
Research-draft status
This page has not completed full editorial review, so it is excluded from search indexing. It becomes eligible only after verifiable public benchmarks or first-party testing are available.
Planned scope
The planned scope for “best ai for emails” is ChatGPT, Claude, Gemini.
What public benchmarks show
Below are results only from public benchmark categories where every model on this page appears together. Scores from different sources cannot be added up, and they do not prove one model is best overall.
Arena · coding preference
Arena asks people to pick the better answer without knowing which model wrote it. Higher Elo means more people preferred that model in that category.
| Product | Exact model version | Rank | Score | Metric |
|---|---|---|---|---|
| Claude | claude-fable-5 | 2 | 1631 | Elo (Elo) |
| ChatGPT | gpt-5.6-sol-xhigh (codex-harness) | 3 | 1618 | Elo (Elo) |
| Gemini | gemini-3.5-flash | 19 | 1504 | Elo (Elo) |
Arena · search-style preference
Arena asks people to pick the better answer without knowing which model wrote it. Higher Elo means more people preferred that model in that category.
| Product | Exact model version | Rank | Score | Metric |
|---|---|---|---|---|
| Claude | claude-opus-4-6-search | 1 | 1255 | Elo (Elo) |
| ChatGPT | gpt-5.5-search | 2 | 1239 | Elo (Elo) |
| Gemini | gemini-3.1-pro-grounding | 7 | 1211 | Elo (Elo) |
Arena · general chat preference
Arena asks people to pick the better answer without knowing which model wrote it. Higher Elo means more people preferred that model in that category.
| Product | Exact model version | Rank | Score | Metric |
|---|---|---|---|---|
| Claude | claude-fable-5 | 1 | 1507 | Elo (Elo) |
| Gemini | gemini-3-pro | 8 | 1486 | Elo (Elo) |
| ChatGPT | gpt-5.6-sol-xhigh | 10 | 1486 | Elo (Elo) |
SWE-bench Verified · real software-issue fixing
SWE-bench Verified measures how often an AI coding setup can fix real GitHub issues. A higher resolved percentage means more issues were fixed in that test.
| Product | Exact model version | Rank | Score | Metric |
|---|---|---|---|---|
| Claude | live-SWE-agent + Claude 4.5 Opus medium (20251101) | 1 | 79.2% | Resolved (%) |
| Gemini | live-SWE-agent + Gemini 3 Pro Preview (2025-11-18) | 4 | 77.4% | Resolved (%) |
| ChatGPT | JoyCode + Claude 4 Sonnet + GPT-4.1 | 17 | 74.6% | Resolved (%) |
Official sources and verification date
The links below are official sources for product identity and plan information, last checked 2026-07-19. Pricing, model availability, quotas, and regional access can change; confirm on the provider site before purchasing or deploying.
- OpenAI ChatGPT plans and pricing - 2026-07-19
- Anthropic plans and pricing - 2026-07-19
- Google Gemini plans - 2026-07-19
Frequently asked questions
Does this page publish an absolute ranking?
No. Conditional findings are published only when test conditions, raw outputs, and a review method are disclosed.
Are pricing and plan details definitive?
This page links to official sources and verification dates. Confirm the provider price shown for your region before purchasing.
Does ModelAny store prompts?
ModelAny is local-first. Prompts are sent only to the AI services you choose and are not uploaded to ModelAny servers.
Compare multiple models with one prompt
ModelAny is available on the Chrome Web Store. The Microsoft Edge Add-ons listing is still under review. Drafts, settings, and history remain in your browser.
Install extension