Skip to main content
📖 The AI Tool Bible

AI tools tagged Editorially Verified

11 tools matching this tag.editorial

All tags →
MI

Midjourney

Featured
Image Generation · Midjourney v7
9.4

The gold standard for aesthetic AI image generation.

Paid· Basic: $20 · Pro: $50 · Enterprise: Contact salesillustrationconcept art
SC

Scite

RAG · Multi-model
8.2

AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.

Freemium· Basic: $20 · Pro: $50 · Team: $50 · Enterprise: Contact usliterature-reviewcitation-analysis
LS

LLM Stats

Evaluation · Multi-model
7.9

Live leaderboard and side-by-side comparison hub for 300+ frontier LLMs across reasoning, coding, and multimodal benchmarks.

Free· Free to browse; underlying model usage billed by each providermodel-comparisonbenchmark-tracking
EM

Emergent Mind

RAG · Undisclosed
7.2

AI-curated arXiv discovery layer that summarizes frontier papers and aggregates social discussion around them.

Freemium· Basic: $0 · Pro: $12 /month · Max: $30 /montharxiv-discoverypaper-summarization
IA

Inspect AI

Evaluation · Multi-model
7.2

Open-source LLM evaluation framework from the UK AI Security Institute with 200+ built-in benchmarks.

Free· Free and open source (MIT-style license); you pay only for underlying model API usage.llm-benchmarkingagent-evaluation
EL

Elicit

RAG · Claude Opus 4.5
7.1

AI research assistant that searches, screens, and extracts data from 138M+ academic papers at scale.

Freemium· Basic: Free · Plus: $11 · Pro: $39 · Scale: $89 · Enterprise: Customliterature-reviewsystematic-review
SL

SEAL Leaderboard

Evaluation · Multi-model (GPT, Claude, Gemini, Llama, etc.)
7.1

Private, expert-graded leaderboards from Scale AI that rank frontier LLMs on domains contaminated public benchmarks can no longer measure.

Free· Free to view; paid custom evals via Scale enterprise salesmodel-selectionbenchmark-tracking
AL

alphaXiv

RAG · Multi-model
7.0

AI reading layer over arXiv with grounded Q&A, auto-summaries, and line-by-line discussion on every preprint.

Free· Free, no signup requiredpaper-qaliterature-review
AA

Artificial Analysis

Evaluation · Multi-model
6.8

Independent benchmarking platform comparing AI models and inference providers across intelligence, speed, and cost.

Freemium· Pro: $417/month per seat · Enterprise: Custom pricingmodel-benchmarkingprovider-comparison
AW

AI World Bakeoff

Evaluation · Ten frontier coding models (specific list not published on landing page; includes at least one Claude Opus generation referenced as 'Opus 5')

Ten AI coding models, three identical briefs, thirty explorable 3D worlds

Free· Free to view. No paid tiers, sign-up, or accounts.One-shot AI coding model comparison3D generative code benchmarking
MO

ModelBias

Evaluation · 100 models across Anthropic, OpenAI, Google, DeepSeek, Meta, xAI, Mistral, Qwen and others (via OpenRouter)

100 models, 100 prompts, 30,000 answers — an interactive look at AI defaults

Free· Free to browse and download the full dataset from GitHub.Comparing default model preferences across vendorsIllustrating RLHF homogenisation in talks and articles