Skip to main content
📖 The AI Tool Bible

AI tools tagged For Researchers

48 tools matching this tag.audience

All tags →
AU

AudioCraft

Audio · MusicGen, AudioGen, EnCodec
8.2

Meta's open-source research toolkit for generating music and sound effects from text via a single autoregressive language model.

Free· Free and open source; self-hostedtext-to-musicsound-effects
LI

LiveBench

Evaluation · Multi-model
8.2

Contamination-free LLM benchmark that refreshes its questions monthly to keep frontier models honest.

Free· Free and open source; self-hosted evaluation runnerllm-benchmarkingmodel-selection
SC

Scite

RAG · Multi-model
8.2

AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.

Freemium· Basic: $20 · Pro: $50 · Team: $50 · Enterprise: Contact usliterature-reviewcitation-analysis
UN

Unsloth

Fine-tuning · Llama, Mistral, Gemma, Qwen, GLM (multi-model)
8.2

Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.

Freemium· Free open-source; Pro and Enterprise contact saleslora-finetuningqlora
BF

Berkeley Function-Calling Leaderboard

Evaluation · Multi-model
8.1

Open benchmark from UC Berkeley that ranks LLMs on real-world tool-use and function-calling accuracy.

Free· Free and open source; you pay only for inference when reproducing runs.function-calling evaltool-use benchmarking
NO

NotebookLM

RAG · Gemini 2.5
8.1

Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.

Freemium· Free tier; Plus via Google One AI Premium ($19.99/mo) or Workspace add-ondocument Q&Aresearch synthesis
OE

OpenAI Evals

Evaluation · OpenAI GPT models (extensible)
8.1

OpenAI's open-source framework for benchmarking LLMs against a shared registry of evaluations.

Free· Free (MIT); you pay OpenAI API costs for eval runsllm-benchmarkingregression-testing
CA

CAMEL-AI

Agents · Multi-model
8.0

Open-source Python framework for building multi-agent systems and synthetic data pipelines.

Free· Free, open-source; pay for the underlying LLM API callsmulti-agent-systemssynthetic-data-generation
RE

Reflect

Writing · OpenAI GPT-4 + Whisper
8.0

Networked note-taking app with GPT-4 and Whisper baked into the writing surface.

Paid· Free: Free · Pro: $20 · Business: $35note-takingvoice transcription
MA

Manus

Agents · Multi-model
7.9

Generalist agent for research, code, and web tasks.

Paid· Credit-based; tiers from $19/moresearchweb tasks
HA

Humata.ai

RAG · Multi-model
7.8

Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.

Freemium· Free: $0 · Expert: $9.99 · Team: $49 / user · Enterprise: customdocument-qaresearch-summarization
ST

STORM

Writing · Multi-model (via LiteLLM)
7.5

Stanford's open-source research agent that turns a topic into a Wikipedia-style article with citations.

Free· Hosted demo free; self-host open-source (pay your own LLM/search API)long-form researchwikipedia-style articles
MA

MathEval

Evaluation · GPT-4 grader / DeepSeek-LLM-7B verifier
7.3

Holistic benchmark suite for evaluating mathematical reasoning in large language models.

Free· Free; open-source benchmark with leaderboard submissions via matheval.aillm-math-benchmarkingmodel-leaderboards
MM

MMagic

Image Generation · Multi-model (Stable Diffusion, ControlNet, StyleGAN, GANs, diffusion)
7.3

OpenMMLab's research-grade toolbox for image and video generation, restoration, and editing.

Free· Free and open source (Apache 2.0)text-to-imagesuper-resolution
RA

Rayyan

RAG
7.3

AI-assisted systematic review platform for screening, deduplicating, and extracting data from large literature corpora.

Freemium· Free forever for basic use; enterprise custom pricingsystematic-reviewliterature-screening
WM

World Monitor

Agents · Multi-model (Claude, GPT via MCP)
7.3

Real-time global intelligence dashboard fusing 56 map layers, 500+ news feeds, and an AI analyst into one situational-awareness console.

Freemium· Free: $0 · Pro: $39.99 · Pro Business: $49.99 · API Starter: $99.99 · API Business: $299.99geopolitical-monitoringsupply-chain-risk
DE

DeerFlow

Agents · Multi-model (Doubao, DeepSeek, OpenAI, Gemini)
7.2

Open-source multi-agent framework from ByteDance for long-running research, coding, and content tasks.

Free· Free, MIT-licensed (self-hosted; you pay only for LLM tokens)deep-researchautonomous-coding
EM

Emergent Mind

RAG · Undisclosed
7.2

AI-curated arXiv discovery layer that summarizes frontier papers and aggregates social discussion around them.

Freemium· Basic: $0 · Pro: $12 /month · Max: $30 /montharxiv-discoverypaper-summarization
IA

Inspect AI

Evaluation · Multi-model
7.2

Open-source LLM evaluation framework from the UK AI Security Institute with 200+ built-in benchmarks.

Free· Free and open source (MIT-style license); you pay only for underlying model API usage.llm-benchmarkingagent-evaluation
LF

LLaMA Factory

Fine-tuning · Multi-model (LLaMA, Mistral, Qwen, Gemma, Phi, LLaVA, ChatGLM, Yi)
7.2

Open-source, no-code WebUI for fine-tuning 100+ open LLMs with LoRA, QLoRA, DPO, and PPO.

Free· Free, open-source (Apache-2.0); self-hostedlora-fine-tuningqlora
LL

LLMEval

Evaluation · Multi-model
7.2

Open academic benchmark suite for stress-testing LLMs on contamination-resistant, domain-specific tasks.

Free· Free; open-source academic benchmarksllm-benchmarkingacademic-evaluation
NO

NoteGen

Writing · Multi-model (Qwen3-8B, BAAI/bge-m3, GLM-4.1V-9B-Thinking, BYO keys)
7.2

Open-source local-first note-taking app that uses AI to turn raw clippings into structured Markdown.

Free· Free and open source (GPL-3.0); bring your own model keys if you want paid LLMsnote-takingknowledge-management
ON

OneKE

RAG · Multi-model (OneKE-13B, LLaMA3, Qwen2.5, GPT, DeepSeek-R1)
7.2

Open-source multi-agent framework for schema-guided knowledge extraction from documents.

Free· Free, MIT-licensed; you pay for LLM API calls or self-hosted computeknowledge-graph-constructionnamed-entity-recognition
OD

Open Deep Research

Agents · o3-mini (default), DeepSeek R1, or any OpenAI-compatible model
7.2

Minimal open-source deep-research agent that iteratively searches, scrapes, and reasons to produce cited markdown reports.

Free· Free (MIT); bring your own Firecrawl + LLM API keysdeep-researchagent-scaffolding
WA

Weco AI

Evaluation · Multi-model (LLM + AIDE tree search)
7.2

Autoresearch engine that iteratively rewrites code to optimize against a numeric evaluation metric.

Freemium· Open-source CLI; hosted/commercial pricing not publishedcode-optimizationgpu-kernel-tuning
AL

AlpacaEval

Evaluation · GPT-4 Preview (Nov 2024) as annotator
7.1

Automatic LLM evaluator and leaderboard that benchmarks instruction-following with length-controlled win rates.

Free· Free and open-source; pay only for the underlying OpenAI annotator API callsllm-benchmarkinginstruction-following eval
EL

Elicit

RAG · Claude Opus 4.5
7.1

AI research assistant that searches, screens, and extracts data from 138M+ academic papers at scale.

Freemium· Basic: Free · Plus: $11 · Pro: $39 · Scale: $89 · Enterprise: Customliterature-reviewsystematic-review
FP

FutureHouse Platform

RAG · Multi-model
7.1

Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.

Freemium· Basic: $10 · Pro: $25 · Enterprise: Contact salesscientific-literature-searchautonomous-research-agent
JA

Jenni AI

Writing
7.1

AI writing workspace built specifically for academic papers, theses, and cited research.

Freemium· Free; Plus $12/mo; Pro $29/mo (up to 60% off annual)academic-writingliterature-review
LA

LangExtract

RAG · Multi-model (Gemini, GPT-4/4o, Ollama-hosted local models)
7.1

Google's open-source Python library for LLM-driven structured extraction from unstructured text, with source-grounded outputs.

Free· Library is free (Apache-2.0); LLM API costs depend on chosen backendstructured-extractiondocument-parsing
RA

RabbitHoles AI

Agents · Multi-model (Anthropic, Google, Groq, Perplexity, custom)
7.1

Infinite-canvas AI chat client that lets you branch, fork, and connect conversations as nodes instead of scrolling a single linear thread.

Paid· Starter: $5/monthdeep researchbranching chat
RU

Runcell

Coding · Multi-model (GPT, Claude, Gemini)
7.1

Jupyter-native AI agent built for multi-week ML and data science projects.

Freemium· Hobby: $0 · Pro: $20/month · Pro+: $60/month · Ultra: $200/month · Teams: $40/seat/monthjupyter-notebooksdata-science
SL

SEAL Leaderboard

Evaluation · Multi-model (GPT, Claude, Gemini, Llama, etc.)
7.1

Private, expert-graded leaderboards from Scale AI that rank frontier LLMs on domains contaminated public benchmarks can no longer measure.

Free· Free to view; paid custom evals via Scale enterprise salesmodel-selectionbenchmark-tracking
VI

VisualWebArena

Evaluation · Model-agnostic (GPT-4V, Gemini, Claude, open VLMs)
7.1

Open benchmark for evaluating multimodal web agents on realistic visual browsing tasks.

Free· Free and open source (MIT-style research release)multimodal-agent-evalweb-browsing-benchmark
AL

alphaXiv

RAG · Multi-model
7.0

AI reading layer over arXiv with grounded Q&A, auto-summaries, and line-by-line discussion on every preprint.

Free· Free, no signup requiredpaper-qaliterature-review
GR

Graphify

RAG · Multi-model
7.0

Open-source on-device knowledge graph engine that turns code, docs, papers, meetings and images into a queryable graph.

Free· MIT-licensed, free forever; cloud tier hinted but unpriced (waitlist)knowledge-graphcode-search
OL

OlympicArena

Evaluation · GPT-4o, Claude-3.5-Sonnet, Doubao-Pro-32k, DeepSeek-Coder-V2, Qwen2-72B-Instruct
7.0

Olympiad-level multi-discipline benchmark for stress-testing reasoning in LLMs and multimodal models.

Free· Free, open-source research benchmarkllm-evaluationmultimodal-eval
PA

PageIndex

RAG
7.0

Vectorless reasoning-based retrieval for long documents, with traceable, auditable answers.

Freemium· Free Try Now tier; enterprise pricing on requestdocument-qalong-pdf-retrieval
SM

SMMRY

Writing
7.0

Veteran web-based summarizer that condenses articles, books, and YouTube videos into digestible bullet points.

Freemium· Free (10 summaries); Essential $7/mo yearly; Advanced $13/mo yearlyarticle-summarizationresearch-reading
CO

CompassRank

Evaluation · Multi-model
6.9

Public leaderboard from the OpenCompass project ranking open and closed LLMs across 100+ benchmarks.

Free· Free leaderboard; OpenCompass toolkit is Apache 2.0 open sourcellm-benchmarkingmodel-selection
IN

InfiBench

Evaluation
6.9

Stack Overflow-derived benchmark for evaluating code LLMs on real-world programming questions.

Free· Free and open source (CC BY-SA 4.0)code-llm-evalmodel-benchmarking
LY

LynxKite

Agents · Multi-model (LLM agents + GNNs + NVIDIA BioNeMo)
6.9

No-code AI orchestration platform built for graph-native pipelines in drug discovery and enterprise analytics.

Enterprise· Contact sales; no public pricingdrug-discoverygraph-neural-networks
MI

MixEval

Evaluation · GPT-3.5-Turbo-0125, GPT-4o-2024-05-13, Claude 3.5 Sonnet, MixEval, MixEval-Hard
6.9

Dynamic LLM benchmark that mixes web queries with existing datasets to mirror Chatbot Arena rankings at a fraction of the cost.

Free· Free and open sourcellm-benchmarkingmodel-ranking
SC

SciSpace

RAG · Multi-model
6.9

AI research assistant that turns dense PDFs and literature reviews into searchable, citation-backed answers.

Freemium· Premium: $20 · Advanced: $90 · Max: $200literature-reviewchat-with-pdf
SH

ShyEditor

Writing · Multi-model (GPT-4, Grok)
6.9

AI-native writing environment that bolts an assistant, citation manager, and knowledge base onto a distraction-free markdown editor.

Freemium· Free Basic (20 AI credits/mo, 100MB); Pro $10/mo (500 credits, 5GB)long-form-writingacademic-writing
YA

Yomu AI

Writing · Multi-model (GPT-4o/5, Claude Sonnet/Opus, Gemini Flash)
6.9

AI writing assistant built specifically for students and academic researchers.

Freemium· Free tools; Pro $11/mo, Ultra $18/mo (annual); $499 lifetimeacademic-writingessay-drafting
YS

YouTube Summary & ChatGPT by Glasp

Agents · Multi-model (GPT-3.5, GPT-4, Claude, Mistral, Gemini)
6.9

Chrome extension that surfaces YouTube transcripts and AI-generated video summaries alongside one-click access to ChatGPT.

Freemium· Basic: $20 · Pro: $50 · Enterprise: Contact salesyoutube-summarizationtranscript-extraction
AI

aiPDF

RAG
6.8

Chat-with-your-documents app that ingests PDFs, EPUBs, web pages and YouTube videos with cited answers.

Freemium· Playful: Free · Dynamic: $9/month · Flagship: $19/monthdocument-chatpdf-summarization