AI tools tagged For Researchers
48 tools matching this tag.audience
AudioCraft
Meta's open-source research toolkit for generating music and sound effects from text via a single autoregressive language model.
LiveBench
Contamination-free LLM benchmark that refreshes its questions monthly to keep frontier models honest.
Scite
AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.
Unsloth
Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.
Berkeley Function-Calling Leaderboard
Open benchmark from UC Berkeley that ranks LLMs on real-world tool-use and function-calling accuracy.
NotebookLM
Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.
OpenAI Evals
OpenAI's open-source framework for benchmarking LLMs against a shared registry of evaluations.
CAMEL-AI
Open-source Python framework for building multi-agent systems and synthetic data pipelines.
Reflect
Networked note-taking app with GPT-4 and Whisper baked into the writing surface.
Manus
Generalist agent for research, code, and web tasks.
Humata.ai
Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.
STORM
Stanford's open-source research agent that turns a topic into a Wikipedia-style article with citations.
MathEval
Holistic benchmark suite for evaluating mathematical reasoning in large language models.
MMagic
OpenMMLab's research-grade toolbox for image and video generation, restoration, and editing.
Rayyan
AI-assisted systematic review platform for screening, deduplicating, and extracting data from large literature corpora.
World Monitor
Real-time global intelligence dashboard fusing 56 map layers, 500+ news feeds, and an AI analyst into one situational-awareness console.
DeerFlow
Open-source multi-agent framework from ByteDance for long-running research, coding, and content tasks.
Emergent Mind
AI-curated arXiv discovery layer that summarizes frontier papers and aggregates social discussion around them.
Inspect AI
Open-source LLM evaluation framework from the UK AI Security Institute with 200+ built-in benchmarks.
LLaMA Factory
Open-source, no-code WebUI for fine-tuning 100+ open LLMs with LoRA, QLoRA, DPO, and PPO.
LLMEval
Open academic benchmark suite for stress-testing LLMs on contamination-resistant, domain-specific tasks.
NoteGen
Open-source local-first note-taking app that uses AI to turn raw clippings into structured Markdown.
OneKE
Open-source multi-agent framework for schema-guided knowledge extraction from documents.
Open Deep Research
Minimal open-source deep-research agent that iteratively searches, scrapes, and reasons to produce cited markdown reports.
Weco AI
Autoresearch engine that iteratively rewrites code to optimize against a numeric evaluation metric.
AlpacaEval
Automatic LLM evaluator and leaderboard that benchmarks instruction-following with length-controlled win rates.
Elicit
AI research assistant that searches, screens, and extracts data from 138M+ academic papers at scale.
FutureHouse Platform
Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.
Jenni AI
AI writing workspace built specifically for academic papers, theses, and cited research.
LangExtract
Google's open-source Python library for LLM-driven structured extraction from unstructured text, with source-grounded outputs.
RabbitHoles AI
Infinite-canvas AI chat client that lets you branch, fork, and connect conversations as nodes instead of scrolling a single linear thread.
Runcell
Jupyter-native AI agent built for multi-week ML and data science projects.
SEAL Leaderboard
Private, expert-graded leaderboards from Scale AI that rank frontier LLMs on domains contaminated public benchmarks can no longer measure.
VisualWebArena
Open benchmark for evaluating multimodal web agents on realistic visual browsing tasks.
alphaXiv
AI reading layer over arXiv with grounded Q&A, auto-summaries, and line-by-line discussion on every preprint.
Graphify
Open-source on-device knowledge graph engine that turns code, docs, papers, meetings and images into a queryable graph.
OlympicArena
Olympiad-level multi-discipline benchmark for stress-testing reasoning in LLMs and multimodal models.
PageIndex
Vectorless reasoning-based retrieval for long documents, with traceable, auditable answers.
SMMRY
Veteran web-based summarizer that condenses articles, books, and YouTube videos into digestible bullet points.
CompassRank
Public leaderboard from the OpenCompass project ranking open and closed LLMs across 100+ benchmarks.
InfiBench
Stack Overflow-derived benchmark for evaluating code LLMs on real-world programming questions.
LynxKite
No-code AI orchestration platform built for graph-native pipelines in drug discovery and enterprise analytics.
MixEval
Dynamic LLM benchmark that mixes web queries with existing datasets to mirror Chatbot Arena rankings at a fraction of the cost.
SciSpace
AI research assistant that turns dense PDFs and literature reviews into searchable, citation-backed answers.
ShyEditor
AI-native writing environment that bolts an assistant, citation manager, and knowledge base onto a distraction-free markdown editor.
Yomu AI
AI writing assistant built specifically for students and academic researchers.
YouTube Summary & ChatGPT by Glasp
Chrome extension that surfaces YouTube transcripts and AI-generated video summaries alongside one-click access to ChatGPT.
aiPDF
Chat-with-your-documents app that ingests PDFs, EPUBs, web pages and YouTube videos with cited answers.