AgentDish directory
Research
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#33
↓ -2
Primus AI Researcher – Free
Primus is an autonomous AI researcher that hypothesizes, reads papers, writes code, runs experiments on compute, and drafts research papers. The page shows example tasks, published-paper claims, waitlist access, and positioning for ML research workflows. |
Research / Knowledge Work | 90 | ↓ -2 | 21 days ago | Details |
|
#53
↓ -2
OpenScience
OpenScience is an open-source AI workbench for scientific research. It can read papers, plan hypotheses, write and run code, manage experiments, query scientific databases, and present results in a browser-based workspace. |
Developer Tools / AI Workbench / Research Agent | 90 | ↓ -2 | 57 days ago | Details |
|
An interactive dashboard that analyzes New York Times coverage since 2000 using the NYT Archive API, with views for reporters, beats, sections, subjects, geography, obituaries, and corrections. |
Research / Data Visualization | 89 | ↑ +454 | 118 days ago | Details |
|
#205
↓ -3
Doubt
Doubt is an open-source, zero-dependency CLI and Agent Skill for turning contested questions into source-grounded evidence maps with supports, contradictions, unknowns, and portable HTML output. |
Developer Tools / AI Development Tools | 88 | ↓ -3 | 30 days ago | Details |
|
#272
↓ -3
Slashspace
Slashspace is a local-first infinite canvas for AI chat and agentic work, designed for research, writing, development, and long-form thinking. It supports multiple models, MCP connections, local storage, and desktop apps for Mac, Windows, and Linux. |
Productivity / AI Workspace | 88 | ↓ -3 | 74 days ago | Details |
|
#302
↓ -3
CAD-Bench
An open benchmark and leaderboard for AI CAD agents, with 308 prompts across 20 categories and layered scoring for geometry, engineering, manufacturability, and cognition. |
Research / Knowledge Work | 88 | ↓ -3 | 114 days ago | Details |
|
#337
↓ -4
FutureSearch
AI forecasting platform with a public track record, API/docs, and integrations for Claude Desktop and Claude Code. The page highlights benchmark results, live market usage, and example forecast types for probability, numeric, date, categorical, and decision questions. |
Developer Tools / AI Forecasting | 87 | ↓ -4 | 27 days ago | Details |
|
A research article from Applied Compute on how agentic, tool-using workloads differ from traditional LLM benchmarks, with production observations, workload profiles, and an open-source harness for replaying traces. |
Research / Knowledge Work | 87 | ↓ -107 | 117 days ago | Details |
|
#477
↑ +2
ThoughtDAG
An infinite canvas that turns LLM conversations, readings, and notes into an editable thought graph. It supports human-in-the-loop workflows, cited passages, replayable dependencies, and local-first backups. |
Writing / Copywriting | 86 | ↑ +2 | 36 days ago | Details |
|
#515
↑ +2
UnderstandDocs
UnderstandDocs is a document analysis tool that summarizes pasted text or uploaded files, flags risks, extracts important dates, and simplifies dense language. The page shows a working analysis form, supported file types, a privacy/no-storage claim, and an example output for a tenancy agreement. |
Research / Knowledge Work | 86 | ↑ +2 | 54 days ago | Details |
|
#530
↑ +2
ZUSE Automat Agent
A deterministic Python project for empirical law discovery in elementary cellular automata, with simulation, discovery, reproducibility guides, and published preprints. |
Research / Scientific Discovery | 86 | ↑ +2 | 63 days ago | Details |
|
#626
↓ -3
LittleLearner
A research site for LittleLearner, a language model trained only on K–5 curriculum data to study how bounded pretraining affects capability, scaling, post-training, and in-context learning. The page includes a live chat demo, model scales, controls, and experiment summaries. |
AI Research Tool / Controlled model sandbox | 85 | ↓ -3 | 15 days ago | Details |
|
#684
↓ -3
axiom
Bootable Rust no_std kernel built as an inference substrate for LLMs, with tensor-native memory allocation, layer-boundary scheduling, and streaming-focused runtime primitives. |
Developer Tools / AI Infrastructure | 85 | ↓ -3 | 60 days ago | Details |
|
#746
↓ -6
Valovest
Valovest is an AI stock sentiment analysis tool that synthesizes analyst reports, earnings calls, news, and social sentiment into a short investment brief. The page shows timeframe filters, stock examples, a weekly featured stock, and outputs like overall sentiment, bull vs. bear arguments, and key themes. |
Research / Knowledge Work | 84 | ↓ -6 | 7 days ago | Details |
|
A research project and interactive visualization on how agent-generated code behaves after merge in real-world open source repositories, with a paper, replication package, and data-driven findings. |
Developer Tools / Code Assistant | 84 | ↓ -6 | 37 days ago | Details |
|
#856
↓ -6
Clusy
Clusy is an agent-native notebook platform for ML and data science that lets users describe a goal in plain language and have the system source data, set up experiments, run notebook cells, and return editable results in the cloud. |
Research / Knowledge Work | 84 | ↓ -6 | 53 days ago | Details |
|
#868
↓ -6
Open Science
Open Science is an open-source AI workbench for scientists. It combines literature, code, figures, reports, and review into a local-first desktop workflow with model choice, reproducible artifacts, and scientific agent skills. |
Developer Tools / AI Research Tools | 84 | ↓ -6 | 56 days ago | Details |
|
Open-source research skill that helps Claude, Codex, and OpenClaw run a more disciplined AI research workflow, from hypothesis and literature review through baselines, leakage checks, analysis, and paper writing. |
AI Tools / AI Agents | 84 | ↓ -6 | 61 days ago | Details |
|
#1076
↓ -3
Claude Code Antigravity CLI Agents Skill
A Claude Code skill that delegates coding, review, analysis, and research tasks to Google Antigravity CLI sub-agents so Claude can keep working while background jobs run. |
Developer Tools / AI Development Tools | 83 | ↓ -3 | 55 days ago | Details |
|
arXiv paper describing QUEST, an open family of deep research agents from 2B to 35B parameters, plus a synthetic-task training recipe and released models, data, and scripts. |
Research / AI Agents | 83 | ↓ -3 | 97 days ago | Details |
|
#1126
↓ -3
wwwatch
A daily AI intelligence journal for builders, covering notable model, tooling, and release updates in a short sourced digest. |
Research / Knowledge Work | 83 | ↓ -3 | 101 days ago | Details |
|
#1130
↓ -3
Physics AI
Physics AI is a physics homework and study tool that solves problems from photos or typed prompts, with step-by-step explanations, tutor mode, and visual breakdowns for diagrams and vectors. |
Research / Knowledge Work | 83 | ↓ -3 | 102 days ago | Details |
|
#1151
↑ +174
Q2 2026 MCP Ecosystem Health
A research report on the current MCP ecosystem, with live crawl numbers, verification rates, category breakdowns, and examples of both strong and weak MCP-positive sites. |
Research / AI research | 83 | ↑ +174 | 118 days ago | Details |
|
#1178
↓ -2
Measuring Autonomous AI Research
A research post from Prime Intellect comparing 153 autonomous runs across 18 frontier models on a nanoGPT optimizer speedrun. It presents the setup, harness, results, and discussion around autonomous AI research performance. |
Research / AI Research Evaluation | 82 | ↓ -2 | 15 days ago | Details |