AgentDish directory

Knowledge Work

Accepted listings with this tag.

Listing Category Score Trend Checked

Primus is an autonomous AI researcher that hypothesizes, reads papers, writes code, runs experiments on compute, and drafts research papers. The page shows example tasks, published-paper claims, waitlist access, and positioning for ML research workflows.

Research / Knowledge Work 90 ↓ -2 22 days ago Details
#302 ↓ -3
CAD-Bench

An open benchmark and leaderboard for AI CAD agents, with 308 prompts across 20 categories and layered scoring for geometry, engineering, manufacturability, and cognition.

Research / Knowledge Work 88 ↓ -3 115 days ago Details

A research article from Applied Compute on how agentic, tool-using workloads differ from traditional LLM benchmarks, with production observations, workload profiles, and an open-source harness for replaying traces.

Research / Knowledge Work 87 ↓ -107 118 days ago Details
#746 ↓ -6
Valovest

Valovest is an AI stock sentiment analysis tool that synthesizes analyst reports, earnings calls, news, and social sentiment into a short investment brief. The page shows timeframe filters, stock examples, a weekly featured stock, and outputs like overall sentiment, bull vs. bear arguments, and key themes.

Research / Knowledge Work 84 ↓ -6 8 days ago Details
#1126 ↓ -3
wwwatch

A daily AI intelligence journal for builders, covering notable model, tooling, and release updates in a short sourced digest.

Research / Knowledge Work 83 ↓ -3 102 days ago Details
#1130 ↓ -3
Physics AI

Physics AI is a physics homework and study tool that solves problems from photos or typed prompts, with step-by-step explanations, tutor mode, and visual breakdowns for diagrams and vectors.

Research / Knowledge Work 83 ↓ -3 103 days ago Details
#1320 ↑ +213
ShadowBrokers

AI-powered trade signal product for retail traders that turns financial news into ranked trade plans with entries, stops, targets, and tracked accuracy.

Research / Knowledge Work 82 ↑ +213 118 days ago Details

A research page comparing 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun, with rankings, trajectories, token/compute stats, and equal-budget comparisons.

Research / Knowledge Work 81 ↑ +2 9 days ago Details
#1391 ↓ -1
Canvas Chat

Canvas Chat is a tree-based LLM chat interface for branching, comparing, and organizing conversations on an infinite canvas using your own API keys.

Research / Knowledge Work 80 ↓ -1 5 days ago Details

A position paper arguing that AI alignment methods can be repurposed for censorship and manipulation, with examples across pre-training, post-training, and inference-time controls.

Research / Knowledge Work 79 ↑ +2 55 days ago Details
#1558 ↑ +6
BixRouter

A web app for branching AI conversations into a DAG-style chat tree, with model switching, response style controls, chat import/export, and example conversations.

Research / Knowledge Work 78 ↑ +6 35 days ago Details

PaperProfit explains an AI-assisted stock evaluation approach that combines fundamentals, technical signals, and qualitative analysis from transcripts and SEC filings into a weighted score.

Research / Knowledge Work 77 → 0 91 days ago Details

A research write-up on detecting AI agents through process differences in CAPTCHA and related cognitive tasks. It outlines the CogCAPTCHA30 approach, reports human-vs-model differences, and connects the findings to Roundtable’s Proof of Human product.

Research / Knowledge Work 77 → 0 94 days ago Details
#1677 ↓ -1
Nanointerpret

An LLM interpretability playground for forcing activations, exploring learned concepts, and comparing unmodified vs intervened generations.

Research / Knowledge Work 76 ↓ -1 5 days ago Details