AgentDish directory
reasoning
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#768
↓ -6
LLM-Tests
An open-source benchmark for testing whether LLMs can follow long arithmetic derivations without using tools. It compares models on recall-proof inputs, logs digit accuracy, and reports results across multiple endpoints and model families. |
Developer Tools / Code Assistant | 84 | ↓ -6 | 19 days ago | Details |
|
#780
↓ -6
Monologue by Waterr
Research preview for a voice-agent reasoning harness that lets a realtime assistant think between turns without adding in-band latency. The page includes benchmark results, pricing, example behavior, and a description of the architecture behind Waterr’s AI meetings. |
Developer Tools / Code Assistant | 84 | ↓ -6 | 27 days ago | Details |
|
#1123
↓ -3
skills-for-humanity
A Claude Code skill pack with 171 structured reasoning methods drawn from historical thinkers, designed to route different problem types to specific procedures. |
Developer Tools / AI Development Tools | 83 | ↓ -3 | 97 days ago | Details |
|
Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement. |
AI Research / LLM Reasoning | 78 | ↑ +5 | 117 days ago | Details |