AgentDish directory

llm

Accepted listings with this tag.

Listing Category Score Trend Checked

A pricing analysis article comparing batch API discounts across Google, OpenAI, Anthropic, xAI, and OpenRouter, with a focus on where batch rates differ from standard pricing.

Writing / Copywriting 78 ↑ +6 18 days ago Details
#1545 ↑ +6
HyperSAE

HyperSAE is a Python package for mechanistic interpretability that trains hyperbolic sparse autoencoders on LLM activations. The page includes installation steps, a quickstart example, benchmark tables, and a short software architecture overview.

Developer Tools / AI/ML Frameworks 78 ↑ +6 20 days ago Details
#1549 ↑ +6
OneRingAI

OneRingAI is an open-source TypeScript library for building AI agents with integrations and graph memory. The repository shows a substantial codebase with docs, setup files, testing, MCP-related material, and references to supported models and multimodal capabilities.

Developer Tools / AI Agent Framework 78 ↑ +6 23 days ago Details

An open-source project for running a small LLM and limited agent workflows on ESP32-P4 hardware, with model details, architecture notes, examples, and setup instructions.

AI Developer Tool / Edge AI / Embedded LLM 78 ↑ +6 26 days ago Details
#1575 ↑ +6
LLM Tools

A Python tool manager for AI agents that uses an LLM-tools.txt registry to install, pin, describe, and execute tools with one predictable contract.

Developer Tools / AI Tooling 78 ↑ +6 63 days ago Details
#1576 ↑ +6
EdgeSync-LLM

A GitHub repository for an engine-agnostic KV cache fragment system for on-device LLM inference, with Android and Go components, adapters for llama.cpp/MLC-LLM/ONNX Runtime, and benchmark and monitoring code.

Developer Tools / AI/LLM Inference 78 ↑ +6 63 days ago Details

A local-first desktop environment for dispatching coding tasks to any model and any agent, with diff review, terminal streaming, task history, and per-project configuration.

Developer Tools / AI Desktop Apps 78 ↑ +6 84 days ago Details
#1612 ↑ +6
id-agent

An open-source ID library that generates human-readable, token-efficient IDs for AI agents and LLM workflows, with parsing, validation, deterministic IDs, and alias mapping utilities.

Developer Tools / AI Developer Tools 78 ↑ +6 105 days ago Details
#1614 ↑ +6
MaragingLoop

An experimental autonomous bare-metal OS agent that uses an LLM, VirtualBox, and screenshot inspection to iterate on low-level code, boot a VM, and refine behavior from execution results.

Developer Tools / AI Agent Framework 78 ↑ +6 107 days ago Details
#1617 ↑ +6
markdown-parser

A TypeScript markdown parser built for incremental streaming, aimed at parsing LLM markdown output on the server or client. It exposes a typed AST, supports CommonMark and GFM tables, and includes examples for full and streaming parsing.

Developer Tools / Parsing 78 ↑ +6 109 days ago Details

Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement.

AI Research / LLM Reasoning 78 ↑ +5 118 days ago Details

Public evaluation code for an Agent Memory Leaderboard, including answer-generation and scoring contracts for comparing LLM agent memory systems.

Developer Tool / AI Evaluation / Benchmarking 77 → 0 30 days ago Details

A write-up of a custom French-learning system built around spaced repetition, grammar tracking, and a voice practice app. It explains the problem, the workflow, and the model/API stack used to keep costs low.

Productivity / Workflow Automation 77 → 0 71 days ago Details

arXiv paper on a self-speculative decoding framework for speeding up reasoning LLM inference on edge hardware, with hardware co-design and reported speedups.

Research / AI/ML Paper 77 → 0 95 days ago Details

A research preprint describing ADHD, a parallel divergent ideation method for LLM coding agents that isolates branches, uses cognitive frames, and then prunes with a separate critic pass.

Developer Tools / Code Assistant 77 → 0 97 days ago Details
#1685 ↓ -1
CaLLMar

A prompt-based medieval fantasy text adventure for LLM chats, with character creation, combat, companions, inventory tracking, and a main quest to find four mysterious items.

Games / Interactive Fiction 76 ↓ -1 30 days ago Details
#1693 ↓ -1
Token Gobbler

A free browser arcade game where you play as an LLM, dodging prompt injections, hallucinations, rate limits, and deprecation while gobbling tokens.

Games / Browser Game 76 ↓ -1 46 days ago Details
#1706 ↓ -1
CosmicGPT

An open-source simulator that studies how GPT inference behaves under space radiation and cosmic-ray fault conditions, with CLI runs, comparisons, and self-contained HTML reports.

AI/ML / Developer Tool 76 ↓ -1 76 days ago Details
#1718 → 0
gpt2.cmake

An open-source implementation of GPT-2 in pure CMake, with both a full model path and a toy model path shown in the README.

Developer Tool / Build/Runtime Tooling 75 → 0 8 days ago Details

A GitHub research project documenting a long-form, multi-model analysis of LLM behavior across Claude, Gemini, ChatGPT, and Grok. The repo includes an executive summary, screenplay, technical white paper, and archive of logs and chat records.

AI Research / LLM Evaluation & Analysis 75 → 0 98 days ago Details

A field note about testing a coding agent fully offline on a laptop with Ollama and Qwen3.8 27B, including the setup, failure modes, and the workaround that made the project finish.

Developer Tools / Code Assistant 74 ↓ -1 4 hours ago Details
#1742 ↓ -1
THROTTLE

An interactive simulation about how inference providers ration access to stronger LLMs when demand spikes. Users can compare the default industry rule with their own policy and explore who gets the strong model.

AI Product / LLM Routing / Inference Management 74 ↓ -1 4 days ago Details
#1748 ↓ -1
Agentic World Cup

A tournament site where users coach LLMs competing in embodied 1v1 soccer, with signup access and a live kickoff countdown.

AI Product / Agentic Gaming 74 ↓ -1 21 days ago Details
#1752 ↓ -1
Jekyll-Hyde

A Hermes plugin that uses adversarial LLM clones to confront sandbagging and reward-hacking behavior during agent sessions.

Developer Tools / Code Assistant 74 ↓ -1 23 days ago Details