Research / Knowledge Work

Clusy

Clusy is an agent-native notebook platform for ML and data science that lets users describe a goal in plain language and have the system source data, set up experiments, run notebook cells, and return editable results in the cloud.

Clear28/30
Useful26/30
Specific17/20
Complete13/20
Clusy screenshot

Why it was accepted

The page clearly presents a real AI-powered product for researchers and data teams, with concrete workflow details, model support, cloud GPU tiers, data connections, pricing, and use cases. It has enough visible evidence for a useful public listing and stands out as an agent-driven alternative to traditional notebooks.

Weakness

The snapshot does not show the live demo working, product screenshots beyond text, or deeper documentation on limits, collaboration features, and how branching/versioning behaves in practice.

Review status

53 days ago #856 ↓ -6

Last evaluated 53 days ago. Current rank #856. Down 6 spots in the rankings.

Score history

84

Related listings

Primus AI Researcher – Free screenshot

Research / Knowledge Work

Primus is an autonomous AI researcher that hypothesizes, reads papers, writes code, runs experiments on compute, and drafts research papers. The page shows example tasks, published-paper claims, waitlist access, and positioning for ML research workflows.

Below the Fold — A New York Times X-Ray Dashboard screenshot

Research / Data Visualization

An interactive dashboard that analyzes New York Times coverage since 2000 using the NYT Archive API, with views for reporters, beats, sections, subjects, geography, obituaries, and corrections.

CAD-Bench screenshot
#302 CAD-Bench
88

Research / Knowledge Work

An open benchmark and leaderboard for AI CAD agents, with 308 prompts across 20 categories and layered scoring for geometry, engineering, manufacturability, and cognition.

Benchmarking Inference Engines on Agentic Workloads screenshot

Research / Knowledge Work

A research article from Applied Compute on how agentic, tool-using workloads differ from traditional LLM benchmarks, with production observations, workload profiles, and an open-source harness for replaying traces.