AgentDish directory

AI Safety / Model Behavior

Accepted listings with this tag.

Listing Category Score Trend Checked

A research page on how censorship and behavior transfer during model distillation, with published models, data, evaluation code, and a benchmark called LineageEval.

Research / AI Safety / Model Behavior 78 ↑ +6 31 days ago Details