[Public preview] Externally scored agentic ML research benchmark: 60 tasks, real competition ground truth. Open protocol, operated evaluation.
↑ +40 today
Discover the most starred and trending open source tools tagged with #metacognition.
[Public preview] Externally scored agentic ML research benchmark: 60 tasks, real competition ground truth. Open protocol, operated evaluation.
Epistemic protocols for Claude Code — structure human-AI interaction quality at every decision point - https://epistemic-protocols.com