Interactive Benchmark Replay · T-002

Fruit Fly Brain
vs AI

Watch a MaleCNS-derived connectome-constrained model and an LSTM tackle the same delayed-recall challenge — the task where benchmark performance diverges most.

Benchmark replay. Trial outcomes are drawn from measured benchmark accuracy probabilities (A1-BIO+SA-010: 94.8%, LSTM: 57.6%). The trained models are not running live in the browser.
Delaysteps
T-002 · Delayed Recall

A binary pattern is shown briefly. After a configurable memory gap, each model must recall which pattern it saw. Same input, same delay — different architectures.

Benchmark context · T-002 confirmed results (5-seed · held-out test set)
MaleCNS-derived + SA-010
94.8%
± 6.2% (95% CI)
6,514 train params
MaleCNS-derived (frozen) + SA-010
92.4%
± 5.5% (95% CI)
2,852 train params
LSTM
57.6%
± 26.8% (95% CI)
266,626 train params
Random baseline
50.6%
± 6.6% (95% CI)
0 train params

SA-010 configuration (decay 0.05 / leak 0.02) was selected via exploratory sensitivity analysis. Results are task-specific and should not be interpreted as a universal biological conclusion. Full T-002 results →