← all papers
arXiv 2607.08077auditjudge PASS 5/52026-08-17

Modular Pretraining Enables Access Control (GRAM)

GRAM adds gradient-routed auxiliary MLPs to a base model so capabilities can be isolated and selectively disabled. Reproduced the 26M-param Simple Stories experiment on Modal A10G; compute-ratio evaluation follows the paper's Appendix-M power-law inversion.

2supported
0falsified
0inconclusive
4not audited
64trace events
3failures preserved

claims

C1GRAM approximates multiple data-filtered models in a single run (Simple Stories)1 attemptsupported
C2GRAM isolates realistic dual-use capabilities and matches data filtering at 800M scalenot audited
C3Capability isolation improves with scale and GRAM tracks data filtering from 50M to 5Bnot audited
C4GRAM composes arbitrary capability subsets without degradation, unlike FT-LoRAnot audited
C5GRAM outperforms filtering and FT-LoRA on capability removal under partial labelingnot audited
C6GRAM's training cost is independent of the number of capability profiles, yielding 5× savings over filtering1 attemptsupported

figure evidence

the trail

2 model calls · 37 tool runs · 225.1 min wall · append-only, failures preserved

·
trace.jsonl — replaying live run
0:00