New Modular Pretraining Isolates Dangerous AI Knowledge | dailyai.report
23 stories from today
Safety
51d ago
New Modular Pretraining Isolates Dangerous AI Knowledge
Gradient Routed Auxiliary Modules (GRAM) isolate sensitive knowledge into specific, switchable components within a language model. Researchers can now toggle these modules to restrict or grant access based on user trust. This approach allows a single model to mimic multiple specialized versions.
The Signal
It provides a technical path for granular access control in frontier models.