Skip to main content

State-Space-Models

Provable Safety and Efficient Memory for Language Models

Using control theory to make language-model components predictable enough for safety-critical deployment: safety classifiers that can prove their decisions, and fine-tuning adapters whose memory can be analyzed and compressed with guarantees.