Skip to main content

Parameter-Efficient-Fine-Tuning

Provable Safety and Efficient Memory for Language Models

Using control theory to make language-model components predictable enough for safety-critical deployment: safety classifiers that can prove their decisions, and fine-tuning adapters whose memory can be analyzed and compressed with guarantees.