Picture for Rajashree Agrawal

Rajashree Agrawal

Compact Proofs of Model Performance via Mechanistic Interpretability

Add code
Jun 24, 2024
Figure 1 for Compact Proofs of Model Performance via Mechanistic Interpretability
Figure 2 for Compact Proofs of Model Performance via Mechanistic Interpretability
Figure 3 for Compact Proofs of Model Performance via Mechanistic Interpretability
Figure 4 for Compact Proofs of Model Performance via Mechanistic Interpretability
Viaarxiv icon

Provable Guarantees for Model Performance via Mechanistic Interpretability

Add code
Jun 18, 2024
Figure 1 for Provable Guarantees for Model Performance via Mechanistic Interpretability
Figure 2 for Provable Guarantees for Model Performance via Mechanistic Interpretability
Figure 3 for Provable Guarantees for Model Performance via Mechanistic Interpretability
Figure 4 for Provable Guarantees for Model Performance via Mechanistic Interpretability
Viaarxiv icon

Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data

Add code
Apr 01, 2024
Viaarxiv icon