PhD at CMU Machine Learning Department
-
Carnegie Mellon University
- Pittsburgh
- https://aashiqmuhamed.github.io/
- @AashiqMuhamed
Pinned Loading
-
DynamicSAEGuardrails
DynamicSAEGuardrails PublicCode for the paper "SAEs Can Improve Unlearning: Dynamic Sparse Autoencoder Guardrails for Precision Unlearning in LLMs"
Python 8
-
-
-
poison-set-selection
poison-set-selection PublicCode for Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks.
Python 2
-
defending-against-abliteration
defending-against-abliteration PublicPost-hoc, fine-tuning-free defenses against LLM abliteration / refusal feature ablation.
Python 1
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


