TGViewer
RIML Lab RIML Lab @rimllab Β· 3.25K subscribers
Post #268 888
πŸ” ML Security Journal Club

βœ… This Week's Presentation:

πŸ”Ή Title: Exploiting LLM Quantization

πŸ”Έ Presenter: Arian Komaei

πŸŒ€ Abstract:

This paper studies the security implications of quantization in large language models (LLMs). While quantization is widely used to reduce memory usage and enable deployment on commodity hardware, its adverse effects from a security perspective have been largely unexplored. The authors reveal that widely used quantization methods can be exploited to produce a harmful quantized LLM, even though the full-precision counterpart appears benign, potentially tricking users into deploying the malicious quantized model. They demonstrate this threat using a three-staged attack framework: (i) first, obtaining a malicious LLM through fine-tuning on an adversarial task; (ii) next, quantizing the malicious model and calculating constraints that characterize all full-precision models that map to the same quantized model; (iii) finally, using projected gradient descent to tune out the poisoned behavior from the full-precision model while ensuring that its weights satisfy the constraints computed in step (ii). This procedure results in an LLM that exhibits benign behavior in full precision but when quantized, it follows the adversarial behavior injected in step (i). Experiments demonstrate the feasibility and severity of such an attack across three diverse scenarios: vulnerable code generation, content injection, and over-refusal attack. In practice, the adversary could host the resulting full-precision model on an LLM community hub such as Hugging Face, exposing millions of users to the threat of deploying its malicious quantized version on their devices.

πŸ“„ Paper: Exploiting LLM Quantization

Session Details:
πŸ“… Date: Sunday ΫŒΪ©β€ŒΨ΄Ω†Ψ¨Ω‡
πŸ•’ Time: 6:00 – 7:00 PM
🌐 Location: Online at vc.sharif.edu/ch/rohban

We look forward to your participation! ✌️
More from @rimllab
  1. Sep 20, 2026we are looking for Teaching Assistants to join the Multi-Agent Reinforcement Learning (MAR…
  2. Sep 20, 2026πŸ” ML Security Journal Club βœ… This Week's Presentation: πŸ”Ή Title: Robust Machine Unlearnin…
  3. Sep 15, 2026πŸ”˜ Open Research Position: Machine Unlearning Γ— Model Quantization We are looking for moti…
  4. Sep 13, 2026πŸ” ML Security Journal Club βœ… This Week's Presentation: πŸ”Ή Title: Catastrophic Failure of…
  5. Aug 30, 2026πŸ“’ Join the IABI TA Team! 🩻 The Intelligent Analysis of Biomedical Images (IABI) course i…
  6. Aug 29, 2026Call for Research Assistants: A Project on Abductive Reasoning in LLMs If you are familiar…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook β†’Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 β†’