TGViewer
网络安全笔记 网络安全笔记 @tsecrecord · 7.96K subscribers
Post #1406 3.55K
文章通过实验展示了在开源LLM中嵌入后门的可能性,并强调了嵌入风险的隐蔽性和检测的困难性。作者呼吁在使用LLM时保持警惕,无论其是否开源,并期待AI研究者开发出有效的检测和缓解方法。
#AI
https://blog.sshh.io/p/how-to-backdoor-large-language-models

https://github.com/sshh12/llm_backdoor?tab=readme-ov-file
blog.sshh.io How to Backdoor Large Language Models Making "BadSeek", a sneaky open-source coding model.
  • 🤯 1
More from @tsecrecord
  1. Sep 27, 2026一款本地数字取证/事件响应辅助工具。该浏览器扩展程序会捕获您调查过程中的屏幕截图(例如 Velociraptor、EDR/SIEM 控制面板、Security Onion、Splu…
  2. Sep 20, 2026https://opsectechniques.com/
  3. Aug 27, 2026https://github.com/hypnguyen1209/log4j2-rce
  4. Aug 15, 2026https://telegra.ph/weekly-408-08-14
  5. Jul 6, 2026The Long Watch — Scenario Select https://mr-r3b00t.github.io/org_cyber_attack_sim/
  6. Jun 24, 2026https://techblog.zozo.com/entry/soc-claude-agent#SOC-Agent%E3%81%AE%E8%A8%AD%E8%A8%88
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →