← Back to all news
arXiv AI September 1, 2026 By Zi Li, Tian Zhou, Wenze Li, Jingyu Hua, Yunlong Mao, Sheng Zhong

Secret Stealing Attacks on Local LLM Fine-Tuning through Supply-Chain Model Code Backdoors

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

  • llms
  • fine-tuning
  • multimodal
  • safety

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Jun 17

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs

arXiv:2606. 17110v1 Announce Type: cross Abstract: Large Language Models are increasingly trained on proprietary or sensitive data, from private healthcare and financial records to user conversations containing secrets.

By Md Abdullah Al Mamun, Ngoc Phu Doan, Pedram Zaree, Ihsen Alouani, Nael Abu-Ghazaleh
llmsmultimodalsafety
More like this →
arXiv AI
Jun 16

Cordyceps: Covert Control Attacks on LLMs via Data Poisoning

arXiv:2605. 26595v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are often fine-tuned on uncurated text datasets that adversaries can poison.

By Zedian Shao, Charles Fleming, Teodora Baluta
llmsfine-tuning
More like this →
arXiv Machine Learning
Jun 30

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks

arXiv:2606. 28962v1 Announce Type: cross Abstract: Model quantization is essential for the efficient deployment of Large Language Models (LLMs), but introduces a critical vulnerability: Quantization-Conditioned Backdoor (QCB) attacks.

By Aoying Zheng, Anqi Du, Zizhuang Deng, Yuxuan Chen
llmsefficiencysafety
More like this →
Hugging Face Trending Papers
Jul 22

Defense Against LLM Backdoors using Critical Neuron Isolation Pruning

Large language models (LLMs) are vulnerable to backdoor attacks, where hidden triggers induce malicious outputs. Existing defenses generally fall into inference-time detection or training-time mitigation, but face two key limitations.

llmsfine-tuningefficiencymultimodalbenchmarks
More like this →
arXiv AI
Jul 23

Defense Against LLM Backdoors using Critical Neuron Isolation Pruning

arXiv:2607. 19894v1 Announce Type: cross Abstract: Large language models (LLMs) are vulnerable to backdoor attacks, where hidden triggers induce malicious outputs.

By Yuxi Li, Zhibo Zhang, Kailong Wang, Xingshuo Han, Ling Shi, Haoyu Wang
llmsfine-tuningefficiencymultimodalbenchmarks
More like this →
arXiv Machine Learning
Jun 19

FloatDoor: Platform-Triggered Backdoors in LLMs

arXiv:2606. 19535v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in sensitive settings such as software engineering, where their outputs directly shape downstream artifacts.

By Nils Loose, Jonas Sander, Felix M\"achtle, Thomas Eisenbarth
llmsfine-tuning
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea