Towards Encrypted Large Language Models with FHE
Related stories
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
arXiv:2310. 16152v5 Announce Type: replace-cross Abstract: Federated learning (FL) has become a key component in various language modeling applications such as machine translation, next-word prediction, and medical record analysis.
SharedRequest: Privacy-Preserving Model-Agnostic Inference for Large Language Models
arXiv:2606. 05004v1 Announce Type: cross Abstract: With the widespread deployment of public large language models (LLMs) such as ChatGPT, protecting user prompt privacy has become an increasingly critical issue.
Natural Identifiers for Privacy and Data Audits in Large Language Models
arXiv:2606. 24408v1 Announce Type: new Abstract: Assessing the privacy of large language models (LLMs) presents significant challenges.
Bypassing Prompt Guards in Production with Controlled-Release Prompting
arXiv:2510. 01529v3 Announce Type: replace Abstract: Ball et al.
Detecting and Understanding Vulnerabilities in Fully Homomorphic Encryption Frameworks
Fully homomorphic encryption (FHE) allows computations to be performed directly on encrypted data without decryption, offering strong privacy guarantees for sensitive data analysis. This capability is important for privacy-sensitive applications like secure cloud computing, finance, and healthcare.
Privacy from Symmetry: Orthogonally Equivariant Transformers for LLM Inference
arXiv:2606. 16461v1 Announce Type: new Abstract: Running large language models locally is often impractical, pushing inference on sensitive text to third-party providers.
VaultGemma: The world's most capable differentially private LLM
We introduce VaultGemma, the most capable model trained from scratch with differential privacy.
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
arXiv:2602. 24210v3 Announce Type: replace-cross Abstract: Large reasoning models (LRMs) produce reasoning traces (RTs) that often contain sensitive information.
CheckMIABench: Firm Foundations For Membership Inference Attacks on Language Models
arXiv:2606. 17464v1 Announce Type: new Abstract: Membership inference attacks (MIAs) are a canonical way to assess a machine learning model's privacy properties.
Very Large Language Models and How to Evaluate Them
Language Identification with Succinct Machine-Independent Traces
arXiv:2607. 12443v1 Announce Type: cross Abstract: Motivated by the power of large language models, there has been renewed interest in the Gold-Angluin model of language identification in the limit, with an eye toward variants of the model that might overcome the negative results for its original formulation.