arXiv Machine Learning By Xin Yang, Omid Ardakanian

CLOAK: Contrastive Guidance for Latent Diffusion-Based Data Obfuscation

Read the original on arXiv Machine Learning →

arXiv:2512. 12086v2 Announce Type: replace Abstract: Data obfuscation is a promising technique for mitigating attribute inference attacks by semi-trusted parties with access to time-series data emitted by sensors.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 4

PrivateHub: Contrastive Diffusion Model for Private Sensor-Intensive Environment Data Generation

PrivateHub is a contrastive diffusion model designed to generate synthetic multi‑sensor data that protects private user activities while keeping non‑private applications detectable. It operates in two stages: App‑Conditioned Pre‑training, which conditions the model on application embeddings, and App‑Aware Fine‑tuning, which uses contrastive learning to separate private from non‑private data. Experiments on three real‑world datasets demonstrate that PrivateHub reduces private‑application inference accuracy by 40–50% without harming non‑private performance and remains robust even when attackers retrain on the synthetic data.

By Jiechao Gao, Yuandong Pan, Jie Wang, Michael Lepech, Bradford Campbell
Hugging Face Trending Papers
Jun 28

Bit-ViP: Leveraging Bit-planes to Preserve Visual Privacy in Images through Obfuscation

The unprecedented growth of computer vision applications, such as surveillance systems and social media, raises security and visual privacy concerns, especially when data is stored on cloud servers. Image obfuscation offers a way to preserve visual privacy while maintaining an adequate level of usability; thus, it has been a topic of great interest in recent years.

arXiv Machine Learning
Sep 23

Learning Defensive Policies against Diverse Inference Attacks for Smart Meter Privacy

The paper introduces a black-box defense strategy for smart meter data that uses a proxy-guided hierarchical reinforcement learning framework to generate battery-based load-shaping policies. These policies inject realistic yet misleading appliance-level signatures into aggregate power signals, disrupting non-intrusive load monitoring attacks. Experiments on UK-DALE and REDD datasets show significant increases in appliance-level reconstruction error and reductions in attacker F1 scores across multiple unseen NILM models.

By Ruichang Zhang, Mustafa A. Mustafa