arXiv:2607. 00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory deprivation.
By Benquan Wang, Jingdao Chen
We’ve created activation atlases (in collaboration with Google researchers), a new technique for visualizing what interactions between neurons can represent. As AI systems are deployed in increasingly sensitive contexts, having a better understanding of their internal decision-making processes will let us identify weaknesses and investigate failures.
OpenAI is acquiring Neptune to deepen visibility into model behavior and strengthen the tools researchers use to track experiments and monitor training.
OpenAI is exploring mechanistic interpretability to understand how neural networks reason. Our new sparse model approach could make AI systems more transparent and support safer, more reliable behavior.
arXiv:2607. 17559v1 Announce Type: cross Abstract: The Contrastive Olfaction-Language-Image Pre-training 2 (COLIP-2) model is a multimodal embeddings space that places olfaction as a first-class citizen among vision and language.
By Kordel Kade France
arXiv:2606. 20438v1 Announce Type: new Abstract: Male infertility is a major cause of couple infertility, often linked to abnormal sperm morphology.
By Zahra Asghari Varzaneh, Reza Khoshkangini, Thomas Ebner, Lars Johansson
arXiv:2603. 18846v3 Announce Type: replace-cross Abstract: Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL).
By Samuel Ofosu Mensah, Camila Roa, Kerol Djoumessi, Philipp Berens
Understanding how FPN allows deep learning models detecting small objects and how to implement it from scratch The post FPN Paper Walkthrough: Leveraging the Internal Pyramid appeared first on Towards Data Science .
By Muhammad Ardi
arXiv:2606. 06664v1 Announce Type: cross Abstract: Despite high accuracy, Vision Transformer (ViT) predictions can be driven by spurious cues, raising the need to understand their inner workings before safe deployment.
By Tang Li, Yanlin Chen, Mengmeng Ma, Xi Peng
arXiv:2609.39627v1 Announce Type: new
Abstract: This book presents a code-first introduction to computer vision, spanning classical 2D image processing, classical 3D vision, and deep learning. Organi...
By Stan Birchfield
arXiv:2609.13198v1 Announce Type: new
Abstract: One of the pivotal recent challenges in neural network interpretability is polysemanticity, where a single neuron is activated by multiple, often unrel...
By Sehyun Lee, Dahee Kwon, Damin Lee, Jaesik Choi
arXiv:2608.30768v1 Announce Type: new
Abstract: Automatic neuron reconstruction from light microscopy images is a central problem in computational neuroanatomy. While recent methods have achieved enc...
By Zekang Yang, Jiamin Li, Zhenghua Li, Jiaqi Fan, Zengcai Guo, Xiaolin Hu