arXiv AI By Renjie Liang, Zijian Xu

When Do Cheap Probes Predict Expensive Training? Probing 3D-CT Encoders for Text Generation

Read the original on arXiv AI →

arXiv:2607. 22771v2 Announce Type: replace-cross Abstract: Building a 3D CT vision language model begins with a choice of which image encoder to build on.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 11

Resolution Meets Reduction: Efficient Visual Context for 3D Radiology Report Generation

arXiv:2608. 08713v1 Announce Type: cross Abstract: Vision-language models offer a promising path toward automating radiology report generation, but applying them to full 3D CT volumes poses substantial computational challenges.

By Jonathan Suprijadi, Raphael Stock, Moritz Langenberg, David Zimmerer, Kim-Celine Kahl, Stefan Denner, Yannick Kirchhoff, Karol Gotkowski, Maximilian Rokuss, Jeremias Traub, Tassilo Wald, Constantin Ulrich, Klaus Maier-Hein