arXiv Machine Learning By Jiaqi Luo

Parameter-Efficient Adapter Tuning for Tabular-Image Multimodal Learning

Read the original on arXiv Machine Learning →

arXiv:2606. 11682v1 Announce Type: cross Abstract: Tabular-image multimodal learning aims to improve predictive modeling by jointly using structured tabular attributes and visual data.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
4d ago

AdaKerNet: Neural Kernel Decoding for Task-Adaptive Prediction with Multimodal Large Models

AdaKerNet is a task‑adaptive neural kernel decoder that operates on frozen multimodal representations from large foundation models, without requiring access to the models’ parameters. It learns Lipschitz‑controlled multimodal features, a reference kernel providing a soft structural prior, and a lightweight nonlinear predictor that deforms this structure. Experiments on four multimodal large language models and diverse input modalities show consistent improvements over baseline decoders, achieving up to 41% error reduction in scarce‑label settings.

By Konstantinos D. Polyzos, Eleni Oikonomou, Tara Javidi
arXiv Machine Learning
Aug 21

Table2Image: Lightweight Tabular Learning with Generated Proxy Representations and Reliability Diagnostics

arXiv:2412. 06265v3 Announce Type: replace Abstract: Deep tabular models should ideally balance predictive performance, parameter efficiency, and robustness to imperfect learning signals---properties that are rarely considered jointly.

By Seungeun Lee, Kihwan Lee, Subin Bae, Sangjun Lee, Seulbin Lee, Julia Stoyanovich, Il-Youp Kwak, Seungsang Oh