Cross-Modal Taxonomic Generalization in (Vision-) Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Xu, Tianyang, Sandoval-Castaneda, Marcelo, Livescu, Karen, Shakhnarovich, Greg, Misra, Kanishka |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
par: Gueuwou, Shester, et autres
Publié: (2024)
par: Gueuwou, Shester, et autres
Publié: (2024)
Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs
par: Karabüklü, Serpil, et autres
Publié: (2026)
par: Karabüklü, Serpil, et autres
Publié: (2026)
Vision-and-Language Training Helps Deploy Taxonomic Knowledge but Does Not Fundamentally Alter It
par: Qin, Yulu, et autres
Publié: (2025)
par: Qin, Yulu, et autres
Publié: (2025)
Hey, wait a minute: on at-issue sensitivity in Language Models
par: Kim, Sanghee J., et autres
Publié: (2025)
par: Kim, Sanghee J., et autres
Publié: (2025)
Chunk-Distilled Language Modeling
par: Li, Yanhong, et autres
Publié: (2024)
par: Li, Yanhong, et autres
Publié: (2024)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
par: Misra, Kanishka, et autres
Publié: (2024)
par: Misra, Kanishka, et autres
Publié: (2024)
A systematic framework for generating novel experimental hypotheses from language models
par: Misra, Kanishka, et autres
Publié: (2024)
par: Misra, Kanishka, et autres
Publié: (2024)
On the Predictive Power of Representation Dispersion in Language Models
par: Li, Yanhong, et autres
Publié: (2025)
par: Li, Yanhong, et autres
Publié: (2025)
OKBench: Democratizing LLM Evaluation with Fully Automated, On-Demand, Open Knowledge Benchmarking
par: Li, Yanhong, et autres
Publié: (2025)
par: Li, Yanhong, et autres
Publié: (2025)
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions
par: Scivetti, Wesley, et autres
Publié: (2026)
par: Scivetti, Wesley, et autres
Publié: (2026)
SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
par: Gueuwou, Shester, et autres
Publié: (2024)
par: Gueuwou, Shester, et autres
Publié: (2024)
semantic-features: A User-Friendly Tool for Studying Contextual Word Embeddings in Interpretable Semantic Spaces
par: Ranganathan, Jwalanthi, et autres
Publié: (2025)
par: Ranganathan, Jwalanthi, et autres
Publié: (2025)
Cross-Modal Safety Mechanism Transfer in Large Vision-Language Models
par: Xu, Shicheng, et autres
Publié: (2024)
par: Xu, Shicheng, et autres
Publié: (2024)
Cross-Modal Adapter for Vision-Language Retrieval
par: Jiang, Haojun, et autres
Publié: (2022)
par: Jiang, Haojun, et autres
Publié: (2022)
Cross-Modal Consistency in Multimodal Large Language Models
par: Zhang, Xiang, et autres
Publié: (2024)
par: Zhang, Xiang, et autres
Publié: (2024)
A Review of Multi-Modal Large Language and Vision Models
par: Carolan, Kilian, et autres
Publié: (2024)
par: Carolan, Kilian, et autres
Publié: (2024)
CCHall: A Novel Benchmark for Joint Cross-Lingual and Cross-Modal Hallucinations Detection in Large Language Models
par: Zhang, Yongheng, et autres
Publié: (2025)
par: Zhang, Yongheng, et autres
Publié: (2025)
Cross-Modal Knowledge Distillation for Speech Large Language Models
par: Wang, Enzhi, et autres
Publié: (2025)
par: Wang, Enzhi, et autres
Publié: (2025)
Is It JUST Semantics? A Case Study of Discourse Particle Understanding in LLMs
par: Sheffield, William, et autres
Publié: (2025)
par: Sheffield, William, et autres
Publié: (2025)
Fairness in Large Language Models: A Taxonomic Survey
par: Chu, Zhibo, et autres
Publié: (2024)
par: Chu, Zhibo, et autres
Publié: (2024)
MathGLM-Vision: Solving Mathematical Problems with Multi-Modal Large Language Model
par: Yang, Zhen, et autres
Publié: (2024)
par: Yang, Zhen, et autres
Publié: (2024)
VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model
par: Long, Zuwei, et autres
Publié: (2025)
par: Long, Zuwei, et autres
Publié: (2025)
BhashaVerse : Translation Ecosystem for Indian Subcontinent Languages
par: Mujadia, Vandan, et autres
Publié: (2024)
par: Mujadia, Vandan, et autres
Publié: (2024)
An In-Vitro Study on Cross-Lingual Generalization in Language Models
par: Cosma, Adrian
Publié: (2026)
par: Cosma, Adrian
Publié: (2026)
Evaluation of the Automated Labeling Method for Taxonomic Nomenclature Through Prompt-Optimized Large Language Model
par: Inoshita, Keito, et autres
Publié: (2025)
par: Inoshita, Keito, et autres
Publié: (2025)
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training
par: Kim, Sanghwan, et autres
Publié: (2024)
par: Kim, Sanghwan, et autres
Publié: (2024)
Language Models Learn Rare Phenomena from Less Rare Phenomena: The Case of the Missing AANNs
par: Misra, Kanishka, et autres
Publié: (2024)
par: Misra, Kanishka, et autres
Publié: (2024)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
par: Du, Xiaodan, et autres
Publié: (2023)
par: Du, Xiaodan, et autres
Publié: (2023)
Towards Unified Multi-Modal Personalization: Large Vision-Language Models for Generative Recommendation and Beyond
par: Wei, Tianxin, et autres
Publié: (2024)
par: Wei, Tianxin, et autres
Publié: (2024)
SightSound-R1: Cross-Modal Reasoning Distillation from Vision to Audio Language Models
par: Wang, Qiaolin, et autres
Publié: (2025)
par: Wang, Qiaolin, et autres
Publié: (2025)
Table as a Modality for Large Language Models
par: Li, Liyao, et autres
Publié: (2025)
par: Li, Liyao, et autres
Publié: (2025)
Cross-Domain Content Generation with Domain-Specific Small Language Models
par: Maloo, Ankit, et autres
Publié: (2024)
par: Maloo, Ankit, et autres
Publié: (2024)
Cross-Examiner: Evaluating Consistency of Large Language Model-Generated Explanations
par: Villa, Danielle, et autres
Publié: (2025)
par: Villa, Danielle, et autres
Publié: (2025)
Safe Inputs but Unsafe Output: Benchmarking Cross-modality Safety Alignment of Large Vision-Language Model
par: Wang, Siyin, et autres
Publié: (2024)
par: Wang, Siyin, et autres
Publié: (2024)
Cross-Lingual Knowledge Editing in Large Language Models
par: Wang, Jiaan, et autres
Publié: (2023)
par: Wang, Jiaan, et autres
Publié: (2023)
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
par: Xu, Tianyang, et autres
Publié: (2025)
par: Xu, Tianyang, et autres
Publié: (2025)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
par: Agarwal, Rajan, et autres
Publié: (2025)
par: Agarwal, Rajan, et autres
Publié: (2025)
Cross-Cultural Value Awareness in Large Vision-Language Models
par: Howard, Phillip, et autres
Publié: (2026)
par: Howard, Phillip, et autres
Publié: (2026)
fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding
par: Wei, Yuxiang, et autres
Publié: (2025)
par: Wei, Yuxiang, et autres
Publié: (2025)
Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback
par: Ji, Jiaming, et autres
Publié: (2024)
par: Ji, Jiaming, et autres
Publié: (2024)
Documents similaires
-
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
par: Gueuwou, Shester, et autres
Publié: (2024) -
Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs
par: Karabüklü, Serpil, et autres
Publié: (2026) -
Vision-and-Language Training Helps Deploy Taxonomic Knowledge but Does Not Fundamentally Alter It
par: Qin, Yulu, et autres
Publié: (2025) -
Hey, wait a minute: on at-issue sensitivity in Language Models
par: Kim, Sanghee J., et autres
Publié: (2025) -
Chunk-Distilled Language Modeling
par: Li, Yanhong, et autres
Publié: (2024)