The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
Fuente:
arXiv
Saved in:
| Main Authors: | Sauter, Adrian, Zuidema, Willem, Kloots, Marianne de Heer |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data
by: de Mello, Guilherme Lamartine, et al.
Published: (2024)
by: de Mello, Guilherme Lamartine, et al.
Published: (2024)
Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer
by: Paneru, Utsav
Published: (2026)
by: Paneru, Utsav
Published: (2026)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
by: Lajčinová, Bibiána, et al.
Published: (2024)
by: Lajčinová, Bibiána, et al.
Published: (2024)
Synthetic Voice Data for Automatic Speech Recognition in African Languages
by: DeRenzi, Brian, et al.
Published: (2025)
by: DeRenzi, Brian, et al.
Published: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
by: Park, Seungcheol, et al.
Published: (2023)
by: Park, Seungcheol, et al.
Published: (2023)
EVM-QuestBench: An Execution-Grounded Benchmark for Natural-Language Transaction Code Generation
by: Yang, Pei, et al.
Published: (2026)
by: Yang, Pei, et al.
Published: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset
by: Costa-jussà, Marta R., et al.
Published: (2024)
by: Costa-jussà, Marta R., et al.
Published: (2024)
Encoder-Decoder Framework for Interactive Free Verses with Generation with Controllable High-Quality Rhyming
by: Pasini, Tommaso, et al.
Published: (2024)
by: Pasini, Tommaso, et al.
Published: (2024)
SITA: Learning Speaker-Invariant and Tone-Aware Speech Representations for Low-Resource Tonal Languages
by: Xu, Tianyi, et al.
Published: (2026)
by: Xu, Tianyi, et al.
Published: (2026)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
by: Klöser, Lars, et al.
Published: (2024)
by: Klöser, Lars, et al.
Published: (2024)
Comparing Complex Concepts with Transformers: Matching Patent Claims Against Natural Language Text
by: Blume, Matthias, et al.
Published: (2024)
by: Blume, Matthias, et al.
Published: (2024)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
by: Levine, Lauren, et al.
Published: (2024)
by: Levine, Lauren, et al.
Published: (2024)
Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
by: Orlikowski, Matthias, et al.
Published: (2025)
by: Orlikowski, Matthias, et al.
Published: (2025)
The Effect of Data Partitioning Strategy on Model Generalizability: A Case Study of Morphological Segmentation
by: Liu, Zoey, et al.
Published: (2024)
by: Liu, Zoey, et al.
Published: (2024)
A Novel Word Pair-based Gaussian Sentence Similarity Algorithm For Bengali Extractive Text Summarization
by: Morshed, Fahim, et al.
Published: (2024)
by: Morshed, Fahim, et al.
Published: (2024)
Steering Language Models in Multi-Token Generation: A Case Study on Tense and Aspect
by: Klerings, Alina, et al.
Published: (2025)
by: Klerings, Alina, et al.
Published: (2025)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
by: Akavarapu, V. S. D. S. Mahesh, et al.
Published: (2025)
by: Akavarapu, V. S. D. S. Mahesh, et al.
Published: (2025)
Multi-Hierarchical Feature Detection for Large Language Model Generated Text
by: Zhang, Luyan, et al.
Published: (2025)
by: Zhang, Luyan, et al.
Published: (2025)
Robustness of Large Language Models to Perturbations in Text
by: Singh, Ayush, et al.
Published: (2024)
by: Singh, Ayush, et al.
Published: (2024)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
by: Evelo, Bart, et al.
Published: (2026)
by: Evelo, Bart, et al.
Published: (2026)
GroundCap: A Visually Grounded Image Captioning Dataset
by: Oliveira, Daniel A. P., et al.
Published: (2025)
by: Oliveira, Daniel A. P., et al.
Published: (2025)
COMET-poly: Machine Translation Metric Grounded in Other Candidates
by: Züfle, Maike, et al.
Published: (2025)
by: Züfle, Maike, et al.
Published: (2025)
Curating Grounded Synthetic Data with Global Perspectives for Equitable AI
by: Törnquist, Elin, et al.
Published: (2024)
by: Törnquist, Elin, et al.
Published: (2024)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
by: Nzeyimana, Antoine, et al.
Published: (2025)
by: Nzeyimana, Antoine, et al.
Published: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
by: Collado-Montañez, Jaime, et al.
Published: (2025)
by: Collado-Montañez, Jaime, et al.
Published: (2025)
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering
by: Muller, Sacha, et al.
Published: (2024)
by: Muller, Sacha, et al.
Published: (2024)
Synthia: Scalable Grounded Persona Generation from Social Media Data
by: Rahimzadeh, Vahid, et al.
Published: (2025)
by: Rahimzadeh, Vahid, et al.
Published: (2025)
Fin-ExBERT: User Intent based Text Extraction in Financial Context using Graph-Augmented BERT and trainable Plugin
by: Sarker, Soumick, et al.
Published: (2025)
by: Sarker, Soumick, et al.
Published: (2025)
Text Summarization With Graph Attention Networks
by: Ardestani, Mohammadreza, et al.
Published: (2026)
by: Ardestani, Mohammadreza, et al.
Published: (2026)
Linguistic Interpretability of Transformer-based Language Models: a systematic review
by: López-Otal, Miguel, et al.
Published: (2025)
by: López-Otal, Miguel, et al.
Published: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
by: Kugler, Kai
Published: (2025)
by: Kugler, Kai
Published: (2025)
Investigating the Impact of Text Summarization on Topic Modeling
by: Khandelwal, Trishia
Published: (2024)
by: Khandelwal, Trishia
Published: (2024)
Active Few-Shot Learning for Text Classification
by: Ahmadnia, Saeed, et al.
Published: (2025)
by: Ahmadnia, Saeed, et al.
Published: (2025)
Normalization of Lithuanian Text Using Regular Expressions
by: Kasparaitis, Pijus
Published: (2023)
by: Kasparaitis, Pijus
Published: (2023)
SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions
by: Wang, Tianyu, et al.
Published: (2026)
by: Wang, Tianyu, et al.
Published: (2026)
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
by: Toker, Michael, et al.
Published: (2024)
by: Toker, Michael, et al.
Published: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
Similar Items
-
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data
by: de Mello, Guilherme Lamartine, et al.
Published: (2024) -
Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer
by: Paneru, Utsav
Published: (2026) -
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
by: Lajčinová, Bibiána, et al.
Published: (2024) -
Synthetic Voice Data for Automatic Speech Recognition in African Languages
by: DeRenzi, Brian, et al.
Published: (2025) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)