ROSA: Addressing text understanding challenges in photographs via ROtated SAmpling
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Maina, Hernán, Ivetta, Guido, Stuto, Mateo Lione, Eisenschlos, Julian Martin, Sánchez, Jorge, Benotti, Luciana |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Selectively Answering Visual Questions
par: Eisenschlos, Julian Martin, et autres
Publié: (2024)
par: Eisenschlos, Julian Martin, et autres
Publié: (2024)
Low-resource domain adaptation while minimizing energy and hardware resource consumption
par: Maina, Hernán, et autres
Publié: (2025)
par: Maina, Hernán, et autres
Publié: (2025)
HESEIA: A community-based dataset for evaluating social biases in large language models, co-designed in real school settings in Latin America
par: Ivetta, Guido, et autres
Publié: (2025)
par: Ivetta, Guido, et autres
Publié: (2025)
Faithful Chart Summarization with ChaTS-Pi
par: Krichene, Syrine, et autres
Publié: (2024)
par: Krichene, Syrine, et autres
Publié: (2024)
Towards culturally-appropriate conversational AI for health in the majority world: An exploratory study with citizens and professionals in Latin America
par: Peters, Dorian, et autres
Publié: (2025)
par: Peters, Dorian, et autres
Publié: (2025)
TANQ: An open domain dataset of table answered questions
par: Akhtar, Mubashara, et autres
Publié: (2024)
par: Akhtar, Mubashara, et autres
Publié: (2024)
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
par: Ivetta, Guido, et autres
Publié: (2025)
par: Ivetta, Guido, et autres
Publié: (2025)
WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts
par: Foroutan, Negar, et autres
Publié: (2025)
par: Foroutan, Negar, et autres
Publié: (2025)
Code-switching in text and speech challenges information-theoretic speaker design
par: Bhattacharya, Debasmita, et autres
Publié: (2024)
par: Bhattacharya, Debasmita, et autres
Publié: (2024)
ROSA-Tuning: Enhancing Long-Context Modeling via Suffix Matching
par: Zheng, Yunao, et autres
Publié: (2026)
par: Zheng, Yunao, et autres
Publié: (2026)
Revisiting Real-Time Digging-In Effects: No Evidence from NP/Z Garden-Paths
par: Maina-Kilaas, Amani, et autres
Publié: (2026)
par: Maina-Kilaas, Amani, et autres
Publié: (2026)
A quantitative analysis of semantic information in deep representations of text and images
par: Acevedo, Santiago, et autres
Publié: (2025)
par: Acevedo, Santiago, et autres
Publié: (2025)
PAGE: Prompt Augmentation for text Generation Enhancement
par: Pacchiotti, Mauro Jose, et autres
Publié: (2025)
par: Pacchiotti, Mauro Jose, et autres
Publié: (2025)
Movie2Story: A framework for understanding videos and telling stories in the form of novel text
par: Li, Kangning, et autres
Publié: (2024)
par: Li, Kangning, et autres
Publié: (2024)
Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding
par: Wang, Zilong, et autres
Publié: (2024)
par: Wang, Zilong, et autres
Publié: (2024)
Advancing Chinese biomedical text mining with community challenges
par: Zong, Hui, et autres
Publié: (2024)
par: Zong, Hui, et autres
Publié: (2024)
Synthetically generated text for supervised text analysis
par: Halterman, Andrew
Publié: (2023)
par: Halterman, Andrew
Publié: (2023)
Synthetic social data: trials and tribulations
par: Ivetta, Guido, et autres
Publié: (2025)
par: Ivetta, Guido, et autres
Publié: (2025)
$\left|\,\circlearrowright\,\boxed{\text{BUS}}\,\right|$: A Large and Diverse Multimodal Benchmark for evaluating the ability of Vision-Language Models to understand Rebus Puzzles
par: Das, Trishanu, et autres
Publié: (2025)
par: Das, Trishanu, et autres
Publié: (2025)
Interpretable and Robust Dialogue State Tracking via Natural Language Summarization with LLMs
par: Carranza, Rafael, et autres
Publié: (2025)
par: Carranza, Rafael, et autres
Publié: (2025)
Causality extraction from medical text using Large Language Models (LLMs)
par: Gopalakrishnan, Seethalakshmi, et autres
Publié: (2024)
par: Gopalakrishnan, Seethalakshmi, et autres
Publié: (2024)
Altogether: Image Captioning via Re-aligning Alt-text
par: Xu, Hu, et autres
Publié: (2024)
par: Xu, Hu, et autres
Publié: (2024)
Identifying attributions of causality in political text
par: Garcia-Corral, Paulina
Publié: (2025)
par: Garcia-Corral, Paulina
Publié: (2025)
How do we measure privacy in text? A survey of text anonymization metrics
par: Ren, Yaxuan, et autres
Publié: (2025)
par: Ren, Yaxuan, et autres
Publié: (2025)
Designing large language model prompts to extract scores from messy text: A shared dataset and challenge
par: Thelwall, Mike
Publié: (2026)
par: Thelwall, Mike
Publié: (2026)
OARelatedWork: A Large-Scale Dataset of Related Work Sections with Full-texts from Open Access Sources
par: Docekal, Martin, et autres
Publié: (2024)
par: Docekal, Martin, et autres
Publié: (2024)
TALENT: Table VQA via Augmented Language-Enhanced Natural-text Transcription
par: Yutong, Guo, et autres
Publié: (2025)
par: Yutong, Guo, et autres
Publié: (2025)
Don't Lose Focus: Activation Steering via Key-Orthogonal Projections
par: Luo, Haoyan, et autres
Publié: (2026)
par: Luo, Haoyan, et autres
Publié: (2026)
Learning Phonotactics from Linguistic Informants
par: Breiss, Canaan, et autres
Publié: (2024)
par: Breiss, Canaan, et autres
Publié: (2024)
Identifying social isolation themes in NVDRS text narratives using topic modeling and text-classification methods
par: Walker, Drew, et autres
Publié: (2025)
par: Walker, Drew, et autres
Publié: (2025)
Qwen it detect machine-generated text?
par: Marchitan, Teodor-George, et autres
Publié: (2025)
par: Marchitan, Teodor-George, et autres
Publié: (2025)
Algorithmic Consequences of Particle Filters for Sentence Processing: Amplified Garden-Paths and Digging-In Effects
par: Maina-Kilaas, Amani, et autres
Publié: (2026)
par: Maina-Kilaas, Amani, et autres
Publié: (2026)
What does it mean to understand language?
par: Casto, Colton, et autres
Publié: (2025)
par: Casto, Colton, et autres
Publié: (2025)
Comparing energy consumption and accuracy in text classification inference
par: Zschache, Johannes, et autres
Publié: (2025)
par: Zschache, Johannes, et autres
Publié: (2025)
CPO: Addressing Reward Ambiguity in Role-playing Dialogue via Comparative Policy Optimization
par: Ye, Xinge, et autres
Publié: (2025)
par: Ye, Xinge, et autres
Publié: (2025)
LUQ: Long-text Uncertainty Quantification for LLMs
par: Zhang, Caiqi, et autres
Publié: (2024)
par: Zhang, Caiqi, et autres
Publié: (2024)
Few-shot text-based emotion detection
par: Marchitan, Teodor-George, et autres
Publié: (2025)
par: Marchitan, Teodor-George, et autres
Publié: (2025)
Transferable text data distillation by trajectory matching
par: Yao, Rong, et autres
Publié: (2025)
par: Yao, Rong, et autres
Publié: (2025)
Leveraging the power of transformers for guilt detection in text
par: Meque, Abdul Gafar Manuel, et autres
Publié: (2024)
par: Meque, Abdul Gafar Manuel, et autres
Publié: (2024)
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts
par: Gokceoglu, Gokcen, et autres
Publié: (2024)
par: Gokceoglu, Gokcen, et autres
Publié: (2024)
Documents similaires
-
Selectively Answering Visual Questions
par: Eisenschlos, Julian Martin, et autres
Publié: (2024) -
Low-resource domain adaptation while minimizing energy and hardware resource consumption
par: Maina, Hernán, et autres
Publié: (2025) -
HESEIA: A community-based dataset for evaluating social biases in large language models, co-designed in real school settings in Latin America
par: Ivetta, Guido, et autres
Publié: (2025) -
Faithful Chart Summarization with ChaTS-Pi
par: Krichene, Syrine, et autres
Publié: (2024) -
Towards culturally-appropriate conversational AI for health in the majority world: An exploratory study with citizens and professionals in Latin America
par: Peters, Dorian, et autres
Publié: (2025)