Recursive Think-Answer Process for LLMs and VLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Byung-Kwan, Chee, Youngchae, Ro, Yong Man |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TroL: Traversal of Layers for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
Causal Unsupervised Semantic Segmentation
di: Kim, Junho, et al.
Pubblicazione: (2023)
di: Kim, Junho, et al.
Pubblicazione: (2023)
What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-modal Models
di: Kim, Junho, et al.
Pubblicazione: (2024)
di: Kim, Junho, et al.
Pubblicazione: (2024)
Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking
di: Chung, Sangyun, et al.
Pubblicazione: (2024)
di: Chung, Sangyun, et al.
Pubblicazione: (2024)
Hide to See: Reasoning-prefix Masking for Visual-anchored Thinking in VLM Distillation
di: Yu, Seonghoon, et al.
Pubblicazione: (2026)
di: Yu, Seonghoon, et al.
Pubblicazione: (2026)
P-CoT: A Pedagogically-motivated Participatory Chain-of-Thought Prompting for Phonological Reasoning in LLMs
di: Jang, Dongjun, et al.
Pubblicazione: (2025)
di: Jang, Dongjun, et al.
Pubblicazione: (2025)
SPARK: Multi-Vision Sensor Perception and Reasoning Benchmark for Large-scale Vision-Language Models
di: Yu, Youngjoon, et al.
Pubblicazione: (2024)
di: Yu, Youngjoon, et al.
Pubblicazione: (2024)
Unlocking Recursive Thinking of LLMs: Alignment via Refinement
di: Zhang, Haoke, et al.
Pubblicazione: (2025)
di: Zhang, Haoke, et al.
Pubblicazione: (2025)
MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models
di: Chung, Sangyun, et al.
Pubblicazione: (2026)
di: Chung, Sangyun, et al.
Pubblicazione: (2026)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
CoLLaVO: Crayon Large Language and Vision mOdel
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
MoAI: Mixture of All Intelligence for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
THREAD: Thinking Deeper with Recursive Spawning
di: Schroeder, Philip, et al.
Pubblicazione: (2024)
di: Schroeder, Philip, et al.
Pubblicazione: (2024)
Acquisition of Recursive Possessives and Recursive Locatives in Mandarin
di: Fu, Chenxi, et al.
Pubblicazione: (2024)
di: Fu, Chenxi, et al.
Pubblicazione: (2024)
RCScore: Quantifying Response Consistency in Large Language Models
di: Jang, Dongjun, et al.
Pubblicazione: (2025)
di: Jang, Dongjun, et al.
Pubblicazione: (2025)
Where Visual Speech Meets Language: VSP-LLM Framework for Efficient and Context-Aware Visual Speech Processing
di: Yeo, Jeong Hun, et al.
Pubblicazione: (2024)
di: Yeo, Jeong Hun, et al.
Pubblicazione: (2024)
Phantom of Latent for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
Think, But Don't Overthink: Reproducing Recursive Language Models
di: Wang, Daren
Pubblicazione: (2026)
di: Wang, Daren
Pubblicazione: (2026)
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
di: Yu, Zhuoran, et al.
Pubblicazione: (2025)
di: Yu, Zhuoran, et al.
Pubblicazione: (2025)
Persona Extraction Through Semantic Similarity for Emotional Support Conversation Generation
di: Han, Seunghee, et al.
Pubblicazione: (2024)
di: Han, Seunghee, et al.
Pubblicazione: (2024)
Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison
di: Yang, Qian, et al.
Pubblicazione: (2024)
di: Yang, Qian, et al.
Pubblicazione: (2024)
Think in Parallel, Answer as One: Logit Averaging for Open-Ended Reasoning
di: Wang, Haonan, et al.
Pubblicazione: (2025)
di: Wang, Haonan, et al.
Pubblicazione: (2025)
Think Together and Work Better: Combining Humans' and LLMs' Think-Aloud Outcomes for Effective Text Evaluation
di: Chu, SeongYeub, et al.
Pubblicazione: (2024)
di: Chu, SeongYeub, et al.
Pubblicazione: (2024)
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
di: Chen, Haoyang, et al.
Pubblicazione: (2026)
di: Chen, Haoyang, et al.
Pubblicazione: (2026)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
di: Zhang, Yue, et al.
Pubblicazione: (2026)
di: Zhang, Yue, et al.
Pubblicazione: (2026)
Test-time Recursive Thinking: Self-Improvement without External Feedback
di: Zhuang, Yufan, et al.
Pubblicazione: (2026)
di: Zhuang, Yufan, et al.
Pubblicazione: (2026)
Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs
di: Carbune, Victor, et al.
Pubblicazione: (2024)
di: Carbune, Victor, et al.
Pubblicazione: (2024)
Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
LLMs Provide Unstable Answers to Legal Questions
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation
di: Kim, Minsu, et al.
Pubblicazione: (2023)
di: Kim, Minsu, et al.
Pubblicazione: (2023)
For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs
di: Deng, Wenlong, et al.
Pubblicazione: (2025)
di: Deng, Wenlong, et al.
Pubblicazione: (2025)
Confidence Estimation in Automatic Short Answer Grading with LLMs
di: Cong, Longwei, et al.
Pubblicazione: (2026)
di: Cong, Longwei, et al.
Pubblicazione: (2026)
VLsI: Verbalized Layers-to-Interactions from Large to Small Vision Language Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
Unified Reinforcement and Imitation Learning for Vision-Language Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
Gold-Switch: Training-Free Superposition of Slow- and Fast- Thinking LLMs
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
Prompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech Recognition
di: Kim, Minsu, et al.
Pubblicazione: (2023)
di: Kim, Minsu, et al.
Pubblicazione: (2023)
Fast, Slow, and Tool-augmented Thinking for LLMs: A Review
di: Jia, Xinda, et al.
Pubblicazione: (2025)
di: Jia, Xinda, et al.
Pubblicazione: (2025)
MMAFFBen: A Multilingual and Multimodal Affective Analysis Benchmark for Evaluating LLMs and VLMs
di: Liu, Zhiwei, et al.
Pubblicazione: (2025)
di: Liu, Zhiwei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TroL: Traversal of Layers for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024) -
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025) -
Causal Unsupervised Semantic Segmentation
di: Kim, Junho, et al.
Pubblicazione: (2023) -
What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-modal Models
di: Kim, Junho, et al.
Pubblicazione: (2024) -
Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking
di: Chung, Sangyun, et al.
Pubblicazione: (2024)