Textual-to-Visual Iterative Self-Verification for Slide Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Yunqing, Ma, Xinbei, Qiu, Jiyang, Zhao, Hai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MEGen: Generative Backdoor into Large Language Models via Model Editing
di: Qiu, Jiyang, et al.
Pubblicazione: (2024)
di: Qiu, Jiyang, et al.
Pubblicazione: (2024)
On the Robustness of Editing Large Language Models
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
Chain-of-Trigger: An Agentic Backdoor that Paradoxically Enhances Agentic Robustness
di: Qiu, Jiyang, et al.
Pubblicazione: (2025)
di: Qiu, Jiyang, et al.
Pubblicazione: (2025)
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
PROM: A Phrase-level Copying Mechanism with Pre-training for Abstractive Summarization
di: Ma, Xinbei, et al.
Pubblicazione: (2023)
di: Ma, Xinbei, et al.
Pubblicazione: (2023)
PDDLEGO: Iterative Planning in Textual Environments
di: Zhang, Li, et al.
Pubblicazione: (2024)
di: Zhang, Li, et al.
Pubblicazione: (2024)
How Deep is Love in LLMs' Hearts? Exploring Semantic Size in Human-like Cognition
di: Yao, Yao, et al.
Pubblicazione: (2025)
di: Yao, Yao, et al.
Pubblicazione: (2025)
Caution for the Environment: Multimodal LLM Agents are Susceptible to Environmental Distractions
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
di: Ma, Xinbei, et al.
Pubblicazione: (2024)
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
di: Cao, Zouying, et al.
Pubblicazione: (2025)
di: Cao, Zouying, et al.
Pubblicazione: (2025)
Wide-Horizon Thinking and Simulation-Based Evaluation for Real-World LLM Planning with Multifaceted Constraints
di: Yang, Dongjie, et al.
Pubblicazione: (2025)
di: Yang, Dongjie, et al.
Pubblicazione: (2025)
Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
di: Wu, Weiqi, et al.
Pubblicazione: (2025)
di: Wu, Weiqi, et al.
Pubblicazione: (2025)
Beyond the Textual: Generating Coherent Visual Options for MCQs
di: Wang, Wanqiang, et al.
Pubblicazione: (2025)
di: Wang, Wanqiang, et al.
Pubblicazione: (2025)
The Automated Verification of Textual Claims (AVeriTeC) Shared Task
di: Schlichtkrull, Michael, et al.
Pubblicazione: (2024)
di: Schlichtkrull, Michael, et al.
Pubblicazione: (2024)
LESA: Learnable LLM Layer Scaling-Up
di: Yang, Yifei, et al.
Pubblicazione: (2025)
di: Yang, Yifei, et al.
Pubblicazione: (2025)
Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
di: Li, Yafu, et al.
Pubblicazione: (2025)
di: Li, Yafu, et al.
Pubblicazione: (2025)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
di: Han, Jiuzhou, et al.
Pubblicazione: (2023)
di: Han, Jiuzhou, et al.
Pubblicazione: (2023)
Unfolding Iterators: Specification and Verification of Higher-Order Iterators, in OCaml
di: Chirica, Ion, et al.
Pubblicazione: (2025)
di: Chirica, Ion, et al.
Pubblicazione: (2025)
Thinking-while-Generating: Interleaving Textual Reasoning throughout Visual Generation
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
TOOLVERIFIER: Generalization to New Tools via Self-Verification
di: Mekala, Dheeraj, et al.
Pubblicazione: (2024)
di: Mekala, Dheeraj, et al.
Pubblicazione: (2024)
Multi-Step Visual Reasoning with Visual Tokens Scaling and Verification
di: Bai, Tianyi, et al.
Pubblicazione: (2025)
di: Bai, Tianyi, et al.
Pubblicazione: (2025)
History-Guided Iterative Visual Reasoning with Self-Correction
di: Yang, Xinglong, et al.
Pubblicazione: (2026)
di: Yang, Xinglong, et al.
Pubblicazione: (2026)
$V_1$: Unifying Generation and Self-Verification for Parallel Reasoners
di: Singh, Harman, et al.
Pubblicazione: (2026)
di: Singh, Harman, et al.
Pubblicazione: (2026)
Mining the Explainability and Generalization: Fact Verification Based on Self-Instruction
di: Lu, Guangyao, et al.
Pubblicazione: (2024)
di: Lu, Guangyao, et al.
Pubblicazione: (2024)
CLEAR: Character Unlearning in Textual and Visual Modalities
di: Dontsov, Alexey, et al.
Pubblicazione: (2024)
di: Dontsov, Alexey, et al.
Pubblicazione: (2024)
LimeAttack: Local Explainable Method for Textual Hard-Label Adversarial Attack
di: Zhu, Hai, et al.
Pubblicazione: (2023)
di: Zhu, Hai, et al.
Pubblicazione: (2023)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
di: Yang, Yunqiao, et al.
Pubblicazione: (2026)
di: Yang, Yunqiao, et al.
Pubblicazione: (2026)
ISSR: Iterative Selection with Self-Review for Vocabulary Test Distractor Generation
di: Liu, Yu-Cheng, et al.
Pubblicazione: (2025)
di: Liu, Yu-Cheng, et al.
Pubblicazione: (2025)
Bridging Textual and Tabular Worlds for Fact Verification: A Lightweight, Attention-Based Model
di: Varnosfaderani, Shirin Dabbaghi, et al.
Pubblicazione: (2024)
di: Varnosfaderani, Shirin Dabbaghi, et al.
Pubblicazione: (2024)
Iterative Multilingual Spectral Attribute Erasure
di: Shao, Shun, et al.
Pubblicazione: (2025)
di: Shao, Shun, et al.
Pubblicazione: (2025)
GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning
di: Wei, Yanbin, et al.
Pubblicazione: (2024)
di: Wei, Yanbin, et al.
Pubblicazione: (2024)
Graph-Guided Textual Explanation Generation Framework
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
ToolGrad: Efficient Tool-use Dataset Generation with Textual "Gradients"
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
EvolveSearch: An Iterative Self-Evolving Search Agent
di: Zhang, Dingchu, et al.
Pubblicazione: (2025)
di: Zhang, Dingchu, et al.
Pubblicazione: (2025)
Textual Self-attention Network: Test-Time Preference Optimization through Textual Gradient-based Attention
di: Mo, Shibing, et al.
Pubblicazione: (2025)
di: Mo, Shibing, et al.
Pubblicazione: (2025)
DIVE: Diversified Iterative Self-Improvement
di: Qin, Yiwei, et al.
Pubblicazione: (2025)
di: Qin, Yiwei, et al.
Pubblicazione: (2025)
Triple Modality Fusion: Aligning Visual, Textual, and Graph Data with Large Language Models for Multi-Behavior Recommendations
di: Ma, Luyi, et al.
Pubblicazione: (2024)
di: Ma, Luyi, et al.
Pubblicazione: (2024)
Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation
di: Liang, Chen, et al.
Pubblicazione: (2024)
di: Liang, Chen, et al.
Pubblicazione: (2024)
Rethinking Verification for LLM Code Generation: From Generation to Testing
di: Ma, Zihan, et al.
Pubblicazione: (2025)
di: Ma, Zihan, et al.
Pubblicazione: (2025)
GraphNarrator: Generating Textual Explanations for Graph Neural Networks
di: Pan, Bo, et al.
Pubblicazione: (2024)
di: Pan, Bo, et al.
Pubblicazione: (2024)
Mismatch Quest: Visual and Textual Feedback for Image-Text Misalignment
di: Gordon, Brian, et al.
Pubblicazione: (2023)
di: Gordon, Brian, et al.
Pubblicazione: (2023)
Documenti analoghi
-
MEGen: Generative Backdoor into Large Language Models via Model Editing
di: Qiu, Jiyang, et al.
Pubblicazione: (2024) -
On the Robustness of Editing Large Language Models
di: Ma, Xinbei, et al.
Pubblicazione: (2024) -
Chain-of-Trigger: An Agentic Backdoor that Paradoxically Enhances Agentic Robustness
di: Qiu, Jiyang, et al.
Pubblicazione: (2025) -
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
di: Ma, Xinbei, et al.
Pubblicazione: (2024) -
PROM: A Phrase-level Copying Mechanism with Pre-training for Abstractive Summarization
di: Ma, Xinbei, et al.
Pubblicazione: (2023)