Textual-to-Visual Iterative Self-Verification for Slide Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Yunqing, Ma, Xinbei, Qiu, Jiyang, Zhao, Hai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MEGen: Generative Backdoor into Large Language Models via Model Editing
por: Qiu, Jiyang, et al.
Publicado: (2024)
por: Qiu, Jiyang, et al.
Publicado: (2024)
On the Robustness of Editing Large Language Models
por: Ma, Xinbei, et al.
Publicado: (2024)
por: Ma, Xinbei, et al.
Publicado: (2024)
Chain-of-Trigger: An Agentic Backdoor that Paradoxically Enhances Agentic Robustness
por: Qiu, Jiyang, et al.
Publicado: (2025)
por: Qiu, Jiyang, et al.
Publicado: (2025)
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
por: Ma, Xinbei, et al.
Publicado: (2024)
por: Ma, Xinbei, et al.
Publicado: (2024)
PROM: A Phrase-level Copying Mechanism with Pre-training for Abstractive Summarization
por: Ma, Xinbei, et al.
Publicado: (2023)
por: Ma, Xinbei, et al.
Publicado: (2023)
PDDLEGO: Iterative Planning in Textual Environments
por: Zhang, Li, et al.
Publicado: (2024)
por: Zhang, Li, et al.
Publicado: (2024)
How Deep is Love in LLMs' Hearts? Exploring Semantic Size in Human-like Cognition
por: Yao, Yao, et al.
Publicado: (2025)
por: Yao, Yao, et al.
Publicado: (2025)
Caution for the Environment: Multimodal LLM Agents are Susceptible to Environmental Distractions
por: Ma, Xinbei, et al.
Publicado: (2024)
por: Ma, Xinbei, et al.
Publicado: (2024)
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
por: Cao, Zouying, et al.
Publicado: (2025)
por: Cao, Zouying, et al.
Publicado: (2025)
Wide-Horizon Thinking and Simulation-Based Evaluation for Real-World LLM Planning with Multifaceted Constraints
por: Yang, Dongjie, et al.
Publicado: (2025)
por: Yang, Dongjie, et al.
Publicado: (2025)
Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
por: Wu, Weiqi, et al.
Publicado: (2025)
por: Wu, Weiqi, et al.
Publicado: (2025)
Beyond the Textual: Generating Coherent Visual Options for MCQs
por: Wang, Wanqiang, et al.
Publicado: (2025)
por: Wang, Wanqiang, et al.
Publicado: (2025)
The Automated Verification of Textual Claims (AVeriTeC) Shared Task
por: Schlichtkrull, Michael, et al.
Publicado: (2024)
por: Schlichtkrull, Michael, et al.
Publicado: (2024)
LESA: Learnable LLM Layer Scaling-Up
por: Yang, Yifei, et al.
Publicado: (2025)
por: Yang, Yifei, et al.
Publicado: (2025)
Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
por: Li, Yafu, et al.
Publicado: (2025)
por: Li, Yafu, et al.
Publicado: (2025)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
por: Han, Jiuzhou, et al.
Publicado: (2023)
por: Han, Jiuzhou, et al.
Publicado: (2023)
Unfolding Iterators: Specification and Verification of Higher-Order Iterators, in OCaml
por: Chirica, Ion, et al.
Publicado: (2025)
por: Chirica, Ion, et al.
Publicado: (2025)
Thinking-while-Generating: Interleaving Textual Reasoning throughout Visual Generation
por: Guo, Ziyu, et al.
Publicado: (2025)
por: Guo, Ziyu, et al.
Publicado: (2025)
TOOLVERIFIER: Generalization to New Tools via Self-Verification
por: Mekala, Dheeraj, et al.
Publicado: (2024)
por: Mekala, Dheeraj, et al.
Publicado: (2024)
Multi-Step Visual Reasoning with Visual Tokens Scaling and Verification
por: Bai, Tianyi, et al.
Publicado: (2025)
por: Bai, Tianyi, et al.
Publicado: (2025)
History-Guided Iterative Visual Reasoning with Self-Correction
por: Yang, Xinglong, et al.
Publicado: (2026)
por: Yang, Xinglong, et al.
Publicado: (2026)
$V_1$: Unifying Generation and Self-Verification for Parallel Reasoners
por: Singh, Harman, et al.
Publicado: (2026)
por: Singh, Harman, et al.
Publicado: (2026)
Mining the Explainability and Generalization: Fact Verification Based on Self-Instruction
por: Lu, Guangyao, et al.
Publicado: (2024)
por: Lu, Guangyao, et al.
Publicado: (2024)
CLEAR: Character Unlearning in Textual and Visual Modalities
por: Dontsov, Alexey, et al.
Publicado: (2024)
por: Dontsov, Alexey, et al.
Publicado: (2024)
LimeAttack: Local Explainable Method for Textual Hard-Label Adversarial Attack
por: Zhu, Hai, et al.
Publicado: (2023)
por: Zhu, Hai, et al.
Publicado: (2023)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
por: Yang, Yunqiao, et al.
Publicado: (2026)
por: Yang, Yunqiao, et al.
Publicado: (2026)
ISSR: Iterative Selection with Self-Review for Vocabulary Test Distractor Generation
por: Liu, Yu-Cheng, et al.
Publicado: (2025)
por: Liu, Yu-Cheng, et al.
Publicado: (2025)
Bridging Textual and Tabular Worlds for Fact Verification: A Lightweight, Attention-Based Model
por: Varnosfaderani, Shirin Dabbaghi, et al.
Publicado: (2024)
por: Varnosfaderani, Shirin Dabbaghi, et al.
Publicado: (2024)
Iterative Multilingual Spectral Attribute Erasure
por: Shao, Shun, et al.
Publicado: (2025)
por: Shao, Shun, et al.
Publicado: (2025)
GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning
por: Wei, Yanbin, et al.
Publicado: (2024)
por: Wei, Yanbin, et al.
Publicado: (2024)
Graph-Guided Textual Explanation Generation Framework
por: Yuan, Shuzhou, et al.
Publicado: (2024)
por: Yuan, Shuzhou, et al.
Publicado: (2024)
ToolGrad: Efficient Tool-use Dataset Generation with Textual "Gradients"
por: Zhou, Zhongyi, et al.
Publicado: (2025)
por: Zhou, Zhongyi, et al.
Publicado: (2025)
EvolveSearch: An Iterative Self-Evolving Search Agent
por: Zhang, Dingchu, et al.
Publicado: (2025)
por: Zhang, Dingchu, et al.
Publicado: (2025)
Textual Self-attention Network: Test-Time Preference Optimization through Textual Gradient-based Attention
por: Mo, Shibing, et al.
Publicado: (2025)
por: Mo, Shibing, et al.
Publicado: (2025)
DIVE: Diversified Iterative Self-Improvement
por: Qin, Yiwei, et al.
Publicado: (2025)
por: Qin, Yiwei, et al.
Publicado: (2025)
Triple Modality Fusion: Aligning Visual, Textual, and Graph Data with Large Language Models for Multi-Behavior Recommendations
por: Ma, Luyi, et al.
Publicado: (2024)
por: Ma, Luyi, et al.
Publicado: (2024)
Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation
por: Liang, Chen, et al.
Publicado: (2024)
por: Liang, Chen, et al.
Publicado: (2024)
Rethinking Verification for LLM Code Generation: From Generation to Testing
por: Ma, Zihan, et al.
Publicado: (2025)
por: Ma, Zihan, et al.
Publicado: (2025)
GraphNarrator: Generating Textual Explanations for Graph Neural Networks
por: Pan, Bo, et al.
Publicado: (2024)
por: Pan, Bo, et al.
Publicado: (2024)
Mismatch Quest: Visual and Textual Feedback for Image-Text Misalignment
por: Gordon, Brian, et al.
Publicado: (2023)
por: Gordon, Brian, et al.
Publicado: (2023)
Ejemplares similares
-
MEGen: Generative Backdoor into Large Language Models via Model Editing
por: Qiu, Jiyang, et al.
Publicado: (2024) -
On the Robustness of Editing Large Language Models
por: Ma, Xinbei, et al.
Publicado: (2024) -
Chain-of-Trigger: An Agentic Backdoor that Paradoxically Enhances Agentic Robustness
por: Qiu, Jiyang, et al.
Publicado: (2025) -
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
por: Ma, Xinbei, et al.
Publicado: (2024) -
PROM: A Phrase-level Copying Mechanism with Pre-training for Abstractive Summarization
por: Ma, Xinbei, et al.
Publicado: (2023)