AutoPresent: Designing Structured Visuals from Scratch
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ge, Jiaxin, Wang, Zora Zhiruo, Zhou, Xuhui, Peng, Yi-Hao, Subramanian, Sanjay, Tan, Qinyue, Sap, Maarten, Suhr, Alane, Fried, Daniel, Neubig, Graham, Darrell, Trevor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inducing Programmatic Skills for Agentic Tasks
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
Agent Workflow Memory
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
Recursive Visual Programming
von: Ge, Jiaxin, et al.
Veröffentlicht: (2023)
von: Ge, Jiaxin, et al.
Veröffentlicht: (2023)
TOM-SWE: User Mental Modeling For Software Engineering Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
Using Language Models to Disambiguate Lexical Choices in Translation
von: Barua, Josh, et al.
Veröffentlicht: (2024)
von: Barua, Josh, et al.
Veröffentlicht: (2024)
OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
Benchmarking Failures in Tool-Augmented Language Models
von: Treviño, Eduardo, et al.
Veröffentlicht: (2025)
von: Treviño, Eduardo, et al.
Veröffentlicht: (2025)
What Are Tools Anyway? A Survey from the Language Model Perspective
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
CodeRAG-Bench: Can Retrieval Augment Code Generation?
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
RAGGED: Towards Informed Design of Scalable and Stable RAG Systems
von: Hsia, Jennifer, et al.
Veröffentlicht: (2024)
von: Hsia, Jennifer, et al.
Veröffentlicht: (2024)
ECCO: Can We Improve Model-Generated Code Efficiency Without Sacrificing Functional Correctness?
von: Waghjale, Siddhant, et al.
Veröffentlicht: (2024)
von: Waghjale, Siddhant, et al.
Veröffentlicht: (2024)
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
Stereotype or Personalization? User Identity Biases Chatbot Recommendations
von: Kantharuban, Anjali, et al.
Veröffentlicht: (2024)
von: Kantharuban, Anjali, et al.
Veröffentlicht: (2024)
Training Software Engineering Agents and Verifiers with SWE-Gym
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning
von: Yin, Shaofeng, et al.
Veröffentlicht: (2026)
von: Yin, Shaofeng, et al.
Veröffentlicht: (2026)
Visually Prompted Benchmarks Are Surprisingly Fragile
von: Feng, Haiwen, et al.
Veröffentlicht: (2025)
von: Feng, Haiwen, et al.
Veröffentlicht: (2025)
Grounding Language in Multi-Perspective Referential Communication
von: Tang, Zineng, et al.
Veröffentlicht: (2024)
von: Tang, Zineng, et al.
Veröffentlicht: (2024)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
von: Huq, Faria, et al.
Veröffentlicht: (2025)
von: Huq, Faria, et al.
Veröffentlicht: (2025)
Training Proactive and Personalized LLM Agents
von: Sun, Weiwei, et al.
Veröffentlicht: (2025)
von: Sun, Weiwei, et al.
Veröffentlicht: (2025)
Evaluating Model Perception of Color Illusions in Photorealistic Scenes
von: Mao, Lingjun, et al.
Veröffentlicht: (2024)
von: Mao, Lingjun, et al.
Veröffentlicht: (2024)
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2023)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2023)
Is the Pope Catholic? Yes, the Pope is Catholic. Generative Evaluation of Non-Literal Intent Resolution in LLMs
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
TraveLER: A Modular Multi-LMM Agent Framework for Video Question-Answering
von: Shang, Chuyi, et al.
Veröffentlicht: (2024)
von: Shang, Chuyi, et al.
Veröffentlicht: (2024)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
von: Sclar, Melanie, et al.
Veröffentlicht: (2023)
von: Sclar, Melanie, et al.
Veröffentlicht: (2023)
Learning Adaptive Parallel Reasoning with Language Models
von: Pan, Jiayi, et al.
Veröffentlicht: (2025)
von: Pan, Jiayi, et al.
Veröffentlicht: (2025)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
Pose Priors from Language Models
von: Subramanian, Sanjay, et al.
Veröffentlicht: (2024)
von: Subramanian, Sanjay, et al.
Veröffentlicht: (2024)
Modeling Distinct Human Interaction in Web Agents
von: Huq, Faria, et al.
Veröffentlicht: (2026)
von: Huq, Faria, et al.
Veröffentlicht: (2026)
TULIP: Towards Unified Language-Image Pretraining
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
von: Fan, Xianzhe, et al.
Veröffentlicht: (2024)
von: Fan, Xianzhe, et al.
Veröffentlicht: (2024)
SOTOPIA-$π$: Interactive Learning of Socially Intelligent Language Agents
von: Wang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiyi, et al.
Veröffentlicht: (2024)
Long Chain-of-Thought Reasoning Across Languages
von: Barua, Josh, et al.
Veröffentlicht: (2025)
von: Barua, Josh, et al.
Veröffentlicht: (2025)
Training Task Experts through Retrieval Based Distillation
von: Ge, Jiaxin, et al.
Veröffentlicht: (2024)
von: Ge, Jiaxin, et al.
Veröffentlicht: (2024)
Is this the real life? Is this just fantasy? The Misleading Success of Simulating Social Interactions With LLMs
von: Zhou, Xuhui, et al.
Veröffentlicht: (2024)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2024)
GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses
von: Mun, Jimin, et al.
Veröffentlicht: (2026)
von: Mun, Jimin, et al.
Veröffentlicht: (2026)
Rethinking Theory of Mind Benchmarks for LLMs: Towards A User-Centered Perspective
von: Wang, Qiaosi, et al.
Veröffentlicht: (2025)
von: Wang, Qiaosi, et al.
Veröffentlicht: (2025)
SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
von: Fan, Xianzhe, et al.
Veröffentlicht: (2025)
von: Fan, Xianzhe, et al.
Veröffentlicht: (2025)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
von: Jain, Devansh, et al.
Veröffentlicht: (2024)
von: Jain, Devansh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Inducing Programmatic Skills for Agentic Tasks
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025) -
Agent Workflow Memory
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024) -
Recursive Visual Programming
von: Ge, Jiaxin, et al.
Veröffentlicht: (2023) -
TOM-SWE: User Mental Modeling For Software Engineering Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025) -
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)