VIGOR+: Iterative Confounder Generation and Validation via LLM-CEVAE Feedback Loop
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhu, JiaWei, Liu, ZiHeng |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Applying Cognitive Design Patterns to General LLM Agents
par: Wray, Robert E., et autres
Publié: (2025)
par: Wray, Robert E., et autres
Publié: (2025)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
par: Annapureddy, Sasank
Publié: (2026)
par: Annapureddy, Sasank
Publié: (2026)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
par: Ray, Aninda
Publié: (2026)
par: Ray, Aninda
Publié: (2026)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
par: Wang, Xiaohua, et autres
Publié: (2026)
par: Wang, Xiaohua, et autres
Publié: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
par: Wang, Yuchen, et autres
Publié: (2026)
par: Wang, Yuchen, et autres
Publié: (2026)
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
par: Wang, Zixu, et autres
Publié: (2026)
par: Wang, Zixu, et autres
Publié: (2026)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
par: Chang, Jiale, et autres
Publié: (2026)
par: Chang, Jiale, et autres
Publié: (2026)
Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay
par: Wang, Xiaohua, et autres
Publié: (2026)
par: Wang, Xiaohua, et autres
Publié: (2026)
CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
par: Lorenzoni, Giuliano, et autres
Publié: (2026)
par: Lorenzoni, Giuliano, et autres
Publié: (2026)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
par: Yan, Lingyong, et autres
Publié: (2026)
par: Yan, Lingyong, et autres
Publié: (2026)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
par: Song, Seyoung
Publié: (2025)
par: Song, Seyoung
Publié: (2025)
Post Hoc Extraction of Pareto Fronts for Continuous Control
par: Thakar, Raghav, et autres
Publié: (2026)
par: Thakar, Raghav, et autres
Publié: (2026)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
par: Sidik, Bronislav, et autres
Publié: (2026)
par: Sidik, Bronislav, et autres
Publié: (2026)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
par: Li, Bowen, et autres
Publié: (2026)
par: Li, Bowen, et autres
Publié: (2026)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
par: Kohl, Jens, et autres
Publié: (2024)
par: Kohl, Jens, et autres
Publié: (2024)
Understanding Multi-Agent LLM Frameworks: A Unified Benchmark and Experimental Analysis
par: Orogat, Abdelghny, et autres
Publié: (2026)
par: Orogat, Abdelghny, et autres
Publié: (2026)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
par: Ueda, Keisuke, et autres
Publié: (2025)
par: Ueda, Keisuke, et autres
Publié: (2025)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
par: Figueiredo, Vanessa
Publié: (2025)
par: Figueiredo, Vanessa
Publié: (2025)
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
par: Resh, William G., et autres
Publié: (2025)
par: Resh, William G., et autres
Publié: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
par: Costa, Rimom
Publié: (2025)
par: Costa, Rimom
Publié: (2025)
Latent Cache Flow: Model-to-Model Communication Without Text
par: Rossi, Maximillian, et autres
Publié: (2026)
par: Rossi, Maximillian, et autres
Publié: (2026)
Eliciting Problem Specifications via Large Language Models
par: Wray, Robert E., et autres
Publié: (2024)
par: Wray, Robert E., et autres
Publié: (2024)
Agent WARPP: Workflow Adherence via Runtime Parallel Personalization
par: Mazzolenis, Maria Emilia, et autres
Publié: (2025)
par: Mazzolenis, Maria Emilia, et autres
Publié: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
par: Cohen, Philip R., et autres
Publié: (2023)
par: Cohen, Philip R., et autres
Publié: (2023)
Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment
par: Jabal, Mohamed Sobhi, et autres
Publié: (2026)
par: Jabal, Mohamed Sobhi, et autres
Publié: (2026)
MMiC: Mitigating Modality Incompleteness in Clustered Federated Learning
par: Yang, Lishan, et autres
Publié: (2025)
par: Yang, Lishan, et autres
Publié: (2025)
FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints
par: Yang, Lishan, et autres
Publié: (2025)
par: Yang, Lishan, et autres
Publié: (2025)
NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation
par: Bose, Joy
Publié: (2026)
par: Bose, Joy
Publié: (2026)
Decoding Fake Narratives in Spreading Hateful Stories: A Dual-Head RoBERTa Model with Multi-Task Learning
par: Bhaskar, Yash, et autres
Publié: (2025)
par: Bhaskar, Yash, et autres
Publié: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
par: Wu, Shuai, et autres
Publié: (2026)
par: Wu, Shuai, et autres
Publié: (2026)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
par: Trooskens, Geert, et autres
Publié: (2026)
par: Trooskens, Geert, et autres
Publié: (2026)
A Simple Architecture for Enterprise Large Language Model Applications based on Role based security and Clearance Levels using Retrieval-Augmented Generation or Mixture of Experts
par: Özgür, Atilla, et autres
Publié: (2024)
par: Özgür, Atilla, et autres
Publié: (2024)
PAVE: A Cognitive Architecture for Legitimate Violation in Generative Agent Societies
par: Yehia, Ahmad, et autres
Publié: (2026)
par: Yehia, Ahmad, et autres
Publié: (2026)
CRAwDAD: Causal Reasoning Augmentation with Dual-Agent Debate
par: Vamosi, Finn G., et autres
Publié: (2025)
par: Vamosi, Finn G., et autres
Publié: (2025)
GSAR: Typed Grounding for Hallucination Detection and Recovery in Multi-Agent LLMs
par: Kamelhar, Federico A.
Publié: (2026)
par: Kamelhar, Federico A.
Publié: (2026)
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
par: Gokdemir, Ozan, et autres
Publié: (2025)
par: Gokdemir, Ozan, et autres
Publié: (2025)
A Scalable Communication Protocol for Networks of Large Language Models
par: Marro, Samuele, et autres
Publié: (2024)
par: Marro, Samuele, et autres
Publié: (2024)
Contrastive Learning-Enhanced Large Language Models for Monolith-to-Microservice Decomposition
par: Sellami, Khaled, et autres
Publié: (2025)
par: Sellami, Khaled, et autres
Publié: (2025)
MALTopic: Multi-Agent LLM Topic Modeling Framework
par: Sharma, Yash
Publié: (2026)
par: Sharma, Yash
Publié: (2026)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
par: Zhang, Ke, et autres
Publié: (2025)
par: Zhang, Ke, et autres
Publié: (2025)
Documents similaires
-
Applying Cognitive Design Patterns to General LLM Agents
par: Wray, Robert E., et autres
Publié: (2025) -
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
par: Annapureddy, Sasank
Publié: (2026) -
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
par: Ray, Aninda
Publié: (2026) -
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
par: Wang, Xiaohua, et autres
Publié: (2026) -
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
par: Wang, Yuchen, et autres
Publié: (2026)