Co-Evolving Agents: Learning from Failures as Hard Negatives
Fuente:
arXiv
Saved in:
| Main Authors: | Jung, Yeonsung, Padhi, Trilok, Shaham, Sina, Khullar, Dipika, Jeong, Joonhyun, Mehrabi, Ninareh, Yang, Eunho |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Playing the Fool: Jailbreaking LLMs and Multimodal LLMs with Out-of-Distribution Strategy
by: Jeong, Joonhyun, et al.
Published: (2025)
by: Jeong, Joonhyun, et al.
Published: (2025)
Hijacking Context in Large Multi-modal Models
by: Jeong, Joonhyun
Published: (2023)
by: Jeong, Joonhyun
Published: (2023)
PruNeRF: Segment-Centric Dataset Pruning via 3D Spatial Consistency
by: Jung, Yeonsung, et al.
Published: (2024)
by: Jung, Yeonsung, et al.
Published: (2024)
Kaleidoscopic Teaming in Multi Agent Simulations
by: Mehrabi, Ninareh, et al.
Published: (2025)
by: Mehrabi, Ninareh, et al.
Published: (2025)
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning
by: Chen, Si, et al.
Published: (2025)
by: Chen, Si, et al.
Published: (2025)
Self-Attribution Bias: When AI Monitors Go Easy on Themselves
by: Khullar, Dipika, et al.
Published: (2026)
by: Khullar, Dipika, et al.
Published: (2026)
FERRET: Framework for Expansion Reliant Red Teaming
by: Mehrabi, Ninareh, et al.
Published: (2026)
by: Mehrabi, Ninareh, et al.
Published: (2026)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
by: Parekh, Tanmay, et al.
Published: (2025)
by: Parekh, Tanmay, et al.
Published: (2025)
Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks
by: Padhi, Trilok, et al.
Published: (2025)
by: Padhi, Trilok, et al.
Published: (2025)
Preserve or Modify? Context-Aware Evaluation for Balancing Preservation and Modification in Text-Guided Image Editing
by: Kim, Yoonjeon, et al.
Published: (2024)
by: Kim, Yoonjeon, et al.
Published: (2024)
Asking Back: Interaction-Layer Antidistillation Watermarks
by: Yang, Guang, et al.
Published: (2026)
by: Yang, Guang, et al.
Published: (2026)
LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding
by: Jang, Doohyuk, et al.
Published: (2024)
by: Jang, Doohyuk, et al.
Published: (2024)
Playing Devil's Advocate: Unmasking Toxicity and Vulnerabilities in Large Vision-Language Models
by: Erol, Abdulkadir, et al.
Published: (2025)
by: Erol, Abdulkadir, et al.
Published: (2025)
Just KIDDIN: Knowledge Infusion and Distillation for Detection of INdecent Memes
by: Garg, Rahul, et al.
Published: (2024)
by: Garg, Rahul, et al.
Published: (2024)
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
by: Kumarage, Tharindu, et al.
Published: (2025)
by: Kumarage, Tharindu, et al.
Published: (2025)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
by: Li, Huihan, et al.
Published: (2025)
by: Li, Huihan, et al.
Published: (2025)
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Tree-of-Traversals: A Zero-Shot Reasoning Algorithm for Augmenting Black-box Language Models with Knowledge Graphs
by: Markowitz, Elan, et al.
Published: (2024)
by: Markowitz, Elan, et al.
Published: (2024)
K-Edit: Language Model Editing with Contextual Knowledge Awareness
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning
by: Li, Zhicong, et al.
Published: (2026)
by: Li, Zhicong, et al.
Published: (2026)
Momentum Contrastive Learning with Enhanced Negative Sampling and Hard Negative Filtering
by: Hoang, Duy, et al.
Published: (2025)
by: Hoang, Duy, et al.
Published: (2025)
Differentially Private Publication of Electricity Time Series Data in Smart Grids
by: Shaham, Sina, et al.
Published: (2024)
by: Shaham, Sina, et al.
Published: (2024)
LFQ: Logit-aware Final-block Quantization for Boosting the Generation Quality of Low-Bit Quantized LLMs
by: Lee, Jung Hyun, et al.
Published: (2026)
by: Lee, Jung Hyun, et al.
Published: (2026)
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data
by: Kursuncu, Ugur, et al.
Published: (2025)
by: Kursuncu, Ugur, et al.
Published: (2025)
Turning Sand to Gold: Recycling Data to Bridge On-Policy and Off-Policy Learning via Causal Bound
by: Fiskus, Tal, et al.
Published: (2025)
by: Fiskus, Tal, et al.
Published: (2025)
CODESKILL: Learning Self-Evolving Skills for Coding Agents
by: Li, Yanzhou, et al.
Published: (2026)
by: Li, Yanzhou, et al.
Published: (2026)
FLIRT: Feedback Loop In-context Red Teaming
by: Mehrabi, Ninareh, et al.
Published: (2023)
by: Mehrabi, Ninareh, et al.
Published: (2023)
SynCo: Synthetic Hard Negatives for Contrastive Visual Representation Learning
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification
by: Zhang, Hanrong, et al.
Published: (2026)
by: Zhang, Hanrong, et al.
Published: (2026)
Multi-Agent Evolve: LLM Self-Improve through Co-evolution
by: Chen, Yixing, et al.
Published: (2025)
by: Chen, Yixing, et al.
Published: (2025)
DeepFact: Co-Evolving Benchmarks and Agents for Deep Research Factuality
by: Huang, Yukun, et al.
Published: (2026)
by: Huang, Yukun, et al.
Published: (2026)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
by: Liu, Youwei, et al.
Published: (2026)
by: Liu, Youwei, et al.
Published: (2026)
HNCSE: Advancing Sentence Embeddings via Hybrid Contrastive Learning with Hard Negatives
by: Liu, Wenxiao, et al.
Published: (2024)
by: Liu, Wenxiao, et al.
Published: (2024)
Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning
by: Padhi, Trilok, et al.
Published: (2024)
by: Padhi, Trilok, et al.
Published: (2024)
Format Inertia: A Failure Mechanism of LLMs in Medical Pre-Consultation
by: Lim, Seungseop, et al.
Published: (2025)
by: Lim, Seungseop, et al.
Published: (2025)
CoMAS: Co-Evolving Multi-Agent Systems via Interaction Rewards
by: Xue, Xiangyuan, et al.
Published: (2025)
by: Xue, Xiangyuan, et al.
Published: (2025)
Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Tasks
by: Wu, Xiyang, et al.
Published: (2026)
by: Wu, Xiyang, et al.
Published: (2026)
Amorphous Fortress Online: Collaboratively Designing Open-Ended Multi-Agent AI and Game Environments
by: Charity, M, et al.
Published: (2025)
by: Charity, M, et al.
Published: (2025)
Towards Reliable Test-Time Adaptation: Style Invariance as a Correctness Likelihood
by: Nam, Gilhyun, et al.
Published: (2025)
by: Nam, Gilhyun, et al.
Published: (2025)
PsychAgent: An Experience-Driven Lifelong Learning Agent for Self-Evolving Psychological Counselor
by: Yang, Yutao, et al.
Published: (2026)
by: Yang, Yutao, et al.
Published: (2026)
Similar Items
-
Playing the Fool: Jailbreaking LLMs and Multimodal LLMs with Out-of-Distribution Strategy
by: Jeong, Joonhyun, et al.
Published: (2025) -
Hijacking Context in Large Multi-modal Models
by: Jeong, Joonhyun
Published: (2023) -
PruNeRF: Segment-Centric Dataset Pruning via 3D Spatial Consistency
by: Jung, Yeonsung, et al.
Published: (2024) -
Kaleidoscopic Teaming in Multi Agent Simulations
by: Mehrabi, Ninareh, et al.
Published: (2025) -
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning
by: Chen, Si, et al.
Published: (2025)