Bot Wars Evolved: Orchestrating Competing LLMs in a Counterstrike Against Phone Scams
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Basta, Nardine, Atkins, Conor, Kaafar, Dali |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DIALEVAL: Automated Type-Theoretic Evaluation of LLM Instruction Following
von: Basta, Nardine, et al.
Veröffentlicht: (2026)
von: Basta, Nardine, et al.
Veröffentlicht: (2026)
ConvoCache: Smart Re-Use of Chatbot Responses
von: Atkins, Conor, et al.
Veröffentlicht: (2024)
von: Atkins, Conor, et al.
Veröffentlicht: (2024)
ASRJam: Human-Friendly AI Speech Jamming to Prevent Automated Phone Scams
von: Grabovski, Freddie, et al.
Veröffentlicht: (2025)
von: Grabovski, Freddie, et al.
Veröffentlicht: (2025)
Orchestrating LLMs with Different Personalizations
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution
von: Maheswaran, Monishwaran, et al.
Veröffentlicht: (2026)
von: Maheswaran, Monishwaran, et al.
Veröffentlicht: (2026)
Multi-Agent Collaboration via Evolving Orchestration
von: Dang, Yufan, et al.
Veröffentlicht: (2025)
von: Dang, Yufan, et al.
Veröffentlicht: (2025)
A Large-Scale Empirical Analysis of Custom GPTs' Vulnerabilities in the OpenAI Ecosystem
von: Ogundoyin, Sunday Oyinlola, et al.
Veröffentlicht: (2025)
von: Ogundoyin, Sunday Oyinlola, et al.
Veröffentlicht: (2025)
Pro-ZD: A Transferable Graph Neural Network Approach for Proactive Zero-Day Threats Mitigation
von: Basta, Nardine, et al.
Veröffentlicht: (2026)
von: Basta, Nardine, et al.
Veröffentlicht: (2026)
PhoneWorld: Scaling Phone-Use Agent Environments
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
Self-Evolved Reward Learning for LLMs
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
NOTAM-Evolve: A Knowledge-Guided Self-Evolving Optimization Framework with LLMs for NOTAM Interpretation
von: Liu, Maoqi, et al.
Veröffentlicht: (2025)
von: Liu, Maoqi, et al.
Veröffentlicht: (2025)
ScamAgents: How AI Agents Can Simulate Human-Level Scam Calls
von: Badhe, Sanket
Veröffentlicht: (2025)
von: Badhe, Sanket
Veröffentlicht: (2025)
When Claims Evolve: Evaluating and Enhancing the Robustness of Embedding Models Against Misinformation Edits
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
ContiGuard: A Framework for Continual Toxicity Detection Against Evolving Evasive Perturbations
von: Kang, Hankun, et al.
Veröffentlicht: (2026)
von: Kang, Hankun, et al.
Veröffentlicht: (2026)
When Facts Change: Probing LLMs on Evolving Knowledge with evolveQA
von: Nakshatri, Nishanth Sridhar, et al.
Veröffentlicht: (2025)
von: Nakshatri, Nishanth Sridhar, et al.
Veröffentlicht: (2025)
CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks
von: Lin, Peiqin, et al.
Veröffentlicht: (2026)
von: Lin, Peiqin, et al.
Veröffentlicht: (2026)
Conversation for Non-verifiable Learning: Self-Evolving LLMs through Meta-Evaluation
von: Sui, Yuan, et al.
Veröffentlicht: (2026)
von: Sui, Yuan, et al.
Veröffentlicht: (2026)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean
von: Kim, SungHo, et al.
Veröffentlicht: (2025)
von: Kim, SungHo, et al.
Veröffentlicht: (2025)
Chatting with Bots: AI, Speech Acts, and the Edge of Assertion
von: Williams, Iwan, et al.
Veröffentlicht: (2024)
von: Williams, Iwan, et al.
Veröffentlicht: (2024)
From Words to Proverbs: Evaluating LLMs Linguistic and Cultural Competence in Saudi Dialects with Absher
von: Al-Monef, Renad, et al.
Veröffentlicht: (2025)
von: Al-Monef, Renad, et al.
Veröffentlicht: (2025)
SpaLLM-Guard: Pairing SMS Spam Detection Using Open-source and Commercial LLMs
von: Salman, Muhammad, et al.
Veröffentlicht: (2025)
von: Salman, Muhammad, et al.
Veröffentlicht: (2025)
Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models
von: Fettach, Yousra, et al.
Veröffentlicht: (2026)
von: Fettach, Yousra, et al.
Veröffentlicht: (2026)
Guided Self-Evolving LLMs with Minimal Human Supervision
von: Yu, Wenhao, et al.
Veröffentlicht: (2025)
von: Yu, Wenhao, et al.
Veröffentlicht: (2025)
When "Competency" in Reasoning Opens the Door to Vulnerability: Jailbreaking LLMs via Novel Complex Ciphers
von: Handa, Divij, et al.
Veröffentlicht: (2024)
von: Handa, Divij, et al.
Veröffentlicht: (2024)
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
von: Abdin, Marah, et al.
Veröffentlicht: (2024)
von: Abdin, Marah, et al.
Veröffentlicht: (2024)
PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs
von: Gungordu, Oguzhan, et al.
Veröffentlicht: (2026)
von: Gungordu, Oguzhan, et al.
Veröffentlicht: (2026)
VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
Confidence is Not Competence
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
Combating Phone Scams with LLM-based Detection: Where Do We Stand?
von: Shen, Zitong, et al.
Veröffentlicht: (2024)
von: Shen, Zitong, et al.
Veröffentlicht: (2024)
xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Agentic AI with Orchestrator-Agent Trust: A Modular Visual Classification Framework with Trust-Aware Orchestration and RAG-Based Reasoning
von: Roumeliotis, Konstantinos I., et al.
Veröffentlicht: (2025)
von: Roumeliotis, Konstantinos I., et al.
Veröffentlicht: (2025)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
Efficient Safety Retrofitting Against Jailbreaking for LLMs
von: Garcia-Gasulla, Dario, et al.
Veröffentlicht: (2025)
von: Garcia-Gasulla, Dario, et al.
Veröffentlicht: (2025)
Learning to Self-Evolve
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
von: Wu, Rong, et al.
Veröffentlicht: (2025)
von: Wu, Rong, et al.
Veröffentlicht: (2025)
EvolvR: Self-Evolving Pairwise Reasoning for Story Evaluation to Enhance Generation
von: Wang, Xinda, et al.
Veröffentlicht: (2025)
von: Wang, Xinda, et al.
Veröffentlicht: (2025)
War and Peace (WarAgent): Large Language Model-based Multi-Agent Simulation of World Wars
von: Hua, Wenyue, et al.
Veröffentlicht: (2023)
von: Hua, Wenyue, et al.
Veröffentlicht: (2023)
LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DIALEVAL: Automated Type-Theoretic Evaluation of LLM Instruction Following
von: Basta, Nardine, et al.
Veröffentlicht: (2026) -
ConvoCache: Smart Re-Use of Chatbot Responses
von: Atkins, Conor, et al.
Veröffentlicht: (2024) -
ASRJam: Human-Friendly AI Speech Jamming to Prevent Automated Phone Scams
von: Grabovski, Freddie, et al.
Veröffentlicht: (2025) -
Orchestrating LLMs with Different Personalizations
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024) -
Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution
von: Maheswaran, Monishwaran, et al.
Veröffentlicht: (2026)