Saved in:
| Main Authors: | Mirowski, Piotr Wojciech, Branch, Boyd, Mathewson, Kory Wallace |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.08474 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Designing and Evaluating Dialogue LLMs for Co-Creative Improvised Theatre
by: Branch, Boyd, et al.
Published: (2024)
by: Branch, Boyd, et al.
Published: (2024)
A Robot Walks into a Bar: Can Language Models Serve as Creativity Support Tools for Comedy? An Evaluation of LLMs' Humour Alignment with Comedians
by: Mirowski, Piotr Wojciech, et al.
Published: (2024)
by: Mirowski, Piotr Wojciech, et al.
Published: (2024)
Universal Conceptual Structure in Neural Translation: Probing NLLB-200's Multilingual Geometry
by: Mathewson, Kyle Elliott
Published: (2026)
by: Mathewson, Kyle Elliott
Published: (2026)
Divergent Creativity in Humans and Large Language Models
by: Bellemare-Pepin, Antoine, et al.
Published: (2024)
by: Bellemare-Pepin, Antoine, et al.
Published: (2024)
From Punchlines to Predictions: A Metric to Assess LLM Performance in Identifying Humor in Stand-Up Comedy
by: Romanowski, Adrianna, et al.
Published: (2025)
by: Romanowski, Adrianna, et al.
Published: (2025)
Language Inclusion for Boundedly-Ambiguous Vector Addition Systems is Decidable
by: Czerwiński, Wojciech, et al.
Published: (2022)
by: Czerwiński, Wojciech, et al.
Published: (2022)
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
by: Hong, Shiwei, et al.
Published: (2026)
by: Hong, Shiwei, et al.
Published: (2026)
CleanComedy: Creating Friendly Humor through Generative Techniques
by: Vikhorev, Dmitry, et al.
Published: (2024)
by: Vikhorev, Dmitry, et al.
Published: (2024)
COMIC: Agentic Sketch Comedy Generation
by: Hong, Susung, et al.
Published: (2026)
by: Hong, Susung, et al.
Published: (2026)
Real-Time Performance Optimization of Travel Reservation Systems Using AI and Microservices
by: Barua, Biman, et al.
Published: (2024)
by: Barua, Biman, et al.
Published: (2024)
StandUp4AI: A New Multilingual Dataset for Humor Detection in Stand-up Comedy Videos
by: Barriere, Valentin, et al.
Published: (2025)
by: Barriere, Valentin, et al.
Published: (2025)
LLM Optimization Unlocks Real-Time Pairwise Reranking
by: Wu, Jingyu, et al.
Published: (2025)
by: Wu, Jingyu, et al.
Published: (2025)
The Oscars of AI Theater: A Survey on Role-Playing with Language Models
by: Chen, Nuo, et al.
Published: (2024)
by: Chen, Nuo, et al.
Published: (2024)
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
by: Boppana, Siddharth, et al.
Published: (2026)
by: Boppana, Siddharth, et al.
Published: (2026)
SimpleTool: Parallel Decoding for Real-Time LLM Function Calling
by: Shi, Xiaoxin, et al.
Published: (2026)
by: Shi, Xiaoxin, et al.
Published: (2026)
LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks
by: Long, Xiang, et al.
Published: (2026)
by: Long, Xiang, et al.
Published: (2026)
Agent Laboratory: Using LLM Agents as Research Assistants
by: Schmidgall, Samuel, et al.
Published: (2025)
by: Schmidgall, Samuel, et al.
Published: (2025)
LiveThinking: Enabling Real-Time Efficient Reasoning for AI-Powered Livestreaming via Reinforcement Learning
by: Sun, Yuhan, et al.
Published: (2025)
by: Sun, Yuhan, et al.
Published: (2025)
LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection
by: Xu, Cheng, et al.
Published: (2026)
by: Xu, Cheng, et al.
Published: (2026)
LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems
by: Moia, Vitor Hugo Galhardo, et al.
Published: (2025)
by: Moia, Vitor Hugo Galhardo, et al.
Published: (2025)
SPORTSQL: An Interactive System for Real-Time Sports Reasoning and Visualization
by: Martinez, Sebastian, et al.
Published: (2025)
by: Martinez, Sebastian, et al.
Published: (2025)
AI-assisted Knowledge Discovery in Biomedical Literature to Support Decision-making in Precision Oncology
by: He, Ting, et al.
Published: (2024)
by: He, Ting, et al.
Published: (2024)
LLM-CAS: Dynamic Neuron Perturbation for Real-Time Hallucination Correction
by: Zhang, Jensen, et al.
Published: (2025)
by: Zhang, Jensen, et al.
Published: (2025)
A Vision for Geo-Temporal Deep Research Systems: Towards Comprehensive, Transparent, and Reproducible Geo-Temporal Information Synthesis
by: Martins, Bruno, et al.
Published: (2025)
by: Martins, Bruno, et al.
Published: (2025)
The Surprising Universality of LLM Outputs: A Real-Time Verification Primitive
by: Bogdan, Alex, et al.
Published: (2026)
by: Bogdan, Alex, et al.
Published: (2026)
Real-Time Trustworthiness Scoring for LLM Structured Outputs and Data Extraction
by: Goh, Hui Wen, et al.
Published: (2026)
by: Goh, Hui Wen, et al.
Published: (2026)
Modeling LLM Agent Reviewer Dynamics in Elo-Ranked Review System
by: Huang, Hsiang-Wei, et al.
Published: (2026)
by: Huang, Hsiang-Wei, et al.
Published: (2026)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
by: Ghosh, Himel, et al.
Published: (2026)
by: Ghosh, Himel, et al.
Published: (2026)
Time Series Language Model for Descriptive Caption Generation
by: Trabelsi, Mohamed, et al.
Published: (2025)
by: Trabelsi, Mohamed, et al.
Published: (2025)
Dynamic Optimizations of LLM Ensembles with Two-Stage Reinforcement Learning Agents
by: Tekin, Selim Furkan, et al.
Published: (2025)
by: Tekin, Selim Furkan, et al.
Published: (2025)
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
by: Chen, Liang, et al.
Published: (2025)
by: Chen, Liang, et al.
Published: (2025)
CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge
by: Bae, Seyun, et al.
Published: (2026)
by: Bae, Seyun, et al.
Published: (2026)
LiveFC: A System for Live Fact-Checking of Audio Streams
by: V, Venktesh, et al.
Published: (2024)
by: V, Venktesh, et al.
Published: (2024)
Evaluating Performance Drift from Model Switching in Multi-Turn LLM Systems
by: Khraishi, Raad, et al.
Published: (2026)
by: Khraishi, Raad, et al.
Published: (2026)
When Agents Trade: Live Multi-Market Trading Benchmark for LLM Agents
by: Qian, Lingfei, et al.
Published: (2025)
by: Qian, Lingfei, et al.
Published: (2025)
LLMLagBench: Identifying Temporal Training Boundaries in Large Language Models
by: Pęzik, Piotr, et al.
Published: (2025)
by: Pęzik, Piotr, et al.
Published: (2025)
SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts
by: Xin, Yuan, et al.
Published: (2026)
by: Xin, Yuan, et al.
Published: (2026)
Reverse Question Answering: Can an LLM Write a Question so Hard (or Bad) that it Can't Answer?
by: Balepur, Nishant, et al.
Published: (2024)
by: Balepur, Nishant, et al.
Published: (2024)
The devil is in discretization discrepancy. Robustifying Differentiable NAS with Single-Stage Searching Protocol
by: Subbotko, Konstanty, et al.
Published: (2024)
by: Subbotko, Konstanty, et al.
Published: (2024)
Controlling Performance and Budget of a Centralized Multi-agent LLM System with Reinforcement Learning
by: Jin, Bowen, et al.
Published: (2025)
by: Jin, Bowen, et al.
Published: (2025)
Similar Items
-
Designing and Evaluating Dialogue LLMs for Co-Creative Improvised Theatre
by: Branch, Boyd, et al.
Published: (2024) -
A Robot Walks into a Bar: Can Language Models Serve as Creativity Support Tools for Comedy? An Evaluation of LLMs' Humour Alignment with Comedians
by: Mirowski, Piotr Wojciech, et al.
Published: (2024) -
Universal Conceptual Structure in Neural Translation: Probing NLLB-200's Multilingual Geometry
by: Mathewson, Kyle Elliott
Published: (2026) -
Divergent Creativity in Humans and Large Language Models
by: Bellemare-Pepin, Antoine, et al.
Published: (2024) -
From Punchlines to Predictions: A Metric to Assess LLM Performance in Identifying Humor in Stand-Up Comedy
by: Romanowski, Adrianna, et al.
Published: (2025)