If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Fan, Siqi, Huang, Xiusheng, Yao, Yiqun, Fang, Xuezhi, Liu, Kang, Han, Peng, Shang, Shuo, Sun, Aixin, Wang, Yequan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
nanoLM: an Affordable LLM Pre-training Benchmark via Accurate Loss Prediction across Scales
by: Yao, Yiqun, et al.
Published: (2023)
by: Yao, Yiqun, et al.
Published: (2023)
The Price of a Second Thought: On the Evaluation of Reasoning Efficiency in Large Language Models
by: Fan, Siqi, et al.
Published: (2025)
by: Fan, Siqi, et al.
Published: (2025)
Position-Aware Depth Decay Decoding ($D^3$): Boosting Large Language Model Inference Efficiency
by: Fan, Siqi, et al.
Published: (2025)
by: Fan, Siqi, et al.
Published: (2025)
EgoMem: Lifelong Memory Agent for Full-duplex Omnimodal Models
by: Yao, Yiqun, et al.
Published: (2025)
by: Yao, Yiqun, et al.
Published: (2025)
Not All Layers of LLMs Are Necessary During Inference
by: Fan, Siqi, et al.
Published: (2024)
by: Fan, Siqi, et al.
Published: (2024)
RoboEgo System Card: An Omnimodal Model with Native Full Duplexity
by: Yao, Yiqun, et al.
Published: (2025)
by: Yao, Yiqun, et al.
Published: (2025)
FLM-101B: An Open LLM and How to Train It with $100K Budget
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Commonsense Knowledge Editing Based on Free-Text in LLMs
by: Huang, Xiusheng, et al.
Published: (2024)
by: Huang, Xiusheng, et al.
Published: (2024)
Sketch: A Toolkit for Streamlining LLM Operations
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
FLM-Audio: Natural Monologues Improves Native Full-Duplex Chatbots via Dual Training
by: Yao, Yiqun, et al.
Published: (2025)
by: Yao, Yiqun, et al.
Published: (2025)
Open-domain Implicit Format Control for Large Language Model Generation
by: Yao, Yiqun, et al.
Published: (2024)
by: Yao, Yiqun, et al.
Published: (2024)
Reasons and Solutions for the Decline in Model Performance after Editing
by: Huang, Xiusheng, et al.
Published: (2024)
by: Huang, Xiusheng, et al.
Published: (2024)
Toward Embodied AGI: A Review of Embodied AI and the Road Ahead
by: Wang, Yequan, et al.
Published: (2025)
by: Wang, Yequan, et al.
Published: (2025)
Evaluating LLM Adaptation to Sociodemographic Factors: User Profile vs. Dialogue History
by: Zhong, Qishuai, et al.
Published: (2025)
by: Zhong, Qishuai, et al.
Published: (2025)
Capability Localization: Capabilities Can be Localized rather than Individual Knowledge
by: Huang, Xiusheng, et al.
Published: (2025)
by: Huang, Xiusheng, et al.
Published: (2025)
Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice
by: Huang, Xiusheng, et al.
Published: (2026)
by: Huang, Xiusheng, et al.
Published: (2026)
GCRE-GPT: A Generative Model for Comparative Relation Extraction
by: Wang, Yequan, et al.
Published: (2023)
by: Wang, Yequan, et al.
Published: (2023)
Masked Structural Growth for 2x Faster Language Model Pre-training
by: Yao, Yiqun, et al.
Published: (2023)
by: Yao, Yiqun, et al.
Published: (2023)
Hint Tuning: Less Data Makes Better Reasoners
by: Fan, Siqi, et al.
Published: (2026)
by: Fan, Siqi, et al.
Published: (2026)
How Would We Know What God is Up To?
by: Eaton, Heather, et al.
Published: (2025)
by: Eaton, Heather, et al.
Published: (2025)
Bring Your Own Character: A Holistic Solution for Automatic Facial Animation Generation of Customized Characters
by: Bai, Zechen, et al.
Published: (2024)
by: Bai, Zechen, et al.
Published: (2024)
Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs
by: Tan, Wenhui, et al.
Published: (2026)
by: Tan, Wenhui, et al.
Published: (2026)
Exploiting Contextual Knowledge in LLMs through V-usable Information based Layer Enhancement
by: Yuan, Xiaowei, et al.
Published: (2025)
by: Yuan, Xiaowei, et al.
Published: (2025)
Kindergarteners Building a Library of Their Own: Using Apps to Make Digital Stories and Work towards Lifelong Learning in Information Literacy
by: Moore, Hilde Terese Drivenes, et al.
Published: (2021)
by: Moore, Hilde Terese Drivenes, et al.
Published: (2021)
Theory-optimal Quantization Based on Flatness
by: Huang, Xiusheng, et al.
Published: (2026)
by: Huang, Xiusheng, et al.
Published: (2026)
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
SimpleMem: Efficient Lifelong Memory for LLM Agents
by: Liu, Jiaqi, et al.
Published: (2026)
by: Liu, Jiaqi, et al.
Published: (2026)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
by: Mayne, Harry, et al.
Published: (2025)
by: Mayne, Harry, et al.
Published: (2025)
Do Language Models Enjoy Their Own Stories? Prompting Large Language Models for Automatic Story Evaluation
by: Chhun, Cyril, et al.
Published: (2024)
by: Chhun, Cyril, et al.
Published: (2024)
KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital Companions
by: Wu, Tingyu, et al.
Published: (2026)
by: Wu, Tingyu, et al.
Published: (2026)
LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners
by: Zheng, Junhao, et al.
Published: (2025)
by: Zheng, Junhao, et al.
Published: (2025)
CatCode: A Comprehensive Evaluation Framework for LLMs On the Mixture of Code and Text
by: Lin, Zhenru, et al.
Published: (2024)
by: Lin, Zhenru, et al.
Published: (2024)
Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement
by: Ding, Peng, et al.
Published: (2025)
by: Ding, Peng, et al.
Published: (2025)
Long Context vs. RAG for LLMs: An Evaluation and Revisits
by: Li, Xinze, et al.
Published: (2024)
by: Li, Xinze, et al.
Published: (2024)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
LLM Evaluators Recognize and Favor Their Own Generations
by: Panickssery, Arjun, et al.
Published: (2024)
by: Panickssery, Arjun, et al.
Published: (2024)
In Their Own Words: Student Stories of Seeking Learning Support
by: Brown, Mark, et al.
Published: (2013)
by: Brown, Mark, et al.
Published: (2013)
We're Different, We're the Same: Creative Homogeneity Across LLMs
by: Wenger, Emily, et al.
Published: (2025)
by: Wenger, Emily, et al.
Published: (2025)
The Stories We're Told: Nation‐Specific Narratives of Race and Racism
by: Brandon A. Jackson
Published: (2025)
by: Brandon A. Jackson
Published: (2025)
Similar Items
-
nanoLM: an Affordable LLM Pre-training Benchmark via Accurate Loss Prediction across Scales
by: Yao, Yiqun, et al.
Published: (2023) -
The Price of a Second Thought: On the Evaluation of Reasoning Efficiency in Large Language Models
by: Fan, Siqi, et al.
Published: (2025) -
Position-Aware Depth Decay Decoding ($D^3$): Boosting Large Language Model Inference Efficiency
by: Fan, Siqi, et al.
Published: (2025) -
EgoMem: Lifelong Memory Agent for Full-duplex Omnimodal Models
by: Yao, Yiqun, et al.
Published: (2025) -
Not All Layers of LLMs Are Necessary During Inference
by: Fan, Siqi, et al.
Published: (2024)