CogLM: Tracking Cognitive Development of Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Xinglin, Yuan, Peiwen, Feng, Shaoxiong, Li, Yiwei, Pan, Boyuan, Wang, Heda, Hu, Yao, Li, Kan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
Focused Large Language Models are Stable Many-Shot Learners
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
Instruction Embedding: Latent Representations of Instructions Towards Task Identification
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
BatchEval: Towards Human-like Text Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2023)
di: Yuan, Peiwen, et al.
Pubblicazione: (2023)
Poor-Supervised Evaluation for SuperLLM via Mutual Consistency
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Generative Dense Retrieval: Memory Can Be a Burden
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules
di: Zhang, Yueqi, et al.
Pubblicazione: (2025)
di: Zhang, Yueqi, et al.
Pubblicazione: (2025)
From Sub-Ability Diagnosis to Human-Aligned Generation: Bridging the Gap for Text Length Control via MARKERGEN
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Revisiting Self-Consistency from Dynamic Distributional Alignment Perspective on Answer Aggregation
di: Li, Yiwei, et al.
Pubblicazione: (2025)
di: Li, Yiwei, et al.
Pubblicazione: (2025)
Speculative Decoding for Multi-Sample Inference
di: Li, Yiwei, et al.
Pubblicazione: (2025)
di: Li, Yiwei, et al.
Pubblicazione: (2025)
On Time, Within Budget: Constraint-Driven Online Resource Allocation for Agentic Workflows
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
UniCBE: An Uniformity-driven Comparing Based Evaluation Framework with Unified Multi-Objective Optimization
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Silencer: From Discovery to Mitigation of Self-Bias in LLM-as-Benchmark-Generator
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
InsBank: Evolving Instruction Subset for Ongoing Alignment
di: Shi, Jiayi, et al.
Pubblicazione: (2025)
di: Shi, Jiayi, et al.
Pubblicazione: (2025)
Beyond One-Size-Fits-All: Tailored Benchmarks for Efficient Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
di: Wang, Xinglin, et al.
Pubblicazione: (2025)
di: Wang, Xinglin, et al.
Pubblicazione: (2025)
Diagnosing and Mitigating System Bias in Self-Rewarding RL
di: Tan, Chuyi, et al.
Pubblicazione: (2025)
di: Tan, Chuyi, et al.
Pubblicazione: (2025)
Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
PatternKV: Flattening KV Representation Expands Quantization Headroom
di: Zhang, Ji, et al.
Pubblicazione: (2025)
di: Zhang, Ji, et al.
Pubblicazione: (2025)
Learning More from Less: Unlocking Internal Representations for Benchmark Compression
di: Zhang, Yueqi, et al.
Pubblicazione: (2026)
di: Zhang, Yueqi, et al.
Pubblicazione: (2026)
CogSteer: Cognition-Inspired Selective Layer Intervention for Efficiently Steering Large Language Models
di: Wang, Xintong, et al.
Pubblicazione: (2024)
di: Wang, Xintong, et al.
Pubblicazione: (2024)
CogGPT: Unleashing the Power of Cognitive Dynamics on Large Language Models
di: Lv, Yaojia, et al.
Pubblicazione: (2024)
di: Lv, Yaojia, et al.
Pubblicazione: (2024)
Dynamic Stochastic Decoding Strategy for Open-Domain Dialogue Generation
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
CogAtom: From Cognitive Atoms to Olympiad-level Mathematical Reasoning in Large Language Models
di: Chen, Zhuofan, et al.
Pubblicazione: (2025)
di: Chen, Zhuofan, et al.
Pubblicazione: (2025)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
di: Li, Zihan, et al.
Pubblicazione: (2026)
di: Li, Zihan, et al.
Pubblicazione: (2026)
FaithLM: Towards Faithful Explanations for Large Language Models
di: Chuang, Yu-Neng, et al.
Pubblicazione: (2024)
di: Chuang, Yu-Neng, et al.
Pubblicazione: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
di: Zhu, Lianghui, et al.
Pubblicazione: (2023)
di: Zhu, Lianghui, et al.
Pubblicazione: (2023)
ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models
di: Guo, Wenbin, et al.
Pubblicazione: (2025)
di: Guo, Wenbin, et al.
Pubblicazione: (2025)
PonderLM: Pretraining Language Models to Ponder in Continuous Space
di: Zeng, Boyi, et al.
Pubblicazione: (2025)
di: Zeng, Boyi, et al.
Pubblicazione: (2025)
LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
di: Han, Chi, et al.
Pubblicazione: (2023)
di: Han, Chi, et al.
Pubblicazione: (2023)
Self-Cognition in Large Language Models: An Exploratory Study
di: Chen, Dongping, et al.
Pubblicazione: (2024)
di: Chen, Dongping, et al.
Pubblicazione: (2024)
CogBench: A Large Language Model Benchmark for Multilingual Speech-Based Cognitive Impairment Assessment
di: Feng, Rui, et al.
Pubblicazione: (2025)
di: Feng, Rui, et al.
Pubblicazione: (2025)
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
di: Li, Qixiu, et al.
Pubblicazione: (2024)
di: Li, Qixiu, et al.
Pubblicazione: (2024)
KG-BiLM: Knowledge Graph Embedding via Bidirectional Language Models
di: Chen, Zirui, et al.
Pubblicazione: (2025)
di: Chen, Zirui, et al.
Pubblicazione: (2025)
ECCoT: A Framework for Enhancing Effective Cognition via Chain of Thought in Large Language Model
di: Duan, Zhenke, et al.
Pubblicazione: (2025)
di: Duan, Zhenke, et al.
Pubblicazione: (2025)
Cognitive Memory in Large Language Models
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
CataLM: Empowering Catalyst Design Through Large Language Models
di: Wang, Ludi, et al.
Pubblicazione: (2024)
di: Wang, Ludi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation
di: Wang, Xinglin, et al.
Pubblicazione: (2024) -
Focused Large Language Models are Stable Many-Shot Learners
di: Yuan, Peiwen, et al.
Pubblicazione: (2024) -
Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning
di: Li, Yiwei, et al.
Pubblicazione: (2024) -
Instruction Embedding: Latent Representations of Instructions Towards Task Identification
di: Li, Yiwei, et al.
Pubblicazione: (2024) -
BatchEval: Towards Human-like Text Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2023)