Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ning, Xuefei, Wang, Zifu, Li, Shiyao, Lin, Zinan, Yao, Peiran, Fu, Tianyu, Blaschko, Matthew B., Dai, Guohao, Yang, Huazhong, Wang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Skeleton-of-Thought: Prompting LLMs for Efficient Parallel Generation
von: Ning, Xuefei, et al.
Veröffentlicht: (2023)
von: Ning, Xuefei, et al.
Veröffentlicht: (2023)
Linear Combination of Saved Checkpoints Makes Consistency and Diffusion Models Better
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
Jaccard Metric Losses: Optimizing the Jaccard Index with Soft Labels
von: Wang, Zifu, et al.
Veröffentlicht: (2023)
von: Wang, Zifu, et al.
Veröffentlicht: (2023)
Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
von: Zhao, Lin, et al.
Veröffentlicht: (2024)
von: Zhao, Lin, et al.
Veröffentlicht: (2024)
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
Evaluating Quantized Large Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
Distilled Decoding 2: One-step Sampling of Image Auto-regressive Models with Conditional Score Distillation
von: Liu, Enshu, et al.
Veröffentlicht: (2025)
von: Liu, Enshu, et al.
Veröffentlicht: (2025)
Mixture of Attention Spans: Optimizing LLM Inference Efficiency with Heterogeneous Sliding-Window Lengths
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
CSKV: Training-Efficient Channel Shrinking for KV Cache in Long-Context Scenarios
von: Wang, Luning, et al.
Veröffentlicht: (2024)
von: Wang, Luning, et al.
Veröffentlicht: (2024)
R2R: Efficiently Navigating Divergent Reasoning Paths with Small-Large Model Token Routing
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Dice Semimetric Losses: Optimizing the Dice Score with Soft Labels
von: Wang, Zifu, et al.
Veröffentlicht: (2023)
von: Wang, Zifu, et al.
Veröffentlicht: (2023)
Distilled Decoding 1: One-step Sampling of Image Auto-regressive Models with Flow Matching
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
von: Liu, Enshu, et al.
Veröffentlicht: (2024)
NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization
von: Liu, Enshu, et al.
Veröffentlicht: (2026)
von: Liu, Enshu, et al.
Veröffentlicht: (2026)
PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
Jigsaw-R1: A Study of Rule-based Visual Reinforcement Learning with Jigsaw Puzzles
von: Wang, Zifu, et al.
Veröffentlicht: (2025)
von: Wang, Zifu, et al.
Veröffentlicht: (2025)
MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
MBQ: Modality-Balanced Quantization for Large Vision-Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
LLMs Can Teach Themselves to Better Predict the Future
von: Turtel, Benjamin, et al.
Veröffentlicht: (2025)
von: Turtel, Benjamin, et al.
Veröffentlicht: (2025)
GenMAC: Compositional Text-to-Video Generation with Multi-Agent Collaboration
von: Huang, Kaiyi, et al.
Veröffentlicht: (2024)
von: Huang, Kaiyi, et al.
Veröffentlicht: (2024)
LV-Eval: A Balanced Long-Context Benchmark with 5 Length Levels Up to 256K
von: Yuan, Tao, et al.
Veröffentlicht: (2024)
von: Yuan, Tao, et al.
Veröffentlicht: (2024)
DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
Accelerating Auto-regressive Text-to-Image Generation with Training-free Speculative Jacobi Decoding
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGAs
von: Zeng, Shulin, et al.
Veröffentlicht: (2024)
von: Zeng, Shulin, et al.
Veröffentlicht: (2024)
A Survey on Efficient Inference for Large Language Models
von: Zhou, Zixuan, et al.
Veröffentlicht: (2024)
von: Zhou, Zixuan, et al.
Veröffentlicht: (2024)
SpeContext: Enabling Efficient Long-context Reasoning with Speculative Context Sparsity in LLMs
von: Xu, Jiaming, et al.
Veröffentlicht: (2025)
von: Xu, Jiaming, et al.
Veröffentlicht: (2025)
Latent Zoning Network: A Unified Principle for Generative Modeling, Representation Learning, and Classification
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
NAB: Neural Adaptive Binning for Sparse-View CT reconstruction
von: Xie, Wangduo, et al.
Veröffentlicht: (2026)
von: Xie, Wangduo, et al.
Veröffentlicht: (2026)
SJD++: Improved Speculative Jacobi Decoding for Training-free Acceleration of Discrete Auto-regressive Text-to-Image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
EchoPrune: Interpreting Redundancy as Temporal Echoes for Efficient VideoLLMs
von: Li, Jiameng, et al.
Veröffentlicht: (2026)
von: Li, Jiameng, et al.
Veröffentlicht: (2026)
Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
FilMaster: Bridging Cinematic Principles and Generative AI for Automated Film Generation
von: Huang, Kaiyi, et al.
Veröffentlicht: (2025)
von: Huang, Kaiyi, et al.
Veröffentlicht: (2025)
MI-Pruner: Crossmodal Mutual Information-guided Token Pruner for Efficient MLLMs
von: Li, Jiameng, et al.
Veröffentlicht: (2026)
von: Li, Jiameng, et al.
Veröffentlicht: (2026)
CLASH: A Benchmark for Cross-Modal Contradiction Detection
von: Popordanoska, Teodora, et al.
Veröffentlicht: (2025)
von: Popordanoska, Teodora, et al.
Veröffentlicht: (2025)
SoftCFG: Uncertainty-guided Stable Guidance for Visual Autoregressive Model
von: Xu, Dongli, et al.
Veröffentlicht: (2025)
von: Xu, Dongli, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Skeleton-of-Thought: Prompting LLMs for Efficient Parallel Generation
von: Ning, Xuefei, et al.
Veröffentlicht: (2023) -
Linear Combination of Saved Checkpoints Makes Consistency and Diffusion Models Better
von: Liu, Enshu, et al.
Veröffentlicht: (2024) -
Jaccard Metric Losses: Optimizing the Jaccard Index with Soft Labels
von: Wang, Zifu, et al.
Veröffentlicht: (2023) -
Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs
von: Liu, Enshu, et al.
Veröffentlicht: (2024) -
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
von: Zhao, Lin, et al.
Veröffentlicht: (2024)