CC-LEARN: Cohort-based Consistency Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Xiao, Shrivastava, Shaswat, Li, Zhaonan, Dineen, Jacob, Lu, Shijie, Ahuja, Avneet, Shen, Ming, Xu, Zhikun, Zhou, Ben |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ToW: Thoughts of Words Improve Reasoning in Large Language Models
by: Xu, Zhikun, et al.
Published: (2024)
by: Xu, Zhikun, et al.
Published: (2024)
BOW: Reinforcement Learning for Bottlenecked Next Word Prediction
by: Shen, Ming, et al.
Published: (2025)
by: Shen, Ming, et al.
Published: (2025)
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
by: Dineen, Jacob, et al.
Published: (2025)
by: Dineen, Jacob, et al.
Published: (2025)
Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution
by: Dineen, Jacob, et al.
Published: (2026)
by: Dineen, Jacob, et al.
Published: (2026)
Evaluating Medical LLMs by Levels of Autonomy: A Survey Moving from Benchmarks to Applications
by: Ye, Xiao, et al.
Published: (2025)
by: Ye, Xiao, et al.
Published: (2025)
Unbiased Visual Reasoning with Controlled Visual Inputs
by: Li, Zhaonan, et al.
Published: (2025)
by: Li, Zhaonan, et al.
Published: (2025)
Echoes of Agreement: Argument Driven Opinion Shifts in Large Language Models
by: Kaur, Avneet
Published: (2025)
by: Kaur, Avneet
Published: (2025)
StockBot 2.0: Vanilla LSTMs Outperform Transformer-based Forecasting for Stock Prices
by: Mohanty, Shaswat
Published: (2026)
by: Mohanty, Shaswat
Published: (2026)
Beyond the Battlefield: Framing Analysis of Media Coverage in Conflict Reporting
by: Kaur, Avneet, et al.
Published: (2025)
by: Kaur, Avneet, et al.
Published: (2025)
Skill Reuse as Compression in Agentic RL
by: Xu, Zhikun, et al.
Published: (2026)
by: Xu, Zhikun, et al.
Published: (2026)
RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems
by: Srinivasan, Adarsh, et al.
Published: (2025)
by: Srinivasan, Adarsh, et al.
Published: (2025)
VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images
by: Li, Zhaonan, et al.
Published: (2026)
by: Li, Zhaonan, et al.
Published: (2026)
ArenaBencher: Automatic Benchmark Evolution via Multi-Model Competitive Evaluation
by: Liu, Qin, et al.
Published: (2025)
by: Liu, Qin, et al.
Published: (2025)
ThinkTuning: Instilling Cognitive Reflections without Distillation
by: RRV, Aswin, et al.
Published: (2025)
by: RRV, Aswin, et al.
Published: (2025)
Reliable Use of Lemmas via Eligibility Reasoning and Section$-$Aware Reinforcement Learning
by: Xu, Zhikun, et al.
Published: (2026)
by: Xu, Zhikun, et al.
Published: (2026)
Contrastive Learning of Emoji-based Representations for Resource-Poor Languages
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Consistency-Aware Parameter-Preserving Knowledge Editing Framework for Multi-Hop Question Answering
by: Deng, Lingwen, et al.
Published: (2025)
by: Deng, Lingwen, et al.
Published: (2025)
CORE: Concept-Oriented Reinforcement for Bridging the Definition-Application Gap in Mathematical Reasoning
by: Gao, Zijun, et al.
Published: (2025)
by: Gao, Zijun, et al.
Published: (2025)
Bridging Latent Reasoning and Target-Language Generation via Retrieval-Transition Heads
by: Patel, Shaswat, et al.
Published: (2026)
by: Patel, Shaswat, et al.
Published: (2026)
ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch
by: Chen, Jiawei, et al.
Published: (2025)
by: Chen, Jiawei, et al.
Published: (2025)
CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
by: Ye, Yangfan, et al.
Published: (2025)
by: Ye, Yangfan, et al.
Published: (2025)
In-Context Learning through the Bayesian Prism
by: Panwar, Madhur, et al.
Published: (2023)
by: Panwar, Madhur, et al.
Published: (2023)
Reinforcement World Model Learning for LLM-based Agents
by: Yu, Xiao, et al.
Published: (2026)
by: Yu, Xiao, et al.
Published: (2026)
SysCaps: Language Interfaces for Simulation Surrogates of Complex Systems
by: Emami, Patrick, et al.
Published: (2024)
by: Emami, Patrick, et al.
Published: (2024)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
by: Ahuja, Sanchit, et al.
Published: (2026)
by: Ahuja, Sanchit, et al.
Published: (2026)
CM-Align: Consistency-based Multilingual Alignment for Large Language Models
by: Zhang, Xue, et al.
Published: (2025)
by: Zhang, Xue, et al.
Published: (2025)
WanJuan-CC: A Safe and High-Quality Open-sourced English Webtext Dataset
by: Qiu, Jiantao, et al.
Published: (2024)
by: Qiu, Jiantao, et al.
Published: (2024)
Segmentation Beyond Defaults: Asymmetrical Byte Pair Encoding for Optimal Machine Translation Performance
by: Yadav, Saumitra, et al.
Published: (2025)
by: Yadav, Saumitra, et al.
Published: (2025)
Get away with less: Need of source side data curation to build parallel corpus for low resource Machine Translation
by: Yadav, Saumitra, et al.
Published: (2026)
by: Yadav, Saumitra, et al.
Published: (2026)
Can Constructions "SCAN" Compositionality ?
by: Katrapati, Ganesh, et al.
Published: (2025)
by: Katrapati, Ganesh, et al.
Published: (2025)
TeClass: A Human-Annotated Relevance-based Headline Classification and Generation Dataset for Telugu
by: Kanumolu, Gopichand, et al.
Published: (2024)
by: Kanumolu, Gopichand, et al.
Published: (2024)
FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge
by: Yang, Bo, et al.
Published: (2026)
by: Yang, Bo, et al.
Published: (2026)
MME-CC: A Challenging Multi-Modal Evaluation Benchmark of Cognitive Capacity
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
Beneath the Surface: Investigating LLMs' Capabilities for Communicating with Subtext
by: Ahuja, Kabir, et al.
Published: (2026)
by: Ahuja, Kabir, et al.
Published: (2026)
Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning
by: Zhang, Xue, et al.
Published: (2025)
by: Zhang, Xue, et al.
Published: (2025)
VI-OOD: A Unified Representation Learning Framework for Textual Out-of-distribution Detection
by: Zhan, Li-Ming, et al.
Published: (2024)
by: Zhan, Li-Ming, et al.
Published: (2024)
Emotions are Universal: Learning Sentiment Based Representations of Resource-Poor Languages using Siamese Networks
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Test-Time Reinforcement Learning for GUI Grounding via Region Consistency
by: Du, Yong, et al.
Published: (2025)
by: Du, Yong, et al.
Published: (2025)
CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing
by: Xu, Zhipeng, et al.
Published: (2026)
by: Xu, Zhipeng, et al.
Published: (2026)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
by: Kholkar, Gauri, et al.
Published: (2025)
by: Kholkar, Gauri, et al.
Published: (2025)
Similar Items
-
ToW: Thoughts of Words Improve Reasoning in Large Language Models
by: Xu, Zhikun, et al.
Published: (2024) -
BOW: Reinforcement Learning for Bottlenecked Next Word Prediction
by: Shen, Ming, et al.
Published: (2025) -
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
by: Dineen, Jacob, et al.
Published: (2025) -
Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution
by: Dineen, Jacob, et al.
Published: (2026) -
Evaluating Medical LLMs by Levels of Autonomy: A Survey Moving from Benchmarks to Applications
by: Ye, Xiao, et al.
Published: (2025)