LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Di, Wu, Jianbo, Lei, Jingdi, Che, Tong, Li, Jiatong, Xie, Tong, Huang, Xiaoshui, Zhang, Shufei, Pavone, Marco, Li, Yuqiang, Ouyang, Wanli, Zhou, Dongzhan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
por: Zhang, Di, et al.
Publicado: (2024)
por: Zhang, Di, et al.
Publicado: (2024)
LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training
por: Zhu, Tong, et al.
Publicado: (2024)
por: Zhu, Tong, et al.
Publicado: (2024)
Control-R: Towards controllable test-time scaling
por: Zhang, Di, et al.
Publicado: (2025)
por: Zhang, Di, et al.
Publicado: (2025)
LLaMA Pro: Progressive LLaMA with Block Expansion
por: Wu, Chengyue, et al.
Publicado: (2024)
por: Wu, Chengyue, et al.
Publicado: (2024)
LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
ECHO-LLaMA: Efficient Caching for High-Performance LLaMA Training
por: Dialameh, Maryam, et al.
Publicado: (2025)
por: Dialameh, Maryam, et al.
Publicado: (2025)
LLaMA-Reg: Using LLaMA 2 for Unsupervised Medical Image Registration
por: Ma, Mingrui, et al.
Publicado: (2024)
por: Ma, Mingrui, et al.
Publicado: (2024)
VisionLLaMA: A Unified LLaMA Backbone for Vision Tasks
por: Chu, Xiangxiang, et al.
Publicado: (2024)
por: Chu, Xiangxiang, et al.
Publicado: (2024)
Enhancing Document-Level Question Answering via Multi-Hop Retrieval-Augmented Generation with LLaMA 3
por: Huang, Xinyue, et al.
Publicado: (2025)
por: Huang, Xinyue, et al.
Publicado: (2025)
Adapting LLaMA Decoder to Vision Transformer
por: Wang, Jiahao, et al.
Publicado: (2024)
por: Wang, Jiahao, et al.
Publicado: (2024)
LLaMAs Have Feelings Too: Unveiling Sentiment and Emotion Representations in LLaMA Models Through Probing
por: Di Palma, Dario, et al.
Publicado: (2025)
por: Di Palma, Dario, et al.
Publicado: (2025)
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning
por: Zhang, Di, et al.
Publicado: (2024)
por: Zhang, Di, et al.
Publicado: (2024)
Tailored-LLaMA: Optimizing Few-Shot Learning in Pruned LLaMA Models with Task-Specific Prompts
por: Aftab, Danyal, et al.
Publicado: (2024)
por: Aftab, Danyal, et al.
Publicado: (2024)
Online Test-time Adaptation for Interatomic Potentials
por: Cui, Taoyong, et al.
Publicado: (2024)
por: Cui, Taoyong, et al.
Publicado: (2024)
Iterative Pretraining Framework for Interatomic Potentials
por: Cui, Taoyong, et al.
Publicado: (2025)
por: Cui, Taoyong, et al.
Publicado: (2025)
How Vocabulary Sharing Facilitates Multilingualism in LLaMA?
por: Yuan, Fei, et al.
Publicado: (2023)
por: Yuan, Fei, et al.
Publicado: (2023)
LogLLaMA: Transformer-based log anomaly detection with LLaMA
por: Yang, Zhuoyi, et al.
Publicado: (2025)
por: Yang, Zhuoyi, et al.
Publicado: (2025)
FedSEA-LLaMA: A Secure, Efficient and Adaptive Federated Splitting Framework for Large Language Models
por: Zhang, Zishuai, et al.
Publicado: (2025)
por: Zhang, Zishuai, et al.
Publicado: (2025)
BanglaLlama: LLaMA for Bangla Language
por: Zehady, Abdullah Khan, et al.
Publicado: (2024)
por: Zehady, Abdullah Khan, et al.
Publicado: (2024)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
por: Cheng, Zebang, et al.
Publicado: (2024)
por: Cheng, Zebang, et al.
Publicado: (2024)
MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts
por: Li, Jiatong, et al.
Publicado: (2024)
por: Li, Jiatong, et al.
Publicado: (2024)
LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning
por: Jahangir, Md. Zihad Bin, et al.
Publicado: (2025)
por: Jahangir, Md. Zihad Bin, et al.
Publicado: (2025)
LLaMA Beyond English: An Empirical Study on Language Capability Transfer
por: Zhao, Jun, et al.
Publicado: (2024)
por: Zhao, Jun, et al.
Publicado: (2024)
Is Bigger and Deeper Always Better? Probing LLaMA Across Scales and Layers
por: Chen, Nuo, et al.
Publicado: (2023)
por: Chen, Nuo, et al.
Publicado: (2023)
Multimodal Medical Disease Classification with LLaMA II
por: Gapp, Christian, et al.
Publicado: (2024)
por: Gapp, Christian, et al.
Publicado: (2024)
LLaMA based Punctuation Restoration With Forward Pass Only Decoding
por: Pang, Yutong, et al.
Publicado: (2024)
por: Pang, Yutong, et al.
Publicado: (2024)
Amharic LLaMA and LLaVA: Multimodal LLMs for Low Resource Languages
por: Andersland, Michael
Publicado: (2024)
por: Andersland, Michael
Publicado: (2024)
LLaMA-Omni: Seamless Speech Interaction with Large Language Models
por: Fang, Qingkai, et al.
Publicado: (2024)
por: Fang, Qingkai, et al.
Publicado: (2024)
360-LLaMA-Factory: Plug & Play Sequence Parallelism for Long Post-Training
por: Zou, Haosheng, et al.
Publicado: (2025)
por: Zou, Haosheng, et al.
Publicado: (2025)
LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement
por: Kang, Boyi, et al.
Publicado: (2025)
por: Kang, Boyi, et al.
Publicado: (2025)
What If We Recaption Billions of Web Images with LLaMA-3?
por: Li, Xianhang, et al.
Publicado: (2024)
por: Li, Xianhang, et al.
Publicado: (2024)
Efficient and Effective Text Encoding for Chinese LLaMA and Alpaca
por: Cui, Yiming, et al.
Publicado: (2023)
por: Cui, Yiming, et al.
Publicado: (2023)
Dynamic Activation Pitfalls in LLaMA Models: An Empirical Study
por: Ma, Chi, et al.
Publicado: (2024)
por: Ma, Chi, et al.
Publicado: (2024)
Faster Speech-LLaMA Inference with Multi-token Prediction
por: Raj, Desh, et al.
Publicado: (2024)
por: Raj, Desh, et al.
Publicado: (2024)
Parameter-Efficient Fine-Tuning of LLaMA for the Clinical Domain
por: Gema, Aryo Pradipta, et al.
Publicado: (2023)
por: Gema, Aryo Pradipta, et al.
Publicado: (2023)
Evaluating LLaMA 3.2 for Software Vulnerability Detection
por: Gonçalves, José, et al.
Publicado: (2025)
por: Gonçalves, José, et al.
Publicado: (2025)
LLaMA-Based Models for Aspect-Based Sentiment Analysis
por: Šmíd, Jakub, et al.
Publicado: (2025)
por: Šmíd, Jakub, et al.
Publicado: (2025)
Me LLaMA: Foundation Large Language Models for Medical Applications
por: Xie, Qianqian, et al.
Publicado: (2024)
por: Xie, Qianqian, et al.
Publicado: (2024)
LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
por: Zhang, Renrui, et al.
Publicado: (2023)
por: Zhang, Renrui, et al.
Publicado: (2023)
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation
por: Li, Siyou, et al.
Publicado: (2024)
por: Li, Siyou, et al.
Publicado: (2024)
Ejemplares similares
-
Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
por: Zhang, Di, et al.
Publicado: (2024) -
LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training
por: Zhu, Tong, et al.
Publicado: (2024) -
Control-R: Towards controllable test-time scaling
por: Zhang, Di, et al.
Publicado: (2025) -
LLaMA Pro: Progressive LLaMA with Block Expansion
por: Wu, Chengyue, et al.
Publicado: (2024) -
LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training
por: Qu, Xiaoye, et al.
Publicado: (2024)