Salvato in:
| Autori principali: | Ren, Houxing, Zhan, Mingjie, Wu, Zhongyuan, Li, Hongsheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2405.17103 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
di: Ren, Houxing, et al.
Pubblicazione: (2024)
di: Ren, Houxing, et al.
Pubblicazione: (2024)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
di: Wang, Ke, et al.
Pubblicazione: (2025)
di: Wang, Ke, et al.
Pubblicazione: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
di: Lu, Zimu, et al.
Pubblicazione: (2024)
di: Lu, Zimu, et al.
Pubblicazione: (2024)
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
di: Lu, Zimu, et al.
Pubblicazione: (2025)
di: Lu, Zimu, et al.
Pubblicazione: (2025)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
di: Lu, Zimu, et al.
Pubblicazione: (2024)
di: Lu, Zimu, et al.
Pubblicazione: (2024)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
di: Wang, Ke, et al.
Pubblicazione: (2025)
di: Wang, Ke, et al.
Pubblicazione: (2025)
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
di: Ren, Houxing, et al.
Pubblicazione: (2026)
di: Ren, Houxing, et al.
Pubblicazione: (2026)
Flexible-length Text Infilling for Discrete Diffusion Models
di: Zhang, Andrew, et al.
Pubblicazione: (2025)
di: Zhang, Andrew, et al.
Pubblicazione: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
di: Lu, Zimu, et al.
Pubblicazione: (2026)
di: Lu, Zimu, et al.
Pubblicazione: (2026)
Token Alignment via Character Matching for Subword Completion
di: Athiwaratkun, Ben, et al.
Pubblicazione: (2024)
di: Athiwaratkun, Ben, et al.
Pubblicazione: (2024)
Explainable Token-level Noise Filtering for LLM Fine-tuning Datasets
di: Yang, Yuchen, et al.
Pubblicazione: (2026)
di: Yang, Yuchen, et al.
Pubblicazione: (2026)
Eliminating Agentic Workflow for Introduction Generation with Parametric Stage Tokens
di: Zhang, Meicong, et al.
Pubblicazione: (2025)
di: Zhang, Meicong, et al.
Pubblicazione: (2025)
Unlocking Prompt Infilling Capability for Diffusion Language Models
di: Fujinuma, Yoshinari, et al.
Pubblicazione: (2026)
di: Fujinuma, Yoshinari, et al.
Pubblicazione: (2026)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
di: Lu, Zimu, et al.
Pubblicazione: (2024)
di: Lu, Zimu, et al.
Pubblicazione: (2024)
Edit-Based Refinement for Parallel Masked Diffusion Language Models
di: Ren, Houxing, et al.
Pubblicazione: (2026)
di: Ren, Houxing, et al.
Pubblicazione: (2026)
Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
di: Wang, Ke, et al.
Pubblicazione: (2024)
di: Wang, Ke, et al.
Pubblicazione: (2024)
Reinforcement Learning with Token-level Feedback for Controllable Text Generation
di: Li, Wendi, et al.
Pubblicazione: (2024)
di: Li, Wendi, et al.
Pubblicazione: (2024)
MAGNET: Augmenting Generative Decoders with Representation Learning and Infilling Capabilities
di: Khosla, Savya, et al.
Pubblicazione: (2025)
di: Khosla, Savya, et al.
Pubblicazione: (2025)
Unlocking the Potential of Diffusion Language Models through Template Infilling
di: Lee, Junhoo, et al.
Pubblicazione: (2025)
di: Lee, Junhoo, et al.
Pubblicazione: (2025)
From Language Models over Tokens to Language Models over Characters
di: Vieira, Tim, et al.
Pubblicazione: (2024)
di: Vieira, Tim, et al.
Pubblicazione: (2024)
SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding
di: Hou, Shuyang, et al.
Pubblicazione: (2026)
di: Hou, Shuyang, et al.
Pubblicazione: (2026)
KL3M Tokenizers: A Family of Domain-Specific and Character-Level Tokenizers for Legal, Financial, and Preprocessing Applications
di: Bommarito, Michael J, et al.
Pubblicazione: (2025)
di: Bommarito, Michael J, et al.
Pubblicazione: (2025)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
di: Yang, Yunqiao, et al.
Pubblicazione: (2026)
di: Yang, Yunqiao, et al.
Pubblicazione: (2026)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
di: Yang, Yunqiao, et al.
Pubblicazione: (2025)
di: Yang, Yunqiao, et al.
Pubblicazione: (2025)
Broken-Token: Filtering Obfuscated Prompts by Counting Characters-Per-Token
di: Zychlinski, Shaked, et al.
Pubblicazione: (2025)
di: Zychlinski, Shaked, et al.
Pubblicazione: (2025)
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds
di: Wang, Lei, et al.
Pubblicazione: (2024)
di: Wang, Lei, et al.
Pubblicazione: (2024)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
di: Li, Yanhong, et al.
Pubblicazione: (2025)
di: Li, Yanhong, et al.
Pubblicazione: (2025)
AgentDropout: Dynamic Agent Elimination for Token-Efficient and High-Performance LLM-Based Multi-Agent Collaboration
di: Wang, Zhexuan, et al.
Pubblicazione: (2025)
di: Wang, Zhexuan, et al.
Pubblicazione: (2025)
EMBRE: Entity-aware Masking for Biomedical Relation Extraction
di: Li, Mingjie, et al.
Pubblicazione: (2024)
di: Li, Mingjie, et al.
Pubblicazione: (2024)
CoMat: Aligning Text-to-Image Diffusion Model with Image-to-Text Concept Matching
di: Jiang, Dongzhi, et al.
Pubblicazione: (2024)
di: Jiang, Dongzhi, et al.
Pubblicazione: (2024)
Revisiting Character-level Adversarial Attacks for Language Models
di: Rocamora, Elias Abad, et al.
Pubblicazione: (2024)
di: Rocamora, Elias Abad, et al.
Pubblicazione: (2024)
Rubrics to Tokens: Bridging Response-level Rubrics and Token-level Rewards in Instruction Following Tasks
di: Xu, Tianze, et al.
Pubblicazione: (2026)
di: Xu, Tianze, et al.
Pubblicazione: (2026)
Text Generation Beyond Discrete Token Sampling
di: Zhuang, Yufan, et al.
Pubblicazione: (2025)
di: Zhuang, Yufan, et al.
Pubblicazione: (2025)
T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
di: Jiang, Dongzhi, et al.
Pubblicazione: (2025)
di: Jiang, Dongzhi, et al.
Pubblicazione: (2025)
Re-Initialization Token Learning for Tool-Augmented Large Language Models
di: Li, Chenghao, et al.
Pubblicazione: (2025)
di: Li, Chenghao, et al.
Pubblicazione: (2025)
Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Models in Clinical Text
di: Xiao, Bushi, et al.
Pubblicazione: (2026)
di: Xiao, Bushi, et al.
Pubblicazione: (2026)
Thunder-Tok: Minimizing Tokens per Word in Tokenizing Korean Texts for Generative Language Models
di: Cho, Gyeongje, et al.
Pubblicazione: (2025)
di: Cho, Gyeongje, et al.
Pubblicazione: (2025)
Token Masking Improves Transformer-Based Text Classification
di: Xu, Xianglong, et al.
Pubblicazione: (2025)
di: Xu, Xianglong, et al.
Pubblicazione: (2025)
Explainability-Based Token Replacement on LLM-Generated Text
di: Mohammadi, Hadi, et al.
Pubblicazione: (2025)
di: Mohammadi, Hadi, et al.
Pubblicazione: (2025)
Training Text-to-Molecule Models with Context-Aware Tokenization
di: Kim, Seojin, et al.
Pubblicazione: (2025)
di: Kim, Seojin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
di: Ren, Houxing, et al.
Pubblicazione: (2024) -
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
di: Wang, Ke, et al.
Pubblicazione: (2025) -
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
di: Lu, Zimu, et al.
Pubblicazione: (2024) -
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
di: Lu, Zimu, et al.
Pubblicazione: (2025) -
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
di: Lu, Zimu, et al.
Pubblicazione: (2024)