Gespeichert in:
| Hauptverfasser: | Ma, Xuezhe, Yang, Xiaomeng, Xiong, Wenhan, Chen, Beidi, Yu, Lili, Zhang, Hao, May, Jonathan, Zettlemoyer, Luke, Levy, Omer, Zhou, Chunting |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2404.08801 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
von: Zhou, Chunting, et al.
Veröffentlicht: (2024)
von: Zhou, Chunting, et al.
Veröffentlicht: (2024)
LMFusion: Adapting Pretrained Language Models for Multimodal Generation
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
Self-Alignment with Instruction Backtranslation
von: Li, Xian, et al.
Veröffentlicht: (2023)
von: Li, Xian, et al.
Veröffentlicht: (2023)
Gecko: An Efficient Neural Architecture Inherently Processing Sequences with Arbitrary Lengths
von: Ma, Xuezhe, et al.
Veröffentlicht: (2026)
von: Ma, Xuezhe, et al.
Veröffentlicht: (2026)
CAT: Content-Adaptive Image Tokenization
von: Shen, Junhong, et al.
Veröffentlicht: (2025)
von: Shen, Junhong, et al.
Veröffentlicht: (2025)
ALMA: Alignment with Minimal Annotation
von: Yasunaga, Michihiro, et al.
Veröffentlicht: (2024)
von: Yasunaga, Michihiro, et al.
Veröffentlicht: (2024)
In-context Pretraining: Language Modeling Beyond Document Boundaries
von: Shi, Weijia, et al.
Veröffentlicht: (2023)
von: Shi, Weijia, et al.
Veröffentlicht: (2023)
LLM The Genius Paradox: A Linguistic and Math Expert's Struggle with Simple Word-based Counting Problems
von: Xu, Nan, et al.
Veröffentlicht: (2024)
von: Xu, Nan, et al.
Veröffentlicht: (2024)
Towards Chapter-to-Chapter Context-Aware Literary Translation via Large Language Models
von: Jin, Linghao, et al.
Veröffentlicht: (2024)
von: Jin, Linghao, et al.
Veröffentlicht: (2024)
SpecExec: Massively Parallel Speculative Decoding for Interactive LLM Inference on Consumer Devices
von: Svirschevski, Ruslan, et al.
Veröffentlicht: (2024)
von: Svirschevski, Ruslan, et al.
Veröffentlicht: (2024)
Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities
von: Wen, Nuan, et al.
Veröffentlicht: (2026)
von: Wen, Nuan, et al.
Veröffentlicht: (2026)
Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models
von: Liang, Weixin, et al.
Veröffentlicht: (2024)
von: Liang, Weixin, et al.
Veröffentlicht: (2024)
GSM-Infinite: How Do Your LLMs Behave over Infinitely Increasing Context Length and Reasoning Complexity?
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
Efficient Pretraining Length Scaling
von: Wu, Bohong, et al.
Veröffentlicht: (2025)
von: Wu, Bohong, et al.
Veröffentlicht: (2025)
Beyond Length: Quantifying Long-Range Information for Long-Context LLM Pretraining Data
von: Deng, Haoran, et al.
Veröffentlicht: (2025)
von: Deng, Haoran, et al.
Veröffentlicht: (2025)
Craw4LLM: Efficient Web Crawling for LLM Pretraining
von: Yu, Shi, et al.
Veröffentlicht: (2025)
von: Yu, Shi, et al.
Veröffentlicht: (2025)
Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity
von: Liang, Weixin, et al.
Veröffentlicht: (2025)
von: Liang, Weixin, et al.
Veröffentlicht: (2025)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
von: Dong, Harry, et al.
Veröffentlicht: (2024)
von: Dong, Harry, et al.
Veröffentlicht: (2024)
DecoPrompt : Decoding Prompts Reduces Hallucinations when Large Language Models Meet False Premises
von: Xu, Nan, et al.
Veröffentlicht: (2024)
von: Xu, Nan, et al.
Veröffentlicht: (2024)
Get More with LESS: Synthesizing Recurrence with KV Cache Compression for Efficient LLM Inference
von: Dong, Harry, et al.
Veröffentlicht: (2024)
von: Dong, Harry, et al.
Veröffentlicht: (2024)
Byte Latent Transformer: Patches Scale Better Than Tokens
von: Pagnoni, Artidoro, et al.
Veröffentlicht: (2024)
von: Pagnoni, Artidoro, et al.
Veröffentlicht: (2024)
Megalodon, mako shark and planktonic foraminifera from the continental shelf off Portugal and their age
von: M.T. ANTUNES
Veröffentlicht: (2015)
von: M.T. ANTUNES
Veröffentlicht: (2015)
Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling
von: Ren, Liliang, et al.
Veröffentlicht: (2024)
von: Ren, Liliang, et al.
Veröffentlicht: (2024)
Squeezed Attention: Accelerating Long Context Length LLM Inference
von: Hooper, Coleman, et al.
Veröffentlicht: (2024)
von: Hooper, Coleman, et al.
Veröffentlicht: (2024)
LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
von: Han, Chi, et al.
Veröffentlicht: (2023)
von: Han, Chi, et al.
Veröffentlicht: (2023)
ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
von: Sun, Hanshi, et al.
Veröffentlicht: (2024)
von: Sun, Hanshi, et al.
Veröffentlicht: (2024)
Art Unlimited?
von: Schultheis, Franz, et al.
Veröffentlicht: (2016)
von: Schultheis, Franz, et al.
Veröffentlicht: (2016)
Detecting Pretraining Data from Large Language Models
von: Shi, Weijia, et al.
Veröffentlicht: (2023)
von: Shi, Weijia, et al.
Veröffentlicht: (2023)
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
Bootstrapping LLM Robustness for VLM Safety via Reducing the Pretraining Modality Gap
von: Yang, Wenhan, et al.
Veröffentlicht: (2025)
von: Yang, Wenhan, et al.
Veröffentlicht: (2025)
Keep Guessing? When Considering Inference Scaling, Mind the Baselines
von: Yona, Gal, et al.
Veröffentlicht: (2024)
von: Yona, Gal, et al.
Veröffentlicht: (2024)
MALI: Unlimited Mandate
Veröffentlicht: (2025)
Veröffentlicht: (2025)
Learning Center Unlimited.
von: Vivrette, Lyndon
Veröffentlicht: (1974)
von: Vivrette, Lyndon
Veröffentlicht: (1974)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
Comparing Hallucination Detection Metrics for Multilingual Generation
von: Kang, Haoqiang, et al.
Veröffentlicht: (2024)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2024)
Multimodal RewardBench: Holistic Evaluation of Reward Models for Vision Language Models
von: Yasunaga, Michihiro, et al.
Veröffentlicht: (2025)
von: Yasunaga, Michihiro, et al.
Veröffentlicht: (2025)
(Mis)Fitting: A Survey of Scaling Laws
von: Li, Margaret, et al.
Veröffentlicht: (2025)
von: Li, Margaret, et al.
Veröffentlicht: (2025)
PatentEdits: Framing Patent Novelty as Textual Entailment
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
von: Zhou, Chunting, et al.
Veröffentlicht: (2024) -
LMFusion: Adapting Pretrained Language Models for Multimodal Generation
von: Shi, Weijia, et al.
Veröffentlicht: (2024) -
Self-Alignment with Instruction Backtranslation
von: Li, Xian, et al.
Veröffentlicht: (2023) -
Gecko: An Efficient Neural Architecture Inherently Processing Sequences with Arbitrary Lengths
von: Ma, Xuezhe, et al.
Veröffentlicht: (2026) -
CAT: Content-Adaptive Image Tokenization
von: Shen, Junhong, et al.
Veröffentlicht: (2025)