BERTer: The Efficient One
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saligram, Pradyumna, Lanpouthakoun, Andrew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PreFT: Prefill-only finetuning for efficient inference
von: Lanpouthakoun, Andrew, et al.
Veröffentlicht: (2026)
von: Lanpouthakoun, Andrew, et al.
Veröffentlicht: (2026)
When Two LLMs Debate, Both Think They'll Win
von: Prasad, Pradyumna Shyama, et al.
Veröffentlicht: (2025)
von: Prasad, Pradyumna Shyama, et al.
Veröffentlicht: (2025)
One Size Fits None: Heuristic Collapse in LLM Investment Advice
von: Ross, Jillian, et al.
Veröffentlicht: (2026)
von: Ross, Jillian, et al.
Veröffentlicht: (2026)
EmissionNet: Air Quality Pollution Forecasting for Agriculture
von: Saligram, Prady, et al.
Veröffentlicht: (2025)
von: Saligram, Prady, et al.
Veröffentlicht: (2025)
H1B-KV: Hybrid One-Bit Caches for Memory-Efficient Large Language Model Inference
von: Vejendla, Harshil
Veröffentlicht: (2025)
von: Vejendla, Harshil
Veröffentlicht: (2025)
DLM-One: Diffusion Language Models for One-Step Sequence Generation
von: Chen, Tianqi, et al.
Veröffentlicht: (2025)
von: Chen, Tianqi, et al.
Veröffentlicht: (2025)
Data Shapley in One Training Run
von: Wang, Jiachen T., et al.
Veröffentlicht: (2024)
von: Wang, Jiachen T., et al.
Veröffentlicht: (2024)
Proofread: Fixes All Errors with One Tap
von: Liu, Renjie, et al.
Veröffentlicht: (2024)
von: Liu, Renjie, et al.
Veröffentlicht: (2024)
One Token to Fool LLM-as-a-Judge
von: Zhao, Yulai, et al.
Veröffentlicht: (2025)
von: Zhao, Yulai, et al.
Veröffentlicht: (2025)
Improving Search Agent with One Line of Code
von: Li, Jian, et al.
Veröffentlicht: (2026)
von: Li, Jian, et al.
Veröffentlicht: (2026)
Flextron: Many-in-One Flexible Large Language Model
von: Cai, Ruisi, et al.
Veröffentlicht: (2024)
von: Cai, Ruisi, et al.
Veröffentlicht: (2024)
Provable Knowledge Acquisition and Extraction in One-Layer Transformers
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
Out of One, Many: Using Language Models to Simulate Human Samples
von: Argyle, Lisa P., et al.
Veröffentlicht: (2022)
von: Argyle, Lisa P., et al.
Veröffentlicht: (2022)
OneGen: Efficient One-Pass Unified Generation and Retrieval for LLMs
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
One size doesn't fit all: Predicting the Number of Examples for In-Context Learning
von: Chandra, Manish, et al.
Veröffentlicht: (2024)
von: Chandra, Manish, et al.
Veröffentlicht: (2024)
BitDelta: Your Fine-Tune May Only Be Worth One Bit
von: Liu, James, et al.
Veröffentlicht: (2024)
von: Liu, James, et al.
Veröffentlicht: (2024)
GOFA: A Generative One-For-All Model for Joint Graph Language Modeling
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
2048: Reinforcement Learning in a Delayed Reward Environment
von: Saligram, Prady, et al.
Veröffentlicht: (2025)
von: Saligram, Prady, et al.
Veröffentlicht: (2025)
On the Semantic and Syntactic Information Encoded in Proto-Tokens for One-Step Text Reconstruction
von: Bondarenko, Ivan, et al.
Veröffentlicht: (2026)
von: Bondarenko, Ivan, et al.
Veröffentlicht: (2026)
Learning Multiplex Representations on Text-Attributed Graphs with One Language Model Encoder
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information
von: Tutnov, Rasul, et al.
Veröffentlicht: (2025)
von: Tutnov, Rasul, et al.
Veröffentlicht: (2025)
Many Minds from One Model: Bayesian-Inspired Transformers for Population Diversity
von: Yang, Diji, et al.
Veröffentlicht: (2025)
von: Yang, Diji, et al.
Veröffentlicht: (2025)
One-Pass to Reason: Token Duplication and Block-Sparse Mask for Efficient Fine-Tuning on Multi-Turn Reasoning
von: Goru, Ritesh, et al.
Veröffentlicht: (2025)
von: Goru, Ritesh, et al.
Veröffentlicht: (2025)
One Pass Streaming Algorithm for Super Long Token Attention Approximation in Sublinear Space
von: Addanki, Raghav, et al.
Veröffentlicht: (2023)
von: Addanki, Raghav, et al.
Veröffentlicht: (2023)
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones
von: Zhmoginov, Andrey, et al.
Veröffentlicht: (2025)
von: Zhmoginov, Andrey, et al.
Veröffentlicht: (2025)
Trust in One Round: Confidence Estimation for Large Language Models via Structural Signals
von: Yang, Pengyue, et al.
Veröffentlicht: (2026)
von: Yang, Pengyue, et al.
Veröffentlicht: (2026)
Adaptive Two Sided Laplace Transforms: A Learnable, Interpretable, and Scalable Replacement for Self-Attention
von: Kiruluta, Andrew
Veröffentlicht: (2025)
von: Kiruluta, Andrew
Veröffentlicht: (2025)
Data-Driven Variational Basis Learning Beyond Neural Networks: A Non-Neural Framework for Adaptive Basis Discovery
von: Kiruluta, Andrew
Veröffentlicht: (2026)
von: Kiruluta, Andrew
Veröffentlicht: (2026)
Spectral Generative Flow Models: A Physics-Inspired Replacement for Vectorized Large Language Models
von: Kiruluta, Andrew
Veröffentlicht: (2026)
von: Kiruluta, Andrew
Veröffentlicht: (2026)
Entropic-Time Inference: Self-Organizing Large Language Model Decoding Beyond Attention
von: Kiruluta, Andrew
Veröffentlicht: (2026)
von: Kiruluta, Andrew
Veröffentlicht: (2026)
From Gradients to Riccati Geometry: Kalman World Models for Single-Pass Learning
von: Kiruluta, Andrew
Veröffentlicht: (2026)
von: Kiruluta, Andrew
Veröffentlicht: (2026)
One Sample to Rule Them All: Extreme Data Efficiency in Multidiscipline Reasoning with Reinforcement Learning
von: Li, Yiyuan, et al.
Veröffentlicht: (2026)
von: Li, Yiyuan, et al.
Veröffentlicht: (2026)
ROSE: Reordered SparseGPT for More Accurate One-Shot Large Language Models Pruning
von: Su, Mingluo, et al.
Veröffentlicht: (2026)
von: Su, Mingluo, et al.
Veröffentlicht: (2026)
How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models
von: Schwethelm, Kristian, et al.
Veröffentlicht: (2026)
von: Schwethelm, Kristian, et al.
Veröffentlicht: (2026)
Fact or Guesswork? Evaluating Large Language Models' Medical Knowledge with Structured One-Hop Judgments
von: Li, Jiaxi, et al.
Veröffentlicht: (2025)
von: Li, Jiaxi, et al.
Veröffentlicht: (2025)
Scaling Efficient LLMs
von: Kausik, B. N.
Veröffentlicht: (2024)
von: Kausik, B. N.
Veröffentlicht: (2024)
Efficient Reasoning on the Edge
von: Bondarenko, Yelysei, et al.
Veröffentlicht: (2026)
von: Bondarenko, Yelysei, et al.
Veröffentlicht: (2026)
Enhancing One-shot Pruned Pre-trained Language Models through Sparse-Dense-Sparse Mechanism
von: Li, Guanchen, et al.
Veröffentlicht: (2024)
von: Li, Guanchen, et al.
Veröffentlicht: (2024)
In-Context Learning of a Linear Transformer Block: Benefits of the MLP Component and One-Step GD Initialization
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PreFT: Prefill-only finetuning for efficient inference
von: Lanpouthakoun, Andrew, et al.
Veröffentlicht: (2026) -
When Two LLMs Debate, Both Think They'll Win
von: Prasad, Pradyumna Shyama, et al.
Veröffentlicht: (2025) -
One Size Fits None: Heuristic Collapse in LLM Investment Advice
von: Ross, Jillian, et al.
Veröffentlicht: (2026) -
EmissionNet: Air Quality Pollution Forecasting for Agriculture
von: Saligram, Prady, et al.
Veröffentlicht: (2025) -
H1B-KV: Hybrid One-Bit Caches for Memory-Efficient Large Language Model Inference
von: Vejendla, Harshil
Veröffentlicht: (2025)