Regression-aware Inference with LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Lukasik, Michal, Narasimhan, Harikrishna, Menon, Aditya Krishna, Yu, Felix, Kumar, Sanjiv |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Language Model Cascades: Token-level uncertainty and beyond
por: Gupta, Neha, et al.
Publicado: (2024)
por: Gupta, Neha, et al.
Publicado: (2024)
Faster Cascades via Speculative Decoding
por: Narasimhan, Harikrishna, et al.
Publicado: (2024)
por: Narasimhan, Harikrishna, et al.
Publicado: (2024)
Bipartite Ranking From Multiple Labels: On Loss Versus Label Aggregation
por: Lukasik, Michal, et al.
Publicado: (2025)
por: Lukasik, Michal, et al.
Publicado: (2025)
Think before you speak: Training Language Models With Pause Tokens
por: Goyal, Sachin, et al.
Publicado: (2023)
por: Goyal, Sachin, et al.
Publicado: (2023)
Tandem Transformers for Inference Efficient LLMs
por: S, Aishwarya P, et al.
Publicado: (2024)
por: S, Aishwarya P, et al.
Publicado: (2024)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
por: Zhou, Yongchao, et al.
Publicado: (2023)
por: Zhou, Yongchao, et al.
Publicado: (2023)
A Critical Study of What Code-LLMs (Do Not) Learn
por: Anand, Abhinav, et al.
Publicado: (2024)
por: Anand, Abhinav, et al.
Publicado: (2024)
Universal Model Routing for Efficient LLM Inference
por: Jitkrittum, Wittawat, et al.
Publicado: (2025)
por: Jitkrittum, Wittawat, et al.
Publicado: (2025)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
por: Banerjee, Tanushree, et al.
Publicado: (2024)
por: Banerjee, Tanushree, et al.
Publicado: (2024)
Activation-aware Probe-Query: Effective Key-Value Retrieval for Long-Context LLMs Inference
por: Xiao, Qingfa, et al.
Publicado: (2025)
por: Xiao, Qingfa, et al.
Publicado: (2025)
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
por: Dutta, Aritra, et al.
Publicado: (2026)
por: Dutta, Aritra, et al.
Publicado: (2026)
MIR: Methodology Inspiration Retrieval for Scientific Research Problems
por: Garikaparthi, Aniketh, et al.
Publicado: (2025)
por: Garikaparthi, Aniketh, et al.
Publicado: (2025)
LLMs cannot spot math errors, even when allowed to peek into the solution
por: Srivatsa, KV Aditya, et al.
Publicado: (2025)
por: Srivatsa, KV Aditya, et al.
Publicado: (2025)
Can LLMs Reliably Simulate Real Students' Abilities in Mathematics and Reading Comprehension?
por: Srivatsa, KV Aditya, et al.
Publicado: (2025)
por: Srivatsa, KV Aditya, et al.
Publicado: (2025)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
por: Kumar, Anantha Padmanaban Krishna
Publicado: (2025)
por: Kumar, Anantha Padmanaban Krishna
Publicado: (2025)
When Benchmarks Leak: Inference-Time Decontamination for LLMs
por: Chai, Jianzhe, et al.
Publicado: (2026)
por: Chai, Jianzhe, et al.
Publicado: (2026)
Concept-aware Data Construction Improves In-context Learning of Language Models
por: Štefánik, Michal, et al.
Publicado: (2024)
por: Štefánik, Michal, et al.
Publicado: (2024)
Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition
por: Juvekar, Kush, et al.
Publicado: (2026)
por: Juvekar, Kush, et al.
Publicado: (2026)
Deep sequence models tend to memorize geometrically; it is unclear why
por: Noroozizadeh, Shahriar, et al.
Publicado: (2025)
por: Noroozizadeh, Shahriar, et al.
Publicado: (2025)
On student-teacher deviations in distillation: does it pay to disobey?
por: Nagarajan, Vaishnavh, et al.
Publicado: (2023)
por: Nagarajan, Vaishnavh, et al.
Publicado: (2023)
LLMs Are Prone to Fallacies in Causal Inference
por: Joshi, Nitish, et al.
Publicado: (2024)
por: Joshi, Nitish, et al.
Publicado: (2024)
Humans and LLMs Diverge on Probabilistic Inferences
por: Kamath, Gaurav, et al.
Publicado: (2026)
por: Kamath, Gaurav, et al.
Publicado: (2026)
Draft-based Approximate Inference for LLMs
por: Galim, Kevin, et al.
Publicado: (2025)
por: Galim, Kevin, et al.
Publicado: (2025)
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders
por: Zhu, Xiaofeng, et al.
Publicado: (2024)
por: Zhu, Xiaofeng, et al.
Publicado: (2024)
Fairness Evaluation and Inference Level Mitigation in LLMs
por: Nadeem, Afrozah, et al.
Publicado: (2025)
por: Nadeem, Afrozah, et al.
Publicado: (2025)
Decompose, Enrich, and Extract! Schema-aware Event Extraction using LLMs
por: Shiri, Fatemeh, et al.
Publicado: (2024)
por: Shiri, Fatemeh, et al.
Publicado: (2024)
POPI: Personalizing LLMs via Optimized Natural Language Preference Inference
por: Chen, Yizhuo, et al.
Publicado: (2025)
por: Chen, Yizhuo, et al.
Publicado: (2025)
Cascade-Aware Training of Language Models
por: Wang, Congchao, et al.
Publicado: (2024)
por: Wang, Congchao, et al.
Publicado: (2024)
Can Language Models Solve Olympiad Programming?
por: Shi, Quan, et al.
Publicado: (2024)
por: Shi, Quan, et al.
Publicado: (2024)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
por: Spiegel, Michal, et al.
Publicado: (2024)
por: Spiegel, Michal, et al.
Publicado: (2024)
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
por: Yao, Shunyu, et al.
Publicado: (2024)
por: Yao, Shunyu, et al.
Publicado: (2024)
Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs
por: Sinha, Aditya, et al.
Publicado: (2026)
por: Sinha, Aditya, et al.
Publicado: (2026)
Path-Consistency with Prefix Enhancement for Efficient Inference in LLMs
por: Zhu, Jiace, et al.
Publicado: (2024)
por: Zhu, Jiace, et al.
Publicado: (2024)
Self-controller: Controlling LLMs with Multi-round Step-by-step Self-awareness
por: Peng, Xiao, et al.
Publicado: (2024)
por: Peng, Xiao, et al.
Publicado: (2024)
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
por: Chehade, Mohamad, et al.
Publicado: (2025)
por: Chehade, Mohamad, et al.
Publicado: (2025)
Agent Context Protocols Enhance Collective Inference
por: Bhardwaj, Devansh, et al.
Publicado: (2025)
por: Bhardwaj, Devansh, et al.
Publicado: (2025)
LoRA Done RITE: Robust Invariant Transformation Equilibration for LoRA Optimization
por: Yen, Jui-Nan, et al.
Publicado: (2024)
por: Yen, Jui-Nan, et al.
Publicado: (2024)
DAIQ: Auditing Demographic Attribute Inference from Question in LLMs
por: Panda, Srikant, et al.
Publicado: (2025)
por: Panda, Srikant, et al.
Publicado: (2025)
Infinite Problem Generator: Verifiably Scaling Physics Reasoning Data with Agentic Workflows
por: Sharan, Aditya, et al.
Publicado: (2026)
por: Sharan, Aditya, et al.
Publicado: (2026)
Culturally Responsive Artificial Intelligence -- Problems, Challenges and Solutions
por: Ożegalska-Łukasik, Natalia, et al.
Publicado: (2023)
por: Ożegalska-Łukasik, Natalia, et al.
Publicado: (2023)
Ejemplares similares
-
Language Model Cascades: Token-level uncertainty and beyond
por: Gupta, Neha, et al.
Publicado: (2024) -
Faster Cascades via Speculative Decoding
por: Narasimhan, Harikrishna, et al.
Publicado: (2024) -
Bipartite Ranking From Multiple Labels: On Loss Versus Label Aggregation
por: Lukasik, Michal, et al.
Publicado: (2025) -
Think before you speak: Training Language Models With Pause Tokens
por: Goyal, Sachin, et al.
Publicado: (2023) -
Tandem Transformers for Inference Efficient LLMs
por: S, Aishwarya P, et al.
Publicado: (2024)