Enregistré dans:
| Auteurs principaux: | Li, Yuchen, Kirchmeyer, Alexandre, Mehta, Aashay, Qin, Yilong, Dadachev, Boris, Papineni, Kishore, Kumar, Sanjiv, Risteski, Andrej |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2407.21046 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
On the Query Complexity of Verifier-Assisted Language Generation
par: Botta, Edoardo, et autres
Publié: (2025)
par: Botta, Edoardo, et autres
Publié: (2025)
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
par: Qin, Yilong, et autres
Publié: (2023)
par: Qin, Yilong, et autres
Publié: (2023)
Exploring and Improving Drafts in Blockwise Parallel Decoding
par: Kim, Taehyeon, et autres
Publié: (2024)
par: Kim, Taehyeon, et autres
Publié: (2024)
CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
par: Li, Shanda, et autres
Publié: (2025)
par: Li, Shanda, et autres
Publié: (2025)
Large Language Models for Stemming: Promises, Pitfalls and Failures
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
Toward Understanding the Transferability of Adversarial Suffixes in Large Language Models
par: Ball, Sarah, et autres
Publié: (2025)
par: Ball, Sarah, et autres
Publié: (2025)
The Promises and Pitfalls of Using Language Models to Measure Instruction Quality in Education
par: Xu, Paiheng, et autres
Publié: (2024)
par: Xu, Paiheng, et autres
Publié: (2024)
A 2-step Framework for Automated Literary Translation Evaluation: Its Promises and Pitfalls
par: Shafayat, Sheikh, et autres
Publié: (2024)
par: Shafayat, Sheikh, et autres
Publié: (2024)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
par: Anzenberg, Eitan, et autres
Publié: (2025)
par: Anzenberg, Eitan, et autres
Publié: (2025)
A computational phase transition for learning-to-sample from Ising models
par: Risteski, Andrej, et autres
Publié: (2026)
par: Risteski, Andrej, et autres
Publié: (2026)
PPLqa: An Unsupervised Information-Theoretic Quality Metric for Comparing Generative Large Language Models
par: Friedland, Gerald, et autres
Publié: (2024)
par: Friedland, Gerald, et autres
Publié: (2024)
A Large Language Model-based Framework for Semi-Structured Tender Document Retrieval-Augmented Generation
par: Zhao, Yilong, et autres
Publié: (2024)
par: Zhao, Yilong, et autres
Publié: (2024)
Benefits and Pitfalls of Reinforcement Learning for Language Model Planning: A Theoretical Perspective
par: Wang, Siwei, et autres
Publié: (2025)
par: Wang, Siwei, et autres
Publié: (2025)
The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives
par: Moitra, Ankur, et autres
Publié: (2026)
par: Moitra, Ankur, et autres
Publié: (2026)
Steering diffusion models with quadratic rewards: a fine-grained analysis
par: Moitra, Ankur, et autres
Publié: (2026)
par: Moitra, Ankur, et autres
Publié: (2026)
The Promises and Pitfalls of LLM Annotations in Dataset Labeling: a Case Study on Media Bias Detection
par: Horych, Tomas, et autres
Publié: (2024)
par: Horych, Tomas, et autres
Publié: (2024)
The Perils & Promises of Fact-checking with Large Language Models
par: Quelle, Dorian, et autres
Publié: (2023)
par: Quelle, Dorian, et autres
Publié: (2023)
Pitfalls of Evaluating Language Models with Open Benchmarks
par: Hasan, Md. Najib, et autres
Publié: (2025)
par: Hasan, Md. Najib, et autres
Publié: (2025)
MOTOR: Multimodal Optimal Transport via Grounded Retrieval in Medical Visual Question Answering
par: Shaaban, Mai A., et autres
Publié: (2025)
par: Shaaban, Mai A., et autres
Publié: (2025)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
par: Sam, Dylan, et autres
Publié: (2025)
par: Sam, Dylan, et autres
Publié: (2025)
From Guidelines to Practice: A New Paradigm for Arabic Language Model Evaluation
par: Sibaee, Serry, et autres
Publié: (2025)
par: Sibaee, Serry, et autres
Publié: (2025)
From Explainability to Action: A Generative Operational Framework for Integrating XAI in Clinical Mental Health Screening
par: Kandala, Ratna, et autres
Publié: (2025)
par: Kandala, Ratna, et autres
Publié: (2025)
The Pitfalls of Defining Hallucination
par: van Deemter, Kees
Publié: (2024)
par: van Deemter, Kees
Publié: (2024)
MultiCheck: Strengthening Web Trust with Unified Multimodal Fact Verification
par: Kishore, Aditya, et autres
Publié: (2025)
par: Kishore, Aditya, et autres
Publié: (2025)
What Makes an Evaluation Useful? Common Pitfalls and Best Practices
par: Gekker, Gil, et autres
Publié: (2025)
par: Gekker, Gil, et autres
Publié: (2025)
Masked Mixers for Language Generation and Retrieval
par: Badger, Benjamin L.
Publié: (2024)
par: Badger, Benjamin L.
Publié: (2024)
Deep sequence models tend to memorize geometrically; it is unclear why
par: Noroozizadeh, Shahriar, et autres
Publié: (2025)
par: Noroozizadeh, Shahriar, et autres
Publié: (2025)
Beware of Reasoning Overconfidence: Pitfalls in the Reasoning Process for Multi-solution Tasks
par: Guan, Jiannan, et autres
Publié: (2025)
par: Guan, Jiannan, et autres
Publié: (2025)
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
par: Lahnala, Allison, et autres
Publié: (2025)
par: Lahnala, Allison, et autres
Publié: (2025)
Pitfalls of Scale: Investigating the Inverse Task of Redefinition in Large Language Models
par: Stringli, Elena, et autres
Publié: (2025)
par: Stringli, Elena, et autres
Publié: (2025)
Pitfalls and Outlooks in Using COMET
par: Zouhar, Vilém, et autres
Publié: (2024)
par: Zouhar, Vilém, et autres
Publié: (2024)
Language Model Cascades: Token-level uncertainty and beyond
par: Gupta, Neha, et autres
Publié: (2024)
par: Gupta, Neha, et autres
Publié: (2024)
Enhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice Guidelines
par: Oniani, David, et autres
Publié: (2024)
par: Oniani, David, et autres
Publié: (2024)
XGrammar: Flexible and Efficient Structured Generation Engine for Large Language Models
par: Dong, Yixin, et autres
Publié: (2024)
par: Dong, Yixin, et autres
Publié: (2024)
Think before you speak: Training Language Models With Pause Tokens
par: Goyal, Sachin, et autres
Publié: (2023)
par: Goyal, Sachin, et autres
Publié: (2023)
Representations as Language: An Information-Theoretic Framework for Interpretability
par: Conklin, Henry, et autres
Publié: (2024)
par: Conklin, Henry, et autres
Publié: (2024)
Refine Medical Diagnosis Using Generation Augmented Retrieval and Clinical Practice Guidelines
par: Li, Wenhao, et autres
Publié: (2025)
par: Li, Wenhao, et autres
Publié: (2025)
Entailed Opinion Matters: Improving the Fact-Checking Performance of Language Models by Relying on their Entailment Ability
par: Kumar, Gaurav, et autres
Publié: (2025)
par: Kumar, Gaurav, et autres
Publié: (2025)
CLA Guidelines on Non-Sexist Language
Publié: (1976)
Publié: (1976)
On the Benefits of Memory for Modeling Time-Dependent PDEs
par: Ruiz, Ricardo Buitrago, et autres
Publié: (2024)
par: Ruiz, Ricardo Buitrago, et autres
Publié: (2024)
Documents similaires
-
On the Query Complexity of Verifier-Assisted Language Generation
par: Botta, Edoardo, et autres
Publié: (2025) -
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
par: Qin, Yilong, et autres
Publié: (2023) -
Exploring and Improving Drafts in Blockwise Parallel Decoding
par: Kim, Taehyeon, et autres
Publié: (2024) -
CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
par: Li, Shanda, et autres
Publié: (2025) -
Large Language Models for Stemming: Promises, Pitfalls and Failures
par: Wang, Shuai, et autres
Publié: (2024)