Gespeichert in:
| Hauptverfasser: | Remfry, Elizabeth, Henkin, Rafael, Barnes, Michael R, Naik, Aakanksha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2412.01331 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating Collaborative Data Practices: a Case Study on Artificial Intelligence for Healthcare Research
von: Henkin, Rafael, et al.
Veröffentlicht: (2023)
von: Henkin, Rafael, et al.
Veröffentlicht: (2023)
Intent-Aware Schema Generation And Refinement For Literature Review Tables
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
Distilling Multi-Scale Knowledge for Event Temporal Relation Extraction
von: Yao, Hao-Ren, et al.
Veröffentlicht: (2022)
von: Yao, Hao-Ren, et al.
Veröffentlicht: (2022)
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
von: Aakanksha, et al.
Veröffentlicht: (2024)
von: Aakanksha, et al.
Veröffentlicht: (2024)
Towards Compositionality in Concept Learning
von: Stein, Adam, et al.
Veröffentlicht: (2024)
von: Stein, Adam, et al.
Veröffentlicht: (2024)
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
von: Naik, Ninad
Veröffentlicht: (2024)
von: Naik, Ninad
Veröffentlicht: (2024)
Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential
von: Samragh, Mohammad, et al.
Veröffentlicht: (2025)
von: Samragh, Mohammad, et al.
Veröffentlicht: (2025)
Do We Need Frontier Models to Verify Mathematical Proofs?
von: Naik, Aaditya, et al.
Veröffentlicht: (2026)
von: Naik, Aaditya, et al.
Veröffentlicht: (2026)
ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution
von: Dong, Zican, et al.
Veröffentlicht: (2026)
von: Dong, Zican, et al.
Veröffentlicht: (2026)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
Evaluating Very Long-Term Conversational Memory of LLM Agents
von: Maharana, Adyasha, et al.
Veröffentlicht: (2024)
von: Maharana, Adyasha, et al.
Veröffentlicht: (2024)
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference
von: Bhendawade, Nikhil, et al.
Veröffentlicht: (2025)
von: Bhendawade, Nikhil, et al.
Veröffentlicht: (2025)
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
von: Aakanksha, et al.
Veröffentlicht: (2024)
von: Aakanksha, et al.
Veröffentlicht: (2024)
Towards a Realistic Long-Term Benchmark for Open-Web Research Agents
von: Mühlbacher, Peter, et al.
Veröffentlicht: (2024)
von: Mühlbacher, Peter, et al.
Veröffentlicht: (2024)
LongPO: Long Context Self-Evolution of Large Language Models through Short-to-Long Preference Optimization
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
Probe and Skip: Self-Predictive Token Skipping for Efficient Long-Context LLM Inference
von: Wu, Zimeng, et al.
Veröffentlicht: (2026)
von: Wu, Zimeng, et al.
Veröffentlicht: (2026)
When Routine Chats Turn Toxic: Unintended Long-Term State Poisoning in Personalized Agents
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2026)
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2026)
Large Language Models are Learnable Planners for Long-Term Recommendation
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
Exploring Contrastive Learning for Long-Tailed Multi-Label Text Classification
von: Audibert, Alexandre, et al.
Veröffentlicht: (2024)
von: Audibert, Alexandre, et al.
Veröffentlicht: (2024)
DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA
von: Yin, Jianing, et al.
Veröffentlicht: (2026)
von: Yin, Jianing, et al.
Veröffentlicht: (2026)
Latent Traits and Cross-Task Transfer: Deconstructing Dataset Interactions in LLM Fine-tuning
von: Krishna, Shambhavi, et al.
Veröffentlicht: (2025)
von: Krishna, Shambhavi, et al.
Veröffentlicht: (2025)
HoPE: A Novel Positional Encoding Without Long-Term Decay for Enhanced Context Awareness and Extrapolation
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
Long Context RAG Performance of Large Language Models
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
Once Upon an Input: Reasoning via Per-Instance Program Synthesis
von: Stein, Adam, et al.
Veröffentlicht: (2025)
von: Stein, Adam, et al.
Veröffentlicht: (2025)
Data Augmentation for Code Translation with Comparable Corpora and Multiple References
von: Xie, Yiqing, et al.
Veröffentlicht: (2023)
von: Xie, Yiqing, et al.
Veröffentlicht: (2023)
Statistical NLP for Optimization of Clinical Trial Success Prediction in Pharmaceutical R&D
von: Doane, Michael R.
Veröffentlicht: (2025)
von: Doane, Michael R.
Veröffentlicht: (2025)
Predicting Evoked Emotions in Conversations
von: Altarawneh, Enas, et al.
Veröffentlicht: (2023)
von: Altarawneh, Enas, et al.
Veröffentlicht: (2023)
Multipole Attention for Efficient Long Context Reasoning
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
Benchmarking Uncertainty Calibration in Large Language Model Long-Form Question Answering
von: Müller, Philip, et al.
Veröffentlicht: (2026)
von: Müller, Philip, et al.
Veröffentlicht: (2026)
Idea2Plan: Exploring AI-Powered Research Planning
von: Huang, Jin, et al.
Veröffentlicht: (2025)
von: Huang, Jin, et al.
Veröffentlicht: (2025)
Relationships are Complicated! An Analysis of Relationships Between Datasets on the Web
von: Lin, Kate, et al.
Veröffentlicht: (2024)
von: Lin, Kate, et al.
Veröffentlicht: (2024)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
LongEmbed: Extending Embedding Models for Long Context Retrieval
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
Long-Short Alignment for Effective Long-Context Modeling in LLMs
von: Du, Tianqi, et al.
Veröffentlicht: (2025)
von: Du, Tianqi, et al.
Veröffentlicht: (2025)
BioCoref: Benchmarking Biomedical Coreference Resolution with LLMs
von: Salem, Nourah M, et al.
Veröffentlicht: (2025)
von: Salem, Nourah M, et al.
Veröffentlicht: (2025)
Mnemosyne: An Unsupervised, Human-Inspired Long-Term Memory Architecture for Edge-Based LLMs
von: Jonelagadda, Aneesh, et al.
Veröffentlicht: (2025)
von: Jonelagadda, Aneesh, et al.
Veröffentlicht: (2025)
Double Equivariance for Inductive Link Prediction for Both New Nodes and New Relation Types
von: Zhou, Jincheng, et al.
Veröffentlicht: (2023)
von: Zhou, Jincheng, et al.
Veröffentlicht: (2023)
ComplicaCode: Enhancing Disease Complication Detection in Electronic Health Records through ICD Path Generation
von: Zhou, Xiaofan
Veröffentlicht: (2023)
von: Zhou, Xiaofan
Veröffentlicht: (2023)
LongReward: Improving Long-context Large Language Models with AI Feedback
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Investigating Collaborative Data Practices: a Case Study on Artificial Intelligence for Healthcare Research
von: Henkin, Rafael, et al.
Veröffentlicht: (2023) -
Intent-Aware Schema Generation And Refinement For Literature Review Tables
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025) -
Distilling Multi-Scale Knowledge for Event Temporal Relation Extraction
von: Yao, Hao-Ren, et al.
Veröffentlicht: (2022) -
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
von: Aakanksha, et al.
Veröffentlicht: (2024) -
Towards Compositionality in Concept Learning
von: Stein, Adam, et al.
Veröffentlicht: (2024)