Set-LLM: A Permutation-Invariant LLM
Fuente:
arXiv
Salvato in:
| Autori principali: | Egressy, Beni, Stühmer, Jan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PICASO: Permutation-Invariant Context Composition with State Space Models
di: Liu, Tian Yu, et al.
Pubblicazione: (2025)
di: Liu, Tian Yu, et al.
Pubblicazione: (2025)
Robustly Improving LLM Fairness in Realistic Settings via Interpretability
di: Karvonen, Adam, et al.
Pubblicazione: (2025)
di: Karvonen, Adam, et al.
Pubblicazione: (2025)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
di: Gupta, Manan, et al.
Pubblicazione: (2026)
di: Gupta, Manan, et al.
Pubblicazione: (2026)
Realistic Synthetic Financial Transactions for Anti-Money Laundering Models
di: Altman, Erik, et al.
Pubblicazione: (2023)
di: Altman, Erik, et al.
Pubblicazione: (2023)
Non-Determinism of "Deterministic" LLM Settings
di: Atil, Berk, et al.
Pubblicazione: (2024)
di: Atil, Berk, et al.
Pubblicazione: (2024)
Communication Compression for Tensor Parallel LLM Inference
di: Hansen-Palmus, Jan, et al.
Pubblicazione: (2024)
di: Hansen-Palmus, Jan, et al.
Pubblicazione: (2024)
Does Context Matter? ContextualJudgeBench for Evaluating LLM-based Judges in Contextual Settings
di: Xu, Austin, et al.
Pubblicazione: (2025)
di: Xu, Austin, et al.
Pubblicazione: (2025)
LLM Chemistry Estimation for Multi-LLM Recommendation
di: Sanchez, Huascar, et al.
Pubblicazione: (2025)
di: Sanchez, Huascar, et al.
Pubblicazione: (2025)
User-LLM: Efficient LLM Contextualization with User Embeddings
di: Ning, Lin, et al.
Pubblicazione: (2024)
di: Ning, Lin, et al.
Pubblicazione: (2024)
Enhancing NLP Robustness and Generalization through LLM-Generated Contrast Sets: A Scalable Framework for Systematic Evaluation and Adversarial Training
di: Lin, Hender
Pubblicazione: (2025)
di: Lin, Hender
Pubblicazione: (2025)
Stylometry recognizes human and LLM-generated texts in short samples
di: Przystalski, Karol, et al.
Pubblicazione: (2025)
di: Przystalski, Karol, et al.
Pubblicazione: (2025)
ToxiGAN: Toxic Data Augmentation via LLM-Guided Directional Adversarial Generation
di: Li, Peiran, et al.
Pubblicazione: (2026)
di: Li, Peiran, et al.
Pubblicazione: (2026)
MultiSoc-4D: A Benchmark for Diagnosing Instruction-Induced Label Collapse in Closed-Set LLM Annotation of Bengali Social Media
di: Pramanik, Souvik, et al.
Pubblicazione: (2026)
di: Pramanik, Souvik, et al.
Pubblicazione: (2026)
VBART: The Turkish LLM
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
di: Fu, Qichen, et al.
Pubblicazione: (2024)
di: Fu, Qichen, et al.
Pubblicazione: (2024)
BPO: Staying Close to the Behavior LLM Creates Better Online LLM Alignment
di: Xu, Wenda, et al.
Pubblicazione: (2024)
di: Xu, Wenda, et al.
Pubblicazione: (2024)
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models
di: Fang, Gongfan, et al.
Pubblicazione: (2024)
di: Fang, Gongfan, et al.
Pubblicazione: (2024)
LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
di: Jin, Hongye, et al.
Pubblicazione: (2024)
di: Jin, Hongye, et al.
Pubblicazione: (2024)
LLM See, LLM Do: Guiding Data Generation to Target Non-Differentiable Objectives
di: Shimabucoro, Luísa, et al.
Pubblicazione: (2024)
di: Shimabucoro, Luísa, et al.
Pubblicazione: (2024)
LLM Pruning and Distillation in Practice: The Minitron Approach
di: Sreenivas, Sharath Turuvekere, et al.
Pubblicazione: (2024)
di: Sreenivas, Sharath Turuvekere, et al.
Pubblicazione: (2024)
Understanding the planning of LLM agents: A survey
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
LLM Assistance for Pediatric Depression
di: Ignashina, Mariia, et al.
Pubblicazione: (2025)
di: Ignashina, Mariia, et al.
Pubblicazione: (2025)
Muon is Scalable for LLM Training
di: Liu, Jingyuan, et al.
Pubblicazione: (2025)
di: Liu, Jingyuan, et al.
Pubblicazione: (2025)
Learning Dynamics of LLM Finetuning
di: Ren, Yi, et al.
Pubblicazione: (2024)
di: Ren, Yi, et al.
Pubblicazione: (2024)
Understanding LLM Embeddings for Regression
di: Tang, Eric, et al.
Pubblicazione: (2024)
di: Tang, Eric, et al.
Pubblicazione: (2024)
To Believe or Not to Believe Your LLM
di: Yadkori, Yasin Abbasi, et al.
Pubblicazione: (2024)
di: Yadkori, Yasin Abbasi, et al.
Pubblicazione: (2024)
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
di: Huang, Wei, et al.
Pubblicazione: (2026)
di: Huang, Wei, et al.
Pubblicazione: (2026)
TransformLLM: Adapting Large Language Models via LLM-Transformed Reading Comprehension Text
di: Arbel, Iftach, et al.
Pubblicazione: (2024)
di: Arbel, Iftach, et al.
Pubblicazione: (2024)
Provably Powerful Graph Neural Networks for Directed Multigraphs
di: Egressy, Béni, et al.
Pubblicazione: (2023)
di: Egressy, Béni, et al.
Pubblicazione: (2023)
New Encoders for German Trained from Scratch: Comparing ModernGBERT with Converted LLM2Vec Models
di: Wunderle, Julia, et al.
Pubblicazione: (2025)
di: Wunderle, Julia, et al.
Pubblicazione: (2025)
Training Proactive and Personalized LLM Agents
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
Steer LLM Latents for Hallucination Detection
di: Park, Seongheon, et al.
Pubblicazione: (2025)
di: Park, Seongheon, et al.
Pubblicazione: (2025)
Multi-LLM Collaboration for Medication Recommendation
di: Sanchez, Huascar, et al.
Pubblicazione: (2025)
di: Sanchez, Huascar, et al.
Pubblicazione: (2025)
Turning LLM Activations Quantization-Friendly
di: Czakó, Patrik, et al.
Pubblicazione: (2025)
di: Czakó, Patrik, et al.
Pubblicazione: (2025)
Measuring all the noises of LLM Evals
di: Wang, Sida
Pubblicazione: (2025)
di: Wang, Sida
Pubblicazione: (2025)
Survey on Evaluation of LLM-based Agents
di: Yehudai, Asaf, et al.
Pubblicazione: (2025)
di: Yehudai, Asaf, et al.
Pubblicazione: (2025)
Explainable LLM Unlearning Through Reasoning
di: Liao, Junfeng, et al.
Pubblicazione: (2026)
di: Liao, Junfeng, et al.
Pubblicazione: (2026)
Collaboratively adding new knowledge to an LLM
di: Lee, Rhui Dih, et al.
Pubblicazione: (2024)
di: Lee, Rhui Dih, et al.
Pubblicazione: (2024)
Token-Budget-Aware LLM Reasoning
di: Han, Tingxu, et al.
Pubblicazione: (2024)
di: Han, Tingxu, et al.
Pubblicazione: (2024)
BAGEN: Are LLM Agents Budget-Aware?
di: Lin, Yuxiang, et al.
Pubblicazione: (2026)
di: Lin, Yuxiang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
PICASO: Permutation-Invariant Context Composition with State Space Models
di: Liu, Tian Yu, et al.
Pubblicazione: (2025) -
Robustly Improving LLM Fairness in Realistic Settings via Interpretability
di: Karvonen, Adam, et al.
Pubblicazione: (2025) -
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
di: Gupta, Manan, et al.
Pubblicazione: (2026) -
Realistic Synthetic Financial Transactions for Anti-Money Laundering Models
di: Altman, Erik, et al.
Pubblicazione: (2023) -
Non-Determinism of "Deterministic" LLM Settings
di: Atil, Berk, et al.
Pubblicazione: (2024)