Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Sorstkins, Andrejs |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
par: Qi, Jinhu, et autres
Publié: (2024)
par: Qi, Jinhu, et autres
Publié: (2024)
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
par: Saji, Alan, et autres
Publié: (2025)
par: Saji, Alan, et autres
Publié: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
par: Ashuach, Tomer, et autres
Publié: (2025)
par: Ashuach, Tomer, et autres
Publié: (2025)
HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools
par: Garg, Aashna, et autres
Publié: (2026)
par: Garg, Aashna, et autres
Publié: (2026)
PersonalLLM: Tailoring LLMs to Individual Preferences
par: Zollo, Thomas P., et autres
Publié: (2024)
par: Zollo, Thomas P., et autres
Publié: (2024)
Surprisingly Fragile: Assessing and Addressing Prompt Instability in Multimodal Foundation Models
par: Stewart, Ian, et autres
Publié: (2024)
par: Stewart, Ian, et autres
Publié: (2024)
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
par: Song, Chenyang, et autres
Publié: (2023)
par: Song, Chenyang, et autres
Publié: (2023)
Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis
par: Nagao, Moe, et autres
Publié: (2026)
par: Nagao, Moe, et autres
Publié: (2026)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
par: Simhi, Adi, et autres
Publié: (2024)
par: Simhi, Adi, et autres
Publié: (2024)
Augmenting Dialog with Think-Aloud Utterances for Modeling Individual Personality Traits by LLM
par: Ishikura, Seiya, et autres
Publié: (2025)
par: Ishikura, Seiya, et autres
Publié: (2025)
Comparing the Performance of LLMs in RAG-based Question-Answering: A Case Study in Computer Science Literature
par: Dayarathne, Ranul, et autres
Publié: (2025)
par: Dayarathne, Ranul, et autres
Publié: (2025)
LLMs Are Not Scorers: Rethinking MT Evaluation with Generation-Based Methods
par: Cui, Hyang
Publié: (2025)
par: Cui, Hyang
Publié: (2025)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
par: Chehbouni, Khaoula, et autres
Publié: (2025)
par: Chehbouni, Khaoula, et autres
Publié: (2025)
MIRIAD: Augmenting LLMs with millions of medical query-response pairs
par: Zheng, Qinyue, et autres
Publié: (2025)
par: Zheng, Qinyue, et autres
Publié: (2025)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
par: Bouchekif, Abdessalam, et autres
Publié: (2026)
par: Bouchekif, Abdessalam, et autres
Publié: (2026)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
par: Evelo, Bart, et autres
Publié: (2026)
par: Evelo, Bart, et autres
Publié: (2026)
ManagerBench: Evaluating the Safety-Pragmatism Trade-off in Autonomous LLMs
par: Simhi, Adi, et autres
Publié: (2025)
par: Simhi, Adi, et autres
Publié: (2025)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
par: CH-Wang, Sky, et autres
Publié: (2025)
par: CH-Wang, Sky, et autres
Publié: (2025)
LLMs and the Human Condition
par: Wallis, Peter
Publié: (2024)
par: Wallis, Peter
Publié: (2024)
Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer
par: Paneru, Utsav
Publié: (2026)
par: Paneru, Utsav
Publié: (2026)
Does Localization Inform Unlearning? A Rigorous Examination of Local Parameter Attribution for Knowledge Unlearning in Language Models
par: Lee, Hwiyeong, et autres
Publié: (2025)
par: Lee, Hwiyeong, et autres
Publié: (2025)
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
par: Šindelář, Pavel, et autres
Publié: (2025)
par: Šindelář, Pavel, et autres
Publié: (2025)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
par: Ewais, Ahmed, et autres
Publié: (2026)
par: Ewais, Ahmed, et autres
Publié: (2026)
Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
par: Zhang, Zhaowei, et autres
Publié: (2025)
par: Zhang, Zhaowei, et autres
Publié: (2025)
MALT: Mechanistic Ablation of Lossy Translation in LLMs for a Low-Resource Language: Urdu
par: Bajwa, Taaha Saleem
Publié: (2025)
par: Bajwa, Taaha Saleem
Publié: (2025)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
par: Simhi, Adi, et autres
Publié: (2025)
par: Simhi, Adi, et autres
Publié: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
par: Bajpai, Ashutosh, et autres
Publié: (2025)
par: Bajpai, Ashutosh, et autres
Publié: (2025)
Improving the OOD Performance of Closed-Source LLMs on NLI Through Strategic Data Selection
par: Stacey, Joe, et autres
Publié: (2025)
par: Stacey, Joe, et autres
Publié: (2025)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
par: Akavarapu, V. S. D. S. Mahesh, et autres
Publié: (2025)
par: Akavarapu, V. S. D. S. Mahesh, et autres
Publié: (2025)
Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs
par: Cui, Wanyun, et autres
Publié: (2025)
par: Cui, Wanyun, et autres
Publié: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025)
par: Peters, Sydney, et autres
Publié: (2025)
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
par: Mamidanna, Siddarth, et autres
Publié: (2025)
par: Mamidanna, Siddarth, et autres
Publié: (2025)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
par: Wang, Yuxia, et autres
Publié: (2024)
par: Wang, Yuxia, et autres
Publié: (2024)
Lisbon Computational Linguists at SemEval-2024 Task 2: Using A Mistral 7B Model and Data Augmentation
par: Guimarães, Artur, et autres
Publié: (2024)
par: Guimarães, Artur, et autres
Publié: (2024)
EduGuardBench: A Holistic Benchmark for Evaluating the Pedagogical Fidelity and Adversarial Safety of LLMs as Simulated Teachers
par: Jiang, Yilin, et autres
Publié: (2025)
par: Jiang, Yilin, et autres
Publié: (2025)
Improving LLMs with a knowledge from databases
par: Máša, Petr
Publié: (2025)
par: Máša, Petr
Publié: (2025)
ARAGOG: Advanced RAG Output Grading
par: Eibich, Matouš, et autres
Publié: (2024)
par: Eibich, Matouš, et autres
Publié: (2024)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
par: Gwak, Jiho, et autres
Publié: (2025)
par: Gwak, Jiho, et autres
Publié: (2025)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
par: Bian, Zhipeng, et autres
Publié: (2026)
par: Bian, Zhipeng, et autres
Publié: (2026)
Documents similaires
-
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
par: Qi, Jinhu, et autres
Publié: (2024) -
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
par: Cacioli, Jon-Paul
Publié: (2026) -
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
par: Saji, Alan, et autres
Publié: (2025) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
par: Ashuach, Tomer, et autres
Publié: (2025) -
HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools
par: Garg, Aashna, et autres
Publié: (2026)