Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA
Fuente:
arXiv
Salvato in:
| Autori principali: | Ciosici, Manuel R., Cecil, Joe, Hedges, Alex, Lee, Dong-Ho, Freedman, Marjorie, Weischedel, Ralph |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Machine-Assisted Script Curation
di: Ciosici, Manuel R., et al.
Pubblicazione: (2021)
di: Ciosici, Manuel R., et al.
Pubblicazione: (2021)
Understanding Multimodal Procedural Knowledge by Sequencing Multimodal Instructional Manuals
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
BookAsSumQA: An Evaluation Framework for Aspect-Based Book Summarization via Question Answering
di: Miyazato, Ryuhei, et al.
Pubblicazione: (2025)
di: Miyazato, Ryuhei, et al.
Pubblicazione: (2025)
eBooks--Ready for School Libraries?
di: Pappas, Marjorie
Pubblicazione: (2009)
di: Pappas, Marjorie
Pubblicazione: (2009)
Multilingual Open QA on the MIA Shared Task
di: Yarrabelly, Navya, et al.
Pubblicazione: (2025)
di: Yarrabelly, Navya, et al.
Pubblicazione: (2025)
Memorization: A Close Look at Books
di: Ma, Iris, et al.
Pubblicazione: (2025)
di: Ma, Iris, et al.
Pubblicazione: (2025)
Back to School: Translation Using Grammar Books
di: Hus, Jonathan, et al.
Pubblicazione: (2024)
di: Hus, Jonathan, et al.
Pubblicazione: (2024)
ExpertGenQA: Open-ended QA generation in Specialized Domains
di: Shahgir, Haz Sameen, et al.
Pubblicazione: (2025)
di: Shahgir, Haz Sameen, et al.
Pubblicazione: (2025)
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Think Before you Write: QA-Guided Reasoning for Character Descriptions in Books
di: Papoudakis, Argyrios, et al.
Pubblicazione: (2026)
di: Papoudakis, Argyrios, et al.
Pubblicazione: (2026)
(Perhaps) Beyond Human Translation: Harnessing Multi-Agent Collaboration for Translating Ultra-Long Literary Texts
di: Wu, Minghao, et al.
Pubblicazione: (2024)
di: Wu, Minghao, et al.
Pubblicazione: (2024)
Books, Readers, and Individuals
di: Sullivan, Marjorie
Pubblicazione: (1971)
di: Sullivan, Marjorie
Pubblicazione: (1971)
RetrievalQA: Assessing Adaptive Retrieval-Augmented Generation for Short-form Open-Domain Question Answering
di: Zhang, Zihan, et al.
Pubblicazione: (2024)
di: Zhang, Zihan, et al.
Pubblicazione: (2024)
Tree Schools and Modular Book Collections.
di: Hallein, Joe
Pubblicazione: (1995)
di: Hallein, Joe
Pubblicazione: (1995)
JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge
di: Cao, Zhihan, et al.
Pubblicazione: (2025)
di: Cao, Zhihan, et al.
Pubblicazione: (2025)
Benchmarking Large Language Models for Quebec Insurance: From Closed-Book to Retrieval-Augmented Generation
di: Beauchemin, David, et al.
Pubblicazione: (2026)
di: Beauchemin, David, et al.
Pubblicazione: (2026)
Navigating Uncertainty: Optimizing API Dependency for Hallucination Reduction in Closed-Book Question Answering
di: Erbacher, Pierre, et al.
Pubblicazione: (2024)
di: Erbacher, Pierre, et al.
Pubblicazione: (2024)
MedConceptsQA: Open Source Medical Concepts QA Benchmark
di: Shoham, Ofir Ben, et al.
Pubblicazione: (2024)
di: Shoham, Ofir Ben, et al.
Pubblicazione: (2024)
Who Should Decide on a Book's Merit?
di: Darkatsh, Manuel
Pubblicazione: (1974)
di: Darkatsh, Manuel
Pubblicazione: (1974)
Back to Basics: Reevaluating Picture Books
di: Lewis, Marjorie
Pubblicazione: (1976)
di: Lewis, Marjorie
Pubblicazione: (1976)
Contextual Breach: Assessing the Robustness of Transformer-based QA Models
di: Saadat, Asir, et al.
Pubblicazione: (2024)
di: Saadat, Asir, et al.
Pubblicazione: (2024)
Teaching to the Heart: An Affective Approach to Literacy Instruction. Second Edition.
di: Cecil, Nancy Lee
Pubblicazione: (1993)
di: Cecil, Nancy Lee
Pubblicazione: (1993)
Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task
di: Jung, JungHo, et al.
Pubblicazione: (2025)
di: Jung, JungHo, et al.
Pubblicazione: (2025)
1993 Children's Book Awards.
di: Horowitz, Marjorie, Comp.
Pubblicazione: (1993)
di: Horowitz, Marjorie, Comp.
Pubblicazione: (1993)
Impact of After-School Nutrition Workshops in a Public Library Setting
di: Freedman, Marjorie R., et al.
Pubblicazione: (2010)
di: Freedman, Marjorie R., et al.
Pubblicazione: (2010)
Selecting Spanish Books in the Elementary School
di: Corona, Elia
Pubblicazione: (2007)
di: Corona, Elia
Pubblicazione: (2007)
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2025)
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2025)
Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2026)
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2026)
Correlating the Classes of Books Taken Out Of and Books Used Within an Open-Stack Library. Research Report.
di: Domas, Ralph E.
Pubblicazione: (1978)
di: Domas, Ralph E.
Pubblicazione: (1978)
DisasterQA: A Benchmark for Assessing the performance of LLMs in Disaster Response
di: Rawat, Rajat
Pubblicazione: (2024)
di: Rawat, Rajat
Pubblicazione: (2024)
Team Anotheroption at SemEval-2025 Task 8: Bridging the Gap Between Open-Source and Proprietary LLMs in Table QA
di: Evkarpidi, Nikolas, et al.
Pubblicazione: (2025)
di: Evkarpidi, Nikolas, et al.
Pubblicazione: (2025)
BookGPT: A General Framework for Book Recommendation Empowered by Large Language Model
di: Zhiyuli, Aakas, et al.
Pubblicazione: (2023)
di: Zhiyuli, Aakas, et al.
Pubblicazione: (2023)
Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks
di: Tong, Chaodong, et al.
Pubblicazione: (2025)
di: Tong, Chaodong, et al.
Pubblicazione: (2025)
CondAmbigQA: A Benchmark and Dataset for Conditional Ambiguous Question Answering
di: Li, Zongxi, et al.
Pubblicazione: (2025)
di: Li, Zongxi, et al.
Pubblicazione: (2025)
Assessing The Potential Of Mid-Sized Language Models For Clinical QA
di: Bolton, Elliot, et al.
Pubblicazione: (2024)
di: Bolton, Elliot, et al.
Pubblicazione: (2024)
Accurate and Nuanced Open-QA Evaluation Through Textual Entailment
di: Yao, Peiran, et al.
Pubblicazione: (2024)
di: Yao, Peiran, et al.
Pubblicazione: (2024)
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation
di: Kim, Kiseung, et al.
Pubblicazione: (2024)
di: Kim, Kiseung, et al.
Pubblicazione: (2024)
Code-Based English Models Surprising Performance on Chinese QA Pair Extraction Task
di: Zheng, Linghan, et al.
Pubblicazione: (2024)
di: Zheng, Linghan, et al.
Pubblicazione: (2024)
Assessing the Performance of Chinese Open Source Large Language Models in Information Extraction Tasks
di: Cai, Yida, et al.
Pubblicazione: (2024)
di: Cai, Yida, et al.
Pubblicazione: (2024)
Access?: Books, Children, and Literature-Based Curriculum in Schools.
di: Guice, Sherry, et al.
Pubblicazione: (1996)
di: Guice, Sherry, et al.
Pubblicazione: (1996)
Documenti analoghi
-
Machine-Assisted Script Curation
di: Ciosici, Manuel R., et al.
Pubblicazione: (2021) -
Understanding Multimodal Procedural Knowledge by Sequencing Multimodal Instructional Manuals
di: Wu, Te-Lin, et al.
Pubblicazione: (2021) -
BookAsSumQA: An Evaluation Framework for Aspect-Based Book Summarization via Question Answering
di: Miyazato, Ryuhei, et al.
Pubblicazione: (2025) -
eBooks--Ready for School Libraries?
di: Pappas, Marjorie
Pubblicazione: (2009) -
Multilingual Open QA on the MIA Shared Task
di: Yarrabelly, Navya, et al.
Pubblicazione: (2025)