Science Hierarchography: Hierarchical Organization of Science Literature
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Muhan, Shah, Jash, Wang, Weiqi, Huang, Kuan-Hao, Khashabi, Daniel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
por: Wang, Weiqi, et al.
Publicado: (2025)
por: Wang, Weiqi, et al.
Publicado: (2025)
ICL CIPHERS: Quantifying "Learning" in In-Context Learning via Substitution Ciphers
por: Fang, Zhouxiang, et al.
Publicado: (2025)
por: Fang, Zhouxiang, et al.
Publicado: (2025)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
GOLD PANNING: Strategic Context Shuffling for Needle-in-Haystack Reasoning
por: Byerly, Adam, et al.
Publicado: (2025)
por: Byerly, Adam, et al.
Publicado: (2025)
Challenging the Evaluator: LLM Sycophancy Under User Rebuttal
por: Kim, Sungwon, et al.
Publicado: (2025)
por: Kim, Sungwon, et al.
Publicado: (2025)
SIMPLEMIX: Frustratingly Simple Mixing of Off- and On-policy Data in Language Model Preference Learning
por: Li, Tianjian, et al.
Publicado: (2025)
por: Li, Tianjian, et al.
Publicado: (2025)
Self-Consistency Falls Short! The Adverse Effects of Positional Bias on Long-Context Problems
por: Byerly, Adam, et al.
Publicado: (2024)
por: Byerly, Adam, et al.
Publicado: (2024)
CHIME: LLM-Assisted Hierarchical Organization of Scientific Studies for Literature Review Support
por: Hsu, Chao-Chun, et al.
Publicado: (2024)
por: Hsu, Chao-Chun, et al.
Publicado: (2024)
Evaluating the Evaluators: Are readability metrics good measures of readability?
por: Cachola, Isabel, et al.
Publicado: (2025)
por: Cachola, Isabel, et al.
Publicado: (2025)
Are Finer Citations Always Better? Rethinking Granularity for Attributed Generation
por: Wang, Hexuan, et al.
Publicado: (2026)
por: Wang, Hexuan, et al.
Publicado: (2026)
The Flaw of Averages: Quantifying Uniformity of Performance on Benchmarks
por: Uzunoglu, Arda, et al.
Publicado: (2025)
por: Uzunoglu, Arda, et al.
Publicado: (2025)
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
por: Uzunoglu, Arda, et al.
Publicado: (2026)
por: Uzunoglu, Arda, et al.
Publicado: (2026)
Feedback Friction: LLMs Struggle to Fully Incorporate External Feedback
por: Jiang, Dongwei, et al.
Publicado: (2025)
por: Jiang, Dongwei, et al.
Publicado: (2025)
Hell or High Water: Evaluating Agentic Recovery from External Failures
por: Wang, Andrew, et al.
Publicado: (2025)
por: Wang, Andrew, et al.
Publicado: (2025)
On Lexical Invariance on Multisets and Graphs
por: Zhang, Muhan
Publicado: (2024)
por: Zhang, Muhan
Publicado: (2024)
Structure-Augmented Reasoning Generation
por: Parekh, Jash Rajesh, et al.
Publicado: (2025)
por: Parekh, Jash Rajesh, et al.
Publicado: (2025)
Language Steering for Multilingual In-Context Learning
por: Kirtane, Neeraja, et al.
Publicado: (2026)
por: Kirtane, Neeraja, et al.
Publicado: (2026)
IA2: Alignment with ICL Activations Improves Supervised Fine-Tuning
por: Mishra, Aayush, et al.
Publicado: (2025)
por: Mishra, Aayush, et al.
Publicado: (2025)
Do pretrained Transformers Learn In-Context by Gradient Descent?
por: Shen, Lingfeng, et al.
Publicado: (2023)
por: Shen, Lingfeng, et al.
Publicado: (2023)
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
por: Ou, Jiefu, et al.
Publicado: (2024)
por: Ou, Jiefu, et al.
Publicado: (2024)
Hierarchical Organization Simulacra in the Investment Sector
por: Chen, Chung-Chi, et al.
Publicado: (2024)
por: Chen, Chung-Chi, et al.
Publicado: (2024)
Crystal: Characterizing Relative Impact of Scholarly Publications
por: Collison, Hannah, et al.
Publicado: (2026)
por: Collison, Hannah, et al.
Publicado: (2026)
Highlighting Case Studies in LLM Literature Review of Interdisciplinary System Science
por: McGinness, Lachlan, et al.
Publicado: (2025)
por: McGinness, Lachlan, et al.
Publicado: (2025)
Can Coding Agents Reproduce Findings in Computational Materials Science?
por: Huang, Ziyang, et al.
Publicado: (2026)
por: Huang, Ziyang, et al.
Publicado: (2026)
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
por: Tan, Weiting, et al.
Publicado: (2024)
por: Tan, Weiting, et al.
Publicado: (2024)
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models
por: Li, Tianjian, et al.
Publicado: (2023)
por: Li, Tianjian, et al.
Publicado: (2023)
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature
por: Yi, Gyeong Hoon, et al.
Publicado: (2024)
por: Yi, Gyeong Hoon, et al.
Publicado: (2024)
Ai2 Scholar QA: Organized Literature Synthesis with Attribution
por: Singh, Amanpreet, et al.
Publicado: (2025)
por: Singh, Amanpreet, et al.
Publicado: (2025)
SHA256 at SemEval-2025 Task 4: Selective Amnesia -- Constrained Unlearning for Large Language Models via Knowledge Isolation
por: Agrawal, Saransh, et al.
Publicado: (2025)
por: Agrawal, Saransh, et al.
Publicado: (2025)
Certified Mitigation of Worst-Case LLM Copyright Infringement
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data
por: Zhang, Jingyu, et al.
Publicado: (2024)
por: Zhang, Jingyu, et al.
Publicado: (2024)
DSBench: How Far Are Data Science Agents from Becoming Data Science Experts?
por: Jing, Liqiang, et al.
Publicado: (2024)
por: Jing, Liqiang, et al.
Publicado: (2024)
Matter-of-Fact: A Benchmark for Verifying the Feasibility of Literature-Supported Claims in Materials Science
por: Jansen, Peter, et al.
Publicado: (2025)
por: Jansen, Peter, et al.
Publicado: (2025)
MARS: Benchmarking the Metaphysical Reasoning Abilities of Language Models with a Multi-task Evaluation Dataset
por: Wang, Weiqi, et al.
Publicado: (2024)
por: Wang, Weiqi, et al.
Publicado: (2024)
Steering Vector Fields for Context-Aware Inference-Time Control in Large Language Models
por: Li, Jiaqian, et al.
Publicado: (2026)
por: Li, Jiaqian, et al.
Publicado: (2026)
Hierarchical Memory Organization for Wikipedia Generation
por: Yu, Eugene J., et al.
Publicado: (2025)
por: Yu, Eugene J., et al.
Publicado: (2025)
Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches
por: Shah, Syed Mehtab Hussain, et al.
Publicado: (2026)
por: Shah, Syed Mehtab Hussain, et al.
Publicado: (2026)
Language Models as Science Tutors
por: Chevalier, Alexis, et al.
Publicado: (2024)
por: Chevalier, Alexis, et al.
Publicado: (2024)
OLMo: Accelerating the Science of Language Models
por: Groeneveld, Dirk, et al.
Publicado: (2024)
por: Groeneveld, Dirk, et al.
Publicado: (2024)
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
por: Lin, Hao, et al.
Publicado: (2025)
por: Lin, Hao, et al.
Publicado: (2025)
Ejemplares similares
-
arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
por: Wang, Weiqi, et al.
Publicado: (2025) -
ICL CIPHERS: Quantifying "Learning" in In-Context Learning via Substitution Ciphers
por: Fang, Zhouxiang, et al.
Publicado: (2025) -
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024) -
GOLD PANNING: Strategic Context Shuffling for Needle-in-Haystack Reasoning
por: Byerly, Adam, et al.
Publicado: (2025) -
Challenging the Evaluator: LLM Sycophancy Under User Rebuttal
por: Kim, Sungwon, et al.
Publicado: (2025)