Saved in:
| Main Authors: | Jayannavar, Prashant, Ren, Liliang, Hudspeth, Marisa, Sidhu, Risham, Lambert, Charlotte, Cordes, Ariel, Kaplan, Elizabeth, Narayan-Chen, Anjali, Hockenmaier, Julia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.10836 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grid Spatial Understanding: A Dataset for Textual Spatial Reasoning over Grids, Embodied Settings, and Coordinate Structures
by: Sidhu, Risham, et al.
Published: (2026)
by: Sidhu, Risham, et al.
Published: (2026)
MDC-R: The Minecraft Dialogue Corpus with Reference
by: Madge, Chris, et al.
Published: (2025)
by: Madge, Chris, et al.
Published: (2025)
RAG-RL: Advancing Retrieval-Augmented Generation via RL and Curriculum Learning
by: Huang, Jerry, et al.
Published: (2025)
by: Huang, Jerry, et al.
Published: (2025)
Analyzing the Performance of Large Language Models on Code Summarization
by: Haldar, Rajarshi, et al.
Published: (2024)
by: Haldar, Rajarshi, et al.
Published: (2024)
Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks
by: Haldar, Rajarshi, et al.
Published: (2025)
by: Haldar, Rajarshi, et al.
Published: (2025)
Evaluating Step-by-step Reasoning Traces: A Survey
by: Lee, Jinu, et al.
Published: (2025)
by: Lee, Jinu, et al.
Published: (2025)
MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory
by: Park, Junyeong, et al.
Published: (2024)
by: Park, Junyeong, et al.
Published: (2024)
Latin Treebanks in Review: An Evaluation of Morphological Tagging Across Time
by: Hudspeth, Marisa, et al.
Published: (2024)
by: Hudspeth, Marisa, et al.
Published: (2024)
Evaluating Morphological Alignment of Tokenizers in 70 Languages
by: Arnett, Catherine, et al.
Published: (2025)
by: Arnett, Catherine, et al.
Published: (2025)
MCPDial: A Minecraft Persona-driven Dialogue Dataset
by: Alavi, Seyed Hossein, et al.
Published: (2024)
by: Alavi, Seyed Hossein, et al.
Published: (2024)
RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering
by: Hudspeth, Marisa, et al.
Published: (2026)
by: Hudspeth, Marisa, et al.
Published: (2026)
Contextual morphologically-guided tokenization for Latin encoder models
by: Hudspeth, Marisa, et al.
Published: (2025)
by: Hudspeth, Marisa, et al.
Published: (2025)
TOD-ProcBench: Benchmarking Complex Instruction-Following in Task-Oriented Dialogues
by: Ghazarian, Sarik, et al.
Published: (2025)
by: Ghazarian, Sarik, et al.
Published: (2025)
ViT-MUL: A Baseline Study on Recent Machine Unlearning Methods Applied to Vision Transformers
by: Cho, Ikhyun, et al.
Published: (2024)
by: Cho, Ikhyun, et al.
Published: (2024)
CausalMACE: Causality Empowered Multi-Agents in Minecraft Cooperative Tasks
by: Chai, Qi, et al.
Published: (2025)
by: Chai, Qi, et al.
Published: (2025)
Mississippi Library Commission, Final Performance Report for Library Services and Construction Act (LSCA) Title VI, Library Literacy Program.
by: Beard, Patricia, et al.
Published: (1994)
by: Beard, Patricia, et al.
Published: (1994)
MineNPC-Task: Task Suite for Memory-Aware Minecraft Agents
by: Doss, Tamil Sudaravan Mohan, et al.
Published: (2026)
by: Doss, Tamil Sudaravan Mohan, et al.
Published: (2026)
Multimedia Generative Script Learning for Task Planning
by: Wang, Qingyun, et al.
Published: (2022)
by: Wang, Qingyun, et al.
Published: (2022)
How Reliable are Causal Probing Interventions?
by: Canby, Marc, et al.
Published: (2024)
by: Canby, Marc, et al.
Published: (2024)
A Multi-Perspective Architecture for Semantic Code Search
by: Haldar, Rajarshi, et al.
Published: (2020)
by: Haldar, Rajarshi, et al.
Published: (2020)
Generalizing Verifiable Instruction Following
by: Pyatkin, Valentina, et al.
Published: (2025)
by: Pyatkin, Valentina, et al.
Published: (2025)
WildIFEval: Instruction Following in the Wild
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Laser‐Fabricated Zinc Salicylate Nanoparticles as Dual‐Functional Agents for the Suppression of Fusarium verticillioides and Growth Enhancement in Rice Seedlings
by: Anjali Sidhu, et al.
Published: (2026)
by: Anjali Sidhu, et al.
Published: (2026)
BAP1-2025-public
by: Hong, Jing Han
Published: (2026)
by: Hong, Jing Han
Published: (2026)
Doubly Separable Spacetimes and Symmetry Constraints on their Self-Gravitating Matter Content
by: Kocherlakota, Prashant, et al.
Published: (2025)
by: Kocherlakota, Prashant, et al.
Published: (2025)
Self-Gravitating Matter in Stationary and Axisymmetric Black Hole Spacetimes
by: Kocherlakota, Prashant, et al.
Published: (2024)
by: Kocherlakota, Prashant, et al.
Published: (2024)
Benchmarking GPUs on SVBRDF Extractor Model
by: Kandel, Narayan, et al.
Published: (2023)
by: Kandel, Narayan, et al.
Published: (2023)
Development of Multistage Machine Learning Classifier using Decision Trees and Boosting Algorithms over Darknet Network Traffic
by: Nair, Anjali Sureshkumar, et al.
Published: (2024)
by: Nair, Anjali Sureshkumar, et al.
Published: (2024)
Mississippi Magic: Summer Library Program, 1999.
by: Hudspeth, Jean, et al.
Published: (1999)
by: Hudspeth, Jean, et al.
Published: (1999)
Fantasy Quest: Summer Library Program 1997.
by: Thieling, Kaileen R., et al.
Published: (1997)
by: Thieling, Kaileen R., et al.
Published: (1997)
Attack and Reset for Unlearning: Exploiting Adversarial Noise toward Machine Unlearning through Parameter Re-initialization
by: Jung, Yoonhwa, et al.
Published: (2024)
by: Jung, Yoonhwa, et al.
Published: (2024)
Evaluating and Designing Sparse Autoencoders by Approximating Quasi-Orthogonality
by: Lee, Sewoong, et al.
Published: (2025)
by: Lee, Sewoong, et al.
Published: (2025)
ReasoningFlow: Semantic Structure of Complex Reasoning Traces
by: Lee, Jinu, et al.
Published: (2025)
by: Lee, Jinu, et al.
Published: (2025)
Hypopigmented Junctional BAP1‐Inactivated Melanocytoma: A Paradigm Shift
by: Ali Gunesch, et al.
Published: (2025)
by: Ali Gunesch, et al.
Published: (2025)
BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks
by: Du, Weihong, et al.
Published: (2025)
by: Du, Weihong, et al.
Published: (2025)
A LLM Benchmark based on the Minecraft Builder Dialog Agent Task
by: Madge, Chris, et al.
Published: (2024)
by: Madge, Chris, et al.
Published: (2024)
Contextual Relevance and Adaptive Sampling for LLM-Based Document Reranking
by: Huang, Jerry, et al.
Published: (2025)
by: Huang, Jerry, et al.
Published: (2025)
Evaluating Legal Reasoning Traces with Legal Issue Tree Rubrics
by: Lee, Jinu, et al.
Published: (2025)
by: Lee, Jinu, et al.
Published: (2025)
Playing in Minecraft: an exploratory study
by: Arlete dos Santos Petry
Published: (2018)
by: Arlete dos Santos Petry
Published: (2018)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
Similar Items
-
Grid Spatial Understanding: A Dataset for Textual Spatial Reasoning over Grids, Embodied Settings, and Coordinate Structures
by: Sidhu, Risham, et al.
Published: (2026) -
MDC-R: The Minecraft Dialogue Corpus with Reference
by: Madge, Chris, et al.
Published: (2025) -
RAG-RL: Advancing Retrieval-Augmented Generation via RL and Curriculum Learning
by: Huang, Jerry, et al.
Published: (2025) -
Analyzing the Performance of Large Language Models on Code Summarization
by: Haldar, Rajarshi, et al.
Published: (2024) -
Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks
by: Haldar, Rajarshi, et al.
Published: (2025)