Salvato in:
| Autori principali: | Anantheswaran, Ujjwala, Gupta, Himanshu, Scaria, Kevin, Verma, Shreyas, Baral, Chitta, Mishra, Swaroop |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.15444 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark
di: Gupta, Himanshu, et al.
Pubblicazione: (2024)
di: Gupta, Himanshu, et al.
Pubblicazione: (2024)
TarGEN: Targeted Data Generation with Large Language Models
di: Gupta, Himanshu, et al.
Pubblicazione: (2023)
di: Gupta, Himanshu, et al.
Pubblicazione: (2023)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
Embarrassingly Simple Unsupervised Aspect Based Sentiment Tuple Extraction
di: Scaria, Kevin, et al.
Pubblicazione: (2024)
di: Scaria, Kevin, et al.
Pubblicazione: (2024)
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
di: Kumbhar, Shrinidhi, et al.
Pubblicazione: (2025)
di: Kumbhar, Shrinidhi, et al.
Pubblicazione: (2025)
Insights into Alignment: Evaluating DPO and its Variants Across Multiple Tasks
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
di: Ouyang, Jialin
Pubblicazione: (2025)
di: Ouyang, Jialin
Pubblicazione: (2025)
Map&Make: Schema Guided Text to Table Generation
di: Ahuja, Naman, et al.
Pubblicazione: (2025)
di: Ahuja, Naman, et al.
Pubblicazione: (2025)
Triple Preference Optimization: Achieving Better Alignment using a Single Step Optimization
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
GETReason: Enhancing Image Context Extraction through Hierarchical Multi-Agent Reasoning
di: Siingh, Shikhhar, et al.
Pubblicazione: (2025)
di: Siingh, Shikhhar, et al.
Pubblicazione: (2025)
The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness
di: Varshney, Neeraj, et al.
Pubblicazione: (2023)
di: Varshney, Neeraj, et al.
Pubblicazione: (2023)
Rethinking Information Synthesis in Multimodal Question Answering A Multi-Agent Perspective
di: Rajput, Krishna Singh, et al.
Pubblicazione: (2025)
di: Rajput, Krishna Singh, et al.
Pubblicazione: (2025)
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Template-Driven LLM-Paraphrased Framework for Tabular Math Word Problem Generation
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2024)
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2024)
Iterative LLM-Based Generation and Refinement of Distracting Conditions in Math Word Problems
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
What Makes Math Word Problems Challenging for LLMs?
di: Srivatsa, KV Aditya, et al.
Pubblicazione: (2024)
di: Srivatsa, KV Aditya, et al.
Pubblicazione: (2024)
Expression Syntax Information Bottleneck for Math Word Problems
di: Xiong, Jing, et al.
Pubblicazione: (2023)
di: Xiong, Jing, et al.
Pubblicazione: (2023)
Self-consistent Reasoning For Solving Math Word Problems
di: Xiong, Jing, et al.
Pubblicazione: (2022)
di: Xiong, Jing, et al.
Pubblicazione: (2022)
ViTaB-A: Evaluating Multimodal Large Language Models on Visual Table Attribution
di: Alqurnawi, Yahia, et al.
Pubblicazione: (2026)
di: Alqurnawi, Yahia, et al.
Pubblicazione: (2026)
Gradient-Controlled Decoding: A Safety Guardrail for LLMs with Dual-Anchor Steering
di: Chiniya, Purva, et al.
Pubblicazione: (2026)
di: Chiniya, Purva, et al.
Pubblicazione: (2026)
ToW: Thoughts of Words Improve Reasoning in Large Language Models
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
Can LLMs Solve longer Math Word Problems Better?
di: Xu, Xin, et al.
Pubblicazione: (2024)
di: Xu, Xin, et al.
Pubblicazione: (2024)
We Need Knowledge Distillation for Solving Math Word Problems
di: Shen, Zhenquan, et al.
Pubblicazione: (2025)
di: Shen, Zhenquan, et al.
Pubblicazione: (2025)
Structured Reasoning with Tree-of-Thoughts for Bengali Math Word Problems
di: Mahmood, Aurprita, et al.
Pubblicazione: (2025)
di: Mahmood, Aurprita, et al.
Pubblicazione: (2025)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
Investigating and Addressing Hallucinations of LLMs in Tasks Involving Negation
di: Varshney, Neeraj, et al.
Pubblicazione: (2024)
di: Varshney, Neeraj, et al.
Pubblicazione: (2024)
FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments
di: Saeidi, Amir, et al.
Pubblicazione: (2026)
di: Saeidi, Amir, et al.
Pubblicazione: (2026)
$λ$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
Can Small Language Models Learn, Unlearn, and Retain Noise Patterns?
di: Scaria, Nicy, et al.
Pubblicazione: (2024)
di: Scaria, Nicy, et al.
Pubblicazione: (2024)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
Elementary Math Word Problem Generation using Large Language Models
di: Ariyarathne, Nimesh, et al.
Pubblicazione: (2025)
di: Ariyarathne, Nimesh, et al.
Pubblicazione: (2025)
EDUMATH: Generating Standards-aligned Educational Math Word Problems
di: Christ, Bryan R., et al.
Pubblicazione: (2025)
di: Christ, Bryan R., et al.
Pubblicazione: (2025)
ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models
di: Patel, Maitreya, et al.
Pubblicazione: (2023)
di: Patel, Maitreya, et al.
Pubblicazione: (2023)
LLM The Genius Paradox: A Linguistic and Math Expert's Struggle with Simple Word-based Counting Problems
di: Xu, Nan, et al.
Pubblicazione: (2024)
di: Xu, Nan, et al.
Pubblicazione: (2024)
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
Data Augmentation with In-Context Learning and Comparative Evaluation in Math Word Problem Solving
di: Yigit, Gulsum, et al.
Pubblicazione: (2024)
di: Yigit, Gulsum, et al.
Pubblicazione: (2024)
Solving Math Word Problems via Cooperative Reasoning induced Language Models
di: Zhu, Xinyu, et al.
Pubblicazione: (2022)
di: Zhu, Xinyu, et al.
Pubblicazione: (2022)
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on $τ$-bench
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
Investigating the Shortcomings of LLMs in Step-by-Step Legal Reasoning
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
Ask-Before-Detection: Identifying and Mitigating Conformity Bias in LLM-Powered Error Detector for Math Word Problem Solutions
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark
di: Gupta, Himanshu, et al.
Pubblicazione: (2024) -
TarGEN: Targeted Data Generation with Large Language Models
di: Gupta, Himanshu, et al.
Pubblicazione: (2023) -
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
di: Parmar, Mihir, et al.
Pubblicazione: (2022) -
Embarrassingly Simple Unsupervised Aspect Based Sentiment Tuple Extraction
di: Scaria, Kevin, et al.
Pubblicazione: (2024) -
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
di: Kumbhar, Shrinidhi, et al.
Pubblicazione: (2025)