Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jin, Rihui, Xin, Zheyu, Xie, Xing, Li, Zuoyi, Qi, Guilin, Chen, Yongrui, Dai, Xinbang, Wu, Tongtong, Haffari, Gholamreza |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Question Answering Over Spatio-Temporal Knowledge Graph
par: Dai, Xinbang, et autres
Publié: (2024)
par: Dai, Xinbang, et autres
Publié: (2024)
After Retrieval, Before Generation: Enhancing the Trustworthiness of Large Language Models in Retrieval-Augmented Generation
par: Dai, Xinbang, et autres
Publié: (2025)
par: Dai, Xinbang, et autres
Publié: (2025)
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge
par: Chen, Yongrui, et autres
Publié: (2025)
par: Chen, Yongrui, et autres
Publié: (2025)
ELAIPBench: A Benchmark for Expert-Level Artificial Intelligence Paper Understanding
par: Dai, Xinbang, et autres
Publié: (2025)
par: Dai, Xinbang, et autres
Publié: (2025)
Pandora: Leveraging Code-driven Knowledge Transfer for Unified Structured Knowledge Reasoning
par: Chen, Yongrui, et autres
Publié: (2025)
par: Chen, Yongrui, et autres
Publié: (2025)
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
par: Dai, Xinbang, et autres
Publié: (2024)
par: Dai, Xinbang, et autres
Publié: (2024)
Continual Speech Learning with Fused Speech Features
par: Wang, Guitao, et autres
Publié: (2025)
par: Wang, Guitao, et autres
Publié: (2025)
Magic Mushroom: A Customizable Benchmark for Fine-grained Analysis of Retrieval Noise Erosion in RAG Systems
par: Zhang, Yuxin, et autres
Publié: (2025)
par: Zhang, Yuxin, et autres
Publié: (2025)
Towards Event Extraction from Speech with Contextual Clues
par: Kang, Jingqi, et autres
Publié: (2024)
par: Kang, Jingqi, et autres
Publié: (2024)
HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table Understanding
par: Jin, Rihui, et autres
Publié: (2024)
par: Jin, Rihui, et autres
Publié: (2024)
Environment-Aware Code Generation: How far are We?
par: Wu, Tongtong, et autres
Publié: (2026)
par: Wu, Tongtong, et autres
Publié: (2026)
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data
par: Min, Dehai, et autres
Publié: (2024)
par: Min, Dehai, et autres
Publié: (2024)
CARD: Towards Conditional Design of Multi-agent Topological Structures
par: Wu, Tongtong, et autres
Publié: (2026)
par: Wu, Tongtong, et autres
Publié: (2026)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
par: Luo, Linhao, et autres
Publié: (2023)
par: Luo, Linhao, et autres
Publié: (2023)
Resurfacing Paralinguistic Awareness in Large Audio Language Models
par: Yang, Hao, et autres
Publié: (2026)
par: Yang, Hao, et autres
Publié: (2026)
Double Mixture: Towards Continual Event Detection from Speech
par: Kang, Jingqi, et autres
Publié: (2024)
par: Kang, Jingqi, et autres
Publié: (2024)
K-DeCore: Facilitating Knowledge Transfer in Continual Structured Knowledge Reasoning via Knowledge Decoupling
par: Chen, Yongrui, et autres
Publié: (2025)
par: Chen, Yongrui, et autres
Publié: (2025)
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
par: Hua, Yuncheng, et autres
Publié: (2024)
par: Hua, Yuncheng, et autres
Publié: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
par: Jin, Rihui, et autres
Publié: (2026)
par: Jin, Rihui, et autres
Publié: (2026)
Evidence-based Distributional Alignment for Large Language Models
par: Pham, Viet-Thanh, et autres
Publié: (2026)
par: Pham, Viet-Thanh, et autres
Publié: (2026)
Continual Learning for Large Language Models: A Survey
par: Wu, Tongtong, et autres
Publié: (2024)
par: Wu, Tongtong, et autres
Publié: (2024)
IntroLM: Introspective Language Models via Prefilling-Time Self-Evaluation
par: Kasnavieh, Hossein Hosseini, et autres
Publié: (2026)
par: Kasnavieh, Hossein Hosseini, et autres
Publié: (2026)
AIPO: Learning to Reason from Active Interaction
par: Liu, Junnan, et autres
Publié: (2026)
par: Liu, Junnan, et autres
Publié: (2026)
StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models
par: Chen, Yongrui, et autres
Publié: (2026)
par: Chen, Yongrui, et autres
Publié: (2026)
Graph-constrained Reasoning: Faithful Reasoning on Knowledge Graphs with Large Language Models
par: Luo, Linhao, et autres
Publié: (2024)
par: Luo, Linhao, et autres
Publié: (2024)
Towards Inference-time Scaling for Continuous Space Reasoning
par: Wang, Minghan, et autres
Publié: (2025)
par: Wang, Minghan, et autres
Publié: (2025)
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
par: Liu, Junnan, et autres
Publié: (2025)
par: Liu, Junnan, et autres
Publié: (2025)
Jigsaw Puzzles: Splitting Harmful Questions to Jailbreak Large Language Models
par: Yang, Hao, et autres
Publié: (2024)
par: Yang, Hao, et autres
Publié: (2024)
Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models
par: Yang, Hao, et autres
Publié: (2025)
par: Yang, Hao, et autres
Publié: (2025)
ChatRule: Mining Logical Rules with Large Language Models for Knowledge Graph Reasoning
par: Luo, Linhao, et autres
Publié: (2023)
par: Luo, Linhao, et autres
Publié: (2023)
Modelling Political Coalition Negotiations Using LLM-based Agents
par: Moghimifar, Farhad, et autres
Publié: (2024)
par: Moghimifar, Farhad, et autres
Publié: (2024)
Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models
par: Wu, Minghao, et autres
Publié: (2024)
par: Wu, Minghao, et autres
Publié: (2024)
Multi-Layer Scheduling for MoE-Based LLM Reasoning
par: Sun, Yifan, et autres
Publié: (2026)
par: Sun, Yifan, et autres
Publié: (2026)
DoG-Instruct: Towards Premium Instruction-Tuning Data via Text-Grounded Instruction Wrapping
par: Chen, Yongrui, et autres
Publié: (2023)
par: Chen, Yongrui, et autres
Publié: (2023)
GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Sampler
par: Wang, Minghan, et autres
Publié: (2026)
par: Wang, Minghan, et autres
Publié: (2026)
An Empirical Analysis on Spatial Reasoning Capabilities of Large Multimodal Models
par: Shiri, Fatemeh, et autres
Publié: (2024)
par: Shiri, Fatemeh, et autres
Publié: (2024)
Towards Probing Speech-Specific Risks in Large Multimodal Models: A Taxonomy, Benchmark, and Insights
par: Yang, Hao, et autres
Publié: (2024)
par: Yang, Hao, et autres
Publié: (2024)
Zero-Shot Privacy-Aware Text Rewriting via Iterative Tree Search
par: Huang, Shuo, et autres
Publié: (2025)
par: Huang, Shuo, et autres
Publié: (2025)
Audio Is the Achilles' Heel: Red Teaming Audio Large Multimodal Models
par: Yang, Hao, et autres
Publié: (2024)
par: Yang, Hao, et autres
Publié: (2024)
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
par: Feng, Tao, et autres
Publié: (2025)
par: Feng, Tao, et autres
Publié: (2025)
Documents similaires
-
Question Answering Over Spatio-Temporal Knowledge Graph
par: Dai, Xinbang, et autres
Publié: (2024) -
After Retrieval, Before Generation: Enhancing the Trustworthiness of Large Language Models in Retrieval-Augmented Generation
par: Dai, Xinbang, et autres
Publié: (2025) -
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge
par: Chen, Yongrui, et autres
Publié: (2025) -
ELAIPBench: A Benchmark for Expert-Level Artificial Intelligence Paper Understanding
par: Dai, Xinbang, et autres
Publié: (2025) -
Pandora: Leveraging Code-driven Knowledge Transfer for Unified Structured Knowledge Reasoning
par: Chen, Yongrui, et autres
Publié: (2025)