Lightweight Spatial Modeling for Combinatorial Information Extraction From Documents
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Dong, Yanfei, Deng, Lambert, Zhang, Jiazheng, Yu, Xiaodong, Lin, Ting, Gelli, Francesco, Poria, Soujanya, Lee, Wee Sun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Consistency Guided Knowledge Retrieval and Denoising in LLMs for Zero-shot Document-level Relation Triplet Extraction
par: Sun, Qi, et autres
Publié: (2024)
par: Sun, Qi, et autres
Publié: (2024)
Towards Robust Instruction Tuning on Multimodal Large Language Models
par: Han, Wei, et autres
Publié: (2024)
par: Han, Wei, et autres
Publié: (2024)
DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling
par: Deep, Pala Tej, et autres
Publié: (2024)
par: Deep, Pala Tej, et autres
Publié: (2024)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
par: Bhardwaj, Rishabh, et autres
Publié: (2024)
par: Bhardwaj, Rishabh, et autres
Publié: (2024)
Sowing the Wind, Reaping the Whirlwind: The Impact of Editing Language Models
par: Hazra, Rima, et autres
Publié: (2024)
par: Hazra, Rima, et autres
Publié: (2024)
Not All Votes Count! Programs as Verifiers Improve Self-Consistency of Language Models for Math Reasoning
par: Toh, Vernon Y. H., et autres
Publié: (2024)
par: Toh, Vernon Y. H., et autres
Publié: (2024)
Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
par: Hazra, Rima, et autres
Publié: (2024)
par: Hazra, Rima, et autres
Publié: (2024)
PREMISE: Matching-based Prediction for Accurate Review Recommendation
par: Han, Wei, et autres
Publié: (2025)
par: Han, Wei, et autres
Publié: (2025)
Understanding the Capabilities and Limitations of Large Language Models for Cultural Commonsense
par: Shen, Siqi, et autres
Publié: (2024)
par: Shen, Siqi, et autres
Publié: (2024)
Inference Time Alignment with Reward-Guided Tree Search
par: Hung, Chia-Yu, et autres
Publié: (2024)
par: Hung, Chia-Yu, et autres
Publié: (2024)
Can-Do! A Dataset and Neuro-Symbolic Grounded Framework for Embodied Planning with Large Multimodal Models
par: Chia, Yew Ken, et autres
Publié: (2024)
par: Chia, Yew Ken, et autres
Publié: (2024)
Ruby Teaming: Improving Quality Diversity Search with Memory for Automated Red Teaming
par: Han, Vernon Toh Yan, et autres
Publié: (2024)
par: Han, Vernon Toh Yan, et autres
Publié: (2024)
Stacked from One: Multi-Scale Self-Injection for Context Window Extension
par: Han, Wei, et autres
Publié: (2026)
par: Han, Wei, et autres
Publié: (2026)
Domain-Expanded ASTE: Rethinking Generalization in Aspect Sentiment Triplet Extraction
par: Chia, Yew Ken, et autres
Publié: (2023)
par: Chia, Yew Ken, et autres
Publié: (2023)
PROEMO: Prompt-Driven Text-to-Speech Synthesis Based on Emotion and Intensity Control
par: Zhang, Shaozuo, et autres
Publié: (2025)
par: Zhang, Shaozuo, et autres
Publié: (2025)
Two are better than one: Context window extension with multi-grained self-injection
par: Han, Wei, et autres
Publié: (2024)
par: Han, Wei, et autres
Publié: (2024)
Self-Adaptive Sampling for Efficient Video Question-Answering on Image--Text Models
par: Han, Wei, et autres
Publié: (2023)
par: Han, Wei, et autres
Publié: (2023)
Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning
par: Sun, Qi, et autres
Publié: (2024)
par: Sun, Qi, et autres
Publié: (2024)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
par: Pala, Tej Deep, et autres
Publié: (2025)
par: Pala, Tej Deep, et autres
Publié: (2025)
Harnessing Large Language Models for Scientific Novelty Detection
par: Liu, Yan, et autres
Publié: (2025)
par: Liu, Yan, et autres
Publié: (2025)
Toward Robust Multimodal Learning using Multimodal Foundational Models
par: Zhao, Xianbing, et autres
Publié: (2024)
par: Zhao, Xianbing, et autres
Publié: (2024)
PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
par: Luo, Meng, et autres
Publié: (2024)
par: Luo, Meng, et autres
Publié: (2024)
Large Language Models for Automated Open-domain Scientific Hypotheses Discovery
par: Yang, Zonglin, et autres
Publié: (2023)
par: Yang, Zonglin, et autres
Publié: (2023)
Ferret: Faster and Effective Automated Red Teaming with Reward-Based Scoring Technique
par: Pala, Tej Deep, et autres
Publié: (2024)
par: Pala, Tej Deep, et autres
Publié: (2024)
A Comprehensive Survey of Sentence Representations: From the BERT Epoch to the ChatGPT Era and Beyond
par: Kashyap, Abhinav Ramesh, et autres
Publié: (2023)
par: Kashyap, Abhinav Ramesh, et autres
Publié: (2023)
Leveraging Parameter-Efficient Transfer Learning for Multi-Lingual Text-to-Speech Adaptation
par: Li, Yingting, et autres
Publié: (2024)
par: Li, Yingting, et autres
Publié: (2024)
The Jumping Reasoning Curve? Tracking the Evolution of Reasoning Performance in GPT-[n] and o-[n] Models on Multimodal Puzzles
par: Toh, Vernon Y. H., et autres
Publié: (2025)
par: Toh, Vernon Y. H., et autres
Publié: (2025)
Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths
par: Chia, Yew Ken, et autres
Publié: (2024)
par: Chia, Yew Ken, et autres
Publié: (2024)
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
par: Hong, Pengfei, et autres
Publié: (2024)
par: Hong, Pengfei, et autres
Publié: (2024)
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
par: Li, Yingting, et autres
Publié: (2024)
par: Li, Yingting, et autres
Publié: (2024)
Chain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous Sources
par: Li, Xingxuan, et autres
Publié: (2023)
par: Li, Xingxuan, et autres
Publié: (2023)
M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework
par: Chia, Yew Ken, et autres
Publié: (2024)
par: Chia, Yew Ken, et autres
Publié: (2024)
PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference
par: Jin, Weisheng, et autres
Publié: (2025)
par: Jin, Weisheng, et autres
Publié: (2025)
Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse
par: Song, Maojia, et autres
Publié: (2024)
par: Song, Maojia, et autres
Publié: (2024)
CM-TTS: Enhancing Real Time Text-to-Speech Synthesis Efficiency through Weighted Samplers and Consistency Models
par: Li, Xiang, et autres
Publié: (2024)
par: Li, Xiang, et autres
Publié: (2024)
Improving Text-To-Audio Models with Synthetic Captions
par: Kong, Zhifeng, et autres
Publié: (2024)
par: Kong, Zhifeng, et autres
Publié: (2024)
LMDX: Language Model-based Document Information Extraction and Localization
par: Perot, Vincent, et autres
Publié: (2023)
par: Perot, Vincent, et autres
Publié: (2023)
DiffPO: Diffusion-styled Preference Optimization for Efficient Inference-Time Alignment of Large Language Models
par: Chen, Ruizhe, et autres
Publié: (2025)
par: Chen, Ruizhe, et autres
Publié: (2025)
UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
par: Ji, Yifan, et autres
Publié: (2026)
par: Ji, Yifan, et autres
Publié: (2026)
Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
par: Majumder, Navonil, et autres
Publié: (2024)
par: Majumder, Navonil, et autres
Publié: (2024)
Documents similaires
-
Consistency Guided Knowledge Retrieval and Denoising in LLMs for Zero-shot Document-level Relation Triplet Extraction
par: Sun, Qi, et autres
Publié: (2024) -
Towards Robust Instruction Tuning on Multimodal Large Language Models
par: Han, Wei, et autres
Publié: (2024) -
DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling
par: Deep, Pala Tej, et autres
Publié: (2024) -
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
par: Bhardwaj, Rishabh, et autres
Publié: (2024) -
Sowing the Wind, Reaping the Whirlwind: The Impact of Editing Language Models
par: Hazra, Rima, et autres
Publié: (2024)