StructLM: Towards Building Generalist Models for Structured Knowledge Grounding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhuang, Alex, Zhang, Ge, Zheng, Tianyu, Du, Xinrun, Wang, Junjie, Ren, Weiming, Huang, Stephen W., Fu, Jie, Yue, Xiang, Chen, Wenhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision
von: Wei, Cong, et al.
Veröffentlicht: (2024)
von: Wei, Cong, et al.
Veröffentlicht: (2024)
ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
MAmmoTH2: Scaling Instructions from the Web
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
KARPA: A Training-free Method of Adapting Knowledge Graph as References for Large Language Model's Reasoning Path Aggregation
von: Fang, Siyuan, et al.
Veröffentlicht: (2024)
von: Fang, Siyuan, et al.
Veröffentlicht: (2024)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
von: Shen, Yanqing, et al.
Veröffentlicht: (2025)
von: Shen, Yanqing, et al.
Veröffentlicht: (2025)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
Towards Generalist Prompting for Large Language Models by Mental Models
von: Guan, Haoxiang, et al.
Veröffentlicht: (2024)
von: Guan, Haoxiang, et al.
Veröffentlicht: (2024)
VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
Grasper: A Generalist Pursuer for Pursuit-Evasion Problems
von: Li, Pengdeng, et al.
Veröffentlicht: (2024)
von: Li, Pengdeng, et al.
Veröffentlicht: (2024)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
von: Ren, Weiming, et al.
Veröffentlicht: (2025)
von: Ren, Weiming, et al.
Veröffentlicht: (2025)
TableLlama: Towards Open Large Generalist Models for Tables
von: Zhang, Tianshu, et al.
Veröffentlicht: (2023)
von: Zhang, Tianshu, et al.
Veröffentlicht: (2023)
TIGERScore: Towards Building Explainable Metric for All Text Generation Tasks
von: Jiang, Dongfu, et al.
Veröffentlicht: (2023)
von: Jiang, Dongfu, et al.
Veröffentlicht: (2023)
StructEval: Benchmarking LLMs' Capabilities to Generate Structural Outputs
von: Yang, Jialin, et al.
Veröffentlicht: (2025)
von: Yang, Jialin, et al.
Veröffentlicht: (2025)
Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
Vision Generalist Model: A Survey
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
GIEBench: Towards Holistic Evaluation of Group Identity-based Empathy for Large Language Models
von: Wang, Leyan, et al.
Veröffentlicht: (2024)
von: Wang, Leyan, et al.
Veröffentlicht: (2024)
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation
von: Ma, Wentao, et al.
Veröffentlicht: (2025)
von: Ma, Wentao, et al.
Veröffentlicht: (2025)
MIDAS: Modeling Ground-Truth Distributions with Dark Knowledge for Domain Generalized Stereo Matching
von: Xu, Peng, et al.
Veröffentlicht: (2025)
von: Xu, Peng, et al.
Veröffentlicht: (2025)
UltraMedical: Building Specialized Generalists in Biomedicine
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
von: Ku, Max, et al.
Veröffentlicht: (2023)
von: Ku, Max, et al.
Veröffentlicht: (2023)
StructMem: Structured Memory for Long-Horizon Behavior in LLMs
von: Xu, Buqiang, et al.
Veröffentlicht: (2026)
von: Xu, Buqiang, et al.
Veröffentlicht: (2026)
Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks
von: Ku, Max, et al.
Veröffentlicht: (2024)
von: Ku, Max, et al.
Veröffentlicht: (2024)
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
StructLayoutFormer:Conditional Structured Layout Generation via Structure Serialization and Disentanglement
von: Hu, Xin, et al.
Veröffentlicht: (2025)
von: Hu, Xin, et al.
Veröffentlicht: (2025)
Long-context LLMs Struggle with Long In-context Learning
von: Li, Tianle, et al.
Veröffentlicht: (2024)
von: Li, Tianle, et al.
Veröffentlicht: (2024)
Towards Building Specialized Generalist AI with System 1 and System 2 Fusion
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
StructRe: Rewriting for Structured Shape Modeling
von: Wang, Jiepeng, et al.
Veröffentlicht: (2023)
von: Wang, Jiepeng, et al.
Veröffentlicht: (2023)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
von: Du, Xinrun, et al.
Veröffentlicht: (2024)
von: Du, Xinrun, et al.
Veröffentlicht: (2024)
StructFlowBench: A Structured Flow Benchmark for Multi-turn Instruction Following
von: Li, Jinnan, et al.
Veröffentlicht: (2025)
von: Li, Jinnan, et al.
Veröffentlicht: (2025)
CollabEdit: Towards Non-destructive Collaborative Knowledge Editing
von: Zheng, Jiamu, et al.
Veröffentlicht: (2024)
von: Zheng, Jiamu, et al.
Veröffentlicht: (2024)
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
von: Ma, Kaijing, et al.
Veröffentlicht: (2024)
von: Ma, Kaijing, et al.
Veröffentlicht: (2024)
StructSR: Refuse Spurious Details in Real-World Image Super-Resolution
von: Li, Yachao, et al.
Veröffentlicht: (2025)
von: Li, Yachao, et al.
Veröffentlicht: (2025)
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
GPT-4V(ision) is a Generalist Web Agent, if Grounded
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
OctoNav: Towards Generalist Embodied Navigation
von: Gao, Chen, et al.
Veröffentlicht: (2025)
von: Gao, Chen, et al.
Veröffentlicht: (2025)
QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design
von: Schneider, Benjamin, et al.
Veröffentlicht: (2025)
von: Schneider, Benjamin, et al.
Veröffentlicht: (2025)
Beyond Clicking:A Step Towards Generalist GUI Grounding via Text Dragging
von: Liao, Zeyi, et al.
Veröffentlicht: (2025)
von: Liao, Zeyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision
von: Wei, Cong, et al.
Veröffentlicht: (2024) -
ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation
von: Ren, Weiming, et al.
Veröffentlicht: (2024) -
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024) -
MAmmoTH2: Scaling Instructions from the Web
von: Yue, Xiang, et al.
Veröffentlicht: (2024) -
KARPA: A Training-free Method of Adapting Knowledge Graph as References for Large Language Model's Reasoning Path Aggregation
von: Fang, Siyuan, et al.
Veröffentlicht: (2024)