Funny-Valen-Tine: Planning Solution Distribution Enhances Machine Abstract Reasoning Ability
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Ruizhuo, Yuan, Beiming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Triple-CFN: Separating Concepts and Features Enhances Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2024)
by: Song, Ruizhuo, et al.
Published: (2024)
DIO: Refining Mutual Information and Causal Chain to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2025)
by: Song, Ruizhuo, et al.
Published: (2025)
Johnny: Structuring Representation Space to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2025)
by: Song, Ruizhuo, et al.
Published: (2025)
D4C: Improving Negative Example Quality to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2024)
by: Song, Ruizhuo, et al.
Published: (2024)
Solving the Clustering Reasoning Problems by Modeling a Deep-Learning-Based Probabilistic Model
by: Song, Ruizhuo, et al.
Published: (2024)
by: Song, Ruizhuo, et al.
Published: (2024)
EiHi Net: Out-of-Distribution Generalization Paradigm
by: Wei, Qinglai, et al.
Published: (2022)
by: Wei, Qinglai, et al.
Published: (2022)
FunnyNet-W: Multimodal Learning of Funny Moments in Videos in the Wild
by: Liu, Zhi-Song, et al.
Published: (2024)
by: Liu, Zhi-Song, et al.
Published: (2024)
FunnyNodules: A Customizable Medical Dataset Tailored for Evaluating Explainable AI
by: Gallée, Luisa, et al.
Published: (2025)
by: Gallée, Luisa, et al.
Published: (2025)
LOLGORITHM: Funny Comment Generation Agent For Short Videos
by: Ouyang, Xuan, et al.
Published: (2026)
by: Ouyang, Xuan, et al.
Published: (2026)
CLGRPO: Reasoning Ability Enhancement for Small VLMs
by: Wang, Fanyi, et al.
Published: (2025)
by: Wang, Fanyi, et al.
Published: (2025)
Enhancing Advanced Visual Reasoning Ability of Large Language Models
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
Skeleton2vec: A Self-supervised Learning Framework with Contextualized Target Representations for Skeleton Sequence
by: Xu, Ruizhuo, et al.
Published: (2024)
by: Xu, Ruizhuo, et al.
Published: (2024)
Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
by: Wang, Weiyun, et al.
Published: (2024)
by: Wang, Weiyun, et al.
Published: (2024)
Automatic Funny Scene Extraction from Long-form Cinematic Videos
by: Paul, Sibendu, et al.
Published: (2026)
by: Paul, Sibendu, et al.
Published: (2026)
VisualQuest: A Benchmark for Abstract Visual Reasoning in MLLMs
by: Xiao, Kelaiti, et al.
Published: (2025)
by: Xiao, Kelaiti, et al.
Published: (2025)
Revisiting FunnyBirds evaluation framework for prototypical parts networks
by: Opłatek, Szymon, et al.
Published: (2024)
by: Opłatek, Szymon, et al.
Published: (2024)
Solution for OOD-CV UNICORN Challenge 2024 Object Detection Assistance LLM Counting Ability Improvement
by: Chi, Zhouyang, et al.
Published: (2024)
by: Chi, Zhouyang, et al.
Published: (2024)
Learning to Wander: Improving the Global Image Geolocation Ability of LMMs via Actionable Reasoning
by: Zheng, Yushuo, et al.
Published: (2026)
by: Zheng, Yushuo, et al.
Published: (2026)
Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Actial: Activate Spatial Reasoning Ability of Multimodal Large Language Models
by: Zhan, Xiaoyu, et al.
Published: (2025)
by: Zhan, Xiaoyu, et al.
Published: (2025)
Depth Map Denoising Network and Lightweight Fusion Network for Enhanced 3D Face Recognition
by: Xu, Ruizhuo, et al.
Published: (2024)
by: Xu, Ruizhuo, et al.
Published: (2024)
Rethinking Multimodal Learning from the Perspective of Mitigating Classification Ability Disproportion
by: Jiang, QingYuan, et al.
Published: (2025)
by: Jiang, QingYuan, et al.
Published: (2025)
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization
by: Hong, Minjie, et al.
Published: (2025)
by: Hong, Minjie, et al.
Published: (2025)
Smooth Operator: Smooth Verifiable Reward Activates Spatial Reasoning Ability of Vision-Language Model
by: Jiao, Siwen, et al.
Published: (2026)
by: Jiao, Siwen, et al.
Published: (2026)
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
by: Chen, Liang, et al.
Published: (2025)
by: Chen, Liang, et al.
Published: (2025)
PuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns
by: Chia, Yew Ken, et al.
Published: (2024)
by: Chia, Yew Ken, et al.
Published: (2024)
EscapeCraft: A 3D Room Escape Environment for Benchmarking Complex Multimodal Reasoning Ability
by: Wang, Ziyue, et al.
Published: (2025)
by: Wang, Ziyue, et al.
Published: (2025)
Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Model
by: Zhang, Wenqi, et al.
Published: (2024)
by: Zhang, Wenqi, et al.
Published: (2024)
CoT-Pose: Chain-of-Thought Reasoning for 3D Pose Generation from Abstract Prompts
by: Cha, Junuk, et al.
Published: (2025)
by: Cha, Junuk, et al.
Published: (2025)
Enhancing Visual Planning with Auxiliary Tasks and Multi-token Prediction
by: Zhang, Ce, et al.
Published: (2025)
by: Zhang, Ce, et al.
Published: (2025)
Slot Abstractors: Toward Scalable Abstract Visual Reasoning
by: Mondal, Shanka Subhra, et al.
Published: (2024)
by: Mondal, Shanka Subhra, et al.
Published: (2024)
Dysca: A Dynamic and Scalable Benchmark for Evaluating Perception Ability of LVLMs
by: Zhang, Jie, et al.
Published: (2024)
by: Zhang, Jie, et al.
Published: (2024)
Enhancing Few-Shot Out-of-Distribution Detection with Gradient Aligned Context Optimization
by: Tong, Baoshun, et al.
Published: (2024)
by: Tong, Baoshun, et al.
Published: (2024)
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
by: Neuhaus, Yannic, et al.
Published: (2026)
by: Neuhaus, Yannic, et al.
Published: (2026)
RULER-Bench: Probing Rule-based Reasoning Abilities of Next-level Video Generation Models for Vision Foundation Intelligence
by: He, Xuming, et al.
Published: (2025)
by: He, Xuming, et al.
Published: (2025)
RePlan: Reasoning-guided Region Planning for Complex Instruction-based Image Editing
by: Qu, Tianyuan, et al.
Published: (2025)
by: Qu, Tianyuan, et al.
Published: (2025)
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
by: Sharan, S P, et al.
Published: (2023)
by: Sharan, S P, et al.
Published: (2023)
Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation
by: Zhang, Xitie, et al.
Published: (2026)
by: Zhang, Xitie, et al.
Published: (2026)
PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis
by: Jin, Chuhao, et al.
Published: (2025)
by: Jin, Chuhao, et al.
Published: (2025)
Reasoning Can Hurt the Inductive Abilities of Large Language Models
by: Jin, Haibo, et al.
Published: (2025)
by: Jin, Haibo, et al.
Published: (2025)
Similar Items
-
Triple-CFN: Separating Concepts and Features Enhances Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2024) -
DIO: Refining Mutual Information and Causal Chain to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2025) -
Johnny: Structuring Representation Space to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2025) -
D4C: Improving Negative Example Quality to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2024) -
Solving the Clustering Reasoning Problems by Modeling a Deep-Learning-Based Probabilistic Model
by: Song, Ruizhuo, et al.
Published: (2024)