Saliency-Aware Multi-Route Thinking: Revisiting Vision-Language Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Shi, Mingjia, He, Yinhan, Zhu, Yaochen, Li, Jundong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Landmark-Aware Visual Navigation Dataset
por: Johnson, Faith, et al.
Publicado: (2024)
por: Johnson, Faith, et al.
Publicado: (2024)
Knowledge Distillation: Enhancing Neural Network Compression with Integrated Gradients
por: Hernandez, David E., et al.
Publicado: (2025)
por: Hernandez, David E., et al.
Publicado: (2025)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
por: Kalušev, Vladimir, et al.
Publicado: (2026)
por: Kalušev, Vladimir, et al.
Publicado: (2026)
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
por: S, Kamal Basha, et al.
Publicado: (2024)
por: S, Kamal Basha, et al.
Publicado: (2024)
Exploiting Precision Mapping and Component-Specific Feature Enhancement for Breast Cancer Segmentation and Identification
por: V, Pandiyaraju, et al.
Publicado: (2024)
por: V, Pandiyaraju, et al.
Publicado: (2024)
Model compression using knowledge distillation with integrated gradients
por: Hernandez, David E., et al.
Publicado: (2025)
por: Hernandez, David E., et al.
Publicado: (2025)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
por: Cao, Bryan Bo, et al.
Publicado: (2024)
por: Cao, Bryan Bo, et al.
Publicado: (2024)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
por: Cao, Deyu, et al.
Publicado: (2025)
por: Cao, Deyu, et al.
Publicado: (2025)
Verification and Validation for Trustworthy Scientific Machine Learning
por: Jakeman, John D., et al.
Publicado: (2025)
por: Jakeman, John D., et al.
Publicado: (2025)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
por: Filus, Katarzyna, et al.
Publicado: (2025)
por: Filus, Katarzyna, et al.
Publicado: (2025)
Exploiting Latent Properties to Optimize Neural Codecs
por: Balcilar, Muhammet, et al.
Publicado: (2025)
por: Balcilar, Muhammet, et al.
Publicado: (2025)
Super-Resolution Enhancement of Medical Images Based on Diffusion Model: An Optimization Scheme for Low-Resolution Gastric Images
por: Jia, Haozhe
Publicado: (2025)
por: Jia, Haozhe
Publicado: (2025)
Ring Artifacts Removal Based on Implicit Neural Representation of Sinogram Data
por: Shi, Ligen, et al.
Publicado: (2024)
por: Shi, Ligen, et al.
Publicado: (2024)
Adversarial Patch Attacks on Vision-Based Cargo Occupancy Estimation via Differentiable 3D Simulation
por: Hedna, Mohamed Rissal, et al.
Publicado: (2025)
por: Hedna, Mohamed Rissal, et al.
Publicado: (2025)
Balanced conic rectified flow
por: Kim, Shin Seong, et al.
Publicado: (2025)
por: Kim, Shin Seong, et al.
Publicado: (2025)
Human Vision Constrained Super-Resolution
por: Karpenko, Volodymyr, et al.
Publicado: (2024)
por: Karpenko, Volodymyr, et al.
Publicado: (2024)
Rendering Anywhere You See: Renderability Field-guided Gaussian Splatting
por: Jin, Xiaofeng, et al.
Publicado: (2025)
por: Jin, Xiaofeng, et al.
Publicado: (2025)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
por: Pepe, Alberto, et al.
Publicado: (2026)
por: Pepe, Alberto, et al.
Publicado: (2026)
TSPE-GS: Probabilistic Depth Extraction for Semi-Transparent Surface Reconstruction via 3D Gaussian Splatting
por: Xu, Zhiyuan, et al.
Publicado: (2025)
por: Xu, Zhiyuan, et al.
Publicado: (2025)
Surrealistic-like Image Generation with Vision-Language Models
por: Ayten, Elif, et al.
Publicado: (2024)
por: Ayten, Elif, et al.
Publicado: (2024)
HOSC: A Periodic Activation with Saturation Control for High-Fidelity Implicit Neural Representations
por: Wlodarczyk, Michal Jan, et al.
Publicado: (2026)
por: Wlodarczyk, Michal Jan, et al.
Publicado: (2026)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
por: Kambhatla, Akhila, et al.
Publicado: (2025)
por: Kambhatla, Akhila, et al.
Publicado: (2025)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
por: Koubaa, Anis, et al.
Publicado: (2025)
por: Koubaa, Anis, et al.
Publicado: (2025)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
por: Gautam, Sushant, et al.
Publicado: (2025)
por: Gautam, Sushant, et al.
Publicado: (2025)
On the Structural Failure of Chamfer Distance in 3D Shape Optimization
por: Song, Chang-Yong, et al.
Publicado: (2026)
por: Song, Chang-Yong, et al.
Publicado: (2026)
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
Ray-driven Spectral CT Reconstruction Based on Neural Base-Material Fields
por: Shi, Ligen, et al.
Publicado: (2024)
por: Shi, Ligen, et al.
Publicado: (2024)
A geometric modelling framework to support the design of heterogeneous lattice structures with non-linearly varying geometry
por: Letov, Nikita, et al.
Publicado: (2026)
por: Letov, Nikita, et al.
Publicado: (2026)
JVLGS: Joint Vision-Language Gas Leak Segmentation
por: Zhao, Xinlong, et al.
Publicado: (2025)
por: Zhao, Xinlong, et al.
Publicado: (2025)
FlatCAD: Fast Curvature Regularization of Neural SDFs for CAD Models
por: Yin, Haotian, et al.
Publicado: (2025)
por: Yin, Haotian, et al.
Publicado: (2025)
ViFiCon: Vision and Wireless Association Via Self-Supervised Contrastive Learning
por: Meegan, Nicholas, et al.
Publicado: (2022)
por: Meegan, Nicholas, et al.
Publicado: (2022)
Unlocking UML Class Diagram Understanding in Vision Language Models
por: Naboichenko, Artem, et al.
Publicado: (2026)
por: Naboichenko, Artem, et al.
Publicado: (2026)
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
por: Nayak, Prabhudarshi, et al.
Publicado: (2026)
por: Nayak, Prabhudarshi, et al.
Publicado: (2026)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
por: Viveiros, André G., et al.
Publicado: (2025)
por: Viveiros, André G., et al.
Publicado: (2025)
Attention Please: What Transformer Models Really Learn for Process Prediction
por: Käppel, Martin, et al.
Publicado: (2024)
por: Käppel, Martin, et al.
Publicado: (2024)
ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model
por: Zhang, Haichao, et al.
Publicado: (2026)
por: Zhang, Haichao, et al.
Publicado: (2026)
RLFTSim: Realistic and Controllable Multi-Agent Traffic Simulation via Reinforcement Learning Fine-Tuning
por: Ahmadi, Ehsan, et al.
Publicado: (2026)
por: Ahmadi, Ehsan, et al.
Publicado: (2026)
TGraphX: Tensor-Aware Graph Neural Network for Multi-Dimensional Feature Learning
por: Sajjadi, Arash, et al.
Publicado: (2025)
por: Sajjadi, Arash, et al.
Publicado: (2025)
Swish-T : Enhancing Swish Activation with Tanh Bias for Improved Neural Network Performance
por: Seo, Youngmin, et al.
Publicado: (2024)
por: Seo, Youngmin, et al.
Publicado: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
por: Merugu, Ranjith, et al.
Publicado: (2025)
por: Merugu, Ranjith, et al.
Publicado: (2025)
Ejemplares similares
-
A Landmark-Aware Visual Navigation Dataset
por: Johnson, Faith, et al.
Publicado: (2024) -
Knowledge Distillation: Enhancing Neural Network Compression with Integrated Gradients
por: Hernandez, David E., et al.
Publicado: (2025) -
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
por: Kalušev, Vladimir, et al.
Publicado: (2026) -
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
por: S, Kamal Basha, et al.
Publicado: (2024) -
Exploiting Precision Mapping and Component-Specific Feature Enhancement for Breast Cancer Segmentation and Identification
por: V, Pandiyaraju, et al.
Publicado: (2024)