Gespeichert in:
| Hauptverfasser: | Yu, Dongjian, Min, Weiqing, Jiang, Qian, Lin, Xing, Jin, Xin, Jiang, Shuqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.12356 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Synthesizing Knowledge-enhanced Features for Real-world Zero-shot Food Detection
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024)
FoodSky: A Food-oriented Large Language Model that Passes the Chef and Dietetic Examination
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024)
Advancing Food Nutrition Estimation via Visual-Ingredient Feature Fusion
von: Qi, Huiyan, et al.
Veröffentlicht: (2025)
von: Qi, Huiyan, et al.
Veröffentlicht: (2025)
ReflexSplit: Single Image Reflection Separation via Layer Fusion-Separation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
Speed of Random Walks in Dirichlet Environment on a Galton-Watson Tree
von: Qian, Dongjian, et al.
Veröffentlicht: (2024)
von: Qian, Dongjian, et al.
Veröffentlicht: (2024)
Scaling limit of the random walk on a Galton-Watson tree with regular varying offspring distribution
von: Qian, Dongjian, et al.
Veröffentlicht: (2023)
von: Qian, Dongjian, et al.
Veröffentlicht: (2023)
Butter: Frequency Consistency and Hierarchical Fusion for Autonomous Driving Object Detection
von: Lin, Xiaojian, et al.
Veröffentlicht: (2025)
von: Lin, Xiaojian, et al.
Veröffentlicht: (2025)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
von: Jin, Yang, et al.
Veröffentlicht: (2024)
von: Jin, Yang, et al.
Veröffentlicht: (2024)
FUSE: Label-Free Image-Event Joint Monocular Depth Estimation via Frequency-Decoupled Alignment and Degradation-Robust Fusion
von: Sun, Pihai, et al.
Veröffentlicht: (2025)
von: Sun, Pihai, et al.
Veröffentlicht: (2025)
EM-DARTS: Hierarchical Differentiable Architecture Search for Eye Movement Recognition
von: Qin, Huafeng, et al.
Veröffentlicht: (2024)
von: Qin, Huafeng, et al.
Veröffentlicht: (2024)
AlignMamba-2: Enhancing Multimodal Fusion and Sentiment Analysis with Modality-Aware Mamba
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
MaeFuse: Transferring Omni Features with Pretrained Masked Autoencoders for Infrared and Visible Image Fusion via Guided Training
von: Li, Jiayang, et al.
Veröffentlicht: (2024)
von: Li, Jiayang, et al.
Veröffentlicht: (2024)
Regular subspaces of symmetric stable processes
von: Qian, Dongjian, et al.
Veröffentlicht: (2022)
von: Qian, Dongjian, et al.
Veröffentlicht: (2022)
Critical Self-Similar Markov Trees
von: Curien, Nicolas, et al.
Veröffentlicht: (2026)
von: Curien, Nicolas, et al.
Veröffentlicht: (2026)
MFDNet: Multi-Frequency Deflare Network for Efficient Nighttime Flare Removal
von: Jiang, Yiguo, et al.
Veröffentlicht: (2024)
von: Jiang, Yiguo, et al.
Veröffentlicht: (2024)
Multilingual Text-to-Image Person Retrieval via Bidirectional Relation Reasoning and Aligning
von: Cao, Min, et al.
Veröffentlicht: (2025)
von: Cao, Min, et al.
Veröffentlicht: (2025)
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
von: Wang, Junke, et al.
Veröffentlicht: (2024)
von: Wang, Junke, et al.
Veröffentlicht: (2024)
From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities
von: Jiang, Shixin, et al.
Veröffentlicht: (2024)
von: Jiang, Shixin, et al.
Veröffentlicht: (2024)
PhaSR: Generalized Image Shadow Removal with Physically Aligned Priors
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
OmniGen2: Towards Instruction-Aligned Multimodal Generation
von: Wu, Chenyuan, et al.
Veröffentlicht: (2025)
von: Wu, Chenyuan, et al.
Veröffentlicht: (2025)
Holistic Dynamic Frequency Transformer for Image Fusion and Exposure Correction
von: Shang, Xiaoke, et al.
Veröffentlicht: (2023)
von: Shang, Xiaoke, et al.
Veröffentlicht: (2023)
Neural Dynamics-Informed Pre-trained Framework for Personalized Brain Functional Network Construction
von: Jiang, Hongjie, et al.
Veröffentlicht: (2026)
von: Jiang, Hongjie, et al.
Veröffentlicht: (2026)
Omni-Fusion of Spatial and Spectral for Hyperspectral Image Segmentation
von: Zhang, Qing, et al.
Veröffentlicht: (2025)
von: Zhang, Qing, et al.
Veröffentlicht: (2025)
Oscillations-Aware Frequency Security Assessment via Efficient Worst-Case Frequency Nadir Computation
von: Jiang, Yan, et al.
Veröffentlicht: (2024)
von: Jiang, Yan, et al.
Veröffentlicht: (2024)
Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
Quadratic Differentials as Stability Conditions of Graded Skew-gentle Algebras
von: Lu, Suiqi, et al.
Veröffentlicht: (2023)
von: Lu, Suiqi, et al.
Veröffentlicht: (2023)
A counterexample to the Jordan-Hölder property for polarizable semiorthogonal decompositions
von: Haiden, Fabian, et al.
Veröffentlicht: (2025)
von: Haiden, Fabian, et al.
Veröffentlicht: (2025)
Stability Conditions on $\mathbb P^3$
von: Wu, Dongjian, et al.
Veröffentlicht: (2024)
von: Wu, Dongjian, et al.
Veröffentlicht: (2024)
The $G$-Noncommutative Minimal Model Program
von: Wu, Dongjian, et al.
Veröffentlicht: (2026)
von: Wu, Dongjian, et al.
Veröffentlicht: (2026)
Relative Stability Conditions on Triangulated Categories
von: Liu, Bowen, et al.
Veröffentlicht: (2024)
von: Liu, Bowen, et al.
Veröffentlicht: (2024)
Stability Conditions and Algebraic Hearts for Acyclic Quivers
von: Otani, Takumi, et al.
Veröffentlicht: (2025)
von: Otani, Takumi, et al.
Veröffentlicht: (2025)
Dendritic Convolution for Noise Image Recognition
von: Xue, Jiarui, et al.
Veröffentlicht: (2025)
von: Xue, Jiarui, et al.
Veröffentlicht: (2025)
OmniFusion: Simultaneous Multilingual Multimodal Translations via Modular Fusion
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
DecAlign: Hierarchical Cross-Modal Alignment for Decoupled Multimodal Representation Learning
von: Qian, Chengxuan, et al.
Veröffentlicht: (2025)
von: Qian, Chengxuan, et al.
Veröffentlicht: (2025)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
Financial Modeling and Predictive Innovation in the Food-Tech Industry: Aligning Profitability with Nutritional Sustainability
von: Audu, Mariam Sha
Veröffentlicht: (2025)
von: Audu, Mariam Sha
Veröffentlicht: (2025)
Frequency-Domain Fusion Transformer for Image Inpainting
von: He, Sijin, et al.
Veröffentlicht: (2025)
von: He, Sijin, et al.
Veröffentlicht: (2025)
Omni-Captioner: Data Pipeline, Models, and Benchmark for Omni Detailed Perception
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
HumanMaterial: Human Material Estimation from a Single Image via Progressive Training
von: Jiang, Yu, et al.
Veröffentlicht: (2025)
von: Jiang, Yu, et al.
Veröffentlicht: (2025)
EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM
von: Zong, Zhuofan, et al.
Veröffentlicht: (2024)
von: Zong, Zhuofan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Synthesizing Knowledge-enhanced Features for Real-world Zero-shot Food Detection
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024) -
FoodSky: A Food-oriented Large Language Model that Passes the Chef and Dietetic Examination
von: Zhou, Pengfei, et al.
Veröffentlicht: (2024) -
Advancing Food Nutrition Estimation via Visual-Ingredient Feature Fusion
von: Qi, Huiyan, et al.
Veröffentlicht: (2025) -
ReflexSplit: Single Image Reflection Separation via Layer Fusion-Separation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026) -
Speed of Random Walks in Dirichlet Environment on a Galton-Watson Tree
von: Qian, Dongjian, et al.
Veröffentlicht: (2024)