Saved in:
| Main Authors: | Zhang, Tianle, Fang, Wanlong, Woo, Jonathan, Latawa, Paridhi, Subramanian, Deepak A., Chan, Alvin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.17552 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Align or Not to Align: Strategic Multimodal Representation Alignment for Optimal Performance
by: Fang, Wanlong, et al.
Published: (2025)
by: Fang, Wanlong, et al.
Published: (2025)
Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition
by: Fang, Wanlong, et al.
Published: (2026)
by: Fang, Wanlong, et al.
Published: (2026)
Unlocking Noisy Real-World Corpora for Foundation Model Pre-Training via Quality-Aware Tokenization
by: Gollwitzer, Arvid E., et al.
Published: (2026)
by: Gollwitzer, Arvid E., et al.
Published: (2026)
CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
How Creative Are Large Language Models in Generating Molecules?
by: Tao, Wen, et al.
Published: (2026)
by: Tao, Wen, et al.
Published: (2026)
Unseen Object Reasoning with Shared Appearance Cues
by: Singh, Paridhi, et al.
Published: (2024)
by: Singh, Paridhi, et al.
Published: (2024)
Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
A Systematic Study of Cross-Modal Typographic Attacks on Audio-Visual Reasoning
by: Chen, Tianle, et al.
Published: (2026)
by: Chen, Tianle, et al.
Published: (2026)
Impact of ESG Disclosures on Corporate Financial Performance: An Industry‐Specific Analysis of Indian Firms
by: Paridhi, et al.
Published: (2025)
by: Paridhi, et al.
Published: (2025)
SLAP: The Semantic Least Action Principle for Variational Video-Language Modeling
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
LLMs Can Evolve Continually on Modality for X-Modal Reasoning
by: Yu, Jiazuo, et al.
Published: (2024)
by: Yu, Jiazuo, et al.
Published: (2024)
The Triangle of Similarity: A Multi-Faceted Framework for Comparing Neural Network Representations
by: Sirikova, Olha, et al.
Published: (2026)
by: Sirikova, Olha, et al.
Published: (2026)
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
by: Chen, Xingwu, et al.
Published: (2025)
by: Chen, Xingwu, et al.
Published: (2025)
Stack-Based Dynamic Context Allocation: Engineering Case Study
by: Mohan, Deepak
Published: (2026)
by: Mohan, Deepak
Published: (2026)
Optimizing Small Language Models for NL2SQL via Chain-of-Thought Fine-Tuning
by: Solanki, Anshul, et al.
Published: (2026)
by: Solanki, Anshul, et al.
Published: (2026)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
by: Wang, Tianle, et al.
Published: (2026)
by: Wang, Tianle, et al.
Published: (2026)
Can Reasons Help Improve Pedestrian Intent Estimation? A Cross-Modal Approach
by: Khindkar, Vaishnavi, et al.
Published: (2024)
by: Khindkar, Vaishnavi, et al.
Published: (2024)
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
by: Hsiao, Teng-Fang, et al.
Published: (2025)
by: Hsiao, Teng-Fang, et al.
Published: (2025)
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
by: Wang, Xunguang, et al.
Published: (2024)
by: Wang, Xunguang, et al.
Published: (2024)
ESG Practices and Financial Stability in Emerging Market Healthcare Companies: Insights Amidst the COVID‐19 Pandemic
by: Paridhi, et al.
Published: (2024)
by: Paridhi, et al.
Published: (2024)
Electrical Current‐Mediated Transformation for Efficient Plant Genome Editing: A Case Study in Faba Bean
by: Sruthy Maria Augustine, et al.
Published: (2025)
by: Sruthy Maria Augustine, et al.
Published: (2025)
Can Slow-thinking LLMs Reason Over Time? Empirical Studies in Time Series Forecasting
by: Cheng, Mingyue, et al.
Published: (2025)
by: Cheng, Mingyue, et al.
Published: (2025)
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper
by: Wu, Daoyuan, et al.
Published: (2024)
by: Wu, Daoyuan, et al.
Published: (2024)
Immuno-VLM: Immunizing Large Vision-Language Models via Generative Semantic Antibodies for Open-World Trustworthiness
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
CTTA-T: Continual Test-Time Adaptation for Text Understanding via Teacher-Student with a Domain-aware and Generalized Teacher
by: Liu, Tianlun, et al.
Published: (2025)
by: Liu, Tianlun, et al.
Published: (2025)
Interpolating Video-LLMs: Toward Longer-sequence LMMs in a Training-free Manner
by: Shang, Yuzhang, et al.
Published: (2024)
by: Shang, Yuzhang, et al.
Published: (2024)
BioVERSE: Representation Alignment of Biomedical Modalities to LLMs for Multi-Modal Reasoning
by: Tsou, Ching-Huei, et al.
Published: (2025)
by: Tsou, Ching-Huei, et al.
Published: (2025)
An Empirical Study of Aegis
by: Saragih, Daniel, et al.
Published: (2024)
by: Saragih, Daniel, et al.
Published: (2024)
TimeGraphs: Graph-based Temporal Reasoning
by: Maheshwari, Paridhi, et al.
Published: (2024)
by: Maheshwari, Paridhi, et al.
Published: (2024)
Improving Physical Object State Representation in Text-to-Image Generative Systems
by: Chen, Tianle, et al.
Published: (2025)
by: Chen, Tianle, et al.
Published: (2025)
Blood PCSK9 Impacts Alzheimer's Disease Risk in an APOE Genotype‐Dependent Manner: A Prospective Cohort Study
by: Qiushan Tao, et al.
Published: (2026)
by: Qiushan Tao, et al.
Published: (2026)
Can Post-Training Transform LLMs into Causal Reasoners?
by: Chen, Junqi, et al.
Published: (2026)
by: Chen, Junqi, et al.
Published: (2026)
AI Founding Fathers: A Case Study of GIS Search in Multi-Agent Pipelines
by: Chauhan, Alvin
Published: (2025)
by: Chauhan, Alvin
Published: (2025)
KiC: Keyword-inspired Cascade for Cost-Efficient Text Generation with LLMs
by: Kim, Woo-Chan, et al.
Published: (2025)
by: Kim, Woo-Chan, et al.
Published: (2025)
Can LLMs Interpret and Leverage Structured Linguistic Representations? A Case Study with AMRs
by: Raut, Ankush, et al.
Published: (2025)
by: Raut, Ankush, et al.
Published: (2025)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
by: Wang, Yibin, et al.
Published: (2025)
by: Wang, Yibin, et al.
Published: (2025)
Is Factuality Enhancement a Free Lunch For LLMs? Better Factuality Can Lead to Worse Context-Faithfulness
by: Bi, Baolong, et al.
Published: (2024)
by: Bi, Baolong, et al.
Published: (2024)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
by: Zengaffinen, Yanick, et al.
Published: (2026)
by: Zengaffinen, Yanick, et al.
Published: (2026)
Similar Items
-
To Align or Not to Align: Strategic Multimodal Representation Alignment for Optimal Performance
by: Fang, Wanlong, et al.
Published: (2025) -
Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition
by: Fang, Wanlong, et al.
Published: (2026) -
Unlocking Noisy Real-World Corpora for Foundation Model Pre-Training via Quality-Aware Tokenization
by: Gollwitzer, Arvid E., et al.
Published: (2026) -
CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning
by: Fang, Xiang, et al.
Published: (2026) -
How Creative Are Large Language Models in Generating Molecules?
by: Tao, Wen, et al.
Published: (2026)