Saved in:
| Main Authors: | Lei, Xiaoxuan, Gomez, Lucas, Bai, Hao Yuan, Bashivan, Pouya |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.14343 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometry of naturalistic object representations in recurrent neural network models of working memory
by: Lei, Xiaoxuan, et al.
Published: (2024)
by: Lei, Xiaoxuan, et al.
Published: (2024)
Building spatial world models from sparse transitional episodic memories
by: He, Zizhan, et al.
Published: (2025)
by: He, Zizhan, et al.
Published: (2025)
Caption This, Reason That: VLMs Caught in the Middle
by: Weng, Zihan, et al.
Published: (2025)
by: Weng, Zihan, et al.
Published: (2025)
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026)
by: Pasand, Ali Saheb, et al.
Published: (2026)
Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain)
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
Enabling robots to follow abstract instructions and complete complex dynamic tasks
by: Mon-Williams, Ruaridh, et al.
Published: (2024)
by: Mon-Williams, Ruaridh, et al.
Published: (2024)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
by: Bhagwatkar, Rishika, et al.
Published: (2024)
by: Bhagwatkar, Rishika, et al.
Published: (2024)
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
by: Han, Yifu, et al.
Published: (2025)
by: Han, Yifu, et al.
Published: (2025)
COMET: "Cone of experience" enhanced large multimodal model for mathematical problem generation
by: Liu, Sannyuya, et al.
Published: (2024)
by: Liu, Sannyuya, et al.
Published: (2024)
Do LLMs estimate uncertainty well in instruction-following?
by: Heo, Juyeon, et al.
Published: (2024)
by: Heo, Juyeon, et al.
Published: (2024)
Do LLMs "know" internally when they follow instructions?
by: Heo, Juyeon, et al.
Published: (2024)
by: Heo, Juyeon, et al.
Published: (2024)
WizardLM: Empowering large pre-trained language models to follow complex instructions
by: Xu, Can, et al.
Published: (2023)
by: Xu, Can, et al.
Published: (2023)
SPECTRE: Spectral Pre-training Embeddings with Cylindrical Temporal Rotary Position Encoding for Fine-Grained sEMG-Based Movement Decoding
by: Weng, Zihan, et al.
Published: (2025)
by: Weng, Zihan, et al.
Published: (2025)
BLEUBERI: BLEU is a surprisingly effective reward for instruction following
by: Chang, Yapei, et al.
Published: (2025)
by: Chang, Yapei, et al.
Published: (2025)
A unified multimodal understanding and generation model for cross-disciplinary scientific research
by: Yang, Xiaomeng, et al.
Published: (2026)
by: Yang, Xiaomeng, et al.
Published: (2026)
Human-like in-group bias in instruction-tuned language model agents
by: Lee, Messi H. J.
Published: (2026)
by: Lee, Messi H. J.
Published: (2026)
A multimodal and temporal foundation model for virtual patient representations at healthcare system scale
by: Zhang, Andrew, et al.
Published: (2026)
by: Zhang, Andrew, et al.
Published: (2026)
A multimodal vision foundation model for generalizable knee pathology
by: Yu, Kang, et al.
Published: (2026)
by: Yu, Kang, et al.
Published: (2026)
Zero-shot cross-lingual transfer in instruction tuning of large language models
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
SLPL SHROOM at SemEval2024 Task 06: A comprehensive study on models ability to detect hallucination
by: Fallah, Pouya, et al.
Published: (2024)
by: Fallah, Pouya, et al.
Published: (2024)
OptimusKG: Unifying biomedical knowledge in a modern multimodal graph
by: Vittor, Lucas, et al.
Published: (2026)
by: Vittor, Lucas, et al.
Published: (2026)
A Semi-supervised Fake News Detection using Sentiment Encoding and LSTM with Self-Attention
by: Shaeri, Pouya, et al.
Published: (2024)
by: Shaeri, Pouya, et al.
Published: (2024)
URL: Universal Referential Knowledge Linking via Task-instructed Representation Compression
by: Li, Zhuoqun, et al.
Published: (2024)
by: Li, Zhuoqun, et al.
Published: (2024)
Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
MedReadCtrl: Personalizing medical text generation with readability-controlled instruction learning
by: Tran, Hieu, et al.
Published: (2025)
by: Tran, Hieu, et al.
Published: (2025)
Safety and accuracy follow different scaling laws in clinical large language models
by: Wind, Sebastian, et al.
Published: (2026)
by: Wind, Sebastian, et al.
Published: (2026)
A multi-scale vision transformer-based multimodal GeoAI model for mapping Arctic permafrost thaw
by: Li, Wenwen, et al.
Published: (2025)
by: Li, Wenwen, et al.
Published: (2025)
FastRM: An efficient and automatic explainability framework for multimodal generative models
by: Stan, Gabriela Ben-Melech, et al.
Published: (2024)
by: Stan, Gabriela Ben-Melech, et al.
Published: (2024)
Representation of the structure of graphs by sequences of instructions
by: Lopez-Rubio, Ezequiel
Published: (2025)
by: Lopez-Rubio, Ezequiel
Published: (2025)
MARCUS: An agentic, multimodal vision-language model for cardiac diagnosis and management
by: O'Sullivan, Jack W, et al.
Published: (2026)
by: O'Sullivan, Jack W, et al.
Published: (2026)
Simulating clinical interventions with a generative multimodal model of human physiology
by: Lutsker, Guy, et al.
Published: (2026)
by: Lutsker, Guy, et al.
Published: (2026)
Mars-PO: Multi-Agent Reasoning System Preference Optimization
by: Lou, Xiaoxuan, et al.
Published: (2024)
by: Lou, Xiaoxuan, et al.
Published: (2024)
Leveraging AI multimodal geospatial foundation models for improved near-real-time flood mapping at a global scale
by: Tulbure, Mirela G., et al.
Published: (2025)
by: Tulbure, Mirela G., et al.
Published: (2025)
Retrieval-augmented in-context learning for multimodal large language models in disease classification
by: Zhan, Zaifu, et al.
Published: (2025)
by: Zhan, Zaifu, et al.
Published: (2025)
Multimodal Trajectory Prediction for Autonomous Driving on Unstructured Roads using Deep Convolutional Network
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
WGRAMMAR: Leverage Prior Knowledge to Accelerate Structured Decoding
by: Wang, Ran, et al.
Published: (2025)
by: Wang, Ran, et al.
Published: (2025)
DA-Mamba: Dialogue-aware selective state-space model for multimodal engagement estimation
by: Kang, Shenwei, et al.
Published: (2025)
by: Kang, Shenwei, et al.
Published: (2025)
AI-driven multi-omics integration for multi-scale predictive modeling of causal genotype-environment-phenotype relationships
by: Wu, You, et al.
Published: (2024)
by: Wu, You, et al.
Published: (2024)
D-RMGPT: Robot-assisted collaborative tasks driven by large multimodal models
by: Forlini, M., et al.
Published: (2024)
by: Forlini, M., et al.
Published: (2024)
MID-L: Matrix-Interpolated Dropout Layer with Layer-wise Neuron Selection
by: Shaeri, Pouya, et al.
Published: (2025)
by: Shaeri, Pouya, et al.
Published: (2025)
Similar Items
-
Geometry of naturalistic object representations in recurrent neural network models of working memory
by: Lei, Xiaoxuan, et al.
Published: (2024) -
Building spatial world models from sparse transitional episodic memories
by: He, Zizhan, et al.
Published: (2025) -
Caption This, Reason That: VLMs Caught in the Middle
by: Weng, Zihan, et al.
Published: (2025) -
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026) -
Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain)
by: Oota, Subba Reddy, et al.
Published: (2025)