SkillRater: Untangling Capabilities in Multimodal Data
Fuente:
arXiv
Saved in:
| Main Authors: | Sahi, Naveen, Dohmann, Jeremy, Aghajanyan, Armen, Shrivastava, Akshat |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving MoE Compute Efficiency by Composing Weight and Data Sparsity
by: Kilian, Maciej, et al.
Published: (2026)
by: Kilian, Maciej, et al.
Published: (2026)
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
by: Lin, Xi Victoria, et al.
Published: (2024)
by: Lin, Xi Victoria, et al.
Published: (2024)
When Worse is Better: Navigating the compression-generation tradeoff in visual tokenization
by: Ramanujan, Vivek, et al.
Published: (2024)
by: Ramanujan, Vivek, et al.
Published: (2024)
Small Molecule Optimization with Large Language Models
by: Guevorguian, Philipp, et al.
Published: (2024)
by: Guevorguian, Philipp, et al.
Published: (2024)
DataRater: Meta-Learned Dataset Curation
by: Calian, Dan A., et al.
Published: (2025)
by: Calian, Dan A., et al.
Published: (2025)
Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence
by: Resnick, Paul, et al.
Published: (2021)
by: Resnick, Paul, et al.
Published: (2021)
CoSMoEs: Compact Sparse Mixture of Experts
by: Huber, Patrick, et al.
Published: (2025)
by: Huber, Patrick, et al.
Published: (2025)
TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation
by: Kirscher, Tristan, et al.
Published: (2026)
by: Kirscher, Tristan, et al.
Published: (2026)
CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries
by: Mu, Ni, et al.
Published: (2025)
by: Mu, Ni, et al.
Published: (2025)
Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck
by: Huang, Zihan, et al.
Published: (2026)
by: Huang, Zihan, et al.
Published: (2026)
A Timeline and Analysis for Representation Plasticity in Large Language Models
by: Kannan, Akshat
Published: (2024)
by: Kannan, Akshat
Published: (2024)
Comparison of Scoring Rationales Between Large Language Models and Human Raters
by: Hua, Haowei, et al.
Published: (2025)
by: Hua, Haowei, et al.
Published: (2025)
Obstacle-aware Gaussian Process Regression
by: Shrivastava, Gaurav
Published: (2024)
by: Shrivastava, Gaurav
Published: (2024)
Text Quality-Based Pruning for Efficient Training of Language Models
by: Sharma, Vasu, et al.
Published: (2024)
by: Sharma, Vasu, et al.
Published: (2024)
Untangling Lariats: Subgradient Following of Variationally Penalized Objectives
by: Mo, Kai-Chia, et al.
Published: (2024)
by: Mo, Kai-Chia, et al.
Published: (2024)
Comparing Human and AI Rater Effects Using the Many-Facet Rasch Model
by: Jiao, Hong, et al.
Published: (2025)
by: Jiao, Hong, et al.
Published: (2025)
Mitigating Attrition: Data-Driven Approach Using Machine Learning and Data Engineering
by: Vijayan, Naveen Edapurath
Published: (2025)
by: Vijayan, Naveen Edapurath
Published: (2025)
Generative Kaleidoscopic Networks
by: Shrivastava, Harsh
Published: (2024)
by: Shrivastava, Harsh
Published: (2024)
Untangling Knots: Leveraging LLM for Error Resolution in Computational Notebooks
by: Grotov, Konstantin, et al.
Published: (2024)
by: Grotov, Konstantin, et al.
Published: (2024)
Correcting Human Labels for Rater Effects in AI Evaluation: An Item Response Theory Approach
by: Casabianca, Jodi M., et al.
Published: (2026)
by: Casabianca, Jodi M., et al.
Published: (2026)
Neural Graph Revealers
by: Shrivastava, Harsh, et al.
Published: (2023)
by: Shrivastava, Harsh, et al.
Published: (2023)
Are uGLAD? Time will tell!
by: Imani, Shima, et al.
Published: (2023)
by: Imani, Shima, et al.
Published: (2023)
LeanQuant: Accurate and Scalable Large Language Model Quantization with Loss-error-aware Grid
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Methods for Recovering Conditional Independence Graphs: A Survey
by: Shrivastava, Harsh, et al.
Published: (2022)
by: Shrivastava, Harsh, et al.
Published: (2022)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Triple Component Matrix Factorization: Untangling Global, Local, and Noisy Components
by: Shi, Naichen, et al.
Published: (2024)
by: Shi, Naichen, et al.
Published: (2024)
ALICE: Combining Feature Selection and Inter-Rater Agreeability for Machine Learning Insights
by: Anasashvili, Bachana, et al.
Published: (2024)
by: Anasashvili, Bachana, et al.
Published: (2024)
Modeling Freight Mode Choice Using Machine Learning Classifiers: A Comparative Study Using the Commodity Flow Survey (CFS) Data
by: Uddin, Majbah, et al.
Published: (2024)
by: Uddin, Majbah, et al.
Published: (2024)
Out-of-Distribution Data: An Acquaintance of Adversarial Examples -- A Survey
by: Karunanayake, Naveen, et al.
Published: (2024)
by: Karunanayake, Naveen, et al.
Published: (2024)
GenMM: Geometrically and Temporally Consistent Multimodal Data Generation for Video and LiDAR
by: Singh, Bharat, et al.
Published: (2024)
by: Singh, Bharat, et al.
Published: (2024)
PrE-Text: Training Language Models on Private Federated Data in the Age of LLMs
by: Hou, Charlie, et al.
Published: (2024)
by: Hou, Charlie, et al.
Published: (2024)
NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training
by: Wu, Fang, et al.
Published: (2026)
by: Wu, Fang, et al.
Published: (2026)
Untangling Component Imbalance in Hybrid Linear Attention Conversion Methods
by: Benfeghoul, Martin, et al.
Published: (2025)
by: Benfeghoul, Martin, et al.
Published: (2025)
Diagnosing Capability Gaps in Fine-Tuning Data
by: Taghanaki, Saeid Asgari, et al.
Published: (2026)
by: Taghanaki, Saeid Asgari, et al.
Published: (2026)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Frontier Models are Capable of In-context Scheming
by: Meinke, Alexander, et al.
Published: (2024)
by: Meinke, Alexander, et al.
Published: (2024)
Cyclical Log Annealing as a Learning Rate Scheduler
by: Naveen, Philip
Published: (2024)
by: Naveen, Philip
Published: (2024)
MuTT: A Multimodal Trajectory Transformer for Robot Skills
by: Kienle, Claudius, et al.
Published: (2024)
by: Kienle, Claudius, et al.
Published: (2024)
Untangling Gaussian Mixtures
by: Fluck, Eva, et al.
Published: (2024)
by: Fluck, Eva, et al.
Published: (2024)
TraCeS: Trajectory Based Credit Assignment From Sparse Safety Feedback
by: Low, Siow Meng, et al.
Published: (2025)
by: Low, Siow Meng, et al.
Published: (2025)
Similar Items
-
Improving MoE Compute Efficiency by Composing Weight and Data Sparsity
by: Kilian, Maciej, et al.
Published: (2026) -
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
by: Lin, Xi Victoria, et al.
Published: (2024) -
When Worse is Better: Navigating the compression-generation tradeoff in visual tokenization
by: Ramanujan, Vivek, et al.
Published: (2024) -
Small Molecule Optimization with Large Language Models
by: Guevorguian, Philipp, et al.
Published: (2024) -
DataRater: Meta-Learned Dataset Curation
by: Calian, Dan A., et al.
Published: (2025)