GenCeption: Evaluate Vision LLMs with Unlabeled Unimodal Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Lele, Buchner, Valentin, Senane, Zineb, Yang, Fangkai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Supervised Learning of Time Series Representation via Diffusion Process and Imputation-Interpolation-Forecasting Mask
von: Senane, Zineb, et al.
Veröffentlicht: (2024)
von: Senane, Zineb, et al.
Veröffentlicht: (2024)
Prompt Tuned Embedding Classification for Multi-Label Industry Sector Allocation
von: Buchner, Valentin Leonhard, et al.
Veröffentlicht: (2023)
von: Buchner, Valentin Leonhard, et al.
Veröffentlicht: (2023)
Unveiling Transformer Perception by Exploring Input Manifolds
von: Benfenati, Alessandro, et al.
Veröffentlicht: (2024)
von: Benfenati, Alessandro, et al.
Veröffentlicht: (2024)
Sycophancy as compositions of Atomic Psychometric Traits
von: Jain, Shreyans, et al.
Veröffentlicht: (2025)
von: Jain, Shreyans, et al.
Veröffentlicht: (2025)
Generative AI for Strategic Plan Development
von: Ponnock, Jesse
Veröffentlicht: (2025)
von: Ponnock, Jesse
Veröffentlicht: (2025)
GPT-4 Generated Narratives of Life Events using a Structured Narrative Prompt: A Validation Study
von: Lynch, Christopher J., et al.
Veröffentlicht: (2024)
von: Lynch, Christopher J., et al.
Veröffentlicht: (2024)
From Rule-Based Models to Deep Learning Transformers Architectures for Natural Language Processing and Sign Language Translation Systems: Survey, Taxonomy and Performance Evaluation
von: Shahin, Nada, et al.
Veröffentlicht: (2024)
von: Shahin, Nada, et al.
Veröffentlicht: (2024)
From Capabilities to Performance: Evaluating Key Functional Properties of LLM Architectures in Penetration Testing
von: Huang, Lanxiao, et al.
Veröffentlicht: (2025)
von: Huang, Lanxiao, et al.
Veröffentlicht: (2025)
Graph Language Models
von: Plenz, Moritz, et al.
Veröffentlicht: (2024)
von: Plenz, Moritz, et al.
Veröffentlicht: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
No Saved Kaleidosope: an 100% Jitted Neural Network Coding Language with Pythonic Syntax
von: da Rosa, Augusto Seben, et al.
Veröffentlicht: (2024)
von: da Rosa, Augusto Seben, et al.
Veröffentlicht: (2024)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
FedDPG: An Adaptive Yet Efficient Prompt-tuning Approach in Federated Learning Settings
von: Shakeri, Ali, et al.
Veröffentlicht: (2025)
von: Shakeri, Ali, et al.
Veröffentlicht: (2025)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
Evaluating Long Range Dependency Handling in Code Generation LLMs
von: Assogba, Yannick, et al.
Veröffentlicht: (2024)
von: Assogba, Yannick, et al.
Veröffentlicht: (2024)
CopySpec: Accelerating LLMs with Speculative Copy-and-Paste Without Compromising Quality
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2025)
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2025)
Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
Synthesizing Behaviorally-Grounded Reasoning Chains: A Data-Generation Framework for Personal Finance LLMs
von: Theerthala, Akhil
Veröffentlicht: (2025)
von: Theerthala, Akhil
Veröffentlicht: (2025)
FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
von: Dua, Karan, et al.
Veröffentlicht: (2025)
von: Dua, Karan, et al.
Veröffentlicht: (2025)
Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs
von: Tytarenko, Stepan, et al.
Veröffentlicht: (2024)
von: Tytarenko, Stepan, et al.
Veröffentlicht: (2024)
ALBA: A European Portuguese Benchmark for Evaluating Language and Linguistic Dimensions in Generative LLMs
von: Vieira, Inês, et al.
Veröffentlicht: (2026)
von: Vieira, Inês, et al.
Veröffentlicht: (2026)
ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answering
von: Ghosh, Shubhra, et al.
Veröffentlicht: (2025)
von: Ghosh, Shubhra, et al.
Veröffentlicht: (2025)
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
von: Pareja, Aldo, et al.
Veröffentlicht: (2024)
von: Pareja, Aldo, et al.
Veröffentlicht: (2024)
Contextual Integrity in LLMs via Reasoning and Reinforcement Learning
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
Evaluating Class Membership Relations in Knowledge Graphs using Large Language Models
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations
von: Kumar, Sachin
Veröffentlicht: (2026)
von: Kumar, Sachin
Veröffentlicht: (2026)
Restoring Rhythm: Punctuation Restoration Using Transformer Models for Bangla, A Low-Resource Language
von: Mamun, Md Obyedullahil, et al.
Veröffentlicht: (2025)
von: Mamun, Md Obyedullahil, et al.
Veröffentlicht: (2025)
FS-DAG: Few Shot Domain Adapting Graph Networks for Visually Rich Document Understanding
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2024)
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2024)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey
von: Vegner, Ivan, et al.
Veröffentlicht: (2025)
von: Vegner, Ivan, et al.
Veröffentlicht: (2025)
More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
Extreme Self-Preference in Language Models
von: Lehr, Steven A., et al.
Veröffentlicht: (2025)
von: Lehr, Steven A., et al.
Veröffentlicht: (2025)
Do Schwartz Higher-Order Values Help Sentence-Level Human Value Detection? A Study of Hierarchical Gating and Calibration
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
SpecExtend: A Drop-in Enhancement for Speculative Decoding of Long Sequences
von: Cha, Jungyoub, et al.
Veröffentlicht: (2025)
von: Cha, Jungyoub, et al.
Veröffentlicht: (2025)
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-Supervised Learning of Time Series Representation via Diffusion Process and Imputation-Interpolation-Forecasting Mask
von: Senane, Zineb, et al.
Veröffentlicht: (2024) -
Prompt Tuned Embedding Classification for Multi-Label Industry Sector Allocation
von: Buchner, Valentin Leonhard, et al.
Veröffentlicht: (2023) -
Unveiling Transformer Perception by Exploring Input Manifolds
von: Benfenati, Alessandro, et al.
Veröffentlicht: (2024) -
Sycophancy as compositions of Atomic Psychometric Traits
von: Jain, Shreyans, et al.
Veröffentlicht: (2025) -
Generative AI for Strategic Plan Development
von: Ponnock, Jesse
Veröffentlicht: (2025)