Whispering to a Blackbox: Bootstrapping Frozen OCR with Visual Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Samandarov, Samandar, Ismoiljonov, Nazirjon, Sattorov, Abdullah, Sabyrbayev, Temirlan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Geometry Preserving Loss Functions Promote Improved Adaptation of Blackbox Generative Model
von: Mitra, Sinjini, et al.
Veröffentlicht: (2026)
von: Mitra, Sinjini, et al.
Veröffentlicht: (2026)
Steering Frozen LLMs: Adaptive Social Alignment via Online Prompt Routing
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
XSub: Explanation-Driven Adversarial Attack against Blackbox Classifiers via Feature Substitution
von: Vu, Kiana, et al.
Veröffentlicht: (2024)
von: Vu, Kiana, et al.
Veröffentlicht: (2024)
Sysformer: Safeguarding Frozen Large Language Models with Adaptive System Prompts
von: Sharma, Kartik, et al.
Veröffentlicht: (2025)
von: Sharma, Kartik, et al.
Veröffentlicht: (2025)
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
SplitFrozen: Split Learning with Device-side Model Frozen for Fine-Tuning LLM on Heterogeneous Resource-Constrained Devices
von: Ma, Jian, et al.
Veröffentlicht: (2025)
von: Ma, Jian, et al.
Veröffentlicht: (2025)
PostMark: A Robust Blackbox Watermark for Large Language Models
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
Bootstrapping Expectiles in Reinforcement Learning
von: Clavier, Pierre, et al.
Veröffentlicht: (2024)
von: Clavier, Pierre, et al.
Veröffentlicht: (2024)
Imitation Bootstrapped Reinforcement Learning
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2025)
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2025)
Trained Persistent Memory for Frozen Decoder-Only LLMs
von: Jeong, Hong
Veröffentlicht: (2026)
von: Jeong, Hong
Veröffentlicht: (2026)
Bootstrap SGD: Algorithmic Stability and Robustness
von: Christmann, Andreas, et al.
Veröffentlicht: (2024)
von: Christmann, Andreas, et al.
Veröffentlicht: (2024)
Frozen Layers: Memory-efficient Many-fidelity Hyperparameter Optimization
von: Carstensen, Timur, et al.
Veröffentlicht: (2025)
von: Carstensen, Timur, et al.
Veröffentlicht: (2025)
From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales
von: Viakhirev, Ivan, et al.
Veröffentlicht: (2026)
von: Viakhirev, Ivan, et al.
Veröffentlicht: (2026)
Learning to Rewrite Prompts for Bootstrapping LLMs on Downstream Tasks
von: Zhou, Qinhao, et al.
Veröffentlicht: (2025)
von: Zhou, Qinhao, et al.
Veröffentlicht: (2025)
WhisperD: Dementia Speech Recognition and Filler Word Detection with Whisper
von: Akinrintoyo, Emmanuel, et al.
Veröffentlicht: (2025)
von: Akinrintoyo, Emmanuel, et al.
Veröffentlicht: (2025)
Value of Information-Enhanced Exploration in Bootstrapped DQN
von: Plataniotis, Stergios, et al.
Veröffentlicht: (2025)
von: Plataniotis, Stergios, et al.
Veröffentlicht: (2025)
Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Models
von: Armstrong, Marcus, et al.
Veröffentlicht: (2026)
von: Armstrong, Marcus, et al.
Veröffentlicht: (2026)
Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies
von: Zhao, Guangyu, et al.
Veröffentlicht: (2024)
von: Zhao, Guangyu, et al.
Veröffentlicht: (2024)
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
von: Otsuka, Hikari, et al.
Veröffentlicht: (2024)
von: Otsuka, Hikari, et al.
Veröffentlicht: (2024)
Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation
von: Hariri, Mohsen, et al.
Veröffentlicht: (2025)
von: Hariri, Mohsen, et al.
Veröffentlicht: (2025)
AgentOCR: Reimagining Agent History via Optical Self-Compression
von: Feng, Lang, et al.
Veröffentlicht: (2026)
von: Feng, Lang, et al.
Veröffentlicht: (2026)
Bootstrapped Model Predictive Control
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
Flattening Hierarchies with Policy Bootstrapping
von: Zhou, John L., et al.
Veröffentlicht: (2025)
von: Zhou, John L., et al.
Veröffentlicht: (2025)
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
von: Meng, Li, et al.
Veröffentlicht: (2022)
von: Meng, Li, et al.
Veröffentlicht: (2022)
Bootstrapping LLMs via Preference-Based Policy Optimization
von: Jia, Chen
Veröffentlicht: (2025)
von: Jia, Chen
Veröffentlicht: (2025)
Trained Persistent Memory for Frozen Encoder--Decoder LLMs: Six Architectural Methods
von: Jeong, Hong
Veröffentlicht: (2026)
von: Jeong, Hong
Veröffentlicht: (2026)
Concepts Whisper While Syntax Shouts: Spectral Anti-Concentration and the Dual Geometry of Transformer Representations
von: Acharya, Pratyush, et al.
Veröffentlicht: (2026)
von: Acharya, Pratyush, et al.
Veröffentlicht: (2026)
TopoReformer: Mitigating Adversarial Attacks Using Topological Purification in OCR Models
von: Kumar, Bhagyesh, et al.
Veröffentlicht: (2025)
von: Kumar, Bhagyesh, et al.
Veröffentlicht: (2025)
Bootstrap Off-policy with World Model
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
Beyond Experience Retrieval: Learning to Generate Utility-Optimized Structured Experience for Frozen LLMs
von: Li, Xuancheng, et al.
Veröffentlicht: (2026)
von: Li, Xuancheng, et al.
Veröffentlicht: (2026)
LANPO: Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
von: Li, Ang, et al.
Veröffentlicht: (2025)
von: Li, Ang, et al.
Veröffentlicht: (2025)
When Denoising Hinders: Revisiting Zero-Shot ASR with SAM-Audio and Whisper
von: Islam, Akif, et al.
Veröffentlicht: (2026)
von: Islam, Akif, et al.
Veröffentlicht: (2026)
Bootstrapping your behavior: a new pretraining strategy for user behavior sequence data
von: Wu, Weichang, et al.
Veröffentlicht: (2025)
von: Wu, Weichang, et al.
Veröffentlicht: (2025)
Affordance-Guided Reinforcement Learning via Visual Prompting
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
Smart Sampling: Self-Attention and Bootstrapping for Improved Ensembled Q-Learning
von: Khan, Muhammad Junaid, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Junaid, et al.
Veröffentlicht: (2024)
Frozen Transformers in Language Models Are Effective Visual Encoder Layers
von: Pang, Ziqi, et al.
Veröffentlicht: (2023)
von: Pang, Ziqi, et al.
Veröffentlicht: (2023)
Bootstrapped Mixed Rewards for RL Post-Training: Injecting Canonical Action Order
von: Gupta, Prakhar, et al.
Veröffentlicht: (2025)
von: Gupta, Prakhar, et al.
Veröffentlicht: (2025)
Learning for Interval Prediction of Electricity Demand: A Cluster-based Bootstrapping Approach
von: Dube, Rohit, et al.
Veröffentlicht: (2023)
von: Dube, Rohit, et al.
Veröffentlicht: (2023)
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
von: Che, Fengdi, et al.
Veröffentlicht: (2024)
von: Che, Fengdi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Geometry Preserving Loss Functions Promote Improved Adaptation of Blackbox Generative Model
von: Mitra, Sinjini, et al.
Veröffentlicht: (2026) -
Steering Frozen LLMs: Adaptive Social Alignment via Online Prompt Routing
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026) -
XSub: Explanation-Driven Adversarial Attack against Blackbox Classifiers via Feature Substitution
von: Vu, Kiana, et al.
Veröffentlicht: (2024) -
Sysformer: Safeguarding Frozen Large Language Models with Adaptive System Prompts
von: Sharma, Kartik, et al.
Veröffentlicht: (2025) -
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)