Guardado en:
| Autor principal: | Lee, Jae Joong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.05366 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Language-Guided Invariance Probing of Vision-Language Models
por: Lee, Jae Joong
Publicado: (2025)
por: Lee, Jae Joong
Publicado: (2025)
Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization
por: Jia, Chenwei, et al.
Publicado: (2026)
por: Jia, Chenwei, et al.
Publicado: (2026)
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
por: Zhang, Jiaji, et al.
Publicado: (2025)
por: Zhang, Jiaji, et al.
Publicado: (2025)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
por: Lee, Dongyeun, et al.
Publicado: (2025)
por: Lee, Dongyeun, et al.
Publicado: (2025)
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation
por: Liu, Xuewen, et al.
Publicado: (2024)
por: Liu, Xuewen, et al.
Publicado: (2024)
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approximation
por: Kim, Suyoung, et al.
Publicado: (2026)
por: Kim, Suyoung, et al.
Publicado: (2026)
ProtoQuant: Quantization of Prototypical Parts For General and Fine-Grained Image Classification
por: Janusz, Mikołaj, et al.
Publicado: (2026)
por: Janusz, Mikołaj, et al.
Publicado: (2026)
DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization
por: Lin, Haokun, et al.
Publicado: (2026)
por: Lin, Haokun, et al.
Publicado: (2026)
CacheQuant: Comprehensively Accelerated Diffusion Models
por: Liu, Xuewen, et al.
Publicado: (2025)
por: Liu, Xuewen, et al.
Publicado: (2025)
PlaneCycle: Training-Free 2D-to-3D Lifting of Foundation Models Without Adapters
por: Yu, Yinghong, et al.
Publicado: (2026)
por: Yu, Yinghong, et al.
Publicado: (2026)
Q-HyViT: Post-Training Quantization of Hybrid Vision Transformers with Bridge Block Reconstruction for IoT Systems
por: Lee, Jemin, et al.
Publicado: (2023)
por: Lee, Jemin, et al.
Publicado: (2023)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
por: Bieri, Valentin, et al.
Publicado: (2025)
por: Bieri, Valentin, et al.
Publicado: (2025)
VGGT-CD: Training-Free Robust Registration for 3D Change Detection
por: Zhang, Wei, et al.
Publicado: (2026)
por: Zhang, Wei, et al.
Publicado: (2026)
Post-Training Quantization for 3D Medical Image Segmentation: A Practical Study on Real Inference Engines
por: Qu, Chongyu, et al.
Publicado: (2025)
por: Qu, Chongyu, et al.
Publicado: (2025)
WebAccessVL: Violation-Aware VLM for Web Accessibility
por: Zheng, Amber Yijia, et al.
Publicado: (2025)
por: Zheng, Amber Yijia, et al.
Publicado: (2025)
Trio-ViT: Post-Training Quantization and Acceleration for Softmax-Free Efficient Vision Transformer
por: Shi, Huihong, et al.
Publicado: (2024)
por: Shi, Huihong, et al.
Publicado: (2024)
Progressive Fine-to-Coarse Reconstruction for Accurate Low-Bit Post-Training Quantization in Vision Transformers
por: Ding, Rui, et al.
Publicado: (2024)
por: Ding, Rui, et al.
Publicado: (2024)
Speed3R: Sparse Feed-forward 3D Reconstruction Models
por: Ren, Weining, et al.
Publicado: (2026)
por: Ren, Weining, et al.
Publicado: (2026)
GTR: Improving Large 3D Reconstruction Models through Geometry and Texture Refinement
por: Zhuang, Peiye, et al.
Publicado: (2024)
por: Zhuang, Peiye, et al.
Publicado: (2024)
TeHOR: Text-Guided 3D Human and Object Reconstruction with Textures
por: Nam, Hyeongjin, et al.
Publicado: (2026)
por: Nam, Hyeongjin, et al.
Publicado: (2026)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
por: Jeon, Yerim, et al.
Publicado: (2025)
por: Jeon, Yerim, et al.
Publicado: (2025)
OuroMamba: A Data-Free Quantization Framework for Vision Mamba
por: Ramachandran, Akshat, et al.
Publicado: (2025)
por: Ramachandran, Akshat, et al.
Publicado: (2025)
PhysQuantAgent: An Inference Pipeline of Mass Estimation for Vision-Language Models
por: Yokomizo, Hisayuki, et al.
Publicado: (2026)
por: Yokomizo, Hisayuki, et al.
Publicado: (2026)
Cross-scale Aligned Supervision for Training GANs
por: Hyun, Sangeek, et al.
Publicado: (2026)
por: Hyun, Sangeek, et al.
Publicado: (2026)
RGB2Point: 3D Point Cloud Generation from Single RGB Images
por: Lee, Jae Joong, et al.
Publicado: (2024)
por: Lee, Jae Joong, et al.
Publicado: (2024)
D3: Training-Free AI-Generated Video Detection Using Second-Order Features
por: Zheng, Chende, et al.
Publicado: (2025)
por: Zheng, Chende, et al.
Publicado: (2025)
FreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstruction
por: Cao, Wei, et al.
Publicado: (2026)
por: Cao, Wei, et al.
Publicado: (2026)
Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients
por: Xiang, Ziwei, et al.
Publicado: (2026)
por: Xiang, Ziwei, et al.
Publicado: (2026)
MedPruner: Training-Free Hierarchical Token Pruning for Efficient 3D Medical Image Understanding in Vision-Language Models
por: Liu, Shengyuan, et al.
Publicado: (2026)
por: Liu, Shengyuan, et al.
Publicado: (2026)
Occlusion-Aware Temporally Consistent Amodal Completion for 3D Human-Object Interaction Reconstruction
por: Doh, Hyungjun, et al.
Publicado: (2025)
por: Doh, Hyungjun, et al.
Publicado: (2025)
S3D: Sketch-Driven 3D Model Generation
por: Song, Hail, et al.
Publicado: (2025)
por: Song, Hail, et al.
Publicado: (2025)
Post-Training Quantization for Video Matting
por: Zhu, Tianrui, et al.
Publicado: (2025)
por: Zhu, Tianrui, et al.
Publicado: (2025)
Image-Conditioned 3D Gaussian Splat Quantization
por: Liu, Xinshuang, et al.
Publicado: (2025)
por: Liu, Xinshuang, et al.
Publicado: (2025)
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
por: Ju, Yeong-Joon, et al.
Publicado: (2024)
por: Ju, Yeong-Joon, et al.
Publicado: (2024)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
por: Kim, Gihwan, et al.
Publicado: (2025)
por: Kim, Gihwan, et al.
Publicado: (2025)
Training-Free Reward-Guided Image Editing via Trajectory Optimal Control
por: Chang, Jinho, et al.
Publicado: (2025)
por: Chang, Jinho, et al.
Publicado: (2025)
ClipTBP: Clip-Pair based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval
por: Kim, Ji-Hyeon, et al.
Publicado: (2026)
por: Kim, Ji-Hyeon, et al.
Publicado: (2026)
FIQ: Fundamental Question Generation with the Integration of Question Embeddings for Video Question Answering
por: Oh, Ju-Young, et al.
Publicado: (2025)
por: Oh, Ju-Young, et al.
Publicado: (2025)
FreeAct: Freeing Activations for LLM Quantization
por: Liu, Xiaohao, et al.
Publicado: (2026)
por: Liu, Xiaohao, et al.
Publicado: (2026)
MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization
por: Su, Le, et al.
Publicado: (2026)
por: Su, Le, et al.
Publicado: (2026)
Ejemplares similares
-
Language-Guided Invariance Probing of Vision-Language Models
por: Lee, Jae Joong
Publicado: (2025) -
Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization
por: Jia, Chenwei, et al.
Publicado: (2026) -
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
por: Zhang, Jiaji, et al.
Publicado: (2025) -
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
por: Lee, Dongyeun, et al.
Publicado: (2025) -
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation
por: Liu, Xuewen, et al.
Publicado: (2024)