Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning
Fuente:
arXiv
Saved in:
| Main Author: | Mandalika, Sriram |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SegXAL: Explainable Active Learning for Semantic Segmentation in Driving Scene Scenarios
by: Mandalika, Sriram, et al.
Published: (2024)
by: Mandalika, Sriram, et al.
Published: (2024)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
by: Ray, Arijit, et al.
Published: (2024)
by: Ray, Arijit, et al.
Published: (2024)
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
by: Pang, Bo, et al.
Published: (2025)
by: Pang, Bo, et al.
Published: (2025)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
by: Wei, Shenxing, et al.
Published: (2025)
by: Wei, Shenxing, et al.
Published: (2025)
Few-Shot Multi-Human Neural Rendering Using Geometry Constraints
by: li, Qian, et al.
Published: (2025)
by: li, Qian, et al.
Published: (2025)
Physics-Based Motion Imitation with Adversarial Differential Discriminators
by: Zhang, Ziyu, et al.
Published: (2025)
by: Zhang, Ziyu, et al.
Published: (2025)
Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
by: Dong, Zhao, et al.
Published: (2025)
by: Dong, Zhao, et al.
Published: (2025)
SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control
by: Mu, Yuxuan, et al.
Published: (2025)
by: Mu, Yuxuan, et al.
Published: (2025)
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
by: Yang, Jianing, et al.
Published: (2025)
by: Yang, Jianing, et al.
Published: (2025)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Free-Moving Object Reconstruction and Pose Estimation with Virtual Camera
by: Shi, Haixin, et al.
Published: (2024)
by: Shi, Haixin, et al.
Published: (2024)
SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes
by: Pfaff, Nicholas, et al.
Published: (2026)
by: Pfaff, Nicholas, et al.
Published: (2026)
Neural Implicit Representation for Building Digital Twins of Unknown Articulated Objects
by: Weng, Yijia, et al.
Published: (2024)
by: Weng, Yijia, et al.
Published: (2024)
AddBiomechanics Dataset: Capturing the Physics of Human Motion at Scale
by: Werling, Keenon, et al.
Published: (2024)
by: Werling, Keenon, et al.
Published: (2024)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
by: Shah, Ansh, et al.
Published: (2024)
by: Shah, Ansh, et al.
Published: (2024)
ImDy: Human Inverse Dynamics from Imitated Observations
by: Liu, Xinpeng, et al.
Published: (2024)
by: Liu, Xinpeng, et al.
Published: (2024)
Leveraging Foundation Models To learn the shape of semi-fluid deformable objects
by: Assal, Omar El, et al.
Published: (2024)
by: Assal, Omar El, et al.
Published: (2024)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
by: Ye, Weicai, et al.
Published: (2024)
by: Ye, Weicai, et al.
Published: (2024)
Few-Shot Unsupervised Implicit Neural Shape Representation Learning with Spatial Adversaries
by: Ouasfi, Amine, et al.
Published: (2024)
by: Ouasfi, Amine, et al.
Published: (2024)
RigidFormer: Learning Rigid Dynamics using Transformers
by: Dou, Zhiyang, et al.
Published: (2026)
by: Dou, Zhiyang, et al.
Published: (2026)
Learning 3D-Gaussian Simulators from RGB Videos
by: Zhobro, Mikel, et al.
Published: (2025)
by: Zhobro, Mikel, et al.
Published: (2025)
CoMAD: A Multiple-Teacher Self-Supervised Distillation Framework
by: Mandalika, Sriram, et al.
Published: (2025)
by: Mandalika, Sriram, et al.
Published: (2025)
SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
MaskAdapt: Learning Flexible Motion Adaptation via Mask-Invariant Prior for Physics-Based Characters
by: Park, Soomin, et al.
Published: (2026)
by: Park, Soomin, et al.
Published: (2026)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
by: Feng, Yuming, et al.
Published: (2024)
by: Feng, Yuming, et al.
Published: (2024)
One-Shot Real-to-Sim via End-to-End Differentiable Simulation and Rendering
by: Zhu, Yifan, et al.
Published: (2024)
by: Zhu, Yifan, et al.
Published: (2024)
Structure from Collision
by: Kaneko, Takuhiro
Published: (2025)
by: Kaneko, Takuhiro
Published: (2025)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
by: Zheng, Shuhong, et al.
Published: (2025)
by: Zheng, Shuhong, et al.
Published: (2025)
IFG: Internet-Scale Guidance for Functional Grasping Generation
by: Liu, Ray Muxin, et al.
Published: (2025)
by: Liu, Ray Muxin, et al.
Published: (2025)
SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control
by: Luo, Zhengyi, et al.
Published: (2025)
by: Luo, Zhengyi, et al.
Published: (2025)
GENMO: A GENeralist Model for Human MOtion
by: Li, Jiefeng, et al.
Published: (2025)
by: Li, Jiefeng, et al.
Published: (2025)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
WHAC: World-grounded Humans and Cameras
by: Yin, Wanqi, et al.
Published: (2024)
by: Yin, Wanqi, et al.
Published: (2024)
Improving Physics-Augmented Continuum Neural Radiance Field-Based Geometry-Agnostic System Identification with Lagrangian Particle Optimization
by: Kaneko, Takuhiro
Published: (2024)
by: Kaneko, Takuhiro
Published: (2024)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
by: Seitzer, Maximilian, et al.
Published: (2023)
by: Seitzer, Maximilian, et al.
Published: (2023)
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026)
by: Liu, Shaowei, et al.
Published: (2026)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
by: Lou, Yuke, et al.
Published: (2025)
by: Lou, Yuke, et al.
Published: (2025)
Few-shot Novel View Synthesis using Depth Aware 3D Gaussian Splatting
by: Kumar, Raja, et al.
Published: (2024)
by: Kumar, Raja, et al.
Published: (2024)
VOODOO XP: Expressive One-Shot Head Reenactment for VR Telepresence
by: Tran, Phong, et al.
Published: (2024)
by: Tran, Phong, et al.
Published: (2024)
Similar Items
-
SegXAL: Explainable Active Learning for Semantic Segmentation in Driving Scene Scenarios
by: Mandalika, Sriram, et al.
Published: (2024) -
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
by: Ray, Arijit, et al.
Published: (2024) -
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
by: Pang, Bo, et al.
Published: (2025) -
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
by: Wei, Shenxing, et al.
Published: (2025) -
Few-Shot Multi-Human Neural Rendering Using Geometry Constraints
by: li, Qian, et al.
Published: (2025)