SGG-R$^{\rm 3}$: From Next-Token Prediction to End-to-End Unbiased Scene Graph Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Feng, Jiaye, Yin, Qixiang, Liu, Yuankun, Mo, Tong, Li, Weiping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Salience-SGG: Enhancing Unbiased Scene Graph Generation with Iterative Salience Estimation
por: Qu, Runfeng, et al.
Publicado: (2026)
por: Qu, Runfeng, et al.
Publicado: (2026)
CAGE-SGG: Counterfactual Active Graph Evidence for Open-Vocabulary Scene Graph Generation
por: Guang, Suiyang, et al.
Publicado: (2026)
por: Guang, Suiyang, et al.
Publicado: (2026)
Ensemble Predicate Decoding for Unbiased Scene Graph Generation
por: Feng, Jiasong, et al.
Publicado: (2024)
por: Feng, Jiasong, et al.
Publicado: (2024)
Hydra-SGG: Hybrid Relation Assignment for One-stage Scene Graph Generation
por: Chen, Minghan, et al.
Publicado: (2024)
por: Chen, Minghan, et al.
Publicado: (2024)
HiKER-SGG: Hierarchical Knowledge Enhanced Robust Scene Graph Generation
por: Zhang, Ce, et al.
Publicado: (2024)
por: Zhang, Ce, et al.
Publicado: (2024)
OED: Towards One-stage End-to-End Dynamic Scene Graph Generation
por: Wang, Guan, et al.
Publicado: (2024)
por: Wang, Guan, et al.
Publicado: (2024)
ReLIC-SGG: Relation Lattice Completion for Open-Vocabulary Scene Graph Generation
por: Hosseini, Amir, et al.
Publicado: (2026)
por: Hosseini, Amir, et al.
Publicado: (2026)
VOST-SGG: VLM-Aided One-Stage Spatio-Temporal Scene Graph Generation
por: Sugandhika, Chinthani, et al.
Publicado: (2025)
por: Sugandhika, Chinthani, et al.
Publicado: (2025)
Generalized Unbiased Scene Graph Generation
por: Lyu, Xinyu, et al.
Publicado: (2023)
por: Lyu, Xinyu, et al.
Publicado: (2023)
RA-SGG: Retrieval-Augmented Scene Graph Generation Framework via Multi-Prototype Learning
por: Yoon, Kanghoon, et al.
Publicado: (2024)
por: Yoon, Kanghoon, et al.
Publicado: (2024)
End-to-End Vision Tokenizer Tuning
por: Wang, Wenxuan, et al.
Publicado: (2025)
por: Wang, Wenxuan, et al.
Publicado: (2025)
GPT4SGG: Synthesizing Scene Graphs from Holistic and Region-specific Narratives
por: Chen, Zuyao, et al.
Publicado: (2023)
por: Chen, Zuyao, et al.
Publicado: (2023)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
por: Peng, Zhenghao, et al.
Publicado: (2025)
por: Peng, Zhenghao, et al.
Publicado: (2025)
Temporally Consistent Dynamic Scene Graphs: An End-to-End Approach for Action Tracklet Generation
por: Ruschel, Raphael, et al.
Publicado: (2024)
por: Ruschel, Raphael, et al.
Publicado: (2024)
SGTR+: End-to-end Scene Graph Generation with Transformer
por: Li, Rongjie, et al.
Publicado: (2024)
por: Li, Rongjie, et al.
Publicado: (2024)
Compositional Feature Augmentation for Unbiased Scene Graph Generation
por: Li, Lin, et al.
Publicado: (2023)
por: Li, Lin, et al.
Publicado: (2023)
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
por: Xu, Mingjie, et al.
Publicado: (2024)
por: Xu, Mingjie, et al.
Publicado: (2024)
DSGG: Dense Relation Transformer for an End-to-end Scene Graph Generation
por: Hayder, Zeeshan, et al.
Publicado: (2024)
por: Hayder, Zeeshan, et al.
Publicado: (2024)
LLM4SGG: Large Language Models for Weakly Supervised Scene Graph Generation
por: Kim, Kibum, et al.
Publicado: (2023)
por: Kim, Kibum, et al.
Publicado: (2023)
Decoupling Scene Perception and Ego Status: A Multi-Context Fusion Approach for Enhanced Generalization in End-to-End Autonomous Driving
por: Tang, Jiacheng, et al.
Publicado: (2025)
por: Tang, Jiacheng, et al.
Publicado: (2025)
Robo-SGG: Exploiting Layout-Oriented Normalization and Restitution Can Improve Robust Scene Graph Generation
por: Lv, Changsheng, et al.
Publicado: (2025)
por: Lv, Changsheng, et al.
Publicado: (2025)
SceneLCM: End-to-End Layout-Guided Interactive Indoor Scene Generation with Latent Consistency Model
por: Lin, Yangkai, et al.
Publicado: (2025)
por: Lin, Yangkai, et al.
Publicado: (2025)
Generating Multimodal Driving Scenes via Next-Scene Prediction
por: Wu, Yanhao, et al.
Publicado: (2025)
por: Wu, Yanhao, et al.
Publicado: (2025)
Rethinking End-to-End 2D to 3D Scene Segmentation in Gaussian Splatting
por: Zhu, Runsong, et al.
Publicado: (2025)
por: Zhu, Runsong, et al.
Publicado: (2025)
GraphAD: Interaction Scene Graph for End-to-end Autonomous Driving
por: Zhang, Yunpeng, et al.
Publicado: (2024)
por: Zhang, Yunpeng, et al.
Publicado: (2024)
End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer
por: Chu, Wenda, et al.
Publicado: (2026)
por: Chu, Wenda, et al.
Publicado: (2026)
Towards Unbiased and Robust Spatio-Temporal Scene Graph Generation and Anticipation
por: Peddi, Rohith, et al.
Publicado: (2024)
por: Peddi, Rohith, et al.
Publicado: (2024)
World Model-Based End-to-End Scene Generation for Accident Anticipation in Autonomous Driving
por: Guan, Yanchen, et al.
Publicado: (2025)
por: Guan, Yanchen, et al.
Publicado: (2025)
Unbiased Scene Graph Generation from Biased Training
por: Tang, Kaihua, et al.
Publicado: (2020)
por: Tang, Kaihua, et al.
Publicado: (2020)
DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation
por: Li, Haoran, et al.
Publicado: (2025)
por: Li, Haoran, et al.
Publicado: (2025)
InsightDrive: Insight Scene Representation for End-to-End Autonomous Driving
por: Song, Ruiqi, et al.
Publicado: (2025)
por: Song, Ruiqi, et al.
Publicado: (2025)
Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving
por: Li, Peidong, et al.
Publicado: (2024)
por: Li, Peidong, et al.
Publicado: (2024)
Active Learning from Scene Embeddings for End-to-End Autonomous Driving
por: Jiang, Wenhao, et al.
Publicado: (2025)
por: Jiang, Wenhao, et al.
Publicado: (2025)
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
por: Xiong, Haomiao, et al.
Publicado: (2025)
por: Xiong, Haomiao, et al.
Publicado: (2025)
Unbiased Scene Graph Generation by Type-Aware Message Passing on Heterogeneous and Dual Graphs
por: Sun, Guanglu, et al.
Publicado: (2024)
por: Sun, Guanglu, et al.
Publicado: (2024)
GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving
por: Xing, Zebin, et al.
Publicado: (2025)
por: Xing, Zebin, et al.
Publicado: (2025)
VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs
por: Zhu, Jiaying, et al.
Publicado: (2025)
por: Zhu, Jiaying, et al.
Publicado: (2025)
From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage
por: Ruan, Cihan, et al.
Publicado: (2026)
por: Ruan, Cihan, et al.
Publicado: (2026)
SparseDrive: End-to-End Autonomous Driving via Sparse Scene Representation
por: Sun, Wenchao, et al.
Publicado: (2024)
por: Sun, Wenchao, et al.
Publicado: (2024)
Polar R-CNN: End-to-End Lane Detection with Fewer Anchors
por: Wang, Shengqi, et al.
Publicado: (2024)
por: Wang, Shengqi, et al.
Publicado: (2024)
Ejemplares similares
-
Salience-SGG: Enhancing Unbiased Scene Graph Generation with Iterative Salience Estimation
por: Qu, Runfeng, et al.
Publicado: (2026) -
CAGE-SGG: Counterfactual Active Graph Evidence for Open-Vocabulary Scene Graph Generation
por: Guang, Suiyang, et al.
Publicado: (2026) -
Ensemble Predicate Decoding for Unbiased Scene Graph Generation
por: Feng, Jiasong, et al.
Publicado: (2024) -
Hydra-SGG: Hybrid Relation Assignment for One-stage Scene Graph Generation
por: Chen, Minghan, et al.
Publicado: (2024) -
HiKER-SGG: Hierarchical Knowledge Enhanced Robust Scene Graph Generation
por: Zhang, Ce, et al.
Publicado: (2024)