Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Junxiao, Dudley, John, Kristensson, Per Ola |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Prompt-Driven Agentic Video Editing System: Autonomous Comprehension of Long-Form, Story-Driven Media
by: Ding, Zihan, et al.
Published: (2025)
by: Ding, Zihan, et al.
Published: (2025)
Large Language Model-assisted Speech and Pointing Benefits Multiple 3D Object Selection in Virtual Reality
by: Chen, Junlong, et al.
Published: (2024)
by: Chen, Junlong, et al.
Published: (2024)
Modeling Subjective Urban Perception with Human Gaze
by: Che, Lin, et al.
Published: (2026)
by: Che, Lin, et al.
Published: (2026)
Generative AI for Accessible and Inclusive Extended Reality
by: Grubert, Jens, et al.
Published: (2024)
by: Grubert, Jens, et al.
Published: (2024)
Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes
by: Chen, Junlong, et al.
Published: (2024)
by: Chen, Junlong, et al.
Published: (2024)
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
Egocentric Co-Pilot: Web-Native Smart-Glasses Agents for Assistive Egocentric AI
by: Yang, Sicheng, et al.
Published: (2026)
by: Yang, Sicheng, et al.
Published: (2026)
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
by: Natalie, Rosiana, et al.
Published: (2025)
by: Natalie, Rosiana, et al.
Published: (2025)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
by: Yakolli, Nivedan, et al.
Published: (2025)
by: Yakolli, Nivedan, et al.
Published: (2025)
Words into World: A Task-Adaptive Agent for Language-Guided Spatial Retrieval in AR
by: Guo, Lixing, et al.
Published: (2025)
by: Guo, Lixing, et al.
Published: (2025)
Towards a Humanized Social-Media Ecosystem: AI-Augmented HCI Design Patterns for Safety, Agency & Well-Being
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
Do Vision Language Models Understand Human Engagement in Games?
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
Proactive Assistant Dialogue Generation from Streaming Egocentric Videos
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
IndEgo: A Dataset of Industrial Scenarios and Collaborative Work for Egocentric Assistants
by: Chavan, Vivek, et al.
Published: (2025)
by: Chavan, Vivek, et al.
Published: (2025)
Photoreal Scene Reconstruction from an Egocentric Device
by: Lv, Zhaoyang, et al.
Published: (2025)
by: Lv, Zhaoyang, et al.
Published: (2025)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
by: Chaudhari, Shravan, et al.
Published: (2025)
by: Chaudhari, Shravan, et al.
Published: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
by: Song, Jookyung, et al.
Published: (2025)
by: Song, Jookyung, et al.
Published: (2025)
Constructive Apraxia: An Unexpected Limit of Instructible Vision-Language Models and Analog for Human Cognitive Disorders
by: Noever, David, et al.
Published: (2024)
by: Noever, David, et al.
Published: (2024)
MP-GUI: Modality Perception with MLLMs for GUI Understanding
by: Wang, Ziwei, et al.
Published: (2025)
by: Wang, Ziwei, et al.
Published: (2025)
Generative Augmented Reality: Paradigms, Technologies, and Future Applications
by: Liang, Chen, et al.
Published: (2025)
by: Liang, Chen, et al.
Published: (2025)
SASG-DA: Sparse-Aware Semantic-Guided Diffusion Augmentation For Myoelectric Gesture Recognition
by: Liu, Chen, et al.
Published: (2025)
by: Liu, Chen, et al.
Published: (2025)
HuLP: Human-in-the-Loop for Prognosis
by: Ridzuan, Muhammad, et al.
Published: (2024)
by: Ridzuan, Muhammad, et al.
Published: (2024)
MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning
by: Bhattarai, Ankit, et al.
Published: (2026)
by: Bhattarai, Ankit, et al.
Published: (2026)
Benchmarking XAI Explanations with Human-Aligned Evaluations
by: Kazmierczak, Rémi, et al.
Published: (2024)
by: Kazmierczak, Rémi, et al.
Published: (2024)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025)
by: Rekimoto, Jun
Published: (2025)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
by: Zhang, Ye, et al.
Published: (2026)
by: Zhang, Ye, et al.
Published: (2026)
Learning User Embeddings from Human Gaze for Personalised Saliency Prediction
by: Strohm, Florian, et al.
Published: (2024)
by: Strohm, Florian, et al.
Published: (2024)
AI Guide Dog: Egocentric Path Prediction on Smartphone
by: Jadhav, Aishwarya, et al.
Published: (2025)
by: Jadhav, Aishwarya, et al.
Published: (2025)
In-Depth Analysis of Emotion Recognition through Knowledge-Based Large Language Models
by: Han, Bin, et al.
Published: (2024)
by: Han, Bin, et al.
Published: (2024)
ThermoHands: A Benchmark for 3D Hand Pose Estimation from Egocentric Thermal Images
by: Ding, Fangqiang, et al.
Published: (2024)
by: Ding, Fangqiang, et al.
Published: (2024)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
by: Luthra, Prerna
Published: (2025)
by: Luthra, Prerna
Published: (2025)
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
by: Kim, Namhee, et al.
Published: (2025)
by: Kim, Namhee, et al.
Published: (2025)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026)
by: Baek, Injun, et al.
Published: (2026)
Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition
by: Choi, Jae Young, et al.
Published: (2026)
by: Choi, Jae Young, et al.
Published: (2026)
Deep Neural Networks and Brain Alignment: Brain Encoding and Decoding (Survey)
by: Oota, Subba Reddy, et al.
Published: (2023)
by: Oota, Subba Reddy, et al.
Published: (2023)
GAITEX: Human motion dataset of impaired gait and rehabilitation exercises using inertial and optical sensors
by: Spilz, Andreas, et al.
Published: (2025)
by: Spilz, Andreas, et al.
Published: (2025)
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning
by: Chaubey, Ashutosh, et al.
Published: (2025)
by: Chaubey, Ashutosh, et al.
Published: (2025)
Similar Items
-
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025) -
Prompt-Driven Agentic Video Editing System: Autonomous Comprehension of Long-Form, Story-Driven Media
by: Ding, Zihan, et al.
Published: (2025) -
Large Language Model-assisted Speech and Pointing Benefits Multiple 3D Object Selection in Virtual Reality
by: Chen, Junlong, et al.
Published: (2024) -
Modeling Subjective Urban Perception with Human Gaze
by: Che, Lin, et al.
Published: (2026) -
Generative AI for Accessible and Inclusive Extended Reality
by: Grubert, Jens, et al.
Published: (2024)