Modeling Real-Time Interactive Conversations as Timed Diarized Transcripts
Fuente:
arXiv
Saved in:
| Main Authors: | Tanzer, Garrett, Ahdritz, Gustaf, Melas-Kyriazi, Luke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The SMeL Test: A simple benchmark for media literacy in language models
by: Ahdritz, Gustaf, et al.
Published: (2025)
by: Ahdritz, Gustaf, et al.
Published: (2025)
A Benchmark for Learning to Translate a New Language from One Grammar Book
by: Tanzer, Garrett, et al.
Published: (2023)
by: Tanzer, Garrett, et al.
Published: (2023)
Fixed Point Diffusion Models
by: Bai, Xingjian, et al.
Published: (2024)
by: Bai, Xingjian, et al.
Published: (2024)
Distinguishing the Knowable from the Unknowable with Language Models
by: Ahdritz, Gustaf, et al.
Published: (2024)
by: Ahdritz, Gustaf, et al.
Published: (2024)
Scaling Sign Language Translation
by: Zhang, Biao, et al.
Published: (2024)
by: Zhang, Biao, et al.
Published: (2024)
FLEURS-ASL: Including American Sign Language in Massively Multilingual Multitask Evaluation
by: Tanzer, Garrett
Published: (2024)
by: Tanzer, Garrett
Published: (2024)
Fingerspelling within Sign Language Translation
by: Tanzer, Garrett
Published: (2024)
by: Tanzer, Garrett
Published: (2024)
Exploring Speech Foundation Models for Speaker Diarization in Child-Adult Dyadic Interactions
by: Xu, Anfeng, et al.
Published: (2024)
by: Xu, Anfeng, et al.
Published: (2024)
Provable Uncertainty Decomposition via Higher-Order Calibration
by: Ahdritz, Gustaf, et al.
Published: (2024)
by: Ahdritz, Gustaf, et al.
Published: (2024)
YouTube-SL-25: A Large-Scale, Open-Domain Multilingual Sign Language Parallel Corpus
by: Tanzer, Garrett, et al.
Published: (2024)
by: Tanzer, Garrett, et al.
Published: (2024)
3D Cardiac Anatomy Generation Using Mesh Latent Diffusion Models
by: Mozyrska, Jolanta, et al.
Published: (2025)
by: Mozyrska, Jolanta, et al.
Published: (2025)
Toward Conversational Agents with Context and Time Sensitive Long-term Memory
by: Alonso, Nick, et al.
Published: (2024)
by: Alonso, Nick, et al.
Published: (2024)
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
The Era of Real-World Human Interaction: RL from User Conversations
by: Jin, Chuanyang, et al.
Published: (2025)
by: Jin, Chuanyang, et al.
Published: (2025)
Domain-Aware Speaker Diarization On African-Accented English
by: Okocha, Chibuzor, et al.
Published: (2025)
by: Okocha, Chibuzor, et al.
Published: (2025)
Real-Time Trustworthiness Scoring for LLM Structured Outputs and Data Extraction
by: Goh, Hui Wen, et al.
Published: (2026)
by: Goh, Hui Wen, et al.
Published: (2026)
GES: Generalized Exponential Splatting for Efficient Radiance Field Rendering
by: Hamdi, Abdullah, et al.
Published: (2024)
by: Hamdi, Abdullah, et al.
Published: (2024)
Can Authorship Attribution Models Distinguish Speakers in Speech Transcripts?
by: Aggazzotti, Cristina, et al.
Published: (2023)
by: Aggazzotti, Cristina, et al.
Published: (2023)
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
by: Shin, Hagyeong, et al.
Published: (2025)
by: Shin, Hagyeong, et al.
Published: (2025)
CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge
by: Bae, Seyun, et al.
Published: (2026)
by: Bae, Seyun, et al.
Published: (2026)
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models
by: Liu, Yong, et al.
Published: (2024)
by: Liu, Yong, et al.
Published: (2024)
Treasure Hunt: Real-time Targeting of the Long Tail using Training-Time Markers
by: D'souza, Daniel, et al.
Published: (2025)
by: D'souza, Daniel, et al.
Published: (2025)
Lyrics Transcription for Humans: A Readability-Aware Benchmark
by: Cífka, Ondřej, et al.
Published: (2024)
by: Cífka, Ondřej, et al.
Published: (2024)
Graph Neural Flows for Unveiling Systemic Interactions Among Irregularly Sampled Time Series
by: Mercatali, Giangiacomo, et al.
Published: (2024)
by: Mercatali, Giangiacomo, et al.
Published: (2024)
ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
by: Wang, Chengsen, et al.
Published: (2024)
by: Wang, Chengsen, et al.
Published: (2024)
Test-Time Scaling with Reflective Generative Model
by: Wang, Zixiao, et al.
Published: (2025)
by: Wang, Zixiao, et al.
Published: (2025)
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
by: Kang, Haoqiang, et al.
Published: (2023)
by: Kang, Haoqiang, et al.
Published: (2023)
The Impact of Automatic Speech Transcription on Speaker Attribution
by: Aggazzotti, Cristina, et al.
Published: (2025)
by: Aggazzotti, Cristina, et al.
Published: (2025)
LiveThinking: Enabling Real-Time Efficient Reasoning for AI-Powered Livestreaming via Reinforcement Learning
by: Sun, Yuhan, et al.
Published: (2025)
by: Sun, Yuhan, et al.
Published: (2025)
Time Series Language Model for Descriptive Caption Generation
by: Trabelsi, Mohamed, et al.
Published: (2025)
by: Trabelsi, Mohamed, et al.
Published: (2025)
On the Power of (Approximate) Reward Models for Inference-Time Scaling
by: Zhu, Youheng, et al.
Published: (2026)
by: Zhu, Youheng, et al.
Published: (2026)
Reactive Transformer (RxT) -- Stateful Real-Time Processing for Event-Driven Reactive Language Models
by: Filipek, Adam
Published: (2025)
by: Filipek, Adam
Published: (2025)
Towards Detecting Contextual Real-Time Toxicity for In-Game Chat
by: Yang, Zachary, et al.
Published: (2023)
by: Yang, Zachary, et al.
Published: (2023)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
by: Obeso, Oscar, et al.
Published: (2025)
by: Obeso, Oscar, et al.
Published: (2025)
Time-MMD: Multi-Domain Multimodal Dataset for Time Series Analysis
by: Liu, Haoxin, et al.
Published: (2024)
by: Liu, Haoxin, et al.
Published: (2024)
Unifying Linear-Time Attention via Latent Probabilistic Modelling
by: Dolga, Rares, et al.
Published: (2024)
by: Dolga, Rares, et al.
Published: (2024)
Test-Time Training on Nearest Neighbors for Large Language Models
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
Bayesian Preference Learning for Test-Time Steerable Reward Models
by: Hong, Jiwoo, et al.
Published: (2026)
by: Hong, Jiwoo, et al.
Published: (2026)
HEART: Emotionally-Driven Test-Time Scaling of Language Models
by: Pinto, Gabriela, et al.
Published: (2025)
by: Pinto, Gabriela, et al.
Published: (2025)
The Curious Language Model: Strategic Test-Time Information Acquisition
by: Cooper, Michael, et al.
Published: (2025)
by: Cooper, Michael, et al.
Published: (2025)
Similar Items
-
The SMeL Test: A simple benchmark for media literacy in language models
by: Ahdritz, Gustaf, et al.
Published: (2025) -
A Benchmark for Learning to Translate a New Language from One Grammar Book
by: Tanzer, Garrett, et al.
Published: (2023) -
Fixed Point Diffusion Models
by: Bai, Xingjian, et al.
Published: (2024) -
Distinguishing the Knowable from the Unknowable with Language Models
by: Ahdritz, Gustaf, et al.
Published: (2024) -
Scaling Sign Language Translation
by: Zhang, Biao, et al.
Published: (2024)