Distilling an End-to-End Voice Assistant Without Instruction Training Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Held, William, Li, Ella, Ryan, Michael, Shi, Weiyan, Zhang, Yanzhe, Yang, Diyi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sketch2Code: Evaluating Vision-Language Models for Interactive Web Design Prototyping
por: Li, Ryan, et al.
Publicado: (2024)
por: Li, Ryan, et al.
Publicado: (2024)
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
por: Li, Minzhi, et al.
Publicado: (2024)
por: Li, Minzhi, et al.
Publicado: (2024)
Searching for Privacy Risks in LLM Agents via Simulation
por: Zhang, Yanzhe, et al.
Publicado: (2025)
por: Zhang, Yanzhe, et al.
Publicado: (2025)
Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants
por: Huang, Vincent, et al.
Publicado: (2025)
por: Huang, Vincent, et al.
Publicado: (2025)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
por: Ryan, Michael J., et al.
Publicado: (2025)
por: Ryan, Michael J., et al.
Publicado: (2025)
Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment
por: Gan, Woody Haosheng, et al.
Publicado: (2026)
por: Gan, Woody Haosheng, et al.
Publicado: (2026)
SynthesizeMe! Inducing Persona-Guided Prompts for Personalized Reward Models in LLMs
por: Ryan, Michael J, et al.
Publicado: (2025)
por: Ryan, Michael J, et al.
Publicado: (2025)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
por: Li, Weizhen, et al.
Publicado: (2025)
por: Li, Weizhen, et al.
Publicado: (2025)
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
por: Zeng, Yi, et al.
Publicado: (2024)
por: Zeng, Yi, et al.
Publicado: (2024)
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
por: Shao, Yijia, et al.
Publicado: (2024)
por: Shao, Yijia, et al.
Publicado: (2024)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
por: Liu, Zijun, et al.
Publicado: (2023)
por: Liu, Zijun, et al.
Publicado: (2023)
CultureBank: An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies
por: Shi, Weiyan, et al.
Publicado: (2024)
por: Shi, Weiyan, et al.
Publicado: (2024)
Generative Interfaces for Language Models
por: Chen, Jiaqi, et al.
Publicado: (2025)
por: Chen, Jiaqi, et al.
Publicado: (2025)
Auditing Gender Presentation Differences in Text-to-Image Models
por: Zhang, Yanzhe, et al.
Publicado: (2023)
por: Zhang, Yanzhe, et al.
Publicado: (2023)
MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning
por: Yuan, Qianhao, et al.
Publicado: (2025)
por: Yuan, Qianhao, et al.
Publicado: (2025)
Thai Semantic End-of-Turn Detection for Real-Time Voice Agents
por: Popit, Thanapol, et al.
Publicado: (2025)
por: Popit, Thanapol, et al.
Publicado: (2025)
WebDS: An End-to-End Benchmark for Web-based Data Science
por: Hsu, Ethan, et al.
Publicado: (2025)
por: Hsu, Ethan, et al.
Publicado: (2025)
Using Large Language Model for End-to-End Chinese ASR and NER
por: Li, Yuang, et al.
Publicado: (2024)
por: Li, Yuang, et al.
Publicado: (2024)
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
por: Si, Chenglei, et al.
Publicado: (2024)
por: Si, Chenglei, et al.
Publicado: (2024)
End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning
por: Zheng, Qiaoyu, et al.
Publicado: (2025)
por: Zheng, Qiaoyu, et al.
Publicado: (2025)
Comparing Data Augmentation Methods for End-to-End Task-Oriented Dialog Systems
por: Vlachos, Christos, et al.
Publicado: (2024)
por: Vlachos, Christos, et al.
Publicado: (2024)
ESAinsTOD: A Unified End-to-End Schema-Aware Instruction-Tuning Framework for Task-Oriented Dialog Modeling
por: Teng, Dechuan, et al.
Publicado: (2026)
por: Teng, Dechuan, et al.
Publicado: (2026)
The End of Manual Decoding: Towards Truly End-to-End Language Models
por: Wang, Zhichao, et al.
Publicado: (2025)
por: Wang, Zhichao, et al.
Publicado: (2025)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
por: Chen, Zhuoen, et al.
Publicado: (2026)
por: Chen, Zhuoen, et al.
Publicado: (2026)
FormalASR: End-to-End Spoken Chinese to Formal Text
por: Ning, Wanyi, et al.
Publicado: (2026)
por: Ning, Wanyi, et al.
Publicado: (2026)
Contextualized Privacy Defense for LLM Agents
por: Wen, Yule, et al.
Publicado: (2026)
por: Wen, Yule, et al.
Publicado: (2026)
End-to-End Graph Flattening Method for Large Language Models
por: Hong, Bin, et al.
Publicado: (2024)
por: Hong, Bin, et al.
Publicado: (2024)
GSQA: An End-to-End Model for Generative Spoken Question Answering
por: Shih, Min-Han, et al.
Publicado: (2023)
por: Shih, Min-Han, et al.
Publicado: (2023)
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
por: Lan, Zhibin, et al.
Publicado: (2024)
por: Lan, Zhibin, et al.
Publicado: (2024)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
por: Huber, Christian, et al.
Publicado: (2023)
por: Huber, Christian, et al.
Publicado: (2023)
Language Model Inversion through End-to-End Differentiation
por: Denamganaï, Kevin Yandoka, et al.
Publicado: (2026)
por: Denamganaï, Kevin Yandoka, et al.
Publicado: (2026)
End-to-End Aspect-Guided Review Summarization at Scale
por: Boytsov, Ilya, et al.
Publicado: (2025)
por: Boytsov, Ilya, et al.
Publicado: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
por: He, Hongliang, et al.
Publicado: (2024)
por: He, Hongliang, et al.
Publicado: (2024)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
por: Du, Weihua, et al.
Publicado: (2025)
por: Du, Weihua, et al.
Publicado: (2025)
End-To-End Clinical Trial Matching with Large Language Models
por: Ferber, Dyke, et al.
Publicado: (2024)
por: Ferber, Dyke, et al.
Publicado: (2024)
End-to-End Long Document Summarization using Gradient Caching
por: Saxena, Rohit, et al.
Publicado: (2025)
por: Saxena, Rohit, et al.
Publicado: (2025)
Meow: End-to-End Outline Writing for Automatic Academic Survey
por: Ma, Zhaoyu, et al.
Publicado: (2025)
por: Ma, Zhaoyu, et al.
Publicado: (2025)
VoiceBench: Benchmarking LLM-Based Voice Assistants
por: Chen, Yiming, et al.
Publicado: (2024)
por: Chen, Yiming, et al.
Publicado: (2024)
Enhancing End-to-End Multi-Task Dialogue Systems: A Study on Intrinsic Motivation Reinforcement Learning Algorithms for Improved Training and Adaptability
por: Kamuni, Navin, et al.
Publicado: (2024)
por: Kamuni, Navin, et al.
Publicado: (2024)
MiCU: End-to-End Smart Home Command Understanding with Large Language Model
por: Han, Haowei, et al.
Publicado: (2026)
por: Han, Haowei, et al.
Publicado: (2026)
Ejemplares similares
-
Sketch2Code: Evaluating Vision-Language Models for Interactive Web Design Prototyping
por: Li, Ryan, et al.
Publicado: (2024) -
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
por: Li, Minzhi, et al.
Publicado: (2024) -
Searching for Privacy Risks in LLM Agents via Simulation
por: Zhang, Yanzhe, et al.
Publicado: (2025) -
Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants
por: Huang, Vincent, et al.
Publicado: (2025) -
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
por: Ryan, Michael J., et al.
Publicado: (2025)