Guardado en:
| Autores principales: | Shen, Lingfeng, Tan, Weiting, Chen, Sihao, Chen, Yunmo, Zhang, Jingyu, Xu, Haoran, Zheng, Boyuan, Koehn, Philipp, Khashabi, Daniel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2401.13136 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
por: Tan, Weiting, et al.
Publicado: (2024)
por: Tan, Weiting, et al.
Publicado: (2024)
Streaming Sequence Transduction through Dynamic Compression
por: Tan, Weiting, et al.
Publicado: (2024)
por: Tan, Weiting, et al.
Publicado: (2024)
Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
por: Xu, Haoran, et al.
Publicado: (2024)
por: Xu, Haoran, et al.
Publicado: (2024)
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models
por: Li, Tianjian, et al.
Publicado: (2023)
por: Li, Tianjian, et al.
Publicado: (2023)
Do pretrained Transformers Learn In-Context by Gradient Descent?
por: Shen, Lingfeng, et al.
Publicado: (2023)
por: Shen, Lingfeng, et al.
Publicado: (2023)
Upsample or Upweight? Balanced Training on Heavily Imbalanced Datasets
por: Li, Tianjian, et al.
Publicado: (2024)
por: Li, Tianjian, et al.
Publicado: (2024)
Learn and Unlearn: Addressing Misinformation in Multilingual LLMs
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
Seeing is Believing: Emotion-Aware Audio-Visual Language Modeling for Expressive Speech Generation
por: Tan, Weiting, et al.
Publicado: (2025)
por: Tan, Weiting, et al.
Publicado: (2025)
Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements
por: Zhang, Jingyu, et al.
Publicado: (2024)
por: Zhang, Jingyu, et al.
Publicado: (2024)
The Translation Barrier Hypothesis: Multilingual Generation with Large Language Models Suffers from Implicit Translation Failure
por: Bafna, Niyati, et al.
Publicado: (2025)
por: Bafna, Niyati, et al.
Publicado: (2025)
It Takes Two: On the Seamlessness between Reward and Policy Model in RLHF
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
GOLD PANNING: Strategic Context Shuffling for Needle-in-Haystack Reasoning
por: Byerly, Adam, et al.
Publicado: (2025)
por: Byerly, Adam, et al.
Publicado: (2025)
Challenging the Evaluator: LLM Sycophancy Under User Rebuttal
por: Kim, Sungwon, et al.
Publicado: (2025)
por: Kim, Sungwon, et al.
Publicado: (2025)
Process-Supervised Reinforcement Learning for Interactive Multimodal Tool-Use Agents
por: Tan, Weiting, et al.
Publicado: (2025)
por: Tan, Weiting, et al.
Publicado: (2025)
Self-Consistency Falls Short! The Adverse Effects of Positional Bias on Long-Context Problems
por: Byerly, Adam, et al.
Publicado: (2024)
por: Byerly, Adam, et al.
Publicado: (2024)
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data
por: Zhang, Jingyu, et al.
Publicado: (2024)
por: Zhang, Jingyu, et al.
Publicado: (2024)
Learning to Retrieve Iteratively for In-Context Learning
por: Chen, Yunmo, et al.
Publicado: (2024)
por: Chen, Yunmo, et al.
Publicado: (2024)
Are Finer Citations Always Better? Rethinking Granularity for Attributed Generation
por: Wang, Hexuan, et al.
Publicado: (2026)
por: Wang, Hexuan, et al.
Publicado: (2026)
SIMPLEMIX: Frustratingly Simple Mixing of Off- and On-policy Data in Language Model Preference Learning
por: Li, Tianjian, et al.
Publicado: (2025)
por: Li, Tianjian, et al.
Publicado: (2025)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
por: Jiang, Dongwei, et al.
Publicado: (2024)
por: Jiang, Dongwei, et al.
Publicado: (2024)
Text Style Transfer with Parameter-efficient LLM Finetuning and Round-trip Translation
por: Liu, Ruoxi, et al.
Publicado: (2026)
por: Liu, Ruoxi, et al.
Publicado: (2026)
Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents
por: Meng, Chutong, et al.
Publicado: (2025)
por: Meng, Chutong, et al.
Publicado: (2025)
X-ALMA: Plug & Play Modules and Adaptive Rejection for Quality Translation at Scale
por: Xu, Haoran, et al.
Publicado: (2024)
por: Xu, Haoran, et al.
Publicado: (2024)
Linguistic Nepotism: Trading-off Quality for Language Preference in Multilingual RAG
por: Ki, Dayeon, et al.
Publicado: (2025)
por: Ki, Dayeon, et al.
Publicado: (2025)
MultiMUC: Multilingual Template Filling on MUC-4
por: Gantt, William, et al.
Publicado: (2024)
por: Gantt, William, et al.
Publicado: (2024)
SemStamp: A Semantic Watermark with Paraphrastic Robustness for Text Generation
por: Hou, Abe Bohan, et al.
Publicado: (2023)
por: Hou, Abe Bohan, et al.
Publicado: (2023)
Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Observations
por: Chen, Nuo, et al.
Publicado: (2023)
por: Chen, Nuo, et al.
Publicado: (2023)
Jailbreak Distillation: Renewable Safety Benchmarking
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
Certified Mitigation of Worst-Case LLM Copyright Infringement
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
Dissecting Multiplication in Transformers: Insights into LLMs
por: Qiu, Luyu, et al.
Publicado: (2024)
por: Qiu, Luyu, et al.
Publicado: (2024)
Pointer-Generator Networks for Low-Resource Machine Translation: Don't Copy That!
por: Bafna, Niyati, et al.
Publicado: (2024)
por: Bafna, Niyati, et al.
Publicado: (2024)
Recovering document annotations for sentence-level bitext
por: Wicks, Rachel, et al.
Publicado: (2024)
por: Wicks, Rachel, et al.
Publicado: (2024)
The Alignment Waltz: Jointly Training Agents to Collaborate for Safety
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
ICL CIPHERS: Quantifying "Learning" in In-Context Learning via Substitution Ciphers
por: Fang, Zhouxiang, et al.
Publicado: (2025)
por: Fang, Zhouxiang, et al.
Publicado: (2025)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
Feedback Friction: LLMs Struggle to Fully Incorporate External Feedback
por: Jiang, Dongwei, et al.
Publicado: (2025)
por: Jiang, Dongwei, et al.
Publicado: (2025)
k-SemStamp: A Clustering-Based Semantic Watermark for Detection of Machine-Generated Text
por: Hou, Abe Bohan, et al.
Publicado: (2024)
por: Hou, Abe Bohan, et al.
Publicado: (2024)
AnaloBench: Benchmarking the Identification of Abstract and Long-context Analogies
por: Ye, Xiao, et al.
Publicado: (2024)
por: Ye, Xiao, et al.
Publicado: (2024)
Many-Tier Instruction Hierarchy in LLM Agents
por: Zhang, Jingyu, et al.
Publicado: (2026)
por: Zhang, Jingyu, et al.
Publicado: (2026)
SafeMT: Multi-turn Safety for Multimodal Language Models
por: Zhu, Han, et al.
Publicado: (2025)
por: Zhu, Han, et al.
Publicado: (2025)
Ejemplares similares
-
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
por: Tan, Weiting, et al.
Publicado: (2024) -
Streaming Sequence Transduction through Dynamic Compression
por: Tan, Weiting, et al.
Publicado: (2024) -
Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
por: Xu, Haoran, et al.
Publicado: (2024) -
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models
por: Li, Tianjian, et al.
Publicado: (2023) -
Do pretrained Transformers Learn In-Context by Gradient Descent?
por: Shen, Lingfeng, et al.
Publicado: (2023)