The Superalignment of Superhuman Intelligence with Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, Minlie, Wang, Yingkang, Cui, Shiyao, Ke, Pei, Tang, Jie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Effective and Efficient Continual Pre-training of Large Language Models
por: Chen, Jie, et al.
Publicado: (2024)
por: Chen, Jie, et al.
Publicado: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
Exploring State Tracking Capabilities of Large Language Models
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
por: Liu, Aiwei, et al.
Publicado: (2024)
por: Liu, Aiwei, et al.
Publicado: (2024)
Distilling Large Language Models for Efficient Clinical Information Extraction
por: Vedula, Karthik S., et al.
Publicado: (2024)
por: Vedula, Karthik S., et al.
Publicado: (2024)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
por: Wang, Zhilin, et al.
Publicado: (2025)
por: Wang, Zhilin, et al.
Publicado: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
por: Bayram, M. Ali, et al.
Publicado: (2024)
por: Bayram, M. Ali, et al.
Publicado: (2024)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
por: Fang, Xi, et al.
Publicado: (2024)
por: Fang, Xi, et al.
Publicado: (2024)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
por: Weigang, Li, et al.
Publicado: (2025)
por: Weigang, Li, et al.
Publicado: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
por: Yao, Ben, et al.
Publicado: (2025)
por: Yao, Ben, et al.
Publicado: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
por: Oehri, Markus, et al.
Publicado: (2025)
por: Oehri, Markus, et al.
Publicado: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
por: Choi, Jonghyeon, et al.
Publicado: (2025)
por: Choi, Jonghyeon, et al.
Publicado: (2025)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
por: Liu, Aiwei, et al.
Publicado: (2025)
por: Liu, Aiwei, et al.
Publicado: (2025)
Evaluating Pixel Language Models on Non-Standardized Languages
por: Muñoz-Ortiz, Alberto, et al.
Publicado: (2024)
por: Muñoz-Ortiz, Alberto, et al.
Publicado: (2024)
ScoreRAG: A Retrieval-Augmented Generation Framework with Consistency-Relevance Scoring and Structured Summarization for News Generation
por: Lin, Pei-Yun, et al.
Publicado: (2025)
por: Lin, Pei-Yun, et al.
Publicado: (2025)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
por: Lee, Yejin, et al.
Publicado: (2026)
por: Lee, Yejin, et al.
Publicado: (2026)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
por: Park, Seungcheol, et al.
Publicado: (2023)
por: Park, Seungcheol, et al.
Publicado: (2023)
WaterSeeker: Pioneering Efficient Detection of Watermarked Segments in Large Documents
por: Pan, Leyi, et al.
Publicado: (2024)
por: Pan, Leyi, et al.
Publicado: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
A Survey on Natural Language Counterfactual Generation
por: Wang, Yongjie, et al.
Publicado: (2024)
por: Wang, Yongjie, et al.
Publicado: (2024)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
por: Delgado, Francisco Jose Cortes, et al.
Publicado: (2025)
por: Delgado, Francisco Jose Cortes, et al.
Publicado: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
por: Nieth, Björn, et al.
Publicado: (2026)
por: Nieth, Björn, et al.
Publicado: (2026)
MELoRA: Mini-Ensemble Low-Rank Adapters for Parameter-Efficient Fine-Tuning
por: Ren, Pengjie, et al.
Publicado: (2024)
por: Ren, Pengjie, et al.
Publicado: (2024)
Math Natural Language Inference: this should be easy!
por: de Paiva, Valeria, et al.
Publicado: (2025)
por: de Paiva, Valeria, et al.
Publicado: (2025)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
Profiling German Text Simplification with Interpretable Model-Fingerprints
por: Klöser, Lars, et al.
Publicado: (2026)
por: Klöser, Lars, et al.
Publicado: (2026)
Large Language Models for Propaganda Span Annotation
por: Hasanain, Maram, et al.
Publicado: (2023)
por: Hasanain, Maram, et al.
Publicado: (2023)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
por: Meng, Shiao, et al.
Publicado: (2024)
por: Meng, Shiao, et al.
Publicado: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
por: Tu, Songjun, et al.
Publicado: (2026)
por: Tu, Songjun, et al.
Publicado: (2026)
Fast Quiet-STaR: Thinking Without Thought Tokens
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
por: Wang, Zhilin, et al.
Publicado: (2026)
por: Wang, Zhilin, et al.
Publicado: (2026)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
por: Luo, Jianwen, et al.
Publicado: (2025)
por: Luo, Jianwen, et al.
Publicado: (2025)
A Semantic Invariant Robust Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
The Impact of Role Design in In-Context Learning for Large Language Models
por: Rouzegar, Hamidreza, et al.
Publicado: (2025)
por: Rouzegar, Hamidreza, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Effective and Efficient Continual Pre-training of Large Language Models
por: Chen, Jie, et al.
Publicado: (2024) -
An Unforgeable Publicly Verifiable Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023) -
Exploring State Tracking Capabilities of Large Language Models
por: Rezaee, Kiamehr, et al.
Publicado: (2025) -
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
por: Pan, Leyi, et al.
Publicado: (2025) -
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
por: Liu, Aiwei, et al.
Publicado: (2024)