StackingNet: Collective Inference Across Independent AI Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Siyang, Liu, Chenhao, Wu, Dongrui, Zeng, Zhigang, Ding, Lieyun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Yi: Open Foundation Models by 01.AI
por: AI, 01., et al.
Publicado: (2024)
por: AI, 01., et al.
Publicado: (2024)
How do Language Models Generate Slang: A Systematic Comparison between Human and Machine-Generated Slang Usages
por: Wu, Siyang, et al.
Publicado: (2025)
por: Wu, Siyang, et al.
Publicado: (2025)
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
por: Chen, Guanxu, et al.
Publicado: (2026)
por: Chen, Guanxu, et al.
Publicado: (2026)
LLM Inference Unveiled: Survey and Roofline Model Insights
por: Yuan, Zhihang, et al.
Publicado: (2024)
por: Yuan, Zhihang, et al.
Publicado: (2024)
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
por: Yang, Junyao, et al.
Publicado: (2026)
por: Yang, Junyao, et al.
Publicado: (2026)
LMFlow: An Extensible Toolkit for Finetuning and Inference of Large Foundation Models
por: Diao, Shizhe, et al.
Publicado: (2023)
por: Diao, Shizhe, et al.
Publicado: (2023)
Hit-RAG: Learning to Reason with Long Contexts via Preference Alignment
por: Liu, Junming, et al.
Publicado: (2026)
por: Liu, Junming, et al.
Publicado: (2026)
LED-Merging: Mitigating Safety-Utility Conflicts in Model Merging with Location-Election-Disjoint
por: Ma, Qianli, et al.
Publicado: (2025)
por: Ma, Qianli, et al.
Publicado: (2025)
Fanar 2.0: Arabic Generative AI Stack
por: FANAR TEAM, et al.
Publicado: (2026)
por: FANAR TEAM, et al.
Publicado: (2026)
From Personal to Collective: On the Role of Local and Global Memory in LLM Personalization
por: Wang, Zehong, et al.
Publicado: (2025)
por: Wang, Zehong, et al.
Publicado: (2025)
T-TIME: Test-Time Information Maximization Ensemble for Plug-and-Play BCIs
por: Li, Siyang, et al.
Publicado: (2024)
por: Li, Siyang, et al.
Publicado: (2024)
COLLEAGUE.SKILL: Automated AI Skill Generation via Expert Knowledge Distillation
por: Zhou, Tianyi, et al.
Publicado: (2026)
por: Zhou, Tianyi, et al.
Publicado: (2026)
A Survey on Human-AI Collaboration with Large Foundation Models
por: Vats, Vanshika, et al.
Publicado: (2024)
por: Vats, Vanshika, et al.
Publicado: (2024)
PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning
por: Zhang, Luan, et al.
Publicado: (2026)
por: Zhang, Luan, et al.
Publicado: (2026)
Entropy-Gradient Inversion: Moving Toward Internal Mechanism of Large Reasoning Models
por: Yang, Junyao, et al.
Publicado: (2026)
por: Yang, Junyao, et al.
Publicado: (2026)
ITERTL: An Iterative Framework for Fine-tuning LLMs for RTL Code Generation
por: Wu, Peiyang, et al.
Publicado: (2024)
por: Wu, Peiyang, et al.
Publicado: (2024)
RvB: Automating AI System Hardening via Iterative Red-Blue Games
por: Huang, Lige, et al.
Publicado: (2026)
por: Huang, Lige, et al.
Publicado: (2026)
A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science
por: Feng, Jie, et al.
Publicado: (2025)
por: Feng, Jie, et al.
Publicado: (2025)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
por: Ding, Junchen, et al.
Publicado: (2025)
por: Ding, Junchen, et al.
Publicado: (2025)
Do AI Models Perform Human-like Abstract Reasoning Across Modalities?
por: Beger, Claas, et al.
Publicado: (2025)
por: Beger, Claas, et al.
Publicado: (2025)
AgentStealth: Reinforcing Large Language Model for Anonymizing User-generated Text
por: Shao, Chenyang, et al.
Publicado: (2025)
por: Shao, Chenyang, et al.
Publicado: (2025)
Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Principles
por: Wei, Qingyan, et al.
Publicado: (2025)
por: Wei, Qingyan, et al.
Publicado: (2025)
Mapping Overlaps in Benchmarks through Perplexity in the Wild
por: Wu, Siyang, et al.
Publicado: (2025)
por: Wu, Siyang, et al.
Publicado: (2025)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
por: Yuan, Hongbang, et al.
Publicado: (2024)
por: Yuan, Hongbang, et al.
Publicado: (2024)
Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models
por: Qian, Chen, et al.
Publicado: (2024)
por: Qian, Chen, et al.
Publicado: (2024)
Rethinking Entropy Regularization in Large Reasoning Models
por: Jiang, Yuxian, et al.
Publicado: (2025)
por: Jiang, Yuxian, et al.
Publicado: (2025)
Conditional Advantage Estimation for Reinforcement Learning in Large Reasoning Models
por: Chen, Guanxu, et al.
Publicado: (2025)
por: Chen, Guanxu, et al.
Publicado: (2025)
LLMs Deceive Unintentionally: Emergent Misalignment in Dishonesty from Misaligned Samples to Biased Human-AI Interactions
por: Hu, Xuhao, et al.
Publicado: (2025)
por: Hu, Xuhao, et al.
Publicado: (2025)
Zero-Shot Continuous Prompt Transfer: Generalizing Task Semantics Across Language Models
por: Wu, Zijun, et al.
Publicado: (2023)
por: Wu, Zijun, et al.
Publicado: (2023)
Focus on Your Question! Interpreting and Mitigating Toxic CoT Problems in Commonsense Reasoning
por: Li, Jiachun, et al.
Publicado: (2024)
por: Li, Jiachun, et al.
Publicado: (2024)
Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning
por: Qian, Chen, et al.
Publicado: (2025)
por: Qian, Chen, et al.
Publicado: (2025)
From Feedback to Checklists: Grounded Evaluation of AI-Generated Clinical Notes
por: Zhou, Karen, et al.
Publicado: (2025)
por: Zhou, Karen, et al.
Publicado: (2025)
Training Foundation Models on a Full-Stack AMD Platform: Compute, Networking, and System Design
por: Anthony, Quentin, et al.
Publicado: (2025)
por: Anthony, Quentin, et al.
Publicado: (2025)
SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoning
por: Wang, Bin, et al.
Publicado: (2023)
por: Wang, Bin, et al.
Publicado: (2023)
Brainstorming Brings Power to Large Language Models of Knowledge Reasoning
por: Qin, Zining, et al.
Publicado: (2024)
por: Qin, Zining, et al.
Publicado: (2024)
How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles
por: Kuai, Chenchen, et al.
Publicado: (2026)
por: Kuai, Chenchen, et al.
Publicado: (2026)
DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence
por: Venkit, Pranav Narayanan, et al.
Publicado: (2025)
por: Venkit, Pranav Narayanan, et al.
Publicado: (2025)
K-ON: Stacking Knowledge On the Head Layer of Large Language Model
por: Guo, Lingbing, et al.
Publicado: (2025)
por: Guo, Lingbing, et al.
Publicado: (2025)
A Survey on the Applications of Frontier AI, Foundation Models, and Large Language Models to Intelligent Transportation Systems
por: Shoaib, Mohamed R., et al.
Publicado: (2024)
por: Shoaib, Mohamed R., et al.
Publicado: (2024)
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
por: Han, Tianyu, et al.
Publicado: (2023)
por: Han, Tianyu, et al.
Publicado: (2023)
Ejemplares similares
-
Yi: Open Foundation Models by 01.AI
por: AI, 01., et al.
Publicado: (2024) -
How do Language Models Generate Slang: A Systematic Comparison between Human and Machine-Generated Slang Usages
por: Wu, Siyang, et al.
Publicado: (2025) -
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
por: Chen, Guanxu, et al.
Publicado: (2026) -
LLM Inference Unveiled: Survey and Roofline Model Insights
por: Yuan, Zhihang, et al.
Publicado: (2024) -
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
por: Yang, Junyao, et al.
Publicado: (2026)