Guardado en:
| Autores principales: | Schoenberg, William, Girard, Davidson, Chung, Saras, O'Neill, Ellen, Velasquez, Janet, Metcalf, Sara |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2503.15580 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Qualitative Engine: Creating and Evaluating an Iterative AI Modeling Tool
por: William Schoenberg, et al.
Publicado: (2026)
por: William Schoenberg, et al.
Publicado: (2026)
BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation
por: Metcalf, Sara, et al.
Publicado: (2026)
por: Metcalf, Sara, et al.
Publicado: (2026)
Regulating AI: Applying insights from behavioural economics and psychology to the application of article 5 of the EU AI Act
por: Zhong, Huixin, et al.
Publicado: (2023)
por: Zhong, Huixin, et al.
Publicado: (2023)
Building and Learning With Models Using AI
por: William Schoenberg
Publicado: (2026)
por: William Schoenberg
Publicado: (2026)
Efficiency Will Not Lead to Sustainable Reasoning AI
por: Wiesner, Philipp, et al.
Publicado: (2025)
por: Wiesner, Philipp, et al.
Publicado: (2025)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
por: Asawa, Parth, et al.
Publicado: (2025)
por: Asawa, Parth, et al.
Publicado: (2025)
Paired Completion: Flexible Quantification of Issue-framing at Scale with LLMs
por: Angus, Simon D, et al.
Publicado: (2024)
por: Angus, Simon D, et al.
Publicado: (2024)
How Can Generative AI Enhance the Well-being of Blind?
por: Bendel, Oliver
Publicado: (2024)
por: Bendel, Oliver
Publicado: (2024)
Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest
por: O'Neill, Abigail, et al.
Publicado: (2026)
por: O'Neill, Abigail, et al.
Publicado: (2026)
How Well Can Transformers Emulate In-context Newton's Method?
por: Giannou, Angeliki, et al.
Publicado: (2024)
por: Giannou, Angeliki, et al.
Publicado: (2024)
How Well Do Models Follow Their Constitutions?
por: Jakkli, Arya, et al.
Publicado: (2026)
por: Jakkli, Arya, et al.
Publicado: (2026)
“Can you help me think this through?” How pediatric hospitalists learn from informal peer consultation
por: Laura B. O'Neill, et al.
Publicado: (2024)
por: Laura B. O'Neill, et al.
Publicado: (2024)
AI reasoning effort predicts human decision time in content moderation
por: Davidson, Thomas
Publicado: (2025)
por: Davidson, Thomas
Publicado: (2025)
How Clinicians Think and What AI Can Learn From It
por: Sengupta, Dipayan, et al.
Publicado: (2026)
por: Sengupta, Dipayan, et al.
Publicado: (2026)
AIBuildAI: An AI Agent for Automatically Building AI Models
por: Zhang, Ruiyi, et al.
Publicado: (2026)
por: Zhang, Ruiyi, et al.
Publicado: (2026)
How Well Can LLMs Negotiate? NegotiationArena Platform and Analysis
por: Bianchi, Federico, et al.
Publicado: (2024)
por: Bianchi, Federico, et al.
Publicado: (2024)
Dukawalla: Voice Interfaces for Small Businesses in Africa
por: Ankrah, Elizabeth, et al.
Publicado: (2025)
por: Ankrah, Elizabeth, et al.
Publicado: (2025)
Can OpenAI o1 Reason Well in Ophthalmology? A 6,990-Question Head-to-Head Evaluation Study
por: Srinivasan, Sahana, et al.
Publicado: (2025)
por: Srinivasan, Sahana, et al.
Publicado: (2025)
Why the Center Can't Hold: A Diagnosis of Puritanized America
por: O’Neill, Tom
Publicado: (2019)
por: O’Neill, Tom
Publicado: (2019)
Towards Measuring and Modeling "Culture" in LLMs: A Survey
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2024)
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2024)
How Well Do Large Language Models Truly Ground?
por: Lee, Hyunji, et al.
Publicado: (2023)
por: Lee, Hyunji, et al.
Publicado: (2023)
How Well Do Multimodal Models Reason on ECG Signals?
por: Xu, Maxwell A., et al.
Publicado: (2026)
por: Xu, Maxwell A., et al.
Publicado: (2026)
How Well Can a Long Sequence Model Model Long Sequences? Comparing Architechtural Inductive Biases on Long-Context Abilities
por: Huang, Jerry
Publicado: (2024)
por: Huang, Jerry
Publicado: (2024)
How Well Can Vison-Language Models Understand Humans' Intention? An Open-ended Theory of Mind Question Evaluation Benchmark
por: Wen, Ximing, et al.
Publicado: (2025)
por: Wen, Ximing, et al.
Publicado: (2025)
MedLoRD: A Medical Low-Resource Diffusion Model for High-Resolution 3D CT Image Synthesis
por: Seyfarth, Marvin, et al.
Publicado: (2025)
por: Seyfarth, Marvin, et al.
Publicado: (2025)
"I know myself better, but not really greatly": How Well Can LLMs Detect and Explain LLM-Generated Texts?
por: Ji, Jiazhou, et al.
Publicado: (2025)
por: Ji, Jiazhou, et al.
Publicado: (2025)
VibeServe: Can AI Agents Build Bespoke LLM Serving Systems?
por: Kamahori, Keisuke, et al.
Publicado: (2026)
por: Kamahori, Keisuke, et al.
Publicado: (2026)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
por: Cheng, Audrey, et al.
Publicado: (2025)
por: Cheng, Audrey, et al.
Publicado: (2025)
How Well Can LLM Agents Simulate End-User Security and Privacy Attitudes and Behaviors?
por: Li, Yuxuan, et al.
Publicado: (2026)
por: Li, Yuxuan, et al.
Publicado: (2026)
The AI Co-Ethnographer: How Far Can Automation Take Qualitative Research?
por: Retkowski, Fabian, et al.
Publicado: (2025)
por: Retkowski, Fabian, et al.
Publicado: (2025)
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
por: Zhang, Qiyuan, et al.
Publicado: (2025)
por: Zhang, Qiyuan, et al.
Publicado: (2025)
SD-MoE: Spectral Decomposition for Effective Expert Specialization
por: Huang, Ruijun, et al.
Publicado: (2026)
por: Huang, Ruijun, et al.
Publicado: (2026)
SD-VLM: Spatial Measuring and Understanding with Depth-Encoded Vision-Language Models
por: Chen, Pingyi, et al.
Publicado: (2025)
por: Chen, Pingyi, et al.
Publicado: (2025)
Two Online Map Matching Algorithms Based on Analytic Hierarchy Process and Fuzzy Logic
por: Lin, Jeremy J., et al.
Publicado: (2024)
por: Lin, Jeremy J., et al.
Publicado: (2024)
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
por: Ferrag, Mohamed Amine, et al.
Publicado: (2026)
por: Ferrag, Mohamed Amine, et al.
Publicado: (2026)
SD$^2$: Self-Distilled Sparse Drafters
por: Lasby, Mike, et al.
Publicado: (2025)
por: Lasby, Mike, et al.
Publicado: (2025)
AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment
por: Wei, Zhenlin, et al.
Publicado: (2026)
por: Wei, Zhenlin, et al.
Publicado: (2026)
How Group Lives Go Well
por: Beverley, John, et al.
Publicado: (2025)
por: Beverley, John, et al.
Publicado: (2025)
How to Build AI Agents by Augmenting LLMs with Codified Human Expert Domain Knowledge? A Software Engineering Framework
por: uulu, Choro Ulan, et al.
Publicado: (2026)
por: uulu, Choro Ulan, et al.
Publicado: (2026)
How Well Does Agent Development Reflect Real-World Work?
por: Wang, Zora Zhiruo, et al.
Publicado: (2026)
por: Wang, Zora Zhiruo, et al.
Publicado: (2026)
Ejemplares similares
-
The Qualitative Engine: Creating and Evaluating an Iterative AI Modeling Tool
por: William Schoenberg, et al.
Publicado: (2026) -
BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation
por: Metcalf, Sara, et al.
Publicado: (2026) -
Regulating AI: Applying insights from behavioural economics and psychology to the application of article 5 of the EU AI Act
por: Zhong, Huixin, et al.
Publicado: (2023) -
Building and Learning With Models Using AI
por: William Schoenberg
Publicado: (2026) -
Efficiency Will Not Lead to Sustainable Reasoning AI
por: Wiesner, Philipp, et al.
Publicado: (2025)