Enregistré dans:
| Auteurs principaux: | Liu, Ruibo, Wei, Jerry, Liu, Fangyu, Si, Chenglei, Zhang, Yanzhe, Rao, Jinmeng, Zheng, Steven, Peng, Daiyi, Yang, Diyi, Zhou, Denny, Dai, Andrew M. |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2404.07503 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
par: Si, Chenglei, et autres
Publié: (2025)
par: Si, Chenglei, et autres
Publié: (2025)
Higher Layers Need More LoRA Experts
par: Gao, Chongyang, et autres
Publié: (2024)
par: Gao, Chongyang, et autres
Publié: (2024)
Searching for Privacy Risks in LLM Agents via Simulation
par: Zhang, Yanzhe, et autres
Publié: (2025)
par: Zhang, Yanzhe, et autres
Publié: (2025)
Scaling Tumor Segmentation: Best Lessons from Real and Synthetic Data
par: Chen, Qi, et autres
Publié: (2025)
par: Chen, Qi, et autres
Publié: (2025)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
par: Liu, Zijun, et autres
Publié: (2023)
par: Liu, Zijun, et autres
Publié: (2023)
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing
par: Yang, Diji, et autres
Publié: (2025)
par: Yang, Diji, et autres
Publié: (2025)
Sketch2Code: Evaluating Vision-Language Models for Interactive Web Design Prototyping
par: Li, Ryan, et autres
Publié: (2024)
par: Li, Ryan, et autres
Publié: (2024)
Attacking Vision-Language Computer Agents via Pop-ups
par: Zhang, Yanzhe, et autres
Publié: (2024)
par: Zhang, Yanzhe, et autres
Publié: (2024)
Towards Execution-Grounded Automated AI Research
par: Si, Chenglei, et autres
Publié: (2026)
par: Si, Chenglei, et autres
Publié: (2026)
Auditing Gender Presentation Differences in Text-to-Image Models
par: Zhang, Yanzhe, et autres
Publié: (2023)
par: Zhang, Yanzhe, et autres
Publié: (2023)
Long-form factuality in large language models
par: Wei, Jerry, et autres
Publié: (2024)
par: Wei, Jerry, et autres
Publié: (2024)
Distilling an End-to-End Voice Assistant Without Instruction Training Data
par: Held, William, et autres
Publié: (2024)
par: Held, William, et autres
Publié: (2024)
Highly Efficient and Stable Narrow Band Green Emitting Phosphor of Sb 3+ /Ce 3+ Sensitized Cs 2 NaTbCl 6 for WLED
par: Changheng Chen, et autres
Publié: (2024)
par: Changheng Chen, et autres
Publié: (2024)
Contextual Experience Replay for Self-Improvement of Language Agents
par: Liu, Yitao, et autres
Publié: (2025)
par: Liu, Yitao, et autres
Publié: (2025)
Generative Interfaces for Language Models
par: Chen, Jiaqi, et autres
Publié: (2025)
par: Chen, Jiaqi, et autres
Publié: (2025)
Real-Time Reasoning Agents in Evolving Environments
par: Wen, Yule, et autres
Publié: (2025)
par: Wen, Yule, et autres
Publié: (2025)
Type-Compliant Adaptation Cascades: Adapting Programmatic LM Workflows to Data
par: Lin, Chu-Cheng, et autres
Publié: (2025)
par: Lin, Chu-Cheng, et autres
Publié: (2025)
Borderline content and platformised speech governance: Mapping TikTok's moderation controversies in South and Southeast Asia
par: Diyi Liu
Publié: (2024)
par: Diyi Liu
Publié: (2024)
SPHERE: An Evaluation Card for Human-AI Systems
par: Ma, Qianou, et autres
Publié: (2025)
par: Ma, Qianou, et autres
Publié: (2025)
Defect‐Engineered Zero‐Dimensional Perovskite Cs 3 LuCl 6 : Tb 3+ Scintillator with Exceptional Thermal Stability for Flexible High‐Temperature X‐Ray Imaging
par: Ruibo Gao, et autres
Publié: (2026)
par: Ruibo Gao, et autres
Publié: (2026)
Achieving Single‐Phased Full Visible Spectrum Broadband White Emission in Ag⁺, Bi 3 ⁺, and Sb 3 ⁺ Tri‐Doped Cs₂NaLuCl₆ Double Perovskite Phosphor
par: Changheng Chen, et autres
Publié: (2025)
par: Changheng Chen, et autres
Publié: (2025)
The Best Instruction-Tuning Data are Those That Fit
par: Zhang, Dylan, et autres
Publié: (2025)
par: Zhang, Dylan, et autres
Publié: (2025)
LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
par: Zhang, Yanzhe, et autres
Publié: (2023)
par: Zhang, Yanzhe, et autres
Publié: (2023)
Contextualized Privacy Defense for LLM Agents
par: Wen, Yule, et autres
Publié: (2026)
par: Wen, Yule, et autres
Publié: (2026)
Security and Innovation in ERP Systems: Best Practices for AI, OIC, and Automation Integration
par: Sreenivasa Rao Sola
Publié: (2023)
par: Sreenivasa Rao Sola
Publié: (2023)
Simple synthetic data reduces sycophancy in large language models
par: Wei, Jerry, et autres
Publié: (2023)
par: Wei, Jerry, et autres
Publié: (2023)
Challenges and Best Practices in Corporate AI Governance:Lessons from the Biopharmaceutical Industry
par: Mökander, Jakob, et autres
Publié: (2024)
par: Mökander, Jakob, et autres
Publié: (2024)
When to Showcase Automated Production Processes? Disclosing Production Processes Increases Evaluation of Low‐End but Decreases Evaluation of High‐End Products
par: Diyi Liu, et autres
Publié: (2025)
par: Diyi Liu, et autres
Publié: (2025)
Deploying Tiny LVLM Judges for Real-World Evaluation of Chart Models: Lessons Learned and Best Practices
par: Laskar, Md Tahmid Rahman, et autres
Publié: (2025)
par: Laskar, Md Tahmid Rahman, et autres
Publié: (2025)
Selecting the Best Optimizing System
par: Si, Nian, et autres
Publié: (2022)
par: Si, Nian, et autres
Publié: (2022)
Robust Output Regulation of Uncertain Linear Time-Varying Systems
par: Zha, Jinmeng, et autres
Publié: (2026)
par: Zha, Jinmeng, et autres
Publié: (2026)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
par: Ryan, Michael J., et autres
Publié: (2025)
par: Ryan, Michael J., et autres
Publié: (2025)
S3Eval: A Synthetic, Scalable, Systematic Evaluation Suite for Large Language Models
par: Lei, Fangyu, et autres
Publié: (2023)
par: Lei, Fangyu, et autres
Publié: (2023)
Evaluation on Aggregate Particle Spalling of Induction Heating‐Based Functional Ultra‐Thin Friction Layer Using Image Processing Based on MATLAB
par: Zhengmengyuan Rao, et autres
Publié: (2026)
par: Zhengmengyuan Rao, et autres
Publié: (2026)
IM-RAG: Multi-Round Retrieval-Augmented Generation Through Learning Inner Monologues
par: Yang, Diji, et autres
Publié: (2024)
par: Yang, Diji, et autres
Publié: (2024)
Tweedie Regression for Video Recommendation System
par: Zheng, Yan, et autres
Publié: (2025)
par: Zheng, Yan, et autres
Publié: (2025)
Relic abundance of dark matter with coannihilation in non-standard cosmological scenarios
par: Liu, Fangyu, et autres
Publié: (2023)
par: Liu, Fangyu, et autres
Publié: (2023)
Constraints on Asymmetric Dark Matter Self Annihilation Cross Sections in Non-standard Cosmological Scenarios
par: Liu, Fangyu, et autres
Publié: (2023)
par: Liu, Fangyu, et autres
Publié: (2023)
Documents similaires
-
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
par: Si, Chenglei, et autres
Publié: (2024) -
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
par: Si, Chenglei, et autres
Publié: (2024) -
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
par: Si, Chenglei, et autres
Publié: (2025) -
Higher Layers Need More LoRA Experts
par: Gao, Chongyang, et autres
Publié: (2024) -
Searching for Privacy Risks in LLM Agents via Simulation
par: Zhang, Yanzhe, et autres
Publié: (2025)