Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Tuo, Yuan, Jinyue, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations
von: Salgado, Henry, et al.
Veröffentlicht: (2026)
von: Salgado, Henry, et al.
Veröffentlicht: (2026)
Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment
von: Favero, Lucile, et al.
Veröffentlicht: (2025)
von: Favero, Lucile, et al.
Veröffentlicht: (2025)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
von: Tian, Yuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuan, et al.
Veröffentlicht: (2025)
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
von: Wu, Tongshuang, et al.
Veröffentlicht: (2023)
von: Wu, Tongshuang, et al.
Veröffentlicht: (2023)
CALYPSO: LLMs as Dungeon Masters' Assistants
von: Zhu, Andrew, et al.
Veröffentlicht: (2023)
von: Zhu, Andrew, et al.
Veröffentlicht: (2023)
Dialogue Language Model with Large-Scale Persona Data Engineering
von: Hong, Mengze, et al.
Veröffentlicht: (2024)
von: Hong, Mengze, et al.
Veröffentlicht: (2024)
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination
von: Rakin, Salman, et al.
Veröffentlicht: (2024)
von: Rakin, Salman, et al.
Veröffentlicht: (2024)
Robots in the Middle: Evaluating LLMs in Dispute Resolution
von: Tan, Jinzhe, et al.
Veröffentlicht: (2024)
von: Tan, Jinzhe, et al.
Veröffentlicht: (2024)
LLMs Get Lost In Multi-Turn Conversation
von: Laban, Philippe, et al.
Veröffentlicht: (2025)
von: Laban, Philippe, et al.
Veröffentlicht: (2025)
Pragmatics beyond humans: meaning, communication, and LLMs
von: Gvoždiak, Vít
Veröffentlicht: (2025)
von: Gvoždiak, Vít
Veröffentlicht: (2025)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews
von: Xu, Huimin, et al.
Veröffentlicht: (2025)
von: Xu, Huimin, et al.
Veröffentlicht: (2025)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
von: Fröhling, Leon, et al.
Veröffentlicht: (2024)
von: Fröhling, Leon, et al.
Veröffentlicht: (2024)
LLMs Corrupt Your Documents When You Delegate
von: Laban, Philippe, et al.
Veröffentlicht: (2026)
von: Laban, Philippe, et al.
Veröffentlicht: (2026)
VoicePilot: Harnessing LLMs as Speech Interfaces for Physically Assistive Robots
von: Padmanabha, Akhil, et al.
Veröffentlicht: (2024)
von: Padmanabha, Akhil, et al.
Veröffentlicht: (2024)
Generate-Then-Validate: A Novel Question Generation Approach Using Small Language Models
von: Wei, Yumou, et al.
Veröffentlicht: (2025)
von: Wei, Yumou, et al.
Veröffentlicht: (2025)
Generating Educational Materials with Different Levels of Readability using LLMs
von: Huang, Chieh-Yang, et al.
Veröffentlicht: (2024)
von: Huang, Chieh-Yang, et al.
Veröffentlicht: (2024)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
von: Huang, Dun-Ming, et al.
Veröffentlicht: (2024)
von: Huang, Dun-Ming, et al.
Veröffentlicht: (2024)
Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts
von: Kim, Seon Gyeom, et al.
Veröffentlicht: (2025)
von: Kim, Seon Gyeom, et al.
Veröffentlicht: (2025)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
von: Choi, Alexander S., et al.
Veröffentlicht: (2024)
von: Choi, Alexander S., et al.
Veröffentlicht: (2024)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
von: Kim, Ahrii, et al.
Veröffentlicht: (2026)
von: Kim, Ahrii, et al.
Veröffentlicht: (2026)
Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows
von: Balashov, Yuri, et al.
Veröffentlicht: (2026)
von: Balashov, Yuri, et al.
Veröffentlicht: (2026)
Sandpiper: Orchestrated AI-Annotation for Educational Discourse at Scale
von: Hedley, Daryl, et al.
Veröffentlicht: (2026)
von: Hedley, Daryl, et al.
Veröffentlicht: (2026)
WordCraft: Scaffolding the Keyword Method for L2 Vocabulary Learning with Multimodal LLMs
von: Shao, Yuheng, et al.
Veröffentlicht: (2026)
von: Shao, Yuheng, et al.
Veröffentlicht: (2026)
Agentic AutoSurvey: Let LLMs Survey LLMs
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
Will the Real Linda Please Stand up...to Large Language Models? Examining the Representativeness Heuristic in LLMs
von: Wang, Pengda, et al.
Veröffentlicht: (2024)
von: Wang, Pengda, et al.
Veröffentlicht: (2024)
Multi-Intent Recognition in Dialogue Understanding: A Comparison Between Smaller Open-Source LLMs
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
von: Vo, Truong, et al.
Veröffentlicht: (2025)
von: Vo, Truong, et al.
Veröffentlicht: (2025)
Using Contextually Aligned Online Reviews to Measure LLMs' Performance Disparities Across Language Varieties
von: Tang, Zixin, et al.
Veröffentlicht: (2025)
von: Tang, Zixin, et al.
Veröffentlicht: (2025)
AutoTAMP: Autoregressive Task and Motion Planning with LLMs as Translators and Checkers
von: Chen, Yongchao, et al.
Veröffentlicht: (2023)
von: Chen, Yongchao, et al.
Veröffentlicht: (2023)
OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
Exploring Personalized Health Support through Data-Driven, Theory-Guided LLMs: A Case Study in Sleep Health
von: Wang, Xingbo, et al.
Veröffentlicht: (2025)
von: Wang, Xingbo, et al.
Veröffentlicht: (2025)
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
Telephone Surveys Meet Conversational AI: Evaluating a LLM-Based Telephone Survey System at Scale
von: Lang, Max M., et al.
Veröffentlicht: (2025)
von: Lang, Max M., et al.
Veröffentlicht: (2025)
DTS-SQL: Decomposed Text-to-SQL with Small Large Language Models
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
Revisiting Active Learning under (Human) Label Variation
von: Gruber, Cornelia, et al.
Veröffentlicht: (2025)
von: Gruber, Cornelia, et al.
Veröffentlicht: (2025)
Voices of Freelance Professional Writers on AI: Limitations, Expectations, and Fears
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
von: Ivey, Jonathan, et al.
Veröffentlicht: (2024)
von: Ivey, Jonathan, et al.
Veröffentlicht: (2024)
Why Robots Are Bad at Detecting Their Mistakes: Limitations of Miscommunication Detection in Human-Robot Dialogue
von: Janssens, Ruben, et al.
Veröffentlicht: (2025)
von: Janssens, Ruben, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations
von: Salgado, Henry, et al.
Veröffentlicht: (2026) -
Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment
von: Favero, Lucile, et al.
Veröffentlicht: (2025) -
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
von: Tian, Yuan, et al.
Veröffentlicht: (2025) -
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
von: Wu, Tongshuang, et al.
Veröffentlicht: (2023) -
CALYPSO: LLMs as Dungeon Masters' Assistants
von: Zhu, Andrew, et al.
Veröffentlicht: (2023)