Gespeichert in:
| Hauptverfasser: | Berkovitch, Omri, Caduri, Sapir, Kahlon, Noam, Efros, Anatoly, Caciularu, Avi, Dagan, Ido |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2406.14314 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Small Models, Big Results: Achieving Superior Intent Extraction through Decomposition
von: Cohen, Danielle, et al.
Veröffentlicht: (2025)
von: Cohen, Danielle, et al.
Veröffentlicht: (2025)
Bi-Fact: A Bidirectional Factorization-based Evaluation of Intent Extraction from UI Trajectories
von: Caduri, Sapir, et al.
Veröffentlicht: (2025)
von: Caduri, Sapir, et al.
Veröffentlicht: (2025)
Agent-Initiated Interaction in Phone UI Automation
von: Kahlon, Noam, et al.
Veröffentlicht: (2025)
von: Kahlon, Noam, et al.
Veröffentlicht: (2025)
Dont Add, dont Miss: Effective Content Preserving Generation from Pre-Selected Text Spans
von: Slobodkin, Aviv, et al.
Veröffentlicht: (2023)
von: Slobodkin, Aviv, et al.
Veröffentlicht: (2023)
PrefixNLI: Detecting Factual Inconsistencies as Soon as They Arise
von: Harary, Sapir, et al.
Veröffentlicht: (2025)
von: Harary, Sapir, et al.
Veröffentlicht: (2025)
Latent Reasoning with Supervised Thinking States
von: Amos, Ido, et al.
Veröffentlicht: (2026)
von: Amos, Ido, et al.
Veröffentlicht: (2026)
Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
von: Ghandeharioun, Asma, et al.
Veröffentlicht: (2024)
von: Ghandeharioun, Asma, et al.
Veröffentlicht: (2024)
Unpacking Tokenization: Evaluating Text Compression and its Correlation with Model Performance
von: Goldman, Omer, et al.
Veröffentlicht: (2024)
von: Goldman, Omer, et al.
Veröffentlicht: (2024)
GenerationPrograms: Fine-grained Attribution with Executable Programs
von: Wan, David, et al.
Veröffentlicht: (2025)
von: Wan, David, et al.
Veröffentlicht: (2025)
Generating Tables from the Parametric Knowledge of Language Models
von: Berkovitch, Yevgeni, et al.
Veröffentlicht: (2024)
von: Berkovitch, Yevgeni, et al.
Veröffentlicht: (2024)
Is It Really Long Context if All You Need Is Retrieval? Towards Genuinely Difficult Long Context NLP
von: Goldman, Omer, et al.
Veröffentlicht: (2024)
von: Goldman, Omer, et al.
Veröffentlicht: (2024)
QAPyramid: Fine-grained Evaluation of Content Selection for Text Summarization
von: Zhang, Shiyue, et al.
Veröffentlicht: (2024)
von: Zhang, Shiyue, et al.
Veröffentlicht: (2024)
Efficient Data Generation for Source-grounded Information-seeking Dialogs: A Use Case for Meeting Transcripts
von: Golany, Lotem, et al.
Veröffentlicht: (2024)
von: Golany, Lotem, et al.
Veröffentlicht: (2024)
Identifying Narrative Patterns and Outliers in Holocaust Testimonies Using Topic Modeling
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs
von: Cattan, Arie, et al.
Veröffentlicht: (2025)
von: Cattan, Arie, et al.
Veröffentlicht: (2025)
$How^{2}$: How to learn from procedural How-to questions
von: Dagan, Gautier, et al.
Veröffentlicht: (2025)
von: Dagan, Gautier, et al.
Veröffentlicht: (2025)
MAIC-UI: Making Interactive Courseware with Generative UI
von: Tu, Shangqing, et al.
Veröffentlicht: (2026)
von: Tu, Shangqing, et al.
Veröffentlicht: (2026)
UI-JEPA: Towards Active Perception of User Intent through Onscreen User Activity
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
Plancraft: an evaluation dataset for planning with LLM agents
von: Dagan, Gautier, et al.
Veröffentlicht: (2024)
von: Dagan, Gautier, et al.
Veröffentlicht: (2024)
Goal Alignment in LLM-Based User Simulators for Conversational AI
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
Generative UI: LLMs are Effective UI Generators
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2026)
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2026)
Identifying & Interactively Refining Ambiguous User Goals for Data Visualization Code Generation
von: İnan, Mert, et al.
Veröffentlicht: (2025)
von: İnan, Mert, et al.
Veröffentlicht: (2025)
Express Your Doubts -- Probabilistic World Modeling Should not be Based on Token logprobs
von: Wagner, Eitan, et al.
Veröffentlicht: (2025)
von: Wagner, Eitan, et al.
Veröffentlicht: (2025)
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
von: Tseytlin, Maria, et al.
Veröffentlicht: (2025)
von: Tseytlin, Maria, et al.
Veröffentlicht: (2025)
A Comparative Study of Task Adaptation Techniques of Large Language Models for Identifying Sustainable Development Goals
von: Cadeddu, Andrea, et al.
Veröffentlicht: (2025)
von: Cadeddu, Andrea, et al.
Veröffentlicht: (2025)
CONTESTS: a Framework for Consistency Testing of Span Probabilities in Language Models
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
Exploring the Learning Capabilities of Language Models using LEVERWORLDS
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
Toward a Benchmark for Controllable Simulation of Imperfect Students with Large Language Models
von: Apartsin, Alexander, et al.
Veröffentlicht: (2026)
von: Apartsin, Alexander, et al.
Veröffentlicht: (2026)
FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection
von: Ouyang, Mingyu, et al.
Veröffentlicht: (2026)
von: Ouyang, Mingyu, et al.
Veröffentlicht: (2026)
Pretrained LLMs Learn Multiple Types of Uncertainty
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding
von: Wu, Qinzhuo, et al.
Veröffentlicht: (2024)
von: Wu, Qinzhuo, et al.
Veröffentlicht: (2024)
UFO: A UI-Focused Agent for Windows OS Interaction
von: Zhang, Chaoyun, et al.
Veröffentlicht: (2024)
von: Zhang, Chaoyun, et al.
Veröffentlicht: (2024)
Measuring Pragmatic Influence in Large Language Model Instructions
von: Geng, Yilin, et al.
Veröffentlicht: (2026)
von: Geng, Yilin, et al.
Veröffentlicht: (2026)
ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark
von: Shalyt, Michael, et al.
Veröffentlicht: (2025)
von: Shalyt, Michael, et al.
Veröffentlicht: (2025)
Artificial Expert Intelligence through PAC-reasoning
von: Shalev-Shwartz, Shai, et al.
Veröffentlicht: (2024)
von: Shalev-Shwartz, Shai, et al.
Veröffentlicht: (2024)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding
von: Tang, Fei, et al.
Veröffentlicht: (2026)
von: Tang, Fei, et al.
Veröffentlicht: (2026)
SelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
From Joy to Fear: A Benchmark of Emotion Estimation in Pop Song Lyrics
von: Dahary, Shay, et al.
Veröffentlicht: (2025)
von: Dahary, Shay, et al.
Veröffentlicht: (2025)
Mind the Goal: Data-Efficient Goal-Oriented Evaluation of Conversational Agents and Chatbots using Teacher Models
von: Piskala, Deepak Babu, et al.
Veröffentlicht: (2025)
von: Piskala, Deepak Babu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Small Models, Big Results: Achieving Superior Intent Extraction through Decomposition
von: Cohen, Danielle, et al.
Veröffentlicht: (2025) -
Bi-Fact: A Bidirectional Factorization-based Evaluation of Intent Extraction from UI Trajectories
von: Caduri, Sapir, et al.
Veröffentlicht: (2025) -
Agent-Initiated Interaction in Phone UI Automation
von: Kahlon, Noam, et al.
Veröffentlicht: (2025) -
Dont Add, dont Miss: Effective Content Preserving Generation from Pre-Selected Text Spans
von: Slobodkin, Aviv, et al.
Veröffentlicht: (2023) -
PrefixNLI: Detecting Factual Inconsistencies as Soon as They Arise
von: Harary, Sapir, et al.
Veröffentlicht: (2025)