Tuning Qwen2.5-VL to Improve Its Web Interaction Skills
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yakovleva, Alexandra, Pärssinen, Henrik, Valpola, Harri, Kannala, Juho, Ilin, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Iffy-Or-Not: Extending the Web to Support the Critical Evaluation of Fallacious Texts
von: Lim, Gionnieve, et al.
Veröffentlicht: (2025)
von: Lim, Gionnieve, et al.
Veröffentlicht: (2025)
WebAccessVL: Violation-Aware VLM for Web Accessibility
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2025)
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2025)
Benchmarking Large Language Models for Diagnosing Students' Cognitive Skills from Handwritten Math Work
von: Kim, Yoonsu, et al.
Veröffentlicht: (2025)
von: Kim, Yoonsu, et al.
Veröffentlicht: (2025)
ChoiceMates: Supporting Unfamiliar Online Decision-Making with Multi-Agent Conversational Interactions
von: Park, Jeongeon, et al.
Veröffentlicht: (2023)
von: Park, Jeongeon, et al.
Veröffentlicht: (2023)
Contrastive Explanations That Anticipate Human Misconceptions Can Improve Human Decision-Making Skills
von: Buçinca, Zana, et al.
Veröffentlicht: (2024)
von: Buçinca, Zana, et al.
Veröffentlicht: (2024)
AVIN-Chat: An Audio-Visual Interactive Chatbot System with Emotional State Tuning
von: Park, Chanhyuk, et al.
Veröffentlicht: (2024)
von: Park, Chanhyuk, et al.
Veröffentlicht: (2024)
Interactions with Prompt Problems: A New Way to Teach Programming with Large Language Models
von: Prather, James, et al.
Veröffentlicht: (2024)
von: Prather, James, et al.
Veröffentlicht: (2024)
"When to Hand Off, When to Work Together": Expanding Human-Agent Co-Creative Collaboration through Concurrent Interaction
von: Son, Kihoon, et al.
Veröffentlicht: (2026)
von: Son, Kihoon, et al.
Veröffentlicht: (2026)
ExpressEdit: Video Editing with Natural Language and Sketching
von: Tilekbay, Bekzat, et al.
Veröffentlicht: (2024)
von: Tilekbay, Bekzat, et al.
Veröffentlicht: (2024)
Human-Based Risk Model for Improved Driver Support in Interactive Driving Scenarios
von: Puphal, Tim, et al.
Veröffentlicht: (2024)
von: Puphal, Tim, et al.
Veröffentlicht: (2024)
Improving Steering and Verification in AI-Assisted Data Analysis with Interactive Task Decomposition
von: Kazemitabaar, Majeed, et al.
Veröffentlicht: (2024)
von: Kazemitabaar, Majeed, et al.
Veröffentlicht: (2024)
2-Factor Retrieval for Improved Human-AI Decision Making in Radiology
von: Solomon, Jim, et al.
Veröffentlicht: (2024)
von: Solomon, Jim, et al.
Veröffentlicht: (2024)
EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined Criteria
von: Kim, Tae Soo, et al.
Veröffentlicht: (2023)
von: Kim, Tae Soo, et al.
Veröffentlicht: (2023)
Explorer: Scaling Exploration-driven Web Trajectory Synthesis for Multimodal Web Agents
von: Pahuja, Vardaan, et al.
Veröffentlicht: (2025)
von: Pahuja, Vardaan, et al.
Veröffentlicht: (2025)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
Improving Student-AI Interaction Through Pedagogical Prompting: An Example in Computer Science Education
von: Xiao, Ruiwei, et al.
Veröffentlicht: (2025)
von: Xiao, Ruiwei, et al.
Veröffentlicht: (2025)
MindSpeech: Continuous Imagined Speech Decoding using High-Density fNIRS and Prompt Tuning for Advanced Human-AI Interaction
von: Zhang, Suyi, et al.
Veröffentlicht: (2024)
von: Zhang, Suyi, et al.
Veröffentlicht: (2024)
AutoS$^2$earch: Unlocking the Reasoning Potential of Large Models for Web-based Source Search
von: Zhu, Zhengqiu, et al.
Veröffentlicht: (2025)
von: Zhu, Zhengqiu, et al.
Veröffentlicht: (2025)
ReaLJam: Real-Time Human-AI Music Jamming with Reinforcement Learning-Tuned Transformers
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2025)
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2025)
The Impact of AI Usage and Informativeness on Skill Development in Logical Reasoning
von: Wu, Shang, et al.
Veröffentlicht: (2026)
von: Wu, Shang, et al.
Veröffentlicht: (2026)
Toward User Comprehension Supports for LLM Agent Skill Specifications
von: Wen, Zikai Alex
Veröffentlicht: (2026)
von: Wen, Zikai Alex
Veröffentlicht: (2026)
ClearFairy: Capturing Creative Workflows through Decision Structuring, In-Situ Questioning, and Rationale Inference
von: Son, Kihoon, et al.
Veröffentlicht: (2025)
von: Son, Kihoon, et al.
Veröffentlicht: (2025)
Towards Scalable Web Accessibility Audit with MLLMs as Copilots
von: Gu, Ming, et al.
Veröffentlicht: (2025)
von: Gu, Ming, et al.
Veröffentlicht: (2025)
Explainable AI for Automated User-specific Feedback in Surgical Skill Acquisition
von: Gomez, Catalina, et al.
Veröffentlicht: (2025)
von: Gomez, Catalina, et al.
Veröffentlicht: (2025)
Hand by Hand: LLM Driving EMS Assistant for Operational Skill Learning
von: Xiang, Wei, et al.
Veröffentlicht: (2025)
von: Xiang, Wei, et al.
Veröffentlicht: (2025)
Intentmaking and Sensemaking: Human Interaction with AI-Guided Mathematical Discovery
von: Bäuerle, Alex, et al.
Veröffentlicht: (2026)
von: Bäuerle, Alex, et al.
Veröffentlicht: (2026)
USER-VLM 360: Personalized Vision Language Models with User-aware Tuning for Social Human-Robot Interactions
von: Rahimi, Hamed, et al.
Veröffentlicht: (2025)
von: Rahimi, Hamed, et al.
Veröffentlicht: (2025)
Improving through Interaction: Searching Behavioral Representation Spaces with CMA-ES-IG
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2026)
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2026)
The Future of Skill: What Is It to Be Skilled at Work?
von: Niklasson, Axel, et al.
Veröffentlicht: (2024)
von: Niklasson, Axel, et al.
Veröffentlicht: (2024)
AppAgent v2: Advanced Agent for Flexible Mobile Interactions
von: Li, Yanda, et al.
Veröffentlicht: (2024)
von: Li, Yanda, et al.
Veröffentlicht: (2024)
Vibe Researching as Wolf Coming: Can AI Agents with Skills Replace or Augment Social Scientists?
von: Zhang, Yongjun
Veröffentlicht: (2026)
von: Zhang, Yongjun
Veröffentlicht: (2026)
Avenir-UX: Automated UX Evaluation via Simulated Human Web Interaction with GUI Grounding
von: Tan, Wee Joe, et al.
Veröffentlicht: (2026)
von: Tan, Wee Joe, et al.
Veröffentlicht: (2026)
The Widening Gap: The Benefits and Harms of Generative AI for Novice Programmers
von: Prather, James, et al.
Veröffentlicht: (2024)
von: Prather, James, et al.
Veröffentlicht: (2024)
Robots and Children that Learn Together : Improving Knowledge Retention by Teaching Peer-Like Interactive Robots
von: Tarakli, Imene, et al.
Veröffentlicht: (2025)
von: Tarakli, Imene, et al.
Veröffentlicht: (2025)
Are You Listening to Me? Fine-Tuning Chatbots for Empathetic Dialogue
von: Knob, Paulo Ricardo, et al.
Veröffentlicht: (2025)
von: Knob, Paulo Ricardo, et al.
Veröffentlicht: (2025)
From Static to Interactive: Authoring Interactive Visualizations via Natural Language
von: Liu, Can, et al.
Veröffentlicht: (2026)
von: Liu, Can, et al.
Veröffentlicht: (2026)
Improving Human-Autonomous Vehicle Interaction in Complex Systems
von: Kaufman, Robert
Veröffentlicht: (2025)
von: Kaufman, Robert
Veröffentlicht: (2025)
Biotic Browser: Applying StreamingLLM as a Persistent Web Browsing Co-Pilot
von: Dunnell, Kevin F., et al.
Veröffentlicht: (2024)
von: Dunnell, Kevin F., et al.
Veröffentlicht: (2024)
Towards Sustainable Web Agents: A Plea for Transparency and Dedicated Metrics for Energy Consumption
von: Krupp, Lars, et al.
Veröffentlicht: (2025)
von: Krupp, Lars, et al.
Veröffentlicht: (2025)
HearHere: Mitigating Echo Chambers in News Consumption through an AI-based Web System
von: Jeon, Youngseung, et al.
Veröffentlicht: (2024)
von: Jeon, Youngseung, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Iffy-Or-Not: Extending the Web to Support the Critical Evaluation of Fallacious Texts
von: Lim, Gionnieve, et al.
Veröffentlicht: (2025) -
WebAccessVL: Violation-Aware VLM for Web Accessibility
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2025) -
Benchmarking Large Language Models for Diagnosing Students' Cognitive Skills from Handwritten Math Work
von: Kim, Yoonsu, et al.
Veröffentlicht: (2025) -
ChoiceMates: Supporting Unfamiliar Online Decision-Making with Multi-Agent Conversational Interactions
von: Park, Jeongeon, et al.
Veröffentlicht: (2023) -
Contrastive Explanations That Anticipate Human Misconceptions Can Improve Human Decision-Making Skills
von: Buçinca, Zana, et al.
Veröffentlicht: (2024)