Evaluating Spatiotemporal Consistency in Automatically Generated Sewing Instructions
Fuente:
arXiv
Salvato in:
| Autori principali: | Geiger, Luisa, Hartmann, Mareike, Sullivan, Michael, Koller, Alexander |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Survey on Complex Tasks for Goal-Directed Interactive Agents
di: Hartmann, Mareike, et al.
Pubblicazione: (2024)
di: Hartmann, Mareike, et al.
Pubblicazione: (2024)
Procedural Environment Generation for Tool-Use Agents
di: Sullivan, Michael, et al.
Pubblicazione: (2025)
di: Sullivan, Michael, et al.
Pubblicazione: (2025)
Fine-grained Controllable Text Generation through In-context Learning with Feedback
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2024)
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2024)
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2026)
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2026)
Recent Advancements and Challenges of Turkic Central Asian Language Processing
di: Veitsman, Yana, et al.
Pubblicazione: (2024)
di: Veitsman, Yana, et al.
Pubblicazione: (2024)
ADaPT: As-Needed Decomposition and Planning with Language Models
di: Prasad, Archiki, et al.
Pubblicazione: (2023)
di: Prasad, Archiki, et al.
Pubblicazione: (2023)
Simple and effective data augmentation for compositional generalization
di: Yao, Yuekun, et al.
Pubblicazione: (2024)
di: Yao, Yuekun, et al.
Pubblicazione: (2024)
Predicting generalization performance with correctness discriminators
di: Yao, Yuekun, et al.
Pubblicazione: (2023)
di: Yao, Yuekun, et al.
Pubblicazione: (2023)
Towards Adaptable and Interactive Image Captioning with Data Augmentation and Episodic Memory
di: Anagnostopoulou, Aliki, et al.
Pubblicazione: (2023)
di: Anagnostopoulou, Aliki, et al.
Pubblicazione: (2023)
Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters
di: Gurgurov, Daniil, et al.
Pubblicazione: (2024)
di: Gurgurov, Daniil, et al.
Pubblicazione: (2024)
LLMs Generate Kitsch
di: Klinge, Xenia, et al.
Pubblicazione: (2026)
di: Klinge, Xenia, et al.
Pubblicazione: (2026)
LLMs syntactically adapt their language use to their conversational partner
di: Kandra, Florian, et al.
Pubblicazione: (2025)
di: Kandra, Florian, et al.
Pubblicazione: (2025)
Collaborative Problem-Solving in an Optimization Game
di: Jeknic, Isidora, et al.
Pubblicazione: (2025)
di: Jeknic, Isidora, et al.
Pubblicazione: (2025)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
di: Lindemann, Matthias, et al.
Pubblicazione: (2023)
di: Lindemann, Matthias, et al.
Pubblicazione: (2023)
A Dialogue Game for Eliciting Balanced Collaboration
di: Jeknić, Isidora, et al.
Pubblicazione: (2024)
di: Jeknić, Isidora, et al.
Pubblicazione: (2024)
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
di: Lindemann, Matthias, et al.
Pubblicazione: (2024)
di: Lindemann, Matthias, et al.
Pubblicazione: (2024)
Automating the Generation of Prompts for LLM-based Action Choice in PDDL Planning
di: Stein, Katharina, et al.
Pubblicazione: (2023)
di: Stein, Katharina, et al.
Pubblicazione: (2023)
How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs?
di: Doostmohammadi, Ehsan, et al.
Pubblicazione: (2024)
di: Doostmohammadi, Ehsan, et al.
Pubblicazione: (2024)
Evaluating Diversity in Automatic Poetry Generation
di: Chen, Yanran, et al.
Pubblicazione: (2024)
di: Chen, Yanran, et al.
Pubblicazione: (2024)
Automatic Answerability Evaluation for Question Generation
di: Wang, Zifan, et al.
Pubblicazione: (2023)
di: Wang, Zifan, et al.
Pubblicazione: (2023)
Less is More for Improving Automatic Evaluation of Factual Consistency
di: Wang, Tong, et al.
Pubblicazione: (2024)
di: Wang, Tong, et al.
Pubblicazione: (2024)
Improved Generalized Planning with LLMs through Strategy Refinement and Reflection
di: Stein, Katharina, et al.
Pubblicazione: (2025)
di: Stein, Katharina, et al.
Pubblicazione: (2025)
Language models can learn implicit multi-hop reasoning, but only if they have lots of training data
di: Yao, Yuekun, et al.
Pubblicazione: (2025)
di: Yao, Yuekun, et al.
Pubblicazione: (2025)
PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
di: Wang, Yidong, et al.
Pubblicazione: (2023)
di: Wang, Yidong, et al.
Pubblicazione: (2023)
AIR: Complex Instruction Generation via Automatic Iterative Refinement
di: Liu, Wei, et al.
Pubblicazione: (2025)
di: Liu, Wei, et al.
Pubblicazione: (2025)
Positional Biases Shift as Inputs Approach Context Window Limits
di: Veseli, Blerta, et al.
Pubblicazione: (2025)
di: Veseli, Blerta, et al.
Pubblicazione: (2025)
Scope-enhanced Compositional Semantic Parsing for DRT
di: Yang, Xiulin, et al.
Pubblicazione: (2024)
di: Yang, Xiulin, et al.
Pubblicazione: (2024)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
di: Liu, Yilun, et al.
Pubblicazione: (2023)
di: Liu, Yilun, et al.
Pubblicazione: (2023)
Are Checklists Really Useful for Automatic Evaluation of Generative Tasks?
di: Furuhashi, Momoka, et al.
Pubblicazione: (2025)
di: Furuhashi, Momoka, et al.
Pubblicazione: (2025)
Does Instruction Tuning Make LLMs More Consistent?
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
di: Ryan, Michael J., et al.
Pubblicazione: (2025)
di: Ryan, Michael J., et al.
Pubblicazione: (2025)
Barriers to Universal Reasoning With Transformers (And How to Overcome Them)
di: Kraus, Oliver, et al.
Pubblicazione: (2026)
di: Kraus, Oliver, et al.
Pubblicazione: (2026)
Exploring Format Consistency for Instruction Tuning
di: Liang, Shihao, et al.
Pubblicazione: (2023)
di: Liang, Shihao, et al.
Pubblicazione: (2023)
Evaluating the Consistency of LLM Evaluators
di: Lee, Noah, et al.
Pubblicazione: (2024)
di: Lee, Noah, et al.
Pubblicazione: (2024)
D-GEN: Automatic Distractor Generation and Evaluation for Reliable Assessment of Generative Model
di: Byun, Grace, et al.
Pubblicazione: (2025)
di: Byun, Grace, et al.
Pubblicazione: (2025)
Automatic Instruction Evolving for Large Language Models
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
di: Bañeras-Roux, Thibault, et al.
Pubblicazione: (2026)
di: Bañeras-Roux, Thibault, et al.
Pubblicazione: (2026)
Improving the Validity of Automatically Generated Feedback via Reinforcement Learning
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
di: Yang, Yixin, et al.
Pubblicazione: (2025)
di: Yang, Yixin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Survey on Complex Tasks for Goal-Directed Interactive Agents
di: Hartmann, Mareike, et al.
Pubblicazione: (2024) -
Procedural Environment Generation for Tool-Use Agents
di: Sullivan, Michael, et al.
Pubblicazione: (2025) -
Fine-grained Controllable Text Generation through In-context Learning with Feedback
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2024) -
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2026) -
Recent Advancements and Challenges of Turkic Central Asian Language Processing
di: Veitsman, Yana, et al.
Pubblicazione: (2024)