Action Controlled Paraphrasing
Fuente:
arXiv
Salvato in:
| Autori principali: | Shi, Ning, Wu, Zijun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
di: Kaneko, Masahiro
Pubblicazione: (2026)
di: Kaneko, Masahiro
Pubblicazione: (2026)
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
Parameter Efficient Diverse Paraphrase Generation Using Sequence-Level Knowledge Distillation
di: Jayawardena, Lasal, et al.
Pubblicazione: (2024)
di: Jayawardena, Lasal, et al.
Pubblicazione: (2024)
Analyzing Persuasive Strategies in Meme Texts: A Fusion of Language Models with Paraphrase Enrichment
di: Nayak, Kota Shamanth Ramanath, et al.
Pubblicazione: (2024)
di: Nayak, Kota Shamanth Ramanath, et al.
Pubblicazione: (2024)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
di: Jung, Jaehun, et al.
Pubblicazione: (2023)
di: Jung, Jaehun, et al.
Pubblicazione: (2023)
ReCode: Unify Plan and Action for Universal Granularity Control
di: Yu, Zhaoyang, et al.
Pubblicazione: (2025)
di: Yu, Zhaoyang, et al.
Pubblicazione: (2025)
ControlMath: Controllable Data Generation Promotes Math Generalist Models
di: Chen, Nuo, et al.
Pubblicazione: (2024)
di: Chen, Nuo, et al.
Pubblicazione: (2024)
AliMark: Enhancing Robustness of Sentence-Level Watermarking Against Text Paraphrasing
di: Li, Yuexin, et al.
Pubblicazione: (2026)
di: Li, Yuexin, et al.
Pubblicazione: (2026)
AMPS: ASR with Multimodal Paraphrase Supervision
di: Gupta, Abhishek, et al.
Pubblicazione: (2024)
di: Gupta, Abhishek, et al.
Pubblicazione: (2024)
ParaFusion: A Large-Scale LLM-Driven English Paraphrase Dataset Infused with High-Quality Lexical and Syntactic Diversity
di: Jayawardena, Lasal, et al.
Pubblicazione: (2024)
di: Jayawardena, Lasal, et al.
Pubblicazione: (2024)
Internalizing LLM Reasoning via Discovery and Replay of Latent Actions
di: Shi, Zhenning, et al.
Pubblicazione: (2026)
di: Shi, Zhenning, et al.
Pubblicazione: (2026)
Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions
di: Li, Yongqi, et al.
Pubblicazione: (2026)
di: Li, Yongqi, et al.
Pubblicazione: (2026)
Branch-and-Browse: Efficient and Controllable Web Exploration with Tree-Structured Reasoning and Action Memory
di: He, Shiqi, et al.
Pubblicazione: (2025)
di: He, Shiqi, et al.
Pubblicazione: (2025)
Diagnosing Training Inference Mismatch in LLM Reinforcement Learning
di: Zhong, Tianle, et al.
Pubblicazione: (2026)
di: Zhong, Tianle, et al.
Pubblicazione: (2026)
Inference-Time Scaling for Generalist Reward Modeling
di: Liu, Zijun, et al.
Pubblicazione: (2025)
di: Liu, Zijun, et al.
Pubblicazione: (2025)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
di: Chen, Jianhui, et al.
Pubblicazione: (2024)
di: Chen, Jianhui, et al.
Pubblicazione: (2024)
Wonderful Matrices: Combining for a More Efficient and Effective Foundation Model Architecture
di: Shi, Jingze, et al.
Pubblicazione: (2024)
di: Shi, Jingze, et al.
Pubblicazione: (2024)
Large Language Models for Controllable Multi-property Multi-objective Molecule Optimization
di: Dey, Vishal, et al.
Pubblicazione: (2025)
di: Dey, Vishal, et al.
Pubblicazione: (2025)
Self-Improving World Modelling with Latent Actions
di: Qiu, Yifu, et al.
Pubblicazione: (2026)
di: Qiu, Yifu, et al.
Pubblicazione: (2026)
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking
di: Liu, Zijun, et al.
Pubblicazione: (2024)
di: Liu, Zijun, et al.
Pubblicazione: (2024)
Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders
di: Jing, Yi, et al.
Pubblicazione: (2026)
di: Jing, Yi, et al.
Pubblicazione: (2026)
Text2Data: Low-Resource Data Generation with Textual Control
di: Wang, Shiyu, et al.
Pubblicazione: (2024)
di: Wang, Shiyu, et al.
Pubblicazione: (2024)
Bone Soups: A Seek-and-Soup Model Merging Approach for Controllable Multi-Objective Generation
di: Xie, Guofu, et al.
Pubblicazione: (2025)
di: Xie, Guofu, et al.
Pubblicazione: (2025)
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL
di: Yang, Rui, et al.
Pubblicazione: (2026)
di: Yang, Rui, et al.
Pubblicazione: (2026)
AIGS: Generating Science from AI-Powered Automated Falsification
di: Liu, Zijun, et al.
Pubblicazione: (2024)
di: Liu, Zijun, et al.
Pubblicazione: (2024)
Learning to Reason as Action Abstractions with Scalable Mid-Training RL
di: Zhang, Shenao, et al.
Pubblicazione: (2025)
di: Zhang, Shenao, et al.
Pubblicazione: (2025)
Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective
di: Chen, Kun, et al.
Pubblicazione: (2026)
di: Chen, Kun, et al.
Pubblicazione: (2026)
Interpreting and Controlling LLM Reasoning through Integrated Policy Gradient
di: Li, Changming, et al.
Pubblicazione: (2026)
di: Li, Changming, et al.
Pubblicazione: (2026)
Recall-Extend Dynamics: Enhancing Small Language Models through Controlled Exploration and Refined Offline Integration
di: Guan, Zhong, et al.
Pubblicazione: (2025)
di: Guan, Zhong, et al.
Pubblicazione: (2025)
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback
di: Hoang, Thai, et al.
Pubblicazione: (2025)
di: Hoang, Thai, et al.
Pubblicazione: (2025)
ASPERA: A Simulated Environment to Evaluate Planning for Complex Action Execution
di: Coca, Alexandru, et al.
Pubblicazione: (2025)
di: Coca, Alexandru, et al.
Pubblicazione: (2025)
Preference Poisoning Attacks on Reward Model Learning
di: Wu, Junlin, et al.
Pubblicazione: (2024)
di: Wu, Junlin, et al.
Pubblicazione: (2024)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
Effectively Controlling Reasoning Models through Thinking Intervention
di: Wu, Tong, et al.
Pubblicazione: (2025)
di: Wu, Tong, et al.
Pubblicazione: (2025)
From Words to Actions: Unveiling the Theoretical Underpinnings of LLM-Driven Autonomous Systems
di: He, Jianliang, et al.
Pubblicazione: (2024)
di: He, Jianliang, et al.
Pubblicazione: (2024)
Learning to Clarify: Multi-turn Conversations with Action-Based Contrastive Self-Training
di: Chen, Maximillian, et al.
Pubblicazione: (2024)
di: Chen, Maximillian, et al.
Pubblicazione: (2024)
WebOperator: Action-Aware Tree Search for Autonomous Agents in Web Environment
di: Dihan, Mahir Labib, et al.
Pubblicazione: (2025)
di: Dihan, Mahir Labib, et al.
Pubblicazione: (2025)
UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches
di: Wang, Chao, et al.
Pubblicazione: (2024)
di: Wang, Chao, et al.
Pubblicazione: (2024)
Mixture of Heterogeneous Grouped Experts for Language Modeling
di: Ma, Zhicheng, et al.
Pubblicazione: (2026)
di: Ma, Zhicheng, et al.
Pubblicazione: (2026)
LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
di: Kaneko, Masahiro
Pubblicazione: (2026) -
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
di: Yadav, Vikas, et al.
Pubblicazione: (2024) -
Parameter Efficient Diverse Paraphrase Generation Using Sequence-Level Knowledge Distillation
di: Jayawardena, Lasal, et al.
Pubblicazione: (2024) -
Analyzing Persuasive Strategies in Meme Texts: A Fusion of Language Models with Paraphrase Enrichment
di: Nayak, Kota Shamanth Ramanath, et al.
Pubblicazione: (2024) -
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
di: Jung, Jaehun, et al.
Pubblicazione: (2023)