Automatic Generation of Behavioral Test Cases For Natural Language Processing Using Clustering and Prompting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Ying, Singh, Rahul, Joshi, Tarun, Sudjianto, Agus |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Skill Learning Using Process Mining for Large Language Model Plan Generation
von: Redis, Andrei Cosmin, et al.
Veröffentlicht: (2024)
von: Redis, Andrei Cosmin, et al.
Veröffentlicht: (2024)
InvestAlign: Overcoming Data Scarcity in Aligning Large Language Models with Investor Decision-Making Processes under Herd Behavior
von: Wang, Huisheng, et al.
Veröffentlicht: (2025)
von: Wang, Huisheng, et al.
Veröffentlicht: (2025)
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
Combinatorial Reasoning: Selecting Reasons in Generative AI Pipelines via Combinatorial Optimization
von: Esencan, Mert, et al.
Veröffentlicht: (2024)
von: Esencan, Mert, et al.
Veröffentlicht: (2024)
Embedding-Aligned Language Models
von: Tennenholtz, Guy, et al.
Veröffentlicht: (2024)
von: Tennenholtz, Guy, et al.
Veröffentlicht: (2024)
Iterative Resolution of Prompt Ambiguities Using a Progressive Cutting-Search Approach
von: Marozzo, Fabrizio
Veröffentlicht: (2025)
von: Marozzo, Fabrizio
Veröffentlicht: (2025)
ProcessTBench: An LLM Plan Generation Dataset for Process Mining
von: Redis, Andrei Cosmin, et al.
Veröffentlicht: (2024)
von: Redis, Andrei Cosmin, et al.
Veröffentlicht: (2024)
LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
Few-Shot Testing: Estimating Uncertainty of Memristive Deep Neural Networks Using One Bayesian Test Vector
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
G-Zero: Self-Play for Open-Ended Generation from Zero Data
von: Huang, Chengsong, et al.
Veröffentlicht: (2026)
von: Huang, Chengsong, et al.
Veröffentlicht: (2026)
Auditing an Automatic Grading Model with deep Reinforcement Learning
von: Condor, Aubrey, et al.
Veröffentlicht: (2024)
von: Condor, Aubrey, et al.
Veröffentlicht: (2024)
Quantifying the Effectiveness of Student Organization Activities using Natural Language Processing
von: Taruc, Lyberius Ennio F., et al.
Veröffentlicht: (2024)
von: Taruc, Lyberius Ennio F., et al.
Veröffentlicht: (2024)
Evaluating Quantized Large Language Models for Code Generation on Low-Resource Language Benchmarks
von: Nyamsuren, Enkhbold
Veröffentlicht: (2024)
von: Nyamsuren, Enkhbold
Veröffentlicht: (2024)
On the Limitations of Compute Thresholds as a Governance Strategy
von: Hooker, Sara
Veröffentlicht: (2024)
von: Hooker, Sara
Veröffentlicht: (2024)
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
von: Gerstgrasser, Matthias, et al.
Veröffentlicht: (2024)
von: Gerstgrasser, Matthias, et al.
Veröffentlicht: (2024)
Toward Large Language Models as a Therapeutic Tool: Comparing Prompting Techniques to Improve GPT-Delivered Problem-Solving Therapy
von: Filienko, Daniil, et al.
Veröffentlicht: (2024)
von: Filienko, Daniil, et al.
Veröffentlicht: (2024)
Learning with Calibration: Exploring Test-Time Computing of Spatio-Temporal Forecasting
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
Self Distillation via Iterative Constructive Perturbations
von: Dave, Maheak, et al.
Veröffentlicht: (2025)
von: Dave, Maheak, et al.
Veröffentlicht: (2025)
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models
von: Ke, Shih-Wen, et al.
Veröffentlicht: (2025)
von: Ke, Shih-Wen, et al.
Veröffentlicht: (2025)
NEURODNAAI: Neural pipeline approaches for the advancing dna-based information storage as a sustainable digital medium using deep learning framework
von: Thakur, Rakesh, et al.
Veröffentlicht: (2025)
von: Thakur, Rakesh, et al.
Veröffentlicht: (2025)
Concurrent Self-testing of Neural Networks Using Uncertainty Fingerprint
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
A Survey of Multimodal Retrieval-Augmented Generation
von: Mei, Lang, et al.
Veröffentlicht: (2025)
von: Mei, Lang, et al.
Veröffentlicht: (2025)
xai_evals : A Framework for Evaluating Post-Hoc Local Explanation Methods
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
Scale-Dropout: Estimating Uncertainty in Deep Neural Networks Using Stochastic Scale
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2023)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2023)
CIRCUITSYNTH: Leveraging Large Language Models for Circuit Topology Synthesis
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2024)
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2024)
MedCodER: A Generative AI Assistant for Medical Coding
von: Baksi, Krishanu Das, et al.
Veröffentlicht: (2024)
von: Baksi, Krishanu Das, et al.
Veröffentlicht: (2024)
The Roles of Generative Artificial Intelligence in Internet of Electric Vehicles
von: Zhang, Hanwen, et al.
Veröffentlicht: (2024)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2024)
Self-Supervised Learning for Time Series: Contrastive or Generative?
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
Human-Calibrated Automated Testing and Validation of Generative Language Models
von: Sudjianto, Agus, et al.
Veröffentlicht: (2024)
von: Sudjianto, Agus, et al.
Veröffentlicht: (2024)
Road User Classification from High-Frequency GNSS Data Using Distributed Edge Intelligence
von: Köpper, Lennart, et al.
Veröffentlicht: (2024)
von: Köpper, Lennart, et al.
Veröffentlicht: (2024)
Nonlinear Processing with Linear Optics
von: Yildirim, Mustafa, et al.
Veröffentlicht: (2023)
von: Yildirim, Mustafa, et al.
Veröffentlicht: (2023)
Virtual Sensor for Real-Time Bearing Load Prediction Using Heterogeneous Temporal Graph Neural Networks
von: Zhao, Mengjie, et al.
Veröffentlicht: (2024)
von: Zhao, Mengjie, et al.
Veröffentlicht: (2024)
Large-Language Memorization During the Classification of United States Supreme Court Cases
von: Ortega, John E., et al.
Veröffentlicht: (2025)
von: Ortega, John E., et al.
Veröffentlicht: (2025)
A Unified Generative-AI Framework for Smart Energy Infrastructure: Intelligent Gas Distribution, Utility Billing, Carbon Analytics, and Quantum-Inspired Optimisation
von: Manjunath, Pavan, et al.
Veröffentlicht: (2026)
von: Manjunath, Pavan, et al.
Veröffentlicht: (2026)
Biothreat Benchmark Generation Framework for Evaluating Frontier AI Models II: Benchmark Generation Process
von: Ackerman, Gary, et al.
Veröffentlicht: (2025)
von: Ackerman, Gary, et al.
Veröffentlicht: (2025)
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
von: Mohanty, Anwesha, et al.
Veröffentlicht: (2025)
von: Mohanty, Anwesha, et al.
Veröffentlicht: (2025)
LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems
von: Moia, Vitor Hugo Galhardo, et al.
Veröffentlicht: (2025)
von: Moia, Vitor Hugo Galhardo, et al.
Veröffentlicht: (2025)
Agentic AI framework for End-to-End Medical Data Inference
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
Which English Do LLMs Prefer? Triangulating Structural Bias Towards American English in Foundation Models
von: Nayeem, Mir Tafseer, et al.
Veröffentlicht: (2026)
von: Nayeem, Mir Tafseer, et al.
Veröffentlicht: (2026)
The Compliance Paradox: Semantic-Instruction Decoupling in Automated Academic Code Evaluation
von: Sahoo, Devanshu, et al.
Veröffentlicht: (2026)
von: Sahoo, Devanshu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Skill Learning Using Process Mining for Large Language Model Plan Generation
von: Redis, Andrei Cosmin, et al.
Veröffentlicht: (2024) -
InvestAlign: Overcoming Data Scarcity in Aligning Large Language Models with Investor Decision-Making Processes under Herd Behavior
von: Wang, Huisheng, et al.
Veröffentlicht: (2025) -
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
von: Liu, Qibang, et al.
Veröffentlicht: (2025) -
Combinatorial Reasoning: Selecting Reasons in Generative AI Pipelines via Combinatorial Optimization
von: Esencan, Mert, et al.
Veröffentlicht: (2024) -
Embedding-Aligned Language Models
von: Tennenholtz, Guy, et al.
Veröffentlicht: (2024)