An LLM Agentic Approach for Legal-Critical Software: A Case Study for Tax Prep Software
Fuente:
arXiv
Saved in:
| Main Authors: | Gogani-Khiabani, Sina, Trivedi, Ashutosh, Saha, Diptikalyan, Tizpaz-Niari, Saeid |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Technical Challenges in Maintaining Tax Prep Software with Large Language Models
by: Gogani-Khiabani, Sina, et al.
Published: (2025)
by: Gogani-Khiabani, Sina, et al.
Published: (2025)
Metamorphic Debugging for Accountable Software
by: Tizpaz-Niari, Saeid, et al.
Published: (2024)
by: Tizpaz-Niari, Saeid, et al.
Published: (2024)
Predicting Fairness of ML Software Configurations
by: Herrera, Salvador Robles, et al.
Published: (2024)
by: Herrera, Salvador Robles, et al.
Published: (2024)
On the Potential and Limitations of Few-Shot In-Context Learning to Generate Metamorphic Specifications for Tax Preparation Software
by: Srinivas, Dananjay, et al.
Published: (2023)
by: Srinivas, Dananjay, et al.
Published: (2023)
Fairness Testing through Extreme Value Theory
by: Monjezi, Verya, et al.
Published: (2025)
by: Monjezi, Verya, et al.
Published: (2025)
Worst-Case Convergence Time of ML Algorithms via Extreme Value Theory
by: Tizpaz-Niari, Saeid, et al.
Published: (2024)
by: Tizpaz-Niari, Saeid, et al.
Published: (2024)
Uncovering Discrimination Clusters: Quantifying and Explaining Systematic Fairness Violations
by: Akash, Ranit Debnath, et al.
Published: (2025)
by: Akash, Ranit Debnath, et al.
Published: (2025)
On the Robustness of Fairness Practices: A Causal Framework for Systematic Evaluation
by: Monjezi, Verya, et al.
Published: (2026)
by: Monjezi, Verya, et al.
Published: (2026)
FairLay-ML: Intuitive Debugging of Fairness in Data-Driven Social-Critical Software
by: Yu, Normen, et al.
Published: (2024)
by: Yu, Normen, et al.
Published: (2024)
NeuFair: Neural Network Fairness Repair with Dropout
by: Dasu, Vishnu Asutosh, et al.
Published: (2024)
by: Dasu, Vishnu Asutosh, et al.
Published: (2024)
Querying Large Automotive Software Models: Agentic vs. Direct LLM Approaches
by: Mazur, Lukasz, et al.
Published: (2025)
by: Mazur, Lukasz, et al.
Published: (2025)
LLM-Based Agentic Systems for Software Engineering: Challenges and Opportunities
by: Tang, Yongjian, et al.
Published: (2026)
by: Tang, Yongjian, et al.
Published: (2026)
Reinforcement Learning Integrated Agentic RAG for Software Test Cases Authoring
by: Hariharan, Mohanakrishnan
Published: (2025)
by: Hariharan, Mohanakrishnan
Published: (2025)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
by: Casserini, Matteo, et al.
Published: (2026)
by: Casserini, Matteo, et al.
Published: (2026)
Agentic AI Software Engineers: Programming with Trust
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
Every Software as an Agent: Blueprint and Case Study
by: Xu, Mengwei
Published: (2025)
by: Xu, Mengwei
Published: (2025)
LLM-Powered Workflow Optimization for Multidisciplinary Software Development: An Automotive Industry Case Study
by: Wang, Shuai, et al.
Published: (2026)
by: Wang, Shuai, et al.
Published: (2026)
Agentic AI for Software: thoughts from Software Engineering community
by: Roychoudhury, Abhik
Published: (2025)
by: Roychoudhury, Abhik
Published: (2025)
Agentic Software Issue Resolution with Large Language Models: A Survey
by: Jiang, Zhonghao, et al.
Published: (2025)
by: Jiang, Zhonghao, et al.
Published: (2025)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
by: Li, Jingyue, et al.
Published: (2026)
by: Li, Jingyue, et al.
Published: (2026)
Agentic Software Engineering: Foundational Pillars and a Research Roadmap
by: Hassan, Ahmed E., et al.
Published: (2025)
by: Hassan, Ahmed E., et al.
Published: (2025)
Agentic AI in 6G Software Businesses: A Layered Maturity Model
by: Zohaib, Muhammad, et al.
Published: (2025)
by: Zohaib, Muhammad, et al.
Published: (2025)
Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering Studies
by: Angermeir, Florian, et al.
Published: (2025)
by: Angermeir, Florian, et al.
Published: (2025)
The Semi-Executable Stack: Agentic Software Engineering and the Expanding Scope of SE
by: Feldt, Robert, et al.
Published: (2026)
by: Feldt, Robert, et al.
Published: (2026)
Feedback Loops and Code Perturbations in LLM-based Software Engineering: A Case Study on a C-to-Rust Translation System
by: Weiss, Martin, et al.
Published: (2025)
by: Weiss, Martin, et al.
Published: (2025)
Toward an Agentic Infused Software Ecosystem
by: Marron, Mark
Published: (2026)
by: Marron, Mark
Published: (2026)
Beyond Human-Readable: Rethinking Software Engineering Conventions for the Agentic Development Era
by: Ustynov, Dmytro
Published: (2026)
by: Ustynov, Dmytro
Published: (2026)
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
by: Naqvi, Saba, et al.
Published: (2026)
by: Naqvi, Saba, et al.
Published: (2026)
Agentic RAG for Software Testing with Hybrid Vector-Graph and Multi-Agent Orchestration
by: Hariharan, Mohanakrishnan, et al.
Published: (2025)
by: Hariharan, Mohanakrishnan, et al.
Published: (2025)
Process-Centric Analysis of Agentic Software Systems
by: Liu, Shuyang, et al.
Published: (2025)
by: Liu, Shuyang, et al.
Published: (2025)
Can LLM Generate Regression Tests for Software Commits?
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
LLM Company Policies and Policy Implications in Software Organizations
by: Khojah, Ranim, et al.
Published: (2025)
by: Khojah, Ranim, et al.
Published: (2025)
RAILS: Retrieval-Augmented Intelligence for Learning Software Development
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
Quality-Driven Agentic Reasoning for LLM-Assisted Software Design: Questions-of-Thoughts (QoT) as a Time-Series Self-QA Chain
by: Liu, Yen-Ku, et al.
Published: (2026)
by: Liu, Yen-Ku, et al.
Published: (2026)
RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades
by: Xu, Xinbo, et al.
Published: (2026)
by: Xu, Xinbo, et al.
Published: (2026)
Automating a Complete Software Test Process Using LLMs: An Automotive Case Study
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Scalable and Efficient Large-Scale Log Analysis with LLMs: An IT Software Support Case Study
by: Gupta, Pranjal, et al.
Published: (2025)
by: Gupta, Pranjal, et al.
Published: (2025)
Multi-modal Summarization in Model-Based Engineering: Automotive Software Development Case Study
by: Petrovic, Nenad, et al.
Published: (2025)
by: Petrovic, Nenad, et al.
Published: (2025)
Orchestrating Human-AI Software Delivery: A Retrospective Longitudinal Field Study of Three Software Modernization Programs
by: Armesto, Maximiliano, et al.
Published: (2026)
by: Armesto, Maximiliano, et al.
Published: (2026)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
by: Wang, Ruiqi, et al.
Published: (2025)
by: Wang, Ruiqi, et al.
Published: (2025)
Similar Items
-
Technical Challenges in Maintaining Tax Prep Software with Large Language Models
by: Gogani-Khiabani, Sina, et al.
Published: (2025) -
Metamorphic Debugging for Accountable Software
by: Tizpaz-Niari, Saeid, et al.
Published: (2024) -
Predicting Fairness of ML Software Configurations
by: Herrera, Salvador Robles, et al.
Published: (2024) -
On the Potential and Limitations of Few-Shot In-Context Learning to Generate Metamorphic Specifications for Tax Preparation Software
by: Srinivas, Dananjay, et al.
Published: (2023) -
Fairness Testing through Extreme Value Theory
by: Monjezi, Verya, et al.
Published: (2025)