APEX-SWE
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kottamasu, Abhi, Mahapatra, Chirag, Lee, Sam, Pan, Ben, Barthwal, Aakash, Datta, Akul, Gupta, Anurag, Mehta, Pranav, Arun, Ajay, Alberti, Silas, Hiremath, Adarsh, Foody, Brendan, Vidgen, Bertie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The AI Productivity Index (APEX)
von: Vidgen, Bertie, et al.
Veröffentlicht: (2025)
von: Vidgen, Bertie, et al.
Veröffentlicht: (2025)
The AI Consumer Index (ACE)
von: Benchek, Julien, et al.
Veröffentlicht: (2025)
von: Benchek, Julien, et al.
Veröffentlicht: (2025)
APEX-Agents
von: Vidgen, Bertie, et al.
Veröffentlicht: (2026)
von: Vidgen, Bertie, et al.
Veröffentlicht: (2026)
The Evolution of RWKV: Advancements in Efficient Language Modeling
von: Datta, Akul
Veröffentlicht: (2024)
von: Datta, Akul
Veröffentlicht: (2024)
SafetyPrompts: a Systematic Review of Open Datasets for Evaluating and Improving Large Language Model Safety
von: Röttger, Paul, et al.
Veröffentlicht: (2024)
von: Röttger, Paul, et al.
Veröffentlicht: (2024)
Classification is a RAG problem: A case study on hate speech detection
von: Willats, Richard, et al.
Veröffentlicht: (2025)
von: Willats, Richard, et al.
Veröffentlicht: (2025)
WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting
von: Styles, Olly, et al.
Veröffentlicht: (2024)
von: Styles, Olly, et al.
Veröffentlicht: (2024)
Early Detection of Latent Microstructure Regimes in Limit Order Books
von: Hiremath, Prakul Sunil, et al.
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil, et al.
Veröffentlicht: (2026)
Why human-AI relationships need socioaffective alignment
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2025)
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2025)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
von: Barthwal, Ankur, et al.
Veröffentlicht: (2025)
von: Barthwal, Ankur, et al.
Veröffentlicht: (2025)
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
Evaluating Contextual Intelligence in Recyclability: A Comprehensive Study of Image-Based Reasoning Systems
von: Park, Eliot, et al.
Veröffentlicht: (2025)
von: Park, Eliot, et al.
Veröffentlicht: (2025)
Ethical AI for Young Digital Citizens: A Call to Action on Privacy Governance
von: Shouli, Austin, et al.
Veröffentlicht: (2025)
von: Shouli, Austin, et al.
Veröffentlicht: (2025)
Ethical AI for Young Digital Citizens: A Call to Action on Privacy Governance
von: Austin Shouli, et al.
Veröffentlicht: (2026)
von: Austin Shouli, et al.
Veröffentlicht: (2026)
BinSparX: Sparsified Binary Neural Networks for Reduced Hardware Non-Idealities in Xbar Arrays
von: Malhotra, Akul, et al.
Veröffentlicht: (2024)
von: Malhotra, Akul, et al.
Veröffentlicht: (2024)
Memory Faults in Activation-sparse Quantized Deep Neural Networks: Analysis and Mitigation using Sharpness-aware Training
von: Malhotra, Akul, et al.
Veröffentlicht: (2024)
von: Malhotra, Akul, et al.
Veröffentlicht: (2024)
Weight Transformations in Bit-Sliced Crossbar Arrays for Fault Tolerant Computing-in-Memory: Design Techniques and Evaluation Framework
von: Malhotra, Akul, et al.
Veröffentlicht: (2025)
von: Malhotra, Akul, et al.
Veröffentlicht: (2025)
ReTern: Exploiting Natural Redundancy and Sign Transformations for Enhanced Fault Tolerance in Compute-in-Memory based Ternary LLMs
von: Malhotra, Akul, et al.
Veröffentlicht: (2025)
von: Malhotra, Akul, et al.
Veröffentlicht: (2025)
PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2026)
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2026)
Neural steering vectors reveal dose and exposure-dependent impacts of human-AI relationships
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2025)
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2025)
SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
von: Vidgen, Bertie, et al.
Veröffentlicht: (2023)
von: Vidgen, Bertie, et al.
Veröffentlicht: (2023)
A hyperbolic model for two-layer thin film flow with a perfectly soluble anti-surfactant
von: Barthwal, Rahul, et al.
Veröffentlicht: (2025)
von: Barthwal, Rahul, et al.
Veröffentlicht: (2025)
Rage Music Classification and Analysis using K-Nearest Neighbour, Random Forest, Support Vector Machine, Convolutional Neural Networks, and Gradient Boosting
von: Kumar, Akul
Veröffentlicht: (2024)
von: Kumar, Akul
Veröffentlicht: (2024)
Investigation of the Privacy Concerns in AI Systems for Young Digital Citizens: A Comparative Stakeholder Analysis
von: Campbell, Molly, et al.
Veröffentlicht: (2025)
von: Campbell, Molly, et al.
Veröffentlicht: (2025)
Data Unlearning in Diffusion Models
von: Alberti, Silas, et al.
Veröffentlicht: (2025)
von: Alberti, Silas, et al.
Veröffentlicht: (2025)
PIGEON: Predicting Image Geolocations
von: Haas, Lukas, et al.
Veröffentlicht: (2023)
von: Haas, Lukas, et al.
Veröffentlicht: (2023)
Shipping route end to end data analysis
von: Sarkar, Abhi
Veröffentlicht: (2026)
von: Sarkar, Abhi
Veröffentlicht: (2026)
The Progression of Transformers from Language to Vision to MOT: A Literature Review on Multi-Object Tracking with Transformers
von: Kamboj, Abhi
Veröffentlicht: (2024)
von: Kamboj, Abhi
Veröffentlicht: (2024)
Dynamical Dark Energy from a Massive Vector Field in Generalized Proca Theory
von: Savaliya, Abhi
Veröffentlicht: (2025)
von: Savaliya, Abhi
Veröffentlicht: (2025)
Charged-particle production in pp collisions at $\sqrt{s}$ = 13.6 TeV and Pb$-$Pb collisions at $\sqrt{s_{\mathrm{NN}}}$ = 5.36 TeV with ALICE
von: Modak, Abhi
Veröffentlicht: (2024)
von: Modak, Abhi
Veröffentlicht: (2024)
Enhancing Inventory Management with Progressive Web Applications (PWAs): A Scalable Solution for Small and Large Enterprises
von: Desai, Abhi
Veröffentlicht: (2025)
von: Desai, Abhi
Veröffentlicht: (2025)
Recent ALICE results from light-ion collision systems
von: Modak, Abhi
Veröffentlicht: (2026)
von: Modak, Abhi
Veröffentlicht: (2026)
Architectural Isolation as a Timing Safety Primitive for Edge AI Medical Devices: Controlled Experimental Evidence on a Shared-Silicon Platform
von: Swami, Akul Mallayya
Veröffentlicht: (2026)
von: Swami, Akul Mallayya
Veröffentlicht: (2026)
Data to Decisions: A Computational Framework to Identify skill requirements from Advertorial Data
von: Singh, Aakash, et al.
Veröffentlicht: (2025)
von: Singh, Aakash, et al.
Veröffentlicht: (2025)
APEX pretrained models
von: Jiménez Castro, Lucía
Veröffentlicht: (2026)
von: Jiménez Castro, Lucía
Veröffentlicht: (2026)
Partial tidal disruption of White Dwarfs in off-equatorial orbits around Kerr black holes
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2024)
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2024)
Tidal disruptions of close white dwarf binaries by intermediate mass black holes
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2025)
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2025)
Tidal encounters of close white dwarf binaries with spinning black holes
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2026)
von: Mahapatra, Aryabrat, et al.
Veröffentlicht: (2026)
Understanding visual attention beehind bee-inspired UAV navigation
von: Rajbhandari, Pranav, et al.
Veröffentlicht: (2025)
von: Rajbhandari, Pranav, et al.
Veröffentlicht: (2025)
Sensitivity Uncertainty Alignment in Large Language Models
von: Hiremath, Prakul Sunil, et al.
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The AI Productivity Index (APEX)
von: Vidgen, Bertie, et al.
Veröffentlicht: (2025) -
The AI Consumer Index (ACE)
von: Benchek, Julien, et al.
Veröffentlicht: (2025) -
APEX-Agents
von: Vidgen, Bertie, et al.
Veröffentlicht: (2026) -
The Evolution of RWKV: Advancements in Efficient Language Modeling
von: Datta, Akul
Veröffentlicht: (2024) -
SafetyPrompts: a Systematic Review of Open Datasets for Evaluating and Improving Large Language Model Safety
von: Röttger, Paul, et al.
Veröffentlicht: (2024)