Remote Labor Index: Measuring AI Automation of Remote Work
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mazeika, Mantas, Gatti, Alice, Menghini, Cristina, Sehwag, Udari Madhushani, Singhal, Shivam, Orlovskiy, Yury, Basart, Steven, Sharma, Manasi, Peskoff, Denis, Lau, Elaine, Lim, Jaehyuk, Carroll, Lachlan, Blair, Alice, Sivakumar, Vinaya, Basu, Sumana, Kenstler, Brad, Ma, Yuntao, Michael, Julian, Li, Xiaoke, Ingebretsen, Oliver, Mehta, Aditya, Mottola, Jean, Teichmann, John, Yu, Kevin, Shaik, Zaina, Khoja, Adam, Ren, Richard, Hausenloy, Jason, Phan, Long, Htet, Ye, Aich, Ankit, Rabbani, Tahseen, Shah, Vivswan, Novykov, Andriy, Binder, Felix, Chugunov, Kirill, Ramirez, Luis, Geralnik, Matias, Mesura, Hernán, Lee, Dean, Cardona, Ed-Yeremai Hernandez, Diamond, Annette, Yue, Summer, Wang, Alexandr, Liu, Bing, Hernandez, Ernesto, Hendrycks, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
von: Ren, Richard, et al.
Veröffentlicht: (2025)
von: Ren, Richard, et al.
Veröffentlicht: (2025)
In-Context Learning with Topological Information for Knowledge Graph Completion
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024)
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024)
Can LLMs be Scammed? A Baseline Measurement Study
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?
von: Ren, Richard, et al.
Veröffentlicht: (2024)
von: Ren, Richard, et al.
Veröffentlicht: (2024)
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
Reducing Political Manipulation with Consistency Training
von: Phan, Long, et al.
Veröffentlicht: (2026)
von: Phan, Long, et al.
Veröffentlicht: (2026)
Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs
von: Mazeika, Mantas, et al.
Veröffentlicht: (2025)
von: Mazeika, Mantas, et al.
Veröffentlicht: (2025)
Continual Learning of Domain Knowledge from Human Feedback in Text-to-SQL
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
TextQuests: How Good are LLMs at Text-Based Video Games?
von: Phan, Long, et al.
Veröffentlicht: (2025)
von: Phan, Long, et al.
Veröffentlicht: (2025)
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
Defensive Refusal Bias: How Safety Alignment Fails Cyber Defenders
von: Campbell, David, et al.
Veröffentlicht: (2026)
von: Campbell, David, et al.
Veröffentlicht: (2026)
Aggressive Compression Enables LLM Weight Theft
von: Brown, Davis, et al.
Veröffentlicht: (2026)
von: Brown, Davis, et al.
Veröffentlicht: (2026)
LHAW: Controllable Underspecification for Long-Horizon Tasks
von: Pu, George, et al.
Veröffentlicht: (2026)
von: Pu, George, et al.
Veröffentlicht: (2026)
EnigmaEval: A Benchmark of Long Multimodal Reasoning Challenges
von: Wang, Clinton J., et al.
Veröffentlicht: (2025)
von: Wang, Clinton J., et al.
Veröffentlicht: (2025)
Leveraging Continuously Differentiable Activation Functions for Learning in Quantized Noisy Environments
von: Shah, Vivswan, et al.
Veröffentlicht: (2024)
von: Shah, Vivswan, et al.
Veröffentlicht: (2024)
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
von: Mazeika, Mantas, et al.
Veröffentlicht: (2024)
von: Mazeika, Mantas, et al.
Veröffentlicht: (2024)
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
Community Rehabilitation for Rural and Remote Australia: Measuring What Matters Based on the International Classification of Functioning, Disability and Health (ICF): A Scoping Review
von: Alice Cairns, et al.
Veröffentlicht: (2025)
von: Alice Cairns, et al.
Veröffentlicht: (2025)
Opportunities and Challenges of Frontier Data Governance With Synthetic Data
von: Thakur, Madhavendra, et al.
Veröffentlicht: (2025)
von: Thakur, Madhavendra, et al.
Veröffentlicht: (2025)
Neural Field Representations of Mobile Computational Photography
von: Chugunov, Ilya
Veröffentlicht: (2025)
von: Chugunov, Ilya
Veröffentlicht: (2025)
O3D: Offline Data-driven Discovery and Distillation for Sequential Decision-Making with Large Language Models
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
Catulo, c. 59. El castigo de Rufa
von: Emilio Zaina
Veröffentlicht: (2016)
von: Emilio Zaina
Veröffentlicht: (2016)
Conocimiento y método en Descartes, Pascal y Leibniz
von: Josep M. Basart Muñoz
Veröffentlicht: (2004)
von: Josep M. Basart Muñoz
Veröffentlicht: (2004)
Private equity acquisitions and product market decisions: Evidence from trademarks
von: Moazzam Khoja
Veröffentlicht: (2025)
von: Moazzam Khoja
Veröffentlicht: (2025)
Introduction to AI Safety, Ethics, and Society
von: Hendrycks, Dan
Veröffentlicht: (2024)
von: Hendrycks, Dan
Veröffentlicht: (2024)
Introduction to AI Safety, Ethics, and Society
von: Hendrycks, Dan
Veröffentlicht: (2024)
von: Hendrycks, Dan
Veröffentlicht: (2024)
The public sector in the Caribbean : issues and reform options / Vinaya Swaroop
von: Swaroop, Vinaya
Veröffentlicht: (1996)
von: Swaroop, Vinaya
Veröffentlicht: (1996)
Gravitational Vacuum Condensate Stars in the Effective Theory of Gravity
von: Mottola, Emil
Veröffentlicht: (2025)
von: Mottola, Emil
Veröffentlicht: (2025)
BUBELLO, Juan Pablo Historia del esoterismo en Argentina. Prácticas, representaciones y persecuciones de curanderos, espiritistas, astrólogos y otros esoteristas, Editorial Biblos, Buenos Aires, 2010, 285 pp - ISBN 978-950-786-809-2
von: Marcelo Móttola
Veröffentlicht: (2011)
von: Marcelo Móttola
Veröffentlicht: (2011)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2025)
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2025)
Pini, M., Más Rocha, S., Gorostiaga, J., Tello, C., y Asprella, G. (coords.) La Educación Secundaria, ¿Modelo en (re)construcción?. Buenos Aires: Editorial Aique, pags. 252.
von: Raúl A. Menghini
Veröffentlicht: (2016)
von: Raúl A. Menghini
Veröffentlicht: (2016)
Presentación
von: Raúl A. Menghini
Veröffentlicht: (2022)
von: Raúl A. Menghini
Veröffentlicht: (2022)
Exact simulation scheme for the Ornstein-Uhlenbeck driven stochastic volatility model with the Karhunen-Loève expansions
von: Choi, Jaehyuk
Veröffentlicht: (2024)
von: Choi, Jaehyuk
Veröffentlicht: (2024)
Forging the Ideal Educated Girl (Volume 1.0)
von: Khoja-Moolji, Shenila
Veröffentlicht: (2020)
von: Khoja-Moolji, Shenila
Veröffentlicht: (2020)
Forging the Ideal Educated Girl
von: Khoja-Moolji, Shenila
Veröffentlicht: (2018)
von: Khoja-Moolji, Shenila
Veröffentlicht: (2018)
Ähnliche Einträge
-
The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
von: Ren, Richard, et al.
Veröffentlicht: (2025) -
In-Context Learning with Topological Information for Knowledge Graph Completion
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024) -
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024) -
Can LLMs be Scammed? A Baseline Measurement Study
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024) -
Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?
von: Ren, Richard, et al.
Veröffentlicht: (2024)