What if I ask in \textit{alia lingua}? Measuring Functional Similarity Across Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mishra, Debangan, Rastogi, Arihant, Negi, Agyeya, Goel, Shashwat, Kumaraguru, Ponnurangam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Cognac Shot To Forget Bad Memories: Corrective Unlearning for Graph Neural Networks
von: Kolipaka, Varshita, et al.
Veröffentlicht: (2024)
von: Kolipaka, Varshita, et al.
Veröffentlicht: (2024)
Measuring Moral Inconsistencies in Large Language Models
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
Representation Surgery: Theory and Practice of Affine Steering
von: Singh, Shashwat, et al.
Veröffentlicht: (2024)
von: Singh, Shashwat, et al.
Veröffentlicht: (2024)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
Great Models Think Alike and this Undermines AI Oversight
von: Goel, Shashwat, et al.
Veröffentlicht: (2025)
von: Goel, Shashwat, et al.
Veröffentlicht: (2025)
Corrective Machine Unlearning
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
Communicating about Space: Language-Mediated Spatial Integration Across Partial Views
von: Sikarwar, Ankur, et al.
Veröffentlicht: (2026)
von: Sikarwar, Ankur, et al.
Veröffentlicht: (2026)
Enhancing AI Safety Through the Fusion of Low Rank Adapters
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2024)
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2024)
LLM Vocabulary Compression for Low-Compute Environments
von: Vennam, Sreeram, et al.
Veröffentlicht: (2024)
von: Vennam, Sreeram, et al.
Veröffentlicht: (2024)
Rethinking Thinking Tokens: Understanding Why They Underperform in Practice
von: Vennam, Sreeram, et al.
Veröffentlicht: (2024)
von: Vennam, Sreeram, et al.
Veröffentlicht: (2024)
Causal Reasoning Favors Encoders: On The Limits of Decoder-Only Models
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
Can Language Models Falsify? Evaluating Algorithmic Reasoning with Counterexample Creation
von: Sinha, Shiven, et al.
Veröffentlicht: (2025)
von: Sinha, Shiven, et al.
Veröffentlicht: (2025)
Long-context Non-factoid Question Answering in Indic Languages
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
Wu's Method can Boost Symbolic AI to Rival Silver Medalists and AlphaGeometry to Outperform Gold Medalists at IMO Geometry
von: Sinha, Shiven, et al.
Veröffentlicht: (2024)
von: Sinha, Shiven, et al.
Veröffentlicht: (2024)
Flying Pigs, FaR and Beyond: Evaluating LLM Reasoning in Counterfactual Worlds
von: Joishy, Anish R, et al.
Veröffentlicht: (2025)
von: Joishy, Anish R, et al.
Veröffentlicht: (2025)
SceneGraMMi: Scene Graph-boosted Hybrid-fusion for Multi-Modal Misinformation Veracity Prediction
von: Joshi, Swarang, et al.
Veröffentlicht: (2024)
von: Joshi, Swarang, et al.
Veröffentlicht: (2024)
Just KIDDIN: Knowledge Infusion and Distillation for Detection of INdecent Memes
von: Garg, Rahul, et al.
Veröffentlicht: (2024)
von: Garg, Rahul, et al.
Veröffentlicht: (2024)
HLDC: Hindi Legal Documents Corpus
von: Kapoor, Arnav, et al.
Veröffentlicht: (2022)
von: Kapoor, Arnav, et al.
Veröffentlicht: (2022)
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
Scaling Open-Ended Reasoning to Predict the Future
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
SPIRIT: Short-term Prediction of solar IRradIance for zero-shot Transfer learning using Foundation Models
von: Mishra, Aditya, et al.
Veröffentlicht: (2025)
von: Mishra, Aditya, et al.
Veröffentlicht: (2025)
Do LLMs Adhere to Label Definitions? Examining Their Receptivity to External Label Definitions
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
Multilingual Coreference Resolution in Low-resource South Asian Languages
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
Multi-Stage Training for Abusive Comment Detection in Indic Languages
von: Rastogi, Pranshu, et al.
Veröffentlicht: (2026)
von: Rastogi, Pranshu, et al.
Veröffentlicht: (2026)
Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models
von: Mammen, Priyanka Mary, et al.
Veröffentlicht: (2026)
von: Mammen, Priyanka Mary, et al.
Veröffentlicht: (2026)
Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs
von: Aneja, Krishak, et al.
Veröffentlicht: (2026)
von: Aneja, Krishak, et al.
Veröffentlicht: (2026)
Can LLMs $\textit{understand}$ Math? -- Exploring the Pitfalls in Mathematical Reasoning
von: Roy, Tiasa Singha, et al.
Veröffentlicht: (2025)
von: Roy, Tiasa Singha, et al.
Veröffentlicht: (2025)
Differentially Private Steering for Large Language Model Alignment
von: Goel, Anmol, et al.
Veröffentlicht: (2025)
von: Goel, Anmol, et al.
Veröffentlicht: (2025)
Speculative Decoding Across Languages
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
FutureSim: Replaying World Events to Evaluate Adaptive Agents
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models
von: Goel, Naman
Veröffentlicht: (2023)
von: Goel, Naman
Veröffentlicht: (2023)
CAFIN: Centrality Aware Fairness inducing IN-processing for Unsupervised Representation Learning on Graphs
von: Arun, Arvindh, et al.
Veröffentlicht: (2023)
von: Arun, Arvindh, et al.
Veröffentlicht: (2023)
Conformal Language Model Reasoning with Coherent Factuality
von: Rubin-Toles, Maxon, et al.
Veröffentlicht: (2025)
von: Rubin-Toles, Maxon, et al.
Veröffentlicht: (2025)
Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity
von: Hinostroza, Cristian, et al.
Veröffentlicht: (2026)
von: Hinostroza, Cristian, et al.
Veröffentlicht: (2026)
Are Models Trained on Indian Legal Data Fair?
von: Girhepuje, Sahil, et al.
Veröffentlicht: (2023)
von: Girhepuje, Sahil, et al.
Veröffentlicht: (2023)
Cat, Rat, Meow: On the Alignment of Language Model and Human Term-Similarity Judgments
von: Linhardt, Lorenz, et al.
Veröffentlicht: (2025)
von: Linhardt, Lorenz, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Cognac Shot To Forget Bad Memories: Corrective Unlearning for Graph Neural Networks
von: Kolipaka, Varshita, et al.
Veröffentlicht: (2024) -
Measuring Moral Inconsistencies in Large Language Models
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024) -
Representation Surgery: Theory and Practice of Affine Steering
von: Singh, Shashwat, et al.
Veröffentlicht: (2024) -
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024) -
Great Models Think Alike and this Undermines AI Oversight
von: Goel, Shashwat, et al.
Veröffentlicht: (2025)