Learning to Reason in 13 Parameters
Fuente:
arXiv
Salvato in:
| Autori principali: | Morris, John X., Mireshghallah, Niloofar, Ibrahim, Mark, Mahloujifar, Saeed |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025)
di: Naseh, Ali, et al.
Pubblicazione: (2025)
Position: Privacy Is Not Just Memorization!
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
Z0-Inf: Zeroth Order Approximation for Data Influence
di: Kokhlikyan, Narine, et al.
Pubblicazione: (2025)
di: Kokhlikyan, Narine, et al.
Pubblicazione: (2025)
Auditing $f$-Differential Privacy in One Run
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2024)
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2024)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
Machine Learning with Privacy for Protected Attributes
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2025)
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2025)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
Privacy Blur: Quantifying Privacy and Utility for Image Data Release
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2025)
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2025)
Smaller Language Models are Better Black-box Machine-Generated Text Detectors
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
Privacy Amplification for the Gaussian Mechanism via Bounded Support
di: Hu, Shengyuan, et al.
Pubblicazione: (2024)
di: Hu, Shengyuan, et al.
Pubblicazione: (2024)
Observational Auditing of Label Privacy
di: Kalemaj, Iden, et al.
Pubblicazione: (2025)
di: Kalemaj, Iden, et al.
Pubblicazione: (2025)
Guarantees of confidentiality via Hammersley-Chapman-Robbins bounds
di: Chaudhuri, Kamalika, et al.
Pubblicazione: (2024)
di: Chaudhuri, Kamalika, et al.
Pubblicazione: (2024)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
di: Panda, Ashwinee, et al.
Pubblicazione: (2022)
di: Panda, Ashwinee, et al.
Pubblicazione: (2022)
Private Fine-tuning of Large Language Models with Zeroth-order Optimization
di: Tang, Xinyu, et al.
Pubblicazione: (2024)
di: Tang, Xinyu, et al.
Pubblicazione: (2024)
SecAlign: Defending Against Prompt Injection with Preference Optimization
di: Chen, Sizhe, et al.
Pubblicazione: (2024)
di: Chen, Sizhe, et al.
Pubblicazione: (2024)
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
Publicly-Detectable Watermarking for Language Models
di: Fairoze, Jaiden, et al.
Pubblicazione: (2023)
di: Fairoze, Jaiden, et al.
Pubblicazione: (2023)
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
di: Chen, Tong, et al.
Pubblicazione: (2024)
di: Chen, Tong, et al.
Pubblicazione: (2024)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
di: Chen, Tong, et al.
Pubblicazione: (2025)
di: Chen, Tong, et al.
Pubblicazione: (2025)
RefGrader: Automated Grading of Mathematical Competition Proofs using Agentic Workflows
di: Mahdavi, Hamed, et al.
Pubblicazione: (2025)
di: Mahdavi, Hamed, et al.
Pubblicazione: (2025)
Extracting Prompts by Inverting LLM Outputs
di: Zhang, Collin, et al.
Pubblicazione: (2024)
di: Zhang, Collin, et al.
Pubblicazione: (2024)
Do language models plan ahead for future tokens?
di: Wu, Wilson, et al.
Pubblicazione: (2024)
di: Wu, Wilson, et al.
Pubblicazione: (2024)
Harnessing the Universal Geometry of Embeddings
di: Jha, Rishi, et al.
Pubblicazione: (2025)
di: Jha, Rishi, et al.
Pubblicazione: (2025)
Early Recognition of Parkinson's Disease Through Acoustic Analysis and Machine Learning
di: Fadavi, Niloofar, et al.
Pubblicazione: (2024)
di: Fadavi, Niloofar, et al.
Pubblicazione: (2024)
Towards Parameter-Free Temporal Difference Learning
di: Li, Yunxiang, et al.
Pubblicazione: (2026)
di: Li, Yunxiang, et al.
Pubblicazione: (2026)
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
di: Xin, Rui, et al.
Pubblicazione: (2025)
di: Xin, Rui, et al.
Pubblicazione: (2025)
Validating Interpretability in siRNA Efficacy Prediction: A Perturbation-Based, Dataset-Aware Protocol
di: Khodagholi, Zahra, et al.
Pubblicazione: (2026)
di: Khodagholi, Zahra, et al.
Pubblicazione: (2026)
SGC: A semi-supervised pipeline for gene clustering using self-training approach in gene co-expression networks
di: Aghaieabiane, Niloofar, et al.
Pubblicazione: (2022)
di: Aghaieabiane, Niloofar, et al.
Pubblicazione: (2022)
Quantifying the Effect of Test Set Contamination on Generative Evaluations
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
Occam's Razor for Self Supervised Learning: What is Sufficient to Learn Good Representations?
di: Ibrahim, Mark, et al.
Pubblicazione: (2024)
di: Ibrahim, Mark, et al.
Pubblicazione: (2024)
BIOGEN: Evidence-Grounded Multi-Agent Reasoning Framework for Transcriptomic Interpretation in Antimicrobial Resistance
di: Hossain, Elias, et al.
Pubblicazione: (2025)
di: Hossain, Elias, et al.
Pubblicazione: (2025)
Lotus at SemEval-2025 Task 11: RoBERTa with Llama-3 Generated Explanations for Multi-Label Emotion Classification
di: Ranjbar, Niloofar, et al.
Pubblicazione: (2025)
di: Ranjbar, Niloofar, et al.
Pubblicazione: (2025)
ReasonCACHE: Teaching LLMs To Reason Without Weight Updates
di: Gupta, Sharut, et al.
Pubblicazione: (2026)
di: Gupta, Sharut, et al.
Pubblicazione: (2026)
Reinforcement Learning Improves Traversal of Hierarchical Knowledge in LLMs
di: Zhang, Renfei, et al.
Pubblicazione: (2025)
di: Zhang, Renfei, et al.
Pubblicazione: (2025)
Learning Stable Predictors from Weak Supervision under Distribution Shift
di: Shoeibi, Mehrdad, et al.
Pubblicazione: (2026)
di: Shoeibi, Mehrdad, et al.
Pubblicazione: (2026)
K2-Think: A Parameter-Efficient Reasoning System
di: Cheng, Zhoujun, et al.
Pubblicazione: (2025)
di: Cheng, Zhoujun, et al.
Pubblicazione: (2025)
PERK: Long-Context Reasoning as Parameter-Efficient Test-Time Learning
di: Chen, Zeming, et al.
Pubblicazione: (2025)
di: Chen, Zeming, et al.
Pubblicazione: (2025)
Lasso Penalization for High-Dimensional Beta Regression Models: Computation, Analysis, and Inference
di: Ramezani, Niloofar, et al.
Pubblicazione: (2025)
di: Ramezani, Niloofar, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025) -
Position: Privacy Is Not Just Memorization!
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025) -
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
di: Zhou, Jijie, et al.
Pubblicazione: (2025) -
Z0-Inf: Zeroth Order Approximation for Data Influence
di: Kokhlikyan, Narine, et al.
Pubblicazione: (2025) -
Auditing $f$-Differential Privacy in One Run
di: Mahloujifar, Saeed, et al.
Pubblicazione: (2024)