Gespeichert in:
| Hauptverfasser: | Rahman, Md Hafizur, Chakraborty, Prabuddha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.18443 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HARP: Measuring Harm Amplification in Multi-Agent LLM Systems
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2026)
How a Bit Becomes a Story: Semantic Steering via Differentiable Fault Injection
von: Haider, Zafaryab, et al.
Veröffentlicht: (2025)
von: Haider, Zafaryab, et al.
Veröffentlicht: (2025)
ILASH: A Predictive Neural Architecture Search Framework for Multi-Task Applications
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2024)
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2024)
LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning
von: Wang, Tuowei, et al.
Veröffentlicht: (2025)
von: Wang, Tuowei, et al.
Veröffentlicht: (2025)
Multi-Objective Hardware Aware Neural Architecture Search using Hardware Cost Diversity
von: Sinha, Nilotpal, et al.
Veröffentlicht: (2024)
von: Sinha, Nilotpal, et al.
Veröffentlicht: (2024)
GraDE: A Graph Diffusion Estimator for Frequent Subgraph Discovery in Neural Architectures
von: Yang, Yikang, et al.
Veröffentlicht: (2026)
von: Yang, Yikang, et al.
Veröffentlicht: (2026)
The Laminar Flow Hypothesis: Detecting Jailbreaks via Semantic Turbulence in Large Language Models
von: Rahman, Md. Hasib Ur
Veröffentlicht: (2025)
von: Rahman, Md. Hasib Ur
Veröffentlicht: (2025)
NCO4CVRP: Neural Combinatorial Optimization for the Capacitated Vehicle Routing Problem
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
Cross-Architecture Model Diffing with Crosscoders: Unsupervised Discovery of Differences Between LLMs
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
LeMoF: Level-guided Multimodal Fusion for Heterogeneous Clinical Data
von: Kim, Jongseok, et al.
Veröffentlicht: (2026)
von: Kim, Jongseok, et al.
Veröffentlicht: (2026)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
Strategic Fusion Optimizes Transformer Compression
von: Rahman, Md Shoaibur
Veröffentlicht: (2025)
von: Rahman, Md Shoaibur
Veröffentlicht: (2025)
Regularized Multi-LLMs Collaboration for Enhanced Score-based Causal Discovery
von: Li, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxuan, et al.
Veröffentlicht: (2024)
A Small Math Model: Recasting Strategy Choice Theory in an LLM-Inspired Architecture
von: Rahman, Roussel, et al.
Veröffentlicht: (2025)
von: Rahman, Roussel, et al.
Veröffentlicht: (2025)
MoRe Fine-Tuning with 10x Fewer Parameters
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
Unified Parameter-Efficient Unlearning for LLMs
von: Ding, Chenlu, et al.
Veröffentlicht: (2024)
von: Ding, Chenlu, et al.
Veröffentlicht: (2024)
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
von: Gao, Jifan, et al.
Veröffentlicht: (2025)
von: Gao, Jifan, et al.
Veröffentlicht: (2025)
Multi-Objective Neural Architecture Search by Learning Search Space Partitions
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
KUET at StanceNakba Shared Task: StanceMoE: Mixture-of-Experts Architecture for Stance Detection
von: Shafi, Abdullah Al, et al.
Veröffentlicht: (2026)
von: Shafi, Abdullah Al, et al.
Veröffentlicht: (2026)
NdLinear: Preserving Multi-Dimensional Structure for Parameter-Efficient Neural Networks
von: Reneau, Alex, et al.
Veröffentlicht: (2025)
von: Reneau, Alex, et al.
Veröffentlicht: (2025)
Hierarchical Multi-Scale Graph Neural Networks: Scalable Heterophilous Learning with Oversmoothing and Oversquashing Mitigation
von: Hossen, Md Sazzad, et al.
Veröffentlicht: (2026)
von: Hossen, Md Sazzad, et al.
Veröffentlicht: (2026)
Exploiting the Experts: Unauthorized Compression in MoE-LLMs
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
M3-JEPA: Multimodal Alignment via Multi-gate MoE based on the Joint-Embedding Predictive Architecture
von: Lei, Hongyang, et al.
Veröffentlicht: (2024)
von: Lei, Hongyang, et al.
Veröffentlicht: (2024)
MoEITS: A Green AI approach for simplifying MoE-LLMs
von: Balderas, Luis, et al.
Veröffentlicht: (2026)
von: Balderas, Luis, et al.
Veröffentlicht: (2026)
Grid2Guide: A* Enabled Small Language Model for Indoor Navigation
von: Haque, Md. Wasiul, et al.
Veröffentlicht: (2025)
von: Haque, Md. Wasiul, et al.
Veröffentlicht: (2025)
Trackly: A Unified SaaS Platform for User Behavior Analytics and Real Time Rule Based Anomaly Detection
von: Haque, Md Zahurul, et al.
Veröffentlicht: (2026)
von: Haque, Md Zahurul, et al.
Veröffentlicht: (2026)
A Fragile Number Sense: Probing the Elemental Limits of Numerical Reasoning in LLMs
von: Rahman, Roussel, et al.
Veröffentlicht: (2025)
von: Rahman, Roussel, et al.
Veröffentlicht: (2025)
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
Can LLMs Leverage Observational Data? Towards Data-Driven Causal Discovery with LLMs
von: Susanti, Yuni, et al.
Veröffentlicht: (2025)
von: Susanti, Yuni, et al.
Veröffentlicht: (2025)
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
von: Wei, Zhao, et al.
Veröffentlicht: (2025)
von: Wei, Zhao, et al.
Veröffentlicht: (2025)
MoSLD: An Extremely Parameter-Efficient Mixture-of-Shared LoRAs for Multi-Task Learning
von: Zhao, Lulu, et al.
Veröffentlicht: (2024)
von: Zhao, Lulu, et al.
Veröffentlicht: (2024)
Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers
von: Sankaran, Shyam, et al.
Veröffentlicht: (2026)
von: Sankaran, Shyam, et al.
Veröffentlicht: (2026)
Symmetry in Neural Network Parameter Spaces
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
CauScale: Neural Causal Discovery at Scale
von: Peng, Bo, et al.
Veröffentlicht: (2026)
von: Peng, Bo, et al.
Veröffentlicht: (2026)
Spectral Manifold Regularization for Stable and Modular Routing in Deep MoE Architectures
von: Delibasoglu, Ibrahim
Veröffentlicht: (2026)
von: Delibasoglu, Ibrahim
Veröffentlicht: (2026)
Emotion Detection From Social Media Posts
von: Rahman, Md Mahbubur, et al.
Veröffentlicht: (2023)
von: Rahman, Md Mahbubur, et al.
Veröffentlicht: (2023)
LatentMoE: Toward Optimal Accuracy per FLOP and Parameter in Mixture of Experts
von: Elango, Venmugil, et al.
Veröffentlicht: (2026)
von: Elango, Venmugil, et al.
Veröffentlicht: (2026)
An Accurate and Low-Parameter Machine Learning Architecture for Next Location Prediction
von: Jary, Calvin, et al.
Veröffentlicht: (2024)
von: Jary, Calvin, et al.
Veröffentlicht: (2024)
LeMoLE: LLM-Enhanced Mixture of Linear Experts for Time Series Forecasting
von: Zhang, Lingzheng, et al.
Veröffentlicht: (2024)
von: Zhang, Lingzheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
HARP: Measuring Harm Amplification in Multi-Agent LLM Systems
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2026) -
How a Bit Becomes a Story: Semantic Steering via Differentiable Fault Injection
von: Haider, Zafaryab, et al.
Veröffentlicht: (2025) -
ILASH: A Predictive Neural Architecture Search Framework for Multi-Task Applications
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2024) -
LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning
von: Wang, Tuowei, et al.
Veröffentlicht: (2025) -
Multi-Objective Hardware Aware Neural Architecture Search using Hardware Cost Diversity
von: Sinha, Nilotpal, et al.
Veröffentlicht: (2024)