Opto-ViT: Architecting a Near-Sensor Region of Interest-Aware Vision Transformer Accelerator with Silicon Photonics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Morsali, Mehrdad, Zhou, Chengwei, Najafi, Deniz, Sarkar, Sreetama, Mercati, Pietro, Khoshavi, Navid, Beerel, Peter, Nikdast, Mahdi, Datta, Gourav, Angizi, Shaahin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Light-Bound Transformers: Hardware-Anchored Robustness for Silicon-Photonic Computer Vision Systems
von: Chen, Xuming, et al.
Veröffentlicht: (2026)
von: Chen, Xuming, et al.
Veröffentlicht: (2026)
Neuro-Photonix: Enabling Near-Sensor Neuro-Symbolic AI Computing on Silicon Photonics Substrate
von: Najafi, Deniz, et al.
Veröffentlicht: (2024)
von: Najafi, Deniz, et al.
Veröffentlicht: (2024)
Lightator: An Optical Near-Sensor Accelerator with Compressive Acquisition Enabling Versatile Image Processing
von: Morsali, Mehrdad, et al.
Veröffentlicht: (2024)
von: Morsali, Mehrdad, et al.
Veröffentlicht: (2024)
FixPix: Fixing Bad Pixels using Deep Learning
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2023)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2023)
OASIS: Optimized Lightweight Autoencoder System for Distributed In-Sensor computing
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
MaskVD: Region Masking for Efficient Video Object Detection
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
SA-DS: A Dataset for Large Language Model-Driven AI Accelerator Design Generation
von: Vungarala, Deepak, et al.
Veröffentlicht: (2024)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2024)
Energy-Efficient & Real-Time Computer Vision with Intelligent Skipping via Reconfigurable CMOS Image Sensors
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
Region Masking to Accelerate Video Processing on Neuromorphic Hardware
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
Linearizing Models for Efficient yet Robust Private Inference
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
Accelerating Neural Networks for Large Language Models and Graph Processing with Silicon Photonics
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
Learning Scalable Temporal Representations in Spiking Neural Networks Without Labels
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
Block Selective Reprogramming for On-device Training of Vision Transformers
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
Silicon Photonic 2.5D Interposer Networks for Overcoming Communication Bottlenecks in Scale-out Machine Learning Hardware Accelerators
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
LIMCA: LLM for Automating Analog In-Memory Computing Architecture Design Exploration
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
Dependability in Embedded Systems: A Survey of Fault Tolerance Methods and Software-Based Mitigation Techniques
von: Solouki, Mohammadreza Amel, et al.
Veröffentlicht: (2024)
von: Solouki, Mohammadreza Amel, et al.
Veröffentlicht: (2024)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
LMUFormer: Low Complexity Yet Powerful Spiking Model With Legendre Memory Units
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
HIVTP: A Training-Free Method to Improve VLMs Efficiency via Hierarchical Visual Token Pruning Using Middle-Layer-Based Importance Score
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
BladderFormer: A Streaming Transformer for Real-Time Urological State Monitoring
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
Rethinking Vision Transformer Depth via Structural Reparameterization
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)
ViT Registers and Fractal ViT
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
Technology-Circuit-Algorithm Tri-Design for Processing-in-Pixel-in-Memory (P2M)
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2023)
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2023)
SPICEPilot: Navigating SPICE Code Generation and Simulation with AI Guidance
von: Vungarala, Deepak, et al.
Veröffentlicht: (2024)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2024)
LightPro: A Linear Photonic Processor with Full Programmability
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
Enabling Scalable Photonic Tensor Cores with Polarization-Domain Photonic Computing
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
LuxNAS: A Coherent Photonic Neural Network Powered by Neural Architecture Search
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
von: Shafiee, Amin, et al.
Veröffentlicht: (2025)
DNN-Defender: A Victim-Focused In-DRAM Defense Mechanism for Taming Adversarial Weight Attack on DNNs
von: Zhou, Ranyang, et al.
Veröffentlicht: (2023)
von: Zhou, Ranyang, et al.
Veröffentlicht: (2023)
RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
von: Pandit, Kartik, et al.
Veröffentlicht: (2025)
von: Pandit, Kartik, et al.
Veröffentlicht: (2025)
Toward High Performance, Programmable Extreme-Edge Intelligence for Neuromorphic Vision Sensors utilizing Magnetic Domain Wall Motion-based MTJ
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
TFS-ViT: Token-Level Feature Stylization for Domain Generalization
von: Noori, Mehrdad, et al.
Veröffentlicht: (2023)
von: Noori, Mehrdad, et al.
Veröffentlicht: (2023)
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
A Low-Computational Video Synopsis Framework with a Standard Dataset
von: Malekpour, Ramtin, et al.
Veröffentlicht: (2024)
von: Malekpour, Ramtin, et al.
Veröffentlicht: (2024)
Compromising the Intelligence of Modern DNNs: On the Effectiveness of Targeted RowPress
von: Zhou, Ranyang, et al.
Veröffentlicht: (2024)
von: Zhou, Ranyang, et al.
Veröffentlicht: (2024)
Arena: A Patch-of-Interest ViT Inference Acceleration System for Edge-Assisted Video Analytics
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Light-Bound Transformers: Hardware-Anchored Robustness for Silicon-Photonic Computer Vision Systems
von: Chen, Xuming, et al.
Veröffentlicht: (2026) -
Neuro-Photonix: Enabling Near-Sensor Neuro-Symbolic AI Computing on Silicon Photonics Substrate
von: Najafi, Deniz, et al.
Veröffentlicht: (2024) -
Lightator: An Optical Near-Sensor Accelerator with Compressive Acquisition Enabling Versatile Image Processing
von: Morsali, Mehrdad, et al.
Veröffentlicht: (2024) -
FixPix: Fixing Bad Pixels using Deep Learning
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2023) -
OASIS: Optimized Lightweight Autoencoder System for Distributed In-Sensor computing
von: Zhou, Chengwei, et al.
Veröffentlicht: (2025)