Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
April 24, 2024Open Access

NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing

View Full Paper
Ask AI
Bookmark
Share

Authors

GHGuseul HeoKorea Advanced Institute of Science and TechnologySLSangyeop LeeSungkyunkwan UniversityJCJaehong ChoKorea Advanced Institute of Science and Technology

Discussion

Loading...

Member takes

Implication

Key Points

Key points are not available for this paper at this time.

Cite This Study

Heo et al. (2024) studied this question.

synapsesocial.com/papers/68e6de61b6db643587659b71https://doi.org/10.1145/3620666.3651380
View Full Paper
Ask AI
Bookmark
Share

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing2024 · 121 citations
  2. 2Towards Floating Point-Based AI Acceleration: Hybrid PIM with Non-Uniform Data Format and Reduced Multiplications2025
  3. 3IMI: In-memory Multi-job Inference Acceleration for Large Language Models2024 · 4 citations
  4. 4HyPIM: LLM Acceleration with A Hybrid ReRAM/SRAM 3D-PIM Architecture2026
  5. 5IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System2024 · 57 citations