"Scalable Processing-Near-Memory for 1M-Token LLM Inference: CXL-Enabled ..."

Dowon Kim et al. (2025)

Details and statistics

DOI: 10.1109/PACT65351.2025.00013

access: closed

type: Conference or Workshop Paper

metadata version: 2026-06-21