본문으로 이동

속성:Evidence citation

S3 연구 메모리

Text

원본을 다시 찾을 수 있는 인용입니다.

( | ) (20 | 50 | 100 | 250 | 500) 보기
이 속성을 사용하는 문서 20개를 보여줍니다.
i
ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices. EuroSys 2023.  +
Changlong Li et al., "ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices", EuroSys 2023.  +
Changlong Li et al., "ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices", EuroSys 2023.  +
Minwoo Ahn, Jeongmin Han, Youngjin Kwon, Jinkyu Jeong, "Identifying On-/Off-CPU Bottlenecks Together with Blocked Samples" in Proceedings of the 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI '24), July 2024  +
Minwoo Ahn, Jeongmin Han, Youngjin Kwon, Jinkyu Jeong, "Identifying On-/Off-CPU Bottlenecks Together with Blocked Samples," in Proceedings of the 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI '24), July 2024  +
Weijian Chen et al., "IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference", FAST 2025.  +
IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference. FAST 2025.  +
Weijian Chen et al., "IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference", FAST 2025.  +
Andre Luiz Nunes Martins, Cesar A. V. Duarte, Jinkyu Jeong, "Improving Application Launch Performance in Smartphones Using Recurrent Neural Network," in Proceedings of the 2018 International Conference on Machine Learning Technologies (ICMLT 2018), Jinan, China, May 26-28, 2018.  +
Andre Luiz Nunes Martins, Cesar A. V. Duarte, Jinkyu Jeong, "Improving Application Launch Performance in Smartphones Using Recurrent Neural Network" in Proceedings of the 2018 International Conference on Machine Learning Technologies (ICMLT 2018), Jinan, China, May 26-28, 2018.  +
INF2: High-Throughput Generative Inference of Large Language Models using Near-Storage Processing. arXiv 2025.  +
Hongsun Jang, Jaeyong Song, Changmin Shin, Si Ung Noh, Jaewon Jung, Jisung Park, and Jinho Lee. “A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs.” ASPLOS ’26. DOI: 10.1145/3779212.3790119. arXiv:2502.09921v2.  +
Hongsun Jang et al., "INF²: High-Throughput Generative Inference of Large Language Models using Near-Storage Processing", arXiv 2025.  +
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management. OSDI 2024.  +
Wonbeom Lee et al., "InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management", OSDI 2024.  +
Wonbeom Lee et al., "InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management", OSDI 2024.  +
Yifan Yuan et al., "Intel Accelerator Ecosystem: An SoC-Oriented Perspective", ISCA 2024.  +
Intel Accelerator Ecosystem: An SoC-Oriented Perspective. ISCA 2024.  +
Yifan Yuan et al., "Intel Accelerator Ecosystem: An SoC-Oriented Perspective", ISCA 2024.  +
Yi Xu et al., "IONIA: High-Performance Replication for Modern Disk-based KV Stores", FAST 2024.  +