속성:Evidence citation
외관
원본을 다시 찾을 수 있는 인용입니다.
i
ice collaborating memory and process management for user experience on resource limited mobile d 312e453f +
ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices. EuroSys 2023. +
ice collaborating memory and process management for user experience on resource limited mobile d 312e453f +
Changlong Li et al., "ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices", EuroSys 2023. +
ice collaborating memory and process management for user experience on resource limited mobile d 312e453f +
Changlong Li et al., "ICE: Collaborating Memory and Process Management for User Experience on Resource-limited Mobile Devices", EuroSys 2023. +
Minwoo Ahn, Jeongmin Han, Youngjin Kwon, Jinkyu Jeong, "Identifying On-/Off-CPU Bottlenecks Together with Blocked Samples" in Proceedings of the 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI '24), July 2024 +
Minwoo Ahn, Jeongmin Han, Youngjin Kwon, Jinkyu Jeong, "Identifying On-/Off-CPU Bottlenecks Together with Blocked Samples," in Proceedings of the 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI '24), July 2024 +
impress an importance informed multi tier prefix kv storage system for large language model infe 51b47015 +
Weijian Chen et al., "IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference", FAST 2025. +
impress an importance informed multi tier prefix kv storage system for large language model infe 51b47015 +
IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference. FAST 2025. +
impress an importance informed multi tier prefix kv storage system for large language model infe 51b47015 +
Weijian Chen et al., "IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference", FAST 2025. +
Andre Luiz Nunes Martins, Cesar A. V. Duarte, Jinkyu Jeong, "Improving Application Launch Performance in Smartphones Using Recurrent Neural Network," in Proceedings of the 2018 International Conference on Machine Learning Technologies (ICMLT 2018), Jinan, China, May 26-28, 2018. +
Andre Luiz Nunes Martins, Cesar A. V. Duarte, Jinkyu Jeong, "Improving Application Launch Performance in Smartphones Using Recurrent Neural Network" in Proceedings of the 2018 International Conference on Machine Learning Technologies (ICMLT 2018), Jinan, China, May 26-28, 2018. +
inf2 high throughput generative inference of large language models using near storage processing 3fcead73 +
INF2: High-Throughput Generative Inference of Large Language Models using Near-Storage Processing. arXiv 2025. +
inf2 high throughput generative inference of large language models using near storage processing 3fcead73 +
Hongsun Jang, Jaeyong Song, Changmin Shin, Si Ung Noh, Jaewon Jung, Jisung Park, and Jinho Lee. “A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs.” ASPLOS ’26. DOI: 10.1145/3779212.3790119. arXiv:2502.09921v2. +
inf2 high throughput generative inference of large language models using near storage processing 3fcead73 +
Hongsun Jang et al., "INF²: High-Throughput Generative Inference of Large Language Models using Near-Storage Processing", arXiv 2025. +
infinigen efficient generative inference of large language models with dynamic kv cache manageme a1ca3228 +
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management. OSDI 2024. +
infinigen efficient generative inference of large language models with dynamic kv cache manageme a1ca3228 +
Wonbeom Lee et al., "InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management", OSDI 2024. +
infinigen efficient generative inference of large language models with dynamic kv cache manageme a1ca3228 +
Wonbeom Lee et al., "InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management", OSDI 2024. +
Yifan Yuan et al., "Intel Accelerator Ecosystem: An SoC-Oriented Perspective", ISCA 2024. +
Intel Accelerator Ecosystem: An SoC-Oriented Perspective. ISCA 2024. +
Yifan Yuan et al., "Intel Accelerator Ecosystem: An SoC-Oriented Perspective", ISCA 2024. +
Yi Xu et al., "IONIA: High-Performance Replication for Modern Disk-based KV Stores", FAST 2024. +