속성:Applicability
외관
이 메모가 맞는 조건과 예외를 적습니다.
k
Memory-constrained, CPU/disk-offloaded MoE batch inference.
Limits: Preprint-only record; gains are hardware/model/batch dependent, and more batches increase KV-cache load and latency. +
온라인 캐시 크기·정책 튜닝.
Limits: 근사 오차와 샘플링, 지원 정책 집합에 제약된다. +
kvaccel a novel write accelerator for lsm tree based kv stores with host ssd collaboration 968f5966 +
LSM-based stores on programmable dual-interface SSDs.
Limits: Requires modified dual-interface SSD firmware; the headline gain is for write-intensive workloads. +
kvcache cache in the wild characterizing and optimizing kvcache cache at a large cloud provider 0b39b53c +
Cross-request prefix/KV reuse가 있는 shared LLM serving cluster에 적용합니다.
Limits: 한 cloud provider의 두 production 기간을 분석했고 논문은 각 trace의 대표 하루 결과를 주로 제시합니다. Reasoning workload와 prefix 이외의 KV reuse, global scheduling, cache fairness는 평가 범위 밖입니다. +
l
Rack-scale disaggregated compute, memory, and storage fabrics.
Limits: Evaluation used an emulated commodity-server environment rather than native disaggregated hardware. +
가상화 데이터센터의 긴급 전력 상한 제어와 VM 재배치 정책. +
조직 내 공용 Hadoop/MapReduce 클러스터와 클라우드 기반 사용자별 분석 환경. +
모바일 SoC의 image/modem 등 전용 CPU subsystem의 빠른 재기동. +
현재는 고성능 storage I/O 및 interrupt delegation 연구의 서지·계보 추적에만 사용한다. 구현 또는 성능 판단의 근거로 사용하지 않는다. +
현재는 모바일 메모리 압박과 앱 실행시간 연구의 서지 추적 및 원문 확보 우선순위 지정에만 사용한다. 메모리 관리 정책 선택에는 사용하지 않는다. +
초저지연 NVMe polling과 SMT/하이퍼스레딩을 함께 사용하는 시스템. +
light dedup a light weight inline deduplication framework for non volatile memory file systems 000ae22e +
NVM 파일시스템 inline dedup.
Limits: 중복·지역성 정도와 NVM 특성에 민감하며 일부 결과는 crafted aging workload다. +
Multi-tenant GPUs running inference and mixed ML jobs.
Limits: Requires a specialized GPU OS/runtime and kernel atomization; gains vary by workload interference. +
Single-host recent high-frequency telemetry and interactive debugging.
Limits: Loom targets recent situational analysis rather than long-term storage, has finite ingest capacity, and may lose the active in-memory block (for example 64 MiB) on machine failure. +
Bloom filter와 block index를 사용하는 LSM-tree 키-값 저장소의 point lookup. +
LSM-tree 기반 키-값 저장소의 초기 적재, 마이그레이션, 대규모 ingest 진단. +
m
CXL DRAM tiers with modifiable controller hardware.
Limits: Requires controller support; gains are measured under a 2–3x CXL/DRAM latency gap. +
Research artifact inventories and promotion pipelines that need machine-verifiable provenance without a self-asserted human credential. This does not make B200 data paper-grade or authorize target-hardware final runs. +
GPU buffer가 system memory를 공유하고 background app caching이 중요한 mobile SoC에 해당한다. +
Virtualized CXL memory pools and multi-tenant cloud hosts.
Limits: Results depend on the evaluated FMM/CXL prototype, workload mix, and estimator/page-coloring assumptions. +