본문으로 이동

속성:Applicability

S3 연구 메모리

Text

이 메모가 맞는 조건과 예외를 적습니다.

( | ) (20 | 50 | 100 | 250 | 500) 보기
이 속성을 사용하는 문서 20개를 보여줍니다.
k
Memory-constrained, CPU/disk-offloaded MoE batch inference. Limits: Preprint-only record; gains are hardware/model/batch dependent, and more batches increase KV-cache load and latency.  +
온라인 캐시 크기·정책 튜닝. Limits: 근사 오차와 샘플링, 지원 정책 집합에 제약된다.  +
LSM-based stores on programmable dual-interface SSDs. Limits: Requires modified dual-interface SSD firmware; the headline gain is for write-intensive workloads.  +
Cross-request prefix/KV reuse가 있는 shared LLM serving cluster에 적용합니다. Limits: 한 cloud provider의 두 production 기간을 분석했고 논문은 각 trace의 대표 하루 결과를 주로 제시합니다. Reasoning workload와 prefix 이외의 KV reuse, global scheduling, cache fairness는 평가 범위 밖입니다.  +
l
Rack-scale disaggregated compute, memory, and storage fabrics. Limits: Evaluation used an emulated commodity-server environment rather than native disaggregated hardware.  +
가상화 데이터센터의 긴급 전력 상한 제어와 VM 재배치 정책.  +
조직 내 공용 Hadoop/MapReduce 클러스터와 클라우드 기반 사용자별 분석 환경.  +
모바일 SoC의 image/modem 등 전용 CPU subsystem의 빠른 재기동.  +
현재는 고성능 storage I/O 및 interrupt delegation 연구의 서지·계보 추적에만 사용한다. 구현 또는 성능 판단의 근거로 사용하지 않는다.  +
현재는 모바일 메모리 압박과 앱 실행시간 연구의 서지 추적 및 원문 확보 우선순위 지정에만 사용한다. 메모리 관리 정책 선택에는 사용하지 않는다.  +
초저지연 NVMe polling과 SMT/하이퍼스레딩을 함께 사용하는 시스템.  +
NVM 파일시스템 inline dedup. Limits: 중복·지역성 정도와 NVM 특성에 민감하며 일부 결과는 crafted aging workload다.  +
Multi-tenant GPUs running inference and mixed ML jobs. Limits: Requires a specialized GPU OS/runtime and kernel atomization; gains vary by workload interference.  +
Single-host recent high-frequency telemetry and interactive debugging. Limits: Loom targets recent situational analysis rather than long-term storage, has finite ingest capacity, and may lose the active in-memory block (for example 64 MiB) on machine failure.  +
Bloom filter와 block index를 사용하는 LSM-tree 키-값 저장소의 point lookup.  +
LSM-tree 기반 키-값 저장소의 초기 적재, 마이그레이션, 대규모 ingest 진단.  +
m
CXL DRAM tiers with modifiable controller hardware. Limits: Requires controller support; gains are measured under a 2–3x CXL/DRAM latency gap.  +
Research artifact inventories and promotion pipelines that need machine-verifiable provenance without a self-asserted human credential. This does not make B200 data paper-grade or authorize target-hardware final runs.  +
GPU buffer가 system memory를 공유하고 background app caching이 중요한 mobile SoC에 해당한다.  +
Virtualized CXL memory pools and multi-tenant cloud hosts. Limits: Results depend on the evaluated FMM/CXL prototype, workload mix, and estimator/page-coloring assumptions.  +