본문으로 이동

속성:Context

S3 연구 메모리

Text

하드웨어, 부하, 버전, 규모처럼 결과에 영향을 줄 수 있는 조건을 적습니다.

( | ) (20 | 50 | 100 | 250 | 500) 보기
이 속성을 사용하는 문서 20개를 보여줍니다.
o
Venue: OSDI. Year: 2022. 요청 단위 배칭은 먼저 끝난 요청 때문에 GPU를 낭비하고 head-of-line blocking을 만든다. Verification: abstract_only; confidence=high.  +
Venue: USENIX ATC. Year: 2023. DRAM 확장은 비싸고 플래시는 느리며 쓰기 수명 제한이 있다. Verification: abstract_only; confidence=high.  +
Venue: SOSP. Year: 2024. Incorrect memory-barrier use creates rare concurrency failures that ordinary schedules and hardware execution reproduce unreliably. Verification: official_abstract; confidence=high.  +
p
Venue: EuroSys. Year: 2026. A single page-cache copy can force remote NUMA accesses for file-heavy multithreaded workloads. Verification: official DOI/EuroSys metadata and author project full-text summary; confidence=high.  +
Venue: ASPLOS. Year: 2026. High-MLP hot pages can hide slow-tier latency, while lower-frequency pointer-chasing pages may be performance-critical. Verification: publication metadata is preserved, but the current fetched object is a landing/cookie page; claim-bearing full text must be reattached before raising confidence; confidence=low.  +
Venue: SOSP. Year: 2023. 동시 모델의 커널 순서를 GPU가 불투명하게 결정해 SLO 제어가 어렵다. Verification: abstract_only; confidence=medium.  +
Venue: ASPLOS. Year: 2025. A fixed placement wastes either GPU compute or PIM bandwidth across different inference kernels. Verification: official DOI metadata and arXiv abstract; confidence=high.  +
Venue: SOSP. Year: 2023. 클라이언트 장애는 refcount 누수·중복 해제·wild pointer를 만들 수 있다. Verification: abstract_only; confidence=high.  +
Venue: OSDI. Year: 2020. 고정 파티셔닝은 소수 핫 객체가 특정 서버를 포화시키면 확장성이 무너진다. Verification: abstract_only; confidence=high.  +
Publication scope: international. Lab publication metadata: 1. Jaehyun Song, Hwanjin Jeong, Jinkyu Jeong, "Performance Optimization of Object Tracking Algorithms in OpenCV on GPUs," Applied Sciences, Vol. 12, Issue 15, pp. 7801, 2022, doi.org/10.3390/app12157801 Verification level: full_text. Sources checked: https://www.mdpi.com/2076-3417/12/15/7801 https://doi.org/10.3390/app12157801  +
Venue: FAST. Year: 2023. fail-slow는 정상 응답을 하면서 꼬리 지연을 키워 단순 장애 탐지로 잡기 어렵다. Verification: abstract_only; confidence=high.  +
Venue: EuroSys. Year: 2025. Page-granular hotness can retain unused bytes in scarce fast memory and cause capacity-cliff slowdowns. Verification: official DOI metadata and accessible publisher abstract; confidence=high.  +
Venue: MSST. Year: 2024. 기존 hot-page 선이동은 DRAM의 page counter가 캐시 공간을 잠식한다. Verification: full_text; confidence=high.  +
Venue: SOSP. Year: 2023. 동적 희소성은 정적 컴파일·일반 sparse kernel 모두에 불규칙성을 만든다. Verification: abstract_only; confidence=high.  +
Three-policy pivot target-generation migration collector, source commit 038c9cd00c25fc83678d4d698f8780bfc5d96369, intended for sequential Qwen then Mistral collection on B200-2.  +
All 8 B200 GPUs were idle before launch. The runner acquired experiments/runs/.weighted_slo_final.global.lock, then immediately reran its clean-tree gate before model preflight or GPU initialization.  +
Sequential Qwen then Mistral target-generation migration collection on B200-2, source commit 8ca2dbb68d202a8ba8b4ae226b64d77ec3e604ca, before any CUDA generation process was launched.  +
Repo dw-kv; protocol weighted_slo_final_v1 sha256=528c4a83e8dac516ab6898f9b73b652c348c14d9eb5c8a2eea1e965e50949c10; host B200-2; 8x NVIDIA B200 all 0 MiB/0% and no compute processes; runtime torch 2.11.0+cu130, CUDA 13.0, vLLM 0.21.0; pinned Qwen and Mistral snapshots available; global concurrency remains 1.  +
B200-to-L40S or RTX PRO migration of target-model length calibration artifacts before any final performance experiment.  +
B200-2 sequential migration launch at source commit 038c9cd00c25fc83678d4d698f8780bfc5d96369.  +