속성:Question
외관
무엇을 확인하려 했는지 적습니다.
g
이 논문은 인과관계 프로파일링의 가상 최적화를 GPU 커널에 어떻게 적용하고, 그 결과 어떤 성능 향상을 예측했는가? 현재 확보한 서지 메타데이터만으로는 이를 검증할 수 없다. +
MHA 품질과 MQA 디코딩 속도 사이의 절충을 기존 체크포인트에서 얻을 수 있는가? +
system-wide memory oversubscription이 깨뜨리는 멀티테넌트 격리를 보존하면서 그룹 내부 메모리 효율을 얻을 수 있는가? +
h
stateful serverless의 장애 허용 로그 비용을 작업별 최솟값으로 줄일 수 있는가? +
Can software safely run useful work during memory-bound CPU stalls when SMT violates latency SLOs? +
How can an LSM store promote hot records without scanning or rewriting large sorted runs? +
모바일 앱의 사용자 체감 지연에서 storage I/O가 어느 경로와 비중으로 영향을 주는가? 현재 확인 가능한 EuroSys 2015 poster 메타데이터만으로는 연구 질문과 답을 검증할 수 없다. +
Can memory copying become a globally optimized asynchronous OS service instead of repeated local memcpy operations? +
i
Linux 범용 블록 계층 비용을 우회하는 유연한 NVMe 경로를 mainline에 넣을 수 있는가? +
Can prior question-answer results be reused as in-context examples to reduce expensive large-model serving? +
ice collaborating memory and process management for user experience on resource limited mobile d 312e453f +
Can mobile memory and process management cooperate to stop background refault activity from disrupting foreground frame rendering? +
CPU를 쓰는 구간과 I/O·동기화·스케줄링으로 막힌 구간을 한 프로파일에서 코드 라인 수준으로 연결하고, off-CPU 코드 최적화의 잠재효과까지 예측할 수 있는가? +
impress an importance informed multi tier prefix kv storage system for large language model infe 51b47015 +
How can disk-tiered prefix KV reuse reduce LLM time to first token when loading every cached token is too slow? +
Android LMK가 곧 사용할 앱을 제거하는 문제를 앱 사용 이력 기반 예측으로 줄일 수 있는가? +
inf2 high throughput generative inference of large language models using near storage processing 3fcead73 +
모델 가중치와 테라바이트 규모의 KV cache가 GPU 메모리는 물론 호스트 메모리 용량까지 넘는 대형·장문맥 LLM의 offline inference에서, 모델 구조를 바꾸거나 attention을 근사하지 않으면서 KV-cache I/O 병목을 줄이고 처리량을 높일 수 있는가? +
infinigen efficient generative inference of large language models with dynamic kv cache manageme a1ca3228 +
긴 문맥 LLM의 KV 캐시를 정확도 손실 적게 동적으로 줄여 offload 추론을 가속할 수 있는가? +
Intel Accelerator Ecosystem: An SoC-Oriented Perspective가 다루는 시스템 문제와 제안 기법·평가 결과는 무엇인가? +
SSD KV 저장소의 순서 보장을 완화하지 않고 복제 병목을 줄일 수 있는가? +
k
kal kernel assisted non invasive memory leak tolerance with a general purpose memory allocator 372cd042 +
기존 general-purpose allocator와 application을 수정·중단하지 않고 leak memory를 찾아 회수할 수 있는가? +
Can a flash cache store billions of roughly 100-byte objects with both tiny DRAM metadata and low flash write amplification? +