Imported from hiyenwong/ai_collection (
collection/skills/nlp-llm/arxiv-2608-22643v1-neuroprefetcher-storage-aware-sparse-llm-inference/SKILL.md). Install upstream withnpx skills add hiyenwong/ai_collection --skill arxiv-2608-22643v1-neuroprefetcher-storage-aware-sparse-llm-inference. Copyright stays with the author.
-- name: arxiv-2608-22643v1-neuroprefetcher-storage-aware-sparse-llm-inference description: 'NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching (arXiv: 2608.22643v1)' metadata: { "arxiv_id": "2608.22643v1", "utility": 1.0, "title": "NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching", "authors": "Nobel Dhar, Md Romyull Islam, Xuechen Zhang, Gongjin Sun, Sahidul Islam, Bobin Deng, Kun Suo", "url": "http://arxiv.org/abs/2608.22643v1" }
NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching
arXiv ID: 2608.22643v1 Authors: Nobel Dhar, Md Romyull Islam, Xuechen Zhang, Gongjin Sun, Sahidul Islam, Bobin Deng, Kun Suo URL: http://arxiv.org/abs/2608.22643v1 Utility Score: 1.00
Summary
This skill was automatically generated from the arXiv paper titled "NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching" (ID: 2608.22643v1).
Usage
This skill can be used to reference the paper's concepts, methodologies, or findings in agent workflows.