You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Differential-KV (DKV) is a sparse KV-cache inference runtime designed for high-efficiency, memory-bounded long-context Large Language Model (LLM) inference across Apple Silicon (MLX) and CUDA GPUs.
KV Memory Orchestrator — a browser-native research simulator for deterministic KV-cache residency, SRAM/HBM tiering, DMA scheduling, and memory-fabric orchestration in future transformer inference systems. Patent No. 202641062302, filed in India.