Awide Labs

Blog

Engineering notes, benchmarks, release reports, and company news from the team building the Awide Data Processor.

Release KV Cache vLLM Security

ADP KV cache connector v6.1.1: tenant isolation, 6× faster model loading, no more engine freezes

Release v6.1.1 of the ADP KV cache connector carries per-tenant cache isolation into the external KV store, loads model weights about six times faster, and keeps generation running when the storage gateway goes away. Verified with a full regression on hardware and already serving production traffic.

September 28, 2026
Read
KV Cache vLLM LMCache Benchmark

The same KV cache speed without 400 GiB of RAM: LMCache's DRAM tier against the ADP connector

We gave LMCache 400 GiB of DRAM, enough for the whole working set, and measured it against the ADP KV cache connector alone, which restores from NVMe drives and keeps no cache in RAM. The two came out roughly level: RAM was 10% faster on a warm restore, the connector faster after a restart, the same decode rate under load. The difference is the 400 GiB of RAM per node that the connector does not need.

September 28, 2026
Read
Security KV Cache vLLM

Isolating the KV cache between tenants: what vLLM does, what the front end must do

Prefix caching shares computed KV blocks between every request that starts the same way, and the time to first token tells anyone whether a prompt was already cached. How cache isolation closes that channel, what vLLM gives you out of the box, what the API front end has to do, and when isolation is worth its cost.

September 28, 2026
Read
Release KV Cache vLLM NVLink

New release: MiniMax M3, the Kimi K3 architecture and NVLink fan-out

MiniMax M3 on vLLM 0.28 verified on hardware with a full regression, support for the Kimi K3 architecture on vLLM 0.29, NVLink fan-out of cache restores, and the first steps toward a connector that new models no longer break.

September 22, 2026
Read
Release KV Cache vLLM

New ADP KV cache connector release: GLM-5.3 and DeepSeek-V4-Flash on vLLM 0.28/0.29

A new release brings hardware-verified support for GLM-5.3 and DeepSeek-V4-Flash on vLLM 0.28/0.29, deployment tooling, and improved stability for context-heavy inference workloads.

September 14, 2026
Read
Release KV Cache HMA MTP vLLM

New ADP KV cache connector release: hybrid Mamba and MTP on vLLM 0.26

A new release of the ADP KV cache connector: vLLM 0.26, plus the model families that used to break KV offload — hybrid Mamba (HMA) and multi-token prediction (MTP).

September 2, 2026
Read
Company Acquisition

Awide Labs Acquires Pliops Technology Portfolio

Awide Labs has acquired the intellectual property license and manufacturing rights for the Pliops technology portfolio, bringing hardware-accelerated data processing and key-value technology together with its PostgreSQL expertise and AI roadmap.

May 18, 2026
Read