Astera Labs targets the KV cache bottleneck as agentic AI outgrows GPU memory
Astera Labs has added three products to its Leo memory controller line, aimed at a problem that arises as AI inference shifts from single queries to agents that run continuous loops: the key-value cache outgrows the HBM on the accelerator, and the usual fallback makes things worse.
Read original article ↗
Related Articles
Delos Data Is Reimagining the Interconnect for the AI Age
Multiple AI models. Heterogenous compute. A variety of network topologies. Copper and optical. The mix of silicon, stora
Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARE
Federated learning (FL) projects often begin with a straightforward setup: one server, a few clients, and one dataset at
Meta’s new One subscriptions put a price on social media and AI
Shortly after launching its new do-everything AI assistant Muse, Meta's launching subscription bundles that pair its sta
PALO ALTO, Calif., Sept. 15, 2026 /PRNewswire/ -- Delos Data today announced it has raised over $100 million from Matrix