Advanced Memory Compression Methods

19h

Breaking the 100M Token Limit: EverMind's MSA Architecture Achieves Efficient End-to-End Long-Term Memory for LLMs

The research introduces a novel memory architecture called MSA (Memory Sparse Attention). Through a combination of the Memory Sparse Attention mechanism, Document-wise RoPE for extreme context ...

EE World Online

How to approach AI hardware design to address the memory wall?

This article outlines the design strategies currently used to address these bottlenecks, ranging from data center systolic ...

Nvidia shrinks LLM memory 20x without changing model weights

Nvidia's KV Cache Transform Coding (KVTC) compresses LLM key-value cache by 20x without model changes, cutting GPU memory costs and time-to-first-token by up to 8x for multi-turn AI applications.

TMCnet

Samsung Unveils HBM4E, Showcasing Comprehensive AI Solutions, NVIDIA Partnership and Vision at NVIDIA GTC 2026

Samsung Electronics Co., Ltd., a global leader in advanced semiconductor technology, today announced the comprehensive AI computing technologies it will showcase at NVIDIA GTC 2026 in San Jose, ...

Korea JoongAng Daily

Samsung unveils HBM4E at Nvidia GTC, raises bar for AI memory

Samsung Electronics debuted its seventh-generation high bandwidth memory, HBM4E, at the Nvidia GTC 2026 developer conference ...

13d

New KV cache compaction technique cuts LLM memory 50x without accuracy loss

MIT researchers developed Attention Matching, a KV cache compaction technique that compresses LLM memory by 50x in seconds — ...

TMCnet

Keysight Demonstrates 5G-Advanced AI-Powered Channel State Information Compression and Paves the Way for 6G

Keysight Technologies, Inc. (NYSE: KEYS) has collaborated with Qualcomm Technologies, Inc. to demonstrate machine learning (ML)-based Channel State Information (CSI) compression to enhance link ...

Business Wire

Keysight Demonstrates 5G-Advanced AI-Powered Channel State Information Compression and Paves the Way for 6G

Joint lab validation shows more than 40 percent downlink throughput gain versus standardized channel feedback in four-layer (rank-4) operation SANTA ROSA, Calif.--(BUSINESS WIRE)--Keysight ...

Seeking Alpha

ASML plans to build tools for advanced packaging for AI chips: report

ASML (ASML) plans to expand its chipmaking equipment portfolio with new products to capture more of the growing market for AI chips, Reuters reported, citing the company's Chief Technology Officer ...

Reuters

Exclusive: ASML plots future of chipmaking tools for AI beyond EUV

ASML plans to expand into advanced packaging for AI chips Company to use AI to enhance tool performance and production speed ASML explores larger chip sizes and new scanner systems SAN JOSE, ...

IEEE

Multi-Scale Feature Compression via Multi-Receptive-Field Convolutional Neural Network for Machine Vision

Abstract: Multi-scale feature compression is essential in machine vision tasks for reducing storage and transmission costs while maintaining task performance. However, existing multi-scale feature ...

Some results have been hidden because they may be inaccessible to you

Show inaccessible results