Discover ANY AI to make more online for less.

select between over 22,900 AI Tool and 17,900 AI News Posts.


thenextweb
Google’s new compression algorithm cut memory stocks within hours of publication

Google published a research blog post on Tuesday about a new compression algorithm for AI models. Within hours, memory stocks were falling. Micron dropped 3 per cent, Western Digital lost 4.7 per cent, and SanDisk fell 5.7 per cent, as investors recalculated how much physical memory the AI industry might actually need. The algorithm is […]
This story continues at The Next Web

Rating

Innovation

Pricing

Technology

Usability

We have discovered similar tools to what you are looking for. Check out our suggestions for similar AI tools.

venturebeat
Nvidia says it can shrink LLM memory 20x without changing model weights

<p>Nvidia researchers have introduced a new technique that dramatically reduces how much memory large language models need to track conversation history — by as much as 20x — without modifyi [...]

Match Score: 144.75

venturebeat
Context compression finally works in production: new research cuts LLM inpu

<p>Context windows are becoming a computational bottleneck. The longer an agent runs, the more tokens accumulate from retrieved documents, reasoning traces and conversation history, and the more [...]

Match Score: 124.43

venturebeat
DeepSeek drops open-source model that compresses text 10x through images, d

<p><a href="https://www.deepseek.com/"><u>DeepSeek</u></a>, the Chinese artificial intelligence research company that has repeatedly challenged assumptions abou [...]

Match Score: 117.88

venturebeat
Google's new TurboQuant algorithm speeds up AI memory 8x, cutting cost

<p>As Large Language Models (LLMs) expand their context windows to process massive documents and intricate conversations, they encounter a brutal hardware reality known as the &quot;Key-Valu [...]

Match Score: 91.63

venturebeat
New KV cache compaction technique cuts LLM memory 50x without accuracy loss

<p>Enterprise AI applications that handle large documents or long-horizon tasks face a severe memory bottleneck. As the context grows longer, so does the KV cache, the area where the model’s w [...]

Match Score: 90.17

venturebeat
'Observational memory' cuts AI agent costs 10x and outscores RAG

<p>RAG isn&#x27;t always fast enough or intelligent enough for modern agentic AI workflows. As teams move from short-lived chatbots to long-running, tool-heavy agents embedded in production [...]

Match Score: 78.35

venturebeat
A 0.12% parameter add-on gives AI agents the working memory RAG can't

<p>AI agents forget. Every time a coding assistant loses track of a debugging thread, or a data analysis agent re-ingests the same context it already processed, the team pays in latency, token c [...]

Match Score: 76.81

X's 'open source' algorithm isn't a win for transparency, researchers say
X's 'open source' algorithm isn't a win for transparenc

<p>When X&#39;s engineering team published <a target="_blank" class="link" href="https://github.com/xai-org/x-algorithm" data-i13n="cpos:1;pos:1"&g [...]

Match Score: 71.07

venturebeat
Tencent's Team Memory shares AI agent memory across a team — with no

<p>A <a href="https://venturebeat.com/data/57-of-enterprises-have-watched-ai-agents-be-confidently-wrong-the-fix-is-an-agentic-context-layer-but-who-has-one">VB Pulse survey this [...]

Match Score: 67.49