Topic: memory
13 stories found
Today
MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents
MemoryForge aims to equip large language models with human-like memory capabilities to enhance their performance in agentic tasks like role-play and user simulation by moving beyond static textual profiles. This advancement is crucial as it allows LLMs to better mimic human behavior and interactions.
Yesterday
AirLLM 70B inference with single 4GB GPU
AirLLM 70B model can now be run using a single 4GB GPU, significantly reducing hardware requirements for large language models. This breakthrough could lower barriers to entry for deploying advanced AI models in resource-constrained environments.
Friday, July 31, 2026
Wednesday, July 29, 2026
ggml/llama.cpp releases: b10181
The ggml-cuda project updated its code to disable Multi-Memory Queue (MMQ) on devices with less than 48 KiB of shared memory, ensuring optimal performance across different hardware configurations. This update is crucial for maintaining compatibility and efficiency in various computational environments.
Neuromorphic Diffusion Language Models: Addressing Compute and Memory Bottlenecks via Sparsity and Block Denoising
A new approach called Neuromorphic Diffusion Language Models aims to address inefficiencies in autoregressive large language models by leveraging sparsity and block denoising techniques, reducing compute and memory demands and potentially lowering energy consumption. This innovation is crucial as it could significantly enhance the operational efficiency of language models, making them more practical for real-world applications.
Tuesday, July 28, 2026
HKUDS/nanobot — Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-

Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI
OpenAI is expanding ChatGPT with features like AGI accessibility and no-code tools to reach 10 million users, aiming to make advanced artificial intelligence more widely available. This effort is crucial as it could democratize access to powerful AI technologies, potentially transforming various industries and daily life.
ggml/llama.cpp releases: b10167
The ggml/llama.cpp project released version b10167, which abstracts llama_memory calls to common_memory. This update is significant for improving memory management and compatibility across different platforms.
Monday, July 27, 2026
Friday, July 24, 2026
13 of 13 items shown. Sources: 76 days indexed.