Topic: gpu

11 stories found

Yesterday

trending59

AirLLM 70B inference with single 4GB GPU

AirLLM 70B model can now be run using a single 4GB GPU, significantly reducing hardware requirements for large language models. This breakthrough could lower barriers to entry for deploying advanced AI models in resource-constrained environments.

github.com

Saturday, August 1, 2026

industry24

AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs

AMD has released Instella-MoE-16B-A3B, an open-source Mixture-of-Experts language model trained on Instinct GPUs, which could advance research by providing transparency in the training process.

marktechpost.com

Friday, July 31, 2026

releases48

ggml/llama.cpp releases: b10215

The ggml/llama.cpp project updated its Vulkan driver with a version check for Windows Intel GPUs to prevent crashes, addressing an issue fixed in driver version 32.0.101.8860 and enhancing compatibility and stability.

github.com

Thursday, July 30, 2026

ai_labs67

GPU Management: Why Idle GPUs Are the New Grounded Aircraft

Companies are now treating idle graphics processing units (GPUs) as a resource to be managed, similar to how grounded aircraft are used more efficiently during downtime. This approach is crucial for optimizing operational costs and improving overall efficiency in industries reliant on GPU usage.

huggingface.co

Tuesday, July 28, 2026

releases48

ggml/llama.cpp releases: b10172

ggml-webgpu updates address binding alias issues and overlapping range problems across different architectures, ensuring compatibility and stability. These fixes are crucial for maintaining the project's reliability across various computing environments.

github.com

Monday, July 27, 2026

Thursday, July 23, 2026

releases54

Ollama releases: v0.32.3

github.com

Wednesday, July 22, 2026

🌿 That's all for now. Come back tomorrow.

11 of 11 items shown. Sources: 76 days indexed.