Topic: models

61 stories found

Today

releases42

vLLM releases: v0.30.0

vLLM released version 0.30.0 with significant contributions from many developers, introducing new models like DeepSeek-V4.1-Flash and DeepGEMM Mega-mHC, marking a substantial update in the model's capabilities.

github.com
research40

Token Signatures of Code: Comparing Coding Behaviors Across Large Language Models

A new study evaluates large language models (LLMs) based on their coding behaviors rather than just performance metrics like pass@k, highlighting that as models improve, traditional evaluation methods become less effective in distinguishing between them.

arxiv.org

Yesterday

trending59

macOS 27: Workaround to avoid downloading AI models and save storage

Apple released an update for macOS that includes a workaround to prevent the automatic download of AI models, helping users conserve storage space. This update is significant as it addresses user concerns about storage usage by AI features without compromising functionality.

reddit.com
newsletters48

Jev: System One models for Prod, not God — with Diogo Almeida, CEO, TypeSafe AI

The Jev podcast discusses the use of System One models in production environments rather than attributing decisions to a divine entity, emphasizing practicality over mysticism. This matters as it clarifies the application and purpose of AI models in real-world scenarios, promoting clearer communication about technology's role.

latent.space
research40

Do small language models know what they don't know?

Researchers investigated ways to enhance the accuracy of small language models (with less than 3 billion parameters) using entropy-based confidence signals, finding potential improvements for models running on consumer hardware. This matters because it could make advanced language capabilities more accessible on standard devices.

arxiv.org
research35

Decoupling Internal Representational Changes and Causal Importance in Fine-Tuned Large Language Models

Researchers have explored how fine-tuning large language models changes their internal representations without affecting their causal importance, aiming to better understand the mechanism behind model adaptation for various tasks. This study is crucial as it helps in optimizing and interpreting the behavior of fine-tuned LLMs more effectively.

arxiv.org

Sunday, September 20, 2026

trending62

Pirate Face Rescues LLM Models from Deletion

A viral social media post featuring a "pirate face" meme saved multiple large language model (LLM) entries from being deleted by highlighting their cultural significance and humor value. This incident underscores the power of internet memes in preserving digital content and challenging mass deletions online.

pirateface.co
open_source62

langgenius/dify — Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cl

Agentic offers a collaborative workspace for building AI workflows and RAG pipelines with extensive model and tool support, enabling seamless deployment options from cloud to self-hosted environments, facilitating rapid transition from prototypes to production.

github.com

Friday, September 18, 2026

research40

Subliminal Prompting Beyond Static Geometry: Causal Depth and Multi-Token Confounds

A recent study suggests that language models can subtly convey hidden traits in their outputs, even when those outputs seem unrelated, challenging current explanations like token entanglement. This finding is significant as it deepens our understanding of subliminal learning and the causal mechanisms within language models.

arxiv.org

Thursday, September 17, 2026

research40

Faking Good and Faking Bad in LLMs: Response Distortion Across Dark Triad Personality Traits

The study explores how large language models (LLMs) are influenced by social desirability and impression management, similar to humans during personality assessments, highlighting the need for better understanding of response distortions in AI.

arxiv.org

61 of 61 items shown. Sources: 123 days indexed.