Topic: under

20 stories found

Yesterday

industry32

The Download: our 35 Innovators Under 35 this year

The MIT Technology Review has announced its 35 Innovators Under 35 list for 2026, highlighting young scientists and technologists driving significant advancements. This list showcases the future leaders in science and technology whose innovations could shape the tech landscape.

technologyreview.com↗

Monday, September 7, 2026

research40

Towards Understanding Pause Token Fine-Tuning Dynamics: A Mode Retention Perspective

Researchers explore the training dynamics of pause-token methods in large language models (LLMs), focusing on how these tokens affect reasoning processes, and argue that understanding their fine-tuning dynamics offers insights beyond just computational expressivity. This matters because it could lead to more effective and efficient ways to enhance LLMs' reasoning capabilities.

arxiv.org↗

Sunday, September 6, 2026

industry24

Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed

Perplexity shared details about its GPU-based embedding stack, including Ivy, Tulip, and ROSE, to enhance retrieval quality in AI search products by improving how cheaply embeddings can be run across indexes. This matters because optimizing embedding serving infrastructure can significantly boost the efficiency and performance of AI applications.

marktechpost.com↗

Saturday, September 5, 2026

releases48

ggml/llama.cpp releases: b10817

The ggml/llama.cpp project released version b10817, which includes enhancements for tracking memory allocations using new environment variables. These changes are crucial for optimizing the memory management in the --fit algorithm and aiding in debugging other parts of the system.

github.com↗
industry24

Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by Up to 88%

Google introduced Agentic Video Understanding in Gemini flash models, enabling the platform to navigate videos efficiently rather than processing them at 1 FPS, thereby reducing video tokens by up to 88% and improving performance. This update is significant as it enhances Gemini's efficiency and responsiveness when handling video content.

marktechpost.com↗

Friday, September 4, 2026

releases48

ggml/llama.cpp releases: b10795

The ggml/llama.cpp project updated its codebase to fuse RMS_NORM+MUL+ADD and ADD+ADD operations under SYCL fusion, enhancing computational efficiency. This update is significant as it optimizes the handling of addition operations in a way that maintains consistency with standalone add() functions, potentially improving performance in large language model training.

github.com↗
research40

Counterexamples as Feedback for Agent Self-Correction

A new framework called A-CEGIS has been developed to help artificial intelligence agents self-correct by using counterexamples as feedback, addressing limitations in current single-turn metrics which fail to assess the agents' ability to repair mistakes. This matters because it enhances the reliability and adaptability of AI systems in real-world applications where initial errors need correction.

arxiv.org↗

Thursday, September 3, 2026

ai_labs75

Safety overview: GPT-6 Astra

GPT-6 Astra has achieved the Critical level of cybersecurity, making it the company's most secure widely used model. This milestone underscores the firm's commitment to enhancing safety in their advanced language models.

openai.com↗

20 of 20 items shown. Sources: 110 days indexed.