Topic: token
15 stories found
Today
Token Signatures of Code: Comparing Coding Behaviors Across Large Language Models
A new study evaluates large language models (LLMs) based on their coding behaviors rather than just performance metrics like pass@k, highlighting that as models improve, traditional evaluation methods become less effective in distinguishing between them.
Saturday, September 19, 2026
TypeSafe AI Releases Jev: A System One Model That Returns Typed, Calibrated Decisions Instead of Text
TypeSafe AI introduced Jev, a System One model that provides typed, calibrated responses with probabilities rather than text, aiming to offer more precise decision-making tools for developers. This release is significant as it could enhance the reliability and utility of AI in applications requiring probabilistic outputs over textual answers.
Friday, September 18, 2026
Subliminal Prompting Beyond Static Geometry: Causal Depth and Multi-Token Confounds
A recent study suggests that language models can subtly convey hidden traits in their outputs, even when those outputs seem unrelated, challenging current explanations like token entanglement. This finding is significant as it deepens our understanding of subliminal learning and the causal mechanisms within language models.
Show HN: Agentgit – a Git host for AI agents, no account, no token, no key
Agentgit is introduced as a new Git hosting service specifically designed for AI agents, offering seamless access without requiring accounts, tokens, or keys. This innovation aims to simplify the development and collaboration process for AI projects by removing traditional authentication barriers.
Wednesday, September 16, 2026
JuliusBrussee/caveman — 🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking l
A viral method reduces the number of tokens needed in coding by mimicking caveman-like speech, cutting usage by 65%, showcasing an efficient alternative to complex language in programming.
The Functionalizer: Lossless Functional Decomposition for Subword Tokenization
A new method called The Functionalizer for subword tokenization has been introduced to address limitations of existing techniques by decomposing words functionally rather than orthographically, thus preserving the embedding space without loss. This advancement matters because it improves the consistency and efficiency of natural language processing models.
Monday, September 14, 2026
Ollama releases: v0.34.1
Ollama released version 0.34.1, which marks the full integration of MLX safetensors support in their `ollama create` command and introduces stricter conditions for detecting runaway repeat tokens. These updates enhance model handling on Apple Silicon and improve overall stability.
Saturday, September 12, 2026
Friday, September 11, 2026
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction
A new latent-space language model called NCP-ArchPreview has been introduced, which extends autoregressive pretraining by incorporating Next Concept Prediction (NCP) alongside standard next-token prediction. This advancement aims to enhance the model's ability to understand and generate more complex linguistic concepts.
Thursday, September 10, 2026
X-CoSD: Communication-Efficient Cross-Vocabulary Collaborative Speculative Decoding
A new paper proposes X-CoSD, a communication-efficient method for collaborative speculative decoding that involves an on-device small language model generating candidates while a server large language model verifies them, aiming to improve distributed inference processes in large language models. This approach is crucial as it could enhance the efficiency and scalability of AI applications across devices.
15 of 15 items shown. Sources: 123 days indexed.