Topic: ci

105 stories found

Today

releases48

ggml/llama.cpp releases: b10276

The ggml/llama.cpp project released version b10276, which includes a preference for using npm ci over install for enhanced security. This update is important as it addresses best practices in package management to reduce potential vulnerabilities.

github.com
research40

Preferred, Not Safer: Pairwise Preference Is a Poor Proxy for Clinical Safety

The study finds that clinicians' pairwise preferences do not reliably indicate the clinical safety of large language models, suggesting that other methods may be needed for accurate safety assessments. This matters because ensuring the safety of AI tools in healthcare is crucial, but current evaluation methods might be insufficient.

arxiv.org

Yesterday

ai_labs75

Third-party cyber evaluations involving OpenAI models

OpenAI addressed security assessments of its AI models, acknowledging recent cybersecurity evaluations and implementing new safety measures. These steps are crucial to enhance the reliability and security of their AI systems.

openai.com
releases48

ggml/llama.cpp releases: b10254

New templates for DeepSeek V4 Flash 0731 were added to ggml/llama.cpp, aligning them with official encoders and updating the chat functionality. This update ensures compatibility and improves parser behavior without altering existing settings.

github.com
research40

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

MemoryForge aims to equip large language models with human-like memory capabilities to enhance their performance in agentic tasks like role-play and user simulation by moving beyond static textual profiles. This advancement is crucial as it allows LLMs to better mimic human behavior and interactions.

arxiv.org

Monday, August 3, 2026

ai_labs75

Circles powers telco personalization with OpenAI technology

Circles implemented OpenAI's API and Codex to enhance personalized telecom services, boosting average revenue per user by 22% and lowering customer churn by 9%, while also improving development processes.

openai.com
newsletters48

Import AI 467: Self-sustaining AI viruses; pacing AI progress; confusion about AI and creativity

The article discusses self-sustaining AI viruses, the need for better pacing in AI development to address societal concerns, and misunderstandings about AI's role in creative processes. The topics highlight challenges and considerations as AI technology advances rapidly.

importai.substack.com
research40

Imbalanced Data Clustering via Targeted Data Augmentation Using GMM and LLM

A new method using Gaussian Mixture Models and Large Language Models has been proposed to address the issue of imbalanced data clustering in Natural Language Processing, particularly for underrepresented topics in unsupervised tasks. This approach aims to improve the accuracy of clustering algorithms by augmenting targeted data, making it crucial for enhancing the handling of minority topics in NLP.

arxiv.org

Sunday, August 2, 2026

open_source31

Show HN: CostPerPrompt – Live AI API pricing and real-workload cost calculators

CostPerPrompt offers live API pricing and real-workload cost calculators for AI, helping developers and businesses estimate costs accurately. This tool is crucial for managing budgets in the rapidly growing AI industry by providing transparent and precise cost information.

costperprompt.com
industry24

End-to-End Forecasting with TimesFM 2.5: Backtesting, Covariates, Anomaly Detection, and Scalable Colab Deployment

A new tutorial details how to create an advanced time-series forecasting system using TimesFM 2.5, focusing on backtesting, covariates, anomaly detection, and scalable Colab deployment; this matters because it provides a comprehensive workflow for improving forecast accuracy in retail datasets.

marktechpost.com

105 of 105 items shown. Sources: 77 days indexed.