Topic: research

27 stories found

Yesterday

open_source62

The-Art-of-Hacking/h4cker — This repository is maintained by Omar Santos (@santosomar) and includes thousands of resources related to ethical hackin

Omar Santos maintains a repository with thousands of resources for ethical hacking and related fields like bug bounties and digital forensics, aiming to support professionals in these areas. The extensive collection covers topics such as AI security and exploit development, highlighting the importance of responsible cybersecurity practices.

github.com

Tuesday, September 22, 2026

ai_labs75

Parallel cut research time and cost in half with GPT‑6 Astra

Parallel reduced research time and costs by half using GPT-6 Astra, cutting the duration and expenses of analyzing labor-market data in half compared to previous models. This achievement highlights the potential of advanced AI in enhancing efficiency and reducing expenditures in data-intensive tasks.

openai.com

Monday, September 21, 2026

releases60

NousResearch releases: Hermes Agent v0.21.4 (v2026.9.21)

Hermes Agent version v0.21.4 was released on September 21, 2026, consolidating approximately 1,800 merged pull requests to provide a stable release for various deployment environments, including Docker images and Hermes Cloud services.

github.com
research40

Boosting Deepresearch and LongContext Ability with Self-Generated Deepresearch Rollouts Traces

A new approach aims to enhance deepresearch and long-context abilities by using self-generated rollouts traces, addressing limitations in current agentic reinforcement learning models which struggle with maintaining context over extended interactions. This matters because it could significantly improve how AI agents handle complex, multi-step tasks in dynamic environments.

arxiv.org

Saturday, September 19, 2026

open_source62

affaan-m/ECC — The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development f

A new agent harness performance optimization system focusing on skills, instincts, memory, security, and research-first development is being implemented for Claude Code, Codex, Opencode, Cursor, and other similar systems. This update aims to enhance overall efficiency and reliability by improving core functionalities and security measures.

github.com
industry24

Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model

Linkup Research has unveiled SPARSEUP, an open-source sparse embedding model with 149 million parameters that outperforms other models in its class on the BEIR-13 benchmark, highlighting advancements in efficient large-scale language modeling.

marktechpost.com

Friday, September 18, 2026

ai_labs67

New experts join Google’s AI & Economy team

Google has expanded its AI & Economy team by adding world-class academic advisors, fellows, and internal researchers. This move aims to enhance the company's research and insights into the intersection of artificial intelligence and the economy.

blog.google
research40

What Users Think of Generative AI: A Cross-Platform NLP Analysis of Trust and Friction in App Store Reviews

A study analyzed app store reviews to understand users' perceptions of generative AI (GenAI) applications, focusing on trust levels and adoption challenges, highlighting a lack of large-scale research in this area.

arxiv.org

Wednesday, September 16, 2026

ai_labs75

How workers are unlocking new ways of working

The study reveals that workers are integrating AI into various tasks beyond their conventional roles, making AI a regular part of many jobs. This shift highlights the growing importance of AI in the workplace and its potential to transform job functions and responsibilities.

openai.com
research40

Comment on arXiv:2607.01233: Survivorship Bias in Published-Paper Baselines for Research-Idea Distributions

Chen, Zhao, and Cohan's study evaluates LLM-generated research ideas but faces criticism for survivorship bias in its human baseline, which includes only published papers, while the LLM baseline considers one-shot responses. This discrepancy highlights a potential flaw in how the LLM's performance is being compared to real-world scenarios.

arxiv.org

27 of 27 items shown. Sources: 124 days indexed.