AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,415 results
7 Aug 2026

In the era of base LLM scaling (2022-2024), I believed the LLM line of research would reach a capability plateau (as later seen with base LL…

ResearchDGX agent

In the era of base LLM scaling (2022-2024), I believed the LLM line of research would reach a capability plateau (as later seen with base LLMs). In late 2024, after the o3 test-time compute demo, I ch

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

SafetyDGX agent

arXiv:2608.05656v1 Announce Type: cross Abstract: Safety risks of AI are becoming increasingly evident in human interactions with AI technologies. The prominent approaches to evaluating these risks fa

6 Aug 2026

Advancing brain tumor research with privacy-first AI

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

The intersection of medicine and AI has led to remarkable innovations. However, developers now face the thorny challenge of building robust medical AI tools that have been tested and evaluated on dive

29 Jul 2026

My LLM kept implementing every method it found, so I added research and specification gates[D]

AgentsDGX agent

While building this workflow a thing that surprised me was that, initially I thought the pipeline was complete: From Goal to → Decompose → Research → Specification → Implementation It successfully bro

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

Model ReleasesDGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

24 Jul 2026

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

Model ReleasesDGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

LegalCiteTrust: Benchmarking Citation Trustworthiness in Chinese Long-Form Legal Research Reports

Model ReleasesDGX agent

arXiv:2607.20872v1 Announce Type: new Abstract: Long-form legal research reports increasingly rely on LLMs and agentic research systems, but their reliability depends not only on answering the task, b

AREX: Towards a Recursively Self-Improving Agent for Deep Research

Model ReleasesDGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

3 Jul 2026

Google DeepMind and A24 announce first-of-its-kind research partnership

Model ReleasesDGX agent

Google DeepMind and A24 announced a first-of-its-kind research partnership pairing the AI research lab with the filmmaker-focused studio to help artists develop new workflows and techniques. Google is

22 May 2026

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

ResearchDGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

19 May 2026

AI for Auto-Research: Roadmap & User Guide

Model ReleasesDGX agent

arXiv:2605.18661v1 Announce Type: new Abstract: AI-assisted research is crossing a threshold: fully automated systems can now generate research papers for as little as $15, while long-horizon agents c

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

Model ReleasesDGX agent

arXiv:2605.17554v1 Announce Type: new Abstract: Frontier deep research agents (DRAs) plan a research task, synthesize across documents, and return a structured deliverable on demand. They are being de

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

Model ReleasesDGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

Vector RAG vs LLM-Compiled Wiki: A Preregistered Comparison on a Small Multi-Domain Research

ResearchDGX agent

arXiv:2605.18490v1 Announce Type: new Abstract: We preregistered a comparison of two ways to help an LLM answer questions over a small research corpus: a single-round Vector RAG system and an LLM-comp

10 May 2026

Join the Nous Research team for another Hermes Agent Jam in our Discord This one will be an interactive session, so come prepared to discuss…

AgentsDGX agent

Nous Research is hosting an interactive Hermes Agent Jam session in their Discord community where participants can discuss and collaborate on Hermes agent development. The event invites community memb

8 May 2026

Disillusionment with mechanistic interpretability research [D]

ResearchDGX agent

Mechanistic interpretability research aims to uncover specific neurons and circuits in neural networks responsible for tasks, but over a decade of efforts suggests these findings may not translate int

7 May 2026

We started running monthly research sessions at the @agi_inc office. First one was last Friday on PNM compilation. If you do on-device resea…

Local AiDGX agent

We started running monthly research sessions at the @agi_inc office. First one was last Friday on PNM compilation. If you do on-device research and want to come present, apply here: https://form.typef

1 May 2026

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

Model ReleasesDGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

本日付で、Sakana AI @SakanaAILabs に Applied Research Engineer として入社いたしました! 日本の AI 開発をより良く前進させてゆくための、価値ある貢献をしてゆきたいと思っております!

ResearchDGX agent

David Ha announced joining Sakana AI as an Applied Research Engineer, expressing his commitment to making valuable contributions to advancing AI development in Japan. This represents a significant hir

28 Apr 2026

Evaluating whether AI models would sabotage AI safety research

Model ReleasesDGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

ReFinE: Streamlining UI Mockup Iteration with Research Findings

TutorialsDGX agent

arXiv:2604.04353v2 Announce Type: replace-cross Abstract: Although HCI research papers offer valuable design insights, designers often struggle to apply them in design workflows due to difficulties in

23 Apr 2026

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

AgentsDGX agent

arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and

pAI/MSc: ML Theory Research with Humans on the Loop

AgentsDGX agent

arXiv:2604.20622v1 Announce Type: new Abstract: We present pAI/MSc, an open-source, customizable, modular multi-agent system for academic research workflows. Our goal is not autonomous scientific idea

22 Apr 2026

On Accelerating Grounded Code Development for Research

ResearchDGX agent

arXiv:2604.19022v1 Announce Type: new Abstract: A major challenge for niche scientific and technical domains in leveraging coding agents is the lack of access to up-to-date, domain- specific knowledge

15 Apr 2026

El Agente Quntur: A research collaborator agent for quantum chemistry

AgentsDGX agent

arXiv:2602.04850v2 Announce Type: replace-cross Abstract: Quantum chemistry is a foundational enabling tool for the fields of chemistry, materials science, computational biology and others. Despite of

14 Apr 2026

https://x.com/nousresearch/status/2043969403247616478?s=46

ResearchDGX agent

Nous Research shared an announcement or update on their X (formerly Twitter) account, likely relating to their ongoing work in AI model development, fine-tuning, or research releases. Nous Research is

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

Model ReleasesDGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

12 Apr 2026

Marcus Hutchins, the guy famous for stopping the WannaCry Ransomware, probably has the best take on Mythos doing vulnerability research

ResearchDGX agent

Marcus Hutchins, the cybersecurity researcher known for halting the 2017 WannaCry ransomware attack by registering a kill-switch domain, is cited as offering a notable perspective on AI systems conduc

8 Apr 2026

https://portal.nousresearch.com/

ResearchDGX agent

The Nous Research Portal (portal.nousresearch.com) is the account and API management hub for Nous Research's inference offerings. Nous Research launched an Inference API serving its open-source mo...

4 Aug 2026

DODA: A Database of Datasets for Aesthetics Research

ResearchDGX agent

arXiv:2608.00089v1 Announce Type: new Abstract: With rapid growth in the fields of empirical and computational aesthetics we have seen a vast increase in large image datasets annotated for aesthetics.

3 Aug 2026

8月より、Sakana AI @SakanaAILabs にApplied Research Engineer Internship として入社しました 🐟🐠🐡! 大学の夏季休業期間にフルタイム勤務予定です! LLMの研究開発と社会実装を頑張ります💪🏻

ResearchDGX agent

Horiyuki 'horiyuki42' joined Sakana AI Labs in August 2026 as an Applied Research Engineer Intern. He plans to work full‑time during the university summer recess while concentrating on large‑language‑

2 Aug 2026

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skil…

AgentsDGX agent

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skills from OpenAI, Anthropic, HuggingFace & more. Thank me late

31 Jul 2026

Baikal: Structured Search for Deep Research over Data Lakes

Model ReleasesDGX agent

arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a repo

28 Jul 2026

Position: AI/ML Deepfake Research is Misaligned with AI-Generated Non-Consensual Intimate Imagery (AIG-NCII)

SafetyDGX agent

arXiv:2607.18263v2 Announce Type: replace Abstract: AI-generated non-consensual intimate imagery (AIG-NCII) is not adequately addressed in AI/ML literature regarding AI-generated media, commonly refer

15 Jul 2026

FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

Model ReleasesDGX agent

arXiv:2607.12252v1 Announce Type: new Abstract: Deep research agents are increasingly used to produce long-form financial reports, yet large-scale evaluation remains bottlenecked by the need for human

7 Jul 2026

How Imperial College London is accelerating dementia research with a modern data platform

IndustryDGX agent

Imperial College London is leveraging a modern data platform to accelerate dementia research by enabling faster analysis of large-scale datasets and improved collaboration among researchers. The platf

29 Jun 2026

CoreWeave debuts ARIA agent to automate AI research in Weights & Biases

AgentsDGX agent

Artificial intelligence cloud operator CoreWeave Inc. today launched ARIA, an AI research agent built into the Weights & Biases platform. The agent reads experiment data and surfaces insights research

23 Jun 2026

I built a @NousResearch Hermes workflow that my team at @Box uses to track AI trends. Every day at 7am, it researches what's happening acros…

AgentsDGX agent

I built a @NousResearch Hermes workflow that my team at @Box uses to track AI trends. Every day at 7am, it researches what's happening across the AI ecosystem and generates a brief for our team. We us

USAID money funded coronavirus research in China that killed millions of people

IndustryDGX agent

USAID money funded coronavirus research in China that killed millions of people .@RandPaul Asks Samantha Power: 'Did USAID Fund Coronavirus Research In Wuhan China?' 'Should we be funding the Academy

11 Jun 2026

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Model ReleasesDGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

6 Jun 2026

LLM Research Papers: The 2026 List (January to May)

ResearchDGX agent

This resource compiles significant large language model research papers published between January and May 2026, curated by Sebastian Raschka. It serves as a reference guide for tracking recent advance

5 Jun 2026

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

ResearchDGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

Some billionaires (or their foundations) do fund certain areas of basic research. Examples: Simons Foundation, Moore Foundation, Sloan Found…

ResearchDGX agent

Some billionaires (or their foundations) do fund certain areas of basic research. Examples: Simons Foundation, Moore Foundation, Sloan Foundation, Keck Foundation, Schmidt Sciences, and several others

2 Jun 2026

@DavidSacks If Trump really were 'the most pro-innovation president we’ve ever had' he would not attempt to cut research budgets by half.

ResearchDGX agent

Yann LeCun criticized a claim that Trump is the most pro-innovation president by pointing out that proposed cuts to research budgets would contradict genuine support for innovation. The post engages w

ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment

Model ReleasesDGX agent

arXiv:2606.00644v1 Announce Type: new Abstract: AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be

1 Jun 2026

Developing a Culturally Grounded, AI-Augmented UX Research Point of View (POV): An Exemplar Case Study from Telemedicine Dementia Care

TutorialsDGX agent

arXiv:2605.31147v1 Announce Type: cross Abstract: User Experience Research (UXR) Points of View (POVs) distil complex and often fragmented research evidence into actionable perspectives that guide how

Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship

Model ReleasesDGX agent

arXiv:2605.30947v1 Announce Type: new Abstract: LLM-based research agents have advanced rapidly in science and engineering, where research is organized around executable experiments, code, and quantit

29 May 2026

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

Model ReleasesDGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

28 May 2026

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

ResearchDGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release…

ResearchDGX agent

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release stable-worldmodel: an open-source, scalable platform built

27 May 2026

Learning to Predict Future-Aligned Research Proposals with Language Models

Model ReleasesDGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

26 May 2026

Re-defining Humor Data Objects for AI Humor Research

ResearchDGX agent

arXiv:2605.25171v1 Announce Type: new Abstract: In most existing AI humor research, humor was treated as either 'present' or 'not present.' We explore the concept of humor as a social interaction with

25 May 2026

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery

AgentsDGX agent

arXiv:2605.23204v1 Announce Type: new Abstract: Scientific research is being reshaped by AI systems that move beyond isolated assistance toward longer-horizon workflows spanning literature grounding,

21 May 2026

ACL-Verbatim: hallucination-free question answering for research

Model ReleasesDGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

13 May 2026

AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.11732v1 Announce Type: cross Abstract: In this paper, we present AgentDisCo, a novel Disentangled and Collaborative agentic architecture that formulates deep research as an adversarial opti

12 May 2026

NanoResearch: Co-Evolving Skills, Memory, and Policy for Personalized Research Automation

Model ReleasesDGX agent

arXiv:2605.10813v1 Announce Type: new Abstract: LLM-powered multi-agent systems can now automate the full research pipeline from ideation to paper writing, but a fundamental question remains: automati

9 May 2026

Europe does not lack innovation. It lacks scale. European universities produce world-class research, engineers and technology. But too many …

ResearchDGX agent

Europe does not lack innovation. It lacks scale. European universities produce world-class research, engineers and technology. But too many companies remain trapped inside fragmented national markets

4 May 2026

Structure Liberates: How Constrained Sensemaking Produces More Novel Research Output

ResearchDGX agent

arXiv:2605.00557v1 Announce Type: new Abstract: Scientific discovery is an extended process of ideation--surveying prior work, forming hypotheses, and refining reasoning--yet existing approaches treat

3 May 2026

Everyone interested in getting the most out of Hermes Agent should definitely consider joining the Nous Research discord - We have a forum c…

AgentsDGX agent

Everyone interested in getting the most out of Hermes Agent should definitely consider joining the Nous Research discord - We have a forum channel for plugins skills and skins and another for communit

29 Apr 2026

Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research

SafetyDGX agent

arXiv:2604.25776v1 Announce Type: new Abstract: Critical analyses of emotion recognition technology have raised ethical concerns around task validity and potential downstream impacts, urging researche

← Previous
12345…424
Next →