AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
21,937 results
Safety

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

DGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

safetyarxiv-cs-ai
11 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Planner-Centric Reinforcement Learning for Deep Research with Structure-Aware Reward

DGX agent

arXiv:2605.30824v1 Announce Type: new Abstract: Deep research tasks require LLMs to plan what to investigate, retrieve evidence, and synthesize long-form answers across multiple branches of inquiry. E

researcharxiv-cs-ai
1 Jun 2026
Safety

Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams

DGX agent

arXiv:2605.27404v1 Announce Type: cross Abstract: The era of Big Science has long been defined by increasingly large and specialized research teams pushing the frontiers of knowledge. However, recent

safetyarxiv-cs-ai
28 May 2026
Model Releases

RMA: an Agentic System for Research-Level Mathematical Problems

DGX agent

arXiv:2605.22875v1 Announce Type: new Abstract: We present extbf{Research Math Agents (RMA)}, an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark

DGX agent

arXiv:2509.26574v4 Announce Type: replace Abstract: While large language models (LLMs) with reasoning capabilities are progressing rapidly on high-school math competitions and coding, can they reason

model-releasesarxiv-cs-ai
12 May 2026
Agents

SciDER: Scientific Data-centric End-to-end Researcher

DGX agent

arXiv:2603.01421v2 Announce Type: replace-cross Abstract: Automated scientific discovery with large language models is transforming the research lifecycle from ideation to experimentation, yet existin

agentsarxiv-cs-cl
29 Apr 2026
Research

Aligning Stuttered-Speech Research with End-User Needs: Scoping Review, Survey, and Guidelines

DGX agent

arXiv:2604.20535v1 Announce Type: new Abstract: Atypical speech is receiving greater attention in speech technology research, but much of this work unfolds with limited interdisciplinary dialogue. For

researcharxiv-cs-cl
23 Apr 2026
Research

Construction of a Battery Research Knowledge Graph using a Global Open Catalog

DGX agent

arXiv:2604.20241v1 Announce Type: new Abstract: Battery research is a rapidly growing and highly interdisciplinary field, making it increasingly difficult to track relevant expertise and identify pote

researcharxiv-cs-cl
23 Apr 2026
Model Releases

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

DGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Coding-Free and Privacy-Preserving MCP Framework for Clinical Agentic Research Intelligence System

DGX agent

arXiv:2604.12258v1 Announce Type: cross Abstract: Clinical research involves labor-intensive processes such as study design, cohort construction, model development, and documentation, requiring domain

agentsarxiv-cs-ai
15 Apr 2026
Agents

Automating and Scaling Behavioral Scientific Research on AI Agents

DGX agent

arXiv:2608.10030v1 Announce Type: new Abstract: As AI agents are increasingly deployed in complex environments, understanding their behaviors becomes critical. Yet behavioral scientific research on AI

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

Nutrition Data Infrastructure for the AI Era: Operationalizing FAIR for Agent-Mediated Research

DGX agent

arXiv:2608.10363v1 Announce Type: new Abstract: AI agents can accelerate nutrition research, but their analyses inherit the identity, semantic, and release ambiguities of the underlying data. We prese

model-releasesarxiv-cs-ai
12 Aug 2026
Safety

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

DGX agent

arXiv:2608.05656v1 Announce Type: cross Abstract: Safety risks of AI are becoming increasingly evident in human interactions with AI technologies. The prominent approaches to evaluating these risks fa

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

DGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

LegalCiteTrust: Benchmarking Citation Trustworthiness in Chinese Long-Form Legal Research Reports

DGX agent

arXiv:2607.20872v1 Announce Type: new Abstract: Long-form legal research reports increasingly rely on LLMs and agentic research systems, but their reliability depends not only on answering the task, b

model-releasesarxiv-cs-cl
24 Jul 2026
Research

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

DGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

researcharxiv-cs-ro
22 May 2026
Model Releases

AI for Auto-Research: Roadmap & User Guide

DGX agent

arXiv:2605.18661v1 Announce Type: new Abstract: AI-assisted research is crossing a threshold: fully automated systems can now generate research papers for as little as $15, while long-horizon agents c

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

DGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Evaluating whether AI models would sabotage AI safety research

DGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

ReFinE: Streamlining UI Mockup Iteration with Research Findings

DGX agent

arXiv:2604.04353v2 Announce Type: replace-cross Abstract: Although HCI research papers offer valuable design insights, designers often struggle to apply them in design workflows due to difficulties in

tutorialsarxiv-cs-ai
28 Apr 2026
Agents

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

DGX agent

arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and

agentsarxiv-cs-ai
23 Apr 2026
Agents

pAI/MSc: ML Theory Research with Humans on the Loop

DGX agent

arXiv:2604.20622v1 Announce Type: new Abstract: We present pAI/MSc, an open-source, customizable, modular multi-agent system for academic research workflows. Our goal is not autonomous scientific idea

agentsarxiv-cs-ai
23 Apr 2026
Research

On Accelerating Grounded Code Development for Research

DGX agent

arXiv:2604.19022v1 Announce Type: new Abstract: A major challenge for niche scientific and technical domains in leveraging coding agents is the lack of access to up-to-date, domain- specific knowledge

researcharxiv-cs-ai
22 Apr 2026
Agents

El Agente Quntur: A research collaborator agent for quantum chemistry

DGX agent

arXiv:2602.04850v2 Announce Type: replace-cross Abstract: Quantum chemistry is a foundational enabling tool for the fields of chemistry, materials science, computational biology and others. Despite of

agentsarxiv-cs-ai
15 Apr 2026
Research

DODA: A Database of Datasets for Aesthetics Research

DGX agent

arXiv:2608.00089v1 Announce Type: new Abstract: With rapid growth in the fields of empirical and computational aesthetics we have seen a vast increase in large image datasets annotated for aesthetics.

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Baikal: Structured Search for Deep Research over Data Lakes

DGX agent

arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a repo

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

DGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

model-releasesarxiv-cs-cl
29 Jul 2026
Safety

Position: AI/ML Deepfake Research is Misaligned with AI-Generated Non-Consensual Intimate Imagery (AIG-NCII)

DGX agent

arXiv:2607.18263v2 Announce Type: replace Abstract: AI-generated non-consensual intimate imagery (AIG-NCII) is not adequately addressed in AI/ML literature regarding AI-generated media, commonly refer

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

AREX: Towards a Recursively Self-Improving Agent for Deep Research

DGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

DGX agent

arXiv:2607.12252v1 Announce Type: new Abstract: Deep research agents are increasingly used to produce long-form financial reports, yet large-scale evaluation remains bottlenecked by the need for human

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

DGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

DGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

researcharxiv-cs-cl
5 Jun 2026
Model Releases

ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment

DGX agent

arXiv:2606.00644v1 Announce Type: new Abstract: AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Developing a Culturally Grounded, AI-Augmented UX Research Point of View (POV): An Exemplar Case Study from Telemedicine Dementia Care

DGX agent

arXiv:2605.31147v1 Announce Type: cross Abstract: User Experience Research (UXR) Points of View (POVs) distil complex and often fragmented research evidence into actionable perspectives that guide how

tutorialsarxiv-cs-ai
1 Jun 2026
Model Releases

Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship

DGX agent

arXiv:2605.30947v1 Announce Type: new Abstract: LLM-based research agents have advanced rapidly in science and engineering, where research is organized around executable experiments, code, and quantit

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

DGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

model-releasesarxiv-cs-lg
29 May 2026
Research

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

DGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

researcharxiv-cs-cl
28 May 2026
Model Releases

Learning to Predict Future-Aligned Research Proposals with Language Models

DGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

model-releasesarxiv-cs-cl
27 May 2026
Research

Re-defining Humor Data Objects for AI Humor Research

DGX agent

arXiv:2605.25171v1 Announce Type: new Abstract: In most existing AI humor research, humor was treated as either 'present' or 'not present.' We explore the concept of humor as a social interaction with

researcharxiv-cs-cl
26 May 2026
Agents

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery

DGX agent

arXiv:2605.23204v1 Announce Type: new Abstract: Scientific research is being reshaped by AI systems that move beyond isolated assistance toward longer-horizon workflows spanning literature grounding,

agentsarxiv-cs-ai
25 May 2026
Model Releases

ACL-Verbatim: hallucination-free question answering for research

DGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

DGX agent

arXiv:2605.17554v1 Announce Type: new Abstract: Frontier deep research agents (DRAs) plan a research task, synthesize across documents, and return a structured deliverable on demand. They are being de

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

DGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

model-releasesarxiv-cs-lg
19 May 2026
Research

Vector RAG vs LLM-Compiled Wiki: A Preregistered Comparison on a Small Multi-Domain Research

DGX agent

arXiv:2605.18490v1 Announce Type: new Abstract: We preregistered a comparison of two ways to help an LLM answer questions over a small research corpus: a single-round Vector RAG system and an LLM-comp

researcharxiv-cs-cl
19 May 2026
Model Releases

AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents

DGX agent

arXiv:2605.11732v1 Announce Type: cross Abstract: In this paper, we present AgentDisCo, a novel Disentangled and Collaborative agentic architecture that formulates deep research as an adversarial opti

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

NanoResearch: Co-Evolving Skills, Memory, and Policy for Personalized Research Automation

DGX agent

arXiv:2605.10813v1 Announce Type: new Abstract: LLM-powered multi-agent systems can now automate the full research pipeline from ideation to paper writing, but a fundamental question remains: automati

model-releasesarxiv-cs-ai
12 May 2026
Research

Structure Liberates: How Constrained Sensemaking Produces More Novel Research Output

DGX agent

arXiv:2605.00557v1 Announce Type: new Abstract: Scientific discovery is an extended process of ideation--surveying prior work, forming hypotheses, and refining reasoning--yet existing approaches treat

researcharxiv-cs-cl
4 May 2026
← Previous
1234…458
Next →