AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,425 results
25 Jun 2026

Heuresis: Search Strategies for Autonomous AI Research Agents Across Quality, Diversity and Novelty

SafetyDGX agent

arXiv:2606.25198v1 Announce Type: new Abstract: Autonomous AI Research promises to accelerate the scientific progress of machine learning. To realise this goal, current Large Language Model (LLM)-base

24 Jun 2026

Agon: An Autonomous Large-Scale Omnidisciplinary Research System Built on Prompt Economy

AgentsDGX agent

arXiv:2606.24177v1 Announce Type: cross Abstract: Large language models are making research production scalable, shifting the bottleneck from producing artifacts to judging claims. We present extsc{Ag

11 Jun 2026

From Awareness to Action: Understanding and Overcoming the Research-Practice Gap in Algorithmic Fairness for Public Health

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

arXiv:2606.11214v1 Announce Type: cross Abstract: Algorithmic fairness is essential for responsible ML-driven public health research, yet its practical implementation remains limited. To investigate t

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

AgentsDGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

9 Jun 2026

GPU MODE has powered much of the public GPU kernel work online, with a permissive license from day one and generous credit from researchers,…

HardwareDGX agent

GPU MODE has powered much of the public GPU kernel work online, with a permissive license from day one and generous credit from researchers, NVIDIA, AMD, and others. Today we’re moving our datasets to

Locked in heated rivalry with researcher, Microsoft fixes 0-day they disclosed

IndustryDGX agent

Microsoft engaged in a public feud with security researcher Nightmare-Eclipse, who released multiple Windows zero-days along with proof-of-concept exploit code. Microsoft initially threatened legal ac

6 Jun 2026

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation

Model ReleasesDGX agent

arXiv:2606.05241v1 Announce Type: cross Abstract: Public benchmarks enable fair and reproducible evaluation of LLM reasoning, but they become fragile for deep research agents that actively search the

5 Jun 2026

Representing Research Attention as Contextually Structured Flows

Model ReleasesDGX agent

arXiv:2606.05895v1 Announce Type: new Abstract: Research attention is widely used as an indicator of visibility, influence, and societal uptake, yet it is typically represented as aggregated counts th

4 Jun 2026

The future of quantum takes center stage at NY Tech Week

ResearchDGX agent

IBM Research highlighted quantum computing developments and applications at NY Tech Week, showcasing the technology's potential impact on industry and research. The event featured discussions on quant

3 Jun 2026

Lingo_Research_Group at SemEval-2026 Task 9: Evaluating Prompt Variants for Polarization Detection

ResearchDGX agent

arXiv:2606.03334v1 Announce Type: new Abstract: Our submission presented in this paper is for SemEval-2026 Task 9: Multilingual Text Classification Challenge - Polarization Detection and it covers all

2 Jun 2026

ADRA-Bank: A Modular Benchmark for Academic Deep Research Agents

Model ReleasesDGX agent

arXiv:2512.00986v3 Announce Type: replace Abstract: A surge in academic publications calls for automated deep research (DR) systems, but accurately evaluating them is still an open problem. First, exi

Places in the Wild: A Large, High-Resolution RAW Photograph Dataset for Ecologically Valid Vision Research

ResearchDGX agent

arXiv:2606.02481v1 Announce Type: new Abstract: Large image datasets have accelerated progress in cognitive neuroscience and computer vision. However, most datasets are low-resolution, internet-source

TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation

Model ReleasesDGX agent

arXiv:2606.02320v1 Announce Type: new Abstract: Deep Research Agents have shown strong capability in multi-step information retrieval, reasoning, and long-form report generation, but existing benchmar

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

Model ReleasesDGX agent

arXiv:2606.02060v1 Announce Type: new Abstract: Deep-research agents solve tasks through long trajectories of search, tool use, evidence inspection, and answer synthesis. Evaluation based on final ans

29 May 2026

AI can give researchers the freedom to pursue “crazier” ideas. For Terence Tao, AI creates more room to experiment, test unexpected paths, a…

Model ReleasesDGX agent

AI tools are enabling researchers, including renowned mathematician Terence Tao, to explore unconventional and high-risk ideas by handling routine computational tasks and verification work. This techn

The Trust Paradox: How CS Researchers Engage LLM Leaderboards

Model ReleasesDGX agent

arXiv:2605.28966v1 Announce Type: new Abstract: Large language model (LLM) leaderboards rank AI models using standardized benchmarks and have become highly visible across computer science, despite kno

28 May 2026

FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research

AgentsDGX agent

arXiv:2605.27864v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in finance, yet most existing work emphasizes trading signals or financial NLP tasks centered on p

ResearchLoop: An Evidence-Gated Control Plane for AI-Assisted Research

ApplicationsDGX agent

arXiv:2605.28282v1 Announce Type: new Abstract: AI-assisted research compresses ideation, implementation, evaluation, and manuscript writing into a single interactive loop. This compression is useful,

Speaking of Language: Reflections on Metalanguage Research in NLP

ResearchDGX agent

arXiv:2604.02645v2 Announce Type: replace-cross Abstract: This work aims to shine a spotlight on the topic of metalanguage. We first define metalanguage, link it to NLP and LLMs, and then discuss our

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

ResearchDGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

27 May 2026

A Multivariate Bernoulli-Based Sampling Method for Multi-Label Data with Application to Meta-Research

ResearchDGX agent

arXiv:2512.08371v4 Announce Type: replace Abstract: Datasets may contain observations with multiple labels. If the labels are not mutually exclusive, and if the labels vary greatly in frequency, obtai

Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis

ResearchDGX agent

arXiv:2512.14561v2 Announce Type: replace Abstract: Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compare

GENESIS: Harnessing AI Agents for Autonomous 6G RAN Synthesis, Research, and Testing

AgentsDGX agent

arXiv:2605.27360v1 Announce Type: cross Abstract: Cellular research and development (R&D) is throttled by six structural processes that each consume months of manual engineering work per iteration: (i

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

Model ReleasesDGX agent

arXiv:2605.26340v1 Announce Type: new Abstract: Autonomous research agents produce competitive solutions and professional-looking manuscripts, yet their outputs contain verifiability failures undetect

26 May 2026

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

AgentsDGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

22 May 2026

DeepWeb-Bench: A Deep Research Benchmark Demanding Massive Cross-Source Evidence and Long-Horizon Derivation

Model ReleasesDGX agent

arXiv:2605.21482v1 Announce Type: new Abstract: Deep research, in which an agent searches the open web, collects evidence, and derives an answer through extended reasoning, is a prominent use case for

MARS: Modular Agent with Reflective Search for Automated AI Research

AgentsDGX agent

arXiv:2602.02660v3 Announce Type: replace Abstract: A critical bottleneck in automating AI research is the execution of complex machine learning engineering (MLE) tasks. MLE differs from general softw

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

Model ReleasesDGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

these guys built a research agent with activegraph (using their @monid_ai tool) and found that every claim was traced to a source - which is…

AgentsDGX agent

these guys built a research agent with activegraph (using their @monid_ai tool) and found that every claim was traced to a source - which is not prompted, but natively baked in to the approach (they a

20 May 2026

Add a Specialized Deep Research Skill to Agent Harnesses

Model ReleasesDGX agent

This article covers the NVIDIA AI-Q Blueprint for building specialized deep research agents that empower AI systems to gather context, synthesize information, and support complex decision-making acros

AffectAI-Capture: A Reproducible Multimodal Protocol for Small-Group Meeting Research

ResearchDGX agent

arXiv:2605.19794v1 Announce Type: cross Abstract: We present AffectAI-Capture, a protocol for collecting synchronized multimodal data in four-person meeting-like interactions, combining eye tracking,

Justifying bio-inspired robotics research: A taxonomy of strategies

ResearchDGX agent

arXiv:2605.19840v1 Announce Type: new Abstract: For most of human history, we have not thought systematically about how and why we incorporate aspects of the natural world into our designs. The lack o

19 May 2026

A Scalable Tool for Measuring Manner and Result Verbs in Developmental Language Research

ResearchDGX agent

arXiv:2605.16654v1 Announce Type: cross Abstract: Manner and result verbs encode different aspects of event structure and have been discussed in developmental work as a potentially informative distinc

The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions

SafetyDGX agent

arXiv:2603.01092v2 Announce Type: replace Abstract: Scientific discovery is constrained not only by what is true, but by what is cognitively available to the researchers currently exploring a field. M

18 May 2026

Argus: Evidence Assembly for Scalable Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.16217v1 Announce Type: cross Abstract: Deep research agents have achieved remarkable progress on complex information seeking tasks. Even long ReAct style rollouts explore only a single traj

Position: Ideas Should be the Center of Machine Learning Research

Model ReleasesDGX agent

arXiv:2605.15253v1 Announce Type: new Abstract: Machine learning research increasingly bifurcates into two disconnected modes: benchmark-driven engineering that prioritizes metrics over understanding,

The Hardness of Achieving Impact in AI for Social Impact Research: A Ground-Level View of Challenges & Opportunities

TutorialsDGX agent

arXiv:2506.14829v2 Announce Type: replace-cross Abstract: AI for Social Impact (AI4SI) is an emergent field harnessing interdisciplinarities between the fields of artificial intelligence (AI), machine

12 May 2026

Agentic MIP Research: Accelerated Constraint Handler Generation

Model ReleasesDGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

Can Deep Research Agents Retrieve and Organize? Evaluating the Synthesis Gap with Expert Taxonomies

Model ReleasesDGX agent

arXiv:2601.12369v3 Announce Type: replace Abstract: Deep Research Agents increasingly automate survey generation, yet whether they match human experts at retrieving essential papers and organizing the

11 May 2026

FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting

Model ReleasesDGX agent

arXiv:2603.19254v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in financial research workflows, where their role is evolving from single-model assistance fo

8 May 2026

Apple Workshop on Privacy-Preserving Machine Learning & AI 2026

ResearchDGX agent

At Apple, we believe privacy is a fundamental human right. As AI capabilities increase and become more integrated into people’s daily lives, advancing research in privacy-preserving techniques is incr

OpenAI introduces GPT‑5.5‑Cyber for high-impact cybersecurity research

Model ReleasesDGX agent

OpenAI Group PBC has developed a version of GPT-5.5 that is specifically optimized for cybersecurity research. GPT‑5.5‑Cyber, as the model is called, made its debut on Thursday. It’s available in limi

7 May 2026

AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research

ApplicationsDGX agent

arXiv:2512.16455v3 Announce Type: replace-cross Abstract: The rapid growth of Artificial Intelligence and Machine Learning in scientific research has highlighted a gap between industry-standard MLOps

Hermes🪽

ResearchDGX agent

Hermes is an AI model developed by Nous Research that focuses on instruction-following and reasoning capabilities. Based on Nous Research's focus, this likely covers the model's architecture, performa

4 May 2026

Even our toughest critics come around eventually

ResearchDGX agent

Nous Research likely discusses how their AI models or research have gained acceptance even among skeptical observers, suggesting that rigorous development and demonstrated capabilities eventually conv

1 May 2026

Intern-Atlas: A Methodological Evolution Graph as Research Infrastructure for AI Scientists

SafetyDGX agent

arXiv:2604.28158v1 Announce Type: new Abstract: Existing research infrastructure is fundamentally document-centric, providing citation links between papers but lacking explicit representations of meth

Mapping the Methodological Space of Classroom Interaction Research: Scale, Duration, and Modality in an Age of AI

TutorialsDGX agent

arXiv:2604.28098v1 Announce Type: new Abstract: Research on classroom interaction has long been divided between large-scale observation and in-depth ethnographic work. We propose a framework mapping t

30 Apr 2026

Thank you to @RugvedSomwanshi and the @lmstudio team for the PR!

ResearchDGX agent

Nous Research is acknowledging and thanking Rugved Somwanshi and the LM Studio team for their contribution via a pull request (PR) to a project, likely related to Nous Research's open-source model dev

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

Model ReleasesDGX agent

arXiv:2510.14438v2 Announce Type: replace Abstract: The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent

29 Apr 2026

CiteRadar: A Citation Intelligence Platform for Researcher Profiling and Geographic Visualization

ResearchDGX agent

arXiv:2604.25057v1 Announce Type: new Abstract: Understanding the geographic reach and community structure of one's scholarly citations is increasingly valuable for career development, grant applicati

28 Apr 2026

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

Model ReleasesDGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

27 Apr 2026

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark…

AgentsDGX agent

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark it) The paper proposes a 'levels × laws' framework. Three c

Presenting DiaData for Research on Type 1 Diabetes

ResearchDGX agent

arXiv:2508.09160v2 Announce Type: replace Abstract: Type 1 diabetes (T1D) is an autoimmune disorder that leads to the destruction of insulin-producing cells, resulting in insulin deficiency, as to why

25 Apr 2026

Great thread from a talented artist on their creative journey using Hermes to unlock new creative coding avenues!

ResearchDGX agent

A Nous Research post on X documents an artist's creative journey and their experience using Hermes (likely Nous Research's language model) to explore new possibilities in creative coding and artistic

23 Apr 2026

A new AI model for probing fusion plasma behavior

ResearchDGX agent

IBM Research has developed an AI model designed to analyze and predict the behavior of fusion plasma, advancing computational capabilities for nuclear fusion research. The model likely leverages machi

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

Model ReleasesDGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

20 Apr 2026

Import AI 454: Automating alignment research; safety study of a Chinese model; HiFloat4

SafetyDGX agent

This newsletter covers three main topics: advances in automating alignment research to improve AI safety processes, a safety evaluation study of a Chinese AI model, and technical details about HiFloat

17 Apr 2026

China's smartphone shipments declined 4% YoY in Q1 amid memory shortages; Huawei's shipments grew 2% YoY for a 20% market share, iPhone grew 20% for a 19% share (Ivan Lam/Counterpoint Research)

IndustryDGX agent

Ivan Lam / Counterpoint Research: China's smartphone shipments declined 4% YoY in Q1 amid memory shortages; Huawei's shipments grew 2% YoY for a 20% market share, iPhone grew 20% for a 19% share — Iva

16 Apr 2026

Exposia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback

Model ReleasesDGX agent

arXiv:2601.06536v2 Announce Type: replace Abstract: We present Exposia, the first public dataset that connects writing and feedback in higher education, enabling research on educationally grounded com

https://x.com/NousResearch/status/2044584517218844775

AgentsDGX agent

Nous Research shared an update or announcement on their official X (Twitter) account, likely relating to their ongoing work in AI model development, fine-tuning, or research releases. Nous Research is

← Previous
1…56789…424
Next →