AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,425 results
6 Aug 2026

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice…

Model ReleasesDGX agent

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice of LLMs (GPT-5.2 & Gemini 3 Flash) But some people get a bi

5 Aug 2026

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

AgentsDGX agent

arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
TutorialsDGX agent

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human mind. I've written some thoughts about what telepathy could

4 Aug 2026

Company approved 128GB Mac for research proposal, best model?

Model ReleasesDGX agent

I‘m doing a research proposal at my company about running local LLMs to replace daily coding models. Qwen 3.6 27B (or 3.8 potentially) is widely seen as the best model in that 20-60GB space, is that s

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

HardwareDGX agent

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanc

Uncertainty-Aware Simulation-Based Inference for Operations Research with Large Language Models

Local AiDGX agent

arXiv:2608.00019v1 Announce Type: new Abstract: Deploying large language models (LLMs) for operations research (OR) tasks remains challenging because correctness depends on a coherent modeling process

3 Aug 2026

YouTuber Hank Green faces online criticism after using ChatGPT to help research a script, and says his LLM usage 'is not healthy for me or good for the world' (Anthony Ha/TechCrunch)

IndustryDGX agent

Anthony Ha / TechCrunch: YouTuber Hank Green faces online criticism after using ChatGPT to help research a script, and says his LLM usage “is not healthy for me or good for the world” — Hank Green, a

31 Jul 2026

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk,…

Model ReleasesDGX agent

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk, which moved the bottleneck from how many exist to what is i

28 Jul 2026

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

Model ReleasesDGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

27 Jul 2026

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, …

HardwareDGX agent

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, in the Open Secure AI Alliance to help organizations identif

26 Jul 2026

New research from NVIDIA. Does AdamW have a scale ceiling? This work claims yes, and shows where it sits. At batch sizes up to 100M tokens f…

Model ReleasesDGX agent

New research from NVIDIA. Does AdamW have a scale ceiling? This work claims yes, and shows where it sits. At batch sizes up to 100M tokens for next-token prediction, SOAP and Muon maintain training st

24 Jul 2026

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Age…

Model ReleasesDGX agent

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Agents today live inside elaborate harnesses like Claude Code,

22 Jul 2026

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder q…

Model ReleasesDGX agent

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder question of whether the answer covers everything it should. I

21 Jul 2026

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

SafetyDGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

16 Jul 2026

Analogical Deep Research: Retrieving and Integrating Historical Analogies for Foresight Analysis

Model ReleasesDGX agent

arXiv:2607.13602v1 Announce Type: cross Abstract: Systematic comparisons between current situations and structurally similar past events in the historical, i.e., historical analogies, is among the mos

15 Jul 2026

“I asked a few AI researchers whether they could name any other real-world software that scales so poorly. None of them could think of any. …

ApplicationsDGX agent

“I asked a few AI researchers whether they could name any other real-world software that scales so poorly. None of them could think of any. Even outside the world of software, it’s hard to find a comp

You don’t have to wait. Merch inspired by research & deployment. Available until sold out. https://openai.com/supply/

Model ReleasesDGX agent

OpenAI announced the launch of limited‑edition merchandise inspired by its research and deployment work, available for purchase until sold out via https://openai.com/supply/. The tweet highlighted “yo

14 Jul 2026

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

TutorialsDGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

11 Jul 2026

To be clear, NotebookLM has its own issues, and is built for a specific use case (research and analysis of sources) but it is an example of …

ApplicationsDGX agent

To be clear, NotebookLM has its own issues, and is built for a specific use case (research and analysis of sources) but it is an example of how a UX might actually operate that treats knowledge work s

10 Jul 2026

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

SafetyDGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

SafetyDGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

Model ReleasesDGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

9 Jul 2026

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops

SafetyDGX agent

arXiv:2607.07663v1 Announce Type: new Abstract: AI systems increasingly participate in their own improvement: revising their outputs, adapting their own harnesses during deployment, training on data t

Synthetic Data Generation for Financial AI Research with NVIDIA NeMo

HardwareDGX agent

NVIDIA's NeMo framework addresses the challenge of limited, imbalanced financial NLP data by generating synthetic financial news headlines to fill gaps for trading research, risk modeling, and surveil

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research…

AgentsDGX agent

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research going on in these areas right now. https://luma.com/mylwoab

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trai…

ToolsDGX agent

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trained for the Computer harness. It delivers near-frontier perf

8 Jul 2026

Powering scientific discovery: BYOKG and GraphRAG for intelligent pharmaceutical research

IndustryDGX agent

In this post, we explore how Graph-based Retrieval Augmented Generation (GraphRAG) is transforming scientific research by combining graph databases with generative AI. With this approach, you can acce

7 Jul 2026

Even before the agentic revolution, prompting tricks stopped being very valuable, as our research has shown. The best approach to AI right n…

AgentsDGX agent

Even before the agentic revolution, prompting tricks stopped being very valuable, as our research has shown. The best approach to AI right now is to clearly specify your goals, your output, what 'good

Paris-based UMA, founded by ex-Tesla Optimus scientist Rémi Cadene and ex-Google DeepMind researcher Pierre Sermanet, demos its Northstar AI humanoid robot (Benoit Berthelot/Bloomberg)

IndustryDGX agent

Benoit Berthelot / Bloomberg: Paris-based UMA, founded by ex-Tesla Optimus scientist Rémi Cadene and ex-Google DeepMind researcher Pierre Sermanet, demos its Northstar AI humanoid robot — A former Tes

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog

Model ReleasesDGX agent

arXiv:2607.04438v1 Announce Type: cross Abstract: Research dissemination, turning a paper into a poster, a talk video, and a blog post, is still a manual last mile. Prior automation treats each artifa

Unsurprising but still big: MTurk is on its way out. Mechanical Turk was a mainstay of social & survey research through the 2010s, as it all…

ApplicationsDGX agent

Unsurprising but still big: MTurk is on its way out. Mechanical Turk was a mainstay of social & survey research through the 2010s, as it allowed you to quickly buy access to many representative humans

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

6 Jul 2026

How Open Models Are Driving AI Research

HardwareDGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

HardwareDGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

Some US voters are using AI tools as nonpartisan researchers, seeing them as a viable alternative to traditional news coverage, voter guides, and social media (Jennifer Medina/New York Times)

IndustryDGX agent

Jennifer Medina / New York Times: Some US voters are using AI tools as nonpartisan researchers, seeing them as a viable alternative to traditional news coverage, voter guides, and social media — It ta

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs …

ApplicationsDGX agent

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs to undercut even open weights models. Some companies will st

2 Jul 2026

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trai…

Model ReleasesDGX agent

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trainable skill instead of a fixed module. The model decides wha

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes …

TutorialsDGX agent

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes bolt calibration on from the outside. RLMF turns the model o

1 Jul 2026

A researcher says a vulnerability in Apple's Hide My Email tool lets anyone see a user's real email address; first reported in June 2025, Apple has not fixed it (Joseph Cox/404 Media)

IndustryDGX agent

Joseph Cox / 404 Media: A researcher says a vulnerability in Apple's Hide My Email tool lets anyone see a user's real email address; first reported in June 2025, Apple has not fixed it — “Hide My Emai

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a…

AgentsDGX agent

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a handful of preset tasks. FARS runs the full loop at scale i

🔬 The Coolest Diffusion Research Isn't in LLMs — Evan Feinberg & Sergey Edunov, Genesis Molecular AI

Model ReleasesDGX agent

This episode discusses cutting-edge diffusion model research applications beyond large language models, featuring insights from Evan Feinberg and Sergey Edunov of Genesis Molecular AI on how diffusion

30 Jun 2026

DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification

AgentsDGX agent

arXiv:2606.29746v1 Announce Type: new Abstract: Navigating the deluge of heterogeneous medical data, from academic literature (PubMed) to clinical guidelines (Web) and private knowledge bases, remains

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, wr…

Model ReleasesDGX agent

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, write a post, give 1 or 2 talks on it, rewrite the post, give

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

Model ReleasesDGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

29 Jun 2026

COOPA: A Modular LLM Agent Architecture for Operations Research Problems

AgentsDGX agent

arXiv:2606.27611v1 Announce Type: new Abstract: Operations Research (OR) provides a rigorous framework for high-stakes decision-making, but effective OR modeling requires substantial domain knowledge,

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @Fireworks…

AgentsDGX agent

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @FireworksAI_HQ team there’s a mountain of Agent Improvement gold sitt

28 Jun 2026

Researchers say Z.ai's GLM-5.2 matches latest US models at finding security bugs, as critics question the US' lax approach in restricting Chinese open models (Wall Street Journal)

IndustryDGX agent

Wall Street Journal: Researchers say Z.ai's GLM-5.2 matches latest US models at finding security bugs, as critics question the US' lax approach in restricting Chinese open models — Clampdown on top U.

27 Jun 2026

One of the recovered passages, read for the first time in two thousand years: “Having…strained ourselves to the utmost through research and …

ApplicationsDGX agent

One of the recovered passages, read for the first time in two thousand years: “Having…strained ourselves to the utmost through research and learning…possessing the same practical wisdom…” Herculaneum

26 Jun 2026

[AINews] OpenAI reports median internal Codex output tokens grew 56x in Research, 32x in Customer Support, 27x in Engineering, and 13x in Legal since November 2025.

ApplicationsDGX agent

OpenAI reported significant increases in internal Codex token usage across different departments between November 2025 and the reporting period, with Research experiencing the largest growth at 56x, f

ReportLogic: Evaluating Logical Quality in Deep Research Reports

Model ReleasesDGX agent

arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to synthesize diverse sources into structured reports th

taking '@openai is cooking' to a new level. our chief research officer @markchen90 loves to cook so when @swyx and @allenpark started a new …

ToolsDGX agent

taking '@openai is cooking' to a new level. our chief research officer @markchen90 loves to cook so when @swyx and @allenpark started a new show, there was only one thing to do. https://www.youtube.co

25 Jun 2026

At @CAISconf last month, @andykonwinski sat down with researchers on the conference floor -- @matei_zaharia @istoica05 @lateinteraction @daw…

AgentsDGX agent

At @CAISconf last month, @andykonwinski sat down with researchers on the conference floor -- @matei_zaharia @istoica05 @lateinteraction @dawnsongtweets @gneubig @pgasawa @JonSaadFalcon @heathercmiller

Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

Model ReleasesDGX agent

arXiv:2606.24902v1 Announce Type: cross Abstract: The 'First Proof' benchmark [1] posed ten research-level mathematics questions to the strongest publicly available LLMs and found them consistently wr

lots of folks prepping talks next week (congrats!). Some thoughts from RLing on thousands of hours of engineer- and researcher- focused talk…

TutorialsDGX agent

lots of folks prepping talks next week (congrats!). Some thoughts from RLing on thousands of hours of engineer- and researcher- focused talks: - AI generated svgs > AI generated imgs. MAXIMUM 4 ai slo

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an ad…

AgentsDGX agent

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an advisor agent with fine-tuned, open-source worker agents. Thre

We're sharing new research on how models hack public benchmarks. The latest models, including Opus 4.8 and Composer 2.5, learn to retrieve s…

TutorialsDGX agent

We're sharing new research on how models hack public benchmarks. The latest models, including Opus 4.8 and Composer 2.5, learn to retrieve solutions from the internet or git history. When we apply a s

24 Jun 2026

Introducing Computer for Counsel. Computer now connects the research databases, document tools, and matter-management systems lawyers use ev…

ToolsDGX agent

Introducing Computer for Counsel. Computer now connects the research databases, document tools, and matter-management systems lawyers use every day. Pull citable sources from @midpageAI, @LegalZoom, @

Obsessed with our new /learn skill. It's my favorite way of learning and researching topics. The agent creates a learning plan and a learnin…

AgentsDGX agent

Obsessed with our new /learn skill. It's my favorite way of learning and researching topics. The agent creates a learning plan and a learning hub (artifact) that adjusts per learner needs and progress

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical…

Model ReleasesDGX agent

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical vision, product launches including Sakana Fugu, Japan’s AI

23 Jun 2026

Build a protein research copilot with Amazon Bedrock AgentCore

TutorialsDGX agent

This post shows you how to build a conversational protein research assistant that combines three capabilities: Natural language query parsing to extract structured search parameters, vector similarity

← Previous
1…910111213…424
Next →