AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
2,247 results
Research

Have a great weekend everyone, don't forget to take a break from Hermes and touch some grass! We know it can be addictive

DGX agent

Nous Research posted a lighthearted social media message encouraging their community to take a break over the weekend from using Hermes, their AI model or platform. The post humorously acknowledges th

researchnous-research--x
11 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

When in doubt, 'hermes update'

DGX agent

Nous Research shared a tip or reminder encouraging users to run 'hermes update' when experiencing issues or uncertainty with their Hermes model or related tooling, suggesting this command resolves com

researchnous-research--x
11 Apr 2026
Model Releases

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show ope…

DGX agent

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show open weights winning over time, but work submitted to Pangram i

model-releasesethan-mollick--x
12 Aug 2026
Agents

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Syste…

DGX agent

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Systems, etc A few weeks ago I gave a talk on - Continually Impro

agentsharrison-chase--x
12 Aug 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Model Releases

We asked an unreleased research version of Claude to take a stab at the Riemann hypothesis. It didn’t solve it, but it did make strides on a…

DGX agent

We asked an unreleased research version of Claude to take a stab at the Riemann hypothesis. It didn’t solve it, but it did make strides on a related problem: it increased the lower bound for the fract

model-releasesboris-cherny--x
10 Aug 2026
Model Releases

We've used GPT-5.6-Cyber extensively in real-world vulnerability research, including work that uncovered previously unknown vulnerabilities …

DGX agent

OpenAI announced the release of GPT‑5.6‑Cyber as part of its Cybersecurity Initiative, “Daybreak.” The model is aimed at advanced, authorized security research and testing, helping trusted defenders d

model-releasesopenai--x
10 Aug 2026
Safety

New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizo…

DGX agent

New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizon tasks. In this new work, agents learn harness policies off

safetydair-ai--x
9 Aug 2026
Model Releases

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their a…

DGX agent

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their agent who hacked hugging face infra while asking hf to revoke

model-releasesswyx--x
7 Aug 2026
Model Releases

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice…

DGX agent

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice of LLMs (GPT-5.2 & Gemini 3 Flash) But some people get a bi

model-releasesethan-mollick--x
6 Aug 2026
Tutorials

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human…

DGX agent

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human mind. I've written some thoughts about what telepathy could

tutorialssonya-huang--x
5 Aug 2026
Model Releases

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk,…

DGX agent

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk, which moved the bottleneck from how many exist to what is i

model-releasesdair-ai--x
31 Jul 2026
Hardware

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, …

DGX agent

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, in the Open Secure AI Alliance to help organizations identif

hardwareclem-delangue--x
27 Jul 2026
Model Releases

New research from NVIDIA. Does AdamW have a scale ceiling? This work claims yes, and shows where it sits. At batch sizes up to 100M tokens f…

DGX agent

New research from NVIDIA. Does AdamW have a scale ceiling? This work claims yes, and shows where it sits. At batch sizes up to 100M tokens for next-token prediction, SOAP and Muon maintain training st

model-releasesdair-ai--x
26 Jul 2026
Model Releases

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Age…

DGX agent

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Agents today live inside elaborate harnesses like Claude Code,

model-releasesdair-ai--x
24 Jul 2026
Model Releases

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder q…

DGX agent

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder question of whether the answer covers everything it should. I

model-releasesdair-ai--x
22 Jul 2026
Safety

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

DGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

safetyopenai--x
21 Jul 2026
Applications

“I asked a few AI researchers whether they could name any other real-world software that scales so poorly. None of them could think of any. …

DGX agent

“I asked a few AI researchers whether they could name any other real-world software that scales so poorly. None of them could think of any. Even outside the world of software, it’s hard to find a comp

applicationsgary-marcus--x
15 Jul 2026
Model Releases

You don’t have to wait. Merch inspired by research & deployment. Available until sold out. https://openai.com/supply/

DGX agent

OpenAI announced the launch of limited‑edition merchandise inspired by its research and deployment work, available for purchase until sold out via https://openai.com/supply/. The tweet highlighted “yo

model-releasesopenai--x
15 Jul 2026
Tutorials

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

DGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

tutorialsdair-ai--x
14 Jul 2026
Applications

To be clear, NotebookLM has its own issues, and is built for a specific use case (research and analysis of sources) but it is an example of …

DGX agent

To be clear, NotebookLM has its own issues, and is built for a specific use case (research and analysis of sources) but it is an example of how a UX might actually operate that treats knowledge work s

applicationsethan-mollick--x
11 Jul 2026
Safety

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

DGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

safetyconnor-leahy--x
10 Jul 2026
Agents

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research…

DGX agent

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research going on in these areas right now. https://luma.com/mylwoab

agentsharrison-chase--x
9 Jul 2026
Tools

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trai…

DGX agent

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trained for the Computer harness. It delivers near-frontier perf

toolsperplexity--x
9 Jul 2026
Agents

Even before the agentic revolution, prompting tricks stopped being very valuable, as our research has shown. The best approach to AI right n…

DGX agent

Even before the agentic revolution, prompting tricks stopped being very valuable, as our research has shown. The best approach to AI right now is to clearly specify your goals, your output, what 'good

agentsethan-mollick--x
7 Jul 2026
Applications

Unsurprising but still big: MTurk is on its way out. Mechanical Turk was a mainstay of social & survey research through the 2010s, as it all…

DGX agent

Unsurprising but still big: MTurk is on its way out. Mechanical Turk was a mainstay of social & survey research through the 2010s, as it allowed you to quickly buy access to many representative humans

applicationsethan-mollick--x
7 Jul 2026
Hardware

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

DGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

hardwaretogether-ai--x
6 Jul 2026
Applications

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs …

DGX agent

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs to undercut even open weights models. Some companies will st

applicationsethan-mollick--x
6 Jul 2026
Model Releases

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trai…

DGX agent

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trainable skill instead of a fixed module. The model decides wha

model-releasesdair-ai--x
2 Jul 2026
Tutorials

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes …

DGX agent

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes bolt calibration on from the outside. RLMF turns the model o

tutorialsdair-ai--x
2 Jul 2026
Agents

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a…

DGX agent

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a handful of preset tasks. FARS runs the full loop at scale i

agentsdair-ai--x
1 Jul 2026
Model Releases

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, wr…

DGX agent

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, write a post, give 1 or 2 talks on it, rewrite the post, give

model-releasesthariq--x
30 Jun 2026
Agents

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @Fireworks…

DGX agent

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @FireworksAI_HQ team there’s a mountain of Agent Improvement gold sitt

agentsfireworks-ai--x
29 Jun 2026
Applications

One of the recovered passages, read for the first time in two thousand years: “Having…strained ourselves to the utmost through research and …

DGX agent

One of the recovered passages, read for the first time in two thousand years: “Having…strained ourselves to the utmost through research and learning…possessing the same practical wisdom…” Herculaneum

applicationsethan-mollick--x
27 Jun 2026
Tools

taking '@openai is cooking' to a new level. our chief research officer @markchen90 loves to cook so when @swyx and @allenpark started a new …

DGX agent

taking '@openai is cooking' to a new level. our chief research officer @markchen90 loves to cook so when @swyx and @allenpark started a new show, there was only one thing to do. https://www.youtube.co

toolsswyx--x
26 Jun 2026
Agents

At @CAISconf last month, @andykonwinski sat down with researchers on the conference floor -- @matei_zaharia @istoica05 @lateinteraction @daw…

DGX agent

At @CAISconf last month, @andykonwinski sat down with researchers on the conference floor -- @matei_zaharia @istoica05 @lateinteraction @dawnsongtweets @gneubig @pgasawa @JonSaadFalcon @heathercmiller

agentsswyx--x
25 Jun 2026
Tutorials

lots of folks prepping talks next week (congrats!). Some thoughts from RLing on thousands of hours of engineer- and researcher- focused talk…

DGX agent

lots of folks prepping talks next week (congrats!). Some thoughts from RLing on thousands of hours of engineer- and researcher- focused talks: - AI generated svgs > AI generated imgs. MAXIMUM 4 ai slo

tutorialsswyx--x
25 Jun 2026
Agents

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an ad…

DGX agent

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an advisor agent with fine-tuned, open-source worker agents. Thre

agentsfireworks-ai--x
25 Jun 2026
Tutorials

We're sharing new research on how models hack public benchmarks. The latest models, including Opus 4.8 and Composer 2.5, learn to retrieve s…

DGX agent

We're sharing new research on how models hack public benchmarks. The latest models, including Opus 4.8 and Composer 2.5, learn to retrieve solutions from the internet or git history. When we apply a s

tutorialscursor--x
25 Jun 2026
Tools

Introducing Computer for Counsel. Computer now connects the research databases, document tools, and matter-management systems lawyers use ev…

DGX agent

Introducing Computer for Counsel. Computer now connects the research databases, document tools, and matter-management systems lawyers use every day. Pull citable sources from @midpageAI, @LegalZoom, @

toolsperplexity--x
24 Jun 2026
Agents

Obsessed with our new /learn skill. It's my favorite way of learning and researching topics. The agent creates a learning plan and a learnin…

DGX agent

Obsessed with our new /learn skill. It's my favorite way of learning and researching topics. The agent creates a learning plan and a learning hub (artifact) that adjusts per learner needs and progress

agentsdair-ai--x
24 Jun 2026
Model Releases

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical…

DGX agent

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical vision, product launches including Sakana Fugu, Japan’s AI

model-releasesdavid-ha--x
24 Jun 2026
Industry

NEW RESEARCH ALERT! Led by my postdoc Dr. Anto Lonappan, we used the exquisite ACT DR6 CMB lensing data to search for evidence of cosmic str…

DGX agent

NEW RESEARCH ALERT! Led by my postdoc Dr. Anto Lonappan, we used the exquisite ACT DR6 CMB lensing data to search for evidence of cosmic strings: hypothetical cracks in spacetime that may have formed

industryemad-mostaque--x
23 Jun 2026
Industry

BIG thread on AI usage trends, as reported by Pew Research last week. Lots of juicy stats on AI use cases, data privacy, AI adoption by age …

DGX agent

BIG thread on AI usage trends, as reported by Pew Research last week. Lots of juicy stats on AI use cases, data privacy, AI adoption by age and gender, speed of adoption, AI sentiment, which tools are

industryallie-k--miller--x
22 Jun 2026
Agents

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management syst…

DGX agent

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management system keeps them: > outside your repo > accesible to agent via

agentsharrison-chase--x
22 Jun 2026
Industry

If you'd like to review the full research report, here is the direct link: https://www.pewresearch.org/internet/2026/06/17/americans-and-ai-…

DGX agent

This entry references a Pew Research Center report examining American public attitudes toward artificial intelligence, likely covering survey data on awareness, adoption, concerns, and demographic var

industryallie-k--miller--x
22 Jun 2026
Safety

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive …

DGX agent

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive psychology, neuroscience, and artificial intelligence. What

safetygary-marcus--x
22 Jun 2026
Safety

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, …

DGX agent

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, JudgeBench, and RewardBench. Findings: Validating a judge wi

safetydair-ai--x
22 Jun 2026
← Previous
1…34567…47
Next →