AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,425 results
23 Jun 2026

LastPass notifies customers that their personal information and customer support case records were stolen during a hack at Canadian market research company Klue (Zack Whittaker/TechCrunch)

IndustryDGX agent

Zack Whittaker / TechCrunch: LastPass notifies customers that their personal information and customer support case records were stolen during a hack at Canadian market research company Klue — Password

NEW RESEARCH ALERT! Led by my postdoc Dr. Anto Lonappan, we used the exquisite ACT DR6 CMB lensing data to search for evidence of cosmic str…

IndustryDGX agent

NEW RESEARCH ALERT! Led by my postdoc Dr. Anto Lonappan, we used the exquisite ACT DR6 CMB lensing data to search for evidence of cosmic strings: hypothetical cracks in spacetime that may have formed

22 Jun 2026

BIG thread on AI usage trends, as reported by Pew Research last week. Lots of juicy stats on AI use cases, data privacy, AI adoption by age …

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

BIG thread on AI usage trends, as reported by Pew Research last week. Lots of juicy stats on AI use cases, data privacy, AI adoption by age and gender, speed of adoption, AI sentiment, which tools are

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management syst…

AgentsDGX agent

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management system keeps them: > outside your repo > accesible to agent via

If you'd like to review the full research report, here is the direct link: https://www.pewresearch.org/internet/2026/06/17/americans-and-ai-…

IndustryDGX agent

This entry references a Pew Research Center report examining American public attitudes toward artificial intelligence, likely covering survey data on awareness, adoption, concerns, and demographic var

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive …

SafetyDGX agent

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive psychology, neuroscience, and artificial intelligence. What

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, …

SafetyDGX agent

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, JudgeBench, and RewardBench. Findings: Validating a judge wi

21 Jun 2026

6️⃣ Things to Know about AI Engineer World's Fair 2026 - It’s bigger than all previous AIEs - 4x Larger Expo with 4 Expo stages - Researcher…

ApplicationsDGX agent

6️⃣ Things to Know about AI Engineer World's Fair 2026 - It’s bigger than all previous AIEs - 4x Larger Expo with 4 Expo stages - Researchers: Poster sessions & Poaster sessions - AI Leadership: Token

We parsed this SpaceX equity research PDF faster than the time it took for Screen Studio to zoom in ⚡️🔥 liteparse is now the best open-sour…

Model ReleasesDGX agent

We parsed this SpaceX equity research PDF faster than the time it took for Screen Studio to zoom in ⚡️🔥 liteparse is now the best open-source document parsing tool out there. There’s no reason to not

20 Jun 2026

🚨🚨🚨A research project idea! How to measure world models? Everyone's talking about world models these days. World model here, world model …

AgentsDGX agent

🚨🚨🚨A research project idea! How to measure world models? Everyone's talking about world models these days. World model here, world model there. We can argue about what 'world model' actually means, an

Asmongold reacts to Pew Research's quiz revealing that 97% of r/politics users lean left with nearly half being the furthest left counted on…

IndustryDGX agent

Asmongold reacts to Pew Research's quiz revealing that 97% of r/politics users lean left with nearly half being the furthest left counted on the poll 'I feel like nobody's surprised..the reason why po

11 Jun 2026

Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude

Model ReleasesDGX agent

Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude Big scoop for Maxwell Zeff at Wired: “We’re changing Fable 5’s safeguards for frontier LLM development to make them

When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?

SafetyDGX agent

arXiv:2510.02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections

10 Jun 2026

Aesthetic Perspectives in Information Systems Research: A Hermeneutic Analysis

TutorialsDGX agent

arXiv:2606.09839v1 Announce Type: cross Abstract: How might implicit aesthetic perspectives shape what Information Systems (IS) scholarship recognises as worthy of study (or not)? In this hermeneutic

Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting 'innocuous tasks' like reading blog posts or performing code reviews (Lorenzo Franceschi-Bicchierai/TechCrunch)

Model ReleasesDGX agent

Lorenzo Franceschi-Bicchierai / TechCrunch: Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting “innocuous tasks” like reading blog posts or performing code rev

if only vetted institutions (big labs, governments, large enterprises) get unrestricted frontier capability, especially for AI research itse…

TutorialsDGX agent

if only vetted institutions (big labs, governments, large enterprises) get unrestricted frontier capability, especially for AI research itself, those players compound their lead while everyone else wo

In film, 'we'll fix it in post' is what you say when something went wrong on set and you don't want to redo it. AI research has made it our …

TutorialsDGX agent

In film, 'we'll fix it in post' is what you say when something went wrong on set and you don't want to redo it. AI research has made it our entire methodology: train the model, then patch whatever com

my weekend hobby: self improvement research

AgentsDGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

recruiting two singer-researchers to stage a dramatic adaptiation of the Muon/Shampoo debate set to the tune of “Your Obedient Servant” from…

ToolsDGX agent

Swyx posted about recruiting singer-researchers to create a dramatic musical adaptation of the Muon/Shampoo debate, using the melody from 'Your Obedient Servant' (likely referencing the Hamilton music

the results are modest (esp compared to parallel & similar research GRASP: https://arxiv.org/abs/2605.29668 - recommended!) the contribution…

AgentsDGX agent

the results are modest (esp compared to parallel & similar research GRASP: https://arxiv.org/abs/2605.29668 - recommended!) the contribution here is not the self improvement approach itself (yet), but

9 Jun 2026

Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL

HardwareDGX agent

NVIDIA FLARE Auto-FL automates federated learning research by constraining agent actions through a control plane, enforcing fixed benchmark contracts, and using an experiment ledger to ensure reproduc

Building at the speed of research: Lambda at CVPR 2026

AgentsDGX agent

Every year, CVPR draws the researchers defining what AI can see, understand, and act on. This year in Denver, more than 9,000 attendees showed up with over 4,000 accepted papers, and one shared proble

Degrading performance on ML research *without telling the user* is shockingly hostile and a terrible look. That could silently damage all so…

TutorialsDGX agent

Degrading performance on ML research *without telling the user* is shockingly hostile and a terrible look. That could silently damage all sorts of work, including some of my own. Also the type of thin

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence

SafetyDGX agent

arXiv:2606.07612v1 Announce Type: cross Abstract: We argue that many Anthropomorphic Misalignment Research (AMR) studies need stronger evidence to ensure that they can provide a robust foundation for

8 Jun 2026

Anthropic researchers say Mythos Preview can now turn publicly disclosed software vulnerabilities, or N-days, into working exploits in hours instead of weeks (Sam Sabin/Axios)

IndustryDGX agent

Sam Sabin / Axios: Anthropic researchers say Mythos Preview can now turn publicly disclosed software vulnerabilities, or N-days, into working exploits in hours instead of weeks — Anthropic's Mythos Pr

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to?

SafetyDGX agent

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to? Part 1 of my 'Pro-Human Manifesto' is now out! I'd love to know what you t

Narrative violation: according to @Stanford research, local models can answer 71.3% of real-world chat and reasoning queries accurately, up …

ApplicationsDGX agent

Narrative violation: according to @Stanford research, local models can answer 71.3% of real-world chat and reasoning queries accurately, up from 23.2% in 2023. Obviously at a fraction of the cost and

Sam Altman and Jakub Pachocki claim OpenAI is 'entering the third phase', aiming to automate AI research, boost the economy, and give everyone a personal AGI (OpenAI)

IndustryDGX agent

OpenAI: Sam Altman and Jakub Pachocki claim OpenAI is “entering the third phase”, aiming to automate AI research, boost the economy, and give everyone a personal AGI — Every few generations, a new tec

Sources: former OpenAI researcher Leopold Aschenbrenner's AI-focused hedge fund, Situational Awareness, has $20B+ AUM, after launching less than two years ago (Peter Rudegeair/Wall Street Journal)

ApplicationsDGX agent

Peter Rudegeair / Wall Street Journal: Sources: former OpenAI researcher Leopold Aschenbrenner's AI-focused hedge fund, Situational Awareness, has $20B+ AUM, after launching less than two years ago —

We published new research with Harvard on the shift from chat interfaces to autonomous agents like Computer. Over 3 months, findings show wo…

AgentsDGX agent

We published new research with Harvard on the shift from chat interfaces to autonomous agents like Computer. Over 3 months, findings show workers using Computer finish tasks in 87% less time at 94% lo

6 Jun 2026

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there…

TutorialsDGX agent

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there is very little progress in measuring it. So the big questio

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge a…

Local AiDGX agent

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge agents, this work argues that the selection layer is a first-

5 Jun 2026

A research team that includes Huawei says it successfully used Huawei's Ascend 910C chips for DeepSeek V4 Pro model's post-training, amid increased US sanctions (Coco Feng/South China Morning Post)

Model ReleasesDGX agent

Coco Feng / South China Morning Post: A research team that includes Huawei says it successfully used Huawei's Ascend 910C chips for DeepSeek V4 Pro model's post-training, amid increased US sanctions —

Cowork is at its best on work that’s too big for a chat: research across dozens of accounts, recurring reports, triaging my inbox and drafti…

ToolsDGX agent

Cowork is at its best on work that’s too big for a chat: research across dozens of accounts, recurring reports, triaging my inbox and drafting replies. If you’ve been curious, this is a good month to

University of Cambridge researchers say they have developed the first vaccine with a key component entirely designed by AI and subsequently trialed it in humans (James Gallagher/BBC)

IndustryDGX agent

James Gallagher / BBC: University of Cambridge researchers say they have developed the first vaccine with a key component entirely designed by AI and subsequently trialed it in humans — Artificial int

4 Jun 2026

Automatic Generation of Titles for Research Papers Using Language Models

Model ReleasesDGX agent

arXiv:2606.05085v1 Announce Type: cross Abstract: The title of a research paper conveys its primary idea and, occasionally, its conclusions in a clear and concise manner. Choosing an appropriate title

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

AgentsDGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time. Today, that work is r…

Model ReleasesDGX agent

We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time. Today, that work is rolling out as a more capable memory system in ChatGPT. https

What happened when one of our models found a counterexample to an 80-year-old Erdős conjecture? Researchers @alexwei_, @HongxunWu, and @wjmz…

Model ReleasesDGX agent

What happened when one of our models found a counterexample to an 80-year-old Erdős conjecture? Researchers @alexwei_, @HongxunWu, and @wjmzbmr1 shared the story on the OpenAI Podcast with @AndrewMayn

3 Jun 2026

AI market research platform AlphaSense raised 350M from Vitruvian, Accenture, and others at a 7.5B valuation, up from $4B in 2024, ahead of a possible IPO (Ben Glickman/Wall Street Journal)

IndustryDGX agent

Ben Glickman / Wall Street Journal: AI market research platform AlphaSense raised 350M from Vitruvian, Accenture, and others at a 7.5B valuation, up from 4B in 2024, ahead of a possible IPO — Firm rai

New research from Google. Just shows the impressive results you can get from custom agent harnesses. LEAP wraps a general-purpose LLM in an …

AgentsDGX agent

New research from Google. Just shows the impressive results you can get from custom agent harnesses. LEAP wraps a general-purpose LLM in an agentic scaffold that grounds every step in the Lean compile

Not to mention that having HIPAA and FERPA compliant AI systems makes thousands of students and researchers using them less risky.

ApplicationsDGX agent

HIPAA and FERPA compliant AI systems reduce institutional and legal risks when used by students and researchers by ensuring sensitive health and educational data are properly protected. Compliance wit

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-…

Model ReleasesDGX agent

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-5.5’s agentic coding and tool use together with stronger int

2 Jun 2026

Analysis: 22 of 24 US executive agencies saw a YoY increase in their average X account engagement during the first year of Trump's second term; @DOGE dominated (Pew Research Center)

IndustryDGX agent

Pew Research Center: Analysis: 22 of 24 US executive agencies saw a YoY increase in their average X account engagement during the first year of Trump's second term; @DOGE dominated — Federal agencies

Feds failing in bid to take a supercomputer from a climate research center

IndustryDGX agent

A federal judge blocked the Trump administration from taking away a supercomputer from the National Center for Atmospheric Research (NCAR) , preventing the National Science Foundation from transferrin

II-Agent is now live on the App Store. Your sovereign AI workspace for building, researching, writing, designing, and automating from one in…

AgentsDGX agent

II-Agent is now live on the App Store. Your sovereign AI workspace for building, researching, writing, designing, and automating from one intelligent interface. Download it. Bring your own key. Build

Microsoft unveils Microsoft Execution Containers for Windows, an OS-level sandbox for AI agents, with OpenAI, Nvidia, Manus, and Nous Research as partners (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Microsoft unveils Microsoft Execution Containers for Windows, an OS-level sandbox for AI agents, with OpenAI, Nvidia, Manus, and Nous Research as partners — For the past t

MIRROR: A Multi-Agent Framework with Iterative Adaptive Revision and Hierarchical Retrieval for Optimization Modeling in Operations Research

AgentsDGX agent

arXiv:2602.03318v3 Announce Type: replace Abstract: Operations Research (OR) relies on expert-driven modeling-a slow and fragile process ill-suited to novel scenarios. While large language models (LLM

Sources: Google DeepMind, Anthropic, and Meta have recently hired experts in psychology, ethics, and philosophy as they expand machine consciousness research (Cristina Criddle/Financial Times)

IndustryDGX agent

Cristina Criddle / Financial Times: Sources: Google DeepMind, Anthropic, and Meta have recently hired experts in psychology, ethics, and philosophy as they expand machine consciousness research — Goog

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1.…

Model ReleasesDGX agent

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1. MSA (MiniMax Sparse Attention) is the star ⭐️. Unlike CSA/HC

1 Jun 2026

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

Model ReleasesDGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

Researchers find packages in the @redhat-cloud-services npm namespace shipped malware that harvests credentials for GitHub Actions, AWS, GCP, Azure, and others (Rohan Prabhu/Step Security Blog)

IndustryDGX agent

Rohan Prabhu / Step Security Blog: Researchers find packages in the @redhat-cloud-services npm namespace shipped malware that harvests credentials for GitHub Actions, AWS, GCP, Azure, and others — Sev

WANDR is our in-house wide benchmark, built to mirror real professional research workloads. Search as Code scores 0.386 to the next best sys…

Model ReleasesDGX agent

WANDR is our in-house wide benchmark, built to mirror real professional research workloads. Search as Code scores 0.386 to the next best system's 0.152, and the benchmark is far from saturated. We're

30 May 2026

Learn more about the latest from @james_y_zou and our Frontier Agents Research team!

AgentsDGX agent

Learn more about the latest from @james_y_zou and our Frontier Agents Research team! To evaluate frontier AI agents, we need more complex tasks. But such tasks are also more prone to have design mista

29 May 2026

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve …

Model ReleasesDGX agent

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve the paths behind discovery, and expands what mathematicians

London-based Inherent, which aims to combine human scientific research with AI to produce innovations, emerges from stealth with $50M led by Index Ventures (Martin Coulter/Sifted)

IndustryDGX agent

Martin Coulter / Sifted: London-based Inherent, which aims to combine human scientific research with AI to produce innovations, emerges from stealth with $50M led by Index Ventures — London-based Inhe

SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing

Model ReleasesDGX agent

arXiv:2605.29468v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to support scientific work, but it is unclear whether they uphold responsible conduct of research (

28 May 2026

AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 (Jake Angelo/Fortune)

Model ReleasesDGX agent

Jake Angelo / Fortune: AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 — Imagine a world

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files fro…

Model ReleasesDGX agent

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files from years ago I had to use GPT-5.5 Pro as a reviewer, it spott

New in Claude Code (research preview): dynamic workflows. Claude writes an orchestration script on the fly, then spins up a large fleet of c…

Model ReleasesDGX agent

New in Claude Code (research preview): dynamic workflows. Claude writes an orchestration script on the fly, then spins up a large fleet of coordinated subagents in parallel to take on your most comple

← Previous
1…1011121314…424
Next →