AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “openai”

GridTimelineEvolution
2,581 results
Tools

if your reaction to this is “haha openclaw bad, see prompt injection is the #1 danger” you: 1) havent sufficiently appreciated the layers to…

DGX agent

if your reaction to this is “haha openclaw bad, see prompt injection is the #1 danger” you: 1) havent sufficiently appreciated the layers to this tweet 2) havent seen enough ai api keys @gilpinskyy @d

toolsswyx--x
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t …

DGX agent

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t posted updates to X in a month, and remains very software fo

model-releasesethan-mollick--x
13 May 2026
Hardware

Richard Socher's Recursive Superintelligence raised 650M+ from GV, Greycroft, Nvidia, AMD, and others at a 4B valuation to pursue 'recursive self-improvement' (Cade Metz/New York Times)

DGX agent

Cade Metz / New York Times: Richard Socher's Recursive Superintelligence raised 650M+ from GV, Greycroft, Nvidia, AMD, and others at a 4B valuation to pursue “recursive self-improvement” — Recursive S

hardwaretechmeme
13 May 2026
Industry

Sources: Microsoft is in discussions to acquire LLM developer Inception; SpaceX also courted Inception, which is looking for a price of over $1B (Reuters)

DGX agent

Reuters: Sources: Microsoft is in discussions to acquire LLM developer Inception; SpaceX also courted Inception, which is looking for a price of over $1B — Microsoft (MSFT.O) is shopping for artificia

industrytechmeme
13 May 2026
Safety

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

DGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

safetyarxiv-cs-lg
13 May 2026
Agents

The most expensive mistake in enterprise AI right now: treating FDEs as your whole transformation plan. Forward deployed engineers (FDEs) ar…

DGX agent

The most expensive mistake in enterprise AI right now: treating FDEs as your whole transformation plan. Forward deployed engineers (FDEs) are important for custom deployments, but they won’t fix the c

agentsallie-k--miller--x
13 May 2026
Model Releases

Byte-Exact Deduplication in Retrieval-Augmented Generation: A Three-Regime Empirical Analysis Across Public Benchmarks

DGX agent

arXiv:2605.09611v1 Announce Type: new Abstract: This preprint presents an empirical analysis of byte-exact chunk-level deduplication in Retrieval-Augmented Generation (RAG) pipelines. We measure conte

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

CHAINTRIX: A multi-pipeline LLM-augmented framework for automated smart-contract security auditing

DGX agent

arXiv:2605.09350v1 Announce Type: new Abstract: Smart-contract exploits have caused billions of USD in cumulative losses, yet audits remain expensive and slow. Automated tools have emerged to close th

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction

DGX agent

arXiv:2605.09973v1 Announce Type: cross Abstract: Reliable detection of personally identifiable information (PII) is increasingly important across modern data-processing systems, yet the task remains

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models

DGX agent

arXiv:2510.08592v3 Announce Type: replace-cross Abstract: Test-Time Scaling (TTS) improves LLM reasoning by exploring multiple candidate responses and then operating over this set to find the best out

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier f…

DGX agent

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier for teams building AI agents on LangChain to run those worklo

model-releasesharrison-chase--x
12 May 2026
Safety

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor.

DGX agent

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor. hot take: unrestricted LLMs are as dangerous as weapons of mass destruction

safetygary-marcus--x
12 May 2026
Model Releases

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark

DGX agent

arXiv:2410.14702v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) exhibit impressive problem-solving abilities in various domains, but their visual comprehension and abstra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology

DGX agent

arXiv:2605.10761v1 Announce Type: new Abstract: Cancer screening is a reasoning task. A radiologist observes findings, compares them to prior scans, integrates clinical context, and reaches a diagnost

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using…

DGX agent

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using AI in assistant mode to study can hurt learning because it

model-releasesethan-mollick--x
12 May 2026
Model Releases

Try Grok Voice

DGX agent

Try Grok Voice Grok Voice Think Fast 1.0 ranks #1 on the Artificial Analysis τ-Voice benchmark for real-world agentic customer service resolution Absolutely outperforming GPT-Realtime-2 (High) and Gem

model-releaseselon-musk--x
12 May 2026
Model Releases

Vapi nabs $50M to make voice AI more human

DGX agent

Voice artificial intelligence startup Vapi Inc. said today it has raised 50 million in new funding to change the way people talk to computers, experience phone calls and interact with customer support

model-releasessiliconangle
12 May 2026
Safety

White Circle raises $11M to help companies secure and monitor AI model behavior

DGX agent

Artificial intelligence guardrail and monitoring startup Pumpkin Intelligence Inc., which operates as White Circle, announced today it raised 11 million in seed funding from a who’s who of AI leadersh

safetysiliconangle
12 May 2026
Model Releases

Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas

DGX agent

arXiv:2605.06673v1 Announce Type: cross Abstract: Aggregate metacognitive quality scores mask within-model variation across MMLU benchmark domains. We administered 1,500 MMLU items (250 per domain, un

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG

DGX agent

arXiv:2605.07273v1 Announce Type: cross Abstract: Multimodal RAG systems increasingly rely on vision-language retrievers to ground visual queries in external textual evidence. Existing adversarial stu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

DGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

model-releasesarxiv-cs-ai
11 May 2026
Safety

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure.

DGX agent

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure. NEW: ChatGPT advised the FSU shooter that a mass shooting w

safetygary-marcus--x
11 May 2026
Agents

okay so from what i can tell, this is probably more about limiting secondary activity to control secondary price, than punishing existing SP…

DGX agent

okay so from what i can tell, this is probably more about limiting secondary activity to control secondary price, than punishing existing SPVs secondaries can be a great price discovery mechanism for

agentsyohei-nakajima--x
11 May 2026
Industry

Alphabet’s Isomorphic Labs reportedly raising $2B+ for its medical research AI

DGX agent

Isomorphic Labs Inc., an Alphabet Inc. unit using artificial intelligence to speed up drug development, is reportedly in talks to raise a new funding round. Sources told Bloomberg today that the inves

industrysiliconangle
9 May 2026
Industry

kicking off a bunch of codex tasks, running around with my kid in the sunshine, and then coming back at naptime to find them all completed m…

DGX agent

Sam Altman posted about starting multiple Codex tasks, spending time outdoors with his child, and returning after naptime to find all the tasks had been completed, suggesting automated or background t

industrysam-altman--x
9 May 2026
Local Ai

I promise this will be the best 20 min you spend today! Robotics: Endgame, the sequel to my last year's Sequoia AI Ascent talk, 'Physical Tu…

DGX agent

I promise this will be the best 20 min you spend today! Robotics: Endgame, the sequel to my last year's Sequoia AI Ascent talk, 'Physical Turing Test'. I laid out the roadmap for solving Physical AGI

local-aijim-fan--x
8 May 2026
Safety

My view on Mythos was and is somewhere in between. It’s real, it’s a wakeup call, and it’s not quite what some of the media coverage suggest…

DGX agent

My view on Mythos was and is somewhere in between. It’s real, it’s a wakeup call, and it’s not quite what some of the media coverage suggested. • Mythos is not going to allow an 8 year old to accident

safetygary-marcus--x
8 May 2026
Safety

Next week at The Trial could be wild. Here’s why: Sam Altman will be on the stand. He is one of the world’s most convincing (though not alwa…

DGX agent

Next week at The Trial could be wild. Here’s why: Sam Altman will be on the stand. He is one of the world’s most convincing (though not always truthful) talkers (as I saw up close when we testified si

safetygary-marcus--x
8 May 2026
Safety

We’ve spent a lot of time on the framework underneath Codex, so it can move quickly on routine work while stopping for review when the risk …

DGX agent

We’ve spent a lot of time on the framework underneath Codex, so it can move quickly on routine work while stopping for review when the risk changes. Here’s how we use sandboxing, approvals, network po

safetysam-altman--x
8 May 2026
Applications

AI product team startup Pit raises $16M from a16z and others to automate enterprise workflows

DGX agent

Legendary venture capital firm Andreessen Horowitz is backing the artificial intelligence-native “product team-as-a-service” startup Pit in its first major funding round. It served as the lead investo

applicationssiliconangle
7 May 2026
Model Releases

MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents

DGX agent

arXiv:2605.03952v1 Announce Type: cross Abstract: Coding agents often pass per-prompt safety review yet ship exploitable code when their tasks are decomposed into routine engineering tickets. The chal

model-releasesarxiv-cs-ai
7 May 2026
Applications

So Mythos was, indeed, not marketing hype. Remember this is a general purpose model that just happens to be good at finding exploits because…

DGX agent

So Mythos was, indeed, not marketing hype. Remember this is a general purpose model that just happens to be good at finding exploits because good models are good at lots of things. Expect similar from

applicationsethan-mollick--x
7 May 2026
Model Releases

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting

DGX agent

arXiv:2605.04426v1 Announce Type: new Abstract: We introduce Telegraph English (TE), a prompt-compression protocol that rewrites natural language into a symbol-rich, formally-structured dialect. Where

model-releasesarxiv-cs-cl
7 May 2026
Safety

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until …

DGX agent

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until it's too late. But it's a very simple set up. Hyperscalers g

safetygary-marcus--x
6 May 2026
Safety

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or…

DGX agent

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or my own writings since fall 2023 They would have realized ho

safetygary-marcus--x
6 May 2026
Model Releases

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, st…

DGX agent

It feels like agent harness evolution runs on two axes that usually get conflated. There’s the temporal axis: simplify as models improve, stripping components that compensated for limitations the new

model-releasesharrison-chase--x
6 May 2026
Model Releases

Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use

DGX agent

arXiv:2605.02964v1 Announce Type: new Abstract: Reinforcement learning (RL) trained language model agents with tool access are increasingly deployed in coding assistants, research tools, and autonomou

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Trojan Hippo: Weaponizing Agent Memory for Data Exfiltration

DGX agent

arXiv:2605.01970v2 Announce Type: cross Abstract: Memory systems enable otherwise-stateless LLM agents to persist user information across sessions, but also introduce a new attack surface. We characte

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Benchmarking Retrieval Strategies for Biomedical Retrieval-Augmented Generation: A Controlled Empirical Study

DGX agent

arXiv:2605.02520v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) offers a well-established path to grounding large language model (LLM) outputs in external knowledge, yet the quest

model-releasesarxiv-cs-cl
5 May 2026
Safety

Elon: Let’s settle. Greg: Nope. Elon: Ok, let’s talk about your diaries, then.

DGX agent

Elon: Let’s settle. Greg: Nope. Elon: Ok, let’s talk about your diaries, then. On the eve of trial, Elon Musk reached out to Greg Brockman about a potential settlement, according to court documents fi

safetygary-marcus--x
5 May 2026
Safety

I predict this will be the most embarrassing cover in the history of this magazine. This is like putting Enron executives on your cover, or …

DGX agent

I predict this will be the most embarrassing cover in the history of this magazine. This is like putting Enron executives on your cover, or doing a gauzy special report on Bernie Madoff. Sam Altman is

safetygary-marcus--x
5 May 2026
Model Releases

Interpretable experiential learning based on state history and global feedback

DGX agent

arXiv:2605.00940v1 Announce Type: new Abstract: A new interpretable experiential learning model based on state history and global feedback is presented. It is capable of learning a behavioral model re

model-releasesarxiv-cs-lg
5 May 2026
Local Ai

Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

DGX agent

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

local-air-ollama
5 May 2026
Model Releases

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

DGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

model-releasesarxiv-cs-cl
5 May 2026
Safety

Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts

DGX agent

arXiv:2510.22628v2 Announce Type: replace-cross Abstract: This paper presents a real-time modular defense system named Sentra-Guard. The system detects and mitigates jailbreak and prompt injection att

safetyarxiv-cs-ai
5 May 2026
Safety

Training Non-Differentiable Networks via Optimal Transport

DGX agent

arXiv:2605.01928v1 Announce Type: new Abstract: Neural networks increasingly embed non-differentiable components (spiking neurons, quantized layers, discrete routing, blackbox simulators, etc.) where

safetyarxiv-cs-lg
5 May 2026
Safety

Ultrasound Vision-Language Alignment via Contrastive Learning

DGX agent

arXiv:2605.02126v1 Announce Type: new Abstract: Ultrasound foundation models have achieved strong performance on structured prediction tasks but remain exclusively vision-based, limiting zero-shot and

safetyarxiv-cs-cv
5 May 2026
Model Releases

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

DGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

model-releasesarxiv-cs-cl
5 May 2026
← Previous
1…4748495051…54
Next →