AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “openai”

GridTimelineEvolution
2,581 results
Agents

Entropy Gate: Entropy Quenching for Near-Lossless Token Compression in LLM Pipelines

DGX agent

arXiv:2606.03739v1 Announce Type: new Abstract: LLM pipelines waste substantial token budgets on low-information content: repeated context, verbose responses, and redundant boilerplate. We introduce E

agentsarxiv-cs-cl
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful s…

DGX agent

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful statement. In hindsight, though, I realized he was lying (und

safetygary-marcus--x
3 Jun 2026
Local Ai

Ideogram 4.0 Just Open Sourced!

DGX agent

Ideogram released version 4.0 of its text-to-image model as an open-weight model with native 2K resolution, bounding box control, and improved text rendering. The model weights are available on GitHub

local-air-stablediffusion
3 Jun 2026
Tools

⚡️Satya Nadella: No Priors x Latent Space Crossover Special at Microsoft Build

DGX agent

This episode features Satya Nadella, CEO of Microsoft, in a crossover special between the 'No Priors' and 'Latent Space' podcasts, likely discussing Microsoft's AI strategy, product developments, and

toolslatent-space
3 Jun 2026
Model Releases

The Impact of Configuring Agentic AI Coding Tools on Build-vs-Buy Decisions: A Study Protocol

DGX agent

arXiv:2606.03907v1 Announce Type: cross Abstract: Agentic AI coding tools write code with increasing autonomy and in doing so decide when to import a library and when to implement functionality from s

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Uber Caps Usage of AI Tools Like Claude Code to Manage Costs

DGX agent

Uber Caps Usage of AI Tools Like Claude Code to Manage Costs I wrote the other day about Uber blowing its 2026 AI budget in four months, and how that wasn't particularly surprising given they would ha

model-releasessimon-willison
3 Jun 2026
Model Releases

Using open models and inference clouds (which serve open models) is a leading indicator of what is to come. The advantage of open weights is…

DGX agent

Using open models and inference clouds (which serve open models) is a leading indicator of what is to come. The advantage of open weights is that you can train, serve, and continually improve your own

model-releasesclem-delangue--x
3 Jun 2026
Model Releases

AgentRedBench: Dynamic Redteaming and Integration-Aware Defense for LLM Agents over SaaS Integrations

DGX agent

arXiv:2606.02240v1 Announce Type: cross Abstract: Indirect prompt injection in tool-use agents is a concrete production threat: LLM agents read from integrations (third-party services such as Gmail, S

model-releasesarxiv-cs-ai
2 Jun 2026
Industry

An interview with Microsoft AI CEO Mustafa Suleyman about its models catching up to the state of the art from months ago, refusing to distill models, and more (Reed Albergotti/Semafor)

DGX agent

Reed Albergotti / Semafor: An interview with Microsoft AI CEO Mustafa Suleyman about its models catching up to the state of the art from months ago, refusing to distill models, and more — THE SCENE —

industrytechmeme
2 Jun 2026
Research

Characterizing Web Search in The Age of Generative AI

DGX agent

arXiv:2510.11560v2 Announce Type: replace-cross Abstract: The advent of LLMs has given rise to generative search, a new search paradigm in which LLMs retrieve information from the web related to a que

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies

DGX agent

arXiv:2606.01151v1 Announce Type: new Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distri

safetyarxiv-cs-lg
2 Jun 2026
Industry

Microsoft’s first advanced reasoning AI is here

DGX agent

Microsoft announced a bunch of new in-house AI models at Build 2026, including a new 'flagship' model: MAI-Thinking-1. It's an ambitious step into model development for Microsoft, which introduced its

industrythe-verge-ai
2 Jun 2026
Model Releases

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

DGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Remind me, @Elonmusk, was GPT-5 really smarter than the smartest humans?

DGX agent

Remind me, @Elonmusk, was GPT-5 really smarter than the smartest humans? @viktaur27 @Teslarati The rate of improvement from original GPT to GPT-3 is impressive. If this rate of improvement continues,

model-releasesgary-marcus--x
2 Jun 2026
Model Releases

Sensor Tower: ChatGPT has become the fastest app to hit 1B global MAUs by far; ChatGPT's MAUs are up 62% YoY in Q2 to date, Claude's MAUs are up 640% YoY to 56M (Harshita Mary Varghese/Reuters)

DGX agent

Harshita Mary Varghese / Reuters: Sensor Tower: ChatGPT has become the fastest app to hit 1B global MAUs by far; ChatGPT's MAUs are up 62% YoY in Q2 to date, Claude's MAUs are up 640% YoY to 56M — Ope

model-releasestechmeme
2 Jun 2026
Research

Evaluating using Mock Tool Calls to Quarantine Untrusted Prompt Inputs

DGX agent

arXiv:2605.30521v1 Announce Type: new Abstract: Large language models must frequently process untrusted inputs, such as judging an answer from another model or running tasks like spam and harm classif

researcharxiv-cs-cl
1 Jun 2026
Model Releases

Fingerprint launches AI Assistant Detection to spot traffic from ChatGPT, Gemini and Claude

DGX agent

Device intelligence company FingerprintJS Inc. today launched a preview of two products built to identify traffic from artificial intelligence assistants, addressing a detection gap that has opened as

model-releasessiliconangle
1 Jun 2026
Research

Incremental BPE Tokenization

DGX agent

arXiv:2605.30813v1 Announce Type: new Abstract: We propose a novel algorithm for incremental Byte Pair Encoding (BPE) tokenization. The algorithm processes each input byte in worst-case O(log^2 t) tim

researcharxiv-cs-cl
1 Jun 2026
Model Releases

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organiza…

DGX agent

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organization to help customize solutions, such as building and tunin

model-releasesandrew-ng--x
1 Jun 2026
Model Releases

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, yo…

DGX agent

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, you should share yours too! https://huggingface.co/datasets?se

model-releasesclem-delangue--x
31 May 2026
Safety

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️

DGX agent

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️ I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthro

safetygary-marcus--x
31 May 2026
Model Releases

BrahmicTokenizer-131K: An Indic-Capable Drop-In Replacement for o200k_base

DGX agent

arXiv:2605.29379v1 Announce Type: new Abstract: We present BrahmicTokenizer-131K, a 131,072-vocabulary byte-level BPE tokenizer that closes the Brahmic compression gap at the 131K-vocabulary class whi

model-releasesarxiv-cs-cl
29 May 2026
Agents

Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents

DGX agent

arXiv:2605.29927v1 Announce Type: cross Abstract: Despite recent advances, LLM-based web agents still struggle with limited exploration, omission of critical steps, and sensitivity to task constraints

agentsarxiv-cs-ai
29 May 2026
Model Releases

Frontier LLM-based agents can overcome the ontology curation bottleneck for natural phenotypes

DGX agent

arXiv:2605.28965v1 Announce Type: new Abstract: Linking free-text phenotype descriptions to ontology terms, typically referred to as phenotype annotation, is essential for the cross-study integration

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis

DGX agent

arXiv:2605.28876v1 Announce Type: cross Abstract: CI failure logs are large (median 5k lines, max 200k in this corpus) and noisy. Coding agents that try to debug them depend on an upstream tool to red

model-releasesarxiv-cs-ai
29 May 2026
Safety

Today I learned that @Grimezsz has more courage in her pinky than Roon does in his entire cowardly body. Roon made excuses; Grimes defended …

DGX agent

Today I learned that @Grimezsz has more courage in her pinky than Roon does in his entire cowardly body. Roon made excuses; Grimes defended her views calmly and respectfully, like grownups should. Ope

safetygary-marcus--x
29 May 2026
Safety

crazy that this was announced on the day tokenmaxxing died.

DGX agent

crazy that this was announced on the day tokenmaxxing died. Anthropic raised 65 billion in a funding round that valued the artificial intelligence company at 965 billion including the new investment,

safetygary-marcus--x
28 May 2026
Safety

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems

DGX agent

arXiv:2605.27766v1 Announce Type: new Abstract: LLM safety evaluations predominantly test models in isolation, yet deployed AI agents increasingly operate within persistent social environments alongsi

safetyarxiv-cs-ai
28 May 2026
Tutorials

Inversely Learning Transferable Rewards via Abstracted States

DGX agent

arXiv:2501.01669v4 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) has progressed significantly toward accurately learning the underlying rewards in both discrete and continuous

tutorialsarxiv-cs-lg
28 May 2026
Model Releases

Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Reproducibility Below the Rerun-Stability Baseline

DGX agent

arXiv:2605.27440v1 Announce Type: cross Abstract: Small changes to how a buyer phrases a question -- 'best CRM' vs 'top CRM' vs 'best CRM for a SaaS startup' -- produce substantially different brand r

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Verifiable Benchmarking of Long-Horizon Spatial Biology

DGX agent

arXiv:2605.28065v1 Announce Type: new Abstract: AI agents are increasingly useful for biological data analysis, but existing benchmarks mostly test broad biological knowledge, executable workflows, or

model-releasesarxiv-cs-ai
28 May 2026
Hardware

AI training data provider Human Archive raises $8.2M

DGX agent

Artificial intelligence training data provider Human Archive Inc. today announced that it has raised 8.2 million in funding. Wing Venture Capital, NVP Capital, Y Combinator headlined the consortium th

hardwaresiliconangle
27 May 2026
Model Releases

E3: Issue-Level Backtesting for Automated Research Critique

DGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

model-releasesarxiv-cs-ai
27 May 2026
Agents

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from pro…

DGX agent

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from production traces to self-improve via detailed tracing tightly

agentslinus-lee--x
27 May 2026
Safety

“Tokens got burned for millions of dollars without any real significant ROI to show for it.” hearing this over and over again

DGX agent

“Tokens got burned for millions of dollars without any real significant ROI to show for it.” hearing this over and over again The same conversation is happening across tech right now and many of us sa

safetygary-marcus--x
27 May 2026
Industry

US law enforcement warns of 'anti-tech extremism' as AI hatred grows

DGX agent

U.S. law enforcement agencies are monitoring growing backlash to AI and have begun classifying anti-technology sentiment as an extremism threat, with unpublished reports from the Department of Homelan

industryars-technica
27 May 2026
Model Releases

AI Content Moderation in Therapy Conversations

DGX agent

arXiv:2605.25454v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly being used for emotional support. They are also being developed for formal therapy purposes. However, LL

model-releasesarxiv-cs-ai
26 May 2026
Research

Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models

DGX agent

arXiv:2605.24053v1 Announce Type: new Abstract: Large Language Models (LLMs) are predominantly governed by probabilistic frameworks in which the sum of outcome probabilities is constrained to unity. T

researcharxiv-cs-ai
26 May 2026
Model Releases

D^2-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing

DGX agent

arXiv:2605.25893v1 Announce Type: new Abstract: Despite the emergence of diffusion large language models (D-LLMs) as an alternative to autoregressive large language models (AR-LLMs), safety monitoring

model-releasesarxiv-cs-ai
26 May 2026
Safety

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection

DGX agent

arXiv:2509.13608v2 Announce Type: replace Abstract: As Large Multimodal Models (LMMs) become integral to daily digital life, understanding their safety architectures is a critical problem for AI Align

safetyarxiv-cs-lg
26 May 2026
Model Releases

MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence

DGX agent

arXiv:2505.23764v3 Announce Type: replace-cross Abstract: Spatial intelligence is essential for multimodal large language models (MLLMs) operating in the complex physical world. Existing benchmarks, h

model-releasesarxiv-cs-cl
26 May 2026
Applications

Nano Banana Pro vs. GPT Image 2 Nano Banana Pro wins on photorealism, 4K output, 14 reference image slots for product/scene consistency, and…

DGX agent

Nano Banana Pro vs. GPT Image 2 Nano Banana Pro wins on photorealism, 4K output, 14 reference image slots for product/scene consistency, and live web search for real-world accuracy. GPT Image 2 wins o

applicationscomfyui--x
26 May 2026
Model Releases

ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs

DGX agent

arXiv:2507.10593v3 Announce Type: replace-cross Abstract: Every LLM tool call is structurally an RPC -- a function name, JSON arguments, and a serialized result -- yet each protocol (native Python, MC

model-releasesarxiv-cs-ai
26 May 2026
Safety

A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies (New York Times)

DGX agent

New York Times: A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies — The government's A.I. Se

safetytechmeme
25 May 2026
Model Releases

An open source model has returned to #1 on the 3D Design leaderboard by Design Arena. Kimi K2.6 has reached the top of the leaderboard for 3…

DGX agent

An open source model has returned to #1 on the 3D Design leaderboard by Design Arena. Kimi K2.6 has reached the top of the leaderboard for 3D Design, ahead of models 10X more expensive like Opus 4.7 b

model-releaseskimi-moonshot--x
25 May 2026
Model Releases

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

DGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

model-releasesarxiv-cs-cl
25 May 2026
Safety

Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control

DGX agent

arXiv:2605.23415v1 Announce Type: cross Abstract: Reinforcement learning has long struggled with poor sample efficiency. One promising approach to mitigate this problem is leveraging group-invariant M

safetyarxiv-cs-ai
25 May 2026
← Previous
1…4546474849…54
Next →