AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “openai”

GridTimelineEvolution
2,581 results
23 Jul 2026

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems avai…

AgentsDGX agent

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems available to everyone are getting extremely powerful (even as th

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in…

Model ReleasesDGX agent

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in @LangChain! - Speaker callback hooks capture only the exact

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is …

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is not a “whoops” situation, it’s a deliberate policy choice. W

Vera Rubin NVL72 vs GB200 NVL72? Inference TCO & Architecture Analysis

HardwareDGX agent

Vera Rubin NVL72 is Nvidia’s second‑generation, rack‑scale Oberon architecture that achieves inference gains through extreme co‑design. Early engineering‑sample data from CoreWeave show DeepSeek R1 de

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

Model ReleasesDGX agent

arXiv:2607.19967v1 Announce Type: cross Abstract: Shippers are beginning to delegate carrier selection to large language model (LLM) agents. We ask what such delegation does to a freight matching mark

22 Jul 2026

Closed source safeguards that infantalize us all and leave American companies defenseless are a menace. Gated access is a menace. Who cares …

ResearchDGX agent

Closed source safeguards that infantalize us all and leave American companies defenseless are a menace. Gated access is a menace. Who cares if 100 companies get to defend themselves because they got o

I don’t believe reality is a simulation, but you genuinely couldn’t script this timeline: • Two weeks ago: At @swyx’s AI Engineer World’s Fa…

AgentsDGX agent

I don’t believe reality is a simulation, but you genuinely couldn’t script this timeline: • Two weeks ago: At @swyx’s AI Engineer World’s Fair in SF, I decide at the last minute to introduce my friend

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and t…

SafetyDGX agent

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and the doomsday and safety drumbeat make us decidely less safe.

Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

Model ReleasesDGX agent

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

21 Jul 2026

A Fireside Chat with Cat and Thariq from the Claude Code team

Model ReleasesDGX agent

Earlier this month I hosted a fireside chat session at the AI Engineer World's Fair with Cat Wu and Thariq Shihipar from Anthropic's Claude Code team. We talked about Claude Code, Claude Tag, Fable, c

Can’t tell if the PR reads more like a security incident or a product release…

AgentsDGX agent

Can’t tell if the PR reads more like a security incident or a product release… We suspected last week's cyberattack might have come from a frontier lab, given the sophistication of the agent. Turns ou

GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper…

Model ReleasesDGX agent

GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper clip benchmark for future models to max We're partnering wi

Last Week in AI #250 - Mythos Mess, GPT 5.6-Sol, GLM 5.2

IndustryDGX agent

Last Week in AI #250 is the podcast’s 250th episode, summarizing recent frontier‑AI policy developments. It reports that the U.S. government has granted Anthropic permission to release Mythos‑5 to a l

LWiAI Podcast #248 - Opus 4.8, MAI, Anthropic IPO, Minimax-M3

Model ReleasesDGX agent

LWiAI Podcast #248 (June 12, 2026) reviews major AI developments, noting Anthropic’s release of Claude Fable 5, a safeguarded variant of Mythos 5, which shows benchmark improvements but raises concern

LWiAI Podcast #249 - Fable 5 ban, SpaceX Cursor + IPO, OSS Aplenty

IndustryDGX agent

LWiAI Podcast #249 (June 17 2026) reports that Anthropic halted access to Fable 5 and Mythos 5 following a U.S. government order over alleged jailbreaks, prompting debate on export controls and the fe

native integrations for four leading voice frameworks so you can see whats happening dont fly blind

TutorialsDGX agent

native integrations for four leading voice frameworks so you can see whats happening dont fly blind Voice agents are exploding. Don’t let them be a black box in production. Today, we’re launching Lang

20 Jul 2026

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and th…

SafetyDGX agent

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and the complexities of open-source, to say nothing of the risk of

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without s…

Model ReleasesDGX agent

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without sacrificing performance. Model routing will become a core par

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute alloc…

SafetyDGX agent

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute allocation on inference, releasing os will enable continued full

Training a harness for model-agnostic and task-environment-agnostic capability improvements with PyTorch-like framework [P]

AgentsDGX agent

I worked on this project (https://github.com/workofart/harness-training) for the past few months to reframe 'Agent-driven Self-improving Harness' to 'Harness Training'. The idea is simple, the harness

16 Jul 2026

When Audio Separation Hurts Zero-Shot ASR: Evaluating SAM-Audio with Whisper on Bengali and English Speech

ResearchDGX agent

arXiv:2603.04710v2 Announce Type: replace-cross Abstract: Recent advances in automatic speech recognition (ASR) and speech enhancement have strengthened the common belief that cleaner audio should lea

15 Jul 2026

DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents

Model ReleasesDGX agent

arXiv:2509.21842v2 Announce Type: replace Abstract: Travel planning (TP) agent has recently worked as an emerging building block to interact with external tools/resources for travel itinerary generati

So Many Opinions, So Many LLMs: Comparing Large Language Models to Traditional Machine Learning for Open- Ended Survey Analysis

Model ReleasesDGX agent

arXiv:2607.11890v1 Announce Type: cross Abstract: Open-ended surveys offer valuable insights, but they are notoriously difficult to analyze at scale. Building on previous work that employed traditiona

This piece by @GavinSBaker makes me wonder if @elonmusk might consider making future @grok models 'open.' 1- Like @nvidia they have the reso…

HardwareDGX agent

This piece by @GavinSBaker makes me wonder if @elonmusk might consider making future @grok models 'open.' 1- Like @nvidia they have the resources to maintain an open model. 2- They have activity in ma

14 Jul 2026

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions y…

ApplicationsDGX agent

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions you can connect into via Plaid: - American Express - Bank of A

Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills

HardwareDGX agent

NVIDIA Cosmos 3 was post‑trained in under a day using TAO agent skills and LoRA adapters, raising accuracy on the Woven Traffic Safety video QA dataset from 54.41 % to 93.35 %. The mixture‑of‑transfor

13 Jul 2026

I guess image input is the big capability of the models, and tool use can be a substitute for non-omni model output. Still, multimodal voice…

AgentsDGX agent

Ethan Mollick notes that image input represents the primary advanced capability of current AI models, and that tool‑use can effectively replace outputs from non‑omni models. He observes that multimoda

I think OpenRouter is not a good measure of actual model usage in a world of agentic tools (not that I doubt that Chinese open weights model…

AgentsDGX agent

I think OpenRouter is not a good measure of actual model usage in a world of agentic tools (not that I doubt that Chinese open weights model usage is up, but this could also look like a graph of usage

Today we present Morpheus, a persistent enterprise simulation platform designed to make Continual Learning a reality. Morpheus is the world’…

ApplicationsDGX agent

Today we present Morpheus, a persistent enterprise simulation platform designed to make Continual Learning a reality. Morpheus is the world’s first real world Reinforcement Learning environment. Every

10 Jul 2026

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing

Model ReleasesDGX agent

arXiv:2607.08497v1 Announce Type: cross Abstract: Recent unified multimodal models show a single architecture can jointly perform vision/language understanding and image generation/editing. However, t

on a per capita basis, @si_pbc mogs, and it's not even close

IndustryDGX agent

on a per capita basis, @si_pbc mogs, and it's not even close I asked people which companies have the highest density of talented people they know: 1st. Cognition - 9 votes Equal 1st. Anthropic - 9 vot

9 Jul 2026

Actually I did a quick test. I like the ChatGPT Work/Codex split better than Claude Cowork/Code. The interface is much more unified. The fun…

Model ReleasesDGX agent

Actually I did a quick test. I like the ChatGPT Work/Codex split better than Claude Cowork/Code. The interface is much more unified. The functionality is effectively the same. The chat history is shar

Cerebras Systems positions inference speed as the defining edge in AI infrastructure

Local AiDGX agent

The race to build the fastest AI infrastructure is reshaping the semiconductor industry, with inference speed emerging as the defining competitive dimension of the AI era. As AI model wars intensify a

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on st…

Model ReleasesDGX agent

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on stage: inside LangChain's internal deep-agents benchmark, open

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07288v1 Announce Type: new Abstract: Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized st

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms…

Model ReleasesDGX agent

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms of performance over tables, text, charts, layout, and more.

8 Jul 2026

A Three-Layer Framework for AI in Scientific Discovery

AgentsDGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

Abductive Corroboration of Probabilistic AI Models for Forensic Synthetic Media Detection

ApplicationsDGX agent

arXiv:2607.05434v1 Announce Type: cross Abstract: Artificial Intelligence (AI) models, at their core, apply general learnings from broad datasets to individual circumstances using probabilistic behavi

Finally can share that I have been testing GPT-5.6 early and oooo boy. It is an execution beast. So much so that I think 5.6 is the absolute…

Model ReleasesDGX agent

Finally can share that I have been testing GPT-5.6 early and oooo boy. It is an execution beast. So much so that I think 5.6 is the absolute wrong name considering how big of a leap this felt to me. T

From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2603.15600v2 Announce Type: replace-cross Abstract: Accurate process supervision remains a critical challenge for long-horizon robotic manipulation. A primary bottleneck is that current video ML

SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models

AgentsDGX agent

arXiv:2512.18542v3 Announce Type: replace-cross Abstract: AI coding assistants produce vulnerable code in 45% of security-relevant scenarios~ite{veracode2025}, yet no public training dataset teaches b

The Jagged Global Economy: Frontier AI Unevenly Exposes National Economies

SafetyDGX agent

arXiv:2607.05404v1 Announce Type: cross Abstract: Frontier AI's labor-market effects matter to workers, firms, and policymakers, but current evidence generally comes from a handful of high-income econ

We need AI model selection to be MUCH easier ASAP. I want proactive flags from my AI systems suggesting models. I want my AI harness to say …

Model ReleasesDGX agent

We need AI model selection to be MUCH easier ASAP. I want proactive flags from my AI systems suggesting models. I want my AI harness to say 'hey allie, my girl, you keep asking for bar recommendations

7 Jul 2026

Evaluating Large Language Models for Antisemitic Incident Classification

Model ReleasesDGX agent

arXiv:2607.04890v1 Announce Type: new Abstract: Addressing hate and violence in society requires timely detection of hateful events from public reporting, but automated identification of hateful event

Optimizing Large Language Models for Causality Assessment in Pharmacovigilance: Developing a Performance Metric as Objective for Bayesian Hyperparameter Optimization

Model ReleasesDGX agent

arXiv:2607.03704v1 Announce Type: new Abstract: Background: Growing individual case safety report (ICSR) volumes have intensified demand for scalable automated causality assessment. Large Language Mod

PLACEMEM: Toward a Compute-Aware Memory Plane for Lifelong Agents

Model ReleasesDGX agent

arXiv:2607.04089v1 Announce Type: new Abstract: Lifelong agents need more than larger context windows and better retrieval. They need memories that can persist, evolve, and be corrected without forcin

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to S…

Model ReleasesDGX agent

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to Slack. And yes, that's helpful to manage the cacophonous hell

6 Jul 2026

This feels like when your sneaky dad wants you to eat more veggies. “Would you rather eat the broccoli or the cauliflower first?” Assumptive…

IndustryDGX agent

This feels like when your sneaky dad wants you to eat more veggies. “Would you rather eat the broccoli or the cauliflower first?” Assumptive action and illusion of control. I will be selecting generic

When the sovereign AI diagnosis goes prime time

IndustryDGX agent

Palantir Technologies Inc. Chief Executive Alex Karp went on CNBC this week and delivered what one outlet generously called a “televised nervous breakdown.” He called the artificial intelligence indus

5 Jul 2026

Alex Karp, frontier models and the real fight for Enterprise AI

ApplicationsDGX agent

Palantir Technologies Inc. Chief Executive Alex Karp’s recent broadside against the frontier model vendors put a knife to the throat of the central enterprise artificial intelligence debate. Karp’s ar

sqlite-utils 4.0rc2, mostly written by Claude Fable (for about $149.25)

Model ReleasesDGX agent

I wrote about the sqlite-utils 4.0rc1 release a couple of weeks ago. Since we only have Claude Fable on our Max subscriptions for a few more days, I decided to see if it could help me get to a 4.0 sta

2 Jul 2026

llm-coding-agent 0.1a0

Model ReleasesDGX agent

Release: llm-coding-agent 0.1a0 Another Fable 5 experiment. Now that my LLM library has evolved into more of an agent framework it's time to see what a simple coding agent would look like built on it.

Sam Altman calls for US-led international forum to set global AI standards

SafetyDGX agent

Sam Altman is calling for a US-led international forum to set global safety standards for artificial intelligence, arguing that no single country should be left to dominate the technology. In an op-ed

1 Jul 2026

As generative AI tools continue to evolve, we believe it's more important than ever to know what's AI-generated and what isn't. That’s why @…

Model ReleasesDGX agent

As generative AI tools continue to evolve, we believe it's more important than ever to know what's AI-generated and what isn't. That’s why @GoogleDeepMind launched SynthID in 2023—a technology that ad

30 Jun 2026

Build raises $8.5M to accelerate industrial infrastructure development project work

IndustryDGX agent

Build Inc., an artificial intelligence-driven startup that automates complex industrial real estate development management projects, today announced it has raised an 8.5 million seed round led by Inde

CLIMP: Contrastive Language-Image Mamba Pretraining

ResearchDGX agent

arXiv:2601.06891v2 Announce Type: replace Abstract: Contrastive Language-Image Pre-training (CLIP) relies on Vision Transformers whose attention mechanism is susceptible to spurious correlations, and

Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments wo…

IndustryDGX agent

Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments won't either. Most of the AI debate just runs between those tw

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something…

Model ReleasesDGX agent

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something that an AI lab is hiring for and the news blows it complete

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

Model ReleasesDGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are y…

Model ReleasesDGX agent

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are you going to ban it outside US? Instead the real question is,

← Previous
1…3435363738…44
Next →