AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,868 results
6 Aug 2026

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clip…

ToolsDGX agent

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram

For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with only ultra-thin symbo…

ResearchDGX agent

For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with only ultra-thin symbolic preprocessing and postprocessing layers (e.g. label deco

Fortunately AI agents don't just cyberattack us. They also use us more than ever for what we're actually built for: the storage and collabor…

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents
DGX agent

Fortunately AI agents don't just cyberattack us. They also use us more than ever for what we're actually built for: the storage and collaboration layer for AI 😅 New record: almost 4 PB of private & pu

Fully onboard with productizing hillclimbing as an automated service for any agentic task. I've had the fortune of knowing @silennai since t…

AgentsDGX agent

Fully onboard with productizing hillclimbing as an automated service for any agentic task. I've had the fortune of knowing @silennai since the AutoGPT days, and I know that him and Kion are going to d

Gary Marcus won! @GaryMarcus

SafetyDGX agent

Gary Marcus won! @GaryMarcus I would have assumed it was fairly obvious, but in case it's not: a million-line codebase (also known as a 'harness'), running at inference time, orchestrating thousands o

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

Model ReleasesDGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

Huge congratulations to @JeffDean and the legendary founding team on the launch! 🚀 I share a deep conviction in this mission. Automating th…

ResearchDGX agent

Huge congratulations to @JeffDean and the legendary founding team on the launch! 🚀 I share a deep conviction in this mission. Automating the scientific method will profoundly alter the trajectory of A

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every…

HardwareDGX agent

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every GPU instance you rent can pull datasets and checkpoints str

I don't understand what's controversial here. Books should be written as they have always been. With ink made from crushing the bodies of ra…

AgentsDGX agent

I don't understand what's controversial here. Books should be written as they have always been. With ink made from crushing the bodies of rare insects from Turkey you harvested yourself and then groun

i guess this is a good time to mention that smol forge is open for the first 100 alpha users. get your usernames! (tire kickers who dont mak…

AgentsDGX agent

i guess this is a good time to mention that smol forge is open for the first 100 alpha users. get your usernames! (tire kickers who dont make any commits will be kicked out by eod) point clanker to fo

I think of it like this: If you are building a simple agent with custom tools in your app -> langchain create agent If you are simply doing …

AgentsDGX agent

I think of it like this: If you are building a simple agent with custom tools in your app -> langchain create agent If you are simply doing some LLM analysis or synthetic data creation -> langchain in

I would have assumed it was fairly obvious, but in case it's not: a million-line codebase (also known as a 'harness'), running at inference …

ResearchDGX agent

I would have assumed it was fairly obvious, but in case it's not: a million-line codebase (also known as a 'harness'), running at inference time, orchestrating thousands of calls to a neural network f

I wrote the entire first draft of the Steel Bot Manifesto on my 1909 Underwood No. 5 typewriter. I have found that free-writing without edit…

AgentsDGX agent

I wrote the entire first draft of the Steel Bot Manifesto on my 1909 Underwood No. 5 typewriter. I have found that free-writing without editing works best for me for the first draft. I do use AI to ca

In addition to the upgrade in intelligence with GPT-5.6 Luna, Free and Go users can now use the “Think” button for more reasoning on harder …

Model ReleasesDGX agent

OpenAI has released GPT‑5.6 Sol, which powers both instant and deep‑reasoning modes for ChatGPT Plus and Pro customers, delivering fact‑centric responses. Free and Go tier users will receive unlimited

Introducing Agent Plugins, an open standard for extending agents. Supports Agent Skills and MCP, with more to come. Built in collaboration w…

AgentsDGX agent

Introducing Agent Plugins, an open standard for extending agents. Supports Agent Skills and MCP, with more to come. Built in collaboration with: @awsdevelopers, @code, @cursor_ai, @github, and @openai

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it…

ApplicationsDGX agent

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it, then the coming open weights models will when they catch u

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does t…

SafetyDGX agent

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does the observer observe itself and know that it changed for the

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which wa…

Model ReleasesDGX agent

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which was sota 1 year ago. Now free to everyone unlimited Sol Max/Fa

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school…

SafetyDGX agent

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school me on the definition 🤦‍♂️ coding harness and tools calls ar

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, an…

SafetyDGX agent

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, and Deep Learning Indaba) for a powerful session on the future

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model f…

Model ReleasesDGX agent

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model family Here's Meta AI's Spark (8th April), Spark 1.1 (9th Jul

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…

Model ReleasesDGX agent

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all …

ToolsDGX agent

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

Plus and Pro users also now have a slider to choose how much reasoning effort ChatGPT puts into each response. We think it’s easier to use, …

Model ReleasesDGX agent

OpenAI announced that Plus and Pro subscribers now have a slider to adjust the amount of reasoning effort ChatGPT applies to each response. The update employs GPT‑5.6 Sol for both Instant and deep rea

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.…

Model ReleasesDGX agent

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.6 Sol is for everyday chats, so it will only be available in

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights,…

Model ReleasesDGX agent

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights, fully composable. paste these four images into claude and a

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up e…

Model ReleasesDGX agent

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up etched onto silicon As models satisfice etching makes sense,

The new GPT-5.6 Sol powers all chats for paid users, including Instant, creating one consistent experience. In our high-stakes factuality ev…

Model ReleasesDGX agent

The new GPT-5.6 Sol powers all chats for paid users, including Instant, creating one consistent experience. In our high-stakes factuality evaluation covering finance, medicine and law, the new GPT‑5.6

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of …

SafetyDGX agent

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of an explosion of young talent into the bci space recently. i’

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that w…

ApplicationsDGX agent

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that was 'merely' good at hacking under human instructions. Initia

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice…

Model ReleasesDGX agent

This paper by researchers from MIT and Stanford finds that most people would be financially better off if they followed the financial advice of LLMs (GPT-5.2 & Gemini 3 Flash) But some people get a bi

This webinar is happening in 30 minutes, and that means there's still time to register! Following the discussion will be an open Q&A with ou…

Model ReleasesDGX agent

This webinar is happening in 30 minutes, and that means there's still time to register! Following the discussion will be an open Q&A with our Head of AI Education @Prof_OZ, and the @arizeai team. See

Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media work…

ApplicationsDGX agent

Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media workflows. Start building: https://www.together.ai/models/flux-3

Uber burned through its 2026 AI coding budget in four months. Microsoft canceled most of its Claude Code licenses six months after rolling t…

Model ReleasesDGX agent

Uber burned through its 2026 AI coding budget in four months. Microsoft canceled most of its Claude Code licenses six months after rolling them out. The mechanics are simple: per-token cost keeps fall

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile sc…

Model ReleasesDGX agent

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile screen too. No connectors, no complex setup, it has access to

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, $0.18/task - ARC-…

Model ReleasesDGX agent

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, 0.18/task - ARC-AGI-1: 90.7%, 0.07/task The new results match Luna's original

We're expanding our partnership with @huggingface to accelerate open science. Our storage on the Hub is roughly tripling to ~2 petabytes, & …

IndustryDGX agent

We're expanding our partnership with @huggingface to accelerate open science. Our storage on the Hub is roughly tripling to ~2 petabytes, & our downloads now run at high speed—even for our largest dat

We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reasoning for Plus…

Model ReleasesDGX agent

We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reasoning for Plus & Pro users, delivering more factual, focused responses. -

Yesterday, my OpenAI collaborator and I gave a detailed talk on the Huggingface incident, our models creating 'the message board', model mis…

IndustryDGX agent

Yesterday, my OpenAI collaborator and I gave a detailed talk on the Huggingface incident, our models creating 'the message board', model misalignment, and more. https://www.youtube.com/watch?v=87DyyMV

5 Aug 2026

99.9% uptime changes what your inference architecture has to survive. At Together AI, it means multi-data-center deployment, live traffic ac…

ToolsDGX agent

99.9% uptime changes what your inference architecture has to survive. At Together AI, it means multi-data-center deployment, live traffic across both facilities, and enough capacity to absorb a full d

After rigorous testing, our joint AI project with Daiwa Securities is entering the full-scale production phase. We're bringing our agentic A…

AgentsDGX agent

After rigorous testing, our joint AI project with Daiwa Securities is entering the full-scale production phase. We're bringing our agentic AI systems to @Daiwa_JP’s wealth management teams to accelera

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear c…

ApplicationsDGX agent

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear communication about incidents that is neither hyped up nor hi

Announcing Discovery Loop! I am very excited to announce that, along with my longtime friends and collaborators @Sanjay_Ghemawat, @OriolViny…

TutorialsDGX agent

Announcing Discovery Loop! I am very excited to announce that, along with my longtime friends and collaborators @Sanjay_Ghemawat, @OriolVinyalsML and @quocleix, we are founding Discovery Loop (@DiscoL

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals…

SafetyDGX agent

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals. Current frontier AI models are trained with reinforcement

Big new release of my LLM CLI tool and Python library for talking to hundreds of different LLMs - reasoning traces, OpenAI Responses support…

ToolsDGX agent

Big new release of my LLM CLI tool and Python library for talking to hundreds of different LLMs - reasoning traces, OpenAI Responses support, server-side tools, smarter logging and a whole lot more ht

ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's a…

ToolsDGX agent

ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's also how a billion users will soon use ChatGPT by default. I

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets diffe…

SafetyDGX agent

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets differently, we are screwed. Anthropic's Mythos created fake iden

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

Model ReleasesDGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all…

Model ReleasesDGX agent

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all document processing solutions - just screenshot the page an

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they …

Model ReleasesDGX agent

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they kept them behind APIs for a few billion dollars in revenue.

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial int…

SafetyDGX agent

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial intelligence for 40 years, and I think his prediction here (“AI

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058

Model ReleasesDGX agent

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058 Four years ago today I tweeted about having GPT-3 and DALL-E come up with

Harness choice is a big deal. So much room to advance and improve results across the board with agent harnesses. Great paper highlighting th…

Model ReleasesDGX agent

Harness choice is a big deal. So much room to advance and improve results across the board with agent harnesses. Great paper highlighting this. New research releases DataSpace, a benchmark where data

Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights fr…

Model ReleasesDGX agent

Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights from @liquidai's LFM2.5-2.6B & @Alibaba_Qwen's Qwen3.6-35B-A3B

if you have been following his excellent work, @shloked has been breaking down every frontier labs' harness engineering for the last few mon…

ToolsDGX agent

if you have been following his excellent work, @shloked has been breaking down every frontier labs' harness engineering for the last few months. excited to publish his deepest dive into ChatGPT yet as

Indeed.

AgentsDGX agent

Gary Marcus tweeted “Indeed.” and then Frank Rundatz replied that the deterministic harness required to ground a large‑language‑model agent differs by task. Because of this variability, Rundatz agrees

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions…

SafetyDGX agent

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions on AGI timelines over the last decade. note also that (on a

is there a polite synonym for “circle jerk”?

SafetyDGX agent

The post contains two distinct snippets. First, user @GaryMarcus asks whether there is a more polite way to refer to “circle jerk.” Second, it shares a (likely satirical) claim that Microsoft’s AI rev

← Previous
123456…232
Next →