AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,867 results
Hardware

Can an open-source model perform like a foundation model? @Osmosis_AI is betting yes, using reinforcement learning and the dedicated @ycombi…

DGX agent

Osmosis_AI claims that an open‑source model can rival a foundation model by leveraging reinforcement learning techniques. To demonstrate this, they will use the Y Combinator‑dedicated GPU cluster on T

hardwaretogether-ai--x
30 Jul 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep t…

DGX agent

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep their infrastructure secure. Everybody deserves access to mod

safetycohere--x
30 Jul 2026
Model Releases

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyU…

DGX agent

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyUI - combining Florence 2, SAM2, and WAN Video into a single

model-releasescomfyui--x
30 Jul 2026
Model Releases

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success …

DGX agent

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success rate. - Unoptimized? - Buggy? - Deliberate (coz that helps i

model-releasessebastian-raschka--x
30 Jul 2026
Agents

Everyone is a real-world AI engineer. They don’t just make cars, but autonomous robots with frontier capabilities. Thanks for their hard wor…

DGX agent

Everyone is a real-world AI engineer. They don’t just make cars, but autonomous robots with frontier capabilities. Thanks for their hard work keeping the line smooth and running. It’s my pleasure work

agentselon-musk--x
30 Jul 2026
Tools

Excited to be on the CNBC live show!

DGX agent

Excited to be on the CNBC live show! Back from vacation and LIVE at 12pm PT / 3pm ET Is AI’s easy-money era ending? We’ll unpack a wild week for the AI trade—big tech earnings, Leopold Aschenbrenner’s

toolsfireworks-ai--x
30 Jul 2026
Local Ai

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3.

DGX agent

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3. Built to bring open-source LLMs to private machines, @Ollama uses Intel Core Ultra Series 3 to run

local-aiollama--x
30 Jul 2026
Agents

Finally a good paper testing if file-system based memory for LLM agents is worth it. First, what does this look like? Deployed agents keep l…

DGX agent

Finally a good paper testing if file-system based memory for LLM agents is worth it. First, what does this look like? Deployed agents keep long-term memory as a folder of markdown files they read and

agentsdair-ai--x
30 Jul 2026
Model Releases

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making th…

DGX agent

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @Goog

model-releasesgoogle-ai--x
30 Jul 2026
Model Releases

goblin-level blog post

DGX agent

goblin-level blog post Turns out GPT-5.6 Sol is actually SoTA on ARC-AGI-3. Just took two setting changes. You just have to allow it to reason and work over multiple context windows with the help of o

model-releasessam-altman--x
30 Jul 2026
Model Releases

good job little bro

DGX agent

good job little bro After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. The results: - 20% lower serving costs from production GPU ke

model-releasessam-altman--x
30 Jul 2026
Model Releases

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of doll…

DGX agent

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of dollars a month in savings at this point? Codex analysed product

model-releasessimon-willison--x
30 Jul 2026
Model Releases

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep,…

DGX agent

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep, I wanted to quickly jot down my thinking here. The basic is

model-releasesgary-marcus--x
30 Jul 2026
Applications

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve w…

DGX agent

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve worked on many of the foundational technologies that power mo

applicationsfireworks-ai--x
30 Jul 2026
Research

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥

DGX agent

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥 Today, we are releasing Inkling-Small. Inkling-Small achieves comparable performance to In

researchsoumith-chintala--x
30 Jul 2026
Agents

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also ac…

DGX agent

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt cache hit rates. Low prici

agentstogether-ai--x
30 Jul 2026
Agents

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. …

DGX agent

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. That is especially true for document parsing 📄🔀: - Some page

agentsjerry-liu--x
30 Jul 2026
Tools

It's wild to me that both Anthropic and OpenAI have products that lean so hard on search, and yet they both obscure the underlying search in…

DGX agent

Simon Willison notes that both Anthropic and OpenAI build products that rely heavily on searching data but keep the underlying search index they use hidden from public view. He finds it surprising how

toolssimon-willison--x
30 Jul 2026
Model Releases

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/207051816739969…

DGX agent

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/2070518167399698490?s=20 'I tried to analyze why Claude Code uses more toke

model-releasessebastian-raschka--x
30 Jul 2026
Model Releases

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of…

DGX agent

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we ar

model-releasesopenai--x
30 Jul 2026
Model Releases

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most pars…

DGX agent

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most parsing pipelines throw all of them at the same parser, forcing

model-releasesjerry-liu--x
30 Jul 2026
Safety

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

DGX agent

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

safetyfireworks-ai--x
30 Jul 2026
Tools

OpenAI had a partnership with Bing, but also run their own crawling and indexing infrastructure. Does ChatGPT decide any Bing at all these d…

DGX agent

OpenAI has collaborated with Microsoft’s Bing while also running its own web‑crawling and indexing systems, and Anthropic similarly relies on search‑derived data. Both firms prominently incorporate se

toolssimon-willison--x
30 Jul 2026
Model Releases

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large…

DGX agent

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large enterprise codebases. We helped them to scale the training

model-releasesfireworks-ai--x
30 Jul 2026
Model Releases

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the b…

DGX agent

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the benchmark or that contain knowledge about the benchmark forma

model-releasesfrancois-chollet--x
30 Jul 2026
Safety

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from …

DGX agent

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from investors after suffering heavy losses during the recent rou

safetygary-marcus--x
30 Jul 2026
Applications

Starting in 15 minutes

DGX agent

Starting in 15 minutes Kimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it. He joins Jue Wang and Zain Hasan from Together A

applicationstogether-ai--x
30 Jul 2026
Model Releases

sure we lose money on every inference but we make it up in volume

DGX agent

sure we lose money on every inference but we make it up in volume We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices f

model-releasesgary-marcus--x
30 Jul 2026
Safety

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that lat…

DGX agent

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that latency by predicting and pre-executing the next call, but exte

safetydair-ai--x
30 Jul 2026
Safety

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, man…

DGX agent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, many forefront members of the AI-Safety community, in their fer

safetyyann-lecun--x
30 Jul 2026
Model Releases

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many b…

DGX agent

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many benchmarks. To test its speed, we plugged it into HF's speech

model-releasessoumith-chintala--x
30 Jul 2026
Model Releases

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three sepa…

DGX agent

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three separate companies back in April without them noticing! In a rev

model-releasessimon-willison--x
30 Jul 2026
Model Releases

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT…

DGX agent

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a f

model-releasesopenai--x
30 Jul 2026
Agents

We have to be careful to not offload our understanding to agents. I think there is also a good opportunity to build agentic applications tha…

DGX agent

We have to be careful to not offload our understanding to agents. I think there is also a good opportunity to build agentic applications that encourage deeper understanding. For example, coding agents

agentsdair-ai--x
30 Jul 2026
Agents

We’re excited to rollout an official batch parsing experience to LlamaParse. ✅ Instead of hitting our APIs one file at a time, create a batc…

DGX agent

We’re excited to rollout an official batch parsing experience to LlamaParse. ✅ Instead of hitting our APIs one file at a time, create a batch of 10k at once. ✅ Get a dedicated UI where you can audit t

agentsjerry-liu--x
30 Jul 2026
Model Releases

Yesterday I cohosted a dinner with @dexhorthy with a wonderful group of founders, to talk about agent loops and loop engineering. Some inter…

DGX agent

Yesterday I cohosted a dinner with @dexhorthy with a wonderful group of founders, to talk about agent loops and loop engineering. Some interesting insights: * Most of our group was *not* actively usin

model-releasesjerry-liu--x
30 Jul 2026
Model Releases

A benchmark score reflects the model as well as the harness and settings used to run it. For long-running agents, retaining reasoning and co…

DGX agent

A benchmark score reflects the model as well as the harness and settings used to run it. For long-running agents, retaining reasoning and compacting context lets the model build on what it has already

model-releasesopenai--x
29 Jul 2026
Model Releases

A new TIL on adding custom MCP servers to both the ChatGPT and Claude regular chat interfaces - it's a little less obvious than I had hoped,…

DGX agent

A new TIL on adding custom MCP servers to both the ChatGPT and Claude regular chat interfaces - it's a little less obvious than I had hoped, but I got there in the end https://til.simonwillison.net/ll

model-releasessimon-willison--x
29 Jul 2026
Model Releases

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models a…

DGX agent

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models are like that. So what changes? The way to interact with Opus

model-releasesdair-ai--x
29 Jul 2026
Model Releases

After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. The results: - 20% lo…

DGX agent

After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. The results: - 20% lower serving costs from production GPU kernel improvements. -

model-releasesopenai--x
29 Jul 2026
Agents

Agentic inference wastes GPUs on KV cache thrashing. ThunderAgent fixes it at the scheduler level: 2.5x higher single-node throughput and ~1…

DGX agent

Agentic inference wastes GPUs on KV cache thrashing. ThunderAgent fixes it at the scheduler level: 2.5x higher single-node throughput and ~10x lower P50 latency at high concurrency. ThunderAgent was a

agentstogether-ai--x
29 Jul 2026
Hardware

agreed. which is part of why coating the world in data centers is a profound mistake.

DGX agent

agreed. which is part of why coating the world in data centers is a profound mistake. AI will get so ridiculously efficient that we will look back at GPU clusters the way we now look at these first ro

hardwaregary-marcus--x
29 Jul 2026
Agents

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a require…

DGX agent

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a requirement. In our new short course, AI Code Review, built in coll

agentsitamar-friedman--x
29 Jul 2026
Agents

Batch parsing used to mean writing API scripts. Now it's a button. 🦙 Parse or extract across up to 10,000 files in one run, straight from t…

DGX agent

Batch parsing used to mean writing API scripts. Now it's a button. 🦙 Parse or extract across up to 10,000 files in one run, straight from the LlamaParse UI. Point it at a folder and go — no code requi

agentsjerry-liu--x
29 Jul 2026
Model Releases

BREAKING: Grok 4.5 (high) ranks #1 on the HighWalk benchmark, which tests how well AI agents update technical specifications from code chang…

DGX agent

BREAKING: Grok 4.5 (high) ranks #1 on the HighWalk benchmark, which tests how well AI agents update technical specifications from code changes. Grok delivered the best combination of quality and opera

model-releaseselon-musk--x
29 Jul 2026
Model Releases

BREAKING: Grok 4.5 just claimed the top spot on the new HighWalk Benchmark. The independent test measures how well AI models update real tec…

DGX agent

BREAKING: Grok 4.5 just claimed the top spot on the new HighWalk Benchmark. The independent test measures how well AI models update real technical specifications from 46 Laravel commits — heavy on cod

model-releaseselon-musk--x
29 Jul 2026
Model Releases

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. Th…

DGX agent

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-world AI agents across conversation,

model-releaseselon-musk--x
29 Jul 2026
Model Releases

BREAKING: SpaceXAI's newly released Grok Voice Think Fast 2.0 beats voice models from OpenAI, Google, Alibaba, and DeepSlate in the Artifici…

DGX agent

SpaceXAI has released its new Grok Voice Think Fast 2.0, which on the Artificial Analysis Speech‑to‑Speech benchmark outperformed leading models from OpenAI, Google, Alibaba and DeepSlate. The claim w

model-releaseselon-musk--x
29 Jul 2026
← Previous
1…910111213…289
Next →