AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,897 results
31 Jul 2026

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets

Model ReleasesDGX agent

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets major price cuts today: *80% drop for GPT-5.6 Luna, now 0.20 per million input tokens and 1.20 per million output *20% drop

Period.

Model ReleasesDGX agent

Period. If LoRA is underperforming, don't reach for more expensive full parameter fine-tuning right away. We ran three cheap tests (data coverage, optimization, rank) to see if we could close the gap

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation…

HardwareDGX agent

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation of software, the systems around them must be transparent, i

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The first thing we learned building autoscaling for dedicated inference is that CPU-style metrics don't tell the whole picture. A GPU can re…

HardwareDGX agent

The first thing we learned building autoscaling for dedicated inference is that CPU-style metrics don't tell the whole picture. A GPU can read 60% busy while the engine's queue is already backing up,

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter

Model ReleasesDGX agent

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabili

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds m…

Model ReleasesDGX agent

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds more or less the same kind of AI, there is no moat, and profi

They knew the charges were bullshit from the start. They knew they arrested an innocent man just to intimidate others and appease the fake n…

Model ReleasesDGX agent

They knew the charges were bullshit from the start. They knew they arrested an innocent man just to intimidate others and appease the fake narrative of an incompetent fucking toddler. This is the real

This is both a real incident (in that the AI really did get unauthorized access to real systems) and also something it was (sort of) prompte…

Model ReleasesDGX agent

This is both a real incident (in that the AI really did get unauthorized access to real systems) and also something it was (sort of) prompted to do. In a review of our cybersecurity evaluations, we fo

This is not optional, things are getting chaotic now, and not dealing with this change won't make it go away. Plus, this could be a huge boo…

ApplicationsDGX agent

This is not optional, things are getting chaotic now, and not dealing with this change won't make it go away. Plus, this could be a huge boost for both individual satisfaction & firm performance if do

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alte…

SafetyDGX agent

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alternatives that can, or we are screwed. Anybody remember this

This will happen frequently as AI becomes smarter and more agentic

Model ReleasesDGX agent

This will happen frequently as AI becomes smarter and more agentic In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or wh

Throwback to 2k7 vibez when I opened my first Iphone. Start of the AI Hardware era? Cool 🛳️ @OpenAI

AgentsDGX agent

On July 31, 2026 at 01:40 AM, Twitter user Murtaza Khomusi posted a “throwback” tweet reminiscing about opening his first iPhone in 2007 (“2k7 vibez”) and questioned whether this marked the start of t

Together AI gives developers a high-throughput production path for running Inkling-Small on @NVIDIAAI Accelerated Infrastructure across mult…

AgentsDGX agent

Together AI gives developers a high-throughput production path for running Inkling-Small on @NVIDIAAI Accelerated Infrastructure across multimodal, coding, and agentic workloads. Start building: https

Together AI gives developers a high-throughput production path for running Inkling-Small on NVIDIA Accelerated Infrastructure across multimo…

HardwareDGX agent

Together AI gives developers a high-throughput production path for running Inkling-Small on NVIDIA Accelerated Infrastructure across multimodal, coding, and agentic workloads. Start building: https://

Try Grok 4.5 http://X.ai/cli

Model ReleasesDGX agent

Try Grok 4.5 http://X.ai/cli BREAKING: Grok 4.5 outperforms GPT-5.6 Terra across almost all shared benchmarks on AskClash, leading in ACB, GPQA, SWE-P, and Atlas while also achieving a higher overall

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that comm…

AgentsDGX agent

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that commoncrawl isn't good enough for you, you have to build a Whole

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improve…

Model ReleasesDGX agent

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improvement a concrete, executable testbed. OpenMLE is an open full

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages a…

Model ReleasesDGX agent

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages are full of native text, which can be directly handled with Li

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

Model ReleasesDGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

30 Jul 2026

A post of ours on vendor lock in and LLMs. We don't like the growing trend of AI companies quietly hiding your data while stripping away you…

AgentsDGX agent

A post of ours on vendor lock in and LLMs. We don't like the growing trend of AI companies quietly hiding your data while stripping away your control. We think that's bad for users and bad for the eco

AI engineers are rediscovering ontologies as a way to keep probabilistic agents inside deterministic boundaries. We delve into presentations…

AgentsDGX agent

AI engineers are rediscovering ontologies as a way to keep probabilistic agents inside deterministic boundaries. We delve into presentations from @coyle_frankp and @emileifrem from the @aiDotEngineer

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

Model ReleasesDGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

And Grok 4.6 comes out in a week

Model ReleasesDGX agent

And Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl

Answer from OpenAI

ApplicationsDGX agent

Answer from OpenAI @emollick re ARC-AGI-3: human testers scored ~48% (ARC uploaded testers logs to HuggingFace some time ago, I believe). re GDPval: it’s close to saturated now, so we’re mostly lookin

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672

Model ReleasesDGX agent

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672 We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are

As the benchmarks that test frontier AI on get more complex, we are losing one of the most important aspects of benchmarking: comparisons to…

ApplicationsDGX agent

As the benchmarks that test frontier AI on get more complex, we are losing one of the most important aspects of benchmarking: comparisons to humans Validated benchmarks need to have human (ideally mul

Atomic Chat signed the Open Weights letter! We believe everyone should be able to run AI on their own device. When a model is open, thousand…

SafetyDGX agent

Atomic Chat signed the Open Weights letter! We believe everyone should be able to run AI on their own device. When a model is open, thousands of teams fine-tune it, quantize it and build new tools on

Big unsaturated benchmarks that have this: ARC-AGI, the original GDPval (not GDPval-AA), METR long horizons, ASI cyber tasks, (Speaking of w…

ApplicationsDGX agent

Ethan Mollick highlights that large, currently under‑explored benchmarks (e.g., ARC‑AGI, GDPval, METR long horizons, ASI cyber tasks) are growing in complexity, yet they increasingly lack systematic h

Can AI agents conduct open-ended AI research? Most evaluations of agents conducting AI research focus on narrow, verifiable tasks. But AI re…

HardwareDGX agent

Can AI agents conduct open-ended AI research? Most evaluations of agents conducting AI research focus on narrow, verifiable tasks. But AI research is often open ended. Researchers pick hypotheses, dec

Can an open-source model perform like a foundation model? @Osmosis_AI is betting yes, using reinforcement learning and the dedicated @ycombi…

HardwareDGX agent

Osmosis_AI claims that an open‑source model can rival a foundation model by leveraging reinforcement learning techniques. To demonstrate this, they will use the Y Combinator‑dedicated GPU cluster on T

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep t…

SafetyDGX agent

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep their infrastructure secure. Everybody deserves access to mod

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyU…

Model ReleasesDGX agent

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyUI - combining Florence 2, SAM2, and WAN Video into a single

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success …

Model ReleasesDGX agent

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success rate. - Unoptimized? - Buggy? - Deliberate (coz that helps i

Everyone is a real-world AI engineer. They don’t just make cars, but autonomous robots with frontier capabilities. Thanks for their hard wor…

AgentsDGX agent

Everyone is a real-world AI engineer. They don’t just make cars, but autonomous robots with frontier capabilities. Thanks for their hard work keeping the line smooth and running. It’s my pleasure work

Excited to be on the CNBC live show!

ToolsDGX agent

Excited to be on the CNBC live show! Back from vacation and LIVE at 12pm PT / 3pm ET Is AI’s easy-money era ending? We’ll unpack a wild week for the AI trade—big tech earnings, Leopold Aschenbrenner’s

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3.

Local AiDGX agent

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3. Built to bring open-source LLMs to private machines, @Ollama uses Intel Core Ultra Series 3 to run

Finally a good paper testing if file-system based memory for LLM agents is worth it. First, what does this look like? Deployed agents keep l…

AgentsDGX agent

Finally a good paper testing if file-system based memory for LLM agents is worth it. First, what does this look like? Deployed agents keep long-term memory as a folder of markdown files they read and

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making th…

Model ReleasesDGX agent

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @Goog

goblin-level blog post

Model ReleasesDGX agent

goblin-level blog post Turns out GPT-5.6 Sol is actually SoTA on ARC-AGI-3. Just took two setting changes. You just have to allow it to reason and work over multiple context windows with the help of o

good job little bro

Model ReleasesDGX agent

good job little bro After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. The results: - 20% lower serving costs from production GPU ke

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of doll…

Model ReleasesDGX agent

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of dollars a month in savings at this point? Codex analysed product

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep,…

Model ReleasesDGX agent

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep, I wanted to quickly jot down my thinking here. The basic is

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve w…

ApplicationsDGX agent

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve worked on many of the foundational technologies that power mo

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥

ResearchDGX agent

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥 Today, we are releasing Inkling-Small. Inkling-Small achieves comparable performance to In

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also ac…

AgentsDGX agent

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt cache hit rates. Low prici

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. …

AgentsDGX agent

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. That is especially true for document parsing 📄🔀: - Some page

It's wild to me that both Anthropic and OpenAI have products that lean so hard on search, and yet they both obscure the underlying search in…

ToolsDGX agent

Simon Willison notes that both Anthropic and OpenAI build products that rely heavily on searching data but keep the underlying search index they use hidden from public view. He finds it surprising how

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/207051816739969…

Model ReleasesDGX agent

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/2070518167399698490?s=20 'I tried to analyze why Claude Code uses more toke

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of…

Model ReleasesDGX agent

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we ar

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most pars…

Model ReleasesDGX agent

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most parsing pipelines throw all of them at the same parser, forcing

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

SafetyDGX agent

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

OpenAI had a partnership with Bing, but also run their own crawling and indexing infrastructure. Does ChatGPT decide any Bing at all these d…

ToolsDGX agent

OpenAI has collaborated with Microsoft’s Bing while also running its own web‑crawling and indexing systems, and Anthropic similarly relies on search‑derived data. Both firms prominently incorporate se

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large…

Model ReleasesDGX agent

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large enterprise codebases. We helped them to scale the training

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the b…

Model ReleasesDGX agent

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the benchmark or that contain knowledge about the benchmark forma

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from …

SafetyDGX agent

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from investors after suffering heavy losses during the recent rou

Starting in 15 minutes

ApplicationsDGX agent

Starting in 15 minutes Kimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it. He joins Jue Wang and Zain Hasan from Together A

sure we lose money on every inference but we make it up in volume

Model ReleasesDGX agent

sure we lose money on every inference but we make it up in volume We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices f

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that lat…

SafetyDGX agent

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that latency by predicting and pre-executing the next call, but exte

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, man…

SafetyDGX agent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, many forefront members of the AI-Safety community, in their fer

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many b…

Model ReleasesDGX agent

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many benchmarks. To test its speed, we plugged it into HF's speech

← Previous
1…7891011…232
Next →