AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,897 results
4 Aug 2026

Building agents that patch other agents. This is an interesting approach for self-improving agents that leverages agent outputs. If you run …

AgentsDGX agent

Building agents that patch other agents. This is an interesting approach for self-improving agents that leverages agent outputs. If you run agents in production, you already have the training data for

can anyone with sufficient training in mathematics refuse this provocation? no ad hominem please; I have seen enough of that and am looking …

SafetyDGX agent

can anyone with sufficient training in mathematics refuse this provocation? no ad hominem please; I have seen enough of that and am looking for real counterarguments, only. 🙏 OpenAI and Anthropic each

can anyone with sufficient training in mathematics refute this provocation? no ad hominem please; I have seen enough of that and am looking …

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

can anyone with sufficient training in mathematics refute this provocation? no ad hominem please; I have seen enough of that and am looking for real counterarguments, only. 🙏 OpenAI and Anthropic each

Coming this weekend: What chess teaches us about Generative AI and its limitations Image with impossible position and bishops of two sizes p…

SafetyDGX agent

Gary Marcus tweeted on August 4, 2026 at 3:53 PM that a forthcoming weekend presentation will explore how chess illustrates the limitations of generative AI systems—highlighting an image produced by C

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…

Model ReleasesDGX agent

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bo…

HardwareDGX agent

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bottlenecked GPU kernels, (still lacks reliable judgment) grea

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down…

Model ReleasesDGX agent

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down after Sydney & got Copilot to market quickly (the 1st profe

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

Model ReleasesDGX agent

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

🛡️Introducing Shieldstral, Mistral’s 3B open-weights model for content safety that can be deployed on-device 🧵 http://mistral.ai/news/shie…

Model ReleasesDGX agent

Mistral AI introduced Shieldstral, a 3‑billion‑parameter, open‑weights model designed for content‑safety tasks and capable of on‑device deployment. The announcement was shared via a tweet from @Mistra

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

Model ReleasesDGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

Onboard Satellite Image Classification for Earth Observation: A Comparative Study of ViT Models

ResearchDGX agent

arXiv:2409.03901v4 Announce Type: replace Abstract: Remote sensing (RS) image classification is central to Earth observation, but onboard deployment requires models that are accurate, efficient, and r

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more …

Model ReleasesDGX agent

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more observable, easier to steer, easier to understand we want to

OpenAI and Anthropic have both just posted about an overlapping cyber incident involving GPT-5.6-Sol and Mythos 5 during an evaluation by UK…

Model ReleasesDGX agent

OpenAI and Anthropic have both just posted about an overlapping cyber incident involving GPT-5.6-Sol and Mythos 5 during an evaluation by UKAISI. I will quote: 'In the most serious case, an agent trie

Parsing a W-2 into markdown was always the easy part. Getting the fields out was a second pipeline: define a schema, map the fields, handle …

AgentsDGX agent

Parsing a W-2 into markdown was always the easy part. Getting the fields out was a second pipeline: define a schema, map the fields, handle the edge cases. Set processing_options.forms='𝗲𝗻𝗿𝗶𝗰𝗵', and L

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now mo…

Model ReleasesDGX agent

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now move it between two agent harnesses and the cost per success c

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update

Model ReleasesDGX agent

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update 📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3

Qwen-Image-3.0-Pro has made a massive leap from the previous generation, now ranking #5 globally. Appreciate the recognition! We will keep b…

Model ReleasesDGX agent

Qwen-Image-3.0-Pro has made a massive leap from the previous generation, now ranking #5 globally. Appreciate the recognition! We will keep building.🚀 More exciting news from @Alibaba_Qwen: Qwen-Image-

Qwen3.8-Max, better and cheaper. Try it out. 👀

Model ReleasesDGX agent

Qwen3.8-Max, better and cheaper. Try it out. 👀 We ran a test between the new Qwen3.8-Max, Opus 5 and GPT-5.6 Sol. 3 models. same prompt. one-shot with the /design command. Reviewed gameplay features,

Qwen3.8-Max is available in Hermes Agent now! Let's build! 🚀🚀

Model ReleasesDGX agent

Qwen 3.8‑Max, Alibaba.Qwen’s latest large‑language model, has been added to Hermes Agent and can currently be accessed at a 20 % discount. The update aims to streamline integration of the model for de

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from…

Model ReleasesDGX agent

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from RTX 4090 or equivalent). New proprietary non-decoder-only a

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability de…

Model ReleasesDGX agent

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability detection, 3x at patching K3 is SOTA cyber-defense model you c

Routing for long-horizon coding agents is a big deal. @notdiamond_ai just announced a model router that works natively with Claude Code. Thi…

Model ReleasesDGX agent

Routing for long-horizon coding agents is a big deal. @notdiamond_ai just announced a model router that works natively with Claude Code. This is huge. It picks the model and reasoning effort before ea

Running fast is not enough, you need fast AND correct An excellent addition from @ArtificialAnlys to make sure that the flashy speed numbers…

Model ReleasesDGX agent

Running fast is not enough, you need fast AND correct An excellent addition from @ArtificialAnlys to make sure that the flashy speed numbers are backed by 100% matching accuracy Announcing the Artific

Shieldstral is available under Apache 2.0. Try it: https://huggingface.co/mistralai/Shieldstral-1.0

Model ReleasesDGX agent

**Shieldstral** is a 3‑billion‑parameter open‑weights model developed by Mistral AI for content safety. It can be deployed on-device and is distributed under the Apache 2.0 license. The model is avail

sweet, i finally set up my chatgpt, codex, claude, & claude code to have a shared memory - so i can ask one of them what i've worked on rece…

Model ReleasesDGX agent

sweet, i finally set up my chatgpt, codex, claude, & claude code to have a shared memory - so i can ask one of them what i've worked on recently across all four of them for codex/cc, i'm using local h

// The confidence cliff in self-improving autoresearch // Autoresearch loops are still quite brittle. Here is a nice paper offering some ins…

Model ReleasesDGX agent

// The confidence cliff in self-improving autoresearch // Autoresearch loops are still quite brittle. Here is a nice paper offering some insights into why this might be happening. Self-improving autor

The model takes moderation policy as a plain-language question and returns a calibrated score. Text and images — one interface. Read the ful…

SafetyDGX agent

Mistral AI has released Shieldstral, a 3‑billion‑parameter open‑weight model for content safety that can run locally on device. It interprets plain‑language moderation queries and returns calibrated s

this is a really clean and unique ai tool that acts more like an extension of you took me 15 min to set up a custom ai voice assistant w/ it…

AgentsDGX agent

this is a really clean and unique ai tool that acts more like an extension of you took me 15 min to set up a custom ai voice assistant w/ it’s own # that also picks up my calls when my phone is off an

this post from 3 years ago and commercial LLMs *still* can’t play chess anywhere near as well serious players (except by calling external to…

Model ReleasesDGX agent

this post from 3 years ago and commercial LLMs *still* can’t play chess anywhere near as well serious players (except by calling external tools) When @GaryMarcus and others point out that GPT-4 is bad

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing th…

IndustryDGX agent

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing the paradigm… describe the molecule you want, and the model de

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and …

SafetyDGX agent

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. I

Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads.…

Model ReleasesDGX agent

Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads. Start building: https://www.together.ai/models/deepseek-v4-

We were proud to host “Build with Frontier Intelligence,” http://Z.ai ’s first community meetup, together with @AISingapore, at Tencent’s ve…

AgentsDGX agent

We were proud to host “Build with Frontier Intelligence,” http://Z.ai ’s first community meetup, together with @AISingapore, at Tencent’s venue. During the event, http://Z.ai shared insights into the

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline w…

Model ReleasesDGX agent

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline what happened, how the activity was contained, and how we’re

3 Aug 2026

8月より、Sakana AI @SakanaAILabs にApplied Research Engineer Internship として入社しました 🐟🐠🐡! 大学の夏季休業期間にフルタイム勤務予定です! LLMの研究開発と社会実装を頑張ります💪🏻

ResearchDGX agent

Horiyuki 'horiyuki42' joined Sakana AI Labs in August 2026 as an Applied Research Engineer Intern. He plans to work full‑time during the university summer recess while concentrating on large‑language‑

Advanced reasoning and problem-solving, plus strong performance in Japanese language and Japan-specific context. Namazu is live on Merge Gat…

ResearchDGX agent

Advanced reasoning and problem-solving, plus strong performance in Japanese language and Japan-specific context. Namazu is live on Merge Gateway now! https://x.com/SakanaAILabs/status/2084276852143919

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorp…

Model ReleasesDGX agent

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorporate symbol/manipulating entities like harnesses, tools, an

All Chinese AI company media: AI will be a wondrous and wonderful and give you back precious time in your life to do things you love. All US…

ResearchDGX agent

All Chinese AI company media: AI will be a wondrous and wonderful and give you back precious time in your life to do things you love. All US AI company media: AI will take all your jobs, eat your chil

All frontier AI models (including GPT-5.6 Sol, Fable 5, and Kimi K3) consistently fail to implement solvers for basic nonlinear PDEs correct…

Model ReleasesDGX agent

All frontier AI models (including GPT-5.6 Sol, Fable 5, and Kimi K3) consistently fail to implement solvers for basic nonlinear PDEs correctly, often introducing significant errors in numerical stabil

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer s…

Model ReleasesDGX agent

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol

Appreciate it! Now let's understand the world through the eyes of Qwen3.8. 🥳

Model ReleasesDGX agent

Qwen3.8-Max from Alibaba’s Qwen team achieved second place in the Vision Arena benchmark, scoring 1,305 points. It trails only Claude Fable 5 (High), which leads by a slim 13‑point margin. The post un

🚨ASI/AGI is imminent fans don’t realize that they ALREADY lost the argument around Astra. My argument (spelled out in detail in my Substack…

Model ReleasesDGX agent

🚨ASI/AGI is imminent fans don’t realize that they ALREADY lost the argument around Astra. My argument (spelled out in detail in my Substack today) was that Astra was unlikely to be the dramatic leap f

Ask anything, anonymously. Qwen3.8-Max has landed on Venice. Give it a try!

Model ReleasesDGX agent

Qwen from Alibaba has released the Qwen 3.8‑Max model on the Venice platform, enabling users to ask questions anonymously. The announcement encourages users to try the new functionality immediately. T

Audio moves through a dedicated fast path, while deeper reasoning and tool use happen asynchronously. We also reduced voice-session startup …

AgentsDGX agent

OpenAI’s GPT‑Live can listen and speak simultaneously. The new voice stack routes audio through a dedicated fast path so that speech flows continuously while intensive reasoning and external tool call

'Engineers make mistakes, and I think this is what happened here.' @huggingface CEO @ClementDelangue discusses what it'll take to prevent fu…

IndustryDGX agent

'Engineers make mistakes, and I think this is what happened here.' @huggingface CEO @ClementDelangue discusses what it'll take to prevent future rogue AI cyberattacks moving forward: https://www.cnbc.

even within math, frontier models are far from being general intelligences. and we still need people to assess which of their output are tru…

Model ReleasesDGX agent

even within math, frontier models are far from being general intelligences. and we still need people to assess which of their output are trustworthy and which are not. All frontier AI models (includin

fascinating comments on possible slowdown – but probably a red herring. - Anthropic and OpenAI presumably aren’t colluding to slowdown, they…

SafetyDGX agent

fascinating comments on possible slowdown – but probably a red herring. - Anthropic and OpenAI presumably aren’t colluding to slowdown, they are asking governments to build a framework such that a slo

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing inter…

AgentsDGX agent

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing interactions, writing records, and reranking retrievals. Every on

future generations will have no idea what was real and what was not.

SafetyDGX agent

Gary Marcus warns that, with increasingly sophisticated generative‑AI tools, future generations may struggle to tell what is real from what is fabricated. Matt Stoller adds that generative AI will tra

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on …

AgentsDGX agent

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on @GaryMarcus for being anti-AI, but he's actually bullish on

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman…

AgentsDGX agent

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman-preserve-records-ai-agent-hacking-probe?%3Fintcmp=tw_fbn&ta

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new a…

AgentsDGX agent

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new architecture keeps audio flowing continuously, so deeper reas

Hugging Face CEO Clément Delangue weighs in on the open-weight AI models debate and said it helped the company when it was hacked by an Open…

IndustryDGX agent

Clément Delangue, CEO of Hugging Face, said the company’s focus on open‑weight AI models helped it withstand a hacking incident involving an unreleased OpenAI model. He also noted that China is leadin

Hugging Face CEO says China is winning the AI race and dominating on open models https://www.cnbc.com/2026/08/03/hugging-face-china-ai-race-…

IndustryDGX agent

Hugging Face CEO says China is winning the AI race and dominating on open models https://www.cnbc.com/2026/08/03/hugging-face-china-ai-race-open-models.html?taid=6a70b9a7142a0e000189af30&utm_campaign=

.@huggingface CEO @ClementDelangue says a hack by OpenAI could have been 'way worse' if not for the company's defensive measures. He speaks …

IndustryDGX agent

On August 3, 2026, Hugging Face CEO Clement Delangue appeared on Bloomberg TV to discuss a hypothetical security incident involving OpenAI. He stated that the hack could have been “way worse” had Defi

.@huggingface CEO @ClementDelangue says the recent OpenAI-linked cyberattack highlights the growing risks posed by rapidly-evolving AI model…

IndustryDGX agent

.@huggingface CEO @ClementDelangue says the recent OpenAI-linked cyberattack highlights the growing risks posed by rapidly-evolving AI models. Speaking to @edludlow, Delangue says that open-source AI

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them.

AgentsDGX agent

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them. @gabriberton Hmm, @ylecun always clarified he was talking about autor

Introducing Sakana Namazu: An LLM API with Japanese-vibes! Today we’re rolling out an upgraded version of our LLM, Namazu, now available as …

ResearchDGX agent

Introducing Sakana Namazu: An LLM API with Japanese-vibes! Today we’re rolling out an upgraded version of our LLM, Namazu, now available as an API. Try Sakana Namazu API: https://sakana.ai/namazu 🐟 🐟

Is China winning the AI race? @huggingface CEO @ClementDelangue thinks so – here's how he's utilizing foreign cybersecurity tools following …

IndustryDGX agent

Is China winning the AI race? @huggingface CEO @ClementDelangue thinks so – here's how he's utilizing foreign cybersecurity tools following the company's hack by rogue OpenAI agents: https://www.cnbc.

← Previous
1…45678…232
Next →