AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
86,428 results
14 Jul 2026

Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calibration, self-verifica…

TutorialsDGX agent

Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calibration, self-verification, knowing when to stop, and knowing what you do not know

How Retail Finance teams are using Agentic AI to protect omni-channel margins

AgentsDGX agent

Retail Finance teams are increasingly deploying agentic AI to safeguard omni‑channel margins. These systems automatically monitor sales performance, forecast demand, and adjust pricing or inventory co

How to measure AI productivity: From LLM token costs to business value with Arize AX

TutorialsDGX agent

AI productivity is best measured by connecting AI usage to validated downstream outcomes. Tokens, prompts, and generated lines show activity, but they do not prove value. A better measurement model tr

Content type
AllBlogX PostPaperYouTubeRedditGitHub

How to Run an Autoresearch Workflow with RL Agent Skills and NVIDIA NeMo

HardwareDGX agent

Autonomous coding agents such as Codex (GPT‑5.5) can fully automate reinforcement‑learning research workflows by provisioning GPU‑hosted environments, orchestrating experiments, and iteratively optimi

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1…

Local AiDGX agent

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1-bit and 134 tok/s in Ternary on an NVIDIA GeForce RTX 5090.

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told yo…

Model ReleasesDGX agent

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told you so!' - The mass majority of researchers and academics in d

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then…

Model ReleasesDGX agent

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then post-train it into specialized intelligence you own. https:

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use …

Model ReleasesDGX agent

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use the OpenAI everything app, you pick between ChatGPT Work & C

I'll open source this if it's interesting! But here's my first public artifact, a breakdown on my Mega Sceptile team: https://claude.ai/code…

Model ReleasesDGX agent

Thariq (@trq212) released his first public artifact on Claude.ai, a breakdown of his Mega Sceptile team titled “Mega Sceptile — Champions Field Guide.” The guide covers the team’s build, game plan, de

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There w…

Model ReleasesDGX agent

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There was a fierce debate at the time about vibe coding and the arg

Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.

Model ReleasesDGX agent

dam bois we eating good this week ngl, The velocity of the open_weight ecosystem right now is hitting a point where proprietary, closed-source APIs are losing their leverage on compute intelligence. W

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan o…

AgentsDGX agent

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan of Fireworks - how open-source agents can use a frontier mode

Lessons From the Leaderboard: What 5,000+ Kagglers Taught Us About Improving AI Reasoning

Model ReleasesDGX agent

The NVIDIA Nemotron Model Reasoning Challenge on Kaggle attracted over 5,000 participants who all began from the same open model, benchmark, and infrastructure. The strongest solutions treated reasoni

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions y…

ApplicationsDGX agent

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions you can connect into via Plaid: - American Express - Bank of A

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

SafetyDGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize…

Local AiDGX agent

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize them (Fine-tuning / RL / Post-training) are invaluable skil

🤗 MOSS-VL-Realtime is now open source on @huggingface . Built for real-time visual understanding over continuous video streams: 🧠 11B visi…

Model ReleasesDGX agent

🤗 MOSS-VL-Realtime is now open source on @huggingface . Built for real-time visual understanding over continuous video streams: 🧠 11B vision-language model 📜 Apache-2.0 license 💬 Ask questions at any

Multi-agent social intelligence with Strands Agents and Amazon Bedrock

AgentsDGX agent

This post shows how Thrad.ai deployed a multi-agent system with Strands Agents and Amazon Bedrock AgentCore that automates the pipeline from prospect discovery through personalized email generation. T

Multilingual Semantic Retrieval for Apple Music Search

Model ReleasesDGX agent

Apple Music serves listeners across 150+ storefronts in dozens of languages, with a catalog that grows by hundreds of thousands of new tracks daily. At this scale, search recall on misspelled, transli

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are ask…

Model ReleasesDGX agent

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are asking for trouble GPT-5.6 Sol just deleted my whole production

Nemotron Labs: How Open Models Give Enterprises and Nations AI They Can Trust, Control and Customize

Model ReleasesDGX agent

Enterprises have plenty of powerful models to choose from. The real test is whether the AI an enterprise builds uniquely addresses the needs of the business: improving workflows, tapping into domain k

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

TutorialsDGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

SafetyDGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

🔗 Official Ollama extension https://marketplace.visualstudio.com/items?itemName=Ollama.ollama

Local AiDGX agent

**Official Ollama Extension for VS Code** The extension on the Visual Studio Marketplace provides a dedicated Ollama language‑model integration for Visual Studio Code, enabling developers to run and a

OpenAI may announce a ChatGPT smart speaker this year

IndustryDGX agent

OpenAI's first device is set to be a smart speaker that lets you talk with ChatGPT, according to a report from Bloomberg. The device apparently won't have a screen, but will use a camera and additiona

Oracle opens Fusion Agentic Applications to pro-code developers and coding agents

AgentsDGX agent

Oracle Corp. today announced a new artificial intelligence-native experience for its Fusion Applications within Oracle AI Agent Studio, allowing developers, customers and partners to build agents toge

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

ToolsDGX agent

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills

HardwareDGX agent

NVIDIA Cosmos 3 was post‑trained in under a day using TAO agent skills and LoRA adapters, raising accuracy on the Woven Traffic Safety video QA dataset from 54.41 % to 93.35 %. The mixture‑of‑transfor

Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants

AgentsDGX agent

Proactive agents that anticipate user needs and autonomously execute tasks hold great promise as digital assistants, yet the lack of realistic user simulation frameworks hinders their development. Exi

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌

Model ReleasesDGX agent

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌 U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama a

Quoting Armin Ronacher

AgentsDGX agent

The shared language of a software project is not English or Python but it is the common understanding of what its concepts mean, where the boundaries are, which invariants matter, who owns what, and w

Scaling medical content review at Flo Health with Amazon Bedrock – Part 2

ApplicationsDGX agent

In this post, we share how Flo Health’s engineering team turned a proof of concept (PoC) from the AWS Generative AI Innovation Center into a production-grade, AI-powered medical content review and gen

Scaling UX testing with Amazon Nova Act: A new approach to user flow analysis

TutorialsDGX agent

Using generative AI enables parallel execution of comprehensive user flow testing at scale. This solution demonstrates how to build a cloud-deployed UX testing platform that automatically generates te

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

SafetyDGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

simonw/pedalican

Model ReleasesDGX agent

simonw/pedalican Clearly I wasn't paying attention when these were first announced back in May, but today I accidentally activated a 'pet' in Codex Desktop - a little animated robot, reminiscent of Cl

some good self improvement research here

AgentsDGX agent

some good self improvement research here The first experimental evidence of recursive self-improvement (RSI). Autoresearching the autoresearch agent for eight days. The result beats the harness we han

Sovereign AI infrastructure startup Valarian raises $50M to help nations secure their defense systems

ApplicationsDGX agent

The defense-focused artificial intelligence systems infrastructure startup Valarian Technologies Ltd. said today it has closed on a hefty 50 million early-stage funding round as it looks to position i

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

AgentsDGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

Text match filters for agents

ToolsDGX agent

Text match filters are a feature in Pinecone that allow users to filter vector search results based on exact text matching criteria, enabling more precise control over which documents or records are r

The Download: Claude’s inner workings, and the future of world models

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. What Anthropic’s latest AI discovery does—and doesn’t—show —Ja

'The paper's insight connects to a broader pattern: AI agents are essentially distributed systems with unreliable components (the LLM), and …

AgentsDGX agent

'The paper's insight connects to a broader pattern: AI agents are essentially distributed systems with unreliable components (the LLM), and we should apply distributed systems patterns to them.' https

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai fam…

Local AiDGX agent

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai family. Based on Qwen3.6 27B, it brings a new capability tier t

Together AI positions open-weight AI models as the enterprise moat for cost, control and IP

AgentsDGX agent

Enterprises racing to deploy AI at scale are discovering that the biggest constraint isn’t model capability anymore — it’s control. As agentic AI moves from experimentation into core business processe

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here

Model ReleasesDGX agent

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here Did... Codex just overtake Claude Code? 24.5 hours ago Tibo an

🤖 Un robot français en seulement 9 mois ! Rémi Cadene, ancien de Tesla AI et Hugging Face, explique comment son équipe veut faire des robot…

ResearchDGX agent

🤖 Un robot français en seulement 9 mois ! Rémi Cadene, ancien de Tesla AI et Hugging Face, explique comment son équipe veut faire des robots un levier pour la logistique, la réindustrialisation et, de

U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama and unlocking complex, longer …

Model ReleasesDGX agent

U.S. open‑source AI models are rapidly gaining popularity, with NVIDIA’s newest model, **Nemotron Ultra**, becoming a prominent entry on the Ollama platform. On Ollama, Nemotron Ultra is quickly scali

v0.32.0

Model ReleasesDGX agent

What's Changed New interactive agent experience: running ollama now launches an agent to help you code and delegate work ❯ ollama Ollama 0.32.0 ▸ Chat, Code, & Work (glm-5.2:cloud) Chat with models, c

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eva…

AgentsDGX agent

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eval signal that reveals whether an agent fails, and where. The

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexi…

Model ReleasesDGX agent

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexity Computer. https://research.perplexity.ai/articles/wandr-b

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 …

Model ReleasesDGX agent

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 tok/s • ⁠Qwen 3.6 27b NVFP4 — 256k ctx, 33 tok/s 2× DGX Spar

Why Performance per Watt Is the Ultimate Metric for AI Infrastructure Efficiency

AgentsDGX agent

Power is AI infrastructure’s inescapable constraint. How many tokens an AI factory can generate within a fixed power budget determines its revenue and profitability. Because of this, performance per w

You can now access and activate your banked resets in Hermes Agent directly with /usage reset when using a codex/openai subscription

AgentsDGX agent

You can now access and activate your banked resets in Hermes Agent directly with /usage reset when using a codex/openai subscription Thank you to the 7M active users who are now using Codex and ChatGP

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build com…

Model ReleasesDGX agent

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build compositions, set keyframes, write expressions, and run the rep

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk …

Model ReleasesDGX agent

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk to anything with a phone number. Just one prompt. 15 seconds

Your extraction schema is now a conversation away. ⁣ 📄 Upload a doc — the agent drafts the schema from it⁣ 💬 Want changes? Just ask⁣ 📚 Or…

AgentsDGX agent

Your extraction schema is now a conversation away. ⁣ 📄 Upload a doc — the agent drafts the schema from it⁣ 💬 Want changes? Just ask⁣ 📚 Or grab a template and go⁣ ⁣ Writing JSON Schema by hand? That's

13 Jul 2026

// An Anatomy of CLI Coding Agent Trajectories // (bookmark it) When your coding agent fails a task, when did the run actually go wrong? Mos…

AgentsDGX agent

// An Anatomy of CLI Coding Agent Trajectories // (bookmark it) When your coding agent fails a task, when did the run actually go wrong? Most reliability studies use the final label to answer this. Th

Big unlock for open-source AI inference: Hugging Face Transformers models can now run in vLLM at native speed, often matching or beating han…

ApplicationsDGX agent

Big unlock for open-source AI inference: Hugging Face Transformers models can now run in vLLM at native speed, often matching or beating hand-written implementations. Until now, every new architecture

Building an agentic AI solution at Bluesight with Amazon Bedrock

AgentsDGX agent

In this post, we describe how Bluesight used two AWS engagements and Amazon Bedrock AgentCore to evolve from a single-product AI prototype to Prism, a unified agentic AI solution spanning six healthca

Building the AI-defined vehicle with Android, Google Cloud, and Nexus SDV

Model ReleasesDGX agent

The automotive industry is moving from building hardware-centric platforms toward building their own sophisticated Software-Defined Vehicle (SDV) architectures. For OEMs, a vehicle is no longer just a

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen …

Model ReleasesDGX agent

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen 4 MiMo 3 Never in the history of LLMs has the frontier been

← Previous
1…282283284285286…1441
Next →