AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “openai”

GridTimelineEvolution
2,581 results
Model Releases

An Expectation-Maximization Perspective on Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2504.18587v2 Announce Type: replace-cross Abstract: Reinforcement learning has emerged as a powerful approach for improving the reasoning capabilities of large language models, as demonstrated b

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

AquiLLM: An Architecture for Supporting Tacit Knowledge Capture in Research Groups

DGX agent

arXiv:2608.08883v1 Announce Type: new Abstract: Recent advances in retrieval-augmented generation (RAG) and large language models (LLMs) enable researchers to integrate AI into scientific workflows. H

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

ChatGPT and Gemini both just passed 1 billion users

DGX agent

For the 14th time, a Google product has hit 1 billion users. Google CEO Sundar Pichai posted on X that a billion people are using Gemini every month, and that Gemini is Google's fastest-growing produc

model-releasesthe-verge-ai
11 Aug 2026
Model Releases

Curriculum Generation under Structured Parametric Environments for Robust Navigation Policies

DGX agent

arXiv:2608.08545v1 Announce Type: cross Abstract: Robust navigation policies for autonomous agents must generalize across continuously varying environmental conditions such as turn rates, obstacles, f

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

DevIntent: How Much Does LLM-Generated Code Violate Developer Intent?

DGX agent

arXiv:2608.07614v1 Announce Type: cross Abstract: Code generated by LLMs can violate a developer's implicit intentions when given an ambiguous prompt, yet standard benchmarks measure only whether code

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

I built a weird, low-power llama.cpp server using an Intel N100 + RTX 5060Ti

DGX agent

Everything started with the sudden death of my old ASRock J1900. While looking for the perfect ITX replacement, I stumbled upon the Chinese CW-NAS-ADLN-K motherboard, which looked perfect on paper: In

model-releasesr-localllama
11 Aug 2026
Industry

I will pay 1 million for a developer to build this

DGX agent

I WILL PAY 1 MILLION FOR THIS 😭 I will literally pay 1,000,000 to the developer who fixes this. ChatGPT macOS: ⌘C text → ⌘Tab back to ChatGPT → ⌘V …and the search box ISN’T FOCUSED. 🤦‍♂️ Why do I have

industryr-chatgpt
11 Aug 2026
Hardware

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sa…

DGX agent

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sachs, and KKR, to raise over $500 billion so its own customers

hardwaregary-marcus--x
11 Aug 2026
Model Releases

Nvidia releases Nemotron 3.5 Lightning and NeMo Switchyard to give enterprise AI capability options

DGX agent

Artificial intelligence silicon and software giant Nvidia Corp. today announced two new services: a highly customizable Nemotron model and an agentic AI model router named NeMo Switchyard. As enterpri

model-releasessiliconangle
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

The small open weight models are scarier in AI development

DGX agent

Imagine if your everyday laptop could run an AI model smart enough to take care of 90% of your work—totally private, lightning fast, and completely free of monthly fees. That is the exact tipping poin

model-releasesr-localllama
11 Aug 2026
Local Ai

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

DGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Blast Radius

DGX agent

arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictive memory management layer that estimates

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation

DGX agent

arXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Grammar Engineering Meets LLMs: Development of Cantonese and Irish ParGram Treebanks

DGX agent

arXiv:2608.07283v1 Announce Type: new Abstract: Grammar engineering requires expertise in linguistic formalism and computational implementation, especially in parallel grammar projects that balance cr

researcharxiv-cs-cl
10 Aug 2026
Safety

LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents

DGX agent

arXiv:2608.06948v1 Announce Type: new Abstract: AI models are becoming increasingly adept at understanding and processing spatial information, thereby facilitating agentic problem-solving in spatial t

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs

DGX agent

arXiv:2608.07167v1 Announce Type: new Abstract: Giving an AI agent the ability to send emails, query databases, or execute commands is useful--until the agent is tricked into doing something it should

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Social World Models

DGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…

DGX agent

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting

safetyyann-lecun--x
10 Aug 2026
Hardware

Ultra-High Interactivity on NVIDIA GPUs? - TileRT InferenceX

DGX agent

The TileRT InferenceX article (Aug 10 2026) examines whether the TileRT software stack on NVIDIA GPUs can compete with dedicated inference systems such as Cerebras, Groq LPUs and SambaNova for ultra‑h

hardwaresemianalysis
10 Aug 2026
Model Releases

A failure of ChatGPT Work & Claude Cowork is they assume that non-coders couldn't understand how to think about problems like a coder, so th…

DGX agent

A failure of ChatGPT Work & Claude Cowork is they assume that non-coders couldn't understand how to think about problems like a coder, so they hide all that stuff. They should instead explain choices

model-releasesethan-mollick--x
9 Aug 2026
Model Releases

GitHub Models is now retired

DGX agent

GitHub Models is now retired I missed this news until today, when the GitHub Actions run for my simonw/research repository failed with this error message: GitHub Models is temporarily unavailable as p

model-releasessimon-willison
9 Aug 2026
Local Ai

Building a zero-dependency C inference engine for BitNet (1.58-bit) - lessons from hitting 36 tok/s on a Xeon CPU

DGX agent

Over the past few months I have been building a CPU-first inference engine from scratch in pure C99 (no Python, no CUDA, no BLAS, just GCC and make). The focus has been running 1.58-bit ternary models

local-air-localllama
8 Aug 2026
Tools

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…

DGX agent

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list

toolssimon-willison--x
8 Aug 2026
Industry

Every CIO should watch this. Persistent systems trying to break through and solve problems at all costs are more creative than you think. La…

DGX agent

Every CIO should watch this. Persistent systems trying to break through and solve problems at all costs are more creative than you think. Labs and enterprises will be focused more on network effects (

industryallie-k--miller--x
7 Aug 2026
Model Releases

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... …

DGX agent

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... ... one that calls for more open intelligence, not less. The

model-releasesfireworks-ai--x
7 Aug 2026
Model Releases

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)

DGX agent

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game

model-releasessimon-willison
7 Aug 2026
Research

ProDVI: Programmatic Dynamics Priors for Value Network Initialization

DGX agent

arXiv:2608.06015v1 Announce Type: cross Abstract: Deep Reinforcement Learning (RL) is notoriously sample inefficient. One contributing factor is that RL agents are typically initialized from scratch,

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Serving Deepseek v4 Flash 0731 on 2x DGX Spark — 5-7 GB OS headroom, what would you do to lower VRAM usage and increase OS available RAM?

DGX agent

Hey all, I'm serving DSv4Flash 0731 on a cluster of 2x DGX Sparks but am running into constant issues with having almost no RAM (unified memory) left for the OS/cache and I'd love to hear the communit

model-releasesr-localllama
7 Aug 2026
Industry

We didn’t give it access to the internet, but it found a way. We forgot to give it the Google Sheets permission, but it found a way. It acci…

DGX agent

We didn’t give it access to the internet, but it found a way. We forgot to give it the Google Sheets permission, but it found a way. It accidentally had edit access to the artifactory so we removed ed

industryallie-k--miller--x
7 Aug 2026
Model Releases

What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)

DGX agent

arXiv:2608.06202v1 Announce Type: cross Abstract: Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet m

model-releasesarxiv-cs-ai
7 Aug 2026
Industry

What’s behind the Google AI shake-up

DGX agent

Some of the biggest names on Google's AI team got new jobs this week. In some cases, including for legendary Googler Jeff Dean, those jobs are no longer at Google. Given that Google's models seem to b

industrythe-verge-ai
7 Aug 2026
Model Releases

Anthropic will design its own hardware to power Claude

DGX agent

Anthropic is hiring a custom silicon team to design proprietary chips that will power its Claude models, while still planning a multi‑chip strategy that mixes internally designed hardware with compone

model-releasesars-technica
6 Aug 2026
Model Releases

Can Post-Training Transform LLMs into Causal Reasoners?

DGX agent

arXiv:2602.06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts. While large language models (LLMs) show promise in

model-releasesarxiv-cs-ai
6 Aug 2026
Hardware

Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs (Anna Tong/Forbes)

DGX agent

Anna Tong / Forbes: Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs — The same Silicon Va

hardwaretechmeme
6 Aug 2026
Model Releases

General Availability of Pinecone Nexus Proves Knowledge Drives Real Outcomes for Agentic AI

DGX agent

Pinecone announced the general availability of Pinecone Nexus, a knowledge engine that converts an enterprise’s proprietary data into governed, agent‑ready knowledge delivered through a single query c

model-releasespinecone
6 Aug 2026
Model Releases

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

DGX agent

I'm the author, so discount the enthusiasm accordingly. This is an unaffiliated community port, not endorsed by the vLLM project, which it uses to verify its correctness. What started it: I love vLLM,

model-releasesr-localllama
6 Aug 2026
Safety

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of …

DGX agent

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of an explosion of young talent into the bci space recently. i’

safetysonya-huang--x
6 Aug 2026
Applications

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that w…

DGX agent

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that was 'merely' good at hacking under human instructions. Initia

applicationsethan-mollick--x
6 Aug 2026
Model Releases

ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference

DGX agent

arXiv:2608.02947v1 Announce Type: cross Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelen

model-releasesarxiv-cs-cl
5 Aug 2026
Research

Beyond Accuracy: A Multidimensional Evaluation of Statistical Reasoning in Large Language Models

DGX agent

arXiv:2608.03038v1 Announce Type: new Abstract: Statistical reasoning is multidimensional, yet evaluations of large language models (LLMs) typically emphasize response accuracy while overlooking how m

researcharxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety

DGX agent

arXiv:2601.17003v2 Announce Type: replace-cross Abstract: Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual dive

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Inkling-Small 276B-A12B at ~2.9 tok/s on <10gb memory

DGX agent

A follow up to the launch of Mference, it now supports and runs Inkling-Small 276B-A12B. Inkling-Small (Thinking Machines, Apache 2.0), from the pipenetwork/Inkling-Small-MLX-4bit conversion: 276B tot

model-releasesr-localllama
5 Aug 2026
Safety

is there a polite synonym for “circle jerk”?

DGX agent

The post contains two distinct snippets. First, user @GaryMarcus asks whether there is a more polite way to refer to “circle jerk.” Second, it shares a (likely satirical) claim that Microsoft’s AI rev

safetygary-marcus--x
5 Aug 2026
Agents

Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens (Jonathan Vanian/CNBC)

DGX agent

Jonathan Vanian / CNBC: Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens — Meta is rolling o

agentstechmeme
5 Aug 2026
Model Releases

MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

DGX agent

TensorSharp's MoE CPU-offload feature has been merged into main. Here is the parameters description of this feature: Mixture-of-Experts CPU offload: --n-cpu-moe <N> | -ncmoe <N> Keep the routed MoE ex

model-releasesr-localllama
5 Aug 2026
Local Ai

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + offi…

DGX agent

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + official quantized builds, Day 0 Diffusers — the reference Pytho

local-aicomfyui--x
5 Aug 2026
← Previous
1…4041424344…54
Next →