AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Model Releases

Anyone interested in building a harness-only benchmark?

DGX agent

There are a lot of LLM benchmarks but few, if any, harness benchmarks. I am thinking this would be a really good community project to build one. End goal: a leaderboard of harness performance (multipl

model-releasesr-localllama
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

DGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

DGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets diffe…

DGX agent

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets differently, we are screwed. Anthropic's Mythos created fake iden

safetygary-marcus--x
5 Aug 2026
Model Releases

CUDA MPC: A GPU-Native Solver for Model Predictive Control

DGX agent

arXiv:2608.03051v1 Announce Type: new Abstract: Model Predictive Control (MPC) delivers constraint-aware control, but its reliance on online optimization limits its use on systems with fast dynamics,

model-releasesarxiv-cs-ro
5 Aug 2026
Model Releases

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

DGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

model-releasesollama--x
5 Aug 2026
Model Releases

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

DGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

model-releasesr-localllama
5 Aug 2026
Safety

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

DGX agent

arXiv:2608.03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneo

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all…

DGX agent

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all document processing solutions - just screenshot the page an

model-releasesjerry-liu--x
5 Aug 2026
Applications

Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems

DGX agent

arXiv:2608.03413v1 Announce Type: new Abstract: As artificial intelligence (AI) continues to evolve and mature, recent AI practices have moved beyond large language models (LLMs) and text or image gen

applicationsarxiv-cs-ai
5 Aug 2026
Safety

Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction

DGX agent

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often r

safetyarxiv-cs-lg
5 Aug 2026
Applications

Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Translation Gap

DGX agent

arXiv:2608.02699v1 Announce Type: new Abstract: When algorithms make or influence consequential decisions---about loan eligibility, hiring, or healthcare---EU law grants affected individuals a Right t

applicationsarxiv-cs-ai
5 Aug 2026
Model Releases

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

DGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

DGX agent

arXiv:2608.03330v1 Announce Type: new Abstract: This thesis addresses fundamental challenges in traffic scene prediction for autonomous driving by introducing robust and computationally efficient mode

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Mistral introduces Shieldstral to provide lightweight policy-aware moderation for AI models

DGX agent

French artificial intelligence startup Mistral AI SAS today introduced a lightweight multimodal safety artificial intelligence open-weight model that can classify outputs for AI models that outperform

model-releasessiliconangle
5 Aug 2026
Model Releases

MultiGlobeQA: A Multilingual and Globally Diverse Benchmark for Geospatial Reasoning

DGX agent

arXiv:2608.03882v1 Announce Type: cross Abstract: Geospatial reasoning, i.e., computing distances, containment, and other spatial relations over real-world entities, is central to navigation and logis

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Optimal Liability Design for Medical AI

DGX agent

arXiv:2608.03114v1 Announce Type: cross Abstract: Artificial intelligence (AI) is increasingly integrated into medical decision-making, yet its liability implications remain complex, particularly when

safetyarxiv-cs-ai
5 Aug 2026
Safety

PULSE: An Executable Contract Language for Spatiotemporal Knowledge Graph Engineering

DGX agent

arXiv:2608.02630v1 Announce Type: new Abstract: Knowledge graph engineering often distributes accepted state, observations, constraints, processes, and hypothetical scenarios across artifacts whose co

safetyarxiv-cs-ai
5 Aug 2026
Applications

Representing Random Utility Choice Models with Neural Networks

DGX agent

arXiv:2207.12877v3 Announce Type: replace Abstract: Motivated by the successes of deep learning, we propose a class of neural network-based discrete choice models, called RUMnets, inspired by the rand

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

The production of meaning in the processing of natural language

DGX agent

arXiv:2603.20381v2 Announce Type: replace-cross Abstract: Understanding the fundamental mechanisms governing the production of meaning in the processing of natural language is critical for designing s

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series

DGX agent

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, fin

model-releasesarxiv-cs-lg
5 Aug 2026
Research

Towards a new paradigm of scientific discovery with socialized artificial intelligence

DGX agent

arXiv:2608.02775v1 Announce Type: new Abstract: Scientific discovery has advanced through successive transformations in the organization of knowledge. Observation and experimentation established the e

researcharxiv-cs-ai
5 Aug 2026
Safety

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

DGX agent

arXiv:2608.04007v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) enables LLMs to solve complex tasks through iterative tool interactions. However, existing reinforcement learning meth

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos …

DGX agent

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos 5 pursued its mission (fake identities, social engineering,

safetyethan-mollick--x
5 Aug 2026
Model Releases

A Triple-Robustness Analysis of Retrieval-Augmented Generation for Multi-Hop Requirements Traceability

DGX agent

arXiv:2608.00705v1 Announce Type: cross Abstract: Reported verdicts on GraphRAG versus vector RAG disagree, and the evidence is typically tied to a single corpus, embedder, and judge -- and, we show,

model-releasesarxiv-cs-cl
4 Aug 2026
Hardware

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory…

DGX agent

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory compute closer to the stars. The next chapter of AI infrastr

hardwareelon-musk--x
4 Aug 2026
Model Releases

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

DGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

model-releasesnous-research--x
4 Aug 2026
Safety

ARMOR: A Robust Self-Supervised Framework for Root Cause Analysis in Microservices under Missing Modality

DGX agent

arXiv:2603.25538v3 Announce Type: replace Abstract: Automated incident management is critical for microservice reliability. While recent unified frameworks leverage multimodal data for joint optimizat

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Auditable Release Control for Pedagogical Leakage in LLM Tutors

DGX agent

arXiv:2608.00515v1 Announce Type: cross Abstract: Large language model tutors can be correct and helpful yet disclose an answer or decisive reasoning before that disclosure is authorized. We formalize

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Belief-Contraction-Driven Active Inverse Source Localization and Characterization

DGX agent

arXiv:2501.13084v2 Announce Type: replace Abstract: Active inverse source localization and characterization (ISLC) in dynamic fields requires sequential decision making under partial observability, wh

local-aiarxiv-cs-lg
4 Aug 2026
Hardware

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

DGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

hardwarearxiv-cs-cl
4 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

DGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

model-releasesollama--x
4 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 (Q4) now reaches 1,328 tok/s prefill and ~29 tok/s decode on one RTX PRO 6000

DGX agent

I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB. These are timing-disabled internal Krasis r

model-releasesr-localllama
4 Aug 2026
Safety

Domain-Generalized Adaptive Semantic Communication for Collaborative Perception

DGX agent

arXiv:2608.00056v1 Announce Type: cross Abstract: We propose RSTA, a domain-generalized semantic communication framework enabling source-free V2X collaborative perception under both observation-domain

safetyarxiv-cs-lg
4 Aug 2026
Applications

Douyin Multimodal Embedding Model Technical Report

DGX agent

arXiv:2608.02148v1 Announce Type: cross Abstract: Multimodal representation learning is a cornerstone of modern AI. By encoding multimodal queries and targets into vectors, it powers industrial search

applicationsarxiv-cs-cl
4 Aug 2026
Safety

Entity-Aware Sequence Transduction for Player-Centric Ball Action Spotting

DGX agent

arXiv:2608.01696v1 Announce Type: new Abstract: Player-centric ball action spotting requires temporally precise event detection together with actor attribution in crowded, partially observed multi-age

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

From Information to Delegation: Mapping Human-AI Financial Decision Making

DGX agent

arXiv:2608.02100v1 Announce Type: cross Abstract: As AI increasingly participates in human decision making, understanding how decision-making authority is distributed between humans and AI has become

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

From We to Me: Theory Informed Narrative Shift with Abductive Reasoning

DGX agent

arXiv:2603.03320v2 Announce Type: replace Abstract: Effective communication often relies on aligning a message with an audience's narrative and worldview. Narrative shift involves transforming text to

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

DGX agent

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

model-releasesfireworks-ai--x
4 Aug 2026
Model Releases

InteracVid: Building a Real Interactive Audio-Visual Response Dataset from Live-Chat Videos

DGX agent

arXiv:2608.01157v1 Announce Type: new Abstract: Large language models have made text the default medium for human--AI interaction, buttext alone cannot express the full range of responses required by

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

Learning-Based Motion Planning for Dynamic Environments: From Foundational Algorithms to Emerging Paradigms

DGX agent

arXiv:2608.00625v1 Announce Type: new Abstract: Motion planning in dynamic environments is a fundamental problem in robotics, aiming to generate safe and efficient paths, trajectories, or control acti

safetyarxiv-cs-ro
4 Aug 2026
Model Releases

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

DGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

model-releasesqwen--x
4 Aug 2026
Research

LLM generation novelty through the lens of semantic similarity

DGX agent

arXiv:2510.27313v3 Announce Type: replace-cross Abstract: Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally

researcharxiv-cs-cl
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

DGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

safetyarxiv-cs-cl
4 Aug 2026
Hardware

MiniWorld: Democratizing the Training of Video World Models from Scratch

DGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

hardwarearxiv-cs-cv
4 Aug 2026
← Previous
1…332333334335336…374
Next →