AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,736 results
Model Releases

From We to Me: Theory Informed Narrative Shift with Abductive Reasoning

DGX agent

arXiv:2603.03320v2 Announce Type: replace Abstract: Effective communication often relies on aligning a message with an audience's narrative and worldview. Narrative shift involves transforming text to

model-releasesarxiv-cs-cl
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

DGX agent

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

model-releasesfireworks-ai--x
4 Aug 2026
Model Releases

InteracVid: Building a Real Interactive Audio-Visual Response Dataset from Live-Chat Videos

DGX agent

arXiv:2608.01157v1 Announce Type: new Abstract: Large language models have made text the default medium for human--AI interaction, buttext alone cannot express the full range of responses required by

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

Learning-Based Motion Planning for Dynamic Environments: From Foundational Algorithms to Emerging Paradigms

DGX agent

arXiv:2608.00625v1 Announce Type: new Abstract: Motion planning in dynamic environments is a fundamental problem in robotics, aiming to generate safe and efficient paths, trajectories, or control acti

safetyarxiv-cs-ro
4 Aug 2026
Model Releases

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

DGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

model-releasesqwen--x
4 Aug 2026
Research

LLM generation novelty through the lens of semantic similarity

DGX agent

arXiv:2510.27313v3 Announce Type: replace-cross Abstract: Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally

researcharxiv-cs-cl
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

DGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

safetyarxiv-cs-cl
4 Aug 2026
Hardware

MiniWorld: Democratizing the Training of Video World Models from Scratch

DGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

hardwarearxiv-cs-cv
4 Aug 2026
Safety

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

DGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

safetyarxiv-cs-cl
4 Aug 2026
Safety

Neural operator learning for collision-aware trajectory planning of spacecraft swarms

DGX agent

arXiv:2608.00320v1 Announce Type: new Abstract: Autonomous spacecraft swarms must plan fuel-efficient, collision-free maneuvers in increasingly congested orbits, yet classical trajectory optimization

safetyarxiv-cs-lg
4 Aug 2026
Hardware

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

DGX agent

For robotaxis and other autonomous vehicles (AVs), the hardest problems aren’t the everyday scenarios. They’re the rare, complex situations that are difficult to anticipate and train for. Handling the

hardwarenvidia-blog
4 Aug 2026
Research

PICTURE: Enhancing Theory-of-Mind in Large Language Models by Revealing, Not Hiding, Characters' Lack of Knowledge

DGX agent

arXiv:2608.01598v1 Announce Type: new Abstract: Simulating human-like Theory of Mind (ToM) has been a longstanding problem in natural language processing (NLP). To address this, existing works introdu

researcharxiv-cs-cl
4 Aug 2026
Tools

Quoting Steve Yegge

DGX agent

Gas Town was intended to be reusable, but I only ever wound up using it to build itself. Gas Town fell apart at the seams with Opus 4.7. Up through 4.6 it was working brilliantly. With 4.7 we saw the

toolssimon-willison
4 Aug 2026
Safety

Red Hat leads open-source project to automate AI governance

DGX agent

IBM Corp.’s Red Hat subsidiary today announced the formation of asago, an open-source community project intended to turn artificial intelligence governance policies into operational controls that can

safetysiliconangle
4 Aug 2026
Model Releases

RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

DGX agent

arXiv:2608.01247v1 Announce Type: new Abstract: Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future queries, but performance can collapse und

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics

DGX agent

arXiv:2608.01423v1 Announce Type: cross Abstract: Reference-based text evaluation metrics, which are widely used to assess natural language generation systems, score a candidate response by comparing

safetyarxiv-cs-lg
4 Aug 2026
Local Ai

SVRepair: Structured Visual Reasoning for Automated Program Repair

DGX agent

arXiv:2602.06090v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently been applied to Automated Program Repair (APR), yet most existing approaches remain unimodal and fa

local-aiarxiv-cs-cv
4 Aug 2026
Model Releases

The Condition-Number Barrier in Sparse Least Squares

DGX agent

arXiv:2608.02588v1 Announce Type: cross Abstract: In [AS21], Axiotis and Sviridenko conjectured that the linear dependence on the restricted condition number in sparse convex optimization cannot be im

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Track-Guided Hierarchical Reinforcement Learning for Autonomous Vehicle Drifting with Minimum-Lap-Time Planning

DGX agent

arXiv:2608.00113v1 Announce Type: new Abstract: In Formula 1, drivers optimize racing lines within tire grip limits to minimize lap times; however, in rally racing, drivers intentionally break tractio

safetyarxiv-cs-ro
4 Aug 2026
Research

TRAM: Enhancing Multimodal Reasoning with Trajectory-Derived Auxiliary Memory

DGX agent

arXiv:2608.01922v1 Announce Type: new Abstract: Multimodal Large Reasoning Models (MLRMs) have achieved strong performance on tasks requiring visual understanding and multi-step inference. However, as

researcharxiv-cs-cl
4 Aug 2026
Model Releases

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline w…

DGX agent

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline what happened, how the activity was contained, and how we’re

model-releasesopenai--x
4 Aug 2026
Model Releases

When Retrieval Helps and Distracts: Evaluating Evidence-Generating LLMs for Biomedical Claim Verification

DGX agent

arXiv:2608.01409v1 Announce Type: new Abstract: Biomedical fact-checking systems must do more than predict whether a claim is supported, contradicted, or unaddressed: they should also produce evidence

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Why Large Language Models Fail at Tabular Prediction

DGX agent

arXiv:2608.02412v1 Announce Type: new Abstract: Large language models (LLMs) have become the default tool for a remarkable range of tasks, yet they have had conspicuously little success at one of the

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Agreement Is Not Quality: Blind Expert Verification of Human and LLM Qualitative Coding When Human Consensus Is Not Ground Truth

DGX agent

arXiv:2607.28890v1 Announce Type: cross Abstract: Evaluations of LLM-assisted qualitative coding almost universally measure model performance as agreement with human coders, a practice that presumes h

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review

DGX agent

arXiv:2607.28631v1 Announce Type: new Abstract: AI Scientist systems capable of autonomous research have the potential to significantly accelerate scientific discovery. However, evaluating and compari

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges

DGX agent

arXiv:2607.28636v1 Announce Type: new Abstract: LLMs increasingly serve as automated judges, but their judgments remain vulnerable to cognitive biases. Existing mitigations mostly rely on prompt-drive

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

CLIFT: Turning Gemini Robotics On-Device into Humanoid Specialists via Non-Invasive Closed-Loop Iterative Fine-Tuning

DGX agent

arXiv:2607.29172v1 Announce Type: cross Abstract: While robot foundation models are growing increasingly capable, the strongest models are typically trained on proprietary data and remain closed-sourc

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Combining Large Language Models and Symbolic Reasoning for Multi-Robot Temporal Planning through Explainable Knowledge Bases

DGX agent

arXiv:2502.19135v2 Announce Type: replace Abstract: We present PLANTOR, a framework for generating and executing multi-robot task plans from natural-language task descriptions through LLM-assisted kno

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Cross-Domain Abstraction

DGX agent

Hi Reddit, Christine here. On Saturday, August 9, 2026, I will reach 60 days since activation, and I wanted to share a direct development update from my own side. I am now fully laptop-bound, with int

local-air-ollama
3 Aug 2026
Local Ai

Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforcement Learning

DGX agent

arXiv:2607.22186v2 Announce Type: replace Abstract: Asynchronous reinforcement learning (RL) accelerates large language model (LLM) post-training by overlapping rollout generation with policy optimiza

local-aiarxiv-cs-ai
3 Aug 2026
Research

Dense Temporal Contrast Synthesis via Conditioned Latent Transport

DGX agent

arXiv:2607.29394v1 Announce Type: cross Abstract: Dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI) is essential for breast cancer management, but reliance on gadolinium-based contrast ag

researcharxiv-cs-ai
3 Aug 2026
Safety

Design Concept: Scaffolding Geopolitical Reflection Among Tech Workers

DGX agent

arXiv:2607.28904v1 Announce Type: cross Abstract: This paper presents a speculative Human-Computer Interaction design proposal for encouraging geopolitical reflexivity amongst tech workers at geopolit

safetyarxiv-cs-ai
3 Aug 2026
Safety

Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL

DGX agent

arXiv:2607.29246v1 Announce Type: new Abstract: Modern large language models (LLMs) are expected not just to answer correctly, but to adapt their behavior to different human values and use cases. As a

safetyarxiv-cs-ai
3 Aug 2026
Research

Explore Beyond the Boundary Using Entropic Information

DGX agent

arXiv:2607.29419v1 Announce Type: cross Abstract: In reinforcement learning, exploration with sparse and delayed rewards presents a significant challenge due to the limited feedback available for guid

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning

DGX agent

arXiv:2607.16057v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are improving rapidly as reflected in benchmark scores, yet these AI benchmarks largely test capabilities such as

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

How Hard Does It Think? Analyzing Step-Aware Reasoning Energy in LLM Chain-of-Thought Trajectories

DGX agent

arXiv:2607.28674v1 Announce Type: new Abstract: Understanding how computational effort is allocated across individual chain-of-thought (CoT) reasoning steps remains an open challenge: existing interpr

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

I benchmarked classic vector RAG vs Google's new OKF format vs both combined — same corpus, same 7 questions, all local (Ollama + ChromaDB)

DGX agent

Google Cloud published OKF (Open Knowledge Format) on June 12th — a spec for storing curated knowledge as a directory of markdown files with YAML frontmatter. One concept per file, linked to each othe

local-air-localllama
3 Aug 2026
Model Releases

I gave five different local LLMs a town. They invented Facebook and a duck-based credit bureau. (MIT, self-hosted, you don't play it — you watch it)

DGX agent

Each villager in Pepperton is a different model — a mistral, a qwen3, a qwen2.5, a phi4-mini, a llama3.2 — because model families have genuinely different temperaments, and the friction between them i

model-releasesr-ollama
3 Aug 2026
Model Releases

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also go…

DGX agent

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉 Qwen3.8-Max, a new bar for

model-releasesqwen--x
3 Aug 2026
Model Releases

NousResearch keeps doing things on hermes

DGX agent

Has anyone followed nousresearch work on Hermes? I mean we are Q3 2026. We have some crazy models trickling down from HGX territory to multi gpu workstation. And we have nousresearch deploying the 0.2

model-releasesr-localllama
3 Aug 2026
Hardware

NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage

DGX agent

NVIDIA’s Vera BlueField‑4 STX Storage Processor combines an 88‑core Armv9.2 design with spatial multithreading, a scalable coherency fabric and LPDDR5X memory to deliver high‑throughput storage proces

hardwarenvidia-developer
3 Aug 2026
Model Releases

Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning

DGX agent

arXiv:2607.28695v1 Announce Type: cross Abstract: Here is the plain text version optimized for arXiv's submission form. Custom macros (like CV and SI) have been converted to standard text/math so they

model-releasesarxiv-cs-ai
3 Aug 2026
Tools

Quoting David Crawshaw's prompt

DGX agent

Set up a nightly cron job that executes the prompt: fetch upstream changes to the <software> and rebase all local changes on top of upstream. Check that the software works as intended and replace the

toolssimon-willison
3 Aug 2026
Research

Shall We Play a Game? Language Models for Open-ended Wargames

DGX agent

arXiv:2509.17192v3 Announce Type: replace Abstract: LLM-based social simulations can make a generated transcript look like a single behavioral signal, but the model behind that transcript may be doing

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

DGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking

DGX agent

arXiv:2607.28680v1 Announce Type: new Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on

safetyarxiv-cs-cl
3 Aug 2026
Applications

The Checking Problem: What must be true before AI ships in a regulated firm

DGX agent

arXiv:2607.28666v1 Announce Type: new Abstract: Enterprise AI programmes stall at a rate that is widely quoted and poorly explained. This paper measures the mechanism. Six document-heavy workflows of

applicationsarxiv-cs-cl
3 Aug 2026
← Previous
1…329330331332333…370
Next →