AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Model Releases

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap!

DGX agent

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap! We ran DeepSeek V4 Flash through 4 more agent harnesses (Hermes Agent, Pi Agent, Pri

model-releasesharrison-chase--x
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition

DGX agent

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predi

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Your VLM Already Knows When: Training-Free Temporal Grounding by Asking Yes or No

DGX agent

arXiv:2608.08315v1 Announce Type: new Abstract: Multimodal LLMs that recognise events reliably still fail to say when they happen. Prompted for timestamps, strong VLMs reach as little as 3.8% R@0.5 on

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Zero-shot 2D Grounding with Novel Affordance Types

DGX agent

arXiv:2608.08929v1 Announce Type: new Abstract: 2D affordance grounding aims to locate the region of an object that a human can interact with. Existing research focuses on recognizing affordance types

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Accounting Graph Transformer for Short-History Multi-KPI Forecasting in Small Businesses

DGX agent

arXiv:2608.07037v1 Announce Type: cross Abstract: Small businesses often have only 12-24 months of accounting history, yet planning and risk workflows require coordinated forecasts across financial st

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallelized Q-Networks

DGX agent

arXiv:2608.07335v1 Announce Type: cross Abstract: Recent advancements in deep reinforcement learning have increasingly favored simplified, highly parallelized paradigms. Notably, the Parallelized Q-Ne

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Beyond Fluency: A Clinical Benchmark and Anomaly-Enhanced Baseline for Spine MRI Report Generation

DGX agent

arXiv:2608.07117v1 Announce Type: new Abstract: Radiology reporting is time-consuming and subject to inter-rater variability, making automated report generation an attractive clinical application for

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Beyond Text Matching: Towards Reference-Free Evaluation for Human-Oriented Binary Reverse Engineering

DGX agent

arXiv:2608.07038v1 Announce Type: cross Abstract: Human-Oriented Binary Reverse Engineering (HOBRE) aims to transform decompiled pseudocode into a more human-friendly representation, thereby reducing

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Coding isn't yet another application domain -- it's the meta-skill required for AI to automatically develop its own training material, via s…

DGX agent

Coding isn't yet another application domain -- it's the meta-skill required for AI to automatically develop its own training material, via symbolic world models. That's how the RSI loop actually kicks

researchfrancois-chollet--x
10 Aug 2026
Model Releases

Comparing how Cline, Kilo, and Qwen Code handle long-task context/state (and why context loops keep happening)

DGX agent

I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside the conversation that gets reinjected on a c

model-releasesr-localllama
10 Aug 2026
Research

Debias in Text, Believe Your Eyes: Text-Anchored Cross-Modal Transfer for Visual Counter-Commonsense Reasoning

DGX agent

arXiv:2608.06938v1 Announce Type: cross Abstract: The visual reasoning ability of multimodal large language models (MLLMs) is crucial for downstream applications, particularly counter-commonsense reas

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Deep Evidential Regression for Sparse Forest Height Estimation from Multimodal Satellite Imagery

DGX agent

arXiv:2608.06406v1 Announce Type: new Abstract: Accurate estimation of forest height from satellite imagery is essential for applications such as carbon accounting, biodiversity monitoring, and ecosys

model-releasesarxiv-cs-cv
10 Aug 2026
Safety

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

DGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Fairis: Fairness-Aware Aggregation with Provable Influence Containment against Fairness Poisoning Attacks in Collaborative Machine Learning

DGX agent

arXiv:2608.06469v1 Announce Type: cross Abstract: Collaborative machine learning among financial institutions must be both group-fair and robust against deliberate adversarial manipulation. Existing f

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

Finding Usable Weight Mechanisms with Tiled SVD

DGX agent

arXiv:2608.06969v1 Announce Type: new Abstract: The dominant approach to mechanistic interpretability trains proxy dictionaries such as sparse autoencoders and labels features from max-activating text

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

DGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

I think this shows 2 key skills w/ AI 1) compute allocation - for most jobs there's not a list of 'X most important problems', you have to d…

DGX agent

I think this shows 2 key skills w/ AI 1) compute allocation - for most jobs there's not a list of 'X most important problems', you have to decide what problems are worth it 2) thought partnership - @_

model-releasesthariq--x
10 Aug 2026
Safety

Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons

DGX agent

arXiv:2506.03392v2 Announce Type: replace Abstract: We propose a new ternary spiking neuron model to improve the representation capacity of binary spiking neurons in deep Q-learning. Although a ternar

safetyarxiv-cs-lg
10 Aug 2026
Research

Is SwiGLU's Open Positive Tail Necessary? Evidence from Closed-Tail Gating with MemGLU

DGX agent

arXiv:2608.07323v1 Announce Type: new Abstract: We test whether decoder-only language-model FFNs require SwiGLU's open positive tail. We introduce MemGLU as a closed-tail comparator derived from a mem

researcharxiv-cs-lg
10 Aug 2026
Model Releases

LifelongCrossNav: Persistent 3D Semantic Memory for Cross-Floor Multi-Object Navigation

DGX agent

arXiv:2608.07079v1 Announce Type: cross Abstract: Object-goal navigation has made substantial progress in semantic perception and exploration, yet persistent memory for multi-object navigation and cro

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Mathematical Principles and Experimental Discoveries of the Emergence of Symbolic Patterns in Artificial Neural Networks

DGX agent

arXiv:2608.06839v1 Announce Type: new Abstract: Artificial Neural networks (ANNs) are often treated as black-box models, making explainability a central challenge in deep learning. Many engineering me

researcharxiv-cs-lg
10 Aug 2026
Model Releases

Most Civitai 10$ checkpoint's are scams. Don't fall for it.

DGX agent

If you read -> huge claims + AI like generated presentation + no negative comment AND '$10 to download on my patreon/whatever' = they're scammers. Period. 1- Anyone leaving a negative comment or tiny

model-releasesr-stablediffusion
10 Aug 2026
Model Releases

Muse Glimmer is now available to run with Ollama. Available today via Ollama’s MLX engine with state-of-the-art-performance on Apple Silicon…

DGX agent

Muse Glimmer is now available to run with Ollama. Available today via Ollama’s MLX engine with state-of-the-art-performance on Apple Silicon, Muse Glimmer can power Claude Code, Codex, and more always

model-releasesollama--x
10 Aug 2026
Model Releases

NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs

DGX agent

arXiv:2608.07167v1 Announce Type: new Abstract: Giving an AI agent the ability to send emails, query databases, or execute commands is useful--until the agent is tricked into doing something it should

model-releasesarxiv-cs-ai
10 Aug 2026
Tutorials

Omni-modal decomposition autoencoders learn full-stack wearable disentangled representations

DGX agent

arXiv:2608.07385v1 Announce Type: cross Abstract: Learning disentangled representations is a key requirement for developing versatile, general-purpose, and sustainable models in multi-modal wearable c

tutorialsarxiv-cs-ai
10 Aug 2026
Model Releases

OpenAI releases GPT-5.6-Cyber, a more cyber-permissive version of GPT-5.6 Sol, to some partners and expands its Daybreak cybersecurity initiative (Sam Sabin/Axios)

DGX agent

Sam Sabin / Axios: OpenAI releases GPT-5.6-Cyber, a more cyber-permissive version of GPT-5.6 Sol, to some partners and expands its Daybreak cybersecurity initiative — OpenAI is introducing a more cybe

model-releasestechmeme
10 Aug 2026
Safety

People Are Not Just Their Countries. Disentangling Social Determinants of LLM Value Alignment Across Europe

DGX agent

arXiv:2608.07367v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly used as a primary source of information and advice, understanding their alignment to humans in terms of

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

PHOENIX: Fine-Tuned SLM-Powered Autonomous Satellite Lifetime Extension via Predictive Self-Healing and Multi-Agent AI Recovery

DGX agent

arXiv:2608.07126v1 Announce Type: cross Abstract: Most CubeSats, small and low-cost satellites roughly the size of a shoebox, do not survive as long as they were designed to: a study of 178 missions f

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Playing with physics is so cool in Minimax H3

DGX agent

Prompt: 'integrated_multimodal_description: [Shot 1] Live-action, ultra-realistic first-person footage at night on a rainy city street, filmed with authentic handheld smartphone qualities. The phone i

model-releasesr-stablediffusion
10 Aug 2026
Safety

Self-Distillation Enables Continual Learning

DGX agent

arXiv:2601.19897v2 Announce Type: replace Abstract: Continual learning, enabling models to acquire new skills and knowledge without degrading existing capabilities, remains a fundamental challenge for

safetyarxiv-cs-lg
10 Aug 2026
Research

TEXAS: Task-Expert-Aware Supervision for Downstream Mixture-of-Experts LLM Adaptation

DGX agent

arXiv:2608.06396v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models route each token through a small subset of experts, making routing patterns useful for identifying task-relev

researcharxiv-cs-ai
10 Aug 2026
Agents

The next frontier of Recursive Self-Improvement is Physical AI. Japan sparked the robotics revolution. We are expanding our RSI Lab to build…

DGX agent

The next frontier of Recursive Self-Improvement is Physical AI. Japan sparked the robotics revolution. We are expanding our RSI Lab to build world models that allow agentic reasoning systems to recurs

agentsdavid-ha--x
10 Aug 2026
Model Releases

TransSLR: A Lightweight Transformer for Sign Language Recognition

DGX agent

arXiv:2608.06407v1 Announce Type: cross Abstract: Automated Sign Language Recognition for under-represented languages remains a largely unsolved problem. Central African Sign Language (CASL) exemplifi

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

We asked an unreleased research version of Claude to take a stab at the Riemann hypothesis. It didn’t solve it, but it did make strides on a…

DGX agent

We asked an unreleased research version of Claude to take a stab at the Riemann hypothesis. It didn’t solve it, but it did make strides on a related problem: it increased the lower bound for the fract

model-releasesboris-cherny--x
10 Aug 2026
Model Releases

Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly …

DGX agent

Grok Build is quickly turning into an all-in-one creation environment It can now also generate images and videos with Grok Imagine directly inside your workflow You can create custom visuals for websi

model-releaseselon-musk--x
9 Aug 2026
Model Releases

My first run of Kimi K3 locally.

DGX agent

Running across 2 clusters using llama.cpp over RPC too. Both clusters are not enough to hold everything in memory, so main cluster still partially offloads to run. Goal will be to get all the GPUs in

model-releasesr-localllama
8 Aug 2026
Model Releases

Weirdly iirc stable diffusion (1.4) finished training around four years ago today too

DGX agent

Emad posted that Stable Diffusion v1.4 reached the end of its training cycle roughly four years before the post was published. Greg Brockman added that GPT‑4 similarly completed training around the sa

model-releasesemad-mostaque--x
8 Aug 2026
Safety

CircuitSteer: Geometrically Aligned Multi-Layer Steering via Sparse Autoencoder Circuits

DGX agent

arXiv:2608.05732v1 Announce Type: new Abstract: Controlling the behavior of large language models (LLMs) remains a critical challenge for AI alignment. Existing steering methods, such as Contrastive A

safetyarxiv-cs-lg
7 Aug 2026
Research

d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation

DGX agent

arXiv:2601.07568v3 Announce Type: replace-cross Abstract: Diffusion large language models (dLLMs) offer capabilities beyond those of autoregressive (AR) LLMs, such as parallel decoding and random-orde

researcharxiv-cs-ai
7 Aug 2026
Research

DBLAST: Dependent Block Drafting for Stochastic Speculative Decoding

DGX agent

arXiv:2608.05448v1 Announce Type: new Abstract: Speculative decoding accelerates large language models' inference by using a lightweight drafter to propose multiple future tokens and a target model to

researcharxiv-cs-cl
7 Aug 2026
Model Releases

Dual-space posterior sampling for Bayesian inference in constrained inverse problems

DGX agent

arXiv:2603.00393v2 Announce Type: replace-cross Abstract: Inverse problems constrained by partial differential equations are often ill-conditioned due to noisy, incomplete data or inherent non-uniquen

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning

DGX agent

arXiv:2608.06197v1 Announce Type: new Abstract: Training large language model agents for long-horizon tool use typically relies on interactions with real or synthesized executable environments, whose

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

I made a simple local voice input extension for pi (nemotron 3.5 0.6B ASR)

DGX agent

There are already plenty of different extensions for voice input, but all I found required having a second server running. I wanted something super simplistic: launching local STT server just for my p

model-releasesr-localllama
7 Aug 2026
Local Ai

Innovation-Residual Auditing of Autonomous Analysis Agents: Localization, Detection Limits, Error Control, and Identifiability

DGX agent

arXiv:2608.05490v1 Announce Type: new Abstract: Autonomous agents now carry out entire data analyses, selecting cohorts, joining tables, and fitting models with little step-by-step supervision. When s

local-aiarxiv-cs-ai
7 Aug 2026
Tutorials

Integrating Implicit and Explicit Relational Biases through Graph-Based Multiple Instance Learning: A Case Study in Skin Lesion Diagnosis

DGX agent

arXiv:2608.06037v1 Announce Type: new Abstract: Relational inductive biases are essential for capturing structural dependencies among data. This study investigates a dual-level relational framework fo

tutorialsarxiv-cs-ai
7 Aug 2026
Model Releases

Invariant Representation Learning for Source-Free Time Series Forecasting with LLM-Centric Proxy Denoising

DGX agent

arXiv:2510.05589v3 Announce Type: replace-cross Abstract: Effective time series forecasting enables various real-world applications, benefiting from the proliferation of mobile devices. However, the v

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

MACRO: Markov Chain Routing of Transformer Layers

DGX agent

arXiv:2608.05872v1 Announce Type: cross Abstract: Standard Large Language Models (LLMs) execute layers sequentially. Dynamic layer routing, i.e. search for a different execution path through layers in

safetyarxiv-cs-ai
7 Aug 2026
← Previous
1…588589590591592…1395
Next →