AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,171 results
Model Releases

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

DGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

model-releasesarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Where To Look? : Causal Tracing of Vision Encoders in VLM

DGX agent

arXiv:2608.10758v1 Announce Type: new Abstract: Vision-language models can describe an image with remarkable accuracy, yet a more fundamental question remains unanswered: what visual information actua

researcharxiv-cs-cv
12 Aug 2026
Model Releases

An Expectation-Maximization Perspective on Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2504.18587v2 Announce Type: replace-cross Abstract: Reinforcement learning has emerged as a powerful approach for improving the reasoning capabilities of large language models, as demonstrated b

model-releasesarxiv-cs-ai
11 Aug 2026
Research

BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes

DGX agent

arXiv:2608.08027v1 Announce Type: cross Abstract: Prompt injection is a critical security threat in large language model (LLM) applications, where attackers hijack model behavior by embedding maliciou

researcharxiv-cs-lg
11 Aug 2026
Model Releases

BDH-CQ: In-Context Learning with Recurrent Latent Reasoning

DGX agent

arXiv:2608.09888v1 Announce Type: cross Abstract: We introduce BDH-CQ, a reasoning model that combines in-context learning with recurrent latent reasoning. Inputs presented at inference time continuou

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation detection in social media posts

DGX agent

arXiv:2608.09510v1 Announce Type: cross Abstract: Detecting machine-generated disinformation on social media is increasingly difficult as large language models (LLMs) make it easier to generate and re

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

DGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

DarwinX: Evolving Agent Harnesses Through Natural Selection

DGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models

DGX agent

arXiv:2608.09109v1 Announce Type: new Abstract: User feedback offers natural supervision for persistent LLM improvement, but a single message may support multiple behavioral changes with different sco

researcharxiv-cs-ai
11 Aug 2026
Research

DiffSafeMerge: Mitigating Backdoor Inheritance in Diffusion Model Merging

DGX agent

arXiv:2608.09445v1 Announce Type: cross Abstract: Unconditional diffusion checkpoint merging assumes benign sources, yet a compromised public checkpoint can transfer a dormant backdoor while clean gen

researcharxiv-cs-cv
11 Aug 2026
Safety

From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition

DGX agent

arXiv:2308.04553v4 Announce Type: replace Abstract: Visual recognition models are prone to learning spurious correlations induced by a biased training set where certain conditions B (eg, Indoors) are

safetyarxiv-cs-cv
11 Aug 2026
Local Ai

Hallucination-Free GUI Grounding via Regression-Free Layout-Aware Matching

DGX agent

arXiv:2608.09654v1 Announce Type: new Abstract: GUI agents are shifting from metadata-dependent large language models to purely visual multimodal large language models (MLLMs) that operate directly on

local-aiarxiv-cs-ai
11 Aug 2026
Research

How Far Do Foundation Models Transfer to Infant Signals? A Cross-Dataset Transfer Audit with a Unified Need Ontology

DGX agent

arXiv:2608.08989v1 Announce Type: cross Abstract: Public infant cry corpora are small, label-incompatible, and almost always evaluated one corpus at a time. We ask what this practice hides and what fi

researcharxiv-cs-ai
11 Aug 2026
Agents

IntelliAudit: Using Large Language Models to Evaluate Audit Controls

DGX agent

arXiv:2608.07688v1 Announce Type: new Abstract: IT audits require auditors to judge whether heterogeneous organizational evidence satisfies semantic security and compliance controls. This judgment is

agentsarxiv-cs-ai
11 Aug 2026
Research

LHSDet: High-Resolution AI-Generated Image Detection via Visual Question Answering

DGX agent

arXiv:2608.07863v1 Announce Type: new Abstract: Driven by advances in diffusion models and autoregressive models, the fidelity and resolution of AI-generated images now rival those of real images. How

researcharxiv-cs-cv
11 Aug 2026
Model Releases

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding

DGX agent

arXiv:2608.07596v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models transform representations from pretrained vision-language models (VLMs) into robot actions, yet the interface that

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

DGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks

DGX agent

arXiv:2507.19634v4 Announce Type: replace-cross Abstract: Recent advances in large language models have laid the foundation for multimodal LLMs (MLLMs), which unify text, speech, and vision within a s

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Rethinking Reasoning with MDLMs: Early Exits, Post-hoc Reasoning, and Beyond

DGX agent

arXiv:2510.19990v2 Announce Type: replace Abstract: The reasoning paradigm, where language models reason before answering, has enabled breakthroughs on tasks such as mathematical problem-solving. Whil

researcharxiv-cs-lg
11 Aug 2026
Model Releases

UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers

DGX agent

arXiv:2608.09209v1 Announce Type: new Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

DGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting

DGX agent

arXiv:2608.09286v1 Announce Type: cross Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

DGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Achievable 253 t/s - unsloth/Muse Glimmer 30B UD-Q5_K_M on a 5090

DGX agent

Benchmarked Muse Glimmer 30B on my RTX 5090 (32GB), 262k context, UD-Q5_K_M + dflash-kquant + mmproj. Workload Stock master + DFlash ngram-simple PR #26842 + DFlash Code patch 78 t/s 57 t/s 220-253 t/

model-releasesr-localllama
10 Aug 2026
Model Releases

An Exploratory Evaluation of LLM-Assisted Rewriting of Moderate-Complexity Financial Sentences for DisCoCat-Based Sentiment Analysis

DGX agent

arXiv:2608.07439v1 Announce Type: new Abstract: Quantum natural language processing (QNLP) provides a grammar-aware framework for text modeling, and Distributional Compositional Categorical (DisCoCat)

model-releasesarxiv-cs-cl
10 Aug 2026
Research

Are Visual Place Recognition Models Recognizing Places or Conditions? Distractor-Augmented Evaluation and Condition Suppression

DGX agent

arXiv:2608.06847v1 Announce Type: cross Abstract: Long-term Visual Place Recognition (VPR) is typically evaluated by matching queries from one condition against a database from another. Crowdsourced m

researcharxiv-cs-cv
10 Aug 2026
Research

Beyond 'AI Language': The case for the idiolectal nature of LLM output

DGX agent

arXiv:2608.06589v1 Announce Type: cross Abstract: While large language model outputs are frequently analysed as a collective super variety termed 'AI language,' this chapter argues that this perspecti

researcharxiv-cs-ai
10 Aug 2026
Safety

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

DGX agent

arXiv:2608.07460v1 Announce Type: cross Abstract: While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively i

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

DATAREEL: Automated Data-Driven Video Story Generation with Animations

DGX agent

arXiv:2604.25220v2 Announce Type: replace Abstract: Data videos combine animated visualizations with synchronized narration to communicate quantitative information and are widely used in journalism, e

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits

DGX agent

arXiv:2608.07430v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mech

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

DGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

From Optimal Actions to World Models: Identifiability of Transition Kernels in Discounted MDPs

DGX agent

arXiv:2608.07301v1 Announce Type: new Abstract: We study what can be recovered about the transition probabilities of a Markov decision process from optimal actions alone. This is closely related to th

safetyarxiv-cs-lg
10 Aug 2026
Model Releases

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance avai…

DGX agent

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. & EU hosting.

model-releasestogether-ai--x
10 Aug 2026
Model Releases

Learning in Deep Networks under Dale's Constraint

DGX agent

arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under the constraints of real neurons. Althou

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Pre-Inference Routing for Cost-Efficient Document Field Extraction

DGX agent

arXiv:2608.06607v1 Announce Type: new Abstract: Most document-extraction systems use a single model for all documents. This is simple but can be costly for easy cases and less effective for difficult

researcharxiv-cs-cl
10 Aug 2026
Research

ReQuant: Fixed-Grid Discrete Refinement for Post-Training Quantization

DGX agent

arXiv:2608.07019v1 Announce Type: new Abstract: Post-training quantization (PTQ) is widely used to reduce the memory and computational cost of large language models. Existing PTQ methods typically obt

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Stoicheia: Character-Level Masked Diffusion for Ancient Greek Textual Restoration, Parsing, and Metrical Scansion

DGX agent

arXiv:2608.07249v1 Announce Type: new Abstract: We introduce Stoicheia, a 405M-parameter character-level masked-diffusion encoder for Ancient Greek whose input factors into five aligned, independently

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Tensor Network Kernel Machines: A JAX Framework for Machine Learning and Nonlinear System Identification

DGX agent

arXiv:2608.07043v1 Announce Type: cross Abstract: Developing nonlinear models that are both expressive and computationally efficient remains a challenge in machine learning and nonlinear system identi

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

CyberKimi just dropped strong results on one of ExploitBench’s hardest V8 bugs , points away from Mythos

DGX agent

Hey everyone ! Quick share from the cyber + local LLM side of things that I found interesting. During this week’s hacker summer camp, an AI researcher and reverse malware engineer veteran 'lordx64' on

model-releasesr-ollama
9 Aug 2026
Model Releases

KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding.

DGX agent

First of all, I'm not a lab, this was a solo summer research project that finally culminated into the github repo and the writeup. The repo includes a much deeper dive with methods, findings about qua

model-releasesr-localllama
9 Aug 2026
Model Releases

~45% lower MiniMax H3 sampler time with new Spectrum settings — degree 1 works surprisingly well (v0.1.8)

DGX agent

Follow-up to my original Spectrum MiniMax H3 post: https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/ In that first post, I released the MiniMax

model-releasesr-stablediffusion
7 Aug 2026
Model Releases

Am I just hallucinating

DGX agent

Or is there any reason why I feel like model output quality seems to be better when I use higher micro-batch values (ub) in llama-cpp? I don't really have any hard numbers or anything (just running th

model-releasesr-localllama
7 Aug 2026
Research

Arbitrage: Efficient Reasoning via Advantage-Aware Speculation

DGX agent

Modern Large Language Models achieve impressive reasoning capabilities with long Chain of Thoughts, but they incur substantial computational cost during inference, and this motivates techniques to imp

researchapple-ml-research
7 Aug 2026
Model Releases

Basically every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be signficantly higher with a bett…

DGX agent

On August 7, 2026 Ethan Mollick tweeted that “every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be significantly higher with a better harness.” The commen

model-releasesethan-mollick--x
7 Aug 2026
Research

BendTwin: Robust Dense-to-Sparse Physical Reconstruction with Bending-Aware Differentiable Spring-Mass Models

DGX agent

arXiv:2608.06164v1 Announce Type: new Abstract: Reconstructing objects with mechanical properties from video observations enables physically consistent dynamic prediction, benefiting robotics planning

researcharxiv-cs-cv
7 Aug 2026
Model Releases

Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

DGX agent

arXiv:2608.05238v1 Announce Type: new Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Evaluating and Improving Pedagogical Fit in LLM-Based AI Tutors with the Pedagogical Suitability Index

DGX agent

arXiv:2608.05411v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as AI tutors, but a correct answer is not always a pedagogically appropriate one. In classroom learni

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…357358359360361…1358
Next →