AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Tutorials

Interpretability Without Tradeoffs: Disentangling Polysemanticity At Equal Predictive Performance

DGX agent

arXiv:2605.31304v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are widely used, but interpreting what they actually learn remains difficult. A major obstacle is that individual neurons

tutorialsarxiv-cs-cv
1 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories

DGX agent

arXiv:2605.31492v1 Announce Type: new Abstract: Large language models (LLMs) often solve reasoning problems by generating intermediate traces that explore and revise partial solutions. From a search p

local-aiarxiv-cs-ai
1 Jun 2026
Model Releases

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

DGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

model-releasesarxiv-cs-ai
1 Jun 2026
Local Ai

Local AI News You Missed - May 2026

DGX agent

This Reddit post from r/StableDiffusion likely curates May 2026 AI news highlights, including releases like Stable Audio 3.0, a model family for artistic experimentation with open-weight models. The p

local-air-stablediffusion
1 Jun 2026
Model Releases

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

DGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

MiniMax M3 imminent. Will be doing deep testing with it on my own coding agent and harness. Review coming soon.

DGX agent

MiniMax M3, an upcoming AI model, is expected to be released soon and will undergo comprehensive testing within a custom coding agent framework. A detailed technical review of the model's performance

agentsdair-ai--x
1 Jun 2026
Local Ai

MiniMax M3 launched!

DGX agent

MiniMax M3 launched on June 1, 2026 as the first open-weights model to combine frontier coding, a 1-million-token context window, and native multimodality. The model achieves top-tier performance on c

local-air-ollama
1 Jun 2026
Model Releases

On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection

DGX agent

arXiv:2605.30544v1 Announce Type: new Abstract: Visual monitoring systems that rely on cloud-based AI inference expose raw image data to external services, creating fundamental tensions with the data-

model-releasesarxiv-cs-cv
1 Jun 2026
Applications

OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning

DGX agent

arXiv:2605.30736v1 Announce Type: cross Abstract: The rapid development of large language models, each with distinct capabilities and inference costs, raises a practical deployment question: given an

applicationsarxiv-cs-ai
1 Jun 2026
Agents

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

DGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

agentsarxiv-cs-ai
1 Jun 2026
Applications

PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

DGX agent

arXiv:2502.12119v4 Announce Type: replace-cross Abstract: Visual instruction tuning adapts pre-trained Multimodal Large Language Models (MLLMs) to follow human instructions for real-world applications

applicationsarxiv-cs-ai
1 Jun 2026
Model Releases

Recognizing Co-Speech Gestures in-the-Wild

DGX agent

arXiv:2605.31589v1 Announce Type: new Abstract: While humans naturally gesture during speech, only a sparse subset of these movements are visually depictive and semantically linked to specific spoken

model-releasesarxiv-cs-cv
1 Jun 2026
Research

Rectified flow-based prediction of post-treatment brain MRI from pre-radiotherapy priors for patients with glioma

DGX agent

arXiv:2603.08385v2 Announce Type: replace-cross Abstract: Brain tumors result in 20 years of lost life on average. Standard therapies induce complex structural changes in the brain that are monitored

researcharxiv-cs-cv
1 Jun 2026
Model Releases

Scaling Conversational Hungarian ASR: The BEA-Dialogue+ Corpus

DGX agent

arXiv:2605.31469v1 Announce Type: cross Abstract: Conversational automatic speech recognition in Hungarian is constrained by the limited amount of publicly available dialogue-style training data. The

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Scaling Higher-Order Graph Learning with Maximal Clique Complexes

DGX agent

arXiv:2605.31373v1 Announce Type: cross Abstract: Graph neural networks (GNNs) are limited to modeling pairwise interactions, while higher-order models based on cell complexes achieve greater expressi

researcharxiv-cs-ai
1 Jun 2026
Model Releases

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

DGX agent

arXiv:2605.31148v1 Announce Type: cross Abstract: Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into ac

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery

DGX agent

arXiv:2605.31156v1 Announce Type: new Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding a

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Targeted Speaker Poisoning Framework in Zero-Shot Text-to-Speech

DGX agent

arXiv:2603.07551v2 Announce Type: replace-cross Abstract: Zero-shot Text-to-Speech (TTS) voice cloning poses severe privacy risks, demanding the removal of specific speaker identities from trained TTS

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy Expansion

DGX agent

arXiv:2601.09633v2 Announce Type: replace Abstract: Taxonomies form the backbone of structured knowledge representation across diverse domains, enabling applications such as e-commerce and semantic se

model-releasesarxiv-cs-cl
1 Jun 2026
Research

The Gaussian-Head OFL Family: One-Shot Federated Learning from Client Global Statistics

DGX agent

arXiv:2602.01186v2 Announce Type: replace-cross Abstract: Classical Federated Learning relies on a multi-round iterative process of model exchange and aggregation between server and clients, with high

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing

DGX agent

arXiv:2605.31367v1 Announce Type: cross Abstract: Token mixing layers play a key role in how language models can learn and generate long-range dependencies. Their efficiency relies on the necessary tr

tutorialsarxiv-cs-cl
1 Jun 2026
Model Releases

TSM-Bench: Detecting LLM-Generated Text in Real-World Wikipedia Editing Practices

DGX agent

arXiv:2605.31113v1 Announce Type: new Abstract: Automatically detecting machine-generated text (MGT) is critical to maintaining the knowledge integrity of user-generated content (UGC) platforms such a

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

VeriGate: Verifier-Gated Step-Level Supervision for GRPO

DGX agent

arXiv:2605.30451v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is an effective recipe for training reasoning models with verifier-based outcome rewards, but its supervision

safetyarxiv-cs-lg
1 Jun 2026
Model Releases

WristCompass: Kinematic Coupling as a Learnable Visual Concept for Ego-Camera Orientation

DGX agent

arXiv:2605.30671v1 Announce Type: new Abstract: Recovering ego-camera orientation from manipulation video is a prerequisite for disentangling hand motion from camera motion, a key step in imitation le

model-releasesarxiv-cs-cv
1 Jun 2026
Research

4DPC^2hat: Towards Dynamic Point Cloud Understanding with Failure-Aware Bootstrapping

DGX agent

arXiv:2602.03890v2 Announce Type: replace Abstract: Point clouds provide a compact and expressive representation of 3D objects, and have recently been integrated into multimodal large language models

researcharxiv-cs-cv
29 May 2026
Model Releases

A Full-Pipeline Framework for Evaluating Membership Inference Attacks in Machine Learning

DGX agent

arXiv:2605.29454v1 Announce Type: new Abstract: While Membership Inference Attacks (MIAs) are the prevailing method for identifying training data, their application has expanded into privacy auditing

model-releasesarxiv-cs-lg
29 May 2026
Research

AdaState: Self-Evolving Anchors for Streaming Video Generation

DGX agent

arXiv:2605.30349v1 Announce Type: new Abstract: Autoregressive video diffusion models generate streaming video by producing frames sequentially, conditioning each chunk on previously generated content

researcharxiv-cs-cv
29 May 2026
Safety

AIRGuard: Guarding Agent Actions with Runtime Authority Control

DGX agent

arXiv:2605.28914v1 Announce Type: cross Abstract: Tool-using language agents turn model decisions into external side effects: they read files, run scripts, call APIs, send messages, and invoke Model C

safetyarxiv-cs-ai
29 May 2026
Model Releases

AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation

DGX agent

arXiv:2512.01334v2 Announce Type: replace Abstract: Text-guided image-to-video generation has made substantial progress, yet it still struggles to execute text-specified edits that require substantial

model-releasesarxiv-cs-cv
29 May 2026
Local Ai

b9391

DGX agent

b9391 is a build release of llama.cpp, an open-source C/C++ inference framework for running large language models locally. llama.cpp provides lightweight, optimized model inference with support for mu

local-aillama-cpp-releases
29 May 2026
Local Ai

Benchmarking at the Edge of Comprehension

DGX agent

arXiv:2602.14307v3 Announce Type: replace Abstract: As frontier Large Language Models (LLMs) increasingly saturate new benchmarks shortly after they are published, benchmarking itself is at a juncture

local-aiarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

DGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
Research

Beyond MSE: Improving Precipitation Nowcasting with Multi-Quantile Regression

DGX agent

arXiv:2605.30122v1 Announce Type: cross Abstract: Deep-learning precipitation nowcasting models are often optimized using pointwise losses such as mean squared error or mean absolute error, which can

researcharxiv-cs-ai
29 May 2026
Model Releases

Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset

DGX agent

arXiv:2605.29365v1 Announce Type: new Abstract: Formality transfer is commonly framed as a symmetric bidirectional task between informal and formal registers. We argue that this framing conceals a sup

model-releasesarxiv-cs-cl
29 May 2026
Research

Causal Label Recovery in Payment Networks

DGX agent

arXiv:2605.29272v1 Announce Type: cross Abstract: Fraud detection models in payment networks train on chargeback labels that are systematically biased. Every label must survive three sequential gates:

researcharxiv-cs-ai
29 May 2026
Model Releases

CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild

DGX agent

arXiv:2605.30241v1 Announce Type: new Abstract: Misinformation verification increasingly occurs in public, fast-moving, and multilingual online settings, where static benchmarks provide an incomplete

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Comparative Evaluation of Machine Translation Systems on Images with Text

DGX agent

arXiv:2605.29476v1 Announce Type: new Abstract: This work presents a comparative evaluation of machine translation systems applied to images containing textual information, a task that lies at the int

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

COMPOSE: Composing Future Theorems from Citations and Formal Structure

DGX agent

arXiv:2605.30333v1 Announce Type: new Abstract: A plausible future mathematical claim must satisfy two constraints: it should follow the direction of prior work and respect the formal dependencies tha

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Demystifying Scientific Problem-Solving in LLMs by Probing Knowledge and Reasoning

DGX agent

arXiv:2508.19202v3 Announce Type: replace Abstract: Scientific problem solving poses unique challenges for LLMs, requiring both deep domain knowledge and the ability to apply such knowledge through co

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Developer's guide to Gemini Enterprise and A2UI integration

DGX agent

If you've built a chatbot, you know this conversation: User: 'Book a table for two tomorrow at 7pm.' Agent: 'Okay, for what day?' User: 'Tomorrow.' Agent: 'What time?' A date picker would have ended t

model-releasesgoogle-cloud-ai
29 May 2026
Model Releases

Diffusion differentiable resampling

DGX agent

arXiv:2512.10401v3 Announce Type: replace-cross Abstract: This paper is concerned with differentiable resampling in the context of sequential Monte Carlo (e.g., particle filtering). Drawing on reparam

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark

DGX agent

arXiv:2605.30027v1 Announce Type: new Abstract: Multimodal documents contain diverse elements, such as tables, figures, and layouts, which can complicate retrieval tasks. While current approaches typi

model-releasesarxiv-cs-cv
29 May 2026
Safety

EAPO: Enhancing Policy Optimization with On-Demand Expert Assistance

DGX agent

arXiv:2509.23730v2 Announce Type: replace Abstract: Large language models (LLMs) have recently advanced in reasoning when optimized with reinforcement learning (RL) under verifiable rewards. Existing

safetyarxiv-cs-ai
29 May 2026
Research

Efficient Training-Free Multi-Token Prediction via Embedding-Space Probing

DGX agent

arXiv:2603.17942v2 Announce Type: replace Abstract: Large Language Models (LLMs) possess latent multi-token prediction (MTP) abilities despite being trained only for next-token generation. We introduc

researcharxiv-cs-cl
29 May 2026
Tutorials

Explaining Concept Shift with Interpretable Feature Attribution

DGX agent

arXiv:2505.20634v2 Announce Type: replace Abstract: Concept shift occurs when the distribution of labels conditioned on the features changes between domains, which can make even a well-tuned ML model

tutorialsarxiv-cs-lg
29 May 2026
← Previous
1…605606607608609…1371
Next →