AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

DGX agent

arXiv:2605.31148v1 Announce Type: cross Abstract: Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into ac

model-releasesarxiv-cs-ai
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery

DGX agent

arXiv:2605.31156v1 Announce Type: new Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding a

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Targeted Speaker Poisoning Framework in Zero-Shot Text-to-Speech

DGX agent

arXiv:2603.07551v2 Announce Type: replace-cross Abstract: Zero-shot Text-to-Speech (TTS) voice cloning poses severe privacy risks, demanding the removal of specific speaker identities from trained TTS

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy Expansion

DGX agent

arXiv:2601.09633v2 Announce Type: replace Abstract: Taxonomies form the backbone of structured knowledge representation across diverse domains, enabling applications such as e-commerce and semantic se

model-releasesarxiv-cs-cl
1 Jun 2026
Research

The Gaussian-Head OFL Family: One-Shot Federated Learning from Client Global Statistics

DGX agent

arXiv:2602.01186v2 Announce Type: replace-cross Abstract: Classical Federated Learning relies on a multi-round iterative process of model exchange and aggregation between server and clients, with high

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing

DGX agent

arXiv:2605.31367v1 Announce Type: cross Abstract: Token mixing layers play a key role in how language models can learn and generate long-range dependencies. Their efficiency relies on the necessary tr

tutorialsarxiv-cs-cl
1 Jun 2026
Model Releases

TSM-Bench: Detecting LLM-Generated Text in Real-World Wikipedia Editing Practices

DGX agent

arXiv:2605.31113v1 Announce Type: new Abstract: Automatically detecting machine-generated text (MGT) is critical to maintaining the knowledge integrity of user-generated content (UGC) platforms such a

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

VeriGate: Verifier-Gated Step-Level Supervision for GRPO

DGX agent

arXiv:2605.30451v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is an effective recipe for training reasoning models with verifier-based outcome rewards, but its supervision

safetyarxiv-cs-lg
1 Jun 2026
Model Releases

WristCompass: Kinematic Coupling as a Learnable Visual Concept for Ego-Camera Orientation

DGX agent

arXiv:2605.30671v1 Announce Type: new Abstract: Recovering ego-camera orientation from manipulation video is a prerequisite for disentangling hand motion from camera motion, a key step in imitation le

model-releasesarxiv-cs-cv
1 Jun 2026
Research

4DPC^2hat: Towards Dynamic Point Cloud Understanding with Failure-Aware Bootstrapping

DGX agent

arXiv:2602.03890v2 Announce Type: replace Abstract: Point clouds provide a compact and expressive representation of 3D objects, and have recently been integrated into multimodal large language models

researcharxiv-cs-cv
29 May 2026
Model Releases

A Full-Pipeline Framework for Evaluating Membership Inference Attacks in Machine Learning

DGX agent

arXiv:2605.29454v1 Announce Type: new Abstract: While Membership Inference Attacks (MIAs) are the prevailing method for identifying training data, their application has expanded into privacy auditing

model-releasesarxiv-cs-lg
29 May 2026
Research

AdaState: Self-Evolving Anchors for Streaming Video Generation

DGX agent

arXiv:2605.30349v1 Announce Type: new Abstract: Autoregressive video diffusion models generate streaming video by producing frames sequentially, conditioning each chunk on previously generated content

researcharxiv-cs-cv
29 May 2026
Safety

AIRGuard: Guarding Agent Actions with Runtime Authority Control

DGX agent

arXiv:2605.28914v1 Announce Type: cross Abstract: Tool-using language agents turn model decisions into external side effects: they read files, run scripts, call APIs, send messages, and invoke Model C

safetyarxiv-cs-ai
29 May 2026
Model Releases

AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation

DGX agent

arXiv:2512.01334v2 Announce Type: replace Abstract: Text-guided image-to-video generation has made substantial progress, yet it still struggles to execute text-specified edits that require substantial

model-releasesarxiv-cs-cv
29 May 2026
Local Ai

Benchmarking at the Edge of Comprehension

DGX agent

arXiv:2602.14307v3 Announce Type: replace Abstract: As frontier Large Language Models (LLMs) increasingly saturate new benchmarks shortly after they are published, benchmarking itself is at a juncture

local-aiarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

DGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
Research

Beyond MSE: Improving Precipitation Nowcasting with Multi-Quantile Regression

DGX agent

arXiv:2605.30122v1 Announce Type: cross Abstract: Deep-learning precipitation nowcasting models are often optimized using pointwise losses such as mean squared error or mean absolute error, which can

researcharxiv-cs-ai
29 May 2026
Model Releases

Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset

DGX agent

arXiv:2605.29365v1 Announce Type: new Abstract: Formality transfer is commonly framed as a symmetric bidirectional task between informal and formal registers. We argue that this framing conceals a sup

model-releasesarxiv-cs-cl
29 May 2026
Research

Causal Label Recovery in Payment Networks

DGX agent

arXiv:2605.29272v1 Announce Type: cross Abstract: Fraud detection models in payment networks train on chargeback labels that are systematically biased. Every label must survive three sequential gates:

researcharxiv-cs-ai
29 May 2026
Model Releases

CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild

DGX agent

arXiv:2605.30241v1 Announce Type: new Abstract: Misinformation verification increasingly occurs in public, fast-moving, and multilingual online settings, where static benchmarks provide an incomplete

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Comparative Evaluation of Machine Translation Systems on Images with Text

DGX agent

arXiv:2605.29476v1 Announce Type: new Abstract: This work presents a comparative evaluation of machine translation systems applied to images containing textual information, a task that lies at the int

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

COMPOSE: Composing Future Theorems from Citations and Formal Structure

DGX agent

arXiv:2605.30333v1 Announce Type: new Abstract: A plausible future mathematical claim must satisfy two constraints: it should follow the direction of prior work and respect the formal dependencies tha

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Demystifying Scientific Problem-Solving in LLMs by Probing Knowledge and Reasoning

DGX agent

arXiv:2508.19202v3 Announce Type: replace Abstract: Scientific problem solving poses unique challenges for LLMs, requiring both deep domain knowledge and the ability to apply such knowledge through co

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Diffusion differentiable resampling

DGX agent

arXiv:2512.10401v3 Announce Type: replace-cross Abstract: This paper is concerned with differentiable resampling in the context of sequential Monte Carlo (e.g., particle filtering). Drawing on reparam

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark

DGX agent

arXiv:2605.30027v1 Announce Type: new Abstract: Multimodal documents contain diverse elements, such as tables, figures, and layouts, which can complicate retrieval tasks. While current approaches typi

model-releasesarxiv-cs-cv
29 May 2026
Safety

EAPO: Enhancing Policy Optimization with On-Demand Expert Assistance

DGX agent

arXiv:2509.23730v2 Announce Type: replace Abstract: Large language models (LLMs) have recently advanced in reasoning when optimized with reinforcement learning (RL) under verifiable rewards. Existing

safetyarxiv-cs-ai
29 May 2026
Research

Efficient Training-Free Multi-Token Prediction via Embedding-Space Probing

DGX agent

arXiv:2603.17942v2 Announce Type: replace Abstract: Large Language Models (LLMs) possess latent multi-token prediction (MTP) abilities despite being trained only for next-token generation. We introduc

researcharxiv-cs-cl
29 May 2026
Tutorials

Explaining Concept Shift with Interpretable Feature Attribution

DGX agent

arXiv:2505.20634v2 Announce Type: replace Abstract: Concept shift occurs when the distribution of labels conditioned on the features changes between domains, which can make even a well-tuned ML model

tutorialsarxiv-cs-lg
29 May 2026
Model Releases

Fairness-Aware Federated Learning with Trajectory Shapley Value

DGX agent

arXiv:2605.30336v1 Announce Type: new Abstract: Federated learning is an emerging distributed paradigm that addresses the challenges posed by heterogeneous, privacy-sensitive data. It enables multiple

model-releasesarxiv-cs-lg
29 May 2026
Tutorials

Faithful Embeddings of Irregular and Asynchronous Data for Online Log-NCDEs

DGX agent

arXiv:2605.30213v1 Announce Type: new Abstract: Continuous-time models are a natural choice for irregular and asynchronous data. A central design choice is how to embed discrete observations into cont

tutorialsarxiv-cs-lg
29 May 2026
Model Releases

First head-to-head comparison of agentic AI applied to the analysis of simulated data of the Einstein Telescope

DGX agent

arXiv:2605.28916v1 Announce Type: cross Abstract: We report a comparison of two state-of-the-art agentic AI systems, Claude Code (Anthropic) and Codex (OpenAI), tasked with autonomously executing a si

model-releasesarxiv-cs-ai
29 May 2026
Agents

GenClaw: Code-Driven Agentic Image Generation

DGX agent

arXiv:2605.30248v1 Announce Type: new Abstract: Image generation models have evolved from text-conditioned pixel synthesis toward multimodal agents endowed with visual comprehension and tool invocatio

agentsarxiv-cs-cv
29 May 2026
Research

Getting to the Point: Pointing Improves LVLMs at Counting

DGX agent

arXiv:2603.21746v2 Announce Type: replace Abstract: Pointing-based methods decompose complex tasks as sequential grounding and reasoning steps. Given a query, the model first grounds the relevant obje

researcharxiv-cs-cv
29 May 2026
Model Releases

Indexing the Unreadable: LLM-Native Recursive Construction and Search of Service Taxonomies

DGX agent

arXiv:2605.29270v1 Announce Type: new Abstract: The era of the Internet of Agents (IoA) is taking shape: LLM agents are expected to fulfill user goals by orchestrating fast-growing populations of Mode

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback

DGX agent

arXiv:2605.29184v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a promising avenue for scientific discovery, yet their application to symbolic regression is often constrained by i

tutorialsarxiv-cs-ai
29 May 2026
Tutorials

Learn from A Rationalist: Distilling Intermediate Interpretable Rationales

DGX agent

arXiv:2601.22531v2 Announce Type: replace-cross Abstract: Because of the pervasive use of deep neural networks (DNNs), especially in high-stakes domains, the interpretability of DNNs has received incr

tutorialsarxiv-cs-ai
29 May 2026
Safety

LoMo: Local Modality Substitution for Deeper Vision-Language Fusion

DGX agent

arXiv:2605.30265v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved substantial progress across a wide range of understanding and reasoning tasks, driven by large-scale image

safetyarxiv-cs-cl
29 May 2026
Tutorials

Make LLM Learn to Synthesize from Streaming Experiences through Feedback

DGX agent

arXiv:2605.29940v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted for synthetic data generation, significantly reducing annotation costs. However, most existing stu

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference

DGX agent

arXiv:2605.30218v1 Announce Type: new Abstract: Temperature-zero BF16 LLM inference is often treated as reproducible, yet the same request can emit different tokens when decoded alone or inside a larg

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions

DGX agent

arXiv:2605.29862v1 Announce Type: cross Abstract: AI-driven respiratory sound classification (RSC) is promising for automated pulmonary disease detection, yet multi-site deployment is hindered by inte

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Multimodal LLMs See Sentiment

DGX agent

arXiv:2508.16873v3 Announce Type: replace Abstract: Understanding how visual content conveys sentiment is increasingly important in a digital landscape dominated by imagery. However, sentiment percept

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software

DGX agent

arXiv:2605.30353v1 Announce Type: new Abstract: Are AI agents tools, co-authors, or researchers? We present a quantified case study (N=1): a physicist supervising an AI coding agent (Claude Code, Sonn

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data

DGX agent

arXiv:2508.15180v3 Announce Type: replace Abstract: High-quality mathematical and logical datasets with verifiable answers are essential for strengthening the reasoning capabilities of large language

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

REPOT: Recoverable Program-of-Thought via Checkpoint Repair

DGX agent

arXiv:2605.30052v1 Announce Type: cross Abstract: One-shot Program-of-Thought (PoT) emits a Python program that prints a primitive-action plan; a single invalid action silently invalidates the traject

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization

DGX agent

arXiv:2603.27758v2 Announce Type: replace Abstract: Metric Cross-View Geo-Localization (MCVGL) aims to estimate the 3-DoF camera pose (position and heading) by matching ground and satellite images. In

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Risk-averse Fair Multi-class Classification

DGX agent

arXiv:2509.05771v2 Announce Type: replace-cross Abstract: We develop a new classification framework based on the theory of coherent risk measures and systemic risk. The proposed approach is suitable f

model-releasesarxiv-cs-lg
29 May 2026
← Previous
1…499500501502503…1109
Next →