AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

A Bitter Lesson for Data Filtering

DGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

DGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

model-releasesarxiv-cs-cv
20 May 2026
Agents

AQuaUI: Visual Token Reduction for GUI Agents with Adaptive Quadtrees

DGX agent

arXiv:2605.19260v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently emerged as promising backbones for GUI-agent models, where high-resolution GUI screenshots are introduced t

agentsarxiv-cs-ai
20 May 2026
Research

Block-Based Double Decoders

DGX agent

arXiv:2605.18807v1 Announce Type: cross Abstract: Encoder-decoder models offer substantial inference-time savings over decoder-only models, but their pretraining objectives suffer from sparse supervis

researcharxiv-cs-ai
20 May 2026
Model Releases

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and …

DGX agent

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and encounter. They’re not really random benchmark tasks, they r

model-releasesharrison-chase--x
20 May 2026
Research

Cross-Paradigm Knowledge Distillation: A Comprehensive Study of Bidirectional Transfer Between Random Forests and Deep Neural Networks for Big Data Applications

DGX agent

arXiv:2605.19299v1 Announce Type: new Abstract: The exponential growth of big data has intensified the need for efficient and interpretable machine learning models that can handle diverse data charact

researcharxiv-cs-lg
20 May 2026
Tutorials

Diffusion and Flow-based Copulas: Forgetting and Remembering Dependencies

DGX agent

arXiv:2509.19707v2 Announce Type: replace-cross Abstract: Copulas are a fundamental tool for modelling multivariate dependencies in data, forming the method of choice in diverse fields and application

tutorialsarxiv-cs-lg
20 May 2026
Model Releases

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

DGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

model-releasesarxiv-cs-ai
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Model Releases

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

DGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

DGX agent

arXiv:2605.19824v1 Announce Type: new Abstract: Recent attempts to support high-level scene interpretation and planning in Autonomous Vehicles (AVs) using ensembles of Large Language Models (LLMs) and

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

GRAB: A Risk Taxonomy--Grounded Benchmark for Unsupervised Topic Discovery in Financial Disclosures

DGX agent

arXiv:2509.21698v2 Announce Type: replace Abstract: Risk categorization in 10-K risk disclosures matters for oversight and investment, yet no public benchmark evaluates unsupervised topic models for t

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Hallucination as Exploit: Evidence-Carrying Multimodal Agents

DGX agent

arXiv:2605.19192v1 Announce Type: new Abstract: Multimodal agents use screenshots, documents, and webpages to choose tool calls. When a false visual claim triggers a click, email, extraction, or trans

model-releasesarxiv-cs-ai
20 May 2026
Safety

How Does Overparameterization Affect Machine Unlearning of Deep Neural Networks?

DGX agent

arXiv:2503.08633v2 Announce Type: replace Abstract: Machine unlearning is the task of updating a trained model to forget specific training data without retraining from scratch. In this paper, we inves

safetyarxiv-cs-lg
20 May 2026
Local Ai

INSIGHTS: Demonstration-Based Summaries of Time Series Predictors

DGX agent

arXiv:2605.18849v1 Announce Type: cross Abstract: Explainability methods have progressed rapidly, but global explanations for time-series models remain underdeveloped, with most approaches focusing on

local-aiarxiv-cs-ai
20 May 2026
Model Releases

Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiency

DGX agent

arXiv:2605.19008v1 Announce Type: new Abstract: Modern language-model training is increasingly exposed to instability, degraded runs, and wasted compute, especially under aggressive learning-rate, sca

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Measuring Safety Alignment Effects in Autonomous Security Agents

DGX agent

arXiv:2605.19722v1 Announce Type: cross Abstract: Do stock safety-aligned language models and their uncensored or abliterated derivatives behave differently when run as autonomous security agents? Sin

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

DGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

DGX agent

arXiv:2605.18825v1 Announce Type: new Abstract: Prefix caching is a key optimization in Large Language Model (LLM) serving, reusing attention Key-Value (KV) states across requests with shared prompt p

model-releasesarxiv-cs-lg
20 May 2026
Local Ai

Perceptual misalignment of texture representations in convolutional neural networks

DGX agent

arXiv:2604.01341v2 Announce Type: replace Abstract: Mathematical modeling of visual textures traces back to Julesz's intuition that texture perception in humans is based on local correlations between

local-aiarxiv-cs-cv
20 May 2026
Model Releases

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, an…

DGX agent

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, and across systems. We’re unlocking new possibilities for deve

model-releasesclem-delangue--x
20 May 2026
Model Releases

Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise

DGX agent

arXiv:2605.18832v1 Announce Type: cross Abstract: The Transformer is the foundational building block of modern AI, yet offers no principled handling of uncertainty, which is prevalent in real applicat

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

PRISM: A Benchmark for Programmatic Spatial-Temporal Reasoning

DGX agent

arXiv:2605.19382v1 Announce Type: new Abstract: Programmatic video generation through code offers geometric precision and temporal coherence beyond pixel-level diffusion models, yet rigorously evaluat

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

DGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

model-releasesallie-k--miller--x
20 May 2026
Research

Robustness and Regularization in Hierarchical Re-Basin

DGX agent

arXiv:2510.09174v3 Announce Type: replace Abstract: This paper takes a closer look at Git Re-Basin, an interesting new approach to merge trained models. We propose a hierarchical model merging scheme

researcharxiv-cs-lg
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions

DGX agent

arXiv:2605.19895v1 Announce Type: new Abstract: Constraint programming practitioners accelerate hard problems through a layered set of techniques applied in order of risk. Standard hardening (symmetry

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding

DGX agent

arXiv:2605.19866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) parse documents end-to-end but frequently break down on layouts unlike those seen in training. We attribute this to a two-

model-releasesarxiv-cs-cv
20 May 2026
Hardware

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

DGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

hardwarearxiv-cs-cl
20 May 2026
Model Releases

Toto 2.0: Time Series Forecasting Enters the Scaling Era

DGX agent

arXiv:2605.20119v1 Announce Type: cross Abstract: We show that time series foundation models scale: a single training recipe produces reliable forecast-quality improvements from 4M to 2.5B parameters.

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs

DGX agent

arXiv:2605.19528v1 Announce Type: new Abstract: 3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera i

model-releasesarxiv-cs-cv
20 May 2026
Safety

Towards Distillation Guarantees under Algorithmic Alignment for Combinatorial Optimization

DGX agent

arXiv:2605.20074v1 Announce Type: new Abstract: Distillation transfers knowledge from a large model trained on broad data to a smaller, more efficient model suitable for deployment. In structured pred

safetyarxiv-cs-lg
20 May 2026
Model Releases

Trust or Abstain? A Self-Aware RAG Approach

DGX agent

arXiv:2605.18792v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves large language models (LLMs) by incorporating external evidence, but it also introduces knowledge confli

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins

DGX agent

arXiv:2603.06740v2 Announce Type: replace-cross Abstract: Protein language models (pLMs) have shown strong potential for zero-shot prediction of missense variant effects, yet systematic benchmarking o

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

WARC-Bench: Web Archive Based Benchmark for GUI Subtask Executions

DGX agent

arXiv:2510.09872v2 Announce Type: replace-cross Abstract: Training web agents to navigate complex, real-world websites requires them to master extit{subtasks} - short-horizon interactions on multiple

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Makes a Representation Good for Single-Cell Perturbation Prediction?

DGX agent

arXiv:2605.19343v1 Announce Type: new Abstract: Single-cell perturbation modeling is fundamental for understanding and predicting cellular responses to genetic perturbations. However, existing approac

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Worldlier: native support for 48 world languages and improved efficiency in non-European languages.

DGX agent

Worldlier is a Cohere language model that natively supports 48 world languages with improved efficiency, particularly for non-European languages. This represents an expansion of language coverage beyo

model-releasescohere--x
20 May 2026
Local Ai

Your Neighbors Know: Leveraging Local Neighborhoods for Backdoor Detection in Decentralized Learning

DGX agent

arXiv:2605.19969v1 Announce Type: new Abstract: Decentralized learning (DL) is an emerging machine learning paradigm where nodes collaboratively train models without a central server. However, the col

local-aiarxiv-cs-lg
20 May 2026
Model Releases

A Machine with Short-Term, Episodic, and Semantic Memory Systems

DGX agent

arXiv:2212.02098v5 Announce Type: replace Abstract: Inspired by the cognitive science theory of the explicit human memory systems, we have modeled an agent with short-term, episodic, and semantic memo

model-releasesarxiv-cs-ai
19 May 2026
Research

A More Word-like Image Tokenization for MLLMs

DGX agent

arXiv:2605.17954v1 Announce Type: cross Abstract: Modern multimodal large language models (MLLMs) typically keep the language model fixed and train a visual projector that maps the pixels into a seque

researcharxiv-cs-ai
19 May 2026
Model Releases

Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring

DGX agent

arXiv:2605.16386v1 Announce Type: new Abstract: Multimodal large language models (LLMs) are increasingly explored as automated evaluators in clinical settings, yet their scoring behavior on ordinal cl

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

BESplit: Bias-Compensated Split Federated Learning with Evidential Aggregation

DGX agent

arXiv:2605.17508v1 Announce Type: cross Abstract: Split Federated Learning (SFL) enables privacy-preserving collaborative training by partitioning models between clients and a server. However, under n

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Point-Wise Matching: Structural Representation Alignment for Accelerating Diffusion Transformers

DGX agent

arXiv:2605.16949v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) demonstrate that aligning noisy latent states with well-trained semantic features-as pioneered by Repre

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Boundedly Rational Meta-Learning in Sequential Consumer Choice

DGX agent

arXiv:2605.16532v1 Announce Type: new Abstract: Many consumer decisions are repeated choices under uncertainty. Standard models capture these decisions using Bayesian learning and dynamic programming:

model-releasesarxiv-cs-lg
19 May 2026
Research

CADS: Conformal Adaptive Decision System for Cost-Efficient Image Classification

DGX agent

arXiv:2605.16401v1 Announce Type: new Abstract: While high-capacity AI models have advanced state-of-the-art performance, their practical deployment is often hindered by high inference costs, environm

researcharxiv-cs-cv
19 May 2026
Model Releases

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean

DGX agent

arXiv:2605.17255v1 Announce Type: new Abstract: Formal theorem-proving benchmarks enable mechanically verifiable evaluation of mathematical reasoning in large language models. However, existing benchm

model-releasesarxiv-cs-ai
19 May 2026
Research

Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks

DGX agent

arXiv:2510.01782v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) should refuse to answer questions beyond their knowledge. This capability, which we term knowledge-aware refusal,

researcharxiv-cs-ai
19 May 2026
← Previous
1…463464465466467…1369
Next →