AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,552 results
Safety

D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models

DGX agent

arXiv:2605.19690v1 Announce Type: new Abstract: Navigation Foundation Models (NFMs) trained on large cross-embodied datasets have demonstrated powerful generalizability in various scenarios. Adopting

safetyarxiv-cs-ro
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

DGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

model-releasesarxiv-cs-cl
20 May 2026
Agents

High-quality generation of dynamic game content via small language models: A proof of concept

DGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

agentsarxiv-cs-ai
20 May 2026
Model Releases

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

DGX agent

arXiv:2605.19301v1 Announce Type: new Abstract: Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastroph

model-releasesarxiv-cs-cv
20 May 2026
Research

Neural Operators for Design-Space Surrogate Modeling of Tendon-Actuated Continuum Robots

DGX agent

arXiv:2605.19104v1 Announce Type: cross Abstract: Continuum robots enable dexterous manipulation in constrained environments, but require accurate and efficient models for real-time manipulation and c

researcharxiv-cs-ai
20 May 2026
Research

Probabilistic Tiny Recursive Model

DGX agent

arXiv:2605.19943v1 Announce Type: new Abstract: Tiny Recursive Models (TRM) solve complex reasoning tasks with a fraction of the parameters of modern large language models (LLMs) by iteratively refini

researcharxiv-cs-ai
20 May 2026
Model Releases

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

DGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

model-releasesarxiv-cs-ai
20 May 2026
Local Ai

Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models

DGX agent

arXiv:2605.20158v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) show promise in medical applications, but their inability to faithfully ground responses in visual evidence raise

local-aiarxiv-cs-ai
20 May 2026
Model Releases

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

DGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

model-releasesarxiv-cs-ai
20 May 2026
Research

Shaping the Prior: How Synthetic Task Distributions Determine Tabular Foundation Model Quality

DGX agent

arXiv:2605.18971v1 Announce Type: cross Abstract: What determines the quality of a tabular foundation model? Unlike language or vision, tabular foundation models acquire their inductive biases almost

researcharxiv-cs-ai
20 May 2026
Model Releases

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

DGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TADA! Tuning Audio Diffusion Models through Activation Steering

DGX agent

arXiv:2602.11910v2 Announce Type: replace-cross Abstract: Audio diffusion models can synthesize high-fidelity music from text, yet achieving fine-grained control over specific musical attributes remai

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular,…

DGX agent

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular, and represents an important milestone for the math and AI c

model-releasesopenai--x
20 May 2026
Industry

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry p…

DGX agent

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry problems. Less than 1 year ago frontier AI models were at IMO

industrysam-altman--x
20 May 2026
Tutorials

Tweedie's Formulae and Diffusion Generative Models Beyond Gaussian

DGX agent

arXiv:2605.19391v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success in generating samples from unknown data distributions. Most popular stochastic differential equation

tutorialsarxiv-cs-lg
20 May 2026
Model Releases

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗

DGX agent

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗 Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it t

model-releasesclem-delangue--x
20 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Safety

Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States

DGX agent

arXiv:2605.17144v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models leverage powerful perceptual priors from web-scale Vision-Language Model (VLM) pre-training, yet they remain surpr

safetyarxiv-cs-ai
19 May 2026
Model Releases

CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark

DGX agent

arXiv:2605.18621v1 Announce Type: cross Abstract: Spatial intelligence requires multimodal large language models (MLLMs) to move beyond single-view perception and reason consistently about objects, vi

model-releasesarxiv-cs-ai
19 May 2026
Research

Data-Driven Dynamic Modeling of a Tendon-Actuated Continuum Robot

DGX agent

arXiv:2605.18720v1 Announce Type: new Abstract: Developing dynamic models for tendon-driven continuum robots is challenging due to their nonlinear, high-dimensional, and friction-dominated dynamics. T

researcharxiv-cs-ro
19 May 2026
Tutorials

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a mode…

DGX agent

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a model? tweaking prompts, tool names, and tool descriptions in ac

tutorialsharrison-chase--x
19 May 2026
Safety

Diffusion Models, Denoiser Architecture and Creativity

DGX agent

arXiv:2605.16415v1 Announce Type: new Abstract: The creativity of diffusion models refers to their ability to generate highly realistic images that are different from their training data. Creativity i

safetyarxiv-cs-cv
19 May 2026
Applications

DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object Dynamics

DGX agent

arXiv:2603.09668v2 Announce Type: replace Abstract: Modeling wind-driven object dynamics from video observations is highly challenging due to the invisibility and spatio-temporal variability of wind,

applicationsarxiv-cs-cv
19 May 2026
Safety

Dual-Space Knowledge Distillation with Key-Query Matching for Large Language Models with Vocabulary Mismatch

DGX agent

arXiv:2603.22056v2 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art (SOTA) performance across language tasks, but are costly to deploy due to their size and resou

safetyarxiv-cs-cl
19 May 2026
Model Releases

Dynamic Elliptical Graph Factor Models via Riemannian Optimization with Geodesic Temporal Regularization

DGX agent

arXiv:2605.18316v1 Announce Type: new Abstract: Inferring time-varying graph structures from high-dimensional nodal observations is a fundamental problem arising in neuroscience, finance, climatology,

model-releasesarxiv-cs-lg
19 May 2026
Safety

ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation

DGX agent

arXiv:2605.17580v1 Announce Type: new Abstract: Electrocardiogram (ECG)-based models have achieved strong performance in diagnostic tasks, yet they remain limited in modeling how cardiac dynamics evol

safetyarxiv-cs-ai
19 May 2026
Model Releases

Embodied Task Planning via Graph-Informed Action Generation with Large Language Models

DGX agent

arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Ensembling Tabular Foundation Models - A Diversity Ceiling And A Calibration Trap

DGX agent

arXiv:2605.18696v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) now match or beat tuned gradient-boosted trees on a growing fraction of tabular tasks, but no single TFM wins on ever

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Everything Google Cloud customers need to know coming out of Google I/O

DGX agent

At Google Cloud Next ‘26, we unveiled the blueprint for the Agentic Enterprise, sharing our eighth-generation TPUs, Gemini Enterprise Agent Platform, a fully reimagined Agentic Data Cloud, Workspace I

model-releasesgoogle-cloud-ai
19 May 2026
Safety

FLAG: Foundation model representation with Latent diffusion Alignment via Graph for spatial gene expression prediction

DGX agent

arXiv:2605.18055v1 Announce Type: cross Abstract: Predicting spatial gene expression from routine H&E enables large-scale molecular profiling, yet current models treat this as isolated pointwise tasks

safetyarxiv-cs-ai
19 May 2026
Model Releases

Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models

DGX agent

arXiv:2508.06038v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) incur substantial computational overhead and inference latency due to the large number of vision tokens introduc

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion

DGX agent

arXiv:2605.17759v1 Announce Type: new Abstract: To circumvent the inherent fidelity bottlenecks and optimization misalignment of VAE-based latent diffusion, pixel-space diffusion models have emerged a

local-aiarxiv-cs-cv
19 May 2026
Model Releases

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

DGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

DGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics

DGX agent

arXiv:2603.02531v2 Announce Type: replace-cross Abstract: Classifier-Free Guidance (CFG) improves sample quality in diffusion models, but its dual-pass inference and reliance on null-condition trainin

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app (Sanuj Bhatia/Android Central)

DGX agent

Sanuj Bhatia / Android Central: Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app — Seriously, who is asking for

model-releasestechmeme
19 May 2026
Local Ai

How Many Visual Tokens Do Multimodal Language Models Need? Scaling Visual Token Pruning with F^3A

DGX agent

arXiv:2605.16359v1 Announce Type: cross Abstract: Vision-language models improve perception by feeding increasingly long visual token sequences into language backbones, but the resulting inference cos

local-aiarxiv-cs-ai
19 May 2026
Safety

Lance: Unified Multimodal Modeling by Multi-Task Synergy

DGX agent

arXiv:2605.18678v1 Announce Type: cross Abstract: We present Lance, a lightweight native unified model supporting multimodal understanding, generation, and editing for both images and videos. Rather t

safetyarxiv-cs-ai
19 May 2026
Research

Large Language Models and Impossible Language Acquisition: 'False Promise' or an Overturn of our Current Perspective towards AI

DGX agent

arXiv:2602.08437v5 Announce Type: replace Abstract: In Chomsky's provocative critique 'The False Promise of CHATGPT,' Large Language Models (LLMs) are characterized as mere pattern predictors that do

researcharxiv-cs-cl
19 May 2026
Agents

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate

DGX agent

arXiv:2601.22297v2 Announce Type: replace Abstract: The reasoning abilities of large language models (LLMs) have been substantially improved by reinforcement learning with verifiable rewards (RLVR). A

agentsarxiv-cs-cl
19 May 2026
Model Releases

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

DGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

DGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

model-releasesarxiv-cs-ai
19 May 2026
Research

Prune, Update and Trim: Robust Structured Pruning for Large Language Models

DGX agent

arXiv:2605.18331v1 Announce Type: new Abstract: Large Language Models (LLMs) have experienced significant growth and development in recent years. However, performing inference on LLMs remains costly,

researcharxiv-cs-lg
19 May 2026
Model Releases

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models

DGX agent

arXiv:2501.17549v2 Announce Type: replace Abstract: Graph-structured data plays a vital role in numerous domains, such as social networks, citation networks, commonsense reasoning graphs and knowledge

model-releasesarxiv-cs-cl
19 May 2026
Applications

Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift

DGX agent

arXiv:2605.16411v1 Announce Type: cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible y

applicationsarxiv-cs-ai
19 May 2026
Safety

Retrieval and competition: how a protein foundation model starts a protein

DGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

safetyarxiv-cs-ai
19 May 2026
Model Releases

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data

DGX agent

arXiv:2605.18287v1 Announce Type: new Abstract: It is infeasible to encompass all possible disturbances within the training dataset. This raises a critical question regarding the robustness of Vision-

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

DGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…140141142143144…1262
Next →