AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,574 results
Safety

EgoExo-WM: Unlocking Exo Video for Ego World Models

DGX agent

arXiv:2605.15477v1 Announce Type: new Abstract: Egocentric world models present a promising direction for enabling agents to predict and plan, but their performance is constrained by the limited avail

safetyarxiv-cs-cv
18 May 2026
Research

Extrapolation Guarantees for Perturbation Modeling Under the Additive Latent Shift Assumption

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2504.18522v3 Announce Type: replace-cross Abstract: We consider the problem of modeling the effects of perturbations like gene knockouts on measurements such as single-cell RNA counts. Given dat

researcharxiv-cs-lg
18 May 2026
Model Releases

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

DGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

model-releasesarxiv-cs-cl
18 May 2026
Research

Few-Step Diffusion Language Models via Trajectory Self-Distillation

DGX agent

arXiv:2602.12262v3 Announce Type: replace Abstract: Diffusion large language models (DLLMs) have emerged as powerful generative models with the promise of fast text generation through parallel decodin

researcharxiv-cs-cl
18 May 2026
Model Releases

Large Language Models Could Be Rote Learners

DGX agent

arXiv:2504.08300v5 Announce Type: replace-cross Abstract: Benchmark-based evaluation, e.g., multiple-choice questions (MCQs) and open-ended questions (OEQs), is widely used for evaluating Large Langua

model-releasesarxiv-cs-ai
18 May 2026
Applications

Metropolis-Scale Road Network Datasets for Fine-Grained Urban Traffic Modeling

DGX agent

arXiv:2510.02278v2 Announce Type: replace Abstract: Modeling traffic dynamics is a critical challenge for urban computing, with applications from real-time traffic management to infrastructure plannin

applicationsarxiv-cs-lg
18 May 2026
Research

PanoWorld: Geometry-Consistent Panoramic Video World Modeling

DGX agent

arXiv:2605.15391v1 Announce Type: cross Abstract: We present PanoWorld, a panoramic video world model that generates geometry-consistent 360egree video from a single image and a caption. Existing pano

researcharxiv-cs-ai
18 May 2026
Research

Prompt Stability Scoring for Text Annotation with Large Language Models

DGX agent

arXiv:2407.02039v3 Announce Type: replace Abstract: Researchers are increasingly using language models (LMs) for text annotation. These approaches rely only on a prompt telling the model to return a g

researcharxiv-cs-cl
18 May 2026
Model Releases

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

DGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

model-releasesarxiv-cs-ai
18 May 2026
Research

Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models

DGX agent

arXiv:2605.15961v1 Announce Type: new Abstract: Large-scale pre-trained vision-language models like CLIP demonstrate remarkable zero-shot performance across diverse tasks. However, fine-tuning these m

researcharxiv-cs-cv
18 May 2026
Research

Testing properties of trees in graphical models with covariance queries

DGX agent

arXiv:2605.15996v1 Announce Type: cross Abstract: We consider the problem of testing properties of graphs underlying high-dimensional graphical models. We adopt the model of covariance queries introdu

researcharxiv-cs-lg
18 May 2026
Industry

the team did an internal test of this model last week the whole company (bar a few exceptions) had all their cursor chats redirected to comp…

DGX agent

the team did an internal test of this model last week the whole company (bar a few exceptions) had all their cursor chats redirected to composer 2.5 for like 2 days. i didn't even notice, which I thin

industryelon-musk--x
18 May 2026
Safety

💯. Way too much focus on language models.

DGX agent

💯. Way too much focus on language models. Fei-Fei Li warns that AI may be staring too hard at language models. The world is not just text on a screen. It is physical, visual, spatial, and always chang

safetygary-marcus--x
16 May 2026
Research

AIM-DDI: A Model-Agnostic Multimodal Integration Module for Drug-Drug Interaction Prediction

DGX agent

arXiv:2605.14327v1 Announce Type: cross Abstract: Drug-drug interaction (DDI) prediction is a critical task in computational biomedicine, as adverse interactions between co-administered drugs can caus

researcharxiv-cs-ai
15 May 2026
Safety

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

DGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

safetyarxiv-cs-ai
15 May 2026
Safety

BiSpikCLM: A Spiking Language Model integrating Softmax-Free Spiking Attention and Spike-Aware Alignment Distillation

DGX agent

arXiv:2605.13859v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer promising energy-efficient alternatives to large language models (LLMs) due to their event-driven nature and ultr

safetyarxiv-cs-ai
15 May 2026
Model Releases

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning

DGX agent

arXiv:2506.05762v5 Announce Type: replace Abstract: Recent advances in offline Reinforcement Learning (RL) have proven that effective policy learning can benefit from imposing conservative constraints

model-releasesarxiv-cs-lg
15 May 2026
Local Ai

Breaking the Reasoning Horizon in Entity Alignment Foundation Models

DGX agent

arXiv:2601.21174v2 Announce Type: replace Abstract: Entity alignment (EA) is critical for knowledge graph (KG) fusion. Existing EA models lack transferability and are incapable of aligning unseen KGs

local-aiarxiv-cs-lg
15 May 2026
Safety

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

DGX agent

arXiv:2605.14544v1 Announce Type: new Abstract: Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this lab

safetyarxiv-cs-ai
15 May 2026
Applications

Enhanced and Efficient Reasoning in Large Learning Models

DGX agent

arXiv:2605.14036v1 Announce Type: new Abstract: In current Large Language Models we can trust the production of smoothly flowing prose on the basis of the principles of machine learning. However, ther

applicationsarxiv-cs-ai
15 May 2026
Tools

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building.

DGX agent

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building. LangSmith Fleet now has a free model powered by @FireworksAI_HQ for Developer and Plus plans. It’

toolsfireworks-ai--x
15 May 2026
Model Releases

GhostCite: A Large-Scale Analysis of Citation Validity in the Age of Large Language Models

DGX agent

arXiv:2602.06718v2 Announce Type: replace-cross Abstract: Citations provide the basis for trusting scientific claims; when they are invalid or fabricated, this trust collapses. With the advent of Larg

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

How Sensitive Are Radiomic AI Models to Acquisition Parameters?

DGX agent

arXiv:2605.14667v1 Announce Type: new Abstract: A main barrier for the deployment of AI radiomic systems in clinical routine is their drop in performance under heterogeneous multicentre acquisition pr

model-releasesarxiv-cs-ai
15 May 2026
Applications

ImmuVis: Hyperconvolutional Foundation Model for Imaging Mass Cytometry

DGX agent

arXiv:2602.04585v2 Announce Type: replace Abstract: We present ImmuVis, a family of efficient foundation models for imaging mass cytometry (IMC), a high-throughput multiplex imaging technology that ha

applicationsarxiv-cs-cv
15 May 2026
Model Releases

K-Models: a Flexible and Interpretable Method for Ordinal Clustering with Application to Antigen-Antibody Interaction Profiles

DGX agent

arXiv:2605.14828v1 Announce Type: cross Abstract: Existing clustering methods for functional data often prioritize partitioning accuracy over interpretability, making it challenging to extract meaning

model-releasesarxiv-cs-lg
15 May 2026
Local Ai

KGPFN: Unlocking the Potential of Knowledge Graph Foundation Model via In-Context Learning

DGX agent

arXiv:2605.14907v1 Announce Type: new Abstract: Knowledge graph (KG) foundation models aim to generalize across graphs with unseen entities and relations by learning transferable relational structure.

local-aiarxiv-cs-ai
15 May 2026
Research

M^2RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling

DGX agent

arXiv:2603.14360v2 Announce Type: replace-cross Abstract: Transformers are highly parallel but are limited to computations in the TC^0 complexity class, excluding tasks such as entity tracking and cod

researcharxiv-cs-ai
15 May 2026
Safety

Measuring and Mitigating Toxicity in Large Language Models: A Comprehensive Replication Study

DGX agent

arXiv:2605.14087v1 Announce Type: new Abstract: Large Language Models (LLMs), when trained on web-scale corpora, inherently absorb toxic patterns from their training data. This leads to ``toxic degene

safetyarxiv-cs-cl
15 May 2026
Safety

Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models

DGX agent

arXiv:2605.14530v1 Announce Type: new Abstract: Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding fo

safetyarxiv-cs-cv
15 May 2026
Research

Pelican-Unified 1.0: A Unified Embodied Intelligence Model for Understanding, Reasoning, Imagination and Action

DGX agent

arXiv:2605.15153v1 Announce Type: cross Abstract: We present Pelican-Unified 1.0, the first embodied foundation model trained according to the principle of unification. Pelican-Unified 1.0 uses a sing

researcharxiv-cs-ai
15 May 2026
Local Ai

Pixal3D: Generate high-fidelity 3D assets from a single image. (TencentARC, locally runnable model)

DGX agent

Pixal3D is a locally runnable AI model developed by TencentARC that generates high-fidelity 3D assets from single 2D images. The tool leverages advanced techniques to convert 2D image inputs into deta

local-air-stablediffusion
15 May 2026
Model Releases

Polaris: A Godel Agent Framework for Small Language Models through Experience-Abstracted Policy Repair

DGX agent

arXiv:2603.23129v2 Announce Type: replace Abstract: Godel agent realize recursive self-improvement: an agent inspects its own policy and traces and then modifies that policy in a tested loop. We intro

model-releasesarxiv-cs-lg
15 May 2026
Safety

RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO

DGX agent

arXiv:2605.15190v1 Announce Type: new Abstract: Causal autoregressive video diffusion models support real-time streaming generation by extrapolating future chunks from previously generated content. Di

safetyarxiv-cs-cv
15 May 2026
Research

Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity

DGX agent

arXiv:2605.14075v1 Announce Type: cross Abstract: Large language models (LLMs) have revolutionized natural language processing. Understanding their internal mechanisms is crucial for developing more i

researcharxiv-cs-cl
15 May 2026
Applications

SurF: A Generative Model for Multivariate Irregular Time Series Forecasting

DGX agent

arXiv:2605.14069v1 Announce Type: new Abstract: Irregularly sampled multivariate event streams remain a stubbornly difficult modality for generative modeling: tokenization-based approaches break down

applicationsarxiv-cs-lg
15 May 2026
Research

TAPIOCA: Why Task- Aware Pruning Improves OOD model Capability

DGX agent

arXiv:2605.14738v1 Announce Type: cross Abstract: Recent work has promoted task-aware layer pruning as a way to improve model performance on particular tasks, as shown by TALE. In this paper, we inves

researcharxiv-cs-ai
15 May 2026
Tutorials

To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model

DGX agent

arXiv:2605.14291v1 Announce Type: cross Abstract: The rapid advancement of Large Vision-Language Models (LVLMs) is increasingly accompanied by unauthorized scraping and training on multimodal web data

tutorialsarxiv-cs-ai
15 May 2026
Tutorials

What Do EEG Foundation Models Capture from Human Brain Signals?

DGX agent

arXiv:2605.11410v2 Announce Type: replace Abstract: Clinical electroencephalogram (EEG) analysis rests on a hand-crafted feature catalog refined over decades, e.g., band power, connectivity, complexit

tutorialsarxiv-cs-ai
15 May 2026
Research

Differences in Text Generated by Diffusion and Autoregressive Language Models

DGX agent

arXiv:2605.12522v1 Announce Type: cross Abstract: Diffusion language models (DLMs) are promising alternatives to autoregressive language models (ARMs), yet the intrinsic differences in their generated

researcharxiv-cs-ai
14 May 2026
Research

Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation

DGX agent

arXiv:2605.13122v1 Announce Type: new Abstract: Instruction-based image editing (IIE) models have recently demonstrated strong capability in modifying specific image regions according to natural langu

researcharxiv-cs-cv
14 May 2026
Research

Exemplar-Free Continual Learning for State Space Models

DGX agent

arXiv:2505.18604v3 Announce Type: replace Abstract: State-Space Models (SSMs) excel at capturing long-range dependencies with structured recurrence, making them well-suited for sequence modeling. Howe

researcharxiv-cs-lg
14 May 2026
Research

Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants

DGX agent

arXiv:2512.10857v2 Announce Type: replace-cross Abstract: Transport-based methods have emerged as a leading paradigm for building generative models from large, clean datasets. However, in many scienti

researcharxiv-cs-ai
14 May 2026
Research

Generative Modeling of Approximately Periodic Time Series by a Posterior-Weighted Gaussian Process

DGX agent

arXiv:2605.13150v1 Announce Type: cross Abstract: Discrete automated processes in industrial and cyber-physical systems often exhibit a repetitive structure in which successive repetitions follow a co

researcharxiv-cs-lg
14 May 2026
Model Releases

Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2605.13632v1 Announce Type: cross Abstract: In this paper, we propose GTA-VLA(Guide, Think, Act), an interactive Vision-Language-Action (VLA) framework that enables spatially steerable embodied

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Inference-Time Machine Unlearning via Gated Activation Redirection

DGX agent

arXiv:2605.12765v1 Announce Type: new Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning

model-releasesarxiv-cs-lg
14 May 2026
Safety

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

DGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

safetyarxiv-cs-ai
14 May 2026
Safety

Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models

DGX agent

arXiv:2605.12725v1 Announce Type: new Abstract: Recent video anomaly detection research has expanded rapidly with an emphasis on general models of normality intended to work across many different scen

safetyarxiv-cs-cv
14 May 2026
Applications

Language Model Networks: Supervision-Efficient Learning through Dense Communication

DGX agent

arXiv:2505.12741v2 Announce Type: replace Abstract: Language models are increasingly used not only as standalone predictors but also as components in larger inference systems, from test-time reasoning

applicationsarxiv-cs-ai
14 May 2026
← Previous
1…171172173174175…1262
Next →