AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
15 May 2026

Polaris: A Godel Agent Framework for Small Language Models through Experience-Abstracted Policy Repair

Model ReleasesDGX agent

arXiv:2603.23129v2 Announce Type: replace Abstract: Godel agent realize recursive self-improvement: an agent inspects its own policy and traces and then modifies that policy in a tested loop. We intro

RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO

SafetyDGX agent

arXiv:2605.15190v1 Announce Type: new Abstract: Causal autoregressive video diffusion models support real-time streaming generation by extrapolating future chunks from previously generated content. Di

Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity

ResearchDGX agent

arXiv:2605.14075v1 Announce Type: cross Abstract: Large language models (LLMs) have revolutionized natural language processing. Understanding their internal mechanisms is crucial for developing more i

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SurF: A Generative Model for Multivariate Irregular Time Series Forecasting

ApplicationsDGX agent

arXiv:2605.14069v1 Announce Type: new Abstract: Irregularly sampled multivariate event streams remain a stubbornly difficult modality for generative modeling: tokenization-based approaches break down

TAPIOCA: Why Task- Aware Pruning Improves OOD model Capability

ResearchDGX agent

arXiv:2605.14738v1 Announce Type: cross Abstract: Recent work has promoted task-aware layer pruning as a way to improve model performance on particular tasks, as shown by TALE. In this paper, we inves

To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model

TutorialsDGX agent

arXiv:2605.14291v1 Announce Type: cross Abstract: The rapid advancement of Large Vision-Language Models (LVLMs) is increasingly accompanied by unauthorized scraping and training on multimodal web data

What Do EEG Foundation Models Capture from Human Brain Signals?

TutorialsDGX agent

arXiv:2605.11410v2 Announce Type: replace Abstract: Clinical electroencephalogram (EEG) analysis rests on a hand-crafted feature catalog refined over decades, e.g., band power, connectivity, complexit

14 May 2026

Differences in Text Generated by Diffusion and Autoregressive Language Models

ResearchDGX agent

arXiv:2605.12522v1 Announce Type: cross Abstract: Diffusion language models (DLMs) are promising alternatives to autoregressive language models (ARMs), yet the intrinsic differences in their generated

Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation

ResearchDGX agent

arXiv:2605.13122v1 Announce Type: new Abstract: Instruction-based image editing (IIE) models have recently demonstrated strong capability in modifying specific image regions according to natural langu

Exemplar-Free Continual Learning for State Space Models

ResearchDGX agent

arXiv:2505.18604v3 Announce Type: replace Abstract: State-Space Models (SSMs) excel at capturing long-range dependencies with structured recurrence, making them well-suited for sequence modeling. Howe

Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants

ResearchDGX agent

arXiv:2512.10857v2 Announce Type: replace-cross Abstract: Transport-based methods have emerged as a leading paradigm for building generative models from large, clean datasets. However, in many scienti

Generative Modeling of Approximately Periodic Time Series by a Posterior-Weighted Gaussian Process

ResearchDGX agent

arXiv:2605.13150v1 Announce Type: cross Abstract: Discrete automated processes in industrial and cyber-physical systems often exhibit a repetitive structure in which successive repetitions follow a co

Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.13632v1 Announce Type: cross Abstract: In this paper, we propose GTA-VLA(Guide, Think, Act), an interactive Vision-Language-Action (VLA) framework that enables spatially steerable embodied

Inference-Time Machine Unlearning via Gated Activation Redirection

Model ReleasesDGX agent

arXiv:2605.12765v1 Announce Type: new Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

SafetyDGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models

SafetyDGX agent

arXiv:2605.12725v1 Announce Type: new Abstract: Recent video anomaly detection research has expanded rapidly with an emphasis on general models of normality intended to work across many different scen

Language Model Networks: Supervision-Efficient Learning through Dense Communication

ApplicationsDGX agent

arXiv:2505.12741v2 Announce Type: replace Abstract: Language models are increasingly used not only as standalone predictors but also as components in larger inference systems, from test-time reasoning

Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models

ResearchDGX agent

arXiv:2601.21975v2 Announce Type: replace Abstract: Recent work identifies a stated-revealed (SvR) preference gap in language models (LMs): a mismatch between the values models endorse and the choices

Neural Surrogate Forward Modelling For Electrocardiology Without Explicit Intracellular Conductivity Tensor

ResearchDGX agent

arXiv:2605.13366v1 Announce Type: new Abstract: Accurate forward modelling is essential for non-invasive cardiac electrophysiology, particularly in atrial fibrillation, where electrical activation is

PaMM: Periodic Motif Memory for Atomistic Models with an Explicit Local-Structure Interface

Model ReleasesDGX agent

arXiv:2605.13297v1 Announce Type: new Abstract: Periodic crystals repeatedly instantiate similar local coordination motifs across translated cells and chemically related structures, but current equiva

Radial Compensation: Fixing Radius Distortion in Chart-Based Generative Models on Riemannian Manifolds

ResearchDGX agent

arXiv:2511.14056v2 Announce Type: replace-cross Abstract: We study the base distribution in chart-based generative models on Riemannian manifolds. Standard methods sample in Euclidean tangent space an

13 May 2026

A Causal Language Modeling Detour Improves Encoder Continued Pretraining

ResearchDGX agent

arXiv:2605.12438v1 Announce Type: new Abstract: When adapting an encoder to a new domain, the standard approach is to continue training with Masked Language Modeling (MLM). We show that temporarily sw

A nonlinear extension of parametric model embedding for dimensionality reduction in parametric shape design

Model ReleasesDGX agent

arXiv:2605.11759v1 Announce Type: cross Abstract: Dimensionality reduction is essential in simulation-based shape design, where high-dimensional parameterizations hinder optimization, surrogate modeli

A Unified Graph Language Model for Multi-Domain Multi-Task Graph Alignment Instruction Tuning

SafetyDGX agent

arXiv:2605.12197v1 Announce Type: new Abstract: Leveraging Graph Neural Networks (GNNs) as graph encoders and aligning the resulting representations with Large Language Models (LLMs) through alignment

Control of Fully Actuated Aerial Vehicles: A Comparison of Model-based and Sensor-based Dynamic Inversion

Model ReleasesDGX agent

arXiv:2605.12071v1 Announce Type: new Abstract: Fully actuated multirotor platforms decouple translational force generation from vehicle attitude, enabling independent control of position and orientat

Do AI models really learn physics, or just learn what physics looks like? Flatiron's @_helenqu, with @PolymathicAI & CDS researchers includi…

TutorialsDGX agent

Do AI models really learn physics, or just learn what physics looks like? Flatiron's @_helenqu, with @PolymathicAI & CDS researchers including @ylecun, finds that models predicting in latent space rec

Enabling Performant and Flexible Model-Internal Observability for LLM Inference

SafetyDGX agent

arXiv:2605.11093v1 Announce Type: new Abstract: Today's inference-time workloads increasingly depend on timely access to a model's internal states. We present DMI-Lib, a high-speed deep model inspecto

From Message-Passing to Linearized Graph Sequence Models

SafetyDGX agent

arXiv:2605.12358v1 Announce Type: new Abstract: Message-passing based approaches form the default backbone of most learning architectures on graph-structured data. However, the rapid progress of moder

Gradient-Boosted Decision Tree for Listwise Context Model in Multimodal Review Helpfulness Prediction

Model ReleasesDGX agent

arXiv:2305.12678v3 Announce Type: replace Abstract: Multimodal Review Helpfulness Prediction (MRHP) aims to rank product reviews based on predicted helpfulness scores and has been widely applied in e-

Grid Games: The Power of Multiple Grids for Quantizing Large Language Models

Model ReleasesDGX agent

arXiv:2605.12327v1 Announce Type: new Abstract: A major recent advance in quantization is given by microscaled 4-bit formats such as NVFP4 and MXFP4, quantizing values into small groups sharing a scal

Hyperbolic Concept Bottleneck Models

ResearchDGX agent

arXiv:2605.06440v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) have become a popular approach to enable interpretability in neural networks by constraining classifier input

Interactive State Space Model with Cross-Modal Local Scanning for Depth Super-Resolution

ResearchDGX agent

arXiv:2605.11934v1 Announce Type: new Abstract: Guided depth super-resolution (GDSR) reconstructs HR depth maps from LR inputs with HR RGB guidance. Existing methods either model each modality indepen

Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation

SafetyDGX agent

arXiv:2510.17062v2 Announce Type: replace Abstract: While reasoning-based large language models excel at complex tasks through an internal, structured thinking process, a concerning phenomenon has eme

Lite3R: A Model-Agnostic Framework for Efficient Feed-Forward 3D Reconstruction

Model ReleasesDGX agent

arXiv:2605.11354v1 Announce Type: new Abstract: Transformer-based 3D reconstruction has emerged as a powerful paradigm for recovering geometry and appearance from multi-view observations, offering str

MAC: Masked Agent Collaboration Boosts Large Language Model Medical Decision-Making

AgentsDGX agent

arXiv:2507.21159v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have proven effective in artificial intelligence, where the multi-agent system (MAS) holds considerable promise f

Modality-Inconsistent Continual Learning of Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2412.13050v2 Announce Type: replace-cross Abstract: In this paper, we introduce Modality-Inconsistent Continual Learning (MICL), a new continual learning scenario for Multimodal Large Language M

more demos on Interaction Models collaboratively doing system design, reading papers, fact-checking with live generative UI

ResearchDGX agent

more demos on Interaction Models collaboratively doing system design, reading papers, fact-checking with live generative UI 1. (System design) - The Interaction Models see your screen and collaborates

Neural ARFIMA model for forecasting BRIC exchange rates with long memory

Model ReleasesDGX agent

arXiv:2509.06697v2 Announce Type: replace-cross Abstract: Accurate forecasting of exchange rates remains a persistent challenge, particularly for emerging economies such as Brazil, Russia, India, and

Oscillators Are All You Need: Irregular Time Series Modelling via Damped Harmonic Oscillators with Closed-Form Solutions

ResearchDGX agent

arXiv:2602.12139v2 Announce Type: replace Abstract: Transformers excel at time series modelling through attention mechanisms that capture long-term temporal patterns. However, they assume uniform time

Partial Model Sharing Improves Byzantine Resilience in Federated Conformal Prediction

ResearchDGX agent

arXiv:2605.11684v1 Announce Type: new Abstract: We propose a Byzantine-resilient federated conformal prediction (FCP) method that leverages partial model sharing, where only a subset of model paramete

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks

SafetyDGX agent

arXiv:2509.06701v2 Announce Type: replace Abstract: We develop a theory of intelligent agency grounded in probabilistic modeling for neural models. Agents are represented as outcome distributions with

ReAD: Reinforcement-Guided Capability Distillation for Large Language Models

ResearchDGX agent

arXiv:2605.11290v1 Announce Type: new Abstract: Capability distillation applies knowledge distillation to selected model capabilities, aiming to compress a large language model (LLM) into a smaller on

Scalable Token-Level Hallucination Detection in Large Language Models

ResearchDGX agent

arXiv:2605.12384v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable capabilities, but they still frequently produce hallucinations. These hallucinations are diffi

SoK: Unlearnability and Unlearning for Model Dememorization

ResearchDGX agent

arXiv:2605.11592v1 Announce Type: new Abstract: Advanced model dememorization methods, including availability poisoning (unlearnability) and machine unlearning, are emerging as key safeguards against

SOMA: Efficient Multi-turn LLM Serving via Small Language Model

Local AiDGX agent

arXiv:2605.11317v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in multi-turn dialogue settings where preserving conversational context across turns is essential

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

SafetyDGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

Variance-aware Reward Modeling with Anchor Guidance

ApplicationsDGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send m…

AgentsDGX agent

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send messages to users, to themselves (CoT) and to tools, and rece

12 May 2026

Active Testing of Large Language Models via Approximate Neyman Allocation

ResearchDGX agent

arXiv:2605.10075v1 Announce Type: new Abstract: Large language models (LLMs) require reliable evaluation from pre-training to test-time scaling, making evaluation a recurring rather than one-off cost.

ALAM: Algebraically Consistent Latent Transitions for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.10819v1 Announce Type: cross Abstract: Vision-language-action (VLA) models remain constrained by the scarcity of action-labeled robot data, whereas action-free videos provide abundant evide

Beyond Majority Voting: Agreement-Based Clustering to Model Annotator Perspectives in Subjective NLP Tasks

ResearchDGX agent

arXiv:2605.09955v1 Announce Type: new Abstract: Disagreement in annotation is a common phenomenon in the development of NLP datasets and serves as a valuable source of insight. While majority voting r

Complete Evidence Extraction with Model Ensembles: A Case Study on Medical Coding

ApplicationsDGX agent

arXiv:2511.07055v3 Announce Type: replace Abstract: High-stakes decisions informed by decision support systems require explicit evidence. While prior work focuses on short sufficient evidence, regulat

Consistent Projection of Langevin Dynamics: Preserving Thermodynamics and Kinetics in Coarse-Grained Models

ResearchDGX agent

arXiv:2512.03706v2 Announce Type: replace-cross Abstract: Coarse graining (CG) is an important task for efficient modeling and simulation of complex multi-scale systems, such as the conformational dyn

Dimensional Coactivation for Representational Consistency in Frozen Vision Foundation Models

ResearchDGX agent

arXiv:2605.08249v1 Announce Type: new Abstract: Frozen vision foundation models do not merely extract features; they organize images through a learned coordinate system. We ask whether that coordinate

Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT

ApplicationsDGX agent

arXiv:2605.09719v1 Announce Type: cross Abstract: Large-scale 3D vision-language models (VLMs) like LLaVA-3D offer strong spatial reasoning but are difficult to deploy due to high computational costs.

Empty SPACE: Cross-Attention Sparsity for Concept Erasure in Diffusion Models

ResearchDGX agent

arXiv:2605.10198v1 Announce Type: cross Abstract: Erasing specific concepts from text-to-image diffusion models is essential for avoiding the generation of copyrighted and explicit content. Closed-for

ExecuTorch -- A Unified PyTorch Solution to Run AI Models On-Device

Local AiDGX agent

arXiv:2605.08195v1 Announce Type: new Abstract: Local execution of AI on edge devices is important for low latency and offline operation. However, deploying models on diverse hardware remains fragment

FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2605.09081v1 Announce Type: cross Abstract: We introduce the first universal pretraining corpus for industrial time-series data: FactoryNet. 51M datapoints across 23k end-to-end task executions

From Syntax to Semantics: Unveiling the Emergence of Chirality in SMILES Translation Models

TutorialsDGX agent

arXiv:2605.09949v1 Announce Type: new Abstract: Understanding how chemical language models (CLMs) learn chemical meaning from molecular string representations, rather than only surface-level string pa

GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction

Model ReleasesDGX agent

arXiv:2605.09973v1 Announce Type: cross Abstract: Reliable detection of personally identifiable information (PII) is increasingly important across modern data-processing systems, yet the task remains

← Previous
1…137138139140141…1009
Next →