AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
Model Releases

M^3-Verse: A 'Spot the Difference' Challenge for Large Multimodal Models

DGX agent

arXiv:2512.18735v2 Announce Type: replace-cross Abstract: Modern Large Multimodal Models (LMMs) have demonstrated extraordinary ability in static image and single-state spatial-temporal understanding.

model-releasesarxiv-cs-ai
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

NSR-Boost: A Neuro-Symbolic Residual Boosting Framework for Industrial Legacy Models

DGX agent

arXiv:2601.10457v3 Announce Type: replace Abstract: Although the Gradient Boosted Decision Trees (GBDTs) dominate industrial tabular applications, upgrading legacy models in high-concurrency productio

applicationsarxiv-cs-ai
26 May 2026
Model Releases

PALoRA: Projection-Adaptive LoRA for Preserving Reasoning in Large Language Models

DGX agent

arXiv:2605.24549v1 Announce Type: new Abstract: Efficiently updating Large Language Models (LLMs) with new or evolving factual knowledge remains a central challenge, as even parameter-efficient adapta

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ParkourFormer: Integrating Predictive Supervision and Sequence Modeling into Parkour Locomotion

DGX agent

arXiv:2605.25782v1 Announce Type: new Abstract: Humanoid parkour requires locomotion policies to coordinate whole-body dynamics across rapidly changing terrains such as stairs, gaps, slopes, and obsta

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

PolySAE: Modeling Feature Interactions in Sparse Autoencoders via Polynomial Decoding

DGX agent

arXiv:2602.01322v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) interpret neural network representations by decomposing activations into sparse combinations of dictionary atoms. H

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Representation Without Control: Testing the Realization Effect in Language Models

DGX agent

arXiv:2605.25151v1 Announce Type: new Abstract: Large language models are increasingly used as behavioral simulators, but it remains unclear when their outputs reflect human-like cognitive mechanisms

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RepSAM: Bridging Foundation Models to Robotic Vision via Representation-Guided Adaptation

DGX agent

arXiv:2605.25495v1 Announce Type: new Abstract: Robotic perception in unstructured environments remains challenging despite the zero-shot capabilities of foundation models such as SAM. This work attri

model-releasesarxiv-cs-ro
26 May 2026
Applications

Simulating Human Memory with Language Models

DGX agent

arXiv:2605.25680v1 Announce Type: cross Abstract: Language models are increasingly being deployed as user simulators, but their memory is far more reliable than that of real users. To measure this gap

applicationsarxiv-cs-ai
26 May 2026
Local Ai

Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow Models

DGX agent

arXiv:2510.01184v2 Announce Type: replace Abstract: We present a mechanism to steer the sampling diversity of denoising diffusion and flow matching models, allowing users to sample from a sharper or b

local-aiarxiv-cs-lg
26 May 2026
Research

The meaning of prompts and the prompts of meaning: Semiotic reflections and modelling

DGX agent

arXiv:2509.14250v2 Announce Type: replace Abstract: This paper explores prompts and prompting in large language models (LLMs) as dynamic semiotic phenomena, drawing on Peirce's triadic model of signs,

researcharxiv-cs-cl
26 May 2026
Model Releases

UtilityMax Prompting: A Formal Framework for Multi-Objective Large Language Model Tasks

DGX agent

arXiv:2603.11583v4 Announce Type: replace-cross Abstract: The success of a Large Language Model (LLM) task depends heavily on its prompt. Most use-cases specify prompts using natural language, which i

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Translation

DGX agent

arXiv:2605.24675v1 Announce Type: cross Abstract: Translating text embedded in Web images is crucial for improving content accessibility and cross-lingual information retrieval, particularly within so

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Visual-Redundancy-Controlled Parallel Decoding for Diffusion-Based Multimodal Large Language Models

DGX agent

arXiv:2605.25820v1 Announce Type: new Abstract: Diffusion-based multimodal large language models (dMLLMs) decode by iteratively predicting tokens at multiple masked positions in parallel. This turns e

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

model-releasesarxiv-cs-ai
25 May 2026
Research

Call for Papers - Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]

DGX agent

This call for papers announces a workshop focused on the growing need for efficient and effective techniques for editing trained models, especially large generative models. The workshop solicits paper

researchr-machinelearning
25 May 2026
Safety

Convergence Without Understanding: When Language Models Agree on Representations but Disagree on Reasoning

DGX agent

arXiv:2605.23315v1 Announce Type: cross Abstract: Large language models trained under diverse objectives and architectures have been shown to develop increasingly similar internal representations, an

safetyarxiv-cs-ai
25 May 2026
Model Releases

Cultural Adaptation in Large Language Models for Political Discourse

DGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

model-releasesarxiv-cs-cl
25 May 2026
Research

Decoding the Critique Mechanism in Large Reasoning Models

DGX agent

arXiv:2603.16331v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) exhibit backtracking and self-verification mechanisms that enable them to revise intermediate steps and reach correct

researcharxiv-cs-lg
25 May 2026
Local Ai

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

DGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

local-aiarxiv-cs-ai
25 May 2026
Tutorials

Learning a Particle Dynamics Model with Real-world Videos

DGX agent

arXiv:2605.23845v1 Announce Type: new Abstract: Data-driven learning approaches for physics simulation, sometimes referred to as world models, have emerged as promising alternatives to traditional phy

tutorialsarxiv-cs-cv
25 May 2026
Research

LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws

DGX agent

arXiv:2605.23901v1 Announce Type: cross Abstract: Existing scaling laws for Large Language Models (LLMs), predominantly monotonic power laws, fail to explain emerging non-monotonic phenomena such as c

researcharxiv-cs-ai
25 May 2026
Model Releases

Model Collapse as Cultural Evolution

DGX agent

arXiv:2605.23054v1 Announce Type: cross Abstract: Model collapse, the progressive degradation of LLMs trained on their own outputs, has been characterized statistically but lacks a linguistic explanat

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

DGX agent

arXiv:2601.15224v2 Announce Type: replace-cross Abstract: Estimating task progress requires reasoning over long-horizon dynamics rather than recognizing static visual content. While modern Vision-Lang

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

TEAM: Temporal-Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration

DGX agent

arXiv:2602.08404v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) have recently gained significant attention due to their inherent support for parallel decoding. Building on

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models

DGX agent

arXiv:2605.22902v1 Announce Type: cross Abstract: Generative Vision-Language Models (VLMs) perform well on multimodal reasoning, but how visual inputs are transformed to text remains poorly understood

model-releasesarxiv-cs-ai
25 May 2026
Agents

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines.…

DGX agent

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines. Frontier models can identify plausible research directions

agentsdair-ai--x
23 May 2026
Model Releases

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

DGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

DGX agent

Nemotron-Labs Diffusion Language Models represent NVIDIA's approach to achieving faster text generation through diffusion-based architectures, potentially offering significant speed improvements over

model-releaseshugging-face
23 May 2026
Research

VRPRM: Process Reward Modeling via Visual Reasoning

DGX agent

arXiv:2508.03556v3 Announce Type: replace Abstract: Process Reward Model (PRM) is widely used in the post-training of Large Language Model (LLM) because it can perform fine-grained evaluation of the r

researcharxiv-cs-lg
23 May 2026
Model Releases

Beyond Euclidean Proximity: Repairing Latent World Models with Horizon-Matched Trajectory Reachability Metrics

DGX agent

arXiv:2605.22164v1 Announce Type: cross Abstract: Latent world models can contain the state needed for control, yet their terminal-cost interface can expose the planner to the wrong decision-relevant

model-releasesarxiv-cs-ro
22 May 2026
Model Releases

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.21854v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have rapidly converged on a small set of architectural patterns: discrete-token autoregression (e.g. OpenVLA) and co

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MAP4TS: A Multi-Aspect Prompting Framework for Time-Series Forecasting with Large Language Models

DGX agent

arXiv:2510.23090v2 Announce Type: replace Abstract: Recent advances have investigated the use of pretrained large language models (LLMs) for time-series forecasting by aligning numerical inputs with L

model-releasesarxiv-cs-cl
22 May 2026
Research

PIU: Proximity-guided Identity Unlearning in ID-Conditioned Diffusion Models

DGX agent

arXiv:2605.22311v1 Announce Type: new Abstract: Identity-conditioned diffusion models enable high-quality and identity-consistent face generation, but they also raise severe privacy concerns, as model

researcharxiv-cs-cv
22 May 2026
Model Releases

Seizure-Semiology-Suite (S3): A Clinically Multimodal Dataset, Benchmark, and Models for Seizure Semiology Understanding

DGX agent

arXiv:2605.21852v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in general video understanding, their capacity to interpret invo

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help …

DGX agent

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help them get stuff done in a cooperative/iterative way. The Clau

model-releasesjeremy-howard--x
22 May 2026
Model Releases

ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models

DGX agent

arXiv:2605.20837v1 Announce Type: new Abstract: Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied i

model-releasesarxiv-cs-cv
21 May 2026
Tutorials

FlowLM: Few-Step Language Modeling via Diffusion-to-Flow Adaptation

DGX agent

arXiv:2605.20199v1 Announce Type: new Abstract: We present FlowLM, a flow matching language model transformed from pre-trained diffusion language models via efficient fine-tuning. By re-aligning the c

tutorialsarxiv-cs-cl
21 May 2026
Model Releases

GradPower: Powering Gradients for Faster Language Model Pre-Training

DGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

model-releasesarxiv-cs-lg
21 May 2026
Research

Matryoshka Concept Bottleneck Models

DGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

researcharxiv-cs-lg
21 May 2026
Model Releases

Parameters as Experts: Adapting Vision Models with Dynamic Parameter Routing

DGX agent

arXiv:2602.06862v2 Announce Type: replace Abstract: Adapting pre-trained vision models using parameter-efficient fine-tuning (PEFT) remains challenging, as it aims to achieve performance comparable to

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

DGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

model-releasesarxiv-cs-lg
21 May 2026
Research

Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models

DGX agent

arXiv:2511.10292v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) typically process visual inputs as a prefix to the language decoder. As the model autoregressively genera

researcharxiv-cs-ai
20 May 2026
Model Releases

Base Models Look Human To AI Detectors

DGX agent

arXiv:2605.19516v1 Announce Type: cross Abstract: As AI-generated text enters the real-world at scale, institutions increasingly use commercial AI-text detectors, especially in education and academic-

model-releasesarxiv-cs-ai
20 May 2026
Applications

CLIF: Concept-Level Influence Functions for Transparent Bottleneck Models

DGX agent

arXiv:2605.19848v1 Announce Type: new Abstract: In recent years, the black-box nature of deep learning models has limited their application in high-stakes domains such as medical diagnosis and finance

applicationsarxiv-cs-cl
20 May 2026
Local Ai

Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds

DGX agent

arXiv:2605.18827v1 Announce Type: cross Abstract: Multiple-choice QA benchmarks usually evaluate small language models (SLMs) as direct answerers, but deployed language-model systems increasingly rely

local-aiarxiv-cs-lg
20 May 2026
Applications

DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models

DGX agent

arXiv:2605.18868v1 Announce Type: cross Abstract: While vision and multimodal foundation models underpin critical tasks from perception to complex reasoning, they remain highly vulnerable to adversari

applicationsarxiv-cs-ai
20 May 2026
Research

Eyes on VLM: Benchmarking Gaze Following and Social Gaze Prediction in Vision Language Models

DGX agent

arXiv:2605.19859v1 Announce Type: new Abstract: Vision-language models (VLMs) have rapidly evolved into general-purpose multimodal reasoners with strong zero-shot generalization. In this context, VLMs

researcharxiv-cs-cv
20 May 2026
Model Releases

Fine-tuning Large Language Model for Automated Algorithm Design

DGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

model-releasesarxiv-cs-ai
20 May 2026
← Previous
1…109110111112113…1261
Next →