AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
26 May 2026

RepSAM: Bridging Foundation Models to Robotic Vision via Representation-Guided Adaptation

Model ReleasesDGX agent

arXiv:2605.25495v1 Announce Type: new Abstract: Robotic perception in unstructured environments remains challenging despite the zero-shot capabilities of foundation models such as SAM. This work attri

Simulating Human Memory with Language Models

ApplicationsDGX agent

arXiv:2605.25680v1 Announce Type: cross Abstract: Language models are increasingly being deployed as user simulators, but their memory is far more reliable than that of real users. To measure this gap

Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow Models

Local AiDGX agent

arXiv:2510.01184v2 Announce Type: replace Abstract: We present a mechanism to steer the sampling diversity of denoising diffusion and flow matching models, allowing users to sample from a sharper or b

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The meaning of prompts and the prompts of meaning: Semiotic reflections and modelling

ResearchDGX agent

arXiv:2509.14250v2 Announce Type: replace Abstract: This paper explores prompts and prompting in large language models (LLMs) as dynamic semiotic phenomena, drawing on Peirce's triadic model of signs,

UtilityMax Prompting: A Formal Framework for Multi-Objective Large Language Model Tasks

Model ReleasesDGX agent

arXiv:2603.11583v4 Announce Type: replace-cross Abstract: The success of a Large Language Model (LLM) task depends heavily on its prompt. Most use-cases specify prompts using natural language, which i

VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Translation

Model ReleasesDGX agent

arXiv:2605.24675v1 Announce Type: cross Abstract: Translating text embedded in Web images is crucial for improving content accessibility and cross-lingual information retrieval, particularly within so

Visual-Redundancy-Controlled Parallel Decoding for Diffusion-Based Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.25820v1 Announce Type: new Abstract: Diffusion-based multimodal large language models (dMLLMs) decode by iteratively predicting tokens at multiple masked positions in parallel. This turns e

25 May 2026

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

Call for Papers - Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]

ResearchDGX agent

This call for papers announces a workshop focused on the growing need for efficient and effective techniques for editing trained models, especially large generative models. The workshop solicits paper

Convergence Without Understanding: When Language Models Agree on Representations but Disagree on Reasoning

SafetyDGX agent

arXiv:2605.23315v1 Announce Type: cross Abstract: Large language models trained under diverse objectives and architectures have been shown to develop increasingly similar internal representations, an

Cultural Adaptation in Large Language Models for Political Discourse

Model ReleasesDGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

Decoding the Critique Mechanism in Large Reasoning Models

ResearchDGX agent

arXiv:2603.16331v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) exhibit backtracking and self-verification mechanisms that enable them to revise intermediate steps and reach correct

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

Local AiDGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

Learning a Particle Dynamics Model with Real-world Videos

TutorialsDGX agent

arXiv:2605.23845v1 Announce Type: new Abstract: Data-driven learning approaches for physics simulation, sometimes referred to as world models, have emerged as promising alternatives to traditional phy

LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws

ResearchDGX agent

arXiv:2605.23901v1 Announce Type: cross Abstract: Existing scaling laws for Large Language Models (LLMs), predominantly monotonic power laws, fail to explain emerging non-monotonic phenomena such as c

Model Collapse as Cultural Evolution

Model ReleasesDGX agent

arXiv:2605.23054v1 Announce Type: cross Abstract: Model collapse, the progressive degradation of LLMs trained on their own outputs, has been characterized statistically but lacks a linguistic explanat

PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.15224v2 Announce Type: replace-cross Abstract: Estimating task progress requires reasoning over long-horizon dynamics rather than recognizing static visual content. While modern Vision-Lang

TEAM: Temporal-Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration

Model ReleasesDGX agent

arXiv:2602.08404v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) have recently gained significant attention due to their inherent support for parallel decoding. Building on

Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.22902v1 Announce Type: cross Abstract: Generative Vision-Language Models (VLMs) perform well on multimodal reasoning, but how visual inputs are transformed to text remains poorly understood

23 May 2026

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines.…

AgentsDGX agent

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines. Frontier models can identify plausible research directions

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

Model ReleasesDGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

Model ReleasesDGX agent

Nemotron-Labs Diffusion Language Models represent NVIDIA's approach to achieving faster text generation through diffusion-based architectures, potentially offering significant speed improvements over

VRPRM: Process Reward Modeling via Visual Reasoning

ResearchDGX agent

arXiv:2508.03556v3 Announce Type: replace Abstract: Process Reward Model (PRM) is widely used in the post-training of Large Language Model (LLM) because it can perform fine-grained evaluation of the r

22 May 2026

Beyond Euclidean Proximity: Repairing Latent World Models with Horizon-Matched Trajectory Reachability Metrics

Model ReleasesDGX agent

arXiv:2605.22164v1 Announce Type: cross Abstract: Latent world models can contain the state needed for control, yet their terminal-cost interface can expose the planner to the wrong decision-relevant

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.21854v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have rapidly converged on a small set of architectural patterns: discrete-token autoregression (e.g. OpenVLA) and co

MAP4TS: A Multi-Aspect Prompting Framework for Time-Series Forecasting with Large Language Models

Model ReleasesDGX agent

arXiv:2510.23090v2 Announce Type: replace Abstract: Recent advances have investigated the use of pretrained large language models (LLMs) for time-series forecasting by aligning numerical inputs with L

PIU: Proximity-guided Identity Unlearning in ID-Conditioned Diffusion Models

ResearchDGX agent

arXiv:2605.22311v1 Announce Type: new Abstract: Identity-conditioned diffusion models enable high-quality and identity-consistent face generation, but they also raise severe privacy concerns, as model

Seizure-Semiology-Suite (S3): A Clinically Multimodal Dataset, Benchmark, and Models for Seizure Semiology Understanding

Model ReleasesDGX agent

arXiv:2605.21852v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in general video understanding, their capacity to interpret invo

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help …

Model ReleasesDGX agent

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help them get stuff done in a cooperative/iterative way. The Clau

21 May 2026

ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.20837v1 Announce Type: new Abstract: Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied i

FlowLM: Few-Step Language Modeling via Diffusion-to-Flow Adaptation

TutorialsDGX agent

arXiv:2605.20199v1 Announce Type: new Abstract: We present FlowLM, a flow matching language model transformed from pre-trained diffusion language models via efficient fine-tuning. By re-aligning the c

GradPower: Powering Gradients for Faster Language Model Pre-Training

Model ReleasesDGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

Matryoshka Concept Bottleneck Models

ResearchDGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

Parameters as Experts: Adapting Vision Models with Dynamic Parameter Routing

Model ReleasesDGX agent

arXiv:2602.06862v2 Announce Type: replace Abstract: Adapting pre-trained vision models using parameter-efficient fine-tuning (PEFT) remains challenging, as it aims to achieve performance comparable to

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

Model ReleasesDGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

20 May 2026

Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models

ResearchDGX agent

arXiv:2511.10292v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) typically process visual inputs as a prefix to the language decoder. As the model autoregressively genera

Base Models Look Human To AI Detectors

Model ReleasesDGX agent

arXiv:2605.19516v1 Announce Type: cross Abstract: As AI-generated text enters the real-world at scale, institutions increasingly use commercial AI-text detectors, especially in education and academic-

CLIF: Concept-Level Influence Functions for Transparent Bottleneck Models

ApplicationsDGX agent

arXiv:2605.19848v1 Announce Type: new Abstract: In recent years, the black-box nature of deep learning models has limited their application in high-stakes domains such as medical diagnosis and finance

Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds

Local AiDGX agent

arXiv:2605.18827v1 Announce Type: cross Abstract: Multiple-choice QA benchmarks usually evaluate small language models (SLMs) as direct answerers, but deployed language-model systems increasingly rely

DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models

ApplicationsDGX agent

arXiv:2605.18868v1 Announce Type: cross Abstract: While vision and multimodal foundation models underpin critical tasks from perception to complex reasoning, they remain highly vulnerable to adversari

Eyes on VLM: Benchmarking Gaze Following and Social Gaze Prediction in Vision Language Models

ResearchDGX agent

arXiv:2605.19859v1 Announce Type: new Abstract: Vision-language models (VLMs) have rapidly evolved into general-purpose multimodal reasoners with strong zero-shot generalization. In this context, VLMs

Fine-tuning Large Language Model for Automated Algorithm Design

Model ReleasesDGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

Improved visual-information-driven model for crowd simulation and its modular application

SafetyDGX agent

arXiv:2504.03758v4 Announce Type: replace-cross Abstract: Crowd movement simulation is crucial for pedestrian safety management and facility design. Data-driven models offer the potential to improve r

Inference-Time Scaling in Diffusion Models through Iterative Partial Refinement

ResearchDGX agent

arXiv:2605.19317v1 Announce Type: cross Abstract: Inference-time scaling has emerged as a major approach for improving reasoning capabilities, and has been increasingly applied to diffusion models. Ho

Jailbreaking on Text-to-Video Models via Scene Splitting Strategy

SafetyDGX agent

arXiv:2509.22292v2 Announce Type: replace-cross Abstract: Along with the rapid advancement of numerous Text-to-Video (T2V) models, growing concerns have emerged regarding their safety risks. While rec

MetaRA: Metamorphic Robustness Assessment for Multimodal Large Language Model-based Visual Question Answering Systems

Model ReleasesDGX agent

arXiv:2605.19307v1 Announce Type: new Abstract: Visual Question Answering (VQA), as the representative multimodal task, serves as a key benchmark for evaluating the reasoning capabilities of Multimoda

MiMuon: Mixed Muon Optimizer with Improved Generalization for Large Models

ResearchDGX agent

arXiv:2605.19619v1 Announce Type: cross Abstract: Matrix-structured parameters frequently appear in many artificial intelligence models such as large language models. More recently, an efficient Muon

No Hard Negatives Required: Concept Centric Learning Leads to Compositionality without Degrading Zero-shot Capabilities of Contrastive Models

Model ReleasesDGX agent

arXiv:2603.25722v2 Announce Type: replace Abstract: Contrastive vision-language (V&L) models remain a popular choice for various applications. However, several limitations have emerged, most notably t

PhyWorld: Physics-Faithful World Model for Video Generation

Model ReleasesDGX agent

arXiv:2605.19242v1 Announce Type: cross Abstract: World simulators can provide safe and scalable environments for training Physical AI systems before real-world deployment. Large video generation mode

Protein Autoregressive Modeling via Multiscale Structure Generation

Model ReleasesDGX agent

arXiv:2602.04883v2 Announce Type: replace-cross Abstract: We present protein autoregressive modeling (PAR), the first multi-scale autoregressive framework for protein backbone generation via coarse-to

Reducing Diffusion Model Memorization with Higher Order Langevin Dynamics

ApplicationsDGX agent

arXiv:2605.19170v1 Announce Type: cross Abstract: Diffusion/score-based models have emerged as powerful generative models, capable of generating high-quality samples that mimic the training data distr

Robust Basis Spline Decoupling for the Compression of Transformer Models

Model ReleasesDGX agent

arXiv:2605.18794v1 Announce Type: cross Abstract: Decoupling is a powerful modeling paradigm for representing multivariate functions as compositions of linear transformations and univariate nonlinear

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

The Growing Pains of Frontier Models: When Leaderboards Stop Separating and What to Measure Next

Model ReleasesDGX agent

arXiv:2605.18840v1 Announce Type: cross Abstract: Leaderboards rank frontier models on independent axes but do not reveal whether capabilities reinforce or trade off across releases -- and at the fron

19 May 2026

A Survey on Foundation Models for Personalized Federated Intelligence

Model ReleasesDGX agent

arXiv:2505.06907v2 Announce Type: replace Abstract: The rise of large language models (LLMs), such as ChatGPT, Gemini, and Grok, has reshaped the AI landscape. As prominent instances of foundational m

A Unified Framework for Structured Flow Modeling: From Continuous Fields to Data-Driven Representations

ResearchDGX agent

arXiv:2605.18250v1 Announce Type: cross Abstract: Many dynamical systems can be described in terms of structured flows combining source/sink behavior, cyclic dynamics, and topology-constrained transpo

Accelerating Rectified Flow Models via Trajectory-Aware Caching

Model ReleasesDGX agent

arXiv:2605.16789v1 Announce Type: new Abstract: Diffusion and rectified flow (RF) models generate high-fidelity images and videos, but their iterative velocity-field evaluations are computationally ex

Accelerating Redshift-Conditioned Galaxy Image Synthesis with One-step Generative Modeling

ResearchDGX agent

arXiv:2605.17546v1 Announce Type: cross Abstract: Understanding galaxy morphology evolution across cosmic time requires models that can generate realistic galaxy populations conditioned on redshift. I

Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap

Model ReleasesDGX agent

arXiv:2605.17903v1 Announce Type: new Abstract: We automatically generate feedback causal fuzzy cognitive maps (FCMs) from text by teaching large-language-model agents to break the text into overlappi

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making

SafetyDGX agent

arXiv:2605.17228v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes domains such as clinical decision support and medical documentation. However, the

← Previous
1…8788899091…1009
Next →