AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
19 May 2026

CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models

ResearchDGX agent

arXiv:2602.17684v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has driven recent progress in code large language models by leveraging execution-based f

Cracks in the Foundation: A Civil Infrastructure Dataset to Challenge Vision Foundation Models

SafetyDGX agent

arXiv:2605.18413v1 Announce Type: new Abstract: Automated structural health monitoring is essential to prevent catastrophic infrastructure failures. Precise, pixel-level defect segmentation is needed

DACA-GRPO: Denoising-Aware Credit Assignment for Reinforcement Learning in Diffusion Language Models

SafetyDGX agent

arXiv:2605.16342v1 Announce Type: cross Abstract: Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DAD4TS: Data-Augmentation-Oriented Diffusion Model for Time-Series Forecasting with Small-Scale Data

ApplicationsDGX agent

arXiv:2605.17866v1 Announce Type: new Abstract: Small-scale data is a critical problem in time-series forecasting tasks. Data augmentation is an effective strategy for this task, but it has a limitati

Data Presentation Over Architecture: Resampling Strategies for Credit Risk Prediction with Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2605.18635v1 Announce Type: cross Abstract: Credit default prediction is a tabular learning problem with severe class imbalance, heterogeneous features, and tight latency budgets. Tabular Founda

Designing streetscapes from street-view imagery using diffusion models

Model ReleasesDGX agent

arXiv:2605.17527v1 Announce Type: new Abstract: Street-view imagery (SVI) is widely used to quantify key indicators of urban environment, such as green- ery, sky, or road view indices. However, existi

DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking

Model ReleasesDGX agent

arXiv:2605.17451v1 Announce Type: new Abstract: Aerial object tracking has broad applications in public safety, emergency rescue, wildlife monitoring, and related fields. However, existing aerial trac

DocReward: A Document Reward Model for Structuring and Stylizing

Model ReleasesDGX agent

arXiv:2510.11391v3 Announce Type: replace-cross Abstract: Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic

Dynamic robotic cloth folding with efficient Koopman operator-based model predictive control

ResearchDGX agent

arXiv:2605.18373v1 Announce Type: cross Abstract: Robotic cloth folding is a challenging task, particularly when considering dynamic folding tasks, which aim at folding cloth by fast motions that leve

Employing Vision-Language Models for Face Image Quality Assessment

Model ReleasesDGX agent

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

EvoQRE: Modeling Bounded Rationality in Safety-Critical Traffic Simulation via Evolutionary Quantal Response Equilibrium

Model ReleasesDGX agent

arXiv:2601.05653v2 Announce Type: replace Abstract: Existing traffic simulation frameworks for autonomous vehicles typically rely on imitation learning or game-theoretic approaches that solve for Nash

FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

Model ReleasesDGX agent

arXiv:2503.16492v3 Announce Type: replace-cross Abstract: ffective Human-Robot Interaction (HRI) is crucial for enhancing accessibility and usability in real-world robotics applications. However, exis

FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints

Model ReleasesDGX agent

arXiv:2509.06984v3 Announce Type: replace-cross Abstract: Federated Learning with LoRA fine-tuning offers an efficient and privacy-aware solution for institutions to collaboratively leverage their lar

GEM: Gaussian Evolution Model for Occupancy Forecasting and Motion Planning

AgentsDGX agent

arXiv:2605.17682v1 Announce Type: new Abstract: Future 3D semantic occupancy forecasting and motion planning are central to autonomous driving, as they require models to reason about how surrounding s

GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics

Model ReleasesDGX agent

arXiv:2506.14009v2 Announce Type: replace Abstract: Autonomous drones capable of interpreting and executing high-level language instructions in unstructured environments remain a long-standing goal. Y

MCQ Difficulty Prediction via Modeling Learner Heterogeneity Using Data-Driven Cognitive Profiling

Model ReleasesDGX agent

arXiv:2605.16290v1 Announce Type: cross Abstract: Predicting the difficulty of multiple-choice questions (MCQs) is important for effective assessment, yet current methods typically assume a unimodal s

Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks

ResearchDGX agent

arXiv:2605.16864v1 Announce Type: cross Abstract: Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-awa

NanoQuant: Efficient Sub-1-Bit Quantization of Large Language Models

HardwareDGX agent

arXiv:2602.06694v2 Announce Type: replace Abstract: Weight-only quantization has become a standard approach for efficiently serving large language models (LLMs). However, existing methods fail to effi

OrbiSim: World Models as Differentiable Physics Engines for Embodied Intelligence

SafetyDGX agent

arXiv:2605.16395v1 Announce Type: cross Abstract: We present OrbiSim, a novel robotic simulation paradigm that redefines world models as a fully differentiable physics engine for embodied intelligence

Self-supervised Hierarchical Visual Reasoning with World Model

Model ReleasesDGX agent

arXiv:2605.17537v1 Announce Type: new Abstract: 3D open-world environments with adversarial opponents remain a core challenge for reinforcement learning due to their vast state spaces. Effective reaso

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

AgentsDGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

Test-Time Hinting for Black-Box Vision-Language Models

ResearchDGX agent

arXiv:2605.16410v1 Announce Type: new Abstract: Test-time scaling (TTS) methods have proven highly effective for LLMs, yet their application to vision-language models (VLMs) remains relatively underex

Traces of Social Competence in Large Language Models

ApplicationsDGX agent

arXiv:2603.04161v2 Announce Type: replace Abstract: The False Belief Test (FBT) has been the main method for assessing Theory of Mind (ToM) and related socio-cognitive competencies. For Large Language

Training data attribution in diffusion models via mirrored unlearning and noise-consistent skew

ApplicationsDGX agent

arXiv:2605.17938v1 Announce Type: cross Abstract: Training data attribution (TDA) should enable generative model interpretability and foster a variety of related downstream tasks. Nonetheless, current

Vision Foundation Models as Generalist Tokenizers for Image Generation

ResearchDGX agent

arXiv:2605.18390v1 Announce Type: new Abstract: In this work, we explore the largely unexplored direction of building a generalist image tokenizer directly on top of a frozen vision foundation model (

WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points

ResearchDGX agent

arXiv:2605.17471v1 Announce Type: new Abstract: Quantization-aware training (QAT) is widely adopted to quantize language models by training full-precision weights using gradients from the quantized mo

18 May 2026

ASRU: Activation Steering Meets Reinforcement Unlearning for Multimodal Large Language Models

SafetyDGX agent

arXiv:2605.15687v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) may memorize sensitive cross-modal information during pretraining, making machine unlearning (MU) crucial. Ex

Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction

Model ReleasesDGX agent

arXiv:2605.16077v1 Announce Type: new Abstract: Accurate assessment of cognitive decline from spontaneous speech remains challenging due to limited dataset size and class imbalance. In this work, we p

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models

ResearchDGX agent

arXiv:2601.21798v2 Announce Type: replace Abstract: Large Language Models(LLMs) have revolutionized text generation and multimodal perception,but their capabilities in 3D content generation remain und

EgoExo-WM: Unlocking Exo Video for Ego World Models

SafetyDGX agent

arXiv:2605.15477v1 Announce Type: new Abstract: Egocentric world models present a promising direction for enabling agents to predict and plan, but their performance is constrained by the limited avail

Extrapolation Guarantees for Perturbation Modeling Under the Additive Latent Shift Assumption

ResearchDGX agent

arXiv:2504.18522v3 Announce Type: replace-cross Abstract: We consider the problem of modeling the effects of perturbations like gene knockouts on measurements such as single-cell RNA counts. Given dat

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

Model ReleasesDGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

Few-Step Diffusion Language Models via Trajectory Self-Distillation

ResearchDGX agent

arXiv:2602.12262v3 Announce Type: replace Abstract: Diffusion large language models (DLLMs) have emerged as powerful generative models with the promise of fast text generation through parallel decodin

Large Language Models Could Be Rote Learners

Model ReleasesDGX agent

arXiv:2504.08300v5 Announce Type: replace-cross Abstract: Benchmark-based evaluation, e.g., multiple-choice questions (MCQs) and open-ended questions (OEQs), is widely used for evaluating Large Langua

Metropolis-Scale Road Network Datasets for Fine-Grained Urban Traffic Modeling

ApplicationsDGX agent

arXiv:2510.02278v2 Announce Type: replace Abstract: Modeling traffic dynamics is a critical challenge for urban computing, with applications from real-time traffic management to infrastructure plannin

PanoWorld: Geometry-Consistent Panoramic Video World Modeling

ResearchDGX agent

arXiv:2605.15391v1 Announce Type: cross Abstract: We present PanoWorld, a panoramic video world model that generates geometry-consistent 360egree video from a single image and a caption. Existing pano

Prompt Stability Scoring for Text Annotation with Large Language Models

ResearchDGX agent

arXiv:2407.02039v3 Announce Type: replace Abstract: Researchers are increasingly using language models (LMs) for text annotation. These approaches rely only on a prompt telling the model to return a g

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

Model ReleasesDGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models

ResearchDGX agent

arXiv:2605.15961v1 Announce Type: new Abstract: Large-scale pre-trained vision-language models like CLIP demonstrate remarkable zero-shot performance across diverse tasks. However, fine-tuning these m

Testing properties of trees in graphical models with covariance queries

ResearchDGX agent

arXiv:2605.15996v1 Announce Type: cross Abstract: We consider the problem of testing properties of graphs underlying high-dimensional graphical models. We adopt the model of covariance queries introdu

the team did an internal test of this model last week the whole company (bar a few exceptions) had all their cursor chats redirected to comp…

IndustryDGX agent

the team did an internal test of this model last week the whole company (bar a few exceptions) had all their cursor chats redirected to composer 2.5 for like 2 days. i didn't even notice, which I thin

16 May 2026

💯. Way too much focus on language models.

SafetyDGX agent

💯. Way too much focus on language models. Fei-Fei Li warns that AI may be staring too hard at language models. The world is not just text on a screen. It is physical, visual, spatial, and always chang

15 May 2026

AIM-DDI: A Model-Agnostic Multimodal Integration Module for Drug-Drug Interaction Prediction

ResearchDGX agent

arXiv:2605.14327v1 Announce Type: cross Abstract: Drug-drug interaction (DDI) prediction is a critical task in computational biomedicine, as adverse interactions between co-administered drugs can caus

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

SafetyDGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

BiSpikCLM: A Spiking Language Model integrating Softmax-Free Spiking Attention and Spike-Aware Alignment Distillation

SafetyDGX agent

arXiv:2605.13859v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer promising energy-efficient alternatives to large language models (LLMs) due to their event-driven nature and ultr

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.05762v5 Announce Type: replace Abstract: Recent advances in offline Reinforcement Learning (RL) have proven that effective policy learning can benefit from imposing conservative constraints

Breaking the Reasoning Horizon in Entity Alignment Foundation Models

Local AiDGX agent

arXiv:2601.21174v2 Announce Type: replace Abstract: Entity alignment (EA) is critical for knowledge graph (KG) fusion. Existing EA models lack transferability and are incapable of aligning unseen KGs

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

SafetyDGX agent

arXiv:2605.14544v1 Announce Type: new Abstract: Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this lab

Enhanced and Efficient Reasoning in Large Learning Models

ApplicationsDGX agent

arXiv:2605.14036v1 Announce Type: new Abstract: In current Large Language Models we can trust the production of smoothly flowing prose on the basis of the principles of machine learning. However, ther

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building.

ToolsDGX agent

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building. LangSmith Fleet now has a free model powered by @FireworksAI_HQ for Developer and Plus plans. It’

GhostCite: A Large-Scale Analysis of Citation Validity in the Age of Large Language Models

Model ReleasesDGX agent

arXiv:2602.06718v2 Announce Type: replace-cross Abstract: Citations provide the basis for trusting scientific claims; when they are invalid or fabricated, this trust collapses. With the advent of Larg

How Sensitive Are Radiomic AI Models to Acquisition Parameters?

Model ReleasesDGX agent

arXiv:2605.14667v1 Announce Type: new Abstract: A main barrier for the deployment of AI radiomic systems in clinical routine is their drop in performance under heterogeneous multicentre acquisition pr

ImmuVis: Hyperconvolutional Foundation Model for Imaging Mass Cytometry

ApplicationsDGX agent

arXiv:2602.04585v2 Announce Type: replace Abstract: We present ImmuVis, a family of efficient foundation models for imaging mass cytometry (IMC), a high-throughput multiplex imaging technology that ha

K-Models: a Flexible and Interpretable Method for Ordinal Clustering with Application to Antigen-Antibody Interaction Profiles

Model ReleasesDGX agent

arXiv:2605.14828v1 Announce Type: cross Abstract: Existing clustering methods for functional data often prioritize partitioning accuracy over interpretability, making it challenging to extract meaning

KGPFN: Unlocking the Potential of Knowledge Graph Foundation Model via In-Context Learning

Local AiDGX agent

arXiv:2605.14907v1 Announce Type: new Abstract: Knowledge graph (KG) foundation models aim to generalize across graphs with unseen entities and relations by learning transferable relational structure.

M^2RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling

ResearchDGX agent

arXiv:2603.14360v2 Announce Type: replace-cross Abstract: Transformers are highly parallel but are limited to computations in the TC^0 complexity class, excluding tasks such as entity tracking and cod

Measuring and Mitigating Toxicity in Large Language Models: A Comprehensive Replication Study

SafetyDGX agent

arXiv:2605.14087v1 Announce Type: new Abstract: Large Language Models (LLMs), when trained on web-scale corpora, inherently absorb toxic patterns from their training data. This leads to ``toxic degene

Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models

SafetyDGX agent

arXiv:2605.14530v1 Announce Type: new Abstract: Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding fo

Pelican-Unified 1.0: A Unified Embodied Intelligence Model for Understanding, Reasoning, Imagination and Action

ResearchDGX agent

arXiv:2605.15153v1 Announce Type: cross Abstract: We present Pelican-Unified 1.0, the first embodied foundation model trained according to the principle of unification. Pelican-Unified 1.0 uses a sing

Pixal3D: Generate high-fidelity 3D assets from a single image. (TencentARC, locally runnable model)

Local AiDGX agent

Pixal3D is a locally runnable AI model developed by TencentARC that generates high-fidelity 3D assets from single 2D images. The tool leverages advanced techniques to convert 2D image inputs into deta

← Previous
1…136137138139140…1009
Next →