AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy

DGX agent

arXiv:2605.15711v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities across various tasks, yet they remain vulnerable to backdoor attacks. Exi

safetyarxiv-cs-cv
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Going Beyond the Edge: Distributed Inference of Transformer Models on Ultra-Low-Power Wireless Devices

DGX agent

arXiv:2605.15694v1 Announce Type: new Abstract: Transformer models are rapidly becoming a cornerstone of modern Internet of Things (IoT) applications, yet their computational and memory demands far ex

applicationsarxiv-cs-lg
18 May 2026
Safety

Improve Large Language Model Systems with User Logs

DGX agent

arXiv:2602.06470v2 Announce Type: replace-cross Abstract: Scaling training data and model parameters has long driven progress in large language models (LLMs), but this paradigm is increasingly constra

safetyarxiv-cs-ai
18 May 2026
Research

Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance

DGX agent

arXiv:2605.15436v1 Announce Type: new Abstract: This paper presents a comprehensive analysis of neural activation patterns across six distinct large language model (LLM) architectures, examining their

researcharxiv-cs-cl
18 May 2026
Model Releases

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models

DGX agent

arXiv:2509.22739v3 Announce Type: replace-cross Abstract: Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the for

model-releasesarxiv-cs-ai
18 May 2026
Research

Privacy Evaluation of Generative Models for Trajectory Generation

DGX agent

arXiv:2605.15246v1 Announce Type: new Abstract: Trajectory data is fundamental to modern urban intelligence, yet its sensitivity raises significant privacy concerns. Generative models such as Generati

researcharxiv-cs-lg
18 May 2026
Research

RoiMAM: Region-of-Interest Medical Attention Model for Efficient Vision-Language Understanding

DGX agent

arXiv:2605.15561v1 Announce Type: new Abstract: Vision-Language Models (VLMs) facilitate medical visual question answering (MedVQA) by jointly interpreting images and text. However, existing models ty

researcharxiv-cs-cv
18 May 2026
Tutorials

Simultaneous State Estimation and Online Model Learning in a Soft Robotic System

DGX agent

arXiv:2602.14092v2 Announce Type: replace-cross Abstract: Operating complex real-world systems, such as soft robots, can benefit from precise predictive control schemes that require accurate state and

tutorialsarxiv-cs-ro
18 May 2026
Applications

Toward World Modeling of Physiological Signals with Chaos-Theoretic Balancing and Latent Dynamics

DGX agent

arXiv:2605.15465v1 Announce Type: new Abstract: Physiological time series signals reflect complex, multi-scale dynamical processes of the human body. Existing modeling studies focus on static tasks su

applicationsarxiv-cs-lg
18 May 2026
Agents

Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

DGX agent

arXiv:2510.10454v2 Announce Type: replace Abstract: Large language models (LLMs) offer a generalizable approach for modeling patient trajectories, but suffer from the long and noisy nature of electron

agentsarxiv-cs-ai
18 May 2026
Model Releases

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

DGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

model-releasesarxiv-cs-lg
18 May 2026
Safety

Video Models Can Reason with Verifiable Rewards

DGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

safetyarxiv-cs-cv
18 May 2026
Model Releases

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

DGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

model-releasesarxiv-cs-ai
15 May 2026
Safety

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

DGX agent

arXiv:2605.14723v1 Announce Type: new Abstract: Sepsis management in the ICU requires sequential treatment decisions under rapidly evolving patient physiology. Although large language models (LLMs) en

safetyarxiv-cs-ai
15 May 2026
Model Releases

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

DGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

model-releasesarxiv-cs-cl
15 May 2026
Applications

Energy-Regularized Sequential Model Editing on Hyperspheres

DGX agent

arXiv:2510.01172v3 Announce Type: replace Abstract: Large language models (LLMs) require constant updates to remain aligned with evolving real-world knowledge. Model editing offers a lightweight alter

applicationsarxiv-cs-cl
15 May 2026
Model Releases

Exploring Vision-Language Models for Online Signature Verification: A Zero-Shot Capability Study

DGX agent

arXiv:2605.14845v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have demonstrated strong capabilities in general visual reasoning, yet their applicability to rigor

model-releasesarxiv-cs-cv
15 May 2026
Research

Generative Bayesian Optimization: Generative Models as Acquisition Functions

DGX agent

arXiv:2510.25240v3 Announce Type: replace-cross Abstract: We present a general strategy for turning generative models into candidate solution samplers for batch Bayesian optimization (BO). The use of

researcharxiv-cs-lg
15 May 2026
Model Releases

Hyperspectral Image Land Cover Captioning Dataset for Vision Language Models

DGX agent

arXiv:2505.12217v2 Announce Type: replace Abstract: We introduce HyperCap, the first large-scale hyperspectral captioning dataset designed to enhance model performance and effectiveness in remote sens

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

DGX agent

arXiv:2509.25826v3 Announce Type: replace Abstract: Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generaliz

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Mechanistic Interpretability of EEG Foundation Models via Sparse Autoencoders

DGX agent

arXiv:2605.13930v1 Announce Type: new Abstract: EEG foundation models achieve state-of-the-art clinical performance, yet the internal computations driving their predictions remain opaque: a barrier to

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

MultiEmo-Bench: Multi-label Visual Emotion Analysis for Multi-modal Large Language Models

DGX agent

arXiv:2605.14635v1 Announce Type: cross Abstract: This paper introduces a multi-label visual emotion analysis benchmark dataset for comprehensively evaluating the ability of multimodal large language

model-releasesarxiv-cs-ai
15 May 2026
Agents

Orchard: An Open-Source Agentic Modeling Framework

DGX agent

arXiv:2605.15040v1 Announce Type: new Abstract: Agentic modeling aims to transform LLMs into autonomous agents capable of solving complex tasks through planning, reasoning, tool use, and multi-turn in

agentsarxiv-cs-ai
15 May 2026
Local Ai

Small, Private Language Models as Teammates for Educational Assessment Design

DGX agent

arXiv:2605.15015v1 Announce Type: new Abstract: Generative AI increasingly supports educational design tasks, e.g., through Large Language Models (LLMs), demonstrating the capability to design assessm

local-aiarxiv-cs-ai
15 May 2026
Safety

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

DGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

safetyarxiv-cs-ro
15 May 2026
Research

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models

DGX agent

arXiv:2605.13010v1 Announce Type: cross Abstract: We study image inpainting with generative diffusion models. Existing methods typically either train dedicated task-specific models, or adapt a pretrai

researcharxiv-cs-ai
14 May 2026
Model Releases

Bias In, Bias Out? Finding Unbiased Subnetworks in Vanilla Models

DGX agent

arXiv:2603.05582v2 Announce Type: replace-cross Abstract: The issue of algorithmic biases in deep learning has led to the development of various debiasing techniques, many of which perform complex tra

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Continual Fine-Tuning of Large Language Models via Program Memory

DGX agent

arXiv:2605.13162v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT), particularly Low-Rank Adaptation (LoRA), has become a standard approach for adapting Large Language Models (LLMs

model-releasesarxiv-cs-lg
14 May 2026
Agents

Earth Science Foundation Models: From Perception to Reasoning and Discovery

DGX agent

arXiv:2605.12542v1 Announce Type: cross Abstract: Large foundation models (FMs) are transforming Earth science by integrating heterogeneous multimodal data, such as multi-platform imagery, gridded rea

agentsarxiv-cs-lg
14 May 2026
Model Releases

Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention

DGX agent

arXiv:2605.12789v1 Announce Type: new Abstract: Large language-vision models (LVLMs) such as CLIP, Flamingo, and BLIP have revolutionized AI by enabling understanding across textual and visual modalit

model-releasesarxiv-cs-ro
14 May 2026
Research

Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems

DGX agent

arXiv:2512.08411v2 Announce Type: replace Abstract: Model-based planning in robotic domains is challenged by the hybrid nature of physical dynamics, where continuous motion is punctuated by discrete e

researcharxiv-cs-ai
14 May 2026
Applications

RotVLA: Rotational Latent Action for Vision-Language-Action Model

DGX agent

arXiv:2605.13403v1 Announce Type: cross Abstract: Latent Action Models (LAMs) have emerged as an effective paradigm for handling heterogeneous datasets during Vision-Language-Action (VLA) model pretra

applicationsarxiv-cs-cv
14 May 2026
Research

Sampling from Flow Language Models via Marginal-Conditioned Bridges

DGX agent

arXiv:2605.13681v1 Announce Type: new Abstract: Flow Language Models (FLMs) are a recently introduced class of language models which adapt continuous flow matching for one-hot encoded token sequences.

researcharxiv-cs-lg
14 May 2026
Research

Three-Stage Learning Unlocks Strong Performance in Simple Models for Long-Term Time Series Forecasting

DGX agent

arXiv:2605.13678v1 Announce Type: new Abstract: Recent studies on long-term time series forecasting have shown that simple linear models and MLP-based predictors can achieve strong performance without

researcharxiv-cs-lg
14 May 2026
Model Releases

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

DGX agent

arXiv:2605.13119v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective robot action executors, but they remain limited on long-horizon tasks due to the dual burden of exte

model-releasesarxiv-cs-ai
14 May 2026
Agents

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context

DGX agent

arXiv:2605.13831v1 Announce Type: new Abstract: Long-context modeling is becoming a core capability of modern large vision-language models (LVLMs), enabling sustained context management across long-do

agentsarxiv-cs-cv
14 May 2026
Research

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone

DGX agent

arXiv:2605.11405v1 Announce Type: new Abstract: Data curation has shifted the quality-compute frontier for language-model and contrastive image-text pretraining, but its role for vision-language model

researcharxiv-cs-lg
13 May 2026
Model Releases

A Proof-of-Concept Simulation-Driven Digital Twin Framework for Decision-Aware Diabetes Modeling

DGX agent

arXiv:2605.11247v1 Announce Type: new Abstract: This paper presents a proof-of-concept digital twin framework for simulation-driven diabetes modeling using benchmark clinical data, synthetic temporal

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

A Theoretical Analysis of Why Masked Diffusion Models Mitigate the Reversal Curse

DGX agent

arXiv:2602.02133v2 Announce Type: replace-cross Abstract: Autoregressive language models (ARMs) suffer from the reversal curse: after learning ''A is B,'' they often fail on the reverse query ''B is A

model-releasesarxiv-cs-cl
13 May 2026
Research

Curriculum Learning-Guided Progressive Distillation in Large Language Models

DGX agent

arXiv:2605.11260v1 Announce Type: new Abstract: Knowledge distillation is a key technique for transferring the capabilities of large language models (LLMs) into smaller, more efficient student models.

researcharxiv-cs-lg
13 May 2026
Safety

Debiased Model-based Representations for Sample-efficient Continuous Control

DGX agent

arXiv:2605.11711v1 Announce Type: new Abstract: Model-based representations recently stand out as a promising framework that embeds latent dynamics information into the representations for downstream

safetyarxiv-cs-lg
13 May 2026
Research

Stop Marginalizing My Dreams: Model Inversion via Laplace Kernel for Continual Learning

DGX agent

arXiv:2605.11804v1 Announce Type: cross Abstract: Data-free continual learning (DFCIL) relies on model inversion to synthesize pseudo-samples and mitigate catastrophic forgetting. However, existing in

researcharxiv-cs-cv
13 May 2026
Research

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling

DGX agent

arXiv:2605.05922v2 Announce Type: replace Abstract: Recent advances in generative video models are increasingly driven by post-training and test-time scaling, both of which critically depend on the qu

researcharxiv-cs-cv
13 May 2026
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Applications

Counterfactual Stress Testing for Image Classification Models

DGX agent

arXiv:2605.10894v1 Announce Type: new Abstract: Deep learning models in medical imaging often fail when deployed in new clinical environments due to distribution shifts in demographics, scanner hardwa

applicationsarxiv-cs-cv
12 May 2026
Model Releases

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving

DGX agent

arXiv:2605.10426v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for end-to-end autonomous driving. However, existing reasoning mechanisms sti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CrystalREPA: Transferring Physical Priors from Universal MLIPs to Crystal Generative Models

DGX agent

arXiv:2605.08960v1 Announce Type: cross Abstract: Crystal generative models mainly learn what stable crystals look like, with little explicit supervision for what makes them stable. We reveal a substa

model-releasesarxiv-cs-lg
12 May 2026
Safety

Data-driven transport modelling without overfit

DGX agent

arXiv:2605.08801v1 Announce Type: new Abstract: Macroscopic transport modelling aims to predict traffic flows after proposed public policy interventions, such as a new road or railway section or a tem

safetyarxiv-cs-lg
12 May 2026
← Previous
1…8889909192…1030
Next →