AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

RAIGen: Rare Attribute Identification in Text-to-Image Generative Models

DGX agent

arXiv:2602.06806v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve impressive generation quality but inherit and amplify training-data biases, skewing coverage of semantic attr

safetyarxiv-cs-cv
2 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ResMerge: Residual-based Spectral Merging of Large Language Models

DGX agent

arXiv:2606.02252v1 Announce Type: new Abstract: Model merging offers a training-free way to combine multiple post-trained expert models, but merging experts obtained through reinforcement learning (RL

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Retrieval-aligned Tabular Foundation Models Enable Robust Clinical Risk Prediction in Electronic Health Records Under Real-world Constraints

DGX agent

arXiv:2604.01841v2 Announce Type: replace Abstract: Clinical prediction from structured electronic health records (EHRs) is challenging due to high dimensionality, heterogeneity, class imbalance, and

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

RuleEdit: Failure-Guided Human-AI Model Editing with Prospective Impact Preview

DGX agent

arXiv:2606.00011v1 Announce Type: cross Abstract: Despite the promise of AI to assist complex decisions, practitioners still lack ways to detect likely failures and inspect the consequences of model e

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation

DGX agent

arXiv:2603.09292v2 Announce Type: replace-cross Abstract: Measurement of task progress through explicit, actionable milestones is critical for robust robotic manipulation. This progress awareness enab

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Single-Channel Tissue Segmentation via Cross-Modal Distillation from Foundation Models

DGX agent

arXiv:2606.00928v1 Announce Type: new Abstract: Multiplexed fluorescence microscopy improves tissue segmentation by providing complementary channels including nuclear (DAPI) and membrane (E-cadherin),

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement

DGX agent

arXiv:2606.00267v1 Announce Type: cross Abstract: Video world models (WMs) have shown promise for policy evaluation and improvement by imagining realistic future observations conditioned on ego-robot

safetyarxiv-cs-ai
2 Jun 2026
Research

Understanding the Effects of Distractors on Reasoning Vision-Language Models

DGX agent

arXiv:2511.21397v2 Announce Type: replace-cross Abstract: How does irrelevant information (i.e., distractors) affect test-time scaling in vision-language models (VLMs)? Prior work on text-only languag

researcharxiv-cs-ai
2 Jun 2026
Model Releases

VESTA: Visual Exploration with Statistical Tool Agents

DGX agent

arXiv:2606.00384v1 Announce Type: new Abstract: Fitting quantitative models to data is a central step in scientific workflows, yet it remains one of the least automated. Recent agent-based systems lev

model-releasesarxiv-cs-ai
2 Jun 2026
Research

AR Forcing: Towards Long-Horizon Robot Navigation World Model

DGX agent

arXiv:2605.31314v1 Announce Type: new Abstract: The diffusion based robot navigation world models are typically trained using parallel supervision, while autoregressive inference is employed during pa

researcharxiv-cs-ro
1 Jun 2026
Model Releases

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

DGX agent

arXiv:2503.08679v5 Announce Type: replace Abstract: Recent studies indicate that when faced with explicit biases in prompts, models often omit mentioning these biases in their Chain-of-Thought (CoT) o

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Configurable Reward Model for Balanced Safety Alignment

DGX agent

arXiv:2605.30487v1 Announce Type: new Abstract: Aligning large language models (LLMs) to heterogeneous and rapidly evolving safety requirements remains a critical challenge. Existing instruction-tuned

safetyarxiv-cs-cl
1 Jun 2026
Safety

Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?

DGX agent

arXiv:2605.31041v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal arc

safetyarxiv-cs-ai
1 Jun 2026
Research

DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training

DGX agent

arXiv:2512.13996v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts architectures are essential for scaling model capacity efficiently, yet the standard Top-k routing imposes a rigid sparsit

researcharxiv-cs-ai
1 Jun 2026
Local Ai

Federated Learning with Enhanced Privacy via Model Splitting and Random Client Participation

DGX agent

arXiv:2509.25906v2 Announce Type: replace Abstract: Federated Learning (FL) often adopts differential privacy (DP) to protect client data, but the added noise required for privacy guarantees can subst

local-aiarxiv-cs-lg
1 Jun 2026
Applications

FlexRank: Nested Low-Rank Knowledge Decomposition for Adaptive Model Deployment

DGX agent

arXiv:2602.02680v2 Announce Type: replace Abstract: The growing scale of deep neural networks, encompassing large language models (LLMs) and vision transformers (ViTs), has made training from scratch

applicationsarxiv-cs-lg
1 Jun 2026
Research

Flow map learning in nonlinear vector autoregressive models: influence of the feature-library structure on the training error

DGX agent

arXiv:2605.31438v1 Announce Type: new Abstract: Time series forecasting often requires learning nonlinear and time-delayed dependencies. A paradigmatic class of forecasting models are nonlinear vector

researcharxiv-cs-lg
1 Jun 2026
Agents

Generalized Intention Modeling in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.31318v1 Announce Type: new Abstract: Modeling an opponent's intent is critical for effective decision-making in non-cooperative, competitive, and general-sum multi-agent reinforcement learn

agentsarxiv-cs-lg
1 Jun 2026
Safety

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model

DGX agent

arXiv:2605.31234v1 Announce Type: new Abstract: Learning generalizable vision-language-action (VLA) models from large-scale human videos is promising but challenging due to cross-embodiment discrepanc

safetyarxiv-cs-ro
1 Jun 2026
Tutorials

How can embedding models bind concepts?

DGX agent

arXiv:2605.31503v1 Announce Type: new Abstract: Humans easily determine which color belongs to which shape in multi-object scenes, an ability known as concept binding. Vision-language embedding models

tutorialsarxiv-cs-cv
1 Jun 2026
Safety

Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty

DGX agent

arXiv:2605.30675v1 Announce Type: cross Abstract: Uncertainty Quantification is a large and growing subfield of large language model behavioral analysis. Primarily to recognize and combat hallucinatio

safetyarxiv-cs-ai
1 Jun 2026
Research

Inferring Events from Time Series using Language Models

DGX agent

arXiv:2503.14190v3 Announce Type: replace Abstract: A common goal in analyzing time series data is to understand how events cause observed variations. We study whether Large Language Models (LLMs) can

researcharxiv-cs-ai
1 Jun 2026
Local Ai

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models

DGX agent

arXiv:2605.31158v1 Announce Type: new Abstract: Interactive video world models generate video chunk by chunk in response to user-controlled camera movements, enabling applications such as real-time ga

local-aiarxiv-cs-cv
1 Jun 2026
Model Releases

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration

DGX agent

arXiv:2605.31196v1 Announce Type: cross Abstract: Safe human--robot collaboration requires more than visual description: a monitor must determine whether the robot body is safely separated, already co

model-releasesarxiv-cs-ai
1 Jun 2026
Tutorials

Riemannian Diffusion Models on General Manifolds via Physics-Informed Neural Networks

DGX agent

arXiv:2605.31106v1 Announce Type: new Abstract: Riemannian diffusion models generalize score-based generative modeling to manifold-supported data via stochastic diffusion equations on the manifold. Ho

tutorialsarxiv-cs-lg
1 Jun 2026
Safety

Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents

DGX agent

arXiv:2605.30723v1 Announce Type: new Abstract: LLM agents increasingly retrieve externally curated skills-procedural instructions retrieved at decision time-to improve performance on long-horizon int

safetyarxiv-cs-cl
1 Jun 2026
Model Releases

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

DGX agent

arXiv:2602.22971v2 Announce Type: replace Abstract: As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchma

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning

DGX agent

arXiv:2605.31404v1 Announce Type: cross Abstract: Large Language Model (LLM)-based navigation systems commonly construct explicit spatial representations (e.g., topological graphs, semantic raster map

safetyarxiv-cs-ai
1 Jun 2026
Research

Unicorn: Scaling High-Dimensional Time Series Forecasting via Universal Correlation Modeling

DGX agent

arXiv:2605.30376v1 Announce Type: cross Abstract: Modern time series architectures face a fundamental trade-off: channel-independent models scale well with increasing data volume but ignore critical i

researcharxiv-cs-ai
1 Jun 2026
Research

Updating the standard neuron model in artificial neural networks

DGX agent

arXiv:2605.30370v1 Announce Type: cross Abstract: From their inception in the 1950s, artificial neural networks (ANNs) started using the so-called point neuron model then prevalent in neuroscience, ho

researcharxiv-cs-ai
1 Jun 2026
Local Ai

Vector Linking via Cross-Model Local Isometric Consistency

DGX agent

arXiv:2605.31100v1 Announce Type: new Abstract: We study Vector Linking: given two embedding clouds produced by different black-box encoders over partially overlapping datasets, recover cross-model ob

local-aiarxiv-cs-ai
1 Jun 2026
Model Releases

A comparative study of transformer-based embeddings for topic coherence

DGX agent

arXiv:2605.28832v1 Announce Type: cross Abstract: Topic modeling is a branch of Natural Language Processing (NLP) that aims to organize large collections of texts into coherent groups according to wor

model-releasesarxiv-cs-ai
29 May 2026
Research

Archon: A Unified Multimodal Model for Holistic Digital Human Generation

DGX agent

arXiv:2605.30311v1 Announce Type: cross Abstract: Digital humans are fundamental to immersive interaction, yet creating a unified model for holistic modalities, including text, audio, motion, and visu

researcharxiv-cs-ai
29 May 2026
Applications

BuilDyn: Excitation-Driven Data Generation for Building Thermal Dynamics Modeling and Control

DGX agent

arXiv:2605.29849v1 Announce Type: cross Abstract: Machine learning (ML) is increasingly used for data-driven modeling of buildings to enable downstream tasks such as fault detection and diagnosis, and

applicationsarxiv-cs-lg
29 May 2026
Research

Calibrating Generative Models to Distributional Constraints

DGX agent

arXiv:2510.10020v4 Announce Type: replace-cross Abstract: Generative models frequently suffer miscalibration, wherein statistics of the sampling distribution, such as the fraction of generations in a

researcharxiv-cs-lg
29 May 2026
Research

Diffusion Models Are Statistically Optimal for Learning Low-Dimensional Multi-Modal Distributions

DGX agent

arXiv:2605.30153v1 Announce Type: cross Abstract: Score-based diffusion models have demonstrated remarkable empirical success in learning high-dimensional distributions, particularly those exhibiting

researcharxiv-cs-lg
29 May 2026
Model Releases

FinVerBench: Benchmark Validity and Calibration in Large Language Model Financial Statement Verification

DGX agent

arXiv:2605.29586v1 Announce Type: new Abstract: We introduce FinVerBench, a benchmark and validity study for financial statement verification: determining whether a set of corporate financial statemen

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

Learning to Feel Materials from Multisensory Tactile Data via Interpretable Models

DGX agent

arXiv:2605.29572v1 Announce Type: new Abstract: Human tactile perception of materials relies on complex multisensory touch cues, yet the relationship between low-level tactile signals and perceptual r

tutorialsarxiv-cs-ro
29 May 2026
Safety

Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance

DGX agent

arXiv:2411.14279v2 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have achieved impressive results in various vision-language tasks. However, despite showing promising per

safetyarxiv-cs-cl
29 May 2026
Research

Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models

DGX agent

arXiv:2601.14758v4 Announce Type: replace-cross Abstract: Post-training pretrained autoregressive models (ARMs) into masked diffusion models (MDMs) has emerged as a cost-effective way to overcome the

researcharxiv-cs-ai
29 May 2026
Safety

Reasoning While Asking: Transforming Reasoning Large Language Models from Passive Solvers to Proactive Inquirers

DGX agent

arXiv:2601.22139v2 Announce Type: replace-cross Abstract: Reasoning-oriented Large Language Models (LLMs) have achieved remarkable progress with Chain-of-Thought (CoT) prompting, yet they remain funda

safetyarxiv-cs-ai
29 May 2026
Safety

Same Evidence, Different Answers: Canonical-Context On-Policy Distillation for Multi-Turn Language Models

DGX agent

arXiv:2605.30251v1 Announce Type: cross Abstract: Large language models (LLMs) often solve a task when all instructions are given in a single prompt, but fail when the same information is revealed gra

safetyarxiv-cs-ai
29 May 2026
Model Releases

The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language Modeling

DGX agent

arXiv:2605.28864v1 Announce Type: new Abstract: The Cognitive Categorical Transformer (CCT) is a 306M-parameter architecture that augments a pretrained GPT-2 Small backbone with cognitively grounded c

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning

DGX agent

arXiv:2605.28842v1 Announce Type: cross Abstract: The success of large language models (LLMs) across diverse NLP tasks has elevated the importance of reasoning chain optimization as a critical step in

tutorialsarxiv-cs-ai
29 May 2026
Research

Unlocking the Working Memory of Large Language Models for Latent Reasoning

DGX agent

arXiv:2605.30343v1 Announce Type: cross Abstract: To improve the reasoning capabilities of large language models, test-time compute is typically scaled by generating intermediate tokens before the fin

researcharxiv-cs-ai
29 May 2026
Research

Unveiling the Visual Counting Bottleneck in Vision-Language Models

DGX agent

arXiv:2605.30170v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) excel at interpolation, they suffer catastrophic failures in systematic generalization, most notably in visu

researcharxiv-cs-cv
29 May 2026
Research

WaterSearch: A Quality-Aware Search-based Watermarking Framework for Large Language Models

DGX agent

arXiv:2512.00837v2 Announce Type: replace Abstract: Watermarking acts as a critical safeguard in text generated by Large Language Models (LLMs). By embedding identifiable signals into model outputs, w

researcharxiv-cs-cl
29 May 2026
Model Releases

Benchmarking and Mechanistic Analysis of Vision-Language Models for Cross-Depiction Assembly Instruction Alignment

DGX agent

arXiv:2604.00913v2 Announce Type: replace-cross Abstract: 2D assembly diagrams are often abstract and hard to follow, creating a need for intelligent assistants that can monitor progress, detect error

model-releasesarxiv-cs-cl
28 May 2026
← Previous
1…132133134135136…1030
Next →