AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Research

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models

DGX agent

arXiv:2607.01774v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive models. Unlike standard diffusion-based approaches,

researcharxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FiLM-Coordinated Dual-Branch Transformer for Global-Local Dependency Modeling in Language Modeling

DGX agent

arXiv:2606.21075v1 Announce Type: cross Abstract: Standard Transformers use a single self-attention pathway to model both global dependencies and local patterns, creating tension between long-range st

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Your Model Already Knows: Attention-Guided Safety Filter for Vision-Language-Action Models

DGX agent

arXiv:2606.09749v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive end-to-end performance across a variety of robotic manipulation tasks. However, these

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

WorldFly: A World-Model-Based Vision-Language-Action Model for UAV Navigation

DGX agent

arXiv:2606.06147v1 Announce Type: new Abstract: End-to-end Vision-Language-Action (VLA) models have shown promise in UAV navigation. However, existing approaches typically rely on historical observati

model-releasesarxiv-cs-ai
6 Jun 2026
Hardware

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

DGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

hardwarearxiv-cs-ro
5 Jun 2026
Model Releases

Pairwise Reference Alignment as a Model-Level Ordinal Observable

DGX agent

arXiv:2605.30758v1 Announce Type: new Abstract: Pairwise preference data is widely used in language-model evaluation and alignment, often for model ranking, reward modeling, or preference optimization

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training

DGX agent

arXiv:2605.26484v1 Announce Type: new Abstract: Model merging has emerged as a lightweight paradigm for enhancing Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood.

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization

DGX agent

arXiv:2605.21751v1 Announce Type: new Abstract: Text-to-optimization requires two separable capabilities: modeling -- choosing the right optimization structure -- and binding -- grounding every coeffi

model-releasesarxiv-cs-lg
23 May 2026
Research

Effective Model Pruning: Measure The Redundancy of Model Components

DGX agent

arXiv:2509.25606v3 Announce Type: replace Abstract: This article initiates the study of a basic question about model pruning. Given a vector s of importance scores assigned to model components, how ma

researcharxiv-cs-lg
21 May 2026
Safety

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey

DGX agent

arXiv:2505.17352v2 Announce Type: replace Abstract: Diffusion models have become a central paradigm for image and multimodal generation, yet their deployment raises persistent questions about alignmen

safetyarxiv-cs-cv
19 May 2026
Model Releases

Solve the Loop: Attractor Models for Language and Reasoning

DGX agent

arXiv:2605.12466v1 Announce Type: cross Abstract: Looped Transformers offer a promising alternative to purely feed-forward computation by iteratively refining latent representations, improving languag

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model

DGX agent

arXiv:2605.01194v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities and generalization in embodied manipulation. However, their decision-makin

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

What Single-Prompt Accuracy Misses: A Multi-Variant Reliability Audit of Language Models

DGX agent

arXiv:2605.02038v1 Announce Type: new Abstract: Single-prompt accuracy is the dominant way to benchmark language models, but it can miss reliability failures that matter. We evaluate a 15-model open-w

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models

DGX agent

arXiv:2604.27124v1 Announce Type: new Abstract: Training stable biological foundation models requires rethinking attention mechanisms: we find that using sigmoid attention as a drop in replacement for

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Improving Vision-language Models with Perception-centric Process Reward Models

DGX agent

arXiv:2604.24583v1 Announce Type: new Abstract: Recent advancements in reinforcement learning with verifiable rewards (RLVR) have significantly improved the complex reasoning ability of vision-languag

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Lightweight Retrieval-Augmented Generation and Large Language Model-Based Modeling for Scalable Patient-Trial Matching

DGX agent

arXiv:2604.22061v1 Announce Type: cross Abstract: Patient-trial matching requires reasoning over long, heterogeneous electronic health records (EHRs) and complex eligibility criteria, posing significa

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models

DGX agent

arXiv:2604.19139v1 Announce Type: cross Abstract: As Large Language Models (LLMs) continue to evolve through alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) and Constitu

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

LiFT: Does Instruction Fine-Tuning Improve In-Context Learning for Longitudinal Modelling by Large Language Models?

DGX agent

arXiv:2604.16382v1 Announce Type: new Abstract: Longitudinal NLP tasks require reasoning over temporally ordered text to detect persistence and change in human behavior and opinions. However, in-conte

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value

DGX agent

arXiv:2506.13763v2 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in generative modeling. Despite more stable training, the loss of diffusion models is not in

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?

DGX agent

arXiv:2601.07220v3 Announce Type: replace Abstract: Multilingual language models (LMs) promise broader NLP access, yet current systems deliver uneven performance across the world's languages. This sur

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers

DGX agent

arXiv:2604.09489v1 Announce Type: cross Abstract: Model poisoning attacks pose a significant security threat to Federated Learning (FL). Most existing model poisoning attacks rely on collusion, requir

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

DGX agent

arXiv:2604.07430v1 Announce Type: new Abstract: We introduce HY-Embodied-0.5, a family of foundation models specifically designed for real-world embodied agents. To bridge the gap between general Visi

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

DGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

You Point, I Learn: Online Adaptation of Interactive Segmentation Models for Handling Distribution Shifts in Medical Imaging

DGX agent

arXiv:2503.06717v3 Announce Type: replace Abstract: Interactive segmentation uses real-time user inputs, such as mouse clicks, to iteratively refine model predictions. Although not originally designed

model-releasesarxiv-cs-cv
10 Apr 2026
Research

ChronoSSM: Training for Temporally Aware Representations in Autoregressive State Space Models

DGX agent

arXiv:2608.10120v1 Announce Type: new Abstract: Modern sequence models, from Transformers to State Space Models, have enabled powerful generative modeling across diverse domains, yet they are typicall

researcharxiv-cs-lg
12 Aug 2026
Model Releases

Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling

DGX agent

arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully crafted formulations. While recent works ha

model-releasesarxiv-cs-ai
10 Aug 2026
Local Ai

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

DGX agent

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localiz

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

A Unified Model for Cross-Domain Clone Detection via Model Merging

DGX agent

arXiv:2608.04215v1 Announce Type: cross Abstract: The growing diversity of code clone types, from syntactic copies to cross-language semantic clones to AI-generated duplicates, has created a fragmenta

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Large-Small Model Collaboration for Enhancing Edge-Deployed Small Models

DGX agent

arXiv:2503.10367v2 Announce Type: replace-cross Abstract: Edge devices host domain-specific small language models (SLMs) with limited resources, while private clouds offer larger LLMs. We propose G-Bo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Language Models Agree With Each Other, Not With Readers

DGX agent

arXiv:2607.29274v1 Announce Type: cross Abstract: Claims that language models homogenise are usually measured against human judgements collected for the study, which makes the human side an artifact o

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

DGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Epistemic diversity across language models mitigates knowledge collapse

DGX agent

arXiv:2512.15011v3 Announce Type: replace Abstract: Artificial intelligence (AI) increasingly generates the very content used to train future AI systems. This feedback loop can degrade model quality,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

DenseOn with the LateOn: Fully Open Dense and Late-Interaction Models for Multilingual, Long-Context, and Code Search

DGX agent

arXiv:2607.27178v1 Announce Type: new Abstract: State-of-the-art retrieval models increasingly rely on closed training data, creating a reproducibility gap. We present an open end-to-end recipe for tr

model-releasesarxiv-cs-cl
30 Jul 2026
Applications

CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer

DGX agent

arXiv:2607.26023v1 Announce Type: new Abstract: Graph foundation models (GFMs) have emerged as a promising paradigm for transferring knowledge across graph domains and tasks. Real-world graphs associa

applicationsarxiv-cs-ai
29 Jul 2026
Research

PersGuard: Preventing Malicious Personalization in Text-to-Image Diffusion Models via Model Backdoors

DGX agent

arXiv:2502.16167v2 Announce Type: replace-cross Abstract: Diffusion models (DMs) have advanced text-to-image (T2I) synthesis, yet their personalization capabilities raise serious privacy and copyright

researcharxiv-cs-ai
16 Jul 2026
Model Releases

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

DGX agent

arXiv:2607.12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer ever

model-releasesarxiv-cs-ai
15 Jul 2026
Applications

Xray-Visual Models: Scaling Vision models on Industry Scale Data

DGX agent

arXiv:2602.16918v2 Announce Type: replace-cross Abstract: We present Xray-Visual, a unified vision model architecture for large-scale image and video understanding trained on industry-scale social med

applicationsarxiv-cs-ai
15 Jul 2026
Tutorials

A Survey of Circuit Foundation Model: Foundation AI Models for VLSI Circuit Design and EDA

DGX agent

arXiv:2504.03711v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI)-driven electronic design automation (EDA) techniques have been extensively explored for VLSI circuit design appli

tutorialsarxiv-cs-lg
3 Jul 2026
Safety

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

DGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

safetyarxiv-cs-ai
3 Jul 2026
Safety

From World Models to World Action Models: A Concise Tutorial for Robotics

DGX agent

arXiv:2607.00836v1 Announce Type: cross Abstract: World models are increasingly used in embodied intelligence and generative simulation, yet their scope remains ambiguous across communities. This tuto

safetyarxiv-cs-ai
2 Jul 2026
Tutorials

Deep probabilistic model synthesis enables unified modeling of whole-brain neural activity across individual subjects

DGX agent

arXiv:2603.14161v2 Announce Type: replace Abstract: Many disciplines need quantitative models that synthesize experimental data across multiple instances of the same general system. For example, neuro

tutorialsarxiv-cs-lg
30 Jun 2026
Agents

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

DGX agent

arXiv:2606.28524v1 Announce Type: new Abstract: Recent work suggests that Large Language Models (LLMs) are sensitive to the belief states of agents described by text, as measured by the false belief t

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

Qwen-AgentWorld: Language World Models for General Agents

DGX agent

arXiv:2606.24597v1 Announce Type: new Abstract: A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning.

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Modeling Complex Behaviors: Multi-Personality Composition and Dynamic Switching in Vision-Language Models

DGX agent

arXiv:2606.11074v1 Announce Type: cross Abstract: With the widespread deployment of Multimodal Large Language Models (MLLMs) in social interaction, understanding and controlling their behavior under c

model-releasesarxiv-cs-ai
10 Jun 2026
Research

MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action Models

DGX agent

arXiv:2606.09827v1 Announce Type: cross Abstract: Temporal modeling is essential for robotic manipulation, as effective control requires both memory of past interactions and imagination of future stat

researcharxiv-cs-cv
9 Jun 2026
Tutorials

Sparse Mixture-of-Experts Reward Models Learn Interpretable and Specialized Experts for Personalized Preference Modeling

DGX agent

arXiv:2606.04284v1 Announce Type: cross Abstract: Preference modeling plays a central role in reinforcement learning from human feedback (RLHF), enabling large language models (LLMs) to align with hum

tutorialsarxiv-cs-ai
4 Jun 2026
Agents

All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning

DGX agent

arXiv:2606.01363v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) infers information about the environment from a learned dynamics model and bears the potential to address open

agentsarxiv-cs-lg
2 Jun 2026
Model Releases

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining

DGX agent

arXiv:2510.06048v4 Announce Type: replace Abstract: Effective data selection is essential for pretraining large language models (LLMs), enhancing efficiency and improving generalization to downstream

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
123456…1012
Next →