AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Research

Operator-on-F complements value-equivalence: a planning-time diagnostic for latent world models

DGX agent

arXiv:2607.04464v1 Announce Type: cross Abstract: World-model evaluation for model-based reinforcement learning typically asks whether the learned model predicts reward and value well, which can leave

researcharxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TESSERA v2: Scaling Pixel-wise Earth Foundation Models

DGX agent

arXiv:2607.03949v1 Announce Type: new Abstract: Pixel-wise Earth-observation (EO) foundation models are now achieving state-of-the-art performance via generated spatial embeddings. However, how these

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Toward Cybersecurity-Expert Small Language Models

DGX agent

arXiv:2510.14113v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming everyday applications, yet deployment in cybersecurity lags due to a lack of high-quality, domai

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

ClawArena-Team: Benchmarking Subagent Orchestration and Dynamic Workflows in Language-Model Agents

DGX agent

arXiv:2606.31174v1 Announce Type: new Abstract: Production large language-model (LLM) agents are increasingly deployed not as lone problem-solvers but as managers: a main model creates specialized sub

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

REMSA: Foundation Model Selection for Remote Sensing via a Constraint-Aware Agent

DGX agent

arXiv:2511.17442v3 Announce Type: replace-cross Abstract: Foundation Models (FMs) are increasingly integrated into remote sensing (RS) pipelines. These models include unimodal vision encoders and mult

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

On the Vulnerability of Parameter-Level Defenses to Model Merging

DGX agent

arXiv:2606.30360v1 Announce Type: cross Abstract: The training-free integration of expert models via model merging has exposed significant security risks, enabling free-riders to combine specialized m

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

StackingNet: Collective Inference Across Independent AI Foundation Models

DGX agent

arXiv:2602.13792v2 Announce Type: replace Abstract: Artificial intelligence built on large foundation models has transformed language understanding, computer vision, and reasoning, yet these systems r

safetyarxiv-cs-ai
30 Jun 2026
Applications

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models

DGX agent

arXiv:2602.22960v2 Announce Type: replace Abstract: World models based on video generation demonstrate remarkable potential for simulating interactive environments yet suffer from persistent difficult

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

Benchmarking Open-Weight Foundation Models for Global AI Technical Governance

DGX agent

arXiv:2606.26099v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in artificial intelligence (AI) governance analysis across national and international organisat

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Nemotron-TwoTower: Diffusion Language Modeling with Pretrained Autoregressive Context

DGX agent

arXiv:2606.26493v1 Announce Type: new Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their potential for parallel and iterative generation. However,

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

The Capability Frontier: Benchmarks Miss 82% of Model Performance

DGX agent

arXiv:2606.26836v1 Announce Type: new Abstract: Existing benchmarks typically report accuracy for a single model on a single run. This systematically understates real-world LLM capabilities, particula

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Do vision-language models search like humans? Reasoning tokens as a reaction-time analog in classic visual-search paradigms

DGX agent

arXiv:2606.25066v1 Announce Type: cross Abstract: Visual search has been one of the most productive paradigms in the study of visual attention: the way reaction time scales with the number of items di

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Are Text-to-Image Models Inductivist Turkeys? A Counterfactual Benchmark for Causal Reasoning

DGX agent

arXiv:2606.24548v1 Announce Type: new Abstract: Text-to-image (T2I) generation models have achieved remarkable progress in producing visually realistic images from natural language prompts. Yet it rem

model-releasesarxiv-cs-cv
24 Jun 2026
Research

Can Reasoning Models Detect Changes to their Chains of Thought?

DGX agent

arXiv:2606.22085v1 Announce Type: cross Abstract: There are many reasons one may want to edit a model's chain of thought (CoT) -- e.g., to prefill it with reasoning from a stronger model or to remove

researcharxiv-cs-lg
23 Jun 2026
Tutorials

CogFormer: Learn All Your Models Once

DGX agent

arXiv:2603.20520v2 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) with neural networks has accelerated and transformed cognitive modeling workflows. SBI enables modelers to fi

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Comparative Evaluation of Machine Learning and Deep Learning Models for Wound-Rotor Synchronous Motor Performance Prediction

DGX agent

arXiv:2606.21230v1 Announce Type: new Abstract: Wound rotor synchronous motors have emerged as a strong alternative that eliminates dependence on REEs. However, WRSM design requires the simultaneous o

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Stealthy World Model Manipulation via Data Poisoning

DGX agent

arXiv:2606.18697v2 Announce Type: replace Abstract: Model-based learning agents use learned world models to predict future states, plan actions, and adapt to new environments. However, the process of

researcharxiv-cs-lg
23 Jun 2026
Agents

UniviewVLA: A Unified Multiview Vision-Language-Action Model with World Modeling

DGX agent

arXiv:2606.21501v1 Announce Type: new Abstract: Occluded tasks remain a bottleneck in robot manipulation. Existing solutions either deploy additional physical cameras requiring training-inference came

agentsarxiv-cs-ro
23 Jun 2026
Model Releases

Vision-language models for chest radiography do not always need the image

DGX agent

arXiv:2606.17710v2 Announce Type: replace Abstract: Medical vision-language models report strong chest radiograph accuracy, and this is increasingly read as evidence that they use the image. That infe

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization

DGX agent

arXiv:2606.12016v1 Announce Type: cross Abstract: Model post-training, and in particular reinforcement learning (RL), is one of the primary mechanisms by which developers can shape models' values and

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Making Models Unmergeable via Scaling-Sensitive Loss Landscape

DGX agent

arXiv:2601.21898v2 Announce Type: replace Abstract: The rise of model hubs has made it easier to access reusable model components, making model merging a practical tool for combining capabilities. Yet

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Mapping Scientific Literature with Large Language Models and Topic Modeling

DGX agent

arXiv:2510.16152v2 Announce Type: replace-cross Abstract: Scientific literature is increasingly fragmented by disciplinary boundaries, specialized terminology, and potentially sparse keyword systems,

researcharxiv-cs-ai
11 Jun 2026
Applications

Physics-Distilled Neural Network enabled by Large Language Models for Manufacturing Process-Property Predictive Modeling

DGX agent

arXiv:2606.11605v1 Announce Type: cross Abstract: Predicting process-property relationships in manufacturing is often challenged by high experimental costs and the limited interpretability of complex

applicationsarxiv-cs-ai
11 Jun 2026
Safety

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

DGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

safetyarxiv-cs-ai
9 Jun 2026
Applications

Zero and Few Shot Load Forecasting with Large Language Models

DGX agent

arXiv:2411.11350v2 Announce Type: replace Abstract: Deep learning models have shown strong performance in load forecasting, but they generally require large amounts of data for model training before b

applicationsarxiv-cs-lg
9 Jun 2026
Research

A Conformation-Centric Generative Foundation Model for Linear Polymer Modeling and Design

DGX agent

arXiv:2510.16023v2 Announce Type: replace Abstract: Linear polymers, macromolecules formed from monomers covalently bonded into continuous chains, underpin countless technologies and are indispensable

researcharxiv-cs-lg
8 Jun 2026
Research

Modeling semantic association in self-paced reading with language model embeddings

DGX agent

arXiv:2606.07066v1 Announce Type: new Abstract: Semantic association between a word and its context has been identified as an important component of reading comprehension, even when word predictabilit

researcharxiv-cs-cl
8 Jun 2026
Research

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

DGX agent

arXiv:2606.06094v1 Announce Type: new Abstract: Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved di

researcharxiv-cs-ai
6 Jun 2026
Model Releases

An Empirical Study of Data Scale, Model Complexity, and Input Modalities in Visual Generalization

DGX agent

arXiv:2606.04409v1 Announce Type: cross Abstract: Modern deep neural networks usually have large parameter scales and nonlinear hierarchical structures, and they have achieved strong performance in co

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Geometry-Preserving Unsupervised Alignment for Heterogeneous Foundation Models

DGX agent

arXiv:2606.04385v1 Announce Type: new Abstract: Foundation models have driven rapid progress in computer vision, yet the two dominant paradigms, vision-language foundation models (VLMs) and vision-onl

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

DGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Cosmos 3: Omnimodal World Models for Physical AI

DGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

How Quantization Changes Interpretable Features: A Sparse Autoencoder Analysis of Language Models

DGX agent

arXiv:2606.03002v1 Announce Type: cross Abstract: Quantization is a standard path to deploying large language models, and a quantized model is typically judged acceptable when its perplexity or downst

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Balancing Accuracy and Efficiency: Adaptive Dynamics Orchestration for Model Predictive Control

DGX agent

arXiv:2606.00085v1 Announce Type: new Abstract: Model Predictive Control (MPC) for autonomous navigation faces a fundamental trade-off between model accuracy and real-time efficiency. High-fidelity dy

safetyarxiv-cs-ro
2 Jun 2026
Safety

Expected Value Alignment for Generative Reward Modeling in Formal Mathematics Verification

DGX agent

arXiv:2606.01160v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used with formal interactive theorem provers such as Lean 4. Scaling these systems with reinforcement lear

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models

DGX agent

arXiv:2504.03635v4 Announce Type: replace Abstract: Reasoning is a core capability of language models (LMs), yet it remains unclear how much model capacity is necessary to support reasoning during pre

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Interpretable Modeling of Driver Attention Shifts with a Vision--Language Model

DGX agent

arXiv:2508.05852v2 Announce Type: replace Abstract: Driver gaze is commonly modeled as a spatial heatmap, but heatmaps alone are difficult for humans to interpret because they do not explain which roa

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

RynnVLA-002: A Unified Vision-Language-Action and World Model

DGX agent

arXiv:2511.17502v3 Announce Type: replace Abstract: We introduce RynnVLA-002, a unified Vision-Language-Action (VLA) and world model. The world model leverages action and visual inputs to predict futu

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Uncovering Competency Gaps in Large Language Models and Their Benchmarks

DGX agent

arXiv:2512.20638v2 Announce Type: replace-cross Abstract: The evaluation of large language models relies heavily on standardized benchmarks. These benchmarks provide useful aggregated metrics, but can

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Dreaming Of Others: Latent Teammate Modeling In World Models For Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.31361v1 Announce Type: cross Abstract: In cooperative multi-agent reinforcement learning (MARL), agents must coordinate with partners whose internal policies and intentions are not directly

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

EvoDefense: Co-Evolving Black-Box Defense with Large Language Models

DGX agent

arXiv:2605.31140v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain highly vulnerable to diverse attacks, particularly in black-box settings where the internals of target models are

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Chess-World-Model: A 10M-Game Benchmark for Exact State Tracking from Chess Move Sequences

DGX agent

arXiv:2605.30100v1 Announce Type: new Abstract: World models require state tracking, which is the ability to maintain a correct latent state across action sequences. Existing benchmarks are often synt

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across Physical Regimes and Distribution Shifts

DGX agent

arXiv:2605.29283v1 Announce Type: cross Abstract: Recent physics foundation models claim general spatiotemporal forecasting ability, yet their evaluations often collapse performance into a single aver

model-releasesarxiv-cs-ai
29 May 2026
Safety

Draft-OPD: On-Policy Distillation for Speculative Draft Models

DGX agent

arXiv:2605.29343v1 Announce Type: new Abstract: Speculative decoding accelerates large language model inference by pairing a target model with a lightweight draft model whose proposed tokens are verif

safetyarxiv-cs-cl
29 May 2026
Model Releases

RightNow-Arabic-0.5B-Turbo: An Open Sub-1B Arabic Language Model via Vocabulary Injection and Edge-First Deployment

DGX agent

arXiv:2605.28827v1 Announce Type: new Abstract: Open Arabic large language models split into two classes: sub-1B multilingual models that treat Arabic as an afterthought (Qwen2.5-0.5B, Falcon-H1-0.5B)

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests

DGX agent

arXiv:2605.28734v1 Announce Type: cross Abstract: A general-purpose language model that answers a harmful question returns text; a coding model that complies with a malicious request can return a work

model-releasesarxiv-cs-cl
28 May 2026
Research

Continual Learning in Modern Hopfield Networks with an Application to Diffusion Models

DGX agent

arXiv:2605.27975v1 Announce Type: new Abstract: Generative models, including diffusion models, are increasingly used as foundation models and adapted through sequential fine-tuning, making continual l

researcharxiv-cs-lg
28 May 2026
← Previous
1…1213141516…1012
Next →