AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Self-supervised Pretraining of Cell Segmentation Models

DGX agent

arXiv:2604.10609v1 Announce Type: new Abstract: Instance segmentation enables the analysis of spatial and temporal properties of cells in microscopy images by identifying the pixels belonging to each

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Symmetry-Aware Generative Modeling through Learned Canonicalization

DGX agent

arXiv:2501.07773v3 Announce Type: replace Abstract: Generative modeling of symmetric densities has a range of applications in AI for science, from drug discovery to physics simulations. The existing g

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

DGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching

DGX agent

arXiv:2604.08574v1 Announce Type: cross Abstract: Large Genomic Foundation Models have recently achieved remarkable results and in-vivo translation capabilities. However these models quickly grow to o

researcharxiv-cs-ai
13 Apr 2026
Safety

Learning Vision-Language-Action World Models for Autonomous Driving

DGX agent

arXiv:2604.09059v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved notable progress in end-to-end autonomous driving by integrating perception, reasoning, and

safetyarxiv-cs-ai
13 Apr 2026
Safety

Post-Selection Distributional Model Evaluation

DGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest

DGX agent

arXiv:2604.08525v1 Announce Type: cross Abstract: Today's large language models (LLMs) are trained to align with user preferences through methods such as reinforcement learning. Yet models are beginni

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Models

DGX agent

arXiv:2604.07749v1 Announce Type: new Abstract: Large language models (LLMs) can shift their answers under pressure in ways that reflect accommodation rather than reasoning. Prior work on sycophancy h

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning

DGX agent

arXiv:2604.07808v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models is constrained by substantial GPU memory requirements. Low-rank adaptation methods mitigate this cha

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

MCLR: Improving Conditional Modeling via Inter-Class Likelihood-Ratio Maximization and Unifying Classifier-Free Guidance with Alignment Objectives

DGX agent

arXiv:2603.22364v2 Announce Type: replace-cross Abstract: Diffusion models have achieved state-of-the-art performance in generative modeling, but their success often relies heavily on classifier-free

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning

DGX agent

arXiv:2604.07944v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong potential for autonomous vehicle motion planning by reformulating trajectory prediction a

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

DGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning

DGX agent

arXiv:2601.04268v2 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Self-Supervised Foundation Model for Calcium-imaging Population Dynamics

DGX agent

arXiv:2604.04958v2 Announce Type: replace-cross Abstract: Recent work suggests that large-scale, multi-animal modeling can significantly improve neural recording analysis. However, for functional calc

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution

DGX agent

arXiv:2604.07725v1 Announce Type: cross Abstract: We show that verifier-free evolution is bottlenecked by both diversity and efficiency: without external correction, repeated evolution accelerates col

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

DGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

VSAS-BENCH: Real-Time Evaluation of Visual Streaming Assistant Models

DGX agent

arXiv:2604.07634v1 Announce Type: new Abstract: Streaming vision-language models (VLMs) continuously generate responses given an instruction prompt and an online stream of input frames. This is a core

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

What do Language Models Learn and When? The Implicit Curriculum Hypothesis

DGX agent

arXiv:2604.08510v1 Announce Type: new Abstract: Large language models (LLMs) can perform remarkably complex tasks, yet the fine-grained details of how these capabilities emerge during pretraining rema

tutorialsarxiv-cs-cl
10 Apr 2026
Research

First-order friction models with bristle dynamics: lumped and distributed formulations

DGX agent

arXiv:2602.09429v3 Announce Type: replace-cross Abstract: Dynamic models, particularly rate-dependent models, have proven effective in capturing the key phenomenological features of frictional process

researcharxiv-cs-ro
13 Aug 2026
Model Releases

How China-Origin Vision-Language Models Move from Refusal to Reframing in State Alignment

DGX agent

arXiv:2608.11816v1 Announce Type: cross Abstract: State-aligned distortion has been documented in China-origin text-based large language models (LLMs), but whether, and in what form, it arises in mult

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Localizing Safety Alignment: MLP Layers and Mid-Network Blocks Encode Refusal Behavior in Large Language Models

DGX agent

arXiv:2608.11583v1 Announce Type: new Abstract: Safety alignment in large language models is often treated as a distributed property of the entire network, yet its practical brittleness suggests that

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets

DGX agent

arXiv:2608.11233v1 Announce Type: cross Abstract: A dense, pretrained language model can be retrofitted with recurrent depth and learn an iterative latent transition that persists after outcome-only a

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

More Accurate, Less Human: Gestalt Grouping in Vision Models

DGX agent

arXiv:2608.10195v1 Announce Type: new Abstract: Human vision organizes what it sees into wholes: same-colored points group into series, similar marks cohere into categories, and shapes complete into r

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Optimal Stopping of Self-Refining Foundation Models

DGX agent

arXiv:2608.10729v1 Announce Type: cross Abstract: Foundation models can improve their outputs through a self-refinement process driven by external feedback. In this process, the model is embedded in a

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes

DGX agent

arXiv:2608.10699v1 Announce Type: cross Abstract: Text-Attributed Graphs (TAGs), endowed with abundant textual content along with topological structures, have emerged as a versatile backbone for real-

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models

DGX agent

arXiv:2509.25143v2 Announce Type: replace-cross Abstract: Existing medical reasoning benchmarks for vision-language models primarily focus on analyzing a patient's condition based on an image from a s

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages

DGX agent

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, fo

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models

DGX agent

arXiv:2603.05868v2 Announce Type: replace Abstract: Despite remarkable progress in Vision-Language-Action models (VLAs) for robot manipulation, these large pre-trained models require fine-tuning to be

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safety Probes Across Model Families

DGX agent

arXiv:2608.08029v1 Announce Type: cross Abstract: Khatri et al. (2026) [DOI: 10.1109/DSN-W70714.2026.00027] show that lightweight MLP probes on final-layer activations of a single 8B model (LLaMA-3.1-

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Evo-Bench: Can Language Models Improve Agent Harness?

DGX agent

arXiv:2608.09096v1 Announce Type: new Abstract: Large Language Models (LLMs) have driven rapid progress in autonomous agents, yet standard evaluations remain confined to static task solving. An emergi

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch

DGX agent

arXiv:2608.09925v1 Announce Type: cross Abstract: Large language models are increasingly being deployed in governmental settings, yet few existing evaluation frameworks jointly reflect the values of p

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

On the use of foundation models in cognitive science

DGX agent

arXiv:2608.07812v1 Announce Type: new Abstract: A host of recent studies have evaluated the cognitive and developmental alignment of Foundation Models (FMs). These investigations include evaluations o

safetyarxiv-cs-cl
11 Aug 2026
Research

Second Order Drifting Models

DGX agent

arXiv:2608.07924v1 Announce Type: cross Abstract: Drifting models are a recent class of one-step generative models that evolve the model distribution during training using a predefined sample-based dr

researcharxiv-cs-ai
11 Aug 2026
Model Releases

SurgWMBench: A Vision-Based Benchmark for World-Modeling Surgical Instrument Motion Planning

DGX agent

arXiv:2608.08070v1 Announce Type: new Abstract: Reliable surgical planning requires models that move beyond recognizing the current surgical step or imitating expert demonstrations, and instead antici

model-releasesarxiv-cs-ro
11 Aug 2026
Tutorials

Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models

DGX agent

arXiv:2608.06994v1 Announce Type: cross Abstract: World Action Models (WAMs) aim to construct a unified architecture capable of understanding world state evolution and guiding to generative motion pla

tutorialsarxiv-cs-ai
10 Aug 2026
Safety

Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models

DGX agent

arXiv:2608.06779v1 Announce Type: cross Abstract: Large Language Models (LLMs) have accelerated drug discovery, particularly in the automated design of antimicrobial peptides (AMPs). However, current

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

DGX agent

arXiv:2608.07463v1 Announce Type: new Abstract: Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Model Confidence Under Answer-Preserving Attacks: An Informativeness-Manipulability Frontier

DGX agent

arXiv:2608.06571v1 Announce Type: cross Abstract: Deployed vision-language systems often gate their answers on confidence, making confidence robustness relevant to oversight. We study confidence reado

model-releasesarxiv-cs-cl
10 Aug 2026
Safety

SoRoMoX: Fast, Differentiable, and Parallelizable Soft Robot Models

DGX agent

arXiv:2608.06650v1 Announce Type: cross Abstract: Reduced-order models based on Cosserat-rod theory are now well established, and modeling theory is no longer the primary bottleneck in soft-robot cont

safetyarxiv-cs-ai
10 Aug 2026
Research

Diff-VF: Training-free High-quality Long Video Generation via Diffusion Model

DGX agent

arXiv:2608.05976v1 Announce Type: new Abstract: Recently, diffusion models have made great progress in video generation. However, most existing video diffusion models are trained with short videos, an

researcharxiv-cs-cv
7 Aug 2026
Model Releases

Zero-Shot Multi-Disease Labeling of Chest, Abdomen, and Pelvis CT Reports Using Open-Weight Large Language Models: The Effect of Labeling Conventions

DGX agent

arXiv:2506.03259v3 Announce Type: replace Abstract: Purpose: To compare five lightweight open-weight large language models (LLMs) with a rule-based algorithm (RBA) and fine-tuned RadBERT for zero-shot

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

DGX agent

arXiv:2608.04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instru

model-releasesarxiv-cs-ai
6 Aug 2026
Research

DASyR-LLM: Domain-Aware Symbolic Regression with LLMs for Kinetic Model Discovery

DGX agent

arXiv:2608.05120v1 Announce Type: new Abstract: Kinetic model discovery is a central challenge in chemical engineering, as accurate rate expressions are essential for understanding and controlling che

researcharxiv-cs-lg
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

SciCode-Verified: How Benchmark Defects Underestimated the Scientific-Coding Ability of Language Models

DGX agent

arXiv:2608.04975v1 Announce Type: cross Abstract: SciCode is the standard measure of the scientific-coding ability of language models: research-level problems that demand both frontier scientific theo

model-releasesarxiv-cs-ai
6 Aug 2026
← Previous
1…2122232425…1021
Next →