AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Tutorials

Leveraging generative models to assist Monte Carlo sampling

DGX agent

arXiv:2608.07648v1 Announce Type: cross Abstract: Sampling high-dimensional probability distributions is a central task in scientific computing, with applications ranging from Bayesian inference to st

tutorialsarxiv-cs-lg
11 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal Language Models

DGX agent

arXiv:2608.09435v1 Announce Type: new Abstract: Understanding dynamic sound sources requires jointly determining what produces a sound, where the source is located, and how it moves over time. Yet exi

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

LLMs Remember First, Forget Last: Dual-Process Interference in Large Language Models

DGX agent

arXiv:2603.00270v3 Announce Type: replace-cross Abstract: Large language models can process millions of tokens, yet how they handle conflicting information within context remains poorly understood. Fr

safetyarxiv-cs-ai
11 Aug 2026
Research

LookME: Lookup-Based Multimodal Embeddings for Layer Injection in Vision-Language Models

DGX agent

arXiv:2607.16305v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved strong progress in multimodal understanding. However, scaling dense or sparse Mixture-of-Experts (

researcharxiv-cs-ai
11 Aug 2026
Tutorials

MaxModShift: Model Privacy via Designed Shifts

DGX agent

arXiv:2608.09328v1 Announce Type: new Abstract: Model learning by an eavesdropper is treated as an estimation problem in a federated environment. The Fisher Information Matrix for the eavesdropper's e

tutorialsarxiv-cs-lg
11 Aug 2026
Model Releases

Multilingual Agent-Based World Modeling for Social Science

DGX agent

arXiv:2512.07195v2 Announce Type: replace-cross Abstract: Multi-agent role-playing has recently shown promise for studying social behavior with language agents, but existing simulations are mostly mon

model-releasesarxiv-cs-ai
11 Aug 2026
Research

One Model to Magnify Them All: Efficient Scale-Invariant Histopathology via Conditional Normalization and Continuous Magnification Training

DGX agent

arXiv:2608.09403v1 Announce Type: new Abstract: Whole slide images (WSIs) in digital histopathology are acquired at discrete magnification levels encoding complementary diagnostic information from glo

researcharxiv-cs-cv
11 Aug 2026
Model Releases

Performance of large language models in the optical diagnosis of colorectal polyps

DGX agent

arXiv:2608.07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ReMIND: Orchestrating Modular Large Language Models for Controllable Serendipity A REM-Inspired System Design for Emergent Creative Ideation

DGX agent

arXiv:2601.07121v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used not only for problem solving but also for creative ideation; however, generating ideas that

researcharxiv-cs-ai
11 Aug 2026
Research

Renormalising Generative Models for Active Inference: Foundations, Derivations, and Verification

DGX agent

arXiv:2608.09512v1 Announce Type: new Abstract: Active inference offers a unified framework for perception, learning, and action, but scaling discrete active-inference models to rich spatial and tempo

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Toward Mask Annotation-Free Surgical Instrument Segmentation from Endoscopic Images Using Text-Prompted Segment Anything Model 3 (SAM3)

DGX agent

arXiv:2608.08844v1 Announce Type: new Abstract: Surgical instrument segmentation is a fundamental task for computer-assisted interventions, yet most existing methods rely on pixel-level annotations or

model-releasesarxiv-cs-cv
11 Aug 2026
Research

CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation Models

DGX agent

arXiv:2608.06659v1 Announce Type: new Abstract: This paper shows that latent-space predictive pretraining can provide a scalable route to foundation models for spatial transcriptomics. Existing spatia

researcharxiv-cs-ai
10 Aug 2026
Research

Convergence of Diffusion Models Under the Manifold Hypothesis in High-Dimensions

DGX agent

arXiv:2409.18804v3 Announce Type: replace-cross Abstract: Denoising Diffusion Probabilistic Models (DDPM) are powerful state-of-the-art methods used to generate synthetic data from high-dimensional da

researcharxiv-cs-lg
10 Aug 2026
Research

Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model

DGX agent

arXiv:2608.07361v1 Announce Type: cross Abstract: Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself

researcharxiv-cs-cv
10 Aug 2026
Model Releases

Faster Query-Key Learning Sharpens Attention in Self-Attention Models

DGX agent

arXiv:2608.06776v1 Announce Type: new Abstract: A standard self-attention layer consists of two interacting circuits: the query-key circuit that governs attention allocation, and the output-value circ

model-releasesarxiv-cs-lg
10 Aug 2026
Research

How Molecular Generative Models Organize Molecular Identity

DGX agent

arXiv:2608.06956v1 Announce Type: new Abstract: Generative models for matter are often evaluated as samplers over output representations, and their latent spaces are commonly used as proxies for navig

researcharxiv-cs-lg
10 Aug 2026
Research

Separating Decision-Rule Misalignment from Readout-Coverage Limitations in Speech Language Models

DGX agent

arXiv:2608.06409v1 Announce Type: cross Abstract: Speech language models are increasingly evaluated on paralinguistic tasks by the accuracy of prompted answers, but answer accuracy combines failures a

researcharxiv-cs-ai
10 Aug 2026
Model Releases

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

DGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

DGX agent

arXiv:2607.18056v2 Announce Type: replace Abstract: Frontier large language models (LLMs) are increasingly integrated into scientific workflows, yet their growing biological capabilities may outpace c

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Safety

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

DGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

safetyarxiv-cs-ai
7 Aug 2026
Safety

LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models

DGX agent

arXiv:2608.05706v1 Announce Type: new Abstract: World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied int

safetyarxiv-cs-cv
7 Aug 2026
Research

On the Anisotropy of Score-Based Generative Models

DGX agent

arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce t

researcharxiv-cs-lg
7 Aug 2026
Model Releases

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

DGX agent

arXiv:2608.06361v1 Announce Type: new Abstract: Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure mo

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Toward Deployable Bangla Sign Language Recognition with Expert-Validated Data and a Lightweight Attention-Based Model

DGX agent

arXiv:2608.06252v1 Announce Type: cross Abstract: Deaf and hard-of-hearing people in Bangladesh communicate mainly through Bangla Sign Language (BdSL). Automatic BdSL recognition on personal devices c

model-releasesarxiv-cs-ai
7 Aug 2026
Applications

Uncertainty-Aware World Model for Aerial Image-Goal Navigation

DGX agent

arXiv:2608.05597v1 Announce Type: new Abstract: Aerial image-goal navigation requires an unmanned aerial vehicle (UAV) to reach a target location specified by a goal image. Existing world-model-based

applicationsarxiv-cs-cv
7 Aug 2026
Safety

XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?

DGX agent

arXiv:2608.05799v1 Announce Type: cross Abstract: Action-conditioned world models are promising learned simulators for robotic manipulation, yet evaluating them exclusively on training robots fails to

safetyarxiv-cs-cv
7 Aug 2026
Research

Convergence and Stability Analysis of Self-Consuming Generative Models with Heterogeneous Human Curation

DGX agent

arXiv:2511.09002v3 Announce Type: replace-cross Abstract: Self-consuming generative models have received significant attention over the last few years. In this paper, we study a self-consuming generat

researcharxiv-cs-lg
6 Aug 2026
Research

Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations

DGX agent

arXiv:2604.00209v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings, yet they frequently violate contextual privacy by disclosing private

researcharxiv-cs-cl
6 Aug 2026
Model Releases

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

DGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning

DGX agent

arXiv:2608.04646v1 Announce Type: new Abstract: Large language models (LLMs) have recently shown strong performance on Theory of Mind (ToM) tests, prompting debate about the nature and validity of the

researcharxiv-cs-cl
6 Aug 2026
Research

Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models

DGX agent

arXiv:2608.04765v1 Announce Type: cross Abstract: Vision-language-action (VLA) models provide a unified paradigm for connecting visual perception, language understanding, and robotic control. However,

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

DGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

model-releasesarxiv-cs-lg
6 Aug 2026
Research

Koopman-Based Nonlinear Identification and Model Predictive Control of a Turbofan Engine

DGX agent

arXiv:2604.01730v2 Announce Type: replace Abstract: This paper investigates Koopman operator-based approaches for multivariable control of a two-spool turbofan engine. A physics-based component-level

researcharxiv-cs-lg
6 Aug 2026
Model Releases

LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content

DGX agent

arXiv:2410.10783v4 Announce Type: replace Abstract: The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required

model-releasesarxiv-cs-cv
6 Aug 2026
Research

MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres

DGX agent

arXiv:2608.05054v1 Announce Type: cross Abstract: We investigate the transferability of Earth weather foundation models to planetary atmospheres by adapting the GraphCast graph neural weather forecast

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2608.04633v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are

model-releasesarxiv-cs-ro
6 Aug 2026
Safety

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

DGX agent

arXiv:2608.04436v1 Announce Type: new Abstract: Text-to-image (T2I) models can produce visually compelling images, yet they remain limited on open-world tasks that require complex semantic understandi

safetyarxiv-cs-cv
6 Aug 2026
Research

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models

DGX agent

arXiv:2608.04591v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked whether something is absent from a record, list, or retrieved context. Yet non-observation licenses a neg

researcharxiv-cs-ai
6 Aug 2026
Local Ai

Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging

DGX agent

arXiv:2608.03316v1 Announce Type: cross Abstract: On-policy distillation, in which a teacher corrects samples that the student itself generates, presupposes that the two models speak the same language

local-aiarxiv-cs-cv
5 Aug 2026
Model Releases

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

DGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

model-releasesarxiv-cs-cl
5 Aug 2026
Safety

BOW: Training Language Models to Reason Over Plausible Next Words

DGX agent

arXiv:2506.13502v3 Announce Type: replace Abstract: Next-word prediction (NWP) trains language models against a single observed continuation, even though many contexts admit multiple plausible next wo

safetyarxiv-cs-cl
5 Aug 2026
Model Releases

Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?

DGX agent

arXiv:2608.03983v1 Announce Type: cross Abstract: Optimizing compilers miss profitable transformations when their enabling semantics are absent from the analyzed program representation. We ask whether

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Can Training Logs Make Model Comparisons More Precise?

DGX agent

arXiv:2608.02705v1 Announce Type: cross Abstract: Comparing stochastically trained models requires estimating both a performance difference and its uncertainty from repeated runs. We study whether tra

researcharxiv-cs-ai
5 Aug 2026
Applications

Cross-Model KV Cache Transfer in LLM Families: A Closed-Form Linear Mapping for Prefill Reuse

DGX agent

arXiv:2608.03893v1 Announce Type: new Abstract: Production deployments often swap between different-sized models in a family for cost-quality cascading, mid-conversation switching, and routing, and ea

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

Cura 1T: Specialized Model for Agentic Healthcare

DGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

model-releasesarxiv-cs-ai
5 Aug 2026
Research

DUD: Decoupled Update Dynamics for Reliable Uncertainty Quantification in Large Language Models

DGX agent

arXiv:2608.03411v1 Announce Type: new Abstract: Accurate Uncertainty Quantification (UQ) is critical for reliable deployment of Large Language Models (LLMs), yet traditional probability-based metrics

researcharxiv-cs-cl
5 Aug 2026
Research

Failure-Informed Image Self-Augmentation for Multimodal Large Language Model Self-Improvement

DGX agent

arXiv:2608.03733v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved remarkable performance across vision-language tasks, but their progress depends heavily on large-

researcharxiv-cs-ai
5 Aug 2026
← Previous
1…122123124125126…1030
Next →