AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
12 May 2026

Deep Dreams Are Made of This: Visualizing Monosemantic Features in Diffusion Models

ResearchDGX agent

arXiv:2605.08218v1 Announce Type: cross Abstract: This paper proposes latent visualization by optimization (LVO), a mechanistic interpretability technique that extends feature visualization by optimiz

Detecting Multi-Agent Collusion Through Multi-Agent Interpretability

Model ReleasesDGX agent

arXiv:2604.01151v2 Announce Type: replace Abstract: As LLM agents are increasingly deployed in multi-agent systems, they introduce risks of covert coordination that may evade standard forms of human o

Deterministic Differentiable Structured Pruning for Large Language Models

ResearchDGX agent

arXiv:2603.08065v2 Announce Type: replace-cross Abstract: Structured pruning reduces LLM inference cost by removing low-importance architectural components. This can be viewed as learning a multiplica

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding

Model ReleasesDGX agent

arXiv:2605.08888v1 Announce Type: new Abstract: Evaluating whether Multimodal Large Language Models can produce trustworthy, verifiable reasoning over long, visually rich documents requires evaluation

Efficient Dataset Distillation for Pre-Trained Self-Supervised Models via Statistical Flow Matching

HardwareDGX agent

arXiv:2602.05391v2 Announce Type: replace Abstract: Dataset distillation seeks to synthesize a highly compact dataset that achieves performance comparable to the original dataset on downstream tasks.

Elite Polarization in European Parliamentary Speeches: a Novel Measurement Approach Using Large Language Models

ApplicationsDGX agent

arXiv:2507.06658v2 Announce Type: replace-cross Abstract: Theories of democratic stability, populism, and party-system crisis often point to a form of polarization that comparative research rarely mea

ERIS: Enhancing Privacy and Scalability in Federated Learning via Federated Shard Aggregation

Model ReleasesDGX agent

arXiv:2602.08617v2 Announce Type: replace Abstract: Scaling Federated Learning (FL) to billion-parameter models forces a challenging trade-off between privacy, scalability, and model utility. Existing

Evading Visual Aphasia: Contrastive Adaptive Semantic Token Pruning for Vision-Language Models

ResearchDGX agent

arXiv:2605.09429v1 Announce Type: cross Abstract: Are low-attention visual tokens truly redundant in vision-language reasoning? Existing pruning methods often assume so, ranking visual tokens by shall

EverydayMMQA: A Multilingual and Multimodal Framework for Culturally Grounded Spoken Visual QA

Model ReleasesDGX agent

arXiv:2510.06371v2 Announce Type: replace-cross Abstract: Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries re

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

Model ReleasesDGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling

Model ReleasesDGX agent

arXiv:2512.02010v5 Announce Type: replace Abstract: As large language models have grown larger, interest has grown in low-precision numerical formats such as NVFP4 as a way to improve speed and reduce

From Passive Reuse to Active Reasoning: Grounding Large Language Models for Neuro-Symbolic Experience Replay

SafetyDGX agent

arXiv:2605.09419v1 Announce Type: new Abstract: While experience replay is essential for data efficiency in reinforcement learning (RL), standard methods treat the replay buffer as a passive memory sy

Functional Stable Model Semantics and Answer Set Programming Modulo Theories

ResearchDGX agent

arXiv:2605.09524v1 Announce Type: new Abstract: Recently there has been an increasing interest in incorporating ``intensional'' functions in answer set programming. Intensional functions are those who

Geometrically Approximated Modeling for Emitter-Centric Ray-Triangle Filtering in Arbitrarily Dynamic LiDAR Simulation

HardwareDGX agent

arXiv:2605.10457v1 Announce Type: cross Abstract: Real-time Light Detection And Ranging (LiDAR) simulation must find, per emitted ray, the closest intersecting triangle even in dynamic scenes containi

GLiNER-Relex: A Unified Framework for Joint Named Entity Recognition and Relation Extraction

Model ReleasesDGX agent

arXiv:2605.10108v1 Announce Type: new Abstract: Joint named entity recognition (NER) and relation extraction (RE) is a fundamental task in natural language processing for constructing knowledge graphs

Identified-Set Geometry of Distributional Model Extraction under Top-K Censored API Access

ResearchDGX agent

arXiv:2605.10407v1 Announce Type: new Abstract: Modern LLM APIs often reveal only top-K logit scores and censor the remaining vocabulary. We study the per-position distribution-recovery limits of this

LLM Translation of Compiler Intermediate Representation

Model ReleasesDGX agent

arXiv:2605.08247v1 Announce Type: cross Abstract: GCC and LLVM underpin much of modern software infrastructure, relying on distinct Intermediate Representations (IRs) to drive optimizations and code g

MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

AgentsDGX agent

arXiv:2605.10763v1 Announce Type: new Abstract: LLMs are increasingly deployed as autonomous agents with access to tools, databases, and external services, yet practitioners (across different sectors)

Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare

Model ReleasesDGX agent

arXiv:2605.08445v1 Announce Type: new Abstract: AI models are increasingly deployed in live clinical environments where they must perform reliably across complex, high-stakes workflows that standard t

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks

Model ReleasesDGX agent

arXiv:2507.23511v3 Announce Type: replace-cross Abstract: While large audio-language models have advanced open-ended audio understanding, they still fall short of nuanced human-level comprehension. Th

MedMeta: A Benchmark for LLMs in Synthesizing Meta-Analysis Conclusion from Medical Studies

Model ReleasesDGX agent

arXiv:2605.09661v1 Announce Type: cross Abstract: Large language models (LLMs) have saturated standard medical benchmarks that test factual recall, yet their ability to perform higher-order reasoning,

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier f…

Model ReleasesDGX agent

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier for teams building AI agents on LangChain to run those worklo

Neuromorphic Monocular Depth Estimation with Uncertainty Modeling

ResearchDGX agent

arXiv:2605.10675v1 Announce Type: new Abstract: Event cameras offer distinct advantages over conventional frame-based sensors, including microsecond-level temporal resolution, high dynamic range, and

Privacy Auditing Synthetic Data Release through Local Likelihood Attacks

Model ReleasesDGX agent

arXiv:2508.21146v2 Announce Type: replace Abstract: Auditing the privacy leakage of synthetic data is an important but unresolved problem. Existing privacy auditing frameworks for synthetic data rely

QT-Net: Rethinking Evaluation of AI Models in Atomic Chemical Space

ResearchDGX agent

arXiv:2605.10458v1 Announce Type: new Abstract: Atomic properties such as partial charges or multipoles encode chemically meaningful information that can inform downstream molecular property predictio

Quantum Circuit Simulation of Compartmental Drug Dynamics: Leveraging Variational Algorithms for Nonlinear Mixed-Effects Population Pharmacokinetics

Model ReleasesDGX agent

arXiv:2605.09691v1 Announce Type: new Abstract: Population pharmacokinetic/pharmacodynamic (PK/PD) modeling traditionally relies on classical ordinary differential equations to simulate drug dynamics.

Retrieval Mechanisms Surpass Long-Context Scaling in Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.08217v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have borrowed the long context paradigm from natural language processing under the premise that feeding more histo

RewardHarness: Self-Evolving Agentic Post-Training

Model ReleasesDGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

Sequential Feature Selection for Efficient Landslide Segmentation from Multi-Spectral Data

Model ReleasesDGX agent

arXiv:2605.09746v1 Announce Type: cross Abstract: Landslide detection from satellite imagery has advanced through deep learning, yet most models rely on large, highly correlated spectral-topographic i

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

Model ReleasesDGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning

Model ReleasesDGX agent

arXiv:2605.10247v1 Announce Type: new Abstract: Using Large Language Models (LLMs) to process graph-structured data is an active research area, yet current state-of-the-art approaches typically rely o

The Geometric Structure of Models Learning Sparse Data

SafetyDGX agent

arXiv:2605.08464v1 Announce Type: new Abstract: The manifold hypothesis (MH) is often used to explain how machine learning can overcome the curse of dimensionality. However, the MH is only applicable

The Geometric Wall: Manifold Structure Predicts Layerwise Sparse Autoencoder Scaling Laws

Model ReleasesDGX agent

arXiv:2605.09887v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) operationalise the linear representation hypothesis: they reconstruct model activations as sparse linear combinations of in

VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation

Model ReleasesDGX agent

arXiv:2605.08553v1 Announce Type: cross Abstract: Large language models can generate useful code from natural language, but their outputs come without correctness guarantees. Verifiable code generatio

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

Model ReleasesDGX agent

arXiv:2605.08146v1 Announce Type: cross Abstract: Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domai

WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.10912v1 Announce Type: new Abstract: Large language and vision-language models increasingly power agents that act on a user's behalf through command-line interface (CLI) harnesses. However,

11 May 2026

A Unified and Controllable Framework for Layered Image Generation with Visual Effects

Model ReleasesDGX agent

arXiv:2601.15507v2 Announce Type: replace Abstract: Recent image generation models produce impressive composites, but often fail to preserve the identity of user-provided content when editing specific

Assumed Density Filtering and Smoothing with Neural Network Surrogate Models

TutorialsDGX agent

arXiv:2511.09016v2 Announce Type: replace-cross Abstract: The Kalman filter and Rauch-Tung-Striebel (RTS) smoother are optimal for state estimation in linear dynamic systems. With nonlinear systems, t

AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models

ApplicationsDGX agent

arXiv:2605.07478v1 Announce Type: new Abstract: Speech-driven facial animation requires accurate correspondence between acoustic signals and facial motion, especially for articulation-related mouth mo

Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs

SafetyDGX agent

arXiv:2605.07806v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in settings where reliable self-assessment is critical. Assessing model reliability has evolved fro

Discovering Multiagent Learning Algorithms with Large Language Models

SafetyDGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment

SafetyDGX agent

arXiv:2605.06885v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have recently demonstrated capabilities that complement standard autoregressive (AR) models, particularly in non-sequ

EmambaIR: Efficient Visual State Space Model for Event-guided Image Reconstruction

TutorialsDGX agent

arXiv:2605.08073v1 Announce Type: cross Abstract: Recent event-based image reconstruction methods predominantly rely on Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs) to process c

Emergent social transmission of model-based representations without inference

SafetyDGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models

SafetyDGX agent

arXiv:2605.07244v1 Announce Type: cross Abstract: We introduce Mutual Reinforcement Learning, a framework for concurrent RL post-training in which heterogeneous LLM policies exchange typed experience

HAIC: Humanoid Agile Object Interaction Control via Dynamics-Aware World Model

SafetyDGX agent

arXiv:2602.11758v2 Announce Type: replace Abstract: Humanoid robots show promise for complex whole-body tasks in unstructured environments. Although Human-Object Interaction (HOI) has advanced, most m

Hierarchical Perfusion Graphs for Tumor Heterogeneity Modeling in Glioma Molecular Subtyping

ResearchDGX agent

arXiv:2605.07156v1 Announce Type: new Abstract: Precise molecular subtyping of gliomas, including isocitrate dehydrogenase (IDH) mutation and 1p/19q codeletion, directly guides surgical and therapeuti

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings

Model ReleasesDGX agent

arXiv:2605.07492v1 Announce Type: new Abstract: The past year has seen over 20 open-source document parsing models, yet thefield still benchmarks almost exclusively on OmniDocBench, a 1,355-pagemanual

Kernel Selection is Model Selection: A Unified Complexity-Penalized Approach for MMD Two-Sample Tests

ResearchDGX agent

arXiv:2605.06883v1 Announce Type: cross Abstract: The Maximum Mean Discrepancy (MMD) is a cornerstone statistic for nonparametric two-sample testing, but its test power is dictated entirely by the cho

MedVIGIL: Evaluating Trustworthy Medical VLMs Under Broken Visual Evidence

Model ReleasesDGX agent

arXiv:2605.07919v1 Announce Type: new Abstract: Medical vision--language models (VLMs) are usually evaluated on intact image--question pairs, but trustworthy clinical use requires a stronger property:

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

SafetyDGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

Retrieval from Within: An Intrinsic Capability of Attention-Based Models

ResearchDGX agent

arXiv:2605.05806v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) typically treats retrieval and generation as separate systems. We ask whether an attention-based encoder-decode

Sat3R: Satellite DSM Reconstruction via RPC-Aware Depth Fine-tuning

Model ReleasesDGX agent

arXiv:2605.07264v1 Announce Type: new Abstract: Accurate Digital Surface Model (DSM) reconstruction from satellite imagery is critical for applications such as disaster response, urban planning, and l

ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs

Model ReleasesDGX agent

arXiv:2605.07338v1 Announce Type: new Abstract: The decline of global shellfish biodiversity poses a severe threat to coastal ecosystems. Although artificial intelligence (AI) technologies show potent

The Coupling Tax: How Shared Token Budgets Undermine Visible Chain-of-Thought Under Fixed Output Limits

Model ReleasesDGX agent

arXiv:2605.07686v1 Announce Type: new Abstract: Chain-of-thought reasoning is often treated as a monotone way to improve language-model accuracy by letting a model think longer. We identify a counterv

The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking

Model ReleasesDGX agent

arXiv:2605.06707v1 Announce Type: cross Abstract: This paper presents an eight-week observational comparison of 68 single-file HTML generations collected across 17 public experiments in the 'HTML AI B

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

SafetyDGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

Training-Free Multimodal Large Language Model Orchestration

SafetyDGX agent

arXiv:2508.10016v3 Announce Type: replace Abstract: Building interactive omni-modal assistants often relies on end-to-end multimodal alignment to fuse heterogeneous modalities, which incurs substantia

9 May 2026

Apple Removes 256GB M3 Ultra Mac Studio Model From Online Store

Local AiDGX agent

Apple has removed its 256GB M3 Ultra Mac Studio from sale, limiting the machine to 96GB of unified memory , following the removal of the 512GB configuration in March . The removal is likely due to a g

7 May 2026

EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation

ResearchDGX agent

arXiv:2605.04062v1 Announce Type: new Abstract: Recent years have witnessed an increasing interest in deploying LLMs on resource-constrained devices, among which quantization has emerged as a promisin

← Previous
1…262263264265266…1034
Next →