AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “concepts”

GridTimelineEvolution
2,151 results
Model Releases

CodeClinic: Evaluating Automation of Coding Skills for Clinical Reasoning Agents

DGX agent

arXiv:2605.09675v1 Announce Type: new Abstract: Clinical reasoning agents based on large language models (LLMs) aim to automate tasks such as intensive care unit (ICU) monitoring and patient state tra

model-releasesarxiv-cs-ai
12 May 2026
Applications
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Count Anything at Any Granularity

DGX agent

arXiv:2605.10887v1 Announce Type: new Abstract: Open-world object counting remains brittle: despite rapid advances in vision-language models (VLMs), reliably counting the objects a user intends is far

applicationsarxiv-cs-cv
12 May 2026
Research

Deep Dreams Are Made of This: Visualizing Monosemantic Features in Diffusion Models

DGX agent

arXiv:2605.08218v1 Announce Type: cross Abstract: This paper proposes latent visualization by optimization (LVO), a mechanistic interpretability technique that extends feature visualization by optimiz

researcharxiv-cs-cv
12 May 2026
Model Releases

Diffusion Models are Evolutionary Algorithms

DGX agent

arXiv:2410.02543v3 Announce Type: replace-cross Abstract: In a convergence of machine learning and biology, we reveal that diffusion models are evolutionary algorithms. By considering evolution as a d

model-releasesarxiv-cs-lg
12 May 2026
Research

Explaining Graph Neural Networks for Node Similarity on Graphs

DGX agent

arXiv:2407.07639v2 Announce Type: replace-cross Abstract: Similarity search is a fundamental task for exploiting information in various applications dealing with graph data, such as citation networks

researcharxiv-cs-ai
12 May 2026
Model Releases

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

DGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

fmxcoders: Factorized Masked Crosscoders for Cross-Layer Feature Discovery

DGX agent

arXiv:2605.09438v1 Announce Type: new Abstract: Many features in pretrained Transformers span multiple layers: they emerge through stages of inference, persist in the residual stream, or are built joi

local-aiarxiv-cs-lg
12 May 2026
Research

Functional Subspace, where language models can use vector algebra to solve problems

DGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

researcharxiv-cs-ai
12 May 2026
Research

Generative Giants, Retrieval Weaklings: Why do Multimodal Large Language Models Fail at Multimodal Retrieval?

DGX agent

arXiv:2512.19115v2 Announce Type: replace Abstract: Despite the remarkable success of multimodal large language models (MLLMs) in generative tasks, we observe that they exhibit a counterintuitive defi

researcharxiv-cs-cv
12 May 2026
Model Releases

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs

DGX agent

arXiv:2508.20325v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly integral to various domains, their potential to generate harmful responses has prompted si

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification

DGX agent

arXiv:2605.08295v1 Announce Type: cross Abstract: While random demonstration labels barely hurt in-context learning (Min et al., 2022), we show that homogeneous labels--even semantically valid ones--c

model-releasesarxiv-cs-ai
12 May 2026
Research

Interpretable Coreference Resolution Evaluation Using Explicit Semantics

DGX agent

arXiv:2605.10627v1 Announce Type: cross Abstract: Coreference resolution is typically evaluated using aggregate statistical metrics such as CoNLL-F1, which measure structural overlap between predicted

researcharxiv-cs-ai
12 May 2026
Research

Language-Conditioned Visual Grounding with CLIP Multilingual

DGX agent

arXiv:2605.09060v1 Announce Type: new Abstract: Multilingual vision-language models exhibit systematic performance gaps across languages, but the mechanism remains ambiguous: cross-language divergence

researcharxiv-cs-cl
12 May 2026
Research

Lecture Notes on Statistical Physics and Neural Networks

DGX agent

arXiv:2605.06394v1 Announce Type: cross Abstract: These lecture notes introduce some topics of classical statistical physics, particularly those that are relevant for neural networks and deep learning

researcharxiv-cs-lg
12 May 2026
Research

LPT: Less-overfitting Prompt Tuning for Vision-Language Model

DGX agent

arXiv:2410.10247v3 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have demonstrated exceptional generalization capabilities for downstream tasks. Due to its efficiency, prompt le

researcharxiv-cs-ai
12 May 2026
Applications

Medical Model Synthesis Architectures: A Case Study

DGX agent

arXiv:2605.09716v1 Announce Type: new Abstract: Medicine is rife with high-stakes uncertainty. Doctors routinely make clinical judgments and decisions that juggle many fundamental unknowns, like predi

applicationsarxiv-cs-ai
12 May 2026
Safety

Mental Health AI Safety Claims Must Preserve Temporal Evidence

DGX agent

arXiv:2605.08827v1 Announce Type: new Abstract: The safety of mental health AI is often judged at the wrong temporal scale. Current evaluations typically score isolated responses, endpoint outcomes, o

safetyarxiv-cs-ai
12 May 2026
Tutorials

MUR: Momentum Uncertainty guided Reasoning

DGX agent

arXiv:2507.14958v2 Announce Type: replace Abstract: Current models have achieved impressive performance on reasoning-intensive tasks, yet optimizing their reasoning efficiency remains an open challeng

tutorialsarxiv-cs-cl
12 May 2026
Safety

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

DGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

safetyarxiv-cs-ai
12 May 2026
Tutorials

Reasoning Trajectories for Socratic Debugging of Student Code: From Misconceptions to Contradictions and Updated Beliefs

DGX agent

arXiv:2511.00371v2 Announce Type: replace Abstract: In Socratic debugging, instructors guide students towards identifying and fixing a bug on their own, instead of providing the bug fix directly. Most

tutorialsarxiv-cs-cl
12 May 2026
Model Releases

Sens-VisualNews: A Benchmark Dataset for Sensational Image Detection

DGX agent

arXiv:2605.10394v1 Announce Type: new Abstract: The detection of sensational content in media items can be a critical filtering mechanism for identifying check-worthy content and flagging potential di

model-releasesarxiv-cs-cv
12 May 2026
Research

Threat Modelling using Domain-Adapted Language Models: Empirical Evaluation and Insights

DGX agent

arXiv:2605.10808v1 Announce Type: cross Abstract: Large Language Models(LLMs) are increasingly explored for cybersecurity applications such as vulnerability detection. In the domain of threat modellin

researcharxiv-cs-ai
12 May 2026
Research

Time-Warping Recurrent Neural Networks for Transfer Learning

DGX agent

arXiv:2604.02474v2 Announce Type: replace Abstract: Dynamical systems describe how a physical system evolves over time. Physical processes can evolve faster or slower in different environmental condit

researcharxiv-cs-lg
12 May 2026
Model Releases

TINS: Test-time ID-prototype-separated Negative Semantics Learning for OOD Detection

DGX agent

arXiv:2605.10756v1 Announce Type: new Abstract: Vision-language models enable OOD detection by comparing image alignment with ID labels and negative semantics. Existing negative-label-based methods ma

model-releasesarxiv-cs-cv
12 May 2026
Research

Toward an Engineering of Science: Rebalancing Generation and Verification in the Age of AI

DGX agent

arXiv:2605.10425v1 Announce Type: cross Abstract: AI systems can now cheaply generate plausible scientific artifacts such as papers, reviews, and surveys. This creates a risk of epistemic pollution in

researcharxiv-cs-ai
12 May 2026
Model Releases

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding

DGX agent

arXiv:2605.10782v1 Announce Type: new Abstract: Urban mobility is naturally expressed both as trajectories in space and as natural-language descriptions of travel intent, constraints, and preferences.

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media

DGX agent

arXiv:2605.05831v2 Announce Type: replace Abstract: The communication of scientific knowledge has become increasingly multimodal, spanning text, visuals, and speech through materials such as research

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

DGX agent

arXiv:2605.09360v1 Announce Type: cross Abstract: Execution-based evaluation of LLM-generated code implicitly treats successful execution as a proxy for correctness. In scientific simulation, this pro

model-releasesarxiv-cs-ai
12 May 2026
Safety

A Generalized Singular Value Theory for Neural Networks

DGX agent

arXiv:2605.06938v1 Announce Type: cross Abstract: Building on the abstract Generalized Singular Value Decomposition (GSVD) theory of Brown et al. [2025], we prove that most modern neural architectures

safetyarxiv-cs-ai
11 May 2026
Tutorials

Active teacher selection for reward learning

DGX agent

arXiv:2310.15288v3 Announce Type: replace Abstract: Reward learning techniques enable machine learning systems to learn objectives from human feedback. A core limitation of these systems is their assu

tutorialsarxiv-cs-ai
11 May 2026
Research

An abstract effective convergence theorem for stochastic processes, with applications to stochastic approximation

DGX agent

arXiv:2504.12922v3 Announce Type: replace-cross Abstract: We provide a general theorem on the asymptotic behavior of stochastic processes that conform to a relaxed supermartingale condition. The disti

researcharxiv-cs-lg
11 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Model Releases

Curvature Beyond Positivity: Greedy Guarantees for Arbitrary Submodular Functions

DGX agent

arXiv:2605.07902v1 Announce Type: new Abstract: Submodular functions -- functions exhibiting diminishing returns -- are central to machine learning. When the objective is monotone and non-negative, th

model-releasesarxiv-cs-lg
11 May 2026
Research

Demystifying Lipschitz verification: positive matrices, negative results

DGX agent

arXiv:2603.28113v2 Announce Type: replace Abstract: The global Lipschitz constant of a neural network is related to robustness and generalization, yet unlike in many classical models, it is not plainl

researcharxiv-cs-lg
11 May 2026
Local Ai

Emergent Manifold Separability during Reasoning in Large Language Models

DGX agent

arXiv:2602.20338v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting significantly improves reasoning in Large Language Models, yet the temporal dynamics of the underlying representati

local-aiarxiv-cs-lg
11 May 2026
Research

Equivalence of Coarse and Fine-Grained Models for Learning with Distribution Shift

DGX agent

arXiv:2605.07005v1 Announce Type: cross Abstract: Recent work on provably efficient algorithms for learning with distribution shift has focused on two models: PQ learning (Goldwasser et al. (2020)) an

researcharxiv-cs-lg
11 May 2026
Research

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems

DGX agent

arXiv:2506.04565v2 Announce Type: replace-cross Abstract: Compound AI Systems (CAIS) are an emerging paradigm that integrates large language models (LLMs) with external components, including retriever

researcharxiv-cs-cl
11 May 2026
Research

Identifiability Challenges in Sparse Linear Ordinary Differential Equations

DGX agent

arXiv:2506.09816v3 Announce Type: replace Abstract: Dynamical systems modeling is a core pillar of scientific inquiry across natural and life sciences. Increasingly, dynamical system models are learne

researcharxiv-cs-lg
11 May 2026
Model Releases

In-Context Credit Assignment via the Core

DGX agent

arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic

model-releasesarxiv-cs-ai
11 May 2026
Research

Information-theoretic Limits of Learning and Estimation

DGX agent

arXiv:2605.06710v1 Announce Type: cross Abstract: Information theory plays a central role in establishing fundamental limits on what any learning or estimation algorithm can -- and cannot -- achieve,

researcharxiv-cs-lg
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Research

Lossy Common Information in a Learnable Gray-Wyner Network

DGX agent

arXiv:2601.21424v3 Announce Type: replace-cross Abstract: Many computer vision tasks share substantial overlapping information, yet conventional codecs tend to ignore this, leading to redundant and in

researcharxiv-cs-cv
11 May 2026
Model Releases

Mage: Multi-Axis Evaluation of LLM-Generated Executable Game Scenes Beyond Compile-Pass Rate

DGX agent

arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl

model-releasesarxiv-cs-ai
11 May 2026
Applications

Multimodal synthesis of MRI and tabular data with diffusion in a joint latent space via cross-attention

DGX agent

arXiv:2605.06699v1 Announce Type: cross Abstract: We propose a multimodal latent diffusion model that jointly synthesizes volumetric magnetic resonance imaging (MRI) and tabular clinical data within a

applicationsarxiv-cs-ai
11 May 2026
Safety

Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching

DGX agent

arXiv:2605.06474v2 Announce Type: replace-cross Abstract: We present a novel theoretical framework, Q-MMR, for off-policy evaluation in finite-horizon MDPs. Q-MMR learns a set of scalar weights, one f

safetyarxiv-cs-ai
11 May 2026
Research

Regret-Oracle Complexity Tradeoffs in Agnostic Online Learning

DGX agent

arXiv:2605.07155v1 Announce Type: new Abstract: Agnostic online learning is classically solved via a reduction to the realizable setting, utilizing Littlestone's Standard Optimal Algorithm (SOA) as a

researcharxiv-cs-lg
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Safety

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

DGX agent

arXiv:2605.08075v1 Announce Type: new Abstract: Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across

safetyarxiv-cs-lg
11 May 2026
← Previous
1…3637383940…45
Next →