AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
Human
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
12 May 2026

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

Model ReleasesDGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

The autoPET3 Challenge: Automated Lesion Segmentation in Whole-Body PET/CT nicode{x2013} Multitracer Multicenter Generalization

Model ReleasesDGX agent

arXiv:2605.05775v2 Announce Type: replace-cross Abstract: We report the design and results of the third autoPET challenge (MICCAI 2024), which benchmarked automated lesion segmentation in whole-body P

The Benefits of Temporal Correlations: SGD Learns k-Juntas from Random Walks Efficiently

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.10237v1 Announce Type: new Abstract: We study how temporal correlations in the data can make certain sparse learning problems efficiently learnable by gradient-based methods. Our focus is o

The Bystander Effect in Multi-Agent Reasoning: Quantifying Cognitive Loafing in Collaborative Interactions

SafetyDGX agent

arXiv:2605.10698v1 Announce Type: cross Abstract: Multi-agent systems (MAS) assume that collaborating inherently improves Large Language Model (LLM) reasoning. We challenge this by demonstrating that

The Cancellation Hypothesis in Critic-Free RL: From Outcome Rewards to Token Credits

ResearchDGX agent

arXiv:2605.08666v1 Announce Type: new Abstract: A commonly accepted explanation of critic-free RL for LLMs, based on sequence-level rewards, is that it reinforces successful rollouts with a positive a

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space

ResearchDGX agent

arXiv:2605.09883v1 Announce Type: cross Abstract: As current Multimodal Large Language Models rapidly saturate canonical visual reasoning benchmarks, a key question emerges: do these strong scores gen

The Differences Between Direct Alignment Algorithms are a Blur

Model ReleasesDGX agent

arXiv:2502.01237v3 Announce Type: replace Abstract: Direct Alignment Algorithms (DAAs) simplify LLM alignment by directly optimizing policies, bypassing reward modeling and RL. While DAAs differ in th

The Direct Integration Theorem: A Rigorous Framework for Consistent Discrete Solutions of the Inverse Radon Problem

ResearchDGX agent

arXiv:2605.09020v1 Announce Type: new Abstract: This paper presents a novel Direct Integration Theorem (DIT), derived as a non-trivial corollary of the classical Central Slice Theorem (CST). The DIT p

The DSA's Blind Spot: Algorithmic Audit of Advertising and Minor Profiling on TikTok

ResearchDGX agent

arXiv:2603.05653v2 Announce Type: replace-cross Abstract: Adolescents spend an increasing amount of their time in digital environments where their still-developing cognitive capacities leave them unab

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

Model ReleasesDGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

The Ensemble Schr{odinger Bridge filter for Nonlinear Data Assimilation

ResearchDGX agent

arXiv:2512.18928v3 Announce Type: replace Abstract: This work introduces a novel nonlinear optimal filtering method, termed the Ensemble Schr{odinger Bridge nonlinear filter. The proposed filter combi

The Extrapolation Cliff in On-Policy Distillation of Near-Deterministic Structured Outputs

Model ReleasesDGX agent

arXiv:2605.08737v1 Announce Type: cross Abstract: On-policy distillation (OPD) is widely used for LLM post-training. When pushed with a reward-extrapolation coefficient lambda > 1, the student can lif

The finite expression method for turbulent dynamics with high-order moment recovery

TutorialsDGX agent

arXiv:2605.10687v1 Announce Type: new Abstract: Turbulent dynamical systems are characterized by nonlinear interactions and stochastic effects that generate coupled statistical quantities, such as non

The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context Reasoning

AgentsDGX agent

arXiv:2605.10828v1 Announce Type: new Abstract: As large language models are increasingly deployed in retrieval-augmented generation and agentic systems that accumulate extensive context, understandin

The Generalized Turing Test: A Foundation for Comparing Intelligence

ResearchDGX agent

arXiv:2605.10851v1 Announce Type: new Abstract: We introduce the Generalized Turing Test (GTT), a formal framework for comparing the capabilities of arbitrary agents via indistinguishability. For agen

The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning

ResearchDGX agent

arXiv:2601.18832v3 Announce Type: replace-cross Abstract: Scaling test-time compute enhances long chain-of-thought (CoT) reasoning, yet existing approaches face a fundamental trade-off between computa

The Geometric Structure of Models Learning Sparse Data

SafetyDGX agent

arXiv:2605.08464v1 Announce Type: new Abstract: The manifold hypothesis (MH) is often used to explain how machine learning can overcome the curse of dimensionality. However, the MH is only applicable

The Geometric Wall: Manifold Structure Predicts Layerwise Sparse Autoencoder Scaling Laws

Model ReleasesDGX agent

arXiv:2605.09887v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) operationalise the linear representation hypothesis: they reconstruct model activations as sparse linear combinations of in

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations

Model ReleasesDGX agent

arXiv:2605.09195v1 Announce Type: new Abstract: Large language models confidently produce outdated answers, and no existing method can detect them. We show this is not an engineering failure but a str

The Global Empirical NTK: Self-Referential Bias and Dimensionality of Gradient Descent Learning

Model ReleasesDGX agent

arXiv:2605.08746v1 Announce Type: new Abstract: In training a neural network with gradient descent (GD), each iteration induces a linear operator that governs first-order updates to a model's internal

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark

Model ReleasesDGX agent

arXiv:2605.09900v1 Announce Type: new Abstract: A vision-language model can look at a knot diagram and report what it sees, yet fail to act on that structure. KnotBench pairs an 858,318-image corpus f

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans

SafetyDGX agent

arXiv:2605.08837v1 Announce Type: cross Abstract: Abstract concepts - justice, theory, availability - have no single perceivable referent; in the human brain, their meaning emerges from a web of exper

The Impact of Editorial Intervention on Detecting Native Language Traces

ResearchDGX agent

arXiv:2605.10216v1 Announce Type: new Abstract: Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writings. With the advent of huma

The Invisible Handshake: Persistent Overpricing by Adaptive Market Agents

ResearchDGX agent

arXiv:2510.15995v3 Announce Type: replace-cross Abstract: We study overpricing in a repeated game between two representative agents: a market maker, who controls market liquidity, and a market taker,

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

Model ReleasesDGX agent

arXiv:2605.10799v1 Announce Type: cross Abstract: Corruption studies, the primary tool for evaluating chain-of-thought (CoT) faithfulness, identify which chain positions are 'computationally important

The Metacognitive Probe: Five Behavioural Calibration Diagnostics for LLMs

Model ReleasesDGX agent

arXiv:2605.09844v1 Announce Type: new Abstract: The Metacognitive Probe is an exploratory five-task, 15-slot diagnostic that decomposes an LLM's confidence behaviour into five behaviourally-distinct d

The Mixing method: low-rank coordinate descent for semidefinite programming with diagonal constraints

ResearchDGX agent

arXiv:1706.00476v4 Announce Type: replace-cross Abstract: In this paper, we propose a low-rank coordinate descent approach to structured semidefinite programming with diagonal constraints. The approac

The Observable Wasserstein Distance

ResearchDGX agent

arXiv:2605.09916v1 Announce Type: cross Abstract: We introduce the observable Wasserstein distance, a framework for deriving lower bounds on the Wasserstein distance between probability measures on Po

The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime

ResearchDGX agent

arXiv:2605.10601v1 Announce Type: new Abstract: AI deployment in sensitive domains such as health care, credit, employment, and criminal justice is often treated as unsafe to authorize until model int

The Pokemon Theorem and other Fairness Impossibility Results

SafetyDGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

The Polynomial Counting Capabilities of Message Passing Neural Networks

ResearchDGX agent

arXiv:2605.10393v1 Announce Type: new Abstract: The counting power of Message Passing Neural Networks (MPNN) has been the subject of many recent papers, showing that they can express logic that involv

The Power of Second Order Methods for Sequence Preconditioning

ResearchDGX agent

arXiv:2605.08390v1 Announce Type: new Abstract: Sequence prediction methods for dynamical systems with long memory, i.e. marginally stable systems, typically achieve regret that grows polynomially wit

The Procrustean Bed of Time Series: The Optimization Bias in Point-wise Loss Functions

SafetyDGX agent

arXiv:2512.18610v3 Announce Type: replace Abstract: Intuitively, a more deterministic time series should be easier to forecast. However, point-wise loss functions (e.g., MSE and MAE), serving as diffe

The Propagation Field: A Geometric Substrate Theory of Deep Learning

ResearchDGX agent

arXiv:2605.08529v1 Announce Type: new Abstract: Modern deep learning treats neural networks primarily as endpoint functions from inputs to outputs. Inspired by the shift from force to geometry in phys

The Realignment Problem: When Right becomes Wrong in LLMs

Model ReleasesDGX agent

arXiv:2511.02623v2 Announce Type: replace Abstract: Post-training alignment of large language models (LLMs) relies on large-scale human annotations guided by policy specifications that change over tim

The Reciprocity Gradient

AgentsDGX agent

arXiv:2605.08323v1 Announce Type: cross Abstract: Communication is fundamental to sustaining reciprocity and cooperation in strategic interactions. We identify and formulate the influence attribution

The Safety-Aware Denoiser for Text Diffusion Models

SafetyDGX agent

arXiv:2605.08116v1 Announce Type: cross Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling their safety remains underexplored.

The Sample Complexity of Uniform Approximation for Multi-Dimensional CDFs and Fixed-Price Mechanisms

ResearchDGX agent

arXiv:2602.10868v2 Announce Type: replace Abstract: We study the sample complexity of learning a uniform approximation of an n-dimensional cumulative distribution function (CDF) within an error epsilo

The Silent Vote: Improving Zero-Shot LLM Reliability by Aggregating Semantic Neighborhoods

Model ReleasesDGX agent

arXiv:2605.09739v1 Announce Type: cross Abstract: Large Language Models are increasingly used as zero-shot classifiers in complex reasoning tasks. However, standard constrained decoding suffers from a

The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory

Model ReleasesDGX agent

arXiv:2605.09330v1 Announce Type: cross Abstract: Agentic memory enables LLMs to persist information beyond a single context window and reuse it in later decisions, but it also introduces a new vulner

The Truth Lies Somewhere in the Middle (of the Generated Tokens)

Local AiDGX agent

arXiv:2605.09969v1 Announce Type: cross Abstract: How should hidden states generated autoregressively be collapsed into a representation that reflects a language model's internal state? Despite tokens

The two clocks and the innovation window: When and how generative models learn rules

TutorialsDGX agent

arXiv:2605.10019v1 Announce Type: cross Abstract: Generative models trained on finite data face a fundamental tension: their score-matching or next-token objective converges to the empirical training

The Value of Mechanistic Priors in Sequential Decision Making

SafetyDGX agent

arXiv:2605.10018v1 Announce Type: new Abstract: Hybrid mechanistic models, physical priors with learned residuals, promise to reduce the data required for good decisions, but have no computable criter

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

SafetyDGX agent

arXiv:2605.09352v1 Announce Type: new Abstract: Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence lea

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2601.02954v3 Announce Type: replace-cross Abstract: Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding

The Wristband Gaussian Loss: Deterministic, Composable Latents via a Sphere-Interval Decomposition

Model ReleasesDGX agent

arXiv:2605.08749v1 Announce Type: new Abstract: We present the Wristband Gaussian Loss, a deterministic batch loss for Gaussianizing point embeddings without sampling, KL terms, or iterative transport

Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection

SafetyDGX agent

arXiv:2605.10130v1 Announce Type: new Abstract: Existing open-vocabulary detectors focus on RGB images and fail to generalize to thermal imagery, where low texture and emissivity variations challenge

Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving

AgentsDGX agent

arXiv:2605.10117v1 Announce Type: cross Abstract: Autonomous driving scenes range from empty highways to dense intersections with dozens of interacting road users, yet current 3D detection models appl

Thinking with Novel Views: A Systematic Analysis of Generative-Augmented Spatial Intelligence

ResearchDGX agent

arXiv:2605.10588v1 Announce Type: new Abstract: Current Large Multimodal Models (LMMs) struggle with spatial reasoning tasks requiring viewpoint-dependent understanding, largely because they are confi

Threat Modelling using Domain-Adapted Language Models: Empirical Evaluation and Insights

ResearchDGX agent

arXiv:2605.10808v1 Announce Type: cross Abstract: Large Language Models(LLMs) are increasingly explored for cybersecurity applications such as vulnerability detection. In the domain of threat modellin

ThreatCore: A Benchmark for Explicit and Implicit Threat Detection

Model ReleasesDGX agent

arXiv:2605.10563v1 Announce Type: cross Abstract: Threat detection in Natural Language Processing lacks consistent definitions and standardized benchmarks, and is often conflated with broader phenomen

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent

AgentsDGX agent

arXiv:2605.09443v1 Announce Type: cross Abstract: The advancement of Multimodal Large Language Models (MLLMs) has expanded Role-Playing Agents (RPAs) into visually grounded environments. However, huma

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning

Model ReleasesDGX agent

arXiv:2605.09544v1 Announce Type: new Abstract: Tool-integrated reasoning has emerged as a promising paradigm for enhancing large language models with external computation, retrieval, and execution ca

TIDES: Implicit Time-Awareness in Selective State Space Models

Model ReleasesDGX agent

arXiv:2605.09742v1 Announce Type: cross Abstract: Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step Tilde{Delta} a learne

TIE: Time Interval Encoding for Video Generation over Events

SafetyDGX agent

arXiv:2605.10543v1 Announce Type: new Abstract: Director-style prompting, robotic action prediction, and interactive video agents demand temporal grounding over concurrent events -- a regime in which

Tight Generalization Bounds for Noiseless Inverse Optimization

Model ReleasesDGX agent

arXiv:2605.08866v1 Announce Type: cross Abstract: Inverse optimization (IO) seeks to infer the parameters of a decision-maker's objective from observed context--action data. We study noiseless IO, whe

Tighter Information-Theoretic Generalization Bounds via a Novel Class of Change of Measure Inequalities

ResearchDGX agent

arXiv:2602.07999v3 Announce Type: replace-cross Abstract: Change of measure inequalities translate divergences between probability measures into explicit bounds on event probabilities, and play an imp

TiledAttention: a CUDA Tile SDPA Kernel for PyTorch

Model ReleasesDGX agent

arXiv:2603.01960v2 Announce Type: replace-cross Abstract: TiledAttention is a scaled dot-product attention (SDPA) forward operator for SDPA research on NVIDIA GPUs. Implemented in cuTile Python (TileI

TileQ: Efficient Low-Rank Quantization of Mixture-of-Experts with 2D Tiling

ResearchDGX agent

arXiv:2605.09281v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models achieve remarkable performance by sparsely activating specialized experts, yet their massive parameters in experts pose

Time-Warping Recurrent Neural Networks for Transfer Learning

ResearchDGX agent

arXiv:2604.02474v2 Announce Type: replace Abstract: Dynamical systems describe how a physical system evolves over time. Physical processes can evolve faster or slower in different environmental condit

← Previous
1…733734735736737…1005
Next →