AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
Human
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
11 May 2026

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings

Model ReleasesDGX agent

arXiv:2605.07492v1 Announce Type: new Abstract: The past year has seen over 20 open-source document parsing models, yet thefield still benchmarks almost exclusively on OmniDocBench, a 1,355-pagemanual

How Log-Barrier Helps Exploration in Policy Optimization

SafetyDGX agent

arXiv:2603.15001v2 Announce Type: replace-cross Abstract: Recently, it has been shown that the Stochastic Gradient Bandit (SGB) algorithm converges to a globally optimal policy with a constant learnin

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

How to Train Your Latent Diffusion Language Model Jointly With the Latent Space

TutorialsDGX agent

arXiv:2605.07933v1 Announce Type: new Abstract: Latent diffusion models offer an attractive alternative to discrete diffusion for non-autoregressive text generation by operating on continuous text rep

How to utilize failure demo data?: Effective data selection for imitation learning using distribution differences in attention mechanism

SafetyDGX agent

arXiv:2605.07560v1 Announce Type: new Abstract: Imitation learning for robotic tasks has relied primarily on policies trained only on successful demonstrations, although failures are unavoidable durin

How Value Induction Reshapes LLM Behaviour

SafetyDGX agent

arXiv:2605.07925v1 Announce Type: new Abstract: Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and em

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem

ResearchDGX agent

arXiv:2605.06882v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved great improvements in recent years. Nevertheless, it still remains unclear how good LLMs are for reasoning ta

Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models

TutorialsDGX agent

arXiv:2508.05803v2 Announce Type: replace Abstract: Human memory is fleeting. As words are processed, the exact wordforms that make up incoming sentences are rapidly lost. Cognitive scientists have lo

HumanNet: Scaling Human-centric Video Learning to One Million Hours

Model ReleasesDGX agent

arXiv:2605.06747v1 Announce Type: new Abstract: Progress in embodied intelligence increasingly depends on scalable data infrastructure. While vision and language have scaled with internet corpora, lea

Hybrid TF--IDF Logistic Regression and MLP Neural Baseline for Indonesian Three-Class Sentiment Analysis on Social Media Text

ApplicationsDGX agent

arXiv:2605.07793v1 Announce Type: new Abstract: This paper presents a compact three-class sentiment analysis study for Indonesian social media text. The task is formulated with positive, negative, and

HYPER: A Foundation Model for Inductive Link Prediction with Knowledge Hypergraphs

TutorialsDGX agent

arXiv:2506.12362v3 Announce Type: replace-cross Abstract: Inductive link prediction with knowledge hypergraphs is the task of predicting missing hyperedges involving completely novel entities (i.e., n

HyperEyes: Dual-Grained Efficiency-Aware Reinforcement Learning for Parallel Multimodal Search Agents

Model ReleasesDGX agent

arXiv:2605.07177v1 Announce Type: cross Abstract: Existing multimodal search agents process target entities sequentially, issuing one tool call per entity and accumulating redundant interaction rounds

ICDAR 2026 Competition on Writer Identification and Pen Classification from Hand-Drawn Circles

ResearchDGX agent

arXiv:2605.07816v1 Announce Type: new Abstract: This paper presents CircleID, a large-scale ICDAR 2026 competition on writer identification and pen classification from scanned hand-drawn circles. The

Identifiability Challenges in Sparse Linear Ordinary Differential Equations

ResearchDGX agent

arXiv:2506.09816v3 Announce Type: replace Abstract: Dynamical systems modeling is a core pillar of scientific inquiry across natural and life sciences. Increasingly, dynamical system models are learne

ImplantMamba: Long-range Sequential Modeling Mamba For Dental Implant Position Prediction

ResearchDGX agent

arXiv:2605.07082v1 Announce Type: new Abstract: In the design of surgical guides for implant placement, determining the precise implant position is a critical step. However, the implant region itself

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

SafetyDGX agent

arXiv:2605.07316v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning

Implicit Multi-Camera System Calibration Using Gaussian Processes

TutorialsDGX agent

arXiv:2605.07491v1 Announce Type: new Abstract: This paper proposes a novel framework for implicit multi-camera system calibration utilizing Gaussian Process (GP) regression. Conventional explicit cal

Implicit Preference Alignment for Human Image Animation

Model ReleasesDGX agent

arXiv:2605.07545v1 Announce Type: cross Abstract: Human image animation has witnessed significant advancements, yet generating high-fidelity hand motions remains a persistent challenge due to their hi

Improved Model-based Reinforcement Learning with Smooth Kernels

ResearchDGX agent

arXiv:2605.07218v1 Announce Type: new Abstract: For continuous state-action space scenarios, classical reinforcement learning (RL) theory predominantly focuses on low-rank Markov decision processes (M

In-Context Credit Assignment via the Core

Model ReleasesDGX agent

arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic

Inductive Power Grid Cascading Failure Analysis with GRU-Gated Graph Attention

TutorialsDGX agent

arXiv:2605.07010v1 Announce Type: new Abstract: Identifying vulnerable transmission lines in power grids before a cascading failure occurs is challenging: existing methods can learn inter-line failure

Inference of Qualitative Models from Steady-State Data via Weighted MaxSMT

ResearchDGX agent

arXiv:2605.07433v1 Announce Type: cross Abstract: Qualitative models provide crucial instruments for modelling complex biological systems. While advances in automated reasoning and symbolic encodings

Inference-Time Attribute Distribution Alignment for Unconditional Diffusion

SafetyDGX agent

arXiv:2605.07456v1 Announce Type: new Abstract: Inference-time controllable generation is essential for real-world applications of unconditional diffusion models. However, most existing techniques foc

Inference Time Causal Probing in LLMs

Model ReleasesDGX agent

arXiv:2605.07631v1 Announce Type: new Abstract: Causal probing methods aim to test and control how internal representations influence the behavior of generative models. In causal probing, an intervent

InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization

SafetyDGX agent

arXiv:2605.07099v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is fundamental for precise localization and navigation in GPS-denied environments, aiming to match ground or UAV imag

Information-theoretic Limits of Learning and Estimation

ResearchDGX agent

arXiv:2605.06710v1 Announce Type: cross Abstract: Information theory plays a central role in establishing fundamental limits on what any learning or estimation algorithm can -- and cannot -- achieve,

INO-SGD: Addressing Utility Imbalance under Individualized Differential Privacy

ResearchDGX agent

arXiv:2605.07930v1 Announce Type: cross Abstract: Differential privacy (DP) is widely employed in machine learning to protect confidential or sensitive training data from being revealed. As data owner

InsHuman: Towards Natural and Identity-Preserving Human Insertion

ResearchDGX agent

arXiv:2605.07402v1 Announce Type: new Abstract: Human insertion aims to naturally place specific individuals into a target background. Although existing image editing models may have such ability, the

Instruction Tuning Changes How Upstream State Conditions Late Readout: A Cross-Patching Diagnostic

Local AiDGX agent

arXiv:2605.07284v1 Announce Type: new Abstract: Recent interpretability work has identified model-internal handles on post-trained behavior, including refusal directions, assistant/persona axes, and s

Integrating Causal DAGs in Deep RL: Activating Minimal Markovian States with Multi-Order Exposure

ApplicationsDGX agent

arXiv:2605.07057v1 Announce Type: new Abstract: Online reinforcement learning (RL) relies on the Markov property for guaranteed performance, but real-world applications often lack well-defined states

Intelligent Truck Matching in Full Truckload Shipments using Ping2Hex approach

ApplicationsDGX agent

arXiv:2605.07733v1 Announce Type: cross Abstract: Accurate truck-to-shipment matching using GPS data is foundational for full truckload supply chain visibility, enabling real-time tracking and accurat

Intent-Driven Semantic ID Generation for Grounded Conversational News Recommendation

Model ReleasesDGX agent

arXiv:2605.07613v1 Announce Type: new Abstract: Conversational news recommendation requires grounding each suggestion in a rapidly evolving article corpus while addressing implicit user intents that l

IntentGrasp: A Comprehensive Benchmark for Intent Understanding

Model ReleasesDGX agent

arXiv:2605.06832v1 Announce Type: cross Abstract: Accurately understanding the intent behind speech, conversation, and writing is crucial to the development of helpful Large Language Model (LLM) assis

Intention assimilation control for accurate tracking with variable impedance in teleoperation

SafetyDGX agent

arXiv:2605.07037v1 Announce Type: new Abstract: Robot systems for teleoperation commonly use a spring-like force pulling the follower robot towards the leader's position to track their movements. With

Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems

AgentsDGX agent

arXiv:2605.07768v1 Announce Type: cross Abstract: We investigate interactive trajectory planning subject to uncertainty in the decisions of surrounding agents. To control the ego-agent, we aim to firs

InterCoG: Towards Spatially Precise Image Editing with Interleaved Chain-of-Grounding Reasoning

SafetyDGX agent

arXiv:2603.01586v3 Announce Type: replace Abstract: Emerging unified editing models have demonstrated strong capabilities in general object editing tasks. However, it remains a significant challenge t

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

Model ReleasesDGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

Interpreting Reinforcement Learning Agents with Susceptibilities

Model ReleasesDGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features

ResearchDGX agent

arXiv:2603.03096v2 Announce Type: replace-cross Abstract: How do speech models trained through self-supervised learning structure their representations? Previous studies have looked at how information

Into the Rabbit Hull: From Task-Relevant Concepts in DINO to Minkowski Geometry

ResearchDGX agent

arXiv:2510.08638v3 Announce Type: replace-cross Abstract: DINOv2 is routinely deployed to recognize objects, scenes, and actions; yet the nature of what it perceives remains unknown. As a working base

Inverse Reinforcement Learning with Just Classification and a Few Regressions

SafetyDGX agent

arXiv:2509.21172v2 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) aims to infer rewards from observed behavior, but rewards are not identified from the policy alone: many reward

InvThink: Premortem Reasoning for Safer Language Models

SafetyDGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization

ResearchDGX agent

arXiv:2512.23032v2 Announce Type: replace-cross Abstract: Recent work, using the Biasing Features metric, labels a CoT as unfaithful if it omits a prompt-injected hint that affected the prediction. We

Is She Even Relevant? When BERT Ignores Explicit Gender Cues

Local AiDGX agent

arXiv:2605.07622v1 Announce Type: new Abstract: Gender bias in large language models has primarily been investigated for English, while languages with grammatical or morphological gender remain compar

Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models

SafetyDGX agent

arXiv:2605.07514v1 Announce Type: cross Abstract: World Action Models (WAMs) enable decision-making through imagined rollouts by predicting future observations and actions. However, the reliability of

Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies

Model ReleasesDGX agent

arXiv:2510.22944v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have become indispensable for automated code generation, yet the quality and security of their outputs remain a c

It Just Takes Two: Scaling Amortized Inference to Large Sets

ResearchDGX agent

arXiv:2605.07972v1 Announce Type: cross Abstract: Neural posterior estimation has emerged as a powerful tool for amortized inference, with growing adoption across scientific and applied domains. In ma

Kernel Selection is Model Selection: A Unified Complexity-Penalized Approach for MMD Two-Sample Tests

ResearchDGX agent

arXiv:2605.06883v1 Announce Type: cross Abstract: The Maximum Mean Discrepancy (MMD) is a cornerstone statistic for nonparametric two-sample testing, but its test power is dictated entirely by the cho

KL for a KL: On-Policy Distillation with Control Variate Baseline

SafetyDGX agent

arXiv:2605.07865v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) has emerged as a dominant post-training paradigm for large language models, especially for reasoning domains. However, OP

Knowing but Not Correcting: Routine Task Requests Suppress Factual Correction in LLMs

Model ReleasesDGX agent

arXiv:2605.05957v2 Announce Type: replace Abstract: LLMs reliably correct false claims when presented in isolation, yet when the same claims are embedded in task-oriented requests, they often comply r

Knowledge Transfer Scaling Laws for 3D Medical Imaging

ResearchDGX agent

arXiv:2605.06859v1 Announce Type: cross Abstract: Vision foundation models are increasingly moving beyond 2D to volumetric domains such as 3D medical imaging, where unified pretraining across differen

Koopman Autoencoders with Continuous-Time Latent Dynamics for Fluid Dynamics Forecasting

ResearchDGX agent

arXiv:2602.02832v3 Announce Type: replace Abstract: Forecasting physical systems over long horizons from irregularly sampled observations demands models that are stable, computationally efficient, and

Kurtosis-Guided Denoising Score Matching for Tabular Anomaly Detection

Local AiDGX agent

arXiv:2605.06955v1 Announce Type: cross Abstract: Denoising score matching (DSM) provides a way to learn data distributions by training a neural network to recover the score function, defined as the g

LAMES: A Large-Scale and Artisanal Mining Environmental Segmentation Dataset

ApplicationsDGX agent

arXiv:2605.07740v1 Announce Type: new Abstract: Mining operations are of utmost importance to the economy of some nations. However, such operations result in land-use change, very high energy consumpt

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

Model ReleasesDGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

Large Video Planner Enables Generalizable Robot Control

ApplicationsDGX agent

arXiv:2512.15840v2 Announce Type: replace-cross Abstract: General-purpose robots require decision-making models that generalize across diverse tasks and environments. Recent works build robot foundati

Latent Order Bandits

ResearchDGX agent

arXiv:2605.07304v1 Announce Type: new Abstract: Bandit algorithms solve diverse sequential decision-making problems, but are often too sample-inefficient for from-scratch personalization. To substanti

Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

ResearchDGX agent

arXiv:2602.01166v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models benefit from chain-of-thought (CoT) reasoning, but existing approaches incur high inference overhead and rely on

Latent-Space Causal Discovery from Indirect Neuroimaging Observations

ResearchDGX agent

arXiv:2602.09034v2 Announce Type: replace-cross Abstract: Neuroimaging does not observe causal variables directly: hemodynamics and volume conduction distort signals so that statistical dependence nee

LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification

ResearchDGX agent

arXiv:2605.07315v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language models (LLMs) on difficult tasks, but it also makes inference expensive because every intermedi

← Previous
1…745746747748749…1005
Next →