AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
Human
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
27 May 2026

Learning to Balance Motor Thermal Safety and Quadrupedal Locomotion Performance with Residual Policy

SafetyDGX agent

arXiv:2605.27046v1 Announce Type: new Abstract: Motor thermal management is often overlooked in the context of electrically-actuated robots, particularly legged robots, but motor overheating is a key

Learning to Diagnose and Correct Errors: Towards Moral Sensitivity Acquisition in Large Language Models

TutorialsDGX agent

arXiv:2601.03079v4 Announce Type: replace Abstract: Moral sensitivity is the most fundamental capability underlying human moral competence. Although many approaches aim to align large language models

Learning to Orchestrate Agents under Uncertainty

SafetyDGX agent

arXiv:2605.27073v1 Announce Type: new Abstract: Adaptive orchestration of heterogeneous agents requires making sequential delegation decisions under uncertain and evolving agent behaviour, e.g., coord

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning to Predict Future-Aligned Research Proposals with Language Models

Model ReleasesDGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

Learning to Reason Efficiently with Discounted Reinforcement Learning

SafetyDGX agent

arXiv:2510.23486v2 Announce Type: replace Abstract: Large reasoning models (LRMs) often consume excessive tokens, inflating computational cost and latency. More broadly, in goal reaching sequential de

Learning When to Think While Listening in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2605.27190v1 Announce Type: cross Abstract: Recent advances in Large Audio-Language Models (LALMs) have made real-time, streaming spoken interaction increasingly practical. In this setting, reas

LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems

ResearchDGX agent

arXiv:2512.01556v3 Announce Type: replace Abstract: Foundation models often generate unreliable answers, while heuristic uncertainty estimators fail to fully distinguish correct from incorrect outputs

Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Spatial-Relation Data

ResearchDGX agent

arXiv:2601.12809v2 Announce Type: replace-cross Abstract: Spatial understanding remains a key challenge in vision-language models. Yet it is still unclear whether such understanding is truly acquired,

LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation

ApplicationsDGX agent

arXiv:2605.26956v1 Announce Type: new Abstract: Entity linking is a key component of many downstream NLP systems, yet existing approaches are often tied to the specific target knowledge bases and doma

Less is More: Early Stopping Rollout for On-Policy Distillation

SafetyDGX agent

arXiv:2605.27028v1 Announce Type: cross Abstract: On-policy distillation has recently emerged as a promising alternative to standard sequence-level imitation, training a student by scoring its own rol

Lessons from Penetration Tests on Large-Scale Agent Systems

AgentsDGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of

Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking

TutorialsDGX agent

arXiv:2605.26933v1 Announce Type: new Abstract: Unsupervised visual object tracking is a challenging task that requires following arbitrary targets in videos without training on ground-truth annotatio

Leveraging Visual Signals for Robust Token-Level Uncertainty in Vision-Language Generation

ApplicationsDGX agent

arXiv:2605.27136v1 Announce Type: new Abstract: Uncertainty quantification (UQ) remains a critical challenge in Large Vision Language Models (LVLMs) for reliable predictions and real-world deployment.

Lifting Data-Tracing Machine Unlearning to Knowledge-Tracing for Foundation Models

TutorialsDGX agent

arXiv:2506.11253v2 Announce Type: replace Abstract: Machine unlearning removes certain training data points and their influence from AI models (e.g., when a data owner revokes their consent to allow m

LiM-YOLO: Less is More with Pyramid Level Shift for Ship Detection in Optical Remote Sensing

ResearchDGX agent

arXiv:2512.09700v3 Announce Type: replace Abstract: General-purpose object detectors face fundamental structural limitations when applied to ship detection in satellite imagery, where the ship scale d

Linear and Neural Dueling Bandits with Delayed Feedback

SafetyDGX agent

arXiv:2605.26554v1 Announce Type: cross Abstract: Contextual dueling bandits form a cornerstone of preference-based decision-making, with critical applications in recommender systems and large languag

LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning

AgentsDGX agent

arXiv:2412.20505v2 Announce Type: replace Abstract: Participatory Urban Planning (PUP) is increasingly supported by LLM-based agents, yet existing methods largely rely on static preference elicitation

LitSeg: Narrative-Aware Document Segmentation for Literary RAG

ResearchDGX agent

arXiv:2605.27156v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by incorporating external knowledge, particularly for long-tail domains suc

LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?

Model ReleasesDGX agent

arXiv:2605.26781v1 Announce Type: new Abstract: Advanced Large Multimodal Models (LMMs) have demonstrated impressive performance in K-12 reasoning tasks, exhibiting great promise as intelligent tutors

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

Model ReleasesDGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring

Model ReleasesDGX agent

arXiv:2605.27088v1 Announce Type: new Abstract: Aligning LLMs for math tutoring typically requires RL-based training with multi-GPU infrastructure. We investigate whether training-free prompt optimiza

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

Model ReleasesDGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

Localizing Memorized Regions in Diffusion Models via Coordinate-Wise Curvature Differences

Local AiDGX agent

arXiv:2605.26756v1 Announce Type: new Abstract: Diffusion models can unintentionally memorize training samples, raising concerns about privacy and copyright. While recent methods can detect memorizati

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

ResearchDGX agent

arXiv:2605.27365v1 Announce Type: cross Abstract: Vision-language models (VLMs) commonly formulate visual grounding and detection as a coordinate-token generation problem, serializing each 2D box into

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV

Model ReleasesDGX agent

arXiv:2605.26244v1 Announce Type: new Abstract: Audio-visual generation is rapidly advancing from short clips to minute-long content, while existing evaluation protocols remain largely confined to sho

LongCat-Video-Avatar 1.5 Technical Report

Model ReleasesDGX agent

arXiv:2605.26486v1 Announce Type: new Abstract: Despite advances in audio-driven video generation, achieving commercial-grade stability remains challenging. We present LongCat-Video-Avatar 1.5, an upg

Look Further: Socially-Compliant Navigation System in Residential Buildings

SafetyDGX agent

arXiv:2605.26710v1 Announce Type: new Abstract: The distance at which a mobile robot reacts to a person strongly impacts various qualities of the human-robot interaction. In this paper, we focus on th

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)

ResearchDGX agent

arXiv:2605.27268v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) are often criticized for producing repetitive and homogeneous text, despite possessing vast latent vocabularies. W

LUCoS: Latent Unsupervised Context Selection for Tabular Foundation Models

ResearchDGX agent

arXiv:2605.27254v1 Announce Type: cross Abstract: Selecting which instances to label is a key challenge in low-label tabular learning. For recent Tabular Foundation Models such as TabPFN, context sele

LURE: Live-Usage Replay Evaluations for Reducing Evaluation Awareness

Model ReleasesDGX agent

arXiv:2605.26438v1 Announce Type: cross Abstract: Large language models can recognize when they are being evaluated (evaluation awareness) and behave differently because of that, which undermines the

LuxRemix: Lighting Decomposition and Remixing for Indoor Scenes

ApplicationsDGX agent

arXiv:2601.15283v2 Announce Type: replace Abstract: We present a novel approach for interactive light editing in indoor scenes from a single multi-view scene capture. Our method leverages a generative

Maat: The Agentic Legal Research Assistant for Competition Protection

Model ReleasesDGX agent

arXiv:2605.27331v1 Announce Type: new Abstract: Competition law experts conducting legal research must review extensive volumes of cases, decisions, and judicial reports to identify precedents and ass

MAIGO: Mitigating Lost-in-Conversation with History-Cleaned On-Policy Self-Distillation

SafetyDGX agent

arXiv:2605.27186v1 Announce Type: new Abstract: Large language models often solve tasks from a fully specified prompt but degrade when the same requirements unfold over multiple turns, known as the lo

Managing Uncertainty in LLM-Generated Procedural Knowledge for Virtual Laboratory Planning

ResearchDGX agent

arXiv:2605.26333v1 Announce Type: new Abstract: Educational virtual laboratories can make experimental training more scala-ble, adaptive, and accessible, especially when students have limited access t

Manipulating Tangible Virtual Object Dynamics to Promote Learning of Precision Force Generation

ResearchDGX agent

arXiv:2605.26782v1 Announce Type: new Abstract: Robotic haptic devices combined with virtual reality offer novel opportunities to train fine force generation, an essential yet overlooked component of

Many Logics, One Methodology: A Plea for Logical Pluralism in Formalised Reasoning (preprint)

ResearchDGX agent

arXiv:2605.27246v1 Announce Type: cross Abstract: This position statement looks back on two decades of work on shallow embeddings of non-classical logics in classical higher-order logic (HOL), a line

MATCHA: Matching Text via Contrastive Semantic Alignment

SafetyDGX agent

arXiv:2605.27345v1 Announce Type: new Abstract: Reliable evaluation is essential for understanding large language model (LLM) performance, yet today's go-to metrics, namely token-overlap scores (e.g.,

MatFormBench: A Benchmarking Evaluation Framework for Target-Driven Materials Formulation

TutorialsDGX agent

arXiv:2605.26741v1 Announce Type: cross Abstract: Inverse design of materials has significantly advanced target-driven formulation optimization, yet existing materials machine learning benchmarks rema

MATT-CTR: Unleashing a Model-Agnostic Test-Time Paradigm for CTR Prediction with Confidence-Guided Inference Paths

Model ReleasesDGX agent

arXiv:2510.08932v2 Announce Type: replace Abstract: Recently, a growing body of research has focused on either optimizing CTR model architectures to better model feature interactions or refining train

Max-Window Scale Estimation for Near-Lossless HiF8 W8A8 Quantization-Aware Training

ResearchDGX agent

arXiv:2605.26189v1 Announce Type: cross Abstract: Quantization-aware training (QAT) with low-bit floating-point formats enables efficient LLM deployment, yet introduces subtle failure modes invisible

Measuring Prediction Uncertainty in Neural Cellular Automata

SafetyDGX agent

arXiv:2605.26726v1 Announce Type: cross Abstract: Neural cellular automata (NCA) provide a lightweight alternative to encoder-decoder segmentation networks. However, it can be difficult to decide when

MechRL: Reinforcement Learning Agents Perform Circuit Discovery for Mechanistic Interpretability

SafetyDGX agent

arXiv:2605.26343v1 Announce Type: new Abstract: Mechanistic interpretability has identified small sets of attention heads that implement specific behaviours in transformer language models, but recover

Med-CoReasoner: Reducing Language Disparities in Medical Reasoning via Language-Informed Co-Reasoning

Model ReleasesDGX agent

arXiv:2601.08267v3 Announce Type: replace Abstract: While reasoning-enhanced large language models perform strongly on English medical tasks, a persistent multilingual gap remains, with substantially

MedCollab: IBIS-Guided Multi-Agent Collaboration with Hierarchical Disease Relation Chains for Clinical Diagnosis

AgentsDGX agent

arXiv:2603.01131v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in clinical diagnosis but remain limited by unreliable report generation, weak evidence ground

MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Models for Clinical Reasoning

ResearchDGX agent

arXiv:2605.26567v1 Announce Type: new Abstract: Clinical practice guidelines (CPGs) encode evidence-based decision logic that clinicians apply by evaluating patient variables, conditional criteria, an

MedVol-R1: Reward-Driven Evidence Grounding for Volumetric Reasoning Segmentation

Model ReleasesDGX agent

arXiv:2605.26621v1 Announce Type: cross Abstract: Volumetric Reasoning Segmentation (VRS) aims to segment a target region in a 3D medical scan from a free-form clinical query, where the referent is of

Membership Inference Risks in Quantized Models: A Theoretical and Empirical Study

ApplicationsDGX agent

arXiv:2502.06567v2 Announce Type: replace-cross Abstract: Quantizing machine learning models has demonstrated its effectiveness in lowering memory and inference costs while maintaining performance lev

MemFail: Stress-Testing Failure Modes of LLM Memory Systems

Model ReleasesDGX agent

arXiv:2605.26667v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external memory systems to remain consistent across long-horizon interactions, but little empiric

MemMorph: Tool Hijacking in LLM Agents via Memory Poisoning

SafetyDGX agent

arXiv:2605.26154v1 Announce Type: cross Abstract: LLM-driven agents are capable of selecting external tools to complete users' tasks. However, attackers could compromise such process, steering agents

Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study

Model ReleasesDGX agent

arXiv:2605.26394v1 Announce Type: new Abstract: Multi-turn Text-to-SQL is central to enterprise analytics yet remains predominantly evaluated in single-turn settings. We introduce EnterpriseMem-Bench,

Memory-Distilled Selection for Noise-Robust Anomaly Detection

ResearchDGX agent

arXiv:2605.26676v1 Announce Type: new Abstract: Anomaly detection (AD) under data contamination is critical for deploying unsupervised defect detection in industrial environments, where curating perfe

MerLean-Prover: A Recursive Looping Harness for End-to-End Lean 4 Theorem Proving

Model ReleasesDGX agent

arXiv:2605.26959v1 Announce Type: cross Abstract: MerLean-Prover is an end-to-end Lean4 theorem prover that replaces sorry declarations with kernel-checkable proofs. It is built from three agent types

Message-Passing State-Space Models: Improving Graph Learning with Modern Sequence Modeling

ResearchDGX agent

arXiv:2505.18728v2 Announce Type: replace-cross Abstract: The recent success of State-Space Models (SSMs) in sequence modeling has motivated their adaptation to graph learning, giving rise to Graph St

MetaGraph: A Large-Scale Meta-Analysis of GenAI in Financial NLP (2022-2025)

ApplicationsDGX agent

arXiv:2509.09544v3 Announce Type: replace Abstract: Financial NLP has evolved rapidly since late 2022, outpacing narrative surveys. We introduce MetaGraph, a methodology for extracting typed knowledge

MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning

ResearchDGX agent

arXiv:2601.18904v2 Announce Type: replace-cross Abstract: Auditory Large Language Models (LLMs) have demonstrated strong performance across a wide range of speech and audio understanding tasks. Nevert

METATR: A Multilingual, Evolving Benchmark for Automatic Text Recognition

Model ReleasesDGX agent

arXiv:2605.26712v1 Announce Type: new Abstract: Benchmarks that reflect the diversity and complexity of real-world documents are essential for accurately evaluating Automatic Text Recognition (ATR) sy

MicroSpec: Accelerating Speculative Decoding with Lightweight In-Context Vocabularies

Local AiDGX agent

arXiv:2605.26444v1 Announce Type: new Abstract: Large language models typically employ vocabularies of over 100k tokens, which creates a major computational bottleneck at the final linear projection l

Mildly Overparameterized ReLU Networks on Orthogonal Data: Incremental Learning and Implicit Bias

SafetyDGX agent

arXiv:2605.27097v1 Announce Type: new Abstract: The successful training of neural networks hinges on the use of first order optimization methods, yet the theoretical characterization of these methods

Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents

AgentsDGX agent

arXiv:2605.26691v1 Announce Type: new Abstract: Medical AI agents increasingly use external tools for diagnosis, treatment recommendation, and evidence retrieval, yet most existing approaches assume t

Minimal surfaces, Knots, and Neural Networks

ResearchDGX agent

arXiv:2605.26234v1 Announce Type: cross Abstract: A recent conjecture by Joel Fine posits a relationship between the coefficients of the HOMFLY polynomial of a knot K in the 3-sphere S^3, and the sign

← Previous
1…581582583584585…1040
Next →