AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
14 Apr 2026

BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection

Model ReleasesDGX agent

arXiv:2604.10389v1 Announce Type: new Abstract: Terminology substitution errors in clinical notes, where one medical term is replaced by a linguistically valid but clinically different term, pose a pe

BMdataset: A Musicologically Curated LilyPond Dataset

ResearchDGX agent

arXiv:2604.10628v1 Announce Type: cross Abstract: Symbolic music research has relied almost exclusively on MIDI-based datasets; text-based engraving formats such as LilyPond remain unexplored for musi

Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation

ApplicationsDGX agent

arXiv:2604.10950v1 Announce Type: new Abstract: Fully supervised Video Semantic Segmentation (VSS) relies heavily on densely annotated video data, limiting practical applicability. Alternatively, appl

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

'bot lane noob' Towards Deployment of NLP-based Toxicity Detectors in Video Games

ResearchDGX agent

arXiv:2604.10175v1 Announce Type: cross Abstract: Toxicity and harassment are widespread in the video-gaming context. Especially in competitive online multiplayer scenarios, gamers oftentimes send har

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

Model ReleasesDGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

Bottleneck Tokens for Unified Multimodal Retrieval

ResearchDGX agent

arXiv:2604.11095v1 Announce Type: cross Abstract: Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rel

Boxes2Pixels: Learning Defect Segmentation from Noisy SAM Masks

Model ReleasesDGX agent

arXiv:2604.11162v1 Announce Type: new Abstract: Accurate defect segmentation is critical for industrial inspection, yet dense pixel-level annotations are rarely available. A common workaround is to co

BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning

ResearchDGX agent

arXiv:2604.11136v1 Announce Type: cross Abstract: Object-level spatial-temporal understanding is essential for video question answering, yet existing multimodal large language models (MLLMs) encode fr

Brain-Grasp: Graph-based Saliency Priors for Improved fMRI-based Visual Brain Decoding

SafetyDGX agent

arXiv:2604.10617v1 Announce Type: cross Abstract: Recent progress in brain-guided image generation has improved the quality of fMRI-based reconstructions; however, fundamental challenges remain in pre

Breaking the KV Cache Bottleneck: Fan Duality Model Achieves O(1) Decode Memory with Superior Associative Recall

ResearchDGX agent

arXiv:2604.07716v2 Announce Type: replace Abstract: We present FDM (Fan Duality Model), a linear sequence architecture that resolves the fundamental tension between memory efficiency and associative r

BRIDGE and TCH-Net: Heterogeneous Benchmark and Multi-Branch Baseline for Cross-Domain IoT Botnet Detection

Model ReleasesDGX agent

arXiv:2604.11324v1 Announce Type: cross Abstract: IoT botnet detection has advanced, yet most published systems are validated on a single dataset and rarely generalise across environments. Heterogeneo

BridgeSim: Unveiling the OL-CL Gap in End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2604.10856v1 Announce Type: cross Abstract: Open-loop (OL) to closed-loop (CL) gap (OL-CL gap) exists when OL-pretrained policies scoring high in OL evaluations fail to transfer effectively in c

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

SafetyDGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

SafetyDGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

Model ReleasesDGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

Brief2Design: A Multi-phased, Compositional Approach to Prompt-based Graphic Design

ResearchDGX agent

arXiv:2604.11019v1 Announce Type: cross Abstract: Professional designers work from client briefs that specify goals and constraints but often lack concrete design details. Translating these abstract r

Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning

ResearchDGX agent

arXiv:2604.10701v1 Announce Type: cross Abstract: Credit assignment is a central challenge in reinforcement learning (RL). Classical actor-critic methods address this challenge through fine-grained ad

Budget-Aware Uncertainty for Radiotherapy Segmentation QA Using nnU-Net

SafetyDGX agent

arXiv:2604.11798v1 Announce Type: cross Abstract: Accurate delineation of the Clinical Target Volume (CTV) is essential for radiotherapy planning, yet remains time-consuming and difficult to assess, e

Byte-level generative predictions for forensics multimedia carving

ResearchDGX agent

arXiv:2604.11010v1 Announce Type: new Abstract: Digital forensic investigations often face significant challenges when recovering fragmented multimedia files that lack file system metadata. While trad

Byzantine-Robust Distributed SGD: A Unified Analysis and Tight Error Bounds

ResearchDGX agent

arXiv:2604.10179v1 Announce Type: cross Abstract: Byzantine-robust distributed optimization relies on robust aggregation rules to mitigate the influence of malicious Byzantine workers. Despite the pro

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

Model ReleasesDGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

C2F-Thinker: Coarse-to-Fine Reasoning with Hint-Guided Reinforcement Learning for Multimodal Sentiment Analysis

SafetyDGX agent

arXiv:2604.00013v2 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to integrate textual, acoustic, and visual information for deep emotional understanding. Despite the progre

CableTract: A Co-Designed Cable-Driven Field Robot for Low-Compaction, Off-Grid Capable Agriculture

ResearchDGX agent

arXiv:2604.09938v1 Announce Type: cross Abstract: Conventional field operations spend most of their energy moving the tractor body, not the implement. Yet feasibility studies for novel agricultural ve

CAGE: Bridging the Accuracy-Aesthetics Gap in Educational Diagrams via Code-Anchored Generative Enhancement

ResearchDGX agent

arXiv:2604.09691v1 Announce Type: cross Abstract: Educational diagrams -- labeled illustrations of biological processes, chemical structures, physical systems, and mathematical concepts -- are essenti

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

SafetyDGX agent

arXiv:2604.11483v1 Announce Type: new Abstract: Goal-directed molecular generation requires satisfying heterogeneous constraints such as protein--ligand compatibility and multi-objective drug-like pro

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

SafetyDGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

Camyla: Scaling Autonomous Research in Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

Can Large Language Models Infer Causal Relationships from Real-World Text?

Model ReleasesDGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

HardwareDGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

Model ReleasesDGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

TutorialsDGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification?

ApplicationsDGX agent

arXiv:2601.06993v2 Announce Type: replace Abstract: Multi-modal large language models (MLLMs) exhibit strong general-purpose capabilities, yet still struggle on Fine-Grained Visual Classification (FGV

CapBench: A Multi-PDK Dataset for Machine-Learning-Based Post-Layout Capacitance Extraction

ResearchDGX agent

arXiv:2604.11202v1 Announce Type: cross Abstract: We present CapBench, a fully reproducible, multi-PDK dataset for capacitance extraction. The dataset is derived from open-source designs, including si

CapyMOA: Efficient Machine Learning for Data Streams and Online Continual Learning in Python

TutorialsDGX agent

arXiv:2502.07432v2 Announce Type: replace Abstract: CapyMOA is an open-source Python library for efficient machine learning on data streams and online continual learning. It provides a structured fram

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

Model ReleasesDGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

Model ReleasesDGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

Model ReleasesDGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

Model ReleasesDGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

CASK: Core-Aware Selective KV Compression for Reasoning Traces

HardwareDGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

ResearchDGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

Catalyst: Out-of-Distribution Detection via Elastic Scaling

ResearchDGX agent

arXiv:2602.02409v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection is critical for the safe deployment of deep neural networks. State-of-the-art post-hoc methods typically derive

CausalGaze: Unveiling Hallucinations via Counterfactual Graph Intervention in Large Language Models

ResearchDGX agent

arXiv:2604.11087v1 Announce Type: new Abstract: Despite the groundbreaking advancements made by large language models (LLMs), hallucination remains a critical bottleneck for their deployment in high-s

Causally Sufficient and Necessary Feature Expansion for Class-Incremental Learning

TutorialsDGX agent

arXiv:2603.09145v2 Announce Type: replace-cross Abstract: Current expansion-based methods for Class Incremental Learning (CIL) effectively mitigate catastrophic forgetting by freezing old features. Ho

CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation

ApplicationsDGX agent

arXiv:2604.11097v1 Announce Type: new Abstract: Monocular depth estimation is a fundamental yet challenging task in computer vision, especially under complex conditions such as textureless surfaces, t

CFMS: A Coarse-to-Fine Multimodal Synthesis Framework for Enhanced Tabular Reasoning

TutorialsDGX agent

arXiv:2604.10973v1 Announce Type: new Abstract: Reasoning over tabular data is a crucial capability for tasks like question answering and fact verification, as it requires models to comprehend both fr

CHAIRO: Contextual Hierarchical Analogical Induction and Reasoning Optimization for LLMs

ApplicationsDGX agent

arXiv:2604.10502v1 Announce Type: new Abstract: Content moderation in online platforms faces persistent challenges due to the evolving complexity of user-generated content and the limitations of tradi

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

HardwareDGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

Model ReleasesDGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

AgentsDGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

Model ReleasesDGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

Choose Your Battles: Distributed Learning Over Multiple Tug of War Games

ResearchDGX agent

arXiv:2509.20147v2 Announce Type: replace-cross Abstract: Consider N players and K games taking place simultaneously. Each of these games is modeled as a Tug-of-War (ToW) game where increasing the act

CID-TKG: Collaborative Historical Invariance and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

SafetyDGX agent

arXiv:2604.09600v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to infer future facts at unseen timestamps from temporally evolving entities and relations. Despite recent

CircuitSynth: Reliable Synthetic Data Generation

ResearchDGX agent

arXiv:2604.10114v1 Announce Type: cross Abstract: The generation of high-fidelity synthetic data is a cornerstone of modern machine learning, yet Large Language Models (LLMs) frequently suffer from ha

City-Wide Low-Altitude Urban Air Mobility: A Scalable Global Path Planning Approach via Risk-Aware Multi-Scale Cell Decomposition

ResearchDGX agent

arXiv:2408.02786v4 Announce Type: replace Abstract: The realization of Urban Air Mobility (UAM) necessitates scalable global path planning algorithms capable of ensuring safe navigation within complex

CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras

SafetyDGX agent

arXiv:2602.18047v3 Announce Type: replace Abstract: City-scale person re-identification across distributed cameras must handle severe appearance changes from viewpoint, occlusion, and domain shift whi

Claim2Vec: Embedding Fact-Check Claims for Multilingual Similarity and Clustering

SafetyDGX agent

arXiv:2604.09812v1 Announce Type: new Abstract: Recurrent claims present a major challenge for automated fact-checking systems designed to combat misinformation, especially in multilingual settings. W

ClaimDB: A Fact Verification Benchmark over Large Structured Data

Model ReleasesDGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

CLASP: Closed-loop Asynchronous Spatial Perception for Open-vocabulary Desktop Object Grasping

ApplicationsDGX agent

arXiv:2604.11320v1 Announce Type: new Abstract: Robot grasping of desktop object is widely used in intelligent manufacturing, logistics, and agriculture.Although vision-language models (VLMs) show str

← Previous
1…938939940941942…989
Next →