AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
Safety

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

DGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

safetyarxiv-cs-ai
14 Apr 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

DGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

DGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Brief2Design: A Multi-phased, Compositional Approach to Prompt-based Graphic Design

DGX agent

arXiv:2604.11019v1 Announce Type: cross Abstract: Professional designers work from client briefs that specify goals and constraints but often lack concrete design details. Translating these abstract r

researcharxiv-cs-ai
14 Apr 2026
Research

Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning

DGX agent

arXiv:2604.10701v1 Announce Type: cross Abstract: Credit assignment is a central challenge in reinforcement learning (RL). Classical actor-critic methods address this challenge through fine-grained ad

researcharxiv-cs-ai
14 Apr 2026
Safety

Budget-Aware Uncertainty for Radiotherapy Segmentation QA Using nnU-Net

DGX agent

arXiv:2604.11798v1 Announce Type: cross Abstract: Accurate delineation of the Clinical Target Volume (CTV) is essential for radiotherapy planning, yet remains time-consuming and difficult to assess, e

safetyarxiv-cs-ai
14 Apr 2026
Research

Byte-level generative predictions for forensics multimedia carving

DGX agent

arXiv:2604.11010v1 Announce Type: new Abstract: Digital forensic investigations often face significant challenges when recovering fragmented multimedia files that lack file system metadata. While trad

researcharxiv-cs-cv
14 Apr 2026
Research

Byzantine-Robust Distributed SGD: A Unified Analysis and Tight Error Bounds

DGX agent

arXiv:2604.10179v1 Announce Type: cross Abstract: Byzantine-robust distributed optimization relies on robust aggregation rules to mitigate the influence of malicious Byzantine workers. Despite the pro

researcharxiv-cs-lg
14 Apr 2026
Model Releases

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

DGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

C2F-Thinker: Coarse-to-Fine Reasoning with Hint-Guided Reinforcement Learning for Multimodal Sentiment Analysis

DGX agent

arXiv:2604.00013v2 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to integrate textual, acoustic, and visual information for deep emotional understanding. Despite the progre

safetyarxiv-cs-ai
14 Apr 2026
Research

CableTract: A Co-Designed Cable-Driven Field Robot for Low-Compaction, Off-Grid Capable Agriculture

DGX agent

arXiv:2604.09938v1 Announce Type: cross Abstract: Conventional field operations spend most of their energy moving the tractor body, not the implement. Yet feasibility studies for novel agricultural ve

researcharxiv-cs-lg
14 Apr 2026
Research

CAGE: Bridging the Accuracy-Aesthetics Gap in Educational Diagrams via Code-Anchored Generative Enhancement

DGX agent

arXiv:2604.09691v1 Announce Type: cross Abstract: Educational diagrams -- labeled illustrations of biological processes, chemical structures, physical systems, and mathematical concepts -- are essenti

researcharxiv-cs-ai
14 Apr 2026
Safety

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

DGX agent

arXiv:2604.11483v1 Announce Type: new Abstract: Goal-directed molecular generation requires satisfying heterogeneous constraints such as protein--ligand compatibility and multi-objective drug-like pro

safetyarxiv-cs-lg
14 Apr 2026
Safety

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

DGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Camyla: Scaling Autonomous Research in Medical Image Segmentation

DGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

DGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

DGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

DGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

tutorialsarxiv-cs-ai
14 Apr 2026
Applications

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification?

DGX agent

arXiv:2601.06993v2 Announce Type: replace Abstract: Multi-modal large language models (MLLMs) exhibit strong general-purpose capabilities, yet still struggle on Fine-Grained Visual Classification (FGV

applicationsarxiv-cs-cv
14 Apr 2026
Research

CapBench: A Multi-PDK Dataset for Machine-Learning-Based Post-Layout Capacitance Extraction

DGX agent

arXiv:2604.11202v1 Announce Type: cross Abstract: We present CapBench, a fully reproducible, multi-PDK dataset for capacitance extraction. The dataset is derived from open-source designs, including si

researcharxiv-cs-lg
14 Apr 2026
Tutorials

CapyMOA: Efficient Machine Learning for Data Streams and Online Continual Learning in Python

DGX agent

arXiv:2502.07432v2 Announce Type: replace Abstract: CapyMOA is an open-source Python library for efficient machine learning on data streams and online continual learning. It provides a structured fram

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

DGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

DGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

DGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

DGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

model-releasesarxiv-cs-cl
14 Apr 2026
Hardware

CASK: Core-Aware Selective KV Compression for Reasoning Traces

DGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

hardwarearxiv-cs-ai
14 Apr 2026
Research

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

DGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

researcharxiv-cs-cl
14 Apr 2026
Research

Catalyst: Out-of-Distribution Detection via Elastic Scaling

DGX agent

arXiv:2602.02409v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection is critical for the safe deployment of deep neural networks. State-of-the-art post-hoc methods typically derive

researcharxiv-cs-cv
14 Apr 2026
Research

CausalGaze: Unveiling Hallucinations via Counterfactual Graph Intervention in Large Language Models

DGX agent

arXiv:2604.11087v1 Announce Type: new Abstract: Despite the groundbreaking advancements made by large language models (LLMs), hallucination remains a critical bottleneck for their deployment in high-s

researcharxiv-cs-lg
14 Apr 2026
Tutorials

Causally Sufficient and Necessary Feature Expansion for Class-Incremental Learning

DGX agent

arXiv:2603.09145v2 Announce Type: replace-cross Abstract: Current expansion-based methods for Class Incremental Learning (CIL) effectively mitigate catastrophic forgetting by freezing old features. Ho

tutorialsarxiv-cs-ai
14 Apr 2026
Applications

CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation

DGX agent

arXiv:2604.11097v1 Announce Type: new Abstract: Monocular depth estimation is a fundamental yet challenging task in computer vision, especially under complex conditions such as textureless surfaces, t

applicationsarxiv-cs-cv
14 Apr 2026
Tutorials

CFMS: A Coarse-to-Fine Multimodal Synthesis Framework for Enhanced Tabular Reasoning

DGX agent

arXiv:2604.10973v1 Announce Type: new Abstract: Reasoning over tabular data is a crucial capability for tasks like question answering and fact verification, as it requires models to comprehend both fr

tutorialsarxiv-cs-ai
14 Apr 2026
Applications

CHAIRO: Contextual Hierarchical Analogical Induction and Reasoning Optimization for LLMs

DGX agent

arXiv:2604.10502v1 Announce Type: new Abstract: Content moderation in online platforms faces persistent challenges due to the evolving complexity of user-generated content and the limitations of tradi

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Hardware

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

DGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

DGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

DGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

agentsarxiv-cs-cl
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

DGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Choose Your Battles: Distributed Learning Over Multiple Tug of War Games

DGX agent

arXiv:2509.20147v2 Announce Type: replace-cross Abstract: Consider N players and K games taking place simultaneously. Each of these games is modeled as a Tug-of-War (ToW) game where increasing the act

researcharxiv-cs-lg
14 Apr 2026
Safety

CID-TKG: Collaborative Historical Invariance and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

DGX agent

arXiv:2604.09600v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to infer future facts at unseen timestamps from temporally evolving entities and relations. Despite recent

safetyarxiv-cs-ai
14 Apr 2026
Research

CircuitSynth: Reliable Synthetic Data Generation

DGX agent

arXiv:2604.10114v1 Announce Type: cross Abstract: The generation of high-fidelity synthetic data is a cornerstone of modern machine learning, yet Large Language Models (LLMs) frequently suffer from ha

researcharxiv-cs-ai
14 Apr 2026
Research

City-Wide Low-Altitude Urban Air Mobility: A Scalable Global Path Planning Approach via Risk-Aware Multi-Scale Cell Decomposition

DGX agent

arXiv:2408.02786v4 Announce Type: replace Abstract: The realization of Urban Air Mobility (UAM) necessitates scalable global path planning algorithms capable of ensuring safe navigation within complex

researcharxiv-cs-ro
14 Apr 2026
Safety

CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras

DGX agent

arXiv:2602.18047v3 Announce Type: replace Abstract: City-scale person re-identification across distributed cameras must handle severe appearance changes from viewpoint, occlusion, and domain shift whi

safetyarxiv-cs-cv
14 Apr 2026
Safety

Claim2Vec: Embedding Fact-Check Claims for Multilingual Similarity and Clustering

DGX agent

arXiv:2604.09812v1 Announce Type: new Abstract: Recurrent claims present a major challenge for automated fact-checking systems designed to combat misinformation, especially in multilingual settings. W

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

ClaimDB: A Fact Verification Benchmark over Large Structured Data

DGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

model-releasesarxiv-cs-cl
14 Apr 2026
Applications

CLASP: Closed-loop Asynchronous Spatial Perception for Open-vocabulary Desktop Object Grasping

DGX agent

arXiv:2604.11320v1 Announce Type: new Abstract: Robot grasping of desktop object is widely used in intelligent manufacturing, logistics, and agriculture.Although vision-language models (VLMs) show str

applicationsarxiv-cs-ro
14 Apr 2026
← Previous
1…11731174117511761177…1236
Next →