AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Safety

Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration

DGX agent

arXiv:2601.06160v2 Announce Type: replace Abstract: Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produce

safetyarxiv-cs-ai
30 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Safety

BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

DGX agent

arXiv:2604.25203v1 Announce Type: new Abstract: Deploying guardrails for custom policies remains challenging, as generic safety models fail to capture task-specific requirements, while prompting LLMs

safetyarxiv-cs-cl
29 Apr 2026
Safety

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

DGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination

DGX agent

arXiv:2604.24978v1 Announce Type: new Abstract: Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

DGX agent

arXiv:2604.25136v1 Announce Type: new Abstract: We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to

safetyarxiv-cs-cl
29 Apr 2026
Safety

How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning

DGX agent

arXiv:2603.01070v2 Announce Type: replace Abstract: Solving complex geometric problems inherently requires interleaved reasoning: a tight alternation between constructing diagrams and performing logic

safetyarxiv-cs-cl
29 Apr 2026
Safety

Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System

DGX agent

arXiv:2604.24921v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into ex

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

M^3-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering

DGX agent

arXiv:2604.25122v1 Announce Type: new Abstract: We present M^3-VQA, a novel knowledge-based Visual Question Answering (VQA) benchmark, to enhance the evaluation of multimodal large language models (ML

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

DGX agent

arXiv:2604.24954v1 Announce Type: cross Abstract: We introduce Nemotron 3 Nano Omni, the latest model in the Nemotron multimodal series and the first to natively support audio inputs alongside text, i

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

DGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Toward Multimodal Conversational AI for Age-Related Macular Degeneration

DGX agent

arXiv:2604.25720v1 Announce Type: cross Abstract: Despite strong performance of deep learning models in retinal disease detection, most systems produce static predictions without clinical reasoning or

researcharxiv-cs-cl
29 Apr 2026
Safety

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

DGX agent

arXiv:2510.15479v2 Announce Type: replace Abstract: We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatmen

safetyarxiv-cs-lg
28 Apr 2026
Safety

Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI

DGX agent

arXiv:2604.22765v1 Announce Type: cross Abstract: The increasing use of artificial intelligence (AI) by public authorities introduces both opportunities for innovation and significant challenges for t

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

DGX agent

arXiv:2604.23195v1 Announce Type: cross Abstract: Analog circuit design relies heavily on reusing existing intellectual property (IP), yet searching across heterogeneous representations such as SPICE

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Beyond Match Maximization and Fairness: Retention-Optimized Two-Sided Matching

DGX agent

arXiv:2602.15752v2 Announce Type: replace Abstract: On two-sided matching platforms such as online dating and recruiting, recommendation algorithms often aim to maximize the total number of matches. H

safetyarxiv-cs-lg
28 Apr 2026
Model Releases

Bridging the Pose-Semantic Gap: A Cascade Framework for Text-Based Person Anomaly Search

DGX agent

arXiv:2604.23282v1 Announce Type: new Abstract: Text-based person anomaly search retrieves specific behavioral events from surveillance archives using natural-language queries. Although recent pose-aw

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

DGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

safetyarxiv-cs-ai
28 Apr 2026
Safety

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

DGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

safetyarxiv-cs-ai
28 Apr 2026
Research

Caries DETR: Tooth Structure-aware Prior and Lesion-aware Dynamic Loss Refinement for DETR Based Caries Detection

DGX agent

arXiv:2604.23718v1 Announce Type: new Abstract: As dental caries appear as subtle, low-contrast lesions in intraoral imaging, existing deep learning models face significant challenges in the early det

researcharxiv-cs-cv
28 Apr 2026
Model Releases

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

DGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

DGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Safety

Evaluating Language Models' Evaluations of Games

DGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

safetyarxiv-cs-ai
28 Apr 2026
Safety

Extending Precipitation Nowcasting Horizons via Spectral Fusion of Radar Observations and Foundation Model Priors

DGX agent

arXiv:2603.21768v3 Announce Type: replace-cross Abstract: Precipitation nowcasting is critical for disaster mitigation and aviation safety. However, radar-only models frequently suffer from a lack of

safetyarxiv-cs-ai
28 Apr 2026
Local Ai

GeoFunFlow-3D: A Physics-Guided Generative Flow Matching Framework for High-Fidelity 3D Aerodynamic Inference over Complex Geometries

DGX agent

arXiv:2604.23350v1 Announce Type: cross Abstract: Deep generative models and neural operators have demonstrated significant potential for 3D aerodynamic inference. However, they often face inherent ch

local-aiarxiv-cs-lg
28 Apr 2026
Model Releases

Green Shielding: A User-Centric Approach Towards Trustworthy AI

DGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Grounding Before Generalizing: How AI Differs from Humans in Causal Transfer

DGX agent

arXiv:2604.24062v1 Announce Type: new Abstract: Extracting abstract causal structures and applying them to novel situations is a hallmark of human intelligence. While Large Language Models (LLMs) and

researcharxiv-cs-ai
28 Apr 2026
Safety

Institutions for the Post-Scarcity of Judgment

DGX agent

arXiv:2604.22966v1 Announce Type: cross Abstract: Each major technological revolution inverts a particular scarcity and rebuilds institutions around the shift. The near-consensus diagnosis of the AI r

safetyarxiv-cs-ai
28 Apr 2026
Research

IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models

DGX agent

arXiv:2604.24002v1 Announce Type: cross Abstract: Improving the effectiveness of human-robot interaction requires social robots to accurately infer human goals through robust intention understanding.

researcharxiv-cs-ai
28 Apr 2026
Hardware

JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training

DGX agent

arXiv:2604.23838v1 Announce Type: new Abstract: We present JigsawRL, a cost-efficient framework that explores Pipeline Multiplexing as a new dimension of RL parallelism. JigsawRL decomposes each pipel

hardwarearxiv-cs-lg
28 Apr 2026
Model Releases

K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology

DGX agent

arXiv:2604.24645v1 Announce Type: cross Abstract: The development of practical (multimodal) large language model assistants for Korean weather forecasters is hindered by the absence of a multidimensio

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Kwai Summary Attention Technical Report

DGX agent

arXiv:2604.24432v1 Announce Type: cross Abstract: Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic unde

local-aiarxiv-cs-ai
28 Apr 2026
Safety

Large Language Model based Interactive Decision-Making for Autonomous Driving

DGX agent

arXiv:2604.23513v1 Announce Type: new Abstract: In high-conflict mixed-traffic scenarios involving human-driven and autonomous vehicles, most existing autonomous driving systems default to overly cons

safetyarxiv-cs-ro
28 Apr 2026
Safety

Learning Under Moral Hazard with Instrumental Regression and Generalized Method of Moments

DGX agent

arXiv:2405.20642v3 Announce Type: replace Abstract: Machine learning has become increasingly popular in informing data-driven policy-making. Policies influence behavior in individuals or populations,

safetyarxiv-cs-lg
28 Apr 2026
Safety

LLM-Auction: Generative Auction towards LLM-Native Advertising

DGX agent

arXiv:2512.10551v2 Announce Type: replace-cross Abstract: The commercialization of LLM applications is the next frontier in online advertising, with LLM-native advertising emerging as a promising para

safetyarxiv-cs-ai
28 Apr 2026
Research

Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing

DGX agent

arXiv:2604.23247v1 Announce Type: new Abstract: Avatar fingerprinting, i.e., verifying who drives a synthetic talking-head video rather than whether it is real, is a critical safeguard for authorized

researcharxiv-cs-cv
28 Apr 2026
Research

Multimodal QUD: Inquisitive Questions from Scientific Figures

DGX agent

arXiv:2604.23733v1 Announce Type: new Abstract: Asking inquisitive questions while reading, and looking for their answers, is an important part in human discourse comprehension, curiosity, and creativ

researcharxiv-cs-cl
28 Apr 2026
Research

NVILA: Efficient Frontier Visual Language Models

DGX agent

arXiv:2412.04468v3 Announce Type: replace Abstract: Visual language models (VLMs) have made significant advances in accuracy in recent years. However, their efficiency has received much less attention

researcharxiv-cs-cv
28 Apr 2026
Safety

OAMVOS:2nd Report for 5th PVUW MOSE Track

DGX agent

arXiv:2604.22837v1 Announce Type: cross Abstract: SAM-based dense trackers provide strong short-term mask propagation but remain fragile under long occlusion, fast motion, viewpoint change, and distra

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

DGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

DGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Perfecting Aircraft Maneuvers with Reinforcement Learning

DGX agent

arXiv:2604.24338v1 Announce Type: new Abstract: This paper evaluates an advanced jet trainer's utilization of artificial intelligence (AI)-based aircraft aerobatic maneuvers with the intention of deve

researcharxiv-cs-lg
28 Apr 2026
Model Releases

PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis

DGX agent

arXiv:2512.22113v2 Announce Type: replace-cross Abstract: Unresolved production cloud incidents cost an average of over $2M per hour. This paper introduces PRAXIS, an orchestrator that manages and dep

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Probing Visual Planning in Image Editing Models

DGX agent

arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m

tutorialsarxiv-cs-ai
28 Apr 2026
Safety

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

DGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

safetyarxiv-cs-ai
28 Apr 2026
Safety

Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes

DGX agent

arXiv:2603.25562v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) is increasingly used in LLM post-training because it can leverage a teacher model to provide dense supervision on

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…221222223224225…230
Next →