AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Agents

Language-Based Agent Control

DGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

agentsarxiv-cs-ai
14 May 2026
Safety

Learning Transferable Latent User Preferences for Human-Aligned Decision Making

DGX agent

arXiv:2605.12682v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as reasoning modules in many applications. While they are efficient in certain tasks, LLMs often stru

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
safetyarxiv-cs-ai
14 May 2026
Safety

Learning with Rare Success but Rich Feedback via Reflection-Enhanced Self-Distillation

DGX agent

arXiv:2605.12741v1 Announce Type: new Abstract: Enabling Large Language Models (LLMs) to continuously improve from environmental interactions is a central challenge in post-training. While on-policy s

safetyarxiv-cs-lg
14 May 2026
Research

LLMs as Implicit Imputers: Uncertainty Should Scale with Missing Information

DGX agent

arXiv:2605.13188v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in settings where the available context is incomplete or degraded. We argue that an LLM generat

researcharxiv-cs-lg
14 May 2026
Local Ai

Local Conformal Calibration of Dynamics Uncertainty from Semantic Images

DGX agent

arXiv:2605.13028v1 Announce Type: new Abstract: We introduce Observation-aware Conformal Uncertainty Local-Calibration (OCULAR), a conformal prediction-based algorithm that uses perception information

local-aiarxiv-cs-ro
14 May 2026
Research

MambaPanoptic: A Vision Mamba-based Structured State Space Framework for Panoptic Segmentation

DGX agent

arXiv:2605.12640v1 Announce Type: new Abstract: Panoptic segmentation requires the simultaneous recognition of countable thing instances and amorphous stuff regions, placing joint demands on long-rang

researcharxiv-cs-cv
14 May 2026
Agents

MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters

DGX agent

arXiv:2605.13496v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly prevalent in cloud-based platforms, propelled by the introduction of AI-based consumer and enter

agentsarxiv-cs-lg
14 May 2026
Safety

MaskPro: Linear-Space Probabilistic Learning for Strict (N:M)-Sparsity on LLMs

DGX agent

arXiv:2506.12876v2 Announce Type: replace Abstract: The rapid scaling of large language models~(LLMs) has made inference efficiency a primary bottleneck in the practical deployment. To address this, s

safetyarxiv-cs-lg
14 May 2026
Local Ai

MorphOPC: Advancing Mask Optimization with Multi-scale Hierarchical Morphological Learning

DGX agent

arXiv:2605.12528v1 Announce Type: cross Abstract: As feature sizes shrink to the nanometer scale, accurately transferring circuit patterns from photomasks to silicon wafers becomes increasingly challe

local-aiarxiv-cs-ai
14 May 2026
Safety

Multi-Rollout On-Policy Distillation via Peer Successes and Failures

DGX agent

arXiv:2605.12652v1 Announce Type: cross Abstract: Large language models are often post-trained with sparse verifier rewards, which indicate whether a sampled trajectory succeeds but provide limited gu

safetyarxiv-cs-ai
14 May 2026
Research

NAACA: Training-Free NeuroAuditory Attentive Cognitive Architecture with Oscillatory Working Memory for Salience-Driven Attention Gating

DGX agent

arXiv:2605.13651v1 Announce Type: cross Abstract: Audio provides critical situational cues, yet current Audio Language Models (ALMs) face an attention bottleneck in long-form recordings where dominant

researcharxiv-cs-ai
14 May 2026
Research

Physics Guided Generative Optimization for Trotter Suzuki Decomposition

DGX agent

arXiv:2605.13268v1 Announce Type: cross Abstract: Product formulas for Trotter Suzuki simulation remain a practical route to Hamiltonian evolution on noisy intermediate scale quantum (NISQ) hardware,

researcharxiv-cs-lg
14 May 2026
Agents

Position: Agentic AI System Is a Foreseeable Pathway to AGI

DGX agent

arXiv:2605.12966v1 Announce Type: new Abstract: Is monolithic scaling the only path to AGI? This paper challenges the dogma that purely scaling a single model is sufficient to achieve Artificial Gener

agentsarxiv-cs-ai
14 May 2026
Safety

Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy

DGX agent

arXiv:2605.13435v1 Announce Type: cross Abstract: There is growing interest in utilizing flow-based models as decision-making policies in reinforcement learning due to their high expressive capacity.

safetyarxiv-cs-ai
14 May 2026
Applications

Quantifying Potential Observation Missingness in Inverse Reinforcement Learning

DGX agent

arXiv:2605.12831v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL), which infers reward functions from demonstrations, is a valuable tool for modeling and understanding decision-maki

applicationsarxiv-cs-lg
14 May 2026
Safety

Real2Sim: A Physics-driven and Editable Gaussian Splatting Framework for Autonomous Driving Scenes

DGX agent

arXiv:2605.13591v1 Announce Type: new Abstract: Reliable autonomous driving relies on large-scale, well-labeled data and robust models. However, manual data collection is resource-intensive, and tradi

safetyarxiv-cs-cv
14 May 2026
Agents

Reinforced Collaboration in Multi-Agent Flow Networks

DGX agent

arXiv:2605.12943v1 Announce Type: new Abstract: Multi-agent systems provide a powerful way to extend large language models (LLMs) by decomposing a complex task into specialized subtasks handled by dif

agentsarxiv-cs-lg
14 May 2026
Tutorials

Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation

DGX agent

arXiv:2605.12975v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a standard approach for knowledge-intensive question answering, but existing systems remain brittle on m

tutorialsarxiv-cs-ai
14 May 2026
Applications

Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations

DGX agent

arXiv:2507.00990v3 Announce Type: replace-cross Abstract: This work introduces Robots Imitating Generated Videos (RIGVid), a system that enables robots to perform complex manipulation tasks--such as p

applicationsarxiv-cs-ai
14 May 2026
Tutorials

Scaling few-shot spoken word classification with generative meta-continual learning

DGX agent

arXiv:2605.13075v1 Announce Type: cross Abstract: Few-shot spoken word classification has largely been developed for applications where a small number of classes is considered, and so the potential of

tutorialsarxiv-cs-ai
14 May 2026
Safety

ScioMind: Cognitively Grounded Multi-Agent Social Simulation with Anchoring-Based Belief Dynamics and Dynamic Profiles

DGX agent

arXiv:2605.13725v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent simulation offers a powerful testbed for studying social opinion dynamics. Yet current approaches often ado

safetyarxiv-cs-ai
14 May 2026
Research

Self-CriTeach: LLM Self-Teaching and Self-Critiquing for Improving Robotic Planning via Automated Domain Generation

DGX agent

arXiv:2509.21543v3 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong promise for robotic task planning, particularly through the automatic generation of symbolic planning

researcharxiv-cs-ro
14 May 2026
Safety

Sharpness-Guided Group Relative Policy Optimization via Probability Shaping

DGX agent

arXiv:2511.00066v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a practical route to improve large language model reasoning, and Group Relative Pol

safetyarxiv-cs-lg
14 May 2026
Agents

SHM-Agents: A Generalist-Specialist Integrated Agent System for Structural Health Monitoring

DGX agent

arXiv:2605.12916v1 Announce Type: cross Abstract: Artificial intelligence is increasingly used to simplify complex tasks. In engineering applications of structural health monitoring (SHM), existing sp

agentsarxiv-cs-lg
14 May 2026
Tutorials

Shortcut Mitigation via Spurious-Positive Samples

DGX agent

arXiv:2605.13340v1 Announce Type: new Abstract: Shortcut mitigation strategies commonly rely on training data annotations, group-balanced held-out data or the presence of all groups, i.e., all combina

tutorialsarxiv-cs-lg
14 May 2026
Safety

SPOT: Selective Prompt Projection via Total Variation for Inference-Only Safe Text-to-Image Generation

DGX agent

arXiv:2602.00616v3 Announce Type: replace Abstract: Text-to-Image (T2I) diffusion models enable high quality open ended synthesis, but practical use requires suppressing unsafe generations while prese

safetyarxiv-cs-ai
14 May 2026
Research

Support-Conditioned Flow Matching Is Kernel Smoothing

DGX agent

arXiv:2605.13386v1 Announce Type: new Abstract: Generative models are often conditioned on a small set of examples via cross-attention. Under the Gaussian optimal-transport path, we show that the exac

researcharxiv-cs-lg
14 May 2026
Applications

Taming the Long Tail: Rebalancing Adversarial Training via Adaptive Perturbation

DGX agent

arXiv:2605.13395v1 Announce Type: cross Abstract: Deep neural networks are highly vulnerable to adversarial examples, i.e.,small perturbations that can significantly degrade model performance. While a

applicationsarxiv-cs-cv
14 May 2026
Safety

Temper and Tilt Lead to SLOP: Reward Hacking Mitigation with Inference-Time Alignment

DGX agent

arXiv:2605.13537v1 Announce Type: cross Abstract: Inference-time alignment techniques offer a lightweight alternative or complement to costly reinforcement learning, while enabling continual adaptatio

safetyarxiv-cs-ai
14 May 2026
Safety

Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents

DGX agent

arXiv:2605.12620v1 Announce Type: new Abstract: Building generalist embodied agents capable of solving complex real-world tasks remains a fundamental challenge in AI. Multimodal Large Language Models

safetyarxiv-cs-ai
14 May 2026
Safety

Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning

DGX agent

arXiv:2602.06475v2 Announce Type: replace Abstract: Large language models (LLMs) excel at complex tasks with advances in reasoning capabilities. However, existing reward mechanisms remain tightly coup

safetyarxiv-cs-lg
14 May 2026
Applications

Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs

DGX agent

arXiv:2605.13530v1 Announce Type: cross Abstract: Surgical scene understanding is a cornerstone of computer-assisted intervention. While recent advances, particularly in surgical image segmentation, h

applicationsarxiv-cs-ai
14 May 2026
Research

Understanding Generalization through Decision Pattern Shift

DGX agent

arXiv:2605.13148v1 Announce Type: cross Abstract: Understanding why deep neural networks (DNNs) fail to generalize to unseen samples remains a long-standing challenge. Existing studies mainly examine

researcharxiv-cs-cv
14 May 2026
Research

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation

DGX agent

arXiv:2605.13277v1 Announce Type: cross Abstract: Visual evidence selection is a critical component of multimodal retrieval-augmented generation (RAG), yet existing methods typically rely on semantic

researcharxiv-cs-ai
14 May 2026
Safety

VERA-MH: Validation of Ethical and Responsible AI in Mental Health

DGX agent

arXiv:2605.13318v1 Announce Type: new Abstract: Chatbot usage has increased, including in fields for which they were never developed for--notably mental health support. To that end, we introduce Valid

safetyarxiv-cs-ai
14 May 2026
Research

4DVGGT-D: 4D Visual Geometry Transformer with Improved Dynamic Depth Estimation

DGX agent

arXiv:2605.12027v1 Announce Type: new Abstract: Reconstructing dynamic 4D scenes from monocular videos is a fundamental yet challenging task. While recent 3D foundation models provide strong geometric

researcharxiv-cs-cv
13 May 2026
Research

A Composite Activation Function for Learning Stable Binary Representations

DGX agent

arXiv:2605.11558v1 Announce Type: new Abstract: Activation functions play a central role in neural networks by shaping internal representations. Recently, learning binary activation representations ha

researcharxiv-cs-lg
13 May 2026
Research

A Formal Comparison Between Chain of Thought and Latent Thought

DGX agent

arXiv:2509.25239v3 Announce Type: replace-cross Abstract: Chain of thought (CoT) elicits reasoning in large language models by explicitly generating intermediate tokens. In contrast, latent thought re

researcharxiv-cs-cl
13 May 2026
Agents

AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

DGX agent

arXiv:2605.11026v1 Announce Type: cross Abstract: Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks

agentsarxiv-cs-cl
13 May 2026
Safety

Aligning Flow Map Policies with Optimal Q-Guidance

DGX agent

arXiv:2605.12416v1 Announce Type: new Abstract: Generative policies based on expressive model classes, such as diffusion and flow matching, are well-suited to complex control problems with highly mult

safetyarxiv-cs-lg
13 May 2026
Research

AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection

DGX agent

arXiv:2605.12430v1 Announce Type: new Abstract: Segmentation models in automated optical inspection of wire-bonded semiconductors are typically device-specific and must be re-trained when new devices

researcharxiv-cs-cv
13 May 2026
Applications

Birds of a Feather Flock Together: Background-Invariant Representations via Linear Structure in VLMs

DGX agent

arXiv:2605.11107v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP and SigLIP 2, are widely used for image classification, yet their vision encoders remain vulnerable to syste

applicationsarxiv-cs-cv
13 May 2026
Research

BronchoLumen: Analysis of recent YOLO-based architectures for real-time bronchial orifice detection in video bronchoscopy

DGX agent

arXiv:2605.11748v1 Announce Type: new Abstract: Bronchoscopy is routinely conducted in pulmonary clinics and intensive care units, but navigating the complex branching of the respiratory tract remains

researcharxiv-cs-cv
13 May 2026
Safety

Causal Bias Detection in Generative Artifical Intelligence

DGX agent

arXiv:2605.11365v1 Announce Type: cross Abstract: Automated systems built on artificial intelligence (AI) are increasingly deployed across high-stakes domains, raising critical concerns about fairness

safetyarxiv-cs-lg
13 May 2026
Research

Context Convergence Improves Answering Inferential Questions

DGX agent

arXiv:2605.12370v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used in open-domain Question Answering (QA), their ability to handle inferential questions-where answers m

researcharxiv-cs-cl
13 May 2026
Research

Deep Learning for Protein Complex Prediction and Design

DGX agent

arXiv:2605.11189v1 Announce Type: new Abstract: Accurately modeling and designing protein complex structures is a central problem in computational structural biology, with broad implications for under

researcharxiv-cs-lg
13 May 2026
Tutorials

DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense Prediction

DGX agent

arXiv:2605.11265v1 Announce Type: new Abstract: Dense prediction tasks in surgical computer vision, such as segmentation and surgical zone prediction, can provide valuable guidance for laparoscopic an

tutorialsarxiv-cs-cv
13 May 2026
Research

Detecting overfitting in Neural Networks during long-horizon grokking using Random Matrix Theory

DGX agent

arXiv:2605.12394v1 Announce Type: new Abstract: Training Neural Networks (NNs) without overfitting is difficult; detecting that overfitting is difficult as well. We present a novel Random Matrix Theor

researcharxiv-cs-lg
13 May 2026
← Previous
1…787788789790791…1038
Next →