AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Safety

MemReward: Graph-Based Experience Memory for LLM Reward Prediction with Limited Labels

DGX agent

arXiv:2603.19310v3 Announce Type: replace Abstract: Reinforcement learning has emerged as a powerful paradigm for improving large language model (LLM) reasoning, where rollouts are sampled from the po

safetyarxiv-cs-lg
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

MMD-Balls as Credal Sets: A PAC-Bayesian Framework for Epistemic Uncertainty in Test-Time Adaptation

DGX agent

arXiv:2605.21783v1 Announce Type: new Abstract: Test-time adaptation (TTA) methods improve model performance under distribution shift but lack formal guarantees connecting shift magnitude to predictio

researcharxiv-cs-lg
23 May 2026
Local Ai

OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning

DGX agent

arXiv:2605.21851v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become the standard recipe for improving LLM reasoning, but the dominant algorithm GRPO assigns a sin

local-aiarxiv-cs-lg
23 May 2026
Applications

Physics-Informed Generative Solver: Bridging Data-Driven Priors and Conservation Laws for Stable Spatiotemporal Field Reconstruction

DGX agent

arXiv:2605.22338v1 Announce Type: new Abstract: Reconstructing continuous physical fields from sparse measurements is a central inverse problem, but data-driven generative models can produce states th

applicationsarxiv-cs-lg
23 May 2026
Safety

Post-Training is About States, Not Tokens: A State Distribution View of SFT, RL, and On-Policy Distillation

DGX agent

arXiv:2605.22731v1 Announce Type: new Abstract: Large language model post-training methods such as supervised fine-tuning (SFT), reinforcement learning (RL), and distillation are often analyzed throug

safetyarxiv-cs-lg
23 May 2026
Research

Regret-Based (epsilon,elta)-optimal Stopping Criteria for Bayesian Optimization

DGX agent

arXiv:2605.22561v1 Announce Type: new Abstract: Bayesian optimization (BO) is a widely used iterative black-box optimization method that utilizes Gaussian process (GP) surrogate models. In practice, B

researcharxiv-cs-lg
23 May 2026
Research

Reinforced Graph of Thoughts: RL-Driven Adaptive Prompting for LLMs

DGX agent

arXiv:2605.22195v1 Announce Type: new Abstract: Graph of Thoughts (GoT), a generalized form of recent prompting paradigms for large language models (LLMs), has been shown to be useful for elaborate pr

researcharxiv-cs-lg
23 May 2026
Safety

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.22263v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is an emerging LLM post-training paradigm in which the model serves as its own teacher: conditioned on privileged inf

safetyarxiv-cs-lg
23 May 2026
Applications

Temporal Contrastive Transformer for Financial Crime Detection: Self-Supervised Sequence Embeddings via Predictive Contrastive Coding

DGX agent

arXiv:2605.21490v1 Announce Type: new Abstract: We introduce the Temporal Contrastive Transformer (TCT), a representation learning framework designed to capture contextual temporal dynamics in sequenc

applicationsarxiv-cs-lg
23 May 2026
Safety

The Signal in the Noise: OOD Detection Through Goodness-of-Fit Testing in Factorised Latent Spaces

DGX agent

arXiv:2605.22496v1 Announce Type: new Abstract: Deep generative models offer a natural foundation for out-of-distribution (OOD) detection, yet prior work has shown that their assigned likelihoods are

safetyarxiv-cs-lg
23 May 2026
Safety

Twice Sequential Monte Carlo for Tree Search

DGX agent

arXiv:2511.14220v3 Announce Type: replace Abstract: Model-based reinforcement learning (RL) methods that leverage search are responsible for many milestone breakthroughs in RL. Sequential Monte Carlo

safetyarxiv-cs-lg
23 May 2026
Safety

What are the Right Symmetries for Formal Theorem Proving?

DGX agent

arXiv:2605.22257v1 Announce Type: new Abstract: Formal theorem provers based on large language models (LLMs) are highly sensitive to superficial variations in problem representation: semantically equi

safetyarxiv-cs-lg
23 May 2026
Research

Winner-Take-All bottlenecks enforce disentangled symbolic representations in multi-task learning

DGX agent

arXiv:2605.22472v1 Announce Type: new Abstract: Winner-take-all (WTA) networks constitute a central circuit motif in cortical networks of the brain. In addition, WTA-like activations are abundant in m

researcharxiv-cs-lg
23 May 2026
Agents

ACC: Compiling Agent Trajectories for Long-Context Training

DGX agent

arXiv:2605.21850v1 Announce Type: new Abstract: Recent development of agents has renewed demand for long-context reasoning capacity of LLMs. However, training LLMs for this capacity requires costly lo

agentsarxiv-cs-cl
22 May 2026
Research

Action with Visual Primitives

DGX agent

arXiv:2605.22183v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for generalist robotic manipulation. A common design in current architectures m

researcharxiv-cs-ro
22 May 2026
Agents

An Application-Layer Multi-Modal Covert-Channel Reference Monitor for LLM Agent Egress

DGX agent

arXiv:2605.20734v1 Announce Type: cross Abstract: A large language model (LLM) agent that sends messages can leak data inside them. Destination allowlists and content scanners do not police whether an

agentsarxiv-cs-ai
22 May 2026
Research

Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation

DGX agent

arXiv:2605.22435v1 Announce Type: new Abstract: Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs)

researcharxiv-cs-cl
22 May 2026
Local Ai

Auction-Consensus Algorithm with Learned Bidding Scheme for Multi-Robot Systems

DGX agent

arXiv:2605.21932v1 Announce Type: new Abstract: Multi-Robot Task Allocation (MRTA) is a central challenge in decentralized multi-agent systems, where teams of robots must cooperatively assign and exec

local-aiarxiv-cs-ro
22 May 2026
Agents

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

DGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

agentsarxiv-cs-ai
22 May 2026
Applications

BodyReLux: Temporally Consistent Full-Body Video Relighting

DGX agent

arXiv:2605.21766v1 Announce Type: new Abstract: Being able to relight human performance is a fundamental task for post production and content creation. We present BodyReLux, a subject-specific video d

applicationsarxiv-cs-cv
22 May 2026
Research

DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA

DGX agent

arXiv:2605.22411v1 Announce Type: new Abstract: Large language model (LLM) agents still struggle with long-term memory question answering, where answer-supporting evidence is often scattered across lo

researcharxiv-cs-cl
22 May 2026
Applications

Demystifying Transition Matching: When and Why It Can Beat Flow Matching

DGX agent

arXiv:2510.17991v3 Announce Type: replace-cross Abstract: Flow Matching (FM) underpins many state-of-the-art generative models, yet recent results indicate that Transition Matching (TM) can achieve hi

applicationsarxiv-cs-cv
22 May 2026
Research

Detecting Synthetic Political Narratives in Cross-Platform Social Media Discourse

DGX agent

arXiv:2605.21540v1 Announce Type: cross Abstract: The proliferation of large language models has introduced a new paradigm of synthetic political communication in which narratives may be generated, se

researcharxiv-cs-cl
22 May 2026
Tutorials

Diverge to Induce Prompting: Multi-Rationale Induction for Zero-Shot Reasoning

DGX agent

arXiv:2602.08028v1 Announce Type: cross Abstract: To address the instability of unguided reasoning paths in standard Chain-of-Thought prompting, recent methods guide large language models (LLMs) by fi

tutorialsarxiv-cs-ai
22 May 2026
Hardware

EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Generation

DGX agent

arXiv:2605.22051v1 Announce Type: new Abstract: Generating high-fidelity visual effects (VFX) typically demands massive datasets and prohibitive computational power due to the intricate coupling of sp

hardwarearxiv-cs-cv
22 May 2026
Agents

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

DGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

agentsarxiv-cs-ai
22 May 2026
Agents

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning

DGX agent

arXiv:2605.22208v1 Announce Type: new Abstract: Multimodal Large Language Model (MLLM)-driven image restoration agent demonstrates effectiveness in degradation coupling scenarios by flexibly selecting

agentsarxiv-cs-cv
22 May 2026
Applications

Exposing Vulnerabilities in Visible-Infrared VLMs: A Unified Geometric Adversarial Framework with Cross-Task Transferability

DGX agent

arXiv:2605.22273v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, but their adversarial robustness in visible-infrared (VI

applicationsarxiv-cs-cv
22 May 2026
Safety

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

DGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

safetyarxiv-cs-cl
22 May 2026
Research

Improving Viewpoint-Invariance and Temporal Consistency for Action Detection

DGX agent

arXiv:2605.22695v1 Announce Type: new Abstract: Viewpoint change invariance and action temporal consistency are critical aspects for the effective deployment of human action detection of untrimmed vid

researcharxiv-cs-cv
22 May 2026
Agents

Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments

DGX agent

arXiv:2605.22189v1 Announce Type: new Abstract: Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches

agentsarxiv-cs-ro
22 May 2026
Agents

LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management

DGX agent

arXiv:2605.12321v2 Announce Type: replace Abstract: Large language models (LLMs) show strong potential for Intelligent Transportation Systems (ITS), particularly in tasks requiring situational reasoni

agentsarxiv-cs-ai
22 May 2026
Research

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering

DGX agent

arXiv:2605.22269v1 Announce Type: new Abstract: Long streaming video QA remains challenging due to growing visual tokens and limited reasoning length of large language models (LLMs). KV-caching stores

researcharxiv-cs-cv
22 May 2026
Hardware

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

DGX agent

arXiv:2605.22015v1 Announce Type: new Abstract: Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spa

hardwarearxiv-cs-cv
22 May 2026
Applications

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought

DGX agent

arXiv:2605.22013v1 Announce Type: new Abstract: Understanding 3D point clouds through language remains a fundamental challenge in computer graphics and visual computing, due to the irregular structure

applicationsarxiv-cs-cv
22 May 2026
Research

REACH: Hand Pose Estimation from Room Corners

DGX agent

arXiv:2605.22231v1 Announce Type: new Abstract: We introduce a novel 3D hand pose estimator that can accurately recover the shape and pose of people's hands in a room from afar, typically from fixed c

researcharxiv-cs-cv
22 May 2026
Safety

Reducing Political Manipulation with Consistency Training

DGX agent

arXiv:2605.22771v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic political bias across a variety of sensitive contexts. We find that LLMs handle counterpart topics from

safetyarxiv-cs-cl
22 May 2026
Safety

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

DGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

safetyarxiv-cs-ro
22 May 2026
Agents

STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking

DGX agent

arXiv:2507.03674v3 Announce Type: replace Abstract: Extracting structured information from scientific literature is critical for accelerating discovery, yet Large Language Models (LLMs) often struggle

agentsarxiv-cs-cl
22 May 2026
Safety

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

DGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

safetyarxiv-cs-ro
22 May 2026
Safety

Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining

DGX agent

arXiv:2605.21906v1 Announce Type: new Abstract: Computed tomography (CT) is a central to three-dimensional medical imaging, yet CT-based artificial intelligence remains fragmented across task-specific

safetyarxiv-cs-cv
22 May 2026
Safety

Value-Gradient Hypothesis of RL for LLMs

DGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

safetyarxiv-cs-cl
22 May 2026
Research

Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning

DGX agent

arXiv:2601.23224v2 Announce Type: replace Abstract: Existing multimodal large language models for long-video understanding predominantly rely on uniform sampling and single-turn inference, limiting th

researcharxiv-cs-cv
22 May 2026
Agents

What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

DGX agent

arXiv:2602.01334v2 Announce Type: replace Abstract: Vision tool-use reinforcement learning (RL) can equip vision language models with visual operators such as crop-and-zoom and achieves strong perform

agentsarxiv-cs-cv
22 May 2026
Local Ai

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs

DGX agent

arXiv:2605.22823v1 Announce Type: new Abstract: Video Large Language Models (Video-LLMs) have made rapid progress on temporal video understanding, yet many fail at a basic perceptual primitive: signed

local-aiarxiv-cs-cv
22 May 2026
Hardware

WorldKV: Efficient World Memory with World Retrieval and Compression

DGX agent

arXiv:2605.22718v1 Announce Type: new Abstract: Autoregressive video diffusion models have enabled real-time, action-conditioned world generation. However, sustaining a persistent world, where revisit

hardwarearxiv-cs-cv
22 May 2026
Safety

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

DGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

safetyarxiv-cs-cl
22 May 2026
Applications

A New Framework to Analyse the Distributional Robustness of Deep Neural Networks

DGX agent

arXiv:2605.21313v1 Announce Type: new Abstract: Deep neural networks have achieved impressive performance on a variety of tasks, but their brittleness to distributional shifts remains a significant ba

applicationsarxiv-cs-lg
21 May 2026
← Previous
1…778779780781782…1038
Next →