AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,046 results
Research

DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation

DGX agent

arXiv:2605.07221v1 Announce Type: new Abstract: Adapting foundation models to medical segmentation typically requires either backbone fine-tuning or high-capacity task-specific decoders, both of which

researcharxiv-cs-cv
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Does Your Neural Network Extrapolate? Feature Engineering as Identifiability Bias for OOD Generalization

DGX agent

arXiv:2605.07483v1 Announce Type: cross Abstract: Successful deep neural networks discover salient features of data. We show when and why they fail to learn out-of-distribution (OOD)-relevant represen

safetyarxiv-cs-ai
11 May 2026
Safety

Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow

DGX agent

arXiv:2605.07727v1 Announce Type: cross Abstract: We propose Drifting Field Policy (DFP), a non-ODE one-step generative policy built on the drifting model paradigm. We frame the policy update as a rev

safetyarxiv-cs-ai
11 May 2026
Research

DVD: Discrete Voxel Diffusion for 3D Generation and Editing

DGX agent

arXiv:2605.07971v1 Announce Type: new Abstract: We introduce Discrete Voxel Diffusion (DVD), a discrete diffusion framework to generate, assess, and edit sparse voxels for SLat (Structured LATent) bas

researcharxiv-cs-cv
11 May 2026
Safety

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

DGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

safetyarxiv-cs-ai
11 May 2026
Safety

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

DGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

safetyarxiv-cs-ai
11 May 2026
Tutorials

Every Feedforward Neural Network Definable in an o-Minimal Structure Has Finite Sample Complexity

DGX agent

arXiv:2605.07097v1 Announce Type: cross Abstract: We show that, in a precise sense, a broad class of feedforward neural networks learn (have finite sample complexity) in the PAC model: every fixed fin

tutorialsarxiv-cs-lg
11 May 2026
Safety

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

DGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

safetyarxiv-cs-ai
11 May 2026
Research

Exposing and Mitigating Temporal Attack in Deepfake Video Detection

DGX agent

arXiv:2605.07398v1 Announce Type: cross Abstract: While spatiotemporal deepfake detectors achieve high AUC, our experiments reveal their susceptibility to evasion attacks. These models tend to overfit

researcharxiv-cs-ai
11 May 2026
Safety

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

DGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

safetyarxiv-cs-ai
11 May 2026
Research

From Canopy to Collision: A Hybrid Predictive Framework for Identifying Risk Factors in Tree-Involved Traffic Crashes

DGX agent

arXiv:2605.06684v1 Announce Type: new Abstract: Tree-involved crashes represent a critical subset of run-off-road (ROR) collisions, often resulting in fatal or severe injuries due to high-energy impac

researcharxiv-cs-lg
11 May 2026
Agents

From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms

DGX agent

arXiv:2605.06716v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents have fundamentally reshaped artificial intelligence by integrating external tools and planning capabilities. Whi

agentsarxiv-cs-ai
11 May 2026
Safety

From Time Series Analysis to Question Answering: A Survey in the LLM Era

DGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

safetyarxiv-cs-ai
11 May 2026
Applications

Geometric Analysis of Neural Regression Collapse via Intrinsic Dimension

DGX agent

arXiv:2510.01105v2 Announce Type: replace Abstract: Neural multivariate regression underpins a wide range of domains, including control, robotics, and finance, yet the geometry of its learned represen

applicationsarxiv-cs-lg
11 May 2026
Safety

GRAPE: Let GRPO Supervise Query Rewriting by Ranking for Retrieval

DGX agent

arXiv:2509.23370v2 Announce Type: replace Abstract: The CLIP model has established itself as a cornerstone of large-scale retrieval systems. However, its performance often degrades under distributiona

safetyarxiv-cs-cv
11 May 2026
Agents

GraphDC: A Divide-and-Conquer Multi-Agent System for Scalable Graph Algorithm Reasoning

DGX agent

arXiv:2605.06671v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong potential for many mathematical problems. However, their performance on graph algorithmic tasks is

agentsarxiv-cs-ai
11 May 2026
Research

GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasks

DGX agent

arXiv:2605.07454v1 Announce Type: new Abstract: In-context learning enables large language models to adapt to new tasks, but their performance is highly sensitive to the selected examples. Finding eff

researcharxiv-cs-cl
11 May 2026
Tools

Here's the full TIL https://til.simonwillison.net/llms/llm-shebang

DGX agent

Simon Willison shares a technique for using Large Language Models directly from the command line using a shebang (#!) syntax, allowing scripts to be executed with LLM processing without explicit comma

toolssimon-willison--x
11 May 2026
Local Ai

HMACE: Heterogeneous Multi-Agent Collaborative Evolution for Combinatorial Optimization

DGX agent

arXiv:2605.07214v1 Announce Type: new Abstract: Large Language Models have recently emerged as a promising paradigm for automated heuristic design for NP-hard combinatorial optimization problems. Desp

local-aiarxiv-cs-ai
11 May 2026
Safety

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

DGX agent

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

safetyarxiv-cs-ai
11 May 2026
Safety

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

DGX agent

arXiv:2605.07316v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning

safetyarxiv-cs-ai
11 May 2026
Research

InsHuman: Towards Natural and Identity-Preserving Human Insertion

DGX agent

arXiv:2605.07402v1 Announce Type: new Abstract: Human insertion aims to naturally place specific individuals into a target background. Although existing image editing models may have such ability, the

researcharxiv-cs-cv
11 May 2026
Safety

InterCoG: Towards Spatially Precise Image Editing with Interleaved Chain-of-Grounding Reasoning

DGX agent

arXiv:2603.01586v3 Announce Type: replace Abstract: Emerging unified editing models have demonstrated strong capabilities in general object editing tasks. However, it remains a significant challenge t

safetyarxiv-cs-cv
11 May 2026
Safety

KL for a KL: On-Policy Distillation with Control Variate Baseline

DGX agent

arXiv:2605.07865v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) has emerged as a dominant post-training paradigm for large language models, especially for reasoning domains. However, OP

safetyarxiv-cs-ai
11 May 2026
Research

Koopman Autoencoders with Continuous-Time Latent Dynamics for Fluid Dynamics Forecasting

DGX agent

arXiv:2602.02832v3 Announce Type: replace Abstract: Forecasting physical systems over long horizons from irregularly sampled observations demands models that are stable, computationally efficient, and

researcharxiv-cs-lg
11 May 2026
Tutorials

LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction

DGX agent

arXiv:2605.06676v1 Announce Type: cross Abstract: Long-context inference in Large Language Models (LLMs) is bottlenecked by the linear growth of Key-Value (KV) cache memory. Existing KV cache compress

tutorialsarxiv-cs-cl
11 May 2026
Agents

LLM-Guided Open Hypothesis Learning from Autonomous Scanning Probe Microscopy Experiments

DGX agent

arXiv:2605.06839v1 Announce Type: cross Abstract: Autonomous experimentation has transformed microscopy and materials discovery by enabling closed-loop optimization including imaging and spectroscopy

agentsarxiv-cs-ai
11 May 2026
Applications

LLM hallucinations in the wild: Large-scale evidence from non-existent citations

DGX agent

arXiv:2605.07723v1 Announce Type: cross Abstract: Large language models (LLMs) are known to generate plausible but false information across a wide range of contexts, yet the real-world magnitude and c

applicationsarxiv-cs-ai
11 May 2026
Local Ai

LoHGNet: Infrared Small Target Detection through Lorentz Geometric Encoding with High-Order Relation Learning

DGX agent

arXiv:2605.07213v1 Announce Type: new Abstract: Infrared small target detection (IRSTD) remains challenging due to the scarcity of useful target cues and the presence of severe background clutter. Mos

local-aiarxiv-cs-cv
11 May 2026
Research

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference

DGX agent

arXiv:2605.05225v2 Announce Type: replace-cross Abstract: Mixture-of-Experts Multimodal Large Language Models (MoE MLLMs) suffer from a significant efficiency bottleneck during Expert Parallelism (EP)

researcharxiv-cs-ai
11 May 2026
Agents

MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.06623v1 Announce Type: cross Abstract: Large language model (LLM)-based Multi-agent systems (MAS) have shown promise in tackling complex collaborative tasks, where agents are typically orch

agentsarxiv-cs-lg
11 May 2026
Agents

MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments

DGX agent

arXiv:2605.07058v1 Announce Type: cross Abstract: Real-world clinical diagnosis is a complex process in which the doctor is required to obtain information from both interaction with the patient and co

agentsarxiv-cs-ai
11 May 2026
Applications

MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes

DGX agent

arXiv:2605.06897v1 Announce Type: cross Abstract: The rise of Internet of Things (IoT) devices in the physical world necessitates voice-based interfaces capable of handling complex user experiences. W

applicationsarxiv-cs-ai
11 May 2026
Safety

MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation

DGX agent

arXiv:2605.08050v1 Announce Type: new Abstract: Talking-head generation requires joint modeling of identity, head pose, facial expression, and mouth dynamics. Existing methods typically address only a

safetyarxiv-cs-cv
11 May 2026
Tutorials

Not All Tokens Learn Alike: Attention Entropy Reveals Heterogeneous Signals in RL Reasoning

DGX agent

arXiv:2605.07660v1 Announce Type: new Abstract: Reinforcement-learning-based post-training has become a key approach for improving the reasoning ability of large language models, but its token-level l

tutorialsarxiv-cs-cl
11 May 2026
Research

Not All Tokens Need 40 Steps: Heterogeneous Step Allocation in Diffusion Transformers for Efficient Video Generation

DGX agent

arXiv:2605.06892v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have achieved state-of-the-art video generation quality, but they incur immense computational cost because standard infere

researcharxiv-cs-cv
11 May 2026
Safety

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

DGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

safetyarxiv-cs-ai
11 May 2026
Industry

On theCUBE Pod: IBM Think goes AI first and Musk buries the hatchet with Anthropic

DGX agent

IBM Think saw Big Blue grab its place in the artificial intelligence spotlight. In a week full of notable, and surprising, AI news, IBM Think saw CEO Arvind Krishna emphasize AI as the operating model

industrysiliconangle
11 May 2026
Applications

OneViewAll: Semantic Prior Guided One-View 6D Pose Estimation for Novel Objects

DGX agent

arXiv:2605.07023v1 Announce Type: new Abstract: In many practical 6D object pose estimation scenarios, we often have access to only a single real-world RGB-D reference view per object, typically witho

applicationsarxiv-cs-cv
11 May 2026
Safety

Online Allocation with Unknown Shared Supply

DGX agent

arXiv:2605.07080v1 Announce Type: new Abstract: Many real-world resource allocation systems, such as humanitarian logistics and vaccine distribution, must preposition limited supply across multiple lo

safetyarxiv-cs-ai
11 May 2026
Research

Online Localized Conformal Prediction

DGX agent

arXiv:2605.05497v2 Announce Type: replace Abstract: Conformal prediction is a framework that provides valid uncertainty quantification for general models with exchangeable data. However, in the online

researcharxiv-cs-lg
11 May 2026
Safety

Physics-Based Benchmarking Metrics for Multimodal Synthetic Images

DGX agent

arXiv:2511.15204v3 Announce Type: replace-cross Abstract: Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accur

safetyarxiv-cs-ai
11 May 2026
Research

Physics-Informed Reduced-Order Operator Learning for Hyperelasticity in Continuum Micromechanics

DGX agent

arXiv:2605.07738v1 Announce Type: cross Abstract: Physics-informed operator learning is an attractive candidate for surrogate modeling of microstructures, especially in multiscale finite-element simul

researcharxiv-cs-lg
11 May 2026
Safety

POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles

DGX agent

arXiv:2605.07775v1 Announce Type: cross Abstract: Balancing exploration and exploitation is a core challenge in sequential decision-making and black-box optimization. We introduce POETS (extbf{Po}licy

safetyarxiv-cs-ai
11 May 2026
Safety

Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning

DGX agent

arXiv:2605.07804v1 Announce Type: cross Abstract: On-policy distillation (OPD) leverages dense teacher rewards to enhance reasoning models. However, scaling OPD to long-horizon tasks exposes a critica

safetyarxiv-cs-ai
11 May 2026
Research

Rebalancing gradient to improve self-supervised co-training of depth, odometry and optical flow predictions

DGX agent

arXiv:2605.07945v1 Announce Type: new Abstract: We present CoopNet, an approach that improves the cooperation of co-trained networks by dynamically adapting the apportionment of gradient, to ensure eq

researcharxiv-cs-cv
11 May 2026
Safety

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation

DGX agent

arXiv:2408.06747v4 Announce Type: replace Abstract: Recent works utilize CLIP to perform the challenging unsupervised semantic segmentation task where only images without annotations are available. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Repeated Deceptive Path Planning against Learnable Observer

DGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

safetyarxiv-cs-ai
11 May 2026
← Previous
1…976977978979980…1272
Next →