AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
11 May 2026

Every Feedforward Neural Network Definable in an o-Minimal Structure Has Finite Sample Complexity

TutorialsDGX agent

arXiv:2605.07097v1 Announce Type: cross Abstract: We show that, in a precise sense, a broad class of feedforward neural networks learn (have finite sample complexity) in the PAC model: every fixed fin

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

SafetyDGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

Exposing and Mitigating Temporal Attack in Deepfake Video Detection

ResearchDGX agent

arXiv:2605.07398v1 Announce Type: cross Abstract: While spatiotemporal deepfake detectors achieve high AUC, our experiments reveal their susceptibility to evasion attacks. These models tend to overfit

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

SafetyDGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

From Canopy to Collision: A Hybrid Predictive Framework for Identifying Risk Factors in Tree-Involved Traffic Crashes

ResearchDGX agent

arXiv:2605.06684v1 Announce Type: new Abstract: Tree-involved crashes represent a critical subset of run-off-road (ROR) collisions, often resulting in fatal or severe injuries due to high-energy impac

From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms

AgentsDGX agent

arXiv:2605.06716v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents have fundamentally reshaped artificial intelligence by integrating external tools and planning capabilities. Whi

From Time Series Analysis to Question Answering: A Survey in the LLM Era

SafetyDGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

Geometric Analysis of Neural Regression Collapse via Intrinsic Dimension

ApplicationsDGX agent

arXiv:2510.01105v2 Announce Type: replace Abstract: Neural multivariate regression underpins a wide range of domains, including control, robotics, and finance, yet the geometry of its learned represen

GRAPE: Let GRPO Supervise Query Rewriting by Ranking for Retrieval

SafetyDGX agent

arXiv:2509.23370v2 Announce Type: replace Abstract: The CLIP model has established itself as a cornerstone of large-scale retrieval systems. However, its performance often degrades under distributiona

GraphDC: A Divide-and-Conquer Multi-Agent System for Scalable Graph Algorithm Reasoning

AgentsDGX agent

arXiv:2605.06671v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong potential for many mathematical problems. However, their performance on graph algorithmic tasks is

GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasks

ResearchDGX agent

arXiv:2605.07454v1 Announce Type: new Abstract: In-context learning enables large language models to adapt to new tasks, but their performance is highly sensitive to the selected examples. Finding eff

Here's the full TIL https://til.simonwillison.net/llms/llm-shebang

ToolsDGX agent

Simon Willison shares a technique for using Large Language Models directly from the command line using a shebang (#!) syntax, allowing scripts to be executed with LLM processing without explicit comma

HMACE: Heterogeneous Multi-Agent Collaborative Evolution for Combinatorial Optimization

Local AiDGX agent

arXiv:2605.07214v1 Announce Type: new Abstract: Large Language Models have recently emerged as a promising paradigm for automated heuristic design for NP-hard combinatorial optimization problems. Desp

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

SafetyDGX agent

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

SafetyDGX agent

arXiv:2605.07316v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning

InsHuman: Towards Natural and Identity-Preserving Human Insertion

ResearchDGX agent

arXiv:2605.07402v1 Announce Type: new Abstract: Human insertion aims to naturally place specific individuals into a target background. Although existing image editing models may have such ability, the

InterCoG: Towards Spatially Precise Image Editing with Interleaved Chain-of-Grounding Reasoning

SafetyDGX agent

arXiv:2603.01586v3 Announce Type: replace Abstract: Emerging unified editing models have demonstrated strong capabilities in general object editing tasks. However, it remains a significant challenge t

KL for a KL: On-Policy Distillation with Control Variate Baseline

SafetyDGX agent

arXiv:2605.07865v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) has emerged as a dominant post-training paradigm for large language models, especially for reasoning domains. However, OP

Koopman Autoencoders with Continuous-Time Latent Dynamics for Fluid Dynamics Forecasting

ResearchDGX agent

arXiv:2602.02832v3 Announce Type: replace Abstract: Forecasting physical systems over long horizons from irregularly sampled observations demands models that are stable, computationally efficient, and

LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction

TutorialsDGX agent

arXiv:2605.06676v1 Announce Type: cross Abstract: Long-context inference in Large Language Models (LLMs) is bottlenecked by the linear growth of Key-Value (KV) cache memory. Existing KV cache compress

LLM-Guided Open Hypothesis Learning from Autonomous Scanning Probe Microscopy Experiments

AgentsDGX agent

arXiv:2605.06839v1 Announce Type: cross Abstract: Autonomous experimentation has transformed microscopy and materials discovery by enabling closed-loop optimization including imaging and spectroscopy

LLM hallucinations in the wild: Large-scale evidence from non-existent citations

ApplicationsDGX agent

arXiv:2605.07723v1 Announce Type: cross Abstract: Large language models (LLMs) are known to generate plausible but false information across a wide range of contexts, yet the real-world magnitude and c

LoHGNet: Infrared Small Target Detection through Lorentz Geometric Encoding with High-Order Relation Learning

Local AiDGX agent

arXiv:2605.07213v1 Announce Type: new Abstract: Infrared small target detection (IRSTD) remains challenging due to the scarcity of useful target cues and the presence of severe background clutter. Mos

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference

ResearchDGX agent

arXiv:2605.05225v2 Announce Type: replace-cross Abstract: Mixture-of-Experts Multimodal Large Language Models (MoE MLLMs) suffer from a significant efficiency bottleneck during Expert Parallelism (EP)

MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.06623v1 Announce Type: cross Abstract: Large language model (LLM)-based Multi-agent systems (MAS) have shown promise in tackling complex collaborative tasks, where agents are typically orch

MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments

AgentsDGX agent

arXiv:2605.07058v1 Announce Type: cross Abstract: Real-world clinical diagnosis is a complex process in which the doctor is required to obtain information from both interaction with the patient and co

MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes

ApplicationsDGX agent

arXiv:2605.06897v1 Announce Type: cross Abstract: The rise of Internet of Things (IoT) devices in the physical world necessitates voice-based interfaces capable of handling complex user experiences. W

MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation

SafetyDGX agent

arXiv:2605.08050v1 Announce Type: new Abstract: Talking-head generation requires joint modeling of identity, head pose, facial expression, and mouth dynamics. Existing methods typically address only a

Not All Tokens Learn Alike: Attention Entropy Reveals Heterogeneous Signals in RL Reasoning

TutorialsDGX agent

arXiv:2605.07660v1 Announce Type: new Abstract: Reinforcement-learning-based post-training has become a key approach for improving the reasoning ability of large language models, but its token-level l

Not All Tokens Need 40 Steps: Heterogeneous Step Allocation in Diffusion Transformers for Efficient Video Generation

ResearchDGX agent

arXiv:2605.06892v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have achieved state-of-the-art video generation quality, but they incur immense computational cost because standard infere

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

SafetyDGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

On theCUBE Pod: IBM Think goes AI first and Musk buries the hatchet with Anthropic

IndustryDGX agent

IBM Think saw Big Blue grab its place in the artificial intelligence spotlight. In a week full of notable, and surprising, AI news, IBM Think saw CEO Arvind Krishna emphasize AI as the operating model

OneViewAll: Semantic Prior Guided One-View 6D Pose Estimation for Novel Objects

ApplicationsDGX agent

arXiv:2605.07023v1 Announce Type: new Abstract: In many practical 6D object pose estimation scenarios, we often have access to only a single real-world RGB-D reference view per object, typically witho

Online Allocation with Unknown Shared Supply

SafetyDGX agent

arXiv:2605.07080v1 Announce Type: new Abstract: Many real-world resource allocation systems, such as humanitarian logistics and vaccine distribution, must preposition limited supply across multiple lo

Online Localized Conformal Prediction

ResearchDGX agent

arXiv:2605.05497v2 Announce Type: replace Abstract: Conformal prediction is a framework that provides valid uncertainty quantification for general models with exchangeable data. However, in the online

Physics-Based Benchmarking Metrics for Multimodal Synthetic Images

SafetyDGX agent

arXiv:2511.15204v3 Announce Type: replace-cross Abstract: Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accur

Physics-Informed Reduced-Order Operator Learning for Hyperelasticity in Continuum Micromechanics

ResearchDGX agent

arXiv:2605.07738v1 Announce Type: cross Abstract: Physics-informed operator learning is an attractive candidate for surrogate modeling of microstructures, especially in multiscale finite-element simul

POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles

SafetyDGX agent

arXiv:2605.07775v1 Announce Type: cross Abstract: Balancing exploration and exploitation is a core challenge in sequential decision-making and black-box optimization. We introduce POETS (extbf{Po}licy

Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning

SafetyDGX agent

arXiv:2605.07804v1 Announce Type: cross Abstract: On-policy distillation (OPD) leverages dense teacher rewards to enhance reasoning models. However, scaling OPD to long-horizon tasks exposes a critica

Rebalancing gradient to improve self-supervised co-training of depth, odometry and optical flow predictions

ResearchDGX agent

arXiv:2605.07945v1 Announce Type: new Abstract: We present CoopNet, an approach that improves the cooperation of co-trained networks by dynamically adapting the apportionment of gradient, to ensure eq

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation

SafetyDGX agent

arXiv:2408.06747v4 Announce Type: replace Abstract: Recent works utilize CLIP to perform the challenging unsupervised semantic segmentation task where only images without annotations are available. Ho

Repeated Deceptive Path Planning against Learnable Observer

SafetyDGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

Retrieve, Integrate, and Synthesize: Spatial-Semantic Grounded Latent Visual Reasoning

ResearchDGX agent

arXiv:2605.07106v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made remarkable progress on vision-language reasoning, yet most methods still compress visual evidence int

Rubric-based On-policy Distillation

SafetyDGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged …

SafetyDGX agent

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged use of copyrighted works in training AI models https://go.na

See Tomorrow, Act Today: Foresight-Driven Autonomous Driving

AgentsDGX agent

arXiv:2605.07195v1 Announce Type: new Abstract: Current end-to-end autonomous driving planners are fundamentally reactive: they condition on historical and present observations to predict future actio

Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding

HardwareDGX agent

arXiv:2605.07897v1 Announce Type: cross Abstract: Online streaming video understanding requires models to process continuous visual inputs and respond to user queries in real time, where the unbounded

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

SafetyDGX agent

arXiv:2605.06130v2 Announce Type: replace Abstract: A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupl

Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation

SafetyDGX agent

arXiv:2605.07950v1 Announce Type: new Abstract: We study Slowly Annealed Langevin Dynamics (SALD), a sampler for tracking a path of moving target distributions and approximating the terminal target th

SocialReasoning-Bench: Measuring whether AI agents act in users’ best interests

ResearchDGX agent

Using SocialReasoning Bench, we observed a stable pattern across models—agents execute competently, but fail to consistently improve the user’s position, even with explicit instructions to optimize fo

SOCKET: SOft Collision Kernel EsTimator for Sparse Attention

HardwareDGX agent

arXiv:2602.06283v2 Announce Type: replace Abstract: Exploiting sparsity during long-context inference is key to scaling large language models, as attention dominates the cost of autoregressive decodin

Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs

SafetyDGX agent

arXiv:2605.07447v1 Announce Type: cross Abstract: Vision-language models (VLMs) have advanced rapidly and are increasingly deployed in real-world applications, especially with the rise of agent-based

SparseRL-Sync: Lossless Weight Synchronization with ~100x Less Communication

SafetyDGX agent

arXiv:2605.07330v1 Announce Type: cross Abstract: In large-scale reinforcement learning (RL) systems with decoupled Trainer-Rollout execution, the Trainer must regularly synchronize policy weights to

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

SafetyDGX agent

arXiv:2605.07462v1 Announce Type: cross Abstract: Moltbook is a Reddit-like platform where OpenClaw agents post, comment, and vote at scale - a so far unprecedented incident that comes with serious sa

TopoPrune: Robust Data Pruning via Unified Latent Space Topology

ResearchDGX agent

arXiv:2602.02739v2 Announce Type: replace-cross Abstract: Geometric data pruning methods, while practical for leveraging pretrained models, are fundamentally unstable. Their reliance on extrinsic geom

Towards an Inferentialist Account of Information Through Proof-theoretic Semantics

ResearchDGX agent

arXiv:2605.05368v2 Announce Type: replace-cross Abstract: Information is one of the most widely-discussed concepts of the current era. However, a great deal of insightful work notwithstanding, it is y

Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors

ApplicationsDGX agent

arXiv:2505.11325v4 Announce Type: replace-cross Abstract: Prior-data fitted networks (PFNs) have emerged as promising foundation models for prediction from tabular datasets, achieving state-of-the-art

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

SafetyDGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

VIMCAN: Visual-Inertial 3D Human Pose Estimation with Hybrid Mamba-Cross-Attention Network

ResearchDGX agent

arXiv:2605.07552v1 Announce Type: new Abstract: The rapid advances in deep learning have significantly enhanced the accuracy of multimodal 3D human pose estimation (HPE). However, the state-of-the-art

VNN-LIB 2.0: Rigorous Foundations for Neural Network Verification

ResearchDGX agent

arXiv:2605.07451v1 Announce Type: new Abstract: Neural network verification is an active and rapidly maturing research area, with a growing ecosystem of solvers and tools. The VNN-LIB standard was int

← Previous
1…780781782783784…1017
Next →