AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
Model Releases

TTCD:Transformer Integrated Temporal Causal Discovery from Non-Stationary Time Series Data

DGX agent

arXiv:2605.08111v1 Announce Type: cross Abstract: The widespread availability of complex time series data in various domains such as environmental science, epidemiology, and economics demands robust c

model-releasesarxiv-cs-ai
12 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Two Ways to De-Bias an LLM-as-a-Judge: A Continuous-Score Comparison of Hierarchical Bayesian Calibration and Neural-ODE Score Transport

DGX agent

arXiv:2605.09227v1 Announce Type: new Abstract: [Abridged] Using a Large Language Model (LLM) as an automatic rater (LLM-as-a-judge) is cheap but potentially biased: some judges run lenient, others st

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

UFO: A Unified Flow-Oriented Framework for Robust Continual Graph Learning

DGX agent

arXiv:2605.09862v1 Announce Type: cross Abstract: Graph learning research has increasingly shifted toward continual graph learning (CGL), which better reflects real-world scenarios where graphs evolve

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

UMEDA: Unified Multi-modal Efficient Data Fusion for Privacy-Preserving Graph Federated Learning via Spectral-Gated Attention and Diffusion-Based Operator Alignment

DGX agent

arXiv:2605.08288v1 Announce Type: cross Abstract: Device-free localization trains models from heterogeneous wireless and visual sensors (e.g., Wi-Fi, LiDAR) distributed across edge devices. Federated

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Understanding Asynchronous Inference Methods for Vision-Language-Action Models

DGX agent

arXiv:2605.08168v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising path to generalist robot control, but their inference latency causes observation staleness when

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Uni-Synergy: Bridging Understanding and Generation for Personalized Reasoning via Co-operative Reinforcement Learning

DGX agent

arXiv:2605.10445v1 Announce Type: new Abstract: Unified Multimodal Models (UMMs) excel in general tasks but struggle to bridge the gap between personalized understanding and generation. Prior works la

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning

DGX agent

arXiv:2605.08911v1 Announce Type: new Abstract: Autonomous vehicles need to perceive not only physical elements in the driving scene, such as lane lines and traffic lights, but also logical elements l

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media

DGX agent

arXiv:2605.05831v2 Announce Type: replace Abstract: The communication of scientific knowledge has become increasingly multimodal, spanning text, visuals, and speech through materials such as research

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning

DGX agent

arXiv:2605.08709v1 Announce Type: new Abstract: Unified face attack detection (UAD) requires recognizing physical spoofing and digital forgery within a shared decision space, yet existing discriminati

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

DGX agent

arXiv:2605.10889v1 Announce Type: cross Abstract: On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this sign

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Unpredictability dissociates from structured control in language agents

DGX agent

arXiv:2605.09692v1 Announce Type: new Abstract: Unpredictable behavior is often taken as evidence of control, yet stochastic dispersion and structured action control need not coincide. This paper test

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Unveiling High-Probability Generalization in Decentralized SGD

DGX agent

arXiv:2605.10205v1 Announce Type: new Abstract: Decentralized stochastic gradient descent (D-SGD) is an efficient method for large-scale distributed learning. Existing generalization studies mainly ad

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Urban-ImageNet: A Large-Scale Multi-Modal Dataset and Evaluation Framework for Urban Space Perception

DGX agent

arXiv:2605.09936v1 Announce Type: new Abstract: We present Urban-ImageNet, a large-scale multi-modal dataset and evaluation benchmark for urban space perception from user-generated social media imager

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

UserGPT Technical Report

DGX agent

arXiv:2605.08766v1 Announce Type: cross Abstract: Personalized user understanding from large-scale digital traces remains a fundamental challenge. Traditional user profiling methods rely on discrimina

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

DGX agent

arXiv:2605.09769v1 Announce Type: new Abstract: This paper describes our system for classifying psychological defense mechanisms in emotional support dialogues using the Defense Mechanism Rating Scale

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

V4FinBench: Benchmarking Tabular Foundation Models, LLMs, and Standard Methods on Corporate Bankruptcy Prediction

DGX agent

arXiv:2605.10896v1 Announce Type: new Abstract: Corporate bankruptcy prediction is a high-stakes financial task characterized by severe class imbalance and multi-horizon forecasting demands. Public da

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Valid Best-Model Identification for LLM Evaluation via Low-Rank Factorization

DGX agent

arXiv:2605.10405v1 Announce Type: new Abstract: Selecting the best large language model (LLM) for a fixed benchmark is often expensive, since exhaustive evaluation requires running every model on ever

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Vapi nabs $50M to make voice AI more human

DGX agent

Voice artificial intelligence startup Vapi Inc. said today it has raised 50 million in new funding to change the way people talk to computers, experience phone calls and interact with customer support

model-releasessiliconangle
12 May 2026
Model Releases

VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models

DGX agent

arXiv:2603.18113v2 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly shape content generation, interaction, and decision-making across the Web, aligning them with hum

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models

DGX agent

arXiv:2605.10485v1 Announce Type: new Abstract: Precise spatial reasoning is fundamental to robotic manipulation, yet the visual backbones of current vision-language-action (VLA) models are predominan

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation

DGX agent

arXiv:2605.08553v1 Announce Type: cross Abstract: Large language models can generate useful code from natural language, but their outputs come without correctness guarantees. Verifiable code generatio

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VFM-SDM: A vision foundation model-based framework for training-free, marker-free, and calibration-free structural displacement measurement

DGX agent

arXiv:2605.09677v1 Announce Type: new Abstract: Reliable displacement measurement is fundamental for structural health monitoring and digital engineering workflows, as it provides direct structural re

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

VidNum-1.4K: A Comprehensive Benchmark for Video-based Numerical Reasoning

DGX agent

arXiv:2604.03701v2 Announce Type: replace Abstract: Video-based numerical reasoning provides a premier arena for testing whether Vision-Language Models (VLMs) truly 'understand' real-world dynamics, a

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

VISOR: A Vision-Language Model-based Test Oracle for Testing Robot

DGX agent

arXiv:2605.10408v1 Announce Type: cross Abstract: Testing robots requires assessing whether they perform their intended tasks correctly, dependably, and with high quality, a challenge known as the tes

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

VISTA: A Benchmark for Real-Time Video Streaming under Network Impairments in Surgical Teleoperation

DGX agent

arXiv:2605.08886v1 Announce Type: cross Abstract: Real-time video streaming is crucial in surgical teleoperation, yet reproducible evaluation under realistic network impairments remains limited. This

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Visual-ERM: Reward Modeling for Visual Equivalence

DGX agent

arXiv:2603.13224v2 Announce Type: replace-cross Abstract: Vision-to-code tasks require models to reconstruct structured visual inputs, such as charts, tables, and SVGs, into executable or structured r

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VLADriver-RAG: Retrieval-Augmented Vision-Language-Action Models for Autonomous Driving

DGX agent

arXiv:2605.08133v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for end-to-end autonomous driving, yet their reliance on implicit parametric

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VORT: Adaptive Power-Law Memory for NLP Transformers

DGX agent

arXiv:2605.08966v1 Announce Type: new Abstract: Standard Transformers impose near-exponential decay on the influence of distant tokens, conflicting with the power-law structure of long-range dependenc

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

DGX agent

arXiv:2605.08146v1 Announce Type: cross Abstract: Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domai

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

WATCH: Wide-Area Archaeological Site Tracking for Change Detection

DGX agent

arXiv:2605.08160v1 Announce Type: cross Abstract: Monitoring archaeological sites at scale is vital for protecting cultural heritage, yet pinpointing when disturbances occur remains difficult because

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

We integrated FrontierCS into Harbor and are releasing a preview long-horizon agent leaderboard (up to 835 turns, ~200K output tokens) with …

DGX agent

We integrated FrontierCS into Harbor and are releasing a preview long-horizon agent leaderboard (up to 835 turns, ~200K output tokens) with Kimi K2.6 @Kimi_Moonshot (score 46.9) and Claude Code Opus 4

model-releaseskimi-moonshot--x
12 May 2026
Model Releases

Weight Pruning Amplifies Bias: A Multi-Method Study of Compressed LLMs for Edge AI

DGX agent

arXiv:2605.08137v1 Announce Type: cross Abstract: Weight pruning is widely advocated for deploying Large Language Models on resource-constrained IoT and edge devices, yet its impact on model fairness

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

What Parameter Golf taught us about AI-assisted research

DGX agent

Parameter Golf brought together 1,000+ participants and 2,000+ submissions to explore AI-assisted machine learning research, coding agents, quantization, and novel model design under strict constraint

model-releasesopenai
12 May 2026
Model Releases

What Will Happen Next: Large Models-Driven Deduction for Emergency Instances

DGX agent

arXiv:2605.08599v1 Announce Type: new Abstract: Traditional simulation methods reproduce occurred emergency instances through presetting to assist people in risk assessment and emergency decision-maki

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

What’s new in Microsoft Foundry | April 2026

DGX agent

April brings Foundry Local GA for local AI development, GPT-5.5 model support with Tier 5 and Tier 6 default quota in Microsoft Foundry, new tracing paths for Microsoft Agent Framework and hosted agen

model-releasesmicrosoft-foundry
12 May 2026
Model Releases

What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering

DGX agent

arXiv:2601.20164v2 Announce Type: replace-cross Abstract: Prior work suggests that language models, while trained on next token prediction, show implicit planning behavior: they may select the next to

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Adaptation Fails: A Gradient-Based Diagnosis of Collapsed Gating in Vision-Language Prompt Learning

DGX agent

arXiv:2605.09549v1 Announce Type: new Abstract: Adaptive prompting mechanisms have been proposed to enhance vision-language models by dynamically tailoring prompts to inputs. However, in frozen few-sh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning

DGX agent

arXiv:2605.09109v1 Announce Type: new Abstract: Many continuous-control problems ship with a competent but suboptimal controller (a tuned PID, a hand-designed gait). A growing family of methods uses s

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Attention Beats Fourier: Multi-Scale Transformers for PDE Solving on Irregular Domains

DGX agent

arXiv:2605.08318v1 Announce Type: cross Abstract: We study the problem of architecture selection for deep learning models trained to solve partial differential equations (PDEs), asking when transforme

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Does Non-Uniform Replay Matter in Reinforcement Learning?

DGX agent

arXiv:2605.10236v1 Announce Type: cross Abstract: Modern off-policy reinforcement learning algorithms often rely on simple uniform replay sampling and it remains unclear when and why non-uniform repla

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last com…

DGX agent

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative

model-releasesgary-marcus--x
12 May 2026
Model Releases

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

DGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

DGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Trust Imagination: Adaptive Action Execution for World Action Models

DGX agent

arXiv:2605.06222v2 Announce Type: replace-cross Abstract: World Action Models (WAMs) have recently emerged as a promising paradigm for robotic manipulation by jointly predicting future visual observat

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing

DGX agent

arXiv:2605.10544v1 Announce Type: new Abstract: Long-context adaptation is often viewed as window scaling, but this misses a token-level supervision mismatch: in packed training with document masking,

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

DGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Why Zeroth-Order Adaptation May Forget Less: A Randomized Shaping Theory

DGX agent

arXiv:2605.10658v1 Announce Type: new Abstract: Continual learning requires new-task adaptation without damaging previously acquired capabilities. Recent forward-pass and zeroth-order (ZO) results sho

model-releasesarxiv-cs-lg
12 May 2026
← Previous
1…336337338339340…471
Next →