AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
29 May 2026

LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents

Model ReleasesDGX agent

arXiv:2605.29559v1 Announce Type: new Abstract: Mastering terminal environments requires language agents capable of multi-step planning, feedback-grounded execution, and dynamic state adaptation. Howe

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

Model ReleasesDGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

LoCoT2V-Bench: Benchmarking Long-Form and Complex Text-to-Video Generation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.26412v3 Announce Type: replace-cross Abstract: Recent advances in text-to-video generation have achieved impressive performance on short clips, yet evaluating long-form generation under com

MIRAGE: Adaptive Multimodal Gating for Whole-Brain fMRI Encoding

ResearchDGX agent

arXiv:2605.29850v1 Announce Type: new Abstract: Recent progress in task-optimized neural networks has established encoding models as a powerful tool for predicting brain responses to naturalistic stim

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing

Model ReleasesDGX agent

arXiv:2605.22100v2 Announce Type: replace Abstract: Document parsing converts visually rich documents into machine-readable structured representations, forming a crucial foundation for information sys

MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization

Model ReleasesDGX agent

arXiv:2605.29951v1 Announce Type: new Abstract: Understanding how harm emerges from interaction between otherwise benign image-text pairs requires intent-aware cross-modal reasoning beyond surface-lev

One Click per Cell Type Suffices: Training-free Group Interaction for Cell Instance Segmentation

ResearchDGX agent

arXiv:2605.29429v1 Announce Type: new Abstract: Cell instance segmentation models trained on cell-specific datasets suffer severe performance drops on out-of-distribution cell types, while interactive

OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

Model ReleasesDGX agent

arXiv:2605.29253v1 Announce Type: new Abstract: Task success can hide process anomalies in real-world agent executions. An agent may pass the final task oracle while still accumulating unresolved ambi

OVA-IB: One vs All Information Bottleneck for Multi-Modal Alignment

Model ReleasesDGX agent

arXiv:2605.29900v1 Announce Type: new Abstract: Contrastive learning is effective for aligning paired views or modalities, but alignment beyond two modalities remains non-trivial and comparatively und

Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring

Model ReleasesDGX agent

arXiv:2605.29852v1 Announce Type: new Abstract: Histological scoring is essential for diagnosing Non-Alcoholic Fatty Liver Disease (NAFLD), yet its automation remains challenging due to the high annot

Persona Conditioning of Brand Recommendations in Retrieval-Augmented Commercial Chat: A Prominence-Stratified Cross-Provider Audit

ApplicationsDGX agent

arXiv:2605.30207v1 Announce Type: new Abstract: The same prompt -- 'best CRM software' -- reaches AI assistants from buyers in widely different contexts: a solo founder, an enterprise VP, a UK SMB own

PhAIL: A Real-Robot VLA Benchmark and Distributional Methodology

Model ReleasesDGX agent

arXiv:2605.29710v1 Announce Type: new Abstract: Real-world evaluation of vision-language-action (VLA) policies still rests on binary success rate at a fixed timeout with N le 25 rollouts per condition

PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

Model ReleasesDGX agent

arXiv:2605.30094v1 Announce Type: new Abstract: Poker is a landmark challenge for artificial intelligence. The dominant approach relies on equilibrium solvers built on counterfactual regret minimizati

Pre-Registering the Detectable Effect: A Paired-MDE Budget for 4-bit Quantization Benchmarks, with a Pilot Audit

Model ReleasesDGX agent

arXiv:2605.28873v1 Announce Type: new Abstract: This is a planning-method note with an unpaired pilot audit. We adapt the classical paired-binary sample-size calculation (Miettinen, 1968) to quantizat

Prescribe-then-Select: Adaptive Policy Selection for Contextual Stochastic Optimization

Model ReleasesDGX agent

arXiv:2509.08194v2 Announce Type: replace Abstract: We address the problem of policy selection in contextual stochastic optimization (CSO), where covariates are available as contextual information and

PTCG-Bench: Can LLM Agents Master Pokemon Trading Card Game?

Model ReleasesDGX agent

arXiv:2605.29653v1 Announce Type: new Abstract: Given a strategically complex board game, human players can quickly learn to devise strategies after playing a few rounds. Autonomous agents require sim

Recurrent Structural Policy Gradient for Partially Observable Mean Field Games

SafetyDGX agent

arXiv:2602.20141v2 Announce Type: replace Abstract: Mean Field Games (MFGs) provide a principled framework for modelling interactions in large population systems. However, algorithmic progress has bee

Reinforcing Few-step Generators via Reward-Tilted Distribution Matching

SafetyDGX agent

arXiv:2605.26108v2 Announce Type: replace Abstract: Recent advances in few-step diffusion distillation have enabled efficient image generation, yet aligning these models with human preferences remains

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

Model ReleasesDGX agent

arXiv:2605.29796v1 Announce Type: new Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these syste

SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation

Model ReleasesDGX agent

arXiv:2507.09266v2 Announce Type: replace Abstract: Gloss-free Sign Language Translation (SLT) has advanced rapidly, achieving strong performances without relying on gloss annotations. However, these

Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance

Model ReleasesDGX agent

arXiv:2605.30056v1 Announce Type: cross Abstract: Recent advances in reinforcement learning (RL) have achieved great successes by leveraging the multimodality and exploration capability of diffusion p

SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification

Model ReleasesDGX agent

arXiv:2603.12588v2 Announce Type: replace Abstract: Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery is fundamentally challenged by the severe radio

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

Model ReleasesDGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

Statistical Embeddings for Similarity, Retrieval, and Interpretable Alignment of Numeric Tabular Datasets

SafetyDGX agent

arXiv:2605.30289v1 Announce Type: new Abstract: Numeric tabular datasets are the dominant data format in scientific practice, yet large language models lack native mechanisms for representing numeric

Streaming Drag-Oriented Interactive Video Manipulation: Drag Anything, Anytime!

ResearchDGX agent

arXiv:2510.03550v4 Announce Type: replace Abstract: Achieving streaming, fine-grained control over the outputs of autoregressive video diffusion models remains challenging, making it difficult to ensu

Striding Across Reynolds Numbers: Representation Geometry in Neural PDE Generalisation

Model ReleasesDGX agent

arXiv:2605.30112v1 Announce Type: new Abstract: Cross-Reynolds generalisation in neural PDE solvers remains poorly characterised. On the canonical forced 2D Navier-Stokes benchmark, a trained Fourier

TAE: Target-aware enhancer for nighttime UAV tracking

Model ReleasesDGX agent

arXiv:2605.29558v1 Announce Type: new Abstract: Severe image degradation under low-light nighttime conditions constitutes a core bottleneck preventing all-day applications for UAV-based single object

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

Model ReleasesDGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

Text-Preserving Lossy Text Compression: A Study of Strategic Deletion and LLM Reconstruction

Model ReleasesDGX agent

arXiv:2605.29000v1 Announce Type: new Abstract: Traditional lossless text compression preserves every byte, but its gains on natural language are often modest in realistic operating regimes. We study

The Little Book of Generative AI Foundations: An Intuitive Mathematical Primer

ResearchDGX agent

arXiv:2605.29713v1 Announce Type: cross Abstract: This book provides a compact, derivation-oriented introduction to the mathematical foundations of modern generative artificial intelligence. Rather th

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There …

Model ReleasesDGX agent

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There are still so many leaders that have never seen an agent run

Towards Consistent Video Geometry Estimation

ResearchDGX agent

arXiv:2605.30060v1 Announce Type: new Abstract: This work presents ViGeo, a feed-forward foundation model for recovering spatially dense and temporally consistent geometry from video sequences. Built

UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering

ResearchDGX agent

arXiv:2605.30076v1 Announce Type: new Abstract: Activation-based control steers large language models (LLMs) by intervening on their internal representations during inference, and has emerged as an ef

ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement

ResearchDGX agent

arXiv:2605.29302v1 Announce Type: new Abstract: The digital media landscape has seen a pervasive shift toward short-form video advertising on TV, social media and e-commerce platforms. The present stu

Video Individual Counting and Tracking from Moving Drones: A Benchmark and Methods

Model ReleasesDGX agent

arXiv:2601.12500v2 Announce Type: replace Abstract: Counting and tracking dense crowds in large-scale scenes is a highly practical yet challenging problem. Existing methods mostly rely on fixed-camera

VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents

Model ReleasesDGX agent

arXiv:2605.30256v1 Announce Type: cross Abstract: Natural human conversation is full-duplex and audio-visual: people simultaneously speak and listen while continuously interpreting and producing nonve

Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks

ResearchDGX agent

arXiv:2605.30167v1 Announce Type: cross Abstract: Predicting a complete spatially correlated field from sparse observations is a fundamental challenge in spatial statistics and environmental modelling

WASHH: An Anchor-Aware Whale-Guided Selection Hyper-Heuristic for Continuous Optimization and SVC Configuration

Model ReleasesDGX agent

arXiv:2605.28844v1 Announce Type: cross Abstract: Learning-assisted algorithm design often has to make reliable search decisions under small evaluation budgets, where committing to a single metaheuris

What drives performance in molecular MPNNs? An operator-level factorial benchmark

Model ReleasesDGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

Local AiDGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

28 May 2026

A Paired Testing Protocol for Batch-Conditioned Refusal Robustness in LLM Serving

Local AiDGX agent

arXiv:2605.27763v1 Announce Type: new Abstract: Safety evaluations of language models often treat serving configuration as fixed background infrastructure, but batch condition is an untested treatment

A Structural Theory of Position Bias in Transformers

SafetyDGX agent

arXiv:2602.16837v2 Announce Type: replace Abstract: Transformer models systematically favor certain token positions, yet the architectural origins of this position bias remain poorly understood. This

Accelerating Diffusion Sampling via Exploiting Local Transition Coherence

ResearchDGX agent

arXiv:2503.09675v3 Announce Type: replace Abstract: Text-based diffusion models have made significant breakthroughs in generating high-quality images and videos from textual descriptions. However, the

Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning

Model ReleasesDGX agent

arXiv:2605.28192v1 Announce Type: new Abstract: Multi-hop audio-visual reasoning remains challenging for Omni-LLMs, as relevant evidence is often sparse, temporally dispersed, and distributed across b

Also out today: You can now directly configure the effort level and adaptive thinking in Code (/effort) and Cowork! Effort allows you to tun…

Model ReleasesDGX agent

Also out today: You can now directly configure the effort level and adaptive thinking in Code (/effort) and Cowork! Effort allows you to tune Claude's intelligence vs token spend, trading off capabili

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

AgentsDGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

AREA: Attribute Extraction and Aggregation for CLIP-Based Class-Incremental Learning

SafetyDGX agent

arXiv:2605.28809v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) is important in building real-world learning systems. In CLIP-based CIL, the model performs classification by comparing

AssertLLM2: A Comprehensive LLM Benchmark for Assertion Generation from Design Specifications

Model ReleasesDGX agent

arXiv:2605.27472v1 Announce Type: cross Abstract: Assertion-based verification (ABV) is a cornerstone of modern hardware design, yet manually translating design intent into formal SystemVerilog Assert

Auditing Stance Asymmetry in Generative Explanations

SafetyDGX agent

arXiv:2605.27988v1 Announce Type: new Abstract: Bias evaluation for language models has made substantial progress on bounded comparisons, such as overt derogation, stereotype association, or label-sen

Bayesian Optimization Parameter Tuning Framework for a Lyapunov Based Path Following Controller

Model ReleasesDGX agent

arXiv:2512.12649v2 Announce Type: replace Abstract: Parameter tuning in real-world experiments is constrained by the limited evaluation budget available on hardware. The path-following controller stud

Benchmarking Fairness in Spiking Neural Networks: Data Bias, Spurious Features, and Hardware Effects

Model ReleasesDGX agent

arXiv:2605.27407v1 Announce Type: cross Abstract: Evaluating fairness in Spiking Neural Networks (SNNs) demands rigorous benchmarks that reflect real-world complexities, yet existing assessments remai

BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law

Model ReleasesDGX agent

arXiv:2605.28183v1 Announce Type: cross Abstract: We introduce the BenGER (Benchmark for German Law) dataset for evaluating LLM systems on subsumption-based legal reasoning in German law. The BenGER d

Can Hallucinations Be Useful? Solving Multi-Hop Questions With SLMs By Chaining System-I/II Reasoning

ResearchDGX agent

arXiv:2605.27596v1 Announce Type: new Abstract: Recently, there has been increased interest in Small Language Models (SLMs), which are fast, show good performance, and have lower hardware demands than

Category-Level 3D Correspondence in Camera Space via Morphable Object Priors

Model ReleasesDGX agent

arXiv:2605.28257v1 Announce Type: new Abstract: Understanding 3D objects from images is fundamental to robotics and AR/VR applications. While recent work has made progress in category-level pose estim

Causal Direct Preference Optimization for Distributionally Robust Generative Recommendation

SafetyDGX agent

arXiv:2603.22335v2 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) guides large language models (LLMs) to generate recommendations aligned with user historical behavior dis

Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation

Model ReleasesDGX agent

arXiv:2501.04144v3 Announce Type: replace Abstract: Understanding and generating the fine-grained structure of objects -- such as birds with species-specific beaks, wings, and tails -- is a long-stand

Chrome Enterprise rolls out AI agents and automation to streamline security management

Model ReleasesDGX agent

Google LLC today launched new enhancements to Chrome Enterprise, the company’s enterprise version of its Chrome browser, designed to provide greater administrative and security control to information

Claude Opus 4.8 is now available in Windsurf and Devin CLI

Model ReleasesDGX agent

Claude Opus 4.8 has been made available for use in both Windsurf (Cognition AI's code editor) and Devin CLI (their command-line interface). This release expands access to Anthropic's latest Claude mod

ClinConsensus: A Physician-Calibrated Benchmark for Evaluating Clinical Rubric Coverage in Chinese Medical LLMs

Model ReleasesDGX agent

arXiv:2603.02097v5 Announce Type: replace Abstract: Open-ended medical LLM evaluation remains weakly grounded in physician-calibrated coverage of clinically relevant response criteria, especially in l

ClothTransformer: Unified Latent-Space Transformers for Scalable Cloth Simulation

ResearchDGX agent

arXiv:2605.27852v1 Announce Type: cross Abstract: Unified and scalable Transformers have recently achieved remarkable success in modeling diverse phenomena traditionally associated with computer graph

← Previous
1…545546547548549…1061
Next →