AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Research

Are VLMs Seeing or Just Saying? Uncovering the Illusion of Visual Re-examination

DGX agent

arXiv:2605.15864v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) often produce self-reflective statements like 'let me check the figure again' during reasoning. Do such statements trigg

researcharxiv-cs-cl
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty

DGX agent

arXiv:2507.16806v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained via reinforcement learning (RL) to generate natural language 'reasoning chains', their performance impr

researcharxiv-cs-ai
18 May 2026
Applications

Calibrating LLMs with Semantic-level Reward

DGX agent

arXiv:2605.15588v1 Announce Type: new Abstract: As large language models (LLMs) are deployed in consequential settings such as medical question answering and legal reasoning, the ability to estimate w

applicationsarxiv-cs-cl
18 May 2026
Model Releases

Confirming Correct, Missing the Rest: LLM Tutoring Agents Struggle Where Feedback Matters Most

DGX agent

arXiv:2605.16207v1 Announce Type: new Abstract: Effective tutoring requires distinguishing optimal, valid but suboptimal, and incorrect student solutions, a distinction central to intelligent tutoring

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Continual Learning of Domain-Invariant Representations

DGX agent

arXiv:2605.15775v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially over multiple domains without forgetting previously learned knowledge. However, existing CL me

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Decentralized LoRA augmented transformer with multi-scale feature learning for secured eye diagnosis

DGX agent

arXiv:2505.06982v3 Announce Type: replace Abstract: Accurate and privacy-preserving diagnosis of ophthalmic diseases remains a critical challenge in medical imaging, particularly given the limitations

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection

DGX agent

arXiv:2605.15518v1 Announce Type: new Abstract: The effective detection and governance of Large Language Model (LLM) generated content has become increasingly critical due to the growing risk of misus

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation

DGX agent

arXiv:2605.16003v1 Announce Type: new Abstract: Autoregressive video diffusion models enable open-ended generation through local attention and KV caching. However, existing training-free long-video op

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

End-to-end plaque counting and virus titration from laboratory plate images with deep learning

DGX agent

arXiv:2605.16008v1 Announce Type: new Abstract: Plaque assays remain the gold standard readout of virus infectivity; however, plaque counting from plate images is labor-intensive and prone to inter-op

model-releasesarxiv-cs-cv
18 May 2026
Research

Enhancing Medical Image Segmentation via Heat Conduction Equation

DGX agent

arXiv:2511.03260v2 Announce Type: replace Abstract: Medical image segmentation models struggle to achieve efficient global context modeling and long-range dependency reasoning under practical computat

researcharxiv-cs-cv
18 May 2026
Model Releases

FormulaCode: Evaluating Agentic Optimization on Large Codebases

DGX agent

arXiv:2603.16011v2 Announce Type: replace-cross Abstract: Large language model (LLM) coding agents increasingly operate at the repository level, motivating benchmarks that evaluate their ability to op

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery

DGX agent

arXiv:2605.15412v1 Announce Type: cross Abstract: Modern quantitative trading increasingly relies on systematic models to extract predictive signals from large-scale financial data, where alpha factor

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GESD: Beyond Outcome-Oriented Fairness

DGX agent

arXiv:2605.15295v1 Announce Type: cross Abstract: Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predi

model-releasesarxiv-cs-ai
18 May 2026
Research

Highly Detailed and Generalizable Broadleaf Tree Crown Instance Segmentation from UAV Imagery

DGX agent

arXiv:2605.15673v1 Announce Type: cross Abstract: We present a highly detailed instance segmentation model for delineating individual tree crowns in natural broadleaf forests using aerial imagery acqu

researcharxiv-cs-cv
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Tutorials

Hypothesis-driven construction of mesoscopic dynamics

DGX agent

arXiv:2605.16211v1 Announce Type: new Abstract: Traditional scientific modeling typically begins with fixed, instance-wise effective equations and then carries out equation-specific analysis and compu

tutorialsarxiv-cs-lg
18 May 2026
Agents

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

DGX agent

arXiv:2605.15224v1 Announce Type: new Abstract: Large language model-based agents make mistakes, yet critique can often guide the same model toward correct behavior. However, when critique is removed,

agentsarxiv-cs-ai
18 May 2026
Model Releases

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

DGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Mask-Morph Graph U-Net: A Generalisable Mesh-Based Surrogate for Crashworthiness Field Prediction under Large Geometric Variation

DGX agent

arXiv:2605.15231v1 Announce Type: cross Abstract: Nonlinear finite element crash simulations are accurate but computationally expensive, limiting their use in iterative design optimisation. Machine-le

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

MyoChallenge 2025: A New Benchmark for Human Athletic Intelligence

DGX agent

arXiv:2605.15650v1 Announce Type: new Abstract: Athletic performance represents the pinnacle of human motor intelligence, demanding rapid choices, precise control, agility, and coordinated physical ex

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

PDRNN: Modular Data-driven Pedestrian Dead Reckoning on Loosely Coupled Radio- and Inertial-Signalstreams

DGX agent

arXiv:2605.15252v1 Announce Type: cross Abstract: Modern pedestrian dead reckoning (PDR) systems rely on fusing noisy and biased estimates of position, velocity, and calibrated orientation derived fro

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

DGX agent

arXiv:2601.12894v2 Announce Type: replace-cross Abstract: Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-s

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

DGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

model-releasesarxiv-cs-ai
15 May 2026
Research

A Systematic Evaluation of Imbalance Handling Methods in Biomedical Binary Classification

DGX agent

arXiv:2605.14147v1 Announce Type: new Abstract: Objective: The primary goal of this study was to systematically examine the impact of commonly used imbalance handling methods (IHMs) on predictive perf

researcharxiv-cs-lg
15 May 2026
Tutorials

AaSP: Aliasing-aware Self-Supervised Pre-Training for Audio Spectrogram Transformers

DGX agent

arXiv:2512.03637v2 Announce Type: replace-cross Abstract: Transformer-based audio self-supervised learning (SSL) models commonly use spectrograms, vision-style Transformers, and masked modeling object

tutorialsarxiv-cs-lg
15 May 2026
Safety

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

DGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

safetyarxiv-cs-lg
15 May 2026
Model Releases

AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

DGX agent

arXiv:2605.13940v1 Announce Type: cross Abstract: Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents

model-releasesarxiv-cs-ai
15 May 2026
Applications

AIMing for Standardised Explainability Evaluation in GNNs: A Framework and Case Study on Graph Kernel Networks

DGX agent

arXiv:2605.14884v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have advanced significantly in handling graph-structured data, but a comprehensive framework for evaluating explainability

applicationsarxiv-cs-lg
15 May 2026
Research

All-atomistic Transferable Neural Potentials for Protein Solvation

DGX agent

arXiv:2605.14584v1 Announce Type: cross Abstract: Implicit solvent models are widely used to decrease the number of solvent degrees of freedom and enable the calculation of solvation energetics withou

researcharxiv-cs-lg
15 May 2026
Model Releases

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

DGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Can Visual Mamba Improve AI-Generated Image Detection? An In-Depth Investigation

DGX agent

arXiv:2605.14799v1 Announce Type: new Abstract: In recent years, computer vision has witnessed remarkable progress, fueled by the development of innovative architectures such as Convolutional Neural N

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

DGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

DGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing

DGX agent

arXiv:2605.14084v1 Announce Type: cross Abstract: Code agents must both reason over long-horizon repository state and obey strict tool-use protocols. In paired Instruct/Thinking checkpoints, these cap

model-releasesarxiv-cs-ai
15 May 2026
Safety

CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation

DGX agent

arXiv:2605.14344v1 Announce Type: new Abstract: Generative modeling has emerged as a promising approach for crystal structure discovery. However, existing LLM-based generative models struggle with low

safetyarxiv-cs-ai
15 May 2026
Model Releases

Discovering Physical Directions in Weight Space: Composing Neural PDE Experts

DGX agent

arXiv:2605.14546v1 Announce Type: new Abstract: Recent advances in neural operators have made partial differential equation (PDE) surrogate modeling increasingly scalable and transferable through larg

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

DVMap: Fine-Grained Pluralistic Value Alignment via High-Consensus Demographic-Value Mapping

DGX agent

arXiv:2605.14420v1 Announce Type: new Abstract: Current Large Language Models (LLMs) typically rely on coarse-grained national labels for pluralistic value alignment. However, such macro-level supervi

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis

DGX agent

arXiv:2605.14842v1 Announce Type: new Abstract: Humans naturally communicate through abstract concepts like 'mood'. However, current image editing benchmarks focus primarily on explicit, literal comma

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Elastic Spiking Transformers for Efficient Gesture Understanding

DGX agent

arXiv:2605.13869v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs), particularly Spiking Transformers, offer energy-efficient processing of event-based sensor data for healthcare applica

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay

DGX agent

arXiv:2605.14237v1 Announce Type: new Abstract: Deploying AI agents for repetitive periodic tasks exposes a critical tension: Large Language Models (LLMs) offer unmatched flexibility in tool orchestra

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration

DGX agent

arXiv:2605.13848v1 Announce Type: new Abstract: Agentic LLM frameworks that rely on prompted orchestration, where the model itself determines workflow transitions, often suffer from hallucinated routi

model-releasesarxiv-cs-ai
15 May 2026
Safety

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution

DGX agent

arXiv:2605.14891v1 Announce Type: new Abstract: We introduce a multi-scale Image Super Resolution (ISR) method building on recent advances in Visual Auto-Regressive (VAR) modeling. VAR models break im

safetyarxiv-cs-cv
15 May 2026
Model Releases

JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation

DGX agent

arXiv:2512.12772v2 Announce Type: replace-cross Abstract: Understanding videos inherently requires reasoning over both visual and auditory information. To properly evaluate Omni-Large Language Models

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Kolmogorov-Arnold Chemical Reaction Neural Networks for learning pressure-dependent kinetic rate laws

DGX agent

arXiv:2511.07686v2 Announce Type: replace-cross Abstract: Chemical Reaction Neural Networks (CRNNs) have emerged as an interpretable machine learning framework for discovering reaction kinetics direct

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Krause Synchronization Transformers

DGX agent

arXiv:2602.11534v3 Announce Type: replace-cross Abstract: Self-attention in Transformers relies on globally normalized softmax weights, causing all tokens to compete for influence at every layer. When

model-releasesarxiv-cs-ai
15 May 2026
Tutorials

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt

DGX agent

arXiv:2510.15849v2 Announce Type: replace Abstract: Accurate tongue segmentation is crucial for reliable TCM analysis. Supervised models require large annotated datasets, while SAM-family models remai

tutorialsarxiv-cs-cv
15 May 2026
← Previous
1…445446447448449…1119
Next →