AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,598Total entries
1Added by human
91,597Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,253 results
Tutorials

Hypothesis-driven construction of mesoscopic dynamics

DGX agent

arXiv:2605.16211v1 Announce Type: new Abstract: Traditional scientific modeling typically begins with fixed, instance-wise effective equations and then carries out equation-specific analysis and compu

tutorialsarxiv-cs-lg
18 May 2026
Agents

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.15224v1 Announce Type: new Abstract: Large language model-based agents make mistakes, yet critique can often guide the same model toward correct behavior. However, when critique is removed,

agentsarxiv-cs-ai
18 May 2026
Model Releases

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

DGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Mask-Morph Graph U-Net: A Generalisable Mesh-Based Surrogate for Crashworthiness Field Prediction under Large Geometric Variation

DGX agent

arXiv:2605.15231v1 Announce Type: cross Abstract: Nonlinear finite element crash simulations are accurate but computationally expensive, limiting their use in iterative design optimisation. Machine-le

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

MyoChallenge 2025: A New Benchmark for Human Athletic Intelligence

DGX agent

arXiv:2605.15650v1 Announce Type: new Abstract: Athletic performance represents the pinnacle of human motor intelligence, demanding rapid choices, precise control, agility, and coordinated physical ex

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

PDRNN: Modular Data-driven Pedestrian Dead Reckoning on Loosely Coupled Radio- and Inertial-Signalstreams

DGX agent

arXiv:2605.15252v1 Announce Type: cross Abstract: Modern pedestrian dead reckoning (PDR) systems rely on fusing noisy and biased estimates of position, velocity, and calibrated orientation derived fro

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can…

DGX agent

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can't wait to release Qwen3.7 series models!Stay tuned! @arena Qw

model-releasesqwen--x
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

DGX agent

arXiv:2601.12894v2 Announce Type: replace-cross Abstract: Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-s

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Cl…

DGX agent

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Claude in 5 min 2026 → train your own model over a weekend hap

model-releasesfireworks-ai--x
17 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the env…

DGX agent

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the environment itself is the verifier. 🏆Among frontier models, @An

model-releaseszhipu-ai--x
17 May 2026
Model Releases

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

DGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

model-releasesarxiv-cs-ai
15 May 2026
Research

A Systematic Evaluation of Imbalance Handling Methods in Biomedical Binary Classification

DGX agent

arXiv:2605.14147v1 Announce Type: new Abstract: Objective: The primary goal of this study was to systematically examine the impact of commonly used imbalance handling methods (IHMs) on predictive perf

researcharxiv-cs-lg
15 May 2026
Tutorials

AaSP: Aliasing-aware Self-Supervised Pre-Training for Audio Spectrogram Transformers

DGX agent

arXiv:2512.03637v2 Announce Type: replace-cross Abstract: Transformer-based audio self-supervised learning (SSL) models commonly use spectrograms, vision-style Transformers, and masked modeling object

tutorialsarxiv-cs-lg
15 May 2026
Safety

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

DGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

safetyarxiv-cs-lg
15 May 2026
Model Releases

AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

DGX agent

arXiv:2605.13940v1 Announce Type: cross Abstract: Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AI radio hosts demonstrate why AI can’t be trusted alone

DGX agent

Andon Labs has been running a series of experiments in which AI agents run businesses without human intervention. Its latest is a quartet of radio stations run by some of the most popular AI models ou

model-releasesthe-verge-ai
15 May 2026
Applications

AIMing for Standardised Explainability Evaluation in GNNs: A Framework and Case Study on Graph Kernel Networks

DGX agent

arXiv:2605.14884v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have advanced significantly in handling graph-structured data, but a comprehensive framework for evaluating explainability

applicationsarxiv-cs-lg
15 May 2026
Research

All-atomistic Transferable Neural Potentials for Protein Solvation

DGX agent

arXiv:2605.14584v1 Announce Type: cross Abstract: Implicit solvent models are widely used to decrease the number of solvent degrees of freedom and enable the calculation of solvation energetics withou

researcharxiv-cs-lg
15 May 2026
Model Releases

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

DGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Can Visual Mamba Improve AI-Generated Image Detection? An In-Depth Investigation

DGX agent

arXiv:2605.14799v1 Announce Type: new Abstract: In recent years, computer vision has witnessed remarkable progress, fueled by the development of innovative architectures such as Convolutional Neural N

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

DGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

DGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing

DGX agent

arXiv:2605.14084v1 Announce Type: cross Abstract: Code agents must both reason over long-horizon repository state and obey strict tool-use protocols. In paired Instruct/Thinking checkpoints, these cap

model-releasesarxiv-cs-ai
15 May 2026
Safety

CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation

DGX agent

arXiv:2605.14344v1 Announce Type: new Abstract: Generative modeling has emerged as a promising approach for crystal structure discovery. However, existing LLM-based generative models struggle with low

safetyarxiv-cs-ai
15 May 2026
Model Releases

Discovering Physical Directions in Weight Space: Composing Neural PDE Experts

DGX agent

arXiv:2605.14546v1 Announce Type: new Abstract: Recent advances in neural operators have made partial differential equation (PDE) surrogate modeling increasingly scalable and transferable through larg

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

DVMap: Fine-Grained Pluralistic Value Alignment via High-Consensus Demographic-Value Mapping

DGX agent

arXiv:2605.14420v1 Announce Type: new Abstract: Current Large Language Models (LLMs) typically rely on coarse-grained national labels for pluralistic value alignment. However, such macro-level supervi

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis

DGX agent

arXiv:2605.14842v1 Announce Type: new Abstract: Humans naturally communicate through abstract concepts like 'mood'. However, current image editing benchmarks focus primarily on explicit, literal comma

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Elastic Spiking Transformers for Efficient Gesture Understanding

DGX agent

arXiv:2605.13869v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs), particularly Spiking Transformers, offer energy-efficient processing of event-based sensor data for healthcare applica

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay

DGX agent

arXiv:2605.14237v1 Announce Type: new Abstract: Deploying AI agents for repetitive periodic tasks exposes a critical tension: Large Language Models (LLMs) offer unmatched flexibility in tool orchestra

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration

DGX agent

arXiv:2605.13848v1 Announce Type: new Abstract: Agentic LLM frameworks that rely on prompted orchestration, where the model itself determines workflow transitions, often suffer from hallucinated routi

model-releasesarxiv-cs-ai
15 May 2026
Safety

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution

DGX agent

arXiv:2605.14891v1 Announce Type: new Abstract: We introduce a multi-scale Image Super Resolution (ISR) method building on recent advances in Visual Auto-Regressive (VAR) modeling. VAR models break im

safetyarxiv-cs-cv
15 May 2026
Model Releases

How business operations teams use Codex

DGX agent

This document from OpenAI describes how business operations teams leverage Codex, OpenAI's code generation model, to automate and streamline their workflows. It likely covers practical applications su

model-releasesopenai
15 May 2026
Model Releases

How sales teams use Codex

DGX agent

This resource from OpenAI's Academy demonstrates practical applications of Codex, their code-generation AI model, within sales team workflows. It likely covers how sales professionals can leverage Cod

model-releasesopenai
15 May 2026
Model Releases

JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation

DGX agent

arXiv:2512.12772v2 Announce Type: replace-cross Abstract: Understanding videos inherently requires reasoning over both visual and auditory information. To properly evaluate Omni-Large Language Models

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Kolmogorov-Arnold Chemical Reaction Neural Networks for learning pressure-dependent kinetic rate laws

DGX agent

arXiv:2511.07686v2 Announce Type: replace-cross Abstract: Chemical Reaction Neural Networks (CRNNs) have emerged as an interpretable machine learning framework for discovering reaction kinetics direct

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Krause Synchronization Transformers

DGX agent

arXiv:2602.11534v3 Announce Type: replace-cross Abstract: Self-attention in Transformers relies on globally normalized softmax weights, causing all tokens to compete for influence at every layer. When

model-releasesarxiv-cs-ai
15 May 2026
Tutorials

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt

DGX agent

arXiv:2510.15849v2 Announce Type: replace Abstract: Accurate tongue segmentation is crucial for reliable TCM analysis. Supervised models require large annotated datasets, while SAM-family models remai

tutorialsarxiv-cs-cv
15 May 2026
Model Releases

Peng's Q(lambda) for Conservative Value Estimation in Offline Reinforcement Learning

DGX agent

arXiv:2605.14779v1 Announce Type: new Abstract: We propose a model-free offline multi-step reinforcement learning (RL) algorithm, Conservative Peng's Q(lambda) (CPQL). Our algorithm adapts the Peng's

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Physics-Grounded Adversarial Stain Augmentation with Calibrated Coverage Guarantees

DGX agent

arXiv:2605.13889v1 Announce Type: cross Abstract: Stain variation across hospitals degrades histopathology models at deployment. Existing augmentation methods perturb color spaces with arbitrary hyper

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

pi-Bench: Evaluating Proactive Personal Assistant Agents in Long-Horizon Workflows

DGX agent

arXiv:2605.14678v1 Announce Type: new Abstract: The rise of personal assistant agents, e.g., OpenClaw, highlights the growing potential of large language models to support users across everyday life a

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

PreFT: Prefill-only finetuning for efficient inference

DGX agent

arXiv:2605.14217v1 Announce Type: cross Abstract: Large language models can now be personalised efficiently at scale using parameter efficient finetuning methods (PEFTs), but serving user-specific PEF

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

RefDecoder: Enhancing Visual Generation with Conditional Video Decoding

DGX agent

arXiv:2605.15196v1 Announce Type: new Abstract: Video generation powers a vast array of downstream applications. However, while the de facto standard, i.e., latent diffusion models, typically employ h

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR)

DGX agent

arXiv:2605.14126v1 Announce Type: cross Abstract: Fast Healthcare Interoperability Resources (FHIR) is the dominant standard for interoperable exchange of healthcare data. In FHIR, electronic health r

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Safe Bayesian Optimization for Complex Control Systems via Additive Gaussian Processes

DGX agent

arXiv:2408.16307v3 Announce Type: replace-cross Abstract: Automatic controller tuning is attractive for robotics and mechatronic systems whose dynamics are difficult to model accurately, but direct bl

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Sat3DGen: Comprehensive Street-Level 3D Scene Generation from Single Satellite Image

DGX agent

arXiv:2605.14984v1 Announce Type: cross Abstract: Generating a street-level 3D scene from a single satellite image is a crucial yet challenging task. Current methods present a stark trade-off: geometr

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…541542543544545…1381
Next →