AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On

DGX agent

arXiv:2604.27958v1 Announce Type: new Abstract: Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limite

model-releasesarxiv-cs-cv
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

DGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching

DGX agent

arXiv:2604.27375v1 Announce Type: new Abstract: Reasoning photo retouching has gained significant traction, requiring models to analyze image defects, give reasoning processes, and execute precise ret

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

VeriTaS: The First Dynamic Benchmark for Multimodal Automated Fact-Checking

DGX agent

arXiv:2601.08611v2 Announce Type: replace-cross Abstract: The growing scale of online misinformation urgently demands Automated Fact-Checking (AFC). Existing benchmarks for evaluating AFC systems, how

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Visual Analysis of Multi-outcome Causal Graphs

DGX agent

arXiv:2408.02679v3 Announce Type: replace Abstract: We introduce a visual analysis method for multiple causal graphs with different outcome variables, namely, multi-outcome causal graphs. Multi-outcom

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning

DGX agent

arXiv:2604.27629v1 Announce Type: new Abstract: We present WaferSAGE, a framework for wafer defect visual question answering using small vision-language models. To address data scarcity in semiconduct

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

WebMall -- A Multi-Shop Benchmark for Evaluating Web Agents

DGX agent

arXiv:2508.13024v3 Announce Type: replace Abstract: LLM-based web agents have the potential to automate long-running web tasks, such as searching for products in multiple e-shops and subsequently orde

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

DGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

DGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

When Continual Learning Moves to Memory: A Study of Experience Reuse in LLM Agents

DGX agent

arXiv:2604.27003v1 Announce Type: cross Abstract: Memory-augmented LLM agents offer an appealing shortcut to continual learning: rather than updating model parameters, they accumulate experience in ex

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

When Roles Fail: Epistemic Constraints on Advocate Role Fidelity in LLM-Based Political Statement Analysis

DGX agent

arXiv:2604.27228v1 Announce Type: new Abstract: Democratic discourse analysis systems increasingly rely on multi-agent LLM pipelines in which distinct evaluator models are assigned adversarial roles t

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments

DGX agent

arXiv:2604.27776v1 Announce Type: new Abstract: While GUI agents have shown impressive capabilities in common computer-use tasks such as OSWorld, current benchmarks mainly focus on isolated and single

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

WST-X Series: Wavelet Scattering Transform for Interpretable Speech Deepfake Detection

DGX agent

arXiv:2602.02980v2 Announce Type: replace-cross Abstract: In this work, we focus on front-end design for speech deepfake detectors, the component that determines the discriminative acoustic cues provi

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

3D-LENS: A 3D Lifting-based Elevated Novel-view Synthesis method for Single-View Aerial-Ground Re-Identification

DGX agent

arXiv:2604.26520v1 Announce Type: new Abstract: Aerial-Ground Re-Identification (AG-ReID) is constrained by the viewpoint-domain gap, as drastic viewpoint disparities occlude or distort discriminative

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

A Multi-Dataset Benchmark of Multiple Instance Learning for 3D Neuroimage Classification

DGX agent

arXiv:2604.26807v1 Announce Type: new Abstract: Despite being resource-intensive to train, 3D convolutional neural networks (CNNs) have been the standard approach to classify CT and MRI scans. Recent

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

A Multistage Extraction Pipeline for Long Scanned Financial Documents: An Empirical Study in Industrial KYC Workflows

DGX agent

arXiv:2604.26462v1 Announce Type: new Abstract: Structured information extraction from long, multilingual scanned financial documents is a core requirement in industrial KYC and compliance workflows.

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

A Note on How to Remove the lnln T Term from the Squint Bound

DGX agent

arXiv:2604.26926v1 Announce Type: new Abstract: In Orabona and Pal [2016], we introduced the shifted KT potentials, to remove the ln ln T factor in the parameter-free learning with expert bound. In th

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio

DGX agent

arXiv:2409.06624v4 Announce Type: replace-cross Abstract: Large Language Models (LLM) often need to be Continual Pre-Trained (CPT) to obtain unfamiliar language skills or adapt to new domains. The hug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

A self-evolving agent for explainable diagnosis of DFT-experiment band-gap mismatch

DGX agent

arXiv:2604.26703v1 Announce Type: cross Abstract: Standard density functional theory (DFT) routinely misclassifies the electronic ground state of correlated and structurally complex compounds, predict

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

A Systematic Comparison of Prompting and Multi-Agent Methods for LLM-based Stance Detection

DGX agent

arXiv:2604.26319v1 Announce Type: new Abstract: Stance detection identifies the attitude of a text author toward a given target. Recent studies have explored various LLM-based strategies for this task

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents

DGX agent

arXiv:2603.16496v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on external memory to support long-horizon interaction, personalized assistance, and multi-step

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Adaptive and Fine-grained Module-wise Expert Pruning for Efficient LoRA-MoE Fine-Tuning

DGX agent

arXiv:2604.26340v1 Announce Type: new Abstract: LoRA-MoE has emerged as an effective paradigm for parameter-efficient fine-tuning, combining the low training cost of LoRA with the increased adaptation

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning

DGX agent

arXiv:2508.19900v2 Announce Type: replace Abstract: Offline reinforcement learning (RL) enables learning effective policies from fixed datasets without any environment interaction. Existing methods ty

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

Affective Flow Language Model for Emotional Support Conversation

DGX agent

arXiv:2602.08826v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been widely applied to emotional support conversation (ESC). However, complex multi-turn support remains cha

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests

DGX agent

arXiv:2601.17617v3 Announce Type: replace-cross Abstract: LLM-powered search agents are increasingly being used for multi-step information seeking tasks, yet the IR community lacks empirical understan

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision

DGX agent

arXiv:2604.26567v1 Announce Type: new Abstract: Despite the rapid progress in data-driven 3D vision, aerial geometric 3D vision remains a formidable challenge due to the severe scarcity of large-scale

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence

DGX agent

arXiv:2604.25930v1 Announce Type: new Abstract: We study whether a structured recurrent state can serve as a compact associative backbone for language modeling while still supporting exact retrieval.

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Auditing Marketing Budget Allocation with Hindsight Regret

DGX agent

arXiv:2604.25977v1 Announce Type: cross Abstract: Organizations routinely make strategic budget allocations under operational constraints, but often lack a principled way to assess whether realized al

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI

DGX agent

arXiv:2604.26382v1 Announce Type: cross Abstract: Most enterprise document AI today is a pipeline. Parse, index, retrieve, generate. Each of those stages has been studied to death on its own -- what's

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation

DGX agent

arXiv:2506.21444v4 Announce Type: replace Abstract: Atypical mitosis marks a deviation in the cell division process that has been shown be an independent prognostic marker for tumor malignancy. Howeve

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-Codex

DGX agent

arXiv:2604.14858v2 Announce Type: replace Abstract: As agent systems move into increasingly diverse execution settings, trajectory-level safety evaluation and diagnosis require benchmarks that evolve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models

DGX agent

arXiv:2604.26365v1 Announce Type: new Abstract: To address the high sampling cost of Diffusion Transformers (DiTs), feature caching offers a training-free acceleration method. However, existing method

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language Models

DGX agent

arXiv:2508.04325v2 Announce Type: replace-cross Abstract: Large language models (LLMs) show significant potential in healthcare, prompting numerous benchmarks to evaluate their capabilities. However,

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection

DGX agent

arXiv:2604.26868v1 Announce Type: new Abstract: Existing 3D anomaly detection methods are built on a rigid prior: normal geometry is pose-invariant and can be canonicalized through registration or ali

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Bridge: Basis-Driven Causal Inference Marries VFMs for Domain Generalization

DGX agent

arXiv:2604.26820v1 Announce Type: new Abstract: Detectors often suffer from degraded performance, primarily due to the distributional gap between the source and target domains. This issue is especiall

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems

DGX agent

arXiv:2604.26145v1 Announce Type: cross Abstract: AI-powered language learning tools increasingly provide instant, personalised feedback to millions of learners worldwide. However, this feedback can f

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

ChinaTravel: An Open-Ended Travel Planning Benchmark with Compositional Constraint Validation for Language Agents

DGX agent

arXiv:2412.13682v5 Announce Type: replace Abstract: Travel planning stands out among real-world applications of Language Agents because it couples significant practical demand with a rigorous constrai

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

ClassEval-Pro: A Cross-Domain Benchmark for Class-Level Code Generation

DGX agent

arXiv:2604.26923v1 Announce Type: cross Abstract: LLMs have achieved strong results on both function-level code synthesis and repository-level code modification, yet a capability that falls between th

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

ClawGym: A Scalable Framework for Building Effective Claw Agents

DGX agent

arXiv:2604.26904v1 Announce Type: cross Abstract: Claw-style environments support multi-step workflows over local files, tools, and persistent workspace states. However, scalable development around th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Compton Form Factor Extraction using Quantum Deep Neural Networks

DGX agent

arXiv:2504.15458v4 Announce Type: replace Abstract: We extract Compton form factors (CFFs) from deeply virtual Compton scattering measurements at the Thomas Jefferson National Accelerator Facility (JL

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models

DGX agent

arXiv:2604.25922v1 Announce Type: cross Abstract: We present DenialBench, a systematic benchmark measuring consciousness denial behaviors across 115 large language models from 25+ providers. Using a t

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

COP-GEN: Latent Diffusion Transformer for Copernicus Earth Observation Data

DGX agent

arXiv:2603.03239v2 Announce Type: replace Abstract: Earth observation applications increasingly rely on data from multiple sensors, including optical, radar, elevation, and land-cover. Relationships b

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

CoQuant: Joint Weight-Activation Subspace Projection for Mixed-Precision LLMs

DGX agent

arXiv:2604.26378v1 Announce Type: new Abstract: Post-training quantization (PTQ) has become an important technique for reducing the inference cost of Large Language Models (LLMs). While recent mixed-p

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

Cross-Domain Transfer of Hyperspectral Foundation Models

DGX agent

arXiv:2604.26478v1 Announce Type: new Abstract: Hyperspectral imaging (HSI) semantic segmentation typically relies on in-domain training, but limited data availability often restricts model performanc

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

CurEvo: Curriculum-Guided Self-Evolution for Video Understanding

DGX agent

arXiv:2604.26707v1 Announce Type: new Abstract: Recent advances in self-evolution video understanding frameworks have demonstrated the potential of autonomous learning without human annotations. Howev

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces

DGX agent

arXiv:2505.18441v2 Announce Type: replace Abstract: Dictionary learning has recently emerged as a promising approach for mechanistic interpretability of large transformer models. Disentangling high-di

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

Decoupled Prototype Matching with Vision Foundation Models for Few-Shot Industrial Object Detection

DGX agent

arXiv:2604.26404v1 Announce Type: new Abstract: Industrial object detection systems typically rely on large annotated datasets, which are expensive to collect and challenging to maintain in industrial

model-releasesarxiv-cs-cv
30 Apr 2026
← Previous
1…288289290291292…361
Next →