AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

AWS launches Agentic Shopping Assistant to help retailers build AI tools

DGX agent

Amazon Web Services Inc. today introduced a new offering designed to help retailers integrate artificial intelligence features into their online stores. AWS Agentic Shopping Assistant, or ASA, combine

model-releasessiliconangle
27 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Axial-Centric Cross-Plane Attention for 3D Medical Image Classification

DGX agent

arXiv:2602.21636v2 Announce Type: replace Abstract: Abridged: Clinicians commonly interpret 3D medical images by examining multiple anatomical planes rather than relying on volumetric views. In clinic

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

DGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

DGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

DGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback

DGX agent

arXiv:2509.21106v2 Announce Type: replace Abstract: Search-augmented large language models (LLMs) have advanced information-seeking tasks by integrating retrieval into generation, reducing users' cogn

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal

DGX agent

arXiv:2605.26772v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate chain-of-thought (CoT) traces before producing final outputs, introducing a dynamic internal state that may compl

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Beyond Questions: Evaluating What Large Language Models (Actually) Know

DGX agent

arXiv:2605.26937v1 Announce Type: cross Abstract: Parametric knowledge in large language models (LLMs) is a cornerstone of their success, yet remains poorly understood. Existing knowledge benchmarks t

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation

DGX agent

arXiv:2601.08146v3 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting their use on diverse natural text. We adapt Co

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

BhashaSetu: A Data-Centric Approach to Low-Resource Machine Translation

DGX agent

arXiv:2605.27050v1 Announce Type: new Abstract: We present BhashaSetu, a linguistically enriched English--Marathi parallel dataset addressing persistent data limitations in low-resource neural machine

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

DGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

DGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

DGX agent

arXiv:2605.27000v1 Announce Type: cross Abstract: Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@K as the canonical metric. Yet the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Causal Representation Learning for Generalisable Recommendation

DGX agent

arXiv:2605.27043v1 Announce Type: cross Abstract: Predictive models trained on observational data often fail to generalise to the distributions they encounter when deployed, especially when the traini

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Cesarean Scar Defect Segmentation in Transvaginal Ultrasound Images: a Dataset and Benchmark

DGX agent

arXiv:2605.26774v1 Announce Type: new Abstract: Cesarean Scar Defect (CSD) is one of the most prevalent complications following cesarean delivery. Transvaginal ultrasonography is widely used for prima

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Chain Of Thought Compression: A Theoretical Analysis

DGX agent

arXiv:2601.21576v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has unlocked advanced reasoning abilities of Large Language Models (LLMs) with intermediate steps, yet incurs prohibitive com

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ChartAct: A Benchmark for Dynamic Chart Understanding

DGX agent

arXiv:2605.26994v1 Announce Type: new Abstract: Charts are widely used to present complex data for analysis and decision making. Existing chart understanding benchmarks mainly focus on static charts,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains

DGX agent

arXiv:2605.26734v1 Announce Type: new Abstract: Existing Multi-Turn Composed Image Retrieval (MTCIR) datasets lack dialogue-history consistency and are restricted to the fashion domain. To address the

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

DGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

DGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Clinically-Grounded Counterfactual Reasoning for Medical Video Diagnosis

DGX agent

arXiv:2605.26483v1 Announce Type: new Abstract: Medical video diagnosis involves inferring clinical decisions from dynamic tissue responses throughout examination processes. Existing methods rely on a

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection

DGX agent

arXiv:2605.26294v1 Announce Type: new Abstract: Skin cancer is a common and fast rising malignancy worldwide. Early detection is critical for improving outcomes. Deep learning models trained on dermos

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

CodecCap: High-Fidelity Codec-Inspired Residual Modeling for Dense Video Captioning

DGX agent

arXiv:2605.26967v1 Announce Type: new Abstract: Existing video captioning methods struggle to balance visual fidelity and redundancy: holistic captions are compact but lose fine-grained evidence, wher

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Cogent Security launches autonomous vulnerability response tools as AI-assisted exploits outpace scanners

DGX agent

Cogent Security Inc., a startup that employs agentic artificial intelligence for vulnerability management, today launched two new platform capabilities aimed at compressing enterprise vulnerability re

model-releasessiliconangle
27 May 2026
Model Releases

Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning

DGX agent

arXiv:2605.26789v1 Announce Type: new Abstract: Post-training is routinely evaluated through aggregate benchmark scores that treat multi-hop reasoning as a single capability -- as if a model that answ

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Constraint acquisition needs better benchmarks

DGX agent

arXiv:2605.26279v1 Announce Type: new Abstract: Constraint Acquisition (CA) and related research on the validation and enhancement of Mathematical Programming (MP) models from domain knowledge artifac

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Constructing Industrial-Scale Optimization Modeling Benchmark

DGX agent

arXiv:2602.10450v2 Announce Type: replace-cross Abstract: Optimization modeling underpins decision-making in logistics, manufacturing, energy, and finance, yet translating natural-language requirement

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ConVer: Using Contracts and Loop Invariant Synthesis for Scalable Formal Software Verification

DGX agent

arXiv:2605.27051v1 Announce Type: cross Abstract: Formal verification of large C programs is impeded by state-space explosion: Bounded Model Checking (BMC) tools must encode the entire state space up

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection

DGX agent

arXiv:2605.27116v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has made significant progress, enabling detectors to generalize from seen to unseen categories. However, real-wor

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Curation and Extraction of Drug-Related Entities from Reddit Platform

DGX agent

arXiv:2605.26445v1 Announce Type: new Abstract: Physicians learn primarily about illicit drugs from clinical overdose cases, limiting their understanding of real-world usage. Meanwhile, drug users sha

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% — For months,

model-releasestechmeme
27 May 2026
Model Releases

Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems

DGX agent

arXiv:2605.27133v1 Announce Type: cross Abstract: Deep unfolding neural networks derived from iterative optimization schemes and numerical ordinary/partial differential equations (ODEs/PDEs) have attr

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DEI: Diversity in Evolutionary Inference for Quality-Diversity Search

DGX agent

arXiv:2605.27130v1 Announce Type: cross Abstract: We present DEI: Diversity in Evolutionary Inference, a distributed Quality-Diversity (QD) search framework that assigns heterogeneous large language m

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DelowlightSplat: Feed-Forward Gaussian Splatting for Lowlight 3D Scene Reconstruction

DGX agent

arXiv:2605.26629v1 Announce Type: new Abstract: Novel-view synthesis and 3D reconstruction from sparse posed images are central to robotics and AR/VR. Yet, feed-forward 3D Gaussian reconstruction fail

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Dense2MoE: Pushing the Pareto Frontier of On-Device LLMs via Unified Pruning and Upcycling

DGX agent

arXiv:2605.26496v1 Announce Type: cross Abstract: The Mixture of Experts MoE architecture is highly promising for resource constrained on device deployments yet training these models from scratch incu

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Developing a Totally Unimodular Linear Program for Optimal Conformance Checking: When and Why It Complements A*

DGX agent

arXiv:2605.26938v1 Announce Type: new Abstract: Alignment-based conformance checking is the state-of-the-art approach for comparing observed process executions with normative process models. The stand

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Device Context Protocol: A Compact, Safety-First Architecture for LLM-Driven Control of Constrained Devices

DGX agent

arXiv:2605.26159v1 Announce Type: cross Abstract: Large language models are increasingly used as orchestrators of external tools via the Model Context Protocol (MCP), but MCP is built for software ser

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

DGLD: Domain-Gated Latent Diffusion for the Discovery of Novel Energetic Materials

DGX agent

arXiv:2605.26540v1 Announce Type: cross Abstract: Energetic-materials performance gains translate directly into reduced propellant mass, smaller warheads, and more efficient civilian gas-generators, y

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

DGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Distribution-Aware Conformal Prediction: A Framework for generating efficient prediction intervals for time series

DGX agent

arXiv:2605.26569v1 Announce Type: new Abstract: We present Distribution-aware Conformal Prediction (DCP), a unified framework integrating probabilistic predictors like Monte Carlo dropout, deep ensemb

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Doppel launches agentic email security to disrupt phishing campaigns at the source

DGX agent

Social engineering defense startup Doppel Inc. today launched Doppel Email Security, an agentic artificial intelligence layer that traces phishing messages back to attacker infrastructure and orchestr

model-releasessiliconangle
27 May 2026
Model Releases

Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving

DGX agent

arXiv:2601.14702v2 Announce Type: replace Abstract: Autonomous driving requires reliable perception and safe decision-making in complex scenarios. Recent vision-language models (VLMs) demonstrate reas

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DunbaaBERT: From Sacrifice to Semantics

DGX agent

arXiv:2605.26935v1 Announce Type: new Abstract: Large language models have achieved strong performance across many NLP tasks, yet Urdu remains comparatively underexplored due to limited resources and

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

E3: Issue-Level Backtesting for Automated Research Critique

DGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning

DGX agent

arXiv:2602.02192v5 Announce Type: replace Abstract: Reinforcement learning (RL) is a critical stage in post-training large language models (LLMs), involving repeated interaction between rollout genera

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models

DGX agent

arXiv:2510.07231v4 Announce Type: replace-cross Abstract: Socio-economic causal effects depend heavily on their institutional and environmental contexts. The same intervention can produce different, e

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…259260261262263…475
Next →