AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Compressed Video Aggregator: Content-driven Module for Efficient Micro-Video Recommendation

DGX agent

arXiv:2605.08810v1 Announce Type: cross Abstract: We propose Compressed Video Aggregator (CVA), a lightweight micro-video recommendation module that decouples video information from preference learnin

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Computer Use at the Edge of the Statistical Precipice

DGX agent

arXiv:2605.08261v1 Announce Type: cross Abstract: Evaluating Computer Use Agents (CUAs) on interactive environments is fraught with methodological pitfalls that the field has yet to systematically add

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Concordia: Self-Improving Synthetic Tables for Federated LLMs

DGX agent

arXiv:2605.09855v1 Announce Type: new Abstract: Federated learning (FL) enables training large language models (LLMs) without sharing raw data, but adapting LLMs under strict data isolation and non-II

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition

DGX agent

arXiv:2605.10916v1 Announce Type: cross Abstract: Recognition of handwritten Bangla compound characters remains a challenging problem due to complex character structures, large intra-class variation,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

ConFit v3: Improving Resume-Job Matching with LLM-based Re-Ranking

DGX agent

arXiv:2605.09760v1 Announce Type: new Abstract: A reliable resume-job matching system helps a company find suitable candidates from a pool of resumes and helps a job seeker find relevant jobs from a l

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ConQuR: Corner Aligned Activation Quantization via Optimized Rotations for LLMs

DGX agent

arXiv:2605.10793v1 Announce Type: new Abstract: Large language models (LLMs) are costly to deploy due to their large memory footprint and high inference cost. Weight-activation quantization can reduce

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%

DGX agent

arXiv:2605.08112v1 Announce Type: cross Abstract: AI coding agents powered by large language models can read codebases and produce functional code, but they routinely violate team-specific product dec

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Continual Harness: Online Adaptation for Self-Improving Foundation Agents

DGX agent

arXiv:2605.09998v1 Announce Type: cross Abstract: Coding harnesses such as Claude Code and OpenHands wrap foundation models with tools, memory, and planning, but no equivalent exists for embodied agen

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Continuous Latent Contexts Enable Efficient Online Learning in Transformers

DGX agent

arXiv:2605.09867v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit a strong capacity for in-context learning: Given labeled examples, they can generate good predictions without par

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Convergence Analysis of Newton's Method for Neural Networks in the Overparameterized Limit

DGX agent

arXiv:2605.08352v1 Announce Type: new Abstract: A convergence analysis is developed for the regularized Newton method for training neural networks (NNs) in the overparameterized limit. As the number o

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation

DGX agent

arXiv:2605.08522v1 Announce Type: new Abstract: The evaluation of Large Language Models (LLMs) faces a critical challenge in construct validity, where fragmented benchmarks and ad hoc metrics frequent

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

CORTEG: Foundation Models Enable Cross-Modality Representation Transfer from Scalp to Intracranial Brain Recordings

DGX agent

arXiv:2605.10337v1 Announce Type: new Abstract: Intracranial electrocorticography (ECoG) offers high-signal-to-noise access to cortical activity for brain-computer interfaces, yet limited per-patient

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Cosine-Gated Adam-Decay: Drop-In Staleness-Aware Outer Optimization for Decoupled DiLoCo

DGX agent

arXiv:2605.09126v1 Announce Type: new Abstract: Asynchronous DiLoCo systems may receive pseudo-gradients computed several outer rounds earlier, yet the standard Nesterov outer optimizer does not expli

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving

DGX agent

arXiv:2605.10426v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for end-to-end autonomous driving. However, existing reasoning mechanisms sti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CrackMeBench: Binary Reverse Engineering for Agents

DGX agent

arXiv:2605.10597v1 Announce Type: cross Abstract: Benchmarks for coding agents increasingly measure source-level software repair, and cybersecurity benchmarks increasingly measure broad capture-the-fl

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CREATE: Testing LLMs for Associative Creativity

DGX agent

arXiv:2603.09970v2 Announce Type: replace Abstract: A key component of creativity is associative reasoning: the ability to draw novel yet meaningful connections between concepts. We introduce CREATE,

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Cross-Family Universality of Behavioral Axes via Anchor-Projected Representations

DGX agent

arXiv:2605.09875v1 Announce Type: new Abstract: Large language models from different families use different hidden dimensions, tokenizers, and training procedures, making behavioral directions difficu

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Cross-Sample Relational Fusion: Unifying Domain Generalization and Class-Incremental Learning

DGX agent

arXiv:2605.08839v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) requires a learning system to learn new classes while retaining previously learned knowledge. However, in real-world sc

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

CrystalREPA: Transferring Physical Priors from Universal MLIPs to Crystal Generative Models

DGX agent

arXiv:2605.08960v1 Announce Type: cross Abstract: Crystal generative models mainly learn what stable crystals look like, with little explicit supervision for what makes them stable. We reveal a substa

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

CT-IDP: Segmentation-Derived Quantitative Phenotypes for Interpretable Abdominal CT Disease Classification

DGX agent

arXiv:2605.09002v1 Announce Type: cross Abstract: In this retrospective multi-institutional study, a quantitative phenotyping framework, CT-IDP (CT Image-Derived Phenotypes) was developed on the MERLI

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CTQWformer: A CTQW-based Transformer for Graph Classification

DGX agent

arXiv:2605.09486v1 Announce Type: cross Abstract: Graph Neural Networks (GNN) and Transformer-based architectures have achieved remarkable progress in graph learning, yet they still struggle to captur

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CUDABeaver: Benchmarking LLM-Based Automated CUDA Debugging

DGX agent

arXiv:2605.08455v1 Announce Type: new Abstract: Debugging CUDA programs has long been challenging because failures often arise from subtle interactions among hardware behavior, compiler decisions, mem

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs

DGX agent

arXiv:2605.08467v1 Announce Type: new Abstract: Large language models show promise for automated CUDA programming, however even the strongest coding models (e.g., Claude-Opus-4.6) may still fall short

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

cuRegOT: A GPU-Accelerated Solver for Entropic-Regularized Optimal Transport

DGX agent

arXiv:2605.08793v1 Announce Type: cross Abstract: Optimal transport (OT) has emerged as a fundamental tool in modern machine learning, yet its computational cost remains a significant bottleneck for l

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

DGX agent

arXiv:2605.10933v1 Announce Type: cross Abstract: While Mixture-of-Experts (MoE) scales model capacity without proportionally increasing computation, its massive total parameter footprint creates sign

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Decomposing and Steering Functional Metacognition in Large Language Models

DGX agent

arXiv:2605.08942v1 Announce Type: new Abstract: Large language models (LLMs) increasingly exhibit behaviors suggesting awareness of their evaluation context, often adapting their reasoning strategies

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Deep Learning under Fractional-Order Differential Privacy

DGX agent

arXiv:2605.09890v1 Announce Type: cross Abstract: Differentially private stochastic gradient descent (DP-SGD) is a standard approach to privacy-preserving learning based on per-example clipping, subsa

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Deepfake Detection that Generalizes Across Benchmarks

DGX agent

arXiv:2508.06248v4 Announce Type: replace Abstract: The generalization of deepfake detectors to unseen manipulation techniques remains a challenge for practical deployment. Although many approaches ad

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving

DGX agent

arXiv:2605.10564v1 Announce Type: new Abstract: End-to-end autonomous driving systems are increasingly integrating Vision-Language Model (VLM) architectures, incorporating text reasoning or visual rea

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents

DGX agent

arXiv:2605.09679v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) and AI agents have made significant progress in learning to analyze and reason about clinical images. However, e

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DeformMaster: An Interactive Physics-Neural World Model for Deformable Objects from Videos

DGX agent

arXiv:2605.09586v1 Announce Type: new Abstract: World models for deformable objects should recover not only geometry and appearance, but also underlying physical dynamics, interaction grounding, and m

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Detect, Localize, and Explain: Interactive Hierarchical Log Anomaly Analytics with LLM Augmentation

DGX agent

arXiv:2605.09222v1 Announce Type: cross Abstract: Logs are ubiquitous in modern systems. Unfortunately, their unstructured nature in flat sequences limits understanding of execution behaviors, hinderi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Detecting Multi-Agent Collusion Through Multi-Agent Interpretability

DGX agent

arXiv:2604.01151v2 Announce Type: replace Abstract: As LLM agents are increasingly deployed in multi-agent systems, they introduce risks of covert coordination that may evade standard forms of human o

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Diagnosing Spectral Ceilings in Equivariant Neural Force Fields

DGX agent

arXiv:2605.08286v1 Announce Type: cross Abstract: We introduce a spectral-injection diagnostic for measuring which angular frequencies a trained equivariant force-field backbone preserves: inject a co

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules

DGX agent

arXiv:2605.08614v1 Announce Type: new Abstract: Monitoring complex industrial assets relies on engineer-authored symbolic rules that trigger based on sensor conditions and prompt technicians to perfor

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DiffATS: Diffusion in Aligned Tensor Space

DGX agent

arXiv:2605.09275v1 Announce Type: new Abstract: Direct diffusion modeling of high-resolution spatiotemporal fields is computationally challenging. Parameter-efficient primitives address this by repres

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Different Prompts, Different Ranks: Prompt-aware Dynamic Rank Selection for SVD-based LLM Compression

DGX agent

arXiv:2605.08568v1 Announce Type: new Abstract: Large language models (LLMs) have rapidly grown in scale, creating substantial memory and computational costs that hinder efficient deployment. Singular

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Diffusion Models are Evolutionary Algorithms

DGX agent

arXiv:2410.02543v3 Announce Type: replace-cross Abstract: In a convergence of machine learning and biology, we reveal that diffusion models are evolutionary algorithms. By considering evolution as a d

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Direct Bethe Free Energy Minimization for Bayesian Neural Ne twork

DGX agent

arXiv:2605.08446v1 Announce Type: new Abstract: We propose training Bayesian neural networks by directly minimizing the Bethe free energy rather than maximizing a variational lower bound. On tree-stru

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Distributional Spectral Diagnostics for Localizing Grokking Transitions

DGX agent

arXiv:2605.08237v1 Announce Type: new Abstract: In grokking, a model first fits the training data while test accuracy remains low, and only later begins to generalize. We ask whether this transition c

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment

DGX agent

arXiv:2605.08462v1 Announce Type: cross Abstract: Hallucination remains a persistent challenge in Large Language Models (LLMs), particularly in context-grounded settings such as RAG and agentic AI sys

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Do Foundation Model Embeddings Improve Cross-Country Crop Yield Generalisation? A Leave-One-Country-Out Evaluation in Sub-Saharan Africa

DGX agent

arXiv:2605.08113v1 Announce Type: cross Abstract: Accurate predictions of smallholder maize yields across national boundaries are critical for food security planning in sub-Saharan Africa, yet most pu

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Do not copy and paste! Rewriting strategies for code retrieval

DGX agent

arXiv:2605.08299v1 Announce Type: cross Abstract: Embedding-based code retrieval often suffers when encoders overfit to surface syntax. Prior work mitigates this by using LLMs to rephrase queries and

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation

DGX agent

arXiv:2605.09315v1 Announce Type: new Abstract: Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and mai

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding

DGX agent

arXiv:2605.08888v1 Announce Type: new Abstract: Evaluating whether Multimodal Large Language Models can produce trustworthy, verifiable reasoning over long, visually rich documents requires evaluation

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents

DGX agent

arXiv:2605.08747v1 Announce Type: new Abstract: Standard embodied evaluations do not independently score whether an agent correctly commits to task completion at episode closure, a capacity we call te

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces

DGX agent

arXiv:2605.09497v1 Announce Type: new Abstract: Vision-language model (VLM) based web agents demonstrate impressive autonomous GUI interaction but remain vulnerable to deceptive interface elements. Ex

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Don't Retrieve, Generate: Prompting LLMs for Synthetic Training Data in Dense Retrieval

DGX agent

arXiv:2504.21015v4 Announce Type: replace-cross Abstract: Training effective dense retrieval models typically relies on hard negative (HN) examples mined from large document corpora using methods such

model-releasesarxiv-cs-cl
12 May 2026
← Previous
1…252253254255256…361
Next →