AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Can Large Language Models Really Recognize Your Name?

DGX agent

arXiv:2505.14549v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used in privacy pipelines to detect and remedy sensitive data leakage. These solutions oft

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Can LLMs Act as Historians? Evaluating Historical Research Capabilities of LLMs via the Chinese Imperial Examination

DGX agent

arXiv:2604.24690v1 Announce Type: new Abstract: While Large Language Models (LLMs) have increasingly assisted in historical tasks such as text processing, their capacity for professional-level histori

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Can Multimodal Large Language Models Truly Understand Small Objects?

DGX agent

arXiv:2604.22884v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown promising potential in diverse understanding tasks, e.g., image and video analysis, math and physi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Cataract-LMM Large-Scale Multi-Source Multi-Task Benchmark for Deep Learning in Surgical Video Analysis

DGX agent

arXiv:2510.16371v2 Announce Type: replace-cross Abstract: The development of computer-assisted surgery systems relies on large-scale, annotated datasets. Existing cataract surgery resources lack the d

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Causal Discovery as Dialectical Aggregation: A Quantitative Argumentation Framework

DGX agent

arXiv:2604.23633v1 Announce Type: new Abstract: Constraint-based causal discovery is brittle in finite-sample regimes because erroneous conditional-independence (CI) decisions can cascade into substan

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics

DGX agent

arXiv:2509.20374v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong performance across general NLP tasks, but their utility in automating numerical experime

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Channel Adaptation for EEG Foundation Models: A Systematic Benchmark Across Architectures, Tasks, and Training Regimes

DGX agent

arXiv:2604.23091v1 Announce Type: new Abstract: Scaling EEG foundation models requires pooling data across heterogeneous electrode montages, a prerequisite both for larger pretraining corpora and for

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads

DGX agent

arXiv:2604.23009v1 Announce Type: new Abstract: Job Skill Named Entity Recognition (JobSkillNER) aims to automatically extract key skill information from large-scale job posting data, which is importa

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Citation-Driven Multi-View Training for Patent Embeddings: QaECTER and Sophia-Bench

DGX agent

arXiv:2604.22897v1 Announce Type: cross Abstract: Patent retrieval underpins critical decisions in innovation, examination, and IP strategy, yet progress has been hampered by the absence of benchmarks

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

DGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents

DGX agent

arXiv:2604.23781v1 Announce Type: new Abstract: Language-model agents are increasingly used as persistent coworkers that assist users across multiple working days. During such workflows, the surroundi

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ClawTrace: Cost-Aware Tracing for LLM Agent Skill Distillation

DGX agent

arXiv:2604.23853v1 Announce Type: new Abstract: Skill-distillation pipelines learn reusable rules from LLM agent trajectories, but they lack a key signal: how much each step costs. Without per-step co

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

DGX agent

arXiv:2604.16922v2 Announce Type: replace Abstract: Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of an

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation

DGX agent

arXiv:2510.22609v2 Announce Type: replace Abstract: Accurate symptom-to-disease classification and clinically grounded treatment recommendations remain challenging, particularly in heterogeneous patie

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Clotho: Measuring Task-Specific Pre-Generation Test Adequacy for LLM Inputs

DGX agent

arXiv:2509.17314v3 Announce Type: replace-cross Abstract: Software increasingly relies on the emergent capabilities of Large Language Models (LLMs), from natural language understanding to program anal

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Comparative Study of Weighted and Coupled Second- and Fourth-Order PDEs for Image Despeckling in Grayscale, Color, SAR, and Ultrasound

DGX agent

arXiv:2604.23612v1 Announce Type: new Abstract: Partial Differential Equation (PDE)-based approaches have gained significant attention in image despeckling due to their strong capability to preserve s

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows

DGX agent

arXiv:2604.22820v1 Announce Type: cross Abstract: Long-horizon tool-using tasks sometimes benefit from revisiting earlier subtasks for recovery and exploration, but added multi-agent workflow flexibil

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

DGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ComplianceNLP: Knowledge-Graph-Augmented RAG for Multi-Framework Regulatory Gap Detection

DGX agent

arXiv:2604.23585v1 Announce Type: new Abstract: Financial institutions must track over 60,000 regulatory events annually, overwhelming manual compliance teams; the industry has paid over USD 300 billi

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Constraint-Based Analysis of Reasoning Shortcuts in Neurosymbolic Learning

DGX agent

arXiv:2604.23377v1 Announce Type: new Abstract: Neurosymbolic systems can satisfy logical constraints during learning without achieving the intended concept-label correspondence; this is a problem kno

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

DGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Cortex-Inspired Continual Learning: Unsupervised Instantiation and Recovery of Functional Task Networks

DGX agent

arXiv:2604.24637v1 Announce Type: cross Abstract: Block-sequential continual learning demands that a single model both protect prior solutions from catastrophic forgetting and efficiently infer at inf

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Coverage-Based Calibration for Post-Training Quantization via Weighted Set Cover over Outlier Channels

DGX agent

arXiv:2604.24008v1 Announce Type: new Abstract: Post-Training Quantization (PTQ) compresses large language models to low bit-widths using a small calibration set, and its quality depends strongly on w

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

DGX agent

arXiv:2508.13650v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preser

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks

DGX agent

arXiv:2510.17687v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong reasoning and perception capabilities but are increasingly vulnerable to jailbreak att

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CT-FineBench: A Diagnostic Fidelity Benchmark for Fine-Grained Evaluation of CT Report Generation

DGX agent

arXiv:2604.24001v1 Announce Type: new Abstract: The evaluation of generated reports remains a critical challenge in Computed Tomography (CT) report generation, due to the large volume of text, the div

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CUB: Benchmarking Context Utilisation Techniques for Language Models

DGX agent

arXiv:2505.16518v3 Announce Type: replace-cross Abstract: Incorporating external knowledge is crucial for knowledge-intensive tasks, such as question answering and fact checking. However, language mod

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DARC-CLIP: Dynamic Adaptive Refinement with Cross-Attention for Meme Understanding

DGX agent

arXiv:2604.23214v1 Announce Type: new Abstract: Memes convey meaning through the interaction of visual and textual signals, often combining humor, irony, and offense in subtle ways. Detecting harmful

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting

DGX agent

arXiv:2604.23968v1 Announce Type: cross Abstract: Accurate time series forecasting in scientific domains such as climate modeling, physiological monitoring, and energy systems benefits from both compe

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Deep Learning-Enabled Dissolved Oxygen Sensing in Biofouling Environments for Ocean Monitoring

DGX agent

arXiv:2604.24236v1 Announce Type: cross Abstract: The escalating climate crisis and ecosystem degradation demand intelligent, low-cost sensors capable of robust, long-term monitoring in real-world env

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DeepImagine: Learning Biomedical Reasoning via Successive Counterfactual Imagining

DGX agent

arXiv:2604.23054v1 Announce Type: cross Abstract: Predicting the outcomes of prospective clinical trials remains a major challenge for large language models. Prior work has shown that both traditional

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DeepTaxon: An Interpretable Retrieval-Augmented Multimodal Framework for Unified Species Identification and Discovery

DGX agent

arXiv:2604.24029v1 Announce Type: cross Abstract: Identifying species in biology among tens of thousands of visually similar taxa while discovering unknown species in open-world environments remains a

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis

DGX agent

arXiv:2604.24703v1 Announce Type: cross Abstract: Large language models are widely used for code generation, yet they rely on an implicit assumption that the task descriptions are sufficiently detaile

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Defusing the Trigger: Plug-and-Play Defense for Backdoored LLMs via Tail-Risk Intrinsic Geometric Smoothing

DGX agent

arXiv:2604.24162v1 Announce Type: cross Abstract: Defending against backdoor attacks in large language models remains a critical practical challenge. Existing defenses mitigate these threats but typic

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Deploy DINO with Many-to-Many Association

DGX agent

arXiv:2604.23670v1 Announce Type: new Abstract: Motivated by the limited generalization of supervised image matching models to unseen image domains, we explore the zero-shot deployment of DINO feature

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models

DGX agent

arXiv:2406.10185v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) are increasingly integral to healthcare applications, including medical visual question answering and imaging r

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

DGHMesh: A Large-scale Dual-radar mmWave Dataset and Generalization-focused Benchmark for Human Mesh Reconstruction

DGX agent

arXiv:2604.22827v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar has shown great potential for contactless, privacy-preserving, and robust human sensing, yet existing mmWave-based human

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Differentiable Faithfulness Alignment for Cross-Model Circuit Transfer

DGX agent

arXiv:2604.24302v1 Announce Type: new Abstract: Mechanistic interpretability has made it possible to localize circuits underlying specific behaviors in language models, but existing methods are expens

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion

DGX agent

arXiv:2604.24351v1 Announce Type: cross Abstract: Controllable diffusion methods have substantially expanded the practical utility of diffusion models, but they are typically developed as isolated, ba

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B

DGX agent

arXiv:2604.24070v1 Announce Type: cross Abstract: Small instruct-tuned LLMs produce degenerate verbal confidence under minimal elicitation: ceiling rates above 95%, near-chance Type-2 AUROC, and Inval

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models

DGX agent

arXiv:2604.22822v1 Announce Type: cross Abstract: Object level hallucination remains a central reliability challenge for vision language models (VLMs), particularly in binary object existence verifica

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Do Quantum Transformers Help? A Systematic VQC Architecture Comparison on Tabular Benchmarks

DGX agent

arXiv:2604.23931v1 Announce Type: cross Abstract: Variational quantum circuits (VQCs) are a leading approach to quantum machine learning on near-term devices, yet it remains unclear which circuit arch

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Domain-Adapted Fine-Tuning of ECG Foundation Models for Multi-Label Structural Heart Disease Screening

DGX agent

arXiv:2604.23385v1 Announce Type: new Abstract: Transthoracic echocardiography is the reference standard for confirming structural heart disease (SHD), but first-line screening is limited by cost, wor

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Domain Fine-Tuning vs. Retrieval-Augmented Generation for Medical Multiple-Choice Question Answering: A Controlled Comparison at the 4B-Parameter Scale

DGX agent

arXiv:2604.23801v1 Announce Type: new Abstract: Practitioners deploying small open-weight large language models (LLMs) for medical question answering face a recurring design choice: invest in a domain

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Don't Make the LLM Read the Graph: Make the Graph Think

DGX agent

arXiv:2604.23057v1 Announce Type: new Abstract: We investigate whether explicit belief graphs improve LLM performance in cooperative multi-agent reasoning. Through 3,000+ controlled trials across four

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Don't Pause! Every prediction matters in a streaming video

DGX agent

arXiv:2604.24317v1 Announce Type: new Abstract: Streaming video models should respond the moment an event unfolds, not after the moment has passed. Yet existing online VideoQA benchmarks remain largel

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

DreamAudio: Customized Text-to-Audio Generation with Diffusion Models

DGX agent

arXiv:2509.06027v3 Announce Type: replace-cross Abstract: With the development of large-scale diffusion-based and language-modeling-based generative models, impressive progress has been achieved in te

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DRIFT: Transferring Reasoning Priors for Efficient MLLM Fine-Tuning

DGX agent

arXiv:2510.15050v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made rapid progress, yet their reasoning ability often lags behind strong text-only LLMs. Bridging thi

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…292293294295296…357
Next →