AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

CSMCIR: CoT-Enhanced Symmetric Alignment with Memory Bank for Composed Image Retrieval

DGX agent

arXiv:2601.03728v3 Announce Type: replace-cross Abstract: Composed Image Retrieval (CIR) enables users to search for target images using both a reference image and manipulation text, offering substant

model-releasesarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CSR: Infinite-Horizon Real-Time Policies with Massive Cached State Representations

DGX agent

arXiv:2605.07325v1 Announce Type: cross Abstract: Deploying massive large language models (LLMs) as continuous cognitive engines for robotics is bottlenecked by the time-to-first-token (TTFT) latency

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Curvature Beyond Positivity: Greedy Guarantees for Arbitrary Submodular Functions

DGX agent

arXiv:2605.07902v1 Announce Type: new Abstract: Submodular functions -- functions exhibiting diminishing returns -- are central to machine learning. When the objective is monotone and non-negative, th

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

CyBiasBench: Benchmarking Bias in LLM Agents for Cyber-Attack Scenarios

DGX agent

arXiv:2605.07830v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in offensive cybersecurity. In this paper, we reveal an interesting phenom

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Data Contamination in Neural Hieroglyphic Translation: A Reproducibility Study

DGX agent

arXiv:2605.07453v1 Announce Type: new Abstract: Ancient and endangered languages pose a unique challenge for NLP: their datasets are inherently scarce, difficult to expand, and built from formulaic co

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Dataset Watermarking for Closed LLMs with Provable Detection

DGX agent

arXiv:2605.06865v1 Announce Type: new Abstract: Large language models (LLMs) are pre-trained and post-trained on vast amounts of loosely curated data, raising the possibility that these models may hav

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Decoding the Pulse of Reasoning VLMs in Multi-Image Understanding Tasks

DGX agent

arXiv:2603.04676v2 Announce Type: replace-cross Abstract: Multi-image reasoning remains a significant challenge for vision-language models (VLMs). We investigate a previously overlooked phenomenon: du

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Delulu: A Verified Multi-Lingual Benchmark for Code Hallucination Detection in Fill-in-the-Middle Tasks

DGX agent

arXiv:2605.07024v1 Announce Type: new Abstract: Large Language Models for code generation frequently produce hallucinations in Fill-in-the-Middle (FIM) tasks -- plausible but incorrect completions suc

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Detecting Distillation Data from Reasoning Models

DGX agent

arXiv:2510.04850v3 Announce Type: replace-cross Abstract: Reasoning distillation has emerged as a prevailing paradigm for transferring reasoning capabilities from large reasoning models to small langu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation

DGX agent

arXiv:2508.20909v2 Announce Type: replace Abstract: Foundation models pre-trained on large-scale natural image datasets offer a powerful paradigm for medical image segmentation. However, effectively t

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Disagreement-Regularized Importance Sampling for Adversarial Label Corruption

DGX agent

arXiv:2605.07551v1 Announce Type: new Abstract: Standard Importance Sampling (IS) collapses under label corruption because high-norm examples, prioritized for variance reduction, are often adversarial

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative Evaluation

DGX agent

arXiv:2605.07323v1 Announce Type: new Abstract: Discovering governing differential equations from observational data is a fundamental challenge in scientific machine learning. Existing symbolic regres

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Divide and Conquer: Object Co-occurrence Helps Mitigate Simplicity Bias in OOD Detection

DGX agent

arXiv:2605.07821v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection is crucial for ensuring the reliability of deep learning models. Existing methods mostly focus on regular entangle

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

DKDS: A Benchmark Dataset of Degraded Kuzushiji Documents with Seals for Detection and Binarization

DGX agent

arXiv:2511.09117v4 Announce Type: replace Abstract: Kuzushiji, a pre-modern Japanese cursive script, can currently be read and understood by only a few thousand trained experts in Japan. With the rapi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Do Joint Audio-Video Generation Models Understand Physics?

DGX agent

arXiv:2605.07061v1 Announce Type: cross Abstract: Joint audio-video generation models are rapidly approaching professional production quality, raising a central question: do they understand audio-visu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas

DGX agent

arXiv:2605.06673v1 Announce Type: cross Abstract: Aggregate metacognitive quality scores mask within-model variation across MMLU benchmark domains. We administered 1,500 MMLU items (250 per domain, un

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain

DGX agent

arXiv:2605.07699v1 Announce Type: cross Abstract: LLM-based agents are increasingly deployed for routine but consequential tasks in real-world domains, where their behavior is governed by inherently a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

DT-PBO: an Interpretable Tree-based Surrogate Model for Preferential Bayesian Optimization

DGX agent

arXiv:2512.14263v2 Announce Type: replace-cross Abstract: Preferential Bayesian Optimization (PBO) aims to find a decision-maker's most preferred solution in as few pairwise comparisons as possible. E

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Echo: KV-Cache-Free Associative Recall with Spectral Koopman Operators

DGX agent

arXiv:2605.06997v1 Announce Type: new Abstract: Long chain-of-thought reasoning and agentic tool-calling produce traces spanning tens of thousands of tokens, yet Transformer KV caches grow linearly wi

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Effective and Memory-Efficient Alternatives to ECC for Reliable Large-Scale DNNs

DGX agent

arXiv:2605.07417v1 Announce Type: cross Abstract: Modern Deep Learning (DL) workloads are increasingly deployed in safety-critical domains, such as automotive systems and hyperscale data centers, wher

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Efficient Data Selection for Multimodal Models via Incremental Optimization Utility

DGX agent

arXiv:2605.07488v1 Announce Type: new Abstract: The scaling of Large Multimodal Models (LMMs) is constrained by the quality-quantity trade-off inherent in synthetic data. Previous approaches, such as

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

EgoPro-Bench: Benchmarking Personalized Proactive Interaction in Egocentric Video Streams

DGX agent

arXiv:2605.07299v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) remain primarily reactive, failing to continuously perceive environments or proactively assist users

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

End-to-end PDDL Planning with Hardcoded and Dynamic Agents

DGX agent

arXiv:2512.09629v2 Announce Type: replace Abstract: We present an end-to-end framework for planning supported by verifiers. An orchestrator receives a human specification written in natural language a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation

DGX agent

arXiv:2605.07247v1 Announce Type: new Abstract: Scalable AI agents training relies on interactive environments that faithfully simulate the consequences of agent actions. Manually crafted environments

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ESSAM: A Novel Competitive Evolution Strategies Approach to Reinforcement Learning for Memory Efficient LLMs Fine-Tuning

DGX agent

arXiv:2602.01003v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a key training step for improving mathematical reasoning in large language models (LLMs), but it often

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Evaluating Large Language Models in Scientific Discovery

DGX agent

arXiv:2512.15567v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs

DGX agent

arXiv:2605.06669v1 Announce Type: cross Abstract: Educational LLM tutors face a core AI alignment challenge: they must follow user intent while preserving pedagogical constraints and safety policies.

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Exact Flow Linear Attention: Exact Solution from Continuous-Time Dynamics

DGX agent

arXiv:2512.12602v4 Announce Type: replace Abstract: In this paper, we introduce Exact Flow Linear Attention~(EFLA), an exact-flow formulation of delta-rule linear attention. We show that the delta-rul

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

DGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Excluding the Target Domain Improves Extrapolation: Deconfounded Hierarchical Physics Constraints

DGX agent

arXiv:2605.07485v1 Announce Type: cross Abstract: Extrapolation to out-of-distribution conditions is a fundamental challenge for physics-constrained deep generative models. Existing methods apply phys

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

FactoryBench: Evaluating Industrial Machine Understanding

DGX agent

arXiv:2605.07675v1 Announce Type: new Abstract: We introduce FactoryBench, a benchmark for evaluating time-series models and LLMs on machine understanding over industrial robotic telemetry. Q&A pairs

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution

DGX agent

arXiv:2605.07208v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to brainstorm and evaluate research ideas, yet assessing such judgments is fundamentally difficult be

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

FANoS-v2: Feedback-Controlled Momentum with Thermostat Damping for Lightweight Neural Optimization

DGX agent

arXiv:2601.00889v2 Announce Type: replace Abstract: FANOS{} is a PyTorch optimizer that augments RMS-preconditioned momentum with a scalar feedback controller over update energy. The public reference

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Fast Byte Latent Transformer

DGX agent

arXiv:2605.08044v1 Announce Type: cross Abstract: Recent byte-level language models (LMs) match the performance of token-level models without relying on subword vocabularies, yet their utility is limi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

FastOmniTMAE: Parallel Clause Learning for Scalable and Hardware-Efficient Tsetlin Embeddings

DGX agent

arXiv:2605.06982v1 Announce Type: new Abstract: Embedding models in natural language processing (NLP) increasingly rely on deep architectures such as BERT, while simpler models such as Word2Vec provid

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Fidel-TS: A High-Fidelity Multimodal Benchmark for Time Series Forecasting

DGX agent

arXiv:2509.24789v4 Announce Type: replace Abstract: The evaluation of time series forecasting models is hindered by a lack of high-quality benchmarks, leading to overestimated assessments of progress.

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Fine-tuning a vision-language model for fracture-surface morphology recognition

DGX agent

arXiv:2605.07145v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong potential for scientific image understanding, but general-purpose models often lack the domain-specifi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting

DGX agent

arXiv:2603.19254v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in financial research workflows, where their role is evolving from single-model assistance fo

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Flat Channels to Infinity in Neural Loss Landscapes

DGX agent

arXiv:2506.14951v4 Announce Type: replace-cross Abstract: The loss landscapes of neural networks contain minima and saddle points that may be connected in flat regions or appear in isolation. We ident

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Frequency-Aware Model Parameter Explorer: A new attribution method for improving explainability

DGX agent

arXiv:2510.03245v2 Announce Type: replace-cross Abstract: State-of-the-art attribution methods rely on adversarial sample generation that applies an all-pass filter across the frequency spectrum, disc

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG

DGX agent

arXiv:2605.07273v1 Announce Type: cross Abstract: Multimodal RAG systems increasingly rely on vision-language retrievers to ground visual queries in external textual evidence. Existing adversarial stu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

From Pixels to Prompts: Vision-Language Models

DGX agent

arXiv:2605.07544v1 Announce Type: new Abstract: When you read a paper about a new Vision-Language Model today, it can be easy to forget how strange this idea would have sounded not so long ago. Teachi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

From Synthetic to Real: Toward Identity-Consistent Makeup Transfer with Synthetic and Real Data

DGX agent

arXiv:2605.07861v1 Announce Type: new Abstract: Makeup transfer aims to apply the makeup style of a reference portrait to a source portrait while preserving identity and background. Early methods form

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

GAD in the Wild: Benchmarking Graph Anomaly Detection under Realistic Deployment Challenges

DGX agent

arXiv:2605.07133v1 Announce Type: cross Abstract: Graph Anomaly Detection (GAD) is a critical task in graph machine learning with vital applications in financial fraud detection and social platform go

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Gated QKAN-FWP: Scalable Quantum-inspired Sequence Learning

DGX agent

arXiv:2605.06734v1 Announce Type: cross Abstract: Fast Weight Programmers (FWPs) encode temporal dependencies through dynamically updated parameters rather than recurrent hidden states. Quantum FWPs (

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

DGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GC-ART: Global Learnable Second-Order Rational Tone Curves for Illumination Robustness

DGX agent

arXiv:2605.07329v1 Announce Type: new Abstract: We introduce GC-ART (Global Curve Adaptive Rational Tone-mapping), a lightweight differentiable pre-processing module for robust image classification. G

model-releasesarxiv-cs-cv
11 May 2026
← Previous
1…264265266267268…361
Next →