AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

UWM-JEPA: Predictive World Models That Imagine in Belief Space

DGX agent

arXiv:2605.25313v1 Announce Type: cross Abstract: World models for partially observed environments must imagine multiple compatible hidden futures and steer between them under counterfactual actions.

model-releasesarxiv-cs-ai
26 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

DGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

model-releasesarxiv-cs-cl
25 May 2026
Applications

Evaluating Customized vs. Generalist Transformer-based Models for Legal Contract Classification

DGX agent

arXiv:2508.07849v2 Announce Type: replace Abstract: Despite advances in legal NLP, no comprehensive evaluation of Transformer-based models customized for legal tasks (referred to as `legal-specific' m

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

DGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

model-releasesarxiv-cs-cv
25 May 2026
Research

InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion

DGX agent

arXiv:2505.13893v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have intensified efforts to fuse heterogeneous open-source models into a unified system that inherit

researcharxiv-cs-cl
25 May 2026
Model Releases

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

DGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

model-releasesarxiv-cs-ai
25 May 2026
Safety

The physics of AI weather models

DGX agent

arXiv:2605.23778v1 Announce Type: cross Abstract: Could it be that AI weather models are solving physical equations, although they may not be the equations used by conventional NWP models? We compute

safetyarxiv-cs-lg
25 May 2026
Model Releases

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models

DGX agent

arXiv:2605.22870v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting is necessary for arithmetic in small language models, yet shuffling its steps preserves most performance. What does C

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

DGX agent

arXiv:2601.21306v2 Announce Type: replace-cross Abstract: This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compoundin

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

DGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

model-releasesarxiv-cs-lg
23 May 2026
Safety

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

DGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

safetyarxiv-cs-lg
23 May 2026
Model Releases

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

DGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

model-releasesarxiv-cs-lg
23 May 2026
Research

Accelerated Test-Time Scaling with Model-Free Speculative Sampling

DGX agent

arXiv:2506.04708v3 Announce Type: replace Abstract: Language models have demonstrated remarkable capabilities in reasoning tasks through test-time scaling techniques like best-of-N sampling and tree s

researcharxiv-cs-cl
22 May 2026
Model Releases

BEiTScore: Reference-free Image Captioning Evaluation with an Efficient Cross-Encoder Model

DGX agent

arXiv:2605.21728v1 Announce Type: cross Abstract: Image captioning evaluation remains a significant challenge, as vision-language models evolve toward more challenging capabilities such as generating

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

DGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing

DGX agent

arXiv:2602.01851v2 Announce Type: replace Abstract: Recent generative models have achieved remarkable progress in image editing. However, existing systems and benchmarks remain largely text-guided. In

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

DGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model

DGX agent

arXiv:2605.22089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework for end-to-end autonomous driving. However, existing VLAs typically rely on sp

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System

DGX agent

arXiv:2605.20368v1 Announce Type: cross Abstract: Organizations that scan documents for sensitive information face a practical problem. Cloud services require data to be sent to external infrastructur

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Names

DGX agent

arXiv:2511.04106v5 Announce Type: replace-cross Abstract: The diffusion of ideas and language in society has conventionally been described by S-shaped models, such as the logistic curve. However, the

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

DGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

model-releasesarxiv-cs-cl
22 May 2026
Research

The Neglected Baseline in Model Interpretation

DGX agent

arXiv:2605.22417v1 Announce Type: new Abstract: We observe that existing model interpretation methods generally ignore the baseline, and such neglect often results in imprecise or even incorrect inter

researcharxiv-cs-cv
22 May 2026
Model Releases

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

DGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Visual-Advantage On-Policy Distillation for Vision-Language Models

DGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

model-releasesarxiv-cs-cv
22 May 2026
Research

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

DGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

researcharxiv-cs-cl
22 May 2026
Model Releases

An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees

DGX agent

arXiv:2605.20521v1 Announce Type: new Abstract: Fine-tuning adapts a pretrained machine learning model to a small, sensitive dataset, but this process risks memorizing individual new data points, maki

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models

DGX agent

arXiv:2605.20777v1 Announce Type: new Abstract: Visual storytelling with diffusion models has made impressive strides in maintaining character consistency across narrative scenes. However, a critical

model-releasesarxiv-cs-cv
21 May 2026
Safety

Bayesian Preference Learning for Test-Time Steerable Reward Models

DGX agent

arXiv:2602.08819v2 Announce Type: replace-cross Abstract: Reward models are central to aligning language models with human preferences via reinforcement learning (RL). As RL is increasingly applied to

safetyarxiv-cs-cl
21 May 2026
Safety

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

DGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

safetyarxiv-cs-cl
21 May 2026
Model Releases

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

DGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

DGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Optimization Hyper-parameter Laws for Large Language Models

DGX agent

arXiv:2409.04777v4 Announce Type: replace Abstract: Large Language Models have driven significant AI advancements, yet their training is resource-intensive and highly sensitive to hyper-parameter sele

model-releasesarxiv-cs-lg
21 May 2026
Applications

Towards the Anonymization of the Language Modeling

DGX agent

arXiv:2501.02407v3 Announce Type: replace Abstract: Rapid advances in Natural Language Processing (NLP) have revolutionized many fields, including healthcare. However, these advances raise significant

applicationsarxiv-cs-cl
21 May 2026
Model Releases

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

DGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents

DGX agent

arXiv:2605.20306v1 Announce Type: new Abstract: We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection

DGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

DGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Backdooring Masked Diffusion Language Models

DGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model

DGX agent

arXiv:2602.23622v2 Announce Type: replace-cross Abstract: Significant progress has been made in the field of Instruction-based Image Editing Models (IIEMs). However, while these models demonstrate pla

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Entry-level guide to the use of large language models for medical research

DGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

model-releasesarxiv-cs-ai
20 May 2026
Research

Neural Network Models for Contextual Regression

DGX agent

arXiv:2603.24400v2 Announce Type: replace-cross Abstract: We propose a neural network model for contextual regression in which the regression model depends on contextual features that determine the ac

researcharxiv-cs-lg
20 May 2026
Safety

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

DGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

safetyarxiv-cs-ai
20 May 2026
Research

Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models

DGX agent

arXiv:2605.19227v1 Announce Type: cross Abstract: Unified autoregressive models (UAMs) are transformer models that generate text as well as image tokens within a single autoregressive pass. Shared par

researcharxiv-cs-ai
20 May 2026
Model Releases

ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models

DGX agent

arXiv:2605.18879v1 Announce Type: cross Abstract: Large language models inevitably retain sensitive information, defined as inputs that may induce harmful generations, due to training on massive web c

model-releasesarxiv-cs-ai
20 May 2026
Research

Better Together: Evaluating the Complementarity of Earth Embedding Models

DGX agent

arXiv:2605.18667v1 Announce Type: new Abstract: Earth embedding models transform Earth observation data into embeddings uniquely tied to locations on the Earth's surface. These models are typically ev

researcharxiv-cs-cv
19 May 2026
Model Releases

CarbonScaling: Extending Neural Scaling Laws for Carbon Footprint in Large Language Models

DGX agent

arXiv:2508.06524v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly follow neural scaling laws that tie performance gains to rapidly expanding computational budgets, ra

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

DGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study

DGX agent

arXiv:2602.12015v2 Announce Type: replace Abstract: Deploying large language models for clinical Text-to-SQL requires distinguishing two qualitatively different causes of output diversity: (i) input a

model-releasesarxiv-cs-cl
19 May 2026
← Previous
1…4950515253…1021
Next →