AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Research

Evaluation of Prompt Injection Defenses in Large Language Models

DGX agent

arXiv:2604.23887v1 Announce Type: cross Abstract: LLM-powered applications routinely embed secrets in system prompts, yet models can be tricked into revealing them. We built an adaptive attacker that

researcharxiv-cs-ai
28 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency

DGX agent

arXiv:2501.02673v4 Announce Type: replace Abstract: Having a sufficient quantity of quality data is a critical enabler of training effective machine learning models. Being able to effectively determin

researcharxiv-cs-lg
28 Apr 2026
Research

Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting

DGX agent

arXiv:2506.19089v5 Announce Type: replace-cross Abstract: We introduce StorySim, a programmable framework for synthetically generating stories to evaluate the theory of mind (ToM) and world modeling (

researcharxiv-cs-ai
28 Apr 2026
Tutorials

Representational Curvature Modulates Behavioral Uncertainty in Large Language Models

DGX agent

arXiv:2604.23985v1 Announce Type: new Abstract: In autoregressive large language models (LLMs), temporal straightening offers an account of how the next-token prediction objective shapes representatio

tutorialsarxiv-cs-ai
28 Apr 2026
Research

SGP-SAM: Self-Gated Prompting for Transferring 3D Segment Anything Models to Lesion Segmentation

DGX agent

arXiv:2604.22825v1 Announce Type: cross Abstract: Large segmentation foundation models such as the Segment Anything Model (SAM) have reshaped promptable segmentation in natural images, and recent effo

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

DGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs

DGX agent

arXiv:2604.23356v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical diagnosis, but real-world deployment remains challenging due to high-stakes clinical decisions and

applicationsarxiv-cs-cl
28 Apr 2026
Tutorials

Are Natural-Domain Foundation Models Effective for Accelerated Cardiac MRI Reconstruction?

DGX agent

arXiv:2604.22557v1 Announce Type: cross Abstract: The emergence of large-scale pretrained foundation models has transformed computer vision, enabling strong performance across diverse downstream tasks

tutorialsarxiv-cs-cv
27 Apr 2026
Model Releases

Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries

DGX agent

arXiv:2603.28258v2 Announce Type: replace-cross Abstract: Categorical perception (CP) -- enhanced discriminability at category boundaries -- is among the most studied phenomena in perceptual psycholog

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

FETS Benchmark: Foundation Models Outperform Dataset-specific Machine Learning in Energy Time Series Forecasting

DGX agent

arXiv:2604.22328v1 Announce Type: cross Abstract: Driven by the transition towards a climate-neutral energy system, accurate energy time series forecasting is critical for planning and operation. Yet,

model-releasesarxiv-cs-ai
27 Apr 2026
Agents

AgenticQwen: Training Small Agentic Language Models with Dual Data Flywheels for Industrial-Scale Tool Use

DGX agent

arXiv:2604.21590v1 Announce Type: new Abstract: Modern industrial applications increasingly demand language models that act as agents, capable of multi-step reasoning and tool use in real-world settin

agentsarxiv-cs-cl
24 Apr 2026
Applications

Behavioral Consistency and Transparency Analysis on Large Language Model API Gateways

DGX agent

arXiv:2604.21083v1 Announce Type: cross Abstract: Third-party Large Language Model (LLM) API gateways are rapidly emerging as unified access points to models offered by multiple vendors. However, the

applicationsarxiv-cs-ai
24 Apr 2026
Research

On the Relationship between Bayesian Networks and Probabilistic Structural Causal Models

DGX agent

arXiv:2603.27406v2 Announce Type: replace Abstract: In this paper, the relationship between probabilistic graphical models, in particular Bayesian networks, and causal diagrams, also called structural

researcharxiv-cs-ai
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

AI models of unstable flow exhibit hallucination

DGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

safetyarxiv-cs-ai
23 Apr 2026
Applications

Cross-Modal Taxonomic Generalization in (Vision-) Language Models

DGX agent

arXiv:2603.07474v2 Announce Type: replace-cross Abstract: What is the interplay between semantic representations learned by language models (LM) from surface form alone to those learned from more grou

applicationsarxiv-cs-ai
23 Apr 2026
Applications

Explainability in Generative Medical Diffusion Models: A Faithfulness-Based Analysis on MRI Synthesis

DGX agent

arXiv:2602.09781v2 Announce Type: replace-cross Abstract: This study investigates the explainability of generative diffusion models in the context of medical imaging, focusing on Magnetic resonance im

applicationsarxiv-cs-ai
23 Apr 2026
Tutorials

Local Diffusion Models and Phases of Data Distributions

DGX agent

arXiv:2508.06614v2 Announce Type: replace Abstract: As a class of generative artificial intelligence frameworks inspired by statistical physics, diffusion models have shown extraordinary performance i

tutorialsarxiv-cs-lg
23 Apr 2026
Safety

Membership Inference for Contrastive Pre-training Models with Text-only PII Queries

DGX agent

arXiv:2603.14222v2 Announce Type: replace-cross Abstract: Contrastive pretraining models such as CLIP and CLAP, serve as the ubiquitous perceptual backbones for modern multimodal large models, yet the

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

DGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

GRASPrune: Global Gating for Budgeted Structured Pruning of Large Language Models

DGX agent

arXiv:2604.19398v1 Announce Type: new Abstract: Large language models (LLMs) are expensive to serve because model parameters, attention computation, and KV caches impose substantial memory and latency

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models

DGX agent

arXiv:2604.19300v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have recently achieved strong performance across various audio-centric tasks. However, hallucination, where models

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights

DGX agent

arXiv:2510.04800v3 Announce Type: replace Abstract: Recent progress in large language models demonstrates that hybrid architectures--combining self-attention mechanisms with structured state space mod

researcharxiv-cs-cl
22 Apr 2026
Tutorials

Learning Lifted Action Models from Unsupervised Visual Traces

DGX agent

arXiv:2604.19043v1 Announce Type: new Abstract: Efficient construction of models capturing the preconditions and effects of actions is essential for applying AI planning in real-world domains. Extensi

tutorialsarxiv-cs-ai
22 Apr 2026
Model Releases

LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification

DGX agent

arXiv:2604.18878v1 Announce Type: new Abstract: We introduce LegalBench-BR, the first public benchmark for evaluating language models on Brazilian legal text classification. The dataset comprises 3,10

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit

DGX agent

arXiv:2510.06133v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) generate text through iterative denoising. In commonly adopted parallel decoding schemes, each step confirms

model-releasesarxiv-cs-cl
21 Apr 2026
Research

CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models

DGX agent

arXiv:2604.16363v1 Announce Type: cross Abstract: Text-to-image models are commercially valuable assets often distributed under restrictive licenses, but such licenses are enforceable only when violat

researcharxiv-cs-cv
21 Apr 2026
Model Releases

D-QRELO: Training- and Data-Free Delta Compression for Large Language Models via Quantization and Residual Low-Rank Approximation

DGX agent

arXiv:2604.16940v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) accelerates taskspecific large language models (LLMs) development, but the resulting proliferation of finetuned models incu

model-releasesarxiv-cs-lg
21 Apr 2026
Research

DGSSM: Diffusion guided state-space models for multimodal salient object detection

DGX agent

arXiv:2604.17585v1 Announce Type: new Abstract: Salient object detection (SOD) requires modeling both long-range contextual dependencies and fine-grained structural details, which remains challenging

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Evaluating Multimodal LLMs for Inpatient Diagnosis: Real-World Performance, Safety, and Cost Across Ten Frontier Models

DGX agent

arXiv:2604.16980v1 Announce Type: new Abstract: Background: Large language models (LLMs) are increasingly proposed for diagnostic support, but few evaluations use real-world multimodal inpatient data,

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Generalization Boundaries of Fine-Tuned Small Language Models for Graph Structural Inference

DGX agent

arXiv:2604.18092v1 Announce Type: new Abstract: Small language models fine-tuned for graph property estimation have demonstrated strong in-distribution performance, yet their generalization capabiliti

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

MultiWorld: Scalable Multi-Agent Multi-View Video World Models

DGX agent

arXiv:2604.18564v1 Announce Type: new Abstract: Video world models have achieved remarkable success in simulating environmental dynamics in response to actions by users or agents. They are modeled as

agentsarxiv-cs-cv
21 Apr 2026
Research

New Fourth-Order Grayscale Indicator-Based Telegraph Diffusion Model for Image Despeckling

DGX agent

arXiv:2509.26010v2 Announce Type: replace Abstract: Second-order PDE models have been widely used for suppressing multiplicative noise, but they often introduce blocky artifacts in the early stages of

researcharxiv-cs-cv
21 Apr 2026
Safety

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL

DGX agent

arXiv:2604.17706v1 Announce Type: new Abstract: Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception,

safetyarxiv-cs-ro
21 Apr 2026
Safety

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling

DGX agent

arXiv:2510.24235v3 Announce Type: replace Abstract: Reward models (RMs) are central to reinforcement learning from human feedback (RLHF), providing the critical supervision signals that align large la

safetyarxiv-cs-lg
21 Apr 2026
Safety

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction

DGX agent

arXiv:2506.01770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved tremendous success in various tasks, yet concerns about their safety and security have emerged. In

safetyarxiv-cs-lg
21 Apr 2026
Research

Rethinking Post-Unlearning Behavior of Large Vision-Language Models

DGX agent

arXiv:2506.02541v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) can recognize individuals in images and disclose sensitive personal information about them, raising criti

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Robust Bias Evaluation with FilBBQ: A Filipino Bias Benchmark for Question-Answering Language Models

DGX agent

arXiv:2602.14466v2 Announce Type: replace Abstract: With natural language generation becoming a popular use case for language models, the Bias Benchmark for Question-Answering (BBQ) has grown to be an

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Sonata: A Hybrid World Model for Inertial Kinematics under Clinical Data Scarcity

DGX agent

arXiv:2604.18058v1 Announce Type: new Abstract: We introduce Sonata, a compact latent world model for six-axis trunk IMU representation learning under clinical data scarcity. Clinical cohorts typicall

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

(Sparse) Attention to the Details: Preserving Spectral Fidelity in ML-based Weather Forecasting Models

DGX agent

arXiv:2604.16429v1 Announce Type: cross Abstract: We introduce Mosaic, a probabilistic weather forecasting model that addresses two principal sources of spectral degradation in ML-based weather predic

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

DGX agent

arXiv:2604.18518v1 Announce Type: new Abstract: Uniform Discrete Diffusion Model (UDM) has recently emerged as a promising paradigm for discrete generative modeling; however, its integration with rein

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators

DGX agent

arXiv:2602.19946v4 Announce Type: replace Abstract: Recent text-to-image (T2I) diffusion models produce visually stunning images and demonstrate excellent prompt following. But do they perform well as

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ECG-Lens: Benchmarking ML & DL Models on PTB-XL Dataset

DGX agent

arXiv:2604.15822v1 Announce Type: cross Abstract: Automated classification of electrocardiogram (ECG) signals is a useful tool for diagnosing and monitoring cardiovascular diseases. This study compare

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Efficient Video Diffusion Models: Advancements and Challenges

DGX agent

arXiv:2604.15911v1 Announce Type: new Abstract: Video diffusion models have rapidly become the dominant paradigm for high-fidelity generative video synthesis, but their practical deployment remains co

applicationsarxiv-cs-cv
20 Apr 2026
Research

Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data

DGX agent

arXiv:2604.15380v1 Announce Type: cross Abstract: We present an exascale workflow for materials discovery using atomistic graph foundation models built on HydraGNN. We jointly train on 16 open first-p

researcharxiv-cs-ai
20 Apr 2026
Model Releases

neuralCAD-Edit: An Expert Benchmark for Multimodal-Instructed 3D CAD Model Editing

DGX agent

arXiv:2604.16170v1 Announce Type: new Abstract: We introduce neuralCAD-Edit, the first benchmark for editing 3D CAD models collected from expert CAD engineers. Instead of text conditioning as in prior

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

TabularMath: Understanding Math Reasoning over Tables with Large Language Models

DGX agent

arXiv:2505.19563v4 Announce Type: replace Abstract: Mathematical reasoning has long been a key benchmark for evaluating large language models. Although substantial progress has been made on math word

model-releasesarxiv-cs-ai
20 Apr 2026
Research

TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models

DGX agent

arXiv:2604.15756v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP exhibit strong Out-of-distribution (OOD) detection capabilities by aligning visual and textual representation

researcharxiv-cs-cl
20 Apr 2026
← Previous
1…3940414243…1021
Next →