AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models

DGX agent

arXiv:2604.15967v1 Announce Type: cross Abstract: Despite the remarkable synthesis capabilities of text-to-image (T2I) models, safeguarding them against content violations remains a persistent challen

model-releasesarxiv-cs-cv
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Contextuality from Single-State Ontological Models: An Information-Theoretic Obstruction

DGX agent

arXiv:2602.16716v3 Announce Type: replace Abstract: Contextuality is a central feature of quantum theory, traditionally understood as the impossibility of reproducing quantum measurement statistics us

researcharxiv-cs-ai
17 Apr 2026
Research

Edge-preserving noise for diffusion models

DGX agent

arXiv:2410.01540v4 Announce Type: replace Abstract: Classical diffusion models typically rely on isotropic Gaussian noise, treating all regions uniformly and overlooking structural information importa

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Fact4ac at the Financial Misinformation Detection Challenge Task: Reference-Free Financial Misinformation Detection via Fine-Tuning and Few-Shot Prompting of Large Language Models

DGX agent

arXiv:2604.14640v1 Announce Type: new Abstract: The proliferation of financial misinformation poses a severe threat to market stability and investor trust, misleading market behavior and creating crit

model-releasesarxiv-cs-cl
17 Apr 2026
Research

GUI-Perturbed: Domain Randomization Reveals Systematic Brittleness in GUI Grounding Models

DGX agent

arXiv:2604.14262v1 Announce Type: new Abstract: GUI grounding models report over 85% accuracy on standard benchmarks, yet drop 27-56 percentage points when instructions require spatial reasoning rathe

researcharxiv-cs-lg
17 Apr 2026
Agents

In-Context Autonomous Network Incident Response: An End-to-End Large Language Model Agent Approach

DGX agent

arXiv:2602.13156v2 Announce Type: replace-cross Abstract: Rapidly evolving cyberattacks demand incident response systems that can autonomously learn and adapt to changing threats. Prior work has exten

agentsarxiv-cs-ai
17 Apr 2026
Agents

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

DGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

agentsarxiv-cs-lg
17 Apr 2026
Safety

Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality

DGX agent

arXiv:2603.17512v4 Announce Type: replace Abstract: Large language models (LLMs) exhibit strong general intelligence, yet their multilingual performance remains highly imbalanced. Although LLMs encode

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems

DGX agent

arXiv:2601.14053v2 Announce Type: replace-cross Abstract: The field of artificial intelligence has undergone a revolution from foundational Transformer architectures to reasoning-capable systems appro

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

MetaDent: Labeling Clinical Images for Vision-Language Models in Dentistry

DGX agent

arXiv:2604.14866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated significant potential in medical image analysis, yet their application in intraoral photography remains

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models

DGX agent

arXiv:2604.14672v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used in urban planning, but since gendered space theory highlights how gender hierarchies are embedd

safetyarxiv-cs-cl
17 Apr 2026
Safety

Switch-KD: Visual-Switch Knowledge Distillation for Vision-Language Models

DGX agent

arXiv:2604.14629v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown remarkable capabilities in joint vision-language understanding, but their large scale poses significant challen

safetyarxiv-cs-cv
17 Apr 2026
Safety

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

DGX agent

arXiv:2604.13715v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) enable general audio understanding and demonstrate remarkable performance across various audio tasks. However, the

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

Weight Patching: Toward Source-Level Mechanistic Localization in LLMs

DGX agent

arXiv:2604.13694v1 Announce Type: new Abstract: Mechanistic interpretability seeks to localize model behavior to the internal components that causally realize it. Prior work has advanced activation-sp

model-releasesarxiv-cs-ai
17 Apr 2026
Tutorials

World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems

DGX agent

arXiv:2604.14732v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action.

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

Auto-FP: An Experimental Study of Automated Feature Preprocessing for Tabular Data

DGX agent

arXiv:2310.02540v2 Announce Type: replace Abstract: Classical machine learning models, such as linear models and tree-based models, are widely used in industry. These models are sensitive to data dist

model-releasesarxiv-cs-lg
16 Apr 2026
Safety

Beyond State Consistency: Behavior Consistency in Text-Based World Models

DGX agent

arXiv:2604.13824v1 Announce Type: new Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

DGX agent

arXiv:2604.13077v1 Announce Type: new Abstract: Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

DGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

F-Actor: Controllable Conversational Behaviour in Full-Duplex Models

DGX agent

arXiv:2601.11329v3 Announce Type: replace Abstract: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding

DGX agent

arXiv:2604.13540v1 Announce Type: new Abstract: Unified Multimodal Models (UMMs) aim to integrate visual understanding and generation within a single structure. However, these models exhibit a notable

researcharxiv-cs-cv
16 Apr 2026
Agents

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch

DGX agent

arXiv:2604.14029v1 Announce Type: new Abstract: While Large Multimodal Models (LMMs) demonstrate impressive visual perception, they remain epistemically constrained by their static parametric knowledg

agentsarxiv-cs-cv
16 Apr 2026
Model Releases

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction

DGX agent

arXiv:2603.07083v2 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) agents operating in high-dimensional observation spaces, such as Dreamer, rely on learning abstract repres

model-releasesarxiv-cs-lg
15 Apr 2026
Research

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning

DGX agent

arXiv:2409.06679v3 Announce Type: replace Abstract: Processing long contexts is increasingly important for Large Language Models (LLMs) in tasks like multi-turn dialogues, code generation, and documen

researcharxiv-cs-cl
15 Apr 2026
Tutorials

KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates

DGX agent

arXiv:2604.12397v1 Announce Type: new Abstract: Standard Large Language Model (LLM) pre-training typically treats corpora as flattened token sequences, often overlooking the real-world context that hu

tutorialsarxiv-cs-cl
15 Apr 2026
Model Releases

KumoRFM-2: Scaling Foundation Models for Relational Learning

DGX agent

arXiv:2604.12596v1 Announce Type: cross Abstract: We introduce KumoRFM-2, the next iteration of a pre-trained foundation model for relational data. KumoRFM-2 supports in-context learning as well as fi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Large Language Models are Powerful Electronic Health Record Encoders

DGX agent

arXiv:2502.17403v5 Announce Type: replace-cross Abstract: Electronic Health Records (EHRs) offer considerable potential for clinical prediction, but their complexity and heterogeneity challenge tradit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

League of LLMs: A Benchmark-Free Paradigm for Mutual Evaluation of Large Language Models

DGX agent

arXiv:2507.22359v4 Announce Type: replace Abstract: Although large language models (LLMs) have shown exceptional capabilities across a wide range of tasks, reliable evaluation remains a critical chall

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner

DGX agent

arXiv:2604.12668v1 Announce Type: new Abstract: The Diffusion Probabilistic Model (DPM) achieves remarkable performance in image generation, while its increasing parameter size and computational overh

model-releasesarxiv-cs-cv
15 Apr 2026
Research

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

DGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model

DGX agent

arXiv:2410.01473v2 Announce Type: replace Abstract: Soil sinkholes significantly influence soil degradation, infrastructure vulnerability, and landscape evolution. However, their irregular shapes, com

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

DGX agent

arXiv:2604.12617v1 Announce Type: cross Abstract: The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL)

safetyarxiv-cs-ai
15 Apr 2026
Local Ai

Towards Interpretable Foundation Models for Retinal Fundus Images

DGX agent

arXiv:2603.18846v2 Announce Type: replace Abstract: Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

A Compact and Efficient 1.251 Million Parameter Machine Learning CNN Model PD36-C for Plant Disease Detection: A Case Study

DGX agent

arXiv:2604.11332v1 Announce Type: cross Abstract: Deep learning has markedly advanced image based plant disease diagnosis as improved hardware and dataset quality have enabled increasingly accurate ne

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model

DGX agent

arXiv:2604.11490v1 Announce Type: new Abstract: While the field of vision-language (VL) has achieved remarkable success in integrating visual and textual information across multiple languages and doma

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

AOP-Smart: A RAG-Enhanced Large Language Model Framework for Adverse Outcome Pathway Analysis

DGX agent

arXiv:2604.10874v1 Announce Type: cross Abstract: Adverse Outcome Pathways (AOPs) are an important knowledge framework in toxicological research and risk assessment. In recent years, large language mo

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation

DGX agent

arXiv:2504.14988v3 Announce Type: replace Abstract: Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal perception capabilities, garnering significant a

model-releasesarxiv-cs-cv
14 Apr 2026
Research

COREY: A Prototype Study of Entropy-Guided Operator Fusion with Hadamard Reparameterization for Selective State Space Models

DGX agent

arXiv:2604.10597v1 Announce Type: cross Abstract: State Space Models (SSMs), represented by the Mamba family, provide linear-time sequence modeling and are attractive for long-context inference. Yet p

researcharxiv-cs-ai
14 Apr 2026
Research

Data-Efficient Surgical Phase Segmentation in Small-Incision Cataract Surgery: A Controlled Study of Vision Foundation Models

DGX agent

arXiv:2604.10514v1 Announce Type: cross Abstract: Surgical phase segmentation is central to computer-assisted surgery, yet robust models remain difficult to develop when labeled surgical videos are sc

researcharxiv-cs-ai
14 Apr 2026
Safety

Demographic and Linguistic Bias Evaluation in Omnimodal Language Models

DGX agent

arXiv:2604.10014v1 Announce Type: cross Abstract: This paper provides a comprehensive evaluation of demographic and linguistic biases in omnimodal language models that process text, images, audio, and

safetyarxiv-cs-ai
14 Apr 2026
Safety

Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models

DGX agent

arXiv:2604.10567v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) have emerged as a promising alternative to autoregressive language models, offering the potential for parallel

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Finetune Like You Pretrain: Boosting Zero-shot Adversarial Robustness in Vision-language Models

DGX agent

arXiv:2604.11576v1 Announce Type: new Abstract: Despite their impressive zero-shot abilities, vision-language models such as CLIP have been shown to be susceptible to adversarial attacks. To enhance i

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

FPBench: A Comprehensive Benchmark of Multimodal Large Language Models for Fingerprint Analysis

DGX agent

arXiv:2512.18073v2 Announce Type: replace Abstract: Multimodal LLMs (MLLMs) are capable of performing complex data analysis, visual question answering, generation, and reasoning tasks. However, their

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

From Topology to Trajectory: LLM-Driven World Models For Supply Chain Resilience

DGX agent

arXiv:2604.11041v1 Announce Type: new Abstract: Semiconductor supply chains face unprecedented resilience challenges amidst global geopolitical turbulence. Conventional Large Language Model (LLM) plan

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

DGX agent

arXiv:2604.10591v1 Announce Type: cross Abstract: Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, ye

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

LoGo-MR: Screening Breast MRI for Cancer Risk Prediction by Efficient Omni-Slice Modeling

DGX agent

arXiv:2604.11348v1 Announce Type: new Abstract: Efficient and explainable breast cancer (BC) risk prediction is critical for large-scale population-based screening. Breast MRI provides functional info

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models

DGX agent

arXiv:2604.11216v1 Announce Type: new Abstract: What values, evidence preferences, and source trust hierarchies do AI systems actually exhibit when facing structured dilemmas? We present the first lar

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

DGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

model-releasesarxiv-cs-cl
14 Apr 2026
← Previous
1…7374757677…1030
Next →