AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction

DGX agent

arXiv:2606.22925v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared t

model-releasesarxiv-cs-lg
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval

DGX agent

arXiv:2606.22955v1 Announce Type: new Abstract: Large-scale pretrained foundation models have revolutionized general medical screening, but often falter on rare diseases because such conditions are un

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Explainable Boosting Machine for Predicting Claim Severity and Frequency in Car Insurance

DGX agent

arXiv:2503.21321v2 Announce Type: replace-cross Abstract: With the rapid development of machine learning and deep learning techniques, actuaries and the broader insurance industry face a persistent tr

model-releasesarxiv-cs-lg
23 Jun 2026
Research

FedOT: Ownership Verification and Leakage Tracing via Watermarks for Federated LDMs

DGX agent

arXiv:2606.22875v1 Announce Type: new Abstract: Training Latent Diffusion Models (LDMs) within Federated Learning (FL) has attracted increasing attention due to its ability to combine the powerful gen

researcharxiv-cs-cv
23 Jun 2026
Model Releases

From Convolution to Transformer: A Comparative Study of U-Net Variants for Brain Tumor and Retinal Vessel Segmentation

DGX agent

arXiv:2606.22168v1 Announce Type: new Abstract: Medical image segmentation plays an important role in computer aided diagnosis, treatment planning, and disease monitoring. U-Net has been widely used f

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Gradient-Descent Steps to Success over Mean Accuracy: A Paradigm Shift for ML

DGX agent

arXiv:2606.22053v1 Announce Type: new Abstract: Traditional evaluation of machine learning (ML) models typically focuses on achieving the maximum possible accuracy irrespective of the computational co

researcharxiv-cs-lg
23 Jun 2026
Model Releases

HaineiFRDM: Structure-Preserving Diffusion for Film Restoration under Fast Motion and Diverse Defects

DGX agent

arXiv:2512.24946v2 Announce Type: replace Abstract: Existing film-restoration methods frequently fail under fast motion, producing limb disappearance and structural distortion due to inaccurate motion

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Hedgementation = Hedgerow Segmentation: A Remote Sensing Benchmark

DGX agent

arXiv:2606.23615v1 Announce Type: new Abstract: We propose Hedgementation: a new benchmark to evaluate machine learning models for hedgerow mapping from remote sensing data at country scale and 10m^2

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do

DGX agent

arXiv:2606.22565v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) has become a standard method for improving reasoning capabilities in large language models (LLMs) by eliciting step-by-step thi

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding

DGX agent

arXiv:2601.07298v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) excel at single-image understanding, they exhibit significantly degraded performance in multi-image r

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Mirage: a Clean-Label Backdoor against LiDAR 3D Object Detection

DGX agent

arXiv:2606.20752v1 Announce Type: new Abstract: Deep neural network-based LiDAR 3D object detection serves as a critical perception component in safety-critical autonomous systems. However, recent stu

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Multigrid Training for Molecular Generation using Graph Neural Networks

DGX agent

arXiv:2606.22377v1 Announce Type: new Abstract: Deep learning has demonstrated significant success for modeling biochemical molecular systems, where inputs are commonly represented as graphs or 3D gri

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

OmniV2X: A Generative Foundation Planner for Efficient End-to-End Cooperative Driving

DGX agent

arXiv:2606.21165v1 Announce Type: new Abstract: We present OmniV2X, a generative foundation model for vehicle-to-everything (V2X) cooperative driving. The model directly interprets independent context

agentsarxiv-cs-ro
23 Jun 2026
Model Releases

ORBIT: Training-Free Multi-Attribute Behavioral Steering via Orthogonal Subspace Rotation

DGX agent

arXiv:2606.22357v1 Announce Type: cross Abstract: Language models are widely used in assistant settings, where controlling behavioral attributes is often essential. Activation steering modifies hidden

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

PROTON: Prototype-Based Test-Time Online OOD Detection for Medical VLMs

DGX agent

arXiv:2606.20913v1 Announce Type: new Abstract: Medical vision-language models (VLMs) enable zero-shot clinical image classification, yet reliably detecting out-of-distribution (OOD) inputs at deploym

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild

DGX agent

arXiv:2603.04205v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) achieve near-perfect scores on digital document benchmarks like OmniDocBench, their performance in the unpredict

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Revisiting the Neural Tangent Kernel: the role of large width and depth

DGX agent

arXiv:2511.07272v2 Announce Type: replace Abstract: Overparameterized fully-connected neural networks have been shown to behave like kernel models when trained with gradient descent, assuming standard

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RS-Gen: A Multi-Stage Agentic Framework for Reasoning and Search-Augmented Image Generation

DGX agent

arXiv:2606.23221v1 Announce Type: new Abstract: Recent years have witnessed remarkable progress in image generation and editing, particularly regarding instruction following and visual fidelity. Howev

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Sakana Fugu Technical Report

DGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

SATURN: Symbolic Spatial Reasoning for Multi-Perspective Grounding

DGX agent

arXiv:2606.22694v1 Announce Type: new Abstract: Vision-Language Models (VLMs) remain unreliable when spatial reasoning requires composing relations whose meanings depend on frames of reference. Existi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Scaling Linear Mode Connectivity and Merging to Billion Parameter Pretrained Transformers

DGX agent

arXiv:2606.23607v1 Announce Type: new Abstract: Linear mode connectivity (LMC) provides a promising foundation for understanding and merging independently trained neural networks, but existing methods

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

DGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

DGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

T-IMPACT: A Severity-Aware Benchmark for Contextual Image-Text Manipulation

DGX agent

arXiv:2606.22339v1 Announce Type: new Abstract: Recent advances in vision-language models and generative editing systems have made it increasingly easy to produce persuasive multimodal misinformation

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Temporal-Spectral Alignment with Frequency Adaptation for Source-Free Time-Series Adaptation

DGX agent

arXiv:2606.23120v1 Announce Type: new Abstract: The goal of source-free domain adaptation (SFDA) for time-series data is to transfer knowledge from a pre-trained source model to an unlabeled target do

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

The Alignment Problem in Constrained Code Generation

DGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction

DGX agent

arXiv:2606.22969v1 Announce Type: new Abstract: Predicting the behavior of dynamical systems (DS) beyond the dynamical and parameter regimes observed in training is a pivotal and essentially unresolve

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Towards Robust Personalized Federated Learning: Vulnerability Assessment and Defense Co-Design

DGX agent

arXiv:2606.22782v1 Announce Type: new Abstract: The proliferation of IoT devices has fueled distributed edge systems to collect vast amounts of sensitive data, creating fertile ground for on-device ma

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

DGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

safetyarxiv-cs-lg
23 Jun 2026
Safety

Using predictive multiplicity to measure individual performance within the AI Act

DGX agent

arXiv:2602.11944v2 Announce Type: replace Abstract: When building AI systems for decision support, one often encounters the phenomenon of predictive multiplicity: a single best model does not exist; i

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Where Does the Signal Live? A Web Data Recipe for Medical Encoder Pretraining

DGX agent

arXiv:2606.22079v1 Announce Type: cross Abstract: Web data curation has been widely studied for decoder Large Language Model (LLM) pretraining. Encoders for dense-terminology domains such as medicine,

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

ALIGNBEAM : Inference-Time Alignment Transfer via Cross-Vocabulary Logit Mixing

DGX agent

arXiv:2606.12342v1 Announce Type: cross Abstract: Domain fine-tuning degrades the safety of large language models: fine-tuned specialists readily comply with harmful prompts framed in domain language.

safetyarxiv-cs-ai
11 Jun 2026
Safety

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

DGX agent

arXiv:2606.11634v1 Announce Type: new Abstract: The rapid progress of reasoning and agentic large language models (LLMs) has increased the demand for long-context inference, but self-attention (SA) sc

safetyarxiv-cs-ai
11 Jun 2026
Safety

Beyond Third-Person Audits: Situated Interaction Auditing for User-Centered LLM Bias Research

DGX agent

arXiv:2606.12247v1 Announce Type: cross Abstract: Research on bias in large language models (LLMs) has predominantly focused on third-person audits, which study how models represent or evaluate demogr

safetyarxiv-cs-cl
11 Jun 2026
Agents

Bootstrapped Monitoring: Leveraging Transparent Reasoning to Oversee Stronger AI Agents

DGX agent

arXiv:2606.11998v1 Announce Type: new Abstract: Trusted monitoring is a cornerstone of AI control. However, as frontier models grow more capable, the increasing capabilities gap between trusted and un

agentsarxiv-cs-lg
11 Jun 2026
Model Releases

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

DGX agent

arXiv:2606.12344v1 Announce Type: cross Abstract: General-purpose agents such as OpenClaw are increasingly used as autonomous tool users, but their coding ability is difficult to measure under SWE-ben

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

DGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training

DGX agent

arXiv:2606.11854v1 Announce Type: cross Abstract: There are two main Parameter-Efficient Fine-Tuning (PEFT) techniques for Large Language Models (LLMs). While Low-Rank Adaptation (LoRA) introduces add

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback

DGX agent

arXiv:2601.04203v2 Announce Type: replace Abstract: We present FronTalk, a benchmark for front-end code generation that pioneers the study of a unique interaction dynamic: conversational code generati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Improving Detection of Rare Nodes in Hierarchical Multi-Label Learning

DGX agent

arXiv:2602.08986v2 Announce Type: replace-cross Abstract: In hierarchical multi-label classification, a persistent challenge is enabling model predictions to reach deeper levels of the hierarchy for m

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends

DGX agent

arXiv:2606.12207v1 Announce Type: cross Abstract: Embodied intelligence now spans navigation, household assistance, manipulation, autonomous driving, aerial agents, and multimodal large-model control.

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Lung-SRAD: Spectral-Aware Regularized Audio DASS with Dual-Axis Patch-Mix Contrastive Learning for Respiratory Sound Classification

DGX agent

arXiv:2606.11922v1 Announce Type: cross Abstract: Recent respiratory sound classification (RSC) studies largely rely on CLS-token driven self-attention architectures such as the Audio Spectrogram Tran

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MARIC: Multi-Agent Reasoning for Image Classification

DGX agent

arXiv:2509.14860v2 Announce Type: replace-cross Abstract: Image classification has traditionally relied on parameter-intensive model training, requiring large-scale annotated datasets and extensive fi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios

DGX agent

arXiv:2602.22638v2 Announce Type: replace Abstract: Route-planning agents powered by large language models (LLMs) have emerged as a promising paradigm for supporting everyday human mobility through na

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Noise-Aware Framework for Correcting Corrupted Labels

DGX agent

arXiv:2606.11695v1 Announce Type: cross Abstract: High-quality labeled data is essential for training reliable ML/DL models. However, real-world datasets often contain a considerable proportion of cor

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

ProGRank: Probe-Gradient Reranking to Defend Dense-Retriever RAG from Corpus Poisoning

DGX agent

arXiv:2603.22934v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves large language model applications by grounding generation in retrieved evidence, but also introduces c

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding

DGX agent

arXiv:2606.12125v1 Announce Type: new Abstract: Long-video understanding remains challenging for multimodal large language models, because temporally extended videos often contain thousands of frames

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

DGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

model-releasesarxiv-cs-cl
11 Jun 2026
← Previous
1…353354355356357…1067
Next →