AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents

DGX agent

arXiv:2602.14257v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents have made remarkable progress on complex reasoning, evaluating them in real-world environments remains

model-releasesarxiv-cs-lg
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

Adversarial Domain Prompt Tuning and Generation for Single Domain Generalization

DGX agent

arXiv:2606.21736v1 Announce Type: new Abstract: Single domain generalization (SDG) aims to learn a robust model, which could perform well on many unseen domains while there is only one single domain a

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Aligning AI-driven discovery with human intuition

DGX agent

arXiv:2410.07397v2 Announce Type: replace Abstract: As data-driven modeling of physical dynamical systems becomes more prevalent, a new challenge is emerging: making these models more compatible and a

researcharxiv-cs-lg
23 Jun 2026
Model Releases

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation

DGX agent

arXiv:2511.22973v2 Announce Type: replace Abstract: Long video generation is a critical step toward building realistic world models, requiring both high visual fidelity and long-range interaction cons

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

BoxCtrl: 3D-Aware Visual Prompting for Geometric Image Editing

DGX agent

arXiv:2606.23270v1 Announce Type: new Abstract: As instruction-based editing models and multimodal large language models advance, diverse image editing tasks have become feasible. However, achieving p

tutorialsarxiv-cs-cv
23 Jun 2026
Model Releases

CAOA -- Completion-Assisted Object-CAD Alignment

DGX agent

arXiv:2606.18429v2 Announce Type: replace Abstract: Accurately aligning CAD models to their corresponding objects in indoor RGB-D scans is a central challenge in 3D semantic reconstruction. The task r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Catching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

DGX agent

arXiv:2606.22699v1 Announce Type: new Abstract: Frontier multimodal models can guess whether a person is lying from a testimony video. To do so, they stream that raw face and voice to a third-party mo

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Causal Discovery in the Era of Agents

DGX agent

arXiv:2606.23608v1 Announce Type: cross Abstract: Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or in

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

DGX agent

arXiv:2606.22862v1 Announce Type: new Abstract: Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CheXpercept: A Benchmark for Evaluating Expert-Level Lesion Perception in Chest X-rays

DGX agent

arXiv:2606.21020v1 Announce Type: new Abstract: The evaluation of vision-language models (VLMs) for chest X-ray (CXR) analysis has largely been limited to disease-presence classification without visua

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization

DGX agent

arXiv:2606.21146v1 Announce Type: new Abstract: Text-to-video (T2V) diffusion models have made it increasingly easy to synthesize realistic and temporally coherent videos, while recent personalization

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Clipping the Price of Adaptivity at the Tail

DGX agent

arXiv:2606.22669v1 Announce Type: new Abstract: Adaptive stochastic convex optimization (SCO) methods face a fundamental ``price of adaptivity'' barrier: under the standard set of assumptions, they ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Disentangling Intrinsic Importance from Emergent Structure in Multi-Expert Orchestration

DGX agent

arXiv:2602.04291v2 Announce Type: replace Abstract: Multi-expert systems, where multiple Large Language Models (LLMs) collaborate to solve complex tasks, are increasingly adopted for high-performance

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies

DGX agent

arXiv:2606.18239v2 Announce Type: replace Abstract: We present EBench, a simulation benchmark that diagnoses generalist mobile manipulation policies beyond a single success-rate scalar. EBench compris

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction

DGX agent

arXiv:2606.22925v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared t

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval

DGX agent

arXiv:2606.22955v1 Announce Type: new Abstract: Large-scale pretrained foundation models have revolutionized general medical screening, but often falter on rare diseases because such conditions are un

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Explainable Boosting Machine for Predicting Claim Severity and Frequency in Car Insurance

DGX agent

arXiv:2503.21321v2 Announce Type: replace-cross Abstract: With the rapid development of machine learning and deep learning techniques, actuaries and the broader insurance industry face a persistent tr

model-releasesarxiv-cs-lg
23 Jun 2026
Research

FedOT: Ownership Verification and Leakage Tracing via Watermarks for Federated LDMs

DGX agent

arXiv:2606.22875v1 Announce Type: new Abstract: Training Latent Diffusion Models (LDMs) within Federated Learning (FL) has attracted increasing attention due to its ability to combine the powerful gen

researcharxiv-cs-cv
23 Jun 2026
Model Releases

From Convolution to Transformer: A Comparative Study of U-Net Variants for Brain Tumor and Retinal Vessel Segmentation

DGX agent

arXiv:2606.22168v1 Announce Type: new Abstract: Medical image segmentation plays an important role in computer aided diagnosis, treatment planning, and disease monitoring. U-Net has been widely used f

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Gradient-Descent Steps to Success over Mean Accuracy: A Paradigm Shift for ML

DGX agent

arXiv:2606.22053v1 Announce Type: new Abstract: Traditional evaluation of machine learning (ML) models typically focuses on achieving the maximum possible accuracy irrespective of the computational co

researcharxiv-cs-lg
23 Jun 2026
Model Releases

HaineiFRDM: Structure-Preserving Diffusion for Film Restoration under Fast Motion and Diverse Defects

DGX agent

arXiv:2512.24946v2 Announce Type: replace Abstract: Existing film-restoration methods frequently fail under fast motion, producing limb disappearance and structural distortion due to inaccurate motion

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Hedgementation = Hedgerow Segmentation: A Remote Sensing Benchmark

DGX agent

arXiv:2606.23615v1 Announce Type: new Abstract: We propose Hedgementation: a new benchmark to evaluate machine learning models for hedgerow mapping from remote sensing data at country scale and 10m^2

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

How GPT-5 helped immunologist Derya Unutmaz solve a 3-year-old mystery

DGX agent

Immunologist Derya Unutmaz leveraged GPT-5 to resolve a complex scientific mystery that had remained unsolved for three years, demonstrating the AI model's capacity to assist in advanced biomedical re

model-releasesopenai
23 Jun 2026
Tutorials

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do

DGX agent

arXiv:2606.22565v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) has become a standard method for improving reasoning capabilities in large language models (LLMs) by eliciting step-by-step thi

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding

DGX agent

arXiv:2601.07298v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) excel at single-image understanding, they exhibit significantly degraded performance in multi-image r

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Mirage: a Clean-Label Backdoor against LiDAR 3D Object Detection

DGX agent

arXiv:2606.20752v1 Announce Type: new Abstract: Deep neural network-based LiDAR 3D object detection serves as a critical perception component in safety-critical autonomous systems. However, recent stu

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Multigrid Training for Molecular Generation using Graph Neural Networks

DGX agent

arXiv:2606.22377v1 Announce Type: new Abstract: Deep learning has demonstrated significant success for modeling biochemical molecular systems, where inputs are commonly represented as graphs or 3D gri

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

OmniV2X: A Generative Foundation Planner for Efficient End-to-End Cooperative Driving

DGX agent

arXiv:2606.21165v1 Announce Type: new Abstract: We present OmniV2X, a generative foundation model for vehicle-to-everything (V2X) cooperative driving. The model directly interprets independent context

agentsarxiv-cs-ro
23 Jun 2026
Model Releases

ORBIT: Training-Free Multi-Attribute Behavioral Steering via Orthogonal Subspace Rotation

DGX agent

arXiv:2606.22357v1 Announce Type: cross Abstract: Language models are widely used in assistant settings, where controlling behavioral attributes is often essential. Activation steering modifies hidden

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

PROTON: Prototype-Based Test-Time Online OOD Detection for Medical VLMs

DGX agent

arXiv:2606.20913v1 Announce Type: new Abstract: Medical vision-language models (VLMs) enable zero-shot clinical image classification, yet reliably detecting out-of-distribution (OOD) inputs at deploym

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild

DGX agent

arXiv:2603.04205v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) achieve near-perfect scores on digital document benchmarks like OmniDocBench, their performance in the unpredict

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Revisiting the Neural Tangent Kernel: the role of large width and depth

DGX agent

arXiv:2511.07272v2 Announce Type: replace Abstract: Overparameterized fully-connected neural networks have been shown to behave like kernel models when trained with gradient descent, assuming standard

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RS-Gen: A Multi-Stage Agentic Framework for Reasoning and Search-Augmented Image Generation

DGX agent

arXiv:2606.23221v1 Announce Type: new Abstract: Recent years have witnessed remarkable progress in image generation and editing, particularly regarding instruction following and visual fidelity. Howev

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Sakana Fugu Technical Report

DGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

SATURN: Symbolic Spatial Reasoning for Multi-Perspective Grounding

DGX agent

arXiv:2606.22694v1 Announce Type: new Abstract: Vision-Language Models (VLMs) remain unreliable when spatial reasoning requires composing relations whose meanings depend on frames of reference. Existi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Scaling Linear Mode Connectivity and Merging to Billion Parameter Pretrained Transformers

DGX agent

arXiv:2606.23607v1 Announce Type: new Abstract: Linear mode connectivity (LMC) provides a promising foundation for understanding and merging independently trained neural networks, but existing methods

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

DGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

DGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

T-IMPACT: A Severity-Aware Benchmark for Contextual Image-Text Manipulation

DGX agent

arXiv:2606.22339v1 Announce Type: new Abstract: Recent advances in vision-language models and generative editing systems have made it increasingly easy to produce persuasive multimodal misinformation

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Temporal-Spectral Alignment with Frequency Adaptation for Source-Free Time-Series Adaptation

DGX agent

arXiv:2606.23120v1 Announce Type: new Abstract: The goal of source-free domain adaptation (SFDA) for time-series data is to transfer knowledge from a pre-trained source model to an unlabeled target do

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

The Alignment Problem in Constrained Code Generation

DGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction

DGX agent

arXiv:2606.22969v1 Announce Type: new Abstract: Predicting the behavior of dynamical systems (DS) beyond the dynamical and parameter regimes observed in training is a pivotal and essentially unresolve

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Towards Robust Personalized Federated Learning: Vulnerability Assessment and Defense Co-Design

DGX agent

arXiv:2606.22782v1 Announce Type: new Abstract: The proliferation of IoT devices has fueled distributed edge systems to collect vast amounts of sensitive data, creating fertile ground for on-device ma

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

DGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

safetyarxiv-cs-lg
23 Jun 2026
Safety

Using predictive multiplicity to measure individual performance within the AI Act

DGX agent

arXiv:2602.11944v2 Announce Type: replace Abstract: When building AI systems for decision support, one often encounters the phenomenon of predictive multiplicity: a single best model does not exist; i

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Where Does the Signal Live? A Web Data Recipe for Medical Encoder Pretraining

DGX agent

arXiv:2606.22079v1 Announce Type: cross Abstract: Web data curation has been widely studied for decoder Large Language Model (LLM) pretraining. Encoders for dense-terminology domains such as medicine,

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

DGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

agentsdavid-ha--x
22 Jun 2026
Tools

Sakana Fugu Ultra now available on AI Gateway

DGX agent

Sakana Fugu Ultra, a new AI model, is now available through Vercel's AI Gateway, expanding the selection of models developers can access via the platform. This addition allows users to integrate Sakan

toolsvercel-blog
22 Jun 2026
← Previous
1…452453454455456…1371
Next →