AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Agents

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

DGX agent

arXiv:2606.23743v1 Announce Type: cross Abstract: Modern video diffusion models achieve higher generation quality through scaling, but this also increases inference cost. Although many acceleration me

agentsarxiv-cs-ai
24 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMs

DGX agent

arXiv:2606.24460v1 Announce Type: cross Abstract: Commercial large language models bill, scale latency, and budget context per token. Yet tokenizers assign more subword tokens to the same meaning in s

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The Degeneracy Distillery

DGX agent

arXiv:2606.23838v1 Announce Type: new Abstract: When two or more parameters or labels produce similar data, they are degenerate, or hard to distinguish. Degeneracies render both label prediction and i

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

VisCritic: Visual State Comparison as Process Reward for GUI Agents

DGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents

DGX agent

arXiv:2602.14257v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents have made remarkable progress on complex reasoning, evaluating them in real-world environments remains

model-releasesarxiv-cs-lg
23 Jun 2026
Tutorials

Adversarial Domain Prompt Tuning and Generation for Single Domain Generalization

DGX agent

arXiv:2606.21736v1 Announce Type: new Abstract: Single domain generalization (SDG) aims to learn a robust model, which could perform well on many unseen domains while there is only one single domain a

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Aligning AI-driven discovery with human intuition

DGX agent

arXiv:2410.07397v2 Announce Type: replace Abstract: As data-driven modeling of physical dynamical systems becomes more prevalent, a new challenge is emerging: making these models more compatible and a

researcharxiv-cs-lg
23 Jun 2026
Model Releases

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation

DGX agent

arXiv:2511.22973v2 Announce Type: replace Abstract: Long video generation is a critical step toward building realistic world models, requiring both high visual fidelity and long-range interaction cons

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

BoxCtrl: 3D-Aware Visual Prompting for Geometric Image Editing

DGX agent

arXiv:2606.23270v1 Announce Type: new Abstract: As instruction-based editing models and multimodal large language models advance, diverse image editing tasks have become feasible. However, achieving p

tutorialsarxiv-cs-cv
23 Jun 2026
Model Releases

CAOA -- Completion-Assisted Object-CAD Alignment

DGX agent

arXiv:2606.18429v2 Announce Type: replace Abstract: Accurately aligning CAD models to their corresponding objects in indoor RGB-D scans is a central challenge in 3D semantic reconstruction. The task r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Catching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

DGX agent

arXiv:2606.22699v1 Announce Type: new Abstract: Frontier multimodal models can guess whether a person is lying from a testimony video. To do so, they stream that raw face and voice to a third-party mo

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Causal Discovery in the Era of Agents

DGX agent

arXiv:2606.23608v1 Announce Type: cross Abstract: Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or in

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

DGX agent

arXiv:2606.22862v1 Announce Type: new Abstract: Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CheXpercept: A Benchmark for Evaluating Expert-Level Lesion Perception in Chest X-rays

DGX agent

arXiv:2606.21020v1 Announce Type: new Abstract: The evaluation of vision-language models (VLMs) for chest X-ray (CXR) analysis has largely been limited to disease-presence classification without visua

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization

DGX agent

arXiv:2606.21146v1 Announce Type: new Abstract: Text-to-video (T2V) diffusion models have made it increasingly easy to synthesize realistic and temporally coherent videos, while recent personalization

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Clipping the Price of Adaptivity at the Tail

DGX agent

arXiv:2606.22669v1 Announce Type: new Abstract: Adaptive stochastic convex optimization (SCO) methods face a fundamental ``price of adaptivity'' barrier: under the standard set of assumptions, they ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Disentangling Intrinsic Importance from Emergent Structure in Multi-Expert Orchestration

DGX agent

arXiv:2602.04291v2 Announce Type: replace Abstract: Multi-expert systems, where multiple Large Language Models (LLMs) collaborate to solve complex tasks, are increasingly adopted for high-performance

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies

DGX agent

arXiv:2606.18239v2 Announce Type: replace Abstract: We present EBench, a simulation benchmark that diagnoses generalist mobile manipulation policies beyond a single success-rate scalar. EBench compris

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction

DGX agent

arXiv:2606.22925v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared t

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval

DGX agent

arXiv:2606.22955v1 Announce Type: new Abstract: Large-scale pretrained foundation models have revolutionized general medical screening, but often falter on rare diseases because such conditions are un

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Explainable Boosting Machine for Predicting Claim Severity and Frequency in Car Insurance

DGX agent

arXiv:2503.21321v2 Announce Type: replace-cross Abstract: With the rapid development of machine learning and deep learning techniques, actuaries and the broader insurance industry face a persistent tr

model-releasesarxiv-cs-lg
23 Jun 2026
Research

FedOT: Ownership Verification and Leakage Tracing via Watermarks for Federated LDMs

DGX agent

arXiv:2606.22875v1 Announce Type: new Abstract: Training Latent Diffusion Models (LDMs) within Federated Learning (FL) has attracted increasing attention due to its ability to combine the powerful gen

researcharxiv-cs-cv
23 Jun 2026
Model Releases

From Convolution to Transformer: A Comparative Study of U-Net Variants for Brain Tumor and Retinal Vessel Segmentation

DGX agent

arXiv:2606.22168v1 Announce Type: new Abstract: Medical image segmentation plays an important role in computer aided diagnosis, treatment planning, and disease monitoring. U-Net has been widely used f

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Gradient-Descent Steps to Success over Mean Accuracy: A Paradigm Shift for ML

DGX agent

arXiv:2606.22053v1 Announce Type: new Abstract: Traditional evaluation of machine learning (ML) models typically focuses on achieving the maximum possible accuracy irrespective of the computational co

researcharxiv-cs-lg
23 Jun 2026
Model Releases

HaineiFRDM: Structure-Preserving Diffusion for Film Restoration under Fast Motion and Diverse Defects

DGX agent

arXiv:2512.24946v2 Announce Type: replace Abstract: Existing film-restoration methods frequently fail under fast motion, producing limb disappearance and structural distortion due to inaccurate motion

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Hedgementation = Hedgerow Segmentation: A Remote Sensing Benchmark

DGX agent

arXiv:2606.23615v1 Announce Type: new Abstract: We propose Hedgementation: a new benchmark to evaluate machine learning models for hedgerow mapping from remote sensing data at country scale and 10m^2

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

How GPT-5 helped immunologist Derya Unutmaz solve a 3-year-old mystery

DGX agent

Immunologist Derya Unutmaz leveraged GPT-5 to resolve a complex scientific mystery that had remained unsolved for three years, demonstrating the AI model's capacity to assist in advanced biomedical re

model-releasesopenai
23 Jun 2026
Tutorials

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do

DGX agent

arXiv:2606.22565v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) has become a standard method for improving reasoning capabilities in large language models (LLMs) by eliciting step-by-step thi

tutorialsarxiv-cs-cv
23 Jun 2026
Research

Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding

DGX agent

arXiv:2601.07298v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) excel at single-image understanding, they exhibit significantly degraded performance in multi-image r

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Mirage: a Clean-Label Backdoor against LiDAR 3D Object Detection

DGX agent

arXiv:2606.20752v1 Announce Type: new Abstract: Deep neural network-based LiDAR 3D object detection serves as a critical perception component in safety-critical autonomous systems. However, recent stu

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Multigrid Training for Molecular Generation using Graph Neural Networks

DGX agent

arXiv:2606.22377v1 Announce Type: new Abstract: Deep learning has demonstrated significant success for modeling biochemical molecular systems, where inputs are commonly represented as graphs or 3D gri

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

OmniV2X: A Generative Foundation Planner for Efficient End-to-End Cooperative Driving

DGX agent

arXiv:2606.21165v1 Announce Type: new Abstract: We present OmniV2X, a generative foundation model for vehicle-to-everything (V2X) cooperative driving. The model directly interprets independent context

agentsarxiv-cs-ro
23 Jun 2026
Model Releases

ORBIT: Training-Free Multi-Attribute Behavioral Steering via Orthogonal Subspace Rotation

DGX agent

arXiv:2606.22357v1 Announce Type: cross Abstract: Language models are widely used in assistant settings, where controlling behavioral attributes is often essential. Activation steering modifies hidden

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

PROTON: Prototype-Based Test-Time Online OOD Detection for Medical VLMs

DGX agent

arXiv:2606.20913v1 Announce Type: new Abstract: Medical vision-language models (VLMs) enable zero-shot clinical image classification, yet reliably detecting out-of-distribution (OOD) inputs at deploym

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild

DGX agent

arXiv:2603.04205v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) achieve near-perfect scores on digital document benchmarks like OmniDocBench, their performance in the unpredict

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Revisiting the Neural Tangent Kernel: the role of large width and depth

DGX agent

arXiv:2511.07272v2 Announce Type: replace Abstract: Overparameterized fully-connected neural networks have been shown to behave like kernel models when trained with gradient descent, assuming standard

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RS-Gen: A Multi-Stage Agentic Framework for Reasoning and Search-Augmented Image Generation

DGX agent

arXiv:2606.23221v1 Announce Type: new Abstract: Recent years have witnessed remarkable progress in image generation and editing, particularly regarding instruction following and visual fidelity. Howev

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Sakana Fugu Technical Report

DGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

SATURN: Symbolic Spatial Reasoning for Multi-Perspective Grounding

DGX agent

arXiv:2606.22694v1 Announce Type: new Abstract: Vision-Language Models (VLMs) remain unreliable when spatial reasoning requires composing relations whose meanings depend on frames of reference. Existi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Scaling Linear Mode Connectivity and Merging to Billion Parameter Pretrained Transformers

DGX agent

arXiv:2606.23607v1 Announce Type: new Abstract: Linear mode connectivity (LMC) provides a promising foundation for understanding and merging independently trained neural networks, but existing methods

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

DGX agent

arXiv:2606.22873v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed in consumer, medical, financial, and enterprise applications. This broad deployment expands the

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

DGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

T-IMPACT: A Severity-Aware Benchmark for Contextual Image-Text Manipulation

DGX agent

arXiv:2606.22339v1 Announce Type: new Abstract: Recent advances in vision-language models and generative editing systems have made it increasingly easy to produce persuasive multimodal misinformation

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Temporal-Spectral Alignment with Frequency Adaptation for Source-Free Time-Series Adaptation

DGX agent

arXiv:2606.23120v1 Announce Type: new Abstract: The goal of source-free domain adaptation (SFDA) for time-series data is to transfer knowledge from a pre-trained source model to an unlabeled target do

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

The Alignment Problem in Constrained Code Generation

DGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction

DGX agent

arXiv:2606.22969v1 Announce Type: new Abstract: Predicting the behavior of dynamical systems (DS) beyond the dynamical and parameter regimes observed in training is a pivotal and essentially unresolve

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Towards Robust Personalized Federated Learning: Vulnerability Assessment and Defense Co-Design

DGX agent

arXiv:2606.22782v1 Announce Type: new Abstract: The proliferation of IoT devices has fueled distributed edge systems to collect vast amounts of sensitive data, creating fertile ground for on-device ma

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

DGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

safetyarxiv-cs-lg
23 Jun 2026
← Previous
1…440441442443444…1338
Next →