AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
Applications

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

DGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

applicationsarxiv-cs-lg
13 May 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

BEExformer: A Fast Inferencing Binarized Transformer with Early Exits

DGX agent

arXiv:2412.05225v3 Announce Type: replace Abstract: Large Language Models (LLMs) based on transformers achieve cutting-edge results on a variety of applications. However, their enormous size and proce

model-releasesarxiv-cs-cl
13 May 2026
Research

Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies

DGX agent

arXiv:2605.11387v1 Announce Type: new Abstract: We address the problem of fine-tuning pre-trained generative policies with reinforcement learning (RL) while preserving the multimodality of their actio

researcharxiv-cs-lg
13 May 2026
Model Releases

Beyond GRPO and On-Policy Distillation: An Empirical Sparse-to-Dense Reward Principle for Language-Model Post-Training

DGX agent

arXiv:2605.12483v1 Announce Type: new Abstract: In settings where labeled verifiable training data is the binding constraint, each checked example should be allocated carefully. The standard practice

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images

DGX agent

arXiv:2605.12413v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) show strong visual perception, yet remain limited in reasoning about space under changing viewpoints. We study

model-releasesarxiv-cs-cv
13 May 2026
Agents

Beyond Manual Curation: Augmenting Targeted Protein Degradation Databases via Agentic Literature Extraction Workflows

DGX agent

arXiv:2605.11221v1 Announce Type: cross Abstract: Predictive models in biomedicine depend on structured assay data locked in the text, tables, and supplements of primary publications. This bottleneck

agentsarxiv-cs-lg
13 May 2026
Research

Beyond Masks: The Case for Medical Image Parsing

DGX agent

arXiv:2605.11438v1 Announce Type: new Abstract: Medical imaging research has spent a decade getting very good at one thing: producing per-voxel masks. Masks tell us size, volume, and location, and a d

researcharxiv-cs-cv
13 May 2026
Model Releases

Beyond Parameter Aggregation: Semantic Consensus for Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.11857v1 Announce Type: new Abstract: Federated fine-tuning of large language models is commonly formulated as a parameter aggregation problem. However, even parameter-efficient methods requ

model-releasesarxiv-cs-lg
13 May 2026
Research

Beyond Point Estimates: Distributional Uncertainty in Machine Learning Performance Evaluation

DGX agent

arXiv:2501.16931v2 Announce Type: replace Abstract: Machine learning models are often evaluated using point estimates of performance metrics such as accuracy, F1 score, or mean squared error. Such sum

researcharxiv-cs-lg
13 May 2026
Safety

Beyond Point-wise Neural Collapse: A Topology-Aware Hierarchical Classifier for Class-Incremental Learning

DGX agent

arXiv:2605.11904v1 Announce Type: new Abstract: The Nearest Class Mean (NCM) classifier is widely favored in Class-Incremental Learning (CIL) for its superior resistance to catastrophic forgetting com

safetyarxiv-cs-cv
13 May 2026
Model Releases

Beyond Prediction: Interval Neural Networks for Uncertainty-Aware System Identification

DGX agent

arXiv:2605.11460v1 Announce Type: new Abstract: System identification (SysID) is critical for modeling dynamical systems from experimental data, yet traditional approaches often fail to capture nonlin

model-releasesarxiv-cs-lg
13 May 2026
Research

Beyond Similarity: Temporal Operator Attention for Time Series Analysis

DGX agent

arXiv:2605.11287v1 Announce Type: new Abstract: A persistent paradox in time-series forecasting is that structurally simple MLP and linear models often outperform high-capacity Transformers. We argue

researcharxiv-cs-lg
13 May 2026
Model Releases

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm

DGX agent

arXiv:2605.12271v1 Announce Type: new Abstract: Humans often specify and create through visual artifacts: typography sheets, sketches, reference images, and annotated scenes. Yet modern visual generat

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Bin Latent Transformer (BiLT): A shift-invariant autoencoder for calibration-free spectral unmixing of turbid media

DGX agent

arXiv:2605.11829v1 Announce Type: cross Abstract: The accurate recovery of constituent-level optical properties from integrating sphere measurements is a central analytical challenge in pharmaceutical

model-releasesarxiv-cs-lg
13 May 2026
Applications

Birds of a Feather Flock Together: Background-Invariant Representations via Linear Structure in VLMs

DGX agent

arXiv:2605.11107v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP and SigLIP 2, are widely used for image classification, yet their vision encoders remain vulnerable to syste

applicationsarxiv-cs-cv
13 May 2026
Research

BitLM: Unlocking Multi-Token Language Generation with Bitwise Continuous Diffusion

DGX agent

arXiv:2605.11577v1 Announce Type: new Abstract: Autoregressive language models generate text one token at a time, yet natural language is inherently structured in multi-token units, including phrases,

researcharxiv-cs-cl
13 May 2026
Research

BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking

DGX agent

arXiv:2602.00767v2 Announce Type: replace Abstract: Emergent misalignment can arise when a language model is fine-tuned on a narrowly scoped supervised objective: the model learns the target behavior,

researcharxiv-cs-lg
13 May 2026
Model Releases

Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models

DGX agent

arXiv:2605.11726v1 Announce Type: new Abstract: Recently, reinforcement learning (RL) has been widely applied during post-training for diffusion large language models (dLLMs) to enhance reasoning with

model-releasesarxiv-cs-lg
13 May 2026
Hardware

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

DGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

hardwarearxiv-cs-lg
13 May 2026
Model Releases

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation

DGX agent

arXiv:2605.12034v1 Announce Type: cross Abstract: Omni-modal language models are intended to jointly understand audio, visual inputs, and language, but benchmark gains can be inflated when visual evid

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents

DGX agent

arXiv:2508.07642v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) poses significant challenges for agents to interpret natural language instructions and navigate complex 3

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Breaking extit{Winner-Takes-All}: Cooperative Policy Optimization Improves Diverse LLM Reasoning

DGX agent

arXiv:2605.11461v1 Announce Type: cross Abstract: Reinforcement learning with verifiers (RLVR) has become a central paradigm for improving LLM reasoning, yet popular group-based optimization algorithm

model-releasesarxiv-cs-lg
13 May 2026
Research

BronchoLumen: Analysis of recent YOLO-based architectures for real-time bronchial orifice detection in video bronchoscopy

DGX agent

arXiv:2605.11748v1 Announce Type: new Abstract: Bronchoscopy is routinely conducted in pulmonary clinics and intensive care units, but navigating the complex branching of the respiratory tract remains

researcharxiv-cs-cv
13 May 2026
Safety

BSO: Safety Alignment Is Density Ratio Matching

DGX agent

arXiv:2605.12339v1 Announce Type: new Abstract: Aligning language models for both helpfulness and safety typically requires complex pipelines-separate reward and cost models, online reinforcement lear

safetyarxiv-cs-lg
13 May 2026
Local Ai

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating

DGX agent

arXiv:2605.11723v1 Announce Type: new Abstract: In this paper, we propose Concentrate and Concentrate (CaC), a coarse-to-fine anomaly reward model based on Vision-Language Models. During inference, it

local-aiarxiv-cs-cv
13 May 2026
Model Releases

CAD-feature enhanced machine learning for manufacturing effort estimation on sheet metal bending parts

DGX agent

arXiv:2605.12266v1 Announce Type: new Abstract: Graph-based machine learning has emerged as a promising approach for manufacturability analysis by learning directly from CAD models represented as Boun

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Calibrated Multimodal Representation Learning with Missing Modalities

DGX agent

arXiv:2511.12034v2 Announce Type: replace Abstract: Multimodal representation learning harmonizes distinct modalities by aligning them into a unified latent space. Recent research generalizes traditio

model-releasesarxiv-cs-cv
13 May 2026
Safety

Can a Single Message Paralyze the AI Infrastructure? The Rise of AbO-DDoS Attacks through Targeted Mobius Injection

DGX agent

arXiv:2605.11442v1 Announce Type: cross Abstract: Large Language Model (LLM) agents have emerged as key intermediaries, orchestrating complex interactions between human users and a wide range of digit

safetyarxiv-cs-cl
13 May 2026
Safety

Can Graphs Help Vision SSMs See Better?

DGX agent

arXiv:2605.11300v1 Announce Type: new Abstract: Vision state space models inherit the efficiency and long-range modeling ability of Mamba-style selective scans. However, their performance depends crit

safetyarxiv-cs-cv
13 May 2026
Research

Can Nano Banana 2 Replace Traditional Image Restoration Models? An Evaluation of Its Performance on Image Restoration Tasks

DGX agent

arXiv:2604.03061v2 Announce Type: replace Abstract: Recent advances in generative AI raise the question of whether general-purpose image editing models can serve as unified solutions for image restora

researcharxiv-cs-cv
13 May 2026
Model Releases

Caraman at SemEval-2026 Task 8: Three-Stage Multi-Turn Retrieval with Query Rewriting, Hybrid Search, and Cross-Encoder Reranking

DGX agent

arXiv:2605.12028v1 Announce Type: new Abstract: We describe our system for SemEval-2026 Task 8 (MTRAGEval), participating in Task A (Retrieval) across four English-language domains. Our approach emplo

model-releasesarxiv-cs-cl
13 May 2026
Research

CAST: Collapse-Aware multi-Scale Topology Fusion for Multimodal Coreset Selection

DGX agent

arXiv:2605.11705v1 Announce Type: new Abstract: The training of large multimodal models fundamentally relies on massive image-text datasets, which inevitably incur prohibitive computational overhead.

researcharxiv-cs-cv
13 May 2026
Model Releases

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration

DGX agent

arXiv:2605.11186v1 Announce Type: new Abstract: Auto-regressive decoding in Large Language Models (LLMs) is inherently memory-bound: every generation step requires loading the model weights and interm

model-releasesarxiv-cs-lg
13 May 2026
Applications

Causal Algorithmic Recourse: Foundations and Methods

DGX agent

arXiv:2605.11373v1 Announce Type: cross Abstract: The trustworthiness of AI decision-making systems is increasingly important. A key feature of such systems is the ability to provide recommendations f

applicationsarxiv-cs-lg
13 May 2026
Safety

Causal Bias Detection in Generative Artifical Intelligence

DGX agent

arXiv:2605.11365v1 Announce Type: cross Abstract: Automated systems built on artificial intelligence (AI) are increasingly deployed across high-stakes domains, raising critical concerns about fairness

safetyarxiv-cs-lg
13 May 2026
Safety

Causal Fairness for Survival Analysis

DGX agent

arXiv:2605.11362v1 Announce Type: new Abstract: In the data-driven era, large-scale datasets are routinely collected and analyzed using machine learning (ML) and artificial intelligence (AI) to inform

safetyarxiv-cs-lg
13 May 2026
Tutorials

CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives

DGX agent

arXiv:2605.12496v1 Announce Type: new Abstract: Autoregressive video generation aims at real-time, open-ended synthesis. Yet, cinematic storytelling is not merely the endless extension of a single sce

tutorialsarxiv-cs-cv
13 May 2026
Safety

Certified Gradient-Based Contact-Rich Manipulation via Smoothing-Error Reachable Tubes

DGX agent

arXiv:2602.09368v2 Announce Type: replace Abstract: Gradient-based methods can efficiently optimize controllers by leveraging differentiable simulation and physical priors. However, contact-rich manip

safetyarxiv-cs-ro
13 May 2026
Safety

Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing

DGX agent

arXiv:2505.05665v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated success in decision-making tasks including planning, control, and prediction, but thei

safetyarxiv-cs-cl
13 May 2026
Model Releases

Checkup2Action: A Multimodal Clinical Check-up Report Dataset for Patient-Oriented Action Card Generation

DGX agent

arXiv:2605.11533v1 Announce Type: new Abstract: Clinical check-up reports are multimodal documents that combine page layouts, tables, numerical biomarkers, abnormality flags, imaging findings, and dom

model-releasesarxiv-cs-cl
13 May 2026
Safety

CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography

DGX agent

arXiv:2605.11304v1 Announce Type: new Abstract: Chest radiograph interpretation requires temporal reasoning over prior and current studies, yet most vision-language models are trained on static image-

safetyarxiv-cs-cv
13 May 2026
Research

Choosing features for classifying multiword expressions

DGX agent

arXiv:2605.11779v1 Announce Type: new Abstract: Multiword expressions (MWEs) are a heterogeneous set with a glaring need for classifications. Designing a satisfactory classification involves choosing

researcharxiv-cs-cl
13 May 2026
Model Releases

Chronicles-OCR: A Cross-Temporal Perception Benchmark for the Evolutionary Trajectory of Chinese Characters

DGX agent

arXiv:2605.11960v1 Announce Type: new Abstract: Vision Large Language Models (VLLMs) have achieved remarkable success in modern text-rich visual understanding. However, their perceptual robustness in

model-releasesarxiv-cs-cv
13 May 2026
Hardware

ChunkFlow: Communication-Aware Chunked Prefetching for Layerwise Offloading in Distributed Diffusion Transformer Inference

DGX agent

arXiv:2605.11335v1 Announce Type: cross Abstract: Layerwise offloading reduces the GPU memory footprint of large diffusion transformer (DiT) inference by prefetching upcoming layers from host memory,

hardwarearxiv-cs-lg
13 May 2026
Safety

Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models

DGX agent

arXiv:2601.21944v2 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, thes

safetyarxiv-cs-lg
13 May 2026
Model Releases

ClinicalBench: Stress-Testing Assertion-Aware Retrieval for Cross-Admission Clinical QA on MIMIC-IV

DGX agent

arXiv:2605.11143v1 Announce Type: new Abstract: Reasoning benchmarks measure clinical performance on clean inputs. We evaluate the step before reasoning: retrieval over real EHR notes, where negation,

model-releasesarxiv-cs-cl
13 May 2026
Research

Closing the Motion Execution Gap: From Semantic Motion Task Constraints to Kinematic Control

DGX agent

arXiv:2605.12053v1 Announce Type: new Abstract: This paper addresses the Motion Execution Gap, the disconnect between high-level symbolic task descriptions using semantic constraints and executable ro

researcharxiv-cs-ro
13 May 2026
Safety

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

DGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

safetyarxiv-cs-cv
13 May 2026
← Previous
1…897898899900901…1290
Next →