AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Fine-tuning language encoding models on slow fMRI improves prediction for fast ECoG

DGX agent

arXiv:2605.19224v1 Announce Type: new Abstract: Neuroscientists have recently turned to intracranial brain recording methods, like electrocorticography (ECoG), for human experiments because of the fin

researcharxiv-cs-cl
20 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Hybrid Training for Vision-Language-Action Models

DGX agent

arXiv:2510.00600v2 Announce Type: replace-cross Abstract: Using Large Language Models to produce intermediate thoughts, a.k.a. Chain-of-thought (CoT), before providing an answer has been a successful

agentsarxiv-cs-ai
20 May 2026
Tutorials

Language models struggle with compartmentalization

DGX agent

arXiv:2605.19284v1 Announce Type: new Abstract: In the training data used by large language models (LLMs), the same latent concept is often presented in multiple distinct ways: the same facts appear i

tutorialsarxiv-cs-cl
20 May 2026
Model Releases

Lightweight and Fast Backdoor Model Detection

DGX agent

arXiv:2605.18907v1 Announce Type: cross Abstract: Deep neural networks (DNN), despite their remarkable performance, are highly vulnerable to backdoor attacks. Existing defenses mainly rely on activati

model-releasesarxiv-cs-ai
20 May 2026
Research

Markov Chain Decoders Overcome the Heavy-Tail Limitations of Lipschitz Generative Models

DGX agent

arXiv:2605.18931v1 Announce Type: cross Abstract: Heavy-tailed distributions are prevalent in performance evaluation, network traffic, and risk modeling. This behavior poses a fundamental challenge fo

researcharxiv-cs-ai
20 May 2026
Model Releases

MSAlign: Aligning Molecule and Mass Spectra Foundation Models for Metabolite Identification

DGX agent

arXiv:2605.19752v1 Announce Type: new Abstract: Accurately identifying metabolites i.e. small molecules from mass spectrometry data remains a core challenge in metabolomics, with broad applications in

model-releasesarxiv-cs-lg
20 May 2026
Applications

Quantized Machine Learning Models for Medical Imaging in Low-Resource Healthcare Settings

DGX agent

arXiv:2605.19207v1 Announce Type: cross Abstract: Deep learning models have shown strong performance in medical image analysis, but deploying them in low-resource clinical environments remains difficu

applicationsarxiv-cs-ai
20 May 2026
Research

Scaling Evaluation-time Compute with Reasoning Models as Evaluators

DGX agent

arXiv:2503.19877v2 Announce Type: replace Abstract: As language model (LM) outputs get more and more natural, it is becoming more difficult than ever to evaluate their quality. Simultaneously, increas

researcharxiv-cs-cl
20 May 2026
Research

Transformers Linearly Represent Highly Structured World Models

DGX agent

arXiv:2605.18847v1 Announce Type: cross Abstract: Do transformers, when trained on sequential reasoning traces, build internal models of the underlying task? And if so, does the structure of those int

researcharxiv-cs-ai
20 May 2026
Safety

When Preference Labels Fall Short: Aligning Diffusion Models from Real Data

DGX agent

arXiv:2605.19839v1 Announce Type: new Abstract: Preference alignment aims to guide generative models by learning from comparisons between preferred and non-preferred samples. In practice, most existin

safetyarxiv-cs-cv
20 May 2026
Safety

When Tabular Foundation Models Meet Strategic Tabular Data: A Prior Alignment Approach

DGX agent

arXiv:2605.19662v1 Announce Type: new Abstract: Tabular foundation models based on pretrained prior-data fitted networks~(PFNs) have shown strong generalization on diverse tabular tasks, but they are

safetyarxiv-cs-ai
20 May 2026
Research

Adversarial Attacks on Downstream Weather Forecasting Models: Application to Tropical Cyclone Trajectory Prediction

DGX agent

arXiv:2510.10140v2 Announce Type: replace Abstract: Deep learning-based weather forecasting (DLWF) models leverage past weather observations to generate future forecasts, supporting a wide range of do

researcharxiv-cs-lg
19 May 2026
Research

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models

DGX agent

arXiv:2605.17310v1 Announce Type: cross Abstract: Existing adversarial attacks on vision-language models (VLMs) can steer model outputs toward attacker-specified target responses, but their effectiven

researcharxiv-cs-ai
19 May 2026
Tutorials

CBT-Audio: Evaluating Audio Language Models for Patient-Side Distress Intensity Estimation in CBT Session Recordings

DGX agent

arXiv:2605.17370v1 Announce Type: new Abstract: Cognitive behavioural therapy is widely used to help patients understand and manage psychological distress. It is often delivered through spoken convers

tutorialsarxiv-cs-ai
19 May 2026
Research

CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models

DGX agent

arXiv:2602.17684v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has driven recent progress in code large language models by leveraging execution-based f

researcharxiv-cs-ai
19 May 2026
Safety

Cracks in the Foundation: A Civil Infrastructure Dataset to Challenge Vision Foundation Models

DGX agent

arXiv:2605.18413v1 Announce Type: new Abstract: Automated structural health monitoring is essential to prevent catastrophic infrastructure failures. Precise, pixel-level defect segmentation is needed

safetyarxiv-cs-cv
19 May 2026
Safety

DACA-GRPO: Denoising-Aware Credit Assignment for Reinforcement Learning in Diffusion Language Models

DGX agent

arXiv:2605.16342v1 Announce Type: cross Abstract: Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps

safetyarxiv-cs-ai
19 May 2026
Applications

DAD4TS: Data-Augmentation-Oriented Diffusion Model for Time-Series Forecasting with Small-Scale Data

DGX agent

arXiv:2605.17866v1 Announce Type: new Abstract: Small-scale data is a critical problem in time-series forecasting tasks. Data augmentation is an effective strategy for this task, but it has a limitati

applicationsarxiv-cs-lg
19 May 2026
Model Releases

Data Presentation Over Architecture: Resampling Strategies for Credit Risk Prediction with Tabular Foundation Models

DGX agent

arXiv:2605.18635v1 Announce Type: cross Abstract: Credit default prediction is a tabular learning problem with severe class imbalance, heterogeneous features, and tight latency budgets. Tabular Founda

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Designing streetscapes from street-view imagery using diffusion models

DGX agent

arXiv:2605.17527v1 Announce Type: new Abstract: Street-view imagery (SVI) is widely used to quantify key indicators of urban environment, such as green- ery, sky, or road view indices. However, existi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking

DGX agent

arXiv:2605.17451v1 Announce Type: new Abstract: Aerial object tracking has broad applications in public safety, emergency rescue, wildlife monitoring, and related fields. However, existing aerial trac

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

DocReward: A Document Reward Model for Structuring and Stylizing

DGX agent

arXiv:2510.11391v3 Announce Type: replace-cross Abstract: Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic

model-releasesarxiv-cs-ai
19 May 2026
Research

Dynamic robotic cloth folding with efficient Koopman operator-based model predictive control

DGX agent

arXiv:2605.18373v1 Announce Type: cross Abstract: Robotic cloth folding is a challenging task, particularly when considering dynamic folding tasks, which aim at folding cloth by fast motions that leve

researcharxiv-cs-lg
19 May 2026
Model Releases

Employing Vision-Language Models for Face Image Quality Assessment

DGX agent

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

EvoQRE: Modeling Bounded Rationality in Safety-Critical Traffic Simulation via Evolutionary Quantal Response Equilibrium

DGX agent

arXiv:2601.05653v2 Announce Type: replace Abstract: Existing traffic simulation frameworks for autonomous vehicles typically rely on imitation learning or game-theoretic approaches that solve for Nash

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

DGX agent

arXiv:2503.16492v3 Announce Type: replace-cross Abstract: ffective Human-Robot Interaction (HRI) is crucial for enhancing accessibility and usability in real-world robotics applications. However, exis

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints

DGX agent

arXiv:2509.06984v3 Announce Type: replace-cross Abstract: Federated Learning with LoRA fine-tuning offers an efficient and privacy-aware solution for institutions to collaboratively leverage their lar

model-releasesarxiv-cs-ai
19 May 2026
Agents

GEM: Gaussian Evolution Model for Occupancy Forecasting and Motion Planning

DGX agent

arXiv:2605.17682v1 Announce Type: new Abstract: Future 3D semantic occupancy forecasting and motion planning are central to autonomous driving, as they require models to reason about how surrounding s

agentsarxiv-cs-cv
19 May 2026
Model Releases

GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics

DGX agent

arXiv:2506.14009v2 Announce Type: replace Abstract: Autonomous drones capable of interpreting and executing high-level language instructions in unstructured environments remain a long-standing goal. Y

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

MCQ Difficulty Prediction via Modeling Learner Heterogeneity Using Data-Driven Cognitive Profiling

DGX agent

arXiv:2605.16290v1 Announce Type: cross Abstract: Predicting the difficulty of multiple-choice questions (MCQs) is important for effective assessment, yet current methods typically assume a unimodal s

model-releasesarxiv-cs-ai
19 May 2026
Research

Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks

DGX agent

arXiv:2605.16864v1 Announce Type: cross Abstract: Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-awa

researcharxiv-cs-ai
19 May 2026
Hardware

NanoQuant: Efficient Sub-1-Bit Quantization of Large Language Models

DGX agent

arXiv:2602.06694v2 Announce Type: replace Abstract: Weight-only quantization has become a standard approach for efficiently serving large language models (LLMs). However, existing methods fail to effi

hardwarearxiv-cs-lg
19 May 2026
Safety

OrbiSim: World Models as Differentiable Physics Engines for Embodied Intelligence

DGX agent

arXiv:2605.16395v1 Announce Type: cross Abstract: We present OrbiSim, a novel robotic simulation paradigm that redefines world models as a fully differentiable physics engine for embodied intelligence

safetyarxiv-cs-lg
19 May 2026
Model Releases

Self-supervised Hierarchical Visual Reasoning with World Model

DGX agent

arXiv:2605.17537v1 Announce Type: new Abstract: 3D open-world environments with adversarial opponents remain a core challenge for reinforcement learning due to their vast state spaces. Effective reaso

model-releasesarxiv-cs-ai
19 May 2026
Agents

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

DGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

agentsarxiv-cs-ai
19 May 2026
Research

Test-Time Hinting for Black-Box Vision-Language Models

DGX agent

arXiv:2605.16410v1 Announce Type: new Abstract: Test-time scaling (TTS) methods have proven highly effective for LLMs, yet their application to vision-language models (VLMs) remains relatively underex

researcharxiv-cs-cv
19 May 2026
Applications

Traces of Social Competence in Large Language Models

DGX agent

arXiv:2603.04161v2 Announce Type: replace Abstract: The False Belief Test (FBT) has been the main method for assessing Theory of Mind (ToM) and related socio-cognitive competencies. For Large Language

applicationsarxiv-cs-cl
19 May 2026
Applications

Training data attribution in diffusion models via mirrored unlearning and noise-consistent skew

DGX agent

arXiv:2605.17938v1 Announce Type: cross Abstract: Training data attribution (TDA) should enable generative model interpretability and foster a variety of related downstream tasks. Nonetheless, current

applicationsarxiv-cs-ai
19 May 2026
Research

Vision Foundation Models as Generalist Tokenizers for Image Generation

DGX agent

arXiv:2605.18390v1 Announce Type: new Abstract: In this work, we explore the largely unexplored direction of building a generalist image tokenizer directly on top of a frozen vision foundation model (

researcharxiv-cs-cv
19 May 2026
Research

WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points

DGX agent

arXiv:2605.17471v1 Announce Type: new Abstract: Quantization-aware training (QAT) is widely adopted to quantize language models by training full-precision weights using gradients from the quantized mo

researcharxiv-cs-lg
19 May 2026
Safety

ASRU: Activation Steering Meets Reinforcement Unlearning for Multimodal Large Language Models

DGX agent

arXiv:2605.15687v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) may memorize sensitive cross-modal information during pretraining, making machine unlearning (MU) crucial. Ex

safetyarxiv-cs-ai
18 May 2026
Model Releases

Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction

DGX agent

arXiv:2605.16077v1 Announce Type: new Abstract: Accurate assessment of cognitive decline from spontaneous speech remains challenging due to limited dataset size and class imbalance. In this work, we p

model-releasesarxiv-cs-cl
18 May 2026
Research

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models

DGX agent

arXiv:2601.21798v2 Announce Type: replace Abstract: Large Language Models(LLMs) have revolutionized text generation and multimodal perception,but their capabilities in 3D content generation remain und

researcharxiv-cs-cv
18 May 2026
Safety

EgoExo-WM: Unlocking Exo Video for Ego World Models

DGX agent

arXiv:2605.15477v1 Announce Type: new Abstract: Egocentric world models present a promising direction for enabling agents to predict and plan, but their performance is constrained by the limited avail

safetyarxiv-cs-cv
18 May 2026
Research

Extrapolation Guarantees for Perturbation Modeling Under the Additive Latent Shift Assumption

DGX agent

arXiv:2504.18522v3 Announce Type: replace-cross Abstract: We consider the problem of modeling the effects of perturbations like gene knockouts on measurements such as single-cell RNA counts. Given dat

researcharxiv-cs-lg
18 May 2026
Model Releases

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

DGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

model-releasesarxiv-cs-cl
18 May 2026
Research

Few-Step Diffusion Language Models via Trajectory Self-Distillation

DGX agent

arXiv:2602.12262v3 Announce Type: replace Abstract: Diffusion large language models (DLLMs) have emerged as powerful generative models with the promise of fast text generation through parallel decodin

researcharxiv-cs-cl
18 May 2026
Model Releases

Large Language Models Could Be Rote Learners

DGX agent

arXiv:2504.08300v5 Announce Type: replace-cross Abstract: Benchmark-based evaluation, e.g., multiple-choice questions (MCQs) and open-ended questions (OEQs), is widely used for evaluating Large Langua

model-releasesarxiv-cs-ai
18 May 2026
← Previous
1…135136137138139…1030
Next →