AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,596 results
Safety

Structured Agent Distillation for Large Language Model

DGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

safetyarxiv-cs-ai
28 May 2026
Safety

The Attentional White Bear Effect in Transformer Language Models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.28639v1 Announce Type: cross Abstract: Instruction-based suppression is widely used to prevent language models from generating prohibited content, yet it remains unclear whether suppression

safetyarxiv-cs-ai
28 May 2026
Research

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

DGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

researcharxiv-cs-ai
28 May 2026
Research

Transfer learning RGB models to hyperspectral images with trainable tensor decompositions

DGX agent

arXiv:2605.28331v1 Announce Type: new Abstract: Transfer learning makes it possible to use large vision networks on a variety of domains, by specializing their models' general filters to new tasks. Ho

researcharxiv-cs-cv
28 May 2026
Safety

VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking

DGX agent

arXiv:2605.28083v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models have emerged as powerful generalist policies, their severe vulnerability to adversarial patches significantly

safetyarxiv-cs-cv
28 May 2026
Local Ai

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models

DGX agent

arXiv:2605.27997v1 Announce Type: cross Abstract: Large language models frequently generate toxic, hateful, or harmful content, yet existing mitigation methods rely on costly retraining or output-leve

local-aiarxiv-cs-ai
28 May 2026
Research

Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

DGX agent

arXiv:2602.07120v2 Announce Type: replace Abstract: Language models (LMs) tend to memorize portions of their training data and emit verbatim spans. When the underlying sources are sensitive or copyrig

researcharxiv-cs-cl
27 May 2026
Safety

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

DGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

safetyarxiv-cs-ai
27 May 2026
Tutorials

Can VLA Models Learn from Real-World Data Continually without Forgetting?

DGX agent

arXiv:2605.26820v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising foundation for general-purpose robotics. However, their successful deployment in real-world scen

tutorialsarxiv-cs-ro
27 May 2026
Industry

Cisco report finds no closed frontier AI model is safe from multi-turn attacks

DGX agent

A new report out today from Cisco Systems Inc. argues that none of the closed flagship large language models it tested can be considered safe once an attacker is allowed to push past a single prompt,

industrysiliconangle
27 May 2026
Research

EEG-FM-Audit: A Systematic Evaluation and Analysis Pipeline for EEG Foundation Models

DGX agent

arXiv:2605.26910v1 Announce Type: cross Abstract: Large EEG Foundation Models (FMs) have shown great potential for decoding EEG signals across diverse cognitive tasks. However, existing EEG-FM studies

researcharxiv-cs-ai
27 May 2026
Research

HydraPrompt: An Adaptive and Asymmetric Framework of Vision-Language Models for Synthetic Image Detection

DGX agent

arXiv:2605.26421v1 Announce Type: new Abstract: The rapid evolution of generative models has precipitated a proliferation of fabricated content, posing significant challenges to existing Synthetic Ima

researcharxiv-cs-cv
27 May 2026
Research

Innovative Silicosis and Pneumonia Classification: Leveraging Graph Transformer Post-hoc Modeling and Ensemble Techniques

DGX agent

arXiv:2501.00520v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the classification and detection of Silicosis-related lung inflammation. Our main contributions include

researcharxiv-cs-cv
27 May 2026
Research

Learning Nonlinear Factor Models with Unknown Monotone Links from Incomplete and Noisy Data

DGX agent

arXiv:2605.26271v1 Announce Type: cross Abstract: We study a nonlinear factor model in which observed responses depend on low-rank latent factors through an unknown monotone link function. This settin

researcharxiv-cs-lg
27 May 2026
Tutorials

Lifting Data-Tracing Machine Unlearning to Knowledge-Tracing for Foundation Models

DGX agent

arXiv:2506.11253v2 Announce Type: replace Abstract: Machine unlearning removes certain training data points and their influence from AI models (e.g., when a data owner revokes their consent to allow m

tutorialsarxiv-cs-cv
27 May 2026
Local Ai

Localizing Memorized Regions in Diffusion Models via Coordinate-Wise Curvature Differences

DGX agent

arXiv:2605.26756v1 Announce Type: new Abstract: Diffusion models can unintentionally memorize training samples, raising concerns about privacy and copyright. While recent methods can detect memorizati

local-aiarxiv-cs-lg
27 May 2026
Research

Model Merging on Loss Landscape: A Geometry Perspective

DGX agent

arXiv:2605.26693v1 Announce Type: cross Abstract: Model merging offers a promising avenue for knowledge integration and parallel development without retraining. Yet, existing methods either ignore the

researcharxiv-cs-ai
27 May 2026
Local Ai

MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training

DGX agent

arXiv:2605.26842v1 Announce Type: cross Abstract: The Muon optimizer has recently offered a promising alternative to AdamW for large language model training, leveraging matrix orthogonalization to pro

local-aiarxiv-cs-cl
27 May 2026
Model Releases

Multi-Agent Causal Discovery Using Large Language Models

DGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

model-releasesarxiv-cs-ai
27 May 2026
Research

PLAID: A Unified Data Model for Machine Learning on Heterogeneous Physics Simulations

DGX agent

arXiv:2505.02974v3 Announce Type: replace Abstract: Machine learning-based surrogate models have emerged as a powerful tool to accelerate simulation-driven scientific workflows, but their adoption is

researcharxiv-cs-lg
27 May 2026
Safety

Real Images, Worse Judgments: Evaluating Vision-Language Models on Concreteness and Imagery

DGX agent

arXiv:2605.27315v1 Announce Type: new Abstract: Visual inputs are often assumed to improve language understanding in multimodal models. We examine this assumption by asking whether vision-language mod

safetyarxiv-cs-cl
27 May 2026
Model Releases

Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models

DGX agent

arXiv:2605.27299v1 Announce Type: cross Abstract: Modern intrusion detection systems generate thousands of alerts daily, but alert fatigue severely limits security operations effectiveness due to too

model-releasesarxiv-cs-ai
27 May 2026
Applications

Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks

DGX agent

arXiv:2506.03627v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable performance across various tasks by effectively utilizing a prompting strategy. Howe

applicationsarxiv-cs-ai
27 May 2026
Model Releases

Self-Ensembling Vision-Language Models for Chart Data Extraction

DGX agent

arXiv:2605.27298v1 Announce Type: new Abstract: Charts effectively convey quantitative information, but the underlying data are often locked in image form, hindering reuse and analysis. Manually digit

model-releasesarxiv-cs-cl
27 May 2026
Research

Targeted Remasking: Replacing Token Editing with Token-to-Mask Refinement in Discrete Diffusion Language Models

DGX agent

arXiv:2605.26436v1 Announce Type: cross Abstract: Discrete masked diffusion language models such as LLaDA generate text through iterative denoising, where mask tokens are progressively replaced with p

researcharxiv-cs-ai
27 May 2026
Local Ai

“The future of AI is going to be local models running on extraordinary desktop hardware.” - @Jason This line from the recent @theallinpod hi…

DGX agent

“The future of AI is going to be local models running on extraordinary desktop hardware.” - @Jason This line from the recent @theallinpod hit hard. For years AI meant sending everything to the cloud,

local-aiyohei-nakajima--x
27 May 2026
Model Releases

Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts

DGX agent

arXiv:2605.26776v1 Announce Type: cross Abstract: In recent years, Deep Reinforcement Learning (DRL) has achieved substantial progress on Vehicle Routing Problems (VRPs). However, existing DRL-based m

model-releasesarxiv-cs-ai
27 May 2026
Agents

Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization

DGX agent

arXiv:2605.26501v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have transformed multi-modal understanding, excelling in tasks like image captioning and visual question answerin

agentsarxiv-cs-ai
27 May 2026
Safety

A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions

DGX agent

arXiv:2605.25536v1 Announce Type: cross Abstract: Context. Large language models (LLMs) are increasingly applied to code-generating tasks (CGTs) in software engineering. While reported results are pro

safetyarxiv-cs-ai
26 May 2026
Model Releases

Automated Benchmark Auditing for AI Agents and Large Language Models

DGX agent

arXiv:2605.26079v1 Announce Type: new Abstract: Modern AI benchmarks operate at a complexity that outpaces traditional verification methods. Tasks authored by domain experts often contain implicit ass

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs

DGX agent

arXiv:2605.21602v2 Announce Type: replace Abstract: Many safety and alignment failures of large language models (LLMs) occur due to out-of-distribution (OOD) situations: unusual prompt or response pat

model-releasesarxiv-cs-ai
26 May 2026
Research

Better, Faster: Harnessing Self-Improvement in Large Reasoning Models

DGX agent

arXiv:2605.24998v1 Announce Type: new Abstract: Self-improvement training enables the large reasoning models (LRMs) to improve themselves by self-generating reasoning trajectories as training data wit

researcharxiv-cs-cl
26 May 2026
Agents

Beyond Predefined Learning Objects: A Thinking-Learning Interaction Model for Up-to-Date Autonomous Robot Learning

DGX agent

arXiv:2605.23987v1 Announce Type: new Abstract: Autonomous robots operating in open and changing environments cannot always rely on predefined inputs, outputs, and action routines. Although existing l

agentsarxiv-cs-ai
26 May 2026
Model Releases

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

DGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

model-releasesarxiv-cs-ai
26 May 2026
Local Ai

Communication-Efficient Hybrid Language Model via Uncertainty-Aware Opportunistic and Compressed Transmission

DGX agent

arXiv:2505.11788v2 Announce Type: replace-cross Abstract: To support emerging language-based applications using dispersed and heterogeneous computing resources, the hybrid language model (HLM) offers

local-aiarxiv-cs-lg
26 May 2026
Research

Coupled Variational Reinforcement Learning for Language Model General Reasoning

DGX agent

arXiv:2512.12576v3 Announce Type: replace-cross Abstract: While reinforcement learning has achieved impressive progress in language model reasoning, it is constrained by the requirement for verifiable

researcharxiv-cs-ai
26 May 2026
Safety

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

DGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

safetyarxiv-cs-ro
26 May 2026
Research

Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Performance and Metacognition

DGX agent

arXiv:2605.25856v1 Announce Type: cross Abstract: Large Language Model interfaces are increasingly verbose, exposing intermediate reasoning traces alongside final answers. Traces are framed as transpa

researcharxiv-cs-ai
26 May 2026
Model Releases

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

DGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

model-releasesarxiv-cs-ai
26 May 2026
Research

Fuzzy PyTorch: Rapid Numerical Variability Evaluation for Deep Learning Models

DGX agent

arXiv:2605.25991v1 Announce Type: new Abstract: We introduce Fuzzy PyTorch, a framework for rapid evaluation of numerical variability in deep learning (DL) models. As DL is increasingly applied to div

researcharxiv-cs-lg
26 May 2026
Research

High-fidelity Modeling of Full-scale Pressurized Water Reactor Flow Fields for Machine Learning Applications

DGX agent

arXiv:2605.24763v1 Announce Type: new Abstract: This work presents a high-fidelity computational fluid dynamics (CFD) and data-driven modeling framework for assembly-level flow characterization in a f

researcharxiv-cs-lg
26 May 2026
Model Releases

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

DGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

model-releasesarxiv-cs-ai
26 May 2026
Safety

How Neural Reward Models Learn Features for Policy Optimization: A Single-Index Analysis

DGX agent

arXiv:2605.24749v1 Announce Type: cross Abstract: Reward modeling is not only a prediction problem: in KL-regularized policy optimization, the learned reward is exponentiated to define the deployed po

safetyarxiv-cs-lg
26 May 2026
Safety

Improving Ensemble CAPE Forecasts with a Diffusion Model Incorporating Aerosol Information

DGX agent

arXiv:2605.24009v1 Announce Type: cross Abstract: Convective available potential energy (CAPE) is an important variable for forecasting severe weather and understanding deep convection and precipitati

safetyarxiv-cs-lg
26 May 2026
Applications

Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation

DGX agent

arXiv:2605.24647v1 Announce Type: new Abstract: Personalized dialogue requires more than recalling explicit user histories: systems also need to infer hidden user states that evolve through interactio

applicationsarxiv-cs-cl
26 May 2026
Tutorials

MultiPUFFIN: A Multimodal Domain-Constrained Foundation Model for Molecular Property Prediction of Small Molecules

DGX agent

arXiv:2603.00857v2 Announce Type: replace-cross Abstract: MultiPUFFIN is a domain-informed multimodal foundation model for predicting thermophysical properties of small molecules, addressing a critica

tutorialsarxiv-cs-ai
26 May 2026
Hardware

Paris 2.0: A Decentralized Diffusion Model for Video Generation

DGX agent

arXiv:2605.26064v1 Announce Type: cross Abstract: We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv

hardwarearxiv-cs-lg
26 May 2026
Safety

PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs

DGX agent

arXiv:2601.20539v3 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled automated heuristic design (AHD) for combinatorial optimization problems (COPs), but existing frameworks'

safetyarxiv-cs-ai
26 May 2026
← Previous
1…197198199200201…1263
Next →