AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
28 May 2026

Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors?

Model ReleasesDGX agent

arXiv:2508.11011v2 Announce Type: replace Abstract: Construction safety inspections typically involve a human inspector identifying safety concerns on-site. With the rise of powerful Vision Language M

Auditable Decision Models with Learned Abstention and Real-Time Steering

SafetyDGX agent

arXiv:2605.27768v1 Announce Type: new Abstract: Production AI systems often operate with incomplete, conflicting, or insufficient evidence. Forced classifiers collapse such cases into action labels, w

Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay

Model ReleasesDGX agent

arXiv:2605.28782v1 Announce Type: new Abstract: Discourse particles, such as extit{well} and extit{kind of}, are crucial components that enable LLMs to ``speak'' more like humans. They are used to con

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HEART: Achieving Timely Multi-Model Training for Vehicle-Edge-Cloud-Integrated Hierarchical Federated Learning

AgentsDGX agent

arXiv:2501.09934v3 Announce Type: replace-cross Abstract: The rapid growth of AI-enabled Internet of Vehicles (IoV) calls for efficient Machine Learning (ML) solutions that can handle high vehicular m

Hierarchical Prompt-Domain Control and Learning for Resource-Constrained Agentic Language Models

AgentsDGX agent

arXiv:2605.27703v1 Announce Type: new Abstract: Large Language Models are increasingly deployed inside agentic systems, where they must follow structured protocols, adapt to evolving states, and opera

Imitating and Finetuning Model Predictive Control for Robust and Symmetric Quadrupedal Locomotion

SafetyDGX agent

arXiv:2311.02304v3 Announce Type: replace Abstract: Control of legged robots is a challenging problem that has been investigated by different approaches, such as model-based control and learning algor

Large Language Models Approach Expert Pedagogical Quality in Math Tutoring but Differ in Instructional and Linguistic Profiles

AgentsDGX agent

arXiv:2512.20780v3 Announce Type: replace Abstract: Recent work has explored the use of large language models (LLMs) to generate tutoring responses in mathematics, yet it remains unclear how closely t

Law of Neural Interaction: Depth-Width Shape, Interaction Efficiency, and Generalization

Model ReleasesDGX agent

arXiv:2605.27989v1 Announce Type: new Abstract: The guidance of scaling laws has increased the resource demands of modern large language models (LLMs), yet it remains questionable whether these models

Localizing Input Uncertainty Quantification for Large Language Models via Shapley Values

Local AiDGX agent

arXiv:2605.28170v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly integrated into high-stakes decision-making, the ability to reliably quantify uncertainty has become a

Mitigating Adaptive Attacks against Reasoning Models with Activation Consistency Training

SafetyDGX agent

arXiv:2605.28467v1 Announce Type: new Abstract: As LLMs gain stronger reasoning capabilities, their extended chain-of-thought introduces new degrees of complexity for defending against adversarial jai

Modeling Vehicle-Type-Specific Pedestrian Crash Avoidance Behavior in Safety-Critical Interactions Using Smooth-Mamba Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.28552v1 Announce Type: new Abstract: As automated vehicles (AVs) increasingly share roadways with human-driven vehicles (HDVs), understanding how pedestrians respond to different vehicle ty

Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models

SafetyDGX agent

arXiv:2603.01766v2 Announce Type: replace Abstract: Despite the rapid progress of vision-language-action (VLA) models, the prevailing practice of predicting action chunks as discrete waypoints remains

Poison with Style: A Practical Poisoning Attack on Code Large Language Models

ResearchDGX agent

arXiv:2605.27631v1 Announce Type: cross Abstract: Code Large Language Models (CLLMs) serve as the core of modern code agents, enabling developers to automate complex software development tasks. In thi

Prompt Codebooks: Discrete Compositional Optimization for Language Model Instruction Refinement

Model ReleasesDGX agent

arXiv:2605.28360v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has driven significant gains in LLM-based agentic workflows. However, existing methods treat each task's prompt as a

Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization

SafetyDGX agent

arXiv:2605.27906v1 Announce Type: new Abstract: Multimodal Large Reasoning Models introduce the reasoning paradigm, demonstrating strong capabilities on complex vision-language tasks. However, they st

REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading

ApplicationsDGX agent

arXiv:2605.27402v1 Announce Type: cross Abstract: Open-ended grading is central to equitable and personalized education, yet manual grading remains time-consuming and costly, underscoring the need for

Refining Multidimensional Video Reward Models via Disentangled Influence Functions

SafetyDGX agent

arXiv:2605.28203v1 Announce Type: new Abstract: As Text-to-Video (T2V) generation models continue to evolve, the complexity of video evaluation necessitates a fine-grained assessment across various ax

Revisiting Anthropomorphic Reflection Markers in Large Language Model Reasoning

ResearchDGX agent

arXiv:2605.28305v1 Announce Type: cross Abstract: Large Language Models (LLMs) often produce explicit reflective traces during complex reasoning, accompanied by anthropomorphic markers such as wait, h

Structured Agent Distillation for Large Language Model

SafetyDGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

The Attentional White Bear Effect in Transformer Language Models

SafetyDGX agent

arXiv:2605.28639v1 Announce Type: cross Abstract: Instruction-based suppression is widely used to prevent language models from generating prohibited content, yet it remains unclear whether suppression

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

ResearchDGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

Transfer learning RGB models to hyperspectral images with trainable tensor decompositions

ResearchDGX agent

arXiv:2605.28331v1 Announce Type: new Abstract: Transfer learning makes it possible to use large vision networks on a variety of domains, by specializing their models' general filters to new tasks. Ho

VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking

SafetyDGX agent

arXiv:2605.28083v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models have emerged as powerful generalist policies, their severe vulnerability to adversarial patches significantly

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models

Local AiDGX agent

arXiv:2605.27997v1 Announce Type: cross Abstract: Large language models frequently generate toxic, hateful, or harmful content, yet existing mitigation methods rely on costly retraining or output-leve

27 May 2026

Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

ResearchDGX agent

arXiv:2602.07120v2 Announce Type: replace Abstract: Language models (LMs) tend to memorize portions of their training data and emit verbatim spans. When the underlying sources are sensitive or copyrig

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

SafetyDGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

Can VLA Models Learn from Real-World Data Continually without Forgetting?

TutorialsDGX agent

arXiv:2605.26820v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising foundation for general-purpose robotics. However, their successful deployment in real-world scen

Cisco report finds no closed frontier AI model is safe from multi-turn attacks

IndustryDGX agent

A new report out today from Cisco Systems Inc. argues that none of the closed flagship large language models it tested can be considered safe once an attacker is allowed to push past a single prompt,

EEG-FM-Audit: A Systematic Evaluation and Analysis Pipeline for EEG Foundation Models

ResearchDGX agent

arXiv:2605.26910v1 Announce Type: cross Abstract: Large EEG Foundation Models (FMs) have shown great potential for decoding EEG signals across diverse cognitive tasks. However, existing EEG-FM studies

HydraPrompt: An Adaptive and Asymmetric Framework of Vision-Language Models for Synthetic Image Detection

ResearchDGX agent

arXiv:2605.26421v1 Announce Type: new Abstract: The rapid evolution of generative models has precipitated a proliferation of fabricated content, posing significant challenges to existing Synthetic Ima

Innovative Silicosis and Pneumonia Classification: Leveraging Graph Transformer Post-hoc Modeling and Ensemble Techniques

ResearchDGX agent

arXiv:2501.00520v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the classification and detection of Silicosis-related lung inflammation. Our main contributions include

Learning Nonlinear Factor Models with Unknown Monotone Links from Incomplete and Noisy Data

ResearchDGX agent

arXiv:2605.26271v1 Announce Type: cross Abstract: We study a nonlinear factor model in which observed responses depend on low-rank latent factors through an unknown monotone link function. This settin

Lifting Data-Tracing Machine Unlearning to Knowledge-Tracing for Foundation Models

TutorialsDGX agent

arXiv:2506.11253v2 Announce Type: replace Abstract: Machine unlearning removes certain training data points and their influence from AI models (e.g., when a data owner revokes their consent to allow m

Localizing Memorized Regions in Diffusion Models via Coordinate-Wise Curvature Differences

Local AiDGX agent

arXiv:2605.26756v1 Announce Type: new Abstract: Diffusion models can unintentionally memorize training samples, raising concerns about privacy and copyright. While recent methods can detect memorizati

Model Merging on Loss Landscape: A Geometry Perspective

ResearchDGX agent

arXiv:2605.26693v1 Announce Type: cross Abstract: Model merging offers a promising avenue for knowledge integration and parallel development without retraining. Yet, existing methods either ignore the

MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training

Local AiDGX agent

arXiv:2605.26842v1 Announce Type: cross Abstract: The Muon optimizer has recently offered a promising alternative to AdamW for large language model training, leveraging matrix orthogonalization to pro

Multi-Agent Causal Discovery Using Large Language Models

Model ReleasesDGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

PLAID: A Unified Data Model for Machine Learning on Heterogeneous Physics Simulations

ResearchDGX agent

arXiv:2505.02974v3 Announce Type: replace Abstract: Machine learning-based surrogate models have emerged as a powerful tool to accelerate simulation-driven scientific workflows, but their adoption is

Real Images, Worse Judgments: Evaluating Vision-Language Models on Concreteness and Imagery

SafetyDGX agent

arXiv:2605.27315v1 Announce Type: new Abstract: Visual inputs are often assumed to improve language understanding in multimodal models. We examine this assumption by asking whether vision-language mod

Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models

Model ReleasesDGX agent

arXiv:2605.27299v1 Announce Type: cross Abstract: Modern intrusion detection systems generate thousands of alerts daily, but alert fatigue severely limits security operations effectiveness due to too

Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks

ApplicationsDGX agent

arXiv:2506.03627v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable performance across various tasks by effectively utilizing a prompting strategy. Howe

Self-Ensembling Vision-Language Models for Chart Data Extraction

Model ReleasesDGX agent

arXiv:2605.27298v1 Announce Type: new Abstract: Charts effectively convey quantitative information, but the underlying data are often locked in image form, hindering reuse and analysis. Manually digit

Targeted Remasking: Replacing Token Editing with Token-to-Mask Refinement in Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2605.26436v1 Announce Type: cross Abstract: Discrete masked diffusion language models such as LLaDA generate text through iterative denoising, where mask tokens are progressively replaced with p

“The future of AI is going to be local models running on extraordinary desktop hardware.” - @Jason This line from the recent @theallinpod hi…

Local AiDGX agent

“The future of AI is going to be local models running on extraordinary desktop hardware.” - @Jason This line from the recent @theallinpod hit hard. For years AI meant sending everything to the cloud,

Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2605.26776v1 Announce Type: cross Abstract: In recent years, Deep Reinforcement Learning (DRL) has achieved substantial progress on Vehicle Routing Problems (VRPs). However, existing DRL-based m

Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization

AgentsDGX agent

arXiv:2605.26501v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have transformed multi-modal understanding, excelling in tasks like image captioning and visual question answerin

26 May 2026

A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions

SafetyDGX agent

arXiv:2605.25536v1 Announce Type: cross Abstract: Context. Large language models (LLMs) are increasingly applied to code-generating tasks (CGTs) in software engineering. While reported results are pro

Automated Benchmark Auditing for AI Agents and Large Language Models

Model ReleasesDGX agent

arXiv:2605.26079v1 Announce Type: new Abstract: Modern AI benchmarks operate at a complexity that outpaces traditional verification methods. Tasks authored by domain experts often contain implicit ass

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs

Model ReleasesDGX agent

arXiv:2605.21602v2 Announce Type: replace Abstract: Many safety and alignment failures of large language models (LLMs) occur due to out-of-distribution (OOD) situations: unusual prompt or response pat

Better, Faster: Harnessing Self-Improvement in Large Reasoning Models

ResearchDGX agent

arXiv:2605.24998v1 Announce Type: new Abstract: Self-improvement training enables the large reasoning models (LRMs) to improve themselves by self-generating reasoning trajectories as training data wit

Beyond Predefined Learning Objects: A Thinking-Learning Interaction Model for Up-to-Date Autonomous Robot Learning

AgentsDGX agent

arXiv:2605.23987v1 Announce Type: new Abstract: Autonomous robots operating in open and changing environments cannot always rely on predefined inputs, outputs, and action routines. Although existing l

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

Model ReleasesDGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

Communication-Efficient Hybrid Language Model via Uncertainty-Aware Opportunistic and Compressed Transmission

Local AiDGX agent

arXiv:2505.11788v2 Announce Type: replace-cross Abstract: To support emerging language-based applications using dispersed and heterogeneous computing resources, the hybrid language model (HLM) offers

Coupled Variational Reinforcement Learning for Language Model General Reasoning

ResearchDGX agent

arXiv:2512.12576v3 Announce Type: replace-cross Abstract: While reinforcement learning has achieved impressive progress in language model reasoning, it is constrained by the requirement for verifiable

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

SafetyDGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Performance and Metacognition

ResearchDGX agent

arXiv:2605.25856v1 Announce Type: cross Abstract: Large Language Model interfaces are increasingly verbose, exposing intermediate reasoning traces alongside final answers. Traces are framed as transpa

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

Model ReleasesDGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

Fuzzy PyTorch: Rapid Numerical Variability Evaluation for Deep Learning Models

ResearchDGX agent

arXiv:2605.25991v1 Announce Type: new Abstract: We introduce Fuzzy PyTorch, a framework for rapid evaluation of numerical variability in deep learning (DL) models. As DL is increasingly applied to div

High-fidelity Modeling of Full-scale Pressurized Water Reactor Flow Fields for Machine Learning Applications

ResearchDGX agent

arXiv:2605.24763v1 Announce Type: new Abstract: This work presents a high-fidelity computational fluid dynamics (CFD) and data-driven modeling framework for assembly-level flow characterization in a f

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

← Previous
1…157158159160161…1010
Next →