AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Safety

Multi-layer attentive probing improves transfer of audio representations for bioacoustics

DGX agent

arXiv:2605.10494v1 Announce Type: cross Abstract: Probing heads map the representations learned from audio by a machine learning model to downstream task labels and are a key component in evaluating r

safetyarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Narrative Landscape: Mapping Narrative Dispositions Across LLMs

DGX agent

arXiv:2605.08742v1 Announce Type: cross Abstract: This study proposes a quantitative framework for profiling LLM dispositions as stable, model-specific regularities in output under repeated, controlle

researcharxiv-cs-ai
12 May 2026
Research

Neuroprobe: Evaluating Intracranial Brain Responses to Naturalistic Stimuli

DGX agent

arXiv:2509.21671v2 Announce Type: replace Abstract: High-resolution neural datasets enable foundation models for the next generation of brain-computer interfaces and neurological treatments. The commu

researcharxiv-cs-lg
12 May 2026
Applications

NoTVLA: Semantics-Preserving Robot Adaptation via Narrative Action Interfaces

DGX agent

arXiv:2510.03895v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models represent a pivotal advance in embodied intelligence, yet they confront critical barriers to real-world de

applicationsarxiv-cs-cv
12 May 2026
Safety

On Variance Reduction in Learning Mean Flows

DGX agent

arXiv:2605.09235v1 Announce Type: cross Abstract: One-step generative modeling has emerged as a leading approach to amortize the inference cost of diffusion and flow-matching models. Among distillatio

safetyarxiv-cs-ai
12 May 2026
Model Releases

OracleTSC: Oracle-Informed Reward Hurdle and Uncertainty Regularization for Traffic Signal Control

DGX agent

arXiv:2605.08516v1 Announce Type: new Abstract: Transparent decision-making is essential for traffic signal control (TSC) systems to earn public trust. However, traditional reinforcement learning-base

model-releasesarxiv-cs-ai
12 May 2026
Research

ORICF -- Open Robotics Inference and Control Framework

DGX agent

arXiv:2605.09656v1 Announce Type: new Abstract: Recent advances in artificial intelligence (AI) have enabled effective perception and language models for robots, but their deployment remains computati

researcharxiv-cs-ro
12 May 2026
Model Releases

Parallel Multi-Circuit Quantum Feature Fusion in Hybrid Quantum-Classical Convolutional Neural Networks for Breast Tumor Classification

DGX agent

arXiv:2512.02066v2 Announce Type: replace-cross Abstract: Quantum machine learning has emerged as a promising approach to improve feature extraction and classification tasks in high-dimensional data d

model-releasesarxiv-cs-ai
12 May 2026
Safety

Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs

DGX agent

arXiv:2605.09422v1 Announce Type: new Abstract: Although Large Multimodal Models (LMMs) have achieved strong performance on general video understanding, their susceptibility to textual prior shortcuts

safetyarxiv-cs-cl
12 May 2026
Safety

PHMForge: Evaluating LLM Agents on Industrial Prognostics through MCP-Native, Algorithm-Grounded Tools

DGX agent

arXiv:2604.01532v2 Announce Type: replace Abstract: LLM agents are beginning to invoke industrial asset-management tools through the Model Context Protocol (MCP), yet whether they can act reliably on

safetyarxiv-cs-ai
12 May 2026
Agents

PnP-Corrector: A Universal Correction Framework for Coupled Spatiotemporal Forecasting

DGX agent

arXiv:2605.08935v1 Announce Type: new Abstract: Coupled spatiotemporal forecasting is important for predicting the future evolution of multiple interacting dynamical systems, such as in climate models

agentsarxiv-cs-ai
12 May 2026
Research

Predictive Radiomics for Evaluation of Cancer Immune SignaturE in Glioblastoma: the PRECISE-GBM study

DGX agent

arXiv:2605.10278v1 Announce Type: new Abstract: Background: Radiogenomics allows identification of radiological biomarkers for genomic phenotypes. In glioblastoma, these biomarkers could potentially c

researcharxiv-cs-lg
12 May 2026
Model Releases

Priority-Driven Control and Communication in Decentralized Multi-Agent Systems via Reinforcement Learning

DGX agent

arXiv:2605.10482v1 Announce Type: cross Abstract: Event-triggered control provides a mechanism for avoiding excessive use of constrained communication bandwidth in networked multi-agent systems. Howev

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ReaMOT: A Benchmark and Framework for Reasoning-based Multi-Object Tracking

DGX agent

arXiv:2505.20381v4 Announce Type: replace Abstract: Referring Multi-Object Tracking (RMOT) aims to track targets specified by language instructions. However, existing RMOT paradigms heavily rely on ex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage

DGX agent

arXiv:2604.01527v3 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

DGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark

DGX agent

arXiv:2605.10921v1 Announce Type: new Abstract: Memory is a critical component of robotic intelligence, as robots must rely on past observations and actions to accomplish long-horizon tasks in partial

model-releasesarxiv-cs-ro
12 May 2026
Safety

RuPLaR : Efficient Latent Compression of LLM Reasoning Chains with Rule-Based Priors From Multi-Step to One-Step

DGX agent

arXiv:2605.09346v1 Announce Type: cross Abstract: The Chain-of-Thought (CoT) paradigm, while enhancing the interpretability of Large Language Models (LLMs), is constrained by the inefficiencies and ex

safetyarxiv-cs-ai
12 May 2026
Model Releases

SciIntegrity-Bench: A Benchmark for Evaluating Academic Integrity in AI Scientist Systems

DGX agent

arXiv:2605.10246v1 Announce Type: new Abstract: AI scientist systems are increasingly deployed for autonomous research, yet their academic integrity has never been systematically evaluated. We introdu

model-releasesarxiv-cs-ai
12 May 2026
Research

SciLT: Long-tailed Image Classification under Scientific Image Domains

DGX agent

arXiv:2604.03687v2 Announce Type: replace Abstract: Long-tailed recognition has benefited from foundation models and fine-tuning paradigms, yet existing studies and benchmarks are mainly confined to n

researcharxiv-cs-cv
12 May 2026
Model Releases

Semi-Supervised Neural Super-Resolution for Mesh-Based Simulations

DGX agent

arXiv:2605.09284v1 Announce Type: cross Abstract: Mesh-based simulations provide high-fidelity solutions to partial differential equations (PDEs), but achieving such accuracy typically requires fine m

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Set Prediction for Next-Day Active Fire Forecasting

DGX agent

arXiv:2605.10298v1 Announce Type: new Abstract: Accurate next-day active fire forecasts can support early warning, disaster response, forest risk assessment, and downstream estimation of fire-related

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

simpleposter: a simple baseline for product poster generation

DGX agent

arXiv:2605.08784v1 Announce Type: new Abstract: Product poster generation poses distinct challenges beyond general poster design, requiring both faithful preservation of product appearance and precise

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SpectraLLM: Uncovering the Ability of LLMs for Molecular Structure Elucidation from Multi-Spectral Data

DGX agent

arXiv:2508.08441v3 Announce Type: replace-cross Abstract: Automated molecular structure elucidation remains challenging, as existing approaches often depend on pre-compiled databases or restrict thems

model-releasesarxiv-cs-lg
12 May 2026
Research

Structured Recurrent Mixers for Massively Parallelized Sequence Generation

DGX agent

arXiv:2605.08696v1 Announce Type: new Abstract: Over the last two decades, language modeling has experienced a shift from predominantly recurrent architectures that process tokens sequentially during

researcharxiv-cs-cl
12 May 2026
Research

Tensor Product Representation Probes Reveal Shared Structure Across Linear Directions

DGX agent

arXiv:2605.09967v1 Announce Type: new Abstract: While researchers are finding concepts represented as linear directions in language models, a bag of linear directions fails to capture relational struc

researcharxiv-cs-lg
12 May 2026
Tutorials

TextBridgeGNN: Pre-training Graph Neural Network for Cross-Domain Recommendation via Text-Guided Transfer

DGX agent

arXiv:2601.02366v2 Announce Type: replace-cross Abstract: Graph-based recommendation has achieved great success in recent years. The classical graph recommendation model utilizes ID embedding to store

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

The Alpha Blending Hypothesis: Compositing Shortcut in Deepfake Detection

DGX agent

arXiv:2605.10334v1 Announce Type: new Abstract: Recent deepfake detection methods demonstrate improved cross-dataset generalization, yet the underlying mechanisms remain underexplored. We introduce th

model-releasesarxiv-cs-cv
12 May 2026
Research

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space

DGX agent

arXiv:2605.09883v1 Announce Type: cross Abstract: As current Multimodal Large Language Models rapidly saturate canonical visual reasoning benchmarks, a key question emerges: do these strong scores gen

researcharxiv-cs-ai
12 May 2026
Safety

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans

DGX agent

arXiv:2605.08837v1 Announce Type: cross Abstract: Abstract concepts - justice, theory, availability - have no single perceivable referent; in the human brain, their meaning emerges from a web of exper

safetyarxiv-cs-ai
12 May 2026
Research

The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime

DGX agent

arXiv:2605.10601v1 Announce Type: new Abstract: AI deployment in sensitive domains such as health care, credit, employment, and criminal justice is often treated as unsafe to authorize until model int

researcharxiv-cs-ai
12 May 2026
Model Releases

ThreatCore: A Benchmark for Explicit and Implicit Threat Detection

DGX agent

arXiv:2605.10563v1 Announce Type: cross Abstract: Threat detection in Natural Language Processing lacks consistent definitions and standardized benchmarks, and is often conflated with broader phenomen

model-releasesarxiv-cs-ai
12 May 2026
Safety

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning

DGX agent

arXiv:2508.20697v3 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in capability, so do the risks of harmful misuse through fine-tuning. While most prior studie

safetyarxiv-cs-cl
12 May 2026
Safety

Towards Customized Multimodal Role-Play

DGX agent

arXiv:2605.08129v1 Announce Type: new Abstract: Unified multimodal understanding and generation models enable richer human-AI interaction. Yet jointly customizing a character's persona, dialogue style

safetyarxiv-cs-lg
12 May 2026
Applications

Towards Robust Sequential Decomposition for Complex Image Editing

DGX agent

arXiv:2605.09233v1 Announce Type: cross Abstract: Recent advances in visual generative models have enabled high-fidelity image editing guided by human instructions. However, these models often struggl

applicationsarxiv-cs-ai
12 May 2026
Safety

TrajTok: Learning Trajectory Tokens enables better Video Understanding

DGX agent

arXiv:2602.22779v2 Announce Type: replace Abstract: Tokenization in video models, typically through patchification, generates an excessive and redundant number of tokens. This severely limits video ef

safetyarxiv-cs-cv
12 May 2026
Model Releases

Transcoda: End-to-End Zero-Shot Optical Music Recognition via Data-Centric Synthetic Training

DGX agent

arXiv:2605.10835v1 Announce Type: new Abstract: Optical Music Recognition (OMR), the task of transcribing sheet music into a structured textual representation, is currently bottlenecked by a lack of l

model-releasesarxiv-cs-cv
12 May 2026
Safety

Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning

DGX agent

arXiv:2605.08765v1 Announce Type: cross Abstract: Unlearning in large language models (LLMs) aims to remove harmful training data while preserving overall utility. However, we find that existing metho

safetyarxiv-cs-ai
12 May 2026
Model Releases

UserGPT Technical Report

DGX agent

arXiv:2605.08766v1 Announce Type: cross Abstract: Personalized user understanding from large-scale digital traces remains a fundamental challenge. Traditional user profiling methods rely on discrimina

model-releasesarxiv-cs-cl
12 May 2026
Safety

Users as Annotators: LLM Preference Learning from Comparison Mode

DGX agent

arXiv:2510.13830v2 Announce Type: replace-cross Abstract: Pairwise preference data have played an important role in the alignment of large language models (LLMs). Each sample of such data consists of

safetyarxiv-cs-ai
12 May 2026
Agents

When Agents Say One Thing and Do Another: Validating Elicited Beliefs from LLMs

DGX agent

arXiv:2602.06286v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings where good decisions require forming beliefs over the probability of

agentsarxiv-cs-ai
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

DGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Hierarchical Ensemble Pipeline for Anomaly Detection in ESA Satellite Telemetry

DGX agent

arXiv:2605.06681v1 Announce Type: cross Abstract: A hierarchical ensemble pipeline is introduced to address anomaly detection in multivariate telemetry data provided by European Space Agency (ESA). Th

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Reproducible Multi-Architecture Baseline for Token-Level Chinese Metaphor Identification under the MIPVU Framework

DGX agent

arXiv:2605.07170v1 Announce Type: new Abstract: Metaphor is pervasive in everyday language, yet token-level computational identification of metaphor-related words in Chinese under the MIPVU framework

model-releasesarxiv-cs-cl
11 May 2026
Safety

A Systematic Investigation of The RL-Jailbreaker in LLMs

DGX agent

arXiv:2605.07032v1 Announce Type: cross Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversa

safetyarxiv-cs-ai
11 May 2026
Research

Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR

DGX agent

arXiv:2605.07137v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a highly effective method for improving the reasoning abilities of Large Language Mod

researcharxiv-cs-ai
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…511512513514515…1109
Next →