AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Safety

GIBLy: Improving 3D Semantic Segmentation through an Architecture-Agnostic Lightweight Geometric Inductive Bias Layer

DGX agent

arXiv:2605.24243v1 Announce Type: cross Abstract: In 3D scene understanding, deep learning models rely on large models and extensive training to capture basic geometric structures that are present in

safetyarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Grouter: Decoupling Routing from Representation for Accelerated MoE Training

DGX agent

arXiv:2603.06626v2 Announce Type: replace-cross Abstract: Traditional Mixture-of-Experts (MoE) training typically proceeds without any structural priors, effectively requiring the model to simultaneou

safetyarxiv-cs-ai
26 May 2026
Model Releases

HiMed: Incentivizing Hindi Reasoning in Medical LLMs

DGX agent

arXiv:2605.24635v1 Announce Type: new Abstract: Medical large language models hold promise for reducing healthcare disparities, yet Hindi remains severely underrepresented. While medical LLMs excel in

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

INDUCTION: Finite-Structure Concept Synthesis in First-Order Logic

DGX agent

arXiv:2602.18956v3 Announce Type: replace Abstract: We introduce INDUCTION, a benchmark for finite structure concept synthesis in first order logic. Given small finite relational worlds with extension

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

DGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

JacQuant: STE-Free Quantization-Aware Training via Learned Jacobian Surrogates

DGX agent

arXiv:2605.25469v1 Announce Type: new Abstract: Quantization-aware training (QAT) is widely deployed but typically relies on the Straight-Through Estimator (STE), which passes gradients through non-di

model-releasesarxiv-cs-lg
26 May 2026
Research

Learning dynamical systems with biochemically informed neural ordinary differential equations

DGX agent

arXiv:2605.24170v1 Announce Type: cross Abstract: Ordinary differential equation models of biochemical reactions are often formulated as stoichiometric systems in which the dynamics arise from a colle

researcharxiv-cs-lg
26 May 2026
Model Releases

Learning Sparse Compositional Functions with Norm-Constrained Neural Networks

DGX agent

arXiv:2605.25608v1 Announce Type: cross Abstract: The ability of deep neural networks to learn hierarchical features is widely regarded as a key mechanism underlying their success in high-dimensional

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional

DGX agent

arXiv:2605.24699v1 Announce Type: new Abstract: Most reported gains on agentic-LLM clinical benchmarks are often attributed to prompt engineering, yet our results suggest that larger improvements can

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression

DGX agent

arXiv:2605.22337v2 Announce Type: replace Abstract: The KV cache used in large language models has linearly growing time complexity, so LLMs face memory blow-up and reduced decoding efficiency when th

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

MR-LiDAR: A Multi-Resolution Roadside LiDAR Benchmark for Perception Diagnostics and Deployment Guidance

DGX agent

arXiv:2605.24777v1 Announce Type: new Abstract: LiDAR model selection is a critical issue in roadside sensing systems, as it directly determines both perception capability and deployment cost. However

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning

DGX agent

arXiv:2605.25842v1 Announce Type: new Abstract: Vision-language models (VLMs) increasingly rely on chain-of-thought (CoT) reasoning to solve complex multimodal tasks, but their large parameter sizes m

model-releasesarxiv-cs-ai
26 May 2026
Agents

Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations

DGX agent

arXiv:2605.25101v1 Announce Type: cross Abstract: In many industrial domains, the Functional Mock-up Interface (FMI) is used to exchange simulation models as Functional Mock-up Units (FMUs) across dif

agentsarxiv-cs-ai
26 May 2026
Model Releases

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing

DGX agent

arXiv:2605.24919v1 Announce Type: new Abstract: Hallucinations in Large Language Models (LLMs) represent a critical barrier to their reliable deployment, a vulnerability heavily exacerbated in non-Eng

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

MuNet: A Mutualistic Network for Joint 3D Human Mesh Recovery and 3D Clothed Human Reconstruction from Single Images

DGX agent

arXiv:2605.25861v1 Announce Type: cross Abstract: 3D human mesh recovery and 3D clothed human reconstruction are inherently related, yet they have long been studied in isolation, thereby overlooking t

model-releasesarxiv-cs-ai
26 May 2026
Research

NITP: Next Implicit Token Prediction for LLM Pre-training

DGX agent

arXiv:2605.24956v1 Announce Type: new Abstract: Standard next-token prediction (NTP) supervises language models solely through discrete labels in the output logit space. We argue that this sparse one-

researcharxiv-cs-cl
26 May 2026
Model Releases

ORACAL: A Robust and Explainable Multimodal Framework for Smart Contract Vulnerability Detection with Causal Graph Enrichment

DGX agent

arXiv:2603.28128v2 Announce Type: replace Abstract: Although Graph Neural Networks (GNNs) have shown promise for smart contract vulnerability detection, they still face significant limitations. Homoge

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

DGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

model-releasesarxiv-cs-ai
26 May 2026
Research

PromptAudit: Auditing Prompt Sensitivity in LLM-Based Vulnerability Detection

DGX agent

arXiv:2605.24171v1 Announce Type: cross Abstract: Large language models are increasingly used for vulnerability detection, yet their reliability under different prompt formulations remains uncharacter

researcharxiv-cs-ai
26 May 2026
Tutorials

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis

DGX agent

arXiv:2601.06870v2 Announce Type: replace-cross Abstract: Multimodal large language models have demonstrated strong ability in capturing semantic representations for multimodal sentiment analysis. The

tutorialsarxiv-cs-ai
26 May 2026
Model Releases

Quantifying the Impact of Translation Errors on Multilingual LLM Evaluation

DGX agent

arXiv:2605.24904v1 Announce Type: new Abstract: Machine-translated benchmarks are widely used to assess the multilingual capabilities of large language models (LLMs), yet translation errors in these b

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Quaternion Self-Attention with Shared Scores

DGX agent

arXiv:2605.24920v1 Announce Type: cross Abstract: Quaternion neural networks are parameter-efficient and model multidimensional dependencies by representing four related features as a single entity. H

model-releasesarxiv-cs-ai
26 May 2026
Research

ReactEmbed: A Plug-and-Play Module for Unifying Protein-Molecule Representations Guided by Biochemical Reaction Networks

DGX agent

arXiv:2501.18278v3 Announce Type: replace Abstract: State-of-the-art models represent proteins and molecules in separate embedding manifolds, limiting the modeling of systemic biological processes. We

researcharxiv-cs-lg
26 May 2026
Model Releases

READER: Reasoning-Enhanced AI-Generated Text Detection

DGX agent

arXiv:2605.25281v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have made it increasingly difficult to distinguish human-written text from AI-generated content. Many

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RealBench: Benchmarking Data-Driven Numerical Weather Forecasting Under Operational Conditions and Extreme Event Challenges

DGX agent

arXiv:2605.24945v1 Announce Type: cross Abstract: Accurate evaluation of weather forecasting models is critical for their reliable deployment in real-world applications. However, existing benchmarks p

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RECTOR: Priority-Aware Rule-Based Reranking for Compliance-Aware Autonomous Driving Trajectory Selection

DGX agent

arXiv:2605.25095v1 Announce Type: new Abstract: Autonomous driving stacks must pick one trajectory from a multi-modal candidate set; choosing by model confidence ignores safety, traffic-law, and comfo

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environments

DGX agent

arXiv:2602.19450v2 Announce Type: replace-cross Abstract: Trusted Execution Environments (TEEs) (e.g., Intel SGX and ArmTrustZone) aim to protect sensitive computation from a compromised operating sys

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark

DGX agent

arXiv:2605.26068v1 Announce Type: cross Abstract: Weakly supervised anomaly detection (WSAD) has developed in three primary directions: incomplete, inexact, and inaccurate supervision. However, these

model-releasesarxiv-cs-ai
26 May 2026
Research

Retrieved In-Context Principles from Previous Mistakes

DGX agent

arXiv:2407.05682v2 Announce Type: replace Abstract: In-context learning (ICL) has been instrumental in adapting Large Language Models (LLMs) to downstream tasks using correct input-output examples. Re

researcharxiv-cs-cl
26 May 2026
Model Releases

Retrying vs Resampling in AI Control

DGX agent

arXiv:2605.26047v1 Announce Type: new Abstract: AI coding scaffolds like Claude Code and Codex use extit{retrying}: blocking actions flagged as risky and continuing the trajectory. We study retrying f

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism

DGX agent

arXiv:2605.25565v1 Announce Type: cross Abstract: While Large Language Models (LLMs) are commonly fine-tuned to handle domain-specific tasks before being applied to vertical applications, adapting the

model-releasesarxiv-cs-cl
26 May 2026
Tutorials

Selective Test-Time Compute Scaling for Click-Through Rate Prediction via Uncertainty-Triggered Feature Path Exploration

DGX agent

arXiv:2605.24989v1 Announce Type: cross Abstract: Scaling test-time compute has proven highly effective for language models, yet this opportunity remains largely unexplored for industrial Click-Throug

tutorialsarxiv-cs-ai
26 May 2026
Local Ai

Signs Beat Floats: Low-Rank Double-Binary Adaptation for On-Device Fine-Tuning

DGX agent

arXiv:2605.24058v1 Announce Type: cross Abstract: On-device adaptation of large language models commonly keeps a quantized base model frozen while training and deploying a small, task-specific LoRA ad

local-aiarxiv-cs-ai
26 May 2026
Model Releases

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

DGX agent

arXiv:2605.25160v1 Announce Type: new Abstract: Mobile GUI agents powered by large language models have progressed rapidly, creating urgent needs for realistic and comprehensive evaluation. Existing b

model-releasesarxiv-cs-ai
26 May 2026
Safety

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

DGX agent

arXiv:2509.12672v2 Announce Type: replace Abstract: The volume of machine-generated content online has grown dramatically due to the widespread use of Large Language Models (LLMs), leading to new chal

safetyarxiv-cs-cl
26 May 2026
Model Releases

URS: A Unified Neural Routing Solver for Cross-Problem Zero-Shot Generalization

DGX agent

arXiv:2509.23413v2 Announce Type: replace Abstract: Multi-task neural routing solvers have emerged as a promising paradigm for their ability to solve multiple vehicle routing problems (VRPs) using a s

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

vAttention: Verified Sparse Attention

DGX agent

arXiv:2510.05688v2 Announce Type: replace-cross Abstract: State-of-the-art sparse attention methods for reducing decoding latency fall into two main categories: approximate top-k (and its extension, t

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Can We Trust Early Warnings? Leakage-Excluded Early Outcome Prediction from LMS Interaction Logs

DGX agent

arXiv:2605.25794v1 Announce Type: new Abstract: Early-warning models built from Learning Management System (LMS) logs aim to predict end-of-course outcomes early enough to enable timely learner suppor

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills

DGX agent

arXiv:2605.25832v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator r

model-releasesarxiv-cs-ai
26 May 2026
Applications

WhisTLE: Deeply Supervised, Text-Only Domain Adaptation for Pretrained Speech Recognition Transformers

DGX agent

arXiv:2509.10452v2 Announce Type: replace Abstract: Pretrained automatic speech recognition (ASR) models such as Whisper perform well but still need domain adaptation to handle unseen parlance. In man

applicationsarxiv-cs-cl
26 May 2026
Model Releases

Who judges the judges? Governance from metrics: a runtime framework for continuous LLM compliance monitoring

DGX agent

arXiv:2605.24737v1 Announce Type: cross Abstract: Current approaches to AI compliance treat conformity as a binary, audit-time verdict rather than a continuous, measurable property of production syste

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

WideDepth: Millimeter-Accurate Benchmark for Fisheye Depth Estimation

DGX agent

arXiv:2605.24074v1 Announce Type: cross Abstract: Fisheye cameras are increasingly adopted in robotics for near-field manipulation, navigation, and immersive perception, yet indoor depth benchmarks wi

model-releasesarxiv-cs-ro
26 May 2026
Applications

WISE: Web Information Satire and Fakeness Evaluation

DGX agent

arXiv:2512.24000v3 Announce Type: replace Abstract: Distinguishing fake or untrue news from satire or humor poses a unique challenge due to their overlapping linguistic features and divergent intent.

applicationsarxiv-cs-cl
26 May 2026
Tutorials

A mathematical theory of balancing relational generalization and memorization

DGX agent

arXiv:2605.22972v1 Announce Type: cross Abstract: Humans, animals, and modern machine learning models exhibit impressive abilities to learn complex behaviors and generalize these behaviors to unseen s

tutorialsarxiv-cs-ai
25 May 2026
Model Releases

Approaching I/O-optimality for Approximate Attention

DGX agent

arXiv:2605.23751v1 Announce Type: new Abstract: We revisit the I/O complexity of attention in large language models. Given query-key-value matrices Q,K,VinR^{nimes d}, and a machine with fast memory s

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

DGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

model-releasesarxiv-cs-ai
25 May 2026
Applications

Coupled Training with Privileged Information and Unlabeled Data

DGX agent

arXiv:2605.23268v1 Announce Type: cross Abstract: In many prediction problems, we have extra information during training (for example, measurements that are expensive or slow to collect) that will not

applicationsarxiv-cs-lg
25 May 2026
Model Releases

CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception

DGX agent

arXiv:2605.23655v1 Announce Type: cross Abstract: High-resolution (HR) image perception presents a key bottleneck for multimodal large language models (MLLMs). While visual search offers a promising s

model-releasesarxiv-cs-ai
25 May 2026
← Previous
1…440441442443444…1119
Next →