AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
26 May 2026

Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environments

Model ReleasesDGX agent

arXiv:2602.19450v2 Announce Type: replace-cross Abstract: Trusted Execution Environments (TEEs) (e.g., Intel SGX and ArmTrustZone) aim to protect sensitive computation from a compromised operating sys

Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark

Model ReleasesDGX agent

arXiv:2605.26068v1 Announce Type: cross Abstract: Weakly supervised anomaly detection (WSAD) has developed in three primary directions: incomplete, inexact, and inaccurate supervision. However, these

Retrieved In-Context Principles from Previous Mistakes

ResearchDGX agent

arXiv:2407.05682v2 Announce Type: replace Abstract: In-context learning (ICL) has been instrumental in adapting Large Language Models (LLMs) to downstream tasks using correct input-output examples. Re

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Retrying vs Resampling in AI Control

Model ReleasesDGX agent

arXiv:2605.26047v1 Announce Type: new Abstract: AI coding scaffolds like Claude Code and Codex use extit{retrying}: blocking actions flagged as risky and continuing the trajectory. We study retrying f

RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism

Model ReleasesDGX agent

arXiv:2605.25565v1 Announce Type: cross Abstract: While Large Language Models (LLMs) are commonly fine-tuned to handle domain-specific tasks before being applied to vertical applications, adapting the

Selective Test-Time Compute Scaling for Click-Through Rate Prediction via Uncertainty-Triggered Feature Path Exploration

TutorialsDGX agent

arXiv:2605.24989v1 Announce Type: cross Abstract: Scaling test-time compute has proven highly effective for language models, yet this opportunity remains largely unexplored for industrial Click-Throug

Signs Beat Floats: Low-Rank Double-Binary Adaptation for On-Device Fine-Tuning

Local AiDGX agent

arXiv:2605.24058v1 Announce Type: cross Abstract: On-device adaptation of large language models commonly keeps a quantized base model frozen while training and deploying a small, task-specific LoRA ad

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

Model ReleasesDGX agent

arXiv:2605.25160v1 Announce Type: new Abstract: Mobile GUI agents powered by large language models have progressed rapidly, creating urgent needs for realistic and comprehensive evaluation. Existing b

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

SafetyDGX agent

arXiv:2509.12672v2 Announce Type: replace Abstract: The volume of machine-generated content online has grown dramatically due to the widespread use of Large Language Models (LLMs), leading to new chal

URS: A Unified Neural Routing Solver for Cross-Problem Zero-Shot Generalization

Model ReleasesDGX agent

arXiv:2509.23413v2 Announce Type: replace Abstract: Multi-task neural routing solvers have emerged as a promising paradigm for their ability to solve multiple vehicle routing problems (VRPs) using a s

vAttention: Verified Sparse Attention

Model ReleasesDGX agent

arXiv:2510.05688v2 Announce Type: replace-cross Abstract: State-of-the-art sparse attention methods for reducing decoding latency fall into two main categories: approximate top-k (and its extension, t

When Can We Trust Early Warnings? Leakage-Excluded Early Outcome Prediction from LMS Interaction Logs

Model ReleasesDGX agent

arXiv:2605.25794v1 Announce Type: new Abstract: Early-warning models built from Learning Management System (LMS) logs aim to predict end-of-course outcomes early enough to enable timely learner suppor

When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills

Model ReleasesDGX agent

arXiv:2605.25832v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator r

WhisTLE: Deeply Supervised, Text-Only Domain Adaptation for Pretrained Speech Recognition Transformers

ApplicationsDGX agent

arXiv:2509.10452v2 Announce Type: replace Abstract: Pretrained automatic speech recognition (ASR) models such as Whisper perform well but still need domain adaptation to handle unseen parlance. In man

Who judges the judges? Governance from metrics: a runtime framework for continuous LLM compliance monitoring

Model ReleasesDGX agent

arXiv:2605.24737v1 Announce Type: cross Abstract: Current approaches to AI compliance treat conformity as a binary, audit-time verdict rather than a continuous, measurable property of production syste

WideDepth: Millimeter-Accurate Benchmark for Fisheye Depth Estimation

Model ReleasesDGX agent

arXiv:2605.24074v1 Announce Type: cross Abstract: Fisheye cameras are increasingly adopted in robotics for near-field manipulation, navigation, and immersive perception, yet indoor depth benchmarks wi

WISE: Web Information Satire and Fakeness Evaluation

ApplicationsDGX agent

arXiv:2512.24000v3 Announce Type: replace Abstract: Distinguishing fake or untrue news from satire or humor poses a unique challenge due to their overlapping linguistic features and divergent intent.

25 May 2026

A mathematical theory of balancing relational generalization and memorization

TutorialsDGX agent

arXiv:2605.22972v1 Announce Type: cross Abstract: Humans, animals, and modern machine learning models exhibit impressive abilities to learn complex behaviors and generalize these behaviors to unseen s

Approaching I/O-optimality for Approximate Attention

Model ReleasesDGX agent

arXiv:2605.23751v1 Announce Type: new Abstract: We revisit the I/O complexity of attention in large language models. Given query-key-value matrices Q,K,VinR^{nimes d}, and a machine with fast memory s

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

Model ReleasesDGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

Coupled Training with Privileged Information and Unlabeled Data

ApplicationsDGX agent

arXiv:2605.23268v1 Announce Type: cross Abstract: In many prediction problems, we have extra information during training (for example, measurements that are expensive or slow to collect) that will not

CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception

Model ReleasesDGX agent

arXiv:2605.23655v1 Announce Type: cross Abstract: High-resolution (HR) image perception presents a key bottleneck for multimodal large language models (MLLMs). While visual search offers a promising s

D2 Actor Critic: Diffusion Actor Meets Distributional Critic

Model ReleasesDGX agent

arXiv:2510.03508v3 Announce Type: replace Abstract: We introduce D2AC, a new model-free reinforcement learning (RL) algorithm designed to train expressive diffusion policies online effectively. At its

DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures

Model ReleasesDGX agent

arXiv:2511.15503v2 Announce Type: replace-cross Abstract: High-performance Host processors can integrate Processing-In-Memory (PIM) devices, which can accelerate memory-intensive kernels of Machine Le

Discontinuous Galerkin Neural Operator for Pathology Defocus Deblurring

Model ReleasesDGX agent

arXiv:2605.23282v1 Announce Type: cross Abstract: Defocus deblurring in pathological microscopy remains challenging due to the spatially varying and locally discontinuous nature of optical blur induce

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.23176v1 Announce Type: new Abstract: Spatiotemporal intelligence in autonomous driving (AD) requires an agent to integrate multi-view observations into a coherent scene representation, main

Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking

Model ReleasesDGX agent

arXiv:2605.23118v1 Announce Type: cross Abstract: Tracking tumor lesions across serial CT scans is essential for oncological response assessment. Existing automated methods face a fundamental trade-of

FAST-ME: Foundation-aware Adaptive Stopping for Motion Estimation for Efficient IoT Video Analysis

Model ReleasesDGX agent

arXiv:2605.23428v1 Announce Type: new Abstract: In modern multimedia systems, efficient video processing is critical, especially in resource-constrained environments such as IoT-based camera networks,

Forget What's Sensitive, Remember What Matters: Token-Level Differential Privacy in Memory Sculpting for Continual Learning

ResearchDGX agent

arXiv:2509.12958v2 Announce Type: replace Abstract: Continual Learning (CL) models, while adept at sequential knowledge acquisition, face significant and often overlooked privacy challenges due to acc

From Residuals to Reasons: LLM-Guided Mechanism Inference from Tabular Data

AgentsDGX agent

arXiv:2605.22897v1 Announce Type: new Abstract: A persistent challenge in machine learning for scientific applications is jointly achieving prediction and understanding. Statistical models excel on st

LFRAG: Layout-oriented Fine-grained Retrieval-Augmented Generation on Multimodal Document Understanding

Model ReleasesDGX agent

arXiv:2605.22829v1 Announce Type: cross Abstract: Multimodal Retrieval-Augmented Generation (RAG) has emerged as an effective paradigm for enhancing Large Language Models (LLMs) with external knowledg

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

ResearchDGX agent

arXiv:2605.23384v1 Announce Type: cross Abstract: Recent RL methods have substantially improved the reasoning abilities of LLMs. Existing reward designs mainly follow two paradigms: (1) Reinforcement

Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions

ResearchDGX agent

arXiv:2605.23885v1 Announce Type: new Abstract: Cross-lingual knowledge transfer is critical for building high-performing multilingual language models for languages with insufficient training data. Wh

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

AgentsDGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

OpenAI, Grupo Folha and Grupo UOL announce strategic content partnership

Model ReleasesDGX agent

OpenAI announced a strategic partnership with Grupo Folha and Grupo UOL, major Brazilian media companies, to integrate their content into OpenAI's AI models and products. The partnership enables OpenA

Parametric Prior Mapping Framework for Non-stationary Probabilistic Time Series Forecasting

ResearchDGX agent

arXiv:2605.23402v1 Announce Type: cross Abstract: Effectively modeling non-stationary dynamics in probabilistic multivariate time series(MTS) forecasting requires balancing expressiveness with robustn

Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study

Model ReleasesDGX agent

arXiv:2605.23108v1 Announce Type: cross Abstract: AI-assisted code review tools typically operate as generic 'expert reviewer' agents, producing homogeneous findings regardless of the analysis type ne

Physics-Informed Machine Learning Regulated by Finite Element Analysis for Simulation Acceleration of Melt Pool Dynamics in Laser Powder Bed Fusion

Model ReleasesDGX agent

arXiv:2506.20537v3 Announce Type: replace Abstract: Efficient simulation of Laser Powder Bed Fusion (LPBF) is crucial for process prediction due to the lasting issue of high computational cost associa

Speak-to-Structure: Evaluating LLMs in Open-domain Natural Language-Driven Molecule Generation

Model ReleasesDGX agent

arXiv:2412.14642v4 Announce Type: replace Abstract: Recently, Large Language Models (LLMs) have demonstrated great potential in natural language-driven molecule discovery. However, existing datasets a

Tabular PDF Information Extraction with Local LLMs and Layout-Aware Parsing: A Reliability Evaluation

Model ReleasesDGX agent

arXiv:2604.00003v2 Announce Type: replace-cross Abstract: Extracting structured information from academic PDF documents is non trivial: a single page typically combines free text metadata with tabular

Test-Time Training Undermines Safety Guardrails

SafetyDGX agent

arXiv:2605.22984v1 Announce Type: cross Abstract: Test-Time Training (TTT) is an emerging paradigm that enables models to adapt their parameters during inference, improving performance on tasks such a

23 May 2026

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. Wi…

Model ReleasesDGX agent

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. With motivation, overview, and GPT-style model reference imple

AutoBaxBuilder: Bootstrapping Code Security Benchmarking

Model ReleasesDGX agent

arXiv:2512.21132v2 Announce Type: replace-cross Abstract: As large language models (LLMs) see wide adoption in software engineering, the reliable assessment of the correctness and security of LLM-gene

Calibration, Uncertainty Communication, and Deployment Readiness in CKD Risk Prediction: A Framework Evaluation Study

ResearchDGX agent

arXiv:2605.21566v1 Announce Type: new Abstract: Machine learning models for chronic kidney disease (CKD) risk prediction often post strong discrimination scores on internal test sets. Calibration and

Characterizing the Fault Response of the Intel Neural Compute Stick 2 Under Single-Pulse Electromagnetic Fault Injection

Model ReleasesDGX agent

arXiv:2605.22437v1 Announce Type: cross Abstract: Vision processing units and other commercial neural-network inference accelerators are increasingly deployed in safety-relevant edge applications, but

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

Model ReleasesDGX agent

arXiv:2605.22502v1 Announce Type: cross Abstract: Agent orchestration frameworks have proliferated, collectively exceeding 290,000 GitHub stars across LangGraph, CrewAI, Google ADK, OpenAI Agents SDK,

Engineering Hybrid Physics-Informed Neural Networks for Next-Generation Electricity Systems: A State-of-the-Art Review

Model ReleasesDGX agent

arXiv:2605.21903v1 Announce Type: cross Abstract: The integration of machine learning with domain-specific physics is transforming the design, monitoring, and control of electricity systems, where dat

GraphFlow: A Graph-Based Workflow Management for Efficient LLM-Agent Serving

Model ReleasesDGX agent

arXiv:2605.22566v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents demonstrate strong reasoning and execution capabilities on complex tasks when guided by structured instructions,

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability

Model ReleasesDGX agent

arXiv:2605.22168v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) map complex visual inputs to semantic spaces, but interpreting the cross-modal reasoning of VLMs currently relies on pos

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

Model ReleasesDGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

Uncertainty-Aware Distribution-to-Distribution Flow Matching for Scientific Imaging

ResearchDGX agent

arXiv:2603.21717v4 Announce Type: replace Abstract: Distribution-to-distribution generative models support scientific imaging tasks ranging from modeling cellular perturbation responses to translating

22 May 2026

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

Model ReleasesDGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment

Model ReleasesDGX agent

arXiv:2512.20538v2 Announce Type: replace Abstract: Single-view RGB model-based object pose estimation methods achieve strong generalization but are fundamentally limited by depth ambiguity, clutter,

Balancing Uncertainty and Diversity of Samples: Leveraging Diversity of Least, High Confidence Samples for Effective Active Learning

TutorialsDGX agent

arXiv:2605.22169v1 Announce Type: new Abstract: Deep learning models, including Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs), have achieved state-of-the-art performance on vario

Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion

Model ReleasesDGX agent

arXiv:2605.22579v1 Announce Type: new Abstract: Recent work has identified a counterintuitive phenomenon termed 'Hyperfitting', where fine-tuning Large Language Models (LLMs) to near-zero training los

DeFacto: Counterfactual Thinking with Images for Enforcing Evidence-Grounded and Faithful Reasoning

Model ReleasesDGX agent

arXiv:2509.20912v4 Announce Type: replace Abstract: Recent advances in multimodal language models (MLLMs) have made thinking with images a dominant paradigm for multimodal reasoning. However, existing

Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR

SafetyDGX agent

arXiv:2602.17517v2 Announce Type: replace Abstract: Augmented reality (AR) guidance in laparoscopic liver surgery requires accurate registration of preoperative 3D models to intraoperative 2D video, b

Diverse Yet Consistent: Context-Guided Diffusion with Energy-Based Joint Refinement for Multi-Agent Motion Prediction

Model ReleasesDGX agent

arXiv:2605.22017v1 Announce Type: new Abstract: Deepgenerative models havebecomeapromisingapproach for human motion prediction due to their ability to capture multimodal distributions and represent di

From Parameters to Data: A Task-Parameter-Guided Fine-Tuning Pipeline for Efficient LLM Alignment

Model ReleasesDGX agent

arXiv:2605.21558v1 Announce Type: cross Abstract: Adapting Large Language Models (LLMs) to specialized domains typically incurs high data and computational overhead. While prior efficiency efforts hav

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks

Model ReleasesDGX agent

arXiv:2605.21917v1 Announce Type: new Abstract: Training Vision Language Models (VLMs) for video event reasoning requires high-quality structured annotations capturing not only what happened, but when

← Previous
1…406407408409410…1053
Next →