AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,561 results
12 May 2026

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

Model ReleasesDGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

Model ReleasesDGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

Model ReleasesDGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

A Deep Risk Estimator for Known Operator Learning

Model ReleasesDGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

Model ReleasesDGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

Model ReleasesDGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

Model ReleasesDGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

A new initialisation to Control Gradients in Sinusoidal Neural network

Model ReleasesDGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

Model ReleasesDGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

A Stability Benchmark of Generative Regularizers for Inverse Problems

Model ReleasesDGX agent

arXiv:2605.10076v1 Announce Type: cross Abstract: Generative (diffusion) priors demonstrate remarkable performance in addressing inverse problems in imaging. Yet, for scientific and medical imaging, i

A startup wants to pull magnesium from seawater without torching the environment. Another wants to take small language models to the podium …

Model ReleasesDGX agent

A startup wants to pull magnesium from seawater without torching the environment. Another wants to take small language models to the podium for enterprise customers. Today @jason and @alex sat down wi

A Unified Representation of Neural Networks Architectures

Model ReleasesDGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

Model ReleasesDGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims

Model ReleasesDGX agent

arXiv:2605.10575v1 Announce Type: cross Abstract: Safe fine-tuning defenses are often endorsed on the basis of a held-out gap reduction, but the same reduction can come from sampling noise, subject ar

Action-Guided Attention for Video Action Anticipation

Model ReleasesDGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models

Model ReleasesDGX agent

arXiv:2605.08567v1 Announce Type: new Abstract: Action-conditioned world models (ACWMs) have shown strong promise for video prediction and decision-making. However, existing benchmarks are largely res

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

Model ReleasesDGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

Model ReleasesDGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2605.08734v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) reparameterizes a weight update as a product of two low-rank factors, but the Jacobian J_{G} of the generator mapping the f

Adaptive Multi-view Graph Contrastive Learning via Fractional-order Neural Diffusion Networks

Model ReleasesDGX agent

arXiv:2511.06216v4 Announce Type: replace Abstract: Graph contrastive learning (GCL) learns node and graph representations by contrasting multiple views of the same graph. Existing methods typically r

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

Model ReleasesDGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

Model ReleasesDGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

Model ReleasesDGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

Agentic MIP Research: Accelerated Constraint Handler Generation

Model ReleasesDGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

Agentic Performance at the Edge: Insights from Benchmarking

Model ReleasesDGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization

Model ReleasesDGX agent

arXiv:2605.08704v1 Announce Type: new Abstract: Multi-agent reasoning has shown promise for improving the problem-solving ability of large language models by allowing multiple agents to explore divers

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Model ReleasesDGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

Model ReleasesDGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored)…

Model ReleasesDGX agent

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored). The property manager said she would 'do everything she cou

Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models

Model ReleasesDGX agent

arXiv:2605.08115v1 Announce Type: cross Abstract: Wepresent Alice v1, a 14-billion parameter open-source video generation model that achieves state-of-the-art quality through consistency distillation

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

Model ReleasesDGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

Model ReleasesDGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment

Model ReleasesDGX agent

arXiv:2603.26680v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into lifelong AI assistants, LLM personalization has become a critical frontier. However, progress is c

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

Model ReleasesDGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

An Annotation Scheme and Classifier for Personal Facts in Dialogue

Model ReleasesDGX agent

arXiv:2605.10339v1 Announce Type: new Abstract: The advancement of Large Language Models (LLMs) has enabled their application in personalized dialogue systems. We present an extended annotation scheme

An Empirical Study of Multi-Agent Collaboration for Automated Research

Model ReleasesDGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation

Model ReleasesDGX agent

arXiv:2605.10397v1 Announce Type: cross Abstract: Visual anomaly detection (VAD) is crucial in many real-world fields, such as industrial inspection, medical imaging, infrastructure monitoring, and re

Anthropic announces 12 Claude plugins for the legal sector, including a 'commercial counsel' tool for reviewing vendor agreements and a bar exam study tool (Rachel Metz/Bloomberg)

Model ReleasesDGX agent

Rachel Metz / Bloomberg: Anthropic announces 12 Claude plugins for the legal sector, including a “commercial counsel” tool for reviewing vendor agreements and a bar exam study tool — Anthropic PBC is

AnyDepth-DETR/-YOLO: Any-depth object detection with a single network

Model ReleasesDGX agent

arXiv:2605.09407v1 Announce Type: new Abstract: Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scen

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

Model ReleasesDGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

Architecture, Not Scale: Circuit Localization in Large Language Models

Model ReleasesDGX agent

arXiv:2605.08853v1 Announce Type: new Abstract: Mechanistic interpretability assumes that circuit analysis becomes harder as models scale. We challenge this assumption by showing that the attention ar

Are vision-language models ready to zero-shot replace supervised classification models in agriculture?

Model ReleasesDGX agent

arXiv:2512.15977v3 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly proposed as general-purpose solutions for visual recognition tasks, yet their reliability for agricul

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

Model ReleasesDGX agent

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

Model ReleasesDGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

AssemPlanner: A Multi-Agent Based Task Planning Framework for Flexible Assembly System

Model ReleasesDGX agent

arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio

ASTRA-QA: A Benchmark for Abstract Question Answering over Documents

Model ReleasesDGX agent

arXiv:2605.10168v1 Announce Type: new Abstract: Document-based question answering (QA) increasingly includes abstract questions that require synthesizing scattered information from long documents or a

Attention Grounded Enhancement for Visual Document Retrieval

Model ReleasesDGX agent

arXiv:2511.13415v2 Announce Type: replace-cross Abstract: Visual document retrieval requires understanding heterogeneous and multi-modal content to satisfy implicit information needs. Recent advances

AUHead: Realistic Emotional Talking Head Generation via Action Units Control

Model ReleasesDGX agent

arXiv:2602.09534v2 Announce Type: replace Abstract: Realistic talking-head video generation is critical for virtual avatars, film production, and interactive systems. Current methods struggle with nua

Automated Approach for Solving Infinite-state Polynomial Reachability Games

Model ReleasesDGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation

Model ReleasesDGX agent

arXiv:2605.10845v1 Announce Type: cross Abstract: As global cross-lingual communication intensifies, language barriers in visually rich documents such as PDFs remain a practical bottleneck. Existing d

BCJR-QAT: A Differentiable Relaxation of Trellis-Coded Weight Quantization

Model ReleasesDGX agent

arXiv:2605.10655v1 Announce Type: new Abstract: Trellis-coded quantization sets the current 2-bit post-training frontier for LLMs (QTIP), but pushing below the PTQ ceiling requires quantization-aware

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

Model ReleasesDGX agent

arXiv:2605.10867v1 Announce Type: cross Abstract: Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and mot

BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD

Model ReleasesDGX agent

arXiv:2605.10865v1 Announce Type: new Abstract: Industrial Computer-Aided Design (CAD) code generation requires models to produce executable parametric programs from visual or textual inputs. Beyond r

BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition

Model ReleasesDGX agent

arXiv:2605.08296v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports broad healthcare and behavior science applications. However, data heterogeneity and the

Benchmarking Compositional Generalisation for Machine Learning Interatomic Potentials

Model ReleasesDGX agent

arXiv:2605.08988v1 Announce Type: cross Abstract: Machine Learning Interatomic Potentials play a fundamental role in computational chemistry and materials science, enabling applications from molecular

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

Model ReleasesDGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

Benchmarking Sensor-Fault Robustness in Forecasting

Model ReleasesDGX agent

arXiv:2605.10822v1 Announce Type: new Abstract: Cyber-physical system (CPS) forecasting models depend on sensor streams with noisy, biased, missing, or temporally misaligned readings, yet standard for

← Previous
1…258259260261262…377
Next →