AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

Entropy-informed Decoding: Adaptive Information-Driven Branching

ResearchDGX agent

arXiv:2605.09745v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable generative performance, yet their output quality is dependent on the decoding strategy. While sampling

EpiGraph: A Knowledge Graph and Benchmark for Evidence-Intensive Reasoning in Epilepsy

Model ReleasesDGX agent

arXiv:2605.09505v1 Announce Type: new Abstract: Epilepsy diagnosis and treatment require evidence-intensive reasoning across heterogeneous clinical knowledge, including biosignal patterns, genetic mec

EquiMem: Calibrating Shared Memory in Multi-Agent Debate via Game-Theoretic Equilibrium

AgentsDGX agent

arXiv:2605.09278v1 Announce Type: new Abstract: Multi-agent debate (MAD) systems increasingly rely on shared memory to support long-horizon reasoning, but this convenience opens a critical vulnerabili


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Equivariant Volumetric Grasping

ApplicationsDGX agent

arXiv:2507.18847v3 Announce Type: replace-cross Abstract: We propose a new volumetric grasp model that is equivariant to rotations around the vertical axis, leading to a significant improvement in sam

Evading Visual Aphasia: Contrastive Adaptive Semantic Token Pruning for Vision-Language Models

ResearchDGX agent

arXiv:2605.09429v1 Announce Type: cross Abstract: Are low-attention visual tokens truly redundant in vision-language reasoning? Existing pruning methods often assume so, ranking visual tokens by shall

Evaluating Developmental Cognition Capabilities of LLMs

ResearchDGX agent

arXiv:2605.08549v1 Announce Type: new Abstract: Conversational AI is increasingly personalized around users' preferences, histories, goals, and knowledge, but much less around how users interpret and

Event Fields: Learning Latent Event Structure for Waveform Foundation Models

Local AiDGX agent

arXiv:2605.08685v1 Announce Type: cross Abstract: We propose a new class of waveform foundation models that departs from conventional sequence based representations by modeling physiological time seri

EventTSF: Event-Aware Non-Stationary Time Series Forecasting

ApplicationsDGX agent

arXiv:2508.13434v2 Announce Type: replace-cross Abstract: Time series forecasting is vital in diverse sectors such as energy and transportation, where non-stationary dynamics are deeply intertwined wi

Every finite group admits a just finite presentation

ResearchDGX agent

arXiv:2605.10402v1 Announce Type: cross Abstract: A finite presentation of a finite group is called `just finite' if removing any relation from R results in a presentation for an infinite group. It ha

EverydayMMQA: A Multilingual and Multimodal Framework for Culturally Grounded Spoken Visual QA

Model ReleasesDGX agent

arXiv:2510.06371v2 Announce Type: replace-cross Abstract: Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries re

Evidence Over Plans: Online Trajectory Verification for Skill Distillation

AgentsDGX agent

arXiv:2605.09192v1 Announce Type: new Abstract: Agent skills can remarkably improve task success rates by using human-written procedural documents, but their quality is difficult to assess without env

EvoDriveVLA: Evolving Driving VLA Models via Collaborative Perception-Planning Distillation

AgentsDGX agent

arXiv:2603.09465v3 Announce Type: replace-cross Abstract: Vision-Language-Action models have shown great promise for autonomous driving, yet they suffer from degraded perception after unfreezing the v

Evolutionary Ensemble of Agents

AgentsDGX agent

arXiv:2605.09018v1 Announce Type: cross Abstract: We introduce Evolutionary Ensemble (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving sys

Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents

ResearchDGX agent

arXiv:2605.10663v1 Announce Type: new Abstract: Experience-driven self-evolving agents aim to overcome the static nature of large language models by distilling reusable experience from past interactio

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent

SafetyDGX agent

arXiv:2605.09777v1 Announce Type: cross Abstract: Gradient-based preference optimization methods for large language model (LLM) alignment suffer from preference collapse, converging to narrow behavior

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

SafetyDGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

SafetyDGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

Expert Evaluation and the Limits of Human Feedback in Mental Health AI Safety Testing

SafetyDGX agent

arXiv:2601.18061v3 Announce Type: replace Abstract: Learning from human feedback~(LHF) assumes that expert judgments, appropriately aggregated, yield valid ground truth for training and evaluating AI

Explainability of Recurrent Neural Networks for Enhancing P300-based Brain-Computer Interfaces

Local AiDGX agent

arXiv:2605.10121v1 Announce Type: cross Abstract: Brain-Computer Interfaces (BCIs) based on P300 event-related potentials offer promising applications in health, education, and assistive technologies.

Explainable Knowledge Tracing via Probabilistic Embeddings and Pattern-based Reasoning

ResearchDGX agent

arXiv:2605.09369v1 Announce Type: new Abstract: Knowledge Tracing (KT) models students' knowledge states based on learning interactions to predict performance. While deep learning-based KT models have

Explainable Machine Learning Framework for Cardiovascular Disease Diagnosis and Prognosis

ApplicationsDGX agent

arXiv:2507.11185v2 Announce Type: replace-cross Abstract: Heart disease continues to pose a critical worldwide health issue, more specifically in areas with insufficient access to healthcare infrastru

Explaining Graph Neural Networks for Node Similarity on Graphs

ResearchDGX agent

arXiv:2407.07639v2 Announce Type: replace-cross Abstract: Similarity search is a fundamental task for exploiting information in various applications dealing with graph data, such as citation networks

Explanation Fairness in Large Language Models: An Empirical Analysis of Disparities in How LLMs Justify Decisions Across Demographic Groups

Model ReleasesDGX agent

arXiv:2605.08671v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed not only to make decisions but to explain them. While AI decision fairness has been studied ext

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

Model ReleasesDGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models

Model ReleasesDGX agent

arXiv:2605.09773v1 Announce Type: cross Abstract: We use sparse autoencoder (SAE) feature steering to amplify Dark Triad personality traits (Machiavellianism, narcissism, and psychopathy) in Llama-3.3

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

Model ReleasesDGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

expo: Exploration-prioritized policy optimization via adaptive kl regulation and gaussian curriculum sampling

Model ReleasesDGX agent

arXiv:2605.09923v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the standard paradigm for LLM mathematical reasoning, where Group Relative Policy Optim

Extrusion Segmentation Strategy to improve CAD Reconstruction from Point Cloud

ApplicationsDGX agent

arXiv:2605.08971v1 Announce Type: cross Abstract: Computer-Aided Design is ubiquitous in todays world, as almost every manufactured object begins as a digital model across industries. At the same time

FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2605.09081v1 Announce Type: cross Abstract: We introduce the first universal pretraining corpus for industrial time-series data: FactoryNet. 51M datapoints across 23k end-to-end task executions

FairHealth: An Open-Source Python Library for Trustworthy Healthcare AI in Low-Resource Settings

SafetyDGX agent

arXiv:2605.08198v1 Announce Type: cross Abstract: We present FairHealth, an open-source Python library that provides a unified, modular framework for trustworthy machine learning in healthcare applica

Fairness of Explanations in Artificial Intelligence (AI): A Unifying Framework, Axioms, and Future Direction toward Responsible AI

SafetyDGX agent

arXiv:2605.09852v1 Announce Type: new Abstract: Machine learning algorithms are being used in high-stakes decisions, including those in criminal justice, healthcare, credit, and employment. The resear

Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems

SafetyDGX agent

arXiv:2605.10604v1 Announce Type: cross Abstract: Designing fair algorithmic decision systems requires balancing model performance with fairness toward affected individuals: More fairness might requir

Fashion Florence: Fine-Tuning Florence-2 for Structured Fashion Attribute Extraction

Model ReleasesDGX agent

arXiv:2605.09827v1 Announce Type: cross Abstract: We present Fashion Florence, a Florence-2 vision-language model fine-tuned with LoRA to extract structured fashion attributes from clothing images. Gi

Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability

SafetyDGX agent

arXiv:2605.09214v1 Announce Type: cross Abstract: Kullback-Leibler (KL) regularization is ubiquitous in reinforcement learning algorithms in the form of reverse or forward KL. Recent studies have demo

Feature Repulsion and Spectral Lock-in: An Empirical Study of Two-Layer Network Grokking

Model ReleasesDGX agent

arXiv:2605.08119v1 Announce Type: cross Abstract: Tian (2025) proves a repulsion theorem (Theorem 6) for the matrix B = (widetilde{F}^op widetilde{F} + eta I)^{-1} during the interactive feature-learn

Field-Localized Forgery Detection for Digital Identity Documents

ResearchDGX agent

arXiv:2605.09089v1 Announce Type: cross Abstract: Digital identity verification systems used in remote onboarding rely on document images to authenticate users, making them vulnerable to localized man

Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs

ResearchDGX agent

arXiv:2605.08894v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance but incur high deployment costs, motivating extremely low-bit but lossy quantization. Existing

Fitting Multilinear Polynomials for Logic Gate Networks

Model ReleasesDGX agent

arXiv:2605.08657v1 Announce Type: cross Abstract: We study learnable logic gate networks that stack layers of 2-input Boolean gates to build combinational circuits. Every 2-input gate has a unique mul

Flag Varieties: A Geometric Framework for Deep Network Alignment

SafetyDGX agent

arXiv:2605.09861v1 Announce Type: cross Abstract: Alignment, the tendency of adjacent weight matrices in deep networks to develop compatible subspace orientations, underlies gradient flow, Neural Coll

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

Model ReleasesDGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

FlashSVD v1.5: Making Low-Rank Transformers Inference Actually Fast

HardwareDGX agent

arXiv:2605.08314v1 Announce Type: cross Abstract: SVD-based Low-rank compression reduces transformer parameters and nominal FLOPs, but these savings often translate poorly into real LLM serving speedu

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

ApplicationsDGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs

Model ReleasesDGX agent

arXiv:2605.08905v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable success on reasoning benchmarks through Reinforcement Learning with Verifiable Rewards (RLVR), exc

Formal Policy Enforcement for Real-World Agentic Systems

SafetyDGX agent

arXiv:2602.16708v3 Announce Type: replace-cross Abstract: Security policy enforcement in contemporary agentic systems predominantly consists of embedding natural-language policies within an agent's sy

Formally Verifying Analog Neural Networks Under Process Variations Using Polynomial Zonotopes

ApplicationsDGX agent

arXiv:2605.10474v1 Announce Type: cross Abstract: Analog neural networks are gaining attention due to their efficiency in terms of power consumption and processing speed. However, since analog neural

FormalRewardBench: A Benchmark for Formal Theorem Proving Reward Models

Model ReleasesDGX agent

arXiv:2605.10141v1 Announce Type: new Abstract: Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. Whil

FORTIS: Benchmarking Over-Privilege in Agent Skills

Model ReleasesDGX agent

arXiv:2605.09163v1 Announce Type: new Abstract: Large language model agents increasingly operate through an intermediate skill layer that mediates between user intent and concrete task execution. This

Fourier Feature Methods for Nonlinear Causal Discovery: FFML Scoring, TRFF Scoring, and FFCI Testing in Mixed Data

ResearchDGX agent

arXiv:2605.05743v2 Announce Type: replace-cross Abstract: Gaussian process (GP) marginal likelihood scores and kernel conditional independence tests are theoretically appealing for nonlinear causal di

FQPDR: Federated Quantum Neural Network for Privacy-preserving Early Detection of Diabetic Retinopathy

ResearchDGX agent

arXiv:2605.08324v1 Announce Type: cross Abstract: Diabetic Retinopathy (DR) is a common complication of diabetes that can lead to blindness of people. Detecting DR at the earliest stage is essential t

FRACTAL: SSM with Fractional Recurrent Architecture for Computational Temporal Analysis of Long Sequences

Model ReleasesDGX agent

arXiv:2605.08833v1 Announce Type: new Abstract: Effective sequence modeling fundamentally requires balancing the retention of unbounded history with the high-resolution detection of abrupt short-term

FragileFlow: Spectral Control of Correct-but-Fragile Predictions for Foundation Model Robustness

ResearchDGX agent

arXiv:2605.08896v1 Announce Type: cross Abstract: Robust adaptation of LLMs and VLMs is often evaluated by average accuracy or average consistency under perturbations. However, these averages can hide

FraudBench: A Multimodal Benchmark for Detecting AI-Generated Fraudulent Refund Evidence

Model ReleasesDGX agent

arXiv:2605.08820v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-generated images have become increasingly realistic and readily adaptable to concrete real-world claims, creating new cha

Free Energy Manifold: Score-Based Inference for Hybrid Bayesian Networks

ResearchDGX agent

arXiv:2605.09839v1 Announce Type: cross Abstract: We introduce the Free Energy Manifold (FEM), a score-trained conditional energy model specialized for inference in hybrid Bayesian networks with discr

Frictional Q-Learning

SafetyDGX agent

arXiv:2509.19771v5 Announce Type: replace-cross Abstract: Off-policy reinforcement learning suffers from extrapolation errors when a learned policy selects actions that are weakly supported in the rep

From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World

ApplicationsDGX agent

arXiv:2605.10834v1 Announce Type: new Abstract: AI pentesting agents are increasingly credible as offensive security systems, but current benchmarks still provide limited guidance on which will perfor

From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs

Local AiDGX agent

arXiv:2605.09370v1 Announce Type: cross Abstract: Large-scale AI training is now fundamentally a distributed systems problem, and hardware failures have become routine operating conditions rather than

From Historical Tabular Image to Knowledge Graphs: A Provenance-Aware Modular Pipeline

ApplicationsDGX agent

arXiv:2605.08222v1 Announce Type: cross Abstract: Handwritten archival tables contain rich historical information, yet transforming them into structured representations, such as Knowledge Graphs, requ

From Holo Pockets to Electron Density: GPT-style Drug Design with Density

SafetyDGX agent

arXiv:2605.08767v1 Announce Type: new Abstract: Recent advances in generative modeling have enabled significant progress in structure-based drug design (SBDD). Existing methods typically condition mol

From Ontology Conformance to Admissible Reconfiguration: A RoSO/SMGI Adequacy Argument for Robotic Service Governance

ResearchDGX agent

arXiv:2605.08185v1 Announce Type: cross Abstract: The Robotic Service Ontology (RoSO) gives service robotics a typed semantic vocabulary for services, functions, interactions, and deployment-sensitive

← Previous
1…266267268269270…358
Next →