AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
26 May 2026

GIBLy: Improving 3D Semantic Segmentation through an Architecture-Agnostic Lightweight Geometric Inductive Bias Layer

SafetyDGX agent

arXiv:2605.24243v1 Announce Type: cross Abstract: In 3D scene understanding, deep learning models rely on large models and extensive training to capture basic geometric structures that are present in

GL-LFGNN:A Global-Local Dual-branch Causal Graph Neural Network Based on Liang-Kleeman Information Flow for EEG Emotion Recognition

Model ReleasesDGX agent

arXiv:2605.25061v1 Announce Type: cross Abstract: EEG-based emotion recognition holds significant promise for objective diagnosis of mood disorders. Graph neural networks (GNNs) have emerged as the do

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2605.24636v1 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios re

Go witheFlow: Real-time Emotion Driven Audio Effects Modulation

ResearchDGX agent

arXiv:2510.02171v3 Announce Type: replace-cross Abstract: Music performance is a distinctly human activity, intrinsically linked to the performer's ability to convey, evoke, or express emotion. Machin

GRAIL: AI translation for scientists application workflow on satellite data

AgentsDGX agent

arXiv:2605.24784v1 Announce Type: new Abstract: Domain scientists increasingly develop Python scripts to analyze satellite imagery but they lack scalability to large-scale data. This paper demonstrate

Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers

Model ReleasesDGX agent

arXiv:2605.24518v1 Announce Type: cross Abstract: The quadratic complexity of self-attention in Transformer models remains a significant bottleneck for processing long sequences and deploying large la

Grouter: Decoupling Routing from Representation for Accelerated MoE Training

SafetyDGX agent

arXiv:2603.06626v2 Announce Type: replace-cross Abstract: Traditional Mixture-of-Experts (MoE) training typically proceeds without any structural priors, effectively requiring the model to simultaneou

Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation

SafetyDGX agent

arXiv:2605.25170v1 Announce Type: cross Abstract: Training data for olfaction is scattered through disparate, non-standardized datasets that limit the ability to build representative world models. Olf

Guarded Repair for Harm-Aware Post-hoc Replacement of LLM Mathematical Reasoning

ResearchDGX agent

arXiv:2605.24613v1 Announce Type: cross Abstract: Post-hoc repair of LLM mathematical reasoning introduces an asymmetric risk: fixing an incorrect reasoning trace is useful, but replacing a trace that

Guess the Unified Model: How Much Can We Recover from Generated Images?

ResearchDGX agent

arXiv:2605.25254v1 Announce Type: cross Abstract: With unified model-generated images now widespread online, attributing their model of origin offers a path toward transparency and deeper insight into

Harnessing AtomisticSkills for Agentic Atomistic Research

AgentsDGX agent

arXiv:2605.24002v1 Announce Type: cross Abstract: Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) hav

HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space

Model ReleasesDGX agent

arXiv:2509.22299v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures in large language models (LLMs) deliver exceptional performance and reduced inference costs compared to

HeartBeatAI: An Interpretable and Robust Deep Learning Framework for Multi-Label ECG Arrhythmia Detection

ResearchDGX agent

arXiv:2605.24588v1 Announce Type: new Abstract: While Deep Learning (DL) enhances automated electrocardiogram (ECG) analysis, clinical deployment is hindered by class imbalance and the generalization

Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents

Local AiDGX agent

arXiv:2605.24598v1 Announce Type: new Abstract: Large language model (LLM) agents excel at solving complex long-horizon tasks through autonomous interaction with environments. However, their real-worl

Hidden-State Privacy Has an Empty Middle

SafetyDGX agent

arXiv:2605.24042v1 Announce Type: cross Abstract: Of 1{,}536 Gaussian release covariances we tested for single-layer hidden-state privacy, zero achieve both moderate utility and moderate privacy again

Hide-and-Shill: A Reinforcement Learning Framework for Market Manipulation Detection in Symphony-a Decentralized Multi-Agent System

SafetyDGX agent

arXiv:2507.09179v3 Announce Type: replace Abstract: Decentralized finance (DeFi) has introduced a new era of permissionless financial innovation but also led to unprecedented market manipulation. With

Hide to Guide: Learning via Semantic Masking

SafetyDGX agent

arXiv:2605.25198v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a powerful paradigm for improving language models on reasoning-intensive tasks, but i

High-Risk AI Systems and the Problem of Identity in the European AI Act

ResearchDGX agent

arXiv:2605.23922v1 Announce Type: cross Abstract: The EU Artificial Intelligence Act (AIA) establishes a lifecycle governance regime for high-risk AI systems built around ex-ante conformity assessment

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

Model ReleasesDGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

HiTeC: Hierarchical Contrastive Learning on Text-Attributed Hypergraph with Semantic-Aware Augmentation

ApplicationsDGX agent

arXiv:2508.03104v3 Announce Type: replace-cross Abstract: Contrastive learning (CL) has become a dominant paradigm for self-supervised hypergraph learning, enabling effective training without costly l

HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

Model ReleasesDGX agent

arXiv:2605.24687v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have made significant strides in visual realism and semantic consistency, yet they often perpetuate and amplify societal bi

How does Bayesian Sampling help Membership Inference Attacks?

ResearchDGX agent

arXiv:2503.07482v2 Announce Type: replace-cross Abstract: Membership Inference Attacks (MIAs) aim to estimate whether a specific data point was used in the training of a given model. Existing state-of

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

Model ReleasesDGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

How Should LLMs Consume High-Quality Data? Optimal Data Scheduling via Quality-Aware Functional Scaling Laws

TutorialsDGX agent

arXiv:2605.25698v1 Announce Type: cross Abstract: High-quality data is scarce in large language model (LLM) training, yet how to schedule its use jointly with training dynamics lacks theoretical guida

How Well Do Models Follow Their Constitutions?

Model ReleasesDGX agent

arXiv:2605.24229v1 Announce Type: new Abstract: Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's

Human-AI Collaboration in Science at Scale: A Global Large-scale Randomized Field Experiment

ApplicationsDGX agent

arXiv:2605.24180v1 Announce Type: cross Abstract: Collaboration is the defining mode of modern science, yet its core mechanism -- feedback -- remains hard to observe, difficult to scale, and unequally

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos

SafetyDGX agent

arXiv:2605.24934v1 Announce Type: cross Abstract: Human egocentric video captures rich manipulation demonstrations without any robot hardware, yet transferring these skills to robots remains challengi

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

AgentsDGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

Hylos: Operability Contracts for Model-Native Spatial Intelligence

Model ReleasesDGX agent

arXiv:2605.24728v1 Announce Type: new Abstract: Foundation models can increasingly describe, reconstruct, and generate 3D objects, assemblies, scenes, and environments, but visually plausible spatial

HyperGuide: Hyperbolic Guidance for Efficient Multi-Step Reasoning in Large Language Models

TutorialsDGX agent

arXiv:2605.24140v1 Announce Type: new Abstract: Multi-step reasoning remains a central challenge for large language models: single-pass generation is efficient but lacks accuracy; tree-search methods

Hypothesis Generation and Inductive Inference in Children and Language Models

ApplicationsDGX agent

arXiv:2605.24528v1 Announce Type: new Abstract: Real world decision-making requires constructing mental models under uncertainty over evidence, over the underlying causal rules, and over the state of

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

SafetyDGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

Improving Labeling Consistency with Detailed Constitutional Definitions and AI-Driven Evaluation

SafetyDGX agent

arXiv:2605.24247v1 Announce Type: cross Abstract: Many automated labeling pipelines classify inputs into categories defined by a written specification, content moderation being a prominent use case. S

In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models

ApplicationsDGX agent

arXiv:2605.23908v1 Announce Type: new Abstract: We are in the midst of large-scale industrial and academic efforts to automate the processes of scientific, technological and creative production throug

IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2605.25475v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly expected to operate over long contexts, yet standard softmax attention incurs a KV cache that grows line

INDUCTION: Finite-Structure Concept Synthesis in First-Order Logic

Model ReleasesDGX agent

arXiv:2602.18956v3 Announce Type: replace Abstract: We introduce INDUCTION, a benchmark for finite structure concept synthesis in first order logic. Given small finite relational worlds with extension

Inference-Time Alignment of Diffusion Models via Trust-Region Iterative Twisted Sequential Monte Carlo

SafetyDGX agent

arXiv:2605.25123v1 Announce Type: cross Abstract: We study inference-time alignment for diffusion-based generative models, aiming to steer a base model toward high-reward outputs without updating its

Inference Time Context Sparsity: Illusion or Opportunity?

HardwareDGX agent

arXiv:2605.24168v1 Announce Type: new Abstract: Sparsity has long been a central theme in LLM efficiency, but its role in context processing remains unresolved. As LLM workloads shift toward longer co

Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization

ResearchDGX agent

arXiv:2605.25203v1 Announce Type: cross Abstract: We apply the influence-adaptive Walsh geometry of a companion theory paper (arXiv:2605.01637) to extreme low-bit weight-only LLM quantization. The rec

INSIGHT: INference-time Sequence Introspection for Generating Help Triggers in Vision-Language-Action Models

ResearchDGX agent

arXiv:2510.01389v2 Announce Type: replace-cross Abstract: Recent Vision-Language-Action (VLA) models show strong generalization capabilities, yet they lack introspective mechanisms for anticipating fa

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

Intent Signal Theory: A Computational Framework for Intent-State Control in Human-AI Interaction

ResearchDGX agent

arXiv:2605.25058v1 Announce Type: cross Abstract: Current AI interaction models treat the prompt as the primary object of exchange, omitting a critical layer: the user's latent source intent, the goal

Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning

SafetyDGX agent

arXiv:2605.05226v2 Announce Type: replace-cross Abstract: The central challenge of reinforcement learning for reasoning lies not only in the sparsity of outcome-level supervision, but more fundamental

Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

Local AiDGX agent

arXiv:2605.24999v1 Announce Type: cross Abstract: An agent must act on the situation before it, learn what it cannot yet represent, and model other agents well enough to coordinate. These faculties ar

Intrinsically Interpretable Attention via Sparse Post-Training

ResearchDGX agent

arXiv:2512.05865v5 Announce Type: replace-cross Abstract: We introduce a simple post-training method that makes transformer attention sparse without sacrificing performance. Applying a flexible sparsi

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications

SafetyDGX agent

arXiv:2605.24883v1 Announce Type: new Abstract: The widespread integration of Large Language Models (LLMs) necessitates rigorous and systematic safety evaluation. Existing paradigms either rely on con

Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization

SafetyDGX agent

arXiv:2605.24960v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) faithfulness, i.e., whether CoTs genuinely reflect large language models' (LLM) underlying behavior, is typically evaluated und

Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol

SafetyDGX agent

arXiv:2605.24538v1 Announce Type: cross Abstract: Every major framework for governing artificial intelligence presupposes an identifiable entity -- a developer, deployer, or operator -- who can be hel

Is Human Annotation Necessary? Iterative MBR Distillation for Error Span Detection in Machine Translation

ResearchDGX agent

arXiv:2603.12983v3 Announce Type: replace-cross Abstract: Error Span Detection (ESD) is a crucial subtask in Machine Translation (MT) evaluation, aiming to identify the location and severity of transl

Iterative Refinement Neural Operators are Learned Fixed-Point Solvers: A Principled Approach to Spectral Bias Mitigation

SafetyDGX agent

arXiv:2605.24041v1 Announce Type: cross Abstract: Neural operators serve as fast, data-driven surrogates for scientific modeling but typically rely on a monolithic, single-pass inference procedure tha

IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning

SafetyDGX agent

arXiv:2605.23997v1 Announce Type: cross Abstract: Multimodal large language models via reinforcement learning (RL) have demonstrated remarkable capabilities in complex visual reasoning tasks, yet they

JAEGER: Joint 3D Audio-Visual Grounding and Reasoning in Simulated Physical Environments

Model ReleasesDGX agent

arXiv:2602.18527v2 Announce Type: replace-cross Abstract: Current audio-visual large language models (AV-LLMs) are predominantly restricted to 2D perception, relying on RGB video and monaural audio. T

Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-Tuning in Large Language Models

SafetyDGX agent

arXiv:2605.24550v1 Announce Type: new Abstract: Fine-tuning-as-a-Service (FaaS) enables personalization of large language models (LLMs), but it can weaken safety-alignment under harmful fine-tuning at

JEPA-DNA: Grounding Genomic Foundation Models through Joint-Embedding Predictive Architectures

Model ReleasesDGX agent

arXiv:2602.17162v2 Announce Type: replace Abstract: Genomic Foundation Models (GFMs) typically rely on Masked Language Modeling (MLM) or Next-Token Prediction (NTP) to learn the 'Laws of Nature'. Whil

JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data

SafetyDGX agent

arXiv:2605.24414v1 Announce Type: new Abstract: We introduce JT-Safe-V2, a large language model designed to advance the safety and trustworthiness of foundation models, extending our previous JT-Safe

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment

Model ReleasesDGX agent

arXiv:2605.25240v1 Announce Type: cross Abstract: Two methodologies dominate current practices of benchmarking: rubric-based scoring evaluates items against predefined criteria, whereas comparative ju

K-U-KAN: Koopman-Enhanced U-KAN for 3D Dental Reconstruction from a Single Panoramic X-ray Radiograph

ResearchDGX agent

arXiv:2605.25163v1 Announce Type: cross Abstract: A panoramic X-ray compresses a 3D jaw into a 2D strip; we aim to recover the missing depth cleanly and fast. Existing implicit neural representations

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI

Model ReleasesDGX agent

arXiv:2510.02327v2 Announce Type: replace-cross Abstract: Real-time speech-to-speech (S2S) models excel at generating natural, low-latency conversational responses but often lack deep knowledge and se

Keep the Proof State Live: Snapshotting for Efficient Tactic Search in Lean 4

Model ReleasesDGX agent

arXiv:2605.25556v1 Announce Type: cross Abstract: Automated theorem proving systems built on Lean 4 increasingly rely on parallel tactic search over partially specified proofs, such as those generated

← Previous
1…214215216217218…358
Next →