AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
1 Jun 2026

From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Backdoors

Model ReleasesDGX agent

arXiv:2605.31042v1 Announce Type: cross Abstract: LLM agents are evolving from conversational chatbots to operational tools in real-world workspaces. In local agentic harnesses, an LLM can read and wr

From Weak Cues to Real Identities: Evaluating Inference-Driven De-Anonymization in LLM Agents

Model ReleasesDGX agent

arXiv:2603.18382v2 Announce Type: replace Abstract: Anonymization is often assumed to protect privacy once explicit identifiers are removed, because re-identification has historically required special

Full-field prediction for engineering-scale three-dimensional aircraft with multigrid-hierarchical learning

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.30375v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics is essential for aerospace design, but engineering-scale simulations of practical three-dimensional aircraf

Functional MRI Time Series Generation via Wavelet-Based Image Transform and Spectral Flow Matching for Brain Disorder Identification

Local AiDGX agent

arXiv:2605.30387v1 Announce Type: cross Abstract: Functional Magnetic Resonance Imaging (fMRI) provides non-invasive access to dynamic brain activity by measuring blood oxygen level-dependent (BOLD) s

Functorial Neural Architectures from Higher Inductive Types

TutorialsDGX agent

arXiv:2603.16123v2 Announce Type: replace-cross Abstract: Neural networks often learn the parts of a task but fail on novel combinations of those parts. We argue that this failure is architectural: a

G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition

Local AiDGX agent

arXiv:2603.10468v2 Announce Type: replace-cross Abstract: We study timestamped speaker-attributed automatic speech recognition (SA-ASR) for long-form, multi-party speech with overlap. In this setting,

GaMi: Geometry-Agnostic Material Identification via Cross-Modal Subtractive Disentanglement

ResearchDGX agent

arXiv:2605.30818v1 Announce Type: cross Abstract: Non-contact material identification enables adaptive interaction for embodied intelligence yet faces challenges from geometry-induced variations (e.g.

Gap-K%: Measuring Top-1 Prediction Gap for Detecting Pretraining Data

Local AiDGX agent

arXiv:2601.19936v2 Announce Type: replace-cross Abstract: The opacity of massive pretraining corpora in Large Language Models (LLMs) raises significant privacy and copyright concerns, making pretraini

Generalistic or Specific Embeddings, Which is Better? An Empirical Study on Search for Clinical Coding in Non-English Languages

Model ReleasesDGX agent

arXiv:2605.30529v1 Announce Type: cross Abstract: Sentence-embedding models for semantic search are overwhelmingly developed and evaluated on English corpora. When applied to clinical retrieval in oth

Generating Graph-like Rules for Knowledge Graph Reasoning via Diffusion Models

Model ReleasesDGX agent

arXiv:2605.30747v1 Announce Type: new Abstract: Logical rules constitute a cornerstone of knowledge graph (KG) reasoning, valued for their interpretability and ability to model relational patterns. Ho

Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report Generation

Model ReleasesDGX agent

arXiv:2605.30984v1 Announce Type: cross Abstract: Modern 3D medical vision-language models (VLMs) can generate fluent radiology-style text while exhibit critically low pathology detection and output d

GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization

Local AiDGX agent

arXiv:2605.31464v1 Announce Type: cross Abstract: GPU kernels are the workhorse of modern deep learning, and optimizing them (via evolutionary search or coding agents) usually requires repeated measur

Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies

ResearchDGX agent

arXiv:2605.30361v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer compelling energy efficiency on neuromorphic hardware, yet their training remains challenging because the discret

Graph-Conditioned Mixture of Graph Neural Network Experts for Traffic Forecasting

Model ReleasesDGX agent

arXiv:2605.30486v1 Announce Type: cross Abstract: Spatio-temporal forecasting on sensor graphs is commonly tackled with a single backbone architecture applied uniformly across all nodes, although grap

Graph Energy Matching: Transport-Aligned Energy-Based Modeling for Graph Generation

Local AiDGX agent

arXiv:2603.23398v2 Announce Type: replace-cross Abstract: Generative modeling of discrete data, such as graphs, underpins many scientific and industrial applications, including molecular discovery and

Graph Machine Learning in the Era of Large Language Models (LLMs)

ResearchDGX agent

arXiv:2404.14928v3 Announce Type: replace-cross Abstract: Graphs play an important role in representing complex relationships in various domains like social networks, knowledge graphs, and molecular d

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

Model ReleasesDGX agent

arXiv:2605.31031v1 Announce Type: new Abstract: Relational reasoning lies at the heart of intelligence, but existing benchmarks are typically confined to formats such as grids or text. We introduce Gr

GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation

SafetyDGX agent

arXiv:2605.30740v1 Announce Type: cross Abstract: Articulated object manipulation is a unique challenge for service robots. Existing methods employ end-to-end policy learning, visionmotion planning, a

HADT: A Heterogeneous Multi-Agent Differential Transformer for Autonomous Earth Observation Satellite Cluster

AgentsDGX agent

arXiv:2605.31023v1 Announce Type: new Abstract: This work addresses the problem of autonomous resource management in heterogeneous satellite cluster conducting Earth Observation (EO) missions includin

Hamiltonian-Inspired Attention Mechanism for Scalable RF Transmitter Fingerprinting

SafetyDGX agent

arXiv:2605.30364v1 Announce Type: cross Abstract: Radio-frequency (RF) fingerprinting identifies wire-less transmitters using hardware-induced imperfections present in baseband I/Q signals. However, d

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2605.30621v1 Announce Type: new Abstract: LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task

Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response

SafetyDGX agent

arXiv:2605.30680v1 Announce Type: new Abstract: Healthcare mechanisms are inseparable from the strategic provider response they induce: existing healthcare AI benchmarks hold this response fixed and s

HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2511.18760v2 Announce Type: replace Abstract: Informal mathematics has been central to modern large language model (LLM) reasoning, offering flexibility and efficient construction of arguments.

Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring

ApplicationsDGX agent

arXiv:2605.30834v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable robots to follow natural language instructions and generalize across diverse tasks, but they remain vulnera

HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

SafetyDGX agent

arXiv:2602.16165v2 Announce Type: replace-cross Abstract: Training LLMs as interactive agents for multi-turn decision-making remains challenging, particularly in long-horizon tasks with sparse and del

How Early Adopters Used Generative AI Worldwide: Variation by Country Income and Language

ResearchDGX agent

arXiv:2605.30685v1 Announce Type: cross Abstract: AI is being used by people globally, but not everyone is using it in the same ways. Using a large-scale dataset of anonymized, de-identified, and priv

Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models

SafetyDGX agent

arXiv:2508.08204v2 Announce Type: replace-cross Abstract: There has been much recent interest in evaluating large language models for uncertainty calibration to facilitate model control and modulate u

Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty

SafetyDGX agent

arXiv:2605.30675v1 Announce Type: cross Abstract: Uncertainty Quantification is a large and growing subfield of large language model behavioral analysis. Primarily to recognize and combat hallucinatio

Human Psychometric Questionnaires Mischaracterize LLM Behavior

SafetyDGX agent

arXiv:2509.10078v4 Announce Type: replace-cross Abstract: We examine whether human psychometric questionnaires can serve as reliable tools for characterizing and predicting LLM behavior in everyday us

HypoAgent: An Agentic Framework for Interactive Abductive Hypothesis Generation over Knowledge Graphs

AgentsDGX agent

arXiv:2605.31370v1 Announce Type: new Abstract: Abductive reasoning over knowledge graphs aims to generate logical hypotheses that explain observed entities or facts. Existing controllable hypothesis

idSCD: Identifying Training Datasets through Semantic Correlation Descriptors

ResearchDGX agent

arXiv:2605.30462v1 Announce Type: cross Abstract: Can a dataset be recognized from the spurious correlations it induces during training? We argue that datasets leave dataset-specific traces in a model

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

AgentsDGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment

SafetyDGX agent

arXiv:2605.30965v1 Announce Type: cross Abstract: Recent advancements in text-guided audio generation have yielded promising results in diverse domains, including sound effects, speech, and music. How

ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law

Model ReleasesDGX agent

arXiv:2605.30589v1 Announce Type: cross Abstract: U.S. immigration law spans thousands of pages of official policy, federal regulations, and procedural guidance that change frequently and carry high s

Improved Distribution Estimation in ell_infty

ResearchDGX agent

arXiv:2605.30509v1 Announce Type: cross Abstract: We present improved bounds for estimating discrete probability distributions under the ell_infty norm. These include minimax bounds in expectation and

Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data

Model ReleasesDGX agent

arXiv:2605.31324v1 Announce Type: cross Abstract: Estimating the generalization gap and developing optimization methods that improve generalization are crucial for deep learning models, for both theor

Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Agentic Systems Evaluation

AgentsDGX agent

arXiv:2605.31278v1 Announce Type: new Abstract: Reliable evaluation of agentic systems requires unbiased estimates with valid uncertainty, but standard practice navigates between costly human annotati

Inferring Events from Time Series using Language Models

ResearchDGX agent

arXiv:2503.14190v3 Announce Type: replace Abstract: A common goal in analyzing time series data is to understand how events cause observed variations. We study whether Large Language Models (LLMs) can

Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach

ResearchDGX agent

arXiv:2605.30903v1 Announce Type: cross Abstract: Inverse reinforcement learning (IRL) typically assumes demonstrations from a single optimal demonstrator, but in many applications data come from mult

Inverting Data Transformations via Diffusion Sampling

ResearchDGX agent

arXiv:2602.08267v2 Announce Type: replace-cross Abstract: We study the problem of transformation inversion on general Lie groups: a datum is transformed by an unknown group element, and the goal is to

Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

AgentsDGX agent

arXiv:2605.30677v1 Announce Type: cross Abstract: Agentic software reverse engineering systems are vulnerable to prompt injection attacks placed into the source code of executable binary files. This r

Joint angle based learning to refine kinematic human pose estimation

ResearchDGX agent

arXiv:2507.11075v2 Announce Type: replace-cross Abstract: Marker-free human pose estimation (HPE) has found increasing applications in various fields. Current HPE suffers from occasional errors in key

Kalimati Vegetable Price Index Forecasting with a Momentum Corrected Online Stacking Ensemble

ResearchDGX agent

arXiv:2605.30720v1 Announce Type: cross Abstract: Forecasting agricultural commodity prices in emerging economies is difficult due to high volatility, frequent supply disruptions, and strong cultural

KnowledgeGain: Evaluating and Optimizing Science News Generation for Reader Learning

TutorialsDGX agent

arXiv:2605.31099v1 Announce Type: cross Abstract: Science news is an important medium to communicate discoveries between the research communities and the public. Yet, most metrics for generated or sum

Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions

Model ReleasesDGX agent

arXiv:2605.31586v1 Announce Type: cross Abstract: Grasping the semantics of rare constructions (form-meaning pairings) has been shown to be a challenging problem that has currently only been solved by

LARK: Learnability-Grounded Trajectory Selection for Efficient Reasoning Distillation

SafetyDGX agent

arXiv:2605.30651v1 Announce Type: cross Abstract: We study trajectory selection for reasoning distillation, where teacher-generated reasoning trajectories are selectively used as supervision for a stu

Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control in Symbolic Music Generation

ResearchDGX agent

arXiv:2605.31295v1 Announce Type: cross Abstract: Transformer-based architectures have significantly advanced the generation of complex symbolic sequences, yet a significant gap remains in achieving f

Learning Agent-Compatible Context Management for Long-Horizon Tasks

AgentsDGX agent

arXiv:2605.30785v1 Announce Type: new Abstract: LLM agents increasingly face long-horizon tasks such as web search and deep research in real-world applications, where accumulated context can cause lon

Learning Cardiac Latent Representations in Vectorcardiogram Space

ResearchDGX agent

arXiv:2605.31249v1 Announce Type: cross Abstract: Electrocardiography (ECG) is a cornerstone of cardiac assessment, making the learning of informative ECG representations fundamental to tasks ranging

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration

AgentsDGX agent

arXiv:2605.31365v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to promising progress in web agents. However, existing web agents often rely on han

Learning to Solve and Optimize by Evolving Code

ApplicationsDGX agent

arXiv:2605.31049v1 Announce Type: cross Abstract: Combinatorial and optimization problems are fundamental to many industrial AI applications. Solving large-scale real-world instances of such problems

LH-Bench: Skill-Grounded Evaluation of Long-Horizon Agents on Subjective Enterprise Tasks

AgentsDGX agent

arXiv:2603.22744v2 Announce Type: replace Abstract: Large language models excel on objectively verifiable tasks such as math and programming, where evaluation reduces to unit tests or a single correct

Linear Ordering Problem: Time for a Change

Model ReleasesDGX agent

arXiv:2605.31051v1 Announce Type: cross Abstract: The Linear Ordering Problem (LOP) is a fundamental combinatorial optimization problem with important applications in areas such as economics, social c

LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories

Local AiDGX agent

arXiv:2605.31492v1 Announce Type: new Abstract: Large language models (LLMs) often solve reasoning problems by generating intermediate traces that explore and revise partial solutions. From a search p

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

Model ReleasesDGX agent

arXiv:2409.14583v4 Announce Type: replace Abstract: LLM bias evaluation is critical as large language models (LLMs) increasingly influence high-stakes decisions. This paper provides a comprehensive as

LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability

Local AiDGX agent

arXiv:2605.31167v1 Announce Type: new Abstract: Assessing whether Large Language Models outputs are factually grounded, epistemically calibrated, and methodologically reproducible is a prerequisite fo

LLMs Lean on Priors, Not Programming Language Semantics

ResearchDGX agent

arXiv:2510.03415v3 Announce Type: replace-cross Abstract: Recent work asks whether large language models (LLMs) condition their reasoning on explicit rules rather than statistical regularities from pr

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

Model ReleasesDGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

Model ReleasesDGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards

AgentsDGX agent

arXiv:2605.31584v1 Announce Type: cross Abstract: Long-context reasoning remains a central challenge for large language models, which often fail to locate and integrate key information in extensive di

← Previous
1…184185186187188…358
Next →