AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
3 Aug 2026

Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

ResearchDGX agent

arXiv:2607.28684v1 Announce Type: new Abstract: Existing benchmarks for scientific equation discovery are largely composed of well-known equations available in the public domain, making it difficult t

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

Model ReleasesDGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

LLM Framework for Discovering Major Mathematical Conjectures: AI's Quest for the Next Riemann Hypothesis

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.28632v1 Announce Type: new Abstract: Major mathematical conjectures still depend heavily on expert intuition, so a unified method for the systematic generation and validation of conjectures

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

Model ReleasesDGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

Model ReleasesDGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

MAGA: Multi-Platform Self-Fusion of GUI Agents via Structured Action Distillation

SafetyDGX agent

arXiv:2607.29320v1 Announce Type: new Abstract: Graphical user interface (GUI) agents based on large language models are increasingly deployed across mobile, web, and desktop environments. However, ex

Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning

TutorialsDGX agent

arXiv:2603.25464v2 Announce Type: replace-cross Abstract: Zero-shot reinforcement learning (RL) algorithms aim to learn a family of policies from a reward-free dataset, and recover optimal policies fo

MBDiff: Multi-view Behavior-aware Diffusion Model for Probabilistic Utility Data Imputation

Local AiDGX agent

arXiv:2607.29177v1 Announce Type: cross Abstract: Utility data (e.g., electricity, water, and gas consumption), collected by ubiquitous sensors and embedded devices, often contains substantial missing

Memory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory

ResearchDGX agent

arXiv:2607.29167v1 Announce Type: cross Abstract: Long-term memory lets large language model(LLM) agents reuse prior preferences and work flows, but it also turns untrusted observations into persisten

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

AgentsDGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

Metaphor-Induced Algorithmic Steering: Cross-Domain Procedural Transfer in LLM Code Generation

ResearchDGX agent

arXiv:2607.28683v1 Announce Type: cross Abstract: Large language models benefit from elements in natural language, such as metaphors and analogies in training data and inference input to achieve gener

metasignal: A Python Package for Comprehensive Metacognitive Analysis and Decision-Making

SafetyDGX agent

arXiv:2607.29093v1 Announce Type: cross Abstract: Metasignal is an open-source Python package for signal detection theory (SDT) and metacognitive measurement. It implements the 17 metacognitive measur

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

Model ReleasesDGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

Model ReleasesDGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

Model ReleasesDGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

Model ReleasesDGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

MoRAE: Flow-Friendly Self-Supervised Latents for Text-to-Motion Generation

TutorialsDGX agent

arXiv:2607.29180v1 Announce Type: cross Abstract: Text-to-motion generation must produce motions that are semantically correct, temporally coherent, and physically plausible. A natural approach is to

MOSAIC: Masked Outsourcing of Secure AI Computations

TutorialsDGX agent

arXiv:2607.29221v1 Announce Type: cross Abstract: We address the challenge of securely and efficiently outsourcing AI computations from a trusted but computationally weak client to an untrusted but po

MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

Local AiDGX agent

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

MPP-GNN: Subject-Adaptive Community Detection for fMRI-Based Alzheimer's Disease Classification

SafetyDGX agent

arXiv:2607.28681v1 Announce Type: cross Abstract: Functional magnetic resonance imaging (fMRI) is a widely used technique for studying the brain. Recent methods that utilize graph neural networks (GNN

Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO

Model ReleasesDGX agent

arXiv:2607.28679v1 Announce Type: new Abstract: Multi-agent planning problems arise in a variety of engineering applications, such as multi-robot wildfire fighting and unmanned aerial inspection in fa

Multi-Granularity Position Embedding of Graphs via Granular-Ball for Link Prediction

ResearchDGX agent

arXiv:2607.29115v1 Announce Type: cross Abstract: Link prediction aims to identify potential or future connections within a given graph structure. Position information is essential for link prediction

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

Local AiDGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

AgentsDGX agent

arXiv:2607.28942v1 Announce Type: new Abstract: Recently Large Language Models (LLMs) have been increasingly deployed as autonomous agents in applications such as self-reflection, retrieval-augmented

NeurOWL: An LLM-Based Neural-symbolic Framework for Incomplete OWL Ontology Reasoning

ApplicationsDGX agent

arXiv:2607.15776v2 Announce Type: replace Abstract: OWL ontologies provide a formal knowledge representation framework that enables semantic reasoning, and have been widely adopted across domains such

'Not in My Backyard': LLMs Uncover Online and Offline Social Biases Against Homelessness

Model ReleasesDGX agent

arXiv:2508.13187v5 Announce Type: replace-cross Abstract: Homelessness is a persistent social challenge, impacting millions worldwide. Over 876,000 people experiencing homelessness (PEH) were recorded

On the Expressive Power of Sparse Geometric MPNNs

ResearchDGX agent

arXiv:2407.02025v5 Announce Type: replace-cross Abstract: Motivated by applications in chemistry and other sciences, we study the expressive power of message-passing neural networks for geometric grap

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness

Model ReleasesDGX agent

arXiv:2607.29062v1 Announce Type: new Abstract: Model capabilities have improved in large part due to scaling chain of thought. This has been a promising development for AI safety--where models verbal

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems

Model ReleasesDGX agent

arXiv:2607.28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural und

OPERA: Online Data Pruning for Efficient Retrieval Model Adaptation

ResearchDGX agent

arXiv:2603.17205v3 Announce Type: replace-cross Abstract: Domain-specific finetuning is essential for dense retrievers, yet not all data pairs contribute equally to the learning process. We introduce

OsteoCAD: A Human-in-the-Loop Cloud-Edge Framework for Bone Tumor Segmentation

HardwareDGX agent

arXiv:2607.29266v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Deep Learning (DL) have notably advanced medical image analysis, yet many health- care organizations struggle to adop

PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learning under Explicit Limits

Model ReleasesDGX agent

arXiv:2607.28982v1 Announce Type: cross Abstract: Recent language models achieve strong performance across a variety of tasks, but conventional adaptation applies updates uniformly across training sam

Patch-Based 3D Variational Autoencoder for Super-Resolution of Turbulent Channel Flow

Model ReleasesDGX agent

arXiv:2507.22082v2 Announce Type: replace-cross Abstract: Direct numerical simulation (DNS) accurately resolves all spatio-temporal scales of wall-bounded turbulence but becomes prohibitively expensiv

Pay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers

ApplicationsDGX agent

arXiv:2511.00847v5 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) through Application Programming Interfaces (APIs) induces a critical vulnerability: th

Point2Radio: A Foundation Model for Cross-Scene Radio Fields from Material-Aware Point Clouds

HardwareDGX agent

arXiv:2607.28994v1 Announce Type: cross Abstract: High-fidelity radio fields are typically simulated for every scene--transmitter configuration or fitted separately to each scene, failing to exploit p

Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning

Model ReleasesDGX agent

arXiv:2607.28695v1 Announce Type: cross Abstract: Here is the plain text version optimized for arXiv's submission form. Custom macros (like CV and SI) have been converted to standard text/math so they

QR-Structured Thermal Triggers for Targeted Semantic Attacks on Infrared Vision-Language Models

SafetyDGX agent

arXiv:2607.29445v1 Announce Type: cross Abstract: Infrared vision-language models (IR-VLMs) extend thermal perception to open-vocabulary classification, image captioning, and visual question answering

RAID: Towards Robust AI-Generated Image Detection with Bit-Reversed Images

ResearchDGX agent

arXiv:2607.28974v1 Announce Type: cross Abstract: The rapid advancement of image generation models has made it increasingly difficult for people to distinguish AI-generated images from real ones. To p

RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving

SafetyDGX agent

arXiv:2602.07339v2 Announce Type: replace Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck fo

RareSense: Rarity-Aware Similarity Search for Anomaly Retrieval in Transactional Data

Model ReleasesDGX agent

arXiv:2607.28879v1 Announce Type: cross Abstract: Similarity search over sparse set-valued data is often dominated by frequent background attributes because classical measures such as Jaccard, cosine,

Reasoning in Real World Clinical Care: Why Large Language Models Are Not Yet Safe for Autonomous Clinical Decision Support

SafetyDGX agent

arXiv:2607.28677v1 Announce Type: new Abstract: LLM now pass medical licensing examinations and, in curated cases, can rival physicians at diagnostic reasoning. These developments have accelerated the

RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems

Local AiDGX agent

arXiv:2607.29241v1 Announce Type: cross Abstract: Optimizing modern recommender models still depends heavily on engineers manually iterating over architectural, objective, and training-strategy change

Reflected UAS: Corrected Deterministic Stability and Direct CTMC Drift Calculation

Model ReleasesDGX agent

arXiv:2607.28688v1 Announce Type: cross Abstract: We analyze Reflected UAS routing for heterogeneous multi-server queues at fixed parameters under subcritical load. The deterministic surrogate is a re

RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment

SafetyDGX agent

arXiv:2507.22580v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) seeks to automatically correct software bugs without requiring human intervention. However, existing tools tend

Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion

AgentsDGX agent

arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, spor

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning

ResearchDGX agent

arXiv:2606.13316v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a central technique for improving long-horizon reasoning in Large Language Models (LLMs). H

Retrieval-Driven Training-Free AI-Generated Video Attribution

Model ReleasesDGX agent

arXiv:2607.28955v1 Announce Type: cross Abstract: AI-generated videos are becoming increasingly realistic and difficult to distinguish from authentic ones, which facilitates malicious misuse and poses

Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations

Model ReleasesDGX agent

arXiv:2410.06665v4 Announce Type: replace-cross Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike tradition

Robust Bidirectional Associative Memory via Regularization Inspired by the Subspace Rotation Algorithm

SafetyDGX agent

arXiv:2511.11902v2 Announce Type: replace-cross Abstract: Bidirectional Associative Memory (BAM) trained with Bidirectional Backpropagation (B-BP) often suffers from poor robustness and high sensitivi

Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing

Model ReleasesDGX agent

arXiv:2607.28814v1 Announce Type: cross Abstract: In Motivational Interviewing (MI), a client's sustain talk (arguments for the status quo) calls for the counselor to roll with resistance, a move that

SAF-OPD: Stable Advantage Fusion for On-Policy Distillation

SafetyDGX agent

arXiv:2607.29209v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) broadcasts a single response-level reward to every token, while on-policy distillation (OPD) sco

Safety, or Just Capability? A Validity Audit of Agent-Safety Benchmarks

Model ReleasesDGX agent

arXiv:2607.28685v1 Announce Type: new Abstract: Agent-safety benchmarks measure different behaviors, and their scores get quoted interchangeably as an agent's safety. We treat four of them (R-Judge, I

SATViz: Real-Time Visualization of Clausal Proofs

ResearchDGX agent

arXiv:2209.05838v2 Announce Type: replace Abstract: Visual layouts of graphs representing SAT instances can highlight the community structure of SAT instances. The community structure of SAT instances

Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks

SafetyDGX agent

arXiv:2607.28630v1 Announce Type: cross Abstract: Generative AI (GenAI) holds significant promise for advancing educational equity among ethnic minority students by broadening access to learning resou

Scaling Scientific Discovery Environments for Turn-Level Agentic RL

AgentsDGX agent

arXiv:2607.28990v1 Announce Type: new Abstract: Large language model agents have shown promising capabilities in data-driven scientific discovery tasks, where an agent interacts with an execution envi

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

Model ReleasesDGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

SCMA: Structure-Conditioned and Metal-Aware Flow Matching for CT Metal Artifact Reduction

ResearchDGX agent

arXiv:2607.28759v1 Announce Type: cross Abstract: In X-ray CT, metallic objects cause beam hardening, photon starvation, and scattering, leading to projection inconsistency, streaks, dark bands, and s

SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Industrial NILM

Model ReleasesDGX agent

arXiv:2607.28693v1 Announce Type: cross Abstract: Industrial NILM remains challenging because measurement noise and widespread concurrent machine operation reduce the generalization of models tuned on

Seeing Differently: Modeling Interpretive Perspectives in Computational Creativity using a Four-World Framework

ResearchDGX agent

arXiv:2607.28644v1 Announce Type: cross Abstract: Creativity in computational systems is often evaluated as an objective property of artifacts, with existing Computational Creativity (CC) frameworks a

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

Model ReleasesDGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

← Previous
1…3738394041…354
Next →