AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Agents

From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale

DGX agent

arXiv:2607.29516v1 Announce Type: cross Abstract: AI coding agents are generating code at volumes that exceed the capacity of traditional peer review. At the same time, existing AI code review tools o

agentsarxiv-cs-ai
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning

DGX agent

arXiv:2607.16057v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are improving rapidly as reflected in benchmark scores, yet these AI benchmarks largely test capabilities such as

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Gated Q-learning: Add Off-Policy Bias to Taste

DGX agent

arXiv:2607.28916v1 Announce Type: cross Abstract: Multistep credit assignment is critical for sample-efficient reinforcement learning, yet managing off-policy bias in Q-learning remains a fundamental

safetyarxiv-cs-ai
3 Aug 2026
Agents

Generative AI in Action: Field Experimental Evidence from Alibaba's Customer Service Operations

DGX agent

arXiv:2603.29888v2 Announce Type: replace-cross Abstract: In collaboration with Alibaba, we study how a generative AI assistant affects service performance in e-commerce after-sales operations. In a l

agentsarxiv-cs-ai
3 Aug 2026
Hardware

GPU-Accelerated ANNS: Quantized for Speed, Built for Change

DGX agent

arXiv:2601.07048v5 Announce Type: replace-cross Abstract: Approximate nearest neighbor search (ANNS) is a core problem in machine learning and information retrieval applications. GPUs offer a promisin

hardwarearxiv-cs-ai
3 Aug 2026
Research

Guarantees on Dynamical System Distinguishability for LLM Token Generation

DGX agent

arXiv:2607.28667v1 Announce Type: cross Abstract: Recent work has shown that classifying large language models (LLMs)' responses can be distinguished by modeling token embeddings as trajectories of a

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Harnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration

DGX agent

arXiv:2607.29087v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in enterprise settings, yet individual models remain bounded by model-specific capability limitat

model-releasesarxiv-cs-ai
3 Aug 2026
Research

Have I Seen You? Embedding Behavior Signals Synthetic Face Dataset Membership

DGX agent

arXiv:2607.29144v1 Announce Type: cross Abstract: Synthetic face datasets are increasingly used to reduce privacy exposure and data access constraints in biometric recognition. Yet the generators that

researcharxiv-cs-ai
3 Aug 2026
Research

HenTwin: A Multimodal Digital Twin Framework for Longitudinal Biological State Monitoring in Laying Hens

DGX agent

arXiv:2607.28652v1 Announce Type: cross Abstract: Early-life monitoring in laying hens remains constrained by fragmented single-modality sensing and the absence of formal system-level state representa

researcharxiv-cs-ai
3 Aug 2026
Local Ai

HERO: History-Enriched Rollout Training for Long-Horizon Autoregressive Neural Operators

DGX agent

arXiv:2607.29135v1 Announce Type: cross Abstract: Neural operators provide fast surrogates for time-dependent partial differential equations (PDEs) by applying a learned evolution operator recursively

local-aiarxiv-cs-ai
3 Aug 2026
Safety

How Hard Does It Think? Analyzing Step-Aware Reasoning Energy in LLM Chain-of-Thought Trajectories

DGX agent

arXiv:2607.28674v1 Announce Type: new Abstract: Understanding how computational effort is allocated across individual chain-of-thought (CoT) reasoning steps remains an open challenge: existing interpr

safetyarxiv-cs-ai
3 Aug 2026
Agents

Human-LLM Collaborative Inductive Coding for Conceptualizing K-12 Educator AI Use

DGX agent

arXiv:2607.28889v1 Announce Type: cross Abstract: Qualitative researchers increasingly encounter interaction corpora whose scale exceeds what manual coding alone can address, and large language models

agentsarxiv-cs-ai
3 Aug 2026
Safety

Hypergradient-based Bilevel Reinforcement Learning with Improved Sample Complexity

DGX agent

arXiv:2607.28849v1 Announce Type: cross Abstract: Bilevel reinforcement learning (RL) is an important framework within the literature of RL that can be used to formalize various categories of problems

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

DGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

model-releasesarxiv-cs-ai
3 Aug 2026
Hardware

Implicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations

DGX agent

arXiv:2607.29158v1 Announce Type: cross Abstract: We introduce implicit machine learning force fields (I-MLFFs), which replace explicit stacks of neural network layers with self-consistent fixed-point

hardwarearxiv-cs-ai
3 Aug 2026
Safety

Implicit Reasoning for Large Language Model-based Generative Recommendation

DGX agent

arXiv:2606.14142v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly adopted as backbones for Generative Recommendation (GR), promising access to pretrained world kn

safetyarxiv-cs-ai
3 Aug 2026
Research

Improving scDiffusion with Sparsity-Biased Classifier-Free Guidance

DGX agent

arXiv:2607.29043v1 Announce Type: cross Abstract: Single-cell RNA sequencing (scRNA-seq) has become an essential tool in modern cellular biology, and generating accurate synthetic scRNA-seq data is be

researcharxiv-cs-ai
3 Aug 2026
Model Releases

InferQ: A Database-Oriented Benchmark for Quantum Circuits Simulation

DGX agent

arXiv:2607.29134v1 Announce Type: cross Abstract: Recent work suggests that relational database management systems (RDBMSs) can execute quantum circuit simulation by compiling the simulation into SQL

model-releasesarxiv-cs-ai
3 Aug 2026
Research

Information Processing by Neuron Populations in the Central Nervous System: A Theory of the Mathematical Structure of Data and Operations

DGX agent

arXiv:2309.02332v3 Announce Type: replace-cross Abstract: In the mammalian central nervous system, neurons are organized into populations communicating by spike trains propagating along axonal bundles

researcharxiv-cs-ai
3 Aug 2026
Research

Knowledge Restoration-driven Prompt Optimization: Unlocking LLM Potential for Open-Domain Relational Triplet Extraction

DGX agent

arXiv:2601.15037v2 Announce Type: replace-cross Abstract: Open-domain Relational Triplet Extraction (ORTE) aims to mine structured knowledge without predefined relation schemas. Large Language Models

researcharxiv-cs-ai
3 Aug 2026
Research

LAWFUL: Law-Aligned Witness for Faithful Use of Latents

DGX agent

arXiv:2607.28672v1 Announce Type: cross Abstract: When a neural network predicts a physical system accurately, has it learned the governing law as formal, structured knowledge, and if so, does the net

researcharxiv-cs-ai
3 Aug 2026
Research

Learning Lookahead Lemmas for Neural Network Verification

DGX agent

arXiv:2607.29051v1 Announce Type: cross Abstract: State-of-the-art neural network verifiers use the branch-and-bound procedure as their core solving mechanism. We introduce an inprocessing framework f

researcharxiv-cs-ai
3 Aug 2026
Model Releases

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

DGX agent

arXiv:2607.29559v1 Announce Type: new Abstract: Reinforcement Learning (RL) systems are typically trained using a single, well-specified scalar reward function. However, real-world decision-making tas

model-releasesarxiv-cs-ai
3 Aug 2026
Applications

Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping

DGX agent

arXiv:2605.05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground surveys are labour-intensive and geographi

applicationsarxiv-cs-ai
3 Aug 2026
Research

Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

DGX agent

arXiv:2607.28684v1 Announce Type: new Abstract: Existing benchmarks for scientific equation discovery are largely composed of well-known equations available in the public domain, making it difficult t

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

DGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

model-releasesarxiv-cs-ai
3 Aug 2026
Research

LLM Framework for Discovering Major Mathematical Conjectures: AI's Quest for the Next Riemann Hypothesis

DGX agent

arXiv:2607.28632v1 Announce Type: new Abstract: Major mathematical conjectures still depend heavily on expert intuition, so a unified method for the systematic generation and validation of conjectures

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

DGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

MAGA: Multi-Platform Self-Fusion of GUI Agents via Structured Action Distillation

DGX agent

arXiv:2607.29320v1 Announce Type: new Abstract: Graphical user interface (GUI) agents based on large language models are increasingly deployed across mobile, web, and desktop environments. However, ex

safetyarxiv-cs-ai
3 Aug 2026
Tutorials

Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning

DGX agent

arXiv:2603.25464v2 Announce Type: replace-cross Abstract: Zero-shot reinforcement learning (RL) algorithms aim to learn a family of policies from a reward-free dataset, and recover optimal policies fo

tutorialsarxiv-cs-ai
3 Aug 2026
Local Ai

MBDiff: Multi-view Behavior-aware Diffusion Model for Probabilistic Utility Data Imputation

DGX agent

arXiv:2607.29177v1 Announce Type: cross Abstract: Utility data (e.g., electricity, water, and gas consumption), collected by ubiquitous sensors and embedded devices, often contains substantial missing

local-aiarxiv-cs-ai
3 Aug 2026
Research

Memory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory

DGX agent

arXiv:2607.29167v1 Announce Type: cross Abstract: Long-term memory lets large language model(LLM) agents reuse prior preferences and work flows, but it also turns untrusted observations into persisten

researcharxiv-cs-ai
3 Aug 2026
Agents

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

DGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

agentsarxiv-cs-ai
3 Aug 2026
Research

Metaphor-Induced Algorithmic Steering: Cross-Domain Procedural Transfer in LLM Code Generation

DGX agent

arXiv:2607.28683v1 Announce Type: cross Abstract: Large language models benefit from elements in natural language, such as metaphors and analogies in training data and inference input to achieve gener

researcharxiv-cs-ai
3 Aug 2026
Safety

metasignal: A Python Package for Comprehensive Metacognitive Analysis and Decision-Making

DGX agent

arXiv:2607.29093v1 Announce Type: cross Abstract: Metasignal is an open-source Python package for signal detection theory (SDT) and metacognitive measurement. It implements the 17 metacognitive measur

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

DGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

DGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

DGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

model-releasesarxiv-cs-ai
3 Aug 2026
Tutorials

MoRAE: Flow-Friendly Self-Supervised Latents for Text-to-Motion Generation

DGX agent

arXiv:2607.29180v1 Announce Type: cross Abstract: Text-to-motion generation must produce motions that are semantically correct, temporally coherent, and physically plausible. A natural approach is to

tutorialsarxiv-cs-ai
3 Aug 2026
Tutorials

MOSAIC: Masked Outsourcing of Secure AI Computations

DGX agent

arXiv:2607.29221v1 Announce Type: cross Abstract: We address the challenge of securely and efficiently outsourcing AI computations from a trusted but computationally weak client to an untrusted but po

tutorialsarxiv-cs-ai
3 Aug 2026
Local Ai

MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

DGX agent

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

local-aiarxiv-cs-ai
3 Aug 2026
Safety

MPP-GNN: Subject-Adaptive Community Detection for fMRI-Based Alzheimer's Disease Classification

DGX agent

arXiv:2607.28681v1 Announce Type: cross Abstract: Functional magnetic resonance imaging (fMRI) is a widely used technique for studying the brain. Recent methods that utilize graph neural networks (GNN

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO

DGX agent

arXiv:2607.28679v1 Announce Type: new Abstract: Multi-agent planning problems arise in a variety of engineering applications, such as multi-robot wildfire fighting and unmanned aerial inspection in fa

model-releasesarxiv-cs-ai
3 Aug 2026
Research

Multi-Granularity Position Embedding of Graphs via Granular-Ball for Link Prediction

DGX agent

arXiv:2607.29115v1 Announce Type: cross Abstract: Link prediction aims to identify potential or future connections within a given graph structure. Position information is essential for link prediction

researcharxiv-cs-ai
3 Aug 2026
Local Ai

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

DGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

local-aiarxiv-cs-ai
3 Aug 2026
Agents

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

DGX agent

arXiv:2607.28942v1 Announce Type: new Abstract: Recently Large Language Models (LLMs) have been increasingly deployed as autonomous agents in applications such as self-reflection, retrieval-augmented

agentsarxiv-cs-ai
3 Aug 2026
← Previous
1…4647484950…443
Next →