AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
11 Aug 2026

Imaginative Generative AI: Crossing the Entropy Wall into Worlds Beyond Imitation

ResearchDGX agent

arXiv:2608.09385v1 Announce Type: cross Abstract: Generative AI models are primarily designed to imitate the data distribution, an objective that neither corrects diversity lost by a learned generator

Improving Constraint Models with LLM Agents

AgentsDGX agent

arXiv:2608.08127v1 Announce Type: new Abstract: The runtime of Constraint Programming (CP) solvers is highly sensitive to modeling choices, such as symmetry breaking, implied constraints, global const

Improving Generalization Robustness of Multimodal RLVR

SafetyDGX agent

arXiv:2608.08802v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) makes Multimodal Large Language Models more accurate, but the gains are brittle: simply paraphrasi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

IndexTTS 2.5 Technical Report

Model ReleasesDGX agent

arXiv:2601.03888v4 Announce Type: replace-cross Abstract: In prior work, we introduced IndexTTS 2, a zero-shot neural text-to-speech foundation model comprising two core components: a transformer-base

InfoOps Bench: A live information operations safety benchmark

Model ReleasesDGX agent

arXiv:2607.28503v3 Announce Type: replace Abstract: In this paper we present an active, constantly updated AI benchmark which measures the integrity of frontier language models against being co-opted

Innovating with Generative AI: A Human Bottleneck Framework

ResearchDGX agent

arXiv:2608.07504v1 Announce Type: cross Abstract: We propose a human bottleneck perspective for understanding how generative AI transforms the innovation process. The central premise is that many cons

Instruction Set and Language for Hypergraphs

Model ReleasesDGX agent

arXiv:2607.10194v2 Announce Type: replace-cross Abstract: We present IsalHG, a method for representing the structure of any finite, connected hypergraph of bounded hyperedge arity as a string over a c

Integrated Multimodal AI System for Retrieval-Augmented Reasoning, Object Sensing, and Damage Analysis

Local AiDGX agent

arXiv:2608.08935v1 Announce Type: new Abstract: This work presents a unified multimodal AI system for damage assessment that integrates retrieval-augmented generation (RAG) models, thermal spectrum pe

IntelliAudit: Using Large Language Models to Evaluate Audit Controls

AgentsDGX agent

arXiv:2608.07688v1 Announce Type: new Abstract: IT audits require auditors to judge whether heterogeneous organizational evidence satisfies semantic security and compliance controls. This judgment is

Intelligence Foundation Model: A New Perspective to Approach Artificial General Intelligence

ResearchDGX agent

arXiv:2511.10119v4 Announce Type: replace Abstract: We propose a new perspective for approaching artificial general intelligence (AGI) through an intelligence foundation model (IFM). Unlike existing f

Interpreting Video Representations with Spatio-Temporal Sparse Autoencoders

SafetyDGX agent

arXiv:2604.03919v2 Announce Type: replace-cross Abstract: We present the first systematic study of Sparse Autoencoders (SAEs) on video representations. Standard SAEs decompose video into interpretable

JaleesBench: Are AI Assistants Good Spiritual Company?

AgentsDGX agent

arXiv:2608.07508v1 Announce Type: cross Abstract: Large language models are already advisors to millions of people of faith who bring them real decisions. The pressing question for a person of faith i

Janus: An Algorithm-Evaluator Co-Evolution Framework for LLM-Driven Discovery under Expensive Evaluation Budgets

ResearchDGX agent

arXiv:2608.08189v1 Announce Type: new Abstract: LLM-driven program discovery relies on rapid evaluator feedback, but many scientific and engineering tasks require high-fidelity simulations, hardware e

JustLLMGRPO: Radiographic Control for Chest X-Ray Generation

SafetyDGX agent

arXiv:2608.08046v1 Announce Type: new Abstract: Text-conditioned chest X-ray generation aims to synthesize realistic radiographs that faithfully depict specified findings. Existing work has primarily

Keep It Simple: Multi-Key Episodic Memory Retrieval for Ultra-Long Video Understanding

AgentsDGX agent

arXiv:2608.07663v1 Announce Type: cross Abstract: When videos extend from hours to days, directly processing them end-to-end becomes impractical for current Multi-modal Large Language Models (MLLMs).

KGCache: Amortized Subgraph Retrieval for KG Reasoning with LLMs

SafetyDGX agent

arXiv:2608.07954v1 Announce Type: new Abstract: Large language models can answer knowledge-intensive questions more reliably when they are grounded with knowledge graphs, but systems such as Think-on-

KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs

Model ReleasesDGX agent

arXiv:2608.09779v1 Announce Type: cross Abstract: Answering complex conditional questions using Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) remains a challenge, particularly

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

Model ReleasesDGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

KumbhDoot: A Scale-Ready, LLM-Bounded Architecture for Mass-Gathering Public-Service Assistants

SafetyDGX agent

arXiv:2608.07520v1 Announce Type: cross Abstract: Mass religious gatherings such as the Kumbh Mela concentrate tens of millions of people into a single region over a few weeks, producing intense, repe

KVDiagnosis: A Diagnostic Benchmark for KV-Cache Compression in Long-Context Language Models

Model ReleasesDGX agent

arXiv:2608.09412v1 Announce Type: new Abstract: KV-cache compression reduces long-context memory, but aggregate task scores reveal neither which correct executions fail nor why. We present KVDiagnosis

Large Language Models Align with the Human Brain during Creative Thinking

Model ReleasesDGX agent

arXiv:2604.03480v2 Announce Type: replace-cross Abstract: Creative thinking is a fundamental aspect of human cognition, and divergent thinking-the capacity to generate novel and varied ideas-is widely

Large Multimodal Agents for Intelligent Transportation Systems: Architectures, Evidence, and Deployment Challenges

Model ReleasesDGX agent

arXiv:2608.08184v1 Announce Type: new Abstract: Large multimodal agents (LMAs) are increasingly proposed for intelligent transportation systems (ITS), but existing studies often conflate multimodality

Latent-Frequency Validity: Fast Spectral Editing with Screened Video-VAE Transfer Operators

ResearchDGX agent

arXiv:2608.07569v1 Announce Type: cross Abstract: Direct spectral editing in video-VAE latents can control noise, flicker, smoothness, and frequency content without a decode--filter--reencode pass. Ho

LatticeMind: A Conflict-Aware Memory Primitive for Multi-Agent Systems

AgentsDGX agent

arXiv:2608.08236v1 Announce Type: new Abstract: Multi-agent LLM systems often fail not for lack of candidate answers, but because they have no persistent mechanism for deciding which incompatible clai

LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns

ResearchDGX agent

arXiv:2601.08856v3 Announce Type: replace-cross Abstract: Unit tests are critical in the hardware design lifecycle to ensure that component design modules are functionally correct and conform to the s

Layerwise goal-oriented adaptivity for neural ODEs: an optimal control perspective

ResearchDGX agent

arXiv:2601.07397v2 Announce Type: replace-cross Abstract: In this work, we propose a novel layerwise adaptive construction method for neural network architectures. Our approach is based on a goal--ori

Learning aligned EEG representations with subject-specific encoders

SafetyDGX agent

arXiv:2606.16462v2 Announce Type: replace-cross Abstract: Cross-subject EEG decoding promises more training data, but it also exposes neural networks to strong inter-subject distribution shifts. We st

Learning an Interior Layout Policy in a Domain Specific Language Action Space

SafetyDGX agent

arXiv:2608.07547v1 Announce Type: cross Abstract: Indoor scene layout generation is a challenging task in interior design. Existing methods often oversimplify the task by reducing room conditions to c

Learning from Consensus and Disagreement: Unsupervised On-Policy Self-Distillation with Minority-Trajectory Contrast

SafetyDGX agent

arXiv:2608.08764v1 Announce Type: cross Abstract: On-policy self-distillation improves language-model reasoning by querying a teacher on states actually visited by the student. Recent methods create a

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

SafetyDGX agent

arXiv:2608.09507v1 Announce Type: cross Abstract: Natural language user preferences provide an interpretable interface for LLM personalization. However, universal preference summaries often contain in

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates

SafetyDGX agent

arXiv:2608.00326v2 Announce Type: replace Abstract: Tool calling allows large language models (LLMs) to invoke external computation during problem solving, a useful capability in various fields includ

Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump Control

SafetyDGX agent

arXiv:2608.09453v1 Announce Type: cross Abstract: On--off cycling is the main cause of compressor wear in residential heat pumps, yet reinforcement learning (RL) controllers for buildings typically op

LEED: Local Embedding Evolution Distance for over-smoothing estimation and virtual node selection in GNN

TutorialsDGX agent

arXiv:2608.09596v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) suffer from two fundamental limitations: over-smoothing, where node representations become indistinguishable with depth,

Legal Responsibilities Using Autonomous Agents For Artificial Intelligence

SafetyDGX agent

arXiv:2608.08022v1 Announce Type: new Abstract: Recent incidents involving Artificial Intelligence (AI) agents, which were reported escaping their containment `unintentionally' to gain unauthorized ac

LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?

Model ReleasesDGX agent

arXiv:2503.19990v4 Announce Type: replace Abstract: Many real-world applications of spatial intelligence, such as robotic control, autonomous driving, and automated assembly, require spatial reasoning

LegoLM: Structured Weight Sharing for Large Language Models

Model ReleasesDGX agent

arXiv:2608.08652v1 Announce Type: cross Abstract: We present LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight s

Length-MAX Tokenizer for Language Models

ApplicationsDGX agent

arXiv:2511.20849v2 Announce Type: replace-cross Abstract: We introduce a new tokenizer for language models that minimizes the average tokens per character, thereby reducing the number of tokens needed

LF{}^{2}AR: Accounting for Layerwise Dynamics to Improve Multimodal Adaptation of Language Models

SafetyDGX agent

arXiv:2503.06211v3 Announce Type: replace-cross Abstract: Text-pretrained language models (LMs) encode rich world knowledge, but adapting them to process and generate perceptual modalities such as aud

LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs

Model ReleasesDGX agent

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems

LibraSpec: Dynamic Diffusion-Based Speculative Decoding via Marginal-Gain-Driven Optimization

ResearchDGX agent

arXiv:2608.08721v1 Announce Type: cross Abstract: Speculative decoding accelerates large language model inference by drafting multiple tokens for parallel verification, with efficiency critically dete

Linearized 2-Simplicial Attention

ResearchDGX agent

arXiv:2608.09307v1 Announce Type: new Abstract: We present a linearized form of 2-simplicial attention by rewriting the trilinear score as an inner product between a composite query and a key, so that

Lingjing: A Simulation Testbed for Multi-Agent Embodied Tasks in Open-Ended Cities

AgentsDGX agent

arXiv:2608.08045v1 Announce Type: new Abstract: Urban embodied intelligence requires coordination among heterogeneous agents (e.g., UAVs, ground robots, and autonomous vehicles) in dynamic cities. Sim

Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal Language Models

Model ReleasesDGX agent

arXiv:2608.09435v1 Announce Type: new Abstract: Understanding dynamic sound sources requires jointly determining what produces a sound, where the source is located, and how it moves over time. Yet exi

LITEWAY: LIghtweight HAR via Temporal Efficient highWAY

ResearchDGX agent

arXiv:2608.09421v1 Announce Type: cross Abstract: Wearable human activity recognition (HAR) remains challenging due to the computational and energy constraints of deep learning models on resource-limi

LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4

Model ReleasesDGX agent

arXiv:2607.15509v2 Announce Type: replace-cross Abstract: We present a fully automated closed-loop AutoML framework that uses GPT-5, GPT-4o, and Claude Sonnet 4 as autonomous neural architecture desig

LLM-Guided Heuristic Design from Simulation Traces: A Case Study in Dynamic Production and AGV Scheduling

Model ReleasesDGX agent

arXiv:2608.09343v1 Announce Type: new Abstract: Simulation-based optimization (SBO) evaluates executable policies under stochastic dynamics, but most methods treat the simulator as a black box: aggreg

LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Routing

SafetyDGX agent

arXiv:2608.08889v1 Announce Type: new Abstract: Recommendation systems thrive on personalization, where ''correctness'' is rarely a binary truth but a matter of subjective human preference. As Large L

LLM within MCP Matters: Measuring Inefficient Resource Utilization Driven by LLMs

Model ReleasesDGX agent

arXiv:2608.08467v1 Announce Type: new Abstract: The Model Context Protocol (MCP) standardizes how servers expose data and tools to Large Language Models (LLMs). A common server design embeds frequentl

LLMs Remember First, Forget Last: Dual-Process Interference in Large Language Models

SafetyDGX agent

arXiv:2603.00270v3 Announce Type: replace-cross Abstract: Large language models can process millions of tokens, yet how they handle conflicting information within context remains poorly understood. Fr

LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving

Model ReleasesDGX agent

arXiv:2608.08382v1 Announce Type: new Abstract: As LLM inference shifts to multi-tenant GPU clusters, co-batching improves throughput but obscures per-tenant usage and limits control. Enabling fractio

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution

Model ReleasesDGX agent

arXiv:2608.07943v1 Announce Type: new Abstract: Multi-page visually-rich document understanding (MP-VRDU) requires managing evidence that is sparse, spread across pages, and often exceeds a model's co

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

Model ReleasesDGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

LookME: Lookup-Based Multimodal Embeddings for Layer Injection in Vision-Language Models

ResearchDGX agent

arXiv:2607.16305v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved strong progress in multimodal understanding. However, scaling dense or sparse Mixture-of-Experts (

LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tasks

Model ReleasesDGX agent

arXiv:2608.07749v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning enables the adaptation of vision foundation models to biomedical tasks under limited computational resources, but a si

M^3Prune: Hierarchical Communication Graph Pruning for Efficient Multi-Modal Multi-Agent Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2511.19969v2 Announce Type: replace Abstract: Recent advancements in multi-modal retrieval-augmented generation (mRAG), which enhance multi-modal large language models (MLLMs) with external know

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

Model ReleasesDGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

MARA: Flow-Matching-Guided Multi-Agent Resource Allocation for Computational Resource Efficient Learning

SafetyDGX agent

arXiv:2608.09130v1 Announce Type: cross Abstract: Allocating limited computation among concurrent learning tasks is difficult when each task must reach a target loss before a deadline but its required

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

Model ReleasesDGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

Matching Supervision to the Student's Learning Capacity: A Unified Framework for On-Policy Self-Distillation

SafetyDGX agent

arXiv:2608.08176v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) improves the reasoning abilities of LLMs by internalizing privileged context into model parameters through self-disti

MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks

Model ReleasesDGX agent

arXiv:2507.03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education. These mod

← Previous
1…89101112…350
Next →