AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
16 Jul 2026

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

SafetyDGX agent

arXiv:2606.16944v2 Announce Type: replace Abstract: Theory of mind (ToM), the capacity to ascribe mental states to others and use those ascriptions for prediction and inference, is widely assumed to b

A Hybrid Mamba for Audio-Visual Navigation

ResearchDGX agent

arXiv:2607.13110v1 Announce Type: cross Abstract: Since the paradigm centered on convolutional neural networks and recurrent architectures was established in 2020, the fundamental backbone networks fo

A novel network for classification of cuneiform tablet metadata

ResearchDGX agent

arXiv:2603.03892v2 Announce Type: replace-cross Abstract: In this paper, we present a network structure for classifying metadata of cuneiform tablets. The problem is of practical importance, as the si


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Self-Evolving Agent for Longitudinal Personal Health Management

Model ReleasesDGX agent

arXiv:2607.13940v1 Announce Type: new Abstract: Personal health management unfolds over repeated encounters, yet most health AI systems treat each request in isolation. We developed HealthClaw, an ope

A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Multi-Agent Systems

AgentsDGX agent

arXiv:2507.19593v3 Announce Type: replace Abstract: Classical game-theoretic models typically assume rational agents, complete information, and common knowledge of payoffs - assumptions that are often

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks

Model ReleasesDGX agent

arXiv:2607.13305v1 Announce Type: cross Abstract: Benchmark accuracy in video large language models (LLMs) is often treated as evidence of visual understanding. We audit this assumption across twenty

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

SafetyDGX agent

arXiv:2607.13160v1 Announce Type: cross Abstract: Active beyond-diagonal reconfigurable intelligent surfaces (BD-RISs) enables hybrid transmitting and reflecting mode to achieve effective signal ampli

Adapting Generalist Vehicle Models for High-Speed MPC Across Terrains

ApplicationsDGX agent

arXiv:2607.13319v1 Announce Type: cross Abstract: High-speed off-road autonomy requires precise closed-loop control for a target vehicle while remaining robust across changing terrains. Recent forward

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

SafetyDGX agent

arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accum

Advancing Multimodal Judge Models through a Capability-Oriented Benchmark and MCTS-Driven Data Generation

Model ReleasesDGX agent

arXiv:2603.00546v2 Announce Type: replace Abstract: Using Multimodal Large Language Models (MLLMs) as judges to achieve precise and consistent evaluations has gradually become an emerging paradigm acr

Adversarial Prompting Framework for AI Safety Assessment

SafetyDGX agent

arXiv:2607.13453v1 Announce Type: cross Abstract: Artificial Intelligence (AI), especially Generative AI (GenAI), adoption has increased in industries significantly in recent years. However, the use o

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

Model ReleasesDGX agent

arXiv:2607.13705v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, the need for unified evaluation infrastructure becomes critical. However, current evaluat

Agile perceptive multi-skill locomotion for quadrupedal robots in the wild

SafetyDGX agent

arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multipl

AI-accelerated End-to-End Framework for Rapid Professional Upskilling

HardwareDGX agent

arXiv:2607.14044v1 Announce Type: new Abstract: By 2030, 59 of every 100 workers will need reskilling or upskilling, yet the average time to close an enterprise skills gap grew from roughly 3 days in

AI advice suppresses people's willingness to say 'I don't know', even when the advice is wrong and accuracy is incentivized

ResearchDGX agent

arXiv:2607.13562v1 Announce Type: new Abstract: Knowing when to say 'I don't know' is fundamental to human judgment, yet AI assistants offer a fluent answer to almost any question. In five experiments

AI-Augmented Human Resource Management? Insights from German companies

ResearchDGX agent

arXiv:2607.13839v1 Announce Type: cross Abstract: This study examines the integration of AI into Human Resource Management in German companies. We ask if and how AI-based technologies are enquote{augm

AI in Cyberpsychology: A systematic literature review of Cybersecurity enhancement by using AI for analyzing psychology of Victims, Attackers, and Defenders

ResearchDGX agent

arXiv:2607.13123v1 Announce Type: cross Abstract: Cybersecurity is the practice of protecting systems, networks, and data from digital attacks. Cyberpsychology (CPSY) is defined as the use of psycholo

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

SafetyDGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

AIMO Interpretability Challenge

Model ReleasesDGX agent

arXiv:2607.13899v1 Announce Type: new Abstract: We propose the AIMO Interpretability Challenge, a competition on distinguishing robust from spurious reasoning in frontier mathematical language models

An Explainable Agentic System for Detection of Conversational Scams with Summary-Based Memory

Model ReleasesDGX agent

arXiv:2607.11707v2 Announce Type: replace-cross Abstract: Following the rapid progress of generative Artificial Intelligence, there is a growing threat posed by conversational scams. These scams often

Analyzing Curricular Pattern Complexity Using AI to Improve On-Time Graduation Rates

ResearchDGX agent

arXiv:2607.13094v1 Announce Type: cross Abstract: The rise of Artificial Intelligence (AI) enables automatic analysis of large amounts of data. Previously time-consuming and labor-intensive tasks can

Anatomically Faithful but Temporally Blind: Auditing Attribution for Left-Ventricular Ejection-Fraction Estimation from Echocardiography

Local AiDGX agent

arXiv:2607.13738v1 Announce Type: cross Abstract: Background and Objective: Deep video models estimate left-ventricular ejection fraction (EF) from echocardiography with near-expert accuracy, and post

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

Model ReleasesDGX agent

arXiv:2607.13036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for decision support in healthcare, but clinical evidence is often incomplete or evolving. When the

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

SafetyDGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

AgentsDGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

Autonomous UAV Route Planning for Coverage Maximization in Environmental Monitoring: A Systematic Literature Review

AgentsDGX agent

arXiv:2607.13054v1 Announce Type: cross Abstract: Environmental monitoring with unmanned aerial vehicles (UAVs) requires route planning methods that maximize covered area while handling energy limits,

Barnamala: Parameter-Efficient Handwritten Devanagari Recognition at Benchmark Saturation

Model ReleasesDGX agent

arXiv:2607.13689v1 Announce Type: cross Abstract: We built a compact convolutional network (1.11 M parameters) for 46-class DHCD Devanagari recognition and reached 99.73%, the highest reported at 15.6

Baselines Before Architecture: Evaluating Coding Agents for Autonomous Penetration Testing

Model ReleasesDGX agent

arXiv:2607.13085v1 Announce Type: cross Abstract: Recent autonomous penetration testing papers report high benchmark scores while adding multi-component security harnesses around frontier LLMs. Becaus

Benefits and Limitations of Communication in Multi-Agent Reasoning

AgentsDGX agent

arXiv:2510.13903v2 Announce Type: replace-cross Abstract: Chain-of-thought prompting has popularized step-by-step reasoning in large language models, yet model performance still degrades as problem co

Beyond Backbone Backpropagation: A Decoupled Strategy for Efficient Transfer Learning

ResearchDGX agent

arXiv:2607.13043v1 Announce Type: cross Abstract: Deep learning models achieve state-of-the-art image classification but face deployment challenges due to computational costs and energy demands. We pr

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

SafetyDGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

Boltzmann MapReduce: A Partition-Function Reduce for Forkable Sandboxes

ResearchDGX agent

arXiv:2607.09689v2 Announce Type: replace Abstract: To leading order under local asymptotic normality (LAN), the confidence density a worker emits over a chunk of size n is a Gibbs--Boltzmann measure

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

AgentsDGX agent

arXiv:2607.13125v1 Announce Type: cross Abstract: We introduce Boogu-Image-0.1, an open-source unified multimodal understanding and generation model family, comprising Base, Turbo, Edit, and Edit-Turb

Can We Steer the Black-Box? Towards Controllability-Centric Evaluation of Recommender Systems with Collaborative Agents

Model ReleasesDGX agent

arXiv:2607.13418v1 Announce Type: cross Abstract: Recommender systems operate as Black-Boxes, leaving users and regulators unable to steer their outputs toward specific intentions or audit their behav

CAS I: A Geometric Coding Theorem

ResearchDGX agent

arXiv:2607.13796v1 Announce Type: cross Abstract: This paper establishes a direct analogue of the classical Coding Theorem in the setting of symmetry groups. We consider computable bijections on the s

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

Model ReleasesDGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

CayleyR: Solving the TopSpin puzzle via cycle intersection

HardwareDGX agent

arXiv:2607.13219v1 Announce Type: new Abstract: We present cayleyR, an R package for solving permutation puzzles by detecting cycle intersections in Cayley graphs. The core algorithm performs an itera

Classifying daily activities needs posture, reconstructing them needs motion

ResearchDGX agent

arXiv:2607.13216v1 Announce Type: cross Abstract: Humans recognize movements effortlessly, even from noisy and complex visual input. But what information in the stimulus allows humans to rapidly class

CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion

Model ReleasesDGX agent

arXiv:2607.13120v1 Announce Type: cross Abstract: Inferring gene regulatory networks (GRNs) from single-cell transcriptomic data is crucial for biological discovery, yet existing approaches suffer fro

Column Generation with Domain-Independent Dynamic Programming

ResearchDGX agent

arXiv:2510.14317v2 Announce Type: replace-cross Abstract: Column generation and branch-and-price (B&P) are leading mathematical optimization methods for large-scale exact optimization, iterating betwe

Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes

Model ReleasesDGX agent

arXiv:2607.13071v1 Announce Type: cross Abstract: Agentic LLM coding tools compress long session histories into compaction summaries that subsequent sessions inherit as ground truth. This paper docume

Consensus as Privileged Context for Label-Free Self-Distillation

ResearchDGX agent

arXiv:2607.13643v1 Announce Type: cross Abstract: Sampling multiple solutions and returning the majority answer is among the most reliable ways to improve the reasoning accuracy of large language mode

Continuously Evolving Deepfake Detection: An Architecture and Public-Benchmark Evaluation of a Dynamic Detection System

Model ReleasesDGX agent

arXiv:2607.13234v1 Announce Type: cross Abstract: Deepfake detectors that achieve near-perfect scores on academic benchmarks collapse on real-world content: recent in-the-wild evaluations report AUC d

Cortical-SSM: A Deep State Space Model for Motor Imagery Decoding from EEG Signals

ResearchDGX agent

arXiv:2510.15371v2 Announce Type: replace-cross Abstract: Classification of electroencephalogram (EEG) signals obtained during motor imagery (MI) has substantial application potential, including commu

Cost-Optimal Foundation Model Deployment Portfolio for Transportation Management

SafetyDGX agent

arXiv:2607.13239v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs) and vision-language models (VLMs), are increasingly used for transportation management center

Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification

ResearchDGX agent

arXiv:2607.13571v1 Announce Type: cross Abstract: Sound event detection relies on frame-level strong labels whose annotation is expensive. Active learning addresses this problem by selecting the audio

Data-Efficient Adaptation of LLMs via Attention Head Reweighting

Model ReleasesDGX agent

arXiv:2607.13425v1 Announce Type: cross Abstract: Learning effectively from limited data is critical in domains like security where labeled examples are scarce. Large language models (LLMs) have demon

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

SafetyDGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models

ResearchDGX agent

arXiv:2607.14049v1 Announce Type: new Abstract: The emergence of Chain-of-Thought (CoT) reasoning has significantly enhanced the ability of large language models (LLMs) to tackle complex, multi-step t

DeepLoop: Depth Scaling for Looped Transformers

Model ReleasesDGX agent

arXiv:2607.13491v1 Announce Type: cross Abstract: Looped Transformers scale sequential computation by applying a compact stack of physical blocks for multiple rounds, increasing unrolled depth without

Designing Safety-Constrained LLM Systems for Public Health Information Access

SafetyDGX agent

arXiv:2607.13038v1 Announce Type: cross Abstract: We present the design and implementation of a safety constrained large language model (LLM) system for public health information access, focusing on m

DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments

Model ReleasesDGX agent

arXiv:2607.13465v1 Announce Type: cross Abstract: LLM-based agents have rapidly improved at operating individual digital environments such as mobile applications, desktop systems, and smart homes. How

Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for Disaster Governance

SafetyDGX agent

arXiv:2607.13260v1 Announce Type: cross Abstract: Policy documents shape governance outcomes, but their reasoning is often implicit. Participatory commitments and managerial control routinely coexist

Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

ResearchDGX agent

arXiv:2607.13431v1 Announce Type: cross Abstract: Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offeri

Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing

SafetyDGX agent

arXiv:2607.13103v1 Announce Type: cross Abstract: Knowledge tracing (KT) aims to predict students' future performance by modeling their evolving knowledge states from historical interactions. Existing

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

Model ReleasesDGX agent

arXiv:2607.14004v1 Announce Type: new Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is repor

Early Adoption of Agentic Coding Tools by GitHub Projects

AgentsDGX agent

arXiv:2607.14037v1 Announce Type: cross Abstract: Agentic coding tools are increasingly capable of generating and submitting pull requests (PRs) to software projects, introducing new forms of human-ag

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education

SafetyDGX agent

arXiv:2607.14046v1 Announce Type: new Abstract: This paper presents Earthquaker-AI, a hybrid educational framework building upon a previously implemented educational robotics project by integrating a

Efficient and Privacy Aware Edge Cloud Collaborative Inference for Large Language Models

Local AiDGX agent

arXiv:2607.13093v1 Announce Type: cross Abstract: On-device LLM inference faces a trilemma of response latency, limited hardware resources and user privacy. Full cloud inference delivers strong comput

Efficient Text-to-Audio Generation via Pruning

Model ReleasesDGX agent

arXiv:2607.13330v1 Announce Type: cross Abstract: Diffusion-based text-to-audio generative models such as AudioLDM achieve high perceptual quality and strong semantic consistency; however, their pract

← Previous
1…6667686970…358
Next →