AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
24 Jul 2026

Benchmarking Large Language Models on Multi-Sensor Physical Hazard Assessment

Model ReleasesDGX agent

arXiv:2607.20476v1 Announce Type: new Abstract: We present an empirical benchmark evaluating how five large language models assess multisensor physical hazard data. Testing 60 scenarios across three c

Benchmarking the Personalization Capabilities of Large Language Models

Model ReleasesDGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

Benchmarking Unlearning for Vision Transformers

Model ReleasesDGX agent

arXiv:2602.20114v2 Announce Type: replace-cross Abstract: Machine unlearning (MU) refers to the post-training capability to remove (the influence of) training examples that are incorrect, biased, or l


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond Heavy Log Curation: Perplexity-Based APT Detection via Unsupervised, Context-Augmented Language Models

ResearchDGX agent

arXiv:2607.20832v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) remain difficult to detect because only a small fraction of events in large-scale logs are attack-related, and inve

Beyond Independent Optimization: Compression, MoE Routing, and Quantization Interactions in Multimodal Edge Intelligence

ResearchDGX agent

arXiv:2607.20981v1 Announce Type: new Abstract: Efficient multimodal inference is increasingly constrained not only by model quality or FLOP count, but also by the cost of preserving, moving, routing,

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

Model ReleasesDGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design

TutorialsDGX agent

arXiv:2607.20550v1 Announce Type: cross Abstract: The traditional 'one drug, one target' paradigm of structure-based drug design (SBDD) frequently proves inadequate for treating multifactorial disease

Beyond Sufficiency: Time Series Explanation with Counterfactual Necessity

ApplicationsDGX agent

arXiv:2607.21573v1 Announce Type: cross Abstract: Faithful explanations of time-series classifiers should identify subsequences that are not only sufficient to preserve a black-box model's prediction,

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

SafetyDGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

Bound-Founded Semantics for Answer Set Programming with Difference Constraints: Preliminary Report

ResearchDGX agent

arXiv:2607.21201v1 Announce Type: new Abstract: While the integration of linear constraints has significantly expanded the reach of Answer Set Programming (ASP), existing hybrid solvers often rely on

Break Through the Compression Bottleneck: From Theory to Practice

Model ReleasesDGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for Dynamic Graph Systems

ResearchDGX agent

arXiv:2607.21421v1 Announce Type: new Abstract: Generative models can support decision-making under uncertainty by producing ensembles of plausible future system trajectories, but statistical plausibi

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

Model ReleasesDGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows

AgentsDGX agent

arXiv:2607.17528v3 Announce Type: replace Abstract: Large language model (LLM) agents are extending electronic design automation (EDA) beyond static RTL generation toward long-horizon, tool-interactiv

Can an AI System Be Creative? A Critical Perspective from Art and Engineering

ResearchDGX agent

arXiv:2607.20796v1 Announce Type: new Abstract: This paper examines the question of whether artificial intelligence (AI) systems can be creative, approached from the dual perspective of a researcher t

Can Generative Recommendation Reach Cold Items? A Temporal Perspective on Semantic-ID Generation

ResearchDGX agent

arXiv:2607.21101v1 Announce Type: new Abstract: Semantic-ID-based generative recommendation represents items as sequences of shared semantic tokens, enabling token recombination beyond isolated item I

Can Valence Reflect Morality in Natural Language? A Preliminary Annotation Study

ResearchDGX agent

arXiv:2607.20461v1 Announce Type: cross Abstract: Present implementations of artificial intelligence (AI) ethics do not adequately take feelings, or affect, into account. If AI should be aligned with

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

Model ReleasesDGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

Case study: proving sqrt(2) irrational with LPTP and an LLM

ApplicationsDGX agent

arXiv:2607.21187v1 Announce Type: cross Abstract: We present the interactions with an LLM (Large Language Model) aiming at proving that the square root of 2 is not a rational number in an LP (Logic Pr

Case study: solving P-99 with LPTP and an LLM

Model ReleasesDGX agent

arXiv:2607.21196v1 Announce Type: cross Abstract: Ninety-Nine Prolog Problems (P-99) is a famous set of Prolog exercises. We solved the first thirty three just by prompting an LLM (Large Language Mode

Chess_db: A framework for working with large chess game datasets

ResearchDGX agent

arXiv:2607.21195v1 Announce Type: cross Abstract: Chess is a two player strategic game that is embedded in classical AI culture as it was once the frontier for intelligent behaviour. There was the sil

ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models

ResearchDGX agent

arXiv:2607.20463v1 Announce Type: new Abstract: This paper presents an AI-driven browser extension that identifies clickbait to help users avoid misleading Internet articles. Moving beyond traditional

CLOE: Christoffel Loss Autoencoder for Anomaly Detection

TutorialsDGX agent

arXiv:2607.20530v1 Announce Type: cross Abstract: Semi-supervised anomaly detection plays a key role in diverse fields such as process monitoring, healthcare, and finance. However, lightweight methods

Clustered Edge Intelligence: Beyond Just Convergence of Edge Computing and AI

ResearchDGX agent

arXiv:2607.20937v1 Announce Type: new Abstract: We are moving from an information age to the age of intelligence. A decade, or possibly less than that, data will not be the gold anymore rather the der

CMI-Mem: Toward Generalizable Long-Term Memory Management via CMI-Augmented Reinforcement Learning

AgentsDGX agent

arXiv:2607.20553v1 Announce Type: new Abstract: Memory Manager models are pivotal in agent systems. Existing methods rely predominantly on LLM-judged synthetic question-answer (QA) pairs, making memor

Code Monitor Red Teaming for Public-Test-Passing Code

ResearchDGX agent

arXiv:2607.20852v1 Announce Type: new Abstract: Visible tests are a common gate for LLM-generated code, but passing them does not certify specification correctness. We study a deployment-like monitori

Codec-Gauge: Learning Compression-Friendly Gauges for Transformer KV Caches

ResearchDGX agent

arXiv:2607.20538v1 Announce Type: cross Abstract: Long-context Transformer inference increasingly relies on KV-cache compression or quantization. Prior rotation and transform-coding results suggest th

Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections

AgentsDGX agent

arXiv:2607.21488v1 Announce Type: cross Abstract: Coordinating autonomous vehicles at unsignalized intersections remains a critical challenge for multi-agent reinforcement learning (MARL) systems, whi

Compile, Then Page: Executable SOP Programs and a Capability-Gated Runtime for Procedural LLM Agents

SafetyDGX agent

arXiv:2607.11346v3 Announce Type: replace Abstract: Enterprise agents must follow long-horizon, conditional, safety-critical standard operating procedures (SOPs). We compile machine-readable SOP const

ConfidenceBench: Evaluating Confidence Calibration in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20526v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings where fluent but incorrect answers can be costly. In these settings, accuracy alone i

Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception

SafetyDGX agent

arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service of a contextually or experimentally induced goal.

CRAG-MM-Diagnostics: Enabling Stage-Wise Analysis of Knowledge-Intensive VQA

Model ReleasesDGX agent

arXiv:2607.21155v1 Announce Type: cross Abstract: Knowledge-Intensive Visual Question Answering (KI-VQA) benchmarks evaluate Vision-Language Models (VLMs) as multimodal knowledge assistants by requiri

CRAWO: Custom Resources for Adaptive Workload Orchestration

ResearchDGX agent

arXiv:2607.20490v1 Announce Type: new Abstract: Edge Intelligence has emerged as a key paradigm for enabling real-time applications in smart cities by shifting computation from centralized cloud data

Cycle-Consistent and Uncertainty-Aware Neural Surrogates for Tokamak Edge Plasmas

Model ReleasesDGX agent

arXiv:2607.21407v1 Announce Type: cross Abstract: The boundary and divertor plasma govern how a tokamak exhausts power and particles, setting heat fluxes, target conditions, and the onset of detachmen

DC-Leap: Training-Free Acceleration of dLLMs via Draft-Guided Contiguous Leaping Decoding

ResearchDGX agent

arXiv:2607.20467v1 Announce Type: new Abstract: While parallel decoding is central to the efficiency of Diffusion Large Language Models (dLLMs), current strategies are often hindered by overly conserv

Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos

Model ReleasesDGX agent

arXiv:2506.19445v4 Announce Type: cross Abstract: We introduce the largest real-world image deblurring dataset constructed from smartphone slow-motion videos. Using 240 frames captured over one second

Declarative Problem Solving in UAM Strategic Deconfliction

ResearchDGX agent

arXiv:2607.21197v1 Announce Type: cross Abstract: The growing demand for Urban Air Mobility (UAM) introduces significant challenges in airspace management, particularly within densely populated metrop

DecodeShare: Tracing the Shared Subspace of LLM Decode-Time Decisions

ResearchDGX agent

arXiv:2607.20469v1 Announce Type: new Abstract: Large language models (LLMs) handle many tasks with one set of parameters, but under KV-cached inference it is unclear what task-general structure, if a

Delivery, Not Storage: Cue-Anchored Working Memory as a Harness Property for Coding Agents

Local AiDGX agent

arXiv:2607.20972v1 Announce Type: new Abstract: Coding agents ship with one kind of memory: documents. Instruction files, plan artifacts, and auto-written memory directories are deliberately authored

Demographically-Informed Heat-Mortality Risk Curves via Risk Graph Neural Networks

ResearchDGX agent

arXiv:2607.21131v1 Announce Type: cross Abstract: Estimating heat-related mortality risk is a core task in environmental epidemiology, typically addressed with Distributed Lag Non-linear Models (DLNMs

Demonstrating GenDB: Instance-Optimized and Customized Query Processing Code Generation via LLM Agents

Model ReleasesDGX agent

arXiv:2607.20630v1 Announce Type: cross Abstract: Traditional query processing engines require continuous development and extensions to support new techniques and user requirements, and in some cases,

Detecting LLM-Generated Tokens in Human--LLM Coauthored Text

Local AiDGX agent

arXiv:2607.21458v1 Announce Type: new Abstract: The rise of human-AI collaborative writing has created a growing need for fine-grained detection methods that support localizing likely LLM-generated co

DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making

Model ReleasesDGX agent

arXiv:2607.20491v1 Announce Type: new Abstract: Standard evaluation benchmarks measure what a tool-using agent decides, not whether it arrives at that decision through the same process each time. We i

Diagnosing Pathological Chain-of-Thought in Reasoning Models

SafetyDGX agent

arXiv:2602.13904v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning is fundamental to modern LLM architectures and represents a critical intervention point for AI safety. However, CoT

Differentiable Logic Programming to Mitigate Reasoning Shortcuts in Neurosymbolic Systems

ResearchDGX agent

arXiv:2607.21185v1 Announce Type: new Abstract: Neurosymbolic (NeSy) systems integrate neural networks with logical reasoning to achieve both generalization and interpretability, but recent work has s

DINOde: Continuous Vision-Text Alignment for Open-Vocabulary Semantic Segmentation

SafetyDGX agent

arXiv:2607.21371v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) leverages textual semantics to segment objects beyond predefined categories. While the self-supervised mo

Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering

ResearchDGX agent

arXiv:2607.20487v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to answer questions about political information, including in election-adjacent information settings

Drive As You Like: Multi-Head Diffusion with Reinforcement Learning for Personalized Driving

Model ReleasesDGX agent

arXiv:2508.16947v2 Announce Type: replace-cross Abstract: Despite significant progress, imitation learning-based autonomous driving planners remain largely restricted to reproducing high-frequency bia

Dropping the Anchor: Statistical Context Summarization for Distributed Systems via Pulsar Attention

Model ReleasesDGX agent

arXiv:2607.20457v1 Announce Type: cross Abstract: Inference with large language models (LLMs) on long sequences is computationally expensive due to the quadratic complexity of self-attention. Distribu

DS@GT ARC at ImageCLEFmed GANs 2026: Geometric Filtering for Privacy-Preserving CT Slice Generation

ResearchDGX agent

arXiv:2607.20692v1 Announce Type: cross Abstract: We present a privacy-preserving framework for synthetic lung CT slice generation developed for the Image-CLEFmed GANs 2026 challenge. The approach com

DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers

SafetyDGX agent

arXiv:2508.21797v2 Announce Type: replace-cross Abstract: Industry 4.0's highly networked Machine Tool Controllers (MTCs) are prime targets for replay attacks that use outdated sensor data to manipula

DynamicMCPBench: A Trace-Grounded, Effect-Scored Benchmark for LLM Agents over Live MCP Servers

Model ReleasesDGX agent

arXiv:2607.20531v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed over Model Context Protocol (MCP) servers, yet the benchmarks used to evaluate them score th

Efficient and Interpretable Body-Based Emotion Recognition with Lightweight Temporal Convolutional Networks

Model ReleasesDGX agent

arXiv:2607.20820v1 Announce Type: new Abstract: Body-based emotion recognition is important for real-time affective systems, but graph-based skeleton models can be computationally expensive. This pape

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing

TutorialsDGX agent

arXiv:2607.21529v1 Announce Type: cross Abstract: Test-Time Tuning (TTT) on pretrained diffusion models has emerged as a powerful paradigm for video editing. However, there exists a foundational misma

Emergent Compositional Skills in Mixture-of-Experts VLAs

TutorialsDGX agent

arXiv:2607.20771v1 Announce Type: cross Abstract: We consider the problem of learning compositional robot policies end-to-end from expert demonstrations, without any pre-specified notion of task decom

EmoAgent-R1: Towards Multimodal Emotion Understanding with Reinforcement Learning-based Dynamic Agent Specialization

SafetyDGX agent

arXiv:2607.21013v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance in multimodal emotion recognition (MER) tasks and lifted MER to a new leve

Enabling Scalable Topology Inference in Distribution Systems via Constrained Multi-Source Inference

Local AiDGX agent

arXiv:2607.20480v1 Announce Type: new Abstract: Accurate distribution system topology is essential for outage localization, voltage analytics, and operation of distribution grids, yet maintaining reli

Encoding Event-B Proof Rules in Prolog: An Interactive Sequent Prover for ProB

ResearchDGX agent

arXiv:2607.21191v1 Announce Type: cross Abstract: Event-B is a formal method rooted in predicate logic and set theory. We encoded over 600 proof rules in Prolog, enabling a systematic, comprehensible

Enhancing Explainable Cardiac Diagnosis with Guide-Grounded Multimodal LLMs

SafetyDGX agent

arXiv:2607.20814v1 Announce Type: new Abstract: The electrocardiogram (ECG) is a cornerstone of cardiac as- sessment, yet clinical deployment of deep learning models remains con- strained by limited i

Equivariant Conditional Diffusion Model for Head and Neck CT Image Synthesis from CBCT

TutorialsDGX agent

arXiv:2509.21913v2 Announce Type: replace-cross Abstract: Background: Cone-beam computed tomography CBCT is a commonly used modality for image guided radiotherapy. It offers real time anatomical visua

← Previous
1…5455565758…354
Next →