AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
Model Releases

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

DGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Can Open-Weight Models Compete on Financial Text Comprehension?

DGX agent

arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can We Optimize the Performance-Carbon Emission Break-Even Point?: The Quest for Greener LLMs

DGX agent

arXiv:2608.08744v1 Announce Type: cross Abstract: The carbon footprint of any deployed Large Language Model (LLM) accumulates during inference, where repeated use of the model substantially exceeds th

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions and Perception

DGX agent

arXiv:2608.08392v1 Announce Type: new Abstract: Large language models are increasingly deployed as autonomous agents that interact with the web through browsers. While recent progress has been driven

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Capability Is Not Propensity: Measuring Pressure-Robust Cooperative Behavior in Civic LLM Agents

DGX agent

arXiv:2608.09485v1 Announce Type: new Abstract: Cooperative capabilities in language models are dual-use. The same social reasoning that supports civic deliberation can also enable strategic omission,

researcharxiv-cs-ai
11 Aug 2026
Agents

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation

DGX agent

arXiv:2608.09790v1 Announce Type: new Abstract: Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions r

agentsarxiv-cs-ai
11 Aug 2026
Applications

Carnot: Interpretable, Interactive, and Optimized Execution of Deep Research Queries

DGX agent

arXiv:2608.09532v1 Announce Type: cross Abstract: Enterprises increasingly seek to query data lakes using natural language via AI-driven tools like semantic operators or deep research agents. However,

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter Shift

DGX agent

arXiv:2608.07809v1 Announce Type: new Abstract: A world model is only useful for physical AI if it changes what the agent does, and only safe if it declines to do so when it is wrong. We study both ha

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

CDGC-Net: 3D Medical Image Segmentation with Cooperative Dual-Scale Self-Attention and Grouped Channel Modeling

DGX agent

arXiv:2608.08575v1 Announce Type: cross Abstract: Accurate 3D medical image segmentation requires the integration of long-range anatomical context with fine boundary detail. Existing methods often mod

safetyarxiv-cs-ai
11 Aug 2026
Agents

CEAA: A Cognitive Embodied Agents Architecture for Interactive Computing Systems

DGX agent

arXiv:2608.09848v1 Announce Type: new Abstract: The development of embodied Intelligent Virtual Agents (IVAs) that have cognitive capabilities in real-time interactive virtual environments remains a c

agentsarxiv-cs-ai
11 Aug 2026
Research

CFD-Guided Detection of Concept Drift in Multimodal Physiologic Signals

DGX agent

arXiv:2608.07759v1 Announce Type: cross Abstract: Cardiovascular AI models can classify clean elec- trocardiogram (ECG) signals, but real wearable signals change because of motion, breathing, posture,

researcharxiv-cs-ai
11 Aug 2026
Model Releases

ChronoState: Hidden Elapsed-Time Conditioning for Temporal-State Action Selection in Frozen-Backbone Language Models

DGX agent

arXiv:2608.09124v1 Announce Type: new Abstract: Temporal decisions in language-model systems often depend on both symbolic task state and elapsed wall-clock time, such as cache expiration, job complet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CIDER: A Dataset of Contextual Disclosure Boundaries for Privacy Preference Alignment

DGX agent

arXiv:2608.09164v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human privacy preferences requires capturing individuals' disclosure boundaries beyond general privacy norms.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CircuitReason-1k: Benchmarking Long-Horizon Visual-to-Symbolic Reasoning inElectrical Circuits

DGX agent

arXiv:2608.09374v1 Announce Type: new Abstract: Electrical circuit analysis requires more than recognizing components in an image. A solver must ground symbols and labels, recover latent topology, sel

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CliniCARE-Bench: Clinical Calibrated Audit of Medical Reasoning in EHR

DGX agent

arXiv:2608.07796v1 Announce Type: new Abstract: Large language models perform strongly on medical knowledge benchmarks, but reliable clinical deployment requires agents to conduct defensible investiga

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

DGX agent

arXiv:2608.07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primari

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Coarse-to-Fine Registration of Jawbone CT and Intraoral Scan Data Using GeDi and ICP with Pseudo-IOS Ground Truth

DGX agent

arXiv:2608.07564v1 Announce Type: cross Abstract: In digital dentistry and oral surgery, the registration of jawbone CT and intraoral scanner (IOS) data is essential for integrating internal bone stru

safetyarxiv-cs-ai
11 Aug 2026
Local Ai

ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

DGX agent

arXiv:2608.09732v1 Announce Type: cross Abstract: Agent skills are emerging as an important attack surface in LLM-based agent systems. Through an empirical study of existing skill scanners, we find th

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Coupons

DGX agent

arXiv:2608.09282v1 Announce Type: new Abstract: Real-world shopping often requires constructing a basket of complementary items rather than retrieving a single product. Such combo-shopping tasks arise

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

COMEX: A Composition-Grounded Benchmark and Learning Framework for Explainable Aesthetic Image Cropping

DGX agent

arXiv:2608.07570v1 Announce Type: cross Abstract: Explainable aesthetic image cropping requires not only localizing a visually pleasing crop but also explaining why it is preferred. Existing crop-and-

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Communication-efficient distributed hazard difference estimation for heterogeneous multi-site survival data

DGX agent

arXiv:2601.14609v2 Announce Type: replace-cross Abstract: Multi-site collaboration can power survival models that no single hospital could fit alone, but privacy rules and protected computing environm

researcharxiv-cs-ai
11 Aug 2026
Agents

Complete, Scalable, and Robust Prioritized Planning for Multi-Robot Ordered Storage and Retrieval at Maximum Capacity

DGX agent

arXiv:2608.07734v1 Announce Type: cross Abstract: Automated warehouses face a fundamental trade-off between maximizing storage density and achieving high retrieval throughput. While puzzle-based stora

agentsarxiv-cs-ai
11 Aug 2026
Tutorials

Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching

DGX agent

arXiv:2608.08135v1 Announce Type: cross Abstract: Cross-modality medical image translation can reduce the burden of multi-modal acquisitions, yet the field remains constrained by two coupled limitatio

tutorialsarxiv-cs-ai
11 Aug 2026
Agents

Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Scenario

DGX agent

arXiv:2608.08131v1 Announce Type: cross Abstract: In the fictional Order 66, catastrophe does not arise from a powerful command alone: a trusted population is preconditioned, a short directive activat

agentsarxiv-cs-ai
11 Aug 2026
Safety

Concept-Guided Spatial Regularization for World Models in Atari Pong

DGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Confusion-Geometry Rebalancing for Long-Tailed Adversarial Training

DGX agent

arXiv:2608.09688v1 Announce Type: cross Abstract: Adversarial training under long tailed distributions suffers from a dual imbalance: the class imbalance skews the training objective toward head class

safetyarxiv-cs-ai
11 Aug 2026
Applications

ConMem: Contribution-Aware Memory for Long-Horizon Manufacturing Inspection Logs

DGX agent

arXiv:2607.28126v2 Announce Type: replace Abstract: Long-horizon steel-equipment inspection requires reasoning over heterogeneous records accumulated across repeated inspection cycles. Existing retrie

applicationsarxiv-cs-ai
11 Aug 2026
Research

Constraining ontology mappings using metaphysical choices

DGX agent

arXiv:2608.08122v1 Announce Type: new Abstract: In this paper we discuss the foundations behind a novel methodology for the validation of semantic mappings between different data sources based upon di

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Contamination Means Overestimation? A Fine-Grained Empirical Study in Code Intelligence

DGX agent

arXiv:2506.02791v4 Announce Type: replace-cross Abstract: In recent years, code intelligence has gained increasing importance in the field of automated software engineering. Meanwhile, the widespread

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

DGX agent

arXiv:2608.09025v1 Announce Type: new Abstract: Financial agents can turn correct context into an unauthorized effect: a customer-facing commitment, trade, or deployed policy. We present SAGE-Fin, a f

safetyarxiv-cs-ai
11 Aug 2026
Safety

Contextual Value Alignment via Multilayer Combinatorial Fusion

DGX agent

arXiv:2608.07642v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF

safetyarxiv-cs-ai
11 Aug 2026
Safety

Control-Oriented Scenario Tree Construction through Reinforcement Learning

DGX agent

arXiv:2608.09335v1 Announce Type: new Abstract: Multistage stochastic model predictive control (MPC) handles uncertainty by optimizing over a scenario tree, a finite branching approximation of future

safetyarxiv-cs-ai
11 Aug 2026
Agents

Controlled Memory Interference in Continual LLM Agents

DGX agent

arXiv:2608.07622v1 Announce Type: new Abstract: Long-term memory enables AI agents to maintain continuity across sessions, personalize behavior, and evolve through accumulated experience. Yet memory e

agentsarxiv-cs-ai
11 Aug 2026
Safety

Coordinated incentives in AI-generated misinformation governance

DGX agent

arXiv:2608.07070v1 Announce Type: cross Abstract: With the rapid diffusion of AI-generated content, AI-driven misinformation is becoming increasingly pervasive and difficult to govern, undermining inf

safetyarxiv-cs-ai
11 Aug 2026
Safety

CoRCi: Cross-Reconstruction of Coherent Interests Modeling in Cross-Domain Sequential Recommendation

DGX agent

arXiv:2608.09580v1 Announce Type: new Abstract: Cross-Domain Sequential Recommendation (CDSR) aims to alleviate data sparsity by transferring dynamic user interests across related domains. A key chall

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

CORDA: A Benchmark for Hierarchical Harm-Centric Moral Reasoning in Large Language Models

DGX agent

arXiv:2608.08061v1 Announce Type: new Abstract: The key question in moral judgement is not simply whether someone chooses the 'right' answer, but how they decide what matters most when moral principle

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CoRE: Consensus Rewards via Equilibrium for Test-Time Reinforcement Learning

DGX agent

arXiv:2608.09324v1 Announce Type: new Abstract: On unlabeled test data, reinforcement learning lacks a ground-truth reward; test-time RL methods derive one from the model's own roll-outs, rewarding th

model-releasesarxiv-cs-ai
11 Aug 2026
Research

CoRe-UIE: Rethinking Coexisting and Region-wise Degradation for Underwater Image Enhancement

DGX agent

arXiv:2608.08965v1 Announce Type: new Abstract: Underwater images often suffer from diverse and coexisting degradations, including color distortion, scattering haze, texture attenuation, and uneven il

researcharxiv-cs-ai
11 Aug 2026
Model Releases

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting

DGX agent

arXiv:2608.07693v1 Announce Type: cross Abstract: Generative traffic video forecasting aims to synthesize long-horizon, temporally coherent future videos of traffic scenes from a short observation his

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Cost Accounting for Reactive Computational Graphs: Exhaustive Sweeps, Sequential Mutation, and the Backward-Locality Gap

DGX agent

arXiv:2607.18323v2 Announce Type: replace-cross Abstract: Exhaustive site-by-site interventions on a neural network's computational graph -- activation-patching sweeps, circuit-discovery searches, sys

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Counterfactual Benchmarking and Training for Factuality Consistency and Order-Robust Grounded Reasoning in LLMs over Heterogeneous Knowledge

DGX agent

arXiv:2608.07838v1 Announce Type: new Abstract: Large language models (LLMs) have increasingly supported response generation grounded in user-provided knowledge spanning heterogeneous structures. Howe

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity

DGX agent

arXiv:2608.09443v1 Announce Type: new Abstract: Large language model (LLM) agents can support medication review between clinical visits, but safe choices for older adults with multimorbidity depend on

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge

DGX agent

arXiv:2604.03374v2 Announce Type: replace-cross Abstract: Creative problem-solving requires combining multiple cognitive abilities, including logical reasoning, lateral thinking, analogy-making, and c

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Critic-Free Deep Reinforcement Learning for Maritime Coverage Path Planning on Irregular Hexagonal Grids

DGX agent

arXiv:2603.28385v2 Announce Type: replace-cross Abstract: Maritime surveillance missions, such as search and rescue and environmental monitoring, rely on the efficient allocation of sensing assets ove

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Cross-Model Humor Preference Modeling with Cards Against Humanity

DGX agent

arXiv:2608.07481v1 Announce Type: cross Abstract: This paper investigates whether one large language model can approximate the humor preferences of another in a controlled Cards Against Humanity-style

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

CRUISE: Vision-Language Model-Guided Uncertainty-Aware Cross-Modal Sensor Fusion for Robust Autonomous Driving

DGX agent

arXiv:2608.09202v1 Announce Type: new Abstract: Modern autonomous vehicles are equipped with multiple sensors, such as cameras, LiDAR, and radar, for comprehensive environmental perception. However, r

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Cultivar: A Contrastive and Locale-Oriented Translation Benchmark for Investigating Contamination and Localisation Robustness

DGX agent

arXiv:2608.09766v1 Announce Type: cross Abstract: Multilingual translation benchmarks are typically sourced in English and translated into other languages, treating language pairs as the unit of evalu

model-releasesarxiv-cs-ai
11 Aug 2026
Research

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents

DGX agent

arXiv:2608.08638v1 Announce Type: cross Abstract: Zero-shot text-to-speech (TTS) now supports interactive assistants, personalized media, and accessibility tools. All TTS systems require faithful ling

researcharxiv-cs-ai
11 Aug 2026
← Previous
1…678910…438
Next →