AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

Beyond Tables: Doc2DB-Bench for Relationally Faithful Document-to-Database Construction

DGX agent

arXiv:2608.08459v1 Announce Type: cross Abstract: Practical AI systems increasingly need to turn long, heterogeneous documents into queryable relational databases, not isolated spreadsheets. In domain

model-releasesarxiv-cs-ai
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Beyond Uniform Restoration: Empowering All-in-One Restoration with Pixel-Level Multimodal Guidance

DGX agent

arXiv:2608.09482v1 Announce Type: cross Abstract: All-in-one image restoration is a unified low-level vision task that aims to effectively recover high-quality images from inputs degraded by various t

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents

DGX agent

arXiv:2608.09555v1 Announce Type: new Abstract: External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effe

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Biologically Informed Representation Learning for Robust Cross-Center Generalization of MALDI-TOF Mass Spectrometry

DGX agent

arXiv:2608.08182v1 Announce Type: cross Abstract: Machine learning models for MALDI-TOF mass spectrometry have shown considerable promise for clinical microbiology tasks such as microbial identificati

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Bounding Hallucinations: Merlin-Arthur Protocols for Mutual-Information Bounds in Language Models

DGX agent

arXiv:2512.11614v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) relies on retrieved context to guide large language models (LLM), yet treats the retrieval as a heuristic

agentsarxiv-cs-ai
11 Aug 2026
Research

BRACE: Taming Sharp Irregularities via Barycentric Rational Forecasting for Fast Diffusion Transformers Inference

DGX agent

arXiv:2608.07572v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have demonstrated exceptional performance in high-fidelity image and video generation. To alleviate their massive comput

researcharxiv-cs-ai
11 Aug 2026
Agents

Branch2Skill: Efficient Skill Evolution Through Reasoning Trees

DGX agent

arXiv:2608.08677v1 Announce Type: new Abstract: Skill evolution improves agent skills through feedback over time, with failed trajectories often providing informative signals by revealing incomplete o

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Bridging the Evaluation Gap: Standardized Benchmarks for Multi-Objective Search

DGX agent

arXiv:2603.24084v2 Announce Type: replace Abstract: Empirical evaluation in multi-objective search (MOS) has historically suffered from fragmentation, relying on heterogeneous problem instances with i

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

Bridging the Gap Between Semantics and Reconstruction:Unifying Sign Language Translation and Production

DGX agent

arXiv:2608.09045v1 Announce Type: cross Abstract: Recent advances in sign language (SL) research have shown a trend toward unifying multiple sign language understanding (SLU) subtasks, such as isolate

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation detection in social media posts

DGX agent

arXiv:2608.09510v1 Announce Type: cross Abstract: Detecting machine-generated disinformation on social media is increasingly difficult as large language models (LLMs) make it easier to generate and re

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Business Arena: Benchmarking LLM Agents in a Realistic Marketplace

DGX agent

arXiv:2608.08621v1 Announce Type: new Abstract: Running a business is a challenging form of intelligent work. Operators must infer opportunities from partial signals, commit capital under uncertainty,

agentsarxiv-cs-ai
11 Aug 2026
Agents

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline

DGX agent

arXiv:2608.09254v1 Announce Type: new Abstract: LLM analytics agents are evaluated on SQL syntax accuracy, but production failures look different: questions with two valid business definitions, questi

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

CADEngBench: It Looks Like CAD, but Does It Work? Evaluating Parametric Design, Assembly Reasoning, and Physics Simulation

DGX agent

arXiv:2608.09296v1 Announce Type: new Abstract: A CAD model is not engineering-grade merely because it looks correct. It must satisfy design requirements, respond predictably to parameter changes, sup

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Calling the Bluff: Detecting Ever-Shifting Harmful Chat Dialogue via Ordered Reasoning Chain Regularization

DGX agent

arXiv:2608.08451v1 Announce Type: cross Abstract: Harmful chat dialogues are ever-shifting through type-shifting and lexical evasion, yet we find they share invariant principles, i.e., an Ordered Reas

researcharxiv-cs-ai
11 Aug 2026
Agents

Can Coding Agents Solve Repository-Level Issues with Rendered Code? An Exploratory Study of Visual Representations

DGX agent

arXiv:2608.09268v1 Announce Type: cross Abstract: Visual modality has recently been explored as a way to compress textual tokens, including rendering code as images for static code understanding. We s

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

DGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can Open-Weight Models Compete on Financial Text Comprehension?

DGX agent

arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can We Optimize the Performance-Carbon Emission Break-Even Point?: The Quest for Greener LLMs

DGX agent

arXiv:2608.08744v1 Announce Type: cross Abstract: The carbon footprint of any deployed Large Language Model (LLM) accumulates during inference, where repeated use of the model substantially exceeds th

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions and Perception

DGX agent

arXiv:2608.08392v1 Announce Type: new Abstract: Large language models are increasingly deployed as autonomous agents that interact with the web through browsers. While recent progress has been driven

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Capability Is Not Propensity: Measuring Pressure-Robust Cooperative Behavior in Civic LLM Agents

DGX agent

arXiv:2608.09485v1 Announce Type: new Abstract: Cooperative capabilities in language models are dual-use. The same social reasoning that supports civic deliberation can also enable strategic omission,

researcharxiv-cs-ai
11 Aug 2026
Agents

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation

DGX agent

arXiv:2608.09790v1 Announce Type: new Abstract: Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions r

agentsarxiv-cs-ai
11 Aug 2026
Applications

Carnot: Interpretable, Interactive, and Optimized Execution of Deep Research Queries

DGX agent

arXiv:2608.09532v1 Announce Type: cross Abstract: Enterprises increasingly seek to query data lakes using natural language via AI-driven tools like semantic operators or deep research agents. However,

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter Shift

DGX agent

arXiv:2608.07809v1 Announce Type: new Abstract: A world model is only useful for physical AI if it changes what the agent does, and only safe if it declines to do so when it is wrong. We study both ha

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

CDGC-Net: 3D Medical Image Segmentation with Cooperative Dual-Scale Self-Attention and Grouped Channel Modeling

DGX agent

arXiv:2608.08575v1 Announce Type: cross Abstract: Accurate 3D medical image segmentation requires the integration of long-range anatomical context with fine boundary detail. Existing methods often mod

safetyarxiv-cs-ai
11 Aug 2026
Agents

CEAA: A Cognitive Embodied Agents Architecture for Interactive Computing Systems

DGX agent

arXiv:2608.09848v1 Announce Type: new Abstract: The development of embodied Intelligent Virtual Agents (IVAs) that have cognitive capabilities in real-time interactive virtual environments remains a c

agentsarxiv-cs-ai
11 Aug 2026
Research

CFD-Guided Detection of Concept Drift in Multimodal Physiologic Signals

DGX agent

arXiv:2608.07759v1 Announce Type: cross Abstract: Cardiovascular AI models can classify clean elec- trocardiogram (ECG) signals, but real wearable signals change because of motion, breathing, posture,

researcharxiv-cs-ai
11 Aug 2026
Model Releases

ChronoState: Hidden Elapsed-Time Conditioning for Temporal-State Action Selection in Frozen-Backbone Language Models

DGX agent

arXiv:2608.09124v1 Announce Type: new Abstract: Temporal decisions in language-model systems often depend on both symbolic task state and elapsed wall-clock time, such as cache expiration, job complet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CIDER: A Dataset of Contextual Disclosure Boundaries for Privacy Preference Alignment

DGX agent

arXiv:2608.09164v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human privacy preferences requires capturing individuals' disclosure boundaries beyond general privacy norms.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CircuitReason-1k: Benchmarking Long-Horizon Visual-to-Symbolic Reasoning inElectrical Circuits

DGX agent

arXiv:2608.09374v1 Announce Type: new Abstract: Electrical circuit analysis requires more than recognizing components in an image. A solver must ground symbols and labels, recover latent topology, sel

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CliniCARE-Bench: Clinical Calibrated Audit of Medical Reasoning in EHR

DGX agent

arXiv:2608.07796v1 Announce Type: new Abstract: Large language models perform strongly on medical knowledge benchmarks, but reliable clinical deployment requires agents to conduct defensible investiga

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

DGX agent

arXiv:2608.07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primari

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Coarse-to-Fine Registration of Jawbone CT and Intraoral Scan Data Using GeDi and ICP with Pseudo-IOS Ground Truth

DGX agent

arXiv:2608.07564v1 Announce Type: cross Abstract: In digital dentistry and oral surgery, the registration of jawbone CT and intraoral scanner (IOS) data is essential for integrating internal bone stru

safetyarxiv-cs-ai
11 Aug 2026
Local Ai

ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

DGX agent

arXiv:2608.09732v1 Announce Type: cross Abstract: Agent skills are emerging as an important attack surface in LLM-based agent systems. Through an empirical study of existing skill scanners, we find th

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Coupons

DGX agent

arXiv:2608.09282v1 Announce Type: new Abstract: Real-world shopping often requires constructing a basket of complementary items rather than retrieving a single product. Such combo-shopping tasks arise

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

COMEX: A Composition-Grounded Benchmark and Learning Framework for Explainable Aesthetic Image Cropping

DGX agent

arXiv:2608.07570v1 Announce Type: cross Abstract: Explainable aesthetic image cropping requires not only localizing a visually pleasing crop but also explaining why it is preferred. Existing crop-and-

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Communication-efficient distributed hazard difference estimation for heterogeneous multi-site survival data

DGX agent

arXiv:2601.14609v2 Announce Type: replace-cross Abstract: Multi-site collaboration can power survival models that no single hospital could fit alone, but privacy rules and protected computing environm

researcharxiv-cs-ai
11 Aug 2026
Agents

Complete, Scalable, and Robust Prioritized Planning for Multi-Robot Ordered Storage and Retrieval at Maximum Capacity

DGX agent

arXiv:2608.07734v1 Announce Type: cross Abstract: Automated warehouses face a fundamental trade-off between maximizing storage density and achieving high retrieval throughput. While puzzle-based stora

agentsarxiv-cs-ai
11 Aug 2026
Tutorials

Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching

DGX agent

arXiv:2608.08135v1 Announce Type: cross Abstract: Cross-modality medical image translation can reduce the burden of multi-modal acquisitions, yet the field remains constrained by two coupled limitatio

tutorialsarxiv-cs-ai
11 Aug 2026
Agents

Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Scenario

DGX agent

arXiv:2608.08131v1 Announce Type: cross Abstract: In the fictional Order 66, catastrophe does not arise from a powerful command alone: a trusted population is preconditioned, a short directive activat

agentsarxiv-cs-ai
11 Aug 2026
Safety

Concept-Guided Spatial Regularization for World Models in Atari Pong

DGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Confusion-Geometry Rebalancing for Long-Tailed Adversarial Training

DGX agent

arXiv:2608.09688v1 Announce Type: cross Abstract: Adversarial training under long tailed distributions suffers from a dual imbalance: the class imbalance skews the training objective toward head class

safetyarxiv-cs-ai
11 Aug 2026
Applications

ConMem: Contribution-Aware Memory for Long-Horizon Manufacturing Inspection Logs

DGX agent

arXiv:2607.28126v2 Announce Type: replace Abstract: Long-horizon steel-equipment inspection requires reasoning over heterogeneous records accumulated across repeated inspection cycles. Existing retrie

applicationsarxiv-cs-ai
11 Aug 2026
Research

Constraining ontology mappings using metaphysical choices

DGX agent

arXiv:2608.08122v1 Announce Type: new Abstract: In this paper we discuss the foundations behind a novel methodology for the validation of semantic mappings between different data sources based upon di

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Contamination Means Overestimation? A Fine-Grained Empirical Study in Code Intelligence

DGX agent

arXiv:2506.02791v4 Announce Type: replace-cross Abstract: In recent years, code intelligence has gained increasing importance in the field of automated software engineering. Meanwhile, the widespread

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

DGX agent

arXiv:2608.09025v1 Announce Type: new Abstract: Financial agents can turn correct context into an unauthorized effect: a customer-facing commitment, trade, or deployed policy. We present SAGE-Fin, a f

safetyarxiv-cs-ai
11 Aug 2026
Safety

Contextual Value Alignment via Multilayer Combinatorial Fusion

DGX agent

arXiv:2608.07642v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF

safetyarxiv-cs-ai
11 Aug 2026
Safety

Control-Oriented Scenario Tree Construction through Reinforcement Learning

DGX agent

arXiv:2608.09335v1 Announce Type: new Abstract: Multistage stochastic model predictive control (MPC) handles uncertainty by optimizing over a scenario tree, a finite branching approximation of future

safetyarxiv-cs-ai
11 Aug 2026
Agents

Controlled Memory Interference in Continual LLM Agents

DGX agent

arXiv:2608.07622v1 Announce Type: new Abstract: Long-term memory enables AI agents to maintain continuity across sessions, personalize behavior, and evolve through accumulated experience. Yet memory e

agentsarxiv-cs-ai
11 Aug 2026
← Previous
1…1112131415…443
Next →