AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

Recursive Synthesis for Long-Horizon Terminal Tasks

DGX agent

arXiv:2608.05466v1 Announce Type: new Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because ea

model-releasesarxiv-cs-ai
7 Aug 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reducing belief in conspiracy theories as they unfold using large language models

DGX agent

arXiv:2608.06151v1 Announce Type: cross Abstract: The emergence of conspiracy theories in the wake of major events is a significant societal challenge. Here we test whether conversational dialogues wi

researcharxiv-cs-ai
7 Aug 2026
Research

Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning

DGX agent

arXiv:2608.05643v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by using additional inference compute, but wider sampling alone can suffer from diminishing returns: new rollou

researcharxiv-cs-ai
7 Aug 2026
Research

Relay, Don't Route: Adaptive Population Handoff for Cost-Efficient LLM-Driven Evolution

DGX agent

arXiv:2608.05651v1 Announce Type: cross Abstract: Large language model (LLM)-driven evolution has shown promise for program search and algorithm discovery, but relying on strong models throughout long

researcharxiv-cs-ai
7 Aug 2026
Safety

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

DGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

safetyarxiv-cs-ai
7 Aug 2026
Applications

Revisiting Black-Box Model Ownership Verification through Information Theory

DGX agent

arXiv:2409.06130v2 Announce Type: replace-cross Abstract: Modern machine learning models require substantial computational resources and data to train, making them valuable intellectual property. Mode

applicationsarxiv-cs-ai
7 Aug 2026
Model Releases

Runtime Observability for Heterogeneous Attention Memory

DGX agent

arXiv:2608.05863v1 Announce Type: new Abstract: Modern models no longer keep a plain KV cache: latent caches, learned sparse selectors and recurrent states each carry the model's memory in a different

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Signals via Time-Frequency Prior Exploitation

DGX agent

arXiv:2608.05669v1 Announce Type: cross Abstract: Divertor heat-flux analysis is essential for understanding plasma-wall interactions and protecting plasma-facing components in magnetic-confinement fu

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Schema-Guided Hierarchical Information Extraction and Semantic Evaluation Using Generative AI

DGX agent

arXiv:2608.06167v1 Announce Type: new Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by aut

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications

DGX agent

arXiv:2608.05439v1 Announce Type: new Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and fo

safetyarxiv-cs-ai
7 Aug 2026
Safety

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

DGX agent

arXiv:2608.05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reachi

safetyarxiv-cs-ai
7 Aug 2026
Agents

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

DGX agent

arXiv:2608.05245v1 Announce Type: new Abstract: Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-e

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents

DGX agent

arXiv:2608.05212v1 Announce Type: new Abstract: Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning error

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Seeing Is Not Deciding: Can Multimodal LLMs Act as Effective CEOs?

DGX agent

arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in executive business decisions, existing benchmarks are l

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Shapes from Examples: Foundations of Shape Learning in Recursive SHACL

DGX agent

arXiv:2607.27934v2 Announce Type: replace Abstract: SHACL shapes enable data graph validation, making automatic shape learning essential for knowledge graph applications. We investigate the well-known

researcharxiv-cs-ai
7 Aug 2026
Tutorials

Shaping Human-AI Interactions to Provide Improvement Pathways and Balance Competing Objectives

DGX agent

arXiv:2608.05710v1 Announce Type: new Abstract: When an AI system is deployed, the individuals who use and or are evaluated by it form beliefs about how the system operates and use those beliefs to st

tutorialsarxiv-cs-ai
7 Aug 2026
Research

Signal or Spurious Cue? A Randomized Audit of Survey-Country Metadata in LLM Social Inference

DGX agent

arXiv:2608.06085v1 Announce Type: new Abstract: Survey-country metadata can improve an LLM's forecast of an individual response when informative, yet the same cue may redirect the forecast when assign

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Simulator-Grounded Large Language Models for Industrial Causal Reasoning: Tool-Use, Structured Injection, and Plant-Portable Retrieval for Wastewater Treatment Decision Support

DGX agent

arXiv:2608.05151v1 Announce Type: cross Abstract: Wastewater operators need answers grounded in how their plant's variables interact and how fast effects propagate, not in generic pretraining text, wh

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation

DGX agent

arXiv:2608.05628v1 Announce Type: new Abstract: Although agent skills equip LLMs with reusable procedural knowledge, manual maintenance suffers from high costs, unscalability, and misalignment. Real-w

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation

DGX agent

arXiv:2608.05970v1 Announce Type: cross Abstract: Embodied visuomotor models, including Diffusion Policy (DP) and Vision-Language-Action (VLA) models, have demonstrated promising performance on roboti

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

DGX agent

arXiv:2608.05204v1 Announce Type: new Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, refere

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

DGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

DGX agent

arXiv:2608.05604v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly act as agents whose procedural knowledge is stored in reusable skill packages and loaded at inference time.

agentsarxiv-cs-ai
7 Aug 2026
Research

Small Foundation Models of Human Cognition and Behaviour

DGX agent

arXiv:2608.05224v1 Announce Type: new Abstract: Large language models fine-tuned on human behavioural data have emerged as general-purpose cognitive proxies, but the scale this requires, and whether t

researcharxiv-cs-ai
7 Aug 2026
Tutorials

Spectral Aliasing Pretext: A novel task for Self-Supervised fault diagnosis in rotating machinery

DGX agent

arXiv:2608.05705v1 Announce Type: cross Abstract: Deep learning is a new way for machinery fault diagnosis but requires extensive labeled data, a scarce resource in industrial settings. We propose Spe

tutorialsarxiv-cs-ai
7 Aug 2026
Research

Stability of Ranking-dependent Pair-wise Comparison Patterns in the Analytic Hierarchy Process

DGX agent

arXiv:2608.05958v1 Announce Type: new Abstract: The paper addresses several ranking-dependent decision support methods. Ordinal information on compared objects can be used to improve the quality of ex

researcharxiv-cs-ai
7 Aug 2026
Model Releases

StepReflect: Structured UI Transition Reflection for Mobile GUI Agents

DGX agent

arXiv:2608.05587v1 Announce Type: new Abstract: Autonomous mobile GUI agents require accurate action reflection for reliable long-horizon execution. Existing approaches rely on open-ended multimodal r

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to Replicate a Human Survey with Synthetic Data

DGX agent

arXiv:2603.00059v3 Announce Type: replace-cross Abstract: How well can AI-derived synthetic research data replicate the responses of human participants? An emerging literature has begun to engage with

model-releasesarxiv-cs-ai
7 Aug 2026
Tutorials

Stochasticity Is Not the Hard Part: Reduction and Complexity in Instructional Sequencing over Prerequisite DAGs

DGX agent

arXiv:2608.05455v1 Announce Type: new Abstract: When a student must learn concepts connected by prerequisite dependencies, when does the order of instruction matter, and what does it cost to find the

tutorialsarxiv-cs-ai
7 Aug 2026
Safety

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

DGX agent

arXiv:2608.05656v1 Announce Type: cross Abstract: Safety risks of AI are becoming increasingly evident in human interactions with AI technologies. The prominent approaches to evaluating these risks fa

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Subliminal Learning is Non-Semantic Distillation

DGX agent

arXiv:2608.05734v1 Announce Type: new Abstract: Subliminal Learning (SL) is a surprising type of generalization displayed by modern language models. It allows the transfer of a bias or behavior from a

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts

DGX agent

arXiv:2510.14538v3 Announce Type: replace Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural cons

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation

DGX agent

arXiv:2608.05785v1 Announce Type: cross Abstract: Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fund

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Temporal Bridges for Spatial Resolution: Enhancing Climate Data Super-Resolution with Bidirectional Alignment

DGX agent

arXiv:2608.05981v1 Announce Type: new Abstract: High-resolution climate data is crucial for meteorological predictions and for informing decision support across diverse domains. However, the acquisiti

safetyarxiv-cs-ai
7 Aug 2026
Safety

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

DGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

safetyarxiv-cs-ai
7 Aug 2026
Safety

The Closing Window: How Governments Could Lose Their Ability to Restrain Advanced AI

DGX agent

arXiv:2608.05173v1 Announce Type: cross Abstract: As AI capabilities advance, AI systems will pose greater risks to national security and potentially humanity as a whole. Governments may eventually co

safetyarxiv-cs-ai
7 Aug 2026
Research

The em-dash em-beds in Congress: A population-level rise in em-dash frequency in U.S. congressional press releases at the dawn of the large-language-model era, 2021-2025

DGX agent

arXiv:2608.05889v1 Announce Type: cross Abstract: Large language models (LLMs) can leave small stylistic traces in text written with their help. The most discussed is the em-dash (U+2014), especially

researcharxiv-cs-ai
7 Aug 2026
Research

The ethics of artificial intelligence in the life sciences: Universality, cultural diversity and an architecture of care

DGX agent

arXiv:2608.05436v1 Announce Type: cross Abstract: The life sciences and health research have started to benefit from artificial intelligence, which raises ethical concerns that are real but, we argue,

researcharxiv-cs-ai
7 Aug 2026
Model Releases

The Ignition Index: Measuring Global Workspace Dynamics in Language Models

DGX agent

arXiv:2608.05160v1 Announce Type: new Abstract: We introduce the Ignition Index (I), a validated scalar metric that operationalizes Global Workspace Theory's (GWT) all-or-none ignition prediction in t

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

DGX agent

arXiv:2608.06270v1 Announce Type: new Abstract: The 'thinking-with-images' paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations o

safetyarxiv-cs-ai
7 Aug 2026
Applications

The Judgment-Consequence Gap: LLM Moral Reasoning in Healthcare Decisions

DGX agent

arXiv:2608.05583v1 Announce Type: cross Abstract: As large language models (LLMs) enter high-stakes domains such as healthcare, understanding their moral reasoning becomes essential. Decisions about s

applicationsarxiv-cs-ai
7 Aug 2026
Model Releases

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

DGX agent

arXiv:2608.06361v1 Announce Type: new Abstract: Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure mo

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Toward Deployable Bangla Sign Language Recognition with Expert-Validated Data and a Lightweight Attention-Based Model

DGX agent

arXiv:2608.06252v1 Announce Type: cross Abstract: Deaf and hard-of-hearing people in Bangladesh communicate mainly through Bangla Sign Language (BdSL). Automatic BdSL recognition on personal devices c

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions

DGX agent

arXiv:2608.05975v1 Announce Type: cross Abstract: In this paper, we present TRACE (Tokenized Robust Attention for Contact-Aware Estimation), an end-to-end learned proprioceptive odometry estimator for

safetyarxiv-cs-ai
7 Aug 2026
Agents

Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering

DGX agent

arXiv:2608.06366v1 Announce Type: new Abstract: Electronic health record (EHR) feature engineering is a major bottleneck in clinical research and AI, accounting for 39-45% of data scientists' workload

agentsarxiv-cs-ai
7 Aug 2026
Safety

Training a Conditioned Video Game Agent on a VLM Annotated Dataset

DGX agent

arXiv:2608.05954v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the gam

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

DGX agent

arXiv:2608.06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging. Crit

model-releasesarxiv-cs-ai
7 Aug 2026
Research

TriQua: Reconciling Granularity and Context in Factuality Evaluation

DGX agent

arXiv:2608.05228v1 Announce Type: new Abstract: The 'decompose-then-verify' paradigm for LLM factuality evaluation faces a fundamental trade-off: atomic facts, i.e., one sentence conveying one unit of

researcharxiv-cs-ai
7 Aug 2026
← Previous
1…3132333435…443
Next →