AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

A Knowledge-Injection Framework for Zero-Shot Adaptation of LLMs to Delirium Prediction

DGX agent

arXiv:2607.20453v1 Announce Type: cross Abstract: Large language models show promise for clinical prediction, but zero-shot performance on specialized tasks is limited by incomplete domain knowledge,

model-releasesarxiv-cs-ai
24 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A New Well-Supported Semantics for Description Logic Programs

DGX agent

arXiv:2607.21203v1 Announce Type: new Abstract: Description logic programs are a powerful formalism for combining rules with ontologies. The well-supported semantics for description logic programs ens

researcharxiv-cs-ai
24 Jul 2026
Model Releases

A Sovereign, Open-Source Foundation Model for German and English

DGX agent

arXiv:2607.09424v3 Announce Type: replace-cross Abstract: We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervision in Video Face Swapping

DGX agent

arXiv:2607.21434v1 Announce Type: cross Abstract: Video face swapping has no natural paired supervision: no real footage exists of one person's face performing another person's video. The strongest cu

researcharxiv-cs-ai
24 Jul 2026
Model Releases

Adaptive Multi-Horizon Reinforcement Learning

DGX agent

arXiv:2607.20656v1 Announce Type: cross Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

DGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

DGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

DGX agent

arXiv:2607.21503v1 Announce Type: new Abstract: Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning co

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

AI Assistants Overassist

DGX agent

arXiv:2607.21306v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as tutors and thought partners, helping users reason through problems. While guidance from AI assis

model-releasesarxiv-cs-ai
24 Jul 2026
Research

AI-Driven Multi-Hop Relay Selection for Smart Urban NR-V2X Networks via Learning-to-Optimize Graph Neural Networks

DGX agent

arXiv:2607.20554v1 Announce Type: new Abstract: Reliable and low-latency NR-V2X communications are essential for smart mobility in dense urban environments. However, limited Road-Side Unit (RSU) densi

researcharxiv-cs-ai
24 Jul 2026
Agents

AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics

DGX agent

arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments. Thi

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

DGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

An LLM-Driven Workflow for Automated Process Control Strategy Generation and Tuning from Dynamic Process Models

DGX agent

arXiv:2607.21292v1 Announce Type: new Abstract: We present a structured large-language-model-driven workflow for automated multi-variable control design from dynamic process models. The workflow decom

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Animation, Verification and Visualisation of Prolog Transition Systems with ProB

DGX agent

arXiv:2607.21192v1 Announce Type: cross Abstract: ProB is a Prolog-based model checker, animator and constraint solver for high-level formal specifications. One can also use ProB to animate transition

researcharxiv-cs-ai
24 Jul 2026
Research

Answer-then-Edit: Reasoning Skeleton Editing for Anti-Distillation with Preserved Utility

DGX agent

arXiv:2607.20440v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) entail substantial intellectual and financial investment, making them valuable intellectual property (IP). Ho

researcharxiv-cs-ai
24 Jul 2026
Research

Anti-Goal Reasoning: Rethinking the Theory of Goal Reasoning in Non-Axiomatic Logic

DGX agent

arXiv:2607.20902v1 Announce Type: cross Abstract: Goal reasoning in Non-Axiomatic Logic (NAL) explains how an adaptive system derives means for realizing desired events under insufficient knowledge an

researcharxiv-cs-ai
24 Jul 2026
Safety

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation

DGX agent

arXiv:2607.18042v2 Announce Type: replace-cross Abstract: End-to-end vision-language navigation (VLN) with causal vision-language models maps instructions and egocentric observations directly to actio

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use

DGX agent

arXiv:2607.20536v1 Announce Type: new Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ArbiGraph: Arbitrarily Scalable Verifiable Task Graphs for Evaluating Context Management

DGX agent

arXiv:2607.20764v1 Announce Type: new Abstract: We introduce ARBIGRAPH, a benchmark generator for evaluating whether tool-assisted language agents can retain, update, compose, and discard task-relevan

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

DGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

safetyarxiv-cs-ai
24 Jul 2026
Research

Are Diversity Metrics Measuring Diversity? A Capability-Controlled Audit of Majority-Vote Gain in LLM Ensembles

DGX agent

arXiv:2607.20768v1 Announce Type: cross Abstract: Majority voting over LLMs is widely assumed to benefit from diversity, and diversity measures are used to choose which models to combine. We ask wheth

researcharxiv-cs-ai
24 Jul 2026
Model Releases

AREX: Towards a Recursively Self-Improving Agent for Deep Research

DGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

model-releasesarxiv-cs-ai
24 Jul 2026
Tutorials

Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure, and how to mitigate it

DGX agent

arXiv:2607.21498v1 Announce Type: cross Abstract: A rhetorical figure that Cicero and Quintilian catalogued two thousand years ago reappears, systematically, in the text of large language models: epan

tutorialsarxiv-cs-ai
24 Jul 2026
Applications

Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Series Forecasting Models

DGX agent

arXiv:2607.20493v1 Announce Type: new Abstract: Deep learning has led to remarkable progress in artificial intelligence, particularly in robotics, imaging and sound processing. However, a major limita

applicationsarxiv-cs-ai
24 Jul 2026
Model Releases

Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Intervention in Large Language Models

DGX agent

arXiv:2607.20524v1 Announce Type: new Abstract: Mean cross-positional attention degradation is widely reported in transformer interpretability, yet whether it causally limits contextual retrieval rema

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

DGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

safetyarxiv-cs-ai
24 Jul 2026
Local Ai

Auditing Evidence Use in Medical LLM Diagnosis

DGX agent

arXiv:2607.20848v1 Announce Type: new Abstract: Medical LLMs are often evaluated by whether they select the correct diagnosis, but diagnostic accuracy alone does not show whether the model used the ca

local-aiarxiv-cs-ai
24 Jul 2026
Local Ai

Auditing Provenance Sensitivity in LLM Agent Action Selection

DGX agent

arXiv:2607.20827v1 Announce Type: new Abstract: LLM agents choose tools and arguments from context that mixes user requests, tool outputs, retrieved records, memory, and untrusted text. Evidence can b

local-aiarxiv-cs-ai
24 Jul 2026
Model Releases

Autonomous disproofs of the sum-product conjecture over mathbb R with GPT-5.5 Pro

DGX agent

arXiv:2607.20525v1 Announce Type: new Abstract: OpenAI's recent disproof of the Erdos unit distance conjecture marked a milestone for AI in mathematics. It also inspired another breakthrough: a human

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants

DGX agent

arXiv:2607.20488v1 Announce Type: new Abstract: Multi-agent LLM frameworks typically fix their team topology at boot time. When an individual agent becomes overloaded at runtime, for example by mixing

model-releasesarxiv-cs-ai
24 Jul 2026
Applications

Backpropagation-Free Test-Time Adaptation for Lightweight EEG-Based Brain-Computer Interfaces

DGX agent

arXiv:2601.07556v2 Announce Type: replace-cross Abstract: Electroencephalogram (EEG)-based brain-computer interfaces (BCIs) face significant deployment challenges due to inter-subject variability, sig

applicationsarxiv-cs-ai
24 Jul 2026
Research

Barzilai-Borwein Fails Superlinear Convergence on an Open Set of Quadratics for Every Dimension ngeq 4

DGX agent

arXiv:2607.21579v1 Announce Type: cross Abstract: Barzilai--Borwein (BB) method has shown strong practical performance in continuous optimization, yet its convergence dynamics remains poorly understoo

researcharxiv-cs-ai
24 Jul 2026
Research

BasketEvent: Understanding Who Did What and When in Basketball Videos

DGX agent

arXiv:2607.21267v1 Announce Type: new Abstract: Comprehensive basketball video understanding requires resolving not only what event occurs, but also who is responsible and when the key evidence appear

researcharxiv-cs-ai
24 Jul 2026
Agents

Bayesian uncertainty estimation improves clinical decision making in medical AI agents

DGX agent

arXiv:2607.20582v1 Announce Type: cross Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases.

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

Benchmarking Large Language Models on Multi-Sensor Physical Hazard Assessment

DGX agent

arXiv:2607.20476v1 Announce Type: new Abstract: We present an empirical benchmark evaluating how five large language models assess multisensor physical hazard data. Testing 60 scenarios across three c

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Benchmarking the Personalization Capabilities of Large Language Models

DGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Benchmarking Unlearning for Vision Transformers

DGX agent

arXiv:2602.20114v2 Announce Type: replace-cross Abstract: Machine unlearning (MU) refers to the post-training capability to remove (the influence of) training examples that are incorrect, biased, or l

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Beyond Heavy Log Curation: Perplexity-Based APT Detection via Unsupervised, Context-Augmented Language Models

DGX agent

arXiv:2607.20832v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) remain difficult to detect because only a small fraction of events in large-scale logs are attack-related, and inve

researcharxiv-cs-ai
24 Jul 2026
Research

Beyond Independent Optimization: Compression, MoE Routing, and Quantization Interactions in Multimodal Edge Intelligence

DGX agent

arXiv:2607.20981v1 Announce Type: new Abstract: Efficient multimodal inference is increasingly constrained not only by model quality or FLOP count, but also by the cost of preserving, moving, routing,

researcharxiv-cs-ai
24 Jul 2026
Model Releases

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

DGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

model-releasesarxiv-cs-ai
24 Jul 2026
Tutorials

Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design

DGX agent

arXiv:2607.20550v1 Announce Type: cross Abstract: The traditional 'one drug, one target' paradigm of structure-based drug design (SBDD) frequently proves inadequate for treating multifactorial disease

tutorialsarxiv-cs-ai
24 Jul 2026
Applications

Beyond Sufficiency: Time Series Explanation with Counterfactual Necessity

DGX agent

arXiv:2607.21573v1 Announce Type: cross Abstract: Faithful explanations of time-series classifiers should identify subsequences that are not only sufficient to preserve a black-box model's prediction,

applicationsarxiv-cs-ai
24 Jul 2026
Safety

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

DGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

safetyarxiv-cs-ai
24 Jul 2026
Research

Bound-Founded Semantics for Answer Set Programming with Difference Constraints: Preliminary Report

DGX agent

arXiv:2607.21201v1 Announce Type: new Abstract: While the integration of linear constraints has significantly expanded the reach of Answer Set Programming (ASP), existing hybrid solvers often rely on

researcharxiv-cs-ai
24 Jul 2026
Model Releases

Break Through the Compression Bottleneck: From Theory to Practice

DGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for Dynamic Graph Systems

DGX agent

arXiv:2607.21421v1 Announce Type: new Abstract: Generative models can support decision-making under uncertainty by producing ensembles of plausible future system trajectories, but statistical plausibi

researcharxiv-cs-ai
24 Jul 2026
Model Releases

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

DGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows

DGX agent

arXiv:2607.17528v3 Announce Type: replace Abstract: Large language model (LLM) agents are extending electronic design automation (EDA) beyond static RTL generation toward long-horizon, tool-interactiv

agentsarxiv-cs-ai
24 Jul 2026
← Previous
1…7273747576…448
Next →