AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

A Hippocampus for Linear Attention: An Exact Memory for What the Recurrent State Forgets

DGX agent

arXiv:2607.02303v1 Announce Type: new Abstract: Linear-attention and state-space language models compress the prefix into a fixed-size recurrent state, yielding O(1) memory at the cost of a lossy exac

researcharxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A Multi-Branch Hierarchy-Aware Framework for Heterogeneous Audio Classification

DGX agent

arXiv:2607.01974v1 Announce Type: cross Abstract: This technical report describes our system for Task 1 of the DCASE 2026 Challenge, which aims to classify heterogeneous audio recordings according to

researcharxiv-cs-ai
3 Jul 2026
Agents

A Practice Auditing Framework for Large Language Model Use: Collective Empiricism, Pseudo-Rational Cognition, and Governance of AI-Generated Content

DGX agent

arXiv:2607.01248v1 Announce Type: cross Abstract: Large language models are increasingly used for knowledge acquisition, code generation, academic writing, and agent-based automation. In these setting

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

A rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks

DGX agent

arXiv:2607.02175v1 Announce Type: new Abstract: Multiple-choice medical benchmarks are increasingly saturated, and recent rubric-based evaluations such as HealthBench have shown that open-ended clinic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A-TMA: Decoupling State-Aware Memory Failures in Long-Term Agent Memory

DGX agent

arXiv:2607.01935v1 Announce Type: new Abstract: Long term memory lets LLM agents act as persistent assistants, but user facts change. A useful memory system must know what is true now, what used to be

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A^{2}utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT Construction

DGX agent

arXiv:2607.02141v1 Announce Type: new Abstract: Most LP-from-text benchmarks are static datasets of word problems written and labeled by hand. Once such a dataset is released, its size is fixed, its d

model-releasesarxiv-cs-ai
3 Jul 2026
Research

ACID: Action Consistency via Inverse Dynamics for Planning with World Models

DGX agent

arXiv:2607.02403v1 Announce Type: cross Abstract: Decision-time planning with action-conditioned world models has become a popular paradigm for embodied control. However, the standard planning cost ju

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence

DGX agent

arXiv:2604.08169v2 Announce Type: replace Abstract: Alignment in LLMs is more brittle than commonly assumed: misalignment can be induced by adversarial prompts, benign fine-tuning, emergent misalignme

model-releasesarxiv-cs-ai
3 Jul 2026
Research

Actual causality in fault trees

DGX agent

arXiv:2607.01840v1 Announce Type: new Abstract: Fault trees are a widely used as effective risk models for complex systems, answering the question 'what can go wrong?', especially through minimal cut

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Adaptive Batch Sizes Using Non-Euclidean Gradient Noise Scales for Stochastic Sign and Spectral Descent

DGX agent

arXiv:2602.03001v2 Announce Type: replace-cross Abstract: To maximize hardware utilization, modern machine learning systems typically employ large constant or manually tuned batch size schedules, rely

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Adaptive Companionship for Group-Following Robots: Handling Dynamically Changing Group Formations

DGX agent

arXiv:2607.01287v1 Announce Type: cross Abstract: Accompanying a group of humans is an essential aspect of developing human-like social cognition in robots. However, human groups typically do not foll

safetyarxiv-cs-ai
3 Jul 2026
Research

Adaptive Contracts for Cost-Effective AI Delegation

DGX agent

arXiv:2603.17212v2 Announce Type: replace-cross Abstract: When organizations delegate text generation tasks to AI providers via pay-for-performance contracts, expected payments rise when evaluation is

researcharxiv-cs-ai
3 Jul 2026
Research

ADMC: Attention-based Diffusion Model for Missing Modalities Feature Completion

DGX agent

arXiv:2507.05624v2 Announce Type: replace Abstract: Multimodal emotion and intent recognition is essential for automated human-computer interaction, It aims to analyze users' speech, text, and visual

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Adoption and Impact of Command-Line AI Coding Agents: A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

DGX agent

arXiv:2607.01418v1 Announce Type: cross Abstract: Organizations rolling out agentic command line tools like Anthropic's Claude Code and GitHub's Copilot CLI need to know who will try them, who will ke

model-releasesarxiv-cs-ai
3 Jul 2026
Tutorials

ADVENT: LLM-Driven Automatic Predicate Invention for ILP

DGX agent

arXiv:2607.01585v1 Announce Type: cross Abstract: Predicate invention (PI), the creation of new predicates to extend the hypothesis space, remains a critical bottleneck in Inductive Logic Programming

tutorialsarxiv-cs-ai
3 Jul 2026
Model Releases

Agent4cs: A Multi-agent System for Code Summarization in Large Hierarchical Codebases

DGX agent

arXiv:2607.01425v1 Announce Type: new Abstract: Understanding large, complex codebases, especially those with obfuscated structures and incomplete documentation, remains a significant challenge. Exist

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AgenticDataBench: A Comprehensive Benchmark for Data Agents

DGX agent

arXiv:2607.01647v1 Announce Type: cross Abstract: Data science aims to derive actionable insights from heterogeneous raw data, unlocking the value of the massive amounts of data generated in modern so

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents

DGX agent

arXiv:2607.02255v1 Announce Type: new Abstract: Memory for a long-horizon LLM agent is a contract about what each future decision is allowed to see. The simplest contract appends past observations, to

model-releasesarxiv-cs-ai
3 Jul 2026
Applications

AI Assistance for Human Review of Default Judgments

DGX agent

arXiv:2607.01256v1 Announce Type: cross Abstract: Overwhelmed courts in the United States review millions of default judgments each year. Unfortunately, such manual reviews are time-consuming and pron

applicationsarxiv-cs-ai
3 Jul 2026
Hardware

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

DGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

hardwarearxiv-cs-ai
3 Jul 2026
Research

AI Virtue: What is 'Good' Knowledge in the Age of Artificial Intelligence?

DGX agent

arXiv:2607.01776v1 Announce Type: cross Abstract: In the age of AI, what will be good knowledge? This article, which is accepted and forthcoming in a special issue of Modern Fiction Studies on 'Cultur

researcharxiv-cs-ai
3 Jul 2026
Model Releases

AIriskEval-edu: New Dataset for Risk Assessment in AI-mediated K-12 Educational Explanations

DGX agent

arXiv:2607.01934v1 Announce Type: cross Abstract: This work introduces AIriskEval-edu-db2, a new dataset designed to train and evaluate auditors based on LLMs for an explainable pedagogical risk asses

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Algebraic Model Counting for Global Analysis of Optimal Decision Trees

DGX agent

arXiv:2607.02069v1 Announce Type: new Abstract: Ensuring model reliability in Explainable AI requires a global assessment of the hypothesis space. We propose a formal framework for the exhaustive anal

safetyarxiv-cs-ai
3 Jul 2026
Hardware

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

DGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

hardwarearxiv-cs-ai
3 Jul 2026
Research

An Exploratory Study on LLM-Generated Code and Comments in Code Repositories

DGX agent

arXiv:2607.01867v1 Announce Type: cross Abstract: The use of LLMs in software development has become increasingly widespread on tasks such as code generation and summarization. Reports from large tech

researcharxiv-cs-ai
3 Jul 2026
Model Releases

An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms

DGX agent

arXiv:2603.29466v2 Announce Type: replace-cross Abstract: Existing methods for quantifying predictive uncertainty in neural networks are either computationally intractable for large language models or

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

DGX agent

arXiv:2607.02269v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated immense promise in Spatio-Temporal Video Grounding (STVG). However, current evaluation protocols are l

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Aria: An Agent For Retrieval and Iterative Auto-Formalization via Dependency Graph

DGX agent

arXiv:2510.04520v2 Announce Type: replace Abstract: Accurate auto-formalization of theorem statements is essential for advancing automated discovery and verification of research-level mathematics, yet

agentsarxiv-cs-ai
3 Jul 2026
Safety

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning

DGX agent

arXiv:2607.02137v1 Announce Type: cross Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Uniform and ha

safetyarxiv-cs-ai
3 Jul 2026
Research

Artificial Intelligence-Enabled Accounting Information Systems and Fraud Detection in Nigeria's Financial Services Sector: The Moderating Role of Natural Language Processing

DGX agent

arXiv:2607.01257v1 Announce Type: cross Abstract: The rapid digitalisation of financial systems has improved operational efficiency and financial inclusion while simultaneously increasing exposure to

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Assessing VLM Reliability for Medical Image Quality Evaluation Under Corruption and Bias

DGX agent

arXiv:2607.01973v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied in medical tasks such as pathology description, report generation, and visual question answerin

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution

DGX agent

arXiv:2607.01942v1 Announce Type: new Abstract: LLM-based agents have shown strong potential for solving complex multi-step tasks, yet existing performance improvements often rely on either scaling to

agentsarxiv-cs-ai
3 Jul 2026
Local Ai

Auto-FL-Research: Agentic Search for Federated Learning Algorithms

DGX agent

arXiv:2607.01366v1 Announce Type: new Abstract: Federated learning (FL) research often depends on many small but consequential algorithmic choices: optimizer variants, server aggregation rules, local

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Automated grading of Linux/bash examinations using large language models: a four-level cognitive taxonomy approach

DGX agent

arXiv:2607.02432v1 Announce Type: new Abstract: Scalable and reliable grading of command-line examinations remains a challenge in computing education, where rising enrolments make manual marking diffi

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Autonomous discovery of traffic laws with AI traffic scientists

DGX agent

arXiv:2607.01639v1 Announce Type: new Abstract: Universal traffic laws describe recurrent patterns in congestion, mobility and driving behavior across cities, providing a scientific basis for transpor

agentsarxiv-cs-ai
3 Jul 2026
Safety

Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring

DGX agent

arXiv:2607.02121v1 Announce Type: cross Abstract: As Large Language Models (LLMs) and agentic systems become integrated into real-world applications, ensuring their safety and security is critical. Gu

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

Benchmarking Federated Learning and Knowledge Distillation for Point Cloud Classification

DGX agent

arXiv:2607.01272v1 Announce Type: cross Abstract: Deploying 3D point cloud analysis in privacy-sensitive, resource-constrained settings faces two barriers: data cannot be centralized, and models must

model-releasesarxiv-cs-ai
3 Jul 2026
Research

Beyond Adam: SOAP and Muon for Faster, Label-Efficient Training of Machine Learning Interatomic Potentials

DGX agent

arXiv:2607.02499v1 Announce Type: cross Abstract: Machine learning interatomic potentials (MLIPs) have become a hallmark of AI for scientific simulation. While efforts on new architectures and dataset

researcharxiv-cs-ai
3 Jul 2026
Safety

Beyond Detection: Redesigning Assessment and Governande of Generative AI at the Universidad Politecnica de Madrid (UPM)

DGX agent

arXiv:2607.01255v1 Announce Type: cross Abstract: Universities have responded to generative artificial intelligence (GenAI) in noticeably different ways, both internationally and within Spain. So far,

safetyarxiv-cs-ai
3 Jul 2026
Research

Beyond Gradient-Based Attacks: Adversarial Robustness and Explainability Stability in Cybersecurity Classifiers

DGX agent

arXiv:2607.01679v1 Announce Type: cross Abstract: Adversarial attacks on cybersecurity classifiers pose a dual threat: degrading predictions and destabilising the SHAP-based explanations that security

researcharxiv-cs-ai
3 Jul 2026
Safety

Beyond Next-Token Prediction: An RLVR Proof of Concept for Tool-Use Agents on Atlassian Workflows

DGX agent

arXiv:2607.01465v1 Announce Type: new Abstract: Large language models are trained to predict the next token, not to act inside a specific API. In niche enterprise SaaS workflows -- where success means

safetyarxiv-cs-ai
3 Jul 2026
Research

Beyond the Performance Illusion: Structure-Aware Stratified Partitioning and Curriculum Distributionally Robust Optimization for Spatially Correlated Domains

DGX agent

arXiv:2607.02055v1 Announce Type: cross Abstract: Performance evaluation in AI systems commonly assumes that random dataset splits produce independent and identically distributed (i.i.d.) subsets. We

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Black-Box Inference of LLM Architectural Properties with Restrictive API Access

DGX agent

arXiv:2607.01313v1 Announce Type: cross Abstract: In practice, most commercial LLM providers do not publicly release details of underlying LLM architectures. However, prior work has shown that given l

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in LLM Alignment

DGX agent

arXiv:2607.01239v1 Announce Type: cross Abstract: Character-level perturbations bypass safety alignment in modern LLMs despite leaving prompts human-readable. We identify and test a central structural

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

BRIDGE: Predicting Human Task Completion Time From Model Performance

DGX agent

arXiv:2602.07267v2 Announce Type: replace Abstract: Evaluating the real-world capabilities of AI systems requires grounding benchmark performance in human-interpretable measures of task difficulty. Ex

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

BuilderBench: The Building Blocks of Intelligent Agents

DGX agent

arXiv:2510.06288v4 Announce Type: replace Abstract: Today's AI models learn primarily through mimicry and refining, so it is not surprising that they struggle to solve problems beyond the limits set b

model-releasesarxiv-cs-ai
3 Jul 2026
Research

CamoNAS: Neural Architecture Search for Enhanced Camouflaged Object Detection

DGX agent

arXiv:2607.01870v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to locate and segment objects that blend into their surroundings, presenting challenges due to weak edge cues an

researcharxiv-cs-ai
3 Jul 2026
Safety

CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

DGX agent

arXiv:2603.22435v2 Announce Type: replace-cross Abstract: 'Code-as-Policy' considers how executable code can complement data-intensive Vision-Language-Action (VLA) methods, yet their effectiveness as

safetyarxiv-cs-ai
3 Jul 2026
← Previous
1…117118119120121…448
Next →