AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

SoftmaxGRPO: Learning to Reason using Softmax Advantage Group Estimation

DGX agent

arXiv:2608.09271v1 Announce Type: cross Abstract: Group-based reinforcement learning objectives such as GRPO can allocate learning signal poorly across prompt difficulty: under binary rewards, group n

researcharxiv-cs-ai
11 Aug 2026
Safety

Software Engineering for and with GUI Agent

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2608.09278v1 Announce Type: cross Abstract: GUI agents have advanced rapidly, producing a growing body of frameworks, benchmarks, and applications. However, this growth has outpaced the maturity

safetyarxiv-cs-ai
11 Aug 2026
Research

SPA: A Simple but Tough-to-Beat Baseline for Knowledge Injection

DGX agent

arXiv:2603.22213v2 Announce Type: replace-cross Abstract: While large language models (LLMs) are pretrained on massive amounts of data, their knowledge coverage remains incomplete in specialized, data

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Spectral Outliers Reveal Dominant Learned Structure in Transformer Attention

DGX agent

arXiv:2608.07921v1 Announce Type: cross Abstract: We apply Marchenko-Pastur (MP) random matrix theory to pre-trained attention weights in order to separate each projection matrix into a random-like bu

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

DGX agent

arXiv:2608.09138v1 Announce Type: cross Abstract: While learned robotic policies hold promise for advancing generalizable manipulation, their practical deployment is often hindered by suboptimal execu

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models

DGX agent

arXiv:2608.07712v1 Announce Type: cross Abstract: A predictive model receives a self-supervised signal whenever the consequence of an action is observed. Using that signal after deployment is difficul

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation

DGX agent

arXiv:2601.09974v2 Announce Type: replace Abstract: Personalizing Large Language Models typically relies on static retrieval or one-time adaptation, assuming user preferences remain invariant over tim

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

SR-OPSD: Self-Referenced On-Policy Self-Distillation

DGX agent

arXiv:2608.09745v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, provi

safetyarxiv-cs-ai
11 Aug 2026
Agents

STAIR: Effective Incident Response Using an End-to-End Agentic Planning Framework

DGX agent

arXiv:2608.09524v1 Announce Type: cross Abstract: Incident response planning is critical for restoring compromised software systems after cyberattacks. Common practice relies on expert-driven playbook

agentsarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Agents

STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in LLMs

DGX agent

arXiv:2608.08164v1 Announce Type: cross Abstract: Knowledge Distillation is a widely adopted technique in the training and fine-tuning of large language models (LLMs) enabling transfer of structured i

agentsarxiv-cs-ai
11 Aug 2026
Research

Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization

DGX agent

arXiv:2307.10053v5 Announce Type: replace-cross Abstract: In this paper, we focus on providing convergence guarantees for stochastic subgradient methods in minimizing nonsmooth nonconvex functions. We

researcharxiv-cs-ai
11 Aug 2026
Safety

StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoning

DGX agent

arXiv:2608.08326v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective approach for improving multimodal reasoning. However, most existing me

safetyarxiv-cs-ai
11 Aug 2026
Local Ai

Structure-Enhanced Features and Quality-Aware Dynamic Anchor Scoring for Robust Lane Detection

DGX agent

arXiv:2608.09610v1 Announce Type: cross Abstract: Lane detection requires recovering thin, elongated, and frequently occluded lane structures under challenging driving conditions. While anchor-based d

local-aiarxiv-cs-ai
11 Aug 2026
Research

Structure-Preserving Uncertainty Propagation in First-Order Proof Search

DGX agent

arXiv:2608.09190v1 Announce Type: new Abstract: GK is a query-directed first-order prover that extends ordinary resolution-based proof search with explicit positive and negative claims, numerical conf

researcharxiv-cs-ai
11 Aug 2026
Model Releases

SUM-AgriVLN: Spatial Understanding Memory for Agricultural Vision-and-Language Navigation

DGX agent

arXiv:2510.14357v2 Announce Type: replace-cross Abstract: Agricultural robots are emerging as powerful assistants across a wide range of agricultural tasks, nevertheless, they are still heavily relyin

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SuperCoder: Assembly Program Superoptimization with Large Language Models

DGX agent

arXiv:2505.11480v4 Announce Type: replace-cross Abstract: Superoptimization is the task of transforming a program into a faster one, and ideally the very fastest possible one, while preserving its inp

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SuperLocalMemory 4.0: The Governed Memory Operating System for AI Agents

DGX agent

arXiv:2608.08253v1 Announce Type: new Abstract: AI agents are becoming shared infrastructure, yet durable memory is commonly assembled from separate retrieval, governance, and operational components.

model-releasesarxiv-cs-ai
11 Aug 2026
Research

SuperNeuroMAT: An Efficient Matrix-based Simulator for Spiking Neural Networks

DGX agent

arXiv:2608.08479v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer a promising pathway to energy-efficient AI and brain-inspired computing. However, their widespread adoption is hi

researcharxiv-cs-ai
11 Aug 2026
Agents

SurgLAT: Surgical Latent Attention Tracking for Depth-Aware Robotic Laparoscope Control

DGX agent

arXiv:2608.07876v1 Announce Type: new Abstract: Autonomous laparoscopic camera control requires continuous understanding of the surgeon's operative intent in dynamic surgical scenes, where the target

agentsarxiv-cs-ai
11 Aug 2026
Safety

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

DGX agent

arXiv:2608.00143v2 Announce Type: replace-cross Abstract: Automated attack chain generation is critical for modern cybersecurity, yet manual construction fails to scale as adversary behaviors expand.

safetyarxiv-cs-ai
11 Aug 2026
Research

SymboUQ: Symbolic Uncertainty Quantification for Spatial Reasoning in LLMs

DGX agent

arXiv:2608.00417v2 Announce Type: replace Abstract: Although large language models (LLMs) can produce fluent spatial reasoning traces, their intermediate relations may fail to support the final conclu

researcharxiv-cs-ai
11 Aug 2026
Local Ai

SymDiag: Explainable Diagnosis for LLM Reasoning via Neuro-Symbolic Verification

DGX agent

arXiv:2608.08786v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as data-driven reasoners, yet their chains-of-thought (CoT) can be unfaithful even when final answers ar

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

Tabular Numeric Stretch Transformation

DGX agent

arXiv:2608.09162v1 Announce Type: cross Abstract: Tabular data presents unique challenges for deep learning due to its heterogeneous nature, where numeric features exhibit diverse distributions, scale

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification

DGX agent

arXiv:2608.08195v1 Announce Type: cross Abstract: Large language models (LLMs) are high-value assets that can be derived through redeployment, fine-tuning, quantization, or further alignment. Because

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability

DGX agent

arXiv:2608.09538v1 Announce Type: cross Abstract: We introduce TCS-Bench, a benchmark for evaluating Large Language Models (LLMs) on research-level Theoretical Computer Science (TCS) proof generation.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?

DGX agent

arXiv:2608.07899v1 Announce Type: new Abstract: Agent systems increasingly expose execution traces, yet telemetry that reveals a failure may still be inadequate for identifying where that failure orig

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Temporal Generalization in fNIRS-Based Autism Classification: A Cross-Time-Window Transfer Benchmark

DGX agent

arXiv:2608.07567v1 Announce Type: cross Abstract: Functional near-infrared spectroscopy (fNIRS) is a promising modality for autism spectrum disorder (ASD) classification, yet existing approaches assum

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Temporal Misgrounding in Legal RAG: A Versioned-Corpus Benchmark for French Tax Law

DGX agent

arXiv:2608.09393v1 Announce Type: cross Abstract: We identify and quantify temporal misgrounding: the systematic retrieval and citation of the currently in-force version of a legal article when the ap

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework

DGX agent

arXiv:2601.21747v4 Announce Type: replace-cross Abstract: Sepsis remains one of the most complex and heterogeneous syndromes in intensive care. While deep learning models achieve competitive performan

researcharxiv-cs-ai
11 Aug 2026
Research

Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched Compute

DGX agent

arXiv:2608.09351v1 Announce Type: cross Abstract: Test-time scaling improves LLM accuracy but multiplies inference cost, making the accuracy gained per unit of compute the metric that matters in deplo

researcharxiv-cs-ai
11 Aug 2026
Model Releases

TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Source Repair

DGX agent

arXiv:2608.07617v1 Announce Type: new Abstract: Scientific and technical writing depends on markup sources that must compile: LaTeX, Typst, and Markdown pipelines fail on missing delimiters, mismatche

model-releasesarxiv-cs-ai
11 Aug 2026
Research

TGIF: Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs

DGX agent

arXiv:2601.03100v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) typically rely on a single late-layer feature from a frozen vision encoder, leaving the encoder's ric

researcharxiv-cs-ai
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Research

The Announcement Carries the Cue: Markup, Boundaries, and the Notation of Pre-Training Corpora

DGX agent

arXiv:2608.09093v1 Announce Type: cross Abstract: How a document's arrangement is written down, its notation, is a training variable that no dataset card records. The field has established that text-e

researcharxiv-cs-ai
11 Aug 2026
Model Releases

The Authority Expectancy Effect in Multi-User Conflict

DGX agent

arXiv:2608.08026v1 Announce Type: new Abstract: We investigate how social authority (SA) signals interact with severity-based prioritization in large language models, operationalizing each axis as a m

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

The Belief-Desire-Intention Ontology for modelling mental reality and agency

DGX agent

arXiv:2511.17162v2 Announce Type: replace Abstract: The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, it

agentsarxiv-cs-ai
11 Aug 2026
Agents

The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in the AI Era

DGX agent

arXiv:2608.07779v1 Announce Type: new Abstract: Artificial intelligence is changing the task composition of computing work faster than curricula and training typically adapt. This is a curriculum-fram

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

The Cell Must Go On: Agar.io for Continual Reinforcement Learning

DGX agent

arXiv:2505.18347v3 Announce Type: replace-cross Abstract: Continual reinforcement learning (RL) concerns agents that are expected to learn continually, rather than converge to a policy that is then fi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation

DGX agent

arXiv:2511.02687v2 Announce Type: replace Abstract: The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independentl

model-releasesarxiv-cs-ai
11 Aug 2026
Research

The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes

DGX agent

arXiv:2608.07566v1 Announce Type: new Abstract: We introduce a continuous metric field framework trained by a single causal contrastive loss. The framework encodes a scene into coefficients of a fixed

researcharxiv-cs-ai
11 Aug 2026
Model Releases

The Knowing-Saying Gap: When Probes See Errors that Confidence Misses

DGX agent

arXiv:2608.07528v1 Announce Type: new Abstract: Linear probes detect corrupted context in language models with near-perfect accuracy, yet this does not translate into reliable failure prediction. The

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hierarchical Games

DGX agent

arXiv:2608.09574v1 Announce Type: new Abstract: LLMs are rapidly embedding themselves into daily life: drafting our emails, managing our schedules, and making decisions on our behalf. As they move fro

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

DGX agent

arXiv:2608.08654v1 Announce Type: new Abstract: How much an AI coding agent costs to run can depend more on the agent scaffolding that drives it than on the interface through which it reaches its tool

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

The Scaling Paradox in Human-AI Collaboration

DGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

DGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

safetyarxiv-cs-ai
11 Aug 2026
Research

The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection

DGX agent

arXiv:2606.23335v2 Announce Type: replace-cross Abstract: Provenance watermarking is increasingly treated as a safeguard for synthetic speech, whether built directly into speech-generation models such

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Theory-Guided Deception Detection: A RAG-Based Artificial Intelligence Exploration

DGX agent

arXiv:2608.08881v1 Announce Type: new Abstract: The current work developed seven Retrieval-Augmented Generation (RAG) models based on leading deception theories and compared how deception judgments we

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…2021222324…443
Next →