AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
7 Aug 2026

Spectral Aliasing Pretext: A novel task for Self-Supervised fault diagnosis in rotating machinery

TutorialsDGX agent

arXiv:2608.05705v1 Announce Type: cross Abstract: Deep learning is a new way for machinery fault diagnosis but requires extensive labeled data, a scarce resource in industrial settings. We propose Spe

Stability of Ranking-dependent Pair-wise Comparison Patterns in the Analytic Hierarchy Process

ResearchDGX agent

arXiv:2608.05958v1 Announce Type: new Abstract: The paper addresses several ranking-dependent decision support methods. Ordinal information on compared objects can be used to improve the quality of ex

StepReflect: Structured UI Transition Reflection for Mobile GUI Agents

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.05587v1 Announce Type: new Abstract: Autonomous mobile GUI agents require accurate action reflection for reliable long-horizon execution. Existing approaches rely on open-ended multimodal r

Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to Replicate a Human Survey with Synthetic Data

Model ReleasesDGX agent

arXiv:2603.00059v3 Announce Type: replace-cross Abstract: How well can AI-derived synthetic research data replicate the responses of human participants? An emerging literature has begun to engage with

Stochasticity Is Not the Hard Part: Reduction and Complexity in Instructional Sequencing over Prerequisite DAGs

TutorialsDGX agent

arXiv:2608.05455v1 Announce Type: new Abstract: When a student must learn concepts connected by prerequisite dependencies, when does the order of instruction matter, and what does it cost to find the

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

SafetyDGX agent

arXiv:2608.05656v1 Announce Type: cross Abstract: Safety risks of AI are becoming increasingly evident in human interactions with AI technologies. The prominent approaches to evaluating these risks fa

Subliminal Learning is Non-Semantic Distillation

Model ReleasesDGX agent

arXiv:2608.05734v1 Announce Type: new Abstract: Subliminal Learning (SL) is a surprising type of generalization displayed by modern language models. It allows the transfer of a bias or behavior from a

Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts

SafetyDGX agent

arXiv:2510.14538v3 Announce Type: replace Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural cons

Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation

Model ReleasesDGX agent

arXiv:2608.05785v1 Announce Type: cross Abstract: Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fund

Temporal Bridges for Spatial Resolution: Enhancing Climate Data Super-Resolution with Bidirectional Alignment

SafetyDGX agent

arXiv:2608.05981v1 Announce Type: new Abstract: High-resolution climate data is crucial for meteorological predictions and for informing decision support across diverse domains. However, the acquisiti

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

SafetyDGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

The Closing Window: How Governments Could Lose Their Ability to Restrain Advanced AI

SafetyDGX agent

arXiv:2608.05173v1 Announce Type: cross Abstract: As AI capabilities advance, AI systems will pose greater risks to national security and potentially humanity as a whole. Governments may eventually co

The em-dash em-beds in Congress: A population-level rise in em-dash frequency in U.S. congressional press releases at the dawn of the large-language-model era, 2021-2025

ResearchDGX agent

arXiv:2608.05889v1 Announce Type: cross Abstract: Large language models (LLMs) can leave small stylistic traces in text written with their help. The most discussed is the em-dash (U+2014), especially

The ethics of artificial intelligence in the life sciences: Universality, cultural diversity and an architecture of care

ResearchDGX agent

arXiv:2608.05436v1 Announce Type: cross Abstract: The life sciences and health research have started to benefit from artificial intelligence, which raises ethical concerns that are real but, we argue,

The Ignition Index: Measuring Global Workspace Dynamics in Language Models

Model ReleasesDGX agent

arXiv:2608.05160v1 Announce Type: new Abstract: We introduce the Ignition Index (I), a validated scalar metric that operationalizes Global Workspace Theory's (GWT) all-or-none ignition prediction in t

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

SafetyDGX agent

arXiv:2608.06270v1 Announce Type: new Abstract: The 'thinking-with-images' paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations o

The Judgment-Consequence Gap: LLM Moral Reasoning in Healthcare Decisions

ApplicationsDGX agent

arXiv:2608.05583v1 Announce Type: cross Abstract: As large language models (LLMs) enter high-stakes domains such as healthcare, understanding their moral reasoning becomes essential. Decisions about s

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

Model ReleasesDGX agent

arXiv:2608.06361v1 Announce Type: new Abstract: Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure mo

Toward Deployable Bangla Sign Language Recognition with Expert-Validated Data and a Lightweight Attention-Based Model

Model ReleasesDGX agent

arXiv:2608.06252v1 Announce Type: cross Abstract: Deaf and hard-of-hearing people in Bangladesh communicate mainly through Bangla Sign Language (BdSL). Automatic BdSL recognition on personal devices c

TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions

SafetyDGX agent

arXiv:2608.05975v1 Announce Type: cross Abstract: In this paper, we present TRACE (Tokenized Robust Attention for Contact-Aware Estimation), an end-to-end learned proprioceptive odometry estimator for

Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering

AgentsDGX agent

arXiv:2608.06366v1 Announce Type: new Abstract: Electronic health record (EHR) feature engineering is a major bottleneck in clinical research and AI, accounting for 39-45% of data scientists' workload

Training a Conditioned Video Game Agent on a VLM Annotated Dataset

SafetyDGX agent

arXiv:2608.05954v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the gam

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

Model ReleasesDGX agent

arXiv:2608.06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging. Crit

TriQua: Reconciling Granularity and Context in Factuality Evaluation

ResearchDGX agent

arXiv:2608.05228v1 Announce Type: new Abstract: The 'decompose-then-verify' paradigm for LLM factuality evaluation faces a fundamental trade-off: atomic facts, i.e., one sentence conveying one unit of

Trust-Based Incentive Mechanisms in Semi-Decentralized Federated Learning Systems

ResearchDGX agent

arXiv:2602.08290v2 Announce Type: replace-cross Abstract: In federated learning (FL), decentralized model training allows multi-ple participants to collaboratively improve a shared machine learning mo

TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized Views

Local AiDGX agent

arXiv:2607.21910v2 Announce Type: replace Abstract: World models let agents plan against predicted physical state, but that state drifts; re-observation is costly and delayed, and repair can fail. We

TS-RAG: Retrieval Augmented Generation for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2608.06223v1 Announce Type: new Abstract: While deep learning models, particularly transformer-based architectures, have shown impressive performance in time series forecasting, the application

Turing's Frist Imitation Game: Design Concepts and a Human-Approximates-Machine Reading

ResearchDGX agent

arXiv:2608.05558v1 Announce Type: cross Abstract: This paper examines Turing's 1948 report, 'Intelligent Machinery', as an important conceptual source for the later imitation games. Its first contribu

Tytan: Interactive Neurosymbolic Construction of Analytic Semantic Schemas from Relational Data

Model ReleasesDGX agent

arXiv:2608.06331v1 Announce Type: cross Abstract: From natural-language query interfaces to automated report generation, data analysis tools need a description of the data: the real-world entities it

Unified Agent: Managing Interactions across Devices

Model ReleasesDGX agent

arXiv:2608.05729v1 Announce Type: new Abstract: As capabilities rapidly increase, AI agents can move from running inside one app to acting across a user's devices over time. Yet existing agent systems

Universal Pathologies, Conditional Consequences: A Triple-Robustness Analysis of RAG for Multi-Hop Traceability

Model ReleasesDGX agent

arXiv:2608.05153v1 Announce Type: cross Abstract: GraphRAG underperforms vector RAG on citation precision in many reports, but where and why have remained corpus-bound. We present a triple-robustness

UniVVT: A Unified End-to-End Framework for High-Fidelity Video Virtual Try-on

SafetyDGX agent

arXiv:2608.05745v1 Announce Type: cross Abstract: Video Virtual Try-On (VVT) synthesizes a video of a person wearing a target garment while preserving identity, motion, and scene dynamics. Dominant ap

Vibe Compiler: A Research-Logic Synthesis Tool That Runs without Prompt Engineering -Toward Enhancing Metacognition for Sustaining Agency in the Age of Generative AI-

Model ReleasesDGX agent

arXiv:2608.05545v1 Announce Type: cross Abstract: Generative AI used as a capable servant has greatly accelerated intellectual work, but it also risks eroding human epistemic agency by encouraging unc

ViSR-KGC: Visual Subgraph Reasoning with Vision-Language Models for Multimodal Knowledge Graph Completion

ResearchDGX agent

arXiv:2608.05833v1 Announce Type: new Abstract: Knowledge graph completion (KGC) aims to infer missing entities or relations from incomplete graph structures, and has evolved into multimodal knowledge

Visual Grounding in Zero-Shot Vision-Language Control

SafetyDGX agent

arXiv:2608.06154v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as zero-shot controllers, but successful trajectories do not necessarily show that decisions are g

VLMs for Videogame Data Annotation

ApplicationsDGX agent

arXiv:2608.05949v1 Announce Type: new Abstract: Vision Language Models (VLMs) and Artificial Intelligence (AI) agents have revolutionized how engineers approach complex problems in real-world applicat

What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)

Model ReleasesDGX agent

arXiv:2608.06202v1 Announce Type: cross Abstract: Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet m

When Agentic AI Meets Integrated Sensing and Communication

AgentsDGX agent

arXiv:2608.05792v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is transforming Integrated Sensing and Communication (ISAC) from a function-oriented physical-layer technology into

When Do Prompt-Side Agent Playbooks Transfer? Accuracy, Cost, and Runtime Shift in Agent Deployment

AgentsDGX agent

arXiv:2608.05778v1 Announce Type: new Abstract: Prompt-side playbooks can improve tool-using language agents without retraining, but their portability beyond the source setting is unclear. We study fr

When Drafts Evolve: Speculative Decoding Meets Online Learning

ResearchDGX agent

arXiv:2603.12617v2 Announce Type: replace-cross Abstract: Speculative decoding has emerged as a widely adopted paradigm for accelerating large language model inference, where a lightweight draft model

When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

AgentsDGX agent

arXiv:2608.05563v1 Announce Type: cross Abstract: Self-evolving skill (SES) systems distill agent trajectories into persistent skills, allowing untrusted experience to become trusted instruction. We i

When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories

Model ReleasesDGX agent

arXiv:2608.06057v1 Announce Type: new Abstract: Tool-calling agents infer task state from accumulated dialogue and tool traces. In persistent interactions, however, historical traces may remain struct

When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents

SafetyDGX agent

arXiv:2608.05219v1 Announce Type: new Abstract: Privileged on-policy distillation provides dense supervision for multi-turn agents by allowing a synchronized teacher to re-score the student's response

When Self-Evolution Backfires: Pre-Commit Gating against Skill Contamination in LLM Agents

Model ReleasesDGX agent

arXiv:2608.05810v1 Announce Type: new Abstract: Self-evolving agents accumulate capability by distilling reusable skills from their execution trajectories, but we find this process is not monotonic: p

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

SafetyDGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

Why the Third Axis Is Freedom

ResearchDGX agent

arXiv:2608.05423v1 Announce Type: cross Abstract: In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that p

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

Local AiDGX agent

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localiz

WorldClaw: Agentic 3D Open-World Generation at Scale

AgentsDGX agent

arXiv:2608.05248v1 Announce Type: new Abstract: Generating large-scale, freely explorable 3D worlds from open-ended text remains challenging because a system must jointly maintain global spatial coher

6 Aug 2026

A 6G Integrated Sensing and Communication Framework for Railway Intrusion Detection and Collision Prediction

ResearchDGX agent

arXiv:2608.04710v1 Announce Type: cross Abstract: Integrated Sensing and Communication (ISAC) combines sensing and communication to efficiently utilize wireless resources and is emerging as a key para

A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Audits in AI

SafetyDGX agent

arXiv:2608.04921v1 Announce Type: cross Abstract: As AI systems become increasingly integrated into diverse interfaces and applications, model-centric audits are insufficient to address risks arising

A General Sufficient Condition for Rewriting Horn-ALCHI Atomic Queries into GQL

ResearchDGX agent

arXiv:2608.04945v1 Announce Type: cross Abstract: The emergence of the ISO standard GQL introduces a powerful query language extending first-order logic with controlled recursion, raising the question

A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age Score (AAS)

ResearchDGX agent

arXiv:2608.04012v1 Announce Type: new Abstract: Artificial intelligence systems are increasingly expected to operate over repeated cycles of interaction, adaptation, and update rather than through iso

A Model Merging Approach for Continual MLLM Unlearning

ResearchDGX agent

arXiv:2608.04548v1 Announce Type: cross Abstract: Multimodal large language model (MLLM) unlearning methods have been proposed to remove private, sensitive, or proprietary information from well-traine

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

SafetyDGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2602.06052v4 Announce Type: replace-cross Abstract: Research in artificial intelligence is shifting from model innovations and benchmark scores towards problem definition and rigorous real-world

A Trust-region Framework for Moment Estimation

ResearchDGX agent

arXiv:2608.04026v1 Announce Type: cross Abstract: In this paper, we develop a trust-region framework for understanding the behavior of adaptive moment estimation mechanisms, such as extsc{Adam}, in st

A Unified Model for Cross-Domain Clone Detection via Model Merging

Model ReleasesDGX agent

arXiv:2608.04215v1 Announce Type: cross Abstract: The growing diversity of code clone types, from syntactic copies to cross-language semantic clones to AI-generated duplicates, has created a fragmenta

A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing

SafetyDGX agent

arXiv:2608.04625v1 Announce Type: new Abstract: Industrial recommendation strategy iteration heavily relies on large-scale A/B experimentation. Traditional tuning requires experts to repeatedly design

ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment

ResearchDGX agent

arXiv:2608.05102v1 Announce Type: new Abstract: Long-horizon search agents must make multiple sequential actions (steps) to search, retrieve, verify, and integrate evidence to reach a final answer. Ho

Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

Model ReleasesDGX agent

arXiv:2608.04682v1 Announce Type: cross Abstract: Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific b

← Previous
1…2526272829…354
Next →