AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
3 Jul 2026

Repair the Amplifier, Not the Symptom: Stable World-Model Correction for Agent Rollouts

Local AiDGX agent

arXiv:2607.01767v1 Announce Type: new Abstract: As agent planning moves from short tool chains toward persistent workflows with thousands or tens of thousands of steps, failures will occur inside larg

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

Model ReleasesDGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

Rethinking Complexity Metrics for LLM-Integrated Applications: Beyond Source Code

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.01903v1 Announce Type: new Abstract: LLM-integrated applications blend natural language prompts with program code, and much of their runtime behavior originates in the prompt layer rather t

Rethinking Generic Object Tracking Toward Human-Level Perceptual Intelligence

Local AiDGX agent

arXiv:2607.01395v1 Announce Type: cross Abstract: At the heart of human visual perception lies the ability to maintain a continuous and coherent understanding of the external world. By integrating obs

Revisiting Chain-of-Thought Reasoning under Limited Supervision: Semi-supervised Chain-of-Thought Learning

ResearchDGX agent

arXiv:2607.01511v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has emerged as an effective approach for activating latent reasoning capabilities in large language models. However, mo

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

SafetyDGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

Robust and Explainable 3D Mode Shape Recognition Using Region-Aware Graph Neural Networks

ResearchDGX agent

arXiv:2607.01522v1 Announce Type: cross Abstract: Mode shape recognition is a fundamental task in automotive NVH development, yet it remains dependent on manual visual inspection by experienced engine

Robust for the Wrong Reasons: The Representational Geometry of LLM Robustness to Science Skepticism

Model ReleasesDGX agent

arXiv:2607.01951v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consulted on contested scientific questions, raising the concern that they will sycophantically retreat

SA-HGNN: Sample-Adaptive Hyperbolic Graph Neural Network for EEG-Based Depression Recognition

ResearchDGX agent

arXiv:2607.02063v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have been widely used to capture spatial functional connectivity patterns to improve electroencephalography (EEG)-based d

SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.01876v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in multimodal understanding, yet their enormous parameter scale and cross-modal

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

SafetyDGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

SafetyDGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

Safeguarding LLM Agents from Misalignment through Provenance Analysis

SafetyDGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

Safety Targeted Embedding Exploit via Refinement

Model ReleasesDGX agent

arXiv:2607.01859v1 Announce Type: new Abstract: Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-r

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

Model ReleasesDGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

Scaling Laws for Grid-Based Approximate Nearest Neighbor Search in High Dimensions

TutorialsDGX agent

arXiv:2607.01283v1 Announce Type: cross Abstract: Grid-based approaches to approximate nearest neighbor (ANN) search have been absent from modern scaling analyses. We present a systematic characteriza

Scaling Trends for Lie Detector Oversight in Preference Learning

Model ReleasesDGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

Model ReleasesDGX agent

arXiv:2607.01612v1 Announce Type: new Abstract: Training large language models (LLMs) with reinforcement learning (RL) has significantly advanced their performance on reasoning and question-answering

Scene-Conditioned PINN-GNN for Multipath RF Maps: Cross-Scene Generation and In-Scene Completion

ResearchDGX agent

arXiv:2607.01777v1 Announce Type: cross Abstract: Radio frequency (RF) maps provide a compact representation of multipath propagation characteristics and are fundamental to channel modeling, coverage

SelectTSL: Prompt-Guided Selective Target Sound Localization in Complex Scenarios

TutorialsDGX agent

arXiv:2607.02343v1 Announce Type: cross Abstract: Humans can selectively attend to a target sound and estimate its direction in complex scenarios, whereas such selective localization remains challengi

Self-Gating Attention for Efficient Time Series Forecasting

Model ReleasesDGX agent

arXiv:2607.02344v1 Announce Type: cross Abstract: Transformer architectures have shown strong potential in time series forecasting, where multi-head self-attention is widely used to capture temporal d

SemHash-LLM: A Multi-Granularity Semantic Hashing Framework for Document Deduplication

ResearchDGX agent

arXiv:2607.01601v1 Announce Type: new Abstract: Large scale document deduplication must preserve semantic equivalence while remaining efficient over massive corpora. We present SemHash LLM, a multi gr

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

Model ReleasesDGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment

Local AiDGX agent

arXiv:2511.01390v2 Announce Type: replace-cross Abstract: Fine-grained cross-modal alignment aims to establish precise local correspondences between vision and language, forming a cornerstone for visu

Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving

SafetyDGX agent

arXiv:2604.03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-cra

SimWorlds: A Multi-Agent System for Dynamic 3D Scene Creation

Model ReleasesDGX agent

arXiv:2607.01766v1 Announce Type: new Abstract: LLM agents are increasingly used to translate natural language into 3D scenes in a procedural way, but existing systems focus on static output. Dynamic

Single-Channel EEG-Based Cognitive Load Assessment in Online Learning: A Hybrid Deep Learning Approach

ResearchDGX agent

arXiv:2607.01795v1 Announce Type: cross Abstract: Monitoring cognitive load during online learning could help instructors identify content that learners find difficult, but remote settings remove the

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

AgentsDGX agent

arXiv:2607.01874v1 Announce Type: new Abstract: Skills are becoming a reusable operational layer for LLM agents, encoding SOPs, domain rules, tool workflows, scripts, and validation routines. In reali

SkillFuzz: Fuzzing Skill Composition for Implicit Intents Discovery in Open Skill Marketplaces

AgentsDGX agent

arXiv:2607.02345v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents increasingly automate software engineering tasks through reusable skills, natural-language instruction documen

Spanning Tree Autoregressive Visual Generation

Local AiDGX agent

arXiv:2511.17089v2 Announce Type: replace-cross Abstract: We present Spanning Tree Autoregressive (STAR) modeling, which can incorporate prior knowledge of images, such as center bias and locality, to

SPARCLE: SPeaker-aware Aligned Representations via Contrastive Language Embeddings

ResearchDGX agent

arXiv:2607.01238v1 Announce Type: cross Abstract: Recent advances in speech synthesis have shifted from phoneme representations to direct grapheme modeling. While phonemes address the one-to-many mapp

Spatial Support Matters: Geometry-Aware Graph Fusion for Rainfall Field Reconstruction

ResearchDGX agent

arXiv:2607.01621v1 Announce Type: new Abstract: Fine-scale rainfall reconstruction is critical for urban flood modeling, but real rainfall sensing systems observe the field through incompatible spatia

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

Model ReleasesDGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

Spin-Weighted Spherical Harmonics Enable Complete and Scalable E(3)-Equivariant Networks

ResearchDGX agent

arXiv:2607.01408v1 Announce Type: cross Abstract: E(3)-equivariant networks are promising for 3D atomistic system modeling, yet their scalability is limited by the O(L^6) complexity of the Clebsch-Gor

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

Model ReleasesDGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

HardwareDGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

Steerability via constraints: a substrate for scalable oversight of coding agents

Model ReleasesDGX agent

arXiv:2607.02389v1 Announce Type: new Abstract: Coding agents are capable; human oversight is the bottleneck. Unconstrained agents introduce security risks, erode codebase scalability, and make human

Structuring the Space of Sociotechnical Alignment

SafetyDGX agent

arXiv:2607.01250v1 Announce Type: cross Abstract: Sociotechnical alignment concerns the social desirability of AI behavior and is thus inherently normative, not merely technical. While NLP research in

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models

ResearchDGX agent

arXiv:2607.01774v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive models. Unlike standard diffusion-based approaches,

SUNTA: Hierarchical Video Prediction with Surprise-based Chunking

ResearchDGX agent

arXiv:2607.02087v1 Announce Type: new Abstract: Hierarchical state-space models (HSSMs) offer a promising approach to long-horizon prediction by segmenting sequences into temporal chunks. However, the

TestEvo-Bench: An Executable and Live Benchmark for Test and Code Co-Evolution

Model ReleasesDGX agent

arXiv:2607.02469v1 Announce Type: cross Abstract: Software tests and code evolve together: a code change should be followed by new or updated tests that record the new software behavior. Yet existing

Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments

Model ReleasesDGX agent

arXiv:2607.02407v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in 3D indoor synthesis for Manhattan environments. However, existing methods ofte

The Agentic Garden of Forking Paths

AgentsDGX agent

arXiv:2607.01507v1 Announce Type: new Abstract: Empirical research rarely admits a unique analysis. Different analytical choices can lead to different conclusions from the same data, yet these hidden

The Dual Nature of LLM Persona: Aggregated Tendencies and Frame-Dependent Geometry

ResearchDGX agent

arXiv:2607.02368v1 Announce Type: cross Abstract: Evaluations of LLM personas via psychometric questionnaires typically rely on aggregate scores, discarding within-instance correlation structure. We t

The Eticas AI Risk Taxonomy: Open Infrastructure for Operationalizing AI Audits

Model ReleasesDGX agent

arXiv:2607.02201v1 Announce Type: cross Abstract: The rapid deployment of AI systems across high-stakes domains has created urgent demand for standardized evaluation, yet the field remains fragmented

The Rising Unsustainability of AI Graphics Cards Production

SafetyDGX agent

arXiv:2607.01258v1 Announce Type: cross Abstract: The rapid advancement of Artificial Intelligence (AI) has been accompanied by significant increases in computational and environmental costs, driven b

The Wiola Architecture for Efficient Small Language Models

Model ReleasesDGX agent

arXiv:2607.01394v1 Announce Type: new Abstract: We present Wiola, a fully original Small Language Model (SLM) architecture built from first principles, sharing no structural lineage with any existing

ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

ResearchDGX agent

arXiv:2512.07843v2 Announce Type: replace-cross Abstract: Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequ

Token Geometry

Model ReleasesDGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

ResearchDGX agent

arXiv:2607.01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and pr

Towards Cellular-Scale Interpretability in Pathology Foundation Models for Biomarker Assessment

ResearchDGX agent

arXiv:2511.05150v2 Announce Type: replace-cross Abstract: Molecular biomarker testing in pathology is often costly and tissue-consuming, limiting scalable clinical deployment. Artificial intelligence

Towards Load-Aware Prefill Deflection for Disaggregated LLM Serving

Model ReleasesDGX agent

arXiv:2607.02043v1 Announce Type: cross Abstract: Disaggregated LLM serving runs prefill and decode on separate GPU pools to keep the two phases from interfering. In practice, this creates a new asymm

Traceable Fault Diagnosis for Battery Energy Storage Systems via Retrieval-Augmented Multi-Agent O&M Assistant

AgentsDGX agent

arXiv:2607.01992v1 Announce Type: new Abstract: Large-scale battery energy storage systems (BESSs) require O&M decisions that combine alarms, cell-level measurements, device topology, diagnostic table

TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B

Model ReleasesDGX agent

arXiv:2607.01927v1 Announce Type: cross Abstract: This paper presents TUDUM (Turkce Dusunen Uretken Model), a project pipeline for adapting a Qwen-family 27B thinking model toward Turkish reasoning. T

TurnNat: Automatic Evaluation of Turn-Taking Naturalness in Dyadic Spoken Dialogue

Model ReleasesDGX agent

arXiv:2607.01345v1 Announce Type: cross Abstract: Turn-taking naturalness is central to full-duplex spoken dialogue systems, yet its automatic evaluation remains limited. Existing evaluations often re

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

Model ReleasesDGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

Uncertain but Useful: Leveraging CNN Training Variability into Data Augmentation

ResearchDGX agent

arXiv:2509.05238v2 Announce Type: replace-cross Abstract: Deep learning (DL) has transformed neuroimaging by delivering state-of-the-art performance with reduced computation times. Yet, the numerical

Understanding Agent-Based Patching of Compiler Missed Optimizations

Model ReleasesDGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

UniSE: A Unified Framework for Decoder-Only Autoregressive LM-Based Speech Enhancement

ResearchDGX agent

arXiv:2510.20441v2 Announce Type: replace-cross Abstract: Neural audio codecs have largely promoted the application of language models (LMs) for speech applications. However, the effectiveness of auto

Verifiable Knowledge Expansion through Retrieval-Grounded Formal Concept Analysis

ResearchDGX agent

arXiv:2607.01773v1 Announce Type: new Abstract: Ontology construction requires deciding which objects, attributes, and structural relations should be accepted as valid knowledge. Language models can p

← Previous
1…979899100101…358
Next →