AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

Beyond Final Answers: Auditing Trajectory-Level Hallucinations in Multi-Agent Industrial Workflows

DGX agent

arXiv:2605.24219v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that reason, use tools, and act over multiple steps. Yet most hallucination

model-releasesarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion

DGX agent

arXiv:2605.24631v1 Announce Type: cross Abstract: Minority sampling aims to generate low-density instances on a data manifold and is of central importance in applications such as medical diagnosis, an

applicationsarxiv-cs-ai
26 May 2026
Model Releases

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

DGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

model-releasesarxiv-cs-ai
26 May 2026
Safety

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

DGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

safetyarxiv-cs-ai
26 May 2026
Agents

Beyond Predefined Learning Objects: A Thinking-Learning Interaction Model for Up-to-Date Autonomous Robot Learning

DGX agent

arXiv:2605.23987v1 Announce Type: new Abstract: Autonomous robots operating in open and changing environments cannot always rely on predefined inputs, outputs, and action routines. Although existing l

agentsarxiv-cs-ai
26 May 2026
Model Releases

Beyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching

DGX agent

arXiv:2605.25558v1 Announce Type: new Abstract: Optimizing the trade-off among predictive performance and computational cost is a central focus in the deployment of Large Language Models (LLMs). Curre

model-releasesarxiv-cs-ai
26 May 2026
Applications

Beyond Static Uncertainty: Modeling Temporal Uncertainty Dynamics for Probabilistic Time Series Forecasting

DGX agent

arXiv:2603.24254v2 Announce Type: replace-cross Abstract: Real-world time series exhibit temporally structured uncertainty: volatility clusters in turbulent regimes, dissipates in stable periods, and

applicationsarxiv-cs-ai
26 May 2026
Model Releases

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

DGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Beyond the Aggregation Dilemma: Prior-Retaining Decoupled Learning for Multimodal Graphs

DGX agent

arXiv:2605.24684v1 Announce Type: cross Abstract: Multimodal Attributed Graph Learning (MAGL) integrates intrinsic node attributes with structural topology via graph aggregation. However, as pretraine

hardwarearxiv-cs-ai
26 May 2026
Research

Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling

DGX agent

arXiv:2605.25143v1 Announce Type: new Abstract: Test-time scaling improves language model reasoning by spending additional compute to explore multiple solution trajectories. The key challenge is to ma

researcharxiv-cs-ai
26 May 2026
Safety

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

DGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

safetyarxiv-cs-ai
26 May 2026
Research

Bilevel Optimization of Synthetic Trajectories for Multi-Turn LLM Fine-Tuning

DGX agent

arXiv:2605.24743v1 Announce Type: cross Abstract: While LLMs excel at single-turn generation, they struggle with long-horizon, multi-turn interactions. Offline reinforcement learning (RL) offers a sca

researcharxiv-cs-ai
26 May 2026
Research

Binding Visual Features Point by Point

DGX agent

arXiv:2605.25427v1 Announce Type: cross Abstract: Despite success on standard benchmarks, vision language models display persistent failures on tasks involving processing of multi-object scenes, inclu

researcharxiv-cs-ai
26 May 2026
Model Releases

BODHI: Precise OS Kernel Specification Inference

DGX agent

arXiv:2605.23931v1 Announce Type: new Abstract: The formal verification of operating system kernels requires precise specifications that capture the intended behavior of system calls. Writing these sp

model-releasesarxiv-cs-ai
26 May 2026
Local Ai

Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models

DGX agent

arXiv:2605.25230v1 Announce Type: new Abstract: Recent work on recursive architectures has shown that tiny neural networks can be surprisingly powerful on structured reasoning tasks. The trick is to m

local-aiarxiv-cs-ai
26 May 2026
Tutorials

BoxLitE: A Faithful Knowledge Base Embedding Based on Convex Optimization

DGX agent

arXiv:2605.23937v1 Announce Type: new Abstract: Knowledge base (KB) embeddings aim at combining the capability of classical knowledge graph embeddings to generalize the information present in facts, t

tutorialsarxiv-cs-ai
26 May 2026
Research

Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models

DGX agent

arXiv:2605.24053v1 Announce Type: new Abstract: Large Language Models (LLMs) are predominantly governed by probabilistic frameworks in which the sum of outcome probabilities is constrained to unity. T

researcharxiv-cs-ai
26 May 2026
Research

Bridging Evolutionary Algorithms and Reinforcement Learning: A Comprehensive Survey on Hybrid Algorithms

DGX agent

arXiv:2401.11963v5 Announce Type: replace-cross Abstract: Evolutionary Reinforcement Learning (ERL), which integrates Evolutionary Algorithms (EAs) and Reinforcement Learning (RL) for optimization, ha

researcharxiv-cs-ai
26 May 2026
Safety

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

DGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

safetyarxiv-cs-ai
26 May 2026
Research

Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference

DGX agent

arXiv:2511.16449v4 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown great potential for embodied AI by integrating visual perception, language understanding, and a

researcharxiv-cs-ai
26 May 2026
Applications

By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode

DGX agent

arXiv:2605.25186v1 Announce Type: cross Abstract: Formalizing legal provisions promises machine-accessible law and automated legal reasoning, and recent LLMs make it tempting to generate such formaliz

applicationsarxiv-cs-ai
26 May 2026
Model Releases

Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning

DGX agent

arXiv:2605.25920v1 Announce Type: cross Abstract: While large language models (LLMs) augmented with agentic search capabilities show promise for legal reasoning, they overlook a fundamental constraint

model-releasesarxiv-cs-ai
26 May 2026
Research

CARL-CXR: Continual Adapter-Based Routing for Task-Unknown Chest Radiograph Classification

DGX agent

arXiv:2602.15811v2 Announce Type: replace-cross Abstract: Clinical deployment of chest radiograph classifiers requires models that can be updated as new datasets become available without retraining on

researcharxiv-cs-ai
26 May 2026
Model Releases

Cascade-KDE: Robust Time-Series Restoration under Out-of-Distribution Impulse Corruptions

DGX agent

arXiv:2605.24055v1 Announce Type: cross Abstract: Real-world time-series data in industrial sensing, healthcare, and energy systems is often corrupted by a mixture of Gaussian noise and occasional lar

model-releasesarxiv-cs-ai
26 May 2026
Local Ai

Catching MRI outliers: unsupervised detection and localization of MRI artefacts and clinical anomalies using deep learning

DGX agent

arXiv:2605.24609v1 Announce Type: cross Abstract: Artificial intelligence is increasingly integrated into radiotherapy workflows, yet such pipelines remain vulnerable to out-of-distribution image data

local-aiarxiv-cs-ai
26 May 2026
Research

Catching The Correct Answer Trap: Characterising AI Tutor Blind Spots When Analysing Student Reasoning

DGX agent

arXiv:2605.23925v1 Announce Type: cross Abstract: Intelligent tutoring systems increasingly provide automated feedback on student work, but robust feedback requires assessing reasoning, not only final

researcharxiv-cs-ai
26 May 2026
Model Releases

Causal Tongue-Tie: LLMs Can Encode Causal Direction, But Their Yes/No Outputs Fail to Express

DGX agent

arXiv:2605.25891v1 Announce Type: cross Abstract: We find a mismatch between what large language models encode about a causal question and what they answer. On anti-commonsense CLadder items, a fixed

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Agents

CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures

DGX agent

arXiv:2605.25338v1 Announce Type: cross Abstract: Large language model (LLM) agents frequently fail on multi-step tasks involving reasoning, tool use, and environment interaction. While such failures

agentsarxiv-cs-ai
26 May 2026
Safety

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

DGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

safetyarxiv-cs-ai
26 May 2026
Model Releases

Chain-of-Thought Hijacking

DGX agent

arXiv:2510.26418v4 Announce Type: replace Abstract: Large Reasoning Models (LRMs) improve task performance through extended inference-time reasoning. Although previous studies suggest that longer reas

model-releasesarxiv-cs-ai
26 May 2026
Research

Channel-wise Vector Quantization

DGX agent

arXiv:2605.26089v1 Announce Type: cross Abstract: We present Channel-wise Vector Quantization (CVQ), a novel image tokenization paradigm that replaces patch-wise tokens with channel-wise tokens. Unlik

researcharxiv-cs-ai
26 May 2026
Model Releases

ChaosBench-Logic v2: Evaluating LLM Logical Reasoning over Dynamical Systems at Scale

DGX agent

arXiv:2605.24305v1 Announce Type: cross Abstract: Standard accuracy on binary reasoning benchmarks hides critical failure modes: prior collapse, inconsistency under paraphrase, and inability to reason

model-releasesarxiv-cs-ai
26 May 2026
Safety

Characterizing Linear Alignment Across Language Models

DGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

safetyarxiv-cs-ai
26 May 2026
Model Releases

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference

DGX agent

arXiv:2510.02361v2 Announce Type: replace-cross Abstract: Transformer-based large models excel in natural language processing and computer vision, but face severe computational inefficiencies due to t

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CITYREP: A Unified Benchmark for Urban Representations Across Cities, Tasks, and Modalities

DGX agent

arXiv:2605.26036v1 Announce Type: new Abstract: Urban representation learning encodes complex urban environments into general-purpose embeddings for diverse downstream tasks and emerging urban foundat

model-releasesarxiv-cs-ai
26 May 2026
Research

Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation

DGX agent

arXiv:2605.25831v1 Announce Type: cross Abstract: Large language models (LLMs) define a distribution over text, which can be viewed as a probabilistic representation of uncertainty: sampling K respons

researcharxiv-cs-ai
26 May 2026
Model Releases

Claw-Anything: Benchmarking Always-On Personal Assistants with Broader Access to User's Digital World

DGX agent

arXiv:2605.26086v1 Announce Type: new Abstract: Large language model agents are increasingly envisioned as always-on personal assistants with access to anything relevant in the user's digital world. Y

model-releasesarxiv-cs-ai
26 May 2026
Research

CLiViS: Unleashing Cognitive Map through Linguistic-Visual Synergy for Embodied Visual Reasoning

DGX agent

arXiv:2506.17629v2 Announce Type: replace-cross Abstract: Embodied Visual Reasoning (EVR) seeks to follow complex, free-form instructions based on egocentric video, enabling semantic understanding and

researcharxiv-cs-ai
26 May 2026
Safety

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

DGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

safetyarxiv-cs-ai
26 May 2026
Research

Coarse-to-Fine Domain Incremental Learning with Attentive Distillation for Mining Footprint Segmentation in Multispectral Imagery

DGX agent

arXiv:2605.24460v1 Announce Type: cross Abstract: Automatically mapping and segmenting global mining footprints using remote sensing and deep learning is critical for monitoring the socio-environmenta

researcharxiv-cs-ai
26 May 2026
Model Releases

Code2UML: Agentic LLMs with context engineering for scalable software visualization

DGX agent

arXiv:2605.24453v1 Announce Type: cross Abstract: Large Language Model (LLM)-based code analysis tools are adopted to automate software documentation tasks. However, the scalability of these approache

model-releasesarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Model Releases

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation

DGX agent

arXiv:2605.25378v1 Announce Type: cross Abstract: Customized image editing aims to equip pre-trained diffusion models with specific visual effects using limited paired data, typically via Low-Rank Ada

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Committed SAE-Feature Traces for Audited-Session Substitution Detection in Hosted LLMs

DGX agent

arXiv:2604.18179v2 Announce Type: replace-cross Abstract: Hosted-LLM providers have a silent-substitution incentive: advertise a stronger model while serving cheaper replies. Probe-after-return scheme

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Complement Submodular Information Measures for Balanced and Robust Data Selection

DGX agent

arXiv:2605.24779v1 Announce Type: cross Abstract: Submodular optimization has become a fundamental paradigm for data selection, retrieval, summarization, and representation learning due to its ability

model-releasesarxiv-cs-ai
26 May 2026
Safety

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

DGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

safetyarxiv-cs-ai
26 May 2026
Model Releases

Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models

DGX agent

arXiv:2605.25765v1 Announce Type: cross Abstract: Concept unlearning aims to erase a target concept from a pretrained text-to-image diffusion model without retraining. Closed-form methods are attracti

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…264265266267268…448
Next →