AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Safety

PULSE: An Executable Contract Language for Spatiotemporal Knowledge Graph Engineering

DGX agent

arXiv:2608.02630v1 Announce Type: new Abstract: Knowledge graph engineering often distributes accepted state, observations, constraints, processes, and hypothetical scenarios across artifacts whose co

safetyarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Quantifying Hallucinations in Language Language Models on Medical Textbooks

DGX agent

arXiv:2603.09986v3 Announce Type: replace-cross Abstract: Hallucinations, the tendency for large language models to provide responses with factually incorrect and unsupported claims, is a serious prob

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Quo Vadis, World Modeling?

DGX agent

arXiv:2608.02713v1 Announce Type: cross Abstract: Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow,

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains

DGX agent

arXiv:2608.03219v1 Announce Type: new Abstract: Benchmark gains are often treated as evidence of greater LLM capability. Yet the same gain can reflect different changes in model behavior. A model may

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rectify Then Diffuse: Disentangling Concepts Before Denoising Trajectory Unfolds

DGX agent

arXiv:2608.03135v1 Announce Type: cross Abstract: Text-to-image diffusion models can generate individual concepts well, but they often omit or merge concepts incorrectly with multiple concepts. We tra

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning

DGX agent

arXiv:2608.03972v1 Announce Type: new Abstract: On-policy training has emerged as a powerful post-training paradigm for improving the reasoning capabilities of large language models, and is often enha

safetyarxiv-cs-ai
5 Aug 2026
Safety

Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations

DGX agent

arXiv:2608.03611v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) integrates text, audio, and vision to infer human affect, yet real-world multimodal observations are often incomplet

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

DGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Reversing Arrows in Large Language Models

DGX agent

arXiv:2608.03512v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance on text-to-knowledge graph generation and related tasks. Nevertheless, it is still unclear

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Risky Business: Measuring The Faithfulness-Safety Tension

DGX agent

arXiv:2608.03745v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output str

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models

DGX agent

arXiv:2608.02893v1 Announce Type: cross Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from laten

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Route-Align-Verify for Functional Correctness in Code Generation

DGX agent

arXiv:2608.03341v1 Announce Type: cross Abstract: Large language models (LLMs) have substantially improved code generation, yet achieving strong functional correctness remains difficult, especially fo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rubrics as Privileged Information for Open-Ended Generation

DGX agent

arXiv:2608.02948v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD), where a single model acts as both student and teacher with different contexts, has shown promise in verifiable dom

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

S^3: Improving Agent Safety through Multi-Stage Defense

DGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

SAGE: Semantic Explainability of Attention-Based Survival Models in Computational Pathology

DGX agent

arXiv:2608.02803v1 Announce Type: cross Abstract: Attention-based multiple instance learning (ABMIL) is the predominant approach for slide-level prediction in computational pathology, yet its attentio

local-aiarxiv-cs-ai
5 Aug 2026
Local Ai

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

DGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

local-aiarxiv-cs-ai
5 Aug 2026
Research

Scaling an Autoregressive Transformer for Single-Cell Generation

DGX agent

arXiv:2608.02961v1 Announce Type: cross Abstract: We study a self-supervised generation task for single-cell gene expression vectors: given a set of vectors from a cell type, we aim to generate additi

researcharxiv-cs-ai
5 Aug 2026
Model Releases

SciRet: A Compute-Aware Empirical Study of Retrieval and Reranking for Scientific RAG

DGX agent

arXiv:2608.03860v1 Announce Type: cross Abstract: We introduce SciRet, a compute-aware empirical study of retrieval-augmented generation for scientific question answering over CORD-19. Rather than pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

DGX agent

arXiv:2608.03327v1 Announce Type: new Abstract: Hybrid computer-use agents can act through screenshots or call text tools. We find that having a tool available does not settle which way the effect goe

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

DGX agent

arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

DGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Secure AI Watermarking Framework for IP Protection in Multi-Tenant Cloud Platforms

DGX agent

arXiv:2608.02656v1 Announce Type: cross Abstract: The Secured data safe guard transaction with multi-tenant environments run on private-protected authenticate platforms runs by secured handed environm

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

DGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Self-Guided Adaptive Safety Alignment: Synthesizing and Internalizing Guidelines in Reasoning Models

DGX agent

arXiv:2511.21214v4 Announce Type: replace-cross Abstract: Explicit safety policies can improve reasoning-model safety, but their effective coverage may lag behind evolving jailbreak strategies. We stu

safetyarxiv-cs-ai
5 Aug 2026
Safety

Self-Organising Digital Circuits

DGX agent

arXiv:2608.02606v1 Announce Type: new Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware redundancy and error-correcting codes. Biological sys

safetyarxiv-cs-ai
5 Aug 2026
Safety

Self-Supervised Representation-Guided Generative Dataset Distillation

DGX agent

arXiv:2608.03218v1 Announce Type: cross Abstract: Dataset distillation compresses a large training set into a compact synthetic set while retaining its downstream utility. Most existing methods target

safetyarxiv-cs-ai
5 Aug 2026
Research

Separating quantum circuits from classical LLMs

DGX agent

arXiv:2608.03962v1 Announce Type: cross Abstract: Modern large language models - transformers and diffusion language models - are built around two canonical algorithmic tasks: prediction and generatio

researcharxiv-cs-ai
5 Aug 2026
Research

Shaping Wind-Tunnel Airflow for Unmanned Aerial Vehicles using Online Learning

DGX agent

arXiv:2608.03378v1 Announce Type: cross Abstract: The development and testing of advanced aerial robots require experiments in controlled environments with tailored airflow profiles. This paper presen

researcharxiv-cs-ai
5 Aug 2026
Safety

Shielding for Higher-Order Safety

DGX agent

arXiv:2608.03662v1 Announce Type: new Abstract: Safety shields are runtime enforcement mechanisms that restrict the actions of a controller to guarantee safety. Classical shields are usually synthesis

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces

DGX agent

arXiv:2608.03401v1 Announce Type: cross Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, b

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

DGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity

DGX agent

arXiv:2608.02665v1 Announce Type: cross Abstract: A benchmark score is a measurement instrument, yet most benchmarks read each item at a single canonical surface form. We ask whether that reading is f

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

SKILL-KD: Contrastive Skill Distillation for LLM Agents

DGX agent

arXiv:2607.28048v2 Announce Type: replace Abstract: Skill-based prompting has become a practical mechanism for improving large language model (LLM) agents, yet existing skill acquisition methods often

local-aiarxiv-cs-ai
5 Aug 2026
Research

SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents

DGX agent

arXiv:2608.02356v2 Announce Type: replace Abstract: Large language model agents increasingly solve complex tasks by composing reusable skills from a library. To address this, the key challenge is not

researcharxiv-cs-ai
5 Aug 2026
Safety

SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation

DGX agent

arXiv:2608.03092v1 Announce Type: cross Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Opti

safetyarxiv-cs-ai
5 Aug 2026
Safety

Socially Grounded Agentic AI: Coordinating Plural Perspectives through Social Theory

DGX agent

arXiv:2608.03910v1 Announce Type: new Abstract: As AI systems are deployed across increasingly diverse social contexts, alignment can no longer be framed as the optimization of a single, unified set o

safetyarxiv-cs-ai
5 Aug 2026
Tutorials

Soft Guidance Starts to Outperform CoT Prompting as LLMs Improve

DGX agent

arXiv:2608.03550v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting remains the standard baseline for evaluating models' reasoning abilities. Originally, this technique was introduced to

tutorialsarxiv-cs-ai
5 Aug 2026
Safety

Solver-Aware Decompositions for Programming-by-Example: When Dividing Requires Knowing how to Conquer

DGX agent

arXiv:2608.03461v1 Announce Type: new Abstract: Decomposition-based Programming-by-example (PBE) scales performance by splitting tasks into subtasks that a learned synthesizer solves: a decomposer pre

safetyarxiv-cs-ai
5 Aug 2026
Safety

SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling

DGX agent

arXiv:2608.02951v1 Announce Type: cross Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods

safetyarxiv-cs-ai
5 Aug 2026
Research

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels

DGX agent

arXiv:2608.02995v1 Announce Type: cross Abstract: Modern large language models (LLMs) exhibit activation sparsity, wherein only a subset of their neurons is activated for given input tokens. Researche

researcharxiv-cs-ai
5 Aug 2026
Local Ai

Spatial proteomics guided by H&E-based AI reveals recurrence-risk niches in triple-negative breast cancer

DGX agent

arXiv:2608.03145v1 Announce Type: new Abstract: Deep learning models can predict cancer recurrence from H&E stained slides, but the localized molecular states underlying these predictions remain large

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them

DGX agent

arXiv:2607.27703v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used in embodied agents to interpret visual inputs, reason about spatial relationships, and make task

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

Speculative Correction: Draft-then-Refine Decoding for Diffusion Language Models

DGX agent

arXiv:2608.02625v1 Announce Type: cross Abstract: Diffusion language models (DLMs) can revise tokens bidirectionally, but standard decoding procedures often adapt them to left-to-right generation by p

local-aiarxiv-cs-ai
5 Aug 2026
Research

Speech LLMs in Low-Resource Scenarios: Data Volume Requirements and the Impact of Pretraining on High-Resource Languages

DGX agent

arXiv:2508.05149v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated potential in handling spoken inputs for high-resource languages, reaching state-of-the-art perf

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Sphere Retraction Normalizations

DGX agent

arXiv:2608.02668v1 Announce Type: cross Abstract: Residual connections are the de facto mechanism for training deep neural networks stably. Geodesic Normalization (GeoNorm) recasts them on a Riemannia

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Standalone DINOv3 for Training-Free Open-Vocabulary Semantic Segmentation in Remote Sensing

DGX agent

arXiv:2608.03023v1 Announce Type: cross Abstract: Remote sensing semantic segmentation is hindered by costly pixel-level annotations, motivating training-free open-vocabulary methods. Recently, the re

researcharxiv-cs-ai
5 Aug 2026
Research

State Propagation Also Satisfies: A Complex-Valued State-Space Model for Deterministic State Tracking

DGX agent

arXiv:2608.03425v1 Announce Type: new Abstract: Transformer-based architectures have dominated sequence modeling, largely due to the expressive power of attention mechanisms. However, for a class of d

researcharxiv-cs-ai
5 Aug 2026
Agents

Steganalysis of Adaptive Covert Collusion in Tool-Using Agent Populations: A Black-Box, Cross-Principal Approach

DGX agent

arXiv:2608.02698v1 Announce Type: cross Abstract: Tool-using agents built on large language models (LLMs) are increasingly deployed not by a single operator but by many, side by side on shared infrast

agentsarxiv-cs-ai
5 Aug 2026
← Previous
1…4243444546…443
Next →