AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

DGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

model-releasesarxiv-cs-ai
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

P2Voxel: Pyramid Pivot Voxelization for 3D Mesh Tokenization

DGX agent

arXiv:2608.07549v1 Announce Type: cross Abstract: Triangle meshes provide explicit and accurate surface geometry, yet their irregular topology connectivity makes 3D mesh tokenization a geometric sampl

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

P^{3}: Joint Program-and-Proof Planning for Verified Code Generation

DGX agent

arXiv:2608.09277v1 Announce Type: new Abstract: Verified code generation asks a large language model (LLM) to generate both an executable program and a machine-checkable proof that the program meets a

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

PACE: A Playback-Aligned Context Engine for LLM-Based Full-Duplex Voice Dialogue

DGX agent

arXiv:2608.07631v1 Announce Type: cross Abstract: LLM-based full-duplex voice services allow users to speak while the assistant is responding. Because servers can generate output and advance dialogue

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Parameter Exploration for RLVR via Variational Learning

DGX agent

arXiv:2608.09805v1 Announce Type: cross Abstract: Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an importan

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distillation

DGX agent

arXiv:2608.08726v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) uses a privileged teacher to supervise a reasoning model on prefixes sampled from its own rollouts. Yet each rollou

safetyarxiv-cs-ai
11 Aug 2026
Research

PATH: Next-Interval Prediction via Autoregressive Tree Hierarchy on Tabular Data

DGX agent

arXiv:2608.08078v1 Announce Type: new Abstract: Interval prediction aims to achieve a target coverage level while producing intervals that are as short as possible. Many conformal regression pipelines

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Performance of large language models in the optical diagnosis of colorectal polyps

DGX agent

arXiv:2608.07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Persistent Semantic Entities in Tool-Augmented LLM Systems

DGX agent

arXiv:2608.07952v1 Announce Type: cross Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language Models

DGX agent

arXiv:2608.08199v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in group decision-making with other LLMs and humans. Yet it remains unclear whether their influen

safetyarxiv-cs-ai
11 Aug 2026
Safety

PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering

DGX agent

arXiv:2608.07509v1 Announce Type: cross Abstract: LLMs are increasingly used for conversational tutoring, but effective tutoring requires more than correct answers. Tutors must choose when to scaffold

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

DGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

PolicyKG: An Agentic LLM Pipeline for Translating Institutional Policies into SHACL Knowledge Graphs

DGX agent

arXiv:2608.09028v1 Announce Type: new Abstract: Institutional policies stay in natural language while the systems that check compliance demand machine-readable constraints. Bridging that gap is still

model-releasesarxiv-cs-ai
11 Aug 2026
Research

PolypSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering

DGX agent

arXiv:2603.07066v2 Announce Type: replace-cross Abstract: Generative diffusion models are increasingly used for medical imaging data augmentation, but text prompting cannot produce causal training dat

researcharxiv-cs-ai
11 Aug 2026
Agents

Population-Scalable Multi-Agent World Modeling

DGX agent

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environment

agentsarxiv-cs-ai
11 Aug 2026
Local Ai

Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale

DGX agent

arXiv:2601.21249v2 Announce Type: replace Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is

local-aiarxiv-cs-ai
11 Aug 2026
Research

Positioning Generative Artificial Intelligence in STEM Assessment: When to Require, Scaffold, or Restrict Its Use

DGX agent

arXiv:2608.07475v1 Announce Type: cross Abstract: Generative Artificial Intelligence (GenAI) presents a governance challenge for STEM assessment. Unrestricted access can enable task outsourcing that u

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Predictive safety filter enhanced curriculum learning control for efficient vehicle dynamics controller

DGX agent

arXiv:2608.09653v1 Announce Type: cross Abstract: Recent advances in learning-based control have enabled impressive achievements in solving complex control problems in various domains. However, since

model-releasesarxiv-cs-ai
11 Aug 2026
Research

PRISM: A Predictive Protocol for Permutation Optimization via Landscape Diagnostics

DGX agent

arXiv:2608.08344v1 Announce Type: cross Abstract: Permutation optimization arises whenever the components of a system are fixed but their ordering affects performance. We introduce PRISM, a predictive

researcharxiv-cs-ai
11 Aug 2026
Safety

Privacy-Preserving Data Drift Detection and Recovery for Large-Scale LLM Applications via Proxy Representations

DGX agent

arXiv:2608.08245v1 Announce Type: cross Abstract: LLM applications deployed at scale face a fundamental challenge: privacy constraints prevent direct inspection of user interactions, making it difficu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Private Anytime Selective-Risk Certification for Federated Retrieval-Augmented Generation: Guarantees and Empirical Limits

DGX agent

arXiv:2608.07913v1 Announce Type: cross Abstract: Selective-risk certificates promise that accepted outputs meet a declared error target. We develop Fed-SRC, a score-agnostic certificate for federated

safetyarxiv-cs-ai
11 Aug 2026
Local Ai

Private Etymology: Designing Relational Reuse of Shared Symbols in Long-Term Human-AI Interaction

DGX agent

arXiv:2608.08443v1 Announce Type: cross Abstract: Previous studies have shown that people can develop shared symbols, partner-specific expressions, personal idioms, inside jokes, and other parts of a

local-aiarxiv-cs-ai
11 Aug 2026
Safety

Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation

DGX agent

arXiv:2608.09263v1 Announce Type: new Abstract: Outcome verifiers score completed reasoning traces but do not assign credit to intermediate tokens. Privileged self-distillation attempts to fill this g

safetyarxiv-cs-ai
11 Aug 2026
Safety

Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation

DGX agent

arXiv:2608.09228v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) is commonly interpreted as the transfer of privileged information: a teacher observes the verified solution to the

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Probabilistic Circuits for Knowledge Graph Completion with Reduced Rule Sets

DGX agent

arXiv:2508.06706v2 Announce Type: replace Abstract: Rule-based methods for knowledge graph completion provide explainable results, but often require tens of thousands of rules to achieve competitive p

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ProbSPARQL: Querying Knowledge Graphs with Multi-dimensional, Uncertain Numeric Data

DGX agent

arXiv:2607.18262v2 Announce Type: replace Abstract: The SFB 1574 Circular Factory is building a shared knowledge graph infrastructure for integrating data about returned products. A central challenge

researcharxiv-cs-ai
11 Aug 2026
Research

Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States

DGX agent

arXiv:2608.08024v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent and useful responses but remain prone to hallucinations. We introduce Prompt Embedding Probes (PEP),

researcharxiv-cs-ai
11 Aug 2026
Model Releases

PROSLEX: A Novel Dataset for Expert-Annotated Legal Statute Prediction for Indian Judiciary

DGX agent

arXiv:2608.08830v1 Announce Type: new Abstract: Legal Statute Prediction (LSP) involves automatically identifying relevant legal statutes given factual descriptions in legal documents, typically frame

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

Protecting patient privacy in clinical foundation models: Technical and legal perspectives

DGX agent

arXiv:2608.07705v1 Announce Type: new Abstract: Clinical foundation models trained on large-scale patient data are increasingly used for decision support, screening, and public health. As deployment e

applicationsarxiv-cs-ai
11 Aug 2026
Safety

Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update

DGX agent

arXiv:2607.11505v2 Announce Type: replace-cross Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behav

safetyarxiv-cs-ai
11 Aug 2026
Research

Quantization Degradation in Large Language Models: A Signal-Noise Perspective

DGX agent

arXiv:2608.08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-wi

researcharxiv-cs-ai
11 Aug 2026
Agents

QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing

DGX agent

arXiv:2608.07743v1 Announce Type: new Abstract: Identifying a meaningful quantum speedup requires more than matching a classical problem to a familiar quantum primitive: the claim must preserve the ta

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

DGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning

DGX agent

arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex tasks. However, manually authored skills often adap

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis

DGX agent

arXiv:2509.21629v4 Announce Type: replace-cross Abstract: Program verification relies on loop invariants, yet automatically discovering strong invariants remains a long-standing challenge. We investig

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation

DGX agent

arXiv:2601.10168v3 Announce Type: replace-cross Abstract: Open-vocabulary 3D Scene Graph (3DSG) can enhance various downstream tasks in robotics by leveraging structured semantic representations, yet

applicationsarxiv-cs-ai
11 Aug 2026
Research

RAG-Audio: Retrieval-Augmented Generation for Faithful Brain-to-Audio Reconstruction

DGX agent

arXiv:2608.09331v1 Announce Type: cross Abstract: Brain-to-audio reconstruction is limited by prior domination: when a pretrained generator is conditioned on a weak neural signal, it produces realisti

researcharxiv-cs-ai
11 Aug 2026
Model Releases

RAG-Based Auto-Configuration for Industrial Fieldbus Devices

DGX agent

arXiv:2608.08618v1 Announce Type: cross Abstract: Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and tra

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges

DGX agent

arXiv:2608.09526v1 Announce Type: cross Abstract: Real-world cyberattacks often require sustained progress across multiple hosts and network segments, making multi-hop cyber ranges essential infrastru

agentsarxiv-cs-ai
11 Aug 2026
Research

RankGuide: Tensor-Rank-Guided Routing and Steering for Efficient Reasoning

DGX agent

arXiv:2604.16694v2 Announce Type: replace Abstract: Large reasoning models (LRMs) enhance problem-solving capabilities by generating explicit multi-step chains of thought (CoT) reasoning; however, the

researcharxiv-cs-ai
11 Aug 2026
Research

RAVEN-Eval: Rubric-Guided Automatic Evaluation for AI Video Generation Models Based on LMM Preference Judgement

DGX agent

arXiv:2608.09111v1 Announce Type: new Abstract: AI video generation has advanced rapidly and entered widespread commercial use. As a result, quality differences among videos produced by state-of-the-a

researcharxiv-cs-ai
11 Aug 2026
Safety

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

DGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills

DGX agent

arXiv:2608.07885v1 Announce Type: new Abstract: Reasoning modes of language models outperform their non-reasoning counterparts on multi-step agentic tasks, but pay a 3-6x premium in output tokens on e

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation

DGX agent

arXiv:2608.09467v1 Announce Type: cross Abstract: Unmanned aerial vehicle vision-language navigation (UAV-VLN) requires agents to translate visual observations and language instructions into reliable

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Reflex First, Reflect Later: Latency-Aware Embodied LLM Agents for Dynamic Response

DGX agent

arXiv:2506.07223v2 Announce Type: replace Abstract: Large language models (LLMs) have substantially improved the planning capabilities of embodied agents, enabling their deployment in dynamic and safe

safetyarxiv-cs-ai
11 Aug 2026
Safety

REIN: Bridging the Gap between Reasoning and Reliability via Reflection and Abstention Alignment

DGX agent

arXiv:2608.07931v1 Announce Type: new Abstract: Large reasoning models (LRMs) are prone to hallucination, which undermines their reliability and poses challenges for safe deployment. Hallucinations in

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation

DGX agent

arXiv:2503.22122v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly for long-horizon tasks that require

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ReMIND: Orchestrating Modular Large Language Models for Controllable Serendipity A REM-Inspired System Design for Emergent Creative Ideation

DGX agent

arXiv:2601.07121v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used not only for problem solving but also for creative ideation; however, generating ideas that

researcharxiv-cs-ai
11 Aug 2026
← Previous
1…1819202122…443
Next →