AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval

DGX agent

arXiv:2606.28365v1 Announce Type: cross Abstract: RAG ingestion pipelines frequently augment search corpus index with semantic enrichment indices (e.g., synthetic queries or summaries generated from c

safetyarxiv-cs-ai
30 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation

DGX agent

arXiv:2511.05852v4 Announce Type: replace-cross Abstract: Knowledge editing (KE) offers a lightweight alternative to retraining for updating large language models (LLMs). Meanwhile, fine-tuning remain

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Can LLMs Rank? A Tale of Triads and Triage

DGX agent

arXiv:2606.30412v1 Announce Type: cross Abstract: From housing allocation for households experiencing homelessness to triage in emergency departments, LLMs are increasingly being considered as judges

researcharxiv-cs-ai
30 Jun 2026
Local Ai

Can Machines Really See Objects in Images? A Study Based on Syntactic Distance and Visual Self-Referential Instances

DGX agent

arXiv:2606.29416v1 Announce Type: cross Abstract: Can a vision model truly see an object, or does it only fit surface-level visual cues? Following Wittgenstein's view that the limits of language are t

local-aiarxiv-cs-ai
30 Jun 2026
Agents

Capability Gates Are Not Authorization: Confused-Deputy Failures in LLM Agent Frameworks

DGX agent

arXiv:2606.28679v1 Announce Type: cross Abstract: Tool-using LLM agents increasingly read untrusted content while holding side-effecting tools such as payments, email, CRM, and infrastructure APIs, ye

agentsarxiv-cs-ai
30 Jun 2026
Agents

CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training

DGX agent

arXiv:2603.23559v2 Announce Type: replace-cross Abstract: GUI agents are rapidly shifting from multi-module pipelines to end-to-end, native vision-language models (VLMs) that perceive raw screenshots

agentsarxiv-cs-ai
30 Jun 2026
Safety

Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

DGX agent

arXiv:2606.28360v1 Announce Type: cross Abstract: University students often struggle to navigate complex academic policies, leading to advising bottlenecks and delayed access to critical information.

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models

DGX agent

arXiv:2501.14940v4 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with human values is essential for their safe deployment and widespread adoption. Current LLM safety ben

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Categorizing Mathematical Concepts with LLM Voting Ensembles in Mathswitch

DGX agent

arXiv:2606.28815v1 Announce Type: cross Abstract: Mathswitch is an open-source project that imports mathematical concept records from sources such as Wikidata, Wikipedia, MathWorld, Encyclopedia of Ma

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework

DGX agent

arXiv:2406.08311v3 Announce Type: replace-cross Abstract: Existing evaluations of tabular synthesis models rely primarily on low-order statistics and downstream task performance, leaving multivariate

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Local Ai

CCRC: A Change-Aware Captioning and Reasoning Chain for Image Change Captioning and Segmentation

DGX agent

arXiv:2606.28724v1 Announce Type: cross Abstract: Understanding and localizing subtle changes between paired images is critical for tasks such as surveillance and image editing. However, traditional I

local-aiarxiv-cs-ai
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Research

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models

DGX agent

arXiv:2606.29897v1 Announce Type: cross Abstract: Voice anonymization aims to protect speaker identity while preserving linguistic content and speech usability. However, most anonymization systems are

researcharxiv-cs-ai
30 Jun 2026
Agents

Choose Your Agent: Tradeoffs in Adopting AI Advisors, Coaches, and Delegates in Multi-Party Negotiation

DGX agent

arXiv:2602.12089v3 Announce Type: replace-cross Abstract: As AI usage becomes more prevalent in social contexts, understanding agent-user interaction is critical to designing systems that imp rove bot

agentsarxiv-cs-ai
30 Jun 2026
Tutorials

CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts

DGX agent

arXiv:2510.09278v2 Announce Type: replace-cross Abstract: Training expert LLMs in domains with scarce data is difficult, often relying on multiple-choice questions (MCQs). However, standard outcome-ba

tutorialsarxiv-cs-ai
30 Jun 2026
Agents

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

DGX agent

arXiv:2606.30246v1 Announce Type: new Abstract: Existing autonomous research agents can support parts of the research process, but most systems still treat research as either an isolated assistant tas

agentsarxiv-cs-ai
30 Jun 2026
Research

Clinical Reasoning Graphs: Structured Evaluation of LLM Diagnostic Reasoning Reveals Competence Without Consistency

DGX agent

arXiv:2606.29876v1 Announce Type: cross Abstract: Modern large language models (LLMs) reach 60-70% diagnostic accuracy on complex clinical case benchmarks, but accuracy alone cannot distinguish stable

researcharxiv-cs-ai
30 Jun 2026
Research

CLMASP: Coupling Large Language Models with Answer Set Programming for Robotic Task Planning

DGX agent

arXiv:2406.03367v2 Announce Type: replace Abstract: Large Language Models (LLMs) possess extensive foundational knowledge and moderate reasoning abilities, making them suitable for general task planni

researcharxiv-cs-ai
30 Jun 2026
Research

Closed-Form Steepest Descent Direction toward Flat Minima: Reducing Upper Bounds on the Loss Hessian Eigenspectrum in Neural Networks

DGX agent

arXiv:2606.28662v1 Announce Type: cross Abstract: The flatness hypothesis suggests that flatness of the loss landscape, as measured by the eigenvalues of the loss Hessian, correlates with better neura

researcharxiv-cs-ai
30 Jun 2026
Model Releases

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation

DGX agent

arXiv:2606.28397v1 Announce Type: cross Abstract: Vision-language navigation (VLN) has recently advanced with large language and multimodal models, enabling agents to follow natural-language instructi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Closing the Activation-Cone Blind Spot: Response-Time Probing and Unified Defense

DGX agent

arXiv:2606.29441v1 Announce Type: cross Abstract: Inference-time safety methods for large language models have proliferated, yet no systematic comparison exists. We evaluate five defense paradigms (no

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents

DGX agent

arXiv:2606.29771v1 Announce Type: new Abstract: LLM agents are increasingly cast as autonomous portfolio managers, and benchmarks have moved from financial question-answering to sequential trading. Ye

model-releasesarxiv-cs-ai
30 Jun 2026
Tutorials

Clustering Unsupervised Representations as Defense against Poisoning Attacks on Speech Commands Classification System

DGX agent

arXiv:2606.28953v1 Announce Type: cross Abstract: Poisoning attacks entail attackers intentionally tampering with training data. In this paper, we consider a dirty-label poisoning attack scenario on a

tutorialsarxiv-cs-ai
30 Jun 2026
Research

CMSL: Constructive Multi-Sequence Learning for Recommendation Systems

DGX agent

arXiv:2606.28533v1 Announce Type: cross Abstract: Sequence learning has emerged as the promising paradigm in recommendation systems, surpassing traditional Deep Learning Recommendation Models (DLRM) b

researcharxiv-cs-ai
30 Jun 2026
Safety

CMTFormer: Marrying Transformer with Hierarchical Information Interaction for RGB-Event Object Detection

DGX agent

arXiv:2606.29136v1 Announce Type: cross Abstract: Event cameras capture sparse brightness changes with high temporal resolution and high dynamic range, compensating for the deficiencies of the convent

safetyarxiv-cs-ai
30 Jun 2026
Agents

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

DGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Cognitive World Models for Process-Level Social Influence Evaluation

DGX agent

arXiv:2606.29495v1 Announce Type: new Abstract: Social influence dialogue changes user behavior by altering internal cognitive states. The central evaluation question is whether the user's beliefs, de

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

DGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Collective cooperation without individual fidelity in LLM agents

DGX agent

arXiv:2606.30454v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as agents in simulations of social systems, yet it remains unclear when their behavior can be inter

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models

DGX agent

arXiv:2606.28719v1 Announce Type: new Abstract: Test-time adaptation (TTA) of vision-language models (VLMs) is essential for their robust deployment in dynamic, real-world environments. However, exist

model-releasesarxiv-cs-ai
30 Jun 2026
Research

COMPASS: Grounding Composition-Intent Guidance in Unified Multimodal Models

DGX agent

arXiv:2606.28696v1 Announce Type: new Abstract: Composition is a high-level visual intent that governs where subjects are placed and how a scene is organized, yet current unified multimodal models rem

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Compositional Dynamics in Learning and Mechanics

DGX agent

arXiv:2606.28984v1 Announce Type: cross Abstract: We give a single compositional setting in which gradient-based learning and Hamiltonian-style mechanics appear as functorial semantics. The syntax is

model-releasesarxiv-cs-ai
30 Jun 2026
Hardware

ConCise: Training-Free Conclusion-Chain State Compression for Cost-Efficient Multi-Step RAG Services

DGX agent

arXiv:2606.28361v1 Announce Type: cross Abstract: Multi-step retrieval-augmented generation (RAG) has been widely deployed as LLM-powered web services for complex question answering, where iterative r

hardwarearxiv-cs-ai
30 Jun 2026
Tutorials

Confidence-feedback-weighted graph matching network: online-offline laser-induced damage site matching under complex interference

DGX agent

arXiv:2606.29255v1 Announce Type: cross Abstract: Online inspection images of final optics in high-power laser facilities contain pseudo-damage sites that closely resemble true damage sites. Determini

tutorialsarxiv-cs-ai
30 Jun 2026
Applications

Constrained Tabular Diffusion for Finance

DGX agent

arXiv:2606.28674v1 Announce Type: cross Abstract: Generative models in finance face the dual challenge of producing realistic data while satisfying strict regulatory and economic objectives, a require

applicationsarxiv-cs-ai
30 Jun 2026
Model Releases

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

DGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Correct codes for the wrong reasons? validating LLMs as measurement instruments for theoretical constructs

DGX agent

arXiv:2606.28574v1 Announce Type: cross Abstract: When a large language model (LLM) codes a construct in text as a human annotator would, that agreement makes the LLM a reliable coder. Yet reliability

researcharxiv-cs-ai
30 Jun 2026
Model Releases

CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents

DGX agent

arXiv:2511.02734v3 Announce Type: replace Abstract: Current evaluations of Large Language Model (LLM) agents primarily emphasize task completion, often overlooking resource efficiency and adaptability

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Counterfactual Residual Data Augmentation for Regression

DGX agent

arXiv:2606.28460v1 Announce Type: cross Abstract: Data-driven modeling in real-world regression tasks often suffers from limited training samples, high collection costs, and noisy observations. Inspir

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM

DGX agent

arXiv:2606.29563v1 Announce Type: cross Abstract: Large language models (LLMs) excel at complex tasks like question answering and summarization, thanks to their ability to handle long-context inputs.

researcharxiv-cs-ai
30 Jun 2026
Research

Covering the Unseen: Information Demand Coverage Optimization for Retrieval-Augmented Generation

DGX agent

arXiv:2606.29328v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) typically treats context selection as ranking chunks against a single query embedding. This assumption breaks dow

researcharxiv-cs-ai
30 Jun 2026
Safety

CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning

DGX agent

arXiv:2606.29476v1 Announce Type: cross Abstract: Self-distilled agentic reinforcement learning augments trajectory-level reward with a token-level distillation loss, using as its teacher the same pol

safetyarxiv-cs-ai
30 Jun 2026
Safety

Critical Interval MSE: Toward Reliable Offline Validation for Robot Manipulation Policies

DGX agent

arXiv:2606.29898v1 Announce Type: cross Abstract: Real-world evaluation is the gold standard for robot policies because it tests them against the physical conditions and deployment challenges they are

safetyarxiv-cs-ai
30 Jun 2026
Research

Curvature-Guided Sheaf Diffusion for Unsupervised Community Detection on Heterophilic Graphs

DGX agent

arXiv:2606.30249v1 Announce Type: cross Abstract: Detecting communities in heterophilic graphs -- where connected nodes often belong to different classes -- is hard for unsupervised methods: classical

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

DGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

CW-B: Class Weighted Boosting Framework for Imbalance Resilient Multi Class Cardiac Phenotyping

DGX agent

arXiv:2606.29907v1 Announce Type: cross Abstract: Cardiac discharge phenotyping informs post-discharge treatment and follow-up, but real-world records are often incomplete and class-imbalanced, increa

applicationsarxiv-cs-ai
30 Jun 2026
Tutorials

CytoCLIP: Learning Cytoarchitectural Characteristics in Developing Human Brain Using Contrastive Language Image Pre-Training

DGX agent

arXiv:2601.12282v2 Announce Type: replace-cross Abstract: The functions of different regions of the human brain are closely linked to their distinct cytoarchitecture, which is defined by the spatial a

tutorialsarxiv-cs-ai
30 Jun 2026
← Previous
1…136137138139140…448
Next →