AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Agents

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

DGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

agentsarxiv-cs-ai
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

DGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

researcharxiv-cs-ai
3 Jun 2026
Safety

Brief Announcement: Generative Markov Model for Distributed Computing Systems

DGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

safetyarxiv-cs-ai
3 Jun 2026
Safety

Building Better Activation Oracles

DGX agent

arXiv:2606.02609v1 Announce Type: cross Abstract: Activation Oracles (AOs) are promising methods for interpreting residual stream activations. However, current AOs face important issues, such as hallu

safetyarxiv-cs-ai
3 Jun 2026
Research

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

DGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

researcharxiv-cs-ai
3 Jun 2026
Applications

Building Trust in Black-box Optimization: A Comprehensive Framework for Explainability

DGX agent

arXiv:2410.14573v2 Announce Type: replace-cross Abstract: Optimizing costly black-box functions within a constrained evaluation budget presents significant challenges in many real-world applications.

applicationsarxiv-cs-ai
3 Jun 2026
Applications

Calibrating Urban Traffic Simulation from Sparse Road Observations via Genetic Optimization

DGX agent

arXiv:2606.03823v1 Announce Type: new Abstract: Urban traffic simulation is a critical tool for infrastructure planning, including the placement of electric vehicle charging stations. However, realist

applicationsarxiv-cs-ai
3 Jun 2026
Model Releases

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

DGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

DGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

agentsarxiv-cs-ai
3 Jun 2026
Agents

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

DGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

agentsarxiv-cs-ai
3 Jun 2026
Tutorials

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

DGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

tutorialsarxiv-cs-ai
3 Jun 2026
Model Releases

Causal Neural Probabilistic Circuits

DGX agent

arXiv:2603.01372v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) enhance the interpretability of end-to-end neural networks by introducing a layer of concepts and predicting

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Causal Preference Elicitation

DGX agent

arXiv:2602.01483v2 Announce Type: replace-cross Abstract: We propose causal preference elicitation, a Bayesian framework for expert-in-the-loop causal discovery that actively queries local edge relati

model-releasesarxiv-cs-ai
3 Jun 2026
Research

CauTion: Knowing When to Trust LLMs for Ensemble Causal Discovery

DGX agent

arXiv:2606.03602v1 Announce Type: cross Abstract: Causal discovery from observational data remains challenging due to the fundamental limitations of purely statistical methods, such as statistical dis

researcharxiv-cs-ai
3 Jun 2026
Model Releases

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

DGX agent

arXiv:2606.02802v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong natural-language reasoning abilities for clinical decision support, but struggle to effectively model struct

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

CL-DMDF:Dynamic Multimodal Data Fusion Model Based on Contrastive Learning

DGX agent

arXiv:2606.02659v1 Announce Type: cross Abstract: Multimodal data fusion involves integrating and analyzing information from multiple modalities to uncover latent correlations and complementary patter

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

DGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Closed-Loop Molecular Design with Calibrated Deference

DGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

agentsarxiv-cs-ai
3 Jun 2026
Research

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

DGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

researcharxiv-cs-ai
3 Jun 2026
Agents

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

DGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

agentsarxiv-cs-ai
3 Jun 2026
Research

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

DGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

researcharxiv-cs-ai
3 Jun 2026
Agents

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

DGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

DGX agent

arXiv:2606.03650v1 Announce Type: cross Abstract: Choosing or ranking language models for a specific application is hardest when no task-specific labeled data exists, and standard public benchmarks ca

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Collab-REC: An LLM-based Agentic Framework for Balancing Recommendations in Tourism

DGX agent

arXiv:2508.15030v5 Announce Type: replace Abstract: We propose COLLAB-REC, a multi-agent framework designed to counteract popularity bias and improve diversity in tourism recommendations. In our setup

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

CoMPAS3D: A Dataset and Benchmark for Interactive Motion

DGX agent

arXiv:2507.19684v2 Announce Type: replace-cross Abstract: Socially interactive humanoid robots must engage with humans through their bodies, adapting in real time to a partner's movement, intent, and

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates

DGX agent

arXiv:2606.03029v1 Announce Type: cross Abstract: A core goal of computational social science is to discover interpretable differences in how language varies across outcomes of interest, such as polit

applicationsarxiv-cs-ai
3 Jun 2026
Research

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

DGX agent

arXiv:2606.03827v1 Announce Type: cross Abstract: In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typ

researcharxiv-cs-ai
3 Jun 2026
Safety

Consistency Training Can Entrench Misalignment

DGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

safetyarxiv-cs-ai
3 Jun 2026
Safety

Constitutional On-Policy Safe Distillation

DGX agent

arXiv:2606.03089v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) has emerged as an efficient post-training paradigm by using a teacher conditioned on privileged information to prov

safetyarxiv-cs-ai
3 Jun 2026
Safety

ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL

DGX agent

arXiv:2606.03017v1 Announce Type: cross Abstract: Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and

safetyarxiv-cs-ai
3 Jun 2026
Research

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

DGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Cosmos 3: Omnimodal World Models for Physical AI

DGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Cost-Aware Query Routing in RAG: Empirical Analysis of Retrieval Depth Tradeoffs

DGX agent

arXiv:2606.02581v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) faces a fundamental three-way tension: deeper retrieval improves factual grounding but inflates token costs and e

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

Coupled Local and Global World Models for Efficient First Order RL

DGX agent

arXiv:2602.06219v2 Announce Type: replace-cross Abstract: World models offer a promising avenue for more faithfully capturing complex dynamics, including contacts and non-rigidity, as well as complex

local-aiarxiv-cs-ai
3 Jun 2026
Safety

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

DGX agent

arXiv:2606.03435v1 Announce Type: new Abstract: Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readout

safetyarxiv-cs-ai
3 Jun 2026
Hardware

CRAM-ER: Error-Resilient Spintronic Computational Random Access Memory for Scalable In-Memory Computation

DGX agent

arXiv:2606.02781v1 Announce Type: cross Abstract: Deep neural networks (DNNs) have achieved state-of-the-art performance across diverse domains. However, typical Von Neumann compute paradigms face sev

hardwarearxiv-cs-ai
3 Jun 2026
Model Releases

Cross-Lingual Token Arbitrage: Optimizing Code Agent Context Windows via Local LLM Preprocessing

DGX agent

arXiv:2606.03618v1 Announce Type: new Abstract: AI-assisted coding agents are bottlenecked by input-token cost. Two pathologies of raw human input drive much of this overhead: tokenization inefficienc

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Cross-Modal Contrastive Learning of ECG and Angiography Representations for Severe Stenosis Classification

DGX agent

arXiv:2606.02605v1 Announce Type: cross Abstract: Coronary artery stenosis is a common cardiovascular disease, with severe, untreated cases posing significant risks of heart attack. Although coronary

researcharxiv-cs-ai
3 Jun 2026
Safety

Curriculum-Adapted Robust Reinforcement Learning for UAV Deconfliction in Adversarial Environments

DGX agent

arXiv:2506.21129v2 Announce Type: replace-cross Abstract: Autonomous unmanned aerial vehicles (UAVs) increasingly rely on reinforcement learning (RL) for navigation. However, global navigation satelli

safetyarxiv-cs-ai
3 Jun 2026
Safety

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting

DGX agent

arXiv:2606.02640v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to large language model (LLM) safety because they exploit feedback from auxiliary judge models to i

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

DDOR: Delta Debugging for Explainable Overrefusal Testing and Repair

DGX agent

arXiv:2606.03601v1 Announce Type: cross Abstract: While safety alignment and guardrails help large language models (LLMs) avoid harmful outputs, they can also induce overrefusal, i.e., unwarranted rej

local-aiarxiv-cs-ai
3 Jun 2026
Research

Decomposing how prompting steers behavior

DGX agent

arXiv:2606.03093v1 Announce Type: new Abstract: Prompting steers large language models (LLMs) and vision-language models (VLMs) without weight updates, but it remains unclear how instruction changes r

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Decoupled Smart Contract Audits: Lightweight LLM Framework via Distillation and Aggregation

DGX agent

arXiv:2606.03128v1 Announce Type: cross Abstract: Smart contracts face critical security challenges that require thorough auditing in decentralized web services. While Large Language Models (LLMs) hav

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

DELTAMEM: Incremental Experience Memory for LLM Agents via Residual Trees

DGX agent

arXiv:2606.03083v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents increasingly rely on memory to learn from experiences over continual interactions. However, storing experiences

agentsarxiv-cs-ai
3 Jun 2026
Research

DeMuon: A Decentralized Muon for Matrix Optimization over Graphs

DGX agent

arXiv:2510.01377v2 Announce Type: replace-cross Abstract: In this paper, we propose DeMuon, a method for decentralized matrix optimization over a given communication topology. DeMuon incorporates matr

researcharxiv-cs-ai
3 Jun 2026
Model Releases

DeskCraft: Benchmarking Desktop Agents on Professional Workflows and Human-in-the-Loop Collaboration

DGX agent

arXiv:2606.03103v1 Announce Type: new Abstract: Real-world professional desktop workflows in specialized creative and engineering software unfold over long horizons and often require human-in-the-loop

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Diagnosing Knowledge Gaps in LLM Tool Use: An Agentic Benchmark for Novel API Acquisition

DGX agent

arXiv:2606.03657v1 Announce Type: new Abstract: Large language models for code generation often need to use APIs that are absent from their pretraining data. This requires more than recalling a functi

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

Distill-then-Replace: Efficient Task-Specific Hybrid Attention Model Construction

DGX agent

arXiv:2601.11667v2 Announce Type: replace-cross Abstract: Transformer architectures deliver state-of-the-art accuracy via dense full-attention, but their quadratic time and memory complexity with resp

local-aiarxiv-cs-ai
3 Jun 2026
← Previous
1…209210211212213…452
Next →