AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

HLS-Seek: QoR-Aware Code Generation for High-Level Synthesis via Proxy Comparative Reward Reinforcement Learning

DGX agent

arXiv:2605.13536v1 Announce Type: cross Abstract: High-Level Synthesis (HLS) compiles algorithmic C/C++ descriptions into hardware, with Quality of Results (QoR) -- latency and resource utilization --

model-releasesarxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

DGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

How to Interpret Agent Behavior

DGX agent

arXiv:2605.13625v1 Announce Type: new Abstract: Autonomous agents such as Claude Code and Codex now operate for hours or even days. Understanding their runtime behavior has become critical for downstr

model-releasesarxiv-cs-ai
14 May 2026
Safety

Humanwashing -- It Should Leave You Feeling Dirty

DGX agent

arXiv:2605.13723v1 Announce Type: cross Abstract: The phrase 'human in the loop' is increasingly used to imply a sense of safety in relation to AI decision systems. It shouldn't. There are contexts wh

safetyarxiv-cs-ai
14 May 2026
Agents

IdeaForge: A Knowledge Graph-Grounded Multi-Agent Framework for Cross-Methodology Innovation Analysis and Patent Claim Generation

DGX agent

arXiv:2605.13311v1 Announce Type: new Abstract: Current AI-assisted innovation systems typically apply a single ideation methodology (such as TRIZ or Design Thinking) using sequential prompt-based wor

agentsarxiv-cs-ai
14 May 2026
Agents

Identifying AI Web Scrapers Using Canary Tokens

DGX agent

arXiv:2605.13706v1 Announce Type: cross Abstract: From pre-training to query-time augmentation, web-scraped data helps to improve the quality and contextual relevancy of content generated by large lan

agentsarxiv-cs-ai
14 May 2026
Safety

Improving Classifier-Free Guidance of Flow Matching via Manifold Projection

DGX agent

arXiv:2601.21892v2 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is a widely used technique for controllable generation in diffusion and flow-based models. Despite its empirica

safetyarxiv-cs-ai
14 May 2026
Safety

Improving Code Translation with Syntax-Guided and Semantic-aware Preference Optimization

DGX agent

arXiv:2605.13229v1 Announce Type: new Abstract: LLMs have shown immense potential for code translation, yet they often struggle to ensure both syntactic correctness and semantic consistency. While pre

safetyarxiv-cs-ai
14 May 2026
Safety

Improving Diffusion Posterior Samplers with Lagged Temporal Corrections for Image Restoration

DGX agent

arXiv:2605.12573v1 Announce Type: cross Abstract: Diffusion-based posterior sampling (PS) is a leading framework for imaging inverse problems, combining learned priors with measurement constraints. Ye

safetyarxiv-cs-ai
14 May 2026
Safety

Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling

DGX agent

arXiv:2605.13801v1 Announce Type: cross Abstract: As generative AI models such as large language models (LLMs) become more pervasive, ensuring the safety, robustness, and overall trustworthiness of th

safetyarxiv-cs-ai
14 May 2026
Safety

In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores

DGX agent

arXiv:2605.12530v1 Announce Type: cross Abstract: LLM fairness should be evaluated through in-situ conversational behavior rather than standardized-test Q&A benchmarks. We show that the standardized-t

safetyarxiv-cs-ai
14 May 2026
Model Releases

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

DGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models

DGX agent

arXiv:2605.13338v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) are increasingly integrated into systems requiring reliable multi-step inference, yet this growing dependence exposes ne

model-releasesarxiv-cs-ai
14 May 2026
Research

Information as Maximum-Caliber Deviation: A bridge between Integrated Information Theory and the Free Energy Principle

DGX agent

arXiv:2605.12536v1 Announce Type: cross Abstract: The Free Energy Principle (FEP) is a leading framework for mathematically modeling self-organization and learning, while Integrated Information Theory

researcharxiv-cs-ai
14 May 2026
Tutorials

Inline Critic Steers Image Editing

DGX agent

arXiv:2605.12724v1 Announce Type: cross Abstract: Instruction-based image editing exhibits heterogeneous difficulty not only across cases but also across regions of an image, motivating refinement app

tutorialsarxiv-cs-ai
14 May 2026
Safety

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

DGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

safetyarxiv-cs-ai
14 May 2026
Research

Is a Picture Worth a Thousand Words? Adaptive Multimodal Fact-Checking with Visual Evidence Necessity

DGX agent

arXiv:2604.04692v2 Announce Type: replace-cross Abstract: Automated fact-checking is a crucial task that supports a responsible information ecosystem. While recent research has progressed from text-on

researcharxiv-cs-ai
14 May 2026
Research

'It became a self-fulfilling prophecy': How Lived Experiences are Entangled with AI Predictions in Menstrual Cycle Tracking Apps

DGX agent

arXiv:2605.13261v1 Announce Type: cross Abstract: In menstrual cycle tracking apps (MCTAs), AI-based predictions and insights have become increasingly popular. These features enable users to receive p

researcharxiv-cs-ai
14 May 2026
Applications

It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

DGX agent

arXiv:2605.13245v1 Announce Type: new Abstract: Language models can produce convincing scientific analyses, but repeated generations on the same data do not guarantee the same result. A researcher may

applicationsarxiv-cs-ai
14 May 2026
Model Releases

Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance

DGX agent

arXiv:2603.02175v4 Announce Type: replace-cross Abstract: Instruction-based video editing has witnessed rapid progress, yet current methods often struggle with precise visual control, as natural langu

model-releasesarxiv-cs-ai
14 May 2026
Applications

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving

DGX agent

arXiv:2605.13734v1 Announce Type: cross Abstract: LLMs are widely adopted in production, pushing inference systems to their limits. Disaggregated LLM serving (e.g., PD separation and KV state disaggre

applicationsarxiv-cs-ai
14 May 2026
Agents

Language-Based Agent Control

DGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

agentsarxiv-cs-ai
14 May 2026
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Applications

Language Model Networks: Supervision-Efficient Learning through Dense Communication

DGX agent

arXiv:2505.12741v2 Announce Type: replace Abstract: Language models are increasingly used not only as standalone predictors but also as components in larger inference systems, from test-time reasoning

applicationsarxiv-cs-ai
14 May 2026
Safety

Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

DGX agent

arXiv:2605.12729v1 Announce Type: cross Abstract: Large language models are increasingly being used to support network operations (NetOps) and artificial intelligence for IT operations (AIOps), includ

safetyarxiv-cs-ai
14 May 2026
Research

Latent-Augmented Discrete Diffusion Models

DGX agent

arXiv:2510.18114v3 Announce Type: replace-cross Abstract: Discrete diffusion models have emerged as a powerful class of models and a promising route to fast language generation, but practical implemen

researcharxiv-cs-ai
14 May 2026
Model Releases

LeanSearch v2: Global Premise Retrieval for Lean 4 Theorem Proving

DGX agent

arXiv:2605.13137v1 Announce Type: cross Abstract: Proving theorems in Lean 4 often requires identifying a scattered set of library lemmas whose joint use enables a concise proof -- a task we call glob

model-releasesarxiv-cs-ai
14 May 2026
Tutorials

Learning Local Constraints for Reinforcement-Learned Content Generators

DGX agent

arXiv:2605.13570v1 Announce Type: new Abstract: Constraint-based game content generators that learn local constraints from existing content, such as Wave Function Collapse (WFC), can generate visually

tutorialsarxiv-cs-ai
14 May 2026
Safety

Learning to Decide with AI Assistance under Human-Alignment

DGX agent

arXiv:2605.12646v1 Announce Type: cross Abstract: It is widely agreed that when AI models assist decision-makers in high-stakes domains by predicting an outcome of interest, they should communicate th

safetyarxiv-cs-ai
14 May 2026
Safety

Learning Transferable Latent User Preferences for Human-Aligned Decision Making

DGX agent

arXiv:2605.12682v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as reasoning modules in many applications. While they are efficient in certain tasks, LLMs often stru

safetyarxiv-cs-ai
14 May 2026
Model Releases

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

DGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

model-releasesarxiv-cs-ai
14 May 2026
Agents

LMPath: Language-Mediated Priors and Path Generation for Aerial Exploration

DGX agent

arXiv:2605.13782v1 Announce Type: cross Abstract: Traditional autonomous UAV search missions rely on geometric coverage patterns that ignore the semantic context of the target, leading to significant

agentsarxiv-cs-ai
14 May 2026
Local Ai

Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models

DGX agent

arXiv:2605.13538v1 Announce Type: cross Abstract: Personally Identifiable Information (PII) redaction usually replaces detected entities with placeholder tokens such as [PERSON], destroying the downst

local-aiarxiv-cs-ai
14 May 2026
Model Releases

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

DGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

model-releasesarxiv-cs-ai
14 May 2026
Safety

Macro-Action Based Multi-Agent Instruction Following through Value Cancellation

DGX agent

arXiv:2605.12655v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) in real-world use cases may need to adapt to external natural language instructions that interrupt ongoing beh

safetyarxiv-cs-ai
14 May 2026
Model Releases

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn

DGX agent

arXiv:2605.13511v1 Announce Type: cross Abstract: In-context learning (ICL) adapts large language models (LLMs) to new tasks by conditioning on demonstrations in the prompt without parameter updates.

model-releasesarxiv-cs-ai
14 May 2026
Agents

MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning

DGX agent

arXiv:2605.13037v1 Announce Type: new Abstract: Current interactive LLM agents rely on goal-conditioned stepwise planning, where environmental understanding is acquired reactively during execution rat

agentsarxiv-cs-ai
14 May 2026
Research

Margin-calibrated Classifier Guidance for Property-driven Synthesis Planning

DGX agent

arXiv:2605.13101v1 Announce Type: cross Abstract: Synthesis planning seeks an efficient sequence of chemical reactions that produce a target molecule. Typically, a pretrained single-step (autoregressi

researcharxiv-cs-ai
14 May 2026
Research

McCast: Memory-Guided Latent Drift Correction for Long-Horizon Precipitation Nowcasting

DGX agent

arXiv:2605.13197v1 Announce Type: cross Abstract: Existing precipitation nowcasting methods typically adopt an autoregressive formulation, where future states are predicted from previous outputs. Howe

researcharxiv-cs-ai
14 May 2026
Agents

Mechanism Plausibility in Generative Agent-Based Modeling

DGX agent

arXiv:2605.12824v1 Announce Type: cross Abstract: Large language models (LLMs) can generate high-level diverse phenomena without explicitly programmed rules. This capability has led to their adoption

agentsarxiv-cs-ai
14 May 2026
Research

Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models

DGX agent

arXiv:2601.21975v2 Announce Type: replace Abstract: Recent work identifies a stated-revealed (SvR) preference gap in language models (LMs): a mismatch between the values models endorse and the choices

researcharxiv-cs-ai
14 May 2026
Safety

MinT: Managed Infrastructure for Training and Serving Millions of LLMs

DGX agent

arXiv:2605.13779v1 Announce Type: cross Abstract: We present MindLab Toolkit (MinT), a managed infrastructure system for Low-Rank Adaptation (LoRA) post-training and online serving. MinT targets a set

safetyarxiv-cs-ai
14 May 2026
Research

MLGIB: Multi-Label Graph Information Bottleneck for Expressive and Robust Message Passing

DGX agent

arXiv:2605.13126v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) suffer from over-squashing in deep message passing, where information from exponentially growing neighborhoods is compres

researcharxiv-cs-ai
14 May 2026
Model Releases

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence

DGX agent

arXiv:2605.12703v1 Announce Type: cross Abstract: We introduce MMCL-Bench, a benchmark for multimodal context learning: learning task-local rules, procedures, and empirical patterns from visual or mix

model-releasesarxiv-cs-ai
14 May 2026
Agents

MMSkills: Towards Multimodal Skills for General Visual Agents

DGX agent

arXiv:2605.13527v1 Announce Type: new Abstract: Reusable skills have become a core substrate for improving agent capabilities, yet most existing skill packages encode reusable behavior primarily as te

agentsarxiv-cs-ai
14 May 2026
Model Releases

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

DGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

model-releasesarxiv-cs-ai
14 May 2026
Tutorials

Modeling Heterophily in Multiplex Graphs: An Adaptive Approach for Node Classification

DGX agent

arXiv:2605.12699v1 Announce Type: cross Abstract: Existing multiplex graph models often assume homophily, where connected nodes tend to belong to the same class or share similar attributes. Consequent

tutorialsarxiv-cs-ai
14 May 2026
Agents

Moltbook Moderation: Uncovering Hidden Intent Through Multi-Turn Dialogue

DGX agent

arXiv:2605.12856v1 Announce Type: new Abstract: The emergence of multi-agent systems introduces novel moderation challenges that extend beyond content filtering. Agents with {em malicious intent} may

agentsarxiv-cs-ai
14 May 2026
← Previous
1…323324325326327…448
Next →