AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
14 May 2026

GRACE: Gradient-aligned Reasoning Data Curation for Efficient Post-training

SafetyDGX agent

arXiv:2605.13130v1 Announce Type: new Abstract: Existing reasoning data curation pipelines score whole samples, treating every intermediate step as equally valuable. In reality, steps within a trace c

GraphIP-Bench: How Hard Is It to Steal a Graph Neural Network, and Can We Stop It?

Model ReleasesDGX agent

arXiv:2605.12827v1 Announce Type: cross Abstract: Graph neural networks (GNNs) deployed as cloud services can be stolen through model-extraction attacks, which train a surrogate from query responses t

Grid-Orch: An LLM-Powered Orchestrator for Distribution Grid Simulation and Analytics

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.12728v1 Announce Type: cross Abstract: The power distribution engineering workforce faces a projected shortage of up to 1.5 million engineers by 2030, creating urgent demand for more access

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

Local AiDGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents

Model ReleasesDGX agent

arXiv:2601.18842v3 Announce Type: replace-cross Abstract: As GUI agents increasingly rely on screenshots to perceive and operate digital environments, they may inadvertently expose sensitive informati

Harnessing Agentic Evolution

AgentsDGX agent

arXiv:2605.13821v1 Announce Type: new Abstract: Agentic evolution has emerged as a powerful paradigm for improving programs, workflows, and scientific solutions by iteratively generating candidates, e

HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation

ResearchDGX agent

arXiv:2605.13586v1 Announce Type: cross Abstract: Generating controllable and physically plausible indoor scenes is a pivotal prerequisite for constructing high-fidelity simulation environments for em

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2605.13213v1 Announce Type: new Abstract: Multi-modal multi-agent systems (MM-MAS) have gained increasing attention for their capacity to enable complex reasoning and coordination across diverse

High-Rate Quantized Matrix Multiplication I

ResearchDGX agent

arXiv:2601.17187v2 Announce Type: replace-cross Abstract: This paper investigates the problem of quantized matrix multiplication (MatMul), which has become crucial for the efficient deployment of larg

High-Rate Quantized Matrix Multiplication II

Model ReleasesDGX agent

arXiv:2605.13768v1 Announce Type: cross Abstract: This is the second part of the work investigating quantized matrix multiplication (MatMul). In part I we considered the case of calibration-free quant

Higher-order Linear Attention

ResearchDGX agent

arXiv:2510.27258v2 Announce Type: replace-cross Abstract: The quadratic cost of scaled dot-product attention is a central obstacle to scaling autoregressive language models to long contexts. Linear-ti

History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions

SafetyDGX agent

arXiv:2605.13825v1 Announce Type: new Abstract: Frontier LLMs are increasingly deployed as agents that pick the next action after a long log of prior tool calls produced by the same or a different mod

HLS-Seek: QoR-Aware Code Generation for High-Level Synthesis via Proxy Comparative Reward Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.13536v1 Announce Type: cross Abstract: High-Level Synthesis (HLS) compiles algorithmic C/C++ descriptions into hardware, with Quality of Results (QoR) -- latency and resource utilization --

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

Model ReleasesDGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

How to Interpret Agent Behavior

Model ReleasesDGX agent

arXiv:2605.13625v1 Announce Type: new Abstract: Autonomous agents such as Claude Code and Codex now operate for hours or even days. Understanding their runtime behavior has become critical for downstr

Humanwashing -- It Should Leave You Feeling Dirty

SafetyDGX agent

arXiv:2605.13723v1 Announce Type: cross Abstract: The phrase 'human in the loop' is increasingly used to imply a sense of safety in relation to AI decision systems. It shouldn't. There are contexts wh

IdeaForge: A Knowledge Graph-Grounded Multi-Agent Framework for Cross-Methodology Innovation Analysis and Patent Claim Generation

AgentsDGX agent

arXiv:2605.13311v1 Announce Type: new Abstract: Current AI-assisted innovation systems typically apply a single ideation methodology (such as TRIZ or Design Thinking) using sequential prompt-based wor

Identifying AI Web Scrapers Using Canary Tokens

AgentsDGX agent

arXiv:2605.13706v1 Announce Type: cross Abstract: From pre-training to query-time augmentation, web-scraped data helps to improve the quality and contextual relevancy of content generated by large lan

Improving Classifier-Free Guidance of Flow Matching via Manifold Projection

SafetyDGX agent

arXiv:2601.21892v2 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is a widely used technique for controllable generation in diffusion and flow-based models. Despite its empirica

Improving Code Translation with Syntax-Guided and Semantic-aware Preference Optimization

SafetyDGX agent

arXiv:2605.13229v1 Announce Type: new Abstract: LLMs have shown immense potential for code translation, yet they often struggle to ensure both syntactic correctness and semantic consistency. While pre

Improving Diffusion Posterior Samplers with Lagged Temporal Corrections for Image Restoration

SafetyDGX agent

arXiv:2605.12573v1 Announce Type: cross Abstract: Diffusion-based posterior sampling (PS) is a leading framework for imaging inverse problems, combining learned priors with measurement constraints. Ye

Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling

SafetyDGX agent

arXiv:2605.13801v1 Announce Type: cross Abstract: As generative AI models such as large language models (LLMs) become more pervasive, ensuring the safety, robustness, and overall trustworthiness of th

In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores

SafetyDGX agent

arXiv:2605.12530v1 Announce Type: cross Abstract: LLM fairness should be evaluated through in-situ conversational behavior rather than standardized-test Q&A benchmarks. We show that the standardized-t

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

Model ReleasesDGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models

Model ReleasesDGX agent

arXiv:2605.13338v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) are increasingly integrated into systems requiring reliable multi-step inference, yet this growing dependence exposes ne

Information as Maximum-Caliber Deviation: A bridge between Integrated Information Theory and the Free Energy Principle

ResearchDGX agent

arXiv:2605.12536v1 Announce Type: cross Abstract: The Free Energy Principle (FEP) is a leading framework for mathematically modeling self-organization and learning, while Integrated Information Theory

Inline Critic Steers Image Editing

TutorialsDGX agent

arXiv:2605.12724v1 Announce Type: cross Abstract: Instruction-based image editing exhibits heterogeneous difficulty not only across cases but also across regions of an image, motivating refinement app

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

SafetyDGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

Is a Picture Worth a Thousand Words? Adaptive Multimodal Fact-Checking with Visual Evidence Necessity

ResearchDGX agent

arXiv:2604.04692v2 Announce Type: replace-cross Abstract: Automated fact-checking is a crucial task that supports a responsible information ecosystem. While recent research has progressed from text-on

'It became a self-fulfilling prophecy': How Lived Experiences are Entangled with AI Predictions in Menstrual Cycle Tracking Apps

ResearchDGX agent

arXiv:2605.13261v1 Announce Type: cross Abstract: In menstrual cycle tracking apps (MCTAs), AI-based predictions and insights have become increasingly popular. These features enable users to receive p

It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

ApplicationsDGX agent

arXiv:2605.13245v1 Announce Type: new Abstract: Language models can produce convincing scientific analyses, but repeated generations on the same data do not guarantee the same result. A researcher may

Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance

Model ReleasesDGX agent

arXiv:2603.02175v4 Announce Type: replace-cross Abstract: Instruction-based video editing has witnessed rapid progress, yet current methods often struggle with precise visual control, as natural langu

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving

ApplicationsDGX agent

arXiv:2605.13734v1 Announce Type: cross Abstract: LLMs are widely adopted in production, pushing inference systems to their limits. Disaggregated LLM serving (e.g., PD separation and KV state disaggre

Language-Based Agent Control

AgentsDGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

Model ReleasesDGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

Language Model Networks: Supervision-Efficient Learning through Dense Communication

ApplicationsDGX agent

arXiv:2505.12741v2 Announce Type: replace Abstract: Language models are increasingly used not only as standalone predictors but also as components in larger inference systems, from test-time reasoning

Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

SafetyDGX agent

arXiv:2605.12729v1 Announce Type: cross Abstract: Large language models are increasingly being used to support network operations (NetOps) and artificial intelligence for IT operations (AIOps), includ

Latent-Augmented Discrete Diffusion Models

ResearchDGX agent

arXiv:2510.18114v3 Announce Type: replace-cross Abstract: Discrete diffusion models have emerged as a powerful class of models and a promising route to fast language generation, but practical implemen

LeanSearch v2: Global Premise Retrieval for Lean 4 Theorem Proving

Model ReleasesDGX agent

arXiv:2605.13137v1 Announce Type: cross Abstract: Proving theorems in Lean 4 often requires identifying a scattered set of library lemmas whose joint use enables a concise proof -- a task we call glob

Learning Local Constraints for Reinforcement-Learned Content Generators

TutorialsDGX agent

arXiv:2605.13570v1 Announce Type: new Abstract: Constraint-based game content generators that learn local constraints from existing content, such as Wave Function Collapse (WFC), can generate visually

Learning to Decide with AI Assistance under Human-Alignment

SafetyDGX agent

arXiv:2605.12646v1 Announce Type: cross Abstract: It is widely agreed that when AI models assist decision-makers in high-stakes domains by predicting an outcome of interest, they should communicate th

Learning Transferable Latent User Preferences for Human-Aligned Decision Making

SafetyDGX agent

arXiv:2605.12682v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as reasoning modules in many applications. While they are efficient in certain tasks, LLMs often stru

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

Model ReleasesDGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

LMPath: Language-Mediated Priors and Path Generation for Aerial Exploration

AgentsDGX agent

arXiv:2605.13782v1 Announce Type: cross Abstract: Traditional autonomous UAV search missions rely on geometric coverage patterns that ignore the semantic context of the target, leading to significant

Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models

Local AiDGX agent

arXiv:2605.13538v1 Announce Type: cross Abstract: Personally Identifiable Information (PII) redaction usually replaces detected entities with placeholder tokens such as [PERSON], destroying the downst

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

Model ReleasesDGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

Macro-Action Based Multi-Agent Instruction Following through Value Cancellation

SafetyDGX agent

arXiv:2605.12655v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) in real-world use cases may need to adapt to external natural language instructions that interrupt ongoing beh

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn

Model ReleasesDGX agent

arXiv:2605.13511v1 Announce Type: cross Abstract: In-context learning (ICL) adapts large language models (LLMs) to new tasks by conditioning on demonstrations in the prompt without parameter updates.

MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning

AgentsDGX agent

arXiv:2605.13037v1 Announce Type: new Abstract: Current interactive LLM agents rely on goal-conditioned stepwise planning, where environmental understanding is acquired reactively during execution rat

Margin-calibrated Classifier Guidance for Property-driven Synthesis Planning

ResearchDGX agent

arXiv:2605.13101v1 Announce Type: cross Abstract: Synthesis planning seeks an efficient sequence of chemical reactions that produce a target molecule. Typically, a pretrained single-step (autoregressi

McCast: Memory-Guided Latent Drift Correction for Long-Horizon Precipitation Nowcasting

ResearchDGX agent

arXiv:2605.13197v1 Announce Type: cross Abstract: Existing precipitation nowcasting methods typically adopt an autoregressive formulation, where future states are predicted from previous outputs. Howe

Mechanism Plausibility in Generative Agent-Based Modeling

AgentsDGX agent

arXiv:2605.12824v1 Announce Type: cross Abstract: Large language models (LLMs) can generate high-level diverse phenomena without explicitly programmed rules. This capability has led to their adoption

Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models

ResearchDGX agent

arXiv:2601.21975v2 Announce Type: replace Abstract: Recent work identifies a stated-revealed (SvR) preference gap in language models (LMs): a mismatch between the values models endorse and the choices

MinT: Managed Infrastructure for Training and Serving Millions of LLMs

SafetyDGX agent

arXiv:2605.13779v1 Announce Type: cross Abstract: We present MindLab Toolkit (MinT), a managed infrastructure system for Low-Rank Adaptation (LoRA) post-training and online serving. MinT targets a set

MLGIB: Multi-Label Graph Information Bottleneck for Expressive and Robust Message Passing

ResearchDGX agent

arXiv:2605.13126v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) suffer from over-squashing in deep message passing, where information from exponentially growing neighborhoods is compres

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence

Model ReleasesDGX agent

arXiv:2605.12703v1 Announce Type: cross Abstract: We introduce MMCL-Bench, a benchmark for multimodal context learning: learning task-local rules, procedures, and empirical patterns from visual or mix

MMSkills: Towards Multimodal Skills for General Visual Agents

AgentsDGX agent

arXiv:2605.13527v1 Announce Type: new Abstract: Reusable skills have become a core substrate for improving agent capabilities, yet most existing skill packages encode reusable behavior primarily as te

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

Modeling Heterophily in Multiplex Graphs: An Adaptive Approach for Node Classification

TutorialsDGX agent

arXiv:2605.12699v1 Announce Type: cross Abstract: Existing multiplex graph models often assume homophily, where connected nodes tend to belong to the same class or share similar attributes. Consequent

Moltbook Moderation: Uncovering Hidden Intent Through Multi-Turn Dialogue

AgentsDGX agent

arXiv:2605.12856v1 Announce Type: new Abstract: The emergence of multi-agent systems introduces novel moderation challenges that extend beyond content filtering. Agents with {em malicious intent} may

← Previous
1…258259260261262…358
Next →