AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
20 May 2026

Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target

SafetyDGX agent

arXiv:2605.18899v1 Announce Type: cross Abstract: Generative LLM-based recommenders (LLM-Rec) require continual post-deployment updates, yet deployment logs provide only policy-shaped contextual bandi

DOTRAG: Retrieval-Time Reasoning Along Paths

TutorialsDGX agent

arXiv:2605.18760v1 Announce Type: cross Abstract: Graph Retrieval-Augmented Generation (GraphRAG) is dominated by a retrieve-then-reason paradigm, where context is retrieved using heuristics and then

Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding

Local AiDGX agent

arXiv:2605.20104v1 Announce Type: cross Abstract: Speculative decoding (SD) accelerates large language model inference by leveraging a draft-then-verify paradigm. To maximize the acceptance rate, rece


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dr.LLM: Dynamic Layer Routing in LLMs

ResearchDGX agent

arXiv:2510.12773v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) process every token through all layers of a transformer stack, causing wasted computation on simple queries and i

DualView: Adaptive Local-Global Fusion for Multi-Hop Document Reranking

Model ReleasesDGX agent

arXiv:2605.18767v1 Announce Type: cross Abstract: Multi-hop question answering requires aggregating information from multiple documents, a critical capability for knowledge-intensive applications. A f

Dynamic Model Merging Made Slim

Model ReleasesDGX agent

arXiv:2605.18904v1 Announce Type: cross Abstract: Model merging enables the reuse of fine-tuned models without joint training or access to original data. Dynamic merging further improves flexibility b

Efficient Elicitation of Collective Disagreements

ResearchDGX agent

arXiv:2605.19521v1 Announce Type: new Abstract: We analyze the structure of the disagreement among a population of voters over a set of alternatives. Surveys typically ask either for pairwise comparis

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

Model ReleasesDGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs

Model ReleasesDGX agent

arXiv:2605.19559v1 Announce Type: cross Abstract: The rapid development of Multimodal Large Language Models (MLLMs) has led to growing interest in egocentric video understanding, specifically the abil

Eliminating Inductive Bias in Reward Models with Information-Theoretic Guidance

Model ReleasesDGX agent

arXiv:2512.23461v2 Announce Type: replace-cross Abstract: Reward models (RMs) are essential in reinforcement learning from human feedback (RLHF) to align large language models (LLMs) with human values

Embedding by Elicitation: Dynamic Representations for Bayesian Optimization of System Prompts

Model ReleasesDGX agent

arXiv:2605.19093v1 Announce Type: new Abstract: System prompts are a central control mechanism in modern AI systems, shaping behavior across conversations, tasks, and user populations. Yet they are di

EmbGen: Teaching with Reassembled Corpora

ResearchDGX agent

arXiv:2605.19394v1 Announce Type: cross Abstract: Adapting small instruction-tuned models to specialized domains often relies on supervised fine-tuning (SFT) on curated instruction-response examples,

Emergence of Frontier Superposition: Mobius attractor and Cascade Supervision

Model ReleasesDGX agent

arXiv:2605.18820v1 Announce Type: cross Abstract: Superposition allows Transformers to reason in depth, carrying an entire reasoning frontier in parallel through a bounded-depth forward pass instead o

EMO-BOOST: Emotion-Augmented Audio-Visual Features for Improved Generalization in Deepfake Detection

ResearchDGX agent

arXiv:2605.19630v1 Announce Type: new Abstract: With every advancement in generative AI models, forensics is under increasing pressure. The constant emergence of new generation techniques makes it imp

Enabling Real-Time Colonoscopic Polyp Segmentation on Commodity CPUs via Ultra-Lightweight Architecture

Model ReleasesDGX agent

arXiv:2602.04381v2 Announce Type: replace-cross Abstract: Real-time polyp segmentation is essential for early colorectal cancer detection, yet clinical deployment remains blocked by GPU dependency. We

EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design

Model ReleasesDGX agent

arXiv:2605.19743v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly applied to engineering design tasks, yet existing evaluation frameworks do not adequately address mul

Entry-level guide to the use of large language models for medical research

Model ReleasesDGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

SafetyDGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

Euclidean Embedding of Data Using Local Distances

ResearchDGX agent

arXiv:2605.19243v1 Announce Type: cross Abstract: We study the problem of recovering a globally consistent Euclidean embedding of data, given only a local distance graph and propose a method that opti

EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly

Model ReleasesDGX agent

arXiv:2605.18872v1 Announce Type: cross Abstract: Robotic assembly in architectural construction faces a persistent bottleneck: existing planners are either highly specialized, requiring prohibitive r

EVA-0: Test-Time Model Evolution with Only Two Forward Passes per Sample

Model ReleasesDGX agent

arXiv:2605.18867v1 Announce Type: cross Abstract: Test-time model evolution offers a promising way for deployed models to improve from unlabeled test-time experience, yet most existing methods depend

Evaluating the Utility of Personal Health Records in Personalized Health AI

Model ReleasesDGX agent

arXiv:2605.18937v1 Announce Type: new Abstract: Patient-managed Personal Health Records (PHRs) promises to empower patients to better understand their health; but information in the record is complex,

EviTrack: Selection over Sampling for Delayed Disambiguation

Model ReleasesDGX agent

arXiv:2605.19283v1 Announce Type: cross Abstract: Sequential prediction is challenging in regimes of delayed disambiguation, where early observations are ambiguous and multiple latent explanations rem

Exact Linear Attention

SafetyDGX agent

arXiv:2605.18848v1 Announce Type: cross Abstract: This paper introduces Exact Linear Attention (ELA), a mechanism that achieves linear computational complexity for Transformer attention by leveraging

ExECG: An Explainable AI Framework for ECG models

ResearchDGX agent

arXiv:2605.19258v1 Announce Type: cross Abstract: Deep learning has enabled ECG diagnostic models with strong performance in tasks such as arrhythmia classification and abnormality detection. However,

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

Model ReleasesDGX agent

arXiv:2605.19826v1 Announce Type: new Abstract: Operators of safety-critical industrial processes increasingly rely on digital twins to screen control interventions, but such simulators rarely carry c

Exploring and Developing a Pre-Model Safeguard with Draft Models

SafetyDGX agent

arXiv:2605.19321v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment remains vulnerable to jailbreak attacks that elicit unsafe responses, motivating pre-model and post-model guards.

Extreme Self-Preference in Language Models

SafetyDGX agent

arXiv:2509.26464v2 Announce Type: replace Abstract: Self-preference is a fundamental feature of biological organisms. Since large language models (LLMs) lack sentience, they might be expected to avoid

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

AgentsDGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

Fast and Featureless Node Representation Learning with Partial Pairwise Supervision

Model ReleasesDGX agent

arXiv:2605.19916v1 Announce Type: cross Abstract: We introduce Contrastive FUSE, a fast and unified framework for scalable node representation learning in graphs with partially available pairwise node

Fast and Lightweight Backdoor Detection via Head Random Probing

Model ReleasesDGX agent

arXiv:2605.18908v1 Announce Type: cross Abstract: Deep neural networks (DNNs) remain critically vulnerable to backdoor attacks. Existing post-training detectors often require clean or surrogate data,

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models

SafetyDGX agent

arXiv:2410.15362v2 Announce Type: replace-cross Abstract: Aligned Large Language Models (LLMs) have attracted significant attention for their safety, particularly in the context of jailbreak attacks t

Features have life history. And we should care

ResearchDGX agent

arXiv:2605.18789v1 Announce Type: cross Abstract: Features in language models have life history: they emerge, persist, and die during training, yet the importance of that history remains largely unexp

Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models

Model ReleasesDGX agent

arXiv:2605.18824v1 Announce Type: cross Abstract: Evaluation of foundation models often rely on aggregate scores from benchmarks that lack comprehensive coverage and metadata for a fine-grained evalua

Fine-tuning Large Language Model for Automated Algorithm Design

Model ReleasesDGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding

Model ReleasesDGX agent

arXiv:2605.19846v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities in general video understanding, yet they often struggle with the fine-grained

First-Passage Prediction of Grokking Delay: ACalibrated Law under AdamW with Causal Validation

Model ReleasesDGX agent

arXiv:2605.18845v1 Announce Type: cross Abstract: We give the first quantitative prediction of grokking delay under AdamW. Treating the delay as a first-passage time, we derive a closed-form law T_gro

Flash PD-SSM: Memory-Optimized Structured Sparse State-Space Models

ResearchDGX agent

arXiv:2605.19150v1 Announce Type: cross Abstract: State-space models (SSMs) face a fundamental trade-off between efficiency and expressivity that is mainly dictated by the structure of the model's tra

FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives

Model ReleasesDGX agent

arXiv:2605.18866v1 Announce Type: cross Abstract: Reconstructing continuous flow fields from sparse surface-mounted sensors is central to aerodynamic design, flow control, and digital-twin instrumenta

FLUXtrapolation: A benchmark on extrapolating ecosystem fluxes

Model ReleasesDGX agent

arXiv:2605.19812v1 Announce Type: cross Abstract: We introduce FLUXtrapolation, a benchmark for extrapolating ecosystem fluxes under progressively harder distribution shifts. Ecosystem fluxes are cent

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

SafetyDGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

FormalASR: End-to-End Spoken Chinese to Formal Text

Local AiDGX agent

arXiv:2605.19266v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems are typically optimized for verbatim transcription, which preserves disfluencies, filler words, and informa

FreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstruction

TutorialsDGX agent

arXiv:2601.18993v2 Announce Type: replace-cross Abstract: Camera redirection aims to replay a dynamic scene from a single monocular video under a user-specified camera trajectory. However, large-angle

From Intent to AI Pipelines: A Controlled Agentic Framework for Non-AI Expert Scientists

AgentsDGX agent

arXiv:2605.18764v1 Announce Type: cross Abstract: Artificial Intelligence (AI) pipelines have become integral to modern research, supporting fields such as Medical Sciences, Agriculture, and Social Sc

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

Model ReleasesDGX agent

arXiv:2605.19824v1 Announce Type: new Abstract: Recent attempts to support high-level scene interpretation and planning in Autonomous Vehicles (AVs) using ensembles of Large Language Models (LLMs) and

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails

SafetyDGX agent

arXiv:2510.13727v2 Announce Type: replace Abstract: Generative AI systems are increasingly assisting and acting on behalf of end users in practical settings, from digital shopping assistants to next-g

From SGD to Muon: Adaptive Optimization via Schatten-p Norms

Model ReleasesDGX agent

arXiv:2605.19781v1 Announce Type: new Abstract: Modern optimizers, like Muon, impose matrix-wise geometry constraints on their updates. These matrix-wise constraints can be unified under Linear Minimi

From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation

Model ReleasesDGX agent

arXiv:2605.18865v1 Announce Type: cross Abstract: Self-attention serves as the core foundation of large-scale transformer pretraining, but its quadratic token interaction cost makes inference expensiv

GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems

HardwareDGX agent

arXiv:2605.19945v1 Announce Type: cross Abstract: Mixture-of-Expert (MoE) models enable efficient inference by employing smaller experts and activating only a subset of them per token. MoE serving eng

GenAI-FDIA: Physics-Informed Generative Models for False Data Injection Attacks

ResearchDGX agent

arXiv:2605.18873v1 Announce Type: cross Abstract: Training and evaluating false data injection attack (FDIA) detectors for power systems is constrained by data scarcity. Operational grid measurements

Generative Auto-Bidding with Unified Modeling and Exploration

SafetyDGX agent

arXiv:2605.19457v1 Announce Type: new Abstract: Automated bidding is central to modern digital advertising. Early rule-based methods lacked adaptability, while subsequent Reinforcement Learning approa

Generative-Evaluative Agreement: A Necessary Validity Criterion for LLM-Enabled Adaptive Assessment

SafetyDGX agent

arXiv:2605.19529v1 Announce Type: new Abstract: When the same LLM generates assessment items, simulates student responses, and scores them, the validation loop is self-referential. We introduce Genera

Generative Recursive Reasoning

ResearchDGX agent

arXiv:2605.19376v1 Announce Type: new Abstract: How should future neural reasoning systems implement extended computation? Recursive Reasoning Models (RRMs) offer a promising alternative to autoregres

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

Model ReleasesDGX agent

arXiv:2605.20006v1 Announce Type: new Abstract: Geospatial reasoning requires solving image-grounded problems over the complex spatial structure of a scene. However, developing this capability is hind

GOAL: Graph-based Objective-Aligned Diffusion Solvers for Dynamic Multi-Objective Optimization

ResearchDGX agent

arXiv:2605.19119v1 Announce Type: cross Abstract: Existing neural combinatorial optimization solvers frame solution search as imitation of optimal decisions, inherently limiting their utility to singl

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South

Local AiDGX agent

arXiv:2605.19190v1 Announce Type: cross Abstract: Despite the global deployment of text-to-image (T2I) models, their safety frameworks are largely calibrated to a Western-centric default, creating sig

Governing Evolving Memory in LLM Agents: Risks, Mechanisms, and the Stability and Safety Governed Memory (SSGM) Framework

SafetyDGX agent

arXiv:2603.11768v2 Announce Type: replace Abstract: Long-term memory has emerged as a foundational component of autonomous Large Language Model (LLM) agents, enabling continuous adaptation, lifelong m

GRALIS: A Unified Canonical Framework for Linear Attribution Methods via Riesz Representation

Local AiDGX agent

arXiv:2605.05480v2 Announce Type: replace-cross Abstract: The main XAI attribution methods for deep neural networks -- GradCAM, SHAP, LIME, Integrated Gradients -- operate on separate theoretical foun

Graph-Driven Cross-Industry Real-Time Monitoring Framework for Anti-Money Laundering Detection in Converged Mobility-Energy Supply Chain Networks

ApplicationsDGX agent

arXiv:2605.18844v1 Announce Type: cross Abstract: With the deep integration of the travel and energy industries, cross-industry supply chain finance has gradually become a high-risk field of hidden mo

GraphPINE: Graph Importance Propagation for Interpretable Drug Response Prediction

ResearchDGX agent

arXiv:2504.05454v2 Announce Type: replace-cross Abstract: Explainability is necessary for many tasks in biomedical research. Recent explainability methods have focused on attention, gradient, and Shap

← Previous
1…227228229230231…358
Next →