AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
28 May 2026

Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study

TutorialsDGX agent

arXiv:2512.15791v2 Announce Type: replace-cross Abstract: In Artificial Intelligence (AI), language models have gained significant importance due to the widespread adoption of systems capable of simul

EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter AdaptationTarget

Model ReleasesDGX agent

arXiv:2605.27390v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Model inference via a draft-then-verify paradigm, yet the output projection layer becomes a bottleneck

Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.28098v1 Announce Type: new Abstract: Multi-agent systems are increasingly deployed to support various tasks where agents interact to achieve individual and collective objectives. Although t

Explaining is Harder Than Predicting Alone: Evaluating Concept-based Explanations of MLLMs as ICL Visual Classifiers

ResearchDGX agent

arXiv:2605.28215v1 Announce Type: new Abstract: In-context learning (ICL) enables multimodal large language models (MLLMs) to classify images from a few labelled examples. Yet, how these models use th

Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts

ResearchDGX agent

arXiv:2605.28042v1 Announce Type: cross Abstract: Modern large language models (LLMs) achieve state-of-the-art machine translation performance, but they do so as broad generalists largely trained for

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL

AgentsDGX agent

arXiv:2605.28751v1 Announce Type: cross Abstract: Linear interpolation between fine-tuned checkpoints has been shown to trace the Pareto front between competing objectives, but whether extrapolative w

FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification

Model ReleasesDGX agent

arXiv:2604.04074v3 Announce Type: replace Abstract: LLM-based reviewing systems typically take only the manuscript as input, leaving literature and code-based claims hard to verify. We present FactRev

FD-RAG: Federated Dual-System Retrieval-Augmented Generation

Local AiDGX agent

arXiv:2605.27432v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has emerged as a paradigm for grounding large language models in external knowledge, yet most existing RAG system

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.28347v1 Announce Type: new Abstract: Multi-Label Recognition (MLR) based on Vision-Language Models (VLMs) aims to leverage their pre-trained knowledge to better adapt complex recognition sc

Fine-Tuned LLM as a Complementary Predictor Improving Ads System

ApplicationsDGX agent

arXiv:2605.27856v1 Announce Type: cross Abstract: Recommendation systems power engagement and monetization across feeds, ads, and short-video platforms, but translating the latest advances in Large La

FinTexTS: Financial Text-Paired Time-Series Dataset via Semantic-Based and Multi-Level Pairing

ResearchDGX agent

arXiv:2603.02702v3 Announce Type: replace Abstract: The financial domain involves a variety of important time-series problems. Recently, time-series analysis methods that jointly leverage textual and

FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales

Model ReleasesDGX agent

arXiv:2605.28174v1 Announce Type: cross Abstract: Foundation models offer a promising route to transferable remote sensing representations, but many current approaches depend on very large pretraining

FLUID: From Ephemeral IDs to Multimodal Semantic Codes for Industrial-Scale Livestreaming Recommendation

ApplicationsDGX agent

arXiv:2605.21832v2 Announce Type: replace Abstract: Modern recommender systems rely heavily on ID-based collaborative filtering: each item is represented by a unique ID embedding that accumulates coll

FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation

Model ReleasesDGX agent

arXiv:2605.27849v1 Announce Type: cross Abstract: Despite rapid progress in LLM-based code generation, existing models are predominantly trained on imperative languages, leaving functional programming

From AR to Diffusion: Efficiently Adapting Large Language Models with Strictly Causal and Elastic Horizons

SafetyDGX agent

arXiv:2605.27387v1 Announce Type: cross Abstract: Diffusion models promise efficient parallel text generation but rely on bidirectional attention, creating a structural mismatch with pre-trained Autor

From Causal Discovery to Dynamic Causal Inference in Neural Time Series

ApplicationsDGX agent

arXiv:2603.20980v2 Announce Type: replace-cross Abstract: Time-varying causal models provide a powerful framework for studying dynamic scientific systems, yet most existing approaches assume that the

From Detection to Mechanism: Cross-Attention Graph Neural Networks Enable Drug-Drug Interaction Type Prediction An Ablation Study with Acetylsalicylic Acid Validation

Model ReleasesDGX agent

arXiv:2605.27861v1 Announce Type: cross Abstract: Predicting whether two drugs interact (binary detection) is a substantially dif- ferent task from predicting the mechanism type of that interaction (m

From Fact Overwriting to Knowledge Evolution: Causal Editing via On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2605.28303v1 Announce Type: new Abstract: While Knowledge Editing (KE) enables efficient updates, its dominant Static Fact Overwriting paradigm treats LLMs as discrete databases, forcibly inject

From Instructor to Collaborator: What a 90-Participant Study Reveals about Human-Agent Collaboration in a Mobile Serious Game

AgentsDGX agent

arXiv:2605.27384v1 Announce Type: cross Abstract: This position paper reflects empirical data collected during my PhD from a large-scale within-subjects study (N = 90). The study compared a highly hum

From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets

Model ReleasesDGX agent

arXiv:2605.28359v1 Announce Type: new Abstract: Evaluating whether large language model (LLM) agents can profit in capital markets is increasingly framed as end-to-end trading: place an agent in a his

From Learning Resources to Competencies: LLM-Based Tagging with Evidence and Graph Constraints

SafetyDGX agent

arXiv:2605.28483v1 Announce Type: new Abstract: Linking learning resources to a structured competency framework is key to enabling competency-based search and curriculum analytics in Learning Manageme

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence

Model ReleasesDGX agent

arXiv:2605.28371v1 Announce Type: new Abstract: Industrial Prognostics and Health Management (PHM) provides a representative case study for a broader challenge in applied machine learning: translating

From Talking to Singing: A New Challenge for Audio-Visual Deepfake Detection

TutorialsDGX agent

arXiv:2605.27944v1 Announce Type: new Abstract: With rapid advances in audio-visual generative models, reliable forgery detection becomes increasingly critical. Existing methods for audio-visual deepf

Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2605.28500v1 Announce Type: cross Abstract: Large language models have shown impressive capabilities in code generation, yet they often produce functionally incorrect code. Uncertainty quantific

FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research

AgentsDGX agent

arXiv:2605.27864v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in finance, yet most existing work emphasizes trading signals or financial NLP tasks centered on p

Generalized Holographic Reduced Representations

ResearchDGX agent

arXiv:2405.09689v2 Announce Type: replace-cross Abstract: Hyperdimensional Computing (HDC) is a computationally and data-efficient paradigm that acts as a bridge between connectionist and symbolic app

Geometry-Correct Diffusion Posterior Sampling with Denoiser-Pullback Curvature Guidance and Manifold-Aligned Damping

ResearchDGX agent

arXiv:2605.27990v1 Announce Type: cross Abstract: Diffusion posterior sampling conditions diffusion priors on measurements, but data-consistency updates are typically scaled by hand-tuned guidance wei

Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

SafetyDGX agent

arXiv:2605.27970v1 Announce Type: new Abstract: While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometr

Global Policy-Space Response Oracles for Two-Player Zero-Sum Games

Model ReleasesDGX agent

arXiv:2605.28273v1 Announce Type: new Abstract: The Policy-Space Response Oracles (PSRO) framework scales equilibrium computation to large zero-sum games by iteratively expanding a restricted strategy

GONDOR to the Rescue: Satisficing Planning with Low Memory

ResearchDGX agent

arXiv:2605.28454v1 Announce Type: new Abstract: Greedy Best-First Search (GBFS) is the dominant approach for solving search problems where the goal can be estimated with a heuristic, such as planning,

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.27766v1 Announce Type: new Abstract: LLM safety evaluations predominantly test models in isolation, yet deployed AI agents increasingly operate within persistent social environments alongsi

GraD-IBD: Graph Representation Learning from Diagnosis Trajectories for Early Detection of Inflammatory Bowel Disease

ApplicationsDGX agent

arXiv:2605.27799v1 Announce Type: new Abstract: International Classification of Diseases (ICD) is a globally recognized coding system that records diagnostic events during each patient encounter, prov

Gradient Step Plug-and-Play Model for Dental Cone-Beam CT Reconstruction

ResearchDGX agent

arXiv:2605.28124v1 Announce Type: new Abstract: The goal of this work is to reduce the effect of photon noise in dental cone-beam CT reconstruction. We consider an inverse problem formulation and deve

GradientStabilizer:Fix the Norm, Not the Gradient

Model ReleasesDGX agent

arXiv:2502.17055v4 Announce Type: replace-cross Abstract: Training instability in modern deep learning systems is frequently triggered by rare but extreme gradient-norm spikes, which can induce oversi

Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills

Model ReleasesDGX agent

arXiv:2604.05333v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on reusable skills, and as they interact with personal applications, web browsers, and other interfaces, skill l

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

SafetyDGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer?

SafetyDGX agent

arXiv:2605.27494v1 Announce Type: cross Abstract: Modern retrieval-augmented generation(RAG) deployments increasingly rely on caching to reduce token cost and time-to-first-token(TTFT). Prefix-level K

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

SafetyDGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

Guaranteed Optimal Compositional Explanations for Neurons

SafetyDGX agent

arXiv:2511.20934v2 Announce Type: replace Abstract: Compositional explanations are a family of methods that aim to describe the spatial alignment between neurons' receptive field activations and conce

GUI Agents for Continual Game Generation

AgentsDGX agent

arXiv:2605.28258v1 Announce Type: cross Abstract: Generating a game is not the same as making one that can be played. Despite advances in code generation, existing approaches treat game generation as

Hallucination Behavior in Multimodal LLMs Across Agricultural Image Interpretation and Generation Tasks

Model ReleasesDGX agent

arXiv:2605.27595v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being rapidly adopted in agricultural imaging applications, ranging from crop interpretation to synthetic field image

Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Model ReleasesDGX agent

arXiv:2605.27922v1 Announce Type: new Abstract: LLM agents are increasingly deployed as executable systems that use tools, modify workspaces, and produce concrete artifacts. In such workflows, perform

HARP: Measuring Harm Amplification in Multi-Agent LLM Systems

Local AiDGX agent

arXiv:2605.27489v1 Announce Type: cross Abstract: Multi-agent LLM systems decompose workflows across agents, tools, shared context, memory, and decision gates. This modularity improves interpretabilit

HEAL: Resilient and Self-* Hub-based Learning

ResearchDGX agent

arXiv:2605.27475v1 Announce Type: cross Abstract: Decentralized learning enhances privacy, scalability, and fault tolerance by distributing data and computation across nodes. A popular approach is Fed

HEART: Achieving Timely Multi-Model Training for Vehicle-Edge-Cloud-Integrated Hierarchical Federated Learning

AgentsDGX agent

arXiv:2501.09934v3 Announce Type: replace-cross Abstract: The rapid growth of AI-enabled Internet of Vehicles (IoV) calls for efficient Machine Learning (ML) solutions that can handle high vehicular m

Heterogeneous Causal Discovery of Repeated Undesirable Health Outcomes

SafetyDGX agent

arXiv:2503.11477v2 Announce Type: replace Abstract: Understanding the factors that trigger or prevent undesirable health outcomes across patient subpopulations is essential for designing targeted inte

Heterogeneous Multi-Agent Modeling for Measurement and Network Analysis of the Data Service Market

AgentsDGX agent

arXiv:2605.27433v1 Announce Type: cross Abstract: With the increasing complexity of collaboration among various social entities and user demands, the factors affecting the stable development of the da

HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational Modeling

ResearchDGX agent

arXiv:2512.23959v3 Announce Type: replace-cross Abstract: Multi-step retrieval-augmented generation (RAG) has become a widely adopted strategy for enhancing large language models (LLMs) on tasks that

Hierarchical Prompt-Domain Control and Learning for Resource-Constrained Agentic Language Models

AgentsDGX agent

arXiv:2605.27703v1 Announce Type: new Abstract: Large Language Models are increasingly deployed inside agentic systems, where they must follow structured protocols, adapt to evolving states, and opera

High-Fidelity Industrial Crash Dynamics Prediction via Geometry-Aware Operator Learning with Memory-Efficient Low-Rank Attention

SafetyDGX agent

arXiv:2605.27758v1 Announce Type: cross Abstract: Automotive crashworthiness optimization remains a safety-critical challenge, requiring the management of large-scale nonlinear structural deformations

HO-SFL: Hybrid-Order Split Federated Learning with Backprop-Free Clients and Dimension-Free Aggregation

ResearchDGX agent

arXiv:2603.14773v2 Announce Type: replace-cross Abstract: Fine-tuning large models on edge devices is severely hindered by the memory-intensive backpropagation (BP) in standard frameworks like federat

How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving

Model ReleasesDGX agent

arXiv:2605.28302v1 Announce Type: cross Abstract: Modern large language model (LLM) inference has progressively disaggregated to keep pace with growing model sizes and tight TTFT and TPOT service-leve

How Much Can a Few Engine Moves Help? Quantifying Limited Cheating in Chess

ResearchDGX agent

arXiv:2601.05386v2 Announce Type: replace Abstract: Cheating in chess, by using advice from powerful software, has become a major problem, reaching the highest levels. As opposed to the large majority

How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks

SafetyDGX agent

arXiv:2605.27662v1 Announce Type: cross Abstract: Equivariant neural networks encode geometric symmetries by construction, yet they are often difficult to optimize and can underperform less constraine

HRBench: Benchmarking and Understanding Thinking-Mode Switch Strategies in Hybrid-Reasoning LLMs

ResearchDGX agent

arXiv:2605.28398v1 Announce Type: new Abstract: Hybrid-reasoning large language models (LLMs) expose explicit controls over reasoning effort, allowing users or systems to trade off answer quality agai

Human-AI Collaboration for Estimating Scientific Replicability

ResearchDGX agent

arXiv:2605.27394v1 Announce Type: cross Abstract: Determining whether published scientific findings can successfully be replicated is a long-standing challenge in the empirical sciences. Existing appr

Human-like in-group bias in instruction-tuned language model agents

SafetyDGX agent

arXiv:2605.28114v1 Announce Type: new Abstract: As autonomous AI agents are deployed in persistent, interacting networks -- coordinating tasks, routing resources, and accumulating reputational histori

HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning

Model ReleasesDGX agent

arXiv:2605.27724v1 Announce Type: cross Abstract: Imitation learning is a promising approach for training humanoid robots to both walk and manipulate, but it requires a large number of demonstrations,

Hurwitz Quaternion Multiplicative Quantization for KV Cache Compression

Model ReleasesDGX agent

arXiv:2605.27646v1 Announce Type: cross Abstract: We propose extbf{Hurwitz Quaternion Multiplicative Quantization (HQMQ)}, a extbf{calibration-free} method for KV cache compression of large language m

Hybrid Neural World Models

ResearchDGX agent

arXiv:2605.28317v1 Announce Type: cross Abstract: Neural surrogates promise large speedups over classical solvers for physical dynamics but fail silently at sharp dynamical events such as shocks, fron

← Previous
1…198199200201202…358
Next →