AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
20 Apr 2026

Puppets or partners? Governing cyborg propaganda in the digital public square

SafetyDGX agent

arXiv:2602.13088v2 Announce Type: replace-cross Abstract: The distinction between genuine grassroots activism and automated influence operations is collapsing. While contemporary policy debates priori

QuantSightBench: Evaluating LLM Quantitative Forecasting with Prediction Intervals

Model ReleasesDGX agent

arXiv:2604.15859v1 Announce Type: cross Abstract: Forecasting has become a natural benchmark for reasoning under uncertainty. Yet existing evaluations of large language models remain limited to judgme

Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.15464v1 Announce Type: cross Abstract: Large Language Model (LLM) deployment is increasingly shifting to cost-efficient accelerators like Google's Tensor Processing Units (TPUs), prioritizi

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams

Model ReleasesDGX agent

arXiv:2604.15994v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at recognizing individual visual elements and reasoning over simple linear diagrams. However, when faced

Reading Between the Lines: The One-Sided Conversation Problem

ApplicationsDGX agent

arXiv:2511.03056v2 Announce Type: replace-cross Abstract: Conversational AI is constrained in many real-world settings where only one side of a dialogue can be recorded, such as telemedicine, call cen

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

Model ReleasesDGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability

SafetyDGX agent

arXiv:2604.16106v1 Announce Type: cross Abstract: The Project of AI is a world-building endeavor, wherein those who fund and develop AI systems both operate through and seek to sustain networks of pow

RelativeFlow: Taming Medical Image Denoising Learning with Noisy Reference

ResearchDGX agent

arXiv:2604.15459v1 Announce Type: cross Abstract: Medical image denoising (MID) lacks absolutely clean images for supervision, leading to a noisy reference problem that fundamentally limits denoising

Rethinking the Necessity of Adaptive Retrieval-Augmented Generation through the Lens of Adaptive Listwise Ranking

ResearchDGX agent

arXiv:2604.15621v1 Announce Type: cross Abstract: Adaptive Retrieval-Augmented Generation aims to mitigate the interference of extraneous noise by dynamically determining the necessity of retrieving s

Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2510.10959v3 Announce Type: replace-cross Abstract: Reasoning ability has become a defining capability of Large Language Models (LLMs), with Reinforcement Learning with Verifiable Rewards (RLVR)

Revisiting the Uniform Information Density Hypothesis in LLM Reasoning

Local AiDGX agent

arXiv:2510.06953v3 Announce Type: replace Abstract: The Uniform Information Density (UID) hypothesis proposes that effective communication is achieved by maintaining a stable flow of information. In t

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

SafetyDGX agent

arXiv:2604.15577v1 Announce Type: cross Abstract: Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vect

Robust Multispectral Semantic Segmentation under Missing or Full Modalities via Structured Latent Projection

SafetyDGX agent

arXiv:2604.15856v1 Announce Type: cross Abstract: Multimodal remote sensing data provide complementary information for semantic segmentation, but in real-world deployments, some modalities may be unav

Robust Synchronisation for Federated Learning in The Face of Correlated Device Failure

SafetyDGX agent

arXiv:2604.16090v1 Announce Type: cross Abstract: Probabilistic Synchronous Parallel (PSP) is a technique in distributed learning systems to reduce synchronization bottlenecks by sampling a subset of

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

Model ReleasesDGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility

SafetyDGX agent

arXiv:2604.16033v1 Announce Type: cross Abstract: Buildings account for approximately 40% of global energy consumption, and with the growing share of intermittent renewable energy sources, enabling de

Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agents

SafetyDGX agent

arXiv:2504.13541v5 Announce Type: replace-cross Abstract: Training resource-constrained autonomous agents on multiple tasks simultaneously is crucial for adapting to diverse real-world environments. R

Scaling Behaviors of LLM Reinforcement Learning Post-Training: An Empirical Study in Mathematical Reasoning

ResearchDGX agent

arXiv:2509.25300v4 Announce Type: replace-cross Abstract: While scaling laws for large language models (LLMs) during pre-training have been extensively studied, their behavior under reinforcement lear

SCRIPT: Implementing an Intelligent Tutoring System for Programming in a German University Context

ApplicationsDGX agent

arXiv:2604.16117v1 Announce Type: cross Abstract: Practice and extensive exercises are essential in programming education. Intelligent tutoring systems (ITSs) are a viable option to provide individual

SecureRouter: Encrypted Routing for Efficient Secure Inference

ApplicationsDGX agent

arXiv:2604.15499v1 Announce Type: cross Abstract: Cryptographically secure neural network inference typically relies on secure computing techniques such as Secure Multi-Party Computation (MPC), enabli

Security Threat Modeling for Emerging AI-Agent Protocols: A Comparative Analysis of MCP, A2A, Agora, and ANP

AgentsDGX agent

arXiv:2602.11327v2 Announce Type: replace-cross Abstract: The rapid development of the AI agent communication protocols, including the Model Context Protocol (MCP), Agent2Agent (A2A), Agora, and Agent

Seed1.8 Model Card: Towards Generalized Real-World Agency

Model ReleasesDGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

Seeing the imagined: a latent functional alignment in visual imagery decoding from fMRI data

Model ReleasesDGX agent

arXiv:2604.15374v1 Announce Type: cross Abstract: Recent progress in visual brain decoding from fMRI has been enabled by large-scale datasets such as the Natural Scenes Dataset (NSD) and powerful diff

Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories

ResearchDGX agent

arXiv:2308.10562v2 Announce Type: cross Abstract: The field of Computer Vision (CV) is increasingly shifting towards ``high-level'' visual sensemaking tasks, yet the exact nature of these tasks remain

SegMix:Shuffle-based Feedback Learning for Semantic Segmentation of Pathology Images

ResearchDGX agent

arXiv:2604.15777v1 Announce Type: cross Abstract: Segmentation is a critical task in computational pathology, as it identifies areas affected by disease or abnormal growth and is essential for diagnos

Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting

SafetyDGX agent

arXiv:2604.15794v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degra

Sequential KV Cache Compression via Probabilistic Language Tries: Beyond the Per-Vector Shannon Limit

ResearchDGX agent

arXiv:2604.15356v1 Announce Type: cross Abstract: Recent work on KV cache quantization, culminating in TurboQuant, has approached the Shannon entropy limit for per-vector compression of transformer ke

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

Model ReleasesDGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

Social-JEPA: Emergent Geometric Isomorphism

Model ReleasesDGX agent

arXiv:2603.02263v2 Announce Type: replace-cross Abstract: World models compress rich sensory streams into compact latent codes that anticipate future observations. We let separate agents acquire such

SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.16022v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from text processors to autonomous agents, evaluating their social reasoning in embodied multi-agent settings

SocialWise: LLM-Agentic Conversation Therapy for Individuals with Autism Spectrum Disorder to Enhance Communication Skills

AgentsDGX agent

arXiv:2604.15347v1 Announce Type: cross Abstract: Autism Spectrum Disorder (ASD) affects more than 75 million people worldwide. However, scalable support for practicing everyday conversation is scarce

Spectral Tempering for Embedding Compression in Dense Passage Retrieval

Local AiDGX agent

arXiv:2603.19339v2 Announce Type: replace-cross Abstract: Dimensionality reduction is critical for deploying dense retrieval systems at scale, yet mainstream post-hoc methods face a fundamental trade-

SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification

ResearchDGX agent

arXiv:2604.15711v1 Announce Type: cross Abstract: Pathological diagnosis is highly reliant on image analysis, where Regions of Interest (ROIs) serve as the primary basis for diagnostic evidence, while

Stein Variational Black-Box Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2604.15837v1 Announce Type: new Abstract: Combinatorial black-box optimization in high-dimensional settings demands a careful trade-off between exploiting promising regions of the search space a

StoSignSGD: Unbiased Structural Stochasticity Fixes SignSGD for Training Large Language Models

ResearchDGX agent

arXiv:2604.15416v1 Announce Type: cross Abstract: Sign-based optimization algorithms, such as SignSGD, have garnered significant attention for their remarkable performance in distributed learning and

Structured Abductive-Deductive-Inductive Reasoning for LLMs via Algebraic Invariants

ResearchDGX agent

arXiv:2604.15727v1 Announce Type: new Abstract: Large language models exhibit systematic limitations in structured logical reasoning: they conflate hypothesis generation with verification, cannot dist

Struggle Premium : How Human Effort and Imperfection Drive Perceived Value in the Age of AI

ResearchDGX agent

arXiv:2604.15324v1 Announce Type: cross Abstract: As AI enters creative practice, audiences face growing uncertainty in judging authenticity and value. This study examines the Struggle Premium, the ad

Stylistic-STORM (ST-STORM) : Perceiving the Semantic Nature of Appearance

AgentsDGX agent

arXiv:2604.16086v1 Announce Type: cross Abstract: One of the dominant paradigms in self-supervised learning (SSL), illustrated by MoCo or DINO, aims to produce robust representations by capturing feat

Subjective and Objective Quality-of-Experience Evaluation Study for Live Video Streaming

Model ReleasesDGX agent

arXiv:2409.17596v2 Announce Type: replace-cross Abstract: In recent years, live video streaming has gained widespread popularity across various social media platforms. Quality of experience (QoE), whi

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

SafetyDGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

SWNet: A Cross-Spectral Network for Camouflaged Weed Detection

ResearchDGX agent

arXiv:2604.16147v1 Announce Type: cross Abstract: This paper presents SWNet, a bimodal end-to-end cross-spectral network specifically engineered for the detection of camouflaged weeds in dense agricul

Symbolic Guardrails for Domain-Specific Agents: Stronger Safety and Security Guarantees Without Sacrificing Utility

SafetyDGX agent

arXiv:2604.15579v1 Announce Type: cross Abstract: AI agents that interact with their environments through tools enable powerful applications, but in high-stakes business settings, unintended actions c

Synthetic data in cryptocurrencies using generative models

ResearchDGX agent

arXiv:2604.16182v1 Announce Type: cross Abstract: Data plays a fundamental role in consolidating markets, services, and products in the digital financial ecosystem. However, the use of real data, espe

TabularMath: Understanding Math Reasoning over Tables with Large Language Models

Model ReleasesDGX agent

arXiv:2505.19563v4 Announce Type: replace Abstract: Mathematical reasoning has long been a key benchmark for evaluating large language models. Although substantial progress has been made on math word

'Taking Stock at FAccT': Using Participatory Design to Co-Create a Vision for the Fairness, Accountability and Transparency Community

SafetyDGX agent

arXiv:2604.16224v1 Announce Type: cross Abstract: As a relatively new forum, ACM FAccT has become a key space for activists and scholars to critically examine emerging AI and ML technologies. It bring

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

HardwareDGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

Targeted Exploration via Unified Entropy Control for Reinforcement Learning

SafetyDGX agent

arXiv:2604.14646v2 Announce Type: replace Abstract: Recent advances in reinforcement learning (RL) have improved the reasoning capabilities of large language models (LLMs) and vision-language models (

Technically Love: The Evolution of Human-AI Romance Discourse on Reddit

ApplicationsDGX agent

arXiv:2604.15333v1 Announce Type: cross Abstract: Human-AI romantic relationships are increasingly common, yet little is understood about how public discourse around them emerges and shifts over time.

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

SafetyDGX agent

arXiv:2604.15383v1 Announce Type: cross Abstract: Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a temporal smoothing bias: transient

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings

ApplicationsDGX agent

arXiv:2604.15460v1 Announce Type: cross Abstract: The rapid evolution of Large Language Models (LLMs) has made them powerful tools for enhancing student writing. This study explores the extent and lim

The Illusion of Equivalence: Systematic FP16 Divergence in KV-Cached Autoregressive Inference

Model ReleasesDGX agent

arXiv:2604.15409v1 Announce Type: cross Abstract: KV caching is a ubiquitous optimization in autoregressive transformer inference, long presumed to be numerically equivalent to cache-free computation.

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning

SafetyDGX agent

arXiv:2603.01283v2 Announce Type: replace Abstract: Deployed RL agents operate in closed-loop systems where reliable performance depends on maintaining coherent coupling between observations, actions,

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

Model ReleasesDGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

Model ReleasesDGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

The Semi-Executable Stack: Agentic Software Engineering and the Expanding Scope of SE

AgentsDGX agent

arXiv:2604.15468v1 Announce Type: cross Abstract: AI-based systems, currently driven largely by LLMs and tool-using agentic harnesses, are increasingly discussed as a possible threat to software engin

The Synthetic Media Shift: Tracking the Rise, Virality, and Detectability of AI-Generated Multimodal Misinformation

ResearchDGX agent

arXiv:2604.15372v1 Announce Type: cross Abstract: As generative AI advances, the distinction between authentic and synthetic media is increasingly blurred, challenging the integrity of online informat

The threat of analytic flexibility in using large language models to simulate human data

HardwareDGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

The World Leaks the Future: Harness Evolution for Future Prediction Agents

AgentsDGX agent

arXiv:2604.15719v1 Announce Type: new Abstract: Many consequential decisions must be made before the relevant outcome is known. Such problems are commonly framed as future prediction, where an LLM age

To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates

ResearchDGX agent

arXiv:2604.15344v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into design and development workflows, yet decisions about their use are rarely binary or pur

← Previous
1…323324325326327…354
Next →