AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
14 Apr 2026

Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control

SafetyDGX agent

arXiv:2601.06540v2 Announce Type: replace-cross Abstract: This paper proposes a novel reinforcement learning framework, named Self-Organizing Dual-buffer Adaptive Clustering Experience Replay (SODACER

SemaCDR: LLM-Powered Transferable Semantics for Cross-Domain Sequential Recommendation

ApplicationsDGX agent

arXiv:2604.09551v1 Announce Type: cross Abstract: Cross-domain recommendation (CDR) addresses the data sparsity and cold-start problems in the target domain by leveraging knowledge from data-rich sour

SemaClaw: A Step Towards General-Purpose Personal AI Agents through Harness Engineering

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.11548v1 Announce Type: new Abstract: The rise of OpenClaw in early 2026 marks the moment when millions of users began deploying personal AI agents into their daily lives, delegating tasks r

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

Model ReleasesDGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

Semantic Manipulation Localization

Model ReleasesDGX agent

arXiv:2604.10132v1 Announce Type: cross Abstract: Image Manipulation Localization (IML) aims to identify edited regions in an image. However, with the increasing use of modern image editing and genera

Semantic Segmentation Algorithm Based on Light Field and LiDAR Fusion

AgentsDGX agent

arXiv:2510.06687v2 Announce Type: replace-cross Abstract: Semantic segmentation serves as a cornerstone of scene understanding in autonomous driving but continues to face significant challenges under

Seven simple steps for log analysis in AI systems

ResearchDGX agent

arXiv:2604.09563v1 Announce Type: new Abstract: AI systems produce large volumes of logs as they interact with tools and users. Analysing these logs can help understand model capabilities, propensitie

ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values

ResearchDGX agent

arXiv:2604.11200v1 Announce Type: cross Abstract: Changes in input distribution can induce shifts in the average predictions of machine learning models. Such prediction shifts may impact downstream bu

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

Model ReleasesDGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

SHE: Stepwise Hybrid Examination Reinforcement Learning Framework for E-commerce Search Relevance

SafetyDGX agent

arXiv:2510.07972v3 Announce Type: replace Abstract: Query-product relevance prediction is vital for AI-driven e-commerce, yet current LLM-based approaches face a dilemma: SFT and DPO struggle with lon

Should We be Pedantic About Reasoning Errors in Machine Translation?

ResearchDGX agent

arXiv:2604.09890v1 Announce Type: cross Abstract: Across multiple language pairings (English o {Spanish, French, German, Mandarin, Japanese, Urdu, Cantonese}), we find reasoning errors in translation.

SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors

Model ReleasesDGX agent

arXiv:2510.17516v4 Announce Type: replace-cross Abstract: Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only i

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

SafetyDGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

SLALOM: Simulation Lifecycle Analysis via Longitudinal Observation Metrics for Social Simulation

SafetyDGX agent

arXiv:2604.11466v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a potentially-transformative path forward for generative social science but face a critical crisis of validity

SMART: When is it Actually Worth Expanding a Speculative Tree?

Model ReleasesDGX agent

arXiv:2604.09731v1 Announce Type: cross Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying a branching tree of draft tokens in a single target-model forward p

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

Model ReleasesDGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

Spatial Competence Benchmark

Model ReleasesDGX agent

arXiv:2604.09594v1 Announce Type: new Abstract: Spatial competence is the quality of maintaining a consistent internal representation of an environment and using it to infer discrete structure and pla

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

Speaking to No One: Ontological Dissonance and the Double Bind of Conversational AI

SafetyDGX agent

arXiv:2604.10833v1 Announce Type: cross Abstract: Recent reports indicate that sustained interaction with conversational artificial intelligence (AI) systems can, in a small subset of users, contribut

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

Model ReleasesDGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

Model ReleasesDGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

Model ReleasesDGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

Model ReleasesDGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

Steered LLM Activations are Non-Surjective

SafetyDGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

STORM: End-to-End Referring Multi-Object Tracking in Videos

Model ReleasesDGX agent

arXiv:2604.10527v1 Announce Type: cross Abstract: Referring multi-object tracking (RMOT) is a task of associating all the objects in a video that semantically match with given textual queries or refer

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

HardwareDGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

StreetDesignAI: A Multi-Persona Evaluation System for Inclusive Infrastructure Design

ResearchDGX agent

arXiv:2601.15671v2 Announce Type: replace-cross Abstract: Designing cycling infrastructure requires balancing the competing needs of diverse user groups, yet designers often struggle to anticipate how

StyleBench: Evaluating thinking styles in Large Language Models

Model ReleasesDGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

Subargument Argumentation Frameworks: Separating Direct Conflict from Structural Dependency

ResearchDGX agent

arXiv:2601.12038v3 Announce Type: replace Abstract: Dung's abstract argumentation frameworks model acceptability solely in terms of an attack relation, thereby conflating two conceptually distinct asp

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

ResearchDGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

SafetyDGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning

TutorialsDGX agent

arXiv:2604.10228v1 Announce Type: new Abstract: Current multimodal models often suffer from shallow reasoning, leading to errors caused by incomplete or inconsistent thought processes. To address this

SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context

AgentsDGX agent

arXiv:2604.11716v1 Announce Type: new Abstract: Prior representative ReAct-style approaches in autonomous Software Engineering (SWE) typically lack the explicit System-2 reasoning required for deep an

SynthAgent: Adapting Web Agents with Synthetic Supervision

ResearchDGX agent

arXiv:2511.06101v3 Announce Type: replace-cross Abstract: Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored

Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo

Model ReleasesDGX agent

arXiv:2604.11563v1 Announce Type: cross Abstract: Providing AI agents with reliable long-term memory that does not hallucinate remains an open problem. Current approaches to memory for LLM agents -- s

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

ApplicationsDGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

Tail-Aware Information-Theoretic Generalization for RLHF and SGLD

Model ReleasesDGX agent

arXiv:2604.10727v1 Announce Type: cross Abstract: Classical information-theoretic generalization bounds typically control the generalization gap through KL-based mutual information and therefore rely

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

SafetyDGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

ResearchDGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings

SafetyDGX agent

arXiv:2604.10849v1 Announce Type: cross Abstract: Federated learning (FL) performance is highly sensitive to heterogeneity across clients, yet practitioners lack reliable methods to anticipate how a f

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

Model ReleasesDGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation

SafetyDGX agent

arXiv:2507.22767v3 Announce Type: replace-cross Abstract: Obtaining human-readable symbolic formulas via genetic programming-based symbolic distillation of a deep neural network trained on a target da

Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings

Model ReleasesDGX agent

arXiv:2305.14299v3 Announce Type: replace-cross Abstract: Learning high quality sentence embeddings from dialogues has drawn increasing attentions as it is essential to solve a variety of dialogue-ori

The Amazing Agent Race: Strong Tool Users, Weak Navigators

Model ReleasesDGX agent

arXiv:2604.10261v1 Announce Type: new Abstract: Existing tool-use benchmarks for LLM agents are overwhelmingly linear: our analysis of six benchmarks shows 55 to 100% of instances are simple chains of

The Augmentation Trap: AI Productivity and the Cost of Cognitive Offloading

SafetyDGX agent

arXiv:2604.03501v2 Announce Type: replace-cross Abstract: Experimental evidence confirms that AI tools raise worker productivity, but also that sustained use can erode the expertise on which those gai

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

Model ReleasesDGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

The Deployment Gap in AI Media Detection: Platform-Aware and Visually Constrained Adversarial Evaluation

ApplicationsDGX agent

arXiv:2604.09706v1 Announce Type: cross Abstract: Recent AI media detectors report near-perfect performance under clean laboratory evaluation, yet their robustness under realistic deployment condition

The Geometry of Knowing: From Possibilistic Ignorance to Probabilistic Certainty -- A Measure-Theoretic Framework for Epistemic Convergence

ResearchDGX agent

arXiv:2604.09614v1 Announce Type: new Abstract: This paper develops a measure-theoretic framework establishing when and how a possibilistic representation of incomplete knowledge contracts into a prob

The Missing Knowledge Layer in Cognitive Architectures for AI Agents

Model ReleasesDGX agent

arXiv:2604.11364v1 Announce Type: new Abstract: The two most influential cognitive architecture frameworks for AI agents, CoALA [21] and JEPA [12], both lack an explicit Knowledge layer with its own p

The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise

ResearchDGX agent

arXiv:2604.09780v1 Announce Type: new Abstract: Mixture of Experts (MoEs) are now ubiquitous in large language models, yet the mechanisms behind their 'expert specialization' remain poorly understood.

The Paradox of Professional Input: How Expert Collaboration with AI Systems Shapes Their Future Value

SafetyDGX agent

arXiv:2504.12654v1 Announce Type: cross Abstract: This perspective paper examines a fundamental paradox in the relationship between professional expertise and artificial intelligence: as domain expert

The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping

SafetyDGX agent

arXiv:2604.11297v1 Announce Type: cross Abstract: Despite the success of reinforcement learning for large language models, a common failure mode is reduced sampling diversity, where the policy repeate

The Phantom of PCIe: Constraining Generative Artificial Intelligences for Practical Peripherals Trace Synthesizing

ResearchDGX agent

arXiv:2411.06376v3 Announce Type: replace-cross Abstract: Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. The development of

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

SafetyDGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

The Rise and Fall of G in AGI

Model ReleasesDGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems

Model ReleasesDGX agent

arXiv:2604.11309v1 Announce Type: cross Abstract: Large Language Models (LLMs) face prominent security risks from jailbreaking, a practice that manipulates models to bypass built-in security constrain

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

ResearchDGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

← Previous
1…336337338339340…350
Next →