AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

DGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

DGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

DGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

DGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

STORM: End-to-End Referring Multi-Object Tracking in Videos

DGX agent

arXiv:2604.10527v1 Announce Type: cross Abstract: Referring multi-object tracking (RMOT) is a task of associating all the objects in a video that semantically match with given textual queries or refer

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

DGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

hardwarearxiv-cs-ai
14 Apr 2026
Research

StreetDesignAI: A Multi-Persona Evaluation System for Inclusive Infrastructure Design

DGX agent

arXiv:2601.15671v2 Announce Type: replace-cross Abstract: Designing cycling infrastructure requires balancing the competing needs of diverse user groups, yet designers often struggle to anticipate how

researcharxiv-cs-ai
14 Apr 2026
Model Releases

StyleBench: Evaluating thinking styles in Large Language Models

DGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Subargument Argumentation Frameworks: Separating Direct Conflict from Structural Dependency

DGX agent

arXiv:2601.12038v3 Announce Type: replace Abstract: Dung's abstract argumentation frameworks model acceptability solely in terms of an attack relation, thereby conflating two conceptually distinct asp

researcharxiv-cs-ai
14 Apr 2026
Research

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

DGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

researcharxiv-cs-ai
14 Apr 2026
Safety

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

DGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

safetyarxiv-cs-ai
14 Apr 2026
Tutorials

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning

DGX agent

arXiv:2604.10228v1 Announce Type: new Abstract: Current multimodal models often suffer from shallow reasoning, leading to errors caused by incomplete or inconsistent thought processes. To address this

tutorialsarxiv-cs-ai
14 Apr 2026
Agents

SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context

DGX agent

arXiv:2604.11716v1 Announce Type: new Abstract: Prior representative ReAct-style approaches in autonomous Software Engineering (SWE) typically lack the explicit System-2 reasoning required for deep an

agentsarxiv-cs-ai
14 Apr 2026
Research

SynthAgent: Adapting Web Agents with Synthetic Supervision

DGX agent

arXiv:2511.06101v3 Announce Type: replace-cross Abstract: Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo

DGX agent

arXiv:2604.11563v1 Announce Type: cross Abstract: Providing AI agents with reliable long-term memory that does not hallucinate remains an open problem. Current approaches to memory for LLM agents -- s

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

DGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

Tail-Aware Information-Theoretic Generalization for RLHF and SGLD

DGX agent

arXiv:2604.10727v1 Announce Type: cross Abstract: Classical information-theoretic generalization bounds typically control the generalization gap through KL-based mutual information and therefore rely

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

DGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

safetyarxiv-cs-ai
14 Apr 2026
Research

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

DGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

researcharxiv-cs-ai
14 Apr 2026
Safety

Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings

DGX agent

arXiv:2604.10849v1 Announce Type: cross Abstract: Federated learning (FL) performance is highly sensitive to heterogeneity across clients, yet practitioners lack reliable methods to anticipate how a f

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

DGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation

DGX agent

arXiv:2507.22767v3 Announce Type: replace-cross Abstract: Obtaining human-readable symbolic formulas via genetic programming-based symbolic distillation of a deep neural network trained on a target da

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings

DGX agent

arXiv:2305.14299v3 Announce Type: replace-cross Abstract: Learning high quality sentence embeddings from dialogues has drawn increasing attentions as it is essential to solve a variety of dialogue-ori

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Amazing Agent Race: Strong Tool Users, Weak Navigators

DGX agent

arXiv:2604.10261v1 Announce Type: new Abstract: Existing tool-use benchmarks for LLM agents are overwhelmingly linear: our analysis of six benchmarks shows 55 to 100% of instances are simple chains of

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

The Augmentation Trap: AI Productivity and the Cost of Cognitive Offloading

DGX agent

arXiv:2604.03501v2 Announce Type: replace-cross Abstract: Experimental evidence confirms that AI tools raise worker productivity, but also that sustained use can erode the expertise on which those gai

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

The Deployment Gap in AI Media Detection: Platform-Aware and Visually Constrained Adversarial Evaluation

DGX agent

arXiv:2604.09706v1 Announce Type: cross Abstract: Recent AI media detectors report near-perfect performance under clean laboratory evaluation, yet their robustness under realistic deployment condition

applicationsarxiv-cs-ai
14 Apr 2026
Research

The Geometry of Knowing: From Possibilistic Ignorance to Probabilistic Certainty -- A Measure-Theoretic Framework for Epistemic Convergence

DGX agent

arXiv:2604.09614v1 Announce Type: new Abstract: This paper develops a measure-theoretic framework establishing when and how a possibilistic representation of incomplete knowledge contracts into a prob

researcharxiv-cs-ai
14 Apr 2026
Model Releases

The Missing Knowledge Layer in Cognitive Architectures for AI Agents

DGX agent

arXiv:2604.11364v1 Announce Type: new Abstract: The two most influential cognitive architecture frameworks for AI agents, CoALA [21] and JEPA [12], both lack an explicit Knowledge layer with its own p

model-releasesarxiv-cs-ai
14 Apr 2026
Research

The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise

DGX agent

arXiv:2604.09780v1 Announce Type: new Abstract: Mixture of Experts (MoEs) are now ubiquitous in large language models, yet the mechanisms behind their 'expert specialization' remain poorly understood.

researcharxiv-cs-ai
14 Apr 2026
Safety

The Paradox of Professional Input: How Expert Collaboration with AI Systems Shapes Their Future Value

DGX agent

arXiv:2504.12654v1 Announce Type: cross Abstract: This perspective paper examines a fundamental paradox in the relationship between professional expertise and artificial intelligence: as domain expert

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping

DGX agent

arXiv:2604.11297v1 Announce Type: cross Abstract: Despite the success of reinforcement learning for large language models, a common failure mode is reduced sampling diversity, where the policy repeate

safetyarxiv-cs-ai
14 Apr 2026
Research

The Phantom of PCIe: Constraining Generative Artificial Intelligences for Practical Peripherals Trace Synthesizing

DGX agent

arXiv:2411.06376v3 Announce Type: replace-cross Abstract: Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. The development of

researcharxiv-cs-ai
14 Apr 2026
Safety

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

DGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

The Rise and Fall of G in AGI

DGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems

DGX agent

arXiv:2604.11309v1 Announce Type: cross Abstract: Large Language Models (LLMs) face prominent security risks from jailbreaking, a practice that manipulates models to bypass built-in security constrain

model-releasesarxiv-cs-ai
14 Apr 2026
Research

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

DGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

researcharxiv-cs-ai
14 Apr 2026
Model Releases

THEIA: Learning Complete Kleene Three-Valued Logic in a Pure-Neural Modular Architecture

DGX agent

arXiv:2604.11284v1 Announce Type: cross Abstract: We present THEIA, a modular neural architecture that learns complete Kleene three-valued logic (K3) end-to-end without any external symbolic solver, a

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Think Before you Write: QA-Guided Reasoning for Character Descriptions in Books

DGX agent

arXiv:2604.11435v1 Announce Type: cross Abstract: Character description generation is an important capability for narrative-focused applications such as summarization, story analysis, and character-dr

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

DGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation

DGX agent

arXiv:2604.10511v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for causal and counterfactual reasoning, yet their reliability in real-world policy evaluation remain

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Thought Branches: Interpreting LLM Reasoning Requires Resampling

DGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

DGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

DGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance

DGX agent

arXiv:2509.26627v2 Announce Type: replace Abstract: Designing dense rewards is crucial for reinforcement learning (RL), yet in robotics it often demands extensive manual effort and lacks scalability.

tutorialsarxiv-cs-ai
14 Apr 2026
← Previous
1…426427428429430…443
Next →