AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
14 Apr 2026

What do your logits know? (The answer may surprise you!)

TutorialsDGX agent

arXiv:2604.09885v1 Announce Type: new Abstract: Recent work has shown that probing model internals can reveal a wealth of information not apparent from the model generations. This poses the risk of un

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

What's In My Human Feedback? Learning Interpretable Descriptions of Preference Data

SafetyDGX agent

arXiv:2510.26202v2 Announce Type: replace-cross Abstract: Human feedback can alter language models in unpredictable and undesirable ways, as practitioners lack a clear understanding of what feedback d


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When More Thinking Hurts: Overthinking in LLM Test-Time Compute Scaling

ResearchDGX agent

arXiv:2604.10739v1 Announce Type: new Abstract: Scaling test-time compute through extended chains of thought has become a dominant paradigm for improving large language model reasoning. However, exist

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

SafetyDGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

When Valid Signals Fail: Regime Boundaries Between LLM Features and RL Trading Policies

SafetyDGX agent

arXiv:2604.10996v1 Announce Type: cross Abstract: Can large language models (LLMs) generate continuous numerical features that improve reinforcement learning (RL) trading agents? We build a modular pi

When Verification Fails: How Compositionally Infeasible Claims Escape Rejection

ResearchDGX agent

arXiv:2604.10990v1 Announce Type: cross Abstract: Scientific claim verification, the task of determining whether claims are entailed by scientific evidence, is fundamental to establishing discoveries

Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text

Model ReleasesDGX agent

arXiv:2601.17172v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly capable of generating personalized, persuasive text at scale, raising new questions about bias a

Why Do Large Language Models Generate Harmful Content?

ResearchDGX agent

arXiv:2604.11663v1 Announce Type: new Abstract: Large Language Models (LLMs) have been shown to generate harmful content. However, the underlying causes of such behavior remain under explored. We prop

Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?

ResearchDGX agent

arXiv:2510.27269v3 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, yet they still exhibit a multilingual reasoning gap, p

Why Smaller Is Slower? Dimensional Misalignment in Compressed LLMs

Model ReleasesDGX agent

arXiv:2604.09595v1 Announce Type: cross Abstract: Post-training compression reduces LLM parameter counts but often produces irregular tensor dimensions that degrade GPU performance -- a phenomenon we

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

Model ReleasesDGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

WisPaper: Your AI Scholar Search Engine

AgentsDGX agent

arXiv:2512.06879v3 Announce Type: replace-cross Abstract: We present extsc{WisPaper}, an end-to-end agent system that transforms how researchers discover, organize, and track academic literature. The

Wolkowicz-Styan Upper Bound on the Hessian Eigenspectrum for Cross-Entropy Loss in Nonlinear Smooth Neural Networks

ResearchDGX agent

arXiv:2604.10202v1 Announce Type: cross Abstract: Neural networks (NNs) are central to modern machine learning and achieve state-of-the-art results in many applications. However, the relationship betw

Woosh: A Sound Effects Foundation Model

Model ReleasesDGX agent

arXiv:2604.01929v2 Announce Type: replace-cross Abstract: The audio research community depends on open generative models as foundational tools for building novel approaches and establishing baselines.

Working Paper: Towards Schema-based Learning from a Category-Theoretic Perspective

AgentsDGX agent

arXiv:2604.10589v1 Announce Type: new Abstract: We introduce a hierarchical categorical framework for Schema-Based Learning (SBL) structured across four interconnected levels. At the schema level, a f

X-SYS: A Reference Architecture for Interactive Explanation Systems

ResearchDGX agent

arXiv:2602.12748v3 Announce Type: replace Abstract: The explainable AI (XAI) research community has proposed numerous technical methods, yet deploying explainability as systems remains challenging: In

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

ResearchDGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass

Model ReleasesDGX agent

arXiv:2604.10966v1 Announce Type: cross Abstract: We present a discriminative multimodal reward model that scores all candidate responses in a single forward pass. Conventional discriminative reward m

Your Model Diversity, Not Method, Determines Reasoning Strategy

Model ReleasesDGX agent

arXiv:2604.10827v1 Announce Type: new Abstract: Compute scaling for LLM reasoning requires allocating budget between exploring solution approaches (breadth) and refining promising solutions (depth). M

Zero-Shot Quantization via Weight-Space Arithmetic

ResearchDGX agent

arXiv:2604.03420v2 Announce Type: replace-cross Abstract: We show that robustness to post-training quantization (PTQ) is a transferable direction in weight space. We call this direction the quantizati

Zero-shot World Models Are Developmentally Efficient Learners

ResearchDGX agent

arXiv:2604.10333v1 Announce Type: new Abstract: Young children demonstrate early abilities to understand their physical world, estimating depth, motion, object coherence, interactions, and many other

ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval

SafetyDGX agent

arXiv:2604.10898v1 Announce Type: new Abstract: Large language models (LLMs) have shown great performance on complex reasoning tasks but often require generating long intermediate thoughts before reac

13 Apr 2026

3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding

ResearchDGX agent

arXiv:2604.08645v1 Announce Type: cross Abstract: Large multimodal models are increasingly used as the reasoning core of embodied agents operating in 3D environments, yet they remain prone to hallucin

A Closer Look at the Application of Causal Inference in Graph Representation Learning

ResearchDGX agent

arXiv:2604.08890v1 Announce Type: cross Abstract: Modeling causal relationships in graph representation learning remains a fundamental challenge. Existing approaches often draw on theories and methods

A Mathematical Framework for Temporal Modeling and Counterfactual Policy Simulation of Student Dropout

SafetyDGX agent

arXiv:2604.08874v1 Announce Type: cross Abstract: This study proposes a temporal modeling framework with a counterfactual policy-simulation layer for student dropout in higher education, using LMS eng

Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring

ResearchDGX agent

arXiv:2604.08718v1 Announce Type: cross Abstract: Geometric Foundation Models (GFMs) have recently advanced monocular SLAM by providing robust, calibration-free 3D priors. However, deploying these mod

Act or Escalate? Evaluating Escalation Behavior in Automation with Language Models

SafetyDGX agent

arXiv:2604.08588v1 Announce Type: cross Abstract: Effective automation hinges on deciding when to act and when to escalate. We model this as a decision under uncertainty: an LLM forms a prediction, es

ActionNex: A Virtual Outage Manager for Cloud Computing

AgentsDGX agent

arXiv:2604.03512v2 Announce Type: replace Abstract: Outage management in large-scale cloud operations remains heavily manual, requiring rapid triage, cross-team coordination, and experience-driven dec

ActivityEditor: Learning to Synthesize Physically Valid Human Mobility

AgentsDGX agent

arXiv:2604.05529v2 Announce Type: replace Abstract: Human mobility modeling is indispensable for diverse urban applications. However, existing data-driven methods often suffer from data scarcity, limi

Adaptive Dual Residual U-Net with Attention Gate and Multiscale Spatial Attention Mechanisms (ADRUwAMS)

ResearchDGX agent

arXiv:2604.08893v1 Announce Type: cross Abstract: Glioma is a harmful brain tumor that requires early detection to ensure better health results. Early detection of this tumor is key for effective trea

Adaptive Planning for Multi-Attribute Controllable Summarization with Monte Carlo Tree Search

Model ReleasesDGX agent

arXiv:2509.26435v2 Announce Type: replace-cross Abstract: Controllable summarization moves beyond generic outputs toward human-aligned summaries guided by specified attributes. In practice, the interd

Adaptive Rigor in AI System Evaluation using Temperature-Controlled Verdict Aggregation via Generalized Power Mean

Model ReleasesDGX agent

arXiv:2604.08595v1 Announce Type: cross Abstract: Existing evaluation methods for LLM-based AI systems, such as LLM-as-a-Judge, verdict systems, and NLI, do not always align well with human assessment

Advantage-Guided Diffusion for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2604.09035v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) with autoregressive world models suffers from compounding errors, whereas diffusion world models mitigate this

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

Model ReleasesDGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

SafetyDGX agent

arXiv:2502.08691v2 Announce Type: replace-cross Abstract: Understanding human behavior and society is a central focus in social sciences, with the rise of generative social science marking a significa

AI Driven Soccer Analysis Using Computer Vision

ApplicationsDGX agent

arXiv:2604.08722v1 Announce Type: cross Abstract: Sport analysis is crucial for team performance since it provides actionable data that can inform coaching decisions, improve player performance, and e

AI-Induced Human Responsibility (AIHR) in AI-Human teams

AgentsDGX agent

arXiv:2604.08866v1 Announce Type: cross Abstract: As organizations increasingly deploy AI as a teammate rather than a standalone tool, morally consequential mistakes often arise from joint human-AI wo

Aligned Agents, Biased Swarm: Measuring Bias Amplification in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.08963v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) are increasingly deployed for complex workflows, their emergent properties-particularly the accumulation of bias-remai

AlphaCast: A Human Wisdom-LLM Intelligence Co-Reasoning Framework for Interactive Time Series Forecasting

AgentsDGX agent

arXiv:2511.08947v5 Announce Type: replace Abstract: Time series forecasting plays a crucial role in decision-making across many real-world applications. Despite substantial progress, most existing met

AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs

Model ReleasesDGX agent

arXiv:2604.08590v1 Announce Type: cross Abstract: We present AlphaLab, an autonomous research harness that leverages frontier LLM agentic capabilities to automate the full experimental cycle in quanti

ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads

Model ReleasesDGX agent

arXiv:2604.05426v2 Announce Type: replace-cross Abstract: Low-Rank Adaptation (LoRA) is now the dominant method for parameter-efficient fine-tuning of large language models, but achieving a high-quali

An Adaptive Model Selection Framework for Demand Forecasting under Horizon-Induced Degradation to Support Business Strategy and Operations

SafetyDGX agent

arXiv:2602.13939v3 Announce Type: replace-cross Abstract: Business environments characterized by intermittent demand, high variability, and multi-step planning horizons require forecasting policies th

AR-KAN: Autoregressive-Weight-Enhanced Kolmogorov-Arnold Network for Time Series Forecasting

ApplicationsDGX agent

arXiv:2509.02967v3 Announce Type: replace-cross Abstract: Traditional neural networks struggle to capture the spectral structure of complex signals. Fourier neural networks (FNNs) attempt to address t

Artifacts as Memory Beyond the Agent Boundary

SafetyDGX agent

arXiv:2604.08756v1 Announce Type: new Abstract: The situated view of cognition holds that intelligent behavior depends not only on internal memory, but on an agent's active use of environmental resour

Artificial intelligence can persuade people to take political actions

ApplicationsDGX agent

arXiv:2604.09200v1 Announce Type: cross Abstract: There is substantial concern about the ability of advanced artificial intelligence to influence people's behaviour. A rapidly growing body of research

ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transfer

SafetyDGX agent

arXiv:2604.08355v2 Announce Type: replace Abstract: Reinforcement Learning (RL) agents often struggle to generalize knowledge to new tasks, even those structurally similar to ones they have mastered.

ASTRA: Adaptive Semantic Tree Reasoning Architecture for Complex Table Question Answering

SafetyDGX agent

arXiv:2604.08999v1 Announce Type: cross Abstract: Table serialization remains a critical bottleneck for Large Language Models (LLMs) in complex table question answering, hindered by challenges such as

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

Model ReleasesDGX agent

arXiv:2604.08867v1 Announce Type: cross Abstract: Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherentl

Automated Standardization of Legacy Biomedical Metadata Using an Ontology-Constrained LLM Agent

AgentsDGX agent

arXiv:2604.08552v1 Announce Type: cross Abstract: Scientific metadata are often incomplete and noncompliant with community standards, limiting dataset findability, interoperability, and reuse. When re

BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning

Model ReleasesDGX agent

arXiv:2604.09378v1 Announce Type: cross Abstract: Agent ecosystems increasingly rely on installable skills to extend functionality, and some skills bundle learned model artifacts as part of their exec

Bayesian Social Deduction with Graph-Informed Language Models

AgentsDGX agent

arXiv:2506.17788v2 Announce Type: replace Abstract: Social reasoning - inferring unobservable beliefs and intentions from partial observations of other agents - remains a challenging task for large la

BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation

ResearchDGX agent

arXiv:2604.09497v1 Announce Type: cross Abstract: Accurate evaluation is central to the large language model (LLM) ecosystem, guiding model selection and downstream adoption across diverse use cases.

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

SafetyDGX agent

arXiv:2603.06665v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) often benefit from chain-of-thought (CoT) prompting in general domains, yet its efficacy in medical vision

Beyond Isolated Clients: Integrating Graph-Based Embeddings into Event Sequence Models

ResearchDGX agent

arXiv:2604.09085v1 Announce Type: cross Abstract: Large-scale digital platforms generate billions of timestamped user-item interactions (events) that are crucial for predicting user attributes in, e.g

Beyond Relevance: Utility-Centric Retrieval in the LLM Era

AgentsDGX agent

arXiv:2604.08920v1 Announce Type: cross Abstract: Information retrieval systems have traditionally optimized for topical relevance-the degree to which retrieved documents match a query. However, relev

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

Model ReleasesDGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

Boosted Distributional Reinforcement Learning: Analysis and Healthcare Applications

AgentsDGX agent

arXiv:2604.04334v2 Announce Type: replace-cross Abstract: Researchers and practitioners are increasingly considering reinforcement learning to optimize decisions in complex domains like robotics and h

Building Better Environments for Autonomous Cyber Defence

AgentsDGX agent

arXiv:2604.08805v1 Announce Type: cross Abstract: In November 2025, the authors ran a workshop on the topic of what makes a good reinforcement learning (RL) environment for autonomous cyber defence (A

Camera Artist: A Multi-Agent Framework for Cinematic Language Storytelling Video Generation

AgentsDGX agent

arXiv:2604.09195v1 Announce Type: new Abstract: We propose Camera Artist, a multi-agent framework that models a real-world filmmaking workflow to generate narrative videos with explicit cinematic lang

← Previous
1…338339340341342…350
Next →