AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
27 Apr 2026

How Supply Chain Dependencies Complicate Bias Measurement and Accountability Attribution in AI Hiring Applications

Local AiDGX agent

arXiv:2604.22679v1 Announce Type: cross Abstract: The increasing adoption of AI systems in hiring has raised concerns about algorithmic bias and accountability, prompting regulatory responses includin

Initial results of the Digital Consciousness Model

ResearchDGX agent

arXiv:2601.17060v2 Announce Type: replace-cross Abstract: Artificially intelligent systems have become remarkably sophisticated. They hold conversations, write essays, and seem to understand context i

Interpretable Deep Learning for Stock Returns: A Consensus-Bottleneck Asset Pricing Model

ResearchDGX agent

arXiv:2512.16251v5 Announce Type: replace-cross Abstract: We introduce the Consensus-Bottleneck Asset Pricing Model (CB-APM), which embeds aggregate analyst consensus as a structural bottleneck, treat


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models

ResearchDGX agent

arXiv:2604.22411v1 Announce Type: new Abstract: Even when decoding with temperature T=0, large language models (LLMs) can produce divergent outputs for identical inputs. Recent work by Thinking Machin

KuaiLive: A Real-time Interactive Dataset for Live Streaming Recommendation

Model ReleasesDGX agent

arXiv:2508.05633v2 Announce Type: replace-cross Abstract: Live streaming platforms have become a dominant form of online content consumption, offering dynamically evolving content, real-time interacti

Learning-augmented robotic automation for real-world manufacturing

SafetyDGX agent

arXiv:2604.22235v1 Announce Type: cross Abstract: Industrial robots are widely used in manufacturing, yet most manipulation still depends on fixed waypoint scripts that are brittle to environmental ch

Learning Evidence Highlighting for Frozen LLMs

SafetyDGX agent

arXiv:2604.22565v1 Announce Type: cross Abstract: Large Language Models (LLMs) can reason well, yet often miss decisive evidence when it is buried in long, noisy contexts. We introduce HiLight, an Evi

Learning from Natural Language Feedback for Personalized Question Answering

Model ReleasesDGX agent

arXiv:2508.10695v2 Announce Type: replace-cross Abstract: Personalization is crucial for enhancing both the effectiveness and user satisfaction of language technologies, particularly in information-se

LeHome: A Simulation Environment for Deformable Object Manipulation in Household Scenarios

ApplicationsDGX agent

arXiv:2604.22363v1 Announce Type: cross Abstract: Household environments present one of the most common, impactful yet challenging application domains for robotics. Within household scenarios, manipul

Lifting Unlabeled Internet-level Data for 3D Scene Understanding

ResearchDGX agent

arXiv:2604.01907v2 Announce Type: replace-cross Abstract: Annotated 3D scene data is scarce and expensive to acquire, while abundant unlabeled videos are readily available on the internet. In this pap

Lightweight Retrieval-Augmented Generation and Large Language Model-Based Modeling for Scalable Patient-Trial Matching

ApplicationsDGX agent

arXiv:2604.22061v1 Announce Type: cross Abstract: Patient-trial matching requires reasoning over long, heterogeneous electronic health records (EHRs) and complex eligibility criteria, posing significa

LLM+Graph@VLDB'2025 Workshop Summary

ResearchDGX agent

arXiv:2604.02861v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) with graph-structured data has become a pivotal and fast evolving research frontier, drawing s

LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines

Model ReleasesDGX agent

arXiv:2411.08027v3 Announce Type: replace-cross Abstract: Most learning-based approaches to complex physical reasoning sidestep the crucial problem of parameter identification (e.g., mass, friction) t

MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction

Local AiDGX agent

arXiv:2604.21957v1 Announce Type: cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state predicti

Math Takes Two: A test for emergent mathematical reasoning in communication

Model ReleasesDGX agent

arXiv:2604.21935v1 Announce Type: new Abstract: Although language models demonstrate remarkable proficiency on mathematical benchmarks, it remains unclear whether this reflects true mathematical reaso

Mechanistic Interpretability of Antibody Language Models Using SAEs

ResearchDGX agent

arXiv:2512.05794v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) are a mechanistic interpretability technique that have been used to provide insight into learned concepts within la

Memanto: Typed Semantic Memory with Information-Theoretic Retrieval for Long-Horizon Agents

AgentsDGX agent

arXiv:2604.22085v1 Announce Type: new Abstract: The transition from stateless language model inference to persistent, multi session autonomous agents has revealed memory to be a primary architectural

Mochi: Aligning Pre-training and Inference for Efficient Graph Foundation Models via Meta-Learning

ApplicationsDGX agent

arXiv:2604.22031v1 Announce Type: cross Abstract: We propose Mochi, a Graph Foundation Model that addresses task unification and training efficiency by adopting a meta-learning based training framewor

Model Predictive Control of Hybrid Dynamical Systems

ResearchDGX agent

arXiv:2604.21989v1 Announce Type: cross Abstract: The problem of controlling hybrid dynamical systems using model predictive control (MPC) is formulated and sufficient conditions for asymptotic stabil

MolClaw: An Autonomous Agent with Hierarchical Skills for Drug Molecule Evaluation, Screening, and Optimization

Model ReleasesDGX agent

arXiv:2604.21937v1 Announce Type: new Abstract: Computational drug discovery, particularly the complex workflows of drug molecule screening and optimization, requires orchestrating dozens of specializ

Motivating Next-Gen Accelerators with Flexible (N:M) Activation Sparsity via Benchmarking Lightweight Post-Training Sparsification Approaches

HardwareDGX agent

arXiv:2509.22166v4 Announce Type: replace-cross Abstract: The demand for efficient large language model (LLM) inference has intensified the focus on sparsification techniques. While semi-structured (N

Multi-Task Optimization over Networks of Tasks

Model ReleasesDGX agent

arXiv:2604.21991v1 Announce Type: cross Abstract: Multi-task optimization is a powerful approach for solving a large number of tasks in parallel. However, existing algorithms face distinct limitations

Multimodal Neural Operators for Real-Time Biomechanical Modelling of Traumatic Brain Injury

HardwareDGX agent

arXiv:2510.03248v3 Announce Type: replace-cross Abstract: Background: Traumatic brain injury modeling requires integrating volumetric neuroimaging, demographic parameters, and acquisition metadata. Fi

Navigating Large-Scale Document Collections: MuDABench for Multi-Document Analytical QA

Model ReleasesDGX agent

arXiv:2604.22239v1 Announce Type: cross Abstract: This paper introduces the task of analytical question answering over large, semi-structured document collections. We present MuDABench, a benchmark fo

OmniOVCD: Streamlining Open-Vocabulary Change Detection with SAM 3

ResearchDGX agent

arXiv:2601.13895v2 Announce Type: replace-cross Abstract: Change Detection (CD) is a fundamental task in remote sensing. It monitors the evolution of land cover over time. Based on this, Open-Vocabula

On the Hybrid Nature of ABPMS Process Frames and its Implications on Automated Process Discovery

AgentsDGX agent

arXiv:2604.22455v1 Announce Type: new Abstract: A core component of any AI-Augmented Business Process Management System (ABPMS) is the process frame, which gives the system process-awareness and defin

On the Properties of Feature Attribution for Supervised Contrastive Learning

SafetyDGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

Optimal Question Selection from a Large Question Bank for Clinical Field Recovery in Conversational Psychiatric Intake

Model ReleasesDGX agent

arXiv:2604.22067v1 Announce Type: cross Abstract: Psychiatric intake is a sequential, high-stakes information-gathering process in which clinicians must decide what to ask, in what order, and how to i

OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping

ResearchDGX agent

arXiv:2510.18999v2 Announce Type: replace-cross Abstract: Reconstructing signed distance functions (SDFs) from point cloud data benefits many robot autonomy capabilities, including localization, mappi

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

SafetyDGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

PoLO: Proof-of-Learning and Proof-of-Ownership at Once with Chained Watermarking

ResearchDGX agent

arXiv:2505.12296v2 Announce Type: replace-cross Abstract: Our evaluation shows that PoLO achieves extbf{99%} watermark detection accuracy for ownership verification, while preserving data privacy and

Pre-trained Large Language Models Learn Hidden Markov Models In-context

TutorialsDGX agent

arXiv:2506.07298v3 Announce Type: replace-cross Abstract: Hidden Markov Models (HMMs) are foundational tools for modeling sequential data with latent Markovian structure, yet fitting them to real-worl

Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning

Local AiDGX agent

arXiv:2604.22229v1 Announce Type: cross Abstract: One-step offline RL actors are attractive because they avoid backpropagating through long iterative samplers and keep inference cheap, but they still

PrivSTRUCT: Untangling Data Purpose Compliance of Privacy Policies in Google Play Store

TutorialsDGX agent

arXiv:2604.22157v1 Announce Type: cross Abstract: Existing research typically treats privacy policies as flat, uniform text, extracting information without regard for the document's logical hierarchy.

Protect the Brain When Treating the Heart: A Convolutional Neural Network for Detecting Emboli

ResearchDGX agent

arXiv:2604.22258v1 Announce Type: cross Abstract: Gaseous microemboli (GME) represent a common complication of cardiac structural interventions across both surgical and transcatheter approaches. Trans

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

Model ReleasesDGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

QDTraj: Exploration of Diverse Trajectory Primitives for Articulated Objects Robotic Manipulation

AgentsDGX agent

arXiv:2604.22551v1 Announce Type: cross Abstract: Thanks to the latest advances in learning and robotics, domestic robots are beginning to enter homes, aiming to execute household chores autonomously.

QuantClaw: Precision Where It Matters for OpenClaw

AgentsDGX agent

arXiv:2604.22577v1 Announce Type: new Abstract: Autonomous agent systems such as OpenClaw introduce significant efficiency challenges due to long-context inputs and multi-turn reasoning. This results

Read the Paper, Write the Code: Agentic Reproduction of Social-Science Results

AgentsDGX agent

arXiv:2604.21965v1 Announce Type: new Abstract: Recent work has used LLM agents to reproduce empirical social science results with access to both the data and code. We broaden this scope by asking: Ca

ReCast: Recasting Learning Signals for Reinforcement Learning in Generative Recommendation

SafetyDGX agent

arXiv:2604.22169v1 Announce Type: cross Abstract: Generic group-based RL assumes that sampled rollout groups are already usable learning signals. We show that this assumption breaks down in sparse-hit

ReLeVAnT: Relevance Lexical Vectors for Accurate Legal Text Classification

ApplicationsDGX agent

arXiv:2604.22292v1 Announce Type: cross Abstract: The classification of legal documents from an unstructured data corpus has several crucial applications in downstream tasks. Documents relevant to cou

Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores

Model ReleasesDGX agent

arXiv:2604.22063v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in clinical reasoning and risk assessment. However, their interpretive reliability in critical

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

Removing Sandbagging in LLMs by Training with Weak Supervision

ResearchDGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

Report for NSF Workshop on AI for Electronic Design Automation

ApplicationsDGX agent

arXiv:2601.14541v4 Announce Type: replace-cross Abstract: This report distills the discussions and recommendations from the NSF Workshop on AI for Electronic Design Automation (EDA), held on December

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

Model ReleasesDGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity

ResearchDGX agent

arXiv:2604.22597v1 Announce Type: new Abstract: Recent advancements in large language models have led to significant improvements across various tasks, including mathematical reasoning, which is used

Rethinking Publication: A Certification Framework for AI-Enabled Research

Model ReleasesDGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problem

Model ReleasesDGX agent

arXiv:2602.18734v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has demonstrated strong effectiveness in knowledge-intensive tasks by grounding language generation in ex

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

SafetyDGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

Semantic Error Correction and Decoding for Short Block Channel Codes

ResearchDGX agent

arXiv:2604.22269v1 Announce Type: cross Abstract: This paper presents a semantic-enhanced receiver framework for transmitting natural language sentences over noisy wireless channels using multiple sho

Sensory-Aware Sequential Recommendation via Review-Distilled Representations

ResearchDGX agent

arXiv:2603.02709v2 Announce Type: replace-cross Abstract: We propose a novel framework for sensory-aware sequential recommendation that enriches item representations with linguistically extracted sens

Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning

ResearchDGX agent

arXiv:2604.22072v1 Announce Type: cross Abstract: Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clien

Shared Lexical Task Representations Explain Behavioral Variability In LLMs

ApplicationsDGX agent

arXiv:2604.22027v1 Announce Type: cross Abstract: One of the most common complaints about large language models (LLMs) is their prompt sensitivity -- that is, the fact that their ability to perform a

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

AgentsDGX agent

arXiv:2604.22558v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) mature, GUI agents are evolving from static interactions to complex navigation. While Reinforcement Learni

Sound Agentic Science Requires Adversarial Experiments

AgentsDGX agent

arXiv:2604.22080v1 Announce Type: new Abstract: LLM-based agents are rapidly being adopted for scientific data analysis, automating tasks once limited by human time and expertise. This capability is o

Spontaneous Persuasion: An Audit of Model Persuasiveness in Everyday Conversations

ResearchDGX agent

arXiv:2604.22109v1 Announce Type: cross Abstract: Large language models (LLMs) possess strong persuasive capabilities that outperform humans in head-to-head comparisons. Users report consulting LLMs t

SSG: Logit-Balanced Vocabulary Partitioning for LLM Watermarking

ResearchDGX agent

arXiv:2604.22438v1 Announce Type: cross Abstract: Watermarking has emerged as a promising technique for tracing the authorship of content generated by large language models (LLMs). Among existing appr

StateX: Enhancing RNN Recall via Post-training State Expansion

ResearchDGX agent

arXiv:2509.22630v3 Announce Type: replace-cross Abstract: Recurrent neural networks (RNNs), such as linear attention and state-space models, have gained popularity due to their constant per-token comp

Superminds Test: Actively Evaluating Collective Intelligence of Agent Society via Probing Agents

AgentsDGX agent

arXiv:2604.22452v1 Announce Type: new Abstract: Collective intelligence refers to the ability of a group to achieve outcomes beyond what any individual member can accomplish alone. As large language m

← Previous
1…305306307308309…354
Next →