AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

OmniOVCD: Streamlining Open-Vocabulary Change Detection with SAM 3

DGX agent

arXiv:2601.13895v2 Announce Type: replace-cross Abstract: Change Detection (CD) is a fundamental task in remote sensing. It monitors the evolution of land cover over time. Based on this, Open-Vocabula

researcharxiv-cs-ai
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

On the Hybrid Nature of ABPMS Process Frames and its Implications on Automated Process Discovery

DGX agent

arXiv:2604.22455v1 Announce Type: new Abstract: A core component of any AI-Augmented Business Process Management System (ABPMS) is the process frame, which gives the system process-awareness and defin

agentsarxiv-cs-ai
27 Apr 2026
Safety

On the Properties of Feature Attribution for Supervised Contrastive Learning

DGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

Optimal Question Selection from a Large Question Bank for Clinical Field Recovery in Conversational Psychiatric Intake

DGX agent

arXiv:2604.22067v1 Announce Type: cross Abstract: Psychiatric intake is a sequential, high-stakes information-gathering process in which clinicians must decide what to ask, in what order, and how to i

model-releasesarxiv-cs-ai
27 Apr 2026
Research

OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping

DGX agent

arXiv:2510.18999v2 Announce Type: replace-cross Abstract: Reconstructing signed distance functions (SDFs) from point cloud data benefits many robot autonomy capabilities, including localization, mappi

researcharxiv-cs-ai
27 Apr 2026
Safety

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

DGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

safetyarxiv-cs-ai
27 Apr 2026
Research

PoLO: Proof-of-Learning and Proof-of-Ownership at Once with Chained Watermarking

DGX agent

arXiv:2505.12296v2 Announce Type: replace-cross Abstract: Our evaluation shows that PoLO achieves extbf{99%} watermark detection accuracy for ownership verification, while preserving data privacy and

researcharxiv-cs-ai
27 Apr 2026
Tutorials

Pre-trained Large Language Models Learn Hidden Markov Models In-context

DGX agent

arXiv:2506.07298v3 Announce Type: replace-cross Abstract: Hidden Markov Models (HMMs) are foundational tools for modeling sequential data with latent Markovian structure, yet fitting them to real-worl

tutorialsarxiv-cs-ai
27 Apr 2026
Local Ai

Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning

DGX agent

arXiv:2604.22229v1 Announce Type: cross Abstract: One-step offline RL actors are attractive because they avoid backpropagating through long iterative samplers and keep inference cheap, but they still

local-aiarxiv-cs-ai
27 Apr 2026
Tutorials

PrivSTRUCT: Untangling Data Purpose Compliance of Privacy Policies in Google Play Store

DGX agent

arXiv:2604.22157v1 Announce Type: cross Abstract: Existing research typically treats privacy policies as flat, uniform text, extracting information without regard for the document's logical hierarchy.

tutorialsarxiv-cs-ai
27 Apr 2026
Research

Protect the Brain When Treating the Heart: A Convolutional Neural Network for Detecting Emboli

DGX agent

arXiv:2604.22258v1 Announce Type: cross Abstract: Gaseous microemboli (GME) represent a common complication of cardiac structural interventions across both surgical and transcatheter approaches. Trans

researcharxiv-cs-ai
27 Apr 2026
Model Releases

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

DGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

model-releasesarxiv-cs-ai
27 Apr 2026
Agents

QDTraj: Exploration of Diverse Trajectory Primitives for Articulated Objects Robotic Manipulation

DGX agent

arXiv:2604.22551v1 Announce Type: cross Abstract: Thanks to the latest advances in learning and robotics, domestic robots are beginning to enter homes, aiming to execute household chores autonomously.

agentsarxiv-cs-ai
27 Apr 2026
Agents

QuantClaw: Precision Where It Matters for OpenClaw

DGX agent

arXiv:2604.22577v1 Announce Type: new Abstract: Autonomous agent systems such as OpenClaw introduce significant efficiency challenges due to long-context inputs and multi-turn reasoning. This results

agentsarxiv-cs-ai
27 Apr 2026
Agents

Read the Paper, Write the Code: Agentic Reproduction of Social-Science Results

DGX agent

arXiv:2604.21965v1 Announce Type: new Abstract: Recent work has used LLM agents to reproduce empirical social science results with access to both the data and code. We broaden this scope by asking: Ca

agentsarxiv-cs-ai
27 Apr 2026
Safety

ReCast: Recasting Learning Signals for Reinforcement Learning in Generative Recommendation

DGX agent

arXiv:2604.22169v1 Announce Type: cross Abstract: Generic group-based RL assumes that sampled rollout groups are already usable learning signals. We show that this assumption breaks down in sparse-hit

safetyarxiv-cs-ai
27 Apr 2026
Applications

ReLeVAnT: Relevance Lexical Vectors for Accurate Legal Text Classification

DGX agent

arXiv:2604.22292v1 Announce Type: cross Abstract: The classification of legal documents from an unstructured data corpus has several crucial applications in downstream tasks. Documents relevant to cou

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores

DGX agent

arXiv:2604.22063v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in clinical reasoning and risk assessment. However, their interpretive reliability in critical

model-releasesarxiv-cs-ai
27 Apr 2026
Safety

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

DGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

safetyarxiv-cs-ai
27 Apr 2026
Research

Removing Sandbagging in LLMs by Training with Weak Supervision

DGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

researcharxiv-cs-ai
27 Apr 2026
Applications

Report for NSF Workshop on AI for Electronic Design Automation

DGX agent

arXiv:2601.14541v4 Announce Type: replace-cross Abstract: This report distills the discussions and recommendations from the NSF Workshop on AI for Electronic Design Automation (EDA), held on December

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

DGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity

DGX agent

arXiv:2604.22597v1 Announce Type: new Abstract: Recent advancements in large language models have led to significant improvements across various tasks, including mathematical reasoning, which is used

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Publication: A Certification Framework for AI-Enabled Research

DGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problem

DGX agent

arXiv:2602.18734v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has demonstrated strong effectiveness in knowledge-intensive tasks by grounding language generation in ex

model-releasesarxiv-cs-ai
27 Apr 2026
Safety

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

DGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

safetyarxiv-cs-ai
27 Apr 2026
Research

Semantic Error Correction and Decoding for Short Block Channel Codes

DGX agent

arXiv:2604.22269v1 Announce Type: cross Abstract: This paper presents a semantic-enhanced receiver framework for transmitting natural language sentences over noisy wireless channels using multiple sho

researcharxiv-cs-ai
27 Apr 2026
Research

Sensory-Aware Sequential Recommendation via Review-Distilled Representations

DGX agent

arXiv:2603.02709v2 Announce Type: replace-cross Abstract: We propose a novel framework for sensory-aware sequential recommendation that enriches item representations with linguistically extracted sens

researcharxiv-cs-ai
27 Apr 2026
Research

Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning

DGX agent

arXiv:2604.22072v1 Announce Type: cross Abstract: Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clien

researcharxiv-cs-ai
27 Apr 2026
Applications

Shared Lexical Task Representations Explain Behavioral Variability In LLMs

DGX agent

arXiv:2604.22027v1 Announce Type: cross Abstract: One of the most common complaints about large language models (LLMs) is their prompt sensitivity -- that is, the fact that their ability to perform a

applicationsarxiv-cs-ai
27 Apr 2026
Agents

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

DGX agent

arXiv:2604.22558v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) mature, GUI agents are evolving from static interactions to complex navigation. While Reinforcement Learni

agentsarxiv-cs-ai
27 Apr 2026
Agents

Sound Agentic Science Requires Adversarial Experiments

DGX agent

arXiv:2604.22080v1 Announce Type: new Abstract: LLM-based agents are rapidly being adopted for scientific data analysis, automating tasks once limited by human time and expertise. This capability is o

agentsarxiv-cs-ai
27 Apr 2026
Research

Spontaneous Persuasion: An Audit of Model Persuasiveness in Everyday Conversations

DGX agent

arXiv:2604.22109v1 Announce Type: cross Abstract: Large language models (LLMs) possess strong persuasive capabilities that outperform humans in head-to-head comparisons. Users report consulting LLMs t

researcharxiv-cs-ai
27 Apr 2026
Research

SSG: Logit-Balanced Vocabulary Partitioning for LLM Watermarking

DGX agent

arXiv:2604.22438v1 Announce Type: cross Abstract: Watermarking has emerged as a promising technique for tracing the authorship of content generated by large language models (LLMs). Among existing appr

researcharxiv-cs-ai
27 Apr 2026
Research

StateX: Enhancing RNN Recall via Post-training State Expansion

DGX agent

arXiv:2509.22630v3 Announce Type: replace-cross Abstract: Recurrent neural networks (RNNs), such as linear attention and state-space models, have gained popularity due to their constant per-token comp

researcharxiv-cs-ai
27 Apr 2026
Agents

Superminds Test: Actively Evaluating Collective Intelligence of Agent Society via Probing Agents

DGX agent

arXiv:2604.22452v1 Announce Type: new Abstract: Collective intelligence refers to the ability of a group to achieve outcomes beyond what any individual member can accomplish alone. As large language m

agentsarxiv-cs-ai
27 Apr 2026
Agents

Teaching an Agent to Sketch One Part at a Time

DGX agent

arXiv:2603.19500v2 Announce Type: replace Abstract: We develop a method for producing vector sketches one part at a time. To do this, we train a multi-modal language model-based agent using a novel mu

agentsarxiv-cs-ai
27 Apr 2026
Research

Tell Me Why: Designing an Explainable LLM-based Dialogue System for Student Problem Behavior Diagnosis

DGX agent

arXiv:2604.22237v1 Announce Type: cross Abstract: Diagnosing student problem behaviors requires teachers to synthesize multifaceted information, identify behavioral categories, and plan intervention s

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models

DGX agent

arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or

model-releasesarxiv-cs-ai
27 Apr 2026
Safety

The Biggest Risk of Embodied AI is Governance Lag

DGX agent

arXiv:2604.21938v1 Announce Type: cross Abstract: Embodied AI is widely discussed as a job-displacement problem. The deeper risk, however, is governance lag: the inability of public institutions to ke

safetyarxiv-cs-ai
27 Apr 2026
Research

The Shape of Adversarial Influence: Characterizing LLM Latent Spaces with Persistent Homology

DGX agent

arXiv:2505.20435v3 Announce Type: replace-cross Abstract: Existing interpretability methods for Large Language Models (LLMs) predominantly capture linear directions or isolated features. This overlook

researcharxiv-cs-ai
27 Apr 2026
Safety

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

DGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

safetyarxiv-cs-ai
27 Apr 2026
Safety

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

DGX agent

arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Research

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

DGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Universal Transformers Need Memory: Depth-State Trade-offs in Adaptive Recursive Reasoning

DGX agent

arXiv:2604.21999v1 Announce Type: cross Abstract: We study learned memory tokens as computational scratchpad for a single-block Universal Transformer (UT) with Adaptive Computation Time (ACT) on Sudok

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen

DGX agent

arXiv:2604.22215v1 Announce Type: cross Abstract: Verbal confidence elicitation is widely used to extract uncertainty estimates from LLMs. We tested whether seven instruction-tuned open-weight models

researcharxiv-cs-ai
27 Apr 2026
← Previous
1…382383384385386…443
Next →