AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Applications

Cognis: Context-Aware Memory for Conversational AI Agents

DGX agent

arXiv:2604.19771v1 Announce Type: cross Abstract: LLM agents lack persistent memory, causing conversations to reset each session and preventing personalization over time. We present Lyzr Cognis, a uni

applicationsarxiv-cs-ai
23 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

DGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

DGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

model-releasesarxiv-cs-ai
23 Apr 2026
Tutorials

Combo-Gait: Unified Transformer Framework for Multi-Modal Gait Recognition and Attribute Analysis

DGX agent

arXiv:2510.10417v2 Announce Type: replace-cross Abstract: Gait recognition is an important biometric for human identification at a distance, particularly under low-resolution or unconstrained environm

tutorialsarxiv-cs-ai
23 Apr 2026
Model Releases

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

DGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Computing the Reachability Value of Posterior-Deterministic POMDPs

DGX agent

arXiv:2602.07473v2 Announce Type: replace Abstract: Partially observable Markov decision processes (POMDPs) are a fundamental model for sequential decision-making under uncertainty. However, many veri

researcharxiv-cs-ai
23 Apr 2026
Research

Context Attribution with Multi-Armed Bandit Optimization

DGX agent

arXiv:2506.19977v2 Announce Type: replace Abstract: Understanding which parts of the retrieved context contribute to a large language model's generated answer is essential for building interpretable a

researcharxiv-cs-ai
23 Apr 2026
Tutorials

Convergent Evolution: How Different Language Models Learn Similar Number Representations

DGX agent

arXiv:2604.20817v1 Announce Type: cross Abstract: Language models trained on natural text learn to represent numbers using periodic features with dominant periods at T=2, 5, 10. In this paper, we iden

tutorialsarxiv-cs-ai
23 Apr 2026
Applications

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment

DGX agent

arXiv:2604.20246v1 Announce Type: cross Abstract: Industrial robotic manipulation demands reliable long-horizon execution across embodiments, tasks, and changing object distributions. While Vision-Lan

applicationsarxiv-cs-ai
23 Apr 2026
Safety

Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation

DGX agent

arXiv:2604.20763v1 Announce Type: cross Abstract: Retrieval quality is the primary bottleneck for accuracy and robustness in retrieval-augmented generation (RAG). Current evaluation relies on heuristi

safetyarxiv-cs-ai
23 Apr 2026
Agents

CreativeGame:Toward Mechanic-Aware Creative Game Generation

DGX agent

arXiv:2604.19926v1 Announce Type: new Abstract: Large language models can generate plausible game code, but turning this capability into iterative creative improvement remains difficult. In practice,

agentsarxiv-cs-ai
23 Apr 2026
Applications

Cross-Modal Taxonomic Generalization in (Vision-) Language Models

DGX agent

arXiv:2603.07474v2 Announce Type: replace-cross Abstract: What is the interplay between semantic representations learned by language models (LM) from surface form alone to those learned from more grou

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

DGX agent

arXiv:2604.20389v1 Announce Type: cross Abstract: The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles

DGX agent

arXiv:2604.20771v1 Announce Type: cross Abstract: The Internet of Vehicles (IoV) is advancing modern transportation by improving safety, efficiency, and intelligence. However, the reliance on the Cont

safetyarxiv-cs-ai
23 Apr 2026
Research

Deconstructing Superintelligence: Identity, Self-Modification and Differance

DGX agent

arXiv:2604.19845v1 Announce Type: new Abstract: Self-modification is often taken as constitutive of artificial superintelligence (SI), yet modification is a relative action requiring a supplement outs

researcharxiv-cs-ai
23 Apr 2026
Research

Degrees, Levels, and Profiles of Contextuality

DGX agent

arXiv:2603.26692v3 Announce Type: replace-cross Abstract: We introduce a new notion, that of a contextuality profile of a system of random variables. Rather than characterizing a system's contextualit

researcharxiv-cs-ai
23 Apr 2026
Research

Depression Risk Assessment in Social Media via Large Language Models

DGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

researcharxiv-cs-ai
23 Apr 2026
Local Ai

Device-Native Autonomous Agents for Privacy-Preserving Negotiations

DGX agent

arXiv:2601.00911v3 Announce Type: replace-cross Abstract: Automated negotiations in insurance and business-to-business (B2B) commerce encounter substantial challenges. Current systems force a trade-of

local-aiarxiv-cs-ai
23 Apr 2026
Safety

Diagnosing CFG Interpretation in LLMs

DGX agent

arXiv:2604.20811v1 Announce Type: new Abstract: As LLMs are increasingly integrated into agentic systems, they must adhere to dynamically defined, machine-interpretable interfaces. We evaluate LLMs as

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

DGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

DISCA: A Digital In-memory Stochastic Computing Architecture Using A Compressed Bent-Pyramid Format

DGX agent

arXiv:2511.17265v2 Announce Type: replace-cross Abstract: Nowadays, we are witnessing an Artificial Intelligence revolution that dominates the technology landscape in various application domains, such

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

DGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs

DGX agent

arXiv:2604.19765v1 Announce Type: cross Abstract: Recent work identifies a sparse set of 'hallucination neurons' (H-neurons), less than 0.1% of feed-forward network neurons, that reliably predict when

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment

DGX agent

arXiv:2604.19781v1 Announce Type: cross Abstract: Automated scoring of student work at scale requires balancing accuracy against cost and latency. In 'cascade' systems, small language models (LMs) han

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models

DGX agent

arXiv:2604.01965v2 Announce Type: replace-cross Abstract: Scientific knowledge discovery increasingly relies on large language models, yet many existing scholarly assistants depend on proprietary syst

researcharxiv-cs-ai
23 Apr 2026
Agents

DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data

DGX agent

arXiv:2604.19859v1 Announce Type: cross Abstract: Edge-scale deep research agents based on small language models are attractive for real-world deployment due to their advantages in cost, latency, and

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Dual Causal Inference: Integrating Backdoor Adjustment and Instrumental Variable Learning for Medical VQA

DGX agent

arXiv:2604.20306v1 Announce Type: cross Abstract: Medical Visual Question Answering (MedVQA) aims to generate clinically reliable answers conditioned on complex medical images and questions. However,

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

DGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

DGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

DGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

safetyarxiv-cs-ai
23 Apr 2026
Research

Emergence Transformer: Dynamical Temporal Attention Matters

DGX agent

arXiv:2604.19816v1 Announce Type: new Abstract: The Transformer, a breakthrough architecture in artificial intelligence, owes its success to the attention mechanism, which utilizes long-range interact

researcharxiv-cs-ai
23 Apr 2026
Tutorials

Enhancing ASR Performance in the Medical Domain for Dravidian Languages

DGX agent

arXiv:2604.19797v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) for low-resource Dravidian languages like Telugu and Kannada faces significant challenges in specialized medical do

tutorialsarxiv-cs-ai
23 Apr 2026
Agents

Enhancing Research Idea Generation through Combinatorial Innovation and Multi-Agent Iterative Search Strategies

DGX agent

arXiv:2604.20548v1 Announce Type: cross Abstract: Scientific progress depends on the continual generation of innovative re-search ideas. However, the rapid growth of scientific literature has greatly

agentsarxiv-cs-ai
23 Apr 2026
Research

Enhancing Speaker Verification with Whispered Speech via Post-Processing

DGX agent

arXiv:2604.20229v1 Announce Type: cross Abstract: Speaker verification is a task of confirming an individual's identity through the analysis of their voice. Whispered speech differs from phonated spee

researcharxiv-cs-ai
23 Apr 2026
Safety

Environmental Understanding Vision-Language Model for Embodied Agent

DGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

safetyarxiv-cs-ai
23 Apr 2026
Safety

Epistemic Constitutionalism Or: how to avoid coherence bias

DGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

safetyarxiv-cs-ai
23 Apr 2026
Safety

Epistemology gives a Future to Complementarity in Human-AI Interactions

DGX agent

arXiv:2601.09871v2 Announce Type: replace Abstract: Human-AI complementarity is the claim that a human supported by an AI system can outperform either alone in a decision-making process. Since its int

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

DGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

DGX agent

arXiv:2604.20133v1 Announce Type: new Abstract: This paper proposes EvoAgent - an evolvable large language model (LLM) agent framework that integrates structured skill learning with a hierarchical sub

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

EvoForest: A Novel Machine-Learning Paradigm via Open-Ended Evolution of Computational Graphs

DGX agent

arXiv:2604.19761v1 Announce Type: new Abstract: Modern machine learning is still largely organized around a single recipe: choose a parameterized model family and optimize its weights. Although highly

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts

DGX agent

arXiv:2604.19835v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for scaling large language models: frontier models routinely decouple total parameters f

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Explainability in Generative Medical Diffusion Models: A Faithfulness-Based Analysis on MRI Synthesis

DGX agent

arXiv:2602.09781v2 Announce Type: replace-cross Abstract: This study investigates the explainability of generative diffusion models in the context of medical imaging, focusing on Magnetic resonance im

applicationsarxiv-cs-ai
23 Apr 2026
Safety

Explainable AML Triage with LLMs: Evidence Retrieval and Counterfactual Checks

DGX agent

arXiv:2604.19755v1 Announce Type: new Abstract: Anti-money laundering (AML) transaction monitoring generates large volumes of alerts that must be rapidly triaged by investigators under strict audit an

safetyarxiv-cs-ai
23 Apr 2026
Safety

Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias

DGX agent

arXiv:2604.19763v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) systems have growing applications in sensitive domains such as mental health and education, where biased predictions

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization

DGX agent

arXiv:2604.20726v1 Announce Type: cross Abstract: This work explores the role of prompt design and judge selection in LLM-as-a-Judge evaluations of free text legal question answering. We examine wheth

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploring Data Augmentation and Resampling Strategies for Transformer-Based Models to Address Class Imbalance in AI Scoring of Scientific Explanations in NGSS Classroom

DGX agent

arXiv:2604.19754v1 Announce Type: new Abstract: Automated scoring of students' scientific explanations offers the potential for immediate, accurate feedback, yet class imbalance in rubric categories p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Fairness Testing of Large Language Models in Role-Playing

DGX agent

arXiv:2411.00585v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become foundational in modern language-driven software applications, profoundly influencing daily life. A cr

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

FeDa4Fair: Client-Level Federated Datasets for Fairness Evaluation

DGX agent

arXiv:2506.21095v4 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative training while preserving privacy, yet it introduces a critical challenge: the 'illusion of fair

model-releasesarxiv-cs-ai
23 Apr 2026
← Previous
1…390391392393394…443
Next →