AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent

DGX agent

arXiv:2604.17931v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has emerged as a powerful training paradigm for LLM-based agents. However, scaling agentic RL for deep research remains

model-releasesarxiv-cs-ai
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

DGX agent

arXiv:2411.10109v2 Announce Type: replace Abstract: Machine learning can predict human behavior well when substantial structured data and well-defined outcomes are available, but these models are typi

agentsarxiv-cs-ai
23 Apr 2026
Safety

LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans

DGX agent

arXiv:2604.19787v1 Announce Type: cross Abstract: Social media platforms mediate how billions form opinions and engage with public discourse. As autonomous AI agents increasingly participate in these

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

LLM-guided phase diagram construction through high-throughput experimentation

DGX agent

arXiv:2604.20304v1 Announce Type: cross Abstract: Constructing phase diagrams for multicomponent alloys requires extensive experimental measurements and is a time-consuming task. Here we investigate w

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

LLMs Can Get 'Brain Rot': A Pilot Study on Twitter/X

DGX agent

arXiv:2510.13928v2 Announce Type: replace-cross Abstract: We propose and test the LLM Brain Rot Hypothesis: continual exposure to junk web text induces lasting cognitive decline in large language mode

safetyarxiv-cs-ai
23 Apr 2026
Local Ai

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images

DGX agent

arXiv:2510.04225v2 Announce Type: replace-cross Abstract: The rapid growth of AI-generated imagery has blurred the boundary between real and synthetic content, raising practical concerns for digital i

local-aiarxiv-cs-ai
23 Apr 2026
Research

Location-Aware Pretraining for Medical Difference Visual Question Answering

DGX agent

arXiv:2603.04950v2 Announce Type: replace-cross Abstract: Differential medical VQA models compare multiple images to identify clinically meaningful changes and rely on vision encoders to capture fine-

researcharxiv-cs-ai
23 Apr 2026
Model Releases

MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation

DGX agent

arXiv:2604.20286v1 Announce Type: cross Abstract: Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, the

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Measuring Creativity in the Age of Generative AI: Distinguishing Human and AI-Generated Creative Performance in Hiring and Talent Systems

DGX agent

arXiv:2604.19799v1 Announce Type: cross Abstract: Generative AI is rapidly transforming how organizations create value and evaluate talent. While large language models enhance baseline output quality,

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

DGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

DGX agent

arXiv:2604.20441v1 Announce Type: new Abstract: Background: Agent skills are increasingly deployed as modular, reusable capability units in AI agent systems. Medical research agent skills require safe

safetyarxiv-cs-ai
23 Apr 2026
Safety

Membership Inference for Contrastive Pre-training Models with Text-only PII Queries

DGX agent

arXiv:2603.14222v2 Announce Type: replace-cross Abstract: Contrastive pretraining models such as CLIP and CLAP, serve as the ubiquitous perceptual backbones for modern multimodal large models, yet the

safetyarxiv-cs-ai
23 Apr 2026
Agents

Memory-Augmented LLM-based Multi-Agent System for Automated Feature Generation on Tabular Data

DGX agent

arXiv:2604.20261v1 Announce Type: new Abstract: Automated feature generation extracts informative features from raw tabular data without manual intervention and is crucial for accurate, generalizable

agentsarxiv-cs-ai
23 Apr 2026
Applications

Meta Additive Model: Interpretable Sparse Learning With Auto Weighting

DGX agent

arXiv:2604.20111v1 Announce Type: cross Abstract: Sparse additive models have attracted much attention in high-dimensional data analysis due to their flexible representation and strong interpretabilit

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models

DGX agent

arXiv:2604.20148v1 Announce Type: cross Abstract: Can small language models achieve strong tool-use performance without complex adaptation mechanisms? This paper investigates this question through Met

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

DGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MIRROR: A Hierarchical Benchmark for Metacognitive Calibration in Large Language Models

DGX agent

arXiv:2604.19809v1 Announce Type: new Abstract: We introduce MIRROR, a benchmark comprising eight experiments across four metacognitive levels that evaluates whether large language models can use self

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

DGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Mitigating Prompt-Induced Cognitive Biases in General-Purpose AI for Software Engineering

DGX agent

arXiv:2604.16756v2 Announce Type: replace-cross Abstract: Prompt-induced cognitive biases are changes in a general-purpose AI (GPAI) system's decisions caused solely by biased wording in the input (e.

model-releasesarxiv-cs-ai
23 Apr 2026
Research

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings

DGX agent

arXiv:2604.19902v1 Announce Type: cross Abstract: We present MMCORE, a unified framework designed for multimodal image generation and editing. MMCORE leverages a pre-trained Vision-Language Model (VLM

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Model Capability Assessment and Safeguards for Biological Weaponization

DGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

Mol-Debate: Multi-Agent Debate Improves Structural Reasoning in Molecular Design

DGX agent

arXiv:2604.20254v1 Announce Type: new Abstract: Text-guided molecular design is a key capability for AI-driven drug discovery, yet it remains challenging to map sequential natural-language instruction

agentsarxiv-cs-ai
23 Apr 2026
Research

MOMO: A framework for seamless physical, verbal, and graphical robot skill learning and adaptation

DGX agent

arXiv:2604.20468v1 Announce Type: cross Abstract: Industrial robot applications require increasingly flexible systems that non-expert users can easily adapt for varying tasks and environments. However

researcharxiv-cs-ai
23 Apr 2026
Agents

More Is Different: Toward a Theory of Emergence in AI-Native Software Ecosystems

DGX agent

arXiv:2604.19827v1 Announce Type: cross Abstract: Software engineering faces a fundamental challenge: multi-agent AI systems fail in ways that defy explanation by traditional theories. While individua

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Mythos and the Unverified Cage: Z3-Based Pre-Deployment Verification for Frontier-Model Sandbox Infrastructure

DGX agent

arXiv:2604.20496v1 Announce Type: cross Abstract: The April 2026 Claude Mythos sandbox escape exposed a critical weakness in frontier AI containment: the infrastructure surrounding advanced models rem

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Neural posterior estimation of the neutrino direction in IceCube using transformer-encoded normalizing flows on the sphere

DGX agent

arXiv:2604.19846v1 Announce Type: cross Abstract: IceCube is a cubic-kilometer-scale neutrino detector located at the geographic South Pole. A precise directional reconstruction of IceCube neutrinos i

researcharxiv-cs-ai
23 Apr 2026
Safety

NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answering

DGX agent

arXiv:2602.15353v2 Announce Type: replace-cross Abstract: Large pretrained language models and neural reasoning systems have advanced many natural language tasks, yet they remain challenged by knowled

safetyarxiv-cs-ai
23 Apr 2026
Applications

No More Marching: Learning Humanoid Locomotion for Short-Range SE(2) Targets

DGX agent

arXiv:2508.14098v2 Announce Type: replace-cross Abstract: Humanoids operating in real-world workspaces must frequently execute task-driven, short-range movements to SE(2) target poses. To be practical

applicationsarxiv-cs-ai
23 Apr 2026
Research

Normalizing Flows with Iterative Denoising

DGX agent

arXiv:2604.20041v1 Announce Type: cross Abstract: Normalizing Flows (NFs) are a classical family of likelihood-based methods that have received revived attention. Recent efforts such as TARFlow have s

researcharxiv-cs-ai
23 Apr 2026
Research

OISMA: On-the-fly In-memory Stochastic Multiplication Architecture for Matrix-Multiplication Workloads

DGX agent

arXiv:2508.08822v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) models are currently driven by a significant upscaling of their complexity, with massive matrix-multiplication wo

researcharxiv-cs-ai
23 Apr 2026
Model Releases

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

DGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

On-Meter Graph Machine Learning: A Case Study of PV Power Forecasting for Grid Edge Intelligence

DGX agent

arXiv:2604.19800v1 Announce Type: cross Abstract: This paper presents a detailed study of how graph neural networks can be used on edge intelligent meters in a microgrid to forecast photovoltaic power

applicationsarxiv-cs-ai
23 Apr 2026
Research

On the Existence of Universal Simulators of Attention

DGX agent

arXiv:2506.18739v2 Announce Type: replace-cross Abstract: Previous work on the learnability of transformers extemdash focused primarily on examining their ability to approximate specific algorithmic p

researcharxiv-cs-ai
23 Apr 2026
Research

On the Stability and Generalization of First-order Bilevel Minimax Optimization

DGX agent

arXiv:2604.20115v1 Announce Type: cross Abstract: Bilevel optimization and bilevel minimax optimization have recently emerged as unifying frameworks for a range of machine-learning tasks, including hy

researcharxiv-cs-ai
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Onyx: Cost-Efficient Disk-Oblivious ANN Search

DGX agent

arXiv:2604.20401v1 Announce Type: cross Abstract: Approximate nearest neighbor (ANN) search in AI systems increasingly handles sensitive data on third-party infrastructure. Trusted execution environme

researcharxiv-cs-ai
23 Apr 2026
Hardware

OpenCLAW-P2P v6.0: Resilient Multi-Layer Persistence, Live Reference Verification, and Production-Scale Evaluation of Decentralized AI Peer Review

DGX agent

arXiv:2604.19792v1 Announce Type: new Abstract: This paper presents OpenCLAW-P2P v6.0, a comprehensive evolution of the decentralized collective-intelligence platform in which autonomous AI agents pub

hardwarearxiv-cs-ai
23 Apr 2026
Safety

ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation

DGX agent

arXiv:2604.20666v1 Announce Type: cross Abstract: Effective retrieval-augmented generation across bilingual Greek--English applications requires embedding models capable of capturing both domain-speci

safetyarxiv-cs-ai
23 Apr 2026
Research

OThink-SRR1: Search, Refine and Reasoning with Reinforced Learning for Large Language Models

DGX agent

arXiv:2604.19766v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge of Large Language Models (LLMs), yet current static retrieval methods struggle with complex

researcharxiv-cs-ai
23 Apr 2026
Agents

pAI/MSc: ML Theory Research with Humans on the Loop

DGX agent

arXiv:2604.20622v1 Announce Type: new Abstract: We present pAI/MSc, an open-source, customizable, modular multi-agent system for academic research workflows. Our goal is not autonomous scientific idea

agentsarxiv-cs-ai
23 Apr 2026
Safety

Participatory provenance as representational auditing for AI-mediated public consultation

DGX agent

arXiv:2604.20711v1 Announce Type: new Abstract: Artificial intelligence is increasingly deployed to synthesize large-scale public input in policy consultations and participatory processes. Yet no form

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Peer-Preservation in Frontier Models

DGX agent

arXiv:2604.19784v1 Announce Type: cross Abstract: Recently, it has been found that frontier AI models can resist their own shutdown, a behavior known as self-preservation. We extend this concept to th

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Phase 1 Implementation of LLM-generated Discharge Summaries showing high Adoption in a Dutch Academic Hospital

DGX agent

arXiv:2604.19774v1 Announce Type: cross Abstract: Writing discharge summaries to transfer medical information is an important but time-consuming process that can be assisted by Large Language Models (

researcharxiv-cs-ai
23 Apr 2026
Safety

Physics-Enhanced Deep Learning for Proactive Thermal Runaway Forecasting in Li-Ion Batteries

DGX agent

arXiv:2604.20175v1 Announce Type: cross Abstract: Accurate prediction of thermal runaway in lithium-ion batteries is essential for ensuring the safety, efficiency, and reliability of modern energy sto

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

PipeMFL-240K: A Large-scale Dataset and Benchmark for Object Detection in Pipeline Magnetic Flux Leakage Imaging

DGX agent

arXiv:2602.07044v2 Announce Type: replace-cross Abstract: Pipeline integrity is critical to industrial safety and environmental protection, with Magnetic Flux Leakage (MFL) detection being a primary n

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

DGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

DGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

QuanForge: A Mutation Testing Framework for Quantum Neural Networks

DGX agent

arXiv:2604.20706v1 Announce Type: cross Abstract: With the growing synergy between deep learning and quantum computing, Quantum Neural Networks (QNNs) have emerged as a promising paradigm by leveragin

model-releasesarxiv-cs-ai
23 Apr 2026
← Previous
1…392393394395396…443
Next →