AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
18 May 2026

ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment

SafetyDGX agent

arXiv:2505.19241v2 Announce Type: replace-cross Abstract: The recent success in using human preferences to align large language models (LLMs) has significantly improved their performance in various do

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making

SafetyDGX agent

arXiv:2605.16054v1 Announce Type: cross Abstract: Recent work has framed decision-making as a sequence modeling problem using generative models such as diffusion models. Although promising, these appr

ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking

ApplicationsDGX agent

arXiv:2510.13842v2 Announce Type: replace-cross Abstract: Knowledge poisoning poses a critical threat to Retrieval-Augmented Generation (RAG) systems by injecting adversarial content into knowledge ba


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Agent4POI: Agentic Context-Conditioned Affordance Reasoning for Multimodal Point-of-Interest Recommendation

AgentsDGX agent

arXiv:2605.15203v1 Announce Type: cross Abstract: We introduce Agent4POI, the first POI recommendation framework that generates context-conditioned multimodal representations at recommendation time, r

Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design

Model ReleasesDGX agent

arXiv:2605.15871v1 Announce Type: new Abstract: Toward recursive self-improvement, we investigate LLM agents autonomously designing foundation models beyond standard Transformers. We introduce a dual-

AgentStop: Terminating Local AI Agents Early to Save Energy in Consumer Devices

Local AiDGX agent

arXiv:2605.15206v1 Announce Type: cross Abstract: Autonomous agents powered by large language models (LLMs) are increasingly used to automate complex, multi-step tasks such as coding or web-based ques

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

HardwareDGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

AI Consciousness and Existential Risk

SafetyDGX agent

arXiv:2511.19115v2 Announce Type: replace Abstract: In AI, the existential risk denotes the hypothetical threat posed by an artificial system that would possess both the capability and the objective,

AI-Mediated Communication Can Steer Collective Opinion

SafetyDGX agent

arXiv:2605.16245v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is increasingly integrated into the online platforms where humans exchange opinions; large language models (LL

alpha-TCAV: A Unified Framework for Testing with Concept Activation Vectors

Model ReleasesDGX agent

arXiv:2605.15688v1 Announce Type: cross Abstract: Concept Activation Vectors (CAVs) are a fundamental tool for concept-based explainability in deep learning, yet their practical utility is limited by

ALSO: Adversarial Online Strategy Optimization for Social Agents

Model ReleasesDGX agent

arXiv:2605.15768v1 Announce Type: new Abstract: Social simulation provides a compelling testbed for studying social intelligence, where agents interact through multi-turn dialogues under evolving cont

Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time

SafetyDGX agent

arXiv:2605.15220v1 Announce Type: cross Abstract: Data mixing decides how to combine different sources or types of data and is a consequential problem throughout language model training. In pretrainin

Amortized Energy-Based Bayesian Inference

ResearchDGX agent

arXiv:2605.15407v1 Announce Type: cross Abstract: We consider amortized Bayesian inference for nonlinear inverse problems in settings where only samples from the joint distribution of parameters and o

An Algebraic Exposition of the Theory of Dyadic Morality

SafetyDGX agent

arXiv:2605.16153v1 Announce Type: new Abstract: This paper provides an algebraic exposition of the theory of dyadic morality (TDM), a psychological model of moral judgment grounded in a simple two-nod

An LLM-RAG Approach for Healthy Eating Index-Informed Personalized Food Recommendations

ResearchDGX agent

arXiv:2605.15213v1 Announce Type: cross Abstract: Diet quality is a leading determinant of chronic disease risk. Advances in artificial intelligence (AI) have enabled food recommendation systems to ad

Antidistillation Fingerprinting

ResearchDGX agent

arXiv:2602.03812v2 Announce Type: replace-cross Abstract: Model distillation enables efficient emulation of frontier large language models (LLMs), creating a need for robust mechanisms to detect when

Approximate and Weighted Data Reconstruction Attack in Federated Learning

Local AiDGX agent

arXiv:2308.06822v3 Announce Type: replace-cross Abstract: Federated Learning (FL) is a distributed learning paradigm that enables multiple clients to collaborate on building a machine learning model w

Argus: Evidence Assembly for Scalable Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.16217v1 Announce Type: cross Abstract: Deep research agents have achieved remarkable progress on complex information seeking tasks. Even long ReAct style rollouts explore only a single traj

Asking the Right Questions: Improving Reasoning with Generated Stepping Stones

ResearchDGX agent

arXiv:2602.19069v2 Announce Type: replace Abstract: Recent years have witnessed tremendous progress in enabling LLMs to solve complex reasoning tasks such as math and coding. As we start to apply LLMs

ASRU: Activation Steering Meets Reinforcement Unlearning for Multimodal Large Language Models

SafetyDGX agent

arXiv:2605.15687v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) may memorize sensitive cross-modal information during pretraining, making machine unlearning (MU) crucial. Ex

AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs

SafetyDGX agent

arXiv:2605.15565v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly used to improve the reasoning, coding, and tool-use capabilities of large language models, but agentic RL

Attention Dispersion in Dynamic Graph Transformers: Diagnosis and a Transferable Fix

ResearchDGX agent

arXiv:2605.16112v1 Announce Type: cross Abstract: Transformer-based architectures have become the dominant paradigm for Continuous-Time Dynamic Graph (CTDG) learning, yet their performance remains lim

Autoguided Online Data Curation for Diffusion Model Training

ResearchDGX agent

arXiv:2509.15267v2 Announce Type: replace-cross Abstract: The costs of generative model compute rekindled promises and hopes for efficient data curation. In this work, we investigate whether recently

Autonomous Intelligent Agents for Natural-Language-Driven Web Execution with Integrated Security Assurance

AgentsDGX agent

arXiv:2605.15281v1 Announce Type: cross Abstract: Modern web test suites rot. A UI refactor breaks locators, a timing change causes race conditions, and within weeks developers abandon the suite entir

Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation

Model ReleasesDGX agent

arXiv:2605.15343v1 Announce Type: new Abstract: LLM-based agents are increasingly used to simulate deliberative interactions such as negotiation, conflict resolution, and multi-turn opinion exchange.

Benchmark of Benchmarks: Unpacking Influence and Code Repository Quality in LLM Safety Benchmarks

Model ReleasesDGX agent

arXiv:2603.04459v3 Announce Type: replace-cross Abstract: The rapid expansion of research in LLM safety presents challenges in tracking advancements, making benchmarks important evaluation infrastruct

Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty

ResearchDGX agent

arXiv:2507.16806v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained via reinforcement learning (RL) to generate natural language 'reasoning chains', their performance impr

Beyond Content: A Comprehensive Speech Toxicity Dataset and Detection Framework Incorporating Paralinguistic Cues

ResearchDGX agent

arXiv:2605.15984v1 Announce Type: cross Abstract: Toxic speech detection has become a crucial challenge in maintaining safe online communication environments. However, existing approaches to toxic spe

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming

AgentsDGX agent

arXiv:2605.15400v1 Announce Type: new Abstract: While AI agents are rapidly advancing from isolated tools to interactive collaborators, data-driven human-machine teaming (HMT) methods remain costly in

Bidirectional Fusion Guided by Cardiac Patterns for Semi-Supervised ECG Segmentation

Model ReleasesDGX agent

arXiv:2605.15722v1 Announce Type: cross Abstract: Accurate delineation of electrocardiogram (ECG), the segmentation of meaningful waveform features, is crucial for cardiovascular diagnostics. However,

BioBlobs: Unsupervised Discovery of Functional Substructures for Protein Function Prediction

ResearchDGX agent

arXiv:2510.01632v3 Announce Type: replace-cross Abstract: Protein function is driven by cohesive substructures, such as catalytic triads, binding pockets, and structural motifs, that occupy only a sma

BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation

Model ReleasesDGX agent

arXiv:2605.15736v1 Announce Type: cross Abstract: Biomedical Vision--Language Models (VLMs) have shown remarkable promise in few-shot medical diagnosis but face a critical bottleneck: extit{fragility

Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling

SafetyDGX agent

arXiv:2507.01679v3 Announce Type: replace-cross Abstract: Existing LLMs-post-training techniques are broadly categorized into supervised fine-tuning (SFT) and reinforcement fine-tuning (RFT). Each par

Brain-OF: An Omnifunctional Foundation Model for fMRI, EEG and MEG

ResearchDGX agent

arXiv:2602.23410v3 Announce Type: replace-cross Abstract: Brain foundation models have achieved remarkable advances across a wide range of neuroscience tasks. However, most existing models are limited

Breakeven complexity: A new perspective on neural partial differential equation solvers

Model ReleasesDGX agent

arXiv:2605.15399v1 Announce Type: cross Abstract: Neural surrogate solvers of partial differential equations (PDEs) promise dramatic speedups over numerical methods, especially in scenarios requiring

Bridging Silicon and the Hippocampus: Algebro-Deterministic Memory 'VaCoAl' as a Substrate for Vector-HaSH and TEM

HardwareDGX agent

arXiv:2605.15652v1 Announce Type: cross Abstract: Vector-HaSH and the Tolman-Eichenbaum Machine (TEM) propose that the hippocampal-entorhinal circuit factorizes content from a prestructured grid-cell

Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study

ApplicationsDGX agent

arXiv:2605.16011v1 Announce Type: cross Abstract: Adaptive learning refers to educational technologies that track learners' learning progress and adapt the instructional process based on individual le

Can We Trust AI-Inferred User States. A Psychometric Framework for Validating the Reliability of Users States Classification by LLMs in Operational Environments

Model ReleasesDGX agent

arXiv:2605.15734v1 Announce Type: new Abstract: The use of large language models to assess user states in conversational and adaptive systems is based on the assumption that the metrics used for such

CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning

ResearchDGX agent

arXiv:2605.15513v1 Announce Type: new Abstract: Parallel reasoning, where a generator samples many candidate solutions and an aggregator selects the best, is one of the most effective forms of test-ti

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

Model ReleasesDGX agent

arXiv:2605.15218v1 Announce Type: new Abstract: Large language models deployed for MAPDL finite-element simulation face practical reliability challenges: without structured execution control, tool enc

Centralized vs Decentralized Federated Learning: A trade-off performance analysis

ResearchDGX agent

arXiv:2605.16089v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for collaborative model training across distributed edge devices while preserving data pri

ChangeFlow -- Latent Rectified Flow for Change Detection in Remote Sensing

Local AiDGX agent

arXiv:2605.15375v1 Announce Type: cross Abstract: Remote sensing change detection (RSCD) aims to localise changes between two images of the same geographic region. In practice, change masks often foll

CHoE: Cross-Domain Heterogeneous Graph Prompt Learning via Structure-Conditioned Experts

ApplicationsDGX agent

arXiv:2605.15888v1 Announce Type: cross Abstract: Heterogeneous Graph Prompt Learning (HGPL)has emerged as a promising paradigm for bridging the gap between the objectives of pre-training foundation m

CIS-BWE: Chaos-Informed Speech Bandwidth Extension

Model ReleasesDGX agent

arXiv:2507.15970v3 Announce Type: replace-cross Abstract: Recovering high-frequency components lost to bandwidth constraints is crucial for applications ranging from telecommunications to high-fidelit

CitePrism: Human-in-the-Loop AI for Citation Auditing and Editorial Integrity

Local AiDGX agent

arXiv:2605.16000v1 Announce Type: cross Abstract: Editors and reviewers are expected to ensure that manuscripts cite relevant, accurate, current, and ethically appropriate literature, yet manuscript-l

COCO-Inpaint: A Benchmark for Detecting and Localizing Inpainting-Based Image Manipulations

Model ReleasesDGX agent

arXiv:2504.18361v2 Announce Type: replace-cross Abstract: Recent advances in image manipulation have enabled highly photorealistic content generation, but also lowered the barrier to arbitrary editing

CodeDistiller: Automatically Generating Code Libraries for Scientific Coding Agents

AgentsDGX agent

arXiv:2512.01089v2 Announce Type: replace Abstract: Automated Scientific Discovery (ASD) systems can help automatically generate and run code-based experiments, but their capabilities are limited by t

ColPackAgent: Agent-Skill-Guided Hard-Particle Monte Carlo Workflows for Colloidal Packing

Model ReleasesDGX agent

arXiv:2605.15625v1 Announce Type: new Abstract: We introduce ColPackAgent, an agent framework that autonomously runs Monte Carlo simulations of colloidal packing through a Model Context Protocol (MCP)

CompactQE: Interpretable Translation Quality Estimation via Small Open-Weight LLMs

ResearchDGX agent

arXiv:2605.15763v1 Announce Type: cross Abstract: Current state-of-the-art Quality Estimation (QE) in machine translation relies on massive, proprietary LLMs, raising data privacy concerns. We demonst

Confirming Correct, Missing the Rest: LLM Tutoring Agents Struggle Where Feedback Matters Most

Model ReleasesDGX agent

arXiv:2605.16207v1 Announce Type: new Abstract: Effective tutoring requires distinguishing optimal, valid but suboptimal, and incorrect student solutions, a distinction central to intelligent tutoring

Constrained latent state modeling: A unifying perspective on representation learning under competing constraints

TutorialsDGX agent

arXiv:2605.15995v1 Announce Type: cross Abstract: Learning latent representations from complex data is central to modern machine learning, spanning temporal, multimodal, and partially observed systems

Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning

AgentsDGX agent

arXiv:2605.15315v1 Announce Type: new Abstract: LLM-powered coding agents spend the majority of their token budget reading repository files, yet much of the retrieved code is irrelevant to the task at

Context, Reasoning, and Hierarchy: A Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP

AgentsDGX agent

arXiv:2605.16205v1 Announce Type: new Abstract: Deploying compound LLM agents in adversarial, partially observable sequential environments requires navigating several design dimensions: (1) what the a

Convergent Representations of Linguistic Constructions in Human and Artificial Neural Systems

ResearchDGX agent

arXiv:2603.29617v2 Announce Type: replace-cross Abstract: Understanding how the brain processes linguistic constructions is a central challenge in cognitive neuroscience and linguistics. Recent comput

CTF4Nuclear: Common Task Framework for Nuclear Fission and Fusion Models

SafetyDGX agent

arXiv:2605.15549v1 Announce Type: cross Abstract: The demand for clean energy is ever increasing, with new nuclear technologies presenting a complementary solution to renewable energies. However, desi

CUBE: Contrastive Understanding by Balanced Experiments

ResearchDGX agent

arXiv:2509.10825v5 Announce Type: replace-cross Abstract: Explaining a trained model requires a clear account of how explanatory evidence is generated. We propose CUBE, a post-hoc explanation framewor

DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2605.16113v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved unprecedented success due to their exceptional generative capabilities. However, because they depend on kno

DecomPose: Disentangling Cross-Category Optimization Contention for Category-Level 6D Object Pose Estimation

ResearchDGX agent

arXiv:2605.15728v1 Announce Type: cross Abstract: Category-level 6D object pose estimation is typically formulated as a multi-category joint learning problem with fully shared model parameters. Howeve

Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation

SafetyDGX agent

arXiv:2605.15942v1 Announce Type: cross Abstract: Open-vocabulary segmentation models often struggle to generalize to unseen combinations of object categories and attributes, because fine-grained desc

Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training

Model ReleasesDGX agent

arXiv:2602.00747v2 Announce Type: replace-cross Abstract: Determining an effective data mixture is a key factor in Large Language Model (LLM) pre-training, where models must balance general competence

← Previous
1…245246247248249…358
Next →