AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
4 Jun 2026

AI from concrete to abstract: demystifying artificial intelligence to the general public

ResearchDGX agent

arXiv:2006.04013v6 Announce Type: cross Abstract: Artificial Intelligence (AI) has been adopted in a wide range of domains. This shows the imperative need to develop means to endow common people with

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety

Model ReleasesDGX agent

arXiv:2606.04867v1 Announce Type: new Abstract: As AI companion platforms such as Replika and Character.AI rapidly grow, concerns about unsafe human-AI interactions have intensified. This study introd

AIP: A Graph Representation for Learning and Governing Agent Skills

Model ReleasesDGX agent

arXiv:2606.04781v1 Announce Type: new Abstract: Agent Skills today consist largely of free-form prose requiring the agent to read, interpret, and re-derive how to act in every session. This imposes tw


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms

Model ReleasesDGX agent

arXiv:2602.09464v2 Announce Type: replace-cross Abstract: Vericoding refers to the generation of formally verified code from rigorous specifications. Recent AI models show promise in vericoding, but a

Aligning Deep Implicit Preferences by Learning to Reason Defensively

Model ReleasesDGX agent

arXiv:2510.11194v3 Announce Type: replace Abstract: Personalized alignment is crucial for enabling Large Language Models (LLMs) to engage effectively in user-centric interactions. However, current met

An Empirical Audit of Input Encoders for Multi-Channel Signal Transformers

Model ReleasesDGX agent

arXiv:2606.04752v1 Announce Type: cross Abstract: Transformers consuming multi-channel scalar signals must embed C simultaneous values into one d_{ext{model}}-dimensional vector per time step. We empi

An Empirical Study of Data Scale, Model Complexity, and Input Modalities in Visual Generalization

Model ReleasesDGX agent

arXiv:2606.04409v1 Announce Type: cross Abstract: Modern deep neural networks usually have large parameter scales and nonlinear hierarchical structures, and they have achieved strong performance in co

An Ensembled Latent Factor Model via Differential Evolution and Gradient Descent Optimization

ApplicationsDGX agent

arXiv:2606.04408v1 Announce Type: cross Abstract: High-dimensional and incomplete (HDI) data are prevalent in many real-world big data scenarios. Latent factor models serve as a common representation

ANN Search: Recall What Matters

Model ReleasesDGX agent

arXiv:2606.04522v1 Announce Type: cross Abstract: Approximate nearest neighbor (ANN) search has become a core primitive in information retrieval and modern machine learning tasks, from classification

Anycast Performance in Context

SafetyDGX agent

arXiv:2606.04298v1 Announce Type: cross Abstract: IP anycast lets a service advertise one address from many physical sites, leaving BGP to map each client to a site. It is central to the DNS root serv

Archi: Agentic Operations at the CMS Experiment

AgentsDGX agent

arXiv:2606.04755v1 Announce Type: cross Abstract: We present Archi, an open-source, end-to-end framework for scientific collaborations that combines the systematic ingestion and organization of hetero

Arithmetic Pedagogy for Language Models

TutorialsDGX agent

arXiv:2606.05106v1 Announce Type: cross Abstract: We investigate whether methods of human mathematics pedagogy can guide the training of language models toward arithmetic reasoning. Building on the GA

AttnRegDeepLab: A Two-Stage Decoupled Framework for Interpretable Embryo Fragmentation Grading

ResearchDGX agent

arXiv:2511.18454v3 Announce Type: replace-cross Abstract: Embryo fragmentation is a morphological indicator critical for evaluating developmental potential in In Vitro Fertilization (IVF). However, ma

Audio Interaction Model

ResearchDGX agent

arXiv:2606.05121v1 Announce Type: cross Abstract: Audio is an inherently interactive modality, yet today's Large Audio Language Models (LALMs) are offline, and streaming audio models each handle only

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

Model ReleasesDGX agent

arXiv:2606.05080v1 Announce Type: new Abstract: Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and c

Automatic Generation of Titles for Research Papers Using Language Models

Model ReleasesDGX agent

arXiv:2606.05085v1 Announce Type: cross Abstract: The title of a research paper conveys its primary idea and, occasionally, its conclusions in a clear and concise manner. Choosing an appropriate title

Bayes-Sufficient Representations in Supervised Learning

ResearchDGX agent

arXiv:2606.04045v1 Announce Type: cross Abstract: Representation learning is often described as preserving the information in an input that is relevant for prediction. This work asks what relevance me

Beyond Objective Equivalence: Constraint Injection for LLM-Based Optimization Modeling on Vehicle Routing Problems

Model ReleasesDGX agent

arXiv:2606.04816v1 Announce Type: new Abstract: Large language models (LLMs) increasingly translate natural-language optimization problems into executable solver code. Yet for constraint-dense operati

Beyond Pixel Histories: World Models with Persistent 3D State

ResearchDGX agent

arXiv:2603.03482v2 Announce Type: replace-cross Abstract: Interactive world models continually generate video by responding to a user's actions, enabling open-ended generation capabilities. However, e

Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System

AgentsDGX agent

arXiv:2606.04494v1 Announce Type: new Abstract: Biomedical agents promise to automate complex biological workflows, yet current systems face two fundamental bottlenecks: bioinformatics tools are highl

Beyond Static Priors: Dynamic Neural Guidance for Large-Scale Ant Colony Optimization

SafetyDGX agent

arXiv:2606.04039v1 Announce Type: cross Abstract: Neural-guided Ant Colony Optimization (ACO) suffers from a fundamental training-inference misalignment: policies are typically trained to generate sta

BiasGRPO: Stabilizing Bias Mitigation in High-Variance Reward Landscapes via Group-Relative Policy Optimization

SafetyDGX agent

arXiv:2606.04807v1 Announce Type: new Abstract: Mitigating social bias in Large Language Models (LLMs) presents a distinct alignment challenge: unlike verifiable tasks, bias lacks a single ground trut

Bilevel Autoresearch: Meta-Autoresearching Itself

Model ReleasesDGX agent

arXiv:2603.23420v2 Announce Type: replace Abstract: If autoresearch is itself a form of research, then autoresearch can be applied to research itself. We present Bilevel Autoresearch, a bilevel framew

BiNSGPS: Geometry Problem Solving via Bidirectional Neuro-Symbolic Interaction

ResearchDGX agent

arXiv:2606.04648v1 Announce Type: new Abstract: Geometry problem solving poses distinct challenges in artificial intelligence. Existing approaches typically fall into two paradigms: symbolic methods,

BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation format

Model ReleasesDGX agent

arXiv:2509.02655v3 Announce Type: replace-cross Abstract: Many AI alignment discussions of 'runaway optimisation' focus on RL agents: unbounded utility maximisers that over-optimise a proxy objective

Bounded Hyperbolic Tangent: A Stable and Efficient Alternative to Pre-Layer Normalization in Large Language Models

ResearchDGX agent

arXiv:2601.09719v3 Announce Type: replace-cross Abstract: Pre-Layer Normalization (Pre-LN) is the de facto choice for large language models (LLMs) and is crucial for stable pretraining and effective t

BRAINCELL-AID: An Agentic AI Created Brain Cell Type Resource for Community Annotation

AgentsDGX agent

arXiv:2510.17064v4 Announce Type: replace Abstract: Single-cell RNA sequencing has transformed our ability to identify diverse cell types and their transcriptomic signatures. However, annotating these

Breaking Bad Molecules: Are MLLMs Ready for Structure-Level Molecular Detoxification?

Model ReleasesDGX agent

arXiv:2506.10912v4 Announce Type: replace Abstract: Toxicity remains a leading cause of early-stage drug development failure. Despite advances in molecular design and property prediction, the task of

Building The Ph(ysical)AI Layer Of Machine Intelligence

Model ReleasesDGX agent

arXiv:2606.04106v1 Announce Type: cross Abstract: Foundation models achieve generalization through massive-scale training on diverse data, but have limitations with transfer to truly unseen domains wi

Can Generalist Agents Automate Data Curation?

Model ReleasesDGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

Can I Take Another Dose? Evaluating LLM Decision-Making Under Temporal Uncertainty in OTC Dosing QA

Model ReleasesDGX agent

arXiv:2606.04262v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for everyday health questions, including whether a user can safely take another dose of an over-the

Can Reasoning Path still be Effective as Input? Bridging Post-Reasoning to Chain-of-Thought Compression

ResearchDGX agent

arXiv:2510.08647v2 Announce Type: replace-cross Abstract: Recent developments have enabled advanced reasoning in Large Language Models (LLMs) via long Chain-of-Thought (CoT), trading efficiency during

Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics

TutorialsDGX agent

arXiv:2506.06006v3 Announce Type: replace-cross Abstract: Can unified vision-language models (VLMs) perform forward dynamics prediction (FDP), i.e., predicting the future state (in image form) given t

Cascading Hallucination in Agentic RAG: The CHARM Framework for Detection and Mitigation

AgentsDGX agent

arXiv:2606.04435v1 Announce Type: new Abstract: Multi-step agentic retrieval-augmented generation (RAG) pipelines have demonstrated significant capability for complex reasoning tasks, yet remain vulne

Caught in the Act(ivation): Toward Pre-Output and Multi-Turn Detection of Credential Exfiltration by LLM Agents

Model ReleasesDGX agent

arXiv:2606.04141v1 Announce Type: cross Abstract: LLM agents often place sensitive credentials in the same context window as untrusted retrieved content, creating a direct path for indirect prompt inj

Channel-Oriented Design for EEG-to-Music Reconstruction

SafetyDGX agent

arXiv:2606.04040v1 Announce Type: cross Abstract: Brain-computer interfaces aim to decode naturalistic stimuli from neural signals, yet most progress to date has focused on vision and language. In thi

Characterizing initial human-AI proof formalization workflows

ResearchDGX agent

arXiv:2606.04273v1 Announce Type: new Abstract: For centuries, human mathematicians have written proofs to substantiate their mathematical arguments; yet, the ability to automatically verify the valid

ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents

ResearchDGX agent

arXiv:2407.03884v4 Announce Type: replace-cross Abstract: Dialogue agents powered by Large Language Models (LLMs) show superior performance in various tasks. Despite the better user understanding and

ChessMimic: Per-Rating Transformer Models for Human Move, Clock, and Outcome Prediction in Online Blitz Chess

Model ReleasesDGX agent

arXiv:2606.04473v1 Announce Type: cross Abstract: We present ChessMimic, a system of three small encoder-only transformers - for move, thinking-time, and outcome prediction - conditioned on the positi

ClustRecNet: A Novel End-to-End Deep Learning Framework for Clustering Algorithm Recommendation

ApplicationsDGX agent

arXiv:2509.25289v4 Announce Type: replace-cross Abstract: Identifying an effective clustering algorithm for a given dataset remains a fundamental unsupervised learning issue. We introduce ClustRecNet,

Coarse-to-fine Hierarchical Architecture with Sequential Mamba for Brain Reconstruction

Local AiDGX agent

arXiv:2606.04772v1 Announce Type: cross Abstract: Understanding the relationship between deep visual representations and the human visual system is a fundamental challenge in computational neuroscienc

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

Model ReleasesDGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

Conditional PED-ANOVA: Hyperparameter Importance in Hierarchical & Dynamic Search Spaces

ResearchDGX agent

arXiv:2601.20800v3 Announce Type: replace-cross Abstract: We propose conditional PED-ANOVA (condPED-ANOVA), a principled framework for estimating hyperparameter importance (HPI) in conditional search

Consensus is Strategically Insufficient: Reasoning-Trace Disagreement as a Knowledge-Representation Signal

AgentsDGX agent

arXiv:2606.04223v1 Announce Type: new Abstract: Multi-agent systems are commonly designed to reduce disagreement through voting, consensus protocols, debate, or fault-tolerant aggregation. We argue th

Constrained Adaptive Rejection Sampling

ResearchDGX agent

arXiv:2510.01902v2 Announce Type: replace Abstract: Language Models (LMs) are increasingly used in applications where generated outputs must satisfy strict semantic or syntactic constraints. Existing

Constraint-Enhanced Physical Search through Correlation Matching

Model ReleasesDGX agent

arXiv:2606.03554v1 Announce Type: cross Abstract: Physical systems do not merely add noise to search processes; they impose constraints that generate structured correlations. We propose a principle of

ContactExplorer: Contact Coverage-Guided Exploration for General-Purpose Dexterous Manipulation

AgentsDGX agent

arXiv:2603.10971v2 Announce Type: replace-cross Abstract: Reinforcement learning has achieved remarkable success in domains such as Atari games, navigation, and locomotion, where exploration can often

Continual Visual and Verbal Learning Through a Child's Egocentric Input

Model ReleasesDGX agent

arXiv:2606.05115v1 Announce Type: cross Abstract: Children learn the meanings of words from a continuous, temporally structured stream of egocentric experience. Recent work shows that neural networks

CoRe-MoE: Contrastive Reweighted Mixture of Experts for Multi-Terrain Humanoid Locomotion with Gait Adaptation

SafetyDGX agent

arXiv:2606.04718v1 Announce Type: cross Abstract: Humans primarily rely on walking and running to traverse complex terrains, without resorting to unnecessarily complex motion patterns. Similarly, huma

CounterFace: A Synthetic Face Dataset for Fine-Grained Counterfactual Evaluation of Face Recognition Systems

ResearchDGX agent

arXiv:2407.13922v3 Announce Type: replace-cross Abstract: Face recognition (FR) systems are widely deployed in critical applications, making their reliability and robustness across diverse populations

Counterfactual Explanations for Deep Two-Sample Testing

ResearchDGX agent

arXiv:2606.04009v1 Announce Type: cross Abstract: Two-sample testing is a fundamental tool for detecting distributional differences across scientific domains, but classical tests (including kernel-bas

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

SafetyDGX agent

arXiv:2601.22396v2 Announce Type: replace-cross Abstract: Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accura

Curvature-aware dynamic precision approach for physics-informed neural networks

Model ReleasesDGX agent

arXiv:2606.04736v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have become a promising framework for simulating partial differential equations (PDEs) by embedding physical

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

Model ReleasesDGX agent

arXiv:2606.04460v1 Announce Type: cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. How

DAR: Deontic Reasoning with Agentic Harnesses

AgentsDGX agent

arXiv:2606.05009v1 Announce Type: cross Abstract: Deontic reasoning is the task of answering questions by applying explicit rules and policies to case-specific facts, for example computing tax liabili

DeliChess: A Multi-party Dialogue Dataset for Deliberation in Chess Puzzle Solving

ResearchDGX agent

arXiv:2606.04987v1 Announce Type: cross Abstract: Multi-party dialogue is a critical setting for studying collaborative reasoning and decision-making, yet existing datasets rarely focus on structured,

Description-Code Inconsistency in Real-world MCP Servers: Measurement, Detection, and Security Implications

AgentsDGX agent

arXiv:2606.04769v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has emerged as a critical standard empowering Large Language Models (LLMs) to utilize external tools. In this ecosyst

DetectZoo: A Unified Toolkit for AI-Generated Content Detection Across Text, Audio, and Image Modalities

Model ReleasesDGX agent

arXiv:2606.04205v1 Announce Type: cross Abstract: The growing popularity and capacity of generative models have eroded the distinction between human and machine-generated content, motivating a growing

DiffAero: A GPU-Accelerated Differentiable Simulation Framework for Efficient Quadrotor Policy Learning

SafetyDGX agent

arXiv:2509.10247v1 Announce Type: cross Abstract: This letter introduces DiffAero, a lightweight, GPU-accelerated, and fully differentiable simulation framework designed for efficient quadrotor contro

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation

Model ReleasesDGX agent

arXiv:2606.04046v1 Announce Type: cross Abstract: In embodied vision-language decision making tasks such as robotic manipulation and navigation, Vision-Language and Vision-Language-Action Models (VLMs

← Previous
1…158159160161162…358
Next →