AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
28 May 2026

I Hear, Therefore I Trust: A Socio-Technical Investigation of Humans as Synthetic Speech Detectors

ResearchDGX agent

arXiv:2605.28064v1 Announce Type: cross Abstract: Automatic deepfake detection has received considerable research attention, yet the socio-technical environment in which humans actually encounter synt

Identifying and Understanding Human Values in Text: A Tailorable LLM-based Architecture

AgentsDGX agent

arXiv:2605.27373v1 Announce Type: new Abstract: As intelligent systems become more autonomous, the scientific community focuses on creating decision-making mechanisms that include ethical and moral co

Identifying Explicit Parsimonious Piece-wise Polynomial Relationships in Industrial time-series: Application to manipulator robots

Local AiDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.28320v1 Announce Type: cross Abstract: This paper addresses the problem of identifying parsimonious explicit piece-wise polynomial relationships that might involve a relatively large number

Improving Evaluation of Recombination-based Cartesian Genetic Programming

ResearchDGX agent

arXiv:2605.28353v1 Announce Type: cross Abstract: Cartesian Genetic Programming has traditionally been using mutation as its main and often sole genetic operator to drive evolutionary search. Despite

Improving Requirements Classification with SMOTE-Tomek Preprocessing

ResearchDGX agent

arXiv:2501.06491v3 Announce Type: replace-cross Abstract: This study emphasizes the domain of requirements engineering by applying the SMOTE-Tomek preprocessing technique, combined with stratified K-f

InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training

ResearchDGX agent

arXiv:2510.15859v4 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has driven recent breakthroughs in large language models (LLMs), especially for tasks where rewards can be compute

Informing AI Policy Assessment using Large-Scale Simulation of Interventions

SafetyDGX agent

arXiv:2605.27395v1 Announce Type: cross Abstract: As the rapid proliferation of AI systems and harms spurs efforts in AI governance around the world, prioritizing among competing policy options has be

Integrated and Cross-Architecture Interpretation of LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.28006v1 Announce Type: cross Abstract: Understanding how LLMs reason is hindered by a practical asymmetry: while their generated outputs are observable, the underlying reasoning patterns re

Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Systems

SafetyDGX agent

arXiv:2605.27628v1 Announce Type: new Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remai

IPO-Mine: A Toolkit and Dataset for Section-Structured Analysis of Long, Multimodal IPO Documents

Model ReleasesDGX agent

arXiv:2605.28714v1 Announce Type: cross Abstract: An Initial Public Offering (IPO) filing is a document released when a private firm goes public, allowing individual (retail) investors to purchase its

IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage

Model ReleasesDGX agent

arXiv:2605.28247v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for en- hancing LLM reasoning, yet its data ineffi- ciency remains a

Isometry pursuit

ResearchDGX agent

arXiv:2411.18502v2 Announce Type: replace-cross Abstract: Isometry pursuit is a convex algorithm for identifying orthonormal column-submatrices of wide matrices. It consists of a novel normalization m

JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Japanese Large Language Models

Model ReleasesDGX agent

arXiv:2601.01627v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare field, it becomes essential to carefully evaluate their medical safety

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

Model ReleasesDGX agent

arXiv:2605.27984v1 Announce Type: cross Abstract: Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, Speec

LACUNA: Safe Agents as Recursive Program Holes

SafetyDGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

Laguna M.1/XS.2 Technical Report

Model ReleasesDGX agent

arXiv:2605.27605v1 Announce Type: new Abstract: We present Laguna M.1 and Laguna XS.2, two Mixture-of-Experts foundation models built for long-horizon, agentic coding: M.1 has 225.8B total parameters

LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation

ResearchDGX agent

arXiv:2605.27570v1 Announce Type: new Abstract: Parallel LLM test-time scaling techniques (e.g., best-of-N) require drawing N>1 sequences conditioned on the same input prompt. These methods boost accu

Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents

AgentsDGX agent

arXiv:2605.28775v1 Announce Type: cross Abstract: Computer-use agents (CUAs) have recently made substantial progress, but deploying a separate large expert for each software domain remains expensive.

Learning after COVID-19 and the ICT career aspirations: Are students entering the AI era with weaker skills?

ApplicationsDGX agent

arXiv:2605.27391v1 Announce Type: cross Abstract: This paper examines whether students are entering the generative AI era with sufficiently strong educational foundations, focusing on the relationship

Learning Compositional Latent Structure with Vector Networks

Model ReleasesDGX agent

arXiv:2605.28007v1 Announce Type: cross Abstract: Deep networks are powerful function approximators, but they typically store many different computations in shared weight matrices, making it difficult

Learning Tangent Bundles and Characteristic Classes with Autoencoder Atlases

ResearchDGX agent

arXiv:2602.22873v2 Announce Type: replace-cross Abstract: We introduce a theoretical framework that connects multi-chart autoencoders in manifold learning with the classical theory of vector bundles a

Learning the Error Patterns of Language Models

TutorialsDGX agent

arXiv:2605.28328v1 Announce Type: cross Abstract: When generating outputs for domains with specific validity constraints (e.g., a program should compile), LLMs often fail in a small number of focused

Learning Theory of the SVRG: Generalization and Convergence Analysis

ResearchDGX agent

arXiv:2605.28513v1 Announce Type: cross Abstract: Variance reduction (VR) methods employ stochastic gradients with decreasing variance, and they have been widely applied to solve large-scale optimizat

Learning to Assign Prediction Tasks to Agents with Capacity Constraints

SafetyDGX agent

arXiv:2605.27999v1 Announce Type: cross Abstract: We address the problem of learning to assign prediction tasks to one agent from a set of available human or AI agents. In particular, we focus on the

Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages

HardwareDGX agent

arXiv:2605.28213v1 Announce Type: new Abstract: LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations a

LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal Reasoning

AgentsDGX agent

arXiv:2605.28120v1 Announce Type: cross Abstract: Graph-based Retrieval-Augmented Generation (GraphRAG) advances flat document retrieval by structuring knowledge as relational graphs, enabling more co

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration

Model ReleasesDGX agent

arXiv:2602.20497v3 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in image and video generation tasks. However, the high computational demands of Diffusion Tr

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

SafetyDGX agent

arXiv:2605.28524v1 Announce Type: new Abstract: In recent years, Large Language Models (LLMs) have shown great capability in processing graph tasks such as fraud detection. However, most existing meth

Let the Results Speak: A Replication-First Paradigm for LLM Behavioral Benchmarking

Model ReleasesDGX agent

arXiv:2605.27914v1 Announce Type: cross Abstract: Subjective evaluation of LLM behavior -- empathy, restraint, calibrated emotional tone -- is hard. Human inter-rater agreement on such qualities satur

LiDDA: Data Driven Attribution at LinkedIn

ResearchDGX agent

arXiv:2505.09861v3 Announce Type: replace-cross Abstract: Data Driven Attribution, which assigns conversion credits to marketing interactions based on causal patterns learned from data, is the foundat

Ligand-Conditioned Discrete Diffusion for Protein Sequence-Structure Co-Design

ResearchDGX agent

arXiv:2605.27413v1 Announce Type: cross Abstract: Proteins perform their biological functions through three-dimensional structures encoded by amino acid sequences, and ligand-binding protein co-design

LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?

Model ReleasesDGX agent

arXiv:2605.28721v1 Announce Type: new Abstract: Are LLM-based search agents genuinely searching, or using the web to verify what they already know? We study this question on BrowseComp with three diag

LLM-assisted sentiment analysis for integrated computational and qualitative mixed methods education research: A case study of students' written reflection assignments

ApplicationsDGX agent

arXiv:2605.27403v1 Announce Type: cross Abstract: Written reflection assignments give students valuable opportunities for critical self-assessment, meaning making, and learning processing. Additionall

LLM Watermark Evasion via Bias Inversion

SafetyDGX agent

arXiv:2509.23019v5 Announce Type: replace-cross Abstract: Watermarking offers a promising solution for detecting LLM-generated content, yet its robustness under realistic query-free (black-box) evasio

Locality-Aware Redundancy Pruning for LLM Depth Compression

Local AiDGX agent

arXiv:2605.27786v1 Announce Type: cross Abstract: Large language models are known to contain representational redundancy across network depth, making depth pruning an effective approach for improving

Localizing Input Uncertainty Quantification for Large Language Models via Shapley Values

Local AiDGX agent

arXiv:2605.28170v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly integrated into high-stakes decision-making, the ability to reliably quantify uncertainty has become a

Look on Demand: A Cognitive Scheduling Framework for Visual Evidence Acquisition in Multimodal Reasoning

ResearchDGX agent

arXiv:2605.28160v1 Announce Type: new Abstract: Existing multimodal reasoning approaches predominantly follow two paradigms: converting visual inputs into text prior to reasoning, or performing end-to

LoSATok: Low-dimensional Semantic-Acoustic Tokenizer for Cross-Domain Audio Understanding and Generation

ResearchDGX agent

arXiv:2605.27840v1 Announce Type: cross Abstract: Audio tokenizers are fundamental to unifying audio understanding and generation. Understanding requires high-level semantics, while generation demands

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

Model ReleasesDGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.27748v1 Announce Type: cross Abstract: Industrial visual anomaly detection is usually one-class: normal images are abundant, while defects are rare, heterogeneous, and often unavailable dur

Manboformer: Learning Gaussian Representations via Spatial-temporal Attention Mechanism

AgentsDGX agent

arXiv:2503.04863v2 Announce Type: replace-cross Abstract: Compared with voxel-based grid prediction, in the field of 3D semantic occupation prediction for autonomous driving, GaussianFormer proposed u

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

SafetyDGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation

ResearchDGX agent

arXiv:2502.12468v2 Announce Type: replace-cross Abstract: The LLM-as-a-Judge paradigm shows promise for evaluating generative content but lacks reliability in reasoning-intensive scenarios, such as pr

Measuring Form and Function in Language Models

ResearchDGX agent

arXiv:2605.28616v1 Announce Type: cross Abstract: We introduce quantitative metrics for child language acquisition to evaluate language models. Our focus is on the formal syntactic and functional disc

Measuring Massive Multitask Chinese Understanding

ApplicationsDGX agent

arXiv:2304.12986v3 Announce Type: replace-cross Abstract: The development of large-scale Chinese language models is flourishing, yet there is a lack of corresponding capability assessments. Therefore,

Measuring Progress Toward AGI: A Cognitive Framework

ResearchDGX agent

arXiv:2605.28405v1 Announce Type: new Abstract: Despite widespread discussion of AGI, there is no clear framework for measuring progress toward it. This ambiguity fuels subjective claims, makes it dif

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs

ResearchDGX agent

arXiv:2605.28388v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) is empirically shown to notably enhance the reasoning performance of large language models (LLMs),

MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents

Model ReleasesDGX agent

arXiv:2605.28046v1 Announce Type: new Abstract: Existing agent memory systems universally follow what we term a Memory-as-Tool paradigm where a single query triggers one-shot retrieval of flat passage

MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models

Model ReleasesDGX agent

arXiv:2605.28009v1 Announce Type: cross Abstract: Memory-augmented large language models extend reasoning beyond a fixed context window by maintaining long-term memory across interactions. However, ex

Memory-Based vs. Context-Only Conditioning Produces Distinct Behavioral Patterns in Stateful Personalization

ResearchDGX agent

arXiv:2605.27389v1 Announce Type: cross Abstract: We study how conditioning context shapes personalization behavior in a teacher-facing educational recommender system. We compare contextual conditioni

MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems

Model ReleasesDGX agent

arXiv:2605.28732v1 Announce Type: cross Abstract: Memory is essential for enabling large language models to support long-horizon reasoning, yet existing memory systems remain unreliable and difficult

MetaboT: An LLM-based Multi-Agent Frameworkfor Interactive Analysis of Mass SpectrometryMetabolomics Knowledge Graphs

Model ReleasesDGX agent

arXiv:2510.01724v2 Announce Type: replace Abstract: Mass spectrometry-based metabolomics generates complex, high-dimensional data that holds vast potential for biological discovery but remains difficu

MGRetrieval: Memory-Guided Reflective Retrieval for Long-Term Dialogue Agents

ResearchDGX agent

arXiv:2605.27437v1 Announce Type: cross Abstract: Large Language Models (LLMs) have made significant progress in dialogue, yet redundant memory contexts severely limit their effectiveness in long-term

Mind the Gap: Mixtures of Gaussians in Approximate Differential Privacy

ResearchDGX agent

arXiv:2605.28078v1 Announce Type: cross Abstract: We design a class of additive noise mechanisms that satisfy ((arepsilon, elta))-differential privacy (DP) for scalar, real-valued query functions with

Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification

SafetyDGX agent

arXiv:2605.28604v1 Announce Type: cross Abstract: Identifying key individuals in video scenes is essential for applications such as automated video editing and intelligent surveillance. Current method

MIRA: A Bilingual Benchmark for Medical Information Response Audit

Model ReleasesDGX agent

arXiv:2605.28025v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to provide public-facing health information, yet existing safety evaluations overlook whether respons

MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content

Model ReleasesDGX agent

arXiv:2605.28116v1 Announce Type: cross Abstract: Mobile graphical user interface (GUI) agents driven by vision-language models (VLMs) perceive the screen as rendered pixels and choose actions from wh

Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images

TutorialsDGX agent

arXiv:2605.28693v1 Announce Type: cross Abstract: Backpropagation is the core learning mechanism underlying deep learning. However, whether and how this algorithm is implemented in the brain remains h

Mitigating Staleness in Asynchronous Pipeline Parallelism via Basis Rotation

Model ReleasesDGX agent

arXiv:2602.03515v2 Announce Type: replace-cross Abstract: Asynchronous pipeline parallelism maximizes hardware utilization by eliminating the pipeline bubbles inherent in synchronous execution, offeri

MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks

Local AiDGX agent

arXiv:2502.17832v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has become a common practice in multimodal large language models (MLLM) to enhance factual grounding and

← Previous
1…199200201202203…358
Next →