AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
29 May 2026

SERC: LDPC-Inspired Semantic Error Correction for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.28837v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have demonstrated remarkable capabilities, their reliability is significantly compromised by hallucinations. Existi

Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization

ResearchDGX agent

arXiv:2605.29547v1 Announce Type: cross Abstract: Deep learning optimization relies heavily on the assumption of smooth loss landscapes, a condition systematically violated by modern architectures due

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.01869v3 Announce Type: replace Abstract: LLM-driven agents excel at sequential decision-making but often rely on on-the-fly reasoning, re-deriving solutions even in recurring scenarios. Thi

SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents

AgentsDGX agent

arXiv:2605.29440v1 Announce Type: cross Abstract: Retrieval-augmented LLM agents increasingly rely on curated skill banks: collections of reusable textual principles that guide decision making on comp

SkillsInjector: Dynamic Skill Context Construction for LLM Agents

Model ReleasesDGX agent

arXiv:2605.29794v1 Announce Type: new Abstract: LLM agents now draw on growing skill libraries to handle complex tasks. However, injecting more skills does not always improve task completion and can e

Small Agent Group is the Future of Digital Health

Model ReleasesDGX agent

arXiv:2602.08013v2 Announce Type: replace Abstract: The rapid adoption of large language models (LLMs) in digital health has been driven by a 'scaling-first' philosophy, i.e., the assumption that clin

Source-Grounded Semantic Reinforcement Learning for Low-Resource Target-Language Generation

ResearchDGX agent

arXiv:2605.29502v1 Announce Type: cross Abstract: Low-resource target-language generation is often limited by scarce parallel data, while high-resource source-language monolingual data is abundant but

Specialty-Specific Medical Language Model for Immune-Mediated Diseases

ApplicationsDGX agent

arXiv:2605.28838v1 Announce Type: cross Abstract: Extracting detailed clinical information from free-text medical narratives remains a practical challenge for researchers and healthcare systems. Termi

Steering at the Source: Style Modulation Heads for Robust Persona Control

Local AiDGX agent

arXiv:2603.13249v2 Announce Type: replace-cross Abstract: Activation steering offers a computationally efficient mechanism for controlling Large Language Models (LLMs) without fine-tuning. While effec

Steering Language Models Before They Speak: Logit-Level Interventions

ResearchDGX agent

arXiv:2601.10960v2 Announce Type: replace-cross Abstract: Controllable generation requires language models to realize output characteristics such as reading level, politeness, and toxicity. Existing s

Stochastic Lifting for Generating Trajectories of Stochastic Physical Systems

ResearchDGX agent

arXiv:2605.29194v1 Announce Type: cross Abstract: Many stochastic physical systems evolve smoothly over time in the sense that the distribution of states changes regularly across time steps. The trans

Structured Prompt Optimization Meets Reinforcement Learning for Global and Local Interpretability over Complex Text

ResearchDGX agent

arXiv:2605.29076v1 Announce Type: cross Abstract: LLMs have advanced text classification, yet existing paradigms face a trade-off: supervised (label only) fine-tuning is scalable but offers limited re

Surfacing Isolated Learners with Outcome-Independent Mediation of Feedback between Teachers and Students Using AI

TutorialsDGX agent

arXiv:2605.29240v1 Announce Type: new Abstract: AI-augmented classrooms generate rich teacher and student feedback before graded outcomes become available, yet these signals can be difficult to transl

SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow

Model ReleasesDGX agent

arXiv:2605.29368v1 Announce Type: cross Abstract: The intricate nature of modern surgical care necessitates intelligent systems that can synthesize extensive patient records, support collaborative dec

Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio

ResearchDGX agent

arXiv:2505.10975v3 Announce Type: replace-cross Abstract: Monaural multi-speaker automatic speech recognition (ASR) remains challenging due to data scarcity and the intrinsic difficulty of recognizing

Sustainable Metal-Organic Framework Water Harvesters in the Artificial Intelligence Era

ResearchDGX agent

arXiv:2605.29179v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are excellent candidates for water harvesting due to their tunable pore environments, which can be precisely engineere

Tailoring the Curriculum: Student-Centered Reasoning Distillation via Dynamic Data-Model Compatibility

ResearchDGX agent

arXiv:2605.29229v1 Announce Type: new Abstract: Reasoning distillation transfers complex reasoning abilities from large language models (LLMs) to smaller ones, yet its success depends on how well the

Taming Data Challenges in ML-based Security Tasks Using Generative AI

ResearchDGX agent

arXiv:2507.06092v4 Announce Type: replace-cross Abstract: Machine learning-based supervised classifiers are widely used for security tasks, and their improvement has been largely focused on algorithmi

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

Model ReleasesDGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

Model ReleasesDGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies

Model ReleasesDGX agent

arXiv:2605.29712v1 Announce Type: cross Abstract: Grounded claim factuality checking is important for large language model (LLM) applications such as retrieval-augmented generation, as it helps users

Teaching Values to Machines: Simulating Human-Like Behavior in LLMs

SafetyDGX agent

arXiv:2605.30036v1 Announce Type: new Abstract: Large Language Models (LLMs) demonstrate a remarkable capacity to adopt different personas and roles; however, it remains unclear whether they can manif

Temporal Motif-aware Graph Test-time Adaptation for OOD Blockchain Anomaly Detection

ApplicationsDGX agent

arXiv:2605.29526v1 Announce Type: cross Abstract: Ever-evolving transaction patterns have significantly hindered anomaly detection on emerging cryptocurrency blockchains due to the vast number of addr

Temporal Stability and Few-Shot Prompting in Math Task Assessment

Model ReleasesDGX agent

arXiv:2605.30151v1 Announce Type: new Abstract: As AI tools become increasingly integrated into educational contexts, questions arise about both their stability over time and their responsiveness to p

Test Time Training for Supervised Causal Learning

ApplicationsDGX agent

arXiv:2605.30015v1 Announce Type: cross Abstract: Supervised Causal Learning (SCL) has shown promise in causal discovery by framing it as a supervised learning problem. However, it suffers from signif

The Best of the Two Worlds: Harmonizing Semantic and Hash IDs for Sequential Recommendation

SafetyDGX agent

arXiv:2512.10388v2 Announce Type: replace-cross Abstract: Conventional Sequential Recommender Systems (SRS) typically assign unique hash IDs (HID) to construct item embeddings, which mainly capture co

The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure

Model ReleasesDGX agent

arXiv:2605.29087v1 Announce Type: new Abstract: Reasoning models are evaluated on single-turn benchmarks but deployed in multi-turn dialogue, where users push back on correct answers. Under sustained

The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language Modeling

Model ReleasesDGX agent

arXiv:2605.28864v1 Announce Type: new Abstract: The Cognitive Categorical Transformer (CCT) is a 306M-parameter architecture that augments a pretrained GPT-2 Small backbone with cognitively grounded c

The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models

SafetyDGX agent

arXiv:2605.29123v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) uniquely support any-order generation, with confidence-based decoding currently serving as the de facto standard

The Curse of Helpfulness: Inverse Scaling Law in Robustness to Distractor Instructions via DistractionIF

Model ReleasesDGX agent

arXiv:2605.29491v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in agentic and retrieval-augmented generation (RAG) systems, where they must execute user-specifi

The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction

Model ReleasesDGX agent

arXiv:2605.29411v1 Announce Type: cross Abstract: Under standard graphical assumptions, the Markov boundary of a target variable is the smallest set of features that renders every other feature redund

The Hamilton-Jacobi Theory of Deep Learning

Model ReleasesDGX agent

arXiv:2605.28983v1 Announce Type: cross Abstract: In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selec

The Impact of Semantic Pairs on Self-Supervised Representation Learning

ResearchDGX agent

arXiv:2510.08722v3 Announce Type: replace-cross Abstract: Instance discrimination learns visual representations by treating different augmented views of the same image as positive pairs. While this en

The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane

SafetyDGX agent

arXiv:2605.29082v1 Announce Type: new Abstract: AI agents are increasingly expected to operate as digital employees: accessing enterprise data, making decisions, and taking actions autonomously. But a

The Little Book of Generative AI Foundations: An Intuitive Mathematical Primer

ResearchDGX agent

arXiv:2605.29713v1 Announce Type: cross Abstract: This book provides a compact, derivation-oriented introduction to the mathematical foundations of modern generative artificial intelligence. Rather th

The New Pro Se: Generative AI and the Surge in Federal Civil Self-Representation

ApplicationsDGX agent

arXiv:2605.29493v1 Announce Type: cross Abstract: Since public access to generative AI tools became widespread, federal civil litigation has seen a marked increase in pro se (self-represented) plainti

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

Model ReleasesDGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

The Sample Complexity of Multiclass and Sparse Contextual Bandits

SafetyDGX agent

arXiv:2605.29645v1 Announce Type: cross Abstract: We study contextual bandits in the stochastic i.i.d. setting, where a learner observes contexts drawn from an unknown distribution, selects actions fr

Think Fast, Talk Smart: Partitioning Deterministic and Neural Computation for Structured Health Text Generation

SafetyDGX agent

arXiv:2605.29652v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being used to generate health text from structured records such as wearable time series, biomarkers, vital

Thinking Before Constraining: A Unified Decoding Framework for Large Language Models

ResearchDGX agent

arXiv:2601.07525v2 Announce Type: replace-cross Abstract: Natural generation allows Large Language Models (LLMs) to produce free-form responses with rich reasoning, yet the lack of structure makes out

Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning

TutorialsDGX agent

arXiv:2605.28842v1 Announce Type: cross Abstract: The success of large language models (LLMs) across diverse NLP tasks has elevated the importance of reasoning chain optimization as a critical step in

TIMEGATE: Sustainable Time-Boxed Promotion Gates for Continual ML Adaptation Under Resource Constraints

Model ReleasesDGX agent

arXiv:2605.29183v1 Announce Type: cross Abstract: As machine learning(ML) systems evolve to continual adaptation, each re-training cycle uses compute, annotation, and energy. We introduce TIMEGATE, a

Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.30344v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have achieved impressive performance across many tasks, yet prior studies report unsatisfactory perform

Token Inflation: How Dishonest Providers Can Overcharge for Large Language Model Usage

ResearchDGX agent

arXiv:2605.30040v1 Announce Type: cross Abstract: Per-token billing is now the standard pricing model for commercial large language models (LLMs), so the honesty of reported token counts directly affe

Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection

Model ReleasesDGX agent

arXiv:2605.30189v1 Announce Type: cross Abstract: We show that LoRA adapters, the dominant distribution format for fine-tuned LLMs, can be reliably backdoored through training data poisoning while pre

Topological Order in Neural Wavefunctions

ResearchDGX agent

arXiv:2512.01863v2 Announce Type: replace-cross Abstract: Topologically ordered states are among the most interesting quantum phases of matter that host emergent quasi-particles having fractional char

Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment

SafetyDGX agent

arXiv:2605.29930v1 Announce Type: new Abstract: Mutual misunderstanding in contemporary society does not arise merely because people hold different opinions or values. Even under the same observations

Toward Ethical Facial Age Estimation: A Generalized Zero-Shot Benchmark Without Training on Children's Data

Model ReleasesDGX agent

arXiv:2605.29230v1 Announce Type: cross Abstract: Age estimation from facial images typically relies on training data that includes images of minors, a practice that raises serious ethical, legal, and

Toward User Preference Alignment in LLM Recommendation via Explicit Context Feedback

SafetyDGX agent

arXiv:2605.29141v1 Announce Type: cross Abstract: Traditional recommender systems (RecSys) primarily infer user preferences from implicit signals (such as clicks, watches, and purchases), often neglec

Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthetic Data and Relative Context Discrepancy

ResearchDGX agent

arXiv:2509.21190v4 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is a critical task, but developing models that generalize to unseen data in a zero-shot manner remains a

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

SafetyDGX agent

arXiv:2605.29430v1 Announce Type: new Abstract: Automatic speech recognition (ASR) is a core component of human--computer interaction and an increasingly important front-end for LLM-based assistants a

Towards Localized and Disentangled Knowledge Editing for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.29826v1 Announce Type: cross Abstract: Existing methods in Multimodal Knowledge Editing (MKE) have advanced the ability to correct outdated or inaccurate knowledge in Multimodal Large Langu

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

AgentsDGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evaluation

Model ReleasesDGX agent

arXiv:2605.29656v1 Announce Type: new Abstract: Evaluating open-ended outputs from large language models (LLMs) remains challenging due to the absence of ground truth. Existing metrics rely on final-a

TRACER: Persistent Regularization for Robust Multimodal Finetuning

SafetyDGX agent

arXiv:2605.29380v1 Announce Type: cross Abstract: Mainstream strategies for finetuning pretrained multimodal models often degrade out-of-distribution (OOD) robustness, a phenomenon known as catastroph

Training Deliberative Monitors for Black-Box Scheming Detection

Model ReleasesDGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

Transcribing Children's Speech: ASR Performance and Obtaining Reliable Orthographic Transcriptions

ResearchDGX agent

arXiv:2605.28833v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) has the potential to substantially reduce manual annotation effort in child speech research by generating automatic

Trends in AI and Human-AI Interaction in Clinical Trials -- A Hybrid Human-AI Exploration

Model ReleasesDGX agent

arXiv:2605.29096v1 Announce Type: new Abstract: This paper examines records retrieved from the ClinicalTrials.gov registry to characterize temporal trends in AI terminology and the geographical distri

UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning

Model ReleasesDGX agent

arXiv:2605.29170v1 Announce Type: cross Abstract: Legal NLP benchmarks are overwhelmingly English-centric, leaving failure modes in morphologically rich, non-Latin-script languages undetected. We intr

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

Local AiDGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

← Previous
1…194195196197198…358
Next →