AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Safety

Think Fast, Talk Smart: Partitioning Deterministic and Neural Computation for Structured Health Text Generation

DGX agent

arXiv:2605.29652v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being used to generate health text from structured records such as wearable time series, biomarkers, vital

safetyarxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Thinking Before Constraining: A Unified Decoding Framework for Large Language Models

DGX agent

arXiv:2601.07525v2 Announce Type: replace-cross Abstract: Natural generation allows Large Language Models (LLMs) to produce free-form responses with rich reasoning, yet the lack of structure makes out

researcharxiv-cs-ai
29 May 2026
Tutorials

Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning

DGX agent

arXiv:2605.28842v1 Announce Type: cross Abstract: The success of large language models (LLMs) across diverse NLP tasks has elevated the importance of reasoning chain optimization as a critical step in

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

TIMEGATE: Sustainable Time-Boxed Promotion Gates for Continual ML Adaptation Under Resource Constraints

DGX agent

arXiv:2605.29183v1 Announce Type: cross Abstract: As machine learning(ML) systems evolve to continual adaptation, each re-training cycle uses compute, annotation, and energy. We introduce TIMEGATE, a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection

DGX agent

arXiv:2605.30344v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have achieved impressive performance across many tasks, yet prior studies report unsatisfactory perform

model-releasesarxiv-cs-ai
29 May 2026
Research

Token Inflation: How Dishonest Providers Can Overcharge for Large Language Model Usage

DGX agent

arXiv:2605.30040v1 Announce Type: cross Abstract: Per-token billing is now the standard pricing model for commercial large language models (LLMs), so the honesty of reported token counts directly affe

researcharxiv-cs-ai
29 May 2026
Model Releases

Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection

DGX agent

arXiv:2605.30189v1 Announce Type: cross Abstract: We show that LoRA adapters, the dominant distribution format for fine-tuned LLMs, can be reliably backdoored through training data poisoning while pre

model-releasesarxiv-cs-ai
29 May 2026
Research

Topological Order in Neural Wavefunctions

DGX agent

arXiv:2512.01863v2 Announce Type: replace-cross Abstract: Topologically ordered states are among the most interesting quantum phases of matter that host emergent quasi-particles having fractional char

researcharxiv-cs-ai
29 May 2026
Safety

Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment

DGX agent

arXiv:2605.29930v1 Announce Type: new Abstract: Mutual misunderstanding in contemporary society does not arise merely because people hold different opinions or values. Even under the same observations

safetyarxiv-cs-ai
29 May 2026
Model Releases

Toward Ethical Facial Age Estimation: A Generalized Zero-Shot Benchmark Without Training on Children's Data

DGX agent

arXiv:2605.29230v1 Announce Type: cross Abstract: Age estimation from facial images typically relies on training data that includes images of minors, a practice that raises serious ethical, legal, and

model-releasesarxiv-cs-ai
29 May 2026
Safety

Toward User Preference Alignment in LLM Recommendation via Explicit Context Feedback

DGX agent

arXiv:2605.29141v1 Announce Type: cross Abstract: Traditional recommender systems (RecSys) primarily infer user preferences from implicit signals (such as clicks, watches, and purchases), often neglec

safetyarxiv-cs-ai
29 May 2026
Research

Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthetic Data and Relative Context Discrepancy

DGX agent

arXiv:2509.21190v4 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is a critical task, but developing models that generalize to unseen data in a zero-shot manner remains a

researcharxiv-cs-ai
29 May 2026
Safety

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

DGX agent

arXiv:2605.29430v1 Announce Type: new Abstract: Automatic speech recognition (ASR) is a core component of human--computer interaction and an increasingly important front-end for LLM-based assistants a

safetyarxiv-cs-ai
29 May 2026
Local Ai

Towards Localized and Disentangled Knowledge Editing for Multimodal Large Language Models

DGX agent

arXiv:2605.29826v1 Announce Type: cross Abstract: Existing methods in Multimodal Knowledge Editing (MKE) have advanced the ability to correct outdated or inaccurate knowledge in Multimodal Large Langu

local-aiarxiv-cs-ai
29 May 2026
Agents

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

DGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

agentsarxiv-cs-ai
29 May 2026
Model Releases

TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evaluation

DGX agent

arXiv:2605.29656v1 Announce Type: new Abstract: Evaluating open-ended outputs from large language models (LLMs) remains challenging due to the absence of ground truth. Existing metrics rely on final-a

model-releasesarxiv-cs-ai
29 May 2026
Safety

TRACER: Persistent Regularization for Robust Multimodal Finetuning

DGX agent

arXiv:2605.29380v1 Announce Type: cross Abstract: Mainstream strategies for finetuning pretrained multimodal models often degrade out-of-distribution (OOD) robustness, a phenomenon known as catastroph

safetyarxiv-cs-ai
29 May 2026
Model Releases

Training Deliberative Monitors for Black-Box Scheming Detection

DGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

model-releasesarxiv-cs-ai
29 May 2026
Research

Transcribing Children's Speech: ASR Performance and Obtaining Reliable Orthographic Transcriptions

DGX agent

arXiv:2605.28833v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) has the potential to substantially reduce manual annotation effort in child speech research by generating automatic

researcharxiv-cs-ai
29 May 2026
Model Releases

Trends in AI and Human-AI Interaction in Clinical Trials -- A Hybrid Human-AI Exploration

DGX agent

arXiv:2605.29096v1 Announce Type: new Abstract: This paper examines records retrieved from the ClinicalTrials.gov registry to characterize temporal trends in AI terminology and the geographical distri

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning

DGX agent

arXiv:2605.29170v1 Announce Type: cross Abstract: Legal NLP benchmarks are overwhelmingly English-centric, leaving failure modes in morphologically rich, non-Latin-script languages undetected. We intr

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

DGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

local-aiarxiv-cs-ai
29 May 2026
Research

Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems

DGX agent

arXiv:2605.28883v1 Announce Type: new Abstract: Tropical forests worldwide are under intense deforestation pressure driven by economic and political interests, and scientific evidence suggests this de

researcharxiv-cs-ai
29 May 2026
Model Releases

Uncertainty-Aware Transfer Learning for Cross-Building Energy Forecasting: Toward Robust and Scalable District-Level Energy Management

DGX agent

arXiv:2605.29733v1 Announce Type: new Abstract: Scaling data-driven energy forecasting to district level requires models that can be re-used across buildings with minimal target-domain data and honest

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization

DGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

local-aiarxiv-cs-ai
29 May 2026
Agents

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

DGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

agentsarxiv-cs-ai
29 May 2026
Research

Unlocking the Working Memory of Large Language Models for Latent Reasoning

DGX agent

arXiv:2605.30343v1 Announce Type: cross Abstract: To improve the reasoning capabilities of large language models, test-time compute is typically scaled by generating intermediate tokens before the fin

researcharxiv-cs-ai
29 May 2026
Research

Unveiling Multi-regime Patterns in SciML: Distinct Failure Modes and Regime-specific Optimization

DGX agent

arXiv:2605.29153v1 Announce Type: cross Abstract: Neural networks trained under different hyperparameter settings can fall into distinct training 'regimes,' with consistent behavior within regimes and

researcharxiv-cs-ai
29 May 2026
Agents

VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis

DGX agent

arXiv:2605.28978v1 Announce Type: new Abstract: Finite Element Analysis (FEA) serves as the cornerstone of modern engineering design. However, its workflow is inherently complex and relies heavily on

agentsarxiv-cs-ai
29 May 2026
Hardware

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

DGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

hardwarearxiv-cs-ai
29 May 2026
Agents

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

DGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

agentsarxiv-cs-ai
29 May 2026
Agents

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies

DGX agent

arXiv:2605.30011v1 Announce Type: cross Abstract: Recent work has begun to equip vision-language-action (VLA) policies with explicit intermediate reasoning. In embodied control, however, textual chain

agentsarxiv-cs-ai
29 May 2026
Model Releases

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

DGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

model-releasesarxiv-cs-ai
29 May 2026
Applications

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

DGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

applicationsarxiv-cs-ai
29 May 2026
Safety

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

DGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

safetyarxiv-cs-ai
29 May 2026
Research

Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment

DGX agent

arXiv:2605.29243v1 Announce Type: cross Abstract: Forecasting conversational derailment is the task of predicting, as the conversation unfolds, whether it will eventually derail into personal attacks.

researcharxiv-cs-ai
29 May 2026
Research

Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data

DGX agent

arXiv:2502.20838v3 Announce Type: replace-cross Abstract: Passive acoustic monitoring (PAM) systems generate continuous recordings spanning months, yet automated bioacoustic analysis of whale calls re

researcharxiv-cs-ai
29 May 2026
Model Releases

What drives performance in molecular MPNNs? An operator-level factorial benchmark

DGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

model-releasesarxiv-cs-ai
29 May 2026
Safety

When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

DGX agent

arXiv:2605.29267v1 Announce Type: new Abstract: Foundation models are increasingly trained on synthetic data generated by prior model iterations rather than exclusively on real data. This self-consumi

safetyarxiv-cs-ai
29 May 2026
Safety

When and How Long? The Readout-Mediator Angle in Temporal Reasoning

DGX agent

arXiv:2605.29126v1 Announce Type: cross Abstract: A linear probe can decode a representation almost perfectly and yet be completely irrelevant to how the model uses it. On calendar-date duration reaso

safetyarxiv-cs-ai
29 May 2026
Local Ai

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

DGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

local-aiarxiv-cs-ai
29 May 2026
Applications

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

DGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

applicationsarxiv-cs-ai
29 May 2026
Research

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

DGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

researcharxiv-cs-ai
29 May 2026
Tutorials

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

DGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

DGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Who can we trust? LLM-as-a-jury for Comparative Assessment

DGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…248249250251252…452
Next →