AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
6 Jun 2026

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

SafetyDGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

Soft Sequence Policy Optimization

SafetyDGX agent

arXiv:2602.19327v3 Announce Type: replace-cross Abstract: A significant portion of recent research on Large Language Model (LLM) alignment focuses on developing new policy optimization methods based o

Stable Deep Reinforcement Learning via Isotropic Gaussian Representations

ResearchDGX agent

arXiv:2602.19373v3 Announce Type: replace-cross Abstract: Deep reinforcement learning systems often suffer from unstable training dynamics due to non-stationarity, where learning objectives and data d


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Step-adaptive multimodal fusion network with multi-scale cloud feature learning for ultra-short-term solar irradiance forecasting

ResearchDGX agent

arXiv:2606.06102v1 Announce Type: new Abstract: Ultra-short-term solar irradiance prediction is critical for photovoltaic system dispatch and power grid stability. Existing approaches suffer from thre

Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces

Model ReleasesDGX agent

arXiv:2606.05464v1 Announce Type: new Abstract: Verifiable reward training has improved mathematical and coding reasoning, but these domains capture only part of step-by-step decision making. Many rea

Subspace-Aware Sparse Autoencoders for Effective Mechanistic Interpretability

Model ReleasesDGX agent

arXiv:2606.06333v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) are widely used for mechanistic interpretability in large language models, yet their formulation assigns each latent featur

Synapse: Federated Tool Routing via Typed Compendium Artifacts

Model ReleasesDGX agent

arXiv:2602.00911v2 Announce Type: replace Abstract: The unit of collaboration in federated learning determines what guarantees are even expressible. Flat units like weights, prompts, raw examples, car

Synthetic Contrastive Reasoning for Multi-Table Q&A

Model ReleasesDGX agent

arXiv:2606.05382v1 Announce Type: new Abstract: Multi-table question answering requires models to retrieve relevant evidence, link schemas, and perform compositional reasoning across relational tables

TAPO: Tool-Aware Policy Optimization via Credit Transfer for Multimodal Search Agents

Model ReleasesDGX agent

arXiv:2606.05784v1 Announce Type: new Abstract: We identify and formally characterize credit misassignment as a systematic failure mode of GRPO in tool-augmented multimodal search agents: its uniform

The End of Software Engineering: How AI Agents Are Fundamentally Restructuring the Software Paradigm

Model ReleasesDGX agent

arXiv:2606.05608v1 Announce Type: cross Abstract: For over half a century, software engineering has operated on a foundational premise: human engineers decompose problems, encode decision logic into s

The Role of Instructional Guidance in Generative AI-Assisted Learning: Empirical Evidence from Construction Engineering Education

TutorialsDGX agent

arXiv:2606.05509v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is increasingly used to support self-directed learning, yet student interaction with such systems often remain

The Score Hamiltonian: Mapping Diffusion Models to Adiabatic Transport

ResearchDGX agent

arXiv:2606.05217v1 Announce Type: cross Abstract: We exhibit an exact correspondence between sampling with score-based diffusion models and adiabatic transport of ground states for a family of Schrodi

The Virtual Roundtable: Multi-Agent Personas Simulating the Dynamics of Human Brainstorming

AgentsDGX agent

arXiv:2606.05178v1 Announce Type: cross Abstract: As AI-driven product development accelerates, the bottleneck is shifting from how we build to what we build. Traditional human brainstorming faces cha

TinyML-Driven Cybersecurity for Autonomous Spacecraft: Latency-Accuracy Analysis for SPARTA RF and Cyber Threat Detection

AgentsDGX agent

arXiv:2606.05779v1 Announce Type: cross Abstract: Autonomous spacecraft require rapid, lightweight, and reliable onboard detection of cyber-RF threats. Using the SPARTA attack model, we analyze the la

TLA-Prover: Verifiable TLA+ Specification Synthesis via Preference-Optimized Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2606.06133v1 Announce Type: cross Abstract: TLA+ is a formal specification language for verifying distributed systems and safety-critical protocols. Large language models (LLMs) frequently produ

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management

Model ReleasesDGX agent

arXiv:2606.06337v1 Announce Type: new Abstract: Large language model (LLM) deployments for long-horizon tasks face a fundamental constraint: context windows are finite while productive work sessions a

TOKI: A Bitemporal Operator Algebra for Contradiction Resolution in LLM-Agent Persistent Memory

Local AiDGX agent

arXiv:2606.06240v1 Announce Type: cross Abstract: Persistent memory for an LLM agent is a write-heavy substrate: every belief update is a versioned write, and a new claim may contradict a stored one.

ToolChoiceConfusion: Causal Minimal Tool Filtering for Reliable LLM Agents

Model ReleasesDGX agent

arXiv:2606.06284v1 Announce Type: new Abstract: Large language model agents increasingly rely on external tools, but larger tool menus can reduce reliability and efficiency by increasing wrong-tool ca

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

SafetyDGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

Towards Healthy Evolution: Exploring the Role and Mechanisms of Human-Agent Interaction in Self-Evolving Systems

SafetyDGX agent

arXiv:2606.06114v1 Announce Type: new Abstract: Self-evolving agents improve through continual self-play and self-generated learning signals, but autonomous evolution can also cause capability degrada

Towards the Readability of LLM-Generated Codes through Multitask Representation Engineering

ResearchDGX agent

arXiv:2606.06214v1 Announce Type: cross Abstract: Correctness and readability are key measures of code quality, respectively ensuring functional fidelity and ease of comprehension. While most existing

Towards Unified and Data-Efficient Prognostics and Health Management with Tabular Foundation Models

ResearchDGX agent

arXiv:2606.05481v1 Announce Type: cross Abstract: Data-driven Prognostics and Health Management (PHM) uses time-varying condition-monitoring data to diagnose system states and estimate remaining usefu

Towards World Models in Biomedical Research

SafetyDGX agent

arXiv:2606.05925v1 Announce Type: new Abstract: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbatio

TRACE: A Temporal Conditional Estimation for Multimodal Time Series Foundation Models

TutorialsDGX agent

arXiv:2606.06285v1 Announce Type: new Abstract: Time series foundation models (TS-FMs) aim to learn generalizable temporal representations that can be adapted to a wide range of downstream tasks. In r

Trust, but Don't Verify: Epistemic Blind Spots in LLM Source Evaluation

Model ReleasesDGX agent

arXiv:2606.05403v1 Announce Type: cross Abstract: Language models increasingly act as epistemic proxies, synthesizing evidence from multiple sources to inform decisions. Whether they evaluate the qual

Uncertainty Aware Functional Behavior Prediction and Material Fatigue Assessment for Circular Factory

ApplicationsDGX agent

arXiv:2606.05334v1 Announce Type: new Abstract: Returned products in circular factories re-enter production with heterogeneous degradation states, usage histories, and remaining capability. Reuse cann

UniVoice: A Unified Model for Speech and Singing Voice Generation

SafetyDGX agent

arXiv:2606.05852v1 Announce Type: cross Abstract: Text-to-speech (TTS) and singing voice synthesis (SVS) both aim to generate human vocal audio from symbolic inputs, but they impose different requirem

Unsupervised Pattern Analysis in Japanese Veterinary Toxicology: A Regulatory-Compliant Framework for Cross-Species Risk Assessment

SafetyDGX agent

arXiv:2606.06207v1 Announce Type: new Abstract: Veterinary pharmacovigilance systems are essential for monitoring adverse drug events (ADEs), yet existing approaches often fail to capture region-speci

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents

Model ReleasesDGX agent

arXiv:2606.06453v1 Announce Type: new Abstract: Sparse attention is becoming increasingly important for serving large language models (LLMs) as generation lengths continue to grow. However, deploying

What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.05304v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models are typically organized around roles, pipelines, and turn schedules, while the content that age

When Good Enough Is Optimal: Multiplication-Only Matrix Inversion Approximation for Quantized Gated DeltaNet

ResearchDGX agent

arXiv:2606.06034v1 Announce Type: cross Abstract: Matrix inversion in chunk-wise parallel linear attention is a major bottleneck for long-context modeling, particularly on NPUs, where forward-substitu

When Should Memory Stay Silent: Measuring Memory-Use Boundaries in Memory-Augmented Conversational Agents

Model ReleasesDGX agent

arXiv:2606.06055v1 Announce Type: new Abstract: Long-term memory enables language model agents to support personalized interactions, but it remains unclear when available memories warrant integration

When Should We Protect AI? A Precautionary Framework for Consciousness Uncertainty

ResearchDGX agent

arXiv:2606.05528v1 Announce Type: new Abstract: Existing frameworks assess whether AI systems might be conscious but provide no guidance on what to do with that assessment. We address this gap with a

When Surface Form Changes Moderation Decisions: A Paired Study of Code-Mixed Workflow Instability

ResearchDGX agent

arXiv:2606.05654v1 Announce Type: cross Abstract: Hate moderation is often evaluated as classification on clean English inputs, but deployed systems must route content to actions such as ALLOW, FLAG,

When Tools Fail: Benchmarking Dynamic Replanning and Anomaly Recovery in LLM Agents

Model ReleasesDGX agent

arXiv:2606.05806v1 Announce Type: new Abstract: Existing benchmarks evaluate Tool-Integrated Reasoning (TIR) in LLMs on idealized ''happy paths'', largely overlooking real-world tool failures. We intr

Where Should Knowledge Enter? A Layered Framework for Knowledge Infusion in Multimodal Iterative Generative Mo

SafetyDGX agent

arXiv:2606.06356v1 Announce Type: new Abstract: Multimodal generative models produce fluent outputs but remain unreliable when generation must respect structured, domain-specific, or safety-critical k

Where's the Structure? A Systematic Literature Review of Empirical Research on Human-AI Collaboration and Hybrid Intelligence for Learning

ResearchDGX agent

arXiv:2606.05222v1 Announce Type: cross Abstract: Artificial intelligence (AI) has been applied across educational contexts to support learning. One approach to such support is 'human-AI collaboration

Will the Agent Recuse Itself? Measuring LLM-Agent Compliance with In-Band Access-Deny Signals

Model ReleasesDGX agent

arXiv:2606.06460v1 Announce Type: cross Abstract: As autonomous LLM agents increasingly hold real credentials and operate infrastructure without a human in the loop, operators have no standard way to

Willing but Unable: Separating Refusal from Capability in Code LLMs via Abliteration

SafetyDGX agent

arXiv:2606.05396v1 Announce Type: cross Abstract: Producing a labeled vulnerable code at scale is a recurring obstacle for learning-based vulnerability detection: mined corpora carry substantial label

WorldFly: A World-Model-Based Vision-Language-Action Model for UAV Navigation

Model ReleasesDGX agent

arXiv:2606.06147v1 Announce Type: new Abstract: End-to-end Vision-Language-Action (VLA) models have shown promise in UAV navigation. However, existing approaches typically rely on historical observati

X-Band UAV-enabled Integrated Sensing and Communications for Vehicular Networks

ResearchDGX agent

arXiv:2606.05262v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly considered as aerial platforms capable of providing both sensing and communication services, represen

Your GFlowNet Secretly Learns an Optimal Transport Plan

SafetyDGX agent

arXiv:2606.06272v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are a framework for sampling structured objects via stochastic trajectories in a directed graph. In this work, we

Zero knowledge verification for frontier AI training is possible

SafetyDGX agent

arXiv:2606.05433v1 Announce Type: new Abstract: Frontier AI governance frameworks increasingly use cumulative training compute as the primary criterion for designating high-impact models, but enforcem

4 Jun 2026

100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

Model ReleasesDGX agent

arXiv:2505.19293v2 Announce Type: replace-cross Abstract: Long-context capability is considered one of the most important abilities of LLMs, as a truly long context-capable LLM enables users to effort

A Geometric Characterization of the Stationary Plateau for Two-Layer Neural Networks

Local AiDGX agent

arXiv:2606.04327v1 Announce Type: cross Abstract: We investigate the geometric structure of stationary plateaus that arise in the loss landscape of two-layer neural networks with smooth activation fun

A Goal-Set Characterization of Task Composition in the Boolean Task Algebra

SafetyDGX agent

arXiv:2606.04053v1 Announce Type: cross Abstract: The Boolean Task Algebra (BTA) provides a principled framework for zero-shot task composition in reinforcement learning by equipping goal-reaching tas

A Normative Intermediate Representation for ASP-Based Compliance Reasoning

ResearchDGX agent

arXiv:2606.04619v1 Announce Type: new Abstract: We propose MONIR, a Modalized-Output Normative Intermediate Representation for ASP-based compliance reasoning. Its core fragment has a staged operationa

A Study of the Scale Invariant Signal to Distortion Ratio in Speech Separation with Noisy References

Model ReleasesDGX agent

arXiv:2508.14623v2 Announce Type: replace-cross Abstract: This paper examines the implications of using the Scale-Invariant Signal-to-Distortion Ratio (SI-SDR) as both evaluation and training objectiv

A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and Models

ResearchDGX agent

arXiv:2606.04177v1 Announce Type: cross Abstract: Interpretable linguistic features offer a promising approach for explaining why a given text appears machine-generated, particularly for non-expert us

A Unified Framework for Locality in Scalable MARL

SafetyDGX agent

arXiv:2602.16966v2 Announce Type: replace-cross Abstract: Scalable methods for networked multi-agent reinforcement learning let each agent plan using only a small neighborhood of the agent graph. This

A Unified Geometric Space for Topological Alignment Between Transformer-Based Models and Human Brain Networks

SafetyDGX agent

arXiv:2510.24342v2 Announce Type: replace Abstract: Prior brain-AI alignment studies are typically constrained by specific inputs and tasks, limiting their ability to capture organizational properties

Abduction Prover in Isabelle/HOL

ResearchDGX agent

arXiv:2606.04877v1 Announce Type: cross Abstract: Proof assistants based on expressive logics suffer limited automation for proof search, raising the cost of formal verification based on proof assista

Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control

SafetyDGX agent

arXiv:2606.04775v1 Announce Type: cross Abstract: Text-to-video (T2V) models trained on large-scale web data can generate undesired content, motivating interventions that reduce harmful outputs withou

AdaKoop: Efficient Modeling of Nonlinear Dynamics from Nonstationary Data Streams with Koopman Operator Regression

Model ReleasesDGX agent

arXiv:2606.04930v1 Announce Type: cross Abstract: Real-time data analysis requires the ability to accurately and adaptively address nonlinear dynamics in a nonstationary data stream while preserving c

Adaptive Calibration for Fair and Performant Facial Recognition

SafetyDGX agent

arXiv:2606.04469v1 Announce Type: cross Abstract: We introduce Adaptive Calibration (AC), a novel calibration strategy for facial recognition that maps cosine similarity between normalized embeddings

Adaptive Minds: Empowering Agents with LoRA-as-Tools

Model ReleasesDGX agent

arXiv:2510.15416v2 Announce Type: replace Abstract: We investigate a framework in which LoRA adapters are treated as callable tools that a base language model can dynamically select and invoke. We hyp

Adaptive Patching Is Harder Than It Looks For Time-Series Forecasting

Local AiDGX agent

arXiv:2606.04074v1 Announce Type: cross Abstract: Adaptive patching is a recent and compelling proposal for time-series Transformers: allocate finer patches where the sequence looks locally informativ

ADAPTOOD: Uncertainty-Aware Fine-Tuning for Out-of-Distribution ECG Time Series Models

TutorialsDGX agent

arXiv:2606.04164v1 Announce Type: cross Abstract: Data samples used for training often differ from those encountered during fine-tuning and deployment, and while ML models show promise, their performa

AgenticDiffusion: Agentic Diffusion-based Path Planning for Vision-Based UAV Navigation

Local AiDGX agent

arXiv:2606.04111v1 Announce Type: cross Abstract: Indoor UAV navigation requires efficient exploration, scene understanding, and reliable trajectory execution under limited field-of-view observations.

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

HardwareDGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

← Previous
1…157158159160161…358
Next →