AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

DGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

model-releasesarxiv-cs-ai
6 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

DGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs

DGX agent

arXiv:2606.05966v1 Announce Type: cross Abstract: Understanding and reasoning about the physical world is the foundation of intelligent behavior, yet state-of-the-art vision-language models (VLMs) sti

model-releasesarxiv-cs-ai
6 Jun 2026
Applications

CausalPOI: Spatio-Temporal Graph-Based Causal Modeling for Cold-Start POI Check-in Forecasting

DGX agent

arXiv:2606.05413v1 Announce Type: cross Abstract: As urban environments continue to evolve rapidly, accurately modeling the dynamic behaviour of Points of Interest is essential for supporting data-dri

applicationsarxiv-cs-ai
6 Jun 2026
Safety

Class-Specific Branch Attention for Mitigating Gradient Interference under Class Imbalance

DGX agent

arXiv:2606.05740v1 Announce Type: new Abstract: Deep neural networks trained under severe class imbalance often exhibit degraded performance, typically attributed to statistical bias. In this work, we

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Closing the Loop on Latent Reasoning via Test-Time Reconstruction

DGX agent

arXiv:2606.06252v1 Announce Type: new Abstract: Recent work moves intermediate reasoning from natural-language traces into latent or cache-level representations to reduce token overhead and avoid a di

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CogManip: Benchmarking Manipulative Behavior in Multi-Turn Interactions with Large Language Model

DGX agent

arXiv:2606.06099v1 Announce Type: new Abstract: Whether Large Language Models (LLMs) exhibit covert psychological manipulation in complex human-AI interactions has garnered increasing safety concerns.

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

Cognitive Threat Intelligence and Explainable Federated Security Analytics for distributed Infrastructure Systems

DGX agent

arXiv:2606.05701v1 Announce Type: cross Abstract: The increasing adoption of distributed infrastructure systems, cloud computing, Internet of Things (IoT) technologies, and edge-based architectures ha

local-aiarxiv-cs-ai
6 Jun 2026
Local Ai

Compositional Boundaries for Density Fusion

DGX agent

arXiv:2606.05871v1 Announce Type: cross Abstract: Distributed uncertainty-management systems often combine local probabilistic models along aggregation trees chosen by communication, privacy, or sched

local-aiarxiv-cs-ai
6 Jun 2026
Safety

Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Frequency-Domain Insights

DGX agent

arXiv:2411.18343v3 Announce Type: replace-cross Abstract: Personalized Federal learning(PFL) allows clients to cooperatively train a personalized model without disclosing their private dataset. Howeve

safetyarxiv-cs-ai
6 Jun 2026
Safety

Conformal Risk-Averse Decision Making with Action Conditional Guarantee

DGX agent

arXiv:2606.05551v1 Announce Type: cross Abstract: Reliable decision making pipelines powered by machine learning models require uncertainty quantification (UQ) methods that come with explicit safety g

safetyarxiv-cs-ai
6 Jun 2026
Safety

Consistency Training Along the Transformer Stack

DGX agent

arXiv:2606.05817v1 Announce Type: cross Abstract: Consistency training encourages models to behave similarly across different contexts, and has shown promise for reducing misalignment. We broaden the

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

DGX agent

arXiv:2606.05704v1 Announce Type: new Abstract: Recent Large Language Models (LLMs) have shown impressive reasoning abilities; but they are still susceptible to hallucinations, intermediate reasoning

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Cross-Epoch Adaptive Rollout Optimization for RL Post-Training

DGX agent

arXiv:2606.05606v1 Announce Type: cross Abstract: LLM post-training often relies on reinforcement learning methods that sample multiple rollouts per prompt, yet most existing approaches use a fixed ro

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CTIConnect: A Benchmark for Retrieval-Augmented LLMs over Heterogeneous Cyber Threat Intelligence

DGX agent

arXiv:2510.11974v2 Announce Type: replace-cross Abstract: Cyber Threat Intelligence (CTI) is foundational to modern cybersecurity, enabling organizations to proactively defend against evolving threats

model-releasesarxiv-cs-ai
6 Jun 2026
Hardware

CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe

DGX agent

arXiv:2604.01489v2 Announce Type: replace-cross Abstract: High-performance GPU kernels are critical to modern machine learning systems, yet developing them remains a manual, expert-driven process. Rec

hardwarearxiv-cs-ai
6 Jun 2026
Agents

DAST: A VLM-LLM Framework for Cross-Interface Anomaly Detection in O-RAN

DGX agent

arXiv:2606.06261v1 Announce Type: cross Abstract: O-RAN enables a disaggregated baseband stack with programmable functions that communicate over standardized open interfaces. The same openness that en

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

Data Flow Control: Data Safety Policies for AI Agents

DGX agent

arXiv:2606.05679v1 Announce Type: cross Abstract: Agents increasingly generate SQL, orchestrate pipelines, and automate data analysis on behalf of users. While recent work improves query correctness,

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Deciphering Two Training Clocks in Grokking via Deep Linear Network Theory with Conditional ReLU Reduction

DGX agent

arXiv:2606.05863v1 Announce Type: cross Abstract: Grokking suggests that fitting the training data and learning a simple underlying rule may occur on different time scales. We formalize this phenomeno

researcharxiv-cs-ai
6 Jun 2026
Local Ai

Design a Reliable LLM-Integrated Interface for Mortality Forecasting

DGX agent

arXiv:2606.06235v1 Announce Type: cross Abstract: Mortality forecasting plays an important role in actuarial and policy decision-making, but its implementation remains technically complex and inaccess

local-aiarxiv-cs-ai
6 Jun 2026
Agents

Detecting Perspective Shifts in Multi-agent Systems

DGX agent

arXiv:2512.05013v2 Announce Type: replace Abstract: Generative models augmented with external tools and update mechanisms (or extit{agents}) have demonstrated capabilities beyond intelligent prompting

agentsarxiv-cs-ai
6 Jun 2026
Safety

Differentiable Efficient Operator Search

DGX agent

arXiv:2606.05232v1 Announce Type: cross Abstract: Efficient multimodal foundation models often rely on manually designed token-reduction operators, such as pruning, merging, pooling, and adaptive rewe

safetyarxiv-cs-ai
6 Jun 2026
Research

Dimensionality Reduction for Cyberattack Classification: A Comparative Evaluation of PCA and Linear Predictive Coding

DGX agent

arXiv:2606.05584v1 Announce Type: cross Abstract: High-dimensional feature representations are widely used in machine learning-based cyberattack detection systems. However, they increase computational

researcharxiv-cs-ai
6 Jun 2026
Model Releases

Do More Agents Help? Controlled and Protocol-Aligned Evaluation of LLM Agent Workflows

DGX agent

arXiv:2606.05670v1 Announce Type: new Abstract: Does adding more agents help an LLM workflow once compared systems share the same benchmark loader, tool access, answer contract, usage accounting, and

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss

DGX agent

arXiv:2606.06418v1 Announce Type: cross Abstract: Many modern applications of deep learning involve training a neural network via a one-step prediction loss (e.g., L^2 regression, cross-entropy), but

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

DPBench: Structural Determinants of Multi-Agent LLM Coordination Under Simultaneous Resource Contention

DGX agent

arXiv:2602.13255v2 Announce Type: replace Abstract: We present DPBench, a benchmark for evaluating coordination in multi-agent systems built from large language models. Existing benchmarks measure tas

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

DGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

ECI: Effective Contrastive Information to Evaluate Hard-Negatives

DGX agent

arXiv:2603.20990v2 Announce Type: replace-cross Abstract: Hard-negative source selection for dense retrieval is usually decided only after fine-tuning and downstream evaluation. We propose Effective C

local-aiarxiv-cs-ai
6 Jun 2026
Model Releases

Edit-R2: Context-Aware Reinforcement Learning for Multi-Turn Image Editing

DGX agent

arXiv:2606.05950v1 Announce Type: new Abstract: Text-guided image editing has advanced rapidly with diffusion models and unified multimodal foundation models. However, most existing methods remain con

model-releasesarxiv-cs-ai
6 Jun 2026
Tutorials

EEGDancer: Dynamic Emotion Latent Space Masked Modeling with Reinforcement Learning for EEG Continuous Emotion Prediction

DGX agent

arXiv:2606.05855v1 Announce Type: cross Abstract: Continuous electroencephalography (EEG) emotion prediction aims to model the temporal evolution of human emotional states from EEG signals. Unlike con

tutorialsarxiv-cs-ai
6 Jun 2026
Local Ai

Efficient Asynchronous Federated Evaluation with Strategy Similarity Awareness for Intent-Based Networking in Industrial Internet of Things

DGX agent

arXiv:2512.20627v2 Announce Type: replace-cross Abstract: Intent-Based Networking (IBN) offers a promising paradigm for intelligent and automated network control in Industrial Internet of Things (IIoT

local-aiarxiv-cs-ai
6 Jun 2026
Model Releases

Enhancing Software Engineering Through Closed-Loop Memory Optimization

DGX agent

arXiv:2606.05646v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled powerful software engineering (SE) agents capable of navigating complex codebases and resolving real-world i

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Escaping the Verifier: Learning to Reason via Demonstrations

DGX agent

arXiv:2511.21667v4 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) to reason often relies on Reinforcement Learning (RL) with task-specific verifiers. However, many real-w

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluating Agentic Configuration Repair for Computer Networks

DGX agent

arXiv:2606.06212v1 Announce Type: new Abstract: Misconfigurations in computer networks remain a major source of critical Internet outages. Research is turning to Large Language Models (LLMs) to automa

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluation of LLMs for Mathematical Formalization in Lean

DGX agent

arXiv:2606.05632v1 Announce Type: new Abstract: Within the past few years, the ability of Large Language Models (LLMs) to generate formal mathematical proofs has improved drastically. We provide a com

model-releasesarxiv-cs-ai
6 Jun 2026
Applications

Explainable AI-Driven Cyber Risk Analytics and Model Reliability Assessment for Intelligent Governance of U.S. Critical Infrastructure: An XGBoost and SHAP-Based Intrusion Detection Framework

DGX agent

arXiv:2606.05710v1 Announce Type: cross Abstract: The increasing penetrations of the critical infrastructure sector in the United States with intelligent digital technologies have greatly increased ex

applicationsarxiv-cs-ai
6 Jun 2026
Model Releases

Exploring LLMs for South Asian Music Understanding and Generation

DGX agent

arXiv:2606.05522v1 Announce Type: cross Abstract: Recent advancements in Large Language Models (LLMs) have shown promising results in music understanding and generation tasks. However, existing works

model-releasesarxiv-cs-ai
6 Jun 2026
Research

F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation

DGX agent

arXiv:2606.06357v1 Announce Type: cross Abstract: Continuous audio autoencoders reconstruct waveforms well but often produce latents with weak structure for understanding, while self-supervised audio

researcharxiv-cs-ai
6 Jun 2026
Safety

FIDES: Faithful Inference via Deep Evidence Signals for Retrieval-Memory Conflict in RAG

DGX agent

arXiv:2606.05644v1 Announce Type: new Abstract: When retrieved evidence contradicts parametric memory, language models frequently ignore context and default to memorized priors -- a failure that under

safetyarxiv-cs-ai
6 Jun 2026
Research

Finite Element-Based Material Learning via Automatic Differentiation: Learning constitutive neural network models from full-field deformation data

DGX agent

arXiv:2606.05199v1 Announce Type: cross Abstract: The identification of constitutive neural network models from heterogeneous full-field deformation data provides a robust alternative to traditional c

researcharxiv-cs-ai
6 Jun 2026
Research

Fix the Mind, Not the Move: Interpretable AI Assistance via Knowledge-Gap Localization

DGX agent

arXiv:2606.05602v1 Announce Type: new Abstract: AI assistants in human-AI collaboration often correct suboptimal human actions through behavioral feedback (e.g., alerts or steering-wheel nudges in ass

researcharxiv-cs-ai
6 Jun 2026
Applications

From Attack Simulation to SIEM Rule: Deterministic Detection-as-Code Synthesis with Probe-Level Traceability

DGX agent

arXiv:2606.05252v1 Announce Type: cross Abstract: Security teams routinely simulate attacks against their own systems to check whether their monitoring would catch a real intruder. These Breach-and-At

applicationsarxiv-cs-ai
6 Jun 2026
Safety

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

DGX agent

arXiv:2606.06223v1 Announce Type: new Abstract: Language-model agents act through repeated cycles of observation, reasoning, and action selection, making safety monitoring depend on both internal mode

safetyarxiv-cs-ai
6 Jun 2026
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

GenTI: Benchmarking LLMs for Autonomous IDPS Rule Generation for Unseen Attacks

DGX agent

arXiv:2606.05844v1 Announce Type: cross Abstract: Rule-based Intrusion Detection and Prevention Systems (IDPS) offer precise attack detection as well as mitigation, however their manually crafted, sig

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Geographic Bias and Diversity in AI Evaluation

DGX agent

arXiv:2606.05187v1 Announce Type: cross Abstract: Among the many challenges hindering the responsible development and deployment of AI, arguably none has faced more intense scrutiny than bias in its v

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

GIPO: Gaussian Importance Sampling Policy Optimization

DGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

GITCO: Gated Inference-Time Context Optimization in TSFMs

DGX agent

arXiv:2606.05332v1 Announce Type: new Abstract: Patch-based Time Series Foundation Models (TSFMs) suffer from context poisoning: structurally anomalous patches capture disproportionate attention and s

model-releasesarxiv-cs-ai
6 Jun 2026
← Previous
1…194195196197198…448
Next →