AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Local Ai

Retrieval-Augmented Large Language Models as Components of Cognitive Computing architecture for Regulatory Knowledge Management

DGX agent

arXiv:2607.24352v1 Announce Type: new Abstract: The aim of this article is to verify whether integrating large language models (LLMs) with the Retrieval-Augmented Generation (RAG) architecture enables

local-aiarxiv-cs-cl
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems

DGX agent

arXiv:2603.03536v2 Announce Type: replace-cross Abstract: Current LLM-based conversational recommender systems (CRS) primarily optimize recommendation accuracy and user satisfaction. We identify an un

model-releasesarxiv-cs-ai
28 Jul 2026
Research

STEER: Steerable Dyadic Head Avatars

DGX agent

arXiv:2607.23840v1 Announce Type: new Abstract: Facial movement and expression are central to face-to-face communication, conveying turn-taking, attention, agreement, and engagement alongside speech.

researcharxiv-cs-cv
28 Jul 2026
Safety

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

DGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

safetyarxiv-cs-ai
28 Jul 2026
Research

Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy

DGX agent

arXiv:2607.24304v1 Announce Type: cross Abstract: We formulate a statistical physics framework to model a networked stochastic dynamical system exhibiting bistability, driven by additive noise and soc

researcharxiv-cs-ai
28 Jul 2026
Research

Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models

DGX agent

arXiv:2607.22575v1 Announce Type: new Abstract: Human episodic memory supports the retrieval of experiences that unfold over extended timescales, yet the computational mechanisms underlying this abili

researcharxiv-cs-ai
28 Jul 2026
Safety

The Curse of Precision: A Data Scaling Law for High-Precision Robotic Manipulation

DGX agent

arXiv:2607.23108v1 Announce Type: new Abstract: While scaling laws for imitation learning have primarily focused on generalization in open-world settings, the relationship between data and precision i

safetyarxiv-cs-ro
28 Jul 2026
Research

Towards High-Level Semantic Intelligence

DGX agent

arXiv:2607.24082v1 Announce Type: new Abstract: Recent advances in AI have substantially expanded its cognitive and reasoning capabilities. From the perspective of semantic complexity, the development

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Using Reinforcement Learning to Optimize the Global and Local Crossing Number

DGX agent

arXiv:2509.06108v3 Announce Type: replace-cross Abstract: Graph drawing concerns the algorithmic visualization of graphs. A good drawing of a graph is easy to read and facilitates solving tasks on the

model-releasesarxiv-cs-lg
28 Jul 2026
Safety

When Every Simulation Counts: Value-Based Reinforcement Learning for Accelerated Photonics Inverse Design

DGX agent

arXiv:2607.23469v1 Announce Type: cross Abstract: Photonic-crystal surface-emitting lasers (PCSELs) can combine high-power operation with narrow-divergence surface emission, but optimizing coupled par

safetyarxiv-cs-ai
28 Jul 2026
Safety

Adaptive Undulatory Locomotion of Snake-like Robots in Dynamic Viscous Environments via Deep Reinforcement Learning

DGX agent

arXiv:2607.21960v1 Announce Type: new Abstract: This paper demonstrates how deep reinforcement learning (DRL) enables adaptive locomotion of snake-like robots in dynamically changing viscous environme

safetyarxiv-cs-ro
27 Jul 2026
Safety

Adjustment Speed as a Safety Constraint for Nonstationary Reinforcement Learning

DGX agent

arXiv:2607.21646v1 Announce Type: new Abstract: Ensuring safety in reinforcement learning under nonstationarity requires determining whether a learning system can safely adapt to forecasted environmen

safetyarxiv-cs-lg
27 Jul 2026
Safety

Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization

DGX agent

arXiv:2607.21619v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive performance, but their safety alignment remains vulnerable to jailbreak attacks. Exist

safetyarxiv-cs-cl
27 Jul 2026
Model Releases

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

DGX agent

arXiv:2607.10428v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) as multi-turn conversational partners requires probing capabilities that single-turn benchmarks miss: person

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

LMEB: Long-horizon Memory Embedding Benchmark

DGX agent

arXiv:2603.12572v5 Announce Type: replace Abstract: Memory embeddings are crucial for memory-augmented systems, such as OpenClaw, but their evaluation is underexplored in current text embedding benchm

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Pixels for Programs? A Cross-Provider Case Study of Input-Token Accounting for Source Code as Text and Images

DGX agent

arXiv:2607.21672v1 Announce Type: cross Abstract: Long source-code contexts consume many text tokens, motivating the proposal to render code as images for vision-language models. Recent work asks whet

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures

DGX agent

arXiv:2607.21612v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like LoRA have become the default for adapting large language models, succeeding across instruction following,

model-releasesarxiv-cs-lg
27 Jul 2026
Local Ai

SCALE: Self-Supervised Constraint-Aware Layout GEneration for Local P&R DRV Fixing at Advanced Nodes

DGX agent

arXiv:2607.21850v1 Announce Type: new Abstract: As semiconductor manufacturing advances toward sub-2nm nodes, local place-and-route (P&R) design-rule violation (DRV) fixing is increasingly limited by

local-aiarxiv-cs-cv
27 Jul 2026
Model Releases

Adaptive Multi-Horizon Reinforcement Learning

DGX agent

arXiv:2607.20656v1 Announce Type: cross Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL

model-releasesarxiv-cs-ai
24 Jul 2026
Applications

Algorithmic Approaches to Sequential Decision-Making and Social Epistemology

DGX agent

arXiv:2607.20636v1 Announce Type: cross Abstract: As humans, we face many decisions that require us to choose between sticking to something and giving up. This thesis uses algorithmic tools to derive

applicationsarxiv-cs-lg
24 Jul 2026
Safety

End-to-End Learning of Safe Optimal Feedback Control in High Dimensions with Control Barrier Function Layers

DGX agent

arXiv:2607.20674v1 Announce Type: new Abstract: We consider the problem of learning high-dimensional semi-global feedback controllers under hard safety constraints enforced by control barrier function

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

Engine-Native Editable 3D World Reconstruction with Objects and Lighting

DGX agent

arXiv:2607.20889v1 Announce Type: new Abstract: Editable 3D scene creation requires object instances and lights that can be inspected, moved, and imported into standard engines, yet existing single-im

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

Expert Behavior Prior Reinforcement Learning

DGX agent

arXiv:2607.21302v1 Announce Type: new Abstract: Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by l

safetyarxiv-cs-ai
24 Jul 2026
Safety

GeoWorldAD: Geometry World Action Model for Autonomous Driving

DGX agent

arXiv:2607.17521v2 Announce Type: replace Abstract: Autonomous driving requires both safe and efficient planning decisions in dynamic 3D environments. Although recent Vision/Video-Action models learn

safetyarxiv-cs-ro
24 Jul 2026
Safety

HERMES: Heterogeneous Edge-Relational Multi-Head Embedded SSM Attention for Traffic Conflict Prediction at Signalized Intersections

DGX agent

arXiv:2607.20505v1 Announce Type: cross Abstract: Surrogate safety measures (SSMs) enable proactive traffic safety assessment, but many existing methods evaluate pairwise interactions independently or

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

Multi-turn RL with Structural and Performance Aware Rewards for CUDA Kernel Generation

DGX agent

arXiv:2607.20908v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a powerful technique to enhance the reasoning capacity of LLMs for optimized code

model-releasesarxiv-cs-ai
24 Jul 2026
Research

PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers

DGX agent

arXiv:2605.06064v2 Announce Type: replace Abstract: We propose PersonaGesture, a diffusion-based pipeline for single-reference co-speech gesture personalization of unseen speakers. Given target speech

researcharxiv-cs-cv
24 Jul 2026
Safety

thaulab@EEUCA 2026: Who Said What to Whom? A Targeting-Aware Neural-Symbolic Pipeline for Gaming Toxicity Detection

DGX agent

arXiv:2607.20447v1 Announce Type: new Abstract: This paper describes our system for the EEUCA 2026 Shared Task on toxicity classification in gaming chat. We implement a three-stage pipeline combining

safetyarxiv-cs-cl
24 Jul 2026
Model Releases

When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion

DGX agent

arXiv:2607.20543v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study thi

model-releasesarxiv-cs-ai
24 Jul 2026
Local Ai

ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU

DGX agent

arXiv:2607.19191v1 Announce Type: new Abstract: We present ABot-World-0, an action-conditioned video world model for real-time, long-horizon closed-loop interaction, supported by a multi-source data i

local-aiarxiv-cs-cv
23 Jul 2026
Research

Distributed Optimization via Energy Conservation Laws in Dilated Coordinates

DGX agent

arXiv:2409.19279v2 Announce Type: replace-cross Abstract: Continuous-time models can reveal accelerated structures in distributed optimization, but their rates need not survive direct discretization.

researcharxiv-cs-ai
23 Jul 2026
Safety

Drift-Aware RL-based Wavelet Denoising for Network-Traffic Anomaly Detection

DGX agent

arXiv:2607.20011v1 Announce Type: cross Abstract: Traffic-utilisation measurements for network monitoring are corrupted by additive noise and statistical drift: time-dependent change in the signal's m

safetyarxiv-cs-ai
23 Jul 2026
Applications

EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness

DGX agent

arXiv:2607.19880v1 Announce Type: new Abstract: Cross-embodiment navigation is a key challenge in embodied intelligence. Due to differences in embodiment, the same visual observation may imply differe

applicationsarxiv-cs-ro
23 Jul 2026
Model Releases

Hybrid LLM-Guided Search for Quantum Reservoir Architecture Design

DGX agent

arXiv:2607.19506v1 Announce Type: cross Abstract: Quantum reservoir computing (QRC) uses fixed quantum dynamics as a high-dimensional temporal feature map and trains only a lightweight classical reado

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning

DGX agent

arXiv:2607.19397v1 Announce Type: new Abstract: Deep Q-networks use target networks to stabilise bootstrapped value learning, but the standard hard copy update also introduces a tradeoff. Holding the

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Model Gateway: Management Platform for Model-Driven Drug Discovery

DGX agent

arXiv:2512.05462v2 Announce Type: replace-cross Abstract: Pharmaceutical drug discovery demands machine learning (ML) infrastructure that goes beyond general-purpose Machine Learning Operations (MLOps

model-releasesarxiv-cs-lg
23 Jul 2026
Safety

MOF-Sleuth: Tool-Grounded Reward Alignment for Explainable Fine-Grained MOF CIF Auditing

DGX agent

arXiv:2607.19935v1 Announce Type: new Abstract: Large metal-organic framework (MOF) databases support simulation, screening, and machine learning through crystallographic information files (CIFs). Sub

safetyarxiv-cs-ai
23 Jul 2026
Research

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

DGX agent

arXiv:2607.20284v1 Announce Type: new Abstract: The rapid development of multimodal large language models (MLLMs) has introduced a flexible paradigm for remote sensing image scene understanding (RSISU

researcharxiv-cs-cv
23 Jul 2026
Safety

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

DGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Safe Remediation as Risk-Constrained Intervention Decision in Microservice Systems

DGX agent

arXiv:2607.20005v1 Announce Type: new Abstract: In modern IT operations (IT-Ops), the cost of an incorrect repair often exceeds the cost of no action at all. Yet existing automated remediation systems

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

SafeGen: Goal-Conditioned Video Diffusion of Safety-Critical Scenarios for VLM-Based Autonomous Driving

DGX agent

arXiv:2607.19701v1 Announce Type: new Abstract: VLMs are increasingly deployed in AD systems, creating an urgent need for rigorous safety evaluation under rare yet safety-critical scenarios. Among the

safetyarxiv-cs-cv
23 Jul 2026
Safety

Self-Explaining Reinforcement Learning for Mobile Network Resource Allocation

DGX agent

arXiv:2509.14925v2 Announce Type: replace Abstract: Deep reinforcement learning (DRL) methods, though powerful, often lack transparency, which limits their adoption in critical domains. We apply Self-

safetyarxiv-cs-lg
23 Jul 2026
Safety

The Mechanism Matters: When Knowledge Graphs Help Reinforcement Learning

DGX agent

arXiv:2607.19616v1 Announce Type: new Abstract: Knowledge graphs (KGs) are widely used to inject prior knowledge into reinforcement learning (RL), yet the literature is dominated by single-domain, pos

safetyarxiv-cs-lg
23 Jul 2026
Safety

Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion

DGX agent

arXiv:2607.18365v1 Announce Type: cross Abstract: Reinforcement learning (RL) for legged robots is advancing locomotion, demonstrating its ability to adapt to new and challenging terrain. Traditionall

safetyarxiv-cs-lg
23 Jul 2026
Model Releases

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks

DGX agent

arXiv:2607.13305v1 Announce Type: cross Abstract: Benchmark accuracy in video large language models (LLMs) is often treated as evidence of visual understanding. We audit this assumption across twenty

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

DGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

safetyarxiv-cs-ai
16 Jul 2026
Safety

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

DGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

safetyarxiv-cs-ai
16 Jul 2026
Local Ai

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

DGX agent

arXiv:2607.13960v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future visual observations, using future scene evolution as den

local-aiarxiv-cs-ro
16 Jul 2026
← Previous
1…207208209210211…233
Next →