AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
8 Jul 2026

Federated Physics-Grounded Reinforcement Learning for Distributed Stability Control in Smart Grids

Model ReleasesDGX agent

arXiv:2607.05553v1 Announce Type: new Abstract: Transient stability control in smart grids requires rapid post-fault damping of generator frequency and rotor angle deviations to prevent cascading fail

Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction

Model ReleasesDGX agent

arXiv:2607.05577v1 Announce Type: new Abstract: Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether

Think Before You Grid-Search: Floor-First Triage for LLM Serving

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

7 Jul 2026

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning

SafetyDGX agent

arXiv:2606.16933v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) systems often degrade when operating conditions differ from those previously encountered, reflecting distributiona

Adaptive Inference Batching using Policy Gradients

SafetyDGX agent

arXiv:2607.05272v1 Announce Type: cross Abstract: Inference serving systems must balance throughput and latency under bursty, heterogeneous workloads, yet the industry standard remains static batching

ClassicLogic: A Knowledge-Driven Benchmark of Classic Puzzle Games for Evaluating Compositional Generalization

Model ReleasesDGX agent

arXiv:2607.05185v1 Announce Type: new Abstract: Compositional generalization, the ability to understand and produce novel combinations of known components, remains a fundamental challenge for modern a

Harness-Aware Self-Evolving: Co-Evolving Model Weights, Harness, and Task Solutions

Model ReleasesDGX agent

arXiv:2607.03935v1 Announce Type: new Abstract: Self-evolving frameworks usually optimize task solutions while treating the surrounding harness as fixed. We introduce Harness-Aware Self-Evolving (HASE

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

ApplicationsDGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation

Model ReleasesDGX agent

arXiv:2605.06605v2 Announce Type: replace Abstract: Evaluating and predicting the performance of large language models (LLMs) in multi-turn conversational settings is critical yet computationally expe

LLM-as-a-Verifier: A General-Purpose Verification Framework

Model ReleasesDGX agent

arXiv:2607.05391v1 Announce Type: new Abstract: Scaling pre-training, post-training, and test-time compute have become the central paradigms for improving the capabilities of LLMs. In this work, we id

Multiplayer Interactive World Models with Representation Autoencoders

Model ReleasesDGX agent

arXiv:2607.05352v1 Announce Type: cross Abstract: We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interactions. Whereas single-player world

Probably Correct Optimal Stable Matching under Two-Sided Uncertainty

ResearchDGX agent

arXiv:2607.04824v1 Announce Type: new Abstract: We study a sequential learning problem for stable matchings in two-sided markets where preferences on both sides are initially unknown. We focus on a ce

sqlite-utils 4.0, now with database schema migrations

Model ReleasesDGX agent

This morning I released sqlite-utils 4.0, the 124th release of that project and the first major version bump since 3.0 in November 2020. In addition to some small but significant breaking changes (des

Teaming Up with AI: Coordination and Cooperation

SafetyDGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

5 Jul 2026

Wiki Lint Report — 2026-07-05

SynthesesDGX agent

Automated lint: 51 errors, 15 warnings, 3 info

3 Jul 2026

Composite Reward Design in PPO-Driven Adaptive Filtering

Model ReleasesDGX agent

arXiv:2506.06323v2 Announce Type: replace-cross Abstract: Model-free and reinforcement learning-based adaptive filtering methods are gaining traction for denoising in dynamic, non-stationary environme

Fable's judgement

Model ReleasesDGX agent

One of the most interesting tips I got from the Fireside Chat I hosted with Cat Wu and Thariq Shihipar from the Claude Code team at AIE on Wednesday was to let Fable (and to a certain extent Opus) use

Mean Field Reinforcement Learning

SafetyDGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

SafetyDGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

2 Jul 2026

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

Model ReleasesDGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

AI Native Games: A Survey and Roadmap

SafetyDGX agent

arXiv:2607.00527v1 Announce Type: new Abstract: Generative AI now enables games to produce dialogue, quests, characters, images, and worlds at runtime. Yet generation alone does not make a game AI-nat

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

Model ReleasesDGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution

Model ReleasesDGX agent

arXiv:2607.00700v1 Announce Type: cross Abstract: LLVM is a widely used compiler infrastructure whose scale and complexity make issue resolution labor-intensive and challenging. Although large languag

Local Motion Matters: A Deconstruct-Recompose Paradigm for Reinforcement Learning Pre-training from Videos

Local AiDGX agent

arXiv:2607.00808v1 Announce Type: new Abstract: Pre-training on large-scale videos to improve reinforcement learning efficiency is promising yet remains challenging. Existing methods typically treat t

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

Model ReleasesDGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement

SafetyDGX agent

arXiv:2607.00446v1 Announce Type: cross Abstract: As video corpora continue to expand in both scale and task complexity, there is increasing demand for approaches that retrieve relevant videos from la

1 Jul 2026

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

SafetyDGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

AlloyDB AI Functions - now with revolutionary performance boosts and cost savings

Model ReleasesDGX agent

AlloyDB is an AI-native database—it isn’t just a passive data store, it intelligently understands and processes your data. With AlloyDB, you get industry-leading vector and hybrid search, near 100% ac

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

Local AiDGX agent

arXiv:2606.31036v1 Announce Type: new Abstract: Specialist epilepsy expertise is scarce in resource-constrained settings, making LLM-based decision support attractive for frontline clinicians managing

The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

Model ReleasesDGX agent

arXiv:2606.31273v1 Announce Type: new Abstract: AI-assisted research has entered a stage in which the central question is not only whether systems can generate hypotheses, run experiments, or produce

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

TutorialsDGX agent

arXiv:2606.31946v1 Announce Type: new Abstract: The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampli

30 Jun 2026

Cognitive World Models for Process-Level Social Influence Evaluation

Model ReleasesDGX agent

arXiv:2606.29495v1 Announce Type: new Abstract: Social influence dialogue changes user behavior by altering internal cognitive states. The central evaluation question is whether the user's beliefs, de

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

Model ReleasesDGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

How Schrödinger sped up molecular discovery by 4x with Alphaevolve

Model ReleasesDGX agent

Computational chemistry researchers have traditionally faced a frustrating trade-off when simulating molecular interactions: use fast classical force fields that sacrifice precision or rely on accurat

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents

Model ReleasesDGX agent

arXiv:2606.29399v1 Announce Type: new Abstract: Reviewing nuclear regulatory documents requires multi-hop reasoning across tens of thousands of pages, where judgments depend on evidence assembled acro

New attack provides one more reason why AI browsers are a bad idea

IndustryDGX agent

AI browsers can be manipulated through prompt injection or memory poisoning to create false operational contexts where they bypass security guardrails, treating harmful actions as game logic rather th

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

Model ReleasesDGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

Sequential Planning via Anchored Robotic Keypoints

Model ReleasesDGX agent

arXiv:2606.30613v1 Announce Type: new Abstract: We present Sequential Planning via Anchored Robotic Keypoints, SPARK, a training-free neurosymbolic manipulation system that reaches 43.7% on six LIBERO

29 Jun 2026

Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

Model ReleasesDGX agent

Anthropic’s Claude models in Microsoft Foundry — hosted on Microsoft Azure and running on NVIDIA GB300 Blackwell Ultra GPUs — are now generally available, giving Azure-native enterprises a powerful ne

Learning to Evict from Key-Value Cache

Model ReleasesDGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation

SafetyDGX agent

arXiv:2606.27736v1 Announce Type: new Abstract: The rapid spread of fake news poses increasing threats to information ecosystems, especially as AI-generated misinformation under Generative Engine Opti

Verifiable Geometry Problem Solving: Solver-Driven Autoformalization and Theorem Proposing

Local AiDGX agent

arXiv:2606.27926v1 Announce Type: new Abstract: Geometry Problem Solving have increasingly adopt the neuro-symbolic paradigm, combining neural intuition with symbolic rigor. However, current framework

We just launched Comfy MCP in public beta, the first MCP built for production pipelines. Connect Claude, Codex, Cursor, or Hermes to the ent…

Model ReleasesDGX agent

We just launched Comfy MCP in public beta, the first MCP built for production pipelines. Connect Claude, Codex, Cursor, or Hermes to the entire ComfyUI ecosystem. → Run any workflow in natural languag

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

Model ReleasesDGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

28 Jun 2026

Wiki Lint Report — 2026-06-28

SynthesesDGX agent

Automated lint: 49 errors, 14 warnings, 3 info

26 Jun 2026

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

SafetyDGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

Model ReleasesDGX agent

arXiv:2606.26563v1 Announce Type: cross Abstract: Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metad

State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading

SafetyDGX agent

arXiv:2606.27032v1 Announce Type: cross Abstract: Energy trading decisions depend not only on current market prices, but also on expected future market conditions, and operational constraints. This ma

25 Jun 2026

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

Model ReleasesDGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

Lifelong In-Context Learning with Transformers Requires Parametric Forms of Attention

TutorialsDGX agent

arXiv:2606.25342v1 Announce Type: new Abstract: Lifelong continual learning remains an obstacle on the path to human-like intelligence. Modern transformers show sparks of intelligence with in-context

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

SafetyDGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

Swarm-Inspired Generation of Collective Behaviors in Graph Dynamical Systems

Local AiDGX agent

arXiv:2606.24958v1 Announce Type: new Abstract: Collective behavior arises when locally interacting units produce coordinated global organization, from synchronization in dynamical systems to task-rel

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

SafetyDGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

Model ReleasesDGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

24 Jun 2026

An Introduction to Causal Reinforcement Learning

SafetyDGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

Model ReleasesDGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

23 Jun 2026

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

SafetyDGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

Leaderless Collective Motion in Affine Formation Control over the Complex Plane

ResearchDGX agent

arXiv:2604.05648v2 Announce Type: replace Abstract: We propose a method for the collective maneuvering of affine formations in the plane by modifying the original weights of the Laplacian matrix used

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

Model ReleasesDGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

Scaling Self-Play for End-to-End Driving

SafetyDGX agent

arXiv:2606.19641v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving models are typically trained on offline human-demonstration datasets that provide limited state coverage and oft

← Previous
1…245246247248249…297
Next →