AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
Safety

Mean Field Reinforcement Learning

DGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

safetyarxiv-cs-lg
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
safetyarxiv-cs-ai
3 Jul 2026
Model Releases

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

DGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

model-releasesarxiv-cs-ai
2 Jul 2026
Safety

AI Native Games: A Survey and Roadmap

DGX agent

arXiv:2607.00527v1 Announce Type: new Abstract: Generative AI now enables games to produce dialogue, quests, characters, images, and worlds at runtime. Yet generation alone does not make a game AI-nat

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

DGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution

DGX agent

arXiv:2607.00700v1 Announce Type: cross Abstract: LLVM is a widely used compiler infrastructure whose scale and complexity make issue resolution labor-intensive and challenging. Although large languag

model-releasesarxiv-cs-ai
2 Jul 2026
Local Ai

Local Motion Matters: A Deconstruct-Recompose Paradigm for Reinforcement Learning Pre-training from Videos

DGX agent

arXiv:2607.00808v1 Announce Type: new Abstract: Pre-training on large-scale videos to improve reinforcement learning efficiency is promising yet remains challenging. Existing methods typically treat t

local-aiarxiv-cs-lg
2 Jul 2026
Model Releases

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

DGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

model-releasesarxiv-cs-cv
2 Jul 2026
Safety

VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement

DGX agent

arXiv:2607.00446v1 Announce Type: cross Abstract: As video corpora continue to expand in both scale and task complexity, there is increasing demand for approaches that retrieve relevant videos from la

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

DGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

AlloyDB AI Functions - now with revolutionary performance boosts and cost savings

DGX agent

AlloyDB is an AI-native database—it isn’t just a passive data store, it intelligently understands and processes your data. With AlloyDB, you get industry-leading vector and hybrid search, near 100% ac

model-releasesgoogle-cloud-ai
1 Jul 2026
Local Ai

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

DGX agent

arXiv:2606.31036v1 Announce Type: new Abstract: Specialist epilepsy expertise is scarce in resource-constrained settings, making LLM-based decision support attractive for frontline clinicians managing

local-aiarxiv-cs-lg
1 Jul 2026
Model Releases

The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

DGX agent

arXiv:2606.31273v1 Announce Type: new Abstract: AI-assisted research has entered a stage in which the central question is not only whether systems can generate hypotheses, run experiments, or produce

model-releasesarxiv-cs-lg
1 Jul 2026
Tutorials

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

DGX agent

arXiv:2606.31946v1 Announce Type: new Abstract: The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampli

tutorialsarxiv-cs-cv
1 Jul 2026
Model Releases

Cognitive World Models for Process-Level Social Influence Evaluation

DGX agent

arXiv:2606.29495v1 Announce Type: new Abstract: Social influence dialogue changes user behavior by altering internal cognitive states. The central evaluation question is whether the user's beliefs, de

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

DGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How Schrödinger sped up molecular discovery by 4x with Alphaevolve

DGX agent

Computational chemistry researchers have traditionally faced a frustrating trade-off when simulating molecular interactions: use fast classical force fields that sacrifice precision or rely on accurat

model-releasesgoogle-cloud-ai
30 Jun 2026
Model Releases

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents

DGX agent

arXiv:2606.29399v1 Announce Type: new Abstract: Reviewing nuclear regulatory documents requires multi-hop reasoning across tens of thousands of pages, where judgments depend on evidence assembled acro

model-releasesarxiv-cs-ai
30 Jun 2026
Industry

New attack provides one more reason why AI browsers are a bad idea

DGX agent

AI browsers can be manipulated through prompt injection or memory poisoning to create false operational contexts where they bypass security guardrails, treating harmful actions as game logic rather th

industryars-technica
30 Jun 2026
Model Releases

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

DGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Sequential Planning via Anchored Robotic Keypoints

DGX agent

arXiv:2606.30613v1 Announce Type: new Abstract: We present Sequential Planning via Anchored Robotic Keypoints, SPARK, a training-free neurosymbolic manipulation system that reaches 43.7% on six LIBERO

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

DGX agent

Anthropic’s Claude models in Microsoft Foundry — hosted on Microsoft Azure and running on NVIDIA GB300 Blackwell Ultra GPUs — are now generally available, giving Azure-native enterprises a powerful ne

model-releasesnvidia-blog
29 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Safety

ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation

DGX agent

arXiv:2606.27736v1 Announce Type: new Abstract: The rapid spread of fake news poses increasing threats to information ecosystems, especially as AI-generated misinformation under Generative Engine Opti

safetyarxiv-cs-ai
29 Jun 2026
Local Ai

Verifiable Geometry Problem Solving: Solver-Driven Autoformalization and Theorem Proposing

DGX agent

arXiv:2606.27926v1 Announce Type: new Abstract: Geometry Problem Solving have increasingly adopt the neuro-symbolic paradigm, combining neural intuition with symbolic rigor. However, current framework

local-aiarxiv-cs-ai
29 Jun 2026
Model Releases

We just launched Comfy MCP in public beta, the first MCP built for production pipelines. Connect Claude, Codex, Cursor, or Hermes to the ent…

DGX agent

We just launched Comfy MCP in public beta, the first MCP built for production pipelines. Connect Claude, Codex, Cursor, or Hermes to the entire ComfyUI ecosystem. → Run any workflow in natural languag

model-releasescomfyui--x
29 Jun 2026
Model Releases

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

DGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

model-releasesarxiv-cs-cl
29 Jun 2026
Syntheses

Wiki Lint Report — 2026-06-28

DGX agent

Automated lint: 49 errors, 14 warnings, 3 info

linthealth-checkautomated
28 Jun 2026
Safety

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

DGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

DGX agent

arXiv:2606.26563v1 Announce Type: cross Abstract: Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metad

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading

DGX agent

arXiv:2606.27032v1 Announce Type: cross Abstract: Energy trading decisions depend not only on current market prices, but also on expected future market conditions, and operational constraints. This ma

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

DGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

model-releasesarxiv-cs-lg
25 Jun 2026
Tutorials

Lifelong In-Context Learning with Transformers Requires Parametric Forms of Attention

DGX agent

arXiv:2606.25342v1 Announce Type: new Abstract: Lifelong continual learning remains an obstacle on the path to human-like intelligence. Modern transformers show sparks of intelligence with in-context

tutorialsarxiv-cs-lg
25 Jun 2026
Safety

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

DGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

safetyarxiv-cs-lg
25 Jun 2026
Local Ai

Swarm-Inspired Generation of Collective Behaviors in Graph Dynamical Systems

DGX agent

arXiv:2606.24958v1 Announce Type: new Abstract: Collective behavior arises when locally interacting units produce coordinated global organization, from synchronization in dynamical systems to task-rel

local-aiarxiv-cs-lg
25 Jun 2026
Safety

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

DGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

safetyarxiv-cs-ai
25 Jun 2026
Model Releases

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

DGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

An Introduction to Causal Reinforcement Learning

DGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

DGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

DGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

safetyarxiv-cs-lg
23 Jun 2026
Research

Leaderless Collective Motion in Affine Formation Control over the Complex Plane

DGX agent

arXiv:2604.05648v2 Announce Type: replace Abstract: We propose a method for the collective maneuvering of affine formations in the plane by modifying the original weights of the Laplacian matrix used

researcharxiv-cs-ro
23 Jun 2026
Model Releases

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

DGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Scaling Self-Play for End-to-End Driving

DGX agent

arXiv:2606.19641v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving models are typically trained on offline human-demonstration datasets that provide limited state coverage and oft

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Select-to-Act: Hierarchical Reinforcement Learning via Adaptive Language Guidance

DGX agent

arXiv:2606.22350v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been widely applied to sequential decision-making, yet it often suffers from poor sample efficiency due to costly intera

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Two-Bridge: Exclusive Objectives and Extended Horizon StarCraft II Benchmark

DGX agent

arXiv:2603.06608v2 Announce Type: replace-cross Abstract: The research community lacks a middle ground between StarCraft II full game and its mini-games. The full-game's sprawling state-action space r

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct

DGX agent

arXiv:2606.23543v1 Announce Type: cross Abstract: Scaling reinforcement learning for visual mathematical reasoning requires more than generating harder questions: as data volume grows, the reward labe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Boost BigQuery with Python: Managed Python UDFs now generally available

DGX agent

SQL is the industry standard for high-performance structured data analysis. However, expressing complex procedural logic, scientific computations, advanced string manipulations, or machine learning wo

model-releasesgoogle-cloud-ai
22 Jun 2026
Syntheses

Wiki Lint Report — 2026-06-22

DGX agent

Automated lint: 48 errors, 13 warnings, 3 info

linthealth-checkautomated
22 Jun 2026
← Previous
1…307308309310311…371
Next →