AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

DGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

model-releasesarxiv-cs-ai
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Adversarial Social Epistemology for Assemblies of Humans and Large Language Models

DGX agent

arXiv:2607.07760v1 Announce Type: new Abstract: We outline an adversarial social epistemology (ASE) for densely interactive communicative landscapes in which public assertions are scaffolded by chains

researcharxiv-cs-ai
10 Jul 2026
Research

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier

DGX agent

arXiv:2607.07779v1 Announce Type: cross Abstract: Recent developments in AI for Mathematics (AI4Math), especially Large Language Model (LLM)-driven theorem provers, has achieved remarkable success in

researcharxiv-cs-ai
10 Jul 2026
Safety

MPFlow: Learning Budgeted Max-Flow Optimization on the Lightning Network with Deep Graph Reinforcement Learning

DGX agent

arXiv:2607.08703v1 Announce Type: new Abstract: We address liquidity placement in the Bitcoin Lightning Network (LN): given a fixed budget, which channels should a node open to maximize its routing ca

safetyarxiv-cs-lg
10 Jul 2026
Model Releases

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

DGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

DGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

DGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

ROAD-Waymo: A Large-Scale Action Awareness Dataset for Autonomous Driving

DGX agent

arXiv:2411.01683v3 Announce Type: replace Abstract: Autonomous Vehicle (AV) perception systems require more than simply seeing, via e.g., object detection or scene segmentation. They need a holistic u

model-releasesarxiv-cs-cv
9 Jul 2026
Safety

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment

DGX agent

arXiv:2607.06522v1 Announce Type: new Abstract: Vision-language models (VLMs) struggle to generalize in interactive physical reasoning, particularly under unseen tasks and environments. Two key failur

safetyarxiv-cs-ai
8 Jul 2026
Safety

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization

DGX agent

arXiv:2607.06563v1 Announce Type: new Abstract: Traditional data physicalization is often static and disconnected from real environments, limiting its ability to convey embodied spatial dynamics and e

safetyarxiv-cs-ro
8 Jul 2026
Model Releases

Federated Physics-Grounded Reinforcement Learning for Distributed Stability Control in Smart Grids

DGX agent

arXiv:2607.05553v1 Announce Type: new Abstract: Transient stability control in smart grids requires rapid post-fault damping of generator frequency and rotor angle deviations to prevent cascading fail

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction

DGX agent

arXiv:2607.05577v1 Announce Type: new Abstract: Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Think Before You Grid-Search: Floor-First Triage for LLM Serving

DGX agent

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning

DGX agent

arXiv:2606.16933v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) systems often degrade when operating conditions differ from those previously encountered, reflecting distributiona

safetyarxiv-cs-ai
7 Jul 2026
Safety

Adaptive Inference Batching using Policy Gradients

DGX agent

arXiv:2607.05272v1 Announce Type: cross Abstract: Inference serving systems must balance throughput and latency under bursty, heterogeneous workloads, yet the industry standard remains static batching

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

ClassicLogic: A Knowledge-Driven Benchmark of Classic Puzzle Games for Evaluating Compositional Generalization

DGX agent

arXiv:2607.05185v1 Announce Type: new Abstract: Compositional generalization, the ability to understand and produce novel combinations of known components, remains a fundamental challenge for modern a

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Harness-Aware Self-Evolving: Co-Evolving Model Weights, Harness, and Task Solutions

DGX agent

arXiv:2607.03935v1 Announce Type: new Abstract: Self-evolving frameworks usually optimize task solutions while treating the surrounding harness as fixed. We introduce Harness-Aware Self-Evolving (HASE

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

DGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

applicationsarxiv-cs-cl
7 Jul 2026
Model Releases

How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation

DGX agent

arXiv:2605.06605v2 Announce Type: replace Abstract: Evaluating and predicting the performance of large language models (LLMs) in multi-turn conversational settings is critical yet computationally expe

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

LLM-as-a-Verifier: A General-Purpose Verification Framework

DGX agent

arXiv:2607.05391v1 Announce Type: new Abstract: Scaling pre-training, post-training, and test-time compute have become the central paradigms for improving the capabilities of LLMs. In this work, we id

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Multiplayer Interactive World Models with Representation Autoencoders

DGX agent

arXiv:2607.05352v1 Announce Type: cross Abstract: We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interactions. Whereas single-player world

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Probably Correct Optimal Stable Matching under Two-Sided Uncertainty

DGX agent

arXiv:2607.04824v1 Announce Type: new Abstract: We study a sequential learning problem for stable matchings in two-sided markets where preferences on both sides are initially unknown. We focus on a ce

researcharxiv-cs-lg
7 Jul 2026
Safety

Teaming Up with AI: Coordination and Cooperation

DGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Composite Reward Design in PPO-Driven Adaptive Filtering

DGX agent

arXiv:2506.06323v2 Announce Type: replace-cross Abstract: Model-free and reinforcement learning-based adaptive filtering methods are gaining traction for denoising in dynamic, non-stationary environme

model-releasesarxiv-cs-lg
3 Jul 2026
Safety

Mean Field Reinforcement Learning

DGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

safetyarxiv-cs-lg
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

DGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

model-releasesarxiv-cs-ai
2 Jul 2026
Safety

AI Native Games: A Survey and Roadmap

DGX agent

arXiv:2607.00527v1 Announce Type: new Abstract: Generative AI now enables games to produce dialogue, quests, characters, images, and worlds at runtime. Yet generation alone does not make a game AI-nat

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

DGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution

DGX agent

arXiv:2607.00700v1 Announce Type: cross Abstract: LLVM is a widely used compiler infrastructure whose scale and complexity make issue resolution labor-intensive and challenging. Although large languag

model-releasesarxiv-cs-ai
2 Jul 2026
Local Ai

Local Motion Matters: A Deconstruct-Recompose Paradigm for Reinforcement Learning Pre-training from Videos

DGX agent

arXiv:2607.00808v1 Announce Type: new Abstract: Pre-training on large-scale videos to improve reinforcement learning efficiency is promising yet remains challenging. Existing methods typically treat t

local-aiarxiv-cs-lg
2 Jul 2026
Model Releases

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

DGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

model-releasesarxiv-cs-cv
2 Jul 2026
Safety

VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement

DGX agent

arXiv:2607.00446v1 Announce Type: cross Abstract: As video corpora continue to expand in both scale and task complexity, there is increasing demand for approaches that retrieve relevant videos from la

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

DGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

safetyarxiv-cs-ai
1 Jul 2026
Local Ai

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

DGX agent

arXiv:2606.31036v1 Announce Type: new Abstract: Specialist epilepsy expertise is scarce in resource-constrained settings, making LLM-based decision support attractive for frontline clinicians managing

local-aiarxiv-cs-lg
1 Jul 2026
Model Releases

The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

DGX agent

arXiv:2606.31273v1 Announce Type: new Abstract: AI-assisted research has entered a stage in which the central question is not only whether systems can generate hypotheses, run experiments, or produce

model-releasesarxiv-cs-lg
1 Jul 2026
Tutorials

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

DGX agent

arXiv:2606.31946v1 Announce Type: new Abstract: The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampli

tutorialsarxiv-cs-cv
1 Jul 2026
Model Releases

Cognitive World Models for Process-Level Social Influence Evaluation

DGX agent

arXiv:2606.29495v1 Announce Type: new Abstract: Social influence dialogue changes user behavior by altering internal cognitive states. The central evaluation question is whether the user's beliefs, de

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

DGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents

DGX agent

arXiv:2606.29399v1 Announce Type: new Abstract: Reviewing nuclear regulatory documents requires multi-hop reasoning across tens of thousands of pages, where judgments depend on evidence assembled acro

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

DGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Sequential Planning via Anchored Robotic Keypoints

DGX agent

arXiv:2606.30613v1 Announce Type: new Abstract: We present Sequential Planning via Anchored Robotic Keypoints, SPARK, a training-free neurosymbolic manipulation system that reaches 43.7% on six LIBERO

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Safety

ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation

DGX agent

arXiv:2606.27736v1 Announce Type: new Abstract: The rapid spread of fake news poses increasing threats to information ecosystems, especially as AI-generated misinformation under Generative Engine Opti

safetyarxiv-cs-ai
29 Jun 2026
Local Ai

Verifiable Geometry Problem Solving: Solver-Driven Autoformalization and Theorem Proposing

DGX agent

arXiv:2606.27926v1 Announce Type: new Abstract: Geometry Problem Solving have increasingly adopt the neuro-symbolic paradigm, combining neural intuition with symbolic rigor. However, current framework

local-aiarxiv-cs-ai
29 Jun 2026
Model Releases

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

DGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

model-releasesarxiv-cs-cl
29 Jun 2026
Safety

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

DGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

DGX agent

arXiv:2606.26563v1 Announce Type: cross Abstract: Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metad

model-releasesarxiv-cs-ai
26 Jun 2026
← Previous
1…189190191192193…233
Next →