AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

DGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

model-releasesarxiv-cs-ai
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

DGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents

DGX agent

arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

DGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web

DGX agent

arXiv:2604.08516v1 Announce Type: new Abstract: Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with t

agentsarxiv-cs-cv
10 Apr 2026
Hardware

OpenPRC: A Unified Open-Source Framework for Physics-to-Task Evaluation in Physical Reservoir Computing

DGX agent

arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi

hardwarearxiv-cs-ro
10 Apr 2026
Research

Syntax Is Easy, Semantics Is Hard: Evaluating LLMs for LTL Translation

DGX agent

arXiv:2604.07321v1 Announce Type: cross Abstract: Propositional Linear Temporal Logic (LTL) is a popular formalism for specifying desirable requirements and security and privacy policies for software,

researcharxiv-cs-ai
10 Apr 2026
Research

Adaptive Online Learning with LSTM Networks for Energy Price Prediction

DGX agent

arXiv:2510.16898v2 Announce Type: replace-cross Abstract: Accurate prediction of electricity prices is crucial for stakeholders in the energy market, particularly for grid operators, energy producers,

researcharxiv-cs-ai
13 Aug 2026
Model Releases

ADEPT: A Unified Framework for Deep Learning Test Adequacy

DGX agent

arXiv:2608.12144v1 Announce Type: cross Abstract: Over the past decade, many test adequacy metrics have been proposed for deep learning that characterize test dataset adequacy from different perspecti

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Advancing MLLM-based UAV Image Understanding and Reasoning: A Benchmark and a Training-Free Multi-Agent System

DGX agent

arXiv:2608.11738v1 Announce Type: cross Abstract: Multimodal Large Language Model (MLLM)-based UAV aerial image understanding and reasoning is essential for aerial intelligence yet poses distinct chal

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Backdoor Decontamination Dynamics in LLM Agents

DGX agent

arXiv:2608.11295v1 Announce Type: cross Abstract: Open-weight LLM agents are vulnerable to backdoors installed during fine-tuning, which may be undetectable if the trigger conditions are never met dur

researcharxiv-cs-ai
13 Aug 2026
Agents

Beyond Memory: A Transactional Continuity Kernel for Long-Lived AI Agents

DGX agent

arXiv:2608.11632v1 Announce Type: cross Abstract: Persistent AI agents accumulate versioned state across long horizons, but storage retention alone does not identify authoritative state. Without an ex

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence

DGX agent

arXiv:2608.12290v1 Announce Type: cross Abstract: Modern black-box Image-to-Video (I2V) models offer powerful capabilities in automated content creation, yet their lack of fine-grained control and rel

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Certifying What Helps Customer-Return Timing: A Screen-and-Confirm Test for Conditioning Signals, and Why Decay Is Nearly Enough

DGX agent

arXiv:2608.11555v1 Announce Type: new Abstract: Practitioners enrich customer-return models with ever more signals (lifetime value, category, recency/frequency, calendar, geography), and the temporal-

researcharxiv-cs-lg
13 Aug 2026
Research

Chemically Meaningful Textualization Enables Explainable Validation of Metal-Organic Frameworks by Large Language Models

DGX agent

arXiv:2608.11283v1 Announce Type: cross Abstract: Computation-ready metal-organic framework (MOF) databases are essential for high-throughput screening, yet many reported crystal structures remain che

researcharxiv-cs-ai
13 Aug 2026
Safety

Co-constructing sociotechnical AI governance: participatory system mapping using algorithm registers

DGX agent

arXiv:2608.12166v1 Announce Type: cross Abstract: Algorithm registers have been championed as a means of providing transparency on the use of algorithms in public services. Yet potential publics diffe

safetyarxiv-cs-ai
13 Aug 2026
Safety

Constructing Dynamic Master Logic Models as Knowledge Graphs for Complex System Diagnostics Using Retrieval-Augmented Large Language Models

DGX agent

arXiv:2608.12304v1 Announce Type: new Abstract: Dynamic Master Logic (DML) provides a hierarchical framework for representing system behavior by linking functional objectives to underlying structural

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Convergent Detour Hijacking: Task-Preserving Resource Amplification in Skill-Based LLM Agents

DGX agent

arXiv:2608.12273v1 Announce Type: cross Abstract: LLM agents increasingly rely on third-party skills, using natural-language descriptions for selection and instruction bodies for planning. This progre

model-releasesarxiv-cs-ai
13 Aug 2026
Local Ai

Deep Activity Model: A Generative Approach for Human Mobility Pattern Synthesis

DGX agent

arXiv:2405.17468v3 Announce Type: replace-cross Abstract: Human mobility plays a crucial role in transportation, urban planning, and public health, but current approaches face important limitations. E

local-aiarxiv-cs-ai
13 Aug 2026
Model Releases

Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration

DGX agent

arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter. Despite decades of methodological work, researchers

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

DGX agent

arXiv:2507.14267v2 Announce Type: replace Abstract: Large language model (LLM) agents can execute long-horizon scientific workflows, but their numerical outputs are difficult to trust: agents lose con

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

DGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs

DGX agent

arXiv:2608.11674v1 Announce Type: cross Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities,

safetyarxiv-cs-ai
13 Aug 2026
Agents

Harness-IF: Evaluating Instruction Following Across Instruction Surfaces in Coding Agents

DGX agent

arXiv:2608.11727v1 Announce Type: new Abstract: When a coding agent obeys a rule, it may simply have been going to do that anyway. Existing instruction-following benchmarks cannot tell the difference:

agentsarxiv-cs-ai
13 Aug 2026
Agents

Local verification cannot detect non-transportability: a cohomological theory of context preservation in agentic reasoning

DGX agent

arXiv:2608.11252v1 Announce Type: new Abstract: Agentic AI systems routinely transport conclusions across biological, clinical and financial contexts, and the emerging safeguard is local verification:

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

OEIS Open: How many conjectures can language models turn into theorems?

DGX agent

arXiv:2608.11941v1 Announce Type: new Abstract: We construct OEIS Open, a benchmark based on 492 open mathematical conjectures from the OEIS, formalized in Lean by Tsoukalas et al. Whereas these conje

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle

DGX agent

arXiv:2608.11241v1 Announce Type: new Abstract: Deploying LLM agents into industrial recommender operations exposes a three-way tension we frame as the autonomy-determinism-efficiency trilemma: genera

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Remote Sensing and Machine Learning-Based Analysis of Land Use and Vegetation Change in Dhaka District, Bangladesh

DGX agent

arXiv:2608.12001v1 Announce Type: cross Abstract: Rapid urbanization in Dhaka District, Bangladesh has triggered substantial alterations in land use and environmental conditions, necessitating systema

researcharxiv-cs-ai
13 Aug 2026
Research

Small-Scale Experiments: Are We There Yet?

DGX agent

arXiv:2608.11859v1 Announce Type: new Abstract: Scaling laws promised cost-effective experiments; six years later, they have yet to fully deliver. Instead, researchers have found them unreliable at sm

researcharxiv-cs-lg
13 Aug 2026
Model Releases

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy

DGX agent

arXiv:2607.11175v2 Announce Type: replace Abstract: The growing ability of large language models and vision-language models to jointly interpret and reason over images and text is reshaping medical im

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Tight Nonasymptotic Local Convergence of Sinkhorn-Knopp

DGX agent

arXiv:2608.11760v1 Announce Type: cross Abstract: We revisit the Sinkhorn-Knopp (SK) algorithm for the matrix scaling problem. Despite extensive literature on the global convergence of SK and its vari

researcharxiv-cs-lg
13 Aug 2026
Model Releases

Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique

DGX agent

arXiv:2608.10430v1 Announce Type: cross Abstract: Large Language Models (LLMs) deployed as AI agents frequently exhibit user specification-grounding failures, executing hallucinated, undesired actions

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

Automating and Scaling Behavioral Scientific Research on AI Agents

DGX agent

arXiv:2608.10030v1 Announce Type: new Abstract: As AI agents are increasingly deployed in complex environments, understanding their behaviors becomes critical. Yet behavioral scientific research on AI

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

DGX agent

arXiv:2608.10915v1 Announce Type: new Abstract: After an older adult misses a medication dose, a software agent can send another reminder and an embodied agent can bring the medication. Yet neither ex

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Comprendia: AI-Augmented Code Comprehension

DGX agent

arXiv:2608.10290v1 Announce Type: cross Abstract: Comprendia is an Eclipse plugin that integrates structural dependency visualization with LLM-powered code explanation on a shared interactive graph fo

researcharxiv-cs-ai
12 Aug 2026
Local Ai

Conversational Orchestration for Organic 6G

DGX agent

arXiv:2608.10714v1 Announce Type: cross Abstract: The Organic 6G vision of a network of networks spanning an edge-cloud continuum complemented by non-terrestrial resources requires, to realize its pro

local-aiarxiv-cs-ai
12 Aug 2026
Agents

DuplexWorld: Can voice agents help you get through the day?

DGX agent

arXiv:2608.10716v1 Announce Type: cross Abstract: Speech-to-speech (S2S) voice agents are increasingly being incorporated into enterprise for customer care and as daily companions for consumers owing

agentsarxiv-cs-ai
12 Aug 2026
Research

Exploring Semantic Stability Across Reviews in the Linux Kernel

DGX agent

arXiv:2608.10101v1 Announce Type: cross Abstract: Code review is credited with substantially changing a patch's code between its first submission and the version that eventually lands. However, prior

researcharxiv-cs-ai
12 Aug 2026
Model Releases

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes

DGX agent

arXiv:2608.10886v1 Announce Type: new Abstract: Robots operating in human environments need memories that capture not only what objects exist and where, but also how people use them over time and how

model-releasesarxiv-cs-cv
12 Aug 2026
Applications

Graphical Models of False Information and Fact Checking Ecosystems

DGX agent

arXiv:2208.11582v2 Announce Type: replace-cross Abstract: The wide spread of false information online, including misinformation and disinformation, has become a major problem for our highly digitised

applicationsarxiv-cs-ai
12 Aug 2026
Model Releases

HoosierHelp: Benchmarking LLM Agents for Social Service Navigation

DGX agent

arXiv:2608.09946v1 Announce Type: cross Abstract: Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM ag

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Lipschitz Dueling Bandits over Continuous Action Spaces

DGX agent

arXiv:2604.00523v2 Announce Type: replace Abstract: We study for the first time, stochastic dueling bandits over continuous action spaces with Lipschitz structure, where feedback is purely comparative

researcharxiv-cs-lg
12 Aug 2026
Agents

MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph

DGX agent

arXiv:2608.10504v1 Announce Type: new Abstract: As coding agents increasingly handle implementation, the central challenge shifts from building individual agents to building an infrastructure that sys

agentsarxiv-cs-ai
12 Aug 2026
Safety

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

DGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

safetyarxiv-cs-lg
12 Aug 2026
Research

Modelling Geographic Atrophy Progression using Implicit Neural Representations

DGX agent

arXiv:2608.10807v1 Announce Type: cross Abstract: Age-related Macular Degeneration (AMD) is the major cause of blindness in the Western world. Its late dry phase is characterised by irreversible atrop

researcharxiv-cs-ai
12 Aug 2026
Research

Narrative Keyframing for Generative Creative Writing

DGX agent

arXiv:2608.10337v1 Announce Type: cross Abstract: We introduce narrative keyframing, an interaction technique for AI-assisted creative writing that lets writers specify different types of narrative co

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents

DGX agent

arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages update dynamically or hide important structure be

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

On Solomonoff Induction in Large Language Models and the Limits of Self-Improving: The Singularity Is Not Near Without Symbolic Model Synthesis

DGX agent

arXiv:2601.05280v3 Announce Type: replace-cross Abstract: On the one hand, the question of whether large language models (LLMs) are Solomonoff induction estimators has become an explicit question at t

model-releasesarxiv-cs-ai
12 Aug 2026
← Previous
1…5354555657…109
Next →