AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Distributed Coordination for Resilient Multi-UAV Remote Sensing: A Photovoltaic Inspection Case Study

DGX agent

arXiv:2607.24482v1 Announce Type: new Abstract: Deploying multiple UAVs for remote sensing enables proportional reductions in mission time, but realizing these benefits requires the fleet to coordinat

agentsarxiv-cs-ro
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Do LLM Debates Repeat Arguments Differently Across Languages?

DGX agent

arXiv:2607.23442v1 Announce Type: new Abstract: LLM debate is usually evaluated by final answers, but transcripts also reveal whether later turns develop new argumentative content or return to earlier

safetyarxiv-cs-cl
28 Jul 2026
Agents

HydroAgent: Formalizing Forecaster Expertise into Skill-Orchestrated Flood Forecasting Workflows

DGX agent

arXiv:2607.23983v1 Announce Type: cross Abstract: Operational flood forecasting depends on tacit forecaster expertise that is difficult to formalize, audit, and transfer. Although artificial intellige

agentsarxiv-cs-lg
28 Jul 2026
Model Releases

PeopleSearchBench: A Multi-Dimensional Benchmark for Evaluating AI-Powered People Search Platforms

DGX agent

arXiv:2603.27476v2 Announce Type: replace Abstract: AI-powered people search platforms are increasingly used in recruiting, sales prospecting, and professional networking, yet no widely accepted bench

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

Schema-Aware Localisation (SAL): Live Schema Grounding and Hallucination Validation for Oracle NL2SQL

DGX agent

arXiv:2607.22572v1 Announce Type: new Abstract: Large language models can generate fluent SQL from natural language, but on real enterprise Oracle databases they frequently fail at execution time: col

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

DGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Active few-shot segmentation by reinforcing data selection

DGX agent

arXiv:2607.22371v1 Announce Type: new Abstract: Few-shot learning enables medical image segmentation models to adapt to new tasks using only a small number of labelled examples. However, adaptation pe

agentsarxiv-cs-cv
27 Jul 2026
Agents

Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration

DGX agent

arXiv:2607.22153v1 Announce Type: cross Abstract: Industrial health management increasingly relies on heterogeneous information sources, including condition monitoring systems, supervisory control and

agentsarxiv-cs-lg
27 Jul 2026
Tutorials

Variance-Reduced Q-Learning over Static and Time-Varying Networks

DGX agent

arXiv:2607.21876v1 Announce Type: new Abstract: We investigate a decentralized reinforcement learning problem involving multiple agents that interact with the same Markov Decision Process (MDP). The a

tutorialsarxiv-cs-lg
27 Jul 2026
Safety

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space

DGX agent

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

DGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning

DGX agent

arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endoge

agentsarxiv-cs-lg
24 Jul 2026
Agents

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

DGX agent

arXiv:2606.14502v2 Announce Type: replace Abstract: Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reason

agentsarxiv-cs-ai
24 Jul 2026
Agents

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

DGX agent

arXiv:2607.20430v1 Announce Type: cross Abstract: We present LLM-INSTRUCT, the winning system for the UZH Shared Task at ArgMining 2026 on paragraph-level argument mining in UN and UNESCO resolutions.

agentsarxiv-cs-ai
24 Jul 2026
Agents

SevDiff: Severity-Conditioned Diffusion for Long-Tail Conflict Trajectory Generation

DGX agent

arXiv:2607.20549v1 Announce Type: new Abstract: Trajectory datasets used in ADAS evaluation are heavily biased toward routine driving; genuine vehicle-to-vehicle conflict events are rare, and the rare

agentsarxiv-cs-lg
24 Jul 2026
Model Releases

Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text

DGX agent

arXiv:2607.21072v1 Announce Type: new Abstract: Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks a

model-releasesarxiv-cs-cv
24 Jul 2026
Agents

Traceable Scholarship: Page Anchors and Ariadne's Thread for Humanistic Inquiry in the Age of Generative AI

DGX agent

arXiv:2607.20916v1 Announce Type: new Abstract: Generative AI lets large language models produce scholarly-looking text within seconds, yet fluency does not equal valid explanation. The deepest risk i

agentsarxiv-cs-ai
24 Jul 2026
Agents

AutoVSR: Automatic Visual-to-Symbolic Reasoning for Symbolic Expression Generation from Circuit Schematic

DGX agent

arXiv:2607.11338v2 Announce Type: replace Abstract: Symbolic expressions can effectively characterize and predict circuit behavior, but deriving them directly from circuit schematics is challenging. T

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization

DGX agent

arXiv:2607.19653v1 Announce Type: cross Abstract: Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature im

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning

DGX agent

arXiv:2607.19400v1 Announce Type: new Abstract: Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data,

agentsarxiv-cs-lg
23 Jul 2026
Model Releases

Analogical Deep Research: Retrieving and Integrating Historical Analogies for Foresight Analysis

DGX agent

arXiv:2607.13602v1 Announce Type: cross Abstract: Systematic comparisons between current situations and structurally similar past events in the historical, i.e., historical analogies, is among the mos

model-releasesarxiv-cs-lg
16 Jul 2026
Agents

Autonomous UAV Route Planning for Coverage Maximization in Environmental Monitoring: A Systematic Literature Review

DGX agent

arXiv:2607.13054v1 Announce Type: cross Abstract: Environmental monitoring with unmanned aerial vehicles (UAVs) requires route planning methods that maximize covered area while handling energy limits,

agentsarxiv-cs-ai
16 Jul 2026
Local Ai

HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Large Language Models

DGX agent

arXiv:2607.13072v1 Announce Type: cross Abstract: Zero-shot object-goal navigation aims to enable an intelligent agent to explore and navigate to objects of unknown categories in an unfamiliar environ

local-aiarxiv-cs-ai
16 Jul 2026
Agents

Lyapunov Exponent as Physics-Informed Dense Reward: RL Discovery of Stabilization Beyond the Kapitza Pendulum

DGX agent

arXiv:2607.14001v1 Announce Type: new Abstract: We suggest using the Lyapunov characteristic exponent (LCE) as a dense reward signal for the reinforcement learning problem of stabilizing the inverted

agentsarxiv-cs-lg
16 Jul 2026
Model Releases

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

DGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation

DGX agent

arXiv:2607.14006v1 Announce Type: cross Abstract: Penetration testing traditionally evaluates whether adversaries can exploit weaknesses in software, infrastructure, configurations, or operational con

agentsarxiv-cs-ai
16 Jul 2026
Agents

Unleashing Multimodal Large Language Models for Training-free HOI Detection in the Wild

DGX agent

arXiv:2607.13881v1 Announce Type: cross Abstract: Human-object interaction detection (HOID) has traditionally been formulated as a supervised detection problem over predefined interaction categories.

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

DGX agent

arXiv:2607.12252v1 Announce Type: new Abstract: Deep research agents are increasingly used to produce long-form financial reports, yet large-scale evaluation remains bottlenecked by the need for human

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement

DGX agent

arXiv:2606.19387v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success in software development. However, they are susceptible to hallucinations, meanin

agentsarxiv-cs-ai
15 Jul 2026
Agents

Practical Judgment, Virtue, and Intuition in the Use of Opaque AI-Enabled Systems

DGX agent

arXiv:2607.12755v1 Announce Type: cross Abstract: AI-enabled systems are seeing increasing deployment across numerous domains, with many being 'black boxes' with respect to core functions and capabili

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

Self in Space: Benchmarking Self-Awareness and Spatial Cognition in UAV Embodied Intelligence

DGX agent

arXiv:2607.12477v1 Announce Type: new Abstract: Autonomous UAV systems increasingly rely on multimodal large language models (MLLMs) to operate in complex real-world environments. Such embodied scenar

model-releasesarxiv-cs-cv
15 Jul 2026
Agents

Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation

DGX agent

arXiv:2607.10744v2 Announce Type: replace Abstract: Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models

agentsarxiv-cs-cv
15 Jul 2026
Agents

ArtMine: Discovering and Formalizing Artistic Processes

DGX agent

arXiv:2607.08331v1 Announce Type: cross Abstract: Understanding how artworks are created requires reasoning about the iterative decisions, material operations, and contextual influences that shape art

agentsarxiv-cs-ai
10 Jul 2026
Agents

LEEVLA: Seeing What Matters in Latent Environment Evolution for Vision-Language-Action

DGX agent

arXiv:2607.08182v1 Announce Type: cross Abstract: Vision-language-action (VLA) models aim to map multimodal inputs to robot actions. However, most existing approaches struggle to cover complex dynamic

agentsarxiv-cs-ai
10 Jul 2026
Agents

ProjAgent: Procedural Similarity Retrieval for Repository-Level Code Generation

DGX agent

arXiv:2607.08691v1 Announce Type: cross Abstract: Repository-level code generation requires implementing target functions while accounting for complex cross-file dependencies and project-specific conv

agentsarxiv-cs-ai
10 Jul 2026
Agents

Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles

DGX agent

arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from norm

agentsarxiv-cs-ai
10 Jul 2026
Agents

When Does Continual Learning Require Learning

DGX agent

arXiv:2607.07847v1 Announce Type: new Abstract: As large language models (LLMs) become increasingly capable, the next question is how can we enable models to continually learn? Today, the field largel

agentsarxiv-cs-lg
10 Jul 2026
Agents

Gimitest: A Comprehensive Tool for Testing Reinforcement Learning Policies

DGX agent

arXiv:2607.07029v1 Announce Type: cross Abstract: Reinforcement learning (RL) policies can be unsafe and vulnerable to attacks. Ensuring their reliability is often a pain point as existing automated t

agentsarxiv-cs-ai
9 Jul 2026
Agents

Measuring Intelligence Beyond Human Scale

DGX agent

arXiv:2607.07040v1 Announce Type: new Abstract: How can we measure intelligence beyond human capability? Human-authored benchmarks saturate, and above human capability, examiners may not know which ta

agentsarxiv-cs-ai
9 Jul 2026
Agents

Neutral Substrates: A Design Constraint for Shared Records Under Persistent Interpretive Disagreement

DGX agent

arXiv:2601.14271v2 Announce Type: replace Abstract: Shared accountability records are often used by parties who may never agree about causation, responsibility, or normative interpretation. For such r

agentsarxiv-cs-ai
9 Jul 2026
Agents

EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems

DGX agent

arXiv:2607.05638v1 Announce Type: cross Abstract: Teams deploying large language models in business contexts need evaluation systems, yet most treat evaluation as static model selection: run benchmark

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Evaluating calibrated refusal and safe usefulness in dual-use biology settings

DGX agent

arXiv:2607.05462v1 Announce Type: cross Abstract: As AI agents are incorporated into life science workflows, the capabilities that speed discovery might also enable misuse. We present BioSecBench-Refu

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Faithful or Findable? Evaluating LLM-Generated Metadata for RDF Dataset Search

DGX agent

arXiv:2607.05970v1 Announce Type: cross Abstract: Dataset search depends heavily on metadata, making LLM-generated metadata a consequential form of synthetic content in retrieval systems. We study six

agentsarxiv-cs-ai
8 Jul 2026
Safety

From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structured Action Space

DGX agent

arXiv:2607.05794v1 Announce Type: new Abstract: Long-term user memory is essential for personalized conversational agents, yet many memory systems still expose memory through passive retrieval interfa

safetyarxiv-cs-ai
8 Jul 2026
Agents

Learning The Minimum Action Distance

DGX agent

arXiv:2506.09276v4 Announce Type: replace-cross Abstract: This paper presents a state representation framework for Markov decision processes (MDPs) that can be learned solely from state trajectories,

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Quantifying Frontier LLM Capabilities for Container Sandbox Escape

DGX agent

arXiv:2603.02277v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly act as autonomous agents, using tools to execute code, read and write files, and access networks, cr

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models

DGX agent

arXiv:2512.18542v3 Announce Type: replace-cross Abstract: AI coding assistants produce vulnerable code in 45% of security-relevant scenarios~ite{veracode2025}, yet no public training dataset teaches b

agentsarxiv-cs-ai
8 Jul 2026
Agents

VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery

DGX agent

arXiv:2607.06374v1 Announce Type: new Abstract: Vision-language models (VLMs) have made interactive digital museums increasingly feasible by connecting 3D digitization with natural-language artifact e

agentsarxiv-cs-cv
8 Jul 2026
← Previous
1…141142143144145…236
Next →