AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation

DGX agent

arXiv:2607.25881v1 Announce Type: new Abstract: We investigate how well large language models (LLMs) can assist scientific project planning and proposal evaluation. One-page project plans were indepen

model-releasesarxiv-cs-cl
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings

DGX agent

arXiv:2607.24814v1 Announce Type: new Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where physician-to-patient ratios can fall below 1:25,000 in

model-releasesarxiv-cs-ai
29 Jul 2026
Tutorials

Balanced Soft mixture-of-expert model for Glaucoma Detection

DGX agent

arXiv:2607.25324v1 Announce Type: cross Abstract: Glaucoma is a group of eye diseases that damage the optic nerve, often caused by elevated intraocular pressure. It is a leading cause of irreversible

tutorialsarxiv-cs-ai
29 Jul 2026
Model Releases

COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution

DGX agent

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify no

model-releasesarxiv-cs-ai
29 Jul 2026
Research

DeepVRegulome: DNABERT-based deep-learning framework for predicting the functional impact of short genomic variants on the human regulome

DGX agent

arXiv:2511.09026v2 Announce Type: replace-cross Abstract: Whole-genome sequencing (WGS) has revealed numerous non-coding short variants whose functional impacts remain poorly understood. Despite recen

researcharxiv-cs-ai
29 Jul 2026
Model Releases

Detecting Knowledge Inconsistencies Across Text, Tables, and Knowledge Graphs

DGX agent

arXiv:2607.25959v1 Announce Type: cross Abstract: Wikipedia and Wikidata are widely used for information access, LLM pre-training, and retrieval-augmented generation. Their knowledge is deeply connect

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

Distilling Temporal Search and Reasoning: Evolving LLMs for Future Prediction via Harness-Assisted Efficient Data Synthesis

DGX agent

arXiv:2607.25554v1 Announce Type: new Abstract: Future event prediction carries broad social impact yet remains challenging. SOTA approaches augment LLMs with external agent frameworks whose predictiv

agentsarxiv-cs-ai
29 Jul 2026
Agents

Distributing Security Controls Through Harness Engineering

DGX agent

arXiv:2607.25890v1 Announce Type: new Abstract: AI coding agents are being adopted at historic speed, yet security and risk concerns remain the primary barrier to scaling agentic AI across organizatio

agentsarxiv-cs-ai
29 Jul 2026
Tutorials

From Idea to Classroom in Days: Using 'Vibe Coding' to Create a Programming Process Visualizer from IDE Activity Logs

DGX agent

arXiv:2607.24757v1 Announce Type: cross Abstract: This paper reports on the rapid development and classroom deployment of a Thonny log visualizer built using AI-assisted ``vibe coding'' to make studen

tutorialsarxiv-cs-ai
29 Jul 2026
Agents

From Naive RAG to Deep Agentic Retrieval: An Evolving Context Engineering Pipeline for Regulatory Compliance

DGX agent

arXiv:2607.24791v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) is the dominant paradigm for applying large language models (LLMs) to enterprise document corpora, yet naive impl

agentsarxiv-cs-ai
29 Jul 2026
Model Releases

GraphRareBench: An Auditable Graph-Evidence Benchmark for Phenotype-Driven Rare-Disease Diagnosis

DGX agent

arXiv:2607.24878v1 Announce Type: cross Abstract: Phenotype-driven diagnostic benchmarks usually report the rank of the reference disease, but they rarely reveal which plausible alternatives are ranke

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM

DGX agent

arXiv:2508.05775v3 Announce Type: replace Abstract: Large Language Models (LLMs) have revolutionized content creation across digital platforms, offering unprecedented capabilities in natural language

safetyarxiv-cs-cl
29 Jul 2026
Model Releases

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following

DGX agent

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context,

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

DGX agent

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as mat

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

Matryoshka Agent: Unfolding Sub-Agents for Long-Horizon Machine Learning Engineering

DGX agent

arXiv:2607.25090v1 Announce Type: new Abstract: Machine learning engineering (MLE) tasks require long-horizon decision making over iterative solution debugging and refinement, under expensive and feed

agentsarxiv-cs-ai
29 Jul 2026
Safety

Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation

DGX agent

arXiv:2602.10508v2 Announce Type: replace Abstract: Modern segmentation models achieve strong predictive performance but remain largely opaque, limiting our ability to diagnose failures, understand da

safetyarxiv-cs-cv
29 Jul 2026
Research

Modular Robotic Catheters for Endovascular Aneurysm Repair

DGX agent

arXiv:2607.25807v1 Announce Type: new Abstract: Fenestrated/Branched endovascular aneurysm repair (FEVAR/BEVAR) require surgeons to navigate catheters and guidewires into various branches of the abdom

researcharxiv-cs-ro
29 Jul 2026
Research

Multiclass Classification without Labels via Posterior Simplex Geometry

DGX agent

arXiv:2607.24943v1 Announce Type: cross Abstract: In many classification problems, reliable instance-level labels are unavailable. However, it is often possible to construct weakly enriched unlabeled

researcharxiv-cs-ai
29 Jul 2026
Model Releases

On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora?

DGX agent

arXiv:2607.24784v1 Announce Type: new Abstract: Specialised translation relies on the use of documentary and terminological resources, including corpora. These resources are particularly useful for te

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation

DGX agent

arXiv:2607.25656v1 Announce Type: new Abstract: Complex tasks often decompose into parallelizable yet interdependent subtasks, making orchestration critical to the performance of multi-agent systems (

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task

DGX agent

arXiv:2607.24850v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning h

safetyarxiv-cs-lg
29 Jul 2026
Safety

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

DGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

safetyarxiv-cs-ai
29 Jul 2026
Agents

Towards an Agent Operating System - Lessons from Classical and Cloud OS

DGX agent

arXiv:2607.25076v1 Announce Type: new Abstract: Every major wave of platform software follows the same arc: an initial period of experimentation with competing frameworks and ad-hoc implementations, f

agentsarxiv-cs-ai
29 Jul 2026
Research

Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment

DGX agent

arXiv:2607.25393v1 Announce Type: new Abstract: Histopathological examination primarily relies on hematoxylin and eosin (H&E) and immunohistochemistry (IHC) staining. Although IHC provides critical mo

researcharxiv-cs-cv
29 Jul 2026
Safety

Tripody: An Overconstrained 3-SPR-like Parallel Robot for High-Reach Construction Tasks

DGX agent

arXiv:2607.25781v1 Announce Type: new Abstract: Many ceiling construction tasks still rely on heavy serial manipulators that are difficult to deploy in cluttered interiors, motivating lightweight, fie

safetyarxiv-cs-ro
29 Jul 2026
Applications

'We'll have to see how it works': An interview study to understand collaborative practices in interdisciplinary artificial intelligence and healthcare research

DGX agent

arXiv:2311.18424v3 Announce Type: replace-cross Abstract: Developing artificial intelligence (AI) algorithms for healthcare is a collaborative effort, bringing data scientists, clinicians, patients an

applicationsarxiv-cs-ai
29 Jul 2026
Safety

When Do Agent Loops Mistake Stagnation for Progress? Self-Evaluation Bias and Externally Grounded Verification in Long-Running Autonomous LLM Agent Loops

DGX agent

arXiv:2607.25152v1 Announce Type: new Abstract: Long-running autonomous agents plan, act, and judge their own completion without human intervention. When an agent grades its own work, self-evaluation

safetyarxiv-cs-ai
29 Jul 2026
Safety

When Thinking Before Retrieval Hurts: TraceBound Diagnostics for Adaptive Knowledge-Graph Retrieval

DGX agent

arXiv:2607.24800v1 Announce Type: cross Abstract: Adaptive retrieval promises to make knowledge-graph question answering more robust by letting a controller search, inspect neighborhoods, revise actio

safetyarxiv-cs-ai
29 Jul 2026
Safety

Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI: Lessons from Policing

DGX agent

arXiv:2607.25648v1 Announce Type: cross Abstract: Public services face growing pressure to adopt artificial intelligence (AI) to close the gap between rising demand and falling resources. That pressur

safetyarxiv-cs-ai
29 Jul 2026
Research

A Diagnostic Gap Framework for Evaluating Reconstruction Fidelity in Weakly Supervised Mammography

DGX agent

arXiv:2607.22740v1 Announce Type: new Abstract: Weakly supervised pipelines for medical imaging have become increasingly popular over the years. These systems often include multiple stages and compone

researcharxiv-cs-cv
28 Jul 2026
Model Releases

A Few Words Go a Long Way: Language Guided Robot Policy Synthesis

DGX agent

arXiv:2607.23784v1 Announce Type: cross Abstract: While vision-language-action models have demonstrated impressive zero-shot manipulation capabilities, they remain fundamentally black box policies tha

model-releasesarxiv-cs-ai
28 Jul 2026
Research

A New Kind of Adversarial Example: Measuring the Human-Model Gap, and Its Relationship to OOD Detection

DGX agent

arXiv:2607.22722v1 Announce Type: cross Abstract: Almost all adversarial attacks add an imperceptible perturbation to fool a model. We instead study the opposite: a large, clearly visible perturbation

researcharxiv-cs-ai
28 Jul 2026
Agents

ACM: Agentic Context Management for Long Horizon Tasks

DGX agent

arXiv:2607.23809v1 Announce Type: new Abstract: Agentic tasks are inherently long-horizon and multi-turn, constantly accumulating context through interactions with the environment. Existing context co

agentsarxiv-cs-ai
28 Jul 2026
Research

Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models

DGX agent

arXiv:2603.05147v2 Announce Type: replace Abstract: Current research on Vision-Language-Action (VLA) models predominantly focuses on enhancing generalization through reasoning techniques. While effect

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents

DGX agent

arXiv:2607.24625v1 Announce Type: cross Abstract: Autonomous LLM agents processing mixed-confidentiality data face severe security risks from prompt injection attacks and reasoning errors. While dynam

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Agentic Reward Modeling: Verifying GUI Agent via Progressive Trajectory-Grounded Interaction

DGX agent

arXiv:2602.00575v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) provides a promising pathway for continuously advancing GUI agents, yet existing reward modeli

agentsarxiv-cs-ro
28 Jul 2026
Tutorials

AI Strategy: How to Choose What AI Product to Implement

DGX agent

arXiv:2607.23733v1 Announce Type: cross Abstract: Firms struggle to choose AI projects that pay off: two projects can look equally promising to smart, motivated stakeholders and yet deserve opposite d

tutorialsarxiv-cs-ai
28 Jul 2026
Research

An adaptive multi-fuzzy logic model for diagnosing transformer faults using dynamic weight optimization

DGX agent

arXiv:2607.23486v1 Announce Type: new Abstract: Dissolved gas analysis (DGA) is crucial for diagnosing early power transformer failures. Traditional DGA interpretation methods like Duval Triangle, IEC

researcharxiv-cs-lg
28 Jul 2026
Safety

An Unofficial FastLAS Tutorial: A Programmer's Guide

DGX agent

arXiv:2607.23557v1 Announce Type: cross Abstract: FastLAS is a scalable system for Inductive Logic Programming (ILP): you give it some background knowledge, a language bias, and a set of examples, and

safetyarxiv-cs-ai
28 Jul 2026
Safety

ARdena: Scenario-driven control of real-time LLM agents

DGX agent

arXiv:2607.22651v1 Announce Type: new Abstract: Large language models (LLMs) have enabled increasingly capable conversational agents, but reliably controlling their behavior in real-time interactive e

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond Shapley: An Influence-Based Data Auditing Pipeline for LLM Alignment and Evaluation

DGX agent

arXiv:2607.22766v1 Announce Type: cross Abstract: The alignment of Large Language Models (LLMs) is increasingly bottlenecked by data quality. As datasets scale, massive preference and instruction-tuni

model-releasesarxiv-cs-ai
28 Jul 2026
Research

CausAdv: A Causal-based Framework for Detecting Adversarial Examples

DGX agent

arXiv:2411.00839v4 Announce Type: replace-cross Abstract: Deep learning has led to tremendous success in computer vision, largely due to Convolutional Neural Networks (CNNs). However, CNNs have been s

researcharxiv-cs-ai
28 Jul 2026
Local Ai

Co-Harness: Co-Evolving Harnesses and Model Weights for LLM Agents

DGX agent

arXiv:2607.22688v1 Announce Type: new Abstract: Post-training agents for automated AI research requires optimizing not only model parameters, but also the runtime harness that shapes how research traj

local-aiarxiv-cs-ai
28 Jul 2026
Safety

Concept-based Visual Counterfactual Explanations with Diffusion Models

DGX agent

arXiv:2607.22544v1 Announce Type: new Abstract: Visual counterfactual explanations aim to answer 'what minimal change to this image would flip the model's prediction?', and are increasingly important

safetyarxiv-cs-ai
28 Jul 2026
Safety

Continual-RL for Generalization in Autonomous Racing on the RoboRacer Platform

DGX agent

arXiv:2607.24320v1 Announce Type: new Abstract: A key challenge in modern robotics is to adapt to changing environments, a challenge that is exacerbated when simulations cannot encompass every possibl

safetyarxiv-cs-ro
28 Jul 2026
Research

Continuous surrogates versus threshold Boolean networks for modeling Arabidopsis ISR gene regulation

DGX agent

arXiv:2607.23289v1 Announce Type: cross Abstract: Gene regulatory network modeling often requires balancing predictive accuracy and mechanistic interpretability. In this work, we compare continuous su

researcharxiv-cs-lg
28 Jul 2026
Research

DataOrchestra: Learning to Orchestrate Per-Example Curation of Pretraining Data

DGX agent

arXiv:2607.24717v1 Announce Type: cross Abstract: Pretraining data processing is critical to the downstream performance of Large Language Models (LLMs). However, many existing approaches define a fixe

researcharxiv-cs-ai
28 Jul 2026
Safety

Directional Influence Function: Estimating Training Data Influence in Constrained Learning

DGX agent

arXiv:2607.23388v1 Announce Type: cross Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustnes

safetyarxiv-cs-ai
28 Jul 2026
← Previous
1…5960616263…109
Next →