AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Agents

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

DGX agent

arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and

agentsarxiv-cs-ai
23 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evaluating Black-Box Vulnerabilities with Wasserstein-Constrained Data Perturbations

DGX agent

arXiv:2603.15867v2 Announce Type: replace Abstract: The growing use of Machine Learning (ML) tools comes with critical challenges, such as limited model explainability. We propose a global explainabil

safetyarxiv-cs-lg
23 Apr 2026
Agents

Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks

DGX agent

arXiv:2508.09958v3 Announce Type: replace Abstract: As large language models (LLMs) become increasingly popular, there is a growing need to predict which out of a set of LLMs will yield a successful a

agentsarxiv-cs-cl
23 Apr 2026
Research

Choose Your Own Adventure: Non-Linear AI-Assisted Programming with EvoGraph

DGX agent

arXiv:2604.18883v1 Announce Type: cross Abstract: Current AI-assisted programming tools are predominantly linear and chat-based, which deviates from the iterative and branching nature of programming i

researcharxiv-cs-ai
22 Apr 2026
Local Ai

Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents

DGX agent

arXiv:2604.00555v2 Announce Type: replace Abstract: Enterprise adoption of Large Language Models (LLMs) is constrained by hallucination, domain drift, and the inability to enforce regulatory complianc

local-aiarxiv-cs-ai
22 Apr 2026
Applications

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

DGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

applicationsarxiv-cs-cl
22 Apr 2026
Agents

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

DGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

agentsarxiv-cs-ai
22 Apr 2026
Applications

RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian

DGX agent

arXiv:2604.19593v1 Announce Type: cross Abstract: The importance of clear and correct text in legal documents cannot be understated, and, consequently, a grammatical error correction tool meant to ass

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

Time Series Augmented Generation for Financial Applications

DGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks

DGX agent

arXiv:2604.17761v1 Announce Type: cross Abstract: Interpretability tools are increasingly used to analyze failures of Large Language Models (LLMs), yet prior work largely focuses on short prompts or t

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs

DGX agent

arXiv:2510.07591v3 Announce Type: replace Abstract: We present a system that uses LLMs as a tool in the development of Constructed Languages -- ConLangs, which we call IASC (Interactive Agentic System

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Designing Explainable Conversational Agentic Systems for Guarani Speakers

DGX agent

arXiv:2603.05743v3 Announce Type: replace Abstract: Although artificial intelligence (AI) and Human-Computer Interaction (HCI) systems are often presented as universal solutions, their design remains

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

DGX agent

arXiv:2604.17159v1 Announce Type: cross Abstract: We present, to our knowledge, the most comprehensive cross-model evaluation of LLM agents on offensive cybersecurity tasks, benchmarking 10 frontier m

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants

DGX agent

arXiv:2503.14281v4 Announce Type: replace-cross Abstract: AI coding assistants are widely used for tasks like code generation. These tools now require large and complex contexts, automatically sourced

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

DGX agent

arXiv:2604.15709v1 Announce Type: new Abstract: Agent exttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform part

agentsarxiv-cs-ai
20 Apr 2026
Tutorials

Discovering quantum phenomena with Interpretable Machine Learning

DGX agent

arXiv:2604.16015v1 Announce Type: cross Abstract: Interpretable machine learning techniques are becoming essential tools for extracting physical insights from complex quantum data. We build on recent

tutorialsarxiv-cs-lg
20 Apr 2026
Agents

The Semi-Executable Stack: Agentic Software Engineering and the Expanding Scope of SE

DGX agent

arXiv:2604.15468v1 Announce Type: cross Abstract: AI-based systems, currently driven largely by LLMs and tool-using agentic harnesses, are increasingly discussed as a possible threat to software engin

agentsarxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Model Releases

Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems

DGX agent

arXiv:2604.14228v1 Announce Type: cross Abstract: Claude Code is an agentic coding tool that can run shell commands, edit files, and call external services on behalf of the user. This study describes

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems

DGX agent

arXiv:2510.14133v2 Announce Type: replace Abstract: Agentic AI systems, which leverage multiple autonomous agents and large language models (LLMs), are increasingly used to address complex, multi-step

safetyarxiv-cs-ai
17 Apr 2026
Safety

Hierarchical Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text

DGX agent

arXiv:2604.14166v1 Announce Type: new Abstract: Mapping Cyber Threat Intelligence (CTI) text to MITRE ATT&CK technique IDs is a critical task for understanding adversary behaviors and automating threa

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation

DGX agent

arXiv:2604.15309v1 Announce Type: cross Abstract: The rapid progress of Artificial Intelligence Generated Content (AIGC) tools enables images, videos, and visualizations to be created on demand for we

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

LoRA-MME: Multi-Model Ensemble of LoRA-Tuned Encoders for Code Comment Classification

DGX agent

arXiv:2603.03959v4 Announce Type: replace-cross Abstract: Code comment classification is a critical task for automated software documentation and analysis. In the context of the NLBSE'26 Tool Competit

model-releasesarxiv-cs-lg
16 Apr 2026
Research

Mobius transforms and Shapley values for vector-valued functions on weighted directed acyclic multigraphs

DGX agent

arXiv:2510.05786v3 Announce Type: replace-cross Abstract: Mobius inversion and Shapley values are two mathematical tools for characterizing and decomposing higher-order structure in complex systems. T

researcharxiv-cs-lg
16 Apr 2026
Model Releases

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

DGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

GCA Framework: A Gulf-Grounded Dataset and Agentic Pipeline for Climate Decision Support

DGX agent

arXiv:2604.12306v1 Announce Type: cross Abstract: Climate decision-making in the Gulf increasingly demands systems that can translate heterogeneous scientific and policy evidence into actionable guida

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

A Complete Decomposition of KL Error using Refined Information and Mode Interaction Selection

DGX agent

arXiv:2410.11964v2 Announce Type: replace Abstract: The log-linear model has received a significant amount of theoretical attention in previous decades and remains the fundamental tool used for learni

applicationsarxiv-cs-lg
14 Apr 2026
Research

DeepSketcher: Internalizing Visual Manipulation for Multimodal Reasoning

DGX agent

arXiv:2509.25866v2 Announce Type: replace Abstract: The 'thinking with images' paradigm represents a pivotal shift in the reasoning of Vision Language Models (VLMs), moving from text-dominant chain-of

researcharxiv-cs-cv
14 Apr 2026
Safety

FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

DGX agent

arXiv:2604.10693v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has improved LLM reasoning, but models often generate explanations that appear coherent while containing unfaithful int

safetyarxiv-cs-ai
14 Apr 2026
Research

LETGAMES: An LLM-Powered Gamified Approach to Cognitive Training for Patients with Cognitive Impairment

DGX agent

arXiv:2604.09566v1 Announce Type: cross Abstract: The application of games as a therapeutic tool for cognitive training is beneficial for patients with cognitive impairments. However, effective game d

researcharxiv-cs-ai
14 Apr 2026
Agents

Machine Learning-Based Detection of MCP Attacks

DGX agent

arXiv:2604.10534v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is a new and emerging technology that extends the functionality of large language models, improving workflows but als

agentsarxiv-cs-ai
14 Apr 2026
Agents

Problem Reductions at Scale: Agentic Integration of Computationally Hard Problems

DGX agent

arXiv:2604.11535v1 Announce Type: new Abstract: Solving an NP-hard optimization problem often requires reformulating it for a specific solver -- quantum hardware, a commercial optimizer, or a domain h

agentsarxiv-cs-ai
14 Apr 2026
Safety

Resilient Write: A Six-Layer Durable Write Surface for LLM Coding Agents

DGX agent

arXiv:2604.10842v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly rely on tool-use protocols such as the Model Context Protocol~(MCP) to read and write files on a developer's wo

safetyarxiv-cs-ai
14 Apr 2026
Agents

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

DGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

agentsarxiv-cs-cl
14 Apr 2026
Safety

The Augmentation Trap: AI Productivity and the Cost of Cognitive Offloading

DGX agent

arXiv:2604.03501v2 Announce Type: replace-cross Abstract: Experimental evidence confirms that AI tools raise worker productivity, but also that sustained use can erode the expertise on which those gai

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL

DGX agent

arXiv:2604.09470v1 Announce Type: new Abstract: Translating natural language into Jira Query Language (JQL) requires resolving ambiguous field references, instance-specific categorical values, and com

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

DGX agent

arXiv:2507.04736v2 Announce Type: replace Abstract: Large Language Models have emerged as powerful tools for automating Register-Transfer Level (RTL) code generation, yet they face critical limitation

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space

DGX agent

arXiv:2604.09029v1 Announce Type: cross Abstract: Large language models have been widely explored as decision-support tools in high-stakes domains due to their contextual understanding and reasoning c

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Many-Tier Instruction Hierarchy in LLM Agents

DGX agent

arXiv:2604.09443v1 Announce Type: cross Abstract: Large language model agents receive instructions from many sources-system messages, user prompts, tool outputs, and more-each carrying different level

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence

DGX agent

arXiv:2604.09104v1 Announce Type: cross Abstract: Scheming, the covert pursuit of misaligned goals by AI systems, represents a potentially catastrophic risk, yet scheming research suffers from signifi

safetyarxiv-cs-ai
13 Apr 2026
Agents

Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects

DGX agent

arXiv:2604.06373v1 Announce Type: cross Abstract: New generation of AI coding tools, including AI-powered IDEs equipped with agentic capabilities, can generate code within the context of the project.

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

DGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Contrastive Decoding Mitigates Score Range Bias in LLM-as-a-Judge

DGX agent

arXiv:2510.18196v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are commonly used as evaluators in various applications, but the reliability of the outcomes remains a challenge.

safetyarxiv-cs-ai
10 Apr 2026
Safety

Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education

DGX agent

arXiv:2604.07253v1 Announce Type: cross Abstract: In gender-restrictive and surveilled contexts, where access to formal education may be restricted for women, pursuing education involves safety and pr

safetyarxiv-cs-ai
10 Apr 2026
Safety

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

DGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

safetyarxiv-cs-ai
10 Apr 2026
Safety

SymptomWise: A Deterministic Reasoning Layer for Reliable and Efficient AI Systems

DGX agent

arXiv:2604.06375v1 Announce Type: new Abstract: AI-driven symptom analysis systems face persistent challenges in reliability, interpretability, and hallucination. End-to-end generative approaches ofte

safetyarxiv-cs-ai
10 Apr 2026
Research

Thinking in Graphs with CoMAP: A Shared Visual Workspace for Designing Project-Based Learning

DGX agent

arXiv:2604.06200v1 Announce Type: cross Abstract: Designing project-based learning (PBL) demands managing highly interdependent components, a task that both traditional linear tools and purely convers

researcharxiv-cs-ai
10 Apr 2026
Safety

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

DGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

safetyarxiv-cs-ai
10 Apr 2026
← Previous
1…2021222324…108
Next →