AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Model Releases

MathlibLemma: Folklore Lemma Generation and Benchmark for Formal Mathematics

DGX agent

arXiv:2602.02561v2 Announce Type: replace-cross Abstract: While the ecosystem of Lean and Mathlib has enjoyed celebrated success in formal mathematical reasoning with the help of large language models

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MCP-Cosmos: World Model-Augmented Agents for Complex Task Execution in MCP Environments

DGX agent

arXiv:2605.09131v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has unified the interface between Large Language Models (LLMs) and external tools, yet a fundamental gap remains in how

model-releasesarxiv-cs-ai
12 May 2026
Research

NaiAD: Initiate Data-Driven Research for LLM Advertising

DGX agent

arXiv:2605.09918v1 Announce Type: cross Abstract: Reconciling platform revenue with user experience in LLM advertising motivates a data-centric foundation. We introduce NaiAD, the first comprehensive

researcharxiv-cs-ai
12 May 2026
Model Releases

Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks

DGX agent

arXiv:2605.10639v1 Announce Type: new Abstract: The rapid adoption of LLMs in both research and industry highlights the challenges of deploying them safely and reveals a gap in the systematic evaluati

model-releasesarxiv-cs-ai
12 May 2026
Agents

NyayaAI: An AI-Powered Legal Assistant Using Multi-Agent Architecture and Retrieval-Augmented Generation

DGX agent

arXiv:2605.10155v1 Announce Type: new Abstract: Legal information in India remains largely inaccessible due to the complexity of legal language and the sheer volume of legal documentation involved in

agentsarxiv-cs-cl
12 May 2026
Agents

Omni-scale Learning-based Sequential Decision Framework for Order Fulfillment of Tote-handling Robotic Systems

DGX agent

arXiv:2605.08758v1 Announce Type: cross Abstract: Driven by the rapid expansion of e-commerce and small-batch production, the size of the intralogistics load unit of finished goods, semi-finished good

agentsarxiv-cs-ai
12 May 2026
Model Releases

OpenSGA: Efficient 3D Scene Graph Alignment in the Open World

DGX agent

arXiv:2605.10484v1 Announce Type: new Abstract: Scene graph alignment establishes object correspondences between two 3D scene graphs constructed from partially overlapping observations. This enables e

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents

DGX agent

arXiv:2605.10341v1 Announce Type: new Abstract: A LaTeX manuscript that compiles without error is not necessarily publication-ready. The resulting PDFs frequently suffer from misplaced floats, overflo

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark

DGX agent

arXiv:2509.26574v4 Announce Type: replace Abstract: While large language models (LLMs) with reasoning capabilities are progressing rapidly on high-school math competitions and coding, can they reason

model-releasesarxiv-cs-ai
12 May 2026
Research

Rapid Forest Fuel Load Estimation via Virtual Remote Sensing and Metric-Scale Feed-Forward 3D Reconstruction

DGX agent

arXiv:2605.10789v1 Announce Type: new Abstract: Accurate quantification of forest coverage and combustible biomass (fuel load) is critical for wildfire risk assessment and ecosystem management. Howeve

researcharxiv-cs-cv
12 May 2026
Research

Raster2Seq: Polygon Sequence Generation for Floorplan Reconstruction

DGX agent

arXiv:2602.09016v2 Announce Type: replace Abstract: Reconstructing a structured vector-graphics representation from a rasterized floorplan image is typically an important prerequisite for computationa

researcharxiv-cs-cv
12 May 2026
Research

Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science

DGX agent

arXiv:2605.10612v1 Announce Type: cross Abstract: Graph neural networks are increasingly adopted in trigger systems for collider experiments, where strict latency and throughput constraints render dep

researcharxiv-cs-lg
12 May 2026
Agents

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

DGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

agentsarxiv-cs-lg
12 May 2026
Model Releases

Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs

DGX agent

arXiv:2509.02372v3 Announce Type: replace-cross Abstract: Large Language Models have become critical to modern software development, but their reliance on uncurated web-scale datasets for training int

model-releasesarxiv-cs-ai
12 May 2026
Agents

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

DGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

agentsarxiv-cs-cl
12 May 2026
Agents

Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments

DGX agent

arXiv:2605.09721v1 Announce Type: cross Abstract: Tool-enabled AI agents are increasingly deployed in cloud-hosted environments and offered as services, where they perform side-effecting operations th

agentsarxiv-cs-ai
12 May 2026
Model Releases

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

DGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

model-releasesarxiv-cs-ai
12 May 2026
Applications

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation

DGX agent

arXiv:2605.10079v1 Announce Type: new Abstract: Video generation has advanced rapidly, producing photorealistic videos from text or image prompts. Meanwhile, film production and social robotics increa

applicationsarxiv-cs-cv
12 May 2026
Research

Speech-based Psychological Crisis Assessment using LLMs

DGX agent

arXiv:2605.10027v1 Announce Type: cross Abstract: Psychological support hotlines provide critical support for individuals experiencing mental health emergencies, yet current assessments largely rely o

researcharxiv-cs-ai
12 May 2026
Local Ai

SymTorch: Symbolic Distillation of Neural Networks

DGX agent

arXiv:2602.21307v2 Announce Type: replace Abstract: What mathematical functions do neural network components learn? Symbolic distillation addresses this question by expressing neural network component

local-aiarxiv-cs-lg
12 May 2026
Agents

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

DGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

agentsarxiv-cs-ai
12 May 2026
Safety

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring

DGX agent

arXiv:2605.09225v1 Announce Type: cross Abstract: Jailbreak attacks -- adversarial prompts that bypass LLM alignment through purely linguistic manipulation -- pose a growing operational security threa

safetyarxiv-cs-ai
12 May 2026
Safety

Upholding Epistemic Agency: A Brouwerian Assertibility Constraint for Responsible AI

DGX agent

arXiv:2603.03971v2 Announce Type: replace-cross Abstract: Generative AI can convert uncertainty into hypersuasive, authoritative-seeming verdicts, displacing the justificatory work on which democratic

safetyarxiv-cs-ai
12 May 2026
Model Releases

VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation

DGX agent

arXiv:2605.08553v1 Announce Type: cross Abstract: Large language models can generate useful code from natural language, but their outputs come without correctness guarantees. Verifiable code generatio

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VFM-SDM: A vision foundation model-based framework for training-free, marker-free, and calibration-free structural displacement measurement

DGX agent

arXiv:2605.09677v1 Announce Type: new Abstract: Reliable displacement measurement is fundamental for structural health monitoring and digital engineering workflows, as it provides direct structural re

model-releasesarxiv-cs-cv
12 May 2026
Agents

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

DGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

agentsarxiv-cs-ai
12 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Safety

A Large-Scale Dataset for Molecular Structure-Language Description via a Rule-Regularized Method

DGX agent

arXiv:2602.02320v3 Announce Type: replace-cross Abstract: Molecular function is largely determined by structure. Accurately aligning molecular structure with natural language is therefore essential fo

safetyarxiv-cs-ai
11 May 2026
Research

A Linear-Transformer Hybrid for SNP-Based Genotype-to-Phenotype Prediction in Grapevine

DGX agent

arXiv:2605.06762v1 Announce Type: cross Abstract: Robust genotype-to-phenotype (G2P) prediction is essential for accelerating breeding decisions and genetic gain. However, it remains challenging to me

researcharxiv-cs-ai
11 May 2026
Agents

A Self-Healing Framework for Reliable LLM-Based Autonomous Agents

DGX agent

arXiv:2605.06737v1 Announce Type: cross Abstract: Autonomous agents based on Large Language Models (LLMs) are increasingly being utilized in complex software systems. However, reliability remains a si

agentsarxiv-cs-ai
11 May 2026
Research

A Unified Framework for the Detection and Classification of Fatty Pancreas in Ultrasound Images

DGX agent

arXiv:2605.07466v1 Announce Type: new Abstract: Non-alcoholic fatty pancreas disease (NAFPD) is an underdiagnosed condition associated with metabolic syndrome, insulin resistance, and increased risk o

researcharxiv-cs-cv
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Agents

ATHENA: Agentic Team for Hierarchical Evolutionary Numerical Algorithms

DGX agent

arXiv:2512.03476v2 Announce Type: replace-cross Abstract: Bridging the gap between theoretical conceptualization and computational implementation is a major bottleneck in Scientific Computing (SciC) a

agentsarxiv-cs-ai
11 May 2026
Hardware

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

DGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

hardwarearxiv-cs-cl
11 May 2026
Research

Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation

DGX agent

arXiv:2510.07926v2 Announce Type: replace Abstract: Despite demonstrating remarkable performance across a wide range of tasks, large language models (LLMs) have also been found to frequently produce o

researcharxiv-cs-cl
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution

DGX agent

arXiv:2605.07208v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to brainstorm and evaluate research ideas, yet assessing such judgments is fundamentally difficult be

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

DGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Narrow Secret Loyalty Dodges Black-Box Audits

DGX agent

arXiv:2605.06846v1 Announce Type: cross Abstract: Recent work identifies secret loyalties as a distinct threat from standard backdoors. A secret loyalty causes a model to covertly advance the interest

model-releasesarxiv-cs-ai
11 May 2026
Safety

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

DGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

safetyarxiv-cs-ai
11 May 2026
Safety

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

DGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

safetyarxiv-cs-cl
11 May 2026
Model Releases

ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs

DGX agent

arXiv:2605.07338v1 Announce Type: new Abstract: The decline of global shellfish biodiversity poses a severe threat to coastal ecosystems. Although artificial intelligence (AI) technologies show potent

model-releasesarxiv-cs-cv
11 May 2026
Research

Statistical Patterns in the Equations of Physics and the Emergence of a Meta-Law of Nature

DGX agent

arXiv:2408.11065v2 Announce Type: replace-cross Abstract: Physics seeks to uncover the laws of Nature and express them through mathematical equations. Despite the vast diversity of natural phenomena,

researcharxiv-cs-cl
11 May 2026
Safety

STDA-Net: Spectrogram-Based Domain Adaptation for cross-dataset Sleep Stage Classification

DGX agent

arXiv:2605.06736v1 Announce Type: cross Abstract: Accurate sleep stage classification across datasets remains challenging due to variability in EEG channel montages, sampling rates, recording environm

safetyarxiv-cs-ai
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

DGX agent

arXiv:2605.07593v1 Announce Type: new Abstract: Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams,

model-releasesarxiv-cs-cv
11 May 2026
← Previous
1…6869707172…82
Next →