AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Local Ai

Domain-Filtered Knowledge Graphs from Sparse Autoencoder Features

DGX agent

arXiv:2604.23829v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) extract millions of interpretable features from a language model, but flat feature inventories aren't very useful on their ow

local-aiarxiv-cs-ai
28 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Rights to Rites: Expectations Management in Smart-Home AI

DGX agent

arXiv:2604.23635v1 Announce Type: cross Abstract: Domestic voice assistants and smart-home devices are increasingly embedded in everyday routines, yet their ethics are often treated as an afterthought

researcharxiv-cs-ai
28 Apr 2026
Research

Interpretable Physics-Informed Load Forecasting for U.S. Grid Resilience: SHAP-Guided Ensemble Validation in Hybrid Deep Learning Under Extreme Weather

DGX agent

arXiv:2604.23500v1 Announce Type: cross Abstract: Accurate short-term electricity load forecasting is a cornerstone of U.S. grid reliability; however, prevailing deep learning models remain opaque, li

researcharxiv-cs-ai
28 Apr 2026
Model Releases

IntrAgent: An LLM Agent for Content-Grounded Information Retrieval through Literature Review

DGX agent

arXiv:2604.22861v1 Announce Type: cross Abstract: Scientific research relies on accurate information retrieval from literature to support analytical decisions. In this work, we introduce a new task, I

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

KLong: Training LLM Agent for Extremely Long-horizon Tasks

DGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation

DGX agent

arXiv:2603.17717v3 Announce Type: replace-cross Abstract: Supervised detection of network attacks has always been a critical part of network intrusion detection systems (NIDS). Nowadays, in a pivotal

researcharxiv-cs-ai
28 Apr 2026
Model Releases

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

DGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

DGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

NeuroClaw Technical Report

DGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

DGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

DGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Scalable LLM-based Coding of Dialogue in Healthcare Simulation: Balancing Coding Performance, Processing Time, and Environmental Impact

DGX agent

arXiv:2604.23255v1 Announce Type: cross Abstract: Research shows that dialogue, the interactive process through which participants articulate their thinking, plays a central role in constructing share

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Self-Admitted Technical Debt Detection Approaches: A Decade Systematic Review

DGX agent

arXiv:2312.15020v4 Announce Type: replace-cross Abstract: Technical debt (TD) refers to the long-term costs associated with suboptimal design or code decisions in software development, often made to m

researcharxiv-cs-ai
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Research

SPLIT: Separating Physical-Contact via Latent Arithmetic in Image-Based Tactile Sensors

DGX agent

arXiv:2604.24449v1 Announce Type: cross Abstract: Training machine learning models for robotic tactile sensing requires vast amounts of data, yet obtaining realistic interaction data remains a challen

researcharxiv-cs-ai
28 Apr 2026
Model Releases

SWE-QA: Can Language Models Answer Repository-level Code Questions?

DGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

TeachMaster: Generative Teaching via Code

DGX agent

arXiv:2601.04204v2 Announce Type: replace-cross Abstract: The scalability of high-quality online education is hindered by the high costs and slow cycles of manual content creation. Despite advancement

agentsarxiv-cs-ai
28 Apr 2026
Agents

The Last Human-Written Paper: Agent-Native Research Artifacts

DGX agent

arXiv:2604.24658v1 Announce Type: new Abstract: Scientific publication compresses a branching, iterative research process into a linear narrative, discarding the majority of what was discovered along

agentsarxiv-cs-lg
28 Apr 2026
Safety

Time-Series Forecasting in Safety-Critical Environments: An EU-AI-Act-Compliant Open-Source Package / Zeitreihenprognose in sicherheitskritischen Umgebungen: Ein KI-VO-konformes Open-Source-Paket

DGX agent

arXiv:2604.23859v1 Announce Type: new Abstract: With spotforecast2-safe we present an integrated Compliance-by-Design approach to Python-based point forecasting of time series in safety-critical envir

safetyarxiv-cs-ai
28 Apr 2026
Agents

Towards Lawful Autonomous Driving: Deriving Scenario-Aware Driving Requirements from Traffic Laws and Regulations

DGX agent

arXiv:2604.24562v1 Announce Type: new Abstract: Driving in compliance with traffic laws and regulations is a basic requirement for human drivers, yet autonomous vehicles (AVs) can violate these requir

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

DGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

DGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

safetyarxiv-cs-ai
28 Apr 2026
Research

Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

DGX agent

arXiv:2604.23706v1 Announce Type: new Abstract: Histologic assessment of ulcerative colitis (UC) activity is an important endpoint in clinical trials and routine care, but manual grading with indices

researcharxiv-cs-cv
28 Apr 2026
Model Releases

xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models

DGX agent

arXiv:2509.13021v2 Announce Type: replace-cross Abstract: This work introduces xOffense, an AI-driven, multi-agent penetration testing framework that shifts the process from labor-intensive, expert-dr

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Your Students Don't Use LLMs Like You Wish They Did

DGX agent

arXiv:2604.23486v1 Announce Type: new Abstract: Educational NLP systems are typically evaluated using engagement metrics and satisfaction surveys, which are at best a proxy for meeting pedagogical goa

safetyarxiv-cs-cl
28 Apr 2026
Safety

ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems

DGX agent

arXiv:2604.23878v1 Announce Type: new Abstract: Despite a century of empirical memory research, existing AI agent memory systems rely on system-engineering metaphors (virtual-memory paging, flat LLM s

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Atlas-Alignment: Making Interpretability Transferable Across Language Models

DGX agent

arXiv:2510.27413v2 Announce Type: replace-cross Abstract: Interpretability is crucial for building safe, reliable, and controllable language models, yet existing interpretability pipelines remain cost

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Can Large Language Models Adequately Perform Symbolic Reasoning Over Time Series?

DGX agent

arXiv:2508.03963v4 Announce Type: replace Abstract: Uncovering hidden symbolic laws from time series data, as an aspiration dating back to Kepler's discovery of planetary motion, remains a core challe

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding

DGX agent

arXiv:2604.22333v1 Announce Type: cross Abstract: Rapid situational awareness is critical in post-disaster response. While remote sensing damage assessment is evolving from pixel-level change detectio

model-releasesarxiv-cs-ai
27 Apr 2026
Research

CRAFT: Clustered Regression for Adaptive Filtering of Training data

DGX agent

arXiv:2604.22693v1 Announce Type: cross Abstract: Selecting a small, high-quality subset from a large corpus for fine-tuning is increasingly important as corpora grow to tens of millions of datapoints

researcharxiv-cs-ai
27 Apr 2026
Safety

Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework

DGX agent

arXiv:2604.22119v1 Announce Type: new Abstract: As reasoning capacity and deployment scope grow in tandem, large language models (LLMs) gain the capacity to engage in behaviors that serve their own ob

safetyarxiv-cs-ai
27 Apr 2026
Applications

FeatEHR-LLM: Leveraging Large Language Models for Feature Engineering in Electronic Health Records

DGX agent

arXiv:2604.22534v1 Announce Type: cross Abstract: Feature engineering for Electronic Health Records (EHR) is complicated by irregular observation intervals, variable measurement frequencies, and struc

applicationsarxiv-cs-ai
27 Apr 2026
Research

FixV2W: Correcting Invalid CVE-CWE Mappings with Knowledge Graph Embeddings

DGX agent

arXiv:2604.22176v1 Announce Type: cross Abstract: Accurate mapping between Common Vulnerabilities and Exposures (CVE) and Common Weakness Enumeration (CWE) entries is critical for effective vulnerabil

researcharxiv-cs-lg
27 Apr 2026
Safety

Learning Evidence Highlighting for Frozen LLMs

DGX agent

arXiv:2604.22565v1 Announce Type: cross Abstract: Large Language Models (LLMs) can reason well, yet often miss decisive evidence when it is buried in long, noisy contexts. We introduce HiLight, an Evi

safetyarxiv-cs-ai
27 Apr 2026
Research

Lifting Unlabeled Internet-level Data for 3D Scene Understanding

DGX agent

arXiv:2604.01907v2 Announce Type: replace-cross Abstract: Annotated 3D scene data is scarce and expensive to acquire, while abundant unlabeled videos are readily available on the internet. In this pap

researcharxiv-cs-ai
27 Apr 2026
Applications

Manifold Learning for Personalized and Label-Free Detection of Cardiac Arrhythmias

DGX agent

arXiv:2506.16494v3 Announce Type: replace Abstract: Electrocardiograms (ECGs) provide non-invasive measurements of heart activity and are established tools for detecting cardiac arrhythmias. Although

applicationsarxiv-cs-lg
27 Apr 2026
Agents

Memanto: Typed Semantic Memory with Information-Theoretic Retrieval for Long-Horizon Agents

DGX agent

arXiv:2604.22085v1 Announce Type: new Abstract: The transition from stateless language model inference to persistent, multi session autonomous agents has revealed memory to be a primary architectural

agentsarxiv-cs-ai
27 Apr 2026
Applications

Predicting Liquidity-Aware Bond Yields using Causal GANs and Deep Reinforcement Learning with LLM Evaluation

DGX agent

arXiv:2502.17011v2 Announce Type: replace-cross Abstract: Financial bond yield forecasting is challenging due to data scarcity, nonlinear macroeconomic dependencies, and evolving market conditions. In

applicationsarxiv-cs-cl
27 Apr 2026
Tutorials

PrivSTRUCT: Untangling Data Purpose Compliance of Privacy Policies in Google Play Store

DGX agent

arXiv:2604.22157v1 Announce Type: cross Abstract: Existing research typically treats privacy policies as flat, uniform text, extracting information without regard for the document's logical hierarchy.

tutorialsarxiv-cs-ai
27 Apr 2026
Safety

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

DGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

safetyarxiv-cs-ai
27 Apr 2026
Agents

Robust Localization for Autonomous Vehicles in Highway Scenes

DGX agent

arXiv:2604.22040v1 Announce Type: new Abstract: Localization for autonomous vehicles on highways remains under-explored compared to urban roads, and state-of-the-art methods for urban scenes degrade w

agentsarxiv-cs-ro
27 Apr 2026
Safety

The Biggest Risk of Embodied AI is Governance Lag

DGX agent

arXiv:2604.21938v1 Announce Type: cross Abstract: Embodied AI is widely discussed as a job-displacement problem. The deeper risk, however, is governance lag: the inability of public institutions to ke

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

DGX agent

arXiv:2604.22068v1 Announce Type: cross Abstract: Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks add

model-releasesarxiv-cs-ro
27 Apr 2026
Research

AI for software engineering: from probable to provable

DGX agent

arXiv:2511.23159v2 Announce Type: replace-cross Abstract: Vibe coding, the much-touted use of AI techniques for programming, faces two overwhelming obstacles: the difficulty of specifying goals ('prom

researcharxiv-cs-ai
24 Apr 2026
Safety

AI Governance under Political Turnover: The Alignment Surface of Compliance Design

DGX agent

arXiv:2604.21103v1 Announce Type: new Abstract: Governments are increasingly interested in using AI to make administrative decisions cheaper, more scalable, and more consistent. But for probabilistic

safetyarxiv-cs-ai
24 Apr 2026
← Previous
1…7273747576…82
Next →