AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,942 results
Safety

State-Dependent Safety Failures in Multi-Turn Language Model Interaction

DGX agent

arXiv:2603.15684v2 Announce Type: replace-cross Abstract: Safety alignment in large language models is typically evaluated under isolated queries, yet real-world use is inherently multi-turn. Although

safetyarxiv-cs-ai
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Wiring diagram extraction and gluing: a case study in classifying figure skating jumps using 3D dataset

DGX agent

arXiv:2607.27598v1 Announce Type: cross Abstract: Hasse clustering is an algorithm that extracts common patterns in sequential data and represents them in graphical forms. As the number of expected cl

applicationsarxiv-cs-lg
31 Jul 2026
Local Ai

AgentGUI: An Interface for Observing and Steering Long-Running AI Agents

DGX agent

arXiv:2607.26300v1 Announce Type: new Abstract: AI agents are increasingly adept at tackling complex, long-running tasks. With the rapid surge of autonomous capabilities, human oversight is systematic

local-aiarxiv-cs-cl
30 Jul 2026
Safety

Controlled Experiments on Lane Changing by Transitional Autonomous Vehicle: Dataset and Behavioral Insights

DGX agent

arXiv:2607.27085v1 Announce Type: new Abstract: This paper presents the North Carolina Transitional Autonomous Vehicle Lane-Changing (NC-tALC) dataset and uses it to characterize mandatory lane-changi

safetyarxiv-cs-ro
30 Jul 2026
Research

Scientific Knowledge Discovery in the Age of Large Language Models

DGX agent

arXiv:2607.26670v1 Announce Type: cross Abstract: The rapid growth of scholarly literature has made identifying relevant publications increasingly difficult, and conventional search systems still depe

researcharxiv-cs-cl
30 Jul 2026
Local Ai

Where Physics Meets Privacy: Federated PINNs for Privacy-Preserving Brain Tumor Biomechanical Modeling

DGX agent

arXiv:2607.26207v1 Announce Type: new Abstract: Brain tumors such as glioma, meningioma, and pituitary adenoma alter the mechanical behavior of soft brain tissue, yet common diagnostic methods rely on

local-aiarxiv-cs-cv
30 Jul 2026
Model Releases

Balancing multiscale similarity and cartographic constraints: A similarity-driven optimization framework for line generalization

DGX agent

arXiv:2607.25474v1 Announce Type: new Abstract: Cartographic generalization is essential for generating multiscale map representations by balancing information preservation and cartographic readabilit

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CogEEGAgent: Toward Autonomous Cognitive EEG Analysis with Grounded Execution and Selection-Aware Verification

DGX agent

arXiv:2607.25045v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis in cognitive studies requires specialized expertise and involves many defensible choices over contrasts, channels,

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

ARdena: Scenario-driven control of real-time LLM agents

DGX agent

arXiv:2607.22651v1 Announce Type: new Abstract: Large language models (LLMs) have enabled increasingly capable conversational agents, but reliably controlling their behavior in real-time interactive e

safetyarxiv-cs-ai
28 Jul 2026
Research

AutoCluster, AutoTopicModeling, AutoTrendAnalysis: A Complete AutoML Pipeline for Predicting Emerging Trends

DGX agent

arXiv:2607.22641v1 Announce Type: cross Abstract: Predicting emerging trends is vital for businesses, researchers, and policymakers; yet traditional approaches often lack scalability and adaptability.

researcharxiv-cs-ai
28 Jul 2026
Research

DataOrchestra: Learning to Orchestrate Per-Example Curation of Pretraining Data

DGX agent

arXiv:2607.24717v1 Announce Type: cross Abstract: Pretraining data processing is critical to the downstream performance of Large Language Models (LLMs). However, many existing approaches define a fixe

researcharxiv-cs-ai
28 Jul 2026
Safety

Designing Service Systems from Textual Evidence

DGX agent

arXiv:2603.10400v2 Announce Type: replace-cross Abstract: Designing service systems requires selecting among alternative configurations -- choosing the best chatbot variant, the optimal routing policy

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Diffusion-Guided Search via Exponential Tilting (DiffTilt): An Application to Falsification of Safety-Critical Systems

DGX agent

arXiv:2607.23134v1 Announce Type: new Abstract: Discovering rare safety-critical failures in autonomous and cyber-physical systems is a fundamental challenge in verification and validation. Existing f

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

DualityCert: Verifier-Gated Language-Model Repair of Broken Duality Claims in Quantum Field Theory

DGX agent

arXiv:2607.23614v1 Announce Type: cross Abstract: We present DualityCert, a symbolic verifier for candidate Seiberg-duality claims in four-dimensional N=1 quiver gauge theories. The verifier evaluates

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Flash-CNNCap: Capacitance Extraction via Image Mapping

DGX agent

arXiv:2607.23877v1 Announce Type: new Abstract: We present Flash-CNNCap, a CNN-based capacitance extractor that reformulates full-matrix capacitance prediction as image-to-image regression over spatia

model-releasesarxiv-cs-lg
28 Jul 2026
Safety

MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents

DGX agent

arXiv:2607.18999v2 Announce Type: replace-cross Abstract: Evaluating multi-turn medical consultation agents requires judging the diagnostic support provided by the histories they elicit through intera

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows

DGX agent

arXiv:2510.24411v3 Announce Type: replace Abstract: Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobi

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

scMIR: a vision-language foundation model for single-cell light microscopy image representation

DGX agent

arXiv:2607.22712v1 Announce Type: cross Abstract: Single-cell light microscopy images have become an important data source for characterizing cell phenotypes, but their complexity and heterogeneity po

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SIREN: Towards End-to-End Extreme-Weather Early Warning with Experience-Grounded LLM Agents

DGX agent

arXiv:2607.24588v1 Announce Type: new Abstract: Early warning of extreme weather is essential for mitigating the societal, economic, and environmental risks posed by hazardous weather events. However,

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents

DGX agent

arXiv:2607.22563v1 Announce Type: new Abstract: Industrial agent benchmarks require realistic evaluation scenarios that integrate telemetry, failure modes, maintenance records, and domain standards. H

safetyarxiv-cs-ai
28 Jul 2026
Research

Text-based Tactile Graphics Generation for the Visually Impaired

DGX agent

arXiv:2607.22674v1 Announce Type: cross Abstract: Tactile graphics are a primary medium for blind and low-vision (BLV) individuals to access non-textual information. However, they are difficult to sca

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

DGX agent

arXiv:2606.13385v2 Announce Type: replace-cross Abstract: LLM-based web agents are increasingly deployed in real-world settings such as e-commerce, where they interact extensively with untrusted web c

model-releasesarxiv-cs-ai
28 Jul 2026
Research

CARE: Anti-entanglement Ultrasound Image Segmentation via Channel-Aware Region Extrication

DGX agent

arXiv:2508.13899v2 Announce Type: replace Abstract: Accurate ultrasound image segmentation is fundamentally challenged by target-context entanglement, where lesion cues are easily mixed with surroundi

researcharxiv-cs-cv
27 Jul 2026
Model Releases

Medical-Checklist: Assessing the Comprehension of Medical Images by Multimodal Models

DGX agent

arXiv:2607.21998v1 Announce Type: new Abstract: This paper introduces a new benchmark test, Medical-Checklist, for assessing medical multimodal models. The recent advancements in multimodal models hav

model-releasesarxiv-cs-cv
27 Jul 2026
Safety

Reliability Scales Inversely: Bigger Language Models Compound Mistakes Faster

DGX agent

arXiv:2607.18292v2 Announce Type: replace-cross Abstract: As language models scale, answers start truer but degrade faster: scaling buys capability but erodes reliability. The knowledge-gap account --

safetyarxiv-cs-cl
27 Jul 2026
Model Releases

Scaling Laws for Classical Machine Learning on Tabular Data: A Benchmark Study

DGX agent

arXiv:2607.21866v1 Announce Type: new Abstract: Prior classical-ML learning-curve work fits power laws to tree, linear, and kernel models on tabular data, but at small scale: typically one curve, one

model-releasesarxiv-cs-lg
27 Jul 2026
Agents

OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining

DGX agent

arXiv:2607.20486v1 Announce Type: new Abstract: Designing optimizers for modern deep learning remains a challenging scientific problem, requiring the joint consideration of optimization geometry, stat

agentsarxiv-cs-ai
24 Jul 2026
Research

The Human-AI Substitution Principle: When will you be replaced by AI in your organization?

DGX agent

arXiv:2607.20781v1 Announce Type: new Abstract: Artificial Intelligence (AI) is rapidly transforming organizations, raising a fundamental organizational and economic question: when will a human employ

researcharxiv-cs-ai
24 Jul 2026
Model Releases

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

DGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Hypothesis-and-Refinement Learning of Organic Structures from Multimodal Spectroscopic Data

DGX agent

arXiv:2607.19816v1 Announce Type: cross Abstract: Determining molecular structures from spectroscopic data remains fundamentally challenging because the inverse problem is intrinsically underdetermine

model-releasesarxiv-cs-lg
23 Jul 2026
Agents

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

DGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

agentsarxiv-cs-ai
16 Jul 2026
Research

Faithful Autoformalization of Natural Language Assertions

DGX agent

arXiv:2607.13303v1 Announce Type: cross Abstract: Formal contracts are essential for software testing and verification, yet writing them remains labor-intensive and error-prone. LLMs offer a promising

researcharxiv-cs-ai
16 Jul 2026
Safety

Reverse to Advance: Teleoperation-Cost Effective Hard Policy Learning from Reversed Easy Tasks

DGX agent

arXiv:2607.13455v1 Announce Type: new Abstract: High-quality teleoperation datasets are costly to collect, particularly for hard tasks. We observe that many tasks exhibit directional asymmetry: comple

safetyarxiv-cs-ro
16 Jul 2026
Agents

Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Framework

DGX agent

arXiv:2607.13091v1 Announce Type: cross Abstract: LLM-based coding agents repeat the same classes of mistakes across sessions because they lack a mechanism to retain corrections from human review feed

agentsarxiv-cs-ai
16 Jul 2026
Safety

A Neurosymbolic Approach to Natural Language Formalization and Verification

DGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

Agentic systems for breast cancer treatment recommendations

DGX agent

arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?

DGX agent

arXiv:2511.09483v3 Announce Type: replace Abstract: While multimodal large language models can describe visual content, their ability to generate executable procedures remains underexplored. CrochetBe

model-releasesarxiv-cs-ai
15 Jul 2026
Research

Good Benchmarks

DGX agent

arXiv:2607.12217v1 Announce Type: new Abstract: Good tasks are correct, solvable, verifiable, well-specified, and hard for interesting reasons. The best tasks describe a real problem an experienced pr

researcharxiv-cs-ai
15 Jul 2026
Model Releases

How Inference Compute Shapes Frontier LLM Evaluation

DGX agent

arXiv:2606.17930v2 Announce Type: replace Abstract: AI evaluations are shifting toward harder tasks that benefit from longer trajectories involving tool use and iterative problem solving. As a result,

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention

DGX agent

arXiv:2601.14776v3 Announce Type: replace Abstract: Recent advances in multi-modal detection have significantly improved detection accuracy in challenging environments (e.g., low light, overexposure).

safetyarxiv-cs-cv
15 Jul 2026
Research

Same Loss, Same Noise, Opposite Schedules: Noise Structure and Optimizer Normalization Jointly Determine Whether Learning-Rate Cooldown Helps

DGX agent

arXiv:2607.12360v1 Announce Type: new Abstract: The cooldown phase of a warmup-stable-decay (WSD) learning-rate schedule, now a default in large-model pretraining, lowers the final training loss in so

researcharxiv-cs-lg
15 Jul 2026
Model Releases

SeqGPT: A Constrained Transformer Agent for the Inverse Designof Multi-Panel Composite Structures

DGX agent

arXiv:2607.11910v1 Announce Type: cross Abstract: Optimizing composite stacking sequences to match continuous targets (e.g., Lamination or Buckling Parameters) with discrete manufacturing constraints

model-releasesarxiv-cs-ai
15 Jul 2026
Local Ai

WanToFight: Real-Time Generative Game Engine for Multi-Player Combat Interaction

DGX agent

arXiv:2607.12592v1 Announce Type: new Abstract: We present WanToFight, a generative game engine that simulates real-time, two-player The King of Fighters '97 (KOF~'97) gameplay from keyboard input. Pr

local-aiarxiv-cs-cv
15 Jul 2026
Local Ai

FPGN: Redefining Ultra-Fast Programmable Gate-based Neural Acceleration with Differentiable LUTs

DGX agent

arXiv:2607.08427v1 Announce Type: cross Abstract: Achieving nanosecond-scale inference latency for deep neural networks (DNNs) has become a primary architectural concern for latency-critical applicati

local-aiarxiv-cs-lg
10 Jul 2026
Research

GRE-Diff: Gaussian Room Embeddings for Structured Layout Diffusion

DGX agent

arXiv:2607.08086v1 Announce Type: new Abstract: Designing functional and aesthetically coherent floor plans requires exploring a vast space of possible room arrangements, a task that quickly becomes o

researcharxiv-cs-cv
10 Jul 2026
Agents

Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles

DGX agent

arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from norm

agentsarxiv-cs-ai
10 Jul 2026
Safety

Agentic Data Environments

DGX agent

arXiv:2607.07397v1 Announce Type: new Abstract: Autonomous agents promise substantial gains in speed, scale, and labor efficiency, but their failures can impose abrupt and often irreversible costs. Th

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

DGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

model-releasesarxiv-cs-ai
9 Jul 2026
← Previous
1…3435363738…83
Next →