AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

DGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback

DGX agent

arXiv:2605.17448v1 Announce Type: cross Abstract: Computer-aided design (CAD) is the backbone of modern industrial design, yet learned CAD generators still fall short of real engineering pipelines: th

model-releasesarxiv-cs-cl
19 May 2026
Agents

SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

DGX agent

arXiv:2602.14211v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to extend LLM agents with task-specific instructions, executable scripts, and auxiliary resources. While im

agentsarxiv-cs-ai
19 May 2026
Agents

Some[Body] Must Receive That Pain for Agent Accountability

DGX agent

arXiv:2605.16872v1 Announce Type: cross Abstract: AI agents increasingly act consequentially in the real world. This creates a problem we call consequence reception: harm occurs, the producing system

agentsarxiv-cs-ai
19 May 2026
Tutorials

Sparse Autoencoders are Topic Models

DGX agent

arXiv:2511.16309v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) are used to analyze embeddings, but their role and practical value are debated. We propose a new perspective on SAEs by d

tutorialsarxiv-cs-cv
19 May 2026
Agents

Spatiotemporal Robustness of Temporal Logic Tasks using Multi-Objective Reasoning

DGX agent

arXiv:2603.29868v2 Announce Type: replace Abstract: The reliability of autonomous systems depends on their robustness, i.e., their ability to meet their objectives under uncertainty. In this paper, we

agentsarxiv-cs-ai
19 May 2026
Applications

STRIDE-AI: A Threat Modeling Framework for Generative AI Security Assessment

DGX agent

arXiv:2605.17163v1 Announce Type: cross Abstract: Traditional cybersecurity methodologies target deterministic systems and fail to address the probabilistic nature of AI, leaving systems vulnerable to

applicationsarxiv-cs-ai
19 May 2026
Agents

SWoMo: Neuro-Symbolic World Model for Cataract Surgery Simulation

DGX agent

arXiv:2605.16530v1 Announce Type: new Abstract: Realistic surgical simulation plays a crucial role in training novice surgeons and in the development of autonomous agents. World models can scale such

agentsarxiv-cs-cv
19 May 2026
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
Safety

The Impact of AI Search on the Online Content Ecosystem: Evidence from Google and Reddit

DGX agent

arXiv:2605.16428v1 Announce Type: cross Abstract: Search engines traditionally complement online content platforms by directing users seeking information to external websites. The emergence of generat

safetyarxiv-cs-ai
19 May 2026
Tutorials

The Learnability Gap in Medical Latent Diffusion

DGX agent

arXiv:2605.17087v1 Announce Type: new Abstract: Generative data augmentation with latent diffusion models is a promising strategy for addressing class imbalance in medical imaging, yet current approac

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

The Normal Distributions Indistinguishability Spectrum and its Application to Privacy-Preserving Machine Learning

DGX agent

arXiv:2309.01243v4 Announce Type: replace-cross Abstract: We investigate the privacy of {em any} algorithm whose outputs have Gaussian distribution. This work is motivated by the prevalence of such al

model-releasesarxiv-cs-lg
19 May 2026
Applications

The Recovery Mechanism: Technology, Education, and What Happens When the Pattern Breaks

DGX agent

arXiv:2605.16283v1 Announce Type: cross Abstract: For centuries, each new technology has automated some layer of cognitive work and been absorbed by education retreating upward to teach the skills mac

applicationsarxiv-cs-ai
19 May 2026
Local Ai

Thinking with Patterns: Breaking the Perceptual Bottleneck in Visual Planning via Pattern Induction

DGX agent

arXiv:2605.16848v1 Announce Type: cross Abstract: Planning from raw visual input remains a significant challenge for current Vision-Language Models (VLMs), when the complexity of input is beyond their

local-aiarxiv-cs-ai
19 May 2026
Research

Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel Methods

DGX agent

arXiv:2506.10959v3 Announce Type: replace-cross Abstract: While in-context learning (ICL) has achieved remarkable success in natural language and vision domains, its theoretical understanding-particul

researcharxiv-cs-ai
19 May 2026
Applications

Unveiling Memorization-Generalization Coexistence: A Case Study on Arithmetic Tasks with Label Noise

DGX agent

arXiv:2605.18022v1 Announce Type: cross Abstract: Highly over-parameterized models can simultaneously memorize noisy labels and generalize well, yet how these behaviors coexist remains poorly understo

applicationsarxiv-cs-ai
19 May 2026
Model Releases

Usenix'23 Extended Version: Smart Learning to Find Dumb Contracts

DGX agent

arXiv:2304.10726v3 Announce Type: replace-cross Abstract: We introduce the Deep Learning Vulnerability Analyzer (DLVA) for Ethereum smart contracts based on neural networks. We train DLVA to judge byt

model-releasesarxiv-cs-lg
19 May 2026
Hardware

VeriCache: Turning Lossy KV Cache into Lossless LLM Inference

DGX agent

arXiv:2605.17613v1 Announce Type: cross Abstract: The large size of the KV cache has become a major bottleneck for serving LLMs with increasing context lengths. In response, many KV cache compression

hardwarearxiv-cs-lg
19 May 2026
Research

Vidya: An AI-Driven Modular Pipeline for Archival Automation and Semantic Metadata Enrichment

DGX agent

arXiv:2605.16338v1 Announce Type: cross Abstract: The large-scale digitization of historical archives has created a paradox: 'dark data'-digital objects lacking metadata for retrieval. Manual archival

researcharxiv-cs-cl
19 May 2026
Local Ai

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

DGX agent

arXiv:2605.18740v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) still struggle with fine-grained visual understanding, where answers often depend on small but decisive evide

local-aiarxiv-cs-ai
19 May 2026
Model Releases

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

DGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

model-releasesarxiv-cs-ai
19 May 2026
Safety

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search

DGX agent

arXiv:2605.16362v1 Announce Type: cross Abstract: Activation steering offers a lightweight way to control LLMs without retraining, but its effectiveness varies sharply across concepts. Prior work ofte

safetyarxiv-cs-ai
19 May 2026
Research

A Statistical Analysis for Per-Instance Evaluation of Stochastic Optimizers: Avoiding Unreliable Conclusions

DGX agent

arXiv:2503.16589v2 Announce Type: replace Abstract: A key trait of stochastic optimizers is that multiple runs of the same optimizer in attempting to solve the same problem can produce different resul

researcharxiv-cs-lg
18 May 2026
Research

A Unified Perturbation Framework for Analyzing Leaderboard Stability and Manipulation

DGX agent

arXiv:2605.15761v1 Announce Type: new Abstract: Evaluation leaderboards such as LMArena play a central role in benchmarking large language models by aggregating pairwise human preferences into model r

researcharxiv-cs-lg
18 May 2026
Agents

Access Timing as Scaffolding: A Reinforcement Learning Approach to GenAI in Education

DGX agent

arXiv:2605.15850v1 Announce Type: cross Abstract: In recent years, generative AI (GenAI) in educational settings has become ubiquitous in students' daily lives, despite its potential to induce over-re

agentsarxiv-cs-ai
18 May 2026
Model Releases

Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging

DGX agent

arXiv:2505.21698v3 Announce Type: replace Abstract: Vision-language foundation models achieve promising performance in natural image classification, yet their direct application to medical imaging is

model-releasesarxiv-cs-cv
18 May 2026
Local Ai

AgentStop: Terminating Local AI Agents Early to Save Energy in Consumer Devices

DGX agent

arXiv:2605.15206v1 Announce Type: cross Abstract: Autonomous agents powered by large language models (LLMs) are increasingly used to automate complex, multi-step tasks such as coding or web-based ques

local-aiarxiv-cs-ai
18 May 2026
Model Releases

Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction

DGX agent

arXiv:2605.16077v1 Announce Type: new Abstract: Accurate assessment of cognitive decline from spontaneous speech remains challenging due to limited dataset size and class imbalance. In this work, we p

model-releasesarxiv-cs-cl
18 May 2026
Applications

Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study

DGX agent

arXiv:2605.16011v1 Announce Type: cross Abstract: Adaptive learning refers to educational technologies that track learners' learning progress and adapt the instructional process based on individual le

applicationsarxiv-cs-ai
18 May 2026
Applications

Defining Cultural Capabilities for AI Evaluation: A Taxonomy Grounded in Intercultural Communication Theory

DGX agent

arXiv:2605.15990v1 Announce Type: new Abstract: Tremendous efforts have been put into evaluating the inclusivity and effectiveness of AI systems across cultures. However, the cultural capabilities con

applicationsarxiv-cs-cl
18 May 2026
Model Releases

DexJoCo: A Benchmark and Toolkit for Task-Oriented Dexterous Manipulation on MuJoCo

DGX agent

arXiv:2605.16257v1 Announce Type: new Abstract: Achieving human-level manipulation requires dexterous robotic hands capable of complex object interactions. Advancing such capabilities further demands

model-releasesarxiv-cs-ro
18 May 2026
Safety

Dynamic Plasma Shape Control with Arbitrary Sensor Subsets

DGX agent

arXiv:2605.15935v1 Announce Type: new Abstract: Plasma shape control in tokamaks requires a real-time controller that tracks dynamically changing shape targets while tolerating diagnostic failures. Cl

safetyarxiv-cs-ro
18 May 2026
Research

EndoGSim: Physics-Aware 4D Dynamic Endoscopic Scene Simulations via MLLM-Guided Gaussian Splatting

DGX agent

arXiv:2605.16022v1 Announce Type: new Abstract: In robot-assisted minimally invasive surgery, high-fidelity dynamic endoscopic scene reconstruction and simulation are crucial to enhancing downstream t

researcharxiv-cs-cv
18 May 2026
Safety

Eskwai for Students: Generative AI Assistant for Legal Education in Ghana

DGX agent

arXiv:2605.15380v1 Announce Type: new Abstract: Recent advances in generative AI have shown their potential to be leveraged for legal education. Yet, work on the development and deployment of such sys

safetyarxiv-cs-cl
18 May 2026
Model Releases

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

DGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

From Layers to Networks: Comparing Neural Representations via Diffusion Geometry

DGX agent

arXiv:2605.15901v1 Announce Type: new Abstract: Diffusion geometry is a manifold learning framework that uses random walks defined by Markov transition matrices to characterize the geometry of a datas

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Frontier Large Language Models Rival State-of-the-Art Planners

DGX agent

arXiv:2511.09378v2 Announce Type: replace Abstract: A series of influential studies established that large language models cannot reliably solve even simple planning tasks. We show that the latest gen

model-releasesarxiv-cs-ai
18 May 2026
Safety

GAP: Geometric Anchor Pre-training for Data-Efficient Visuomotor Learning of Manipulation Tasks

DGX agent

arXiv:2605.15836v1 Announce Type: cross Abstract: Learning visuomotor policies from scarce expert demonstrations remains a core challenge in robotic manipulation. A primary hurdle lies in distilling h

safetyarxiv-cs-ai
18 May 2026
Model Releases

GESD: Beyond Outcome-Oriented Fairness

DGX agent

arXiv:2605.15295v1 Announce Type: cross Abstract: Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

HAI-Eval: Measuring Human-AI Synergy in Collaborative Coding

DGX agent

arXiv:2512.04111v2 Announce Type: replace-cross Abstract: LLM-powered coding agents are reshaping the development paradigm. However, existing evaluation systems, neither traditional tests for humans n

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

LASER: Language Model Regression for Semi-Structured Workflow Resource and Runtime Estimation

DGX agent

arXiv:2512.19701v2 Announce Type: replace-cross Abstract: Accurate prediction of resource consumption and runtime for cloud workflow jobs is critical for scheduling efficiency, yet remains challenging

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

DGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

model-releasesarxiv-cs-cv
18 May 2026
Agents

NIMO Controller: a self-driving laboratory orchestrator based on the Model Context Protocol

DGX agent

arXiv:2605.15227v1 Announce Type: new Abstract: Self-driving laboratories (SDLs) have attracted increasing attention as a means of accelerating scientific discovery; however, developing SDL software r

agentsarxiv-cs-ai
18 May 2026
Safety

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

DGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

safetyarxiv-cs-lg
18 May 2026
Agents

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

DGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

agentsarxiv-cs-ai
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Tutorials

Simultaneous State Estimation and Online Model Learning in a Soft Robotic System

DGX agent

arXiv:2602.14092v2 Announce Type: replace-cross Abstract: Operating complex real-world systems, such as soft robots, can benefit from precise predictive control schemes that require accurate state and

tutorialsarxiv-cs-ro
18 May 2026
Safety

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

DGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

safetyarxiv-cs-lg
18 May 2026
← Previous
1…8788899091…109
Next →