AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
Model Releases

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration

DGX agent

arXiv:2605.24636v1 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios re

model-releasesarxiv-cs-ai
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework

DGX agent

arXiv:2507.19219v2 Announce Type: replace Abstract: Overestimation in evaluating large language models (LLMs) has become an increasing concern. Due to the contamination of public benchmarks or imbalan

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

How we evolved Google’s global and data center networks for the AI era

DGX agent

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

model-releasesgoogle-cloud-ai
26 May 2026
Safety

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications

DGX agent

arXiv:2605.24883v1 Announce Type: new Abstract: The widespread integration of Large Language Models (LLMs) necessitates rigorous and systematic safety evaluation. Existing paradigms either rely on con

safetyarxiv-cs-ai
26 May 2026
Agents

Iterate Until Retrieved: Factual Nugget Optimization for Discoverable Continual Corrections in Agentic RAG

DGX agent

arXiv:2605.25641v1 Announce Type: new Abstract: Agentic retrieval-augmented generation (RAG) systems in complex B2B (business-to-business) settings may often receive free-form response feedback. Rathe

agentsarxiv-cs-cl
26 May 2026
Safety

IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning

DGX agent

arXiv:2605.23997v1 Announce Type: cross Abstract: Multimodal large language models via reinforcement learning (RL) have demonstrated remarkable capabilities in complex visual reasoning tasks, yet they

safetyarxiv-cs-ai
26 May 2026
Applications

Lake Detection and Water Quality Estimation in Sentinel-2 Data

DGX agent

arXiv:2605.24515v1 Announce Type: new Abstract: With climate change and increasing human pressure on natural landscapes, inland water resources are becoming progressively scarcer, more vulnerable, and

applicationsarxiv-cs-lg
26 May 2026
Tutorials

Leveraging Spreading Activation for Improved Document Retrieval in Knowledge-Graph-Based RAG Systems

DGX agent

arXiv:2512.15922v3 Announce Type: replace Abstract: Despite initial successes and a variety of architectures, retrieval-augmented generation systems still struggle to reliably retrieve and connect the

tutorialsarxiv-cs-ai
26 May 2026
Model Releases

Memory-Induced Tool-Drift in LLM Agents

DGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

model-releasesarxiv-cs-lg
26 May 2026
Agents

Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report

DGX agent

arXiv:2605.25665v1 Announce Type: cross Abstract: AI-native software development is often evaluated at the level of individual models, prompts, or generated artifacts. This framing is insufficient for

agentsarxiv-cs-ai
26 May 2026
Research

Methodology for Creating a Clinically Verified Dermoscopic Image Dataset

DGX agent

arXiv:2605.25168v1 Announce Type: cross Abstract: This study presents a methodology for constructing a clinically verified dataset of dermatoscopic images for medical informatics research. The relevan

researcharxiv-cs-ai
26 May 2026
Model Releases

MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence

DGX agent

arXiv:2505.23764v3 Announce Type: replace-cross Abstract: Spatial intelligence is essential for multimodal large language models (MLLMs) operating in the complex physical world. Existing benchmarks, h

model-releasesarxiv-cs-cl
26 May 2026
Safety

Multi-Agent Coordination Adaptation via Structure-Guided Orchestration

DGX agent

arXiv:2605.25746v1 Announce Type: cross Abstract: As large language model (LLM)-based multi-agent systems scale to handle increasingly complex tasks, balancing structural stability and dynamic adaptab

safetyarxiv-cs-ai
26 May 2026
Safety

PrivFusion: A Privacy-preserving Multi-Agent Framework for Harmonizing Distributed Datasets

DGX agent

arXiv:2605.24249v1 Announce Type: new Abstract: The growing availability of clinical data has increased the use of machine learning, yet centralized data aggregation is often infeasible for sensitive

safetyarxiv-cs-lg
26 May 2026
Safety

ProActor: Timing-Aware Reinforcement Learning for Proactive Task Scheduling Agents

DGX agent

arXiv:2605.24900v1 Announce Type: new Abstract: Proactive task-oriented agents must autonomously anticipate user needs, identify actionable opportunities, and trigger software actions at appropriate m

safetyarxiv-cs-ai
26 May 2026
Tutorials

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis

DGX agent

arXiv:2601.06870v2 Announce Type: replace-cross Abstract: Multimodal large language models have demonstrated strong ability in capturing semantic representations for multimodal sentiment analysis. The

tutorialsarxiv-cs-ai
26 May 2026
Local Ai

Retrieval-Augmented Detection of Potentially Abusive Clauses in Chilean Terms of Service

DGX agent

arXiv:2605.26019v1 Announce Type: cross Abstract: Online Terms of Service often function as contracts of adhesion, creating asymmetries that may expose consumers to potentially abusive clauses. In Chi

local-aiarxiv-cs-ai
26 May 2026
Model Releases

StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs

DGX agent

arXiv:2605.25534v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at structural reasoning yet suffer from a sharp logical brittleness in structural consistency. We term th

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

the basic trick to using Claude Code for non-technical work is to put a bunch of files in a folder and tell it can write scripts + make HTML

DGX agent

Claude Code can be used for non-technical work by organizing files in a folder and instructing it to write scripts and create HTML documents. This approach allows users without programming expertise t

model-releasesthariq--x
26 May 2026
Model Releases

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

DGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

model-releasesarxiv-cs-ai
26 May 2026
Safety

The OpenAI insider @thsottiaux has a warning for everyone offloading their thinking to agents.

DGX agent

An OpenAI insider (@thsottiaux) raises concerns about the risks of over-relying on AI agents to handle cognitive tasks, warning against wholesale delegation of thinking to autonomous systems. The warn

safetygary-marcus--x
26 May 2026
Model Releases

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

DGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

model-releasesarxiv-cs-lg
26 May 2026
Safety

TopoAlign: Topology-Aware Visual Representation Alignment

DGX agent

arXiv:2605.25541v1 Announce Type: cross Abstract: Neural networks encode inputs as high-dimensional vectors, known as representations, that capture how models process data by encoding task-relevant st

safetyarxiv-cs-ai
26 May 2026
Research

When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges

DGX agent

arXiv:2605.26046v1 Announce Type: cross Abstract: Customizing an LLM judge to a specific task or domain often involves optimizing its prompt across multiple evaluation criteria simultaneously. Textual

researcharxiv-cs-ai
26 May 2026
Research

Advanced AI Service Provisioning in O-RAN through LLM Engine Integration

DGX agent

arXiv:2605.23809v1 Announce Type: cross Abstract: The Open Radio Access Network (O-RAN) architecture allows AI to be embedded directly into the RAN through modular xApps and rApps, yet creating these

researcharxiv-cs-lg
25 May 2026
Model Releases

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

DGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models

DGX agent

arXiv:2605.23694v1 Announce Type: new Abstract: Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As

model-releasesarxiv-cs-cl
25 May 2026
Safety

Controlled Personalization in Legacy Media Online Services: A Case Study in News Recommendation

DGX agent

arXiv:2510.09136v2 Announce Type: replace-cross Abstract: Personalized news recommendations have become a standard feature of large news aggregation services, optimizing user engagement through automa

safetyarxiv-cs-ai
25 May 2026
Safety

EquiSumm : A Gender Bias-Aware Framework for Inclusive Tweet Summarization

DGX agent

arXiv:2605.23412v1 Announce Type: new Abstract: While social media platforms, such as Twitter, provide a medium for large-scale opinion sharing during news events, it is manually impossible for indivi

safetyarxiv-cs-cl
25 May 2026
Agents

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

DGX agent

arXiv:2605.23271v1 Announce Type: cross Abstract: The rapid evolution of generative video foundation models has propelled the field toward professional-grade cinematic synthesis. To achieve such deman

agentsarxiv-cs-ai
25 May 2026
Local Ai

From Activation to Causality: Discovery of Causal Visual Representations in the Human Brain

DGX agent

arXiv:2605.23895v1 Announce Type: new Abstract: Identifying which brain regions represent a visual concept in the human brain is a central challenge in neuroscience. Existing approaches have localized

local-aiarxiv-cs-cv
25 May 2026
Tutorials

GlyTwin: Digital Twin for Glucose Control in Type 1 Diabetes Through Optimal Behavioral Modifications Using Patient-Centric Counterfactuals

DGX agent

arXiv:2504.09846v2 Announce Type: replace-cross Abstract: Frequent and long-term exposure to hyperglycemia increases the risk of chronic complications, including neuropathy, nephropathy, and cardiovas

tutorialsarxiv-cs-ai
25 May 2026
Model Releases

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

DGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures

DGX agent

arXiv:2511.03882v2 Announce Type: replace-cross Abstract: Imitation learning-based robot control policies are enjoying renewed interest in video-based robotics. However, it remains unclear whether thi

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

LLM-driven design of physics-constrained constitutive models: two agents are better than one

DGX agent

arXiv:2605.23754v1 Announce Type: new Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics

model-releasesarxiv-cs-lg
25 May 2026
Research

Machine learning applied to emerald gemstone grading: framework proposal and creation of a public dataset

DGX agent

arXiv:2605.23777v1 Announce Type: new Abstract: The grading of gemstones is currently a manual procedure performed by gemologists. A popular approach uses reference stones, where those are visually in

researcharxiv-cs-cv
25 May 2026
Safety

MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems

DGX agent

arXiv:2602.04431v2 Announce Type: replace Abstract: LLM-based multi-agent systems have demonstrated impressive capabilities, but they also introduce significant safety risks when individual agents fai

safetyarxiv-cs-lg
25 May 2026
Applications

MedSAE: Dissecting MedCLIP Representations with Sparse Autoencoders

DGX agent

arXiv:2510.26411v2 Announce Type: replace Abstract: Artificial intelligence in healthcare requires models that are accurate and interpretable. We advance mechanistic interpretability in medical vision

applicationsarxiv-cs-ai
25 May 2026
Industry

My favorite prompt: a) make a plan for <task> b) orchestrate and launch sub-agents to execute the plan c) validate the results from the sub-…

DGX agent

My favorite prompt: a) make a plan for <task> b) orchestrate and launch sub-agents to execute the plan c) validate the results from the sub-agents d) repeat b and c until you finish the plan Grok Buil

industryelon-musk--x
25 May 2026
Agents

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

DGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

agentsarxiv-cs-ai
25 May 2026
Tools

Notes on Pope Leo XIV's encyclical on AI

DGX agent

Dropped this morning by the Vatican: Magnifica Humanitas of His Holiness Pope Leo XIV on Safeguarding the Human Person in the Time of Artificial Intelligence. This is a very interesting document. It's

toolssimon-willison
25 May 2026
Model Releases

Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems

DGX agent

arXiv:2605.23297v1 Announce Type: new Abstract: AI-enabled services deployed in critical digital infrastructure are subject to governance obligations spanning transparency, accountability, fairness, a

model-releasesarxiv-cs-ai
25 May 2026
Research

USIM and U0: A Vision-Language-Action Dataset and Model for General Underwater Robots

DGX agent

arXiv:2510.07869v4 Announce Type: replace Abstract: Underwater environments pose unique challenges for robotic navigation and manipulation. While existing research has primarily focused on task-specif

researcharxiv-cs-ro
25 May 2026
Industry

We've expanded the Beta to many more people. Go to https://x.ai/cli to give it a try. Our entire engineering team will be responding to (and…

DGX agent

We've expanded the Beta to many more people. Go to https://x.ai/cli to give it a try. Our entire engineering team will be responding to (and fixing) any issues you encounter, so please share feedback

industryelon-musk--x
25 May 2026
Applications

A Reproducible Log-Driven AutoML Framework for Interpretable Pipeline Optimization in Healthcare Risk Prediction

DGX agent

arXiv:2605.21528v1 Announce Type: new Abstract: Accurate and reproducible disease risk prediction remains challenging due to heterogeneous features, limited samples, and severe class imbalance. This s

applicationsarxiv-cs-lg
23 May 2026
Model Releases

AutoBaxBuilder: Bootstrapping Code Security Benchmarking

DGX agent

arXiv:2512.21132v2 Announce Type: replace-cross Abstract: As large language models (LLMs) see wide adoption in software engineering, the reliable assessment of the correctness and security of LLM-gene

model-releasesarxiv-cs-lg
23 May 2026
Hardware

AutoMCU: Feasibility-First MCU Neural Network Customization via LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.21560v1 Announce Type: new Abstract: Deploying neural networks on microcontroller units (MCUs) is critical for edge intelligence but remains challenging due to tight memory, storage, and co

hardwarearxiv-cs-lg
23 May 2026
Safety

MambaGaze: Bidirectional Mamba with Explicit Missing Data Modeling for Cognitive Load Assessment from Eye-Gaze Tracking Data

DGX agent

arXiv:2605.22775v1 Announce Type: new Abstract: Real-time cognitive load assessment from eye-tracking signals could potentially enable adaptive human-centered-AI such as safety-critical applications s

safetyarxiv-cs-lg
23 May 2026
← Previous
1…8081828384…104
Next →