AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,013 results
20 May 2026

Guiding Neuro-Symbolic Scenario Generation with Spatio-Temporal Logic

SafetyDGX agent

arXiv:2605.19038v1 Announce Type: cross Abstract: The rapid advancement of autonomous driving (AD) technologies has outpaced the development of robust safety evaluation methods. Conventional testing r

Local LLM - privacy first - doctor

Local AiDGX agent

A discussion from the r/ollama community about using local large language models for privacy-sensitive applications, particularly for handling sensitive documents like medical records. Local LLMs proc

LTX 2.3 IC LoRAs (LipDub and Video Utilities) https://x.com/i/broadcasts/1qKDzPZeLjVJV

Local AiDGX agent

LTX 2.3 IC introduces specialized LoRA (Low-Rank Adaptation) modules for video generation, including LipDub functionality for synchronized lip movements and Video Utilities for enhanced video processi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling

Model ReleasesDGX agent

arXiv:2605.18838v1 Announce Type: cross Abstract: Scaling laws predict loss from compute but not how capabilities interact. We measure the coupling between reasoning and truthfulness across 63 base mo

Managed Agents through the Gemini API is @GoogleAI's response to Anthropic Managed Agents Since it's powered by the new Antigravity agent bu…

Model ReleasesDGX agent

Managed Agents through the Gemini API is @GoogleAI's response to Anthropic Managed Agents Since it's powered by the new Antigravity agent built on Gemini 3.5 Flash, it is the most cost-effective gener

Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges

Model ReleasesDGX agent

arXiv:2605.19723v1 Announce Type: cross Abstract: Mathematical reasoning is essential for problem-solving in education, science, and industry, serving as a crucial benchmark for evaluating artificial

Mechanisms of Object Localization in Vision-Language Models

Local AiDGX agent

arXiv:2605.19792v1 Announce Type: new Abstract: Visually-grounded language models (VLMs) are highly effective in linking visual and textual information, yet they often struggle with basic classificati

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

Model ReleasesDGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

Model ReleasesDGX agent

arXiv:2605.18825v1 Announce Type: new Abstract: Prefix caching is a key optimization in Large Language Model (LLM) serving, reusing attention Key-Value (KV) states across requests with shared prompt p

OpenCompass: A Universal Evaluation Platform for Large Language Models

Model ReleasesDGX agent

arXiv:2605.19276v1 Announce Type: new Abstract: In recent years, the field of artificial intelligence has undergone a paradigm shift from task-specific small-scale models to general-purpose large lang

OpenComputer: Verifiable Software Worlds for Computer-Use Agents

ResearchDGX agent

arXiv:2605.19769v1 Announce Type: new Abstract: We present OpenComputer, a verifier-grounded framework for constructing verifiable software worlds for computer-use agents. OpenComputer integrates four

PASC: Pipeline-Aware Conformal Prediction with Joint Coverage Guarantees for Multi-Stage NLP and LLM Pipelines

AgentsDGX agent

arXiv:2605.18812v1 Announce Type: cross Abstract: Modern NLP and LLM systems are pipelines: named entity recognition (NER) -> entity disambiguation (NED) -> entity typing, retrieval-augmented generati

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

Model ReleasesDGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

Model ReleasesDGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

Robust Basis Spline Decoupling for the Compression of Transformer Models

Model ReleasesDGX agent

arXiv:2605.18794v1 Announce Type: cross Abstract: Decoupling is a powerful modeling paradigm for representing multivariate functions as compositions of linear transformations and univariate nonlinear

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

AgentsDGX agent

arXiv:2605.18857v1 Announce Type: cross Abstract: For most of the history of information retrieval (IR), search results were designed for human consumers who could scan, filter, and discard irrelevant

Toward an AI-Powered Computational Testbed for Workforce Policy

SafetyDGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

Towards Discovery of Polymers for Insulin Delivery via Physics-Grounded Agentic Workflows

AgentsDGX agent

arXiv:2605.18831v1 Announce Type: cross Abstract: Cold-chain storage limits access to insulin for hundreds of millions of people; a thermally protective patch polymer could help, but the design space

Transformers Linearly Represent Highly Structured World Models

ResearchDGX agent

arXiv:2605.18847v1 Announce Type: cross Abstract: Do transformers, when trained on sequential reasoning traces, build internal models of the underlying task? And if so, does the structure of those int

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation

TutorialsDGX agent

arXiv:2602.09023v4 Announce Type: replace Abstract: Despite strong generalization capabilities, Vision-Language-Action (VLA) models remain constrained by the high cost of expert demonstrations and lim

Urban Outfitters achieves major cost savings by moving Sterling OMS to AlloyDB for PostgreSQL

TutorialsDGX agent

Editor’s note: Urban Outfitters, Inc. (URBN) recently completed a major infrastructure upgrade, migrating its IBM Sterling Order Management System (Sterling OMS) from an Oracle database to Google Clou

ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins

Model ReleasesDGX agent

arXiv:2603.06740v2 Announce Type: replace-cross Abstract: Protein language models (pLMs) have shown strong potential for zero-shot prediction of missense variant effects, yet systematic benchmarking o

What is the best image anime upscaler currently?

Local AiDGX agent

LetsEnhance with its Digital Art model is considered best for quality, delivering sharp lines and clearest eye detail with the highest output resolution ceiling. For a free desktop option, Upscayl off

Where Does Authorship Signal Emerge in Encoder-Based Language Models?

ResearchDGX agent

arXiv:2605.19908v1 Announce Type: new Abstract: Authorship attribution models fine-tuned with the same pretrained encoder, data, and loss can differ four-fold in performance depending only on their sc

WisdomAI’s new analytics agents go beyond insights, automating business work through autonomous action

AgentsDGX agent

WisdomAI Inc., creator of an artificial intelligence-native business intelligence platform, is jumping on the agentic AI bandwagon with the latest update to its flagship Federated Agentic Intelligence

XNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception

Model ReleasesDGX agent

arXiv:2603.22453v2 Announce Type: replace Abstract: Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance

19 May 2026

3D Densification for Multi-Map Monocular VSLAM in Endoscopy

ResearchDGX agent

arXiv:2503.14346v3 Announce Type: replace Abstract: Multi-map Sparse Monocular visual Simultaneous Localization and Mapping applied to monocular endoscopic sequences has proven efficient to robustly r

A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks

ResearchDGX agent

arXiv:2512.04329v2 Announce Type: replace Abstract: Reusing existing neural-network components is central to research efficiency, yet discovering, extracting, and validating such modules across thousa

A Unified Framework for Structured Flow Modeling: From Continuous Fields to Data-Driven Representations

ResearchDGX agent

arXiv:2605.18250v1 Announce Type: cross Abstract: Many dynamical systems can be described in terms of structured flows combining source/sink behavior, cyclic dynamics, and topology-constrained transpo

ADR: An Agentic Detection System for Enterprise Agentic AI Security

Model ReleasesDGX agent

arXiv:2605.17380v1 Announce Type: new Abstract: We present the Agentic AI Detection and Response (ADR) system, the first large-scale, production-proven enterprise framework for securing AI agents oper

Adversarial Agent Collaboration for Correctness Improvements of C to Safe Rust Translation

Model ReleasesDGX agent

arXiv:2510.03879v3 Announce Type: replace-cross Abstract: Translating C to memory-safe languages, like Rust, prevents critical memory safety vulnerabilities that are prevalent in legacy C software. Ev

Agentic AI Governance and Lifecycle Management in Healthcare

Model ReleasesDGX agent

arXiv:2601.15630v2 Announce Type: replace Abstract: Healthcare organizations are beginning to embed agentic AI into routine workflows, including clinical documentation support and early-warning monito

Agentmw: Open-source middleware for AI agents — catches mid-run failures,compresses stale context, and grows a reasoning library across runs. Any model, any framework.

Local AiDGX agent

Agentmw is an open-source middleware framework designed to enhance AI agent reliability and efficiency across different models and frameworks. It addresses key operational challenges including mid-run

ALIGN: A Vision-Language Framework for High-Accuracy Accident Location Inference through Geo-Spatial Neural Reasoning

Local AiDGX agent

arXiv:2511.06316v3 Announce Type: replace Abstract: In low- and middle-income countries, public safety and urban planning initiatives frequently face a critical shortage of accurate, location-specific

AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment

SafetyDGX agent

arXiv:2605.18529v1 Announce Type: new Abstract: The alignment of Large Language Models (LLMs) for complex reasoning heavily relies on Reinforcement Learning with Verifiable Rewards (RLVR). However, st

Announcing Claude Managed Agents on Cloudflare

Model ReleasesDGX agent

Cloudflare has integrated with Anthropic's Claude Managed Agents to provide a fast, isolated execution environment for autonomous code delivery. This means builders can scale agent workflows globally

Are Researchers Being Replaced by Artificial Intelligence?

ResearchDGX agent

arXiv:2605.16294v1 Announce Type: cross Abstract: A Nature survey from 2023 involving 1,600 researchers shows that scientists are ``concerned, as well as excited, by the increasing use of artificial-i

As a Soviet historian who has spent years writing about the extreme, repressive control Soviet Communism exercised over its unfortunate citi…

SafetyDGX agent

As a Soviet historian who has spent years writing about the extreme, repressive control Soviet Communism exercised over its unfortunate citizens, I find it really hard to bring a similar accusation ag

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

Model ReleasesDGX agent

arXiv:2605.17324v1 Announce Type: cross Abstract: Clarification-seeking behavior is widely regarded as a desirable property of LLM agents, enabling them to resolve ambiguity before acting on underspec

BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting

Model ReleasesDGX agent

arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers and limited scalability. While

Benchmarking Mythos-Linked Bug Rediscovery

Model ReleasesDGX agent

arXiv:2605.17416v1 Announce Type: cross Abstract: Anthropic's April 2026 Mythos materials combine benchmark claims with concrete bug-finding stories across OpenBSD, FreeBSD, Linux, FFmpeg, and browser

Beyond Morphology: Quantifying the Diagnostic Power of Color Features in Cancer Classification

ResearchDGX agent

arXiv:2605.18522v1 Announce Type: cross Abstract: In histopathology, human experts primarily rely on color as a means of enhancing contrast to interpret tissue morphology, whereas machine vision model

Bridging the Gap: Converting Read Text to Conversational Dialogue

ResearchDGX agent

arXiv:2605.18001v1 Announce Type: new Abstract: In recent advancements within speech processing, converting read speech to conversational speech has gained significant attention. The primary challenge

Building Reliable Arithmetic Multipliers Under NBTI Aging and Process Variations

SafetyDGX agent

arXiv:2605.18444v1 Announce Type: cross Abstract: Hardware aging poses a significant challenge for integrated circuits (ICs), leading to performance degradation and eventual failure. In this work, we

Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows

Model ReleasesDGX agent

arXiv:2605.18327v1 Announce Type: new Abstract: AI agents deployed into SRE workflows currently derive their understanding of environment state from raw observability telemetry at query time, paying a

Code as Agent Harness

SafetyDGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

Convex Dataset Valuation for Post-Training

SafetyDGX agent

arXiv:2605.16704v1 Announce Type: new Abstract: Improving LLM performance on downstream tasks sometimes requires leveraging auxiliary datasets during post-training. In practice, however, developers fa

CVE-Factory: Scaling Expert-Level Agentic Tasks for Code Security Vulnerability

Model ReleasesDGX agent

arXiv:2602.03012v2 Announce Type: replace-cross Abstract: Evaluating and improving the security capabilities of code agents requires high-quality, executable vulnerability tasks. However, existing wor

DBES: A Systematic Benchmark and Metric Suite for Evaluating Expert Specialization in Large-Scale MoEs

Model ReleasesDGX agent

arXiv:2605.18498v1 Announce Type: cross Abstract: Expert specialization in Mixture-of-Experts (MoE) models remains poorly understood, with traditional evaluations conflating architectural load-balanci

EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis

Model ReleasesDGX agent

arXiv:2508.07292v3 Announce Type: replace Abstract: Endoscopic diagnosis is an iterative process in which clinicians progressively acquire, compare, and verify local visual evidence before reaching a

Estimating Item Difficulty with Large Language Models as Experts

SafetyDGX agent

arXiv:2605.18562v1 Announce Type: cross Abstract: Accurate estimates of item difficulty are essential for valid assessment and effective adaptive learning. However, for newly created tasks, response d

Evaluating Cognitive Age Alignment in Interactive AI Agents

Model ReleasesDGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

Event-Grounded Sparse Autoencoders for Vision-Language-Action Policies

SafetyDGX agent

arXiv:2605.17204v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies translate language and visual inputs into robot actions, where their hidden representations directly shape close

Extending conversational memory in Kiro CLI using Amazon Bedrock AgentCore Memory

AgentsDGX agent

In this post, we demonstrate how you can extend the conversational memory of Kiro CLI by implementing a custom Model Context Protocol (MCP) server that integrates with Amazon Bedrock AgentCore Memory.

From months of work at ILM to a single compositor in ComfyUI. That's not a workflow change. That's a shift in who gets to do this work. The …

Local AiDGX agent

From months of work at ILM to a single compositor in ComfyUI. That's not a workflow change. That's a shift in who gets to do this work. The @ActionVFX team just launched Advanced AI Workflows for VFX

Generalized Functional ANOVA in Closed-Form: A Unified View of Additive Explanations

ResearchDGX agent

arXiv:2605.18422v1 Announce Type: cross Abstract: The functional ANOVA, or Hoeffding decomposition, provides a principled framework for interpretability by decomposing a model prediction into main eff

Generative Artificial Intelligence for Literature Reviews

Model ReleasesDGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

Geometry-Aware Surrogate for Real-Time Hydrodynamics Estimation of Autonomous Ground Vehicles in Amphibious Environments

AgentsDGX agent

arXiv:2605.18543v1 Announce Type: new Abstract: Autonomous ground vehicles operating in shallow water or flood-prone terrains require dynamic models that account for hydrodynamic forces. However, the

Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app (Sanuj Bhatia/Android Central)

Model ReleasesDGX agent

Sanuj Bhatia / Android Central: Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app — Seriously, who is asking for

Google wants to compete with Anthropic’s Mythos

AgentsDGX agent

Google is making a big push into cybersecurity. At I/O, the company announced that it was inviting select groups of experts to test the API for CodeMender, an 'AI agent for code security' it debuted l

← Previous
1…138139140141142…167
Next →