AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
19 May 2026

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey

Model ReleasesDGX agent

arXiv:2409.10102v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has quickly grown into a pivotal paradigm in the development of Large Language Models (LLMs). Although ex

Try Anthropic Claude in Comfy today: https://links.comfy.org/4eWCvFk

Model ReleasesDGX agent

ComfyUI announced the availability of Anthropic's Claude AI model integrated into their platform, allowing users to utilize Claude's capabilities within the ComfyUI interface. The announcement include

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16638v1 Announce Type: new Abstract: Recent research has demonstrated that Universal Multimodal Embedding (UME) benefits significantly from Chain-of-Thought (CoT) reasoning. In this paradig

TusoAI: Agentic Optimization for Scientific Methods

Model ReleasesDGX agent

arXiv:2509.23986v2 Announce Type: replace Abstract: Scientific discovery is often slowed by the manual development of computational tools needed to analyze complex experimental data. Building such too

UAVFF3D: A Geometry-Aware Benchmark for Feed-Forward UAV 3D Reconstruction

Model ReleasesDGX agent

arXiv:2605.17942v1 Announce Type: new Abstract: Feed-forward 3D reconstruction has recently demonstrated strong generalization across diverse scenes, yet its performance in UAV imagery remains underex

UbuntuGuard: A Culturally-Grounded Policy Benchmark for Equitable AI Safety in African Languages

Model ReleasesDGX agent

arXiv:2601.12696v2 Announce Type: replace Abstract: Current guardian models are predominantly Western-centric and optimized for high-resource languages, leaving low-resource African languages vulnerab

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

Model ReleasesDGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

UniER: A Unified Benchmark for Item-level and Path-level Exercise Recommendation

Model ReleasesDGX agent

arXiv:2605.16750v1 Announce Type: cross Abstract: Personalized exercise recommendation dynamically aligns pedagogical resources with individual knowledge mastery, which is crucial for satisfying stude

UniPPTBench: A Unified Benchmark for Presentation Generation Across Diverse Input Settings

Model ReleasesDGX agent

arXiv:2605.17356v1 Announce Type: new Abstract: Existing works typically focus on presentation generation under isolated input settings, whereas real-world use cases span diverse scenarios, including

Universal Dynamics of Punctuated Progress

Model ReleasesDGX agent

arXiv:2605.16719v1 Announce Type: cross Abstract: Scientific and technological frontiers advance through punctuated dynamics, yet the principles governing these dynamics remain poorly understood. Here

Usenix'23 Extended Version: Smart Learning to Find Dumb Contracts

Model ReleasesDGX agent

arXiv:2304.10726v3 Announce Type: replace-cross Abstract: We introduce the Deep Learning Vulnerability Analyzer (DLVA) for Ethereum smart contracts based on neural networks. We train DLVA to judge byt

UTOPYA: A Multimodal Deep Learning Framework for Physics-Informed Anomaly Detection and Time-Series Prediction

Model ReleasesDGX agent

arXiv:2605.18188v1 Announce Type: new Abstract: Anomaly detection in batch processes is hindered by transient dynamics, scarce fault labels, and reliance on single-modality sensor data. This work intr

UVTran: Accurate Hole-Filling Parameterization with Transformers

Model ReleasesDGX agent

arXiv:2605.16306v1 Announce Type: cross Abstract: In industrial design, N-sided hole filling is typically formulated as the construction of a single trimmed B-spline surface by minimizing a fairness e

Validate Your Authority: Benchmarking LLMs on Multi-Label Precedent Treatment Classification

Model ReleasesDGX agent

arXiv:2605.17691v1 Announce Type: cross Abstract: Automating the classification of negative treatment in legal precedent is a critical yet nuanced NLP task where misclassification carries significant

Verifier-Guided Code Translation via Meta-Step Decoding

Model ReleasesDGX agent

arXiv:2605.17626v1 Announce Type: new Abstract: Test-time scaling is an important mechanism for improving large language models, especially on tasks with deterministic verifiers. Code translation is a

Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study

Model ReleasesDGX agent

arXiv:2605.17998v1 Announce Type: cross Abstract: As multi-agent systems move from short interactions to tool-using workflows with specialized roles and persistent state, completion becomes a runtime-

VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.17467v1 Announce Type: new Abstract: Large language model-driven multi-agent systems (LLM-MAS) excel at complex tasks, yet unreliable agents remain a key bottleneck to system-level reliabil

VGGT-CD: Training-Free Robust Registration for 3D Change Detection

Model ReleasesDGX agent

arXiv:2605.16859v1 Announce Type: cross Abstract: 3D change detection from multi-view images is essential for urban monitoring, disaster assessment, and autonomous driving. However, existing methods p

VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction

Model ReleasesDGX agent

arXiv:2605.16911v1 Announce Type: new Abstract: 3D semantic occupancy prediction requires accurate 2D-to-3D feature lifting, yet current methods restrict camera geometry to initial projections. Subseq

[video] why we need a new continuity layer for long-running agents (claude did this video! all except the voice which was @elevenlabs)

Model ReleasesDGX agent

This video discusses the architectural need for a continuity layer in long-running AI agents, explaining how agents require persistent memory and state management mechanisms to maintain coherence acro

Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.18160v1 Announce Type: cross Abstract: In recent years, multimodal large language models (MLLMs) have achieved remarkable progress, primarily attributed to effective paradigms for integrati

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

Model ReleasesDGX agent

arXiv:2602.04802v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have achieved impressive performance in cross-modal understanding across textual and visual inputs, yet existing bench

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

Model ReleasesDGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

Model ReleasesDGX agent

arXiv:2605.18172v1 Announce Type: new Abstract: Leveraging the universal representations of pre-trained LLMs and MLLMs offers a promising path toward brain foundation models. However, visually-evoked

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.18313v1 Announce Type: cross Abstract: Small vision-language models (2-8B) are well-suited for clin- ical deployment due to privacy constraints, limited connectivity, and low-latency requir

Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning

Model ReleasesDGX agent

arXiv:2601.06943v2 Announce Type: replace-cross Abstract: In real-world video question answering scenarios, videos often provide only localized visual cues, while verifiable answers are distributed ac

WavFlow: Audio Generation in Waveform Space

Model ReleasesDGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong

Model ReleasesDGX agent

arXiv:2509.22510v3 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) is the ability to satisfy desired objectives during generation, which is critical for trustworthy deployme

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories…

Model ReleasesDGX agent

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories, memorable moments, and many, many (occasionally embarrassi

WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games

Model ReleasesDGX agent

arXiv:2605.17637v1 Announce Type: new Abstract: Coding agents are increasingly used as application builders, yet many evaluations still focus on source code, repository-level tests, or intermediate tr

WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale

Model ReleasesDGX agent

arXiv:2510.16252v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for web agents demands environments that are both effective for evaluation and efficient enough for large-scale on

Weighted Flow Matching and Physics-Informed Nonlinear Filtering for Parameter Estimation in Digital Twins

Model ReleasesDGX agent

arXiv:2605.17146v1 Announce Type: cross Abstract: Digital twins (DTs) rely on continuous synchronization between physical systems and their virtual counterparts through online parameter estimation und

WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing

Model ReleasesDGX agent

arXiv:2510.15221v2 Announce Type: replace Abstract: Affective computing has matured rapidly in laboratory settings, yet no prior dataset combines (i) months-to-years of duration, (ii) a naturalistic w

We’re adding new ways for people to identify AI-generated images and understand where they came from. In addition to C2PA Content Credential…

Model ReleasesDGX agent

We’re adding new ways for people to identify AI-generated images and understand where they came from. In addition to C2PA Content Credentials, images now also contain a SynthID watermark, and can be i

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and …

Model ReleasesDGX agent

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and 3⃣Self-Speculation decoding by simply changing the attention

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

Model ReleasesDGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

What Google I/O '26 means for developing agents on Google Cloud

Model ReleasesDGX agent

At Google I/O, we introduced a unified development toolkit featuring Antigravity 2.0 and the Managed Agents API, giving developers better ways to build locally and deploy securely to the cloud on a sh

When Accuracy Is Not Enough: Uncertainty Collapse between Noisy Label Learning and Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2605.17795v1 Announce Type: cross Abstract: Learning with noisy labels (LNL) is typically benchmarked by closed-set classification accuracy, yet deployment often requires classifiers to reject o

When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings

Model ReleasesDGX agent

arXiv:2605.16288v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in clinical and care settings. This exploratory study investigates whether LLMs exhibit sycophantic

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

Model ReleasesDGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents

Model ReleasesDGX agent

arXiv:2601.17887v2 Announce Type: replace Abstract: Long-term memory enables large language model (LLM) agents to support personalized and sustained interactions. However, most work on personalized ag

Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers

Model ReleasesDGX agent

arXiv:2602.05813v2 Announce Type: replace Abstract: We study adaptive learning rate scheduling for norm-constrained optimizers (e.g., Muon and Lion). We introduce a generalized smoothness assumption u

WhiteTesseract: Reframing the Interpretation of Cultural Heritage through XR and Conversational AI

Model ReleasesDGX agent

arXiv:2605.16972v1 Announce Type: cross Abstract: Cultural heritage exhibitions often struggle to sustain attention and support reflective engagement. Physical exhibitions rely on fixed interpretive a

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

Model ReleasesDGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

Model ReleasesDGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens

Model ReleasesDGX agent

arXiv:2605.18115v1 Announce Type: new Abstract: Building a unified visual tokenizer is essential for bridging the gap between visual understanding and generation. Yet existing approaches struggle with

With expanded Antigravity platform, Google accelerates agent-native software development

Model ReleasesDGX agent

Google Cloud is enhancing its “agent-first” coding platform for developers with the launch of Antigravity 2.0, a new standalone desktop application that enables a full “agent-optimized” user experienc

Wordle 1,794 3/6 ⬛🟨🟩⬛⬛ 🟩⬛⬛⬛🟨 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a completed Wordle puzzle (#1,794) where the player successfully identified the target word in 3 attempts. The emoji grid shows the color-coded feedback from each guess: gray (inco

Wordle 1,795 4/6 🟨🟨⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle puzzle solution (puzzle #1,795) completed in 4 attempts, showing the progression of letter placements and eliminations across guesses before arriving at the final correct

World understanding: Gemini Omni is built on Gemini's vast knowledge of history, science, and culture, so it can produce videos that are gro…

Model ReleasesDGX agent

Gemini Omni is built on Gemini's extensive knowledge base of history, science, and culture, enabling it to generate videos with sophisticated contextual understanding. The capability leverages Google'

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

Model ReleasesDGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

Would you let robots spend your money? Google is betting on it

Model ReleasesDGX agent

Google is going all in on AI-driven shopping even as some competitors back off. At Google I/O, the company unveiled the latest iteration of its AI commerce tools: a 'Universal Cart' that works across

WOW-Seg: A Word-free Open World Segmentation Model

Model ReleasesDGX agent

arXiv:2605.16903v1 Announce Type: new Abstract: Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open

xAI has Released a blog on using Grok in OpenClaw

Model ReleasesDGX agent

xAI has published a blog post detailing how to utilize Grok, their AI assistant, in conjunction with OpenClaw, likely a development framework or tool for integrating AI capabilities. The post was shar

YOLO-NAS-Bench: A Surrogate Benchmark with Self-Evolving Predictors for YOLO Architecture Search

Model ReleasesDGX agent

arXiv:2603.09405v2 Announce Type: replace Abstract: Neural Architecture Search (NAS) for object detection is severely bottlenecked by high evaluation cost, as fully training each candidate YOLO archit

Your SaaS Is an Insurance Product: A Modeling Framework

Model ReleasesDGX agent

arXiv:2605.16699v1 Announce Type: new Abstract: Capped-usage SaaS products -- LLM subscriptions such as Claude Code and ChatGPT, cloud platforms such as Vercel and Cloudflare Workers, corporate benefi

Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions

Model ReleasesDGX agent

arXiv:2605.16877v1 Announce Type: new Abstract: Zero-shot textual explanations aim to make image classifiers more transparent by probing their internal representations, without relying on task-specifi

18 May 2026

3D Segmentation Using Viewpoint-Dependent Spatial Relationships

Model ReleasesDGX agent

arXiv:2605.15708v1 Announce Type: new Abstract: Recent advances in 3D datasets and multimodal models have greatly improved natural language 3D scene understanding. However, most 3D referring segmentat

3DEditSafe: Defending 3D Editing Pipelines from Unsafe Generation

Model ReleasesDGX agent

arXiv:2605.15398v1 Announce Type: cross Abstract: Recent advances in 3D generative editing, particularly pipelines based on 3D Gaussian Splatting (3DGS), have achieved high-fidelity, multi-view-consis

A Causally Grounded Taxonomy for Image Degradation Robustness Evaluation

Model ReleasesDGX agent

arXiv:2605.15906v1 Announce Type: new Abstract: Image degradations can occur during acquisition, processing, and transmission, altering visual appearance and affecting downstream vision tasks. They ar

← Previous
1…240241242243244…377
Next →