AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
Model Releases

VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction

DGX agent

arXiv:2605.16911v1 Announce Type: new Abstract: 3D semantic occupancy prediction requires accurate 2D-to-3D feature lifting, yet current methods restrict camera geometry to initial projections. Subseq

model-releasesarxiv-cs-cv
19 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

[video] why we need a new continuity layer for long-running agents (claude did this video! all except the voice which was @elevenlabs)

DGX agent

This video discusses the architectural need for a continuity layer in long-running AI agents, explaining how agents require persistent memory and state management mechanisms to maintain coherence acro

model-releasesyohei-nakajima--x
19 May 2026
Model Releases

Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models

DGX agent

arXiv:2605.18160v1 Announce Type: cross Abstract: In recent years, multimodal large language models (MLLMs) have achieved remarkable progress, primarily attributed to effective paradigms for integrati

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

DGX agent

arXiv:2602.04802v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have achieved impressive performance in cross-modal understanding across textual and visual inputs, yet existing bench

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

DGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

DGX agent

arXiv:2605.18172v1 Announce Type: new Abstract: Leveraging the universal representations of pre-trained LLMs and MLLMs offers a promising path toward brain foundation models. However, visually-evoked

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

DGX agent

arXiv:2605.18313v1 Announce Type: cross Abstract: Small vision-language models (2-8B) are well-suited for clin- ical deployment due to privacy constraints, limited connectivity, and low-latency requir

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning

DGX agent

arXiv:2601.06943v2 Announce Type: replace-cross Abstract: In real-world video question answering scenarios, videos often provide only localized visual cues, while verifiable answers are distributed ac

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WavFlow: Audio Generation in Waveform Space

DGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong

DGX agent

arXiv:2509.22510v3 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) is the ability to satisfy desired objectives during generation, which is critical for trustworthy deployme

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories…

DGX agent

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories, memorable moments, and many, many (occasionally embarrassi

model-releasesgoogle-ai--x
19 May 2026
Model Releases

WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games

DGX agent

arXiv:2605.17637v1 Announce Type: new Abstract: Coding agents are increasingly used as application builders, yet many evaluations still focus on source code, repository-level tests, or intermediate tr

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale

DGX agent

arXiv:2510.16252v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for web agents demands environments that are both effective for evaluation and efficient enough for large-scale on

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Weighted Flow Matching and Physics-Informed Nonlinear Filtering for Parameter Estimation in Digital Twins

DGX agent

arXiv:2605.17146v1 Announce Type: cross Abstract: Digital twins (DTs) rely on continuous synchronization between physical systems and their virtual counterparts through online parameter estimation und

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing

DGX agent

arXiv:2510.15221v2 Announce Type: replace Abstract: Affective computing has matured rapidly in laboratory settings, yet no prior dataset combines (i) months-to-years of duration, (ii) a naturalistic w

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

We’re adding new ways for people to identify AI-generated images and understand where they came from. In addition to C2PA Content Credential…

DGX agent

We’re adding new ways for people to identify AI-generated images and understand where they came from. In addition to C2PA Content Credentials, images now also contain a SynthID watermark, and can be i

model-releasesopenai--x
19 May 2026
Model Releases

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and …

DGX agent

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and 3⃣Self-Speculation decoding by simply changing the attention

model-releasesemad-mostaque--x
19 May 2026
Model Releases

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

DGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

What Google I/O '26 means for developing agents on Google Cloud

DGX agent

At Google I/O, we introduced a unified development toolkit featuring Antigravity 2.0 and the Managed Agents API, giving developers better ways to build locally and deploy securely to the cloud on a sh

model-releasesgoogle-cloud-ai
19 May 2026
Model Releases

When Accuracy Is Not Enough: Uncertainty Collapse between Noisy Label Learning and Out-of-Distribution Detection

DGX agent

arXiv:2605.17795v1 Announce Type: cross Abstract: Learning with noisy labels (LNL) is typically benchmarked by closed-set classification accuracy, yet deployment often requires classifiers to reject o

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings

DGX agent

arXiv:2605.16288v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in clinical and care settings. This exploratory study investigates whether LLMs exhibit sycophantic

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

DGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents

DGX agent

arXiv:2601.17887v2 Announce Type: replace Abstract: Long-term memory enables large language model (LLM) agents to support personalized and sustained interactions. However, most work on personalized ag

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers

DGX agent

arXiv:2602.05813v2 Announce Type: replace Abstract: We study adaptive learning rate scheduling for norm-constrained optimizers (e.g., Muon and Lion). We introduce a generalized smoothness assumption u

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

WhiteTesseract: Reframing the Interpretation of Cultural Heritage through XR and Conversational AI

DGX agent

arXiv:2605.16972v1 Announce Type: cross Abstract: Cultural heritage exhibitions often struggle to sustain attention and support reflective engagement. Physical exhibitions rely on fixed interpretive a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

DGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens

DGX agent

arXiv:2605.18115v1 Announce Type: new Abstract: Building a unified visual tokenizer is essential for bridging the gap between visual understanding and generation. Yet existing approaches struggle with

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

With expanded Antigravity platform, Google accelerates agent-native software development

DGX agent

Google Cloud is enhancing its “agent-first” coding platform for developers with the launch of Antigravity 2.0, a new standalone desktop application that enables a full “agent-optimized” user experienc

model-releasessiliconangle
19 May 2026
Model Releases

Wordle 1,794 3/6 ⬛🟨🟩⬛⬛ 🟩⬛⬛⬛🟨 🟩🟩🟩🟩🟩

DGX agent

This post documents a completed Wordle puzzle (#1,794) where the player successfully identified the target word in 3 attempts. The emoji grid shows the color-coded feedback from each guess: gray (inco

model-releasesanthropic--x
19 May 2026
Model Releases

Wordle 1,795 4/6 🟨🟨⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle puzzle solution (puzzle #1,795) completed in 4 attempts, showing the progression of letter placements and eliminations across guesses before arriving at the final correct

model-releasesanthropic--x
19 May 2026
Model Releases

World understanding: Gemini Omni is built on Gemini's vast knowledge of history, science, and culture, so it can produce videos that are gro…

DGX agent

Gemini Omni is built on Gemini's extensive knowledge base of history, science, and culture, enabling it to generate videos with sophisticated contextual understanding. The capability leverages Google'

model-releasesgoogle-ai--x
19 May 2026
Model Releases

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

DGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Would you let robots spend your money? Google is betting on it

DGX agent

Google is going all in on AI-driven shopping even as some competitors back off. At Google I/O, the company unveiled the latest iteration of its AI commerce tools: a 'Universal Cart' that works across

model-releasesthe-verge-ai
19 May 2026
Model Releases

WOW-Seg: A Word-free Open World Segmentation Model

DGX agent

arXiv:2605.16903v1 Announce Type: new Abstract: Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

xAI has Released a blog on using Grok in OpenClaw

DGX agent

xAI has published a blog post detailing how to utilize Grok, their AI assistant, in conjunction with OpenClaw, likely a development framework or tool for integrating AI capabilities. The post was shar

model-releaseselon-musk--x
19 May 2026
Model Releases

YOLO-NAS-Bench: A Surrogate Benchmark with Self-Evolving Predictors for YOLO Architecture Search

DGX agent

arXiv:2603.09405v2 Announce Type: replace Abstract: Neural Architecture Search (NAS) for object detection is severely bottlenecked by high evaluation cost, as fully training each candidate YOLO archit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Your SaaS Is an Insurance Product: A Modeling Framework

DGX agent

arXiv:2605.16699v1 Announce Type: new Abstract: Capped-usage SaaS products -- LLM subscriptions such as Claude Code and ChatGPT, cloud platforms such as Vercel and Cloudflare Workers, corporate benefi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions

DGX agent

arXiv:2605.16877v1 Announce Type: new Abstract: Zero-shot textual explanations aim to make image classifiers more transparent by probing their internal representations, without relying on task-specifi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

3D Segmentation Using Viewpoint-Dependent Spatial Relationships

DGX agent

arXiv:2605.15708v1 Announce Type: new Abstract: Recent advances in 3D datasets and multimodal models have greatly improved natural language 3D scene understanding. However, most 3D referring segmentat

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

3DEditSafe: Defending 3D Editing Pipelines from Unsafe Generation

DGX agent

arXiv:2605.15398v1 Announce Type: cross Abstract: Recent advances in 3D generative editing, particularly pipelines based on 3D Gaussian Splatting (3DGS), have achieved high-fidelity, multi-view-consis

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

A Causally Grounded Taxonomy for Image Degradation Robustness Evaluation

DGX agent

arXiv:2605.15906v1 Announce Type: new Abstract: Image degradations can occur during acquisition, processing, and transmission, altering visual appearance and affecting downstream vision tasks. They ar

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

A Cross-Modal Prompt Injection Attack against Large Vision-Language Models with Image-Only Perturbation

DGX agent

arXiv:2605.16090v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have emerged as a powerful paradigm for multimodal intelligence, but their growing deployment also expands the at

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

A Reproducible and Physically Feasible Dynamic Parameter Identification Framework for a Low-Cost Robot Arm

DGX agent

arXiv:2605.15949v1 Announce Type: new Abstract: This paper presents a reproducible and physically feasible dynamic parameter identification framework for CRANE-X7, a low-cost robot arm driven by modul

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement

DGX agent

arXiv:2507.14200v2 Announce Type: replace-cross Abstract: Existing multi-LLM collaboration systems often encounter scalability challenges when integrating new LLMs and tasks, leading to suboptimal per

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

A3D: Agentic AI flow for autonomous Accelerator Design

DGX agent

arXiv:2605.15237v1 Announce Type: cross Abstract: Accelerating applications through the design of hardware accelerators can significantly enhance system performance and energy efficiency. Despite adva

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging

DGX agent

arXiv:2505.21698v3 Announce Type: replace Abstract: Vision-language foundation models achieve promising performance in natural image classification, yet their direct application to medical imaging is

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

AGC: Adaptive Geodesic Correction for Adversarial Robustness on Vision-Language Models

DGX agent

arXiv:2605.15584v1 Announce Type: new Abstract: Vision-language models like CLIP have demonstrated remarkable zero-shot transfer capabilities. However, their susceptibility to imperceptible adversaria

model-releasesarxiv-cs-cv
18 May 2026
← Previous
1…301302303304305…471
Next →