AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,616 results
13 May 2026

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

ApplicationsDGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

ResearchDGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti

Unlocking Compositional Generalization in Continual Few-Shot Learning

ResearchDGX agent

arXiv:2605.11710v1 Announce Type: cross Abstract: Object-centric representations promise a key property for few-shot learning: Rather than treating a scene as a single unit, a model can decompose it i

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Unlocking LLM Creativity in Science through Analogical Reasoning

AgentsDGX agent

arXiv:2605.11258v1 Announce Type: cross Abstract: Autonomous science promises to augment scientific discovery, particularly in complex fields like biomedicine. However, this requires AI systems that c

Unlocking UML Class Diagram Understanding in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives

ResearchDGX agent

arXiv:2605.11166v1 Announce Type: new Abstract: Political and social identities structure how people evaluate political information, a finding decades deep in political science and routinely discarded

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

Model ReleasesDGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation

ResearchDGX agent

arXiv:2605.11131v1 Announce Type: new Abstract: Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global infor

v0.23.4

Local AiDGX agent

Ollama v0.23.4 is a release version of Ollama, a tool for running large language models locally. This patch release likely includes bug fixes, performance improvements, and refinements to existing fea

Variance-aware Reward Modeling with Anchor Guidance

ApplicationsDGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

ResearchDGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization

ResearchDGX agent

arXiv:2605.11913v1 Announce Type: new Abstract: Differentiable vector graphics have enabled powerful gradient-based optimization of vector primitives directly from raster images. However, existing fra

Veeam introduces new backup management, cybersecurity features

IndustryDGX agent

Veeam Software Group GmbH today debuted new features that will make it easier for enterprises to back up their records and protect them from hackers. Some of the capabilities are rolling to the compan

Veeam’s big pivot on display at VeeamON 2026

AgentsDGX agent

Veeam Software Group GmbH used VeeamON 2026 in New York City this week to punctuate its shift from “the backup company” to a data and artificial intelligence trust platform for the agentic era. With a

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

Model ReleasesDGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization

ResearchDGX agent

arXiv:2605.10974v1 Announce Type: new Abstract: Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing ver

Very Efficient Listwise Multimodal Reranking for Long Documents

Model ReleasesDGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

ResearchDGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors

ResearchDGX agent

arXiv:2605.11424v1 Announce Type: new Abstract: Gaussian Splatting has achieved remarkable progress in multi-view surface reconstruction, yet it exhibits notable degradation when only few views are av

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference

SafetyDGX agent

arXiv:2605.12325v1 Announce Type: new Abstract: Pursuing training-free open-vocabulary semantic segmentation in an efficient and generalizable manner remains challenging due to the deep-seated spatial

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

ResearchDGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

Model ReleasesDGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

Model ReleasesDGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

VNDUQE: Information-Theoretic Novelty Detection using Deep Variational Information Bottleneck

SafetyDGX agent

arXiv:2605.11551v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) samples is critical for safe deployment of neural networks in safety-critical applications. While maximum softmax

Vox-style match-cut montage. This workflow splits your input word by word, renders each one as a bold editorial headline on aged newsprint v…

Local AiDGX agent

Vox-style match-cut montage. This workflow splits your input word by word, renders each one as a bold editorial headline on aged newsprint via gpt-image-1, then assembles the frames into a stop-motion

Watch Falcon 9 launch Dragon to the @Space_Station https://x.com/i/broadcasts/1DxleEdlAQjKL

IndustryDGX agent

This post links to a live broadcast of a SpaceX Falcon 9 rocket launching a Dragon spacecraft to the International Space Station. The broadcast was shared by Elon Musk on X (formerly Twitter), allowin

Watch the full episode: Spotify: https://seq.vc/b4u Apple: https://seq.vc/3xf YouTube: https://seq.vc/syf

IndustryDGX agent

I cannot provide an accurate summary as the title consists only of platform links without descriptive content. Based on the shortened URLs and source attribution to Sonya Huang, this appears to be a p

way ahead of its time:

SafetyDGX agent

way ahead of its time: Three questions for @sama that the public deserves to better understand: 👉 What is current value of your indirect stake in OpenAI? (Note that you told the senate that you had no

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates…

IndustryDGX agent

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates more risk than openness. 'Six, seven years ago, at the time

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

Model ReleasesDGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

We just shipped tons of new products to accelerate the full agent development lifecycle: https://www.langchain.com/blog TLDR: ✅ LangSmith En…

AgentsDGX agent

We just shipped tons of new products to accelerate the full agent development lifecycle: https://www.langchain.com/blog TLDR: ✅ LangSmith Engine ✅ SmithDB ✅ Sandboxes ✅ Managed Deep Agents ✅ LLM Gatew

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a…

Model ReleasesDGX agent

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a human would never try on another human. In one run, GPT-5.1

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

Model ReleasesDGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

ResearchDGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

We're building a database 🤠

AgentsDGX agent

We're building a database 🤠 We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data prob

🔴 We're LIVE from the SaaStr floor, Day 2! 🎙️🚀 Joined by an amazing lineup today: • Lindsay Wise + Julia Holm (Green Security) • Kody Low…

ToolsDGX agent

🔴 We're LIVE from the SaaStr floor, Day 2! 🎙️🚀 Joined by an amazing lineup today: • Lindsay Wise + Julia Holm (Green Security) • Kody Low (Replit Field Engineering) • Philipp Dietz (Replit Brand & Cre

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send m…

AgentsDGX agent

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send messages to users, to themselves (CoT) and to tools, and rece

What Does It Mean for a Medical AI System to Be Right?

Model ReleasesDGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

TutorialsDGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

What we find most useful about TST is the decoupling. The training-time efficiency is fully separated from the inference-time architecture, …

ResearchDGX agent

What we find most useful about TST is the decoupling. The training-time efficiency is fully separated from the inference-time architecture, which makes TST a clean addition on top of other pretraining

What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization

SafetyDGX agent

arXiv:2605.12021v1 Announce Type: new Abstract: Many image understanding tasks involve identifying what is present and where it appears. However, tasks that address where, such as object discovery, de

When and How to Canonize: A Generalization Perspective

TutorialsDGX agent

arXiv:2605.11008v1 Announce Type: new Abstract: While invariant architectures are standard for processing symmetric data, there is growing interest in achieving invariance by applying group averaging

When Brains Disagree: Biological Ambiguity Underlies the Challenge of Amyloid PET Synthesis from Structural MRI

ResearchDGX agent

arXiv:2605.11867v1 Announce Type: new Abstract: Structural MRI-to-amyloid PET synthesis has been proposed as a non-invasive alternative for amyloid assessment in Alzheimer's disease (AD). However, rep

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

SafetyDGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

ResearchDGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

ResearchDGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

SafetyDGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

When the Gold Standard Isn't Necessarily Standard: Challenges of Evaluating the Translation of User-Generated Content

ResearchDGX agent

arXiv:2512.17738v2 Announce Type: replace Abstract: User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, ch

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

SafetyDGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

Model ReleasesDGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference

ApplicationsDGX agent

arXiv:2605.12255v1 Announce Type: cross Abstract: When people share the same documents and observations yet reach different conclusions, the disagreement often shifts into a judgment that the other pa

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter …

Model ReleasesDGX agent

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter most in long-context inference: → KV-cache pressure during de

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

Model ReleasesDGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

Will Ollama come out with a non-cloud version of Deepseek-v4 Flash?

Model ReleasesDGX agent

DeepSeek-v4 Flash through Ollama is currently available as a cloud model, where Ollama's CLI sends API calls to Ollama's hosted version rather than running locally . Local support for DeepSeek V4 Flas

World Action Models: The Next Frontier in Embodied AI

SafetyDGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

WorldComp2D: Spatio-semantic Representations of Object Identity and Location from Local Views

ResearchDGX agent

arXiv:2605.11743v1 Announce Type: new Abstract: Learning latent representations that capture both semantic and spatial information is central to efficient spatio-semantic reasoning. However, many exis

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction

ApplicationsDGX agent

arXiv:2605.12162v1 Announce Type: new Abstract: Effectively handling the interplay between spatial perception and action generation remains a critical bottleneck in robotic manipulation. Existing meth

xi-DPO: Direct Preference Optimization via Ratio Reward Margin

ResearchDGX agent

arXiv:2605.10981v1 Announce Type: new Abstract: Reference-free preference optimization has emerged as an efficient alternative to reinforcement learning from human feedback, with Simple Preference Opt

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

Model ReleasesDGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

Yann LeCun says you cannot build a reliable agentic system without a world model LLMs don't have world models. They can't predict the conseq…

AgentsDGX agent

Yann LeCun says you cannot build a reliable agentic system without a world model LLMs don't have world models. They can't predict the consequences of their actions before taking them 'they just act, a

← Previous
1…9969979989991000…1461
Next →