AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
25 May 2026

RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.23068v1 Announce Type: new Abstract: Reliable visual understanding in robot-assisted and minimally invasive surgery (RMIS/MIS) demands more than accurate masks: in clinical practice, clinic

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

Model ReleasesDGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

SciHorizon-GENE: Benchmarking LLM for Life Sciences Inference from Gene Knowledge to Functional Understanding

Model ReleasesDGX agent

arXiv:2601.12805v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown growing promise in biomedical research, particularly for knowledge-driven interpretation tasks. Howeve

Seeing without Looking: Do Vision-Language Benchmarks Really Test Vision?

Model ReleasesDGX agent

arXiv:2605.22903v1 Announce Type: cross Abstract: Benchmark accuracy is often implicitly assumed to reflect grounded visual understanding in vision-language models (VLMs), yet it remains unclear to wh

Semantically Structured Mixture-of-Experts for Compositional Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.23477v1 Announce Type: new Abstract: Diffusion-based policies have established a new standard for precise robotic manipulation but face a critical scalability bottleneck: high-performance m

SemEval-2026 Task 6: CLARITY -- Unmasking Political Question Evasions

Model ReleasesDGX agent

arXiv:2603.14027v2 Announce Type: replace Abstract: Political speakers often avoid answering questions directly while maintaining the appearance of responsiveness. Despite its importance for public di

SkillOpt: Executive Strategy for Self-Evolving Agent Skills

Model ReleasesDGX agent

arXiv:2605.23904v1 Announce Type: new Abstract: Agent skills today are hand-crafted, generated one-shot, or evolved through loosely controlled self-revision, none of which behaves like a deep-learning

Sparse Autoencoders Map Brain-LLM Alignment onto Cortical Semantic Topography

Model ReleasesDGX agent

arXiv:2605.23035v1 Announce Type: cross Abstract: Intermediate layers of large language models (LLMs) best predict human brain responses to language, one of the most robust findings in computational n

Speak-to-Structure: Evaluating LLMs in Open-domain Natural Language-Driven Molecule Generation

Model ReleasesDGX agent

arXiv:2412.14642v4 Announce Type: replace Abstract: Recently, Large Language Models (LLMs) have demonstrated great potential in natural language-driven molecule discovery. However, existing datasets a

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding

Model ReleasesDGX agent

arXiv:2605.23137v1 Announce Type: cross Abstract: Electroencephalography (EEG) visual decoding remains challenging due to the modality gap between low-SNR neural signals and highly structured vision--

StereoGenBench: A Synthetic Multi-Camera Benchmark for Stereo Generation under Controlled Baseline Regimes

Model ReleasesDGX agent

arXiv:2605.23237v1 Announce Type: new Abstract: Stereo image and video generation, stereo geometry estimation, and condition-controlled view synthesis require paired data in which the variables that d

Strategic Coercion Within Alliances: The Greenland Sovereignty Game as an AI Stress Test

Model ReleasesDGX agent

arXiv:2605.22841v1 Announce Type: cross Abstract: What happens when the strongest alliance member pressures a weaker member over territory and strategic control? We examine the Greenland sovereignty c

Tabular PDF Information Extraction with Local LLMs and Layout-Aware Parsing: A Reliability Evaluation

Model ReleasesDGX agent

arXiv:2604.00003v2 Announce Type: replace-cross Abstract: Extracting structured information from academic PDF documents is non trivial: a single page typically combines free text metadata with tabular

Targeted Regularization for Causal Effect Estimation with Exponential Dispersion Family Outcomes

Model ReleasesDGX agent

arXiv:2502.07295v2 Announce Type: replace Abstract: Neural Networks (NNs) for causal effect estimation have shown strong empirical performance, yet endowing them with desirable semiparametric properti

TEAM: Temporal-Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration

Model ReleasesDGX agent

arXiv:2602.08404v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) have recently gained significant attention due to their inherent support for parallel decoding. Building on

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

Model ReleasesDGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models

Model ReleasesDGX agent

arXiv:2605.22870v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting is necessary for arithmetic in small language models, yet shuffling its steps preserves most performance. What does C

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.21306v2 Announce Type: replace-cross Abstract: This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compoundin

Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.22902v1 Announce Type: cross Abstract: Generative Vision-Language Models (VLMs) perform well on multimodal reasoning, but how visual inputs are transformed to text remains poorly understood

Understanding and Improving Noisy Embedding Techniques in Instruction Finetuning

Model ReleasesDGX agent

arXiv:2605.23171v1 Announce Type: cross Abstract: Recent advancements in instructional fine-tuning have injected noise into embeddings, with NEFTune (Jain et al., 2024) setting benchmarks using unifor

Unextractable Protocol Models: Collaborative Training and Inference without Weight Materialization

Model ReleasesDGX agent

arXiv:2605.23464v1 Announce Type: new Abstract: We consider a decentralized setup in which the participants collaboratively train and serve a large neural network, and where each participant only proc

Using Ensemble Diffusion to Estimate Uncertainty for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2506.00560v2 Announce Type: replace-cross Abstract: End-to-end planning systems for autonomous driving are rapidly improving, especially in closed-loop simulation environments like CARLA. Many s

VDE: Training-Free Accelerating Rectified Flow Model via Velocity Decomposition and Estimation

Model ReleasesDGX agent

arXiv:2605.23381v1 Announce Type: new Abstract: Though rectified flow models have achieved remarkable performance in image, video, and 3D generation, their practical deployments are challenged by slow

Vector Retrieval with Similarity and Diversity: How Hard Is It?

Model ReleasesDGX agent

arXiv:2407.04573v4 Announce Type: replace-cross Abstract: Dense vector retrieval is an important building block of modern machine learning systems, underlying applications ranging from semantic search

VideoOdyssey: A Benchmark for Ultra-Long-Context and Omni-Modal Video Understanding

Model ReleasesDGX agent

arXiv:2605.22907v1 Announce Type: new Abstract: Real-world long video understanding requires models to perform continuous tracking, information integration and memory retention over massive temporal s

VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos

Model ReleasesDGX agent

arXiv:2602.07801v4 Announce Type: replace-cross Abstract: In long-video understanding, conventional uniform frame sampling often fails to capture key visual evidence, leading to degraded performance a

VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset

Model ReleasesDGX agent

arXiv:2605.23518v1 Announce Type: new Abstract: Directly editing ultra-high-resolution (UHR) images is valuable but underexplored, primarily due to the lack of high-quality data and the challenge in m

VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images

Model ReleasesDGX agent

arXiv:2605.23141v1 Announce Type: new Abstract: A useful test of visual concept learning is not just whether a model can recognize a concept in a single image, but whether it can preserve and manipula

What Linear Probes Miss: Multi-View Probing for Weight-Space Learning

Model ReleasesDGX agent

arXiv:2605.23410v1 Announce Type: cross Abstract: The explosive growth of open-source model repositories has created a Model Jungle, where checkpoints are frequently shared without adequate documentat

What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA

Model ReleasesDGX agent

arXiv:2605.23067v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a viable recipe for training LLM agents to reason over external memory banks in multi-session dialogue. Exist

When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization

Model ReleasesDGX agent

arXiv:2605.23272v1 Announce Type: cross Abstract: Symbolic Regression (SR) plays a central role in scientific knowledge discovery by distilling mathematical equations from observational data. Most exi

When Symptoms Are Not Enough: Evidence-Weighting Patterns in Large Language Model Psychiatric Screening

Model ReleasesDGX agent

arXiv:2605.23148v1 Announce Type: new Abstract: As demand for mental health care outpaces clinician-delivered assessment, scalable screening tools are increasingly needed. Large language models (LLMs)

Wordle 1,801 4/6 🟨🟨⬛⬛⬛ ⬛🟩⬛⬛⬛ ⬛🟩🟩🟨⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result where the player solved puzzle #1,801 in four attempts using color-coded feedback (yellow for correct letters in wrong positions, green for correct letters in

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

Model ReleasesDGX agent

arXiv:2502.04230v3 Announce Type: replace-cross Abstract: The rapid proliferation of generative audio synthesis and editing technologies has raised serious concerns about copyright infringement, data

24 May 2026

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to…

Model ReleasesDGX agent

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to their European counterparts. What’s especially notable is that

Aleph 2.0 will blow your mind

Model ReleasesDGX agent

Aleph 2.0 will blow your mind Just tested Runway Aleph 2.0 and this blew my mind a bit lol I saw a video like this when Aleph first released and with the new Aleph 2.0 I wanted to create my own versio

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects,…

Model ReleasesDGX agent

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects, with a diverse dataset and new benchmark. 🌐 https://physx-o

Built an AI screen memory using llama.cpp + Gemma 4 — remembers everything you do on your computer,search/chat or make agents over it. 100% local

Model ReleasesDGX agent

This project demonstrates a local AI system built with llama.cpp and Gemma 4 that captures and analyzes screen activity to create persistent memory of user computer interactions, enabling search, chat

DeepSeek says it will lower V4 Pro API prices by 75% to 0.435/1M input and 0.87/1M output tokens, making permanent the discount prices set to expire on May 31 (Bloomberg)

Model ReleasesDGX agent

Bloomberg: DeepSeek says it will lower V4 Pro API prices by 75% to 0.435/1M input and 0.87/1M output tokens, making permanent the discount prices set to expire on May 31 — DeepSeek said it will make p

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk t…

Model ReleasesDGX agent

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk to ChatGPT and Claude pretending to be DoofWarrior123 on X wi

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new law…

Model ReleasesDGX agent

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new lawsuit, two authors allege that Guillaume Lample, while an AI

Mad House — Usborne Creepy Computer Games

Model ReleasesDGX agent

Tool: Mad House — Usborne Creepy Computer Games Via Hacker News I learned that UK publisher Usborne published free PDFs of their 1980s Computer Books, some of which I remember working through on my Co

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no m…

Model ReleasesDGX agent

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no more permission prompts. It is the key building block for mul

quick summary of someone's github, cool! here's me

Model ReleasesDGX agent

quick summary of someone's github, cool! here's me I always wanted a GitHub dashboard: See my repos, open Issues/PRs, what version I released last, how many commits since last release. So I built one

Wordle 1,799 4/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛🟨⬛⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a Wordle game result where the player solved puzzle #1,799 in 4 attempts, with the final answer being a five-letter word where all letters are in the correct positions (indicated by th

Wordle 1,800 4/6 ⬛⬛⬛⬛🟩 ⬛🟩🟨⬛⬛ 🟩🟩🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result where the player solved puzzle #1,800 in 4 attempts, using the color-coded feedback system (gray for wrong letters, yellow for correct letters in wrong positio

23 May 2026

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, a…

Model ReleasesDGX agent

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, and will pose problems for real-world systems that aren’t wel

A Boundary-Layer Mechanism for One-Third Scaling in Online Softmax Classification

Model ReleasesDGX agent

arXiv:2605.22341v1 Announce Type: new Abstract: Hard-label classification is usually trained with smooth surrogate losses, most prominently softmax cross-entropy. We isolate an asymptotic mechanism by

A2QTGN: Adaptive Amplitude Quantum-Integrated Temporal Graph Network for Dynamic Link Prediction

Model ReleasesDGX agent

arXiv:2605.21916v1 Announce Type: cross Abstract: Dynamic link prediction is important for modeling evolving interactions in complex systems, including social, communication, financial, and transporta

Adaptive RBF-KAN: A Comparative Evaluation of Dynamic Shape Parameters in Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2605.21534v1 Announce Type: cross Abstract: Kolmogorov-Arnold Networks (KANs) approximate multivariate functions using learnable univariate edge functions, typically parameterized by B-spline ba

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. Wi…

Model ReleasesDGX agent

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. With motivation, overview, and GPT-style model reference imple

Alike Parts: A Feature-Informed Approach to Local and Global Prototype Explanations

Model ReleasesDGX agent

arXiv:2605.21646v1 Announce Type: new Abstract: Prototype-based explanations offer an intuitive, example-based approach to support the interpretability of machine learning black box classifiers but of

An entropy formula for the Deep Linear Network

Model ReleasesDGX agent

arXiv:2509.09088v3 Announce Type: replace Abstract: We study the Riemannian geometry of the Deep Linear Network (DLN) as a foundation for a thermodynamic description of the learning process. The main

An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning

Model ReleasesDGX agent

arXiv:2605.21968v1 Announce Type: new Abstract: Optimization is essential in deep learning. The foundational method upon which most optimizers are built is momentum-based stochastic gradient descent.

ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage

Model ReleasesDGX agent

arXiv:2605.21615v1 Announce Type: cross Abstract: Existing binary corpora typically capture only one or two axes of binary variation: they either provide cross-compiler builds without a temporal axis,

AutoBaxBuilder: Bootstrapping Code Security Benchmarking

Model ReleasesDGX agent

arXiv:2512.21132v2 Announce Type: replace-cross Abstract: As large language models (LLMs) see wide adoption in software engineering, the reliable assessment of the correctness and security of LLM-gene

Automatic Contextual Audio Denoising

Model ReleasesDGX agent

arXiv:2605.22262v1 Announce Type: cross Abstract: Audio context determines which sound components and sources are relevant and which can be perceived as irrelevant (noise) by listeners. For example, t

Billion-Scale Graph Foundation Models

Model ReleasesDGX agent

arXiv:2602.04768v2 Announce Type: replace Abstract: Graph-structured data underpins many critical applications. While foundation models have transformed language and vision via large-scale pretraining

Characterizing the Fault Response of the Intel Neural Compute Stick 2 Under Single-Pulse Electromagnetic Fault Injection

Model ReleasesDGX agent

arXiv:2605.22437v1 Announce Type: cross Abstract: Vision processing units and other commercial neural-network inference accelerators are increasingly deployed in safety-relevant edge applications, but

← Previous
1…217218219220221…377
Next →