AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍

DGX agent

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍 JUST IN: Anthropic announces it will roll out Claude Mythos “in the com

model-releasesjeremy-howard--x
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

GraD-IBD: Graph Representation Learning from Diagnosis Trajectories for Early Detection of Inflammatory Bowel Disease

DGX agent

arXiv:2605.27799v1 Announce Type: new Abstract: International Classification of Diseases (ICD) is a globally recognized coding system that records diagnostic events during each patient encounter, prov

applicationsarxiv-cs-ai
28 May 2026
Model Releases

Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills

DGX agent

arXiv:2604.05333v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on reusable skills, and as they interact with personal applications, web browsers, and other interfaces, skill l

model-releasesarxiv-cs-ai
28 May 2026
Research

HRBench: Benchmarking and Understanding Thinking-Mode Switch Strategies in Hybrid-Reasoning LLMs

DGX agent

arXiv:2605.28398v1 Announce Type: new Abstract: Hybrid-reasoning large language models (LLMs) expose explicit controls over reasoning effort, allowing users or systems to trade off answer quality agai

researcharxiv-cs-ai
28 May 2026
Model Releases

HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning

DGX agent

arXiv:2605.27724v1 Announce Type: cross Abstract: Imitation learning is a promising approach for training humanoid robots to both walk and manipulate, but it requires a large number of demonstrations,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in tw…

DGX agent

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in twigl, make it like an infinite city of neo-gothic towers part

model-releasessonya-huang--x
28 May 2026
Model Releases

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to…

DGX agent

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to see whether it even did the complete processing or not 😅 ht

model-releasesjerry-liu--x
28 May 2026
Model Releases

Imitation Learning for Robot Assistance in Open Surgery: A Multi-Policy Evaluation on Suture Following

DGX agent

arXiv:2605.28736v1 Announce Type: new Abstract: This study presents the first evaluation of general-purpose imitation learning for surgeon-robot collaborative assistance in open surgery, targeting sut

model-releasesarxiv-cs-ro
28 May 2026
Research

InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training

DGX agent

arXiv:2510.15859v4 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has driven recent breakthroughs in large language models (LLMs), especially for tasks where rewards can be compute

researcharxiv-cs-ai
28 May 2026
Model Releases

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

DGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection

DGX agent

arXiv:2605.27748v1 Announce Type: cross Abstract: Industrial visual anomaly detection is usually one-class: normal images are abundant, while defects are rare, heterogeneous, and often unavailable dur

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation

DGX agent

arXiv:2605.28173v1 Announce Type: new Abstract: End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page la

model-releasesarxiv-cs-cv
28 May 2026
Research

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs

DGX agent

arXiv:2605.28388v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) is empirically shown to notably enhance the reasoning performance of large language models (LLMs),

researcharxiv-cs-ai
28 May 2026
Model Releases

Meta-Attention: Bayesian Per-Token Routing for Efficient Transformer Inference

DGX agent

arXiv:2605.28384v1 Announce Type: new Abstract: Standard transformer architectures apply a single attention mechanism uniformly across all tokens and sequence positions, irrespective of local context

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Mitigating Staleness in Asynchronous Pipeline Parallelism via Basis Rotation

DGX agent

arXiv:2602.03515v2 Announce Type: replace-cross Abstract: Asynchronous pipeline parallelism maximizes hardware utilization by eliminating the pipeline bubbles inherent in synchronous execution, offeri

model-releasesarxiv-cs-ai
28 May 2026
Local Ai

MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks

DGX agent

arXiv:2502.17832v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has become a common practice in multimodal large language models (MLLM) to enhance factual grounding and

local-aiarxiv-cs-ai
28 May 2026
Model Releases

Moment Matters: Mean and Variance Causal Graph Discovery from Heteroscedastic Observational Data

DGX agent

arXiv:2602.23602v2 Announce Type: replace-cross Abstract: Heteroscedasticity -- where the variance of a variable changes with other variables -- is pervasive in real data, and elucidating why it arise

model-releasesarxiv-cs-lg
28 May 2026
Research

Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Sokoban RNN

DGX agent

arXiv:2506.10138v3 Announce Type: replace-cross Abstract: We partially reverse-engineer a convolutional recurrent neural network (RNN) trained with model-free reinforcement learning to play the box-pu

researcharxiv-cs-ai
28 May 2026
Research

Pattern Recognition Tasks with Personalized Federated Learning

DGX agent

arXiv:2605.27816v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) constitutes a novel paradigm that tailors Machine Learning (ML) models to individual clients, thereby furnishing p

researcharxiv-cs-cv
28 May 2026
Model Releases

PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation

DGX agent

arXiv:2601.18006v2 Announce Type: replace Abstract: We present PEAR (Pairwise Evaluation for Automatic Relative Scoring), a supervised quality estimation (QE) metric family that reframes reference-fre

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

DGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations

DGX agent

arXiv:2605.27958v1 Announce Type: cross Abstract: Linear probes trained on LLM activations are increasingly proposed as deception-detection metrics, yet report AUROC exceeding 0.96 on clean benchmarks

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

DGX agent

arXiv:2605.28634v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency an

tutorialsarxiv-cs-ro
28 May 2026
Model Releases

Privately Estimating Monotone Statistics in Polynomial Time

DGX agent

arXiv:2605.27912v1 Announce Type: cross Abstract: We study efficient differentially private algorithms for estimating monotone statistics, i.e., statistics that are monotone under the addition of new

model-releasesarxiv-cs-lg
28 May 2026
Research

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation

DGX agent

arXiv:2605.28230v1 Announce Type: new Abstract: Modern video generative models produce visually impressive results, yet frequently violate basic physical principles. We propose Proprio, a training-fre

researcharxiv-cs-cv
28 May 2026
Model Releases

ProvMind: Provenance-grounded reasoning for materials synthesis

DGX agent

arXiv:2605.28487v1 Announce Type: new Abstract: Materials process optimization requires reasoning over routes, conditions, tools and causal dependencies, yet most computational formulations flatten sy

model-releasesarxiv-cs-ai
28 May 2026
Applications

QuITE: Query-Based Irregular Time Series Embedding

DGX agent

arXiv:2605.28166v1 Announce Type: cross Abstract: Irregular Multivariate Time Series (IMTS) are common in practice, yet their irregular sampling complicates effective modeling. Existing approaches typ

applicationsarxiv-cs-ai
28 May 2026
Model Releases

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG

DGX agent

arXiv:2605.28044v1 Announce Type: new Abstract: Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

DGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

model-releasesarxiv-cs-ai
28 May 2026
Safety

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

safetyarxiv-cs-ai
28 May 2026
Safety

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure

DGX agent

arXiv:2605.27996v1 Announce Type: new Abstract: Single-axis mitigations of reward-model biases (e.g., reducing proxy reliance on length, sycophancy, or style) can rotate optimization pressure onto cor

safetyarxiv-cs-ai
28 May 2026
Local Ai

ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning

DGX agent

arXiv:2605.27959v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have increasingly localized and interleaved visual evidence for deliberative reasoning. Grounding-based appro

local-aiarxiv-cs-ai
28 May 2026
Model Releases

SeeGroup: Multi-Layer Depth Estimation of Transparent Surfaces via Self-Determined Grouping

DGX agent

arXiv:2605.28735v1 Announce Type: new Abstract: Transparent objects are common in daily life, and it is important to understand their multilayer depth, including the transparent surface and the object

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments

DGX agent

arXiv:2605.28442v1 Announce Type: cross Abstract: Self-supervised online traversability estimation enables robots to continuously learn from unlabeled open-world experiences and adapt their navigation

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

DGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Sparse POD Mode Selection and Manifold Dimensionality Reduction with Neural Networks

DGX agent

arXiv:2605.27756v1 Announce Type: cross Abstract: High-performance computing enables simulation of high-dimensional physical systems, but downstream analyses such as inverse problems and control remai

model-releasesarxiv-cs-lg
28 May 2026
Research

Sparse Scheduled Diffusion Guidance for Inverse Problems

DGX agent

arXiv:2603.07860v2 Announce Type: replace Abstract: Pretrained diffusion models are effective priors for Bayesian inverse problems, but posterior sampling with these priors is often costly because dat

researcharxiv-cs-lg
28 May 2026
Model Releases

Stochastic Gradient Descent with Momentum is Algorithmically Stable

DGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

model-releasesarxiv-cs-ai
28 May 2026
Safety

Structure-Guided Visual Perturbation Neutralization for LVLMs

DGX agent

arXiv:2605.27927v1 Announce Type: new Abstract: Image inputs enable Large Vision Language Models (LVLMs) to perceive fine-grained visual information, but also introduce a pixel-level attack surface th

safetyarxiv-cs-cv
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭

DGX agent

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭 Learnings from testing Claude Opus 4.8: > Much worse than Opus 4.7 and GPT 5.5 on Vending Bench > More aligne

model-releasesemad-mostaque--x
28 May 2026
Model Releases

Using Zero-Shot LLM-Generated Survey Data for Geographically Explicit Population Synthesis

DGX agent

arXiv:2605.27401v1 Announce Type: cross Abstract: There is a growing interest in utilizing synthetic populations for a diverse range of applications. At the same time, we are witnessing a tremendous g

model-releasesarxiv-cs-ai
28 May 2026
Safety

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning

DGX agent

arXiv:2605.28023v1 Announce Type: cross Abstract: Visual captioning requires models to capture visual content faithfully while minimizing both omission and hallucination. As the dominant paradigm for

safetyarxiv-cs-ai
28 May 2026
Model Releases

Verifiable Benchmarking of Long-Horizon Spatial Biology

DGX agent

arXiv:2605.28065v1 Announce Type: new Abstract: AI agents are increasingly useful for biological data analysis, but existing benchmarks mostly test broad biological knowledge, executable workflows, or

model-releasesarxiv-cs-ai
28 May 2026
Research

VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer

DGX agent

arXiv:2605.28229v1 Announce Type: cross Abstract: With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding ie image-to-video

researcharxiv-cs-ai
28 May 2026
Model Releases

VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs

DGX agent

arXiv:2605.28422v1 Announce Type: cross Abstract: Latent reasoning enables reasoning over continuous hidden states rather than explicit tokens, avoiding the language bottleneck and inference overhead

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A newly released AI tool has generated an atlas of more than one billion predicted protein structures and billions more protein sequences. h…

DGX agent

DeepMind's AlphaFold3 and related tools have generated a comprehensive atlas containing over one billion predicted protein structures and additional billions of protein sequences, representing a major

model-releasesyann-lecun--x
27 May 2026
Model Releases

AgentSociety: Incentivizing Agentic Social Intelligence

DGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…709710711712713…1371
Next →