AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
28 May 2026

CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning

Model ReleasesDGX agent

arXiv:2605.28056v1 Announce Type: new Abstract: Portrait animation methods have achieved substantial visual quality and lip synchronization, but fine-grained manipulation of the eye region still faces

Conditionally Site-Independent Neural Evolution of Antibody Sequences

ResearchDGX agent

arXiv:2602.18982v4 Announce Type: replace Abstract: Common deep learning approaches for antibody engineering focus on modeling the marginal distribution of sequences. By treating sequences as independ

Context Features Are Cheap: Rank-Aware Decomposition for Efficient Feature Interaction in Recommender Systems

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.27450v1 Announce Type: cross Abstract: Modern industrial recommender systems use a deep ranking model to score N candidates against the same user and context features. Standard implementati

CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning

ResearchDGX agent

arXiv:2605.28742v1 Announce Type: new Abstract: Language models can use verifiable rewards to improve at a wide variety of reasoning tasks. However, both parametric (e.g. RLVR) and non-parametric (e.g

COTTA: Context-Aware Transfer Adaptation for Trajectory Prediction in Autonomous Driving

SafetyDGX agent

arXiv:2604.00402v2 Announce Type: replace-cross Abstract: Developing robust models to accurately predict the trajectories of surrounding agents is fundamental to autonomous driving safety. However, mo

DAISI: Data Assimilation with Inverse Sampling using Stochastic Interpolants

ResearchDGX agent

arXiv:2512.00252v4 Announce Type: replace-cross Abstract: Data assimilation (DA) is a cornerstone of scientific and engineering applications, combining model forecasts with sparse and noisy observatio

Dark Quest II: A Wide-Coverage Neural Network Emulator of the Nonlinear Matter Power Spectrum Across Extended Cosmologies

Model ReleasesDGX agent

arXiv:2605.28596v1 Announce Type: cross Abstract: extsc{DarkEmulator2} is a neural network emulator of the nonlinear matter power spectrum in a nine-dimensional w_0 w_a nu o CDM parameter space, devel

DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification

Model ReleasesDGX agent

arXiv:2605.27858v1 Announce Type: cross Abstract: Claim verification splits between end-to-end classifiers that are accurate but yields no inspectable traces, and decomposition-based methods produce i

DREAM-R: Multimodal Speculative Reasoning with RL-Based Refined Drafting, Precise Verification, and Fully Parallel Execution

SafetyDGX agent

arXiv:2605.28678v1 Announce Type: new Abstract: Speculative reasoning has recently been proposed as a means to accelerate reasoning-intensive generation in large multimodal models, but its effectivene

DRTriton: Large-Scale Synthetic Data Driven Reinforcement Learning for Triton Kernel Generation

Model ReleasesDGX agent

arXiv:2603.21465v2 Announce Type: replace Abstract: Developing efficient CUDA kernels is a fundamental yet challenging task in the generative AI industry. Recent research leverages Large Language Mode

E^3-Agent: An Executable and Evolving Agent for Resource Management of Edge Generative Inference

AgentsDGX agent

arXiv:2605.27428v1 Announce Type: new Abstract: Edge deployments of generative inference increasingly face two practical realities: per-device per-model performance is often unknown at deployment time

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

Energy-Structured Low-Rank Adaptation for Continual Learning

Model ReleasesDGX agent

arXiv:2605.27482v1 Announce Type: cross Abstract: While orthogonal subspace methods try to mitigate task interference in Continual Learning (CL), they often suffer from energy diffusion across the bas

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

Every9D-21M: Large-Scale Real-World 9D Canonicalization of Everyday Objects

Model ReleasesDGX agent

arXiv:2605.28270v1 Announce Type: new Abstract: Estimating the 9D pose of everyday objects from a single real-world image remains challenging. This is largely due to the lack of large-scale supervisio

From Causal Discovery to Dynamic Causal Inference in Neural Time Series

ApplicationsDGX agent

arXiv:2603.20980v2 Announce Type: replace-cross Abstract: Time-varying causal models provide a powerful framework for studying dynamic scientific systems, yet most existing approaches assume that the

From Fact Overwriting to Knowledge Evolution: Causal Editing via On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2605.28303v1 Announce Type: new Abstract: While Knowledge Editing (KE) enables efficient updates, its dominant Static Fact Overwriting paradigm treats LLMs as discrete databases, forcibly inject

Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

SafetyDGX agent

arXiv:2605.27970v1 Announce Type: new Abstract: While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometr

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍

Model ReleasesDGX agent

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍 JUST IN: Anthropic announces it will roll out Claude Mythos “in the com

GraD-IBD: Graph Representation Learning from Diagnosis Trajectories for Early Detection of Inflammatory Bowel Disease

ApplicationsDGX agent

arXiv:2605.27799v1 Announce Type: new Abstract: International Classification of Diseases (ICD) is a globally recognized coding system that records diagnostic events during each patient encounter, prov

Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills

Model ReleasesDGX agent

arXiv:2604.05333v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on reusable skills, and as they interact with personal applications, web browsers, and other interfaces, skill l

HRBench: Benchmarking and Understanding Thinking-Mode Switch Strategies in Hybrid-Reasoning LLMs

ResearchDGX agent

arXiv:2605.28398v1 Announce Type: new Abstract: Hybrid-reasoning large language models (LLMs) expose explicit controls over reasoning effort, allowing users or systems to trade off answer quality agai

HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning

Model ReleasesDGX agent

arXiv:2605.27724v1 Announce Type: cross Abstract: Imitation learning is a promising approach for training humanoid robots to both walk and manipulate, but it requires a large number of demonstrations,

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in tw…

Model ReleasesDGX agent

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in twigl, make it like an infinite city of neo-gothic towers part

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to…

Model ReleasesDGX agent

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to see whether it even did the complete processing or not 😅 ht

Imitation Learning for Robot Assistance in Open Surgery: A Multi-Policy Evaluation on Suture Following

Model ReleasesDGX agent

arXiv:2605.28736v1 Announce Type: new Abstract: This study presents the first evaluation of general-purpose imitation learning for surgeon-robot collaborative assistance in open surgery, targeting sut

InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training

ResearchDGX agent

arXiv:2510.15859v4 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has driven recent breakthroughs in large language models (LLMs), especially for tasks where rewards can be compute

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

Model ReleasesDGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.27748v1 Announce Type: cross Abstract: Industrial visual anomaly detection is usually one-class: normal images are abundant, while defects are rare, heterogeneous, and often unavailable dur

MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation

Model ReleasesDGX agent

arXiv:2605.28173v1 Announce Type: new Abstract: End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page la

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs

ResearchDGX agent

arXiv:2605.28388v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) is empirically shown to notably enhance the reasoning performance of large language models (LLMs),

Meta-Attention: Bayesian Per-Token Routing for Efficient Transformer Inference

Model ReleasesDGX agent

arXiv:2605.28384v1 Announce Type: new Abstract: Standard transformer architectures apply a single attention mechanism uniformly across all tokens and sequence positions, irrespective of local context

Mitigating Staleness in Asynchronous Pipeline Parallelism via Basis Rotation

Model ReleasesDGX agent

arXiv:2602.03515v2 Announce Type: replace-cross Abstract: Asynchronous pipeline parallelism maximizes hardware utilization by eliminating the pipeline bubbles inherent in synchronous execution, offeri

MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks

Local AiDGX agent

arXiv:2502.17832v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has become a common practice in multimodal large language models (MLLM) to enhance factual grounding and

Moment Matters: Mean and Variance Causal Graph Discovery from Heteroscedastic Observational Data

Model ReleasesDGX agent

arXiv:2602.23602v2 Announce Type: replace-cross Abstract: Heteroscedasticity -- where the variance of a variable changes with other variables -- is pervasive in real data, and elucidating why it arise

Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Sokoban RNN

ResearchDGX agent

arXiv:2506.10138v3 Announce Type: replace-cross Abstract: We partially reverse-engineer a convolutional recurrent neural network (RNN) trained with model-free reinforcement learning to play the box-pu

Pattern Recognition Tasks with Personalized Federated Learning

ResearchDGX agent

arXiv:2605.27816v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) constitutes a novel paradigm that tailors Machine Learning (ML) models to individual clients, thereby furnishing p

PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation

Model ReleasesDGX agent

arXiv:2601.18006v2 Announce Type: replace Abstract: We present PEAR (Pairwise Evaluation for Automatic Relative Scoring), a supervised quality estimation (QE) metric family that reframes reference-fre

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

Model ReleasesDGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations

Model ReleasesDGX agent

arXiv:2605.27958v1 Announce Type: cross Abstract: Linear probes trained on LLM activations are increasingly proposed as deception-detection metrics, yet report AUROC exceeding 0.96 on clean benchmarks

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

TutorialsDGX agent

arXiv:2605.28634v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency an

Privately Estimating Monotone Statistics in Polynomial Time

Model ReleasesDGX agent

arXiv:2605.27912v1 Announce Type: cross Abstract: We study efficient differentially private algorithms for estimating monotone statistics, i.e., statistics that are monotone under the addition of new

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation

ResearchDGX agent

arXiv:2605.28230v1 Announce Type: new Abstract: Modern video generative models produce visually impressive results, yet frequently violate basic physical principles. We propose Proprio, a training-fre

ProvMind: Provenance-grounded reasoning for materials synthesis

Model ReleasesDGX agent

arXiv:2605.28487v1 Announce Type: new Abstract: Materials process optimization requires reasoning over routes, conditions, tools and causal dependencies, yet most computational formulations flatten sy

QuITE: Query-Based Irregular Time Series Embedding

ApplicationsDGX agent

arXiv:2605.28166v1 Announce Type: cross Abstract: Irregular Multivariate Time Series (IMTS) are common in practice, yet their irregular sampling complicates effective modeling. Existing approaches typ

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG

Model ReleasesDGX agent

arXiv:2605.28044v1 Announce Type: new Abstract: Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wo

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

Model ReleasesDGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure

SafetyDGX agent

arXiv:2605.27996v1 Announce Type: new Abstract: Single-axis mitigations of reward-model biases (e.g., reducing proxy reliance on length, sycophancy, or style) can rotate optimization pressure onto cor

ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning

Local AiDGX agent

arXiv:2605.27959v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have increasingly localized and interleaved visual evidence for deliberative reasoning. Grounding-based appro

SeeGroup: Multi-Layer Depth Estimation of Transparent Surfaces via Self-Determined Grouping

Model ReleasesDGX agent

arXiv:2605.28735v1 Announce Type: new Abstract: Transparent objects are common in daily life, and it is important to understand their multilayer depth, including the transparent surface and the object

Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments

Model ReleasesDGX agent

arXiv:2605.28442v1 Announce Type: cross Abstract: Self-supervised online traversability estimation enables robots to continuously learn from unlabeled open-world experiences and adapt their navigation

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

Model ReleasesDGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

Sparse POD Mode Selection and Manifold Dimensionality Reduction with Neural Networks

Model ReleasesDGX agent

arXiv:2605.27756v1 Announce Type: cross Abstract: High-performance computing enables simulation of high-dimensional physical systems, but downstream analyses such as inverse problems and control remai

Sparse Scheduled Diffusion Guidance for Inverse Problems

ResearchDGX agent

arXiv:2603.07860v2 Announce Type: replace Abstract: Pretrained diffusion models are effective priors for Bayesian inverse problems, but posterior sampling with these priors is often costly because dat

Stochastic Gradient Descent with Momentum is Algorithmically Stable

Model ReleasesDGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

Structure-Guided Visual Perturbation Neutralization for LVLMs

SafetyDGX agent

arXiv:2605.27927v1 Announce Type: new Abstract: Image inputs enable Large Vision Language Models (LVLMs) to perceive fine-grained visual information, but also introduce a pixel-level attack surface th

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

Model ReleasesDGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭

Model ReleasesDGX agent

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭 Learnings from testing Claude Opus 4.8: > Much worse than Opus 4.7 and GPT 5.5 on Vending Bench > More aligne

Using Zero-Shot LLM-Generated Survey Data for Geographically Explicit Population Synthesis

Model ReleasesDGX agent

arXiv:2605.27401v1 Announce Type: cross Abstract: There is a growing interest in utilizing synthetic populations for a diverse range of applications. At the same time, we are witnessing a tremendous g

← Previous
1…546547548549550…1061
Next →