AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
1 Jun 2026

LARK: Learnability-Grounded Trajectory Selection for Efficient Reasoning Distillation

SafetyDGX agent

arXiv:2605.30651v1 Announce Type: cross Abstract: We study trajectory selection for reasoning distillation, where teacher-generated reasoning trajectories are selectively used as supervision for a stu

Learning from Fine-Grained Visual Discrepancies: Mitigating Multimodal Hallucinations via In-Context Visual Contrastive Optimization

ResearchDGX agent

arXiv:2605.31312v1 Announce Type: cross Abstract: Multimodal hallucination remains a persistent challenge for Vision-Language Models (VLMs). Standard textual Direct Preference Optimization (DPO) often

Learning Global Motion with Compact Gaussians for Feed-Forward 4D Reconstruction

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.31595v1 Announce Type: new Abstract: Dynamic scene reconstruction from monocular video remains a fundamental challenge in computer vision. Existing feed-forward methods predict 3D Gaussians

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration

AgentsDGX agent

arXiv:2605.31365v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to promising progress in web agents. However, existing web agents often rely on han

LH-Bench: Skill-Grounded Evaluation of Long-Horizon Agents on Subjective Enterprise Tasks

AgentsDGX agent

arXiv:2603.22744v2 Announce Type: replace Abstract: Large language models excel on objectively verifiable tasks such as math and programming, where evaluation reduces to unit tests or a single correct

LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability

Local AiDGX agent

arXiv:2605.31167v1 Announce Type: new Abstract: Assessing whether Large Language Models outputs are factually grounded, epistemically calibrated, and methodologically reproducible is a prerequisite fo

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks

AgentsDGX agent

arXiv:2603.02630v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved great success in many real-world applications, especially the one serving as the cognitive backbone

Mechanistic Interpretability as Statistical Estimation: A Variance Analysis

ResearchDGX agent

arXiv:2510.00845v4 Announce Type: replace-cross Abstract: Mechanistic Interpretability (MI) aims to reverse-engineer model behaviors by identifying functional sub-networks. Yet, the scientific validit

MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation

AgentsDGX agent

arXiv:2602.07905v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong potential in complex medical reasoning yet face diminishing gains under inference scaling laws. While

MiniMax M3 is live and Together AI is powering its inference 🚀 Tomorrow at 6pm PT we're going live on X Spaces with the teams behind the mo…

ToolsDGX agent

MiniMax M3 is live and Together AI is powering its inference 🚀 Tomorrow at 6pm PT we're going live on X Spaces with the teams behind the model and the infrastructure to give you a deep dive. https://x

MiniMax-M3 will by arrive on HuggingFace openweight at next week!

AgentsDGX agent

MiniMax-M3 will by arrive on HuggingFace openweight at next week! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Ben

MoE-dqINR: A Unified Mixture-of-Experts Implicit Neural Representation Framework for Scan-Specific Dynamic and Quantitative MRI Reconstruction

ResearchDGX agent

arXiv:2605.31302v1 Announce Type: cross Abstract: Undersampled magnetic resonance imaging (MRI) reconstruction seeks to recover temporally or contrast-varying image series from incomplete multicoil k-

MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.31010v1 Announce Type: new Abstract: Retrieval-augmented generation is intensively studied to ground large language models on external evidence. However, retrieving from a unified knowledge

Multilingual and Cross-Lingual Citation Needed Detection on Wikipedia for Lower-Resource Languages

ResearchDGX agent

arXiv:2605.31136v1 Announce Type: new Abstract: In automated fact-checking (AFC), check-worthiness detection identifies claims requiring verification based on domain-specific criteria. On Wikipedia, t

Multivariate Distributional Reinforcement Learning Using Sliced Divergences

ResearchDGX agent

arXiv:2605.31222v1 Announce Type: new Abstract: Distributional reinforcement learning (DRL) models the full return distribution rather than expectations, but extending it to multivariate settings rema

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making

Local AiDGX agent

arXiv:2605.31460v1 Announce Type: new Abstract: Reasoning-based robotic policies using large language and vision-language models achieve strong semantic planning capabilities but mostly suffer from a

On Revisiting Entropy for Identifying Mislabeled Images

ResearchDGX agent

arXiv:2605.31090v1 Announce Type: cross Abstract: Mislabeled samples in training datasets severely degrade the performance of deep networks, as overparameterized models tend to memorize erroneous labe

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment

SafetyDGX agent

arXiv:2605.30991v1 Announce Type: cross Abstract: Inference-time reward alignment steers pretrained diffusion and flow-based generative models to satisfy user-specified rewards without retraining. Rec

PithTrain: A Compact and Agent-Native MoE Training System

HardwareDGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

Primitive Subspaces Mediate Few-Shot Transfer in VLAs

SafetyDGX agent

arXiv:2605.30695v1 Announce Type: new Abstract: Deploying vision-language-action (VLA) policies in industrial environments requires the ability to teach new tasks at low cost, a property current VLAs

Rank-Factorized Implicit Neural Bias: Scaling Super-Resolution Transformer with FlashAttention

Local AiDGX agent

arXiv:2603.06738v2 Announce Type: replace-cross Abstract: Recent Super-Resolution~(SR) methods mainly adopt Transformers for their strong long-range modeling capability and exceptional representationa

RDGen: Demonstration Generation for High-Quality Robot Learning via Reinforcement Learning

SafetyDGX agent

arXiv:2605.30957v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for general-purpose robot control. However, their performance remains fundament

Read the full announcement: https://cursor.com/blog/teams-pricing-june-2026

ToolsDGX agent

Cursor announced updates to its Teams pricing model in June 2026, as detailed in a blog post linked from their official X account. The announcement covers pricing changes or new features related to Cu

Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't

ResearchDGX agent

arXiv:2605.30523v1 Announce Type: cross Abstract: Recent work describes what transformers can and cannot compute through connections to boolean circuits, but existing results lack exact characterizati

S^3LDBO: A Snapshot Single-Loop Algorithm for Decentralized Bilevel Optimization

AgentsDGX agent

arXiv:2605.31311v1 Announce Type: cross Abstract: Networked AI systems increasingly rely on multiple agents that collaboratively learn and adapt models over communication networks. In such systems, bi

SCOPE: Selective Conformal Optimized Pairwise LLM Judging

SafetyDGX agent

arXiv:2602.13110v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as scalable judges in pairwise evaluation, but they remain prone to miscalibration and bias

Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence

SafetyDGX agent

arXiv:2605.30698v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on visual question answering (VQA). To mitigate individual hallucinations and blind spo

Skill Reuse as Compression in Agentic RL

AgentsDGX agent

arXiv:2605.31509v1 Announce Type: cross Abstract: Large language model agents trained with reinforcement learning (RL) often learn brittle, task-specific shortcuts. We hypothesize that agents generali

SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation

ResearchDGX agent

arXiv:2605.31033v1 Announce Type: new Abstract: Streaming video generation models typically rely on temporal-centric memory, which organizes historical context as raw frames, chunk segments, or unclus

Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding

ResearchDGX agent

arXiv:2602.06161v2 Announce Type: replace-cross Abstract: Parallel diffusion decoding can accelerate diffusion language model inference by unmasking multiple tokens per step, but aggressive parallelis

The Inclusion Depth of Pattern Languages: An Open Problem in Algorithmic Learning Theory

ResearchDGX agent

arXiv:2605.30389v1 Announce Type: cross Abstract: Pattern languages are a classical model in formal language theory and algorithmic learning theory. This note formulates the problem of computing the i

The Inference Tax: How Prefix-Aware Routing Eliminates the Hidden Cost of LLMs at Scale

IndustryDGX agent

This article discusses how prefix-aware routing and prefix caching techniques can reduce the computational overhead and costs associated with running large language models at scale by eliminating redu

To Grok Grokking: Provable Grokking in Ridge Regression

ResearchDGX agent

arXiv:2601.19791v3 Announce Type: replace Abstract: We study grokking, the onset of generalization long after overfitting, in a classical ridge regression setting. We prove end-to-end grokking results

Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining

ApplicationsDGX agent

arXiv:2605.31069v1 Announce Type: cross Abstract: Accurately predicting future events is fundamental to content understanding and decision-making across various domains. While prior research has prima

Ubiquity of Emergent Hebbian Dynamics in Regularized Learning

SafetyDGX agent

arXiv:2505.18069v3 Announce Type: replace Abstract: Hebbian and anti-Hebbian plasticity are widely observed in the brain and are classically modeled as mechanistic, local homosynaptic rules stabilized

UniMedVL: Unifying Medical Multimodal Understanding and Generation through Observation-Knowledge-Analysis

ResearchDGX agent

arXiv:2510.15710v3 Announce Type: replace Abstract: Medical workflows routinely combine reading images with producing visual and textual outputs, making both image understanding and generation central

Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack

ResearchDGX agent

arXiv:2506.01318v4 Announce Type: replace-cross Abstract: Machine unlearning (MU) aims to expunge a designated forget set from a trained model without costly retraining, yet the existing techniques ov

Weights to Code: Extracting Interpretable Algorithms from the Discrete Transformer

ResearchDGX agent

arXiv:2601.05770v3 Announce Type: replace-cross Abstract: Algorithm extraction aims to synthesize executable programs directly from models trained on algorithmic tasks, enabling de novo recovery of ex

When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?

SafetyDGX agent

arXiv:2605.30719v1 Announce Type: cross Abstract: We study when large language models (LLMs) can serve as effective black-box policy optimizers for reinforcement learning (RL) tasks, i.e., when can we

which is NOT new; see this quote from 5 years ago. the fact that all this is still true says a lot.

SafetyDGX agent

which is NOT new; see this quote from 5 years ago. the fact that all this is still true says a lot. This was right five years ago, and still is: “Large scale pretrained models are certainly likely to

XOResNet: Exclusive-OR Meta-Residuals Facilitate Deep Spiking Neural Networks Learning

ResearchDGX agent

arXiv:2605.30362v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) hold promise for demonstrating superior learning and representation capabilities in deep models. Given the tremendous s

Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation

SafetyDGX agent

arXiv:2605.30833v1 Announce Type: cross Abstract: On-policy distillation transfers reasoning capabilities by training a student model on its own generated trajectories using token-level feedback from

31 May 2026

Flux Identity Adjuster V2

Local AiDGX agent

Flux Identity Adjuster V2 is a ComfyUI node designed to improve identity consistency for FLUX.2 klein 9b models . It automatically throttles text strength when facial formation struggles, preventing c

your upbringing is your system prompt

IndustryDGX agent

This post likely argues that an individual's upbringing and early life experiences function similarly to a system prompt for an AI model—shaping foundational values, behaviors, and responses that infl

30 May 2026

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without…

AgentsDGX agent

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without calling a model. It can sit as a default step in agent and

b9434

Local AiDGX agent

B9434 is a release build identifier from the llama.cpp project, an open-source C/C++ implementation for large language model inference. Llama.cpp releases include updates, optimizations, and bug fixes

29 May 2026

A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search

ResearchDGX agent

arXiv:2602.11171v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) with Low-Rank Adaptation (LoRA) offers a resource-efficient way to personalize or specialize. However

A Predictive Law for On-Policy Self-Distillation From World Feedback

SafetyDGX agent

arXiv:2605.30070v1 Announce Type: cross Abstract: Moving beyond simple scalar rewards toward richer world feedback is a natural path to more scalable RL post-training. On-policy self-distillation (OPS

A unified deeplearning framework for contrast-phase-specific virtual monochromatic imaging

ResearchDGX agent

arXiv:2605.29753v1 Announce Type: cross Abstract: Dual-energy CT (DECT) enables virtual monochromatic imaging (VMI) and improved contrast resolution, but its clinical adoption is limited by hardware c

AG-REPA: Causal Layer Selection for Representation Alignment in Audio Flow Matching

SafetyDGX agent

arXiv:2603.01006v2 Announce Type: replace-cross Abstract: REPresentation Alignment (REPA) improves the training of generative flow models by aligning intermediate hidden states with pretrained teacher

AI Disruptors: How the Next Generation of Business is Being Built

IndustryDGX agent

This DigitalOcean blog post examines how emerging AI technologies are transforming business models and enabling startups to challenge traditional industries. It likely covers practical examples of AI-

Analyzing Persona Effects in Generated Explanations from Multimodal LLM Agents in Urban Perception

ResearchDGX agent

arXiv:2605.29064v1 Announce Type: new Abstract: We study how persona prompting shapes language generated by multimodal large language models in an urban perception setting. Using 59,808 annotations fr

Anytime-Valid Federated Conformal RAG for LLM Swarms

SafetyDGX agent

arXiv:2605.29139v1 Announce Type: cross Abstract: Federated Conformal RAG (FC-RAG) provides distribution-free coverage for a bandwidth-limited swarm of weak language models, but only at a fixed horizo

Aryabhata 2: Scaling Reinforcement Learning for Advanced STEM Reasoning

ApplicationsDGX agent

arXiv:2605.28829v1 Announce Type: cross Abstract: Competitive STEM examinations such as JEE and NEET require multi-step symbolic reasoning, precise numerical computation, and deep conceptual understan

b9393

Local AiDGX agent

B9393 is a build release of llama.cpp, the C/C++ inference framework for running large language models locally. llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a

b9401

Local AiDGX agent

b9401 is a release build of llama.cpp, an LLM inference framework in C/C++ that provides tools for running large language models locally. As an intermediate build in the llama.cpp project, it includes

b9402

Local AiDGX agent

B9402 is a release of llama.cpp, an open-source library that performs inference on large language models such as Llama and was developed in pure C/C++ with no dependencies. The release includes comman

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

HardwareDGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

Beyond Attack Success Rate: Temporal Logit Observability for LLM Safety Failures

SafetyDGX agent

arXiv:2605.29629v1 Announce Type: new Abstract: Attack Success Rate (ASR) evaluates each jailbreak with a single yes/no label at the end of generation, telling us whether a failure happened but not ho

Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors

ResearchDGX agent

arXiv:2605.30065v1 Announce Type: new Abstract: In this work, we focus on zero-shot 3D style transfer that can generate multi-view consistent stylized views of the 3D scene given an arbitrary style im

← Previous
1…759760761762763…1018
Next →