AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
12 May 2026

Workspace Optimization: How to Train Your Agent

AgentsDGX agent

arXiv:2605.09650v1 Announce Type: new Abstract: Modern agents built on frontier language models often cannot adapt their weights. What, then, remains trainable? We argue it is the agent's workspace, t

Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

Model ReleasesDGX agent

arXiv:2605.09360v1 Announce Type: cross Abstract: Execution-based evaluation of LLM-generated code implicitly treats successful execution as a proxy for correctness. In scientific simulation, this pro

Z-Erase: Enabling Concept Erasure in Single-Stream Diffusion Transformers

SafetyDGX agent

arXiv:2603.25074v2 Announce Type: replace Abstract: Concept erasure serves as a vital safety mechanism for removing unwanted concepts from text-to-image (T2I) models. While extensively studied in U-Ne

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Zero-Shot Chinese Character Recognition via Global-Local Dual-Branch Alignment and Hierarchical Inference

Model ReleasesDGX agent

arXiv:2605.08814v1 Announce Type: new Abstract: Chinese character categories are extremely large, and unseen characters frequently arise in open-world scenarios, making zero-shot Chinese character rec

11 May 2026

Adaptive Regularization for Sparsity Control in Bregman-Based Optimizers

Model ReleasesDGX agent

arXiv:2605.07892v1 Announce Type: new Abstract: Sparse training reduces the memory and computational costs of deep neural networks. However, sparse optimization methods, e.g., those adding an ell_1 pe

AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

Model ReleasesDGX agent

arXiv:2605.06607v2 Announce Type: replace-cross Abstract: Recent LLM-based agents have closed substantial portions of the scientific discovery loop in software-only machine-learning research, in chemi

@Alibaba_Qwen Sign up to get free access to Qwen 3.6 Plus and much more at http://portal.nousresearch.com/manage-subscription!

Model ReleasesDGX agent

Nous Research announced free access to Qwen 3.6 Plus and additional features available through their subscription portal at portal.nousresearch.com. This promotion likely provides users with complimen

Amortized Molecular Optimization via Group Relative Policy Optimization

Model ReleasesDGX agent

arXiv:2602.12162v3 Announce Type: replace Abstract: In structurally constrained molecular optimization, state-of-the-art methods restart an expensive oracle-driven search from scratch for every new in

Attention Transfer Is Not Universally Effective for Vision Transformers

Model ReleasesDGX agent

arXiv:2605.07191v1 Announce Type: new Abstract: A recent work shows that Attention Transfer, which transfers only the attention patterns from a pre-trained teacher Vision Transformer (ViT) to a random

Attribution-Based Neuron Utility for Plasticity Restoration in Deep Networks

Model ReleasesDGX agent

arXiv:2605.06834v1 Announce Type: new Abstract: Continual learning research attempts to conserve two fundamental capabilities: new knowledge acquisition and the preservation of previously acquired kno

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2

Local AiDGX agent

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2 discusses comparisons between these lightweight AI image generation models, with Flux.2 Klein 9B showing good prompt adherence and natural composi

Beyond 'I cannot fulfill this request': Alleviating Rigid Rejection in LLMs via Label Enhancement

SafetyDGX agent

arXiv:2605.07883v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on safety alignment to obey safe requests while refusing harmful ones. However, traditional refusal mechanisms often l

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

Model ReleasesDGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation

Model ReleasesDGX agent

arXiv:2605.08057v1 Announce Type: cross Abstract: While recent advancements in inference-time learning have improved LLM reasoning on Text-to-SQL tasks, current solutions still struggle to perform wel

CASCADE: Context-Aware Relaxation for Speculative Image Decoding

ResearchDGX agent

arXiv:2605.07230v1 Announce Type: cross Abstract: Autoregressive generation is a powerful approach for high-fidelity image synthesis, but it remains computationally demanding and slow even on the most

CommandSwarm: Safety-Aware Natural Language-to-Behavior-Tree Generation for Robotic Swarms

Model ReleasesDGX agent

arXiv:2605.07764v1 Announce Type: new Abstract: Natural-language interfaces can make swarm robotics more accessible to non-expert operators, but they must translate ambiguous user intent into executab

Congrats to @thinkymachines on the release of TML-Interaction-Small and tying for the top spot on our Audio MC S2S leaderboard! 🥇 Their int…

ResearchDGX agent

Congrats to @thinkymachines on the release of TML-Interaction-Small and tying for the top spot on our Audio MC S2S leaderboard! 🥇 Their interaction model scores a 43.4% APR, demonstrating an impressiv

Convergent Stochastic Training of Attention and Understanding LoRA

ResearchDGX agent

arXiv:2605.07959v1 Announce Type: new Abstract: Transformers have revolutionized machine learning and deploying attention layers in the model is increasingly standard across a myriad of applications.

Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences

SafetyDGX agent

arXiv:2605.07724v1 Announce Type: cross Abstract: Recursive retraining of generative models poses a critical representation challenge: when synthetic outputs are curated based on a fixed reward signal

DeepSeek V4 Flash is ~90% cheaper than GPT 5.4 Mini and ~70% cheaper than Gemini 3.1 Flash Lite For devs pushing ~500M tok/month, this is th…

Model ReleasesDGX agent

DeepSeek V4 Flash is ~90% cheaper than GPT 5.4 Mini and ~70% cheaper than Gemini 3.1 Flash Lite For devs pushing ~500M tok/month, this is the difference between: GPT 5.4 Mini: ~394/mo Gemini 3.1 Flash

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

SafetyDGX agent

arXiv:2605.07503v1 Announce Type: new Abstract: Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent dis

Disagreement-Regularized Importance Sampling for Adversarial Label Corruption

Model ReleasesDGX agent

arXiv:2605.07551v1 Announce Type: new Abstract: Standard Importance Sampling (IS) collapses under label corruption because high-norm examples, prioritized for variance reduction, are often adversarial

Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative Evaluation

Model ReleasesDGX agent

arXiv:2605.07323v1 Announce Type: new Abstract: Discovering governing differential equations from observational data is a fundamental challenge in scientific machine learning. Existing symbolic regres

DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain

Model ReleasesDGX agent

arXiv:2605.07699v1 Announce Type: cross Abstract: LLM-based agents are increasingly deployed for routine but consequential tasks in real-world domains, where their behavior is governed by inherently a

EditRefiner: A Human-Aligned Agentic Framework for Image Editing Refinement

Local AiDGX agent

arXiv:2605.07457v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have made remarkable progress, yet edited images still frequently suffer from fine-grained issues such as

Enhancing Federated Quadruplet Learning: Stochastic Client Selection and Embedding Stability Analysis

ResearchDGX agent

arXiv:2605.07888v1 Announce Type: cross Abstract: Federated Learning (FL) enables decentralised model training across distributed clients without requiring data centralisation. However, the generalisa

EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation

Model ReleasesDGX agent

arXiv:2605.07247v1 Announce Type: new Abstract: Scalable AI agents training relies on interactive environments that faithfully simulate the consequences of agent actions. Manually crafted environments

Exact Flow Linear Attention: Exact Solution from Continuous-Time Dynamics

Model ReleasesDGX agent

arXiv:2512.12602v4 Announce Type: replace Abstract: In this paper, we introduce Exact Flow Linear Attention~(EFLA), an exact-flow formulation of delta-rule linear attention. We show that the delta-rul

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

Model ReleasesDGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

ExpThink: Experience-Guided Reinforcement Learning for Adaptive Chain-of-Thought Compression

ResearchDGX agent

arXiv:2605.07501v1 Announce Type: cross Abstract: Large reasoning models (LRMs) achieve strong performance via extended chain-of-thought (CoT) reasoning, yet suffer from excessive token consumption an

From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG

Model ReleasesDGX agent

arXiv:2605.07273v1 Announce Type: cross Abstract: Multimodal RAG systems increasingly rely on vision-language retrievers to ground visual queries in external textual evidence. Existing adversarial stu

Globally Optimal Training of Spiking Neural Networks via Parameter Reconstruction

Model ReleasesDGX agent

arXiv:2605.08022v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have been proposed as biologically plausible and energy-efficient alternatives to conventional Artificial Neural Networ

Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.06347v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in robotics faces significant hurdles regarding sample efficiency and generalization across varying goals. While O

Graph-Structured Hyperdimensional Computing for Data-Efficient and Explainable Process-Structure-Property Prediction

Model ReleasesDGX agent

arXiv:2605.07999v1 Announce Type: cross Abstract: Multiphoton photoreduction enables high-fidelity fabrication of complex 3D microstructures, yet reliable process-structure-property (PSP) prediction r

GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations

Model ReleasesDGX agent

arXiv:2605.07053v1 Announce Type: cross Abstract: Benchmarks like GSM8K are popular measures of mathematical reasoning, but leaderboard gains can overstate true capability due to memorization of fixed

Hierarchical Task Network Planning with LLM-Generated Heuristics

Model ReleasesDGX agent

arXiv:2605.07707v1 Announce Type: new Abstract: HTN planning is a variation of classical planning where, instead of searching for a linear sequence of actions, an algorithm decomposes higher-level tas

Implicit Preference Alignment for Human Image Animation

Model ReleasesDGX agent

arXiv:2605.07545v1 Announce Type: cross Abstract: Human image animation has witnessed significant advancements, yet generating high-fidelity hand motions remains a persistent challenge due to their hi

Inference-Time Attribute Distribution Alignment for Unconditional Diffusion

SafetyDGX agent

arXiv:2605.07456v1 Announce Type: new Abstract: Inference-time controllable generation is essential for real-world applications of unconditional diffusion models. However, most existing techniques foc

Instruction Tuning Changes How Upstream State Conditions Late Readout: A Cross-Patching Diagnostic

Local AiDGX agent

arXiv:2605.07284v1 Announce Type: new Abstract: Recent interpretability work has identified model-internal handles on post-trained behavior, including refusal directions, assistant/persona axes, and s

Intent-Driven Semantic ID Generation for Grounded Conversational News Recommendation

Model ReleasesDGX agent

arXiv:2605.07613v1 Announce Type: new Abstract: Conversational news recommendation requires grounding each suggestion in a rapidly evolving article corpus while addressing implicit user intents that l

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

Model ReleasesDGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

Knowledge Transfer Scaling Laws for 3D Medical Imaging

ResearchDGX agent

arXiv:2605.06859v1 Announce Type: cross Abstract: Vision foundation models are increasingly moving beyond 2D to volumetric domains such as 3D medical imaging, where unified pretraining across differen

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

Model ReleasesDGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

Learning Minimal-Deviation Corrections for Multi-Dimensional Mismodelling in HEP Simulations

ResearchDGX agent

arXiv:2605.07460v1 Announce Type: new Abstract: Accurate Monte Carlo (MC) modelling in high-energy physics is challenging, particularly in complex scenarios where simulations fail to reproduce observe

Learning to Pose Problems: Reasoning-Driven and Solver-Adaptive Data Synthesis

ResearchDGX agent

arXiv:2511.09907v5 Announce Type: replace Abstract: Data synthesis for training large reasoning models offers a scalable alternative to limited, human-curated datasets, enabling the creation of high-q

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning

Local AiDGX agent

arXiv:2605.07505v1 Announce Type: new Abstract: Developing lightweight, on-device vision-language GUI agents is essential for efficient cross-platform automated interaction. However, current on-device

Local open-weight AI on a laptop has been improving more than twice as fast as Moore's Law! Between May 2024 and May 2026, the most expensiv…

Model ReleasesDGX agent

Local open-weight AI on a laptop has been improving more than twice as fast as Moore's Law! Between May 2024 and May 2026, the most expensive MacBook Pro you could buy stayed at 128 GB of unified memo

MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge

SafetyDGX agent

arXiv:2507.21183v5 Announce Type: replace-cross Abstract: As the era of large language models (LLMs) unfolds, Preference Optimization (PO) methods have become a central approach to aligning LLMs with

Mask2Cause: Causal Discovery via Adjacency Constrained Causal Attention

Model ReleasesDGX agent

arXiv:2605.07280v1 Announce Type: cross Abstract: Leveraging deep learning for causal discovery in time series remains challenging because existing neural methods predominantly rely on component-wise

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators

Model ReleasesDGX agent

arXiv:2605.07600v1 Announce Type: cross Abstract: Recent methods for improving LLM mathematical reasoning, whether through MCTS-based test-time search or causal graph-guided knowledge injection, canno

McNdroid: A Longitudinal Multimodal Benchmark for Robust Drift Detection in Android Malware

Model ReleasesDGX agent

arXiv:2605.06894v1 Announce Type: cross Abstract: Machine learning (ML) in real-world systems must contend with concept drift, adversarial actors, and a spectrum of potential features with varying cos

Meet the latest Database Center, now with Gemini-powered fleet intelligence

Model ReleasesDGX agent

Managing a modern database fleet is both a scale and cognitive problem. As database estates grow, the effort required to monitor, troubleshoot, and optimize them often outpaces teams’ capacity, who fi

MIPIAD: Multilingual Indirect Prompt Injection Attack Defense with Qwen -- TF-IDF Hybrid and Meta-Ensemble Learning

Model ReleasesDGX agent

arXiv:2605.07269v1 Announce Type: new Abstract: Indirect prompt injection remains a persistent weakness in retrieval-augmented and tool-using LLM systems, and the problem becomes harder to characteris

Multi-Objective Constraint Inference using Inverse reinforcement learning

Model ReleasesDGX agent

arXiv:2605.06951v1 Announce Type: new Abstract: Constraint inference is widely considered essential to align reinforcement learning agents with safety boundaries and operational guidelines by observin

Neural Operators as Efficient Function Interpolators

Model ReleasesDGX agent

arXiv:2605.07792v1 Announce Type: cross Abstract: Neural operators (NOs) are designed to learn maps between infinite-dimensional function spaces. We propose a novel reframing of their use. By introduc

New TIL: I figured out how to use my LLM CLI tool in a shebang line, which means you can write executable scripts in English, or hook up mor…

Model ReleasesDGX agent

Simon Willison discovered how to use an LLM command-line interface tool in Unix shebang lines, enabling the creation of executable scripts written in English or natural language. This technique allows

NPMixer: Hierarchical Neighboring Patch Mixing for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.07476v1 Announce Type: new Abstract: Multivariate time series forecasting remains a challenge due to the complexity of local temporal dynamics and global dependencies across multiple variab

On the Robustness of Distribution Support under Diffusion Guidance

ResearchDGX agent

arXiv:2605.07220v1 Announce Type: new Abstract: Diffusion guidance is a powerful technique that enables controllable and high-fidelity sample generation with diffusion models. At a high level, it modi

On Time, Within Budget: Constraint-Driven Online Resource Allocation for Agentic Workflows

AgentsDGX agent

arXiv:2605.06110v2 Announce Type: replace Abstract: Agentic systems increasingly solve complex user requests by executing orchestrated workflows, where subtasks are assigned to specialized models or t

OpenAI Campus Network: Student club interest form

Model ReleasesDGX agent

The OpenAI Campus Network is a program that facilitates student engagement with OpenAI's technology and research on college campuses. This interest form allows students to express interest in starting

← Previous
1…563564565566567…1061
Next →