AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
9 Jun 2026

Inverse design of bespoke interatomic potentials via active learning by information-matching

Model ReleasesDGX agent

arXiv:2606.08148v1 Announce Type: cross Abstract: Interatomic potentials (IPs) enable large-scale atomistic simulations beyond the reach of first-principles methods, but their predictive reliability d

MaskAlign: Token-Subset Representation Alignment for Efficient Diffusion Training

SafetyDGX agent

arXiv:2606.08788v1 Announce Type: new Abstract: Representation alignment with pretrained vision models has recently shown strong potential for accelerating diffusion transformer training. By aligning

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

Model ReleasesDGX agent

arXiv:2605.22664v2 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet ente

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT

HardwareDGX agent

arXiv:2601.20408v2 Announce Type: replace-cross Abstract: Enterprise LLM deployment faces a critical scalability challenge: organizations must optimize models systematically to scale AI initiatives wi

Mesh Graph Neural Network Framework for Accelerating Finite Element Simulation for Arbitrary Geometries

ResearchDGX agent

arXiv:2606.08287v1 Announce Type: new Abstract: Finite element analysis (FEA) is essential for structural design but remains computationally expensive, particularly when evaluating multiple design ite

MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

AgentsDGX agent

arXiv:2606.07603v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning capabilities, yet most LLM-based agents are statically deployed and unable to improve through ta

mythos will be bad ON PURPOSE on ai 'frontier llm research' tasks, this is very very sad for the research community also the fact that this …

Model ReleasesDGX agent

mythos will be bad ON PURPOSE on ai 'frontier llm research' tasks, this is very very sad for the research community also the fact that this is un purpose not visible to the user is crazy Introducing C

Normality Calibration in Semi-supervised Graph Anomaly Detection

SafetyDGX agent

arXiv:2510.02014v3 Announce Type: replace Abstract: Graph anomaly detection (GAD) has attracted growing interest for its crucial ability to uncover irregular patterns in broad applications. Semi-super

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

Model ReleasesDGX agent

arXiv:2606.07550v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-er

OmniFaceRig: Fully Automatic Inner-Mouth-Aware Face Rigging Across Diverse 3D Character Topologies

Model ReleasesDGX agent

arXiv:2606.08043v1 Announce Type: cross Abstract: Facial rigging - creating FACS-based blendshapes together with inner-mouth geometry (teeth, gums, and tongue) - remains a major bottleneck in 3D chara

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

ResearchDGX agent

arXiv:2606.09585v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves the performance of Large Language Models (LLMs) and has been extended to Multimodal Large Language Models (MLLMs). More

PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

Model ReleasesDGX agent

arXiv:2606.08481v1 Announce Type: cross Abstract: Enterprise property graphs vary widely in schema structure, internal terminology, domain assumptions, governance constraints, and user interaction pat

Post-training is (Massive) Supervised Learning

TutorialsDGX agent

arXiv:2606.07527v1 Announce Type: cross Abstract: The prevailing paradigm for training LLMs has evolved to rely on a massive post-training phase consisting of SFT and RL. In this position paper, we ar

Real-time body pose non-verbal communication with a consistency-based reliability measure

Model ReleasesDGX agent

arXiv:2606.09390v1 Announce Type: cross Abstract: Body movement communicates intent at distances and in conditions where neither the face, nor speech can be captured. We study the recognition of commu

Reformulate LLM Reinforcement Learning for Efficient Training under Black-box Discrepancy

Model ReleasesDGX agent

arXiv:2606.08779v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a pivotal post-training paradigm, yet it frequently suffers from unpredictable sub-optimum performance or eve

ReTabSyn: Realistic Tabular Data Synthesis via Reinforcement Learning

TutorialsDGX agent

arXiv:2603.10823v2 Announce Type: replace-cross Abstract: Deep generative models can help with data scarcity and privacy by producing synthetic training data, but they struggle in low-data, imbalanced

RetroReasoner: A Reasoning LLM for Strategic Retrosynthesis Prediction

ResearchDGX agent

arXiv:2603.12666v2 Announce Type: replace-cross Abstract: Retrosynthesis prediction aims to identify reactants that can synthesize a given product molecule. Although molecular large language models (L

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

SafetyDGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

SAILS: Surrogate-based Analysis of Interactions via Local Effect Smooths

Local AiDGX agent

arXiv:2606.09404v1 Announce Type: cross Abstract: Feature interactions drive much of the predictive power of machine learning models, yet existing explanation methods only detect and quantify interact

scCBGM: Interpretable Single-Cell Counterfactual Editing

Model ReleasesDGX agent

arXiv:2606.07760v1 Announce Type: new Abstract: Understanding cellular phenotypes and how they respond to perturbations is critical for disease biology and therapeutic design. Single-cell RNA sequenci

SurfDesign: Effective Protein Design on Molecular Surfaces

Model ReleasesDGX agent

arXiv:2606.07567v1 Announce Type: cross Abstract: Protein function is largely determined by molecular surface geometry and physicochemical complementarity, yet most protein design methods condition on

Temporal-Aware Reasoning Optimization for Video Temporal Grounding

Local AiDGX agent

arXiv:2606.09248v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) have achieved remarkable progress in video temporal grounding with reinforcement learning for generating reaso

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

Model ReleasesDGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation

AgentsDGX agent

arXiv:2606.08310v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on

Transfer learning for causal forest

ApplicationsDGX agent

arXiv:2606.07693v1 Announce Type: cross Abstract: Transfer learning addresses the challenge of transfering knowledge from one domain to another. Traditional transfer learning focuses on adapting model

TT-DAC-PS: Twin-Target Deterministic Actor-Critic with Policy Smoothing for Optimal Trade Execution

Model ReleasesDGX agent

arXiv:2606.08379v1 Announce Type: new Abstract: This study addresses the optimal execution of large stock sell programs by introducing TT-DAC-PS (Twin-Target Deterministic Actor-Critic with Policy Smo

Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data

SafetyDGX agent

arXiv:2606.08520v1 Announce Type: new Abstract: Vision-language models (VLMs) are powerful general-purpose reasoners, yet converting them into robot control policies (VLAs) is surprisingly difficult.

Ultra Flash: Scaling Real-Time Streaming Video Generation to High Resolutions

HardwareDGX agent

arXiv:2606.09150v1 Announce Type: new Abstract: While recent autoregressive video diffusion models achieve remarkable streaming quality, they remain confined to low resolutions (e.g., 480P), leaving e

Understanding Quantization-Aware Training: Gradients at Quantized Weights Bias to the Low-Loss Basin

Local AiDGX agent

arXiv:2606.09012v1 Announce Type: cross Abstract: Post-training quantization (PTQ) converts a trained full-precision model into low-bit weights without task-level retraining, while quantization-aware

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation

Model ReleasesDGX agent

arXiv:2606.08091v1 Announce Type: new Abstract: Recent agent frameworks such as Claude Code, Codex, and OpenClaw are strong at tool use and orchestration, but whether they can handle long video genera

Why Limit the Residual Stream to Layers and Not Tokens? Persistent Memory for Continuous Latent Reasoning

ResearchDGX agent

arXiv:2606.07720v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable reasoning abilities on mathematical and multi-hop planning tasks. The CoCoNuT (Chain of Contin

8 Jun 2026

A French engineer who lives quietly in Paris has spent 30 years writing software that the entire internet now runs on without knowing his na…

Model ReleasesDGX agent

A French engineer who lives quietly in Paris has spent 30 years writing software that the entire internet now runs on without knowing his name. He wrote the code that streams every YouTube video, ever

Adversarial Creation and Detection of AI-Generated Social Bot Content

ApplicationsDGX agent

arXiv:2606.07219v1 Announce Type: new Abstract: The convergence of large language models and social bots allows malicious actors to manipulate the information ecosystem by generating human-like conten

Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin

Model ReleasesDGX agent

arXiv:2606.05050v1 Announce Type: cross Abstract: Theoretical heterogeneous catalysis promises rapid catalyst discovery, yet computational and machine-learning predictions often deviate from experimen

Bradley-Terry Rankings for Recommender Systems Across Dataset Taxonomies

ResearchDGX agent

arXiv:2606.07492v1 Announce Type: cross Abstract: The ranking of recommendation algorithms is a challenging problem since model performance is sensitive to dataset characteristics such as sparsity, se

DEFINED: A Data-Efficient Computational Framework for Fine-Grained Creativity Assessment in Debate Scenarios

SafetyDGX agent

arXiv:2606.07226v1 Announce Type: cross Abstract: Human creativity has emerged as a critical competency in the era of large language models. Assessing creativity in complex, open-ended environments is

DiBS: Diffusion-Informed Branch Selection

Model ReleasesDGX agent

arXiv:2606.06518v1 Announce Type: new Abstract: Sudoku is a representative constraint satisfaction problem that requires global structural reasoning under strict discrete constraints. The existing wor

Differences in Detection: Explainability Where it Matters

TutorialsDGX agent

arXiv:2606.07503v1 Announce Type: new Abstract: We propose Differences in Detection (DnD), an intuitive method to compare two object detection models. Based on the same matching algorithm, it compleme

DisPOSE: Projected Polystochastic Diffusion for Self-Supervised Multi-View 3D Human Pose Estimation

Model ReleasesDGX agent

arXiv:2606.07419v1 Announce Type: new Abstract: Recovering 3D human poses for multiple individuals from different camera views is a fundamental bottleneck for analyzing interacting behaviors. Existing

EASE-TTT: Evidence-Aligned Selective Test-Time Training for Long-Context Question Answering

Local AiDGX agent

arXiv:2606.06906v1 Announce Type: cross Abstract: Long-context question answering (QA) remains challenging for smaller language models even when answer-bearing evidence is already present in the input

Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

SafetyDGX agent

arXiv:2603.22327v2 Announce Type: replace-cross Abstract: Systematic literature reviews (SLRs) are a demanding and high-stakes form of scientific knowledge synthesis that remains underspecified as an

Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software

Model ReleasesDGX agent

arXiv:2602.04894v4 Announce Type: replace-cross Abstract: LLMs are increasingly used for code generation, but their outputs often follow recurring templates that can induce predictable vulnerabilities

FIGMA: Towards FIne-Grained Music retrievAl

SafetyDGX agent

arXiv:2606.06615v1 Announce Type: cross Abstract: Retrieving music using natural language descriptions has improved with contrastive audio-text models such as CLAP, but current systems remain limited

Forecasting as Rendering: A 2D Gaussian Splatting Framework for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2603.02220v2 Announce Type: replace-cross Abstract: Time series forecasting remains a challenging problem due to the intricate entanglement of intra-period fluctuations and inter-period trends.

From Correctness to Utility: Gain-Based Prefix Evaluation for LLM Reasoning

Local AiDGX agent

arXiv:2606.07190v1 Announce Type: new Abstract: Reasoning prefixes shape the future trajectory of LLM problem solving, yet existing process reward models usually evaluate them through local step corre

HKJudge: A Legal Discourse-Annotated Corpus for Interpreting What Courts Find, How They Reason, and What They Rule

Model ReleasesDGX agent

arXiv:2606.06679v1 Announce Type: cross Abstract: Court judgments are central to legal practice and jurisprudence, yet discourse analysis of Hong Kong judgments has received limited attention, owing l

How reliable are LLMs when it comes to playing dice?

SafetyDGX agent

arXiv:2606.07515v1 Announce Type: cross Abstract: We investigate the probabilistic reasoning capabilities of large language models through a controlled benchmarking study on discrete probability probl

HybridCodec: Fast Dual-Stream, Semantically Enhanced Neural Audio Codec

ResearchDGX agent

arXiv:2606.06743v1 Announce Type: cross Abstract: The popularity of neural audio codecs as speech tokenizers has surged with the advent of Multimodal Large Language Models. New codec architectures wit

MMAE: A Massive Multitask Audio Editing Benchmark

Model ReleasesDGX agent

arXiv:2606.07229v1 Announce Type: cross Abstract: We introduce MMAE, a Massive Multitask Audio Editing benchmark, serving as the first comprehensive evaluation testbed designed for general-purpose ins

Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consistent Video-Sourced Datasets

Model ReleasesDGX agent

arXiv:2606.07032v1 Announce Type: cross Abstract: Zero-Shot Composed Image Retrieval (ZS-CIR) aims to retrieve a target image based on a query composed of a reference image and a relative caption with

Oh my God! @METR_Evals’s coding benchmarks are saturated! 🤯 Mythos broke the METR graph 🤯 4 weeks later, out comes a new coding task, this…

Model ReleasesDGX agent

Oh my God! @METR_Evals’s coding benchmarks are saturated! 🤯 Mythos broke the METR graph 🤯 4 weeks later, out comes a new coding task, this time from @cognition: “FrontierCode Diamond remains unsaturat

On the importance of multiple training seeds for evaluating machine unlearning

TutorialsDGX agent

arXiv:2510.26714v5 Announce Type: replace-cross Abstract: Machine unlearning aims to remove the influence of certain data points from a trained model without costly retraining. Most practical unlearni

Queen-Bee Agents: A BeeSpec-Centered Architecture for Governed Enterprise MCP Orchestration

Local AiDGX agent

arXiv:2606.06545v1 Announce Type: cross Abstract: Enterprise agent systems increasingly need to connect large language models to private tools, internal knowledge, and Model Context Protocol (MCP) int

Style or Content? Evaluating Style Classifiers with Controlled Content Overlap

Model ReleasesDGX agent

arXiv:2606.07103v1 Announce Type: new Abstract: Style classifiers can use content cues that correlate with style labels in naturally collected data, yet we lack a systematic way to measure this relian

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

Model ReleasesDGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

SVHighlights: Towards Extremely Long Sport Video Highlight Detection

Model ReleasesDGX agent

arXiv:2606.06926v1 Announce Type: new Abstract: While highlight detection for long-form videos is of great practical importance, most existing methods remain limited to short-form content, largely due

'Testing LCM on a GTX 750 Ti 4GB: Surprisingly Usable for Low-VRAM AI Image Generation'

Local AiDGX agent

This post documents testing Latent Consistency Models (LCM) on a GTX 750 Ti graphics card with 4GB of VRAM, demonstrating that this older, lower-end GPU can still run AI image generation models with a

The Geography of Algorithmic Judgment: LLM Intermediaries, Place Identity, and Racial Steering in Housing Search

Local AiDGX agent

arXiv:2606.06694v1 Announce Type: cross Abstract: Large language models (LLMs) are rapidly assuming an intermediary role in housing search through the integration of listing platforms within conversat

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook

ResearchDGX agent

arXiv:2604.02029v2 Announce Type: replace Abstract: Latent space is rapidly emerging as a native substrate for language-based models. While modern systems are still commonly understood through explici

Think Like a Pilot: Fine-Grained Long-Horizon UAV Navigation

Model ReleasesDGX agent

arXiv:2606.06836v1 Announce Type: cross Abstract: Language-guided UAV agents must execute long-horizon semantic instructions while producing smooth, physically feasible continuous flight commands, yet

← Previous
1…461462463464465…1060
Next →