AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
13 May 2026

FERMI: Exploiting Relations for Membership Inference Against Tabular Diffusion Models

TutorialsDGX agent

arXiv:2605.11527v1 Announce Type: new Abstract: Diffusion models are the leading approach for tabular data synthesis and are increasingly used to share sensitive records. Whether they actually protect

G^2TR: Generation-Guided Visual Token Reduction for Separate-Encoder Unified Multimodal Models

ResearchDGX agent

arXiv:2605.12309v1 Announce Type: new Abstract: The development of separate-encoder Unified multimodal models (UMMs) comes with a rapidly growing inference cost due to dense visual token processing. I

GridSFM: A new, small foundation model for the electric grid

TutorialsDGX agent

Introducing GridSFM, a small foundation model that can predict AC optimal power flow in milliseconds, boosting efficiency and unlocking cost savings. Learn how GridSFM gives grid operators direct visi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation

ResearchDGX agent

arXiv:2605.11596v1 Announce Type: new Abstract: Closed-loop driving simulation requires real-time interaction beyond short offline clips, pushing current driving world models toward autoregressive (AR

Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.12258v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress, yet the object hallucination remains a critical challenge for reliable deplo

Interpretable rainfall modelling reveals rapid reorganisation of Amazonian rainfall under vegetation loss

ResearchDGX agent

arXiv:2605.10948v1 Announce Type: cross Abstract: Understanding how vegetation loss alters rainfall remains a major challenge in climate and hydrological science, as deforestation modifies precipitati

Local and Mixing-Based Algorithms for Gaussian Graphical Model Selection from Glauber Dynamics

Local AiDGX agent

arXiv:2412.18594v3 Announce Type: replace Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications observations arise from dependent dynamic

Make It Long, Keep It Fast: End-to-End 10k-Sequence Modeling at Billion Scale on Douyin Recommendation

ApplicationsDGX agent

arXiv:2511.06077v2 Announce Type: replace Abstract: Short-video recommenders such as Douyin must exploit extremely long user histories without breaking latency or cost budgets. We present an end-to-en

Measuring Accuracy and Energy-to-Solution of Quantum Fine-Tuning of Foundational AI Models

ApplicationsDGX agent

arXiv:2605.02798v1 Announce Type: cross Abstract: We present an experimental study of energy-to-solution (ETS) of hybrid quantum-classical applications, enabled by direct instrumentation of power cons

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has…

SafetyDGX agent

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has been doubling every few months, and this rate has become fa

Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference

Model ReleasesDGX agent

arXiv:2510.05497v5 Announce Type: replace-cross Abstract: Large-scale Mixture of Experts (MoE) Large Language Models (LLMs) have recently become the frontier open-weight models, achieving remarkable m

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

Model ReleasesDGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

SafetyDGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

Reconstruction of Personally Identifiable Information from Supervised Finetuned Models

ApplicationsDGX agent

arXiv:2605.12264v1 Announce Type: cross Abstract: Supervised Finetuning (SFT) has become one of the primary methods for adapting a large language model (LLM) with extensive pre-trained knowledge to do

Recursive Superintelligence raises $650M to build self-improving AI models

HardwareDGX agent

Recursive Superintelligence Inc., a startup that hopes to develop self-improving artificial intelligence models, launched today with 650 million in funding. Alphabet Inc.’s GV fund and Greycroft led t

The DAWN of World-Action Interactive Models

SafetyDGX agent

arXiv:2605.11550v1 Announce Type: new Abstract: A plausible scene evolution depends on the maneuver being considered, while a good maneuver depends on how the scene may evolve. Existing World Action M

Your Model Doesn't Matter. Your Infrastructure Does.

IndustryDGX agent

This article argues that infrastructure quality and architecture are more critical to AI/ML project success than the choice of underlying model. It likely explores how proper deployment, scaling, moni

12 May 2026

A Breast Vision Pathology Foundation Model for Real-world Clinical Utility

ApplicationsDGX agent

arXiv:2605.08207v1 Announce Type: new Abstract: Pathology foundation models have shown strong retrospective performance, but whether such systems can support clinically relevant use remains unclear. T

A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models

ResearchDGX agent

arXiv:2605.08504v1 Announce Type: new Abstract: We investigate the origins of massive activations in large language models (LLMs) and identify a specific layer named the extbf{Massive Emergence Layer

Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization

ResearchDGX agent

arXiv:2510.04988v3 Announce Type: replace Abstract: The vast majority of modern deep learning models are trained with momentum-based first-order optimizers. The momentum term governs the optimizer's m

Advancing AI for materials with MatterSim: experimental synthesis, faster simulation, and multi-task models

ResearchDGX agent

MatterSim is expanding what AI can do for materials science—from faster large-scale simulations to MatterSim-MT, a new multi-task model for simulating properties beyond potential energy surfaces alone

Agentic Performance at the Edge: Insights from Benchmarking

Model ReleasesDGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

[AINews] Thinking Machines' Native Interaction Models - TML-Interaction-Small 276B-A12B - advances SOTA Realtime Voice and kills standard VAD

ToolsDGX agent

Thinking Machines released TML-Interaction-Small, a 276B parameter model with a 12B active subset designed for real-time voice interactions that reportedly surpasses standard voice activity detection

Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis

ResearchDGX agent

arXiv:2407.12173v2 Announce Type: replace-cross Abstract: Generative diffusion models have emerged as a powerful tool for high-quality image synthesis, yet their iterative nature demands significant c

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models

ResearchDGX agent

arXiv:2605.09134v1 Announce Type: new Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually f

C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries During Discharge

Local AiDGX agent

arXiv:2605.08653v1 Announce Type: new Abstract: Accurate state-of-charge (SOC) estimation is critical for the safe and efficient operation of lithium-ion batteries in battery management systems (BMS).

CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models

ResearchDGX agent

arXiv:2605.08735v1 Announce Type: new Abstract: Recent 'Thinking with Video' approaches use Video Generation Models (VGMs) for visual reasoning by producing temporally coherent Chain-of-Frames as reas

Contextual Plackett-Luce: An Efficient Neural Model for Probabilistic Sequence Selection under Ambiguity

ResearchDGX agent

arXiv:2605.09112v1 Announce Type: cross Abstract: Selecting a coherent sequence or subset of elements is a fundamental problem in structured prediction, arising in tasks such as detection, trajectory

Cross-Family Universality of Behavioral Axes via Anchor-Projected Representations

Model ReleasesDGX agent

arXiv:2605.09875v1 Announce Type: new Abstract: Large language models from different families use different hidden dimensions, tokenizers, and training procedures, making behavioral directions difficu

DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models

SafetyDGX agent

arXiv:2605.08902v1 Announce Type: cross Abstract: In recent years, pre-trained visual-linguistic models have demonstrated tremendous potential, becoming a crucial foundational framework for numerous d

Elucidating Representation Degradation Problem in Diffusion Model Training

ResearchDGX agent

arXiv:2605.10790v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success, yet their training remains inefficient due to a severe optimization bottleneck, which we term Represe

Energy-based models for diagnostic reconstruction and analysis in a laboratory plasma device

ResearchDGX agent

arXiv:2605.08645v1 Announce Type: cross Abstract: Energy-based models (EBMs) provide a powerful and flexible way of learning a joint probability distribution over data by constructing an energy surfac

Enhancing Consistency Models for Multi-Agent Trajectory Prediction

AgentsDGX agent

arXiv:2605.08572v1 Announce Type: new Abstract: Diffusion models for multi-agent trajectory prediction are limited by iterative denoising, which causes inference latency that hinders their use in time

Exploration-Driven Optimization for Test-Time Large Language Model Reasoning

SafetyDGX agent

arXiv:2605.09853v1 Announce Type: new Abstract: Post-training techniques combined with inference-time scaling significantly enhance the reasoning and alignment capabilities of large language models (L

ExtraVAR: Stage-Aware RoPE Remapping for Resolution Extrapolation in Visual Autoregressive Models

ResearchDGX agent

arXiv:2605.10045v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models have emerged as a strong alternative to diffusion for image synthesis, yet their fixed training resolution prevents d

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

Model ReleasesDGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

From Spark to Fire: Modeling and Mitigating Error Cascades in LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2603.04474v2 Announce Type: replace-cross Abstract: Large Language Model-based Multi-Agent Systems (LLM-MAS) are increasingly applied to complex collaborative scenarios. However, their collabora

Gate-and-Merge: Zero-shot Compositional Personalization of Vision Language Models

ResearchDGX agent

arXiv:2605.08702v1 Announce Type: cross Abstract: This paper tackles compositional personalization of vision-language models (VLMs). In this problem, multiple user-defined concepts must be recognized

Geometry-Aware Discretization Error of Diffusion Models

Model ReleasesDGX agent

arXiv:2605.08392v1 Announce Type: new Abstract: Practical diffusion sampling is a numerical approximation problem: under a fixed inference budget, one must simulate a reverse-time ODE or SDE using onl

Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model

Model ReleasesDGX agent

arXiv:2605.10739v1 Announce Type: cross Abstract: We introduce SMART-HC-VQA, a Sentinel-2-based visual question answering dataset derived from the IARPA SMART Heavy Construction dataset, designed for

i like that there are models called bert and ernie, but in all seriousness, this update looks impressive

AgentsDGX agent

i like that there are models called bert and ernie, but in all seriousness, this update looks impressive ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total pa

In KAME, a fast speech model starts replying instantly, while a backend LLM runs in parallel to inject deep knowledge on the fly. It’s a com…

ResearchDGX agent

In KAME, a fast speech model starts replying instantly, while a backend LLM runs in parallel to inject deep knowledge on the fly. It’s a completely different way to approach conversational AI, making

Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning

SafetyDGX agent

arXiv:2605.09009v1 Announce Type: cross Abstract: Large language models (LLMs) have shown remarkable in-context learning (ICL) capabilities, yet their potential for sequential decision-making remains

Learning Less Is More: Premature Upper-Layer Attention Specialization Hurts Language Model Pretraining

Model ReleasesDGX agent

arXiv:2605.10504v1 Announce Type: new Abstract: A causal-decoder block is hierarchical: lower layers build the residual basis that upper layers attend over. We identify a failure mode in GPT pretraini

LLM Agents Already Know When to Call Tools -- Even Without Reasoning

Model ReleasesDGX agent

arXiv:2605.09252v1 Announce Type: new Abstract: Tool-augmented LLM agents tend to call tools indiscriminately, even when the model can answer directly. Each unnecessary call wastes API fees and latenc

Metropolis-Adjusted Diffusion Models

SafetyDGX agent

arXiv:2605.09654v1 Announce Type: cross Abstract: Sampling from score-based diffusion models incurs bias due to both time discretisation and the approximation of the score function. A common strategy

Mismatch-Aware Adaptive Constraint Tightening for Bicycle-Model Trajectory Optimization

SafetyDGX agent

arXiv:2605.09376v1 Announce Type: new Abstract: Trajectory optimization for autonomous vehicles usually relies on the kinematic bicycle model because of its computational simplicity. However, when the

Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models

ResearchDGX agent

arXiv:2602.01698v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have recently achieved strong mathematical and code reasoning performance through Reinforcement Learning (RL) post-tra

Rethinking Event-Based Object Dtection through Representation-Level Temporal Aggregation and Model-Level Hypergraph Reasoning

ResearchDGX agent

arXiv:2605.08825v1 Announce Type: new Abstract: Event cameras provide microsecond-level temporal resolution, low latency, and high dynamic range, offering potential for perception under fast motion an

Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models

ResearchDGX agent

arXiv:2605.08145v1 Announce Type: cross Abstract: Current vision language models face hallucination and robustness issues against ambiguous or corrupted modalities. We hypothesize that these issues ca

Talked to a friend at a top AI lab. Their whole team is former journalists, training models on poems, summaries, and creative writing. I use…

IndustryDGX agent

Talked to a friend at a top AI lab. Their whole team is former journalists, training models on poems, summaries, and creative writing. I use AI every day and can see that it tends to flatten my writin

TARO: Temporal Adversarial Rectification Optimization Using Diffusion Models as Purifiers

ResearchDGX agent

arXiv:2605.08440v1 Announce Type: cross Abstract: Adversarial purification with diffusion models seeks to project adversarial examples back toward the data manifold, but balancing semantic preservatio

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

Model ReleasesDGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

Though the smartness comes with a cost: all of the prompts that were written for the old realtime voice model now need to be revised for a m…

ApplicationsDGX agent

Ethan Mollick discusses a tradeoff in OpenAI's newer realtime voice model, where improved capabilities require developers to revise prompts that were written for the previous version. The post highlig

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

AgentsDGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

SafetyDGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing

ResearchDGX agent

arXiv:2601.08321v3 Announce Type: replace Abstract: With the rapid advancement of image generation, visual text editing using natural language instructions has received increasing attention. The main

Unlocking air traffic flow prediction through microscopic aircraft-state modeling

ApplicationsDGX agent

arXiv:2605.10083v1 Announce Type: new Abstract: Short-term air traffic flow prediction in terminal airspace is essential for proactive air traffic management. Existing approaches predominantly model t

UxSID: Semantic-Aware User Interests Modeling for Ultra-Long Sequence

ResearchDGX agent

arXiv:2605.09040v1 Announce Type: new Abstract: Modeling ultra-long user sequences involves a difficult trade-off between efficiency and effectiveness. While current paradigms rely on either item-spec

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models

AgentsDGX agent

arXiv:2605.10106v1 Announce Type: cross Abstract: Recent advances in Multi-modal Large Language Models (MLLMs) target 3D spatial intelligence, yet the progress has been largely driven by post-training

← Previous
1…188189190191192…1018
Next →