AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,515 results
19 May 2026

Data-Driven Dynamic Modeling of a Tendon-Actuated Continuum Robot

ResearchDGX agent

arXiv:2605.18720v1 Announce Type: new Abstract: Developing dynamic models for tendon-driven continuum robots is challenging due to their nonlinear, high-dimensional, and friction-dominated dynamics. T

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a mode…

TutorialsDGX agent

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a model? tweaking prompts, tool names, and tool descriptions in ac

Diffusion Models, Denoiser Architecture and Creativity

SafetyDGX agent

arXiv:2605.16415v1 Announce Type: new Abstract: The creativity of diffusion models refers to their ability to generate highly realistic images that are different from their training data. Creativity i

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object Dynamics

ApplicationsDGX agent

arXiv:2603.09668v2 Announce Type: replace Abstract: Modeling wind-driven object dynamics from video observations is highly challenging due to the invisibility and spatio-temporal variability of wind,

Dual-Space Knowledge Distillation with Key-Query Matching for Large Language Models with Vocabulary Mismatch

SafetyDGX agent

arXiv:2603.22056v2 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art (SOTA) performance across language tasks, but are costly to deploy due to their size and resou

Dynamic Elliptical Graph Factor Models via Riemannian Optimization with Geodesic Temporal Regularization

Model ReleasesDGX agent

arXiv:2605.18316v1 Announce Type: new Abstract: Inferring time-varying graph structures from high-dimensional nodal observations is a fundamental problem arising in neuroscience, finance, climatology,

ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation

SafetyDGX agent

arXiv:2605.17580v1 Announce Type: new Abstract: Electrocardiogram (ECG)-based models have achieved strong performance in diagnostic tasks, yet they remain limited in modeling how cardiac dynamics evol

Embodied Task Planning via Graph-Informed Action Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam

Ensembling Tabular Foundation Models - A Diversity Ceiling And A Calibration Trap

Model ReleasesDGX agent

arXiv:2605.18696v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) now match or beat tuned gradient-boosted trees on a growing fraction of tabular tasks, but no single TFM wins on ever

Everything Google Cloud customers need to know coming out of Google I/O

Model ReleasesDGX agent

At Google Cloud Next ‘26, we unveiled the blueprint for the Agentic Enterprise, sharing our eighth-generation TPUs, Gemini Enterprise Agent Platform, a fully reimagined Agentic Data Cloud, Workspace I

FLAG: Foundation model representation with Latent diffusion Alignment via Graph for spatial gene expression prediction

SafetyDGX agent

arXiv:2605.18055v1 Announce Type: cross Abstract: Predicting spatial gene expression from routine H&E enables large-scale molecular profiling, yet current models treat this as isolated pointwise tasks

Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models

Model ReleasesDGX agent

arXiv:2508.06038v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) incur substantial computational overhead and inference latency due to the large number of vision tokens introduc

FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion

Local AiDGX agent

arXiv:2605.17759v1 Announce Type: new Abstract: To circumvent the inherent fidelity bottlenecks and optimization misalignment of VAE-based latent diffusion, pixel-space diffusion models have emerged a

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

Model ReleasesDGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

Model ReleasesDGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics

Model ReleasesDGX agent

arXiv:2603.02531v2 Announce Type: replace-cross Abstract: Classifier-Free Guidance (CFG) improves sample quality in diffusion models, but its dual-pass inference and reliance on null-condition trainin

Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app (Sanuj Bhatia/Android Central)

Model ReleasesDGX agent

Sanuj Bhatia / Android Central: Google adds a conversational search feature to YouTube and rolls out the new Gemini Omni model in YouTube Shorts Remix and the Create app — Seriously, who is asking for

How Many Visual Tokens Do Multimodal Language Models Need? Scaling Visual Token Pruning with F^3A

Local AiDGX agent

arXiv:2605.16359v1 Announce Type: cross Abstract: Vision-language models improve perception by feeding increasingly long visual token sequences into language backbones, but the resulting inference cos

Lance: Unified Multimodal Modeling by Multi-Task Synergy

SafetyDGX agent

arXiv:2605.18678v1 Announce Type: cross Abstract: We present Lance, a lightweight native unified model supporting multimodal understanding, generation, and editing for both images and videos. Rather t

Large Language Models and Impossible Language Acquisition: 'False Promise' or an Overturn of our Current Perspective towards AI

ResearchDGX agent

arXiv:2602.08437v5 Announce Type: replace Abstract: In Chomsky's provocative critique 'The False Promise of CHATGPT,' Large Language Models (LLMs) are characterized as mere pattern predictors that do

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate

AgentsDGX agent

arXiv:2601.22297v2 Announce Type: replace Abstract: The reasoning abilities of large language models (LLMs) have been substantially improved by reinforcement learning with verifiable rewards (RLVR). A

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

Model ReleasesDGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

Model ReleasesDGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

Prune, Update and Trim: Robust Structured Pruning for Large Language Models

ResearchDGX agent

arXiv:2605.18331v1 Announce Type: new Abstract: Large Language Models (LLMs) have experienced significant growth and development in recent years. However, performing inference on LLMs remains costly,

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models

Model ReleasesDGX agent

arXiv:2501.17549v2 Announce Type: replace Abstract: Graph-structured data plays a vital role in numerous domains, such as social networks, citation networks, commonsense reasoning graphs and knowledge

Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift

ApplicationsDGX agent

arXiv:2605.16411v1 Announce Type: cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible y

Retrieval and competition: how a protein foundation model starts a protein

SafetyDGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data

Model ReleasesDGX agent

arXiv:2605.18287v1 Announce Type: new Abstract: It is infeasible to encompass all possible disturbances within the training dataset. This raises a critical question regarding the robustness of Vision-

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

Model ReleasesDGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

HardwareDGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.17577v1 Announce Type: new Abstract: Large-scale pre-trained Vision-Language models (VLMs), such as CLIP, exhibit strong zero-shot generalization, yet remain highly vulnerable to impercepti

The last six months in LLMs in five minutes

Model ReleasesDGX agent

I put together these annotated slides from my five minute lightning talk at PyCon US 2026, using the latest iteration of my annotated presentation tool. # I presented this lightning talk at PyCon US 2

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2602.22801v2 Announce Type: replace-cross Abstract: Diffusion models have become a popular choice for decision-making tasks in robotics, and more recently, are also being considered for solving

Unlocking the Potential of Diffusion Language Models through Template Infilling

ResearchDGX agent

arXiv:2510.13870v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) have emerged as a promising alternative to Autoregressive Language Models, yet their inference strategies rem

Use your LM Studio models to code locally in @zeddotdev 🚀

Local AiDGX agent

Use your LM Studio models to code locally in @zeddotdev 🚀 Local model usage grew 3x in Zed's agent in the last 10 weeks. Cameron Mcloughlin on why he prefers local: 'I worry about over-reliance on pro

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories…

Model ReleasesDGX agent

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories, memorable moments, and many, many (occasionally embarrassi

WOW-Seg: A Word-free Open World Segmentation Model

Model ReleasesDGX agent

arXiv:2605.16903v1 Announce Type: new Abstract: Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open

18 May 2026

Beyond the Query: 5 Scenarios Laying the Foundation for the Agentic Era

Model ReleasesDGX agent

Accessing enterprise data is shifting from static reports to dynamic use by autonomous systems. To keep up, organizations must route fragmented data from SaaS, IoT, and legacy sources into secure, sca

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

Model ReleasesDGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

DiscussLLM: Teaching Large Language Models When to Speak

TutorialsDGX agent

arXiv:2508.18167v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in understanding and generating human-like text, yet they largely operate as

Dynamic Chunking for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.15676v1 Announce Type: new Abstract: Block discrete diffusion language models factorize a sequence autoregressively over fixed-size positional blocks, decoupling within-block parallel denoi

Dynamics-Level Watermarking of Flow Matching Models with Random Codes

ResearchDGX agent

arXiv:2605.16239v1 Announce Type: new Abstract: We introduce a dynamics-level approach to watermarking generative models. Rather than embedding signals into model weights or outputs, we embed the wate

Enabling Adversarial Robustness in AI Models through Kubeflow MLOps

ResearchDGX agent

arXiv:2605.15249v1 Announce Type: cross Abstract: AI models are increasingly deployed in cloud-native environments to support scalable and automated services. However, while platforms such as Kubernet

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

Model ReleasesDGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s →…

Model ReleasesDGX agent

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s → 45 tok/s (+78%). Two flags on llama-server: --spec-type draf

Measuring Maximum Activations in Open Large Language Models

Model ReleasesDGX agent

arXiv:2605.15572v1 Announce Type: new Abstract: The dynamic range of activations is a first-order constraint for low-bit quantization, activation scaling, and stable LLM inference. Prior work characte

MIND: Decoupling Model-Induced Label Noise via Latent Manifold Disentanglement

Local AiDGX agent

arXiv:2605.16081v1 Announce Type: cross Abstract: The paradigm of learning from automatic annotations driven by pre-trained experts and Foundation Models dominates data-hungry applications. However, i

Offline Reinforcement Learning with Universal Horizon Models

SafetyDGX agent

arXiv:2605.15603v1 Announce Type: cross Abstract: Model-based reinforcement learning (RL) offers a compelling approach to offline RL by enabling value learning on imagined on-policy trajectories. Howe

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers sa…

Model ReleasesDGX agent

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers saying things like 'Better, more targeted version' if you aske

Overfitting has a limitation: a model-independent generalization gap bound based on Renyi entropy

ResearchDGX agent

arXiv:2506.00182v3 Announce Type: replace-cross Abstract: Will further scaling up of machine learning models continue to bring success? A significant challenge in answering this question lies in under

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

Model ReleasesDGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

ReactiveGWM: Steering NPC in Reactive Game World Models

SafetyDGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

Reasoning Models Don't Just Think Longer, They Move Differently

ResearchDGX agent

arXiv:2605.15454v1 Announce Type: new Abstract: Reasoning-trained language models often spend more tokens on harder problems, but longer chains of thought do not show whether a model is merely computi

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers

Local AiDGX agent

arXiv:2505.12601v2 Announce Type: replace Abstract: As large language models (LLMs) grow in scale and specialization, routing--selecting the best model for a given input--has become essential for effi

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

SafetyDGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language

ResearchDGX agent

arXiv:2605.15607v1 Announce Type: new Abstract: Large language models (LLMs) achieve high pass rates on code generation benchmarks, yet whether they can transfer this ability to languages absent from

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

Model ReleasesDGX agent

arXiv:2602.08324v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning successfully enhances the reasoning capabilities of Large Language Models (LLMs), yet it incurs substantial computa

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣

Model ReleasesDGX agent

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣 @bo_wangbo stealth releasing probably the strongest open multilingual ColBERT (and it'

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, Mi…

Model ReleasesDGX agent

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, MiniMax M2.7 & DeepSeek V4 Flash are now one click away on Dell

← Previous
1…112113114115116…1009
Next →