AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
25 May 2026

ComfyUI-Angelo now supports Qwen Edit

Model ReleasesDGX agent

ComfyUI-Angelo now supports Qwen-Image-Edit, an advanced image editing model that provides text editing features and the ability to edit both semantics and appearance of images. The model applies Qwen

Complete-muE: Optimal Hyperparameter Transfer and Scaling for MoE Models

Model ReleasesDGX agent

arXiv:2605.23893v1 Announce Type: new Abstract: We propose Complete-muE, a framework which targets hyperparameter transfer across dense FFN and any Mixture-of-Experts (MoE) setups in transformer block

Computable Fairness: Boltzmann-Softmax Control for AI Resource Allocation

Model ReleasesDGX agent

arXiv:2605.22827v1 Announce Type: cross Abstract: In large-scale AI systems, allocating scarce resources such as GPU compute time and bandwidth among multiple agents is a critical challenge. Conventio


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Convex Optimization for Alignment and Preference Learning on a Single GPU

Model ReleasesDGX agent

arXiv:2605.23244v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) to align with human preferences has driven the success of systems such as Gemini and ChatGPT. However, approach

Cost-Effective Model Evaluation with Meta-Learning

Model ReleasesDGX agent

arXiv:2605.23595v1 Announce Type: cross Abstract: The rapid growth of machine learning has produced an ever-expanding ecosystem of models, making it increasingly challenging to verify the reliability

Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization

Model ReleasesDGX agent

arXiv:2605.23391v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) for coupled multiphysics systems suffer systematic accuracy degradation as inter-equation coupling strengthens.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models

Model ReleasesDGX agent

arXiv:2605.23699v1 Announce Type: new Abstract: Video prediction is increasingly viewed as a path toward generalizable world models, yet it remains unclear whether these systems learn underlying causa

Cultural Adaptation in Large Language Models for Political Discourse

Model ReleasesDGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception

Model ReleasesDGX agent

arXiv:2605.23655v1 Announce Type: cross Abstract: High-resolution (HR) image perception presents a key bottleneck for multimodal large language models (MLLMs). While visual search offers a promising s

D2 Actor Critic: Diffusion Actor Meets Distributional Critic

Model ReleasesDGX agent

arXiv:2510.03508v3 Announce Type: replace Abstract: We introduce D2AC, a new model-free reinforcement learning (RL) algorithm designed to train expressive diffusion policies online effectively. At its

DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures

Model ReleasesDGX agent

arXiv:2511.15503v2 Announce Type: replace-cross Abstract: High-performance Host processors can integrate Processing-In-Memory (PIM) devices, which can accelerate memory-intensive kernels of Machine Le

DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs

Model ReleasesDGX agent

arXiv:2605.23629v1 Announce Type: new Abstract: Medical diagnosis is not a single prediction from a fully specified vignette. It is a sequential workup: clinicians decide what evidence to obtain, revi

Decomposing and Measuring Evaluation Awareness

Model ReleasesDGX agent

arXiv:2605.23055v1 Announce Type: cross Abstract: Frontier language models sometimes recognize that they are being evaluated and adjust their behavior, undermining validity of benchmark results. Yet t

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

Model ReleasesDGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

Decomposition-Based Modular Conformal Prediction for Two-Stage Modeling

Model ReleasesDGX agent

arXiv:2510.04406v2 Announce Type: replace-cross Abstract: Conformal prediction offers finite-sample coverage guarantees under minimal assumptions. However, existing methods treat the entire modeling p

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

Model ReleasesDGX agent

arXiv:2605.23355v1 Announce Type: new Abstract: Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professiona

DeepSeek V4 Flash IS BACK on Nous Portal for FREE for use in Hermes Agent! Check it out at https://portal.nousresearch.com/manage-subscripti…

Model ReleasesDGX agent

DeepSeek V4 Flash model has been made available again on the Nous Research portal at no cost for use with Hermes Agent applications. Users can access and utilize this model through the Nous portal's s

DepthAgent: Towards Better Universal Depth Estimation via Sample-wise Expert Selection

Model ReleasesDGX agent

arXiv:2605.23281v1 Announce Type: new Abstract: Monocular metric depth estimation has achieved strong progress with large-scale training and universal-camera modeling, yet robust deployment across div

Design and Report Benchmarks for Knowledge Work

Model ReleasesDGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge

Model ReleasesDGX agent

arXiv:2605.23069v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used across diverse linguistic and cultural contexts, yet their cultural knowledge remains uneven across r

Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation

Model ReleasesDGX agent

arXiv:2605.22950v1 Announce Type: cross Abstract: Score matching is an alternative to maximum likelihood estimation when the normalizing constant is unknown or too costly to evaluate. However, vanilla

Discontinuous Galerkin Neural Operator for Pathology Defocus Deblurring

Model ReleasesDGX agent

arXiv:2605.23282v1 Announce Type: cross Abstract: Defocus deblurring in pathological microscopy remains challenging due to the spatially varying and locally discontinuous nature of optical blur induce

Do Synthetic Brain MRIs Reliably Improve Tumour Classification? A StyleGAN2-ADA Class-Plane Augmentation Study on BRISC 2025

Model ReleasesDGX agent

arXiv:2605.23094v1 Announce Type: cross Abstract: Generative augmentation is often proposed as a remedy for small medical-image datasets, but synthetic images are only useful when they improve downstr

DreamerNLplus: Interpretable Modeling of Mental Health Dynamics from Social Media Timelines using Hybrid Rule-Based and RAG Methods

Model ReleasesDGX agent

arXiv:2605.23052v1 Announce Type: cross Abstract: We present DreamerNLplus, a hybrid framework for modeling mental health dynamics from social media timelines in the CLPsych 2026 shared task. Our syst

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.23176v1 Announce Type: new Abstract: Spatiotemporal intelligence in autonomous driving (AD) requires an agent to integrate multi-view observations into a coherent scene representation, main

Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.26383v5 Announce Type: replace-cross Abstract: Knowledge-graph retrieval-augmented generation (KG-RAG) couples large language models (LLMs) with structured, verifiable knowledge graphs (KGs

Efficient Gradient Estimation for Parameterized Quantum Systems with Lie Algebraic Symmetries

Model ReleasesDGX agent

arXiv:2404.05108v3 Announce Type: replace-cross Abstract: Gradient estimation is a central challenge in training parameterized quantum circuits (PQCs) for hybrid quantum-classical optimization and lea

Efficient One-Step Diffusion Restoration Model with Compact Token Compression and Linear Attention

Model ReleasesDGX agent

arXiv:2605.23451v1 Announce Type: new Abstract: Real-world image super-resolution aims to recover high-quality images from complex and unknown real-world degradations. However, existing generative Rea

Entropy Equivalence Testing

Model ReleasesDGX agent

arXiv:2605.23225v1 Announce Type: cross Abstract: We introduce the problem of entropy equivalence testing for probability distributions, a relaxation of the well-studied closeness testing problem, whe

ETCHR: Editing To Clarify and Harness Reasoning

Model ReleasesDGX agent

arXiv:2605.23897v1 Announce Type: cross Abstract: Multimodal Large Language Models have advanced visual reasoning, yet a purely textual chain of thought remains a bottleneck for questions that require

Evaluating Large Language Models in a Complex Hidden Role Game

Model ReleasesDGX agent

arXiv:2605.22826v1 Announce Type: cross Abstract: Quantifying the deceptive potential of Large Language Models (LLMs) is critical for AI safety, yet difficult to achieve in uncontrolled environments.

Evaluating Memory Structure in LLM Agents

Model ReleasesDGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

Exploitation of KnowledgeDeliver via ViewState Deserialization Vulnerability

Model ReleasesDGX agent

Written by: Takahiro Sugiyama, Peter Revelant, Mathew Potaczek Introduction In late 2025, Mandiant responded to a security incident involving a compromised web server running KnowledgeDeliver. Knowled

Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking

Model ReleasesDGX agent

arXiv:2605.23118v1 Announce Type: cross Abstract: Tracking tumor lesions across serial CT scans is essential for oncological response assessment. Existing automated methods face a fundamental trade-of

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

Model ReleasesDGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

FAST-ME: Foundation-aware Adaptive Stopping for Motion Estimation for Efficient IoT Video Analysis

Model ReleasesDGX agent

arXiv:2605.23428v1 Announce Type: new Abstract: In modern multimedia systems, efficient video processing is critical, especially in resource-constrained environments such as IoT-based camera networks,

FastKernels: Benchmarking GPU Kernel Generation in Production

Model ReleasesDGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2510.08945v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has emerged as a promising paradigm for improving factual accuracy in large language models (LLMs). We introduc

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches

Model ReleasesDGX agent

arXiv:2512.12677v2 Announce Type: replace-cross Abstract: We explore efficient strategies to fine-tune decoder-only Large Language Models (LLMs) for downstream text classification under resource const

FuRA: Full-Rank Parameter-Efficient Fine-Tuning with Spectral Preconditioning

Model ReleasesDGX agent

arXiv:2605.22869v1 Announce Type: new Abstract: Both full fine-tuning (Full FT) and parameter-efficient fine-tuning methods such as LoRA introduce weight updates without accounting for the spectral st

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would…

Model ReleasesDGX agent

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would be very much in @OpenAI's interest to publish the architect

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-reali…

Model ReleasesDGX agent

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-realistic, multi-location, vertical 9:16 on a holiday deadline. A

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

Model ReleasesDGX agent

arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev

HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retrieval

Model ReleasesDGX agent

arXiv:2605.23572v1 Announce Type: cross Abstract: In the competitive landscape of sponsored search, balancing retrieval quality with production latency is a critical challenge. While large retrieval m

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official…

Model ReleasesDGX agent

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official/autonomous-ai-agents/openhands` Reminder: You can already d

Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence

Model ReleasesDGX agent

arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua

How Far Are We from Generating Missing Modalities with Foundation Models?

Model ReleasesDGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness

Model ReleasesDGX agent

arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate

HTMuon: Improving Muon via Heavy-Tailed Spectral Correction

Model ReleasesDGX agent

arXiv:2603.10067v2 Announce Type: replace-cross Abstract: Muon has recently shown promising results in LLM training. In this work, we study how to further improve Muon. We argue that Muon's orthogonal

Human-Centered Learning Mechanics: A Dynamical Framework for Entropy-Regulated Representation Learning

Model ReleasesDGX agent

arXiv:2605.22940v1 Announce Type: cross Abstract: Deep learning is increasingly viewed as a dynamical process in parameter space, yet many existing theories still treat training as a closed optimizati

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more…

Model ReleasesDGX agent

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more deterministic hit rates? Try explicit caching instead. 🙌 🔗B

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

Model ReleasesDGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried 'new forms of slavery' in AI and tech supply chains (Joshua McElwee/Reuters)

Model ReleasesDGX agent

Joshua McElwee / Reuters: In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried “new forms of slavery” in AI and tech supply chains — Pope Leo urged govern

Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems

Model ReleasesDGX agent

arXiv:2605.23109v1 Announce Type: new Abstract: AI agents increasingly excel at generating, testing, and refining code. However, they fall short on tasks requiring formal guarantees of full coverage t

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction

Model ReleasesDGX agent

arXiv:2605.23187v1 Announce Type: new Abstract: Existing object navigation benchmarks usually tell an embodied agent which object category to find, such as microwave or chair. Human-facing embodied AI

Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures

Model ReleasesDGX agent

arXiv:2511.03882v2 Announce Type: replace-cross Abstract: Imitation learning-based robot control policies are enjoying renewed interest in video-based robotics. However, it remains unclear whether thi

Is Capability a Liability? More Capable Language Models Make Worse Forecasts When It Matters Most

Model ReleasesDGX agent

arXiv:2605.22672v2 Announce Type: replace Abstract: We document inverse scaling in LLMs on forecasting problems whose underlying time series exhibit superlinear growth and tail risk of regime change,

It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt

Model ReleasesDGX agent

arXiv:2605.23825v1 Announce Type: cross Abstract: It has generally been assumed that geopolitical bias in language models originates from the training data used during the pre-training phase. We teste

I’ve just released MiMo V2.5-Coder. If you have 128 GB of RAM, this is one of the best models you can run locally. It’s fast, and in all my …

Model ReleasesDGX agent

I’ve just released MiMo V2.5-Coder. If you have 128 GB of RAM, this is one of the best models you can run locally. It’s fast, and in all my experiments it outperformed Qwen 3.6 and DeepSeek 4-Flash. h

← Previous
1…216217218219220…377
Next →