AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
8 Apr 2026

Control ComfyUI with Claude Code https://x.com/i/broadcasts/1yxBeMkeYOjJN

Model ReleasesDGX agent

ComfyUI hosted a live broadcast on April 8, 2025, titled 'Control ComfyUI with Claude Code,' demonstrating how to queue workflows, tweak parameters, and automate an entire generation pipeline with...

He’s now indistinguishable from Colin Jost’s parody of him.

Model ReleasesDGX agent

Author Kurt Andersen posted on X (formerly Twitter) that Pete Hegseth has become 'indistinguishable' from SNL's Colin Jost parody of him. SNL's Colin Jost has repeatedly parodied Pete Hegseth, Tru...

It seems like the day has come to leave Anthropic. Initially, I loved Claude Code. It was a good harness and a simple TUI... and I had learn…

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

It seems like the day has come to leave Anthropic. Initially, I loved Claude Code. It was a good harness and a simple TUI... and I had learned to eat my tokens with a sauce of subsidy. Before joining

Let's see how MiMo V2 Pro evolves in Hermes Agent!

AgentsDGX agent

Let's see how MiMo V2 Pro evolves in Hermes Agent! We have partnered with @Xiaomi to bring their excellent MiMo V2 Pro model to Hermes Agent via the Nous Portal - completely free to use for the next 2

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agen…

AgentsDGX agent

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agent today → https://portal.nousresearch.com #MiniMax #NousRese

PS: I finally got around to trying out @randal_olson 's Tufte Test tool to prettify the benchmark plot. Great tool 👌! https://www.goodeyela…

Model ReleasesDGX agent

Sebastian Raschka (rasbt) used Randal Olson's Tufte Test tool, developed by Goodeye Labs, to improve the visual quality of a machine learning benchmark plot. The Tufte Test encodes seven of Tufte'...

'The priority for defenders is to start building now: the scaffolds, the pipelines, the maintainer relationships, the integration into devel…

IndustryDGX agent

'The priority for defenders is to start building now: the scaffolds, the pipelines, the maintainer relationships, the integration into development workflows. The models are ready. The question is whet

7 Apr 2026

【速報】中国のAI企業http://Z.ai(旧Zhipu AI)が最新AIモデル「GLM-5.1」をリリースしました🚀 コーディング性能を測る「SWE-Bench Pro」で58.4%を記録し、オープンソース(誰でも無料で使える形)モデルとして世界1位。全体でも3位(GPT-…

Model ReleasesDGX agent

【速報】中国のAI企業http://Z.ai(旧Zhipu AI)が最新AIモデル「GLM-5.1」をリリースしました🚀 コーディング性能を測る「SWE-Bench Pro」で58.4%を記録し、オープンソース(誰でも無料で使える形)モデルとして世界1位。全体でも3位(GPT-5.4の57.7%を上回る)という強力なスコアです📊 最大の特徴は「長時間タスクで力を発揮する」点👇 🧠 8時間にわたって

I was told about the Mythos release, but didn't have access, so have no personal experience to add. Two points from brief: 1) It is not buil…

ApplicationsDGX agent

I was told about the Mythos release, but didn't have access, so have no personal experience to add. Two points from brief: 1) It is not built for IT security, it is just a good enough model that it is

The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is wh…

Model ReleasesDGX agent

The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is what trained it. Zero Nvidia GPUs. 100,000 Huawei Ascend 910B

21 Aug 2026

A knowledge-guided agentic framework for mitigating patient-context ambiguity in health queries

SafetyDGX agent

arXiv:2608.19875v1 Announce Type: cross Abstract: Patients often submit short, underspecified queries to healthcare chatbots that lack the patient-specific information needed to determine an appropria

A Layered Simplex Architecture for Large Alphabets

Model ReleasesDGX agent

arXiv:2608.19908v1 Announce Type: cross Abstract: Probability estimation over large alphabets under log loss is a well-studied problem, with celebrated methods such as the Good-Turing estimator. We in

Active Inference as Context Acquisition for AI Agents

Model ReleasesDGX agent

arXiv:2608.19202v1 Announce Type: new Abstract: Interactive AI agents must acquire the right context as efficiently as possible. When a user omits a constraint, preference, file, or task variable, an

An Inclusive and Lightweight Approach to Federated Continual Learning for Cultural Heritage

Model ReleasesDGX agent

arXiv:2608.20038v1 Announce Type: cross Abstract: Artificial intelligence can support cultural heritage and digital humanities through large-scale retrieval and analysis of digitized collections. Howe

CarBench: A Comprehensive Benchmark for Neural Surrogates on High-Fidelity 3D Car Aerodynamics

Model ReleasesDGX agent

arXiv:2512.07847v2 Announce Type: replace Abstract: Benchmarking has been the cornerstone of progress in computer vision, natural language processing, and the broader deep learning domain, driving alg

Causal Inference under Interference with Learned Exposure Mappings

Local AiDGX agent

arXiv:2608.19224v1 Announce Type: cross Abstract: Exposure mappings are often assumed to be known in causal spillover analyses. In environmental settings, however, they are typically induced by transp

CharTool: Tool-Integrated Visual Reasoning for Chart Understanding

Local AiDGX agent

arXiv:2604.02794v2 Announce Type: replace Abstract: Charts are ubiquitous in scientific and financial literature for presenting structured data. However, chart reasoning remains challenging for multim

CLaST: Context-aware Contrastive VAE for Probabilistic Time Series Forecasting

ApplicationsDGX agent

arXiv:2608.20025v1 Announce Type: new Abstract: Probabilistic forecasting models are widely used for time series forecasting in domains such as energy systems, finance, medicine, and transportation. I

DecoVAE: a Lightweight Interpretable Trend-Seasonal VAE Framework for Efficient Probabilistic Time Series Forecasting

ApplicationsDGX agent

arXiv:2608.20052v1 Announce Type: new Abstract: Probabilistic time series forecasting remains challenging, largely because modeling distinct trend and seasonal dynamics requires specialized approaches

Deep-MKV-TS: Path-Dependent McKean--Vlasov Control for Financial Time Series Generation

ResearchDGX agent

arXiv:2608.19394v1 Announce Type: cross Abstract: We introduce Deep-MKV-TS, a path-dependent McKean-Vlasov framework for financial scenario generation. The stochastic dynamics are chosen by matching s

Discrete Diffusion Inference-Time Control with Nested Sequential Monte Carlo

ResearchDGX agent

arXiv:2608.20123v1 Announce Type: cross Abstract: We study inference-time control for text generation in discrete diffusion language models, where the goal is to steer sampling toward sequence-level r

DraftFM: A FoundationModel for Day-Zero Drafting in Magic: The Gathering

Model ReleasesDGX agent

arXiv:2608.19568v1 Announce Type: cross Abstract: Drafting a new Magic: The Gathering expansion begins before any pick from it has been observed: the complete card list is public, but the draft logs t

Exploiting Completeness Perception with Diffusion Transformer for Unified 3D MRI Synthesis

TutorialsDGX agent

arXiv:2602.18400v3 Announce Type: replace-cross Abstract: Missing data problems, such as missing modalities in multi-modal brain MRI and missing slices in cardiac MRI, pose significant challenges in c

FAR-DPO: Feasibility-Aware and Robust Direct Preference Optimization for Cyclic Peptide Design

Model ReleasesDGX agent

arXiv:2608.19808v1 Announce Type: new Abstract: Cyclic peptides are emerging as promising molecular scaffolds in drug discovery due to their high binding affinity and structural stability. However, ex

Fastest NVFP4 quant of Qwen3.8 27B out there

Model ReleasesDGX agent

Here's a brand new Blackwell-native, prefill-optimized 4-bit quant that runs 50% faster on compatible hardware than a Q4 quant of the same memory footprint. And it runs 4-7% faster than other NVFP4 qu

From Atari to EVE Online: Building on 15 Years of AI Research in Games

Model ReleasesDGX agent

This paper traces the expansion of AI research in games over a 15‑year period, from early Atari benchmarks to contemporary MMORPGs such as EVE Online. It introduces SIMA 2, an agent designed to play,

From Noise to Signal: Improving Security Log Anomaly Detection Using LLMs with Endpoint-Specific Logs

Model ReleasesDGX agent

arXiv:2608.19938v1 Announce Type: cross Abstract: Existing approaches to anomalous behaviour log detection, such as Wazuh rely primarily on predefined detection rules, while statistical anomaly detect

Gallileo-4D: Frozen Backbone Ensemble for Dynamic 4D Reconstruction

Model ReleasesDGX agent

arXiv:2608.19743v1 Announce Type: new Abstract: We describe our entry to the PhysAI Dynamic 4D Reconstruction Challenge, which placed third of 27 teams at 0.58356 APD on the final leaderboard, without

Inter-X++: A Comprehensive Benchmark for Multimodal Human-Human Interaction Analysis

Model ReleasesDGX agent

arXiv:2608.20312v1 Announce Type: new Abstract: The capability to perceive and synthesize human-human interactions is fundamental to developing intelligent digital human systems. However, existing dat

Kahler landscapes for complex neural network descents and guarantees including a search and destroy of the Calabi-Yau manifold

Model ReleasesDGX agent

arXiv:2608.19584v1 Announce Type: new Abstract: We study landscapes for complex-parameterized networks. Our approach is motivated with an information-theoretic manifold perspective of the parameter an

Learning to Beat: Phenotype-Guided Latent Flow with Regional Motion Priors for Biventricular Motion Synthesis

Model ReleasesDGX agent

arXiv:2608.19738v1 Announce Type: cross Abstract: Full-cycle biventricular geometry is essential for characterizing cardiac function. However, dense and temporally consistent 3D+t biventricular meshes

Llama.cpp DSpark PC Tree Fork (up to 3%-29.5% faster!)

Model ReleasesDGX agent

Hello gang, I made an implementation of DSpark PC Tree (Parent conditioned drafting tree). This is an implementation of this research paper: https://arxiv.org/abs/2608.02123 Unaffiliated, just found i

MaliciousSkillBench: A Comprehensive Benchmark for Malicious Agent Skill Detection

Model ReleasesDGX agent

arXiv:2608.19901v1 Announce Type: cross Abstract: Agent Skills extend LLM agents with reusable instruction packages that may also include scripts, resources, and service configuration. This creates a

Maximum Likelihood Reinforcement Learning

ResearchDGX agent

arXiv:2602.02710v2 Announce Type: replace Abstract: Reinforcement learning (RL) is the method of choice for training models in setups where the objective function can only be evaluated by sampling fro

PEA-DPO: Perception-Enhanced Alignment Direct Preference Optimization for MLLMs Alignment

SafetyDGX agent

arXiv:2608.19598v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has emerged as an effective approach for aligning large language models (LLMs) with human preferences. However, i

Planning-Oriented End-to-End Autonomous Driving: Architectures, Evaluation, and Emerging Paradigms

Model ReleasesDGX agent

arXiv:2608.20111v1 Announce Type: new Abstract: End-to-end autonomous driving has evolved from camera-to-control regression toward planning-oriented systems that use structured representations, trajec

Quantum Gaussian processes for prediction of channel observations

Model ReleasesDGX agent

arXiv:2608.19306v1 Announce Type: cross Abstract: Given a set of input states, we consider the task of predicting the expectation value of a Pauli observable at the output of an unknown quantum evolut

Qworld: Question-Specific Evaluation Criteria for LLMs

ResearchDGX agent

arXiv:2603.23522v2 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on open-ended questions is difficult because response quality depends on the question's context. Binar

Remember, Verify, or Ask? Cross-Family Evaluation of Memory Commitment in LLM Agents

Model ReleasesDGX agent

arXiv:2608.19564v1 Announce Type: new Abstract: Persistent memory can personalize an LLM agent, but an incorrect durable update can silently distort future behavior. We study the memory-clarification

SafeBranch: Branch-Pair Safety Alignment for Embodied Agents

SafetyDGX agent

arXiv:2608.19729v1 Announce Type: new Abstract: Vision-language-model-based embodied agents can complete instructed tasks but often violate safety constraints in the process, a problem recently framed

SAPO: Single-Rollout Autoregressive Policy Optimization for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.19842v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become a critical stage in the post-training of large language models. Existing critic-free, group-relative meth

Spike-based Belief Propagation in Nonlinear Dynamical Systems

Model ReleasesDGX agent

arXiv:2608.19907v1 Announce Type: new Abstract: This paper presents a Bayesian control framework that integrates spike-based dynamics with probabilistic inference for adaptive control. Bayesian infere

The Asymmetric Harms of LLM Compression

SafetyDGX agent

arXiv:2608.19670v1 Announce Type: new Abstract: Large language models (LLMs) compression reduces deployment costs, but standard aggregate metrics like perplexity and accuracy often mask underlying beh

The Thousand Brains Theory 2.0: An Extension for the Long-Range Connections of the Neocortical Heterarchy

TutorialsDGX agent

arXiv:2507.05888v2 Announce Type: replace-cross Abstract: Vernon Mountcastle hypothesized that the basis for intelligence in mammals is the replication of a general computational unit, the cortical co

Video Evidence to Reasoning Efficient Video Understanding via Explicit Evidence Grounding

SafetyDGX agent

arXiv:2601.07761v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) face a fundamental dilemma in video reasoning: they are caught between the prohibitive computational costs of v

XDen-1K: A Density Field Dataset of Real-World Objects

Model ReleasesDGX agent

arXiv:2512.10668v2 Announce Type: replace Abstract: A deep understanding of the physical world is essential for robotic manipulation and physically realistic simulation. While current methods, includi

20 Aug 2026

Automated Computational Energy Minimization of ML Algorithms using Constrained Bayesian Optimization

ResearchDGX agent

arXiv:2407.05788v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is an efficient framework for optimization of black-box objectives when function evaluations are costly and gradien

Autonomous Agricultural Tractor: Integrated Weed Detection and LiDAR Navigation for Precision Paddy Farming

Model ReleasesDGX agent

arXiv:2608.19004v1 Announce Type: cross Abstract: Site-specific weed management in paddy farming offers substantial reductions in herbicide use over conventional broadcast spraying, but field deployme

Beyond Predictive Fairness: Quantifying Attribution Consistency Across Demographic Groups in Diabetic Retinopathy Screening

SafetyDGX agent

arXiv:2608.18759v1 Announce Type: cross Abstract: Fairness in medical imaging is commonly evaluated through subgroup performance metrics, yet it remains unclear whether models rely on consistent visua

Beyond receptive fields: sequence-pooled normalization can supply most of a sequence labeler's context

Local AiDGX agent

arXiv:2608.18576v1 Announce Type: new Abstract: A convolutional sequence labeler's receptive field is routinely treated as the extent of the model's usable context: it sets dilation schedules, bounds

Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning

Model ReleasesDGX agent

arXiv:2608.19181v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own responses using dense token-level guidance from a stronger teacher. In long-context tasks, ho

Continual Reasoning Gym: Diagnosing and Harnessing Shared Reasoning in Continual RLVR

SafetyDGX agent

arXiv:2608.18574v2 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) commonly post-trains reasoning models on multiple tasks, while rerunning multitask RLVR (MTRL) as

Coupled-cluster molecular properties across the main group that extrapolate beyond training size

Local AiDGX agent

arXiv:2608.18346v1 Announce Type: cross Abstract: Coupled-cluster theory defines the accuracy standard for molecular electronic-structure properties but scales too steeply for routine application, whe

Different Facets of Verbalised Overconfidence: an Interpretability Study

ResearchDGX agent

arXiv:2608.18106v1 Announce Type: cross Abstract: Large language models tend to overconfidence, giving assertive answers when the evidence suggests hedging or abstention. Using controlled reasoning sc

Does Mapping Non-Maximal Probabilities to GMM Components Matter for S-JEPA Encoder Representations?

ResearchDGX agent

arXiv:2608.19084v1 Announce Type: new Abstract: S-JEPA uses soft Gaussian mixture model (GMM) posteriors instead of hard cluster labels to preserve uncertainty. It remains unclear whether the probabil

Expanding Google Antigravity for enterprise customers

Model ReleasesDGX agent

Since announcing Google Antigravity in Gemini Enterprise Agent Platform at I/O in May, we’ve heard helpful feedback from our customers. Your developers want easy access to coding agents across surface

get the most out of your agent traces w @agnostai

AgentsDGX agent

get the most out of your agent traces w @agnostai Stop using your agent logs just for debugging. Use them to train your own model. Today we're launching Agnost AI (YC S26)'s first model: agnost-******

H^2EDL: Hyper Evidential Deep Learning for Hierarchical Classification

Local AiDGX agent

arXiv:2608.18185v1 Announce Type: cross Abstract: Fine-grained recognition often involves hierarchical label spaces, where a model may be confident about a coarse semantic concept while remaining unce

How Generative Recommenders Are Redefining RecSys at Scale

HardwareDGX agent

Generative recommenders (GRs) replace traditional embedding‑based methods with transformer‑style sequence models like HSTU and Semantic IDs, addressing scalability, cold‑start, and long‑tail issues in

I pushed Qwen3.8-27B to 381 tps for a single request on a RTX 3090

Model ReleasesDGX agent

Four days ago I released a hyper-optimized Qwen3.8-27B inference engine for an RTX 3090 (82 tps single request, 672 peak). Since then it went to ~114, then ~138 tps single-user with DFlash2 drafting a

← Previous
1…435436437438439…1060
Next →