AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
Tutorials

Exploring Time Conditioning in Diffusion Generative Models from Disjoint Noisy Data Manifolds

DGX agent

arXiv:2604.25289v1 Announce Type: cross Abstract: Practically, training diffusion models typically requires explicit time conditioning to guide the network through the denoising sampling process. Espe

tutorialsarxiv-cs-cv
29 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Making AI-Assisted Grant Evaluation Auditable without Exposing the Model

DGX agent

arXiv:2604.25200v1 Announce Type: cross Abstract: Public agencies are beginning to consider large language models (LLMs) as decision-support tools for grant evaluation. This creates a practical govern

researcharxiv-cs-lg
29 Apr 2026
Research

Principled Detection of Hallucinations in Large Language Models via Multiple Testing

DGX agent

arXiv:2508.18473v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have emerged as powerful foundational models to solve a variety of tasks, they have also been shown to be prone t

researcharxiv-cs-cl
29 Apr 2026
Applications

Robustness Evaluation of a Foundation Segmentation Model Under Simulated Domain Shifts in Abdominal CT: Implications for Health Digital Twin Deployment

DGX agent

arXiv:2604.25685v1 Announce Type: cross Abstract: Foundation segmentation models such as the Segment Anything Model (SAM) have demonstrated strong generalization across natural images; however, their

applicationsarxiv-cs-cv
29 Apr 2026
Research

Sketch2Arti: Sketch-based Articulation Modeling of CAD Objects

DGX agent

arXiv:2604.25781v1 Announce Type: new Abstract: Articulation modeling aims to infer movable parts and their motion parameters for a 3D object, enabling interactive animation, simulation, and shape edi

researcharxiv-cs-cv
29 Apr 2026
Hardware

Tendon-Actuated Robots with a Tapered, Flexible Polymer Backbone: Design, Fabrication, and Modeling

DGX agent

arXiv:2603.19124v2 Announce Type: replace Abstract: This paper presents the design, modeling, and fabrication of 3D-printed, tendon-actuated continuum robots featuring a flexible, tapered backbone con

hardwarearxiv-cs-ro
29 Apr 2026
Model Releases

Domain-Adapted Fine-Tuning of ECG Foundation Models for Multi-Label Structural Heart Disease Screening

DGX agent

arXiv:2604.23385v1 Announce Type: new Abstract: Transthoracic echocardiography is the reference standard for confirming structural heart disease (SHD), but first-line screening is limited by cost, wor

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models

DGX agent

arXiv:2502.04424v4 Announce Type: replace-cross Abstract: With the integration of multimodal large language models (MLLMs) into robotic systems and AI applications, embedding emotional intelligence (E

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

feels like forever ago, but had to include openclaw in q1 trends coding models improved greatly in Q4 of 2025, early jan was ppl running cla…

DGX agent

feels like forever ago, but had to include openclaw in q1 trends coding models improved greatly in Q4 of 2025, early jan was ppl running claude codes in parallel, and clawdbot blew up late jan models

model-releasesyohei-nakajima--x
28 Apr 2026
Model Releases

'If they ever tell my story let them say that I walked with giants'--Troy I am humbled&excited the model we released last week is trending #…

DGX agent

'If they ever tell my story let them say that I walked with giants'--Troy I am humbled&excited the model we released last week is trending #2 on @huggingface, between giant models such as DeepSeek,Qwe

model-releasesclem-delangue--x
28 Apr 2026
Research

Impact of Age Specialized Models for Hypoglycemia Classification

DGX agent

arXiv:2604.23732v1 Announce Type: cross Abstract: Disease progression varies with age and is influenced by underlying genetic, biochemical, and hormonal etiologies, suggesting the need for tailored mo

researcharxiv-cs-ai
28 Apr 2026
Research

KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models

DGX agent

arXiv:2603.01581v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models build a token-domain robot control paradigm, yet suffer from low speed. Speculative Decoding (SD) is an op

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models

DGX agent

arXiv:2604.24542v1 Announce Type: cross Abstract: Large language models deployed at runtime can misbehave in ways that clean-data validation cannot anticipate: training-time backdoors lie dormant unti

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

M^2-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills

DGX agent

arXiv:2604.24182v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models predominantly rely on end-to-end fine-tuning. While effective, this paradigm compromises the inherent genera

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big…

DGX agent

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big release. Local models eat well. https://github.com/openclaw/open

model-releasesollama--x
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is co…

DGX agent

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is competitive with many closed-weight models, at a fraction of t

model-releaseskimi-moonshot--x
28 Apr 2026
Research

Using Language Models as Closed-Loop High-Level Planners for Robotics Applications: A Brief Overview and Benchmarks

DGX agent

arXiv:2511.07410v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) and Vision Language Models (VLMs) have become popular tools for embodied high-level planning. However, their depl

researcharxiv-cs-ai
28 Apr 2026
Model Releases

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

DGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models

DGX agent

arXiv:2510.08618v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) offer a promising end-to-end solution for slide-enhanced speech recognition due to their inherent mul

model-releasesarxiv-cs-cv
28 Apr 2026
Local Ai

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond

DGX agent

arXiv:2604.22748v1 Announce Type: new Abstract: As AI systems move from generating text to accomplishing goals through sustained interaction, the ability to model environment dynamics becomes a centra

local-aiarxiv-cs-ai
27 Apr 2026
Model Releases

Calibrating Behavioral Parameters with Large Language Models

DGX agent

arXiv:2602.01022v2 Announce Type: replace-cross Abstract: Behavioral parameters such as loss aversion, herding, and extrapolation are central to asset pricing models but remain difficult to measure re

model-releasesarxiv-cs-ai
27 Apr 2026
Local Ai

Here's a uv one-liner that downloads and runs the MLX model against a local mp3 file uv run --with mlx-audio python -m mlx_audio.stt.generat…

DGX agent

Here's a uv one-liner that downloads and runs the MLX model against a local mp3 file uv run --with mlx-audio python -m mlx_audio.stt.generate --model mlx-community/VibeVoice-ASR-4bit --audio lenny.mp3

local-aisimon-willison--x
27 Apr 2026
Research

Large Language Models Decide Early and Explain Later

DGX agent

arXiv:2604.22266v1 Announce Type: new Abstract: Large Language Models often achieve strong performance by generating long intermediate chain-of-thought reasoning. However, it remains unclear when a mo

researcharxiv-cs-cl
27 Apr 2026
Research

Mixed Membership sub-Gaussian Models

DGX agent

arXiv:2604.22633v1 Announce Type: cross Abstract: The Gaussian mixture model is widely used in unsupervised learning, owing to its simplicity and interpretability. However, a fundamental limitation of

researcharxiv-cs-lg
27 Apr 2026
Model Releases

On Benchmark Hacking in ML Contests: Modeling, Insights and Design

DGX agent

arXiv:2604.22230v1 Announce Type: cross Abstract: Benchmark hacking refers to tuning a machine learning model to score highly on certain evaluation criteria without improving true generalization or fa

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution

DGX agent

arXiv:2604.22464v1 Announce Type: new Abstract: Continual Model Merging (CMM) sequentially integrates task-specific models into a unified architecture without intensive retraining. However, existing C

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

No more updating needed for new model releases via Nous Portal and OpenRouter!

DGX agent

No more updating needed for new model releases via Nous Portal and OpenRouter! Hermes will no longer have to be updated to receive model list curation updates for several providers, including Nous Por

model-releasesnous-research--x
26 Apr 2026
Research

Algebraic Language Models for Inverse Design of Metamaterials via Diffusion Transformers

DGX agent

arXiv:2507.15753v2 Announce Type: replace-cross Abstract: Generative machine learning models have revolutionized material discovery by capturing complex structure-property relationships, yet extending

researcharxiv-cs-ai
24 Apr 2026
Model Releases

API is Available Today! 🔹 Keep base_url, just update model to deepseek-v4-pro or deepseek-v4-flash. 🔹 Supports OpenAI ChatCompletions & An…

DGX agent

API is Available Today! 🔹 Keep base_url, just update model to deepseek-v4-pro or deepseek-v4-flash. 🔹 Supports OpenAI ChatCompletions & Anthropic APIs. 🔹 Both models support 1M context & dual modes (T

model-releasesdeepseek--x
24 Apr 2026
Research

Context Unrolling in Omni Models

DGX agent

arXiv:2604.21921v1 Announce Type: new Abstract: We present Omni, a unified multimodal model natively trained on diverse modalities, including text, images, videos, 3D geometry, and hidden representati

researcharxiv-cs-cv
24 Apr 2026
Model Releases

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-fl…

DGX agent

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:cloud Try it with OpenClaw: ollama launch openclaw --mod

model-releasesollama--x
24 Apr 2026
Model Releases

DenoiseRank: Learning to Rank by Diffusion Models

DGX agent

arXiv:2604.20852v1 Announce Type: cross Abstract: Learning to rank (LTR) is one of the core tasks in Machine Learning. Traditional LTR models have made great progress, but nearly all of them are imple

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Feature request for @huggingface - add a repository size option to the sort menu, I want to see the DeepSeek quantized models that take up t…

DGX agent

Simon Willison requested that Hugging Face add a repository size sorting option to help users find and filter models by storage requirements, specifically mentioning interest in locating DeepSeek quan

model-releasessimon-willison--x
24 Apr 2026
Model Releases

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, thi…

DGX agent

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, this model is key for long-horizon tasks — it excels at underst

model-releaseswindsurf--x
24 Apr 2026
Model Releases

HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping

DGX agent

arXiv:2604.21127v1 Announce Type: new Abstract: The NASA PACE mission provides unprecedented hyperspectral observations of ocean color, aerosols, and clouds, offering new insights into how these compo

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Latent Denoising Improves Visual Alignment in Large Multimodal Models

DGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

DGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

DGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control

DGX agent

arXiv:2604.20867v1 Announce Type: cross Abstract: Recent events surrounding the relationship between frontier AI suppliers and national-security customers have made a structural problem newly visible:

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

DGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Projected Gradient Unlearning for Text-to-Image Diffusion Models: Defending Against Concept Revival Attacks

DGX agent

arXiv:2604.21041v1 Announce Type: new Abstract: Machine unlearning for text-to-image diffusion models aims to selectively remove undesirable concepts from pre-trained models without costly retraining.

researcharxiv-cs-cv
24 Apr 2026
Research

Schoenfeld's Anatomy of Mathematical Reasoning by Language Models

DGX agent

arXiv:2512.19995v2 Announce Type: replace-cross Abstract: Large language models increasingly expose reasoning traces, yet their underlying cognitive structure and steps remain difficult to identify an

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

DGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

StyleVAR: Controllable Image Style Transfer via Visual Autoregressive Modeling

DGX agent

arXiv:2604.21052v1 Announce Type: cross Abstract: We build on the Visual Autoregressive Modeling (VAR) framework and formulate style transfer as conditional discrete sequence modeling in a learned lat

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Thinking Like a Botanist: Challenging Multimodal Language Models with Intent-Driven Chain-of-Inquiry

DGX agent

arXiv:2604.20983v1 Announce Type: cross Abstract: Vision evaluations are typically done through multi-step processes. In most contemporary fields, experts analyze images using structured, evidence-bas

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

DGX agent

arXiv:2604.21860v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This pape

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…6970717273…1259
Next →