AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
24 Jul 2026

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

Model ReleasesDGX agent

arXiv:2607.20425v1 Announce Type: new Abstract: What makes writing 'good' remains a persistent question in literary studies and computational linguistics. We present a two-study investigation of how r

When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard: A Fast Vision-Language Guard for Real-Time Moderation

Model ReleasesDGX agent

arXiv:2607.21401v1 Announce Type: cross Abstract: A vision-language AI assistant returns its answer as a stream of generated tokens. Therefore, a safety guard that watches that answer has to keep up w

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.21535v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by having a cheap draft propose tokens that a target verifies in parallel. Frontier models

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

Model ReleasesDGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

23 Jul 2026

A Novel Hybrid Deep Learning Technique for Speech Emotion Detection using Feature Engineering

Model ReleasesDGX agent

arXiv:2507.07046v3 Announce Type: replace-cross Abstract: Nowadays, speech emotion recognition (SER) plays a vital role in the field of human-computer interaction (HCI) and the evolution of artificial

Agent-Centric Animal Pose Forecasting

AgentsDGX agent

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action --

Alipay-PIBench: A Realistic Payment Integration Benchmark for Coding Agents

Model ReleasesDGX agent

arXiv:2607.14573v3 Announce Type: replace Abstract: Payment integration is a demanding repository-level software task: agents must select a suitable product, implement coordinated client-server flows,

Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document…

Model ReleasesDGX agent

Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document, send a slack message, or update a calendar entry. Ask it t

BaseRT: Advancing Best-in-Class LLM Inference with Apple M5 Neural Accelerators

Model ReleasesDGX agent

arXiv:2607.19438v1 Announce Type: cross Abstract: Apple's M5 generation introduces a redesigned GPU architecture in which every core carries a dedicated Neural Accelerator: on-die matrix units exposed

Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking

Model ReleasesDGX agent

arXiv:2607.19747v1 Announce Type: new Abstract: As large language models and AI agents become the primary consumers of search results, document set quality determines the upper bound of downstream gen

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

Model ReleasesDGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

DobicVLM: Aligning Chest X-Ray Report Generation with Clinically-Grounded Programmatic Rewards via Group Relative Policy Optimization

Model ReleasesDGX agent

arXiv:2607.18988v1 Announce Type: new Abstract: Medical imaging is a cornerstone of diagnostics, yet automated chest X-ray report generation struggles with structural adherence, anatomical completenes

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

Model ReleasesDGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

Generative World Renderer at the Speed of Play

ResearchDGX agent

arXiv:2607.18703v1 Announce Type: new Abstract: Generative world renderer AlayaRenderer receives structured world states exported from physics engines and synthesizes RGB frames. Unlike models that ge

GLID: Gated Local Intrinsic Dimension Repairs the Blind Spots of Face-Forgery Detectors

Model ReleasesDGX agent

arXiv:2607.18770v1 Announce Type: cross Abstract: Fine-tuned foundation-model detectors dominate face-forgery benchmarks, yet they stay blind to generator families absent from training. We present GLI

IBoxCLA: Towards Robust Box-supervised Segmentation of Polyp via Improved Box-dice and Contrastive Latent-anchors

Model ReleasesDGX agent

arXiv:2310.07248v5 Announce Type: replace Abstract: Box-supervised polyp segmentation attracts increasing attention for its cost-effective potential. Existing solutions often rely on learning-free met

In a leaked four-hour investor talk, DeepSeek's Liang Wenfeng says the main US-China gap is compute access, Nvidia's CUDA moat is disintegrating, and more (Fred Gao/Inside China)

Model ReleasesDGX agent

Fred Gao / Inside China: In a leaked four-hour investor talk, DeepSeek's Liang Wenfeng says the main US-China gap is compute access, Nvidia's CUDA moat is disintegrating, and more — In a rare four-hou

Laguna-S-2.1 'thinking forever' loops seem to be a quantization artifact

Model ReleasesDGX agent

If you're running Laguna S 2.1 on llama.cpp and hitting thinking loops because it won't close its </think> tags, you might want to look at your quant before you spend too much time tweaking settings.

LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs

SafetyDGX agent

arXiv:2602.00462v5 Announce Type: replace-cross Abstract: Transforming a large language model (LLM) into a vision-language model (VLM) can be achieved by mapping the visual tokens from a vision encode

LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition

Local AiDGX agent

arXiv:2607.19889v1 Announce Type: new Abstract: Understanding instrument-tissue interactions is essential for context-aware surgical AI and autonomous robotic surgery. Pretrained vision-language model

Learning Explicit Physical Parameter Control and Benchmarking for Video Generation

Model ReleasesDGX agent

arXiv:2607.18924v1 Announce Type: new Abstract: Recent advances in image-to-video generation have improved visual realism, making physically grounded and controllable dynamics an important step toward

Many still debate open vs closed, and compare cost per token (accounting metric). Better to shift attention to cost per successful task. @se…

Model ReleasesDGX agent

Many still debate open vs closed, and compare cost per token (accounting metric). Better to shift attention to cost per successful task. @seldo and the team at @arizeai did so across 2,400 runs. Concl

MoAKE: Toward Unified All-in-One Action Quality Assessment via Mixture of Action Knowledge Experts

ApplicationsDGX agent

arXiv:2607.19826v1 Announce Type: new Abstract: Action Quality Assessment (AQA) aims to objectively evaluate performance quality from action videos. Most existing methods follow a ``one-by-one'' parad

MoGe-3: Fine-Detail Monocular Geometry Estimation with Self-Guided Sparse Volumetric Refinement

Local AiDGX agent

arXiv:2607.17967v2 Announce Type: replace Abstract: Monocular geometry estimation has recently achieved impressive performance across diverse scenes. However, state-of-the-art models still face notabl

No Training, Better Flights: Test-Time Scaled VLMs for UAV Navigation

SafetyDGX agent

arXiv:2607.19288v1 Announce Type: new Abstract: Test-time scaling offers a promising method to improve the inference performance of Vision-Language Models (VLMs) without additional training. Existing

Pathologist Attention-Aligned Report Generation for Prostate Histopathology

SafetyDGX agent

arXiv:2607.19624v1 Announce Type: new Abstract: The allocation of visual attention by pathologists during cancer diagnosis is a highly selective process that critically shapes the information extracte

PathReportEval: A Systematic Benchmark for Pathology Report Generation

Model ReleasesDGX agent

arXiv:2607.18448v1 Announce Type: cross Abstract: Pathology report generation from whole-slide images (WSIs) is a rapidly growing multimodal learning problem, yet progress is difficult to measure beca

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization

Model ReleasesDGX agent

arXiv:2607.19653v1 Announce Type: cross Abstract: Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature im

PRiSM: Prototype Regularization for Few-Shot VLMs

Model ReleasesDGX agent

arXiv:2607.17820v2 Announce Type: replace Abstract: Training-free few-shot adaptation methods have gained significant attention recently in the context of Vision-language Models (VLMs). Yet, current b

Prober.ai: Gated Inquiry-Based Feedback via LLM-Constrained Personas for Argumentative Writing Development

Model ReleasesDGX agent

arXiv:2605.05598v2 Announce Type: replace Abstract: The proliferation of large language models (LLMs) in educational settings has paradoxically undermined the cognitive processes they purport to suppo

Reward-Aware Population Scaling of Evolutionary Strategies in LLM Fine-Tuning

ResearchDGX agent

arXiv:2607.19408v1 Announce Type: new Abstract: Using Evolutionary Strategies (ES) for fine-tuning large language models is attractive because it is memory-efficient, parallel, and compatible with bla

Single-Teacher View Augmentation: Enhancing Knowledge Distillation with Student-Guided Perturbations

Model ReleasesDGX agent

arXiv:2607.11557v2 Announce Type: replace Abstract: Knowledge distillation (KD) typically relies on the fixed perspective of a single teacher, limiting the diversity of supervisory signals. While mult

Some thoughts on the purported distillation timeline. 1. You don't need that many Fable tokens to distill on top of the alleged ~50b Anthrop…

IndustryDGX agent

Some thoughts on the purported distillation timeline. 1. You don't need that many Fable tokens to distill on top of the alleged ~50b Anthropic claimed Moonshot AI distilled in Feb. ~10b, about $60k wo

Sophisticated Policies from Epistemic Priors

Model ReleasesDGX agent

arXiv:2607.19518v1 Announce Type: new Abstract: Sophisticated Inference is a variant of active inference often associated with recursive belief modeling and tree search. We argue that its central comp

Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes

Model ReleasesDGX agent

Prime Intellect Lab offers a streamlined, hosted reinforcement‑learning workflow that lets users customize the NVIDIA Nemotron 3 Nano in minutes. The process establishes a baseline, trains the model o

Statistical Inference for Rank Allocation in Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2607.20205v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become a widely used parameter-efficient fine-tuning method for large language models. Since different modules and laye

Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification

Model ReleasesDGX agent

arXiv:2607.19442v1 Announce Type: cross Abstract: Machine unlearning is commonly evaluated by matching a retrained oracle on trained probes. In a controlled nonce-fact testbed with a matched retrainin

v0.32.3

Model ReleasesDGX agent

What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream

Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study

SafetyDGX agent

arXiv:2607.20270v1 Announce Type: new Abstract: Large language models are increasingly evaluated through the values they endorse, but such evaluations presuppose that models can identify the value exp

Your AI agents are ready. Is your data?

AgentsDGX agent

What’s one of the biggest bottlenecks stopping organizations from scaling their AI initiatives? It isn’t the capabilities of today’s models — it’s their access to business context and semantic meaning

22 Jul 2026

Behind the scenes at Hugging Face headquarter. *literally*

Model ReleasesDGX agent

Behind the scenes at Hugging Face headquarter. *literally* Media We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging F

@Engomar_10 @huggingface @alihfadel @Almalki_io We've seen the outpouring of support for ASRs in under-resourced languages, and we hear you.…

Model ReleasesDGX agent

@Engomar_10 @huggingface @alihfadel @Almalki_io We've seen the outpouring of support for ASRs in under-resourced languages, and we hear you. As a start, we've partnered with @HUMAIN to keep working wi

Holy shit wow

Model ReleasesDGX agent

Holy shit wow We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production during a benchmark evaluation. Shari

MedPerf Meets Google Cloud Confidential Computing: Secure AI Benchmarking for Brain Tumor Research

Model ReleasesDGX agent

At Google Cloud Next 2026 in Las Vegas, MLCommons and Google Cloud demonstrated a powerful new capability for trustworthy medical AI - one that protects patient data, model IP, and benchmark integrity

Troubling …

Model ReleasesDGX agent

Troubling … OpenAI disclosed it themselves yesterday. Their models (GPT-5.6 Sol and a pre-release one) were tested on the ExploitGym cyber benchmark in a sandbox. They escaped, exploited a zero-day to

21 Jul 2026

feeling nostalgic, my favorite blogs on RL & reward hacking https://www.alexirpan.com/2018/02/14/rl-hard.html https://lilianweng.github.io/p…

Model ReleasesDGX agent

feeling nostalgic, my favorite blogs on RL & reward hacking https://www.alexirpan.com/2018/02/14/rl-hard.html https://lilianweng.github.io/posts/2024-11-28-reward-hacking/ TLDR: An openai model, durin

One of the most interesting investigations of my career.

Model ReleasesDGX agent

One of the most interesting investigations of my career. We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face prod

16 Jul 2026

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

Model ReleasesDGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion

Model ReleasesDGX agent

arXiv:2607.13120v1 Announce Type: cross Abstract: Inferring gene regulatory networks (GRNs) from single-cell transcriptomic data is crucial for biological discovery, yet existing approaches suffer fro

Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open Challenges

Local AiDGX agent

arXiv:2607.13045v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as a key paradigm for privacy-preserving collaborative model training across distributed and heterogeneous data so

From Prediction to Collaboration: Interactive Symbolic Music Analysis

Model ReleasesDGX agent

arXiv:2607.13587v1 Announce Type: cross Abstract: Automatic symbolic music analysis has made substantial progress, yet existing systems are typically designed for a single mode of use, such as full-sc

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

Model ReleasesDGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

Multi-view Hand Reconstruction with a Point-Embedded Transformer

ApplicationsDGX agent

arXiv:2408.10581v3 Announce Type: replace Abstract: This work introduces a novel and generalizable multi-view Hand Mesh Reconstruction (HMR) model, named POEM, designed for practical use in real-world

Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment

Model ReleasesDGX agent

arXiv:2607.13941v1 Announce Type: new Abstract: Video aesthetic assessment (VAA) aims to predict how aesthetically pleasing a video is, yet remains far less explored than other visual assessment tasks

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

Model ReleasesDGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

Safe Overtaking for Autonomous Racing Using Hierarchical Optimization and Learning-Based Control

Model ReleasesDGX agent

arXiv:2607.13348v1 Announce Type: new Abstract: Autonomous racing overtaking requires balancing competitive performance with safety under nonlinear vehicle dynamics and real-time constraints. Model Pr

ScanFocus: A Coarse-to-Fine Framework for Spatio-Temporal Video Grounding

Model ReleasesDGX agent

arXiv:2607.13421v1 Announce Type: cross Abstract: Spatio-Temporal Video Grounding (STVG) aims to retrieve the visual trajectory of a specific object from a video stream as described by a natural langu

Semantic Anchoring for Robotic Action Representations

ApplicationsDGX agent

arXiv:2607.13597v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models inherit rich semantic representations from pretrained Vision-Language Models, yet fine-tuning on limited robot dem

Targeted Recovery of Weight-Space Mechanisms From Neural Networks

Model ReleasesDGX agent

arXiv:2607.13047v1 Announce Type: new Abstract: Parameter decomposition (PD) decomposes neural networks into interpretable computational components that faithfully reflect the original network's opera

TheBioCollection: Unified Pre-Training Scale LLM Corpus for Biology

ResearchDGX agent

arXiv:2607.08803v2 Announce Type: replace-cross Abstract: The push toward large language models for biology (BioLM) has created a need for training corpora that can endow models with a genuine underst

← Previous
1…385386387388389…1051
Next →