AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
30 Jun 2026

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

Model ReleasesDGX agent

arXiv:2507.07445v3 Announce Type: replace Abstract: Autonomous agents navigating human society must master both production activities and social interactions, yet existing benchmarks rarely evaluate t

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

Model ReleasesDGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs f…

Model ReleasesDGX agent

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs for the support. “A year ago @GeoffreyHuntley released the Ra

What's new in Claude Sonnet 5

Model ReleasesDGX agent

What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets

ResearchDGX agent

arXiv:2606.29248v1 Announce Type: new Abstract: Vegetable prices in Sri Lanka are highly volatile because the market is largely import-isolated, so supply disruptions quickly drive prices up. This stu

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection

SafetyDGX agent

arXiv:2606.30587v1 Announce Type: cross Abstract: Researchers and practitioners increasingly apply Large Language Models (LLMs) for automated vulnerability detection. Recent work has shown that LLMs a

XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2412.15529v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) synergizes the retrieval of pertinent data with the generative capabilities of Large Language Models (LLM

29 Jun 2026

A Multi-Attribute Latent Space for Visual Analysis of Watches

Model ReleasesDGX agent

arXiv:2606.27897v1 Announce Type: new Abstract: We present a design rationale, embedding model, and interactive visual-analysis system for exploring large wristwatch collections through heterogeneous

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

Model ReleasesDGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

Any new features we must have in the next version of glm?

Model ReleasesDGX agent

This post discusses requested or required features for the next version of GLM (likely referring to Zhipu AI's large language model). The content appears to be a community discussion or announcement o

Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence Cloud-Native Simulation Infrastructure for Embodied Intelligence Training, Evaluation, and Data Collection

Model ReleasesDGX agent

arXiv:2606.27962v1 Announce Type: new Abstract: This paper presents a cloud-native simulation infrastructure framework for embodied intelligence that supports large-scale training, standardized evalua

Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training

SafetyDGX agent

arXiv:2602.16065v2 Announce Type: replace-cross Abstract: As artificial intelligence (AI)-generated content proliferates, models are increasingly trained on their own outputs, risking progressive degr

CBD: API-Only LLM Black-Box Unlearning through Controlled Behavioral Divergence

TutorialsDGX agent

arXiv:2606.27683v1 Announce Type: cross Abstract: Edge devices increasingly invoke large language models (LLMs) through API services for context aware edge intelligence, while edge generated data may

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for…

Model ReleasesDGX agent

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for chess broadcasts. I trained this model on my Nvidia RTX 508

DiScoFormer: One transformer for density and score, across distributions

ToolsDGX agent

DiScoFormer is a unified transformer architecture designed to handle both density estimation and score-based modeling across different probability distributions. The model enables a single framework t

DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain

Model ReleasesDGX agent

arXiv:2504.16116v4 Announce Type: replace-cross Abstract: The Web3 ecosystem, underpinned by cryptographic primitives and decentralized consensus, represents a high-stakes environment where software v

FoggyTrust: Robust Federated Learning with Hierarchical Trust Networks

Model ReleasesDGX agent

arXiv:2606.27622v1 Announce Type: new Abstract: Byzantine-robust federated learning seeks to protect distributed model training from malicious or corrupted clients without requiring access to their pr

Freshness and the Limits of Heuristic Trend Detection in Temporal RAG

Model ReleasesDGX agent

arXiv:2509.19376v2 Announce Type: replace-cross Abstract: We present a lightweight, model-agnostic temporal layer for RAG and use cybersecurity data to separate two problems that are usually conflated

HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-Speech

TutorialsDGX agent

arXiv:2606.28249v1 Announce Type: cross Abstract: Recently, Large Language Model (LLM)-based Text-to-Speech (TTS) models have achieved remarkable naturalness. However, the standard Supervised Fine-Tun

June Launches | Desktop, MCP & Core Engine Improvements — Live Demo & Q&A https://x.com/i/broadcasts/1aJbddORnaoKX

Model ReleasesDGX agent

ComfyUI announced June product launches featuring desktop application improvements, MCP (Model Context Protocol) enhancements, and core engine upgrades during a live demonstration with audience Q&A se

Monocular Avatar Reconstruction via Cascaded Diffusion Priors and UV-Space Differentiable Shading

Model ReleasesDGX agent

arXiv:2606.28144v1 Announce Type: new Abstract: Reconstructing high-fidelity, relightable 3D avatars from a single in-the-wild image is a challenging ill-posed problem, primarily hindered by the scarc

Pair Nova 2 Lite with Claude for cost-optimized document processing

Model ReleasesDGX agent

In this post, we show how pairing Amazon Nova 2 Lite with Anthropic’s Claude Sonnet 4.6 delivers an efficient solution for digitizing scanned documents at scale. We built a two-model pipeline on Amazo

Parameter-Efficient Continuous-Variable Photonic Quantum Neural Networks for Edge Quantum AI: Demonstration in Oral Cancer Detection

Model ReleasesDGX agent

arXiv:2606.28252v1 Announce Type: cross Abstract: Early detection of oral cancer markedly improves clinical outcomes, yet specialized diagnostic tools remain scarce in low-resource settings. Smartphon

Parameter-Efficient Quantum-Inspired Fast Weight Programmers for Traffic-Matrix Forecasting

Model ReleasesDGX agent

arXiv:2606.27821v1 Announce Type: cross Abstract: Traffic matrices (TMs) capture network-wide origin-destination demand and are central to traffic engineering, yet accurate whole-matrix forecasting re

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation

SafetyDGX agent

arXiv:2606.28128v1 Announce Type: cross Abstract: Video generation models have emerged as a promising paradigm for embodied world simulation. However, both general-domain video generators and robot-sp

Prism Transformer: Progressive Head Schedules for Hierarchical Attention Processing

Model ReleasesDGX agent

arXiv:2606.27449v1 Announce Type: new Abstract: Multi-head attention conventionally partitions the hidden dimension equally across all heads at every layer, enforcing an identical representational sub

QuantV2X: A Fully Quantized Multi-Agent System for Cooperative Perception

Model ReleasesDGX agent

arXiv:2509.03704v2 Announce Type: replace Abstract: Cooperative perception through Vehicle-to-Everything (V2X) communication offers significant potential for enhancing vehicle perception by mitigating

RANSAC Scoring Done Right

Model ReleasesDGX agent

arXiv:2606.27385v1 Announce Type: cross Abstract: The most widely used RANSAC variants score candidate models by counting inliers or summing per-point scores that saturate beyond a residual threshold.

RelBall: Relation Ball with Quaternion Rotation for Knowledge Graph Completion

ApplicationsDGX agent

arXiv:2606.27967v1 Announce Type: new Abstract: Real-world knowledge graphs are often incomplete, lacking many valid facts. Knowledge Graph Completion (KGC) aims to predict missing links using known t

RS-Diffuser: Risk-Sensitive Diffusion Planning with Distributional Value Guidance

Model ReleasesDGX agent

arXiv:2606.27766v1 Announce Type: cross Abstract: Offline reinforcement learning enables policy learning from fixed datasets without additional environment interaction, making it appealing for safety-

TA-SparseMG: Trend-Aware Sparse Forecasting via Multi-Scale Gating for Long-Term Time Series

Model ReleasesDGX agent

arXiv:2606.27908v1 Announce Type: new Abstract: Long-term time series forecasting finds extensive applications in domains such as power demand, traffic flow, meteorological observation, and renewable

ToolPrivacyBench: Benchmarking Purpose-Bound Privacy in Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2606.28061v1 Announce Type: cross Abstract: Large language models (LLMs) have increasingly moved from standalone text generation systems to agents that invoke external tools, access environments

TreeLoRA: Efficient Continual Learning via Layer-Wise LoRAs Guided by a Hierarchical Gradient-Similarity Tree

Model ReleasesDGX agent

arXiv:2506.10355v2 Announce Type: replace Abstract: Many real-world applications collect data in a streaming environment, where learning tasks are encountered sequentially. This necessitates continual

When One Adapter Speaks for Many: Discovering Low-Rank Redundancy in Continual Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.28117v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become the standard tool for parameter-efficient fine-tuning of large pretrained models. When applied sequentially across

28 Jun 2026

Is Gemini 3.5 Pro being export controlled? Because if not...

Model ReleasesDGX agent

Ethan Mollick raises questions about whether Google's Gemini 3.5 Pro model should be subject to export controls, suggesting concerns about its capabilities and potential regulatory implications. The p

Like if you have been using Qwen & Kimi & MiniMax, it feels like GLM is right on the curve. Which is itself impressive, and suggests that My…

Model ReleasesDGX agent

Like if you have been using Qwen & Kimi & MiniMax, it feels like GLM is right on the curve. Which is itself impressive, and suggests that Mythos class models are coming in 6-12 months (if they are all

Looks like Llama CPP just merged DFLASH support into main! I wonder how this will stack up against MTP 👀 https://github.com/ggml-org/llama.…

Model ReleasesDGX agent

Llama.cpp has merged DFLASH support into its main branch, which is a development update related to optimization or acceleration technology for running large language models locally. The post expresses

The more serious answer

Model ReleasesDGX agent

The more serious answer @emollick I actually wouldn't be surprised if we're generally in an AI model release moratorium (regardless of capabilities) until the cybersecurity benchmark is finalized. Not

27 Jun 2026

America is about to lose the AI race, and it will not happen at the frontier. It will happen at the floor. Everyone is watching who ships th…

HardwareDGX agent

America is about to lose the AI race, and it will not happen at the frontier. It will happen at the floor. Everyone is watching who ships the smartest model. The actual war is over the 80% of tokens n

An interesting way to take Noam at his word in regards to always keeping a constant inference budget for any eval reporting - is that open m…

ToolsDGX agent

An interesting way to take Noam at his word in regards to always keeping a constant inference budget for any eval reporting - is that open models have a lot more dollar per token mileage than closed m

Fugu-Ultra is now available on Vercel AI Gateway https://vercel.com/changelog/sakana-fugu-ultra-now-available-on-ai-gateway ✨

ResearchDGX agent

Sakana AI's Fugu-Ultra model is now available through Vercel's AI Gateway, expanding access to this language model through Vercel's infrastructure. This integration allows developers to use Fugu-Ultra

26 Jun 2026

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents

Model ReleasesDGX agent

arXiv:2606.26627v1 Announce Type: cross Abstract: Large language model agents increasingly query databases, search document collections, call external APIs, remember past interactions, and act on a us

auto-psych: Automating the science of mind using agent-driven theory discovery and experimentation

Model ReleasesDGX agent

arXiv:2606.26460v1 Announce Type: new Abstract: AI-based scientific automation is increasingly possible by using agents to generate hypotheses, design experiments, and analyze data. Data collection is

Detecting and Controlling Sycophancy with Cascading Linear Features

ResearchDGX agent

arXiv:2606.26155v1 Announce Type: new Abstract: Interpreting and controlling model behaviors through activation steering methods requires many pairs of contrastive samples that clearly exhibit desired

Disco-LoRA: Disentangled Composition of Content, Style, and Motion for Multi-concept Video Customization

Model ReleasesDGX agent

arXiv:2606.26668v1 Announce Type: cross Abstract: Video customization based on Text-to-Video (T2V) models aims to learn specific features from reference data to generate controllable videos. While sig

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM

ResearchDGX agent

arXiv:2606.26120v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) offer a promising alternative to autoregressive models, excelling in text generation tasks due to their bidirect

Error-Conditioned Neural Solvers

Model ReleasesDGX agent

arXiv:2606.27354v1 Announce Type: cross Abstract: Neural surrogate models offer fast approximate mappings from PDE parameters to solutions, but they typically treat solving as a purely statistical tas

From Weights to Features: SAE-Guided Activation Regularization for LLM Continual Learning

Model ReleasesDGX agent

arXiv:2606.26629v1 Announce Type: cross Abstract: Weight-space regularization methods such as Elastic Weight Consolidation (EWC) are the standard approach to catastrophic forgetting in continual learn

GPT-5.6 Sol matches Mythos Preview on ExploitBench, adds Ultra mode with subagents for complex workflows, and max reasoning for deep problem-solving (OpenAI)

Model ReleasesDGX agent

OpenAI: GPT-5.6 Sol matches Mythos Preview on ExploitBench, adds Ultra mode with subagents for complex workflows, and max reasoning for deep problem-solving — We're beginning a limited preview of the

https://huggingface.co/nvidia/GLM-5.2-NVFP4

HardwareDGX agent

NVIDIA's GLM-5.2-NVFP4 is a quantized version of a large language model optimized for inference efficiency using NVIDIA's proprietary quantization format. The model is hosted on Hugging Face and repre

I can personally attest: OpenClaude using GLM 5.2 is now performing on par with Claude Code powered by Opus 4.8.

Model ReleasesDGX agent

I cannot verify the claims in this post as the URL format appears invalid and the specific version numbers (GLM 5.2, Claude Code/Opus 4.8) don't correspond to publicly documented model releases as of

Information-Aware KV Cache Compression for Long Reasoning

Model ReleasesDGX agent

arXiv:2606.26875v1 Announce Type: cross Abstract: Reasoning capability has advanced rapidly in large language models (LLMs), leading to an increasing size of key-value (KV) cache in both prefilling an

Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities Invisible to Standard Evaluation

Model ReleasesDGX agent

arXiv:2606.27091v1 Announce Type: cross Abstract: LLMs fine-tuned for security classification are usually evaluated on held-out examples from the same distribution as their training data. We show that

Learning from Equivalence Queries, Revisited

ResearchDGX agent

arXiv:2604.04535v2 Announce Type: replace Abstract: Modern machine learning systems, such as generative models and recommendation systems, often evolve through a cycle of deployment, user interaction,

Learning to Select Maximum Clique Algorithms: From Traditional Machine Learning to a Dual-Channel Hybrid Neural Architecture

Model ReleasesDGX agent

arXiv:2508.08005v4 Announce Type: replace-cross Abstract: The Maximum Clique Problem (MCP) is an NP-hard problem with wide-ranging applications in fields such as bioinformatics, network science, and s

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds

Model ReleasesDGX agent

arXiv:2606.26964v1 Announce Type: new Abstract: As embodied AI and world models increasingly operate in dynamic 3D environments, visual perception must move beyond passively interpreting given observa

Mask to Concept: Auto-Promptable SAM3 via Efficient Test-Time Concept Embedding Search for Few-Shot Annotation

Model ReleasesDGX agent

arXiv:2606.26711v1 Announce Type: new Abstract: Transforming foundation segmentation models from human-prompted tools into auto-promptable annotators is critical for scalable medical data annotation.

MKG-RAG-Bench: Benchmarking Retrieval in Multimodal Knowledge Graph-Augmented Generation

Model ReleasesDGX agent

arXiv:2606.26458v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) over knowledge graphs has emerged as a promising approach for grounding large language models, yet existing benchma

← Previous
1…392393394395396…1051
Next →