AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

What's new in Claude Sonnet 5

DGX agent

What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official

model-releasessimon-willison
30 Jun 2026
Model Releases

When AI Reviews Its Own Code: Recursive Self-Training Collapse in Code LLMs

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.28438v1 Announce Type: cross Abstract: Recursive self-training can degrade neural generative models when generated data is reused without fresh human data or external quality control. We st

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries

DGX agent

arXiv:2606.28332v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical and health-related questions, yet their safety in high-risk medical scenarios remains p

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling

DGX agent

arXiv:2606.28661v1 Announce Type: cross Abstract: People overthink; language models over-sample, and the extra effort can talk both into a worse answer. Reasoning systems answer a hard question by sam

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

DGX agent

arXiv:2606.30383v1 Announce Type: new Abstract: A rapidly growing class of LLM agents is multi-party: the agent acts for a principal (who briefs it, sends follow-ups, and receives results) while also

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Why Do We Need Warm-up? A Theoretical Perspective

DGX agent

arXiv:2510.03164v2 Announce Type: replace Abstract: Learning rate warm-up -- increasing the learning rate at the beginning of training -- has become a ubiquitous heuristic in modern deep learning, yet

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression

DGX agent

arXiv:2606.29712v1 Announce Type: new Abstract: Large language models achieve high reasoning performance via explicit chain-of-thought and reinforcement learning, but require long output sequences and

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Wordle 1,836 4/6 ⬛⬛⬛🟨🟩 ⬛⬛⬛⬛⬛ ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result (puzzle #1,836) solved in 4 attempts, showing the guess sequence and color-coded feedback (gray for incorrect letters, yellow for correct letters in wrong posi

model-releasesanthropic--x
30 Jun 2026
Model Releases

XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation

DGX agent

arXiv:2412.15529v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) synergizes the retrieval of pertinent data with the generative capabilities of Large Language Models (LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

XYZ-IBD: Benchmarking Robust 6D Object Pose Estimation under Real-World Industrial Complexity

DGX agent

arXiv:2506.00599v3 Announce Type: replace Abstract: While current 6D pose estimation benchmarks have reached near-saturation on household objects, they often fail to capture the stochastic and optical

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

You asked, we listened. Claude Desktop on Linux is here! Download link: https://code.claude.com/docs/en/desktop-linux

DGX agent

You asked, we listened. Claude Desktop on Linux is here! Download link: https://code.claude.com/docs/en/desktop-linux Claude Desktop is now available on Linux (Ubuntu and Debian) in beta. Alongside th

model-releasesboris-cherny--x
30 Jun 2026
Model Releases

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are y…

DGX agent

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are you going to ban it outside US? Instead the real question is,

model-releasesclem-delangue--x
30 Jun 2026
Model Releases

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

DGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot.

DGX agent

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot. FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr)

model-releasesclem-delangue--x
30 Jun 2026
Model Releases

Zenith took GPT-5.5 from 5th to 1st on Frontier SWE. @EMostaque and @PeterDiamandis talk it through, and more, on Moonshots. https://youtu.b…

DGX agent

Zenith took GPT-5.5 from 5th to 1st on Frontier SWE. @EMostaque and @PeterDiamandis talk it through, and more, on Moonshots. https://youtu.be/-H7J_-zr7pA?is=Oeo7gg2ZevKCw2IE Media You don't need Fable

model-releasesemad-mostaque--x
30 Jun 2026
Model Releases

Zero-Gated Language-conditioned Human Motion Prediction

DGX agent

arXiv:2606.29208v1 Announce Type: new Abstract: Pose histories provide the core kinematic evidence for 3D human motion prediction, but they lack explicit high-level semantic guidance. This paper intro

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Zero-Shot Depth from Defocus

DGX agent

arXiv:2603.26658v2 Announce Type: replace Abstract: Depth from Defocus (DfD) is the task of estimating a dense metric depth map from a focus stack. Unlike previous works overfitting to a certain datas

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

DGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

A Multi-Attribute Latent Space for Visual Analysis of Watches

DGX agent

arXiv:2606.27897v1 Announce Type: new Abstract: We present a design rationale, embedding model, and interactive visual-analysis system for exploring large wristwatch collections through heterogeneous

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

DGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

A Unified Framework for Vision Transformers Equivariant to Discrete Subgroups of O(2)

DGX agent

arXiv:2606.27864v1 Announce Type: new Abstract: Vision transformers have become a dominant architecture for visual recognition. However, standard models do not explicitly encode the planar symmetries

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Accelerating Attention with Basis Decomposition

DGX agent

arXiv:2510.01718v2 Announce Type: replace Abstract: Attention is a core operation in large language models (LLMs). We present BD Attention (BDA), a lossless algorithmic reformulation of attention. BDA

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Adaptive Momentum and Nonlinear Damping for Neural Network Training

DGX agent

arXiv:2602.00334v2 Announce Type: replace Abstract: Momentum Stochastic Gradient Descent (mSGD) relies on a fixed momentum coefficient shared across all parameters, failing to account for the heteroge

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Advancing Speaker-Based Vocal Effort Classification with WavLM and Data Augmentation in Naturalistic Non-Calibrated Speech Recordings

DGX agent

arXiv:2606.27543v1 Announce Type: cross Abstract: The variations in vocal effort range (e.g. whisper, soft, neutral, loud, shout) alter production and speech acoustics, reducing intelligibility and li

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Agentic Hardware Design as Repository-Level Code Evolution

DGX agent

arXiv:2606.28279v1 Announce Type: cross Abstract: We present HORIZON, a self-evolving agent framework that treats hardware design as repository-level code evolution. A Markdown harness is compiled int

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

AirGroundBench: Probing Spatial Intelligence in Multimodal Large Models under Heterogeneous Multi-View Embodied Collaboration

DGX agent

arXiv:2606.28049v1 Announce Type: new Abstract: In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometric

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

DGX agent

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

model-releasesemad-mostaque--x
29 Jun 2026
Model Releases

Aloe-Vision: Robust Vision-Language Models for Healthcare

DGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

An Empirical Analysis of Factual Errors in Human-Written Text and its Application

DGX agent

arXiv:2606.27959v1 Announce Type: new Abstract: Factual Error Detection (FED), which is the task of identifying factually incorrect spans in a given text, has long been recognized as an important rese

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Any new features we must have in the next version of glm?

DGX agent

This post discusses requested or required features for the next version of GLM (likely referring to Zhipu AI's large language model). The content appears to be a community discussion or announcement o

model-releaseszhipu-ai--x
29 Jun 2026
Model Releases

Applicability of memorization indicators for early spotting of overfitting while recalibrating sEMG-decoders on low sample sizes

DGX agent

arXiv:2606.27855v1 Announce Type: cross Abstract: Deep learning models for surface electromyography (sEMG) can benefit substantially from subject-specific (re-)calibration, since no sufficiently large

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Aurora: A Leverage-Aware Spectral Optimizer

DGX agent

arXiv:2606.27715v1 Announce Type: new Abstract: We show that for tall matrix parameters, like projection matrices in the MLP layers, the Muon update can have row norms that are arbitrarily non-uniform

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Benchmarking Multi-Modal Graph-based Social Media Popularity Prediction

DGX agent

arXiv:2606.27539v1 Announce Type: cross Abstract: Social media popularity prediction aims to forecast the future reach or influence of online content from early-stage observations. Accurate prediction

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Benchmarking on Tasks That Matter: Dataset Selection for Preserving Model Rankings

DGX agent

arXiv:2606.27997v1 Announce Type: new Abstract: Benchmarks of machine learning models often include many datasets, making evaluation expensive. For efficiency, it is preferable to perform evaluations

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Bridging Ab Initio Symmetries and Global Nuclear Masses with Interpretable Neural Networks

DGX agent

arXiv:2606.28287v1 Announce Type: cross Abstract: Ab initio modeling has established Wigner's SU(4) and Elliott's SU(3) as dominant symmetries of the nuclear force in light and intermediate-mass nucle

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence Cloud-Native Simulation Infrastructure for Embodied Intelligence Training, Evaluation, and Data Collection

DGX agent

arXiv:2606.27962v1 Announce Type: new Abstract: This paper presents a cloud-native simulation infrastructure framework for embodied intelligence that supports large-scale training, standardized evalua

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

CalBrief: A Pilot Diagnostic Benchmark for Evidence-Calibrated Scientific Briefing with Large Language Models

DGX agent

arXiv:2606.27383v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as research assistants, yet it remains unclear whether they can calibrate research takeaways to the

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

California strikes a deal with Anthropic to expand the use of Claude products across state agencies and local governments at a 50% discount (Christine Mui/Politico)

DGX agent

Christine Mui / Politico: California strikes a deal with Anthropic to expand the use of Claude products across state agencies and local governments at a 50% discount — “A lot of departments are going

model-releasestechmeme
29 Jun 2026
Model Releases

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026

DGX agent

arXiv:2606.27446v1 Announce Type: new Abstract: This paper describes team HSA_CORAL's submission to the FinCausal 2026 shared task on extracting cause-effect relations from financial narratives via ex

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching

DGX agent

arXiv:2602.20094v2 Announce Type: replace Abstract: As large language models (LLMs) witness increasing deployment in complex, high-stakes decision-making scenarios, it becomes imperative to ground the

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for…

DGX agent

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for chess broadcasts. I trained this model on my Nvidia RTX 508

model-releasesnous-research--x
29 Jun 2026
Model Releases

Claude in Microsoft Foundry is now generally available, hosted on Azure. Azure customers get Claude Opus 4.8 and Claude Haiku 4.5, with Azur…

DGX agent

Claude models are now generally available through Microsoft Foundry on Azure infrastructure, providing Azure customers access to Claude Opus 4.8 and Claude Haiku 4.5. This integration allows enterpris

model-releasesboris-cherny--x
29 Jun 2026
Model Releases

Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

DGX agent

Anthropic’s Claude models in Microsoft Foundry — hosted on Microsoft Azure and running on NVIDIA GB300 Blackwell Ultra GPUs — are now generally available, giving Azure-native enterprises a powerful ne

model-releasesnvidia-blog
29 Jun 2026
Model Releases

Cloud CISO Perspectives: How Google Cloud Security uses AI internally

DGX agent

Welcome to the second Cloud CISO Perspectives for June 2026. Today, we’re discussing how we use AI to chart a path to autonomous software development lifecycle security.As with all Cloud CISO Perspect

model-releasesgoogle-cloud-ai
29 Jun 2026
Model Releases

Complex-Valued 2D Gaussian Representation for Computer-Generated Holography

DGX agent

arXiv:2511.15022v2 Announce Type: replace Abstract: Complex-valued Gaussian primitives have recently been explored for representing holographic radiance fields in 3D novel view synthesis. In this work

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Compression-Driven Anomaly Detection in Brain MRI Using an Interpretable Quantum Autoencoder

DGX agent

arXiv:2606.27411v1 Announce Type: cross Abstract: We study a quantum autoencoder (QAE) for compression-driven anomaly detection in brain MRI data. The approach leverages angle encoding to map image pa

model-releasesarxiv-cs-ai
29 Jun 2026
← Previous
1…156157158159160…472
Next →