AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,574 results
Model Releases

Gradient Extrapolation-Based Policy Optimization

DGX agent

arXiv:2605.06755v1 Announce Type: cross Abstract: Reinforcement learning is widely used to improve the reasoning ability of large language models, especially when answers can be automatically checked.

model-releasesarxiv-cs-ai
11 May 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Gradient Starvation in Binary-Reward GRPO: Why Group-Mean Centering Fails and Why the Simplest Fix Works

DGX agent

arXiv:2605.07689v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a standard algorithm for reinforcement learning from verifiable rewards, but its group-mean-centered advant

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Graph-Structured Hyperdimensional Computing for Data-Efficient and Explainable Process-Structure-Property Prediction

DGX agent

arXiv:2605.07999v1 Announce Type: cross Abstract: Multiphoton photoreduction enables high-fidelity fabrication of complex 3D microstructures, yet reliable process-structure-property (PSP) prediction r

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GraphReAct: Reasoning and Acting for Multi-step Graph Inference

DGX agent

arXiv:2605.07357v1 Announce Type: new Abstract: Reasoning-acting frameworks enhance large language models (LLMs) by interleaving reasoning with actions for dynamic information acquisition. However, ex

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Grok just completely changed the game for real-time financial research. We just got the preview to 'Grok Skills,' and they look 10x more pow…

DGX agent

Grok just completely changed the game for real-time financial research. We just got the preview to 'Grok Skills,' and they look 10x more powerful than Claude Skills. In seconds, you can create workflo

model-releaseselon-musk--x
11 May 2026
Model Releases

GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations

DGX agent

arXiv:2605.07053v1 Announce Type: cross Abstract: Benchmarks like GSM8K are popular measures of mathematical reasoning, but leaderboard gains can overstate true capability due to memorization of fixed

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GTIG AI Threat Tracker: Adversaries Leverage AI for Vulnerability Exploitation, Augmented Operations, and Initial Access

DGX agent

Executive Summary Since our February 2026 report on AI-related threat activity, Google Threat Intelligence Group (GTIG) has continued to track a maturing transition from nascent AI-enabled operations

model-releasesgoogle-cloud-ai
11 May 2026
Model Releases

Hallucination Detection via Activations of Open-Weight Proxy Analyzers

DGX agent

arXiv:2605.07209v1 Announce Type: cross Abstract: We introduce a proxy-analyzer framework for detecting hallucinations in large language models. Instead of looking inside the generating model, our sys

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Have Graph -- Will Lift? The Case for Higher-Order Benchmarks

DGX agent

arXiv:2605.07397v1 Announce Type: new Abstract: After a somewhat rocky start, geometry and topology have established a foothold in machine learning. Message passing, either on graphs or higher-order c

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition

DGX agent

arXiv:2605.07766v1 Announce Type: new Abstract: Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are mi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

HiDream-Studio v.01 has been released! It is fast and powerful and open-sourced on Github | Easy Install

DGX agent

HiDream-Studio v.01 was open-sourced on May 8, 2026, releasing the HiDream-O1-Image model (8B parameters) with both undistilled and distilled variants. HiDream-O1-Image is a unified image generative f

model-releasesr-stablediffusion
11 May 2026
Model Releases

Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07512v1 Announce Type: new Abstract: Class-incremental learning aims to continuously acquire new knowledge while preserving previously learned information, thereby mitigating catastrophic f

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Hierarchical Task Network Planning with LLM-Generated Heuristics

DGX agent

arXiv:2605.07707v1 Announce Type: new Abstract: HTN planning is a variation of classical planning where, instead of searching for a linear sequence of actions, an algorithm decomposes higher-level tas

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Big Should a Wireless Foundation Model Be?

DGX agent

arXiv:2605.07266v1 Announce Type: cross Abstract: Wireless foundation models are rapidly emerging as a key enabler of AI-native communication systems, yet a fundamental question remains unanswered: ho

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

How ChatGPT adoption broadened in early 2026

DGX agent

In early 2026, ChatGPT expanded its user base across broader demographics and industry sectors, moving beyond early adopters to mainstream adoption. The report likely documents growth metrics, new use

model-releasesopenai
11 May 2026
Model Releases

How enterprises are scaling AI

DGX agent

This guide from OpenAI outlines strategies and best practices for enterprises implementing and scaling artificial intelligence across their organizations. It likely covers topics such as infrastructur

model-releasesopenai
11 May 2026
Model Releases

How Far Are VLMs from Privacy Awareness in the Physical World? An Empirical Study

DGX agent

arXiv:2605.05340v2 Announce Type: replace-cross Abstract: As Vision-Language Models (VLMs) are increasingly deployed as autonomous cognitive cores for embodied assistants, evaluating their privacy awa

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings

DGX agent

arXiv:2605.07492v1 Announce Type: new Abstract: The past year has seen over 20 open-source document parsing models, yet thefield still benchmarks almost exclusively on OmniDocBench, a 1,355-pagemanual

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

HumanNet: Scaling Human-centric Video Learning to One Million Hours

DGX agent

arXiv:2605.06747v1 Announce Type: new Abstract: Progress in embodied intelligence increasingly depends on scalable data infrastructure. While vision and language have scaled with internet corpora, lea

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

HyperEyes: Dual-Grained Efficiency-Aware Reinforcement Learning for Parallel Multimodal Search Agents

DGX agent

arXiv:2605.07177v1 Announce Type: cross Abstract: Existing multimodal search agents process target entities sequentially, issuing one tool call per entity and accumulating redundant interaction rounds

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

🤯 i need all of you to stop what you're doing and look at this @webassembly + transformers.js + @googlegemma powered robot, running complet…

DGX agent

🤯 i need all of you to stop what you're doing and look at this @webassembly + transformers.js + @googlegemma powered robot, running completely offline I think Reachy is the one who needs chess lessons

model-releasesclem-delangue--x
11 May 2026
Model Releases

I'm starting office hours for Claude Code's cloud environments. If you use any of our cloud related features on desktop/web/mobile, come tal…

DGX agent

I'm starting office hours for Claude Code's cloud environments. If you use any of our cloud related features on desktop/web/mobile, come talk to me! Bring feature requests, bug reports, or anything in

model-releasesthariq--x
11 May 2026
Model Releases

Implicit Preference Alignment for Human Image Animation

DGX agent

arXiv:2605.07545v1 Announce Type: cross Abstract: Human image animation has witnessed significant advancements, yet generating high-fidelity hand motions remains a persistent challenge due to their hi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

In-Context Credit Assignment via the Core

DGX agent

arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Inference Time Causal Probing in LLMs

DGX agent

arXiv:2605.07631v1 Announce Type: new Abstract: Causal probing methods aim to test and control how internal representations influence the behavior of generative models. In causal probing, an intervent

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Intent-Driven Semantic ID Generation for Grounded Conversational News Recommendation

DGX agent

arXiv:2605.07613v1 Announce Type: new Abstract: Conversational news recommendation requires grounding each suggestion in a rapidly evolving article corpus while addressing implicit user intents that l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

IntentGrasp: A Comprehensive Benchmark for Intent Understanding

DGX agent

arXiv:2605.06832v1 Announce Type: cross Abstract: Accurately understanding the intent behind speech, conversation, and writing is crucial to the development of helpful Large Language Model (LLM) assis

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

DGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Introducing Claude Platform on AWS: Anthropic’s native platform, through your AWS account

DGX agent

Today, we're excited to announce the general availability of Claude Platform on AWS. Claude Platform on AWS is a new service that gives customers direct access to Anthropic's native Claude Platform ex

model-releasesaws-ml-blog
11 May 2026
Model Releases

Introducing Daybreak: frontier AI for cyber defenders. Daybreak brings together the most capable OpenAI models, Codex, and our security part…

DGX agent

Introducing Daybreak: frontier AI for cyber defenders. Daybreak brings together the most capable OpenAI models, Codex, and our security partners to accelerate cyber defense and continuously secure sof

model-releasesopenai--x
11 May 2026
Model Releases

Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies

DGX agent

arXiv:2510.22944v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have become indispensable for automated code generation, yet the quality and security of their outputs remain a c

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Knowing but Not Correcting: Routine Task Requests Suppress Factual Correction in LLMs

DGX agent

arXiv:2605.05957v2 Announce Type: replace Abstract: LLMs reliably correct false claims when presented in isolation, yet when the same claims are embedded in task-oriented requests, they often comply r

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

DGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Learning Agent Routing From Early Experience

DGX agent

arXiv:2605.07180v1 Announce Type: new Abstract: LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capa

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Learning Material-Aware Hamiltonian Risk Fields for Safe Navigation

DGX agent

arXiv:2605.07038v1 Announce Type: new Abstract: Risk-aware navigation should be selective: a policy should expose evasive degrees of freedom only when the local scene admits a lower-risk feasible mane

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation

DGX agent

arXiv:2605.07640v1 Announce Type: cross Abstract: Remote sensing lithology interpretation is fundamental to geological surveys, mineral exploration, and regional geological mapping. Unlike general lan

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

DGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Local open-weight AI on a laptop has been improving more than twice as fast as Moore's Law! Between May 2024 and May 2026, the most expensiv…

DGX agent

Local open-weight AI on a laptop has been improving more than twice as fast as Moore's Law! Between May 2024 and May 2026, the most expensive MacBook Pro you could buy stayed at 128 GB of unified memo

model-releasesclem-delangue--x
11 May 2026
Model Releases

Mage: Multi-Axis Evaluation of LLM-Generated Executable Game Scenes Beyond Compile-Pass Rate

DGX agent

arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System

DGX agent

arXiv:2605.05949v2 Announce Type: replace Abstract: Algorithmic problem solving serves as a rigorous testbed for evaluating structured reasoning in AI coding systems, as it directly reflects a model's

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mask2Cause: Causal Discovery via Adjacency Constrained Causal Attention

DGX agent

arXiv:2605.07280v1 Announce Type: cross Abstract: Leveraging deep learning for causal discovery in time series remains challenging because existing neural methods predominantly rely on component-wise

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators

DGX agent

arXiv:2605.07600v1 Announce Type: cross Abstract: Recent methods for improving LLM mathematical reasoning, whether through MCTS-based test-time search or causal graph-guided knowledge injection, canno

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries

DGX agent

arXiv:2605.07147v1 Announce Type: cross Abstract: The ecosystem of Lean and Mathlib has become the de facto standard for large language model (LLM) assisted formal reasoning with remarkable successes

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning

DGX agent

arXiv:2605.07850v1 Announce Type: cross Abstract: With the rise in scale for deep learning models to billions of parameters, the computational cost of fine-tuning remains a significant barrier to depl

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing

DGX agent

arXiv:2605.07646v1 Announce Type: cross Abstract: While explicit reasoning trajectories enhance model interpretability, existing paradigms often rely on monolithic chains that lack intermediate verifi

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…341342343344345…471
Next →