AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,060 results
Model Releases

Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning

DGX agent

arXiv:2602.01058v2 Announce Type: replace-cross Abstract: Post-training of reasoning LLMs is a holistic process that typically consists of an offline SFT stage followed by an online reinforcement lear

model-releasesarxiv-cs-ai
29 May 2026
Agents

Governing Technical Debt in Agentic AI Systems

X Post
Paper
YouTube
Reddit
GitHub
DGX agent

arXiv:2605.29129v1 Announce Type: new Abstract: Agentic AI systems are increasingly being explored as production infrastructure: they reason over multiple steps, call tools, act through workflows, and

agentsarxiv-cs-ai
29 May 2026
Model Releases

GPF-LiveNews: A Streaming Evaluation Protocol for Group-Conditioned Framing in Large Language Models

DGX agent

arXiv:2605.28848v1 Announce Type: cross Abstract: Deployed language models are evaluated in a non-stationary environment: model versions, retrieval layers, safety systems, and real-world inputs all ch

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

GPIC: A Giant Permissive Image Corpus for Visual Generation

DGX agent

arXiv:2605.30341v1 Announce Type: cross Abstract: Studying scalable methods for visual generative modeling requires large, accessible, and stable datasets. We introduce GPIC, a Giant Permissive Image

model-releasesarxiv-cs-ai
29 May 2026
Research

GPS-Enhanced Tourist Mobility Modeling with Seasonal Spatial Priors and LLM-Based Activity Chain Generation

DGX agent

arXiv:2605.29578v1 Announce Type: new Abstract: Tourist mobility poses a distinct challenge for urban transportation planning. Unlike resident commuting, tourist travel is largely non-routine, attract

researcharxiv-cs-ai
29 May 2026
Model Releases

GPT-5.5 Pro is the model producing many of the novel math proofs, and also the model you should have reviewing any technical or academic pap…

DGX agent

I cannot verify this claim as it references a future date (the URL timestamp appears invalid) and describes a model version (GPT-5.5 Pro) that doesn't exist in current public releases. The post appear

model-releasesethan-mollick--x
29 May 2026
Model Releases

Gradient Perturbation: Learning to Perturb Gradients for Adaptive Training

DGX agent

arXiv:2605.29494v1 Announce Type: new Abstract: Deep neural network training involves both forward propagation (from features through logits to loss) and backward propagation (from loss through gradie

model-releasesarxiv-cs-lg
29 May 2026
Research

Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation

DGX agent

arXiv:2602.08646v2 Announce Type: replace Abstract: We propose a gradient preconditioning method that makes reward-guided generation with one-step generative models both efficient and reliable. Test-t

researcharxiv-cs-lg
29 May 2026
Model Releases

Gram: Assessing sabotage propensities via automated alignment auditing

DGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

model-releasesarxiv-cs-ai
29 May 2026
Safety

Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop

DGX agent

arXiv:2601.17670v2 Announce Type: replace-cross Abstract: Mathematical programming is widely employed across various sectors - such as logistics, energy, and workforce planning - to model and solve in

safetyarxiv-cs-ai
29 May 2026
Safety

Graph-Enhanced Policy Optimization in LLM Agent Training

DGX agent

arXiv:2510.26270v2 Announce Type: replace Abstract: Multi-step LLM agents in interactive environments represent a crucial step toward long-horizon decision-making. To train such agents, group-based re

safetyarxiv-cs-ai
29 May 2026
Model Releases

GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents

DGX agent

arXiv:2605.29668v1 Announce Type: new Abstract: LLM agents acting in structured environments fail in operational rather than conversational ways, and reliability depends on procedural knowledge of the

model-releasesarxiv-cs-ai
29 May 2026
Research

GRASP: Plan-Guided Graph Retrieval with Adaptive Fusion and Reranking on Semi-Structured Knowledge Bases

DGX agent

arXiv:2605.30237v1 Announce Type: cross Abstract: Semi-structured knowledge bases (SKBs) embed textual documents in a typed graph of entities and relations, and underpin applications such as product s

researcharxiv-cs-cl
29 May 2026
Safety

GrepSeek: Training Search Agents for Direct Corpus Interaction

DGX agent

arXiv:2605.29307v1 Announce Type: cross Abstract: Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and inf

safetyarxiv-cs-ai
29 May 2026
Agents

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic …

DGX agent

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic coding. Priced at 1/m input and 2/m output, it’s extremely c

agentselon-musk--x
29 May 2026
Model Releases

GroundAct: Can LLM Agents Ground Actions in Environmental States?

DGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

model-releasesarxiv-cs-ai
29 May 2026
Safety

Grounded 3D-Aware Spatial Vision-Language Modeling

DGX agent

arXiv:2605.30307v1 Announce Type: new Abstract: We present GR3D, a spatial vision language model equipped with three complementary grounding capabilities--explicit 2D grounding, implicit 2D grounding,

safetyarxiv-cs-cv
29 May 2026
Model Releases

GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human

DGX agent

arXiv:2605.28882v1 Announce Type: cross Abstract: With the rapid advancement of large language models, evaluating human-likeness in open-ended conversation has become increasingly important. However,

model-releasesarxiv-cs-ai
29 May 2026
Safety

GRPO is Secretly a Process Reward Model

DGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

safetyarxiv-cs-ai
29 May 2026
Safety

GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German

DGX agent

arXiv:2605.30214v1 Announce Type: new Abstract: Third-person singular pronouns have long been used to study stereotypical biases in language models and to test their abilities to reason about referenc

safetyarxiv-cs-cl
29 May 2026
Model Releases

GTA: Generating Long-Horizon Tasks for Web Agents at Scale

DGX agent

arXiv:2605.29218v1 Announce Type: new Abstract: Web agents, which couple language models with browsing and tool-use capabilities, show promise as open web assistants. Yet progress is increasingly limi

model-releasesarxiv-cs-ai
29 May 2026
Safety

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization

DGX agent

arXiv:2605.29198v1 Announce Type: new Abstract: Group-advantage-based reinforcement learning methods, such as GRPO and DAPO, have demonstrated strong performance across diverse domains, including math

safetyarxiv-cs-cv
29 May 2026
Model Releases

GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing

DGX agent

arXiv:2605.29532v1 Announce Type: cross Abstract: Exploratory GUI testing is a particularly demanding setting for MLLM agents: without predefined test scripts, an agent must autonomously navigate an a

model-releasesarxiv-cs-ai
29 May 2026
Agents

Had a blast yesterday attending at @techeurope_'s Applied AI Conference in Berlin! I had a talk about building document agents and agentic d…

DGX agent

Had a blast yesterday attending at @techeurope_'s Applied AI Conference in Berlin! I had a talk about building document agents and agentic development in general, that you can find here: https://astra

agentsjerry-liu--x
29 May 2026
Model Releases

Hallucination Detection-Guided Preference Optimization for Clinical Summarization

DGX agent

arXiv:2605.28910v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise on summarization tasks, but they often produce hallucinations, which are unsupported or incorrect stat

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Hallucination Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Semantic Caching

DGX agent

arXiv:2605.29055v1 Announce Type: new Abstract: Hallucination remains a major reliability barrier for production LLM systems, particularly in multi-agent pipelines where unsupported claims can propaga

model-releasesarxiv-cs-ai
29 May 2026
Research

HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering

DGX agent

arXiv:2512.24562v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong question answering (QA) performance but can produce fluent answers unsupported by available evidence. Ex

researcharxiv-cs-cl
29 May 2026
Model Releases

Hands-on with Gemini Spark beta rolling out to AI Ultra subs: planned a birthday party from emails and calendar, but called a live-in boyfriend a 'close friend' (Reece Rogers/Wired)

DGX agent

Reece Rogers / Wired: Hands-on with Gemini Spark beta rolling out to AI Ultra subs: planned a birthday party from emails and calendar, but called a live-in boyfriend a “close friend” — Google's new AI

model-releasestechmeme
29 May 2026
Industry

Hardware’s back: AI supercharges server, PC and memory sales

DGX agent

Hardware firms are cleaning up bigtime as enterprises and cloud providers can’t get enough computing power for their artificial intelligence dreams. Dell Technology’s stock rocketed an incredible 31%

industrysiliconangle
29 May 2026
Agents

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

DGX agent

arXiv:2605.29354v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly participate in software development workflows by generating code, selecting dependencies, and producing package

agentsarxiv-cs-lg
29 May 2026
Safety

Harmonizing Real-Time Constraints and Long-Horizon Reasoning: An Asynchronous Agentic Framework for Dynamic Scheduling

DGX agent

arXiv:2605.29262v1 Announce Type: new Abstract: The Dynamic Flexible Job Shop Scheduling Problem (DFJSP) necessitates a trade-off between instant reaction to stochastic disturbances and global optimiz

safetyarxiv-cs-ai
29 May 2026
Safety

Harnessing non-adversarial robustness in large language models

DGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

safetyarxiv-cs-ai
29 May 2026
Research

HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization

DGX agent

arXiv:2605.29843v1 Announce Type: cross Abstract: Post-training quantization (PTQ) is essential for deploying LLMs under memory and bandwidth constraints. However, extreme low-bit quantization remains

researcharxiv-cs-ai
29 May 2026
Tutorials

Has @AnthropicAI completely given up on making API usage reasonably-priced? Following the token-usage changes recently they announced variou…

DGX agent

Has @AnthropicAI completely given up on making API usage reasonably-priced? Following the token-usage changes recently they announced various updates to *subscription* usage to make it more reasonable

tutorialsjeremy-howard--x
29 May 2026
Tutorials

Having Grok Build sub-agents to iterate several ideas for me on data loading, batching, inference, and writing results to files for dense da…

DGX agent

Having Grok Build sub-agents to iterate several ideas for me on data loading, batching, inference, and writing results to files for dense datasets before I went to sleep. It gave me a nice summary of

tutorialselon-musk--x
29 May 2026
Tutorials

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens

DGX agent

arXiv:2512.15133v2 Announce Type: replace-cross Abstract: Proteins inherently possess a consistent sequence-structure duality. The abundance of protein sequence data, which can be readily represented

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, …

DGX agent

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer sit down on came

model-releasesgoogle-ai--x
29 May 2026
Model Releases

HEART-Bench: Do LLM Agents Exhibit Human-like Psychology?

DGX agent

arXiv:2605.30058v1 Announce Type: new Abstract: While LLM agents have demonstrated remarkable task-oriented abilities such as planning, reasoning, and action, few works have treated them as complete h

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Here's an extended edit of the quote that includes a following fragment where Andrew Macdonald called the trade 'harder to justify' - full, …

DGX agent

Here's an extended edit of the quote that includes a following fragment where Andrew Macdonald called the trade 'harder to justify' - full, unedited transcript is here: https://gist.github.com/simonw/

model-releasessimon-willison--x
29 May 2026
Tools

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end work…

DGX agent

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end workflow from UI to deployment → Real-time team collaboration wi

toolsreplit--x
29 May 2026
Industry

Here's why the failure of Blue Origin's New Glenn rocket is so catastrophic

DGX agent

Blue Origin's New Glenn rocket exploded during an engine-firing test at Cape Canaveral on May 28, 2026 , which is catastrophic because Blue Origin only has one New Glenn pad and it was damaged in the

industryars-technica
29 May 2026
Agents

Hermes Agent now has Tool Search, so your agent only loads what it needs

DGX agent

Hermes Agent has been updated with a Tool Search feature that enables agents to dynamically identify and load only the tools necessary for a given task, rather than loading all available tools upfront

agentsnous-research--x
29 May 2026
Agents

Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction

DGX agent

arXiv:2605.29960v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly leverage long term memory to support persistent and autonomous task execution. However, this capability

agentsarxiv-cs-ai
29 May 2026
Research

HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering

DGX agent

arXiv:2605.29606v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) for document-based Open-domain Question Answering (ODQA) on large-scale industrial corpora faces two critical bottl

researcharxiv-cs-ai
29 May 2026
Model Releases

Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning

DGX agent

arXiv:2605.29782v1 Announce Type: cross Abstract: Reinforcement learning (RL) refines large language models (LLMs) by directly optimizing model behavior through reward signals. While accurate state va

model-releasesarxiv-cs-ai
29 May 2026
Research

HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis

DGX agent

arXiv:2508.10566v3 Announce Type: replace Abstract: Audio-driven talking head generation faces a fundamental trade-off between personalization and generalization, limiting its practical application. I

researcharxiv-cs-cv
29 May 2026
Research

HoliTok:A Coutinuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding

DGX agent

arXiv:2605.29948v1 Announce Type: cross Abstract: Unified speech foundation models require a holistic tokenization space that is both learnable by language models and decodable into high-quality wavef

researcharxiv-cs-ai
29 May 2026
Research

Honest Lying: Understanding Memory Confabulation in Reflexive Agents

DGX agent

arXiv:2605.29463v1 Announce Type: cross Abstract: Reflexion-style agents rely on self-generated reflections as memory, implicitly assuming that agents can accurately diagnose their own failures.We sho

researcharxiv-cs-ai
29 May 2026
← Previous
1…994995996997998…1898
Next →