AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
Human
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,315 results
29 May 2026

Hallucination Detection-Guided Preference Optimization for Clinical Summarization

Model ReleasesDGX agent

arXiv:2605.28910v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise on summarization tasks, but they often produce hallucinations, which are unsupported or incorrect stat

Hallucination Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Semantic Caching

Model ReleasesDGX agent

arXiv:2605.29055v1 Announce Type: new Abstract: Hallucination remains a major reliability barrier for production LLM systems, particularly in multi-agent pipelines where unsupported claims can propaga

HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.24562v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong question answering (QA) performance but can produce fluent answers unsupported by available evidence. Ex

Hands-on with Gemini Spark beta rolling out to AI Ultra subs: planned a birthday party from emails and calendar, but called a live-in boyfriend a 'close friend' (Reece Rogers/Wired)

Model ReleasesDGX agent

Reece Rogers / Wired: Hands-on with Gemini Spark beta rolling out to AI Ultra subs: planned a birthday party from emails and calendar, but called a live-in boyfriend a “close friend” — Google's new AI

Hardware’s back: AI supercharges server, PC and memory sales

IndustryDGX agent

Hardware firms are cleaning up bigtime as enterprises and cloud providers can’t get enough computing power for their artificial intelligence dreams. Dell Technology’s stock rocketed an incredible 31%

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

AgentsDGX agent

arXiv:2605.29354v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly participate in software development workflows by generating code, selecting dependencies, and producing package

Harmonizing Real-Time Constraints and Long-Horizon Reasoning: An Asynchronous Agentic Framework for Dynamic Scheduling

SafetyDGX agent

arXiv:2605.29262v1 Announce Type: new Abstract: The Dynamic Flexible Job Shop Scheduling Problem (DFJSP) necessitates a trade-off between instant reaction to stochastic disturbances and global optimiz

Harnessing non-adversarial robustness in large language models

SafetyDGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization

ResearchDGX agent

arXiv:2605.29843v1 Announce Type: cross Abstract: Post-training quantization (PTQ) is essential for deploying LLMs under memory and bandwidth constraints. However, extreme low-bit quantization remains

Has @AnthropicAI completely given up on making API usage reasonably-priced? Following the token-usage changes recently they announced variou…

TutorialsDGX agent

Has @AnthropicAI completely given up on making API usage reasonably-priced? Following the token-usage changes recently they announced various updates to *subscription* usage to make it more reasonable

Having Grok Build sub-agents to iterate several ideas for me on data loading, batching, inference, and writing results to files for dense da…

TutorialsDGX agent

Having Grok Build sub-agents to iterate several ideas for me on data loading, batching, inference, and writing results to files for dense datasets before I went to sleep. It gave me a nice summary of

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens

TutorialsDGX agent

arXiv:2512.15133v2 Announce Type: replace-cross Abstract: Proteins inherently possess a consistent sequence-structure duality. The abundance of protein sequence data, which can be readily represented

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, …

Model ReleasesDGX agent

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer sit down on came

HEART-Bench: Do LLM Agents Exhibit Human-like Psychology?

Model ReleasesDGX agent

arXiv:2605.30058v1 Announce Type: new Abstract: While LLM agents have demonstrated remarkable task-oriented abilities such as planning, reasoning, and action, few works have treated them as complete h

Here's an extended edit of the quote that includes a following fragment where Andrew Macdonald called the trade 'harder to justify' - full, …

Model ReleasesDGX agent

Here's an extended edit of the quote that includes a following fragment where Andrew Macdonald called the trade 'harder to justify' - full, unedited transcript is here: https://gist.github.com/simonw/

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end work…

ToolsDGX agent

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end workflow from UI to deployment → Real-time team collaboration wi

Here's why the failure of Blue Origin's New Glenn rocket is so catastrophic

IndustryDGX agent

Blue Origin's New Glenn rocket exploded during an engine-firing test at Cape Canaveral on May 28, 2026 , which is catastrophic because Blue Origin only has one New Glenn pad and it was damaged in the

Hermes Agent now has Tool Search, so your agent only loads what it needs

AgentsDGX agent

Hermes Agent has been updated with a Tool Search feature that enables agents to dynamically identify and load only the tools necessary for a given task, rather than loading all available tools upfront

Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction

AgentsDGX agent

arXiv:2605.29960v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly leverage long term memory to support persistent and autonomous task execution. However, this capability

HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering

ResearchDGX agent

arXiv:2605.29606v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) for document-based Open-domain Question Answering (ODQA) on large-scale industrial corpora faces two critical bottl

Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.29782v1 Announce Type: cross Abstract: Reinforcement learning (RL) refines large language models (LLMs) by directly optimizing model behavior through reward signals. While accurate state va

HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis

ResearchDGX agent

arXiv:2508.10566v3 Announce Type: replace Abstract: Audio-driven talking head generation faces a fundamental trade-off between personalization and generalization, limiting its practical application. I

HoliTok:A Coutinuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding

ResearchDGX agent

arXiv:2605.29948v1 Announce Type: cross Abstract: Unified speech foundation models require a holistic tokenization space that is both learnable by language models and decodable into high-quality wavef

Honest Lying: Understanding Memory Confabulation in Reflexive Agents

ResearchDGX agent

arXiv:2605.29463v1 Announce Type: cross Abstract: Reflexion-style agents rely on self-generated reflections as memory, implicitly assuming that agents can accurately diagnose their own failures.We sho

Honeyval: A Comprehensive Evaluation Framework for LLM-powered HTTP Honeypots

AgentsDGX agent

arXiv:2605.29963v1 Announce Type: cross Abstract: Honeypots are decoy systems mimicking real system components designed to defend against cyber attacks. Recently, LLMs increasingly serve as simulation

Horizon Activation Mapping for Neural Networks in Time Series Forecasting

ResearchDGX agent

arXiv:2601.02094v4 Announce Type: replace Abstract: Neural networks for time series forecasting have relied on error metrics and architecture-specific interpretability approaches for model selection t

Hot take on what comes next, after the sudden decline of tokenmaxxing: - OpenAI will struggle - with the decline of tokenmaxxing Anthropic w…

HardwareDGX agent

Hot take on what comes next, after the sudden decline of tokenmaxxing: - OpenAI will struggle - with the decline of tokenmaxxing Anthropic will struggle (aside from this quarter) to make a profit - Go

How Braintrust turns customer requests into code with Codex

Model ReleasesDGX agent

Braintrust leverages OpenAI's Codex model to automatically convert customer requests and natural language specifications into functional code, streamlining the software development process. This appli

How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

Model ReleasesDGX agent

arXiv:2605.29442v1 Announce Type: cross Abstract: AI coding agents increasingly act directly within software environments, yet existing analyses of their failures rely on benchmark trajectories that m

How Consistent Are LLM Agents? Measuring Behavioral Reproducibility in Multi-Step Tool-Calling Pipelines

AgentsDGX agent

arXiv:2605.28840v1 Announce Type: cross Abstract: Large language model (LLM) agents with tool-calling capabilities are increasingly deployed in production systems, yet a fundamental reliability questi

How do I install wan 2.2 into Forge Neo?

Local AiDGX agent

I don't have current information about this specific Reddit discussion or the installation process for WAN 2.2 into Forge Neo. This appears to be a technical support question from the StableDiffusion

How Far Ahead Do LLMs Plan? Uncovering the Latent Horizon in Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2602.02103v2 Announce Type: replace-cross Abstract: Chain-of-thought (CoT) reasoning has become a central mechanism for eliciting multi-step reasoning in Large Language Models (LLMs). Yet recent

How LoRA Remembers? A Parametric Memory Law for LLM Finetuning

Model ReleasesDGX agent

arXiv:2605.30260v1 Announce Type: cross Abstract: Large Language Models (LLMs) must continuously learn and update knowledge to remain effective in dynamic real-world environments. While Low-Rank Adapt

How lucky are you to have been born when and where you are? Had Opus 4.8 in Claude Code whip up a new visualization of all humans who ever l…

Model ReleasesDGX agent

How lucky are you to have been born when and where you are? Had Opus 4.8 in Claude Code whip up a new visualization of all humans who ever lived. In addition to being neat, it is an interesting test o

How massive bonuses for Samsung's employees in the memory division have sparked debate over how companies and governments should share profits from the AI boom (Bloomberg)

TutorialsDGX agent

Bloomberg: How massive bonuses for Samsung's employees in the memory division have sparked debate over how companies and governments should share profits from the AI boom — Payouts at Samsung have rai

How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions

ResearchDGX agent

arXiv:2605.29448v1 Announce Type: cross Abstract: Neural scaling laws appraise data through dataset size, while the Vendi Score uses quantum entropy to measure dataset value. We show both that common

How Reliable Are AI Attackers Against a Fixed Vulnerable Target? A 400-Run Empirical Study of LLM Penetration Testing Consistency

Model ReleasesDGX agent

arXiv:2605.30096v1 Announce Type: cross Abstract: Large language models (LLMs) can autonomously conduct multi-stage cyber attacks, but the consistency of their offensive behavior under repeated trials

How the Pope’s Magnifica Humanitas offers a template for individuals to meet the AI moment

ResearchDGX agent

Pope Leo XIV’s new encyclical on artificial intelligence includes a statement that warrants serious attention from technologists and policymakers: “Technology is never neutral.” Magnifica Humanitas (“

How to build a better agent harness with traces and evals

AgentsDGX agent

Agents are easy to prototype and hard to improve. A repeatable loop of traces, evals, failed-span inspection, and targeted harness changes makes agent behavior easier to debug and improve. The post Ho

How to Relieve Distribution Shifts in Semantic Segmentation for Off-Road Environments

AgentsDGX agent

arXiv:2605.29599v1 Announce Type: cross Abstract: Semantic segmentation is crucial for autonomous navigation in off-road environments, enabling precise classification of surroundings to identify trave

How Together AI built the world’s fastest speech-to-text stack

HardwareDGX agent

Together AI developed an optimized speech-to-text system focused on achieving the fastest processing speeds through technical innovations in their inference stack and model optimization. The approach

How's it going? Reinforcement learning in language models recruits a functional welfare axis

SafetyDGX agent

arXiv:2605.30232v1 Announce Type: cross Abstract: How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation

HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime

SafetyDGX agent

arXiv:2605.30201v1 Announce Type: cross Abstract: We investigate a narrow but common failure mode of GRPO-style reinforcement learning in the context of sparse verifiable rewards: early updates contai

HTAM: Hierarchical Transition-Attended Memory for Operator Optimization

Local AiDGX agent

arXiv:2605.29734v1 Announce Type: new Abstract: High-performance GPU kernels are essential for efficient LLM deployment, yet optimizing them remains expertise-intensive. Recent LLM-based code generati

https://x.com/huntlovell/status/2060399973506924612

AgentsDGX agent

I cannot provide an accurate summary because the URL appears to be invalid or the tweet is no longer accessible (the status ID seems implausible for the current date). To create a reliable knowledge b

Human-in-the-Loop Swarms: A Bionic Swarm Approach to Real-World Soil Mapping

AgentsDGX agent

arXiv:2605.29091v1 Announce Type: new Abstract: Swarm and field robotics face significant barriers to real-world validation due to the high cost and development time to deploy hardware. This paper int

I developed a brutalist GUI to interact with Ollama models and PI. It is open-sourced and available for arm64 and intel Macs. Repository in comments

Local AiDGX agent

A developer created an open-source GUI application with a brutalist design for interacting with Ollama models and the Personal Iris (PI) platform. The application is compatible with both ARM64 and Int

I GOT THE DOMAIN! I FINALLY GOT IT!!!!!!!!!!1 🥳🎉 Paint​.NET is now at https://paint.net! Well, it will be just as soon as I push all the b…

Model ReleasesDGX agent

I GOT THE DOMAIN! I FINALLY GOT IT!!!!!!!!!!1 🥳🎉 Paint​.NET is now at https://paint.net! Well, it will be just as soon as I push all the buttons to migrate content and set up redirects from getpaint​.

I had the experience of playing against Sony AI’s “Project Ace,” the most advanced high-speed autonomous table tennis robot system, which ha…

AgentsDGX agent

I had the experience of playing against Sony AI’s “Project Ace,” the most advanced high-speed autonomous table tennis robot system, which has defeated elite human athletes. I managed to win a point. F

i had to pack away my coding agents on monday cuz I knew I’d stay up too late if I played w activegraph during the week (yay, it’s Friday!) …

AgentsDGX agent

i had to pack away my coding agents on monday cuz I knew I’d stay up too late if I played w activegraph during the week (yay, it’s Friday!) gautham kept playing with it and just showed me a custom UI

i have strong reason to believe this is cope. we may find out soon…

SafetyDGX agent

i have strong reason to believe this is cope. we may find out soon… Im calling BS on this story. 1. That would be 100,000 employees spending 5k/mo each or 10,000 employees averaging 50k/mo each. No wa

I literally haven’t typed anything in weeks Since I started using Grok’s speech-to-text in Hermes Agent... I just talk It transcribes everyt…

AgentsDGX agent

I literally haven’t typed anything in weeks Since I started using Grok’s speech-to-text in Hermes Agent... I just talk It transcribes everything perfectly Every word. Every single time My thoughts flo

- I still get nonstop questions about use cases, even though I think thinking about things in the use case framework is the wrong way of thi…

Model ReleasesDGX agent

- I still get nonstop questions about use cases, even though I think thinking about things in the use case framework is the wrong way of thinking about it - Enterprises are starting to see actual busi

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI wou…

HardwareDGX agent

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI would become a commodity. now they think they have made some bi

I watched the LangChain keynote expecting product releases. Got something better — the clearest breakdown I've heard of why most agents neve…

ApplicationsDGX agent

I watched the LangChain keynote expecting product releases. Got something better — the clearest breakdown I've heard of why most agents never make it past demo. One thesis lands hard. Here's what sepa

“I’ll kill your whole f–king family. Your whole f–king family is dead. Your children, your wife, all dead.' A left-wing activist in Newark w…

IndustryDGX agent

“I’ll kill your whole f–king family. Your whole f–king family is dead. Your children, your wife, all dead.' A left-wing activist in Newark was caught on camera shouting those words at an unmasked ICE

iLoRA: Bayesian Low-Rank Adaptation with Latent Interaction Graphs for Microbiome Diagnosis

Model ReleasesDGX agent

arXiv:2605.30179v1 Announce Type: cross Abstract: Parameter-efficient adaptation has made LLMs practical for domain prediction, but standard LoRA still relies on a static low-rank update and does not

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that th…

Model ReleasesDGX agent

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that there would be a collapse of the AI bubble (which in my judgem

I'm suspicious of that that whole story about Uber blowing their AI budget and being disappointed in the results - I dug into it and it appe…

Model ReleasesDGX agent

Simon Willison expresses skepticism about reports claiming Uber overspent on AI and was disappointed with the results, indicating he investigated the story and found issues with its accuracy or framin

Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content

ResearchDGX agent

arXiv:2605.29245v1 Announce Type: cross Abstract: This paper presents a survey and taxonomy of LLM fingerprinting and watermarking for identity, ownership verification, provenance, and generated-conte

← Previous
1…783784785786787…1506
Next →