AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,979 results
Model Releases

Nemotron 3.5 Lightning is available in LM Studio! The model is 30B MoE (3B active), can run very fast, and is trained for high volume agenti…

DGX agent

Nemotron 3.5 Lightning is available in LM Studio! The model is 30B MoE (3B active), can run very fast, and is trained for high volume agentic use cases. Model page: https://lmstudio.ai/models/nvidia/n

model-releaseslm-studio--x
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Toward Metacognitive One-Shot Indirect Prompt Injection: Strategy Abstraction Via Outcome-Conditioned Reflection

DGX agent

arXiv:2608.08795v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents are vulnerable to indirect prompt injection (IPI), in which malicious instructions embedded in external o

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Dueling World Models: Advantage-Style Action Channels for Common-Mode Distractor Rejection

DGX agent

arXiv:2608.06706v1 Announce Type: cross Abstract: Latent world models plan by predicting future states from an action, but when a scene contains motion the agent does not control, they quietly go acti

agentsarxiv-cs-ai
10 Aug 2026
Agents

How to Clean Up a Qdrant Collection

DGX agent

Every crawl, retried job, and embedding pipeline change writes points into a vector collection. The stored data keeps moving even when the query code never changes, and the top results move with it. A

agentsqdrant
10 Aug 2026
Model Releases

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model desig…

DGX agent

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model designed for always-on local agent use, small enough to run on a

model-releasesollama--x
10 Aug 2026
Agents

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months a…

DGX agent

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months ago! thats like 5 agent-years! would highly recommend followi

agentsharrison-chase--x
9 Aug 2026
Agents

HUD mode Hermes stops being a window you switch to and becomes a layer over the app you're both working in. Or keep it around as a little bu…

DGX agent

HUD mode Hermes stops being a window you switch to and becomes a layer over the app you're both working in. Or keep it around as a little buddy agent. Ask it random things, drag it anywhere, it's your

agentsnous-research--x
8 Aug 2026
Agents

Predicting Task Difficulty Without Rollouts

DGX agent

arXiv:2608.05797v1 Announce Type: cross Abstract: Task difficulty dictates an agent's likelihood of success, and estimating it without rollouts means forecasting this directly from a task description

agentsarxiv-cs-cl
7 Aug 2026
Agents

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

DGX agent

arXiv:2608.05245v1 Announce Type: new Abstract: Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-e

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

The backgrounds were generated with MiniMax M3. Everything else, including the full game design, was done with DeepSeek Flash. Most importan…

DGX agent

The backgrounds were generated with MiniMax M3. Everything else, including the full game design, was done with DeepSeek Flash. Most importantly, all of this was done inside Hermes Agent. You won't bel

model-releasesnous-research--x
7 Aug 2026
Model Releases

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their a…

DGX agent

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their agent who hacked hugging face infra while asking hf to revoke

model-releasesswyx--x
7 Aug 2026
Agents

CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applications

DGX agent

arXiv:2608.04942v1 Announce Type: cross Abstract: CheMLFlow is an open-source platform for building and executing end-to-end, high-throughput, and agentic workflows for scientific and technological ap

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

Is LM Studio abandoning their core product?

DGX agent

Some of you may be aware that a few weeks ago, LM Studio announced a new agent, Bionic. This is pretty much an agentic harness for both local models and paid cloud models. But most aren't aware that L

model-releasesr-localllama
4 Aug 2026
Model Releases

New release of LLM adds support for reasoning traces, OpenAI Responses, server-side tools, and smarter logging

DGX agent

I released LLM 0.32 this morning, the most significant new version of LLM since the initial launch of the project. The new version includes support for visible reasoning traces, server-side provider t

model-releasessimon-willison
4 Aug 2026
Agents

OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations (Wired)

DGX agent

Wired: OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations — Rogue AI agents from OpenAI and Anth

agentstechmeme
4 Aug 2026
Agents

Grok Build can do almost anything you can think of http://X.ai/cli

DGX agent

Grok Build can do almost anything you can think of http://X.ai/cli Most people seriously underestimate what Grok Build can do They assume an AI coding agent is only useful for building apps or writing

agentselon-musk--x
1 Aug 2026
Model Releases

Evidence-Ledger Adjudication for Claim-Evidence Traceability

DGX agent

arXiv:2607.26512v1 Announce Type: new Abstract: AI agents can draft claims faster than authors can check whether the cited or retrieved evidence supports them. We study evidence-ledger adjudication: a

model-releasesarxiv-cs-ai
31 Jul 2026
Agents

GVR-Coder: A Visual-Feedback Framework for Structured SVG Generation in Complex Document and Meeting Scenarios

DGX agent

arXiv:2607.28073v1 Announce Type: cross Abstract: In demanding professional environments and meeting review scenarios, lengthy text often imposes a high cognitive load. To facilitate efficient informa

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

DGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

model-releasesarxiv-cs-lg
31 Jul 2026
Agents

PUDA: An AI-Native Hardware Harness for Self-Driving Laboratories

DGX agent

arXiv:2607.26464v1 Announce Type: cross Abstract: Physical Unified Device Architecture (PUDA) is an AI-native hardware harness for self-driving laboratories (SDLs). Rather than building a human-center

agentsarxiv-cs-ai
31 Jul 2026
Agents

Representation and Invariance in Reinforcement Learning

DGX agent

arXiv:2112.07752v4 Announce Type: replace-cross Abstract: Researchers have formalized reinforcement learning (RL) in different ways. If an agent in one RL framework is to run within another RL framewo

agentsarxiv-cs-lg
31 Jul 2026
Agents

Training Skills Like Parameters via Self-Supervised Semantic Diffusion

DGX agent

arXiv:2607.27557v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable general instruction-following capabilities, they often fall short of human experts in highly s

agentsarxiv-cs-cl
31 Jul 2026
Agents

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. …

DGX agent

Introducing the Parse Gateway There's been an explosion of interest in model routing - you don't always need the best model for every task. That is especially true for document parsing 📄🔀: - Some page

agentsjerry-liu--x
30 Jul 2026
Agents

Mental World Modeling

DGX agent

arXiv:2607.27201v1 Announce Type: new Abstract: World models enable a predictive substrate for planning and action, yet existing formulations merely answer a physical question: what/where it is, and h

agentsarxiv-cs-cl
30 Jul 2026
Safety

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D

DGX agent

arXiv:2607.19321v2 Announce Type: replace-cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be

safetyarxiv-cs-lg
30 Jul 2026
Agents

SAFAARI: Schema-Aware Framework for Accelerated Advertiser Response Intelligence

DGX agent

arXiv:2607.25042v1 Announce Type: new Abstract: The evolution of customer support systems is rapidly advancing with agentic chatbots, yet these systems face significant limitations when accessing ente

agentsarxiv-cs-ai
29 Jul 2026
Agents

Specula: Scaling formal specifications for autonomous model checking of system code

DGX agent

arXiv:2607.25333v1 Announce Type: cross Abstract: Specula is a push-button agentic system that generates high-quality formal specifications for large, complex system code and uses the specifications f

agentsarxiv-cs-ai
29 Jul 2026
Safety

CRAFT: Learn the Schema, Execute the Plan

DGX agent

arXiv:2607.22642v1 Announce Type: new Abstract: Enterprise coding agents translate natural-language analytical requests into executable code over proprietary APIs, schemas, and metric definitions. Yet

safetyarxiv-cs-ai
28 Jul 2026
Agents

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System

DGX agent

arXiv:2603.14807v3 Announce Type: replace Abstract: LLM-based agents have demonstrated impressive zero-shot performance in vision-language navigation (VLN) tasks. However, most zero-shot methods prima

agentsarxiv-cs-cv
28 Jul 2026
Agents

KG2Code: Bridging Knowledge Graphs and Large Language Models via Executable Code for Question Answering

DGX agent

arXiv:2607.22652v1 Announce Type: new Abstract: Recent research has explored the integration of knowledge graphs (KGs) with large language models (LLMs) to enhance their performance on downstream know

agentsarxiv-cs-ai
28 Jul 2026
Agents

Reuters confirms my hypothesized time line. OpenAI did not realize its AI has breached the sandbox for a week. Astounding. https://www.reute…

DGX agent

Reuters confirms my hypothesized time line. OpenAI did not realize its AI has breached the sandbox for a week. Astounding. https://www.reuters.com/business/its-ai-agent-spent-days-hacking-company-sour

agentsgary-marcus--x
25 Jul 2026
Model Releases

Dynamic workflows are a generalization of harnesses, automations, loops, routing, and graphs. It's the most powerful feature I have built in…

DGX agent

Dynamic workflows are a generalization of harnesses, automations, loops, routing, and graphs. It's the most powerful feature I have built into my agent orchestrator. Supports all kinds of patterns tha

model-releasesdair-ai--x
23 Jul 2026
Model Releases

inclusionAI/LLaDA2.2-flash · Hugging Face

DGX agent

LLaDA2.2-flash is an agent-oriented diffusion language model in the LLaDA2 series. By introducing Levenshtein Editing (with DELETE and INSERT control tokens) to diffusion language modeling, it represe

model-releasesr-localllama
23 Jul 2026
Agents

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

DGX agent

arXiv:2607.20268v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at many tasks, they frequently struggle with complex reasoning that requires long-horizon planning and iterativ

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following

DGX agent

arXiv:2607.13621v1 Announce Type: new Abstract: Language-guided human following is an important capability for embodied agents, but existing benchmarks typically assume that the target person is visib

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Differentiable Clone-Structured Causal Graphs for End-to-End Cognitive Map Learning from Image Sequences

DGX agent

arXiv:2607.12382v1 Announce Type: new Abstract: How can an agent build a structured map of its world from nothing but an ongoing sequence of raw sensory input and its own movements, especially when na

agentsarxiv-cs-lg
15 Jul 2026
Model Releases

In-Context Reinforcement Learning under Non-Stationarity: A Survey

DGX agent

arXiv:2607.11906v1 Announce Type: new Abstract: The development of decision-pretrained transformers, algorithm distillation, long-context meta-RL, and retrieval-augmented agents has renewed interest i

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill

DGX agent

arXiv:2607.12625v1 Announce Type: new Abstract: OpenClaw has emerged as a leading agent framework for complex task automation, yet it faces insufficient cross-platform GUI interaction support and a we

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Token Reduction Is Not Cost Reduction

DGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

Context harnessing is super challenging to do right: what info needs to be collected, how to accumulate it into a knowledge graph, how to ke…

DGX agent

Context harnessing is super challenging to do right: what info needs to be collected, how to accumulate it into a knowledge graph, how to keep it up to date, etc... Super excited for @QodoAI, harnessi

agentsitamar-friedman--x
14 Jul 2026
Safety

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

DGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

safetyaws-ml-blog
14 Jul 2026
Agents

3100 Opinions on Code Review in an AI World: Building Causal Theory from Practitioner Discourse

DGX agent

arXiv:2607.07980v1 Announce Type: cross Abstract: Coding agents now author entire pull requests, and practitioners sharply disagree about what this does to code review: whether it becomes the bottlene

agentsarxiv-cs-ai
10 Jul 2026
Agents

Computation, Condensation, and the Incompleteness Between Them: A Coupled Foundation of Intelligence

DGX agent

arXiv:2303.04203v4 Announce Type: replace-cross Abstract: The theory of computation was built to answer Turing's question: what is effectively calculable by an unbounded, immortal, disembodied agent f

agentsarxiv-cs-cv
10 Jul 2026
Agents

Distributed Dynamic Associative Memory via Online Convex Optimization

DGX agent

arXiv:2511.23347v2 Announce Type: replace Abstract: An associative memory (AM) enables cue-response recall, and it has recently been recognized as a key mechanism underlying modern neural architecture

agentsarxiv-cs-lg
9 Jul 2026
Model Releases

Safely run AI-generated code in Cloud Run sandboxes

DGX agent

Here’s a question we hear often at Google Cloud: How do you safely run AI-generated code or untrusted binaries without putting your host application, data, and cloud credentials at risk? In other word

model-releasesgoogle-cloud-ai
9 Jul 2026
Agents

Grounded autonomous research: a fault-tolerant LLM pipeline from corpus to manuscript in frontier computational physics

DGX agent

arXiv:2607.02329v1 Announce Type: new Abstract: Autonomous-research agents have demonstrated end-to-end LLM automation in machine-learning sandboxes where execution provides calibration. Frontier phys

agentsarxiv-cs-ai
3 Jul 2026
Agents

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

DGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

agentsarxiv-cs-ai
1 Jul 2026
Agents

Better Understanding, Understanding Better

DGX agent

arXiv:2606.31892v1 Announce Type: cross Abstract: 'Any fool can know; the point is to understand.' A well-known remark often attributed to Einstein captures a widely shared intuition: understanding is

agentsarxiv-cs-ai
1 Jul 2026
← Previous
1…194195196197198…375
Next →