AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,005 results
Agents

VDGR-RAG: Vectors, Directories, Graphs, and Reflection Are All You Need for Unified Reasoning over Hierarchical Enterprise Knowledge

DGX agent

arXiv:2608.07994v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is essential for enterprise knowledge question answering (QA), particularly in domains with complex product documen

agentsarxiv-cs-ai
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

DGX agent

arXiv:2608.07925v1 Announce Type: new Abstract: EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How Google Cloud detects, contains, and protects against emerging threats

DGX agent

At Google Cloud, securing your data and business systems is our foundational commitment. We empower our customers with the tools, governance, and infrastructure needed to securely deploy workloads and

model-releasesgoogle-cloud-ai
7 Aug 2026
Model Releases

Who Checks the Citations? Benchmarking Legal Hallucination Detection

DGX agent

arXiv:2606.21155v2 Announce Type: replace Abstract: Attorneys, judges, and pro se filers increasingly use AI to draft legal documents, yet these tools frequently fabricate citations. Despite predictio

model-releasesarxiv-cs-cl
7 Aug 2026
Local Ai

AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection

DGX agent

arXiv:2601.19138v2 Announce Type: replace-cross Abstract: Secure code review is critical during pre-integration, where Atlassian developers rely on lightweight analysis tools, while deep security asse

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

DGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

DGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Prompt-Induced Waste in Large Reasoning Models: A Preregistered Two-Harness Benchmark of Coding Agents

DGX agent

arXiv:2608.01347v1 Announce Type: new Abstract: Large reasoning models used as coding agents incur costs from deliberation, tool calls, and repeated agent turns, yet the causal effect of prompt wordin

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorp…

DGX agent

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorporate symbol/manipulating entities like harnesses, tools, an

model-releasesgary-marcus--x
3 Aug 2026
Safety

Beyond Component Testing: Validating Agentic AI Systems

DGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

'Data center in a Box (on Wheels)' 256Gb VRAM/512Gb RAM AI Server 6-8 Month Operational Review, Stability Write Up, Benchmarks

DGX agent

I've been out of these forums for awhile but I figured I would provide a formal update on how this has been going now that it has some operation time under its belt, just to put the information out th

model-releasesr-localllama
3 Aug 2026
Agents

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

DGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

DGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

model-releasesgoogle-cloud-ai
3 Aug 2026
Safety

Unanticipated Effects of Generative AI on Expertise Pathways and Performance Perception in System Administration

DGX agent

arXiv:2607.28650v1 Announce Type: cross Abstract: While industry discourse often emphasizes immediate productivity gains and frames GenAI primarily as a tool for automation, the integration of GenAI i

safetyarxiv-cs-ai
3 Aug 2026
Agents

Unifying public and private data: Scale knowledge graphs with Data Commons on Spanner

DGX agent

To make informed decisions, businesses often need to connect their internal data with public reference data, to create a knowledge graph that connects real-world things and their relationships. Howeve

agentsgoogle-cloud-ai
3 Aug 2026
Agents

datasette-apps 0.2a0

DGX agent

Release: datasette-apps 0.2a0 Changes that improve Datasette Apps when created and edited using Datasette Agent: New app_debug() tool allowing agent to open an app (invisibly) and test it using JavaSc

agentssimon-willison
1 Aug 2026
Agents

CRMWeaver: Building Powerful Business Agent via Agentic RL and Shared Memories

DGX agent

arXiv:2510.25333v2 Announce Type: replace Abstract: Recent years have witnessed the rapid development of LLM-based agents, which shed light on using language agents to solve complex real-world problem

agentsarxiv-cs-cl
31 Jul 2026
Agents

datasette-agent 0.4a0

DGX agent

Release: datasette-agent 0.4a0 New await context.browser_task() mechanism allowing agent tools to run code directly in the user's browser. #33 This is an exciting new capability: it makes it easy for

agentssimon-willison
31 Jul 2026
Safety

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

DGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

safetyarxiv-cs-ai
31 Jul 2026
Agents

Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability

DGX agent

arXiv:2607.26637v1 Announce Type: new Abstract: Deployed LLM agents increasingly keep their long-term memory as a filesystem: a directory tree of markdown files that the agent itself reads, writes, an

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Context Assembly as the Controlled Variable: A Control-Theoretic View of Harness Policies for Frozen LLM Agents

DGX agent

arXiv:2607.25408v1 Announce Type: new Abstract: A growing body of 2026 work applies control theory to LLM agents: Lyapunov-certified stability for tool-mediated controllers (Prinos et al., 'Stable Age

safetyarxiv-cs-ai
29 Jul 2026
Agents

Are You Still the Agent I Authorized? Earned Authority under a Fixed Ceiling for Evolving Agents

DGX agent

arXiv:2607.23586v1 Announce Type: new Abstract: Long-lived AI agents increasingly evolve after deployment by retaining experience, acquiring skills and tools, revising workflows, delegating work, and

agentsarxiv-cs-ai
28 Jul 2026
Agents

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents

DGX agent

arXiv:2607.22711v1 Announce Type: cross Abstract: LLM coding agents operate by constructing trajectories that accumulate reasoning, tool calls, and results to enable multi-step decision-making. Howeve

agentsarxiv-cs-ai
28 Jul 2026
Safety

CRAFT: Learn the Schema, Execute the Plan

DGX agent

arXiv:2607.22642v1 Announce Type: new Abstract: Enterprise coding agents translate natural-language analytical requests into executable code over proprietary APIs, schemas, and metric definitions. Yet

safetyarxiv-cs-ai
28 Jul 2026
Research

From Vibe to Code -- and Back: Lexical Oscillation in the Formation of Design Intent with Generative AI

DGX agent

arXiv:2607.23126v1 Announce Type: cross Abstract: Generative AI design tools make natural-language prompts a starting point for design, placing new articulation demands on designers. Rather than treat

researcharxiv-cs-ai
28 Jul 2026
Agents

Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents

DGX agent

arXiv:2607.23670v1 Announce Type: cross Abstract: Plan Modes have become standard features in agentic programming tools, allowing users to gain transparency and control by working with the agent to de

agentsarxiv-cs-ai
28 Jul 2026
Tools

really grateful to @realbasilchatha for helping me curate a survey of the entire field of Forward Deployed Engineering in one track! From FD…

DGX agent

really grateful to @realbasilchatha for helping me curate a survey of the entire field of Forward Deployed Engineering in one track! From FDE 101 by @zkevinbai (Anthropic, Palantir, Rippling Founding

toolsswyx--x
28 Jul 2026
Model Releases

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

DGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

model-releasesarxiv-cs-ai
28 Jul 2026
Tools

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt th…

DGX agent

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt that workflow based on what we’ve learned serving more than 40

toolstogether-ai--x
28 Jul 2026
Local Ai

Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

DGX agent

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

local-air-ollama
27 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Agents

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in…

DGX agent

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in their context. Conversation history, big prompts, huge tool

agentsdair-ai--x
24 Jul 2026
Research

All of AI benchmarking at your fingertips

DGX agent

IBM, as part of a global effort, is developing tools to streamline AI benchmarking by making results easier to compare, validate, and reuse. The initiative is detailed in a blog post titled 'All of AI

researchibm-research
23 Jul 2026
Safety

The Ethics of Autonomous AI Agents for Offensive Security

DGX agent

arXiv:2607.20255v1 Announce Type: cross Abstract: LLM-driven autonomous agents are reshaping offensive security. Unlike traditional penetration-testing tooling -- deterministic, narrowly scoped, and o

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

DGX agent

Scientists today face challenges of extraordinary scale and complexity. From shaping and simulating the intricate dynamics of fusion plasma, to exploring the vast search space of new materials, to mak

model-releasesgoogle-cloud-ai
22 Jul 2026
Model Releases

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are…

DGX agent

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are released on schedule by July 27, it will become the #1 open

model-releaseskimi-moonshot--x
20 Jul 2026
Model Releases

Mermaid to Unicode box art (grok-mermaid)

DGX agent

Tool: Mermaid to Unicode box art (grok-mermaid) While exploring the codebase for the newly open-sourced Grok CLI coding agent I came across xai-grok-markdown/src/mermaid.rs, a 'self-contained terminal

model-releasessimon-willison
16 Jul 2026
Model Releases

How to Analyze and Govern Gemini Enterprise App Usage at Scale with BigQuery

DGX agent

Deploying the Gemini Enterprise app across an organization marks a transformative leap forward in workforce productivity, providing employees with an amazing, high-performance suite of agentic AI tool

model-releasesgoogle-cloud-ai
15 Jul 2026
Local Ai

Linus Torvalds tells people to stop attacking others for using AI

DGX agent

The full quote: I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as the top-level maintainer. Linux is not one of those anti-AI projec

local-air-localllama
15 Jul 2026
Model Releases

Lost in the Maze: Overcoming Context Limitations in Long-Horizon Agentic Search

DGX agent

arXiv:2510.18939v2 Announce Type: replace Abstract: Long-horizon agentic search requires iteratively exploring the web over long trajectories and synthesizing information across many sources, enabling

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Empowering India’s next generation of innovators with ATL Saathi

DGX agent

ATL Saathi is an initiative aimed at boosting scientific innovation and learning in India through AI-powered tools and programs. Launched in February 2026, the program focuses on accelerating research

model-releasesgoogle-deepmind
13 Jul 2026
Model Releases

CausalDS: Benchmarking Causal Reasoning in Data-Science Agents

DGX agent

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant b

model-releasesarxiv-cs-ai
10 Jul 2026
Tools

We've also simplified the project and repo pickers and made them more powerful. Use the pickers to launch agents in fewer clicks.

DGX agent

Cursor has streamlined its project and repository selection interface to improve user efficiency, allowing developers to launch AI agents with fewer clicks. The update combines simplified navigation w

toolscursor--x
10 Jul 2026
Tools

whoever does AEO for @resend needs to get a raise, all the leading frontier models keep trying to use resend for emails even when i have exi…

DGX agent

Swyx praises Resend's Account Executive Organization (AEO) team for effective positioning, noting that major frontier AI models frequently attempt to use Resend for email functionality even when alter

toolsswyx--x
10 Jul 2026
Tools

GPT 5.6 Sol, Luna, and Terra now available on AI Gateway

DGX agent

Vercel has made GPT 5.6 models (Sol, Luna, and Terra variants) available through its AI Gateway service, expanding the model options developers can access and deploy through the platform. This update

toolsvercel-blog
9 Jul 2026
Model Releases

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

DGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

model-releasesarxiv-cs-ai
8 Jul 2026
Tools

Chat SDK adds Photon support

DGX agent

Vercel's Chat SDK has been updated to include support for Photon, likely expanding real-time communication or messaging capabilities for developers building on the Vercel platform. This addition enabl

toolsvercel-blog
8 Jul 2026
← Previous
1…2930313233…209
Next →