AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,005 results
11 Aug 2026

VDGR-RAG: Vectors, Directories, Graphs, and Reflection Are All You Need for Unified Reasoning over Hierarchical Enterprise Knowledge

AgentsDGX agent

arXiv:2608.07994v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is essential for enterprise knowledge question answering (QA), particularly in domains with complex product documen

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

Model ReleasesDGX agent

arXiv:2608.07925v1 Announce Type: new Abstract: EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong

7 Aug 2026

How Google Cloud detects, contains, and protects against emerging threats

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

At Google Cloud, securing your data and business systems is our foundational commitment. We empower our customers with the tools, governance, and infrastructure needed to securely deploy workloads and

Who Checks the Citations? Benchmarking Legal Hallucination Detection

Model ReleasesDGX agent

arXiv:2606.21155v2 Announce Type: replace Abstract: Attorneys, judges, and pro se filers increasingly use AI to draft legal documents, yet these tools frequently fabricate citations. Despite predictio

5 Aug 2026

AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection

Local AiDGX agent

arXiv:2601.19138v2 Announce Type: replace-cross Abstract: Secure code review is critical during pre-integration, where Atlassian developers rely on lightweight analysis tools, while deep security asse

4 Aug 2026

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

Model ReleasesDGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

SafetyDGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

Prompt-Induced Waste in Large Reasoning Models: A Preregistered Two-Harness Benchmark of Coding Agents

Model ReleasesDGX agent

arXiv:2608.01347v1 Announce Type: new Abstract: Large reasoning models used as coding agents incur costs from deliberation, tool calls, and repeated agent turns, yet the causal effect of prompt wordin

3 Aug 2026

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorp…

Model ReleasesDGX agent

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorporate symbol/manipulating entities like harnesses, tools, an

Beyond Component Testing: Validating Agentic AI Systems

SafetyDGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

'Data center in a Box (on Wheels)' 256Gb VRAM/512Gb RAM AI Server 6-8 Month Operational Review, Stability Write Up, Benchmarks

Model ReleasesDGX agent

I've been out of these forums for awhile but I figured I would provide a formal update on how this has been going now that it has some operation time under its belt, just to put the information out th

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

AgentsDGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

Model ReleasesDGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

Unanticipated Effects of Generative AI on Expertise Pathways and Performance Perception in System Administration

SafetyDGX agent

arXiv:2607.28650v1 Announce Type: cross Abstract: While industry discourse often emphasizes immediate productivity gains and frames GenAI primarily as a tool for automation, the integration of GenAI i

Unifying public and private data: Scale knowledge graphs with Data Commons on Spanner

AgentsDGX agent

To make informed decisions, businesses often need to connect their internal data with public reference data, to create a knowledge graph that connects real-world things and their relationships. Howeve

1 Aug 2026

datasette-apps 0.2a0

AgentsDGX agent

Release: datasette-apps 0.2a0 Changes that improve Datasette Apps when created and edited using Datasette Agent: New app_debug() tool allowing agent to open an app (invisibly) and test it using JavaSc

31 Jul 2026

CRMWeaver: Building Powerful Business Agent via Agentic RL and Shared Memories

AgentsDGX agent

arXiv:2510.25333v2 Announce Type: replace Abstract: Recent years have witnessed the rapid development of LLM-based agents, which shed light on using language agents to solve complex real-world problem

datasette-agent 0.4a0

AgentsDGX agent

Release: datasette-agent 0.4a0 New await context.browser_task() mechanism allowing agent tools to run code directly in the user's browser. #33 This is an exciting new capability: it makes it easy for

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

SafetyDGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

30 Jul 2026

Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability

AgentsDGX agent

arXiv:2607.26637v1 Announce Type: new Abstract: Deployed LLM agents increasingly keep their long-term memory as a filesystem: a directory tree of markdown files that the agent itself reads, writes, an

29 Jul 2026

Addressable Recall Compaction for Long Context-Window Control in AI Agents

Model ReleasesDGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

Context Assembly as the Controlled Variable: A Control-Theoretic View of Harness Policies for Frozen LLM Agents

SafetyDGX agent

arXiv:2607.25408v1 Announce Type: new Abstract: A growing body of 2026 work applies control theory to LLM agents: Lyapunov-certified stability for tool-mediated controllers (Prinos et al., 'Stable Age

28 Jul 2026

Are You Still the Agent I Authorized? Earned Authority under a Fixed Ceiling for Evolving Agents

AgentsDGX agent

arXiv:2607.23586v1 Announce Type: new Abstract: Long-lived AI agents increasingly evolve after deployment by retaining experience, acquiring skills and tools, revising workflows, delegating work, and

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents

AgentsDGX agent

arXiv:2607.22711v1 Announce Type: cross Abstract: LLM coding agents operate by constructing trajectories that accumulate reasoning, tool calls, and results to enable multi-step decision-making. Howeve

CRAFT: Learn the Schema, Execute the Plan

SafetyDGX agent

arXiv:2607.22642v1 Announce Type: new Abstract: Enterprise coding agents translate natural-language analytical requests into executable code over proprietary APIs, schemas, and metric definitions. Yet

From Vibe to Code -- and Back: Lexical Oscillation in the Formation of Design Intent with Generative AI

ResearchDGX agent

arXiv:2607.23126v1 Announce Type: cross Abstract: Generative AI design tools make natural-language prompts a starting point for design, placing new articulation demands on designers. Rather than treat

Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents

AgentsDGX agent

arXiv:2607.23670v1 Announce Type: cross Abstract: Plan Modes have become standard features in agentic programming tools, allowing users to gain transparency and control by working with the agent to de

really grateful to @realbasilchatha for helping me curate a survey of the entire field of Forward Deployed Engineering in one track! From FD…

ToolsDGX agent

really grateful to @realbasilchatha for helping me curate a survey of the entire field of Forward Deployed Engineering in one track! From FDE 101 by @zkevinbai (Anthropic, Palantir, Rippling Founding

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

Model ReleasesDGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt th…

ToolsDGX agent

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt that workflow based on what we’ve learned serving more than 40

27 Jul 2026

Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

Local AiDGX agent

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

Model ReleasesDGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

24 Jul 2026

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in…

AgentsDGX agent

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in their context. Conversation history, big prompts, huge tool

23 Jul 2026

All of AI benchmarking at your fingertips

ResearchDGX agent

IBM, as part of a global effort, is developing tools to streamline AI benchmarking by making results easier to compare, validate, and reuse. The initiative is detailed in a blog post titled 'All of AI

The Ethics of Autonomous AI Agents for Offensive Security

SafetyDGX agent

arXiv:2607.20255v1 Announce Type: cross Abstract: LLM-driven autonomous agents are reshaping offensive security. Unlike traditional penetration-testing tooling -- deterministic, narrowly scoped, and o

22 Jul 2026

Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

Model ReleasesDGX agent

Scientists today face challenges of extraordinary scale and complexity. From shaping and simulating the intricate dynamics of fusion plasma, to exploring the vast search space of new materials, to mak

20 Jul 2026

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are…

Model ReleasesDGX agent

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are released on schedule by July 27, it will become the #1 open

16 Jul 2026

Mermaid to Unicode box art (grok-mermaid)

Model ReleasesDGX agent

Tool: Mermaid to Unicode box art (grok-mermaid) While exploring the codebase for the newly open-sourced Grok CLI coding agent I came across xai-grok-markdown/src/mermaid.rs, a 'self-contained terminal

15 Jul 2026

How to Analyze and Govern Gemini Enterprise App Usage at Scale with BigQuery

Model ReleasesDGX agent

Deploying the Gemini Enterprise app across an organization marks a transformative leap forward in workforce productivity, providing employees with an amazing, high-performance suite of agentic AI tool

Linus Torvalds tells people to stop attacking others for using AI

Local AiDGX agent

The full quote: I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as the top-level maintainer. Linux is not one of those anti-AI projec

Lost in the Maze: Overcoming Context Limitations in Long-Horizon Agentic Search

Model ReleasesDGX agent

arXiv:2510.18939v2 Announce Type: replace Abstract: Long-horizon agentic search requires iteratively exploring the web over long trajectories and synthesizing information across many sources, enabling

13 Jul 2026

Empowering India’s next generation of innovators with ATL Saathi

Model ReleasesDGX agent

ATL Saathi is an initiative aimed at boosting scientific innovation and learning in India through AI-powered tools and programs. Launched in February 2026, the program focuses on accelerating research

10 Jul 2026

CausalDS: Benchmarking Causal Reasoning in Data-Science Agents

Model ReleasesDGX agent

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant b

We've also simplified the project and repo pickers and made them more powerful. Use the pickers to launch agents in fewer clicks.

ToolsDGX agent

Cursor has streamlined its project and repository selection interface to improve user efficiency, allowing developers to launch AI agents with fewer clicks. The update combines simplified navigation w

whoever does AEO for @resend needs to get a raise, all the leading frontier models keep trying to use resend for emails even when i have exi…

ToolsDGX agent

Swyx praises Resend's Account Executive Organization (AEO) team for effective positioning, noting that major frontier AI models frequently attempt to use Resend for email functionality even when alter

9 Jul 2026

GPT 5.6 Sol, Luna, and Terra now available on AI Gateway

ToolsDGX agent

Vercel has made GPT 5.6 models (Sol, Luna, and Terra variants) available through its AI Gateway service, expanding the model options developers can access and deploy through the platform. This update

8 Jul 2026

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

Model ReleasesDGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

Chat SDK adds Photon support

ToolsDGX agent

Vercel's Chat SDK has been updated to include support for Photon, likely expanding real-time communication or messaging capabilities for developers building on the Vercel platform. This addition enabl

Chat SDK now supports Vercel Connect

ToolsDGX agent

Vercel has announced support for Vercel Connect within its Chat SDK, enabling developers to integrate real-time chat functionality with Vercel's infrastructure and deployment services. This update lik

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

Rewriting Bun in Rust https://bun.com/blog/bun-in-rust

ToolsDGX agent

Bun's development team announced a significant architectural rewrite of their JavaScript runtime from Zig to Rust, aiming to improve performance, maintainability, and developer experience. This transi

Use any Chat SDK adapter with eve

ToolsDGX agent

Eve's Chat SDK now supports multiple adapter options, allowing developers to integrate various chat platforms and services without being locked into a single solution. This flexibility enables teams t

7 Jul 2026

DualView: Preventing Indirect Prompt Injection in Personal AI Agents

Model ReleasesDGX agent

arXiv:2607.03821v1 Announce Type: cross Abstract: Personal AI agents that run on the user's local machine, such as OpenClaw, automate daily tasks including web search, email, and file management. Thei

got the raw footage of this talk from @swyx - 59 gb going to see how Fable does editing it

ToolsDGX agent

Thariq shared raw footage from a talk by @swyx, totaling 59 GB in size, and plans to test Fable's video editing capabilities on the material. The post indicates an experiment with AI-assisted video ed

More granular observability for Vercel Sandbox

ToolsDGX agent

Vercel has enhanced observability capabilities for Vercel Sandbox, enabling developers to gain more detailed insights into their sandbox environments. This update likely provides granular monitoring a

Vercel acquires Better Auth to accelerate open source auth

ToolsDGX agent

Vercel has acquired Better Auth, an open source authentication library, to strengthen its authentication capabilities and support for developers building on its platform. The acquisition aims to accel

6 Jul 2026

Building the Document Context Layer for AI Agents AI Agents are the new knowledge workers, but the vast majority of knowledge work depends o…

AgentsDGX agent

Building the Document Context Layer for AI Agents AI Agents are the new knowledge workers, but the vast majority of knowledge work depends on unstructured documents. If you don’t build the proper tool

🤗 Kernels: Major Updates

ToolsDGX agent

Hugging Face announced major updates to their Kernels platform, which is their cloud-based computational environment for machine learning and data science projects. The revamp likely includes improvem

3 Jul 2026

Causal Explanations for Image Classifiers

ResearchDGX agent

arXiv:2411.08875v4 Announce Type: replace Abstract: Existing algorithms for explaining the output of image classifiers use different definitions of explanations and a variety of techniques to find the

I’ve found the most important part of working with Fable is discovering my own unknowns so I can prompt it better, heres how I do that.

ToolsDGX agent

This post discusses strategies for effectively prompting Fable by first identifying gaps in one's own knowledge and understanding. The author suggests that self-discovery of unknowns is a prerequisite

← Previous
1…2324252627…167
Next →