AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,013 results
Safety

Designing for Ethical AI: HCI Feature Considerations to Improve Fairness and User Experience in AutoML use for Human Resources

DGX agent

arXiv:2608.07477v1 Announce Type: cross Abstract: This thesis examines the fairness of Automated Machine Learning (AutoML) tools in human resource hiring systems through the combined lenses of regulat

safetyarxiv-cs-ai
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Evidence-Calibrated Runtime Reconstruction for Agent Skills Across Heterogeneous Coding Agents

DGX agent

arXiv:2608.08793v1 Announce Type: new Abstract: Agent Skills package reusable instructions and assets for tool-using language-model agents. Progressive loading creates failure boundaries poorly repres

local-aiarxiv-cs-cl
11 Aug 2026
Local Ai

Explainable Machine Learning in Healthcare: Methods, Interpretation, and Applications for Clinical Research

DGX agent

arXiv:2608.07522v1 Announce Type: cross Abstract: We present a structured review of commonly used Explainable machine learning (XML) methodologies, including global and local interpretability tools su

local-aiarxiv-cs-lg
11 Aug 2026
Model Releases

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

DGX agent

arXiv:2608.09316v1 Announce Type: new Abstract: Offline context optimization improves an agent by revising its instructions and examples while keeping the model frozen. This approach learns from rollo

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

OBLIVION: Workflow-Level Operational Skill Unlearning for Deployed Agents

DGX agent

arXiv:2608.08264v1 Announce Type: new Abstract: Large language model agents are becoming operational interfaces to files, memories, registries, and external tools. This deployment shift creates a new

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

OmnilingualGAIA2: Evaluating the Multilingual Gap in Frontier AI Agents

DGX agent

arXiv:2608.08775v1 Announce Type: new Abstract: Agentic benchmarks aim to measure how well AI agents plan, search, execute, and recover within realistic multi-tool environments, but they are almost ex

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

DGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

SAKE: Structured Agentic Knowledge Extrapolation for Complex LLM Reasoning via Reinforcement Learning

DGX agent

arXiv:2505.15062v5 Announce Type: replace-cross Abstract: Knowledge extrapolation is the process of inferring novel information by combining and extending existing knowledge that is explicitly availab

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

SHE: Trajectory-driven Safety Harness Evolution for LLM Agents

DGX agent

arXiv:2608.09885v1 Announce Type: new Abstract: The safety of large language model (LLM) agents depends not only on model weights but also on the agent harness that manages context, memory, tools, per

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TRACE: TRajectory Attribution for Automated Context Engineering

DGX agent

arXiv:2608.09153v1 Announce Type: new Abstract: Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory

DGX agent

arXiv:2608.07169v1 Announce Type: new Abstract: Memory systems have shown promise for improving agent performance, but their potential remains largely unexplored for small language models, which strug

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

DGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Google named a Leader in The Forrester Wave™: AI Platforms, Q3 2026

DGX agent

At Google Cloud, we help organizations of all sizes build and operationalize complex agentic workflows with total confidence. By combining world-class AI research with an open, fully integrated AI pla

model-releasesgoogle-cloud-ai
10 Aug 2026
Model Releases

NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs

DGX agent

arXiv:2608.07167v1 Announce Type: new Abstract: Giving an AI agent the ability to send emails, query databases, or execute commands is useful--until the agent is tricked into doing something it should

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Quantization Damage Is Multiplicative, Not Additive

DGX agent

arXiv:2608.06564v1 Announce Type: cross Abstract: Quantization is how large language models are actually deployed, and below four bits it is known to hurt. What nobody can say is which of the model's

model-releasesarxiv-cs-cl
10 Aug 2026
Tools

Quoting OpenClaw

DGX agent

The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3 a

toolssimon-willison
10 Aug 2026
Agents

Strategy-first synthesis planning for complex natural products

DGX agent

arXiv:2608.07454v1 Announce Type: cross Abstract: The total synthesis of a complex molecule is among the most demanding intellectual and experimental feats in chemistry: a chemist must plan many steps

agentsarxiv-cs-ai
10 Aug 2026
Agents

The Perils of Agency: How Developers Perceive, Prioritize, and Address Risks in Agentic AI Products

DGX agent

arXiv:2606.15485v2 Announce Type: replace-cross Abstract: Agentic AI systems act autonomously, use tools, adapt to context, and operate in complex real-world environments. However, these same characte

agentsarxiv-cs-ai
10 Aug 2026
Tools

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Em…

DGX agent

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Embed, retrieve, rerank, generate: your full retrieval pipelin

toolsfireworks-ai--x
10 Aug 2026
Model Releases

Doom Loop: Anyone Else Having DeepSeek v4 Flash 0731 Issues on ollama cloud?

DGX agent

Am I the only one having issues with DeepSeek V4 Flash? It gets stuck in a loop, as if it can't call the tools, and keeps repeating the same things endlessly without moving forward. Is it a poorly wri

model-releasesr-ollama
9 Aug 2026
Tools

Don't miss the bit where OpenAI first found out they were responsible for the Hugging Face attack when they reached out to HF to get one of …

DGX agent

Don't miss the bit where OpenAI first found out they were responsible for the Hugging Face attack when they reached out to HF to get one of their credentials revoked and HF told them it had already be

toolssimon-willison--x
8 Aug 2026
Tools

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…

DGX agent

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list

toolssimon-willison--x
8 Aug 2026
Agents

ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study

DGX agent

arXiv:2608.05201v1 Announce Type: cross Abstract: Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lac

agentsarxiv-cs-ai
7 Aug 2026
Tools

Autoscaling peaky LLM inference workloads is completely different than autoscaling something like a web service. I wrote a deepdive covering…

DGX agent

Zain (@zainhas) published a detailed article on August 7, 2026 explaining that autoscaling for highly peaky large‑language‑model (LLM) inference is fundamentally different from autoscaling conventiona

toolstogether-ai--x
7 Aug 2026
Safety

EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning

DGX agent

arXiv:2608.06197v1 Announce Type: new Abstract: Training large language model agents for long-horizon tool use typically relies on interactions with real or synthesized executable environments, whose

safetyarxiv-cs-ai
7 Aug 2026
Safety

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

DGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

safetyarxiv-cs-cl
7 Aug 2026
Safety

Negotiating Risk Boundaries in AI for Policing Through Mixed-Stakeholder Deliberation

DGX agent

arXiv:2608.05418v1 Announce Type: new Abstract: AI tools are being increasingly adopted in policing in the UK and worldwide. Racial bias is a known and well-documented risk, yet representatives of aff

safetyarxiv-cs-ai
7 Aug 2026
Tools

Thanks to the video from the Black Hat security conference of OpenAI's presentation about 'The Hugging Face Incident' we now have a detailed…

DGX agent

Thanks to the video from the Black Hat security conference of OpenAI's presentation about 'The Hugging Face Incident' we now have a detailed timeline of what happened from OpenAI's perspective - I wro

toolssimon-willison--x
7 Aug 2026
Model Releases

Advancing brain tumor research with privacy-first AI

DGX agent

The intersection of medicine and AI has led to remarkable innovations. However, developers now face the thorny challenge of building robust medical AI tools that have been tested and evaluated on dive

model-releasesgoogle-cloud-ai
6 Aug 2026
Local Ai

Architectural Implications of Agentic AI Workflows

DGX agent

arXiv:2608.04458v1 Announce Type: new Abstract: Agentic AI is emerging in datacenters, but its architectural implications remain unexplored. We organize agentic workflows in a taxonomy and present its

local-aiarxiv-cs-ai
6 Aug 2026
Safety

Breadcrumbing Search Agents

DGX agent

arXiv:2608.04565v1 Announce Type: cross Abstract: LLM-based search agents are widely used for information-seeking tasks, but their reliance on external tool returns introduces a critical security risk

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

DGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

model-releasesarxiv-cs-ai
6 Aug 2026
Tools

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clip…

DGX agent

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram

toolstogether-ai--x
6 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming

DGX agent

arXiv:2608.04018v1 Announce Type: cross Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to per

safetyarxiv-cs-ai
6 Aug 2026
Safety

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school…

DGX agent

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school me on the definition 🤦‍♂️ coding harness and tools calls ar

safetygary-marcus--x
6 Aug 2026
Tools

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all …

DGX agent

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine

toolssimon-willison--x
6 Aug 2026
Safety

SafeCommit: Certifying When Memory-Grounded Agents May Safely Act

DGX agent

arXiv:2608.04289v1 Announce Type: new Abstract: Long-horizon agents increasingly use persistent memory and tools to take actions with external side effects. A central failure mode is premature commitm

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Your agentic summer: No-cost lessons from Google experts to build and scale agents

DGX agent

I’ve talked to developers, IT leaders, and builders who all ask the same question: How do we actually get agents into production? The answer isn't theoretical — it's hands-on. Whether it’s designing a

model-releasesgoogle-cloud-ai
6 Aug 2026
Tools

99.9% uptime changes what your inference architecture has to survive. At Together AI, it means multi-data-center deployment, live traffic ac…

DGX agent

99.9% uptime changes what your inference architecture has to survive. At Together AI, it means multi-data-center deployment, live traffic across both facilities, and enough capacity to absorb a full d

toolstogether-ai--x
5 Aug 2026
Safety

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

DGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLMs Test Terminal User Interfaces?

DGX agent

arXiv:2608.03743v1 Announce Type: cross Abstract: Terminal User Interfaces (TUIs) combine the stateful, screen-oriented behaviour of GUIs with terminal deployment and are now common in developer tools

model-releasesarxiv-cs-ai
5 Aug 2026
Tools

ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's a…

DGX agent

ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's also how a billion users will soon use ChatGPT by default. I

toolsswyx--x
5 Aug 2026
Tools

if you have been following his excellent work, @shloked has been breaking down every frontier labs' harness engineering for the last few mon…

DGX agent

if you have been following his excellent work, @shloked has been breaking down every frontier labs' harness engineering for the last few months. excited to publish his deepest dive into ChatGPT yet as

toolsswyx--x
5 Aug 2026
Model Releases

Introducing Muse Code and Muse Spark 1.2

DGX agent

Introducing Muse Code and Muse Spark 1.2 Yet more evidence that the most important characteristic of any model these days is long-sequence agentic tool calling. Meta shipped their own coding agent as

model-releasessimon-willison
5 Aug 2026
Tools

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, …

DGX agent

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, MMMU Pro Vision, and DeepSWE. Open models like Kimi K3 have

toolstogether-ai--x
5 Aug 2026
Model Releases

Qwen Developers' responses from their recent Twitter/X AMA

DGX agent

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

model-releasesr-localllama
5 Aug 2026
Model Releases

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

DGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…3738394041…209
Next →