AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,088 results
Model Releases

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

DGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

model-releasesarxiv-cs-lg
4 Aug 2026
Agents
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Practical Online KV Cache Compaction for LLM Agents: An Empirical Study

DGX agent

arXiv:2608.00902v1 Announce Type: new Abstract: LLM agents accumulate long trajectories of reasoning steps, tool calls, and environment feedback, making the KV cache a major inference bottleneck. KV c

agentsarxiv-cs-cl
4 Aug 2026
Applications

PyDPF: A Python Package for Differentiable Particle Filtering

DGX agent

arXiv:2510.25693v3 Announce Type: replace-cross Abstract: State-space models (SSMs) are a widely used tool in time series analysis. In the complex systems that arise from real-world data, it is common

applicationsarxiv-cs-lg
4 Aug 2026
Model Releases

ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

DGX agent

arXiv:2608.02358v1 Announce Type: new Abstract: To operate robustly in open-world environments, autonomous agents should be able to infer the behavior of unfamiliar systems through interaction alone,

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Thermalizing Stochastic Programs

DGX agent

arXiv:2608.01615v1 Announce Type: cross Abstract: We present a set of tools for mapping general stochastic programs to thermodynamic hardware designed for energy-efficient stochastic sampling. Given a

researcharxiv-cs-lg
4 Aug 2026
Model Releases

this post from 3 years ago and commercial LLMs *still* can’t play chess anywhere near as well serious players (except by calling external to…

DGX agent

this post from 3 years ago and commercial LLMs *still* can’t play chess anywhere near as well serious players (except by calling external tools) When @GaryMarcus and others point out that GPT-4 is bad

model-releasesgary-marcus--x
4 Aug 2026
Model Releases

Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety

DGX agent

arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequenc

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

An analysis of machine learning approaches for enhancing decision-making in complex discrete choice tasks

DGX agent

arXiv:2607.28854v1 Announce Type: new Abstract: Discrete choice modeling is a common tool used for preference elicitation during policy-making, but this is typically done through parametric models. Ma

safetyarxiv-cs-lg
3 Aug 2026
Model Releases

CrowdStrike finds AI systems under direct attack as exploit windows shrink

DGX agent

Artificial intelligence has become a target for attackers rather than only a tool they use, according to CrowdStrike Holdings Inc.’s “2026 Threat Hunting Report,” released today. The annual report dra

model-releasessiliconangle
3 Aug 2026
Research

Fracture Risk Prediction in Adults Over 50 Years Old Using DXA and EHR: Comparison of Traditional and Machine Learning Models in Two Large Cohorts

DGX agent

arXiv:2607.28671v1 Announce Type: cross Abstract: Accurate fracture risk prediction is important for osteoporosis management, but commonly used clinical tools may not fully use information available i

researcharxiv-cs-lg
3 Aug 2026
Agents

From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale

DGX agent

arXiv:2607.29516v1 Announce Type: cross Abstract: AI coding agents are generating code at volumes that exceed the capacity of traditional peer review. At the same time, existing AI code review tools o

agentsarxiv-cs-ai
3 Aug 2026
Safety

future generations will have no idea what was real and what was not.

DGX agent

Gary Marcus warns that, with increasingly sophisticated generative‑AI tools, future generations may struggle to tell what is real from what is fabricated. Matt Stoller adds that generative AI will tra

safetygary-marcus--x
3 Aug 2026
Research

Improving scDiffusion with Sparsity-Biased Classifier-Free Guidance

DGX agent

arXiv:2607.29043v1 Announce Type: cross Abstract: Single-cell RNA sequencing (scRNA-seq) has become an essential tool in modern cellular biology, and generating accurate synthetic scRNA-seq data is be

researcharxiv-cs-ai
3 Aug 2026
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic to…

DGX agent

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic tools. 開発者の皆様、大変お待たせしました!Sakana Chatのモデルが遂にAPIとして公開です。ぜひお試しください

agentsdavid-ha--x
3 Aug 2026
Applications

The Checking Problem: What must be true before AI ships in a regulated firm

DGX agent

arXiv:2607.28666v1 Announce Type: new Abstract: Enterprise AI programmes stall at a rate that is widely quoted and poorly explained. This paper measures the mechanism. Six document-heavy workflows of

applicationsarxiv-cs-cl
3 Aug 2026
Local Ai

I built a self-hosted studio that turns one reference photo into a curated, captioned, trained and tested LoRA — one browser tab, open source, MIT

DGX agent

I shared this tool here a week ago and the feedback shaped a big new version, so here's the full tour of what it does today. Screenshots of every screen: github.com/perfectgf/lora-dataset-studio — plu

local-air-stablediffusion
2 Aug 2026
Model Releases

Real-world reality check on Qwen for autonomous coding agents

DGX agent

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

model-releasesr-localllama
2 Aug 2026
Local Ai

Try handling complex tasks to your local models with GraphARC, graph engineering yes !

DGX agent

🚀 We just built our first real-time implementation of Graph Engineering, inspired by our experience building graph tooling used by 4,000+ developers. 🔗 Repo: https://github.com/CodeGraphContext/grapha

local-air-localllama
2 Aug 2026
Model Releases

b10217

DGX agent

chat : enable tool call in thinking for DS4 (#26269) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFr

model-releasesllama-cpp-releases
1 Aug 2026
Agents

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on…

DGX agent

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on the Trajectory 1. Measure the state of the environment befo

agentsharrison-chase--x
1 Aug 2026
Agents

AI as Friction for Reflection Support in Ideation

DGX agent

arXiv:2607.26827v1 Announce Type: cross Abstract: Generative AI tools for creative work tend to be designed around the goal of removing friction, on the assumption that smoother iteration and faster o

agentsarxiv-cs-ai
31 Jul 2026
Industry

AI-native software development requires a new engineering model

DGX agent

Artificial intelligence has quickly become a standard part of modern software development. Coding assistants, code completion tools and AI-powered integrated development environments are now widely av

industrysiliconangle
31 Jul 2026
Research

An analysis of binary isotonic regression: degrees of freedom and implications for calibration

DGX agent

arXiv:2607.27301v1 Announce Type: cross Abstract: Isotonic regression is a canonical tool for estimating monotone functions and calibrating probabilistic predictors. We provide a fully sharp finite-sa

researcharxiv-cs-lg
31 Jul 2026
Agents

Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents

DGX agent

arXiv:2607.15715v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used for complex information-extraction tasks, yet it remains unclear whether agentic components

agentsarxiv-cs-ai
31 Jul 2026
Model Releases

Beyond Similarity: Grounded Agentic Extraction and Expert-Adjudicated Evaluation of Intertextuality in Classical Chinese Histories

DGX agent

arXiv:2607.27595v1 Announce Type: new Abstract: Computational approaches to intertextuality have advanced from string matching to neural retrieval, yet their outputs, similarity scores and parallel-pa

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments

DGX agent

arXiv:2607.28591v1 Announce Type: cross Abstract: Scaling coding agents requires a continuing supply of executable data for training, benchmarking, and continuous evaluation. Each task must couple a r

agentsarxiv-cs-cl
31 Jul 2026
Local Ai

Could we all crowdsource a dataset/model/finetune?

DGX agent

I know it’s been discussed to try to make our own model through crowdsourcing, but finetuning seems like it would be even easier. We could edit and proofread and write our own datasets at a large scal

local-air-localllama
31 Jul 2026
Research

Expanding Data-Agnostic Pivotal Instances Selection Models with Proximity Trees and Ensemble Learning

DGX agent

arXiv:2607.27522v1 Announce Type: new Abstract: As decision-making processes grow more complex, machine learning tools have become essential for tackling business and societal challenges. However, man

researcharxiv-cs-lg
31 Jul 2026
Research

FADEx: Feature Attribution and Distortion-based Explanation of Dimensionality Reduction

DGX agent

arXiv:2607.27463v1 Announce Type: new Abstract: Dimensionality Reduction (DR) is a fundamental tool for high-dimensional data exploration, reducing the complexity of latent spaces of machine learning

researcharxiv-cs-lg
31 Jul 2026
Research

How does downsampling affect needle electromyography signals? A generalisable workflow for understanding downsampling effects on high-frequency time series

DGX agent

arXiv:2601.10191v2 Announce Type: replace Abstract: Automated analysis of needle electromyography (nEMG) signals is emerging as a tool to support the detection of neuromuscular diseases (NMDs), yet th

researcharxiv-cs-ai
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

On a joint simultaneous learning of relevant feature subsets and subspaces in regression-like problems

DGX agent

arXiv:2607.28080v1 Announce Type: cross Abstract: We extend a recently introduced Entropy-Optimal Manifold Clustering (EOMC) to allow for a joint simultaneous identification of subsets and subspaces o

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Optimal Realistic Local AI for Most

DGX agent

So you’ve got a 3090 or maybe even a 5090? Or more likely a 4060 8GB Ti. You wanna try local AI, you don’t know what it can/can’t do. 1) Install the best model you can. If you have a 3090 or a 5090, t

model-releasesr-localllama
31 Jul 2026
Applications

Prompt Chaining in Practice: A Case Study in Automated Scholarly Report Generation

DGX agent

arXiv:2607.27210v1 Announce Type: new Abstract: The exponential growth of scholarly publications requires automated tools for effective information synthesis. However, simple, single-shot prompting me

applicationsarxiv-cs-cl
31 Jul 2026
Agents

SciDataSailor: Deep Scientific Data Exploring

DGX agent

arXiv:2607.28098v1 Announce Type: cross Abstract: Scientific datasets are commonly organized as hierarchical repositories containing heterogeneous and interdependent files, making their inspection, in

agentsarxiv-cs-cl
31 Jul 2026
Model Releases

smevals - a small eval suite for evaluating models, prompts, and harnesses

DGX agent

smevals - a small eval suite for evaluating models, prompts, and harnesses I've been working with Jesse Vincent's Prime Radiant applied AI research lab building out this evals framework to help answer

model-releasessimon-willison
31 Jul 2026
Research

Sparsity Induced Identifiability in Matrix Tri-Factorisation

DGX agent

arXiv:2607.27507v1 Announce Type: new Abstract: Matrix factorisation is a fundamental tool for exploiting low-dimensional structure in high-dimensional data, with applications such as data compression

researcharxiv-cs-lg
31 Jul 2026
Tutorials

The Social Cost of an AI Teammate: How an Artificial Teammate Reshapes Human-Human Communication in Small-Team Decision-Making

DGX agent

arXiv:2607.27179v1 Announce Type: cross Abstract: Conversational AI is increasingly positioned as a teammate rather than a tool, yet we know little about how its presence reshapes communication among

tutorialsarxiv-cs-ai
31 Jul 2026
Research

Theatre Chapbooks At Scale: A Statistical Comparative Analysis of Typography

DGX agent

arXiv:2607.27266v1 Announce Type: new Abstract: We propose a statistical methodology that quantifies the similarity of typefaces between printed historical books. This provides a tool that accelerates

researcharxiv-cs-cv
31 Jul 2026
Model Releases

What's your local AI coding setup on a MacBook Pro M4?

DGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

model-releasesr-ollama
31 Jul 2026
Safety

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

DGX agent

arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by

safetyarxiv-cs-cl
30 Jul 2026
Safety

Atomic Chat signed the Open Weights letter! We believe everyone should be able to run AI on their own device. When a model is open, thousand…

DGX agent

Atomic Chat signed the Open Weights letter! We believe everyone should be able to run AI on their own device. When a model is open, thousands of teams fine-tune it, quantize it and build new tools on

safetyai21-labs--x
30 Jul 2026
Agents

Do more with less: How GKE can reduce your cost per agent by 75%

DGX agent

In today’s agentic era, modern cloud applications are evolving from a set of passive tools to fleets of autonomous digital workers that reason, plan, and take action across a wide range of tasks. For

agentsgoogle-cloud-ai
30 Jul 2026
Model Releases

Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration

DGX agent

Gemini Robotics ER 2 helps robots reason, collaborate, and solve real-world tasks. It represents a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic app

model-releasesgoogle-deepmind
30 Jul 2026
Applications

See2Think: Do Multimodal Models Really Use Intermediate Visual States?

DGX agent

arXiv:2607.26769v1 Announce Type: new Abstract: Multimodal large language models increasingly use sketches, annotations, tools, and intermediate images during reasoning, but it remains unclear whether

applicationsarxiv-cs-cv
30 Jul 2026
Safety

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

DGX agent

arXiv:2607.25489v1 Announce Type: new Abstract: Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and

safetyarxiv-cs-cv
29 Jul 2026
Agents

Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response

DGX agent

arXiv:2607.25379v1 Announce Type: new Abstract: Cyber-capable AI agents combine language models with tools, memory, and execution en- vironments to perform multi-step offensive-security tasks. Existin

agentsarxiv-cs-ai
29 Jul 2026
← Previous
1…7778798081…211
Next →