AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,082 results
Safety

Distributed Optimization with Streaming Data: A Temporal Weighting Perspective

DGX agent

arXiv:2608.09565v1 Announce Type: cross Abstract: Optimization theory is a widely used tool for intelligent decision-making. While classical optimization deals with fixed, time-invariant objective fun

safetyarxiv-cs-ai
11 Aug 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Efficient identification of critical regions via Flow Matching-based Monte Carlo initialization

DGX agent

arXiv:2508.15318v5 Announce Type: replace-cross Abstract: Markov chain Monte Carlo (MCMC) is a standard tool for studying many-body systems, but its practical cost can become substantial, especially w

researcharxiv-cs-lg
11 Aug 2026
Research

EvalConvoLearn: An Open-Source Framework for Evaluating Grounded Learner Simulations in Tutoring Conversations

DGX agent

arXiv:2608.07497v1 Announce Type: cross Abstract: Conversational learner simulations are valuable tools for testing learning theories, evaluating instructional materials and automated tutors, or power

researcharxiv-cs-cl
11 Aug 2026
Model Releases

Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment

DGX agent

arXiv:2608.08212v1 Announce Type: new Abstract: In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter answers to unrelated questions. Existing prompts

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

DGX agent

arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harne

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

DGX agent

arXiv:2608.07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by t

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

Idea Search: Guiding Tree Search with Ideas to Explore Diverse Scientific Methods

DGX agent

arXiv:2608.08958v1 Announce Type: cross Abstract: Tree Search-based test-time scaling of LLMs is a powerful tool for automated scientific coding. However, pure Tree Search sometimes struggles with sys

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

DGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Logarithmic-Free Moment and Generalization Bounds for Uniformly Stable Algorithms

DGX agent

arXiv:2608.09870v1 Announce Type: cross Abstract: Uniform stability is a classical tool for controlling the generalization error of a learning algorithm. Bousquet, Klochkov, and Zhivotovskiy (2020) sh

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

DGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

OpenLoopEvolve: A Verifiable Self-Evolution Framework for Loop Policies in Long-Horizon Complex Tasks

DGX agent

arXiv:2608.09380v1 Announce Type: new Abstract: Long-horizon complex tasks require agents to repeatedly observe states, formulate plans, invoke tools, verify results, and recover from failures in cont

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

DGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

model-releasesarxiv-cs-ai
11 Aug 2026
Research

RoboSeg: Online Part-Level Semantic Reconstruction for Robotic Manipulation via a Single Eye-in-Hand Camera

DGX agent

arXiv:2608.09778v1 Announce Type: new Abstract: Robotic manipulation requires perception systemsthat identify actionable parts such as handles, rims, triggers,and tool tips, not merely object categori

researcharxiv-cs-ro
11 Aug 2026
Safety

Safety Cost of Steering Vectors Is Separable and Reducible

DGX agent

arXiv:2608.08383v1 Announce Type: new Abstract: Steering vectors are a lightweight tool for controlling LLM behavior. However, emerging evidence shows that steering vectors can unintentionally comprom

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance

DGX agent

arXiv:2608.09253v1 Announce Type: new Abstract: LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable pr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sparse corruption in low-rank matrix inference: the PCA benchmark

DGX agent

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

This is part of working with the EU AI Act, other labs are adding similar watermarking. It’s hard to identify AI-generated text, and this gi…

DGX agent

This is part of working with the EU AI Act, other labs are adding similar watermarking. It’s hard to identify AI-generated text, and this gives people better tools to do that. We’ll also be a shipping

model-releasesthariq--x
11 Aug 2026
Tutorials

Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and Experimentation

DGX agent

arXiv:2408.04619v2 Announce Type: replace-cross Abstract: The Transformer architecture underpins modern large language models powering state-of-the-art text generation and AI applications. However, it

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations

DGX agent

arXiv:2608.07540v1 Announce Type: new Abstract: AI systems increasingly operate between flexible input representations and formal objects used by downstream tools. A key challenge is recognizing when

model-releasesarxiv-cs-ai
11 Aug 2026
Research

TSPORec: Token Selection via Preference Optimization for LLM-Based Sequential Recommendation

DGX agent

arXiv:2608.09605v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for improving recommendation systems. The effectiveness of LLMs arises from their ability

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Two-Layer Linear Auto-Regressive Models Estimate Latent States

DGX agent

arXiv:2606.12691v2 Announce Type: replace-cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models lear

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

DGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps

DGX agent

arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine judges provide roughly the effective inform

safetyarxiv-cs-ai
10 Aug 2026
Agents

CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows

DGX agent

arXiv:2608.06961v1 Announce Type: new Abstract: Early-stage molecular design is an iterative process, not just a task of generating molecules. Researchers turn broad goals into design strategies, refi

agentsarxiv-cs-ai
10 Aug 2026
Research

CAS2UML: A Handwritten Sketch-to-PlantUML Dataset for Class and Activity Diagrams

DGX agent

arXiv:2608.07036v1 Announce Type: cross Abstract: Automated UML generation from sketches and images is gaining renewed attention with the rise of large language models and multimodal AI. However, repr

researcharxiv-cs-cv
10 Aug 2026
Research

EpiFlow: A framework for improving the utility of wastewater signals for disease forecasting

DGX agent

arXiv:2608.06671v1 Announce Type: new Abstract: Wastewater-based surveillance is an effective tool for disease monitoring and can provide early warning of outbreaks. Although wastewater viral loads (W

researcharxiv-cs-lg
10 Aug 2026
Research

From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning

DGX agent

arXiv:2608.06894v1 Announce Type: new Abstract: Neural operators have become a central tool for solving partial differential equations (PDEs), with spectral operators offering efficient global mixing

researcharxiv-cs-ai
10 Aug 2026
Model Releases

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

DGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Introducing Muse Glimmer

DGX agent

Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to

model-releasessimon-willison
10 Aug 2026
Safety

KnifeHunter: Structured Local Representation Learning for Fine-Grained Knife Image Retrieval in Law Enforcement

DGX agent

arXiv:2608.07057v1 Announce Type: new Abstract: Knife-enabled violence presents a major public safety challenge, and law enforcement agencies require scalable tools for catalogue-level knife identific

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

DGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

DGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

DGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called…

DGX agent

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called Kai. Their own words: a coding agent for non-engineers. You

agentsharrison-chase--x
9 Aug 2026
Model Releases

The Gemma team will host a special event on August 20

DGX agent

Tweet by u/hackerllama Could be copium, but I would love to see Gemma 4.1 there with unified audio input for all model sizes perhaps even up to 120B, much improved tool calling (even with the latest t

model-releasesr-localllama
9 Aug 2026
Model Releases

b10331

DGX agent

server: report the isolate working directory from get_info (#26773) server: report the isolate working directory from get_info Without an explicit cwd, get_info fell back to the server process working

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

DGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

DGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

model-releasesarxiv-cs-ai
7 Aug 2026
Research

CohortHijack: Robustness of Single Cell Annotation to Companion Cell Removal

DGX agent

arXiv:2608.05900v1 Announce Type: new Abstract: Many single-cell annotation tools refine an initial cell label using nearby cells or cluster-level voting. We study whether this refinement can be manip

researcharxiv-cs-lg
7 Aug 2026
Research

Confidence matters: Leveraging Multi-view Geometric Priors for GS-based Reconstruction

DGX agent

arXiv:2608.06117v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) has emerged as a widely-used tool for novel view synthesis, offering real-time rendering in a sparse representation. Howeve

researcharxiv-cs-cv
7 Aug 2026
Safety

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

DGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

safetyarxiv-cs-ai
7 Aug 2026
Local Ai

ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation

DGX agent

arXiv:2608.05893v1 Announce Type: new Abstract: Electrocardiography (ECG) is one of the most widely used non-invasive tools for diagnosing cardiovascular disease, but transforming multi-lead ECG recor

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

From Passive Mirrors to Active Agents: Holonic Digital Twins for Physical AI over Networks

DGX agent

arXiv:2608.06227v1 Announce Type: cross Abstract: Despite advances in artificial intelligence (AI) across multiple sectors, today's AI tools, including deep learning and generative AI, still fail when

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

DGX agent

arXiv:2608.06012v1 Announce Type: new Abstract: Search-agent rewards mix answer quality, citation grounding, tool cost, and anti-hacking terms; a high score therefore need not imply that cited evidenc

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore

DGX agent

In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through AgentCore Gateway, Agen

safetyaws-ml-blog
7 Aug 2026
Agents

Managing AI Coding Costs at Scale

DGX agent

**Managing AI Coding Costs at Scale** This article addresses the economic challenges of deploying AI coding tools across large teams or organizations. It explores strategies for tracking resource usag

agentsdatabricks
7 Aug 2026
← Previous
1…7576777879…211
Next →