AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,023 results
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

DGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Human Oversight and Overload: Two Hidden and Costly Burdens of AI-Assisted Software Engineering

DGX agent

arXiv:2606.05770v1 Announce Type: cross Abstract: AI is changing how software engineers work, but it often comes with hidden burdens and costs. In this paper, we characterize two such often-overlooked

researcharxiv-cs-ai
6 Jun 2026
Hardware

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — a…

DGX agent

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — and absolutely devastate all the data infrastructure investme

hardwaregary-marcus--x
6 Jun 2026
Research

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

DGX agent

arXiv:2606.06094v1 Announce Type: new Abstract: Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved di

researcharxiv-cs-ai
6 Jun 2026
Agents

Knowledge Activation: AI Skills as the Institutional Knowledge Primitive for Agentic Software Development

DGX agent

arXiv:2603.14805v2 Announce Type: replace Abstract: Enterprise software organizations accumulate critical institutional knowledge - architectural decisions, deployment procedures, compliance policies,

agentsarxiv-cs-ai
6 Jun 2026
Research

Metamorphic Testing with the Rashomon Set: Explanation Faithfulness in Machine Learning

DGX agent

arXiv:2606.06056v1 Announce Type: cross Abstract: Multiple machine learning models can achieve near-equivalent predictive performance on the same task, yet provide divergent feature-based explanations

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

The End of Software Engineering: How AI Agents Are Fundamentally Restructuring the Software Paradigm

DGX agent

arXiv:2606.05608v1 Announce Type: cross Abstract: For over half a century, software engineering has operated on a foundational premise: human engineers decompose problems, encode decision logic into s

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

The good news: agentic is leading to lots of new apps! The bad news: ain’t nobody adopting them. Slop FTL [for the loss]

DGX agent

Gary Marcus discusses a paradox in the agentic AI market: while developers are creating numerous new applications powered by agentic AI systems, these applications are failing to achieve meaningful us

agentsgary-marcus--x
6 Jun 2026
Research

This is really stupid, and it’s not getting enough attention. The Trump administration is pulling a working $368 million ocean monitoring sy…

DGX agent

This is really stupid, and it’s not getting enough attention. The Trump administration is pulling a working $368 million ocean monitoring system out of the water, equipment taxpayers already bought, b

researchyann-lecun--x
6 Jun 2026
Applications

Uncertainty Aware Functional Behavior Prediction and Material Fatigue Assessment for Circular Factory

DGX agent

arXiv:2606.05334v1 Announce Type: new Abstract: Returned products in circular factories re-enter production with heterogeneous degradation states, usage histories, and remaining capability. Reuse cann

applicationsarxiv-cs-ai
6 Jun 2026
Research

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing

DGX agent

arXiv:2606.05330v1 Announce Type: new Abstract: Large language models can shift human beliefs across high-stakes domains, but most persuasion studies rely on pre/post belief change. These endpoint mea

researcharxiv-cs-cl
5 Jun 2026
Research

A New Quaternion-Joint Cable-Driven Redundant Manipulator Configuration and its Control Through FABRIK and Residual Reinforcement Learning

DGX agent

arXiv:2606.05236v1 Announce Type: new Abstract: Robotic arms capable of traversing arbitrary spatial paths, especially in highly obstructed workspaces, are highly desired across several industries. Qu

researcharxiv-cs-ro
5 Jun 2026
Model Releases

Coding with 'Enemy': Can Human Developers Detect AI Agent Sabotage?

DGX agent

arXiv:2606.05647v1 Announce Type: cross Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to cod

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

CollabSim: A CSCW-Grounded Methodology for Investigating Collaborative Competence of LLM Agents through Controlled Multi-Agent Experiments

DGX agent

arXiv:2606.06399v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models have shown growing promise, with their effectiveness resting on agents' ability to coordinate t

agentsarxiv-cs-cl
5 Jun 2026
Research

CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning

DGX agent

arXiv:2509.04027v3 Announce Type: replace-cross Abstract: Test-time scaling, primarily manifested through multi-step Chain-of-Thought (CoT) reasoning via Reinforcement Learning (RL), has emerged as a

researcharxiv-cs-cl
5 Jun 2026
Agents

Emergent Language as an Approach to Conscious AI

DGX agent

arXiv:2606.06380v1 Announce Type: new Abstract: The question of whether artificial systems can be conscious remains open, in part because existing approaches either evaluate systems against theory-der

agentsarxiv-cs-cl
5 Jun 2026
Agents

Executable Schema Contracts: From Automatic Ingestion to Multi-Source Retrieval

DGX agent

arXiv:2606.05415v1 Announce Type: new Abstract: Real-world data spans tables, documents, and semi-structured files with implicit semantics. Querying this data requires integrating evidence across inco

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

Gemma 4 Quantization-Aware Training (QAT) weights are now available on Ollama! They reduce memory requirements while maintaining model quali…

DGX agent

Gemma 4 Quantization-Aware Training (QAT) weights are now available on Ollama! They reduce memory requirements while maintaining model quality. E2B: ollama run gemma4:e2b-it-qat E4B: ollama run gemma4

model-releasesollama--x
5 Jun 2026
Agents

Harnessing Generalist Agents for Contextualized Time Series

DGX agent

arXiv:2606.05404v1 Announce Type: cross Abstract: Time series are often embedded in rich contexts that are essential for holistic modeling. Moreover, real-world practitioners often require end-to-end

agentsarxiv-cs-cl
5 Jun 2026
Agents

here's an activegraph based deep research agent that gives you full graph/trace of claims, sources, agent activity...

DGX agent

This post describes an AI research agent built on ActiveGraph that provides complete visibility into its reasoning process through detailed graphs and traces of claims, sources, and internal agent act

agentsyohei-nakajima--x
5 Jun 2026
Safety

HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

DGX agent

arXiv:2602.16705v3 Announce Type: replace-cross Abstract: Visual loco-manipulation of arbitrary in-the-wild objects requires accurate end-effector (EE) control and a generalizable understanding of the

safetyarxiv-cs-cv
5 Jun 2026
Hardware

I absolutely agree that there really is this 10x opportunity for companies to be $40 trillion in market cap and beyond—perhaps Nvidia, Googl…

DGX agent

I absolutely agree that there really is this 10x opportunity for companies to be 40 trillion in market cap and beyond—perhaps Nvidia, Google, and beyond. Really fascinating to consider what that could

hardwareswyx--x
5 Jun 2026
Local Ai

Join us on a live interview with the CEO of ComfyUI!

DGX agent

Join us on a live interview with the CEO of ComfyUI! Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft, @Comf

local-aicomfyui--x
5 Jun 2026
Model Releases

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

DGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

DGX agent

arXiv:2606.05486v1 Announce Type: new Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, whi

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

Personal AI Agent for Camera Roll VQA

DGX agent

arXiv:2606.05275v1 Announce Type: new Abstract: We study the personal camera roll visual question answering setting. In this setting, a conversational AI assistant can access a user's personal camera

agentsarxiv-cs-cv
5 Jun 2026
Tutorials

Question from a beginner.

DGX agent

I don't have the ability to access or retrieve the content of specific Reddit posts from URLs. To write an accurate summary for your knowledge base, I would need you to either: 1. Share the text conte

tutorialsr-ollama
5 Jun 2026
Research

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

DGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Seeing Time: Benchmarking Chronological Reasoning and Shortcut Biases in Vision-Language Models

DGX agent

arXiv:2606.05702v1 Announce Type: cross Abstract: Recent advancements in Vision-Language Models (VLMs) have significantly enhanced their ability to interpret complex visual semantics, yet their capaci

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Text to Audiobook ?

DGX agent

A discussion from the StableDiffusion subreddit addressing whether text-to-speech or text-to-audiobook capabilities could be implemented with Stable Diffusion models. The post likely explores technica

local-air-stablediffusion
5 Jun 2026
Agents

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

DGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

DGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft…

DGX agent

Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft, @ComfyUI is one of the fastest growing platforms in creati

local-aicomfyui--x
5 Jun 2026
Model Releases

Using Large Language Models to Support High Volume Application Review for an Undergraduate Research Program

DGX agent

arXiv:2606.05564v1 Announce Type: new Abstract: Undergraduate research programs such as the Summer Undergraduate Research Fellowship (SURF) at Purdue University receive thousands of applications every

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that …

DGX agent

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that the search space itself changes, and an AI scientist must pe

model-releasesgary-marcus--x
5 Jun 2026
Safety

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

DGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

safetyarxiv-cs-cl
5 Jun 2026
Research

3D Temporal Analysis for Autism Spectrum Disorder Screening During Attention Tasks

DGX agent

arXiv:2606.04836v1 Announce Type: new Abstract: Accurate Autism Spectrum Disorder (ASD) screening for school-age children is crucial to identify cases that may have been missed earlier and to enable t

researcharxiv-cs-cv
4 Jun 2026
Model Releases

AIP: A Graph Representation for Learning and Governing Agent Skills

DGX agent

arXiv:2606.04781v1 Announce Type: new Abstract: Agent Skills today consist largely of free-form prose requiring the agent to read, interpret, and re-derive how to act in every session. This imposes tw

model-releasesarxiv-cs-ai
4 Jun 2026
Industry

Another wild customer story: a major ad agency was able to replicate a 300K–600K campaign for about $3K, delivering a 99%+ cost reduction …

DGX agent

An ad agency successfully replicated a campaign originally costing 300K-600K for approximately $3K, achieving a 99%+ cost reduction. This case study, shared by Cristobal Valenzuela, likely demonstrate

industrycristobal-valenzuela--x
4 Jun 2026
Model Releases

Asana launches AI-powered products to help organizations manage human and agent work

DGX agent

Asana Inc. announced today during the company’s Work Innovation Summit in London the launch of a new product suite that helps organizations manage work by humans and artificial intelligence agents usi

model-releasessiliconangle
4 Jun 2026
Research

Bagged Polynomial Regression and Neural Networks

DGX agent

arXiv:2205.08609v3 Announce Type: replace-cross Abstract: Climate and environmental applications increasingly rely on high-dimensional prediction from remote sensing and other scientific data. Neural

researcharxiv-cs-lg
4 Jun 2026
Research

Best Visual Reasoning Model in 2026 (Including APIs) [D]

DGX agent

Gemini 3.1 Pro and Gemini 3-Pro lead visual reasoning benchmarks , with GPT-5.2, Kimi-K2.5, and GPT-5.2-Pro following . A 2026 evaluation benchmarked 15 leading multimodal models on visual reasoning a

researchr-machinelearning
4 Jun 2026
Agents

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

DGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

agentsarxiv-cs-cl
4 Jun 2026
Research

CaloTrilogy: Toward a Breakthrough in One-Step, End-to-End, Physics-Guided Shower Generation for Modern Calorimeters

DGX agent

arXiv:2606.04165v1 Announce Type: cross Abstract: High-precision calorimeter simulation at current and future colliders imposes rapidly growing computational demands, motivating the development of mac

researcharxiv-cs-lg
4 Jun 2026
Safety

Covert Influence Between Language Models

DGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

safetyarxiv-cs-cl
4 Jun 2026
Agents

Deliberate Evolution: Agentic Reasoning for Sample-Efficient Symbolic Regression with LLMs

DGX agent

arXiv:2606.04360v1 Announce Type: new Abstract: Symbolic regression (SR) discovers compact mathematical expressions from data, yet recent LLM-based evolutionary methods remain sample-inefficient becau

agentsarxiv-cs-cl
4 Jun 2026
← Previous
1…159160161162163…209
Next →