AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,023 results
6 Jun 2026

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

SafetyDGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

Model ReleasesDGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

Human Oversight and Overload: Two Hidden and Costly Burdens of AI-Assisted Software Engineering

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.05770v1 Announce Type: cross Abstract: AI is changing how software engineers work, but it often comes with hidden burdens and costs. In this paper, we characterize two such often-overlooked

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — a…

HardwareDGX agent

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — and absolutely devastate all the data infrastructure investme

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

ResearchDGX agent

arXiv:2606.06094v1 Announce Type: new Abstract: Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved di

Knowledge Activation: AI Skills as the Institutional Knowledge Primitive for Agentic Software Development

AgentsDGX agent

arXiv:2603.14805v2 Announce Type: replace Abstract: Enterprise software organizations accumulate critical institutional knowledge - architectural decisions, deployment procedures, compliance policies,

Metamorphic Testing with the Rashomon Set: Explanation Faithfulness in Machine Learning

ResearchDGX agent

arXiv:2606.06056v1 Announce Type: cross Abstract: Multiple machine learning models can achieve near-equivalent predictive performance on the same task, yet provide divergent feature-based explanations

SentinelBench: A Benchmark for Long-Running Monitoring Agents

Model ReleasesDGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

The End of Software Engineering: How AI Agents Are Fundamentally Restructuring the Software Paradigm

Model ReleasesDGX agent

arXiv:2606.05608v1 Announce Type: cross Abstract: For over half a century, software engineering has operated on a foundational premise: human engineers decompose problems, encode decision logic into s

The good news: agentic is leading to lots of new apps! The bad news: ain’t nobody adopting them. Slop FTL [for the loss]

AgentsDGX agent

Gary Marcus discusses a paradox in the agentic AI market: while developers are creating numerous new applications powered by agentic AI systems, these applications are failing to achieve meaningful us

This is really stupid, and it’s not getting enough attention. The Trump administration is pulling a working $368 million ocean monitoring sy…

ResearchDGX agent

This is really stupid, and it’s not getting enough attention. The Trump administration is pulling a working $368 million ocean monitoring system out of the water, equipment taxpayers already bought, b

Uncertainty Aware Functional Behavior Prediction and Material Fatigue Assessment for Circular Factory

ApplicationsDGX agent

arXiv:2606.05334v1 Announce Type: new Abstract: Returned products in circular factories re-enter production with heterogeneous degradation states, usage histories, and remaining capability. Reuse cann

5 Jun 2026

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing

ResearchDGX agent

arXiv:2606.05330v1 Announce Type: new Abstract: Large language models can shift human beliefs across high-stakes domains, but most persuasion studies rely on pre/post belief change. These endpoint mea

A New Quaternion-Joint Cable-Driven Redundant Manipulator Configuration and its Control Through FABRIK and Residual Reinforcement Learning

ResearchDGX agent

arXiv:2606.05236v1 Announce Type: new Abstract: Robotic arms capable of traversing arbitrary spatial paths, especially in highly obstructed workspaces, are highly desired across several industries. Qu

Coding with 'Enemy': Can Human Developers Detect AI Agent Sabotage?

Model ReleasesDGX agent

arXiv:2606.05647v1 Announce Type: cross Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to cod

CollabSim: A CSCW-Grounded Methodology for Investigating Collaborative Competence of LLM Agents through Controlled Multi-Agent Experiments

AgentsDGX agent

arXiv:2606.06399v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models have shown growing promise, with their effectiveness resting on agents' ability to coordinate t

CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning

ResearchDGX agent

arXiv:2509.04027v3 Announce Type: replace-cross Abstract: Test-time scaling, primarily manifested through multi-step Chain-of-Thought (CoT) reasoning via Reinforcement Learning (RL), has emerged as a

Emergent Language as an Approach to Conscious AI

AgentsDGX agent

arXiv:2606.06380v1 Announce Type: new Abstract: The question of whether artificial systems can be conscious remains open, in part because existing approaches either evaluate systems against theory-der

Executable Schema Contracts: From Automatic Ingestion to Multi-Source Retrieval

AgentsDGX agent

arXiv:2606.05415v1 Announce Type: new Abstract: Real-world data spans tables, documents, and semi-structured files with implicit semantics. Querying this data requires integrating evidence across inco

Gemma 4 Quantization-Aware Training (QAT) weights are now available on Ollama! They reduce memory requirements while maintaining model quali…

Model ReleasesDGX agent

Gemma 4 Quantization-Aware Training (QAT) weights are now available on Ollama! They reduce memory requirements while maintaining model quality. E2B: ollama run gemma4:e2b-it-qat E4B: ollama run gemma4

Harnessing Generalist Agents for Contextualized Time Series

AgentsDGX agent

arXiv:2606.05404v1 Announce Type: cross Abstract: Time series are often embedded in rich contexts that are essential for holistic modeling. Moreover, real-world practitioners often require end-to-end

here's an activegraph based deep research agent that gives you full graph/trace of claims, sources, agent activity...

AgentsDGX agent

This post describes an AI research agent built on ActiveGraph that provides complete visibility into its reasoning process through detailed graphs and traces of claims, sources, and internal agent act

HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

SafetyDGX agent

arXiv:2602.16705v3 Announce Type: replace-cross Abstract: Visual loco-manipulation of arbitrary in-the-wild objects requires accurate end-effector (EE) control and a generalizable understanding of the

I absolutely agree that there really is this 10x opportunity for companies to be $40 trillion in market cap and beyond—perhaps Nvidia, Googl…

HardwareDGX agent

I absolutely agree that there really is this 10x opportunity for companies to be 40 trillion in market cap and beyond—perhaps Nvidia, Google, and beyond. Really fascinating to consider what that could

Join us on a live interview with the CEO of ComfyUI!

Local AiDGX agent

Join us on a live interview with the CEO of ComfyUI! Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft, @Comf

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

Model ReleasesDGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

Model ReleasesDGX agent

arXiv:2606.05486v1 Announce Type: new Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, whi

Personal AI Agent for Camera Roll VQA

AgentsDGX agent

arXiv:2606.05275v1 Announce Type: new Abstract: We study the personal camera roll visual question answering setting. In this setting, a conversational AI assistant can access a user's personal camera

Question from a beginner.

TutorialsDGX agent

I don't have the ability to access or retrieve the content of specific Reddit posts from URLs. To write an accurate summary for your knowledge base, I would need you to either: 1. Share the text conte

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

ResearchDGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

Seeing Time: Benchmarking Chronological Reasoning and Shortcut Biases in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.05702v1 Announce Type: cross Abstract: Recent advancements in Vision-Language Models (VLMs) have significantly enhanced their ability to interpret complex visual semantics, yet their capaci

Text to Audiobook ?

Local AiDGX agent

A discussion from the StableDiffusion subreddit addressing whether text-to-speech or text-to-audiobook capabilities could be implemented with Stable Diffusion models. The post likely explores technica

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

AgentsDGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

Model ReleasesDGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft…

Local AiDGX agent

Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft, @ComfyUI is one of the fastest growing platforms in creati

Using Large Language Models to Support High Volume Application Review for an Undergraduate Research Program

Model ReleasesDGX agent

arXiv:2606.05564v1 Announce Type: new Abstract: Undergraduate research programs such as the Summer Undergraduate Research Fellowship (SURF) at Purdue University receive thousands of applications every

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that …

Model ReleasesDGX agent

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that the search space itself changes, and an AI scientist must pe

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

SafetyDGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

4 Jun 2026

3D Temporal Analysis for Autism Spectrum Disorder Screening During Attention Tasks

ResearchDGX agent

arXiv:2606.04836v1 Announce Type: new Abstract: Accurate Autism Spectrum Disorder (ASD) screening for school-age children is crucial to identify cases that may have been missed earlier and to enable t

AIP: A Graph Representation for Learning and Governing Agent Skills

Model ReleasesDGX agent

arXiv:2606.04781v1 Announce Type: new Abstract: Agent Skills today consist largely of free-form prose requiring the agent to read, interpret, and re-derive how to act in every session. This imposes tw

Another wild customer story: a major ad agency was able to replicate a 300K–600K campaign for about $3K, delivering a 99%+ cost reduction …

IndustryDGX agent

An ad agency successfully replicated a campaign originally costing 300K-600K for approximately $3K, achieving a 99%+ cost reduction. This case study, shared by Cristobal Valenzuela, likely demonstrate

Asana launches AI-powered products to help organizations manage human and agent work

Model ReleasesDGX agent

Asana Inc. announced today during the company’s Work Innovation Summit in London the launch of a new product suite that helps organizations manage work by humans and artificial intelligence agents usi

Bagged Polynomial Regression and Neural Networks

ResearchDGX agent

arXiv:2205.08609v3 Announce Type: replace-cross Abstract: Climate and environmental applications increasingly rely on high-dimensional prediction from remote sensing and other scientific data. Neural

Best Visual Reasoning Model in 2026 (Including APIs) [D]

ResearchDGX agent

Gemini 3.1 Pro and Gemini 3-Pro lead visual reasoning benchmarks , with GPT-5.2, Kimi-K2.5, and GPT-5.2-Pro following . A 2026 evaluation benchmarked 15 leading multimodal models on visual reasoning a

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

CaloTrilogy: Toward a Breakthrough in One-Step, End-to-End, Physics-Guided Shower Generation for Modern Calorimeters

ResearchDGX agent

arXiv:2606.04165v1 Announce Type: cross Abstract: High-precision calorimeter simulation at current and future colliders imposes rapidly growing computational demands, motivating the development of mac

Covert Influence Between Language Models

SafetyDGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

Deliberate Evolution: Agentic Reasoning for Sample-Efficient Symbolic Regression with LLMs

AgentsDGX agent

arXiv:2606.04360v1 Announce Type: new Abstract: Symbolic regression (SR) discovers compact mathematical expressions from data, yet recent LLM-based evolutionary methods remain sample-inefficient becau

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation

Model ReleasesDGX agent

arXiv:2606.04046v1 Announce Type: cross Abstract: In embodied vision-language decision making tasks such as robotic manipulation and navigation, Vision-Language and Vision-Language-Action Models (VLMs

Efficient Brood Cell Detection in Layer Trap Nests for Bees and Wasps: Balancing Labeling Effort and Species Coverage

ResearchDGX agent

arXiv:2603.16652v2 Announce Type: replace Abstract: Monitoring cavity-nesting wild bees and wasps is vital for biodiversity research and conservation. Layer trap nests (LTNs) are emerging as a valuabl

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just g…

AgentsDGX agent

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just getting started. Come build with us: https://fireworks.ai/car

From idea to live store, in minutes. 🛍️ Tomorrow on the Friday Showcase: → Replit's new Shopify partnership, with @Davidizek . Tell Replit …

AgentsDGX agent

From idea to live store, in minutes. 🛍️ Tomorrow on the Friday Showcase: → Replit's new Shopify partnership, with @Davidizek . Tell Replit Agent what you want to sell and it builds your storefront, cr

Handwriting Extraction and Analysis of Signature Lists in Swiss Popular Initiatives

ResearchDGX agent

arXiv:2606.05018v1 Announce Type: new Abstract: Popular initiatives and referendums are central to Swiss democracy, yet the validation of handwritten signature lists remains a labor-intensive manual p

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

Model ReleasesDGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

Model ReleasesDGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

How dynamic workflows allow Claude Code to handle whole new types of tasks https://x.com/trq212/status/2061907337154367865

Model ReleasesDGX agent

Dynamic workflows in Claude Code enable the model to handle complex, multi-step tasks by allowing execution flows to adapt based on intermediate results rather than following fixed paths. This capabil

Interfaze: The Future of AI is built on Task-Specific Small Models

Model ReleasesDGX agent

arXiv:2602.04101v2 Announce Type: replace Abstract: We present Interfaze, a native hybrid model that fuses task-specific deep neural networks (CNNs and DNNs) directly into a transformer decoder throug

Learning symplectic model reduction based on a approximation theorem of symplectic embeddings

ResearchDGX agent

arXiv:2606.04623v1 Announce Type: new Abstract: High-dimensional Hamiltonian systems play a central role in many scientific and engineering disciplines, with dynamics evolving on symplectic manifolds.

Look closely. There’s more in the Showcase.

Model ReleasesDGX agent

OpenAI's developer account posted this message on X (formerly Twitter), likely encouraging developers to explore additional features, updates, or resources available in OpenAI's Showcase platform or d

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

AgentsDGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

← Previous
1…127128129130131…168
Next →