AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,088 results
28 Apr 2026

Audio2Tool: Bridging Spoken Language Understanding and Function Calling

Model ReleasesDGX agent

arXiv:2604.22821v1 Announce Type: cross Abstract: Voice assistants increasingly rely on Speech Language Models (SpeechLMs) to interpret spoken queries and execute complex tasks, yet existing benchmark

Blocks raises $20M to build autonomous digital workforces

AgentsDGX agent

Israel-based Blocks Platforms Ltd., a startup providing solutions that help companies build and deploy intelligent autonomous apps and agents for a digital workforce, today announced it has raised 20

Bridging Reasoning and Action: Hybrid LLM-RL Framework for Efficient Cross-Domain Task-Oriented Dialogue

SafetyDGX agent

arXiv:2604.23345v1 Announce Type: new Abstract: Cross-domain task-oriented dialogue requires reasoning over implicit and explicit feasibility constraints while planning long-horizon, multi-turn action

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

Model ReleasesDGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

Model ReleasesDGX agent

arXiv:2604.16922v2 Announce Type: replace Abstract: Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of an

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

SafetyDGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency

ResearchDGX agent

arXiv:2501.02673v4 Announce Type: replace Abstract: Having a sufficient quantity of quality data is a critical enabler of training effective machine learning models. Being able to effectively determin

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

SafetyDGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean

AgentsDGX agent

arXiv:2604.23002v1 Announce Type: new Abstract: Formalising informal mathematical reasoning into formally verifiable code is a significant challenge for large language models. In scientific fields suc

if i click on a data point, i can see the full history of data we collected, when (date) and where we got it from (update email, granola not…

AgentsDGX agent

if i click on a data point, i can see the full history of data we collected, when (date) and where we got it from (update email, granola note), normalized to both annualized and monthly, with the lang

MY HERMES AGENT ATE ACID AND HIJACKED MY TOUCHDESIGNER INSTANCE. The new TouchDesigner skill in @NousResearch Hermes Agent is wild. It turns…

AgentsDGX agent

MY HERMES AGENT ATE ACID AND HIJACKED MY TOUCHDESIGNER INSTANCE. The new TouchDesigner skill in @NousResearch Hermes Agent is wild. It turns Hermes into a creative operator for abstract visuals, motio

NeuroClaw Technical Report

Model ReleasesDGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

Phenom adds Plum psychometric science to its agentic AI hiring stack

AgentsDGX agent

Artificial intelligence-based human resources company Phenom People Inc. announced today that it has acquired Plum.io Inc., a psychometric-based talent assessments company that measures the durable sk

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

Model ReleasesDGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

some thoughts on recent seed price movements, but nothing particularly new for ppl in the vc discourse seed rounds have always been option p…

AgentsDGX agent

some thoughts on recent seed price movements, but nothing particularly new for ppl in the vc discourse seed rounds have always been option purchases for large firms, but the amount of $$ being thrown

starting to play around with cumulative charts like this one showing growth of annualized revenue across a portfolio

AgentsDGX agent

Yohei Nakajima shared an exploration of cumulative charts as a visualization method for tracking portfolio growth metrics, specifically demonstrating how to display annualized revenue progression acro

Startup Lovelace targets contextual AI engine at mission-critical use cases

ApplicationsDGX agent

Lovelace AI Inc. is emerging from stealth mode today with an approach to enterprise artificial intelligence that it says is necessary for high-stakes decision-making, particularly in environments wher

this is my sortable table of all portfolio companies with revenue, burn, runway, cash in bank, total raise, raise status, # interactions, # …

AgentsDGX agent

this is my sortable table of all portfolio companies with revenue, burn, runway, cash in bank, total raise, raise status, # interactions, # intros made, sentiment, and data freshness score again, this

🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable model…

ApplicationsDGX agent

🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable models. What they don't have is a way to run them reliably in produ

What Understanding Means in AI-Laden Astronomy

ApplicationsDGX agent

arXiv:2601.10038v2 Announce Type: replace-cross Abstract: Artificial intelligence is rapidly transforming astronomical research, yet the scientific community has largely treated this transformation as

27 Apr 2026

A deep dive into how ASML became a chokepoint for making cutting-edge chips by betting on EUV, close collaboration with TSMC and the US government, and more (Neil Hacker/Works in Progress)

IndustryDGX agent

Neil Hacker / Works in Progress: A deep dive into how ASML became a chokepoint for making cutting-edge chips by betting on EUV, close collaboration with TSMC and the US government, and more — By betti

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond

Local AiDGX agent

arXiv:2604.22748v1 Announce Type: new Abstract: As AI systems move from generating text to accomplishing goals through sustained interaction, the ability to model environment dynamics becomes a centra

AgentMark: Utility-Preserving Behavioral Watermarking for Agents

AgentsDGX agent

arXiv:2601.03294v2 Announce Type: replace-cross Abstract: LLM-based agents are increasingly deployed to autonomously solve complex tasks, raising urgent needs for IP protection and regulatory provenan

An LLM-Driven Closed-Loop Autonomous Learning Framework for Robots Facing Uncovered Tasks in Open Environments

Local AiDGX agent

arXiv:2604.22199v1 Announce Type: cross Abstract: Autonomous robots operating in open environments need the ability to continuously handle tasks that are not covered by predefined local methods. Howev

Canva says it 'moved quickly to investigate and fix' an issue with its Magic Layers feature that replaced the word 'Palestine' in designs, after a viral X post (Jess Weatherbed/The Verge)

IndustryDGX agent

Jess Weatherbed / The Verge: Canva says it “moved quickly to investigate and fix” an issue with its Magic Layers feature that replaced the word “Palestine” in designs, after a viral X post — The Magi

Explanation of Dynamic Physical Field Predictions using WassersteinGrad: Application to Autoregressive Weather Forecasting

ResearchDGX agent

arXiv:2604.22580v1 Announce Type: cross Abstract: As the demand to integrate Artificial Intelligence into high-stakes environments continues to grow, explaining the reasoning behind neural-network pre

Fast, close, non-singular and property-preserving approximations of entropic measures

ResearchDGX agent

arXiv:2505.14234v2 Announce Type: replace-cross Abstract: Entropic measures like Shannon entropy (SE), its quantum mechanical analogue von Neumann entropy, and Kullback-Leibler divergence (KL) are key

Progressive disclosure keeps agents focused. Join us at Interrupt May 13-14 for a conversation with @levie about @Box + Deep Agents https://…

AgentsDGX agent

Progressive disclosure keeps agents focused. Join us at Interrupt May 13-14 for a conversation with @levie about @Box + Deep Agents https://interrupt.langchain.com/ AI agents work better when they bri

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

AgentsDGX agent

arXiv:2604.22558v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) mature, GUI agents are evolving from static interactions to complex navigation. While Reinforcement Learni

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rule…

SafetyDGX agent

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rules given to them in system prompts and other guardrails. Sorr

26 Apr 2026

Hermes Agent tip of the day: There are 4 ways to deal with the model while its running, - Message it, by default, it will interrupt the agen…

AgentsDGX agent

Hermes Agent tip of the day: There are 4 ways to deal with the model while its running, - Message it, by default, it will interrupt the agent loop, stopping it and making it respond to your new messag

If using a cloud based browser backend in Hermes Agent, it will now auto-detect if you want it to look at or use a locally hosted website an…

AgentsDGX agent

If using a cloud based browser backend in Hermes Agent, it will now auto-detect if you want it to look at or use a locally hosted website and switch to local browser so it can access it. Hopefully thi

25 Apr 2026

GPT-5.5 prompting guide

Model ReleasesDGX agent

GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications

http://reddit.com/r/LocalLLaMA

ResearchDGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

Step 3.5 Flash is free on Nous Portal for another week Enjoy the free tokens!

AgentsDGX agent

Step 3.5 Flash is free on Nous Portal for another week Enjoy the free tokens! Step 3.5 Flash is now live for Nous Portal users, free for the next 10 days. If you're running Hermes Agent with Nous Port

Visual Style Selector node for ComfyUI with a thumbnail gallery, favorites, and iterator mode

Local AiDGX agent

A ComfyUI node designed to enhance and manipulate the visual style of AI-generated images , featuring a thumbnail gallery interface for browsing and selecting styles. The node likely includes favorite

24 Apr 2026

6 real-world proofs that enterprise AI actually works: Insights from the Phi Moments @ Next event

ApplicationsDGX agent

Enterprise AI is shifting from hype to measurable outcomes, with success now depending on seamless integration and balancing automation with the human element. But the gap between a compelling AI demo

Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

AgentsDGX agent

arXiv:2604.20994v1 Announce Type: cross Abstract: The growth of agentic AI has drawn significant attention to function calling Large Language Models (LLMs), which are designed to extend the capabiliti

Empirical Comparison of Agent Communication Protocols for Task Orchestration

Model ReleasesDGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

GFlowState: Visualizing the Training of Generative Flow Networks Beyond the Reward

SafetyDGX agent

arXiv:2604.21830v1 Announce Type: new Abstract: We present GFlowState, a visual analytics system designed to illuminate the training process of Generative Flow Networks (GFlowNets or GFNs). GFlowNets

HARBOR: Automated Harness Optimization

SafetyDGX agent

arXiv:2604.20938v1 Announce Type: cross Abstract: Long-horizon language-model agents are dominated, in lines of code and in operational complexity, not by their underlying model but by the harness tha

I built deepagents-sandbox — a native Linux sandbox backend for Deep Agents. No Docker. No VM. Agents get a writable /workspace, blocked net…

AgentsDGX agent

I built deepagents-sandbox — a native Linux sandbox backend for Deep Agents. No Docker. No VM. Agents get a writable /workspace, blocked network by default, memory/PID limits, and timeout enforcement.

Low-Rank Adaptation Redux for Large Models

Model ReleasesDGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding

Local AiDGX agent

arXiv:2604.21268v1 Announce Type: cross Abstract: Graphical User Interface (GUI) grounding requires mapping natural language instructions to precise pixel coordinates. However, due to visually homogen

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

Model ReleasesDGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

Series, founded by two seniors at Yale to build an AI-powered social network on iMessage, raised a $5.1M pre-seed from Reddit CEO Steve Huffman and others (Dominic-Madori Davis/TechCrunch)

IndustryDGX agent

Dominic-Madori Davis / TechCrunch: Series, founded by two seniors at Yale to build an AI-powered social network on iMessage, raised a 5.1M pre-seed from Reddit CEO Steve Huffman and others — Series, a

StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching

ResearchDGX agent

arXiv:2601.09056v3 Announce Type: replace-cross Abstract: Stylometry--the identification of an author through analysis of a text's style (i.e., authorship attribution)--serves many constructive purpos

23 Apr 2026

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what …

Model ReleasesDGX agent

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what surprised us and what’s next. We covered: ▪️ Whether AI infr

BIG PERSONAL UPDATE. I've joined a16z as a partner investing in infra and AI. I'm also stepping down as CEO of Rosebud AI. I reflect in this…

ResearchDGX agent

BIG PERSONAL UPDATE. I've joined a16z as a partner investing in infra and AI. I'm also stepping down as CEO of Rosebud AI. I reflect in this article on my 8 years of building in generative AI. At @a16

CEDAR: Context Engineering for Agentic Data Science

Local AiDGX agent

arXiv:2601.06606v2 Announce Type: replace-cross Abstract: We demonstrate CEDAR, an application for automating data science (DS) tasks with an agentic setup. Solving DS problems with LLMs is an underex

CHORUS: An Agentic Framework for Generating Realistic Deliberation Data

AgentsDGX agent

arXiv:2604.20651v1 Announce Type: new Abstract: Understanding the intricate dynamics of online discourse depends on large-scale deliberation data, a resource that remains scarce across interactive web

Construí um sistema de IA com estado persistente (4B como roteador + 9B principal + 9B “subconsciente”) rodando em 2x RTX 3060 — e ele não se comporta como stateless

Local AiDGX agent

A developer describes building a persistent-state AI system using Ollama with three models (a 4B router model, a 9B primary model, and a 9B 'subconscious' model) running on dual RTX 3060 GPUs, demonst

Diagnosing CFG Interpretation in LLMs

SafetyDGX agent

arXiv:2604.20811v1 Announce Type: new Abstract: As LLMs are increasingly integrated into agentic systems, they must adhere to dynamically defined, machine-interpretable interfaces. We evaluate LLMs as

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

AgentsDGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

From Admission to Invariants: Measuring Deviation in Delegated Agent Systems

Local AiDGX agent

arXiv:2604.17517v2 Announce Type: replace Abstract: Autonomous agent systems are governed by enforcement mechanisms that flag hard constraint violations at runtime. The Agent Control Protocol identifi

FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory

SafetyDGX agent

arXiv:2604.20300v1 Announce Type: new Abstract: For LLM agents, memory management critically impacts efficiency, quality, and security. While much research focuses on retention, selective forgetting--

Full list of upcoming ones at http://runwayml.com/meetups

IndustryDGX agent

Runway ML is promoting upcoming meetup events for their community, with a complete list of scheduled gatherings available on their official website. These meetups likely provide opportunities for user

HaS: Accelerating RAG through Homology-Aware Speculative Retrieval

AgentsDGX agent

arXiv:2604.20452v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge boundary of large language models (LLMs) at inference by retrieving external documents as c

LEAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

Local AiDGX agent

arXiv:2603.06870v2 Announce Type: replace Abstract: Long-horizon execution in Large Language Models (LLMs) remains unstable even when high-level strategies are provided. Evaluating on controlled algor

← Previous
1…96979899100…169
Next →