AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
21 Apr 2026

MoRI: Learning Motivation-Grounded Reasoning for Scientific Ideation in Large Language Models

AgentsDGX agent

arXiv:2603.19044v2 Announce Type: replace Abstract: Scientific ideation aims to propose novel solutions within a given scientific context. Existing LLM-based agentic approaches emulate human research

Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding

AgentsDGX agent

arXiv:2510.15253v3 Announce Type: replace Abstract: Document understanding is critical for applications from financial analysis to scientific discovery. Current approaches, whether OCR-based pipelines

The Thin Line Between Comprehension and Persuasion in LLMs

AgentsDGX agent

arXiv:2507.01936v3 Announce Type: replace Abstract: Large language models (LLMs) are excellent at maintaining high-level, convincing dialogue, but it remains unclear whether their persuasive success r

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Will People Enjoy a Robot Trainer? A Case Study with Snoopie the Pacerbot

AgentsDGX agent

arXiv:2604.18331v1 Announce Type: new Abstract: The physicality of exercise makes the role of athletic trainers unique. Their physical presence allows them to guide a student through a motion, demonst

20 Apr 2026

CSLE: A Reinforcement Learning Platform for Autonomous Security Management

AgentsDGX agent

arXiv:2604.15590v1 Announce Type: cross Abstract: Reinforcement learning is a promising approach to autonomous and adaptive security management in networked systems. However, current reinforcement lea

Dynamic Tool Dependency Retrieval for Lightweight Function Calling

Model ReleasesDGX agent

arXiv:2512.17052v4 Announce Type: replace Abstract: Function calling agents powered by Large Language Models (LLMs) select external tools to automate complex tasks. On-device agents typically use a re

Mind DeepResearch Technical Report

Model ReleasesDGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

some interesting trends in terms of LLM adoption 👀

AgentsDGX agent

some interesting trends in terms of LLM adoption 👀 📊 The latest from the LangSmith Signal: We looked at API call volume and developer adoption trends across LLM providers using LangChain open source t

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-t…

AgentsDGX agent

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-trained and gets the best out of models by being “in-distribu

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

AgentsDGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

We have been named one of the 40 Most Innovative AI-Native Prosumer Companies by @notablecap. The Prosumer AI 40 recognizes companies buildi…

Model ReleasesDGX agent

We have been named one of the 40 Most Innovative AI-Native Prosumer Companies by @notablecap. The Prosumer AI 40 recognizes companies building the tools that blur the line between professional and con

19 Apr 2026

Awesome post by @LangChain team on “How We Made Our Docs Test Themselves”. Devtools companies that grok how important *technically accurate*…

AgentsDGX agent

Awesome post by @LangChain team on “How We Made Our Docs Test Themselves”. Devtools companies that grok how important *technically accurate* docs are both to humans and AI coding agents outperform on

17 Apr 2026

Beyond Literal Mapping: Benchmarking and Improving Non-Literal Translation Evaluation

AgentsDGX agent

arXiv:2601.07338v2 Announce Type: replace Abstract: Large Language Models (LLMs) have significantly advanced Machine Translation (MT), applying them to linguistically complex domains-such as Social Ne

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

AgentsDGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees

SafetyDGX agent

arXiv:2604.14243v1 Announce Type: new Abstract: Real-world decision-making systems operate in environments where state transitions depend not only on the agent's actions, but also on extbf{exogenous f

Towards AI-assisted Neutrino Flavor Theory Design

AgentsDGX agent

arXiv:2506.08080v2 Announce Type: replace-cross Abstract: Particle physics theories, such as those which explain neutrino flavor mixing, arise from a vast landscape of model-building possibilities. A

WybeCoder: Verified Imperative Code Generation

AgentsDGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

16 Apr 2026

A closer look at how large language models trust humans: patterns and biases

ResearchDGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation

Model ReleasesDGX agent

arXiv:2604.13946v1 Announce Type: cross Abstract: Automated code generation remains a persistent challenge in software engineering, as conventional multi-agent frameworks are often constrained by stat

From Instruction to Event: Sound-Triggered Mobile Manipulation

SafetyDGX agent

arXiv:2601.21667v2 Announce Type: replace-cross Abstract: Current mobile manipulation research predominantly follows an instruction-driven paradigm, where agents rely on predefined textual commands to

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

Model ReleasesDGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

AgentsDGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

15 Apr 2026

Beyond Majority Voting: Efficient Best-Of-N with Radial Consensus Score

AgentsDGX agent

arXiv:2604.12196v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate multiple candidate responses for a given prompt, yet selecting the most reliable one remains challengin

Defining and Evaluation Method for External Human-Machine Interfaces

AgentsDGX agent

arXiv:2604.12293v1 Announce Type: new Abstract: As the number of fatalities involving Autonomous Vehicles increase, the need for a universal method of communicating between vehicles and other agents o

PrivacyReasoner: Can LLM Emulate a Human-like Privacy Mind?

AgentsDGX agent

arXiv:2601.09152v2 Announce Type: replace Abstract: Prior work on LLM-based privacy focuses on norm judgment over synthetic vignettes, rather than how people think about a specific data practice and f

The Stackelberg Speaker: Optimizing Persuasive Communication in Social Deduction Games

SafetyDGX agent

arXiv:2510.09087v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown remarkable progress in social deduction games (SDGs). However, existing approaches primarily focus on i

‘You better have a lot of trust’: Oracle’s urgent case for rebuilding AI from the data up

AgentsDGX agent

AI can now generate thousands of lines of working code in minutes — but the question of whether enterprises can trust what those systems build has become the defining challenge of the current moment.

14 Apr 2026

Consensus-based Recursive Multi-Output Gaussian Process

AgentsDGX agent

arXiv:2604.10146v1 Announce Type: new Abstract: Multi-output Gaussian Processes provide principled uncertainty-aware learning of vector-valued fields but are difficult to deploy in large-scale, distri

Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation

Local AiDGX agent

arXiv:2604.10741v1 Announce Type: cross Abstract: Recent agentic search frameworks enable deep research via iterative planning and retrieval, reducing hallucinations and enhancing factual grounding. H

Fountain launches Cue to manage modern-day frontline workforce hiring and scheduling

Model ReleasesDGX agent

Frontline workforce hiring and management platform Fountain today launched Cue, an artificial intelligence enterprise-ready management capability designed to assist with hiring and labor operations au

I had such a great time at @aiDotEngineer Europe last week ! Except for one thing: My workshop went terribly because I vibe-slopped a little…

AgentsDGX agent

I had such a great time at @aiDotEngineer Europe last week ! Except for one thing: My workshop went terribly because I vibe-slopped a little bit too hard half an hour before it started and was too fra

Introducing BigQuery Graph: Unlock hidden relationships in your data

AgentsDGX agent

Today, we're thrilled to announce that BigQuery Graph is now available in preview. With BigQuery Graph, we’ve built an easy-to-use, highly scalable graph analytics solution for data engineers, data an

M^3KG-RAG: Multi-hop Multimodal Knowledge Graph-enhanced Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2512.20136v3 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has recently been extended to multimodal settings, connecting multimodal large language models (MLLMs) wi

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

AgentsDGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

13 Apr 2026

AgentZ — SOC Level AI With Ollama

Local AiDGX agent

AgentZ is a locally-run, SOC (Security Operations Center) level AI agent built on top of Ollama, designed to assist with cybersecurity tasks such as threat analysis, alert triage, and incident respons

AI-Induced Human Responsibility (AIHR) in AI-Human teams

AgentsDGX agent

arXiv:2604.08866v1 Announce Type: cross Abstract: As organizations increasingly deploy AI as a teammate rather than a standalone tool, morally consequential mistakes often arise from joint human-AI wo

Contribution of task-irrelevant stimuli to drift of neural representations

AgentsDGX agent

arXiv:2510.21588v2 Announce Type: replace-cross Abstract: Biological and artificial learners are inherently exposed to a stream of data and experience throughout their lifetimes and must constantly ad

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

Model ReleasesDGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation

AgentsDGX agent

arXiv:2604.08782v1 Announce Type: new Abstract: Large language models (LLMs) suffer significant performance degradation when user instructions and context are distributed over multiple conversational

12 Apr 2026

A look at the escalating global AI arms race, as the US, China, Russia, and others rush to build AI-backed autonomous weapons and defense systems (New York Times)

AgentsDGX agent

New York Times: A look at the escalating global AI arms race, as the US, China, Russia, and others rush to build AI-backed autonomous weapons and defense systems — China, the U.S., Russia and others h

Amazing article. If you own the harness you own your memories. Else you get locked into an API model which keeps memory behind APIs and you …

AgentsDGX agent

Harrison Chase, co-founder of LangChain, shared or engaged with commentary emphasizing the importance of owning your own memory harness in AI applications rather than relying on third-party API-based

harness engineering > prompt engineering prompts get you 80% there the last 20% — tool routing, memory retrieval, retry logic, evals — thats…

AgentsDGX agent

Harrison Chase, co-founder of LangChain, argues that while prompt engineering is foundational and can get you 80% of the way to a functional AI system, the remaining 20% of engineering effort involves

Have an idea for an experiment? Hermes can now write conference-grade research papers alongside you.

AgentsDGX agent

Have an idea for an experiment? Hermes can now write conference-grade research papers alongside you. Media introducing Autoreason, a reasoning method inspired by @karpathy's AutoResearch which extends

I fed The Godfather into a structured knowledge graph, here's what the MCP tools surface

AgentsDGX agent

A Reddit post from the r/ollama community demonstrating a practical experiment in which a user ingested the narrative content of *The Godfather* into a structured knowledge graph and then queried it u

Memory, the next toe-hold for closed AI platforms; increasing user ergonomics whilst playing the long game against customers.

AgentsDGX agent

Memory features in AI platforms represent a strategic mechanism for increasing user retention and platform lock-in, as personalized context and learned preferences become increasingly difficult to mig

Not your memory not your life

AgentsDGX agent

Harrison Chase, co-founder of LangChain, likely discusses the concept that AI agents and systems lacking persistent memory cannot truly learn, adapt, or maintain continuity across interactions — makin

The clarity of this piece by @hwchase17 is immense. No jargon.

AgentsDGX agent

Harrison Chase, the creator of LangChain, shared or was referenced in a post praising a piece of writing attributed to him for its clarity and accessibility, notably avoiding technical jargon. The con

This. And i know it works very well.

AgentsDGX agent

Harrison Chase, the co-founder and CEO of LangChain, shared or endorsed a post on X (formerly Twitter) expressing strong confidence in a particular approach, tool, or method, stating it 'works very we

11 Apr 2026

🔗 Codex App: https://chatgpt.com/codex/

Model ReleasesDGX agent

OpenAI's Codex App (available at chatgpt.com/codex) is a dedicated command center for agentic coding, enabling developers to manage multiple AI coding agents working in parallel across projects. T...

Most illuminating graph I have seen for definition of harness. Memory and context are deeply coupled with harness. In my humble opinion, how…

AgentsDGX agent

Most illuminating graph I have seen for definition of harness. Memory and context are deeply coupled with harness. In my humble opinion, how they are injected and managed are one of most important pie

The most important post I've read all week by @hwchase17 . Memory, particularly memory consistency is the biggest performance inhibitor to A…

AgentsDGX agent

The most important post I've read all week by @hwchase17 . Memory, particularly memory consistency is the biggest performance inhibitor to AI agents today. As we interact with different harnesses and

This is why you need model agnostic harnesses

Model ReleasesDGX agent

I was unable to retrieve the content of that specific X (Twitter) post, as web search results did not surface the tweet or its content. X.com posts are generally not indexed in a way that makes the...

What if owning the memory layer goes far beyond escaping lock-in? What if memory becomes fully editable intelligence? You could debug bad pa…

AgentsDGX agent

What if owning the memory layer goes far beyond escaping lock-in? What if memory becomes fully editable intelligence? You could debug bad patterns, reinforce good ones, and crucially synthesize new on

10 Apr 2026

middleware is underrated

AgentsDGX agent

I wasn't able to retrieve the content of that specific tweet or URL from the search results. X (Twitter) posts are generally not indexed in a way that allows direct retrieval of their content, and ...

My X bookmarks have an 'intellectual shape' Apparently mine is: Technique-heavy, Tool-obsessed, Light on Opinion I didn't know until I ran t…

AgentsDGX agent

My X bookmarks have an 'intellectual shape' Apparently mine is: Technique-heavy, Tool-obsessed, Light on Opinion I didn't know until I ran the numbers OSS I've been building w/@LangChain to explore Ge

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

AgentsDGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

PriPG-RL: Privileged Planner-Guided Reinforcement Learning for Partially Observable Systems with Anytime-Feasible MPC

SafetyDGX agent

arXiv:2604.08036v1 Announce Type: cross Abstract: This paper addresses the problem of training a reinforcement learning (RL) policy under partial observability by exploiting a privileged, anytime-feas

Speed of open source! Anthropic adds advisor strategy yesterday -> Emanuele adds an implementation as middleware less than 24 hours later!

AgentsDGX agent

Speed of open source! Anthropic adds advisor strategy yesterday -> Emanuele adds an implementation as middleware less than 24 hours later! Just shipped advisor-middleware: an open-source implementatio

The Illusion of Stochasticity in LLMs

AgentsDGX agent

arXiv:2604.06543v1 Announce Type: cross Abstract: In this work, we demonstrate that reliable stochastic sampling is a fundamental yet unfulfilled requirement for Large Language Models (LLMs) operating

We have @liadyosef and @idosal1 at @aiDotEngineer talking MCP apps and why we need them.

AgentsDGX agent

Nick Taylor (@nickytonline) shared a post highlighting a talk at AI Engineer (@aiDotEngineer) featuring Liad Yosef (@liadyosef) and Ido Salomon (@idosal1) on the topic of MCP apps and their importa...

← Previous
1…158159160161162…300
Next →