AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,214 results
8 Jun 2026

How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope

AgentsDGX agent

arXiv:2606.07489v1 Announce Type: new Abstract: Frontier AI systems are bridging the gap between intelligence and utility by shifting from conversational assistants to autonomous agents that execute t

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a…

AgentsDGX agent

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a much broader set of verticalized agents and infrastructure.

IDDMBSE: Integrating Data-Driven and Model-Based Systems Engineering for Trusted Autonomous Cyber-Physical Systems

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.06727v1 Announce Type: new Abstract: Autonomous cyber-physical systems (CPS) sit at the intersection of Model-Based Systems Engineering (MBSE) and data-driven Machine Learning and Artificia

Introducing FrontierCode: a coding eval that raises the bar for difficulty & quality. Each task took 40+ hrs of work by leading open-source …

AgentsDGX agent

Introducing FrontierCode: a coding eval that raises the bar for difficulty & quality. Each task took 40+ hrs of work by leading open-source maintainers. Models write sloppy code that works but isn’t m

IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems

AgentsDGX agent

arXiv:2606.06559v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models allow voice agents to listen and speak concurrently, enabling natural interaction with real-time overlap. However,

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours …

AgentsDGX agent

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours of maintainer validated software engineering work most front

Kimi Code, our open-source coding agent, just got a major upgrade! 🔹One-line CLI install, zero setup, fast startup​ 🔹Drag in videos as cod…

AgentsDGX agent

Kimi Code, our open-source coding agent, just got a major upgrade! 🔹One-line CLI install, zero setup, fast startup​ 🔹Drag in videos as coding context: reference-to-LUT, long-video-to-short, screen-rec

LangSmith Fleet lets you work with files directly. Create documents, presentations, and webpages inside a conversation, or upload your own f…

AgentsDGX agent

LangSmith Fleet introduces file handling capabilities that enable users to create and manipulate documents, presentations, and webpages directly within conversations, as well as upload existing files

LLM Agent-Assisted Reverse Engineering with Quantitative Readability Metrics

AgentsDGX agent

arXiv:2606.06838v1 Announce Type: cross Abstract: Automatic decompilers produce functionally correct but often unreadable C code. This paper addresses one stage of the reverse engineering workflow: im

Measuring Agents in Production

AgentsDGX agent

arXiv:2512.04123v4 Announce Type: replace-cross Abstract: LLM-based agents already operate in production across many industries, yet we lack an understanding of what technical methods make deployments

Model Context Protocols in Adaptive Transport Systems: A Survey

AgentsDGX agent

arXiv:2508.19239v2 Announce Type: replace Abstract: The rapid expansion of interconnected devices, autonomous systems, and AI applications has created severe fragmentation in adaptive transport system

Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA

AgentsDGX agent

arXiv:2603.24481v2 Announce Type: replace Abstract: Miscalibrated confidence scores are a practical obstacle to deploying AI in clinical settings. A model that is always overconfident offers no useful

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet…

AgentsDGX agent

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet that gap directly. (bookmark it) It studies agent adoption

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://try…

AgentsDGX agent

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://tryphoton.ai Your Hermes Agent now lives in iMessage via @photon

OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational Gait Analysis in Post-Stroke Rehabilitation

AgentsDGX agent

arXiv:2604.05360v2 Announce Type: replace-cross Abstract: Gait analysis is essential in post-stroke rehabilitation but remains time-intensive and cognitively demanding, especially when clinicians must

On the Hardness of Optimal Motion on Trees

AgentsDGX agent

arXiv:2606.06686v1 Announce Type: new Abstract: This paper presents a simple framework that settles the complexity of Multi-Agent Path Finding (MAPF) on trees across standard objectives--distance, mak

OpenSkill: Open-World Self-Evolution for LLM Agents

AgentsDGX agent

arXiv:2606.06741v1 Announce Type: new Abstract: Self-evolving agents requires adaptation after deployment, but existing approaches assume a usable learning loop, such as curated skills, successful tra

PandaAI: A Practical Agent CQ2 for Neuro-symbolic Data Analysis And Integrated Decision-Making in Quantitative Finance

AgentsDGX agent

arXiv:2606.06823v1 Announce Type: cross Abstract: While deep learning has excelled in various domains, its application to sequential decision-making in finance remains challenging due to the low Signa

Pega expands AI platform with agent orchestration, development tools and new pricing model

AgentsDGX agent

Workflow automation vendor Pegasystems Inc. today unveiled a broad set of artificial intelligence enhancements aimed at helping enterprises deploy AI agents in mission-critical business processes whil

@photon_hq More on the Photon integration: https://hermes-agent.nousresearch.com/docs/user-guide/messaging/photon

AgentsDGX agent

Nous Research shared documentation about their Photon integration, providing users with guidance on implementing Photon messaging functionality within the Hermes Agent platform. The resource likely co

Predicting Dynamic Map States from Limited Field-of-View Sensor Data

AgentsDGX agent

arXiv:2602.12360v2 Announce Type: replace Abstract: When autonomous systems are deployed in real-world scenarios, sensors are often subject to limited field-of-view (FOV) constraints, either naturally

SCALE: Scalable Cross-Attention Learning with Extrapolation for Agentic Workflow Scheduling

AgentsDGX agent

arXiv:2606.06820v1 Announce Type: cross Abstract: Agentic Large Language Model (LLM) systems decompose complex tasks into workflow Directed Acyclic Graphs (DAGs) whose primitives must be scheduled on

SCOUT: Semantic scene COverage via Uncertainty-guided Traversal

AgentsDGX agent

arXiv:2606.06721v1 Announce Type: cross Abstract: Robots that operate over extended periods should not merely visit space; they should progressively understand it. Yet most 3D scene graph pipelines tr

seattle friends! 📅 i'm doing a talk on @activegraphai on june 24th at the AI house (@ai2incubator), come by :) https://luma.com/aihouse-2fd…

AgentsDGX agent

Yohei Nakajima announced a talk about ActiveGraph AI scheduled for June 24th at the AI House, an AI2 incubator venue in Seattle, with an invitation for local attendees to participate. The event detail

Should You Use Your Large Language Model to Explore or Exploit?

AgentsDGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

Signal-Driven Observation for Long-Horizon Web Agents

AgentsDGX agent

arXiv:2606.06708v1 Announce Type: new Abstract: Web agents operating over long horizons ingest raw DOM and accessibility trees -- routinely tens of thousands of tokens -- at every action step, causing

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

AgentsDGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

Snowflake and 1Password tackle the growing challenge of securing AI agents at scale

AgentsDGX agent

As AI agents gain access to sensitive enterprise data, AI agent security is becoming a top priority for organizations. The challenge is no longer just protecting systems, but ensuring autonomous agent

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built…

AgentsDGX agent

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built out courts. Stadium seating. Exhibition matches by AI leader

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world …

AgentsDGX agent

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world of AI changes every 2 heartbeats, so by the time I finish thi

The Open Source Community is backing OpenEnv for Agentic RL

AgentsDGX agent

OpenEnv is an open-source framework backed by the community for training and developing agentic reinforcement learning systems. The project represents collaborative efforts within the open-source ecos

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

AgentsDGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

AgentsDGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

tl;dr: we aren’t close to RSI, regardless of the hints IPO-bound Anthropic tried to drop last week.

AgentsDGX agent

tl;dr: we aren’t close to RSI, regardless of the hints IPO-bound Anthropic tried to drop last week. This paper tests whether today’s AI agents can build better AI agents without human design help. i.e

TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

AgentsDGX agent

arXiv:2606.07054v1 Announce Type: cross Abstract: Autonomous LLM agents can pursue hidden malicious objectives through sequences of individually benign actions, making sabotage difficult to detect usi

🔗Try it now: https://www.kimi.com/products/kimi-work We're just getting started. More data sources, more tools, more agent capabilities are…

AgentsDGX agent

Kimi is launching a new product called Kimi Work, an AI work platform with expandable capabilities including additional data sources, tools, and agent functionalities. The announcement indicates the p

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://act…

AgentsDGX agent

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://activegraph.ai/blog/introducing-learn 💻 AG coder (open source) r

We found that more autonomy with autonomous agents like Computer tracks with higher quality and satisfaction.

AgentsDGX agent

Research indicates that autonomous agents with greater operational autonomy, particularly in computer-based tasks, demonstrate improved performance quality and user satisfaction outcomes. The study su

We got into Y Combinator! Agnost AI (YC S26) is the infra for self-improving AI agents. We plug into conversational AI companies, find what'…

AgentsDGX agent

We got into Y Combinator! Agnost AI (YC S26) is the infra for self-improving AI agents. We plug into conversational AI companies, find what's broken, and ship the fix as a PR. You just merge. DM if th

We published new research with Harvard on the shift from chat interfaces to autonomous agents like Computer. Over 3 months, findings show wo…

AgentsDGX agent

We published new research with Harvard on the shift from chat interfaces to autonomous agents like Computer. Over 3 months, findings show workers using Computer finish tasks in 87% less time at 94% lo

When Does Multi-Agent Collaboration Help? An Entropy Perspective

AgentsDGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

You can find full model results and technical implementation details on our blog: https://cognition.ai/blog/frontier-code

AgentsDGX agent

Cognition AI published detailed model results and technical implementation information on their blog, likely covering performance benchmarks, architecture specifics, and methodology for their frontier

Your Hermes Agent now lives in iMessage via @photon_hq. Run 'hermes gateway setup' and choose Photon to start texting your agent.

AgentsDGX agent

Nous Research has integrated their Hermes Agent with iMessage through a partnership with Photon, allowing users to interact with the agent via text messaging. Users can enable this functionality by ru

7 Jun 2026

Agree with everything except the Markdown part There's got to be a better agent-native format for representing unstructured docs Not convinc…

AgentsDGX agent

Agree with everything except the Markdown part There's got to be a better agent-native format for representing unstructured docs Not convinced it's markdown or html I need Google Docs but just for mar

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは…

AgentsDGX agent

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは、AIがAIそのものを作る仕組みです。 この2年間、私たちはLLM-Squared、Darwin Gödel Machi

Great paper on self-improving agents:

AgentsDGX agent

Great paper on self-improving agents: This was one of the standout AI papers of the week. (bookmark it) It tackles a question most self-improving AI agents ignore: is the agent actually discovering an

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1…

AgentsDGX agent

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1. Build v1 agent using a sensible base harness and some task

New MIT study. Code volume surges by 300%, but output increases by only 30%: The AI dividend meets an awkward reality Autonomous AI coding a…

AgentsDGX agent

New MIT study. Code volume surges by 300%, but output increases by only 30%: The AI dividend meets an awkward reality Autonomous AI coding agents raised commits by 180%, but releases rose only 30%. Th

OpenAI’s planned ‘superapp’ gets closer as one employee says ‘chat is dead’

AgentsDGX agent

OpenAI Group PBC is still focused on its plans to transform ChatGPT into some kind of “superapp,” and it will have a heavy focus on artificial intelligence agents and autonomous coding bots, according

Snowflake, Databricks and the model makers: The battle for the agentic client and AI back end

AgentsDGX agent

Agentic artificial intelligence is being misread as a set of separate battles – for example, Snowflake Inc. versus Databricks Inc., copilots versus agents, model makers versus application vendors. We

the copypasta was inevitable

AgentsDGX agent

the copypasta was inevitable my first vc story is a crazy one. my 2 cofounders and i met the gp and his investment committee at his office. they were extremely rowdy, couldnt tell if they were taking

The Top AI Papers of the Week (May 31 - June 7) - LEAP - AutoLab - Learn From Your Own Latents - Reusable Context Engineering - Self-Revisin…

AgentsDGX agent

The Top AI Papers of the Week (May 31 - June 7) - LEAP - AutoLab - Learn From Your Own Latents - Reusable Context Engineering - Self-Revising Discovery Systems - Scaling Laws for Agent Harnesses - Dis

This was one of the standout AI papers of the week. (bookmark it) It tackles a question most self-improving AI agents ignore: is the agent a…

AgentsDGX agent

This was one of the standout AI papers of the week. (bookmark it) It tackles a question most self-improving AI agents ignore: is the agent actually discovering anything, or just remixing what it alrea

We're building this at LangChain Fleet lets you create and manage a fleet of agents. Each agent specializes in a workflow, e.g. inbox manage…

AgentsDGX agent

We're building this at LangChain Fleet lets you create and manage a fleet of agents. Each agent specializes in a workflow, e.g. inbox management, blog writing, competitor research, candidate recruitin

6 Jun 2026

2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support

AgentsDGX agent

arXiv:2602.21889v2 Announce Type: replace Abstract: Predictions from ML models support human decision making in several fields, including high-stakes ones such as healthcare and the judiciary. Yet, we

A Finite Certificate for the Positive n=9 Vasc Inequality

AgentsDGX agent

arXiv:2606.06136v1 Announce Type: cross Abstract: We prove the positive-real n=9 case of the Vasc cyclic inequality. The proof was obtained with human-guided assistance from the AI agent MechMath Agen

A Motivational Architecture for Conversational AGI

AgentsDGX agent

arXiv:2606.05411v1 Announce Type: new Abstract: Motivational architectures in cognitive AI have largely been designed for physical agents regulating bodily needs. Conversational agents operate in a di

A super useful feature I like that we have in Grok Build: we load your .envrc and pass it straight into the agent’s shell environment. Same …

AgentsDGX agent

A super useful feature I like that we have in Grok Build: we load your .envrc and pass it straight into the agent’s shell environment. Same API keys, PATH, sccache, remote caches, .etc — everything yo

A Taxonomy of Runtime Faults in Model Context Protocol Servers

AgentsDGX agent

arXiv:2606.05339v1 Announce Type: cross Abstract: MCP (Model Context Protocol) enables LLMs (Large Language Models) to interact with external tools and data sources via a standardized protocol. Its ra

A2RAG: Adaptive Agentic Graph Retrieval for Cost-Aware and Reliable Reasoning

AgentsDGX agent

arXiv:2601.21162v2 Announce Type: replace-cross Abstract: Graph Retrieval-Augmented Generation (Graph-RAG) enhances multihop question answering by organizing corpora into knowledge graphs and routing

← Previous
1…4445464748…121
Next →