AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
22 Jun 2026

I’m running a 3 hour advanced workshop at AI Engineer World’s Fair! 🚀 2026 has greatly changed how one should learn lower-level technicals …

AgentsDGX agent

I’m running a 3 hour advanced workshop at AI Engineer World’s Fair! 🚀 2026 has greatly changed how one should learn lower-level technicals like kernels, agentic RL, reward hacking, cont learning. What

it's official, I'm on community at NousResearch. I'll be shipping tested walkthroughs, answering questions and carrying what the community b…

AgentsDGX agent

it's official, I'm on community at NousResearch. I'll be shipping tested walkthroughs, answering questions and carrying what the community builds back to the team. if you're on Hermes Agent, the Disco

OMG! Fugu Ultra is ridiculously good at these 3D renders.

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OMG! Fugu Ultra is ridiculously good at these 3D renders. Media Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the p

Porting the Moebius 0.2B image inpainting model to run in the browser with Claude Code

Model ReleasesDGX agent

This morning on Hacker News I saw Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance, describing a small but effective inpainting model - a model where you can mark region

The @aiDotEngineer schedule wins for having the most possible ways that developers can interact with it: JSON, llms.txt, MCP server, iCal, e…

AgentsDGX agent

The @aiDotEngineer schedule wins for having the most possible ways that developers can interact with it: JSON, llms.txt, MCP server, iCal, embeddings, agent skill https://ai.engineer/worldsfair/schedu

20 Jun 2026

Had the pleasure of meeting @jerryjliu0 at a @llama_index meetup in SF. We talked about document parsing and vector search market. 🌉 DM if …

AgentsDGX agent

Had the pleasure of meeting @jerryjliu0 at a @llama_index meetup in SF. We talked about document parsing and vector search market. 🌉 DM if you wanna meet in SF. Would love to chat about search, agents

11 Jun 2026

Towards Responsibly Non-Compliant Machines

AgentsDGX agent

arXiv:2606.12147v1 Announce Type: new Abstract: We consider the problem of engineering autonomous intelligent agents that are capable to responsibly not comply with user requests. We argue that machin

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

AgentsDGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

10 Jun 2026

EstRTL: Functional Estimation Guided RTL Code Generation

AgentsDGX agent

arXiv:2606.09867v1 Announce Type: cross Abstract: Optimizing register transfer level (RTL) code is of vital importance in hardware design. Large language models (LLMs) provide new methods for the auto

first in a series of technical blogs of how we build llm infra

AgentsDGX agent

first in a series of technical blogs of how we build llm infra How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latenc

i showcase 'controlled' self improvement with a novel regime-to-seam approach where failures are categorized and allowed to fix targeted are…

AgentsDGX agent

i showcase 'controlled' self improvement with a novel regime-to-seam approach where failures are categorized and allowed to fix targeted areas of the agent while interesting, it's more to showcase the

Lium raises $5.5M to unlock complex scientific data for AI models

AgentsDGX agent

Lium, a startup formerly known as Astromind, today announced the launch of an “agentic harness” that helps large language models dig into the most complex and messiest datasets. The launch comes after

my weekend hobby: self improvement research

AgentsDGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation

AgentsDGX agent

arXiv:2606.10875v1 Announce Type: new Abstract: Large language models (LLMs) rely on tool use to act as autonomous agents, yet often fail in multi-step execution due to insufficient tool-related knowl

The Distributed Detectability Band Against Marginal-Preserving Attacks

AgentsDGX agent

arXiv:2606.10456v1 Announce Type: cross Abstract: AI-control monitors score individual agent actions to detect misbehavior, but real harm can be distributed across many benign-looking steps, each indi

Zscaler unveils ZAgent Framework to automate zero-trust SASE operations

AgentsDGX agent

Zscaler Inc. today unveiled a major expansion of its zero-trust SASE platform, adding an agentic framework that lets administrators manage the system through natural-language prompts and extending its

9 Jun 2026

Advancing Mathematics Research with AI-Driven Formal Proof Search

AgentsDGX agent

arXiv:2605.22763v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly excel at mathematical reasoning, but their unreliability limits their utility in mathematics research. A m

Data center modernization unlocks AI budget headroom as enterprises fund new AI workloads

AgentsDGX agent

As enterprise AI budgets hit their limits earlier each year, the pressure to fund new agentic and inference workloads without expanding total spend is forcing a fundamental rethink of infrastructure a

Engagement Process: Rethinking the Temporal Interface of Action and Observation

AgentsDGX agent

arXiv:2605.11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over

GIFT: LLM-Guided State-Reward Interface for Financial Reinforcement Learning

AgentsDGX agent

arXiv:2606.08450v1 Announce Type: new Abstract: Financial portfolio trading is naturally formulated as a reinforcement learning problem, where an agent sequentially rebalances assets under changing ma

Motion planning for hundreds of floating robots

AgentsDGX agent

arXiv:2606.09620v1 Announce Type: new Abstract: Planning collision-free motion for large robot fleets is difficult because collision avoidance induces strong inter-agent coupling that grows rapidly wi

puffy plays pickleball june 30 in SF, co-hosted with good friends http://luma.com/the-agent-open

AgentsDGX agent

Puffy participated in a pickleball event on June 30 in San Francisco called 'The Agent Open,' which was co-hosted with Good Friends. The event appears to have been organized or promoted through Luma,

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

Model ReleasesDGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

Self-Paced Curriculum Reinforcement Learning for Autonomous Superbike Racing in Simulation

AgentsDGX agent

arXiv:2606.09236v1 Announce Type: cross Abstract: Autonomous Racing has seen remarkable progress through deep Reinforcement Learning (RL), primarily for four-wheeled vehicles. However, motorbikes intr

Semantic Quorum Assurance: Collective Certification for Non-Deterministic AI Infrastructure

SafetyDGX agent

arXiv:2606.08021v1 Announce Type: cross Abstract: As large language model (LLM) agents are integrated into autonomous cloud operations, distributed systems face a semantic reliability problem: propose

this is sick haha.

AgentsDGX agent

this is sick haha. The Agent Open 🎾🏓 Everyone loves pickleball. We’re hosting a massive pickleball tournament during the AI Engineer World Fair. I’m so excited to see this event come together. This is

8 Jun 2026

@cognition funny you should say that https://x.com/manuelsampedrop/status/2063746243180773656?s=20

AgentsDGX agent

@cognition funny you should say that https://x.com/manuelsampedrop/status/2063746243180773656?s=20 @swyx @cognition hope it scores scope discipline too, half my agent failures are perfectly fine code

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a…

AgentsDGX agent

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a much broader set of verticalized agents and infrastructure.

OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational Gait Analysis in Post-Stroke Rehabilitation

AgentsDGX agent

arXiv:2604.05360v2 Announce Type: replace-cross Abstract: Gait analysis is essential in post-stroke rehabilitation but remains time-intensive and cognitively demanding, especially when clinicians must

Should You Use Your Large Language Model to Explore or Exploit?

AgentsDGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

7 Jun 2026

OpenAI’s planned ‘superapp’ gets closer as one employee says ‘chat is dead’

AgentsDGX agent

OpenAI Group PBC is still focused on its plans to transform ChatGPT into some kind of “superapp,” and it will have a heavy focus on artificial intelligence agents and autonomous coding bots, according

6 Jun 2026

Enhancing Software Engineering Through Closed-Loop Memory Optimization

Model ReleasesDGX agent

arXiv:2606.05646v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled powerful software engineering (SE) agents capable of navigating complex codebases and resolving real-world i

5 Jun 2026

ABBEL: Learning Natural-Language Belief States for Memory-Efficient Interaction

AgentsDGX agent

arXiv:2512.20111v2 Announce Type: replace Abstract: As the time horizons of sequential decision-making tasks grow, keeping full interaction histories in model context becomes increasingly costly. Rece

AI Innovation Takes Center Stage at Milan AI Week’s Hackathon

AgentsDGX agent

Meet the winners of the Milan AI Week Hackathon. Discover how Cascade AI, ContextGuard, and Deals Machine are using agentic AI, workflow automation, and real-time orchestration on Vultr infrastructure

Building AI that Builds AI: Introducing the Sakana AI RSI Lab 🚀 https://sakana.ai/rsi-lab Today, we are announcing the Sakana AI Recursive …

AgentsDGX agent

Building AI that Builds AI: Introducing the Sakana AI RSI Lab 🚀 https://sakana.ai/rsi-lab Today, we are announcing the Sakana AI Recursive Self-Improvement (RSI) Lab: a dedicated research group in Tok

Emergent Language as an Approach to Conscious AI

AgentsDGX agent

arXiv:2606.06380v1 Announce Type: new Abstract: The question of whether artificial systems can be conscious remains open, in part because existing approaches either evaluate systems against theory-der

4 Jun 2026

CADENCE: Predicting Realized MAPF Execution Time Beyond Sum of Costs

AgentsDGX agent

arXiv:2606.04746v1 Announce Type: new Abstract: Multi-Agent Path Finding (MAPF) algorithms are increasingly used to plan motion for robot teams in industrial warehouses and robotic shared workspaces,

DPDL: Towards Differential Privacy Preservation in Decentralized Stochastic Learning on Non-IID Data

Local AiDGX agent

arXiv:2606.04399v1 Announce Type: new Abstract: In the paradigm of decentralized learning, a group of agents collaborate to train a global model using distributed datasets without a central server. Al

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

Model ReleasesDGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

Position: Deployed Reinforcement Learning should be Continual

AgentsDGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

Stateful Visual Encoders for Vision-Language Models

AgentsDGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

3 Jun 2026

Decentralized Stochastic Nonconvex Optimization under the (L_0,L_1)-Smoothness

AgentsDGX agent

arXiv:2509.08726v3 Announce Type: replace-cross Abstract: This paper focuses on the decentralized stochastic optimization problem f(mathbf{x})=frac{1}{m}sum_{i=1}^m f_i(mathbf{x}) over a connected net

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation

Model ReleasesDGX agent

arXiv:2604.17220v2 Announce Type: replace-cross Abstract: Modeling coordination among generative agents in complex multi-round decision-making presents a core challenge for AI and operations managemen

latest iteration of middleware is rad: there was a subagent that we had that was taking some pretty wild trajectories and was costing wayyyy…

AgentsDGX agent

latest iteration of middleware is rad: there was a subagent that we had that was taking some pretty wild trajectories and was costing wayyyy too much We adapted a 60 line middleware from a different a

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or dist…

AgentsDGX agent

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or distillation from previous models. this means reasoning, agentic

Minimax Optimal Strategy for Delayed Observations in Online Reinforcement Learning

AgentsDGX agent

arXiv:2603.03480v2 Announce Type: replace Abstract: We study reinforcement learning with delayed state observation, where the agent observes the current state after some random number of time steps. W

Psi-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

Model ReleasesDGX agent

arXiv:2606.02754v1 Announce Type: new Abstract: Personalization is a crucial capability of modern language agents. However, current research primarily positions personalized agents as passive responde

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

AgentsDGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engi…

AgentsDGX agent

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engineering problem. Outbound = agent design. Enrichment = retri

You shipped your app. Now what? Your app may look great, but if no one can find it, it stays invisible Publishing is only the beginning Meet…

AgentsDGX agent

You shipped your app. Now what? Your app may look great, but if no one can find it, it stays invisible Publishing is only the beginning Meet SEO Agent. It runs a scan for you and suggests fixes to hel

2 Jun 2026

Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation

Model ReleasesDGX agent

arXiv:2606.02528v1 Announce Type: cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. W

Automated Conjecture Resolution with Formal Verification

AgentsDGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation

SafetyDGX agent

arXiv:2606.00827v1 Announce Type: cross Abstract: Strategic classification (SC) investigates scenarios where agents manipulate their features to obtain favorable decisions from predictive models. Exis

Everyone talks about 1M context. The harder part is making 1M context actually usable. Serving MiniMax M3 required optimizing for long-conte…

AgentsDGX agent

Everyone talks about 1M context. The harder part is making 1M context actually usable. Serving MiniMax M3 required optimizing for long-context, multimodal, and agentic workloads simultaneously. Excite

Fiduciary grade AI sets the bar as Thomson Reuters and Snowflake bring governed intelligence to the professions

AgentsDGX agent

Professionals who carry personal liability for their decisions — lawyers, tax accountants, auditors — cannot afford AI that gets it wrong. As enterprises accelerate deployment of agentic systems, the

From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language Model

Model ReleasesDGX agent

arXiv:2512.05277v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as the perception and reasoning backbone of autonomous agents acting in the wild, with

GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing

Model ReleasesDGX agent

arXiv:2606.00860v1 Announce Type: cross Abstract: Self-report questionnaires remain the prevailing tool for probing the psychological states of persona-conditioned agents (PC-Agents). However, classic

micropython-wasm 0.1a1

AgentsDGX agent

micropython-wasm 0.1a1 is a Python library for running a MicroPython sandbox using WebAssembly . This alpha release includes fixes for limitations discovered while building datasette-agent-micropython

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

Model ReleasesDGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

AgentsDGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

← Previous
1…167168169170171…300
Next →