AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
22 Jul 2026

We dropped some new LlamaDrip 🧢 Fear of Docs LlamaParse

AgentsDGX agent

We dropped some new LlamaDrip 🧢 Fear of Docs LlamaParse The team flew into SF for a week onsite. 🌉 2x'd in size since we last did this — first time this many of us have been in the same room. The reca

21 Jul 2026

Can’t tell if the PR reads more like a security incident or a product release…

AgentsDGX agent

Can’t tell if the PR reads more like a security incident or a product release… We suspected last week's cyberattack might have come from a frontier lab, given the sophistication of the agent. Turns ou

16 Jul 2026

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

arXiv:2607.13125v1 Announce Type: cross Abstract: We introduce Boogu-Image-0.1, an open-source unified multimodal understanding and generation model family, comprising Base, Turbo, Edit, and Edit-Turb

Flow-aware Optimal Navigation in Unsteady Flows through Reinforcement Learning

AgentsDGX agent

arXiv:2607.13553v1 Announce Type: cross Abstract: Autonomous robotic navigation in nonstationary time-varying fluid flows remains a fundamental challenge due to partial observability and the unpredict

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

Model ReleasesDGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools

AgentsDGX agent

arXiv:2607.13115v1 Announce Type: new Abstract: Small language models (SLMs) have shown promise for zero-shot molecular property prediction from SMILES strings, yet they often suffer from structural b

15 Jul 2026

Bulkhead: Automated Semantic Detection and Remediation of Container Escape Vulnerabilities

AgentsDGX agent

arXiv:2607.12723v1 Announce Type: cross Abstract: Filesystem isolation in container ecosystems is often weakened by cross-boundary path misresolution, causing path traversal (PaTra) vulnerabilities. T

Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?

SafetyDGX agent

arXiv:2607.12631v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents in high-stakes domains, understanding contextual factors that may modul

Graph Feedback Controls Consensus and Clique Formation in Open-Weight Language-Model Populations

Local AiDGX agent

arXiv:2607.12077v1 Announce Type: new Abstract: Multi-agent language-model systems increasingly route local interactions, yet the runtime interaction graph is often treated as an implementation detail

Interrupt is hitting the road this fall: 📍 NYC- September 24 📍 London- October 13 Info + RSVP: https://interrupt.langchain.com/

AgentsDGX agent

Interrupt is hitting the road this fall: 📍 NYC- September 24 📍 London- October 13 Info + RSVP: https://interrupt.langchain.com/ At Interrupt, @RCUmmadisetti and @kordelfrance from @Toyota's enterprise

PFAdapter: Hierarchical LoRA Decomposition for Personalized Federated MLLMs

Model ReleasesDGX agent

arXiv:2607.12111v1 Announce Type: cross Abstract: Agentic AI systems are reshaping communications and networking by deploying autonomous intelligent agents capable of collaborative learning while main

Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap

SafetyDGX agent

arXiv:2607.12113v1 Announce Type: cross Abstract: One year ago, the AISLE roadmap argued that autonomous laboratories operated as isolated islands and proposed a grassroots network organized around fi

13 Jul 2026

datasette code-frequency chart on GitHub

Model ReleasesDGX agent

datasette code-frequency chart on GitHub Out of curiosity I decided to see if I could find a useful illustration of the impact of coding agents and Opus 4.5 class models on my own output. The best I'v

11 Jul 2026

This Week in Replit Three big drops this week: 1) Community Profiles launched, proof of work for vibe coders 2) A free custom domain, on us,…

AgentsDGX agent

This Week in Replit Three big drops this week: 1) Community Profiles launched, proof of work for vibe coders 2) A free custom domain, on us, through July 17 3) Ramp for Agents, incorporate and run you

10 Jul 2026

Grok Build gets better every day and we love hearing user feedback for improvements

AgentsDGX agent

Grok Build gets better every day and we love hearing user feedback for improvements Grok Build is crazy. 先不管GPT-5.6是不是release到底好不好用。 先來大大稱讚一下 @grok 的 Grok Build,目前唯一集大成的 coding agentic workflow。 Grok

Grok Build improves almost every day

AgentsDGX agent

Grok Build improves almost every day some special features in Grok Build if you're new /dashboard: shows you every agent running in your TUI. no need to tab jump. can click and respond. /imagine: crea

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning

AgentsDGX agent

arXiv:2607.08647v1 Announce Type: cross Abstract: As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions

9 Jul 2026

Learning Spatiotemporal Tubes for Full Class of Signal Temporal Logic Tasks for Control of Unknown Systems under Input Constraints

AgentsDGX agent

arXiv:2607.07136v1 Announce Type: new Abstract: This paper presents a Spatiotemporal Tube (STT)-based control framework for general unknown nonlinear Euler-Lagrange (EL) systems subject to input const

Setup evals with… terraform??

AgentsDGX agent

Setup evals with… terraform?? You've heard of Infrastructure as Code- but agent evals can now ride your existing Terraform setup! I've been using the new LangSmith Terraform provider to auto-provision

8 Jul 2026

Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations

AgentsDGX agent

arXiv:2607.05863v1 Announce Type: new Abstract: Negotiation is a fundamental strategic interaction in management science, characterized by agents attempting to reach agreements while protecting privat

When Assisting One Disempowers Another

AgentsDGX agent

arXiv:2511.04177v2 Announce Type: replace Abstract: Personal AI agents are increasingly deployed in shared environments, where their actions affect not just the primary user they are assisting, but by

7 Jul 2026

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

AgentsDGX agent

arXiv:2607.05155v1 Announce Type: new Abstract: Pretraining scaling laws reveal that model capability improves predictably with data and compute. But learning from real world environments after deploy

FOI-O: An NZ-first ontology and verification methods package for Freedom of Information process modelling

AgentsDGX agent

arXiv:2607.02947v1 Announce Type: cross Abstract: Public official-information request records contain process signals. They can support research, workflow review, and human-supervised agent help. Yet

Resilient by Design -- Active Inference for Distributed Continuum Intelligence

AgentsDGX agent

arXiv:2511.07202v3 Announce Type: replace-cross Abstract: Failures are the norm in highly complex and heterogeneous devices spanning the distributed computing continuum (DCC), from resource-constraine

The Language of Bargaining: Linguistic Effects in LLM Negotiations

AgentsDGX agent

arXiv:2601.04387v2 Announce Type: replace Abstract: Negotiation is a core component of social intelligence, requiring agents to balance strategic reasoning, cooperation, and social norms. Recent work

5 Jul 2026

私もICMLに行きます!もし行く方がいらっしゃったらぜひ現地でお会いしましょう😀

AgentsDGX agent

私もICMLに行きます!もし行く方がいらっしゃったらぜひ現地でお会いしましょう😀 Sakana AI is heading to #ICML2026 in Seoul (July 6–11)! 🐟🇰🇷 Our team will present 11 papers spanning multi-agent coordination, sparse and efficient LLMs, test-

3 Jul 2026

A Practice Auditing Framework for Large Language Model Use: Collective Empiricism, Pseudo-Rational Cognition, and Governance of AI-Generated Content

AgentsDGX agent

arXiv:2607.01248v1 Announce Type: cross Abstract: Large language models are increasingly used for knowledge acquisition, code generation, academic writing, and agent-based automation. In these setting

Prompt Coverage Adequacy

AgentsDGX agent

arXiv:2607.02057v1 Announce Type: cross Abstract: In recent years, it has become increasingly evident that large language models (LLMs) and autonomous agents raise the level of abstraction in software

2 Jul 2026

15 examples of real-world challenges: Insights from the AWS Summit Washington, D.C. event

AgentsDGX agent

As organizations race to operationalize generative and agentic artificial intelligence, the conversation is shifting from pilots and proofs of concept to real-world AI deployments that deliver measura

A Contextual-Bandit Oversight Game with Two-Sided Informational Asymmetry

AgentsDGX agent

arXiv:2607.00155v1 Announce Type: new Abstract: We study runtime human oversight of an AI agent when private information runs in both directions: the human privately knows her reward function, while t

🎧A look at how we built LangSmith Engine with @hwchase17 + @bentannyhill

AgentsDGX agent

🎧A look at how we built LangSmith Engine with @hwchase17 + @bentannyhill Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hu

Fun chat with @hwchase17 talking about our work on Engine!

AgentsDGX agent

Fun chat with @hwchase17 talking about our work on Engine! Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hunts through yo

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @…

AgentsDGX agent

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @badlogicgames! Coding agents are real users of the @huggingf

1 Jul 2026

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

Model ReleasesDGX agent

arXiv:2606.31551v1 Announce Type: new Abstract: Training language models (LMs) remains a highly human-intensive process, even as frontier language model agents become increasingly capable at software

Can LLMs Imagine Moral Alternatives Beyond Binary Dilemmas?

AgentsDGX agent

arXiv:2606.31213v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as moral advisors and agents, they need to address dilemmas between two competing values. Ho

ComAct: Reframing Professional Software Manipulation via COM-as-Action Paradigm

Model ReleasesDGX agent

arXiv:2606.13239v2 Announce Type: replace-cross Abstract: Existing computer-use agents remain fundamentally limited in professional software manipulation: GUI-based agents suffer from fragile visual g

DDIAgents: Mechanism-Conditioned Context Flow for Drug-Drug Interaction Prediction

SafetyDGX agent

arXiv:2606.31085v1 Announce Type: new Abstract: Drug-drug interaction (DDI) prediction is essential for medication safety, yet it requires reasoning over heterogeneous biomedical evidence whose releva

Get started with the Claude apps gateway for Google Cloud

Model ReleasesDGX agent

Anthropic's agentic coding tool Claude Code has worked with Google Cloud for a while now. An individual developer could easily point CLAUDE_CODE_USE_VERTEX=1 at a Google Cloud (GCP) project, grant the

Introducing Devin Security Swarm A more cost effective and accurate way to find security vulnerabilities in complex codebases, based on a ne…

AgentsDGX agent

Cognition AI announced Devin Security Swarm, a new approach to vulnerability detection that uses multiple AI agents working together to identify security issues in complex codebases more cost-effectiv

LLM-Driven Personalities for Decision Making in Emergency Simulations

AgentsDGX agent

arXiv:2606.31038v1 Announce Type: cross Abstract: For virtual humans to appear believable, they must exhibit agency and spatial awareness while interacting with their environment in ways that reflect

What Drives Interactive Improvement from Feedback?

AgentsDGX agent

arXiv:2606.30774v1 Announce Type: new Abstract: We study when natural-language feedback produces improvement beyond the gains obtainable from repeated attempts alone. In multi-turn language agent sett

30 Jun 2026

Fuzzing Large Language Models to Elicit Hidden Behaviours

AgentsDGX agent

arXiv:2606.29646v1 Announce Type: cross Abstract: Sleeper agents are the canonical model organism of deception: models trained to behave normally but to emit an unsafe behaviour on a specific trigger.

Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects

AgentsDGX agent

arXiv:2603.03915v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown remarkable potential in developing role-playing agents (RPAs). However, current evaluation frameworks

29 Jun 2026

Dynamic subagents in deepagents! Lets you spin up subagents programmatically We highlight 6 diff use cases for this

AgentsDGX agent

DeepAgents introduces dynamic subagents that can be created programmatically at runtime, enabling more flexible and adaptive agent architectures. The post highlights six different practical use cases

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

Model ReleasesDGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

28 Jun 2026

omg it's basically pi, but more minimal for educational purposes and in python. @marlene_zw is gonna love this. great educational resource, …

AgentsDGX agent

omg it's basically pi, but more minimal for educational purposes and in python. @marlene_zw is gonna love this. great educational resource, alejandro! introducing tau τ — an educational agent harness

26 Jun 2026

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agen…

AgentsDGX agent

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agent.' - Manus AI prompt caching is important! read about how w

Incident Report: CVE-2026-LGTM

AgentsDGX agent

Incident Report: CVE-2026-LGTM Spectacular hypothetical incident report by Andrew Nesbitt. Day 2, 16:00 UTC --- Two AI review agents from competing vendors, both attached to a downstream pull request

Localizing RL-Induced Tool Use to a Single Crosscoder Feature

AgentsDGX agent

arXiv:2606.26474v1 Announce Type: cross Abstract: Fine-tuning through RL reshapes the internal representations of language models to enable agentic behaviors such as tool use, yet the mechanistic basi

25 Jun 2026

This is a fascinating and important set of data which shows us where things are going, using OpenAI as a canary in the coal mine. The chatbo…

AgentsDGX agent

This is a fascinating and important set of data which shows us where things are going, using OpenAI as a canary in the coal mine. The chatbot era is over, and agentic systems are coming to tasks beyon

Transitioning from using a single video model to generate a clip to achieving a finished video is like the transition from coding models tha…

AgentsDGX agent

Transitioning from using a single video model to generate a clip to achieving a finished video is like the transition from coding models that merely autocomplete to coding models that write working so

23 Jun 2026

3D Vessel Reconstruction from Sparse-View Dynamic DSA Images via Vessel Probability Guided Attenuation Learning

AgentsDGX agent

arXiv:2405.10705v3 Announce Type: replace-cross Abstract: Digital Subtraction Angiography (DSA) is one of the gold standards for vascular disease diagnosis. With the help of a contrast agent, time-res

auto-research style proposal loops should be data driven! they largely work best only when Data/Evals/Feedback give a useful gradient to hil…

AgentsDGX agent

auto-research style proposal loops should be data driven! they largely work best only when Data/Evals/Feedback give a useful gradient to hill-climb against increasingly auto-research is a very good to

Empowering Embodied AI in 6G Networks: Architecture, Enablers, and Open Challenges

AgentsDGX agent

arXiv:2606.20592v1 Announce Type: cross Abstract: Embodied artificial intelligence (AI) is emerging as a key driver of the sixth-generation (6G) wireless networks by enabling agents that continuously

open sourced activegraph on May 20: https://x.com/yoheinakajima/status/2057099245430222926

AgentsDGX agent

open sourced activegraph on May 20: https://x.com/yoheinakajima/status/2057099245430222926 i'm excited to open source Active Graph: an event-sourced reactive graph runtime for long-running, agents 🔄🧠

Sim2O: Efficient Offline-to-Online MARL via Joint Action Composition

AgentsDGX agent

arXiv:2606.21085v1 Announce Type: new Abstract: Offline-to-online adaptation serves as a pivotal paradigm for mitigating the prohibitive cost of online exploration by bootstrapping reinforcement learn

SIMSplat: Language-Aligned 4D Gaussian Splatting for Driving Scenario Generation

AgentsDGX agent

arXiv:2510.02469v2 Announce Type: replace-cross Abstract: Driving scene manipulation using real-world sensor data has emerged as a promising alternative to traditional driving simulators. Despite adva

22 Jun 2026

@trycua Computer Use docs: https://hermes-agent.nousresearch.com/docs/user-guide/features/computer-use

AgentsDGX agent

Nous Research's Computer Use documentation provides a user guide for the @trycua tool, detailing features and functionality for AI agents to interact with computer interfaces. The guide likely covers

11 Jun 2026

Resource-Aware LLM Reasoning for Mobile Edge General Intelligence

AgentsDGX agent

arXiv:2509.23248v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has enabled an emergence of agentic artificial intelligence (AI) with powerful reasoning and a

StatefulDiscovery: Evidence-Calibrated Claim Formation in Open-Ended Scientific Discovery

AgentsDGX agent

arXiv:2606.11851v1 Announce Type: new Abstract: Open-ended scientific discovery asks agents to move beyond executing analyses for predefined questions. Across multiple rounds of exploration, a discove

← Previous
1…160161162163164…300
Next →