AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
7 Apr 2026

tldr: everyone is converging on the same product shape: a general harness that takes a goal, uses tools, and does knowledge work. once every…

AgentsDGX agent

tldr: everyone is converging on the same product shape: a general harness that takes a goal, uses tools, and does knowledge work. once every product is a harness, the next frontier is the feedback loo

14 Aug 2026

Heterogeneity-Aware Belief Synchronization for Semantic Communication in AI-Native 6G Networks

Local AiDGX agent

arXiv:2608.13394v1 Announce Type: cross Abstract: 6G networks will not be serving as communication infrastructures only; rather, they are expected to evolve into intelligent systems, where thousands o

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty

TsuGO: Probing Search Efficiency in LLM Reasoning via Go Life-and-Death Problems

Model ReleasesDGX agent

arXiv:2608.13221v1 Announce Type: new Abstract: The evaluation of LLM reasoning is moving from final-answer accuracy to process-level assessment, yet existing methods still fail to capture how models

13 Aug 2026

GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Settings

AgentsDGX agent

arXiv:2608.12133v1 Announce Type: new Abstract: Enterprise guideline documents are heterogeneous and multimodal, combining narrative text, complex tables, and embedded images. Existing LLM and VLM sys

OpenAg: Democratizing Agricultural Intelligence

AgentsDGX agent

arXiv:2506.04571v3 Announce Type: replace Abstract: Agriculture is undergoing a major transformation driven by artificial intelligence (AI), machine learning, and knowledge representation technologies

Preference Tree Optimization: Enhancing Goal-Oriented Dialogue with Look-Ahead Simulations

AgentsDGX agent

arXiv:2608.12062v1 Announce Type: cross Abstract: Developing dialogue systems capable of engaging in multi-turn, goal-oriented conversations remains a significant challenge, especially in specialized

REVERE: Reflective Evolving Research Engineer

AgentsDGX agent

arXiv:2603.20667v2 Announce Type: replace-cross Abstract: Existing prompt-optimization techniques rely on local signals, causing poor generalization across tasks. In addition, they also rely on weak u

Self-Harness: Harnesses That Improve Themselves

Model ReleasesDGX agent

arXiv:2606.09498v2 Announce Type: replace Abstract: The performance of LLM-based agents is jointly shaped by their base models and the harnesses that mediate their interaction with the environment. Be

Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)

AgentsDGX agent

arXiv:2608.11229v1 Announce Type: new Abstract: Comparative feedback, asking people which of two behaviors they prefer, has become a standard way to align robot and agent behavior with human intent wh

The @latentspacepod team and I will be speaking at Mongodb's .local buildfest today 10a-4pmish, come by! https://www.mongodb.com/events/mong…

AgentsDGX agent

The @latentspacepod team and I will be speaking at Mongodb's .local buildfest today 10a-4pmish, come by! https://www.mongodb.com/events/mongodb-local/build-fest We'll be covering: - tba Agent infra wi

12 Aug 2026

CodeRabbit bags $143M to help companies get a grip on the explosion of AI-generated code

AgentsDGX agent

CodeRabbit Inc., the creator of a popular tool that automatically reviews artificial intelligence-generated code, is becoming more ambitious after closing on its latest 143 million Series C round of f

Inferential Capability Does Not Determine Legal Scope

AgentsDGX agent

arXiv:2608.10601v1 Announce Type: cross Abstract: Two instruments of EU digital law place inference at their centre and mean different things by it. Article 3(1) of the AI Act uses the capability to i

Operationalising Relative Causal Knowledge: Backbone Identifiability from Private Reports on a Shared Outcome

SafetyDGX agent

arXiv:2608.10664v1 Announce Type: new Abstract: The Relativity of Causal Knowledge (RCK) explains how a network of agents with different structural causal models can exchange causal knowledge through

Risk-Aware Kinodynamic Motion Planning Under Uncertainty For Safe Navigation on Planetary Environments

AgentsDGX agent

arXiv:2608.11175v1 Announce Type: new Abstract: For autonomous space exploration, robotic agents need to perform motion planning in which environmental interactions may be unknown. Learning these inte

SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy Distillation

Model ReleasesDGX agent

arXiv:2608.10775v1 Announce Type: new Abstract: Computer-using agents can perceive rich software interfaces, yet their decisions often lack visual procedural memory: they may recognize individual cont

11 Aug 2026

From Prompt to Harness: Coderlet from Scratch

AgentsDGX agent

arXiv:2608.09480v1 Announce Type: new Abstract: A model alone does not determine how a programming agent acts. What the model sees, how actions enter the environment, how feedback returns, and how one

SemPIC: Learning Semantic Position-Independent KV Caches

AgentsDGX agent

arXiv:2607.28069v2 Announce Type: replace Abstract: Long-context retrieval and agentic workloads repeatedly reuse the same documents under changing instructions, histories, and document orders. Prefix

SiriusDeliver: Automating Data Warehouse Delivery at Tencent

AgentsDGX agent

arXiv:2608.09185v1 Announce Type: cross Abstract: Enterprise data warehouses (DWs) support business-critical analytics, but warehouse task delivery remains a complicated production process involving c

The Belief-Desire-Intention Ontology for modelling mental reality and agency

AgentsDGX agent

arXiv:2511.17162v2 Announce Type: replace Abstract: The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, it

10 Aug 2026

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence

Model ReleasesDGX agent

arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inherently iterative: each action reshapes

Counterfactual Shapley Credit Assignment

SafetyDGX agent

arXiv:2607.16999v2 Announce Type: replace-cross Abstract: The Credit Assignment Problem (CAP) is fundamental to developing efficient and explainable Reinforcement Learning (RL) agents. Existing framew

Learning Suffers More Than the Policy Class Under Partial Observability: A Closed-Form Analysis

SafetyDGX agent

arXiv:2608.07228v1 Announce Type: new Abstract: When a reinforcement learning agent cannot observe the full state, we usually blame its policies: it cannot see enough to represent a good one. We show

ResidencyRL: Reinforcement Learning in Simulated Clinical Environments

Model ReleasesDGX agent

arXiv:2608.07418v1 Announce Type: new Abstract: In medical education, physicians convert academic knowledge into clinical expertise through residency: years of training across thousands of encounters,

Strategy-first synthesis planning for complex natural products

AgentsDGX agent

arXiv:2608.07454v1 Announce Type: cross Abstract: The total synthesis of a complex molecule is among the most demanding intellectual and experimental feats in chemistry: a chemist must plan many steps

9 Aug 2026

i shipped first set of llm-as-judge evals for the kill my saas competition tonight. people can run this to check if their solutions at least…

AgentsDGX agent

i shipped first set of llm-as-judge evals for the kill my saas competition tonight. people can run this to check if their solutions at least pass the sniff test. 10,000 kill my saas in a weekend compe

not wrong!

AgentsDGX agent

In a recent Twitter exchange, Luca Ambrogioni stated that large‑language models (LLMs) likely possess fundamental limitations that may be obscured by progress in reasoning and agentic pipelines, thoug

7 Aug 2026

Unifying Structured and Unstructured Data Insights with BQ Search Innovations

Model ReleasesDGX agent

Modern enterprises possess a vast amount of unstructured data, yet they frequently encounter significant challenges in managing and extracting value from it. Historically, unlocking the insights hidde

6 Aug 2026

Deliberate Before You Fly: Vision-Guided Spatial Deliberation for UAV See-and-Reach Navigation

AgentsDGX agent

arXiv:2608.04825v1 Announce Type: new Abstract: UAV see-and-reach navigation requires an aerial agent to approach a language-specified target visible in its initial view and stop reliably near it. Exi

5 Aug 2026

AWS partners with Anthropic and OpenAI to bring Continuum into coding tools

Model ReleasesDGX agent

Amazon Web Services Inc. today said it has partnered with Anthropic PBC and OpenAI Group PBC to wire AWS Continuum for code vulnerabilities directly into the tools developers write code in. The integr

Cloudflare launches Identity-Aware AI Gateway to track who is using AI

Model ReleasesDGX agent

Cloudflare Inc. today launched Identity-Aware AI Gateway, a service that attaches a verified identity to every artificial intelligence request leaving a company network. Information technology and sec

Most AI tools only do one part of building a product. One researches. One designs. One writes code. One generates images. One deploys. You s…

AgentsDGX agent

Most AI tools only do one part of building a product. One researches. One designs. One writes code. One generates images. One deploys. You still have to connect everything yourself. II-Agent is differ

ToolLIFT: Lifting Tool-Specific Trajectories into Function-Level Graphs for Generalizable Tool Planning

AgentsDGX agent

arXiv:2608.03468v1 Announce Type: new Abstract: Historical tool-use trajectories provide valuable experience for large language model (LLM) agents to plan and coordinate tool usage. Existing approache

4 Aug 2026

A Few Neurons Reveal When LLMs Misuse Tools: Sparse Detection and Selective Steering for Reliable Tool Use

Model ReleasesDGX agent

arXiv:2608.00218v1 Announce Type: new Abstract: Agentic LLMs exhibit three consequential tool-use failures: invalid arguments (validity), unnecessary calls (over-calling), and omitted calls when tools

CompanionBench: A Theory-Anchored, Real-World-Grounded Benchmark for AI Emotional Companionship

Model ReleasesDGX agent

arXiv:2608.02046v1 Announce Type: new Abstract: LLM companions are deployed at scale in personally consequential settings, yet poorly evaluated. Existing benchmarks use hand-authored scenarios and pro

DocNavRAG: Document-Structured Graph RAG with Stateful Evidence Construction for Complex Document Question Answering

AgentsDGX agent

arXiv:2608.01565v1 Announce Type: new Abstract: Answering complex questions over large document collections requires assembling complementary evidence across sections and documents. GraphRAG offers st

OmniAI: A Surface-Adaptive Aerial Projection Interface for Human--Drone Interaction

AgentsDGX agent

arXiv:2608.00721v1 Announce Type: new Abstract: Drones in human environments often lack spatially grounded in- terfaces for situated communication. We present OmniAI, an em- bodied aerial agent that s

SyncPlan: Long-Horizon LLM Coordination with Explicit Synchronization and Adaptive Correction

Model ReleasesDGX agent

arXiv:2608.01652v1 Announce Type: new Abstract: LLM-based multi-agent coordination faces a fundamental trade-off between efficiency and adaptivity in dynamic environments. Existing approaches typicall

UEmbed: Unified Sparse and Dense Multimodal Embeddings

AgentsDGX agent

arXiv:2608.02583v1 Announce Type: cross Abstract: Sparse retrieval underpins modern search systems, from web search to retrieval-augmented generation. Existing work has introduced Learned Sparse Retri

3 Aug 2026

Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery

Model ReleasesDGX agent

arXiv:2603.03322v2 Announce Type: replace-cross Abstract: Recent advancements in Large Language Model (LLM) agents have demonstrated remarkable potential in automatic knowledge discovery. However, rig

Learning Stateful Predictive Knowledge From Experience

SafetyDGX agent

arXiv:2607.28638v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly learn from experience, they primarily rely on trajectory-level reflection to extract insights. Viewed

Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion

AgentsDGX agent

arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, spor

2 Aug 2026

How do you test your setup?

AgentsDGX agent

We all have been there, tinkering around with models is fun but we rarely do it with research precision and issues are often subtle and hard to reproduce. There are a lot of benchmarks but running the

1 Aug 2026

@FredKSchott @cramforce @matei_zaharia i am making clanker blog all decisions going forward https://forge.smol.ai/blog/every-repository-gets…

AgentsDGX agent

On July 24, @swyx announced that he had started work on 'forge agents' and outlined four new features for SmolForge: customizable skins and spritesheet animations. He also referenced an upcoming blog

31 Jul 2026

Strategy, Not Payoffs: A Behavioural Embedding of Normal-Form Games

AgentsDGX agent

arXiv:2607.27536v1 Announce Type: cross Abstract: Learning a strategic task changes more than what is directly taught: fine-tuning on one game can either enhance or degrade an agent's ability to reaso

The Role of Causality in Algorithmic Recourse

AgentsDGX agent

arXiv:2607.28497v1 Announce Type: new Abstract: Algorithmic recourse aims to provide individuals with actionable changes to improve their predicted outcomes in high-stakes classification settings, suc

Together AI gives developers a high-throughput production path for running Inkling-Small on @NVIDIAAI Accelerated Infrastructure across mult…

AgentsDGX agent

Together AI gives developers a high-throughput production path for running Inkling-Small on @NVIDIAAI Accelerated Infrastructure across multimodal, coding, and agentic workloads. Start building: https

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that comm…

AgentsDGX agent

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that commoncrawl isn't good enough for you, you have to build a Whole

30 Jul 2026

Minimal Markovization via Stable Quotients in Holonomy-Cover Decision Processes

AgentsDGX agent

arXiv:2607.27132v1 Announce Type: new Abstract: An agent acting under partial observability must retain a recursively updateable statistic of history that restores the Markov property, but the smalles

Top-k Pareto Bandits: Hypervolume Regret for Multi-Objective Slate Selection

AgentsDGX agent

arXiv:2607.26273v1 Announce Type: new Abstract: We consider a stochastic multi-objective bandit problem where, at each round, the agent selects a slate of k arms and observes their d-dimensional rewar

29 Jul 2026

Distilling Temporal Search and Reasoning: Evolving LLMs for Future Prediction via Harness-Assisted Efficient Data Synthesis

AgentsDGX agent

arXiv:2607.25554v1 Announce Type: new Abstract: Future event prediction carries broad social impact yet remains challenging. SOTA approaches augment LLMs with external agent frameworks whose predictiv

Pictura: Perspective-View Self-Play at Scale for Driving

SafetyDGX agent

arXiv:2607.26005v1 Announce Type: cross Abstract: Self-play in simulation produces robust driving policies at scale. Demonstrations of such behavior have been made using privileged vectorized observat

Replit Design reimagines the UX for AI design with Ambient Intelligence. You don’t need to prompt. You don’t need design language. At every …

AgentsDGX agent

Replit Design reimagines the UX for AI design with Ambient Intelligence. You don’t need to prompt. You don’t need design language. At every step the agent suggests next best actions that you can take

28 Jul 2026

A funny companion: Distinct neural responses to AI- versus human-attributed humor

AgentsDGX agent

arXiv:2509.10847v3 Announce Type: replace-cross Abstract: As artificial intelligence (AI) companions become capable of human-like communication, including telling jokes, understanding how people cogni

Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines

AgentsDGX agent

arXiv:2607.24692v1 Announce Type: cross Abstract: Inference systems increasingly combine a fast path that returns predictions within the application's latency deadline together with a higher-accuracy

Eviction as Estimation: A Fixed-Lag Smoothing View of Test-Time Memory, and When Measuring Beats Accumulating

SafetyDGX agent

arXiv:2607.24667v1 Announce Type: new Abstract: A language model with a bounded working memory must repeatedly decide which stored items to keep. Every deployed method decides the moment an item arriv

Exclusive: Dymium introduces single gateway to govern enterprise AI use

Model ReleasesDGX agent

Secure artificial intelligence infrastructure startup Dymium Inc. today introduced GhostAI, a gateway designed to apply security and governance policies across the models, data, context and tools used

Greedy dynamical meta-learning

AgentsDGX agent

arXiv:2607.23925v1 Announce Type: new Abstract: Gradient descent scales well to large models, but becomes unstable over long time horizons. Gradient-free optimizers can scale to arbitrary timespans, b

HiLLTS: Zero-Shot Hierarchical LLM-Guided Traffic Signal Control for Sustainable Transportation

AgentsDGX agent

arXiv:2607.22691v1 Announce Type: new Abstract: Urban traffic congestion significantly increases fuel consumption, greenhouse gas emissions, and commuter delays, resulting in substantial economic loss

@mitsuhiko Turns out the 'third-party provider' with the sandbox that was used for the attack was Modal, though they blame one of their cust…

AgentsDGX agent

@mitsuhiko Turns out the 'third-party provider' with the sandbox that was used for the attack was Modal, though they blame one of their customers for deploying an endpoint without authentication: http

← Previous
1…164165166167168…300
Next →