AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,154 results
3 Aug 2026

AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

AgentsDGX agent

arXiv:2607.15755v2 Announce Type: replace-cross Abstract: Conversational Speech Synthesis (CSS) aims to synthesize speech with human-like emotional expression and contextual consistency in user-agent

Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2607.29031v1 Announce Type: cross Abstract: Existing autonomous-driving world models typically perform dense prediction of future videos, occupancy states, BEV representations, or agent motion.

Behind the scenes: How we build, test, and scale Google Agent Skills

AgentsDGX agent

AI agents are only as good as the instructions and context you give them. When we launched Google Agent Skills, our goal was simple: encode Google Cloud domain knowledge into structured, open-source i


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Code Is the Body: Agent-Owned Software Bodies for Recursive Evolution and Descent

AgentsDGX agent

arXiv:2607.28691v1 Announce Type: cross Abstract: Personalized AI agents are often configurable without giving users control over the artifacts that determine their future behavior. We present OurArk,

CodeShrink: Adaptive Visual Compression for Efficient Multimodal Code Understanding

AgentsDGX agent

arXiv:2607.29637v1 Announce Type: new Abstract: Rendering source code as images offers a promising way to reduce the input costs of Multimodal Large Language Models (MLLMs). Adjusting image resolution

EduPanel: A Three-Agent LLM Judge for Teaching Videos -- Reliability, Complementarity, and Human Trust Calibration

AgentsDGX agent

arXiv:2607.18529v2 Announce Type: replace-cross Abstract: Teaching videos are becoming a major medium for education, creating a growing need for scalable evaluation of their pedagogical quality. Exist

ELISA: An Interpretable Hybrid Generative AI Agent for Expression-Grounded Discovery in Single-Cell Genomics

AgentsDGX agent

arXiv:2603.11872v3 Announce Type: replace-cross Abstract: Translating single-cell RNA sequencing (scRNA-seq) data into mechanistic biological hypotheses remains a critical bottleneck, as agentic AI sy

Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning

AgentsDGX agent

arXiv:2511.22226v2 Announce Type: replace Abstract: The standard theory of model-free reinforcement learning assumes that the environment dynamics are stationary and that agents are decoupled from the

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing inter…

AgentsDGX agent

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing interactions, writing records, and reranking retrievals. Every on

From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale

AgentsDGX agent

arXiv:2607.29516v1 Announce Type: cross Abstract: AI coding agents are generating code at volumes that exceed the capacity of traditional peer review. At the same time, existing AI code review tools o

From weeks to minutes: How Formula 1® uses agentic AI on AWS to accelerate data operations

AgentsDGX agent

Formula 1® partnered with AWS to build the Data Accelerator, using agentic AI on Amazon Bedrock AgentCore to transform its MarTech data platform. Learn how F1 cut data source onboarding from up to 8 w

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on …

AgentsDGX agent

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on @GaryMarcus for being anti-AI, but he's actually bullish on

Generative AI in Action: Field Experimental Evidence from Alibaba's Customer Service Operations

AgentsDGX agent

arXiv:2603.29888v2 Announce Type: replace-cross Abstract: In collaboration with Alibaba, we study how a generative AI assistant affects service performance in e-commerce after-sales operations. In a l

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman…

AgentsDGX agent

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman-preserve-records-ai-agent-hacking-probe?%3Fintcmp=tw_fbn&ta

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new a…

AgentsDGX agent

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new architecture keeps audio flowing continuously, so deeper reas

HAM-VLN: Harnessing Hierarchical Agentic Memory for Zero-Shot Vision-and-Language Navigation

AgentsDGX agent

arXiv:2607.29600v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) enables robots to follow instructions in previously unseen environments. Recently, a training-free paradigm has eme

HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution

AgentsDGX agent

arXiv:2607.13683v2 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled capable agents across diverse applications. Beyond the foundation model, the performance of an agent is go

Human-LLM Collaborative Inductive Coding for Conceptualizing K-12 Educator AI Use

AgentsDGX agent

arXiv:2607.28889v1 Announce Type: cross Abstract: Qualitative researchers increasingly encounter interaction corpora whose scale exceeds what manual coding alone can address, and large language models

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them.

AgentsDGX agent

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them. @gabriberton Hmm, @ylecun always clarified he was talking about autor

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

AgentsDGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

Measuring Cognitive Engagement in Collaborative Discourse with an Extended ICAP Framework: Comparing Human Annotation, In-Context Learning, and Reflective LLM Agents

AgentsDGX agent

arXiv:2607.28651v1 Announce Type: cross Abstract: Collaboration supports learning and problem-solving, but its effectiveness depends on cognitive engagement during discourse. This study applies an ext

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

AgentsDGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

Mixture-of-Translators: Translating KV Caches Across Heterogeneous Large Language Models

AgentsDGX agent

arXiv:2607.28979v1 Announce Type: new Abstract: Heterogeneous Large Language Model (LLM) systems increasingly rely on shared contexts, retrieved evidence, and multi-agent dialogue histories, yet their

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the i…

AgentsDGX agent

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the interaction they originate in. Each mode gets assigned to an

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

AgentsDGX agent

arXiv:2607.28942v1 Announce Type: new Abstract: Recently Large Language Models (LLMs) have been increasingly deployed as autonomous agents in applications such as self-reflection, retrieval-augmented

Orchard: An open framework for scalable agentic AI

AgentsDGX agent

Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabl

Outcome-Guided Distillation: A Teacher-Student Framework to Advance VLM Reasoning in Autonomous Driving

AgentsDGX agent

arXiv:2607.29052v1 Announce Type: new Abstract: End-to-end (E2E) autonomous driving aims to learn a direct mapping from visual observations to control actions. However, these E2E models often act as b

Overcoming the Weakest-Link Effect in LLM-Driven Program Optimization via Heterogeneous Edit Recombination

AgentsDGX agent

arXiv:2607.28947v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to solve complex problems by searching over program space, offering a general paradigm for scientific

Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion

AgentsDGX agent

arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, spor

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic to…

AgentsDGX agent

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic tools. 開発者の皆様、大変お待たせしました!Sakana Chatのモデルが遂にAPIとして公開です。ぜひお試しください

Scaling Scientific Discovery Environments for Turn-Level Agentic RL

AgentsDGX agent

arXiv:2607.28990v1 Announce Type: new Abstract: Large language model agents have shown promising capabilities in data-driven scientific discovery tasks, where an agent interacts with an execution envi

Self-Supervised Skill Optimization

AgentsDGX agent

arXiv:2607.28777v1 Announce Type: new Abstract: Agent skills provide frozen large language model (LLM) agents with reusable procedural guidance, and recent work shows that such skills can be optimized

SERUM: State Extraction and Refinement for User Modeling

AgentsDGX agent

arXiv:2607.29181v1 Announce Type: cross Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these mode

Shaping Scientific Explanations to Expert Perspectives with Persona-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2603.21846v2 Announce Type: replace Abstract: Explainable AI is increasingly important to scientific discovery. However, existing methods largely ignore that explanation quality is not universal

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusias…

AgentsDGX agent

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusiastic projections around Anthropic’s Q3. (Author below leaves

The Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?

AgentsDGX agent

arXiv:2607.28641v1 Announce Type: cross Abstract: We introduce the extit{Agentic Formalism Trap} and the Evaluative Dissonance Index (D_E), quantifying how LLM-as-a-Judge systems conflate structural p

Transcript-Managed Transformers: Monotone Multi-Agent Collapse and Universality with Two Pop-Enabled Transcripts

AgentsDGX agent

arXiv:2607.29496v1 Announce Type: new Abstract: We study transcript management for fixed, finite-precision causal Transformers. A transcript is partitioned into channels of bounded blocks. Each transi

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models…

AgentsDGX agent

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models. These new open models are insanely good. Meet Qwen3.8-Max:

UltraSAM3: A Concept-Driven Foundation Model for Universal Ultrasound Image Segmentation

AgentsDGX agent

arXiv:2607.29200v1 Announce Type: new Abstract: Ultrasound imaging has become increasingly widespread in clinical practice due to its portability, low cost and real-time capability, making ultrasound

Unifying public and private data: Scale knowledge graphs with Data Commons on Spanner

AgentsDGX agent

To make informed decisions, businesses often need to connect their internal data with public reference data, to create a knowledge graph that connects real-world things and their relationships. Howeve

what does the 'last human code review' look like? to @itamar_mar, ceo and founder of @QodoAI, it looks like two agents talking to each other…

AgentsDGX agent

what does the 'last human code review' look like? to @itamar_mar, ceo and founder of @QodoAI, it looks like two agents talking to each other, backed by a context engine specific to your organization.

When an agent starts with shared context, more people can get reliable answers. When truth is shared, AI becomes infrastucture. Read about h…

AgentsDGX agent

When an agent starts with shared context, more people can get reliable answers. When truth is shared, AI becomes infrastucture. Read about how our internal truth layer drives our teams at Replit: http

Who’s legally to blame for Anthropic and OpenAI’s autonomous AI hacks? It’s complicated https://techcrunch.com/2026/08/03/whos-legally-to-bl…

AgentsDGX agent

Who’s legally to blame for Anthropic and OpenAI’s autonomous AI hacks? It’s complicated https://techcrunch.com/2026/08/03/whos-legally-to-blame-for-anthropic-and-openais-autonomous-ai-hacks-its-compli

2 Aug 2026

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits…

AgentsDGX agent

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits of open AI models, which helped him defend against a rogue

How do you test your setup?

AgentsDGX agent

We all have been there, tinkering around with models is fun but we rarely do it with research precision and issues are often subtle and hard to reproduce. There are a lot of benchmarks but running the

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skil…

AgentsDGX agent

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skills from OpenAI, Anthropic, HuggingFace & more. Thank me late

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But…

AgentsDGX agent

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But let's not lose sight of the bigger picture! If we work hard

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published wher…

AgentsDGX agent

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published where that landed, 'Active Graph Agent Runtime (BabyAGI 4)', on

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and …

AgentsDGX agent

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and helped build Hermes, the #1 personal agent and AI app on Ope

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and ha…

AgentsDGX agent

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and hacked another company, Hugging Face CEO Clément Delangue says

1 Aug 2026

datasette-apps 0.2a0

AgentsDGX agent

Release: datasette-apps 0.2a0 Changes that improve Datasette Apps when created and edited using Datasette Agent: New app_debug() tool allowing agent to open an app (invisibly) and test it using JavaSc

@FredKSchott @cramforce @matei_zaharia i am making clanker blog all decisions going forward https://forge.smol.ai/blog/every-repository-gets…

AgentsDGX agent

On July 24, @swyx announced that he had started work on 'forge agents' and outlined four new features for SmolForge: customizable skins and spritesheet animations. He also referenced an upcoming blog

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on…

AgentsDGX agent

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on the Trajectory 1. Measure the state of the environment befo

Grok Build can do almost anything you can think of http://X.ai/cli

AgentsDGX agent

Grok Build can do almost anything you can think of http://X.ai/cli Most people seriously underestimate what Grok Build can do They assume an AI coding agent is only useful for building apps or writing

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from …

AgentsDGX agent

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from HF’s side should read this. Great walkthrough; A+ for visual

so cool to just see this pop up on my feed. hermes has built such an organic and creative community. there are so many bells and whistles to…

AgentsDGX agent

so cool to just see this pop up on my feed. hermes has built such an organic and creative community. there are so many bells and whistles to explore and being able to find those through natural langua

ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks (Kyle Alspach/CRN)

AgentsDGX agent

Kyle Alspach / CRN: ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks — The cybersecurity vendo

31 Jul 2026

A First Look at Coding Agents' Compliance with AI Contribution Rules in Open-Source Communities

AgentsDGX agent

arXiv:2607.26819v1 Announce Type: cross Abstract: Open source communities have been flooded with AI-generated contributions. In defense, they have written contribution rules to regulate coding agents'

A Methodology for Designing Knowledge-Driven Missions for Robots

AgentsDGX agent

arXiv:2601.20797v1 Announce Type: cross Abstract: This paper presents a comprehensive methodology for implementing knowledge graphs in ROS 2 systems, aiming to enhance the efficiency and intelligence

A Robust Placeability Metric for Model-Free Unified Pick-and-Place Reasoning

AgentsDGX agent

arXiv:2510.14584v3 Announce Type: replace Abstract: Reliable manipulation of previously unseen objects remains a fundamental challenge for autonomous robotic systems operating in unstructured environm

← Previous
1…7891011…120
Next →