AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
4 Aug 2026

From Pixels to PCells: A Neurosymbolic Approach to Photonic Component Creation

ResearchDGX agent

arXiv:2608.00084v1 Announce Type: new Abstract: We present PixCell, a neurosymbolic system in which multimodal agents convert a visually presented photonic component into a parametric program over a s

Local-Canonicalization Equivariant Graph Neural Networks for Sample-Efficient and Generalizable Swarm Robot Control

SafetyDGX agent

arXiv:2509.14431v2 Announce Type: replace Abstract: Multi-agent reinforcement learning (MARL) policies for swarm control often learn inefficiently and generalize poorly across coordinate frames, team

No One Wins in Nuclear War: A Social Simulation of Military Decision-making

SafetyDGX agent

arXiv:2608.01868v1 Announce Type: cross Abstract: WOPR is a social-simulation environment for studying how organizations make high-stakes decisions, built on a deterministic, replay-validated rules en

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stress-Relief Annealing: Polynomial-Time Simulation-Free Layout Optimization for Automated Warehouses

AgentsDGX agent

arXiv:2608.01024v1 Announce Type: cross Abstract: We study the problem of optimizing physical layouts for automated warehouses, where hundreds to thousands of robots are coordinated to transport packa

this is a really clean and unique ai tool that acts more like an extension of you took me 15 min to set up a custom ai voice assistant w/ it…

AgentsDGX agent

this is a really clean and unique ai tool that acts more like an extension of you took me 15 min to set up a custom ai voice assistant w/ it’s own # that also picks up my calls when my phone is off an

3 Aug 2026

CodeShrink: Adaptive Visual Compression for Efficient Multimodal Code Understanding

AgentsDGX agent

arXiv:2607.29637v1 Announce Type: new Abstract: Rendering source code as images offers a promising way to reduce the input costs of Multimodal Large Language Models (MLLMs). Adjusting image resolution

Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation

SafetyDGX agent

arXiv:2604.05150v2 Announce Type: replace-cross Abstract: We study compiled AI, a paradigm in which large language models generate executable code artifacts during a compilation phase, after which wor

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on …

AgentsDGX agent

@gabriberton Hmm, @ylecun always clarified he was talking about autoregressive symbol prediction. In a similar vein, people like to dump on @GaryMarcus for being anti-AI, but he's actually bullish on

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

Model ReleasesDGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them.

AgentsDGX agent

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them. @gabriberton Hmm, @ylecun always clarified he was talking about autor

Overcoming the Weakest-Link Effect in LLM-Driven Program Optimization via Heterogeneous Edit Recombination

AgentsDGX agent

arXiv:2607.28947v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to solve complex problems by searching over program space, offering a general paradigm for scientific

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusias…

AgentsDGX agent

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusiastic projections around Anthropic’s Q3. (Author below leaves

2 Aug 2026

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But…

AgentsDGX agent

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But let's not lose sight of the bigger picture! If we work hard

1 Aug 2026

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the mode…

Model ReleasesDGX agent

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the model. Same intelligence, wildly different accuracy and cost, de

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stoppe…

Model ReleasesDGX agent

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stopped using it are wrong - not wrong forever, just giving up on

31 Jul 2026

AI DOOMERS BE LIKE: 'GLM 5.1 WILL WIPE OUT HUMANS IN 2030'

Local AiDGX agent

I swear some AI doomers have never actually used a local model. They watched one flashy keynote, one YouTube thumbnail with a guy making this face 😱, read three headlines, and suddenly civilization is

AI Security Priorities: A Field-Wide Agenda

SafetyDGX agent

arXiv:2607.26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI securit

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

Model ReleasesDGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

Model ReleasesDGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

Model ReleasesDGX agent

arXiv:2607.28568v1 Announce Type: new Abstract: Recursive self-improvement (RSI) requires AI systems that improve the process of building AI (i.e., AI4AI); machine learning engineering (MLE) offers a

Towards Autonomous Aircraft Surveillance from Nanosatellites through On-Board Inference and Generative Data Augmentation

AgentsDGX agent

arXiv:2607.28470v1 Announce Type: cross Abstract: Airborne surveillance from low Earth orbit is hindered by two interconnected bottlenecks: nanosatellites have a limited downlink budget, yet the conve

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

Model ReleasesDGX agent

arXiv:2607.26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benc

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

Model ReleasesDGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

30 Jul 2026

GPT-Red: Automated Red Teaming via Self-Play at Scale

Model ReleasesDGX agent

arXiv:2607.26115v1 Announce Type: cross Abstract: We introduce extbf{GPT-Red}, an automated red-teaming agent that is trained to discover novel prompt injection attacks against frontier LLMs. The goal

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also ac…

AgentsDGX agent

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt cache hit rates. Low prici

Practice Makes Policies: Bootstrapping and Consolidating Robotic Capabilities from Zero Human Demonstrations

AgentsDGX agent

arXiv:2607.26809v1 Announce Type: new Abstract: General-purpose robotic manipulation requires robots to perform diverse tasks in open-world environments while improving their skills over time. Despite

29 Jul 2026

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a require…

AgentsDGX agent

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a requirement. In our new short course, AI Code Review, built in coll

Authenticate with Private Key JWT using Amazon Bedrock AgentCore Identity

AgentsDGX agent

This post explains how Private Key JWT client authentication works in AgentCore Identity and reviews the supported grant flows. We then walk through creating an AWS KMS signing key, registering its pu

Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines

AgentsDGX agent

arXiv:2607.25620v1 Announce Type: new Abstract: Quattrociocchi and colleagues warn that the fluent outputs of large language models may allow linguistic plausibility to substitute for epistemic evalua

F(AI)2R: Who Did What, and Who Checked? Verifiable AI Provenance as an Executable Skill

AgentsDGX agent

arXiv:2607.25637v1 Announce Type: cross Abstract: F(AI)2R is FAIR research with AI in the loop, twice: an AI-assisted authoring pass and a machine-readable audit pass over every artefact. AI systems n

Machine Intelligence on the Edge: Interpretable Cardiac Pattern Localisation Using Reinforcement Learning

AgentsDGX agent

arXiv:2508.21652v2 Announce Type: replace-cross Abstract: Matched filters are widely used to localise signal patterns due to their high efficiency and interpretability. However, their effectiveness de

OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis

SafetyDGX agent

arXiv:2607.25108v1 Announce Type: cross Abstract: Biomedical image analysis spans diverse modalities and tasks, yet real-world deployment is hindered by severe distribution shifts across scanners, pro

What Gets Lost When Memory Becomes Media? Evaluating AI-Generated Oral History Visualization

AgentsDGX agent

arXiv:2607.24756v1 Announce Type: cross Abstract: What gets lost when memory becomes media? Diaspora oral-history interviews require a double transformation; first-person recollection to third-person

28 Jul 2026

DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport

AgentsDGX agent

arXiv:2603.08111v2 Announce Type: replace Abstract: Generalizing decentralized multi-robot cooperative transport across objects with diverse shapes and physical properties remains a fundamental challe

Distributed Coordination for Resilient Multi-UAV Remote Sensing: A Photovoltaic Inspection Case Study

AgentsDGX agent

arXiv:2607.24482v1 Announce Type: new Abstract: Deploying multiple UAVs for remote sensing enables proportional reductions in mission time, but realizing these benefits requires the fleet to coordinat

Do LLM Debates Repeat Arguments Differently Across Languages?

SafetyDGX agent

arXiv:2607.23442v1 Announce Type: new Abstract: LLM debate is usually evaluated by final answers, but transcripts also reveal whether later turns develop new argumentative content or return to earlier

HydroAgent: Formalizing Forecaster Expertise into Skill-Orchestrated Flood Forecasting Workflows

AgentsDGX agent

arXiv:2607.23983v1 Announce Type: cross Abstract: Operational flood forecasting depends on tacit forecaster expertise that is difficult to formalize, audit, and transfer. Although artificial intellige

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-sou…

AgentsDGX agent

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-source model that we have tested to date. Try it today in Cosmo

PeopleSearchBench: A Multi-Dimensional Benchmark for Evaluating AI-Powered People Search Platforms

Model ReleasesDGX agent

arXiv:2603.27476v2 Announce Type: replace Abstract: AI-powered people search platforms are increasingly used in recruiting, sales prospecting, and professional networking, yet no widely accepted bench

Schema-Aware Localisation (SAL): Live Schema Grounding and Hallucination Validation for Oracle NL2SQL

Local AiDGX agent

arXiv:2607.22572v1 Announce Type: new Abstract: Large language models can generate fluent SQL from natural language, but on real enterprise Oracle databases they frequently fail at execution time: col

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

Model ReleasesDGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

27 Jul 2026

Active few-shot segmentation by reinforcing data selection

AgentsDGX agent

arXiv:2607.22371v1 Announce Type: new Abstract: Few-shot learning enables medical image segmentation models to adapt to new tasks using only a small number of labelled examples. However, adaptation pe

Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

Local AiDGX agent

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day!

AgentsDGX agent

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day! Kimi K3 is now live on Together AI. We’re proud to be a Day 0 launch pa

Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration

AgentsDGX agent

arXiv:2607.22153v1 Announce Type: cross Abstract: Industrial health management increasingly relies on heterogeneous information sources, including condition monitoring systems, supervisory control and

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes.

AgentsDGX agent

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes. .@Kimi_Moonshot K3 from Moonshot AI is now live on DigitalOcean

Microsoft says MAI-Cyber-1-Flash and MDASH, its vulnerability identification harness, deliver 'world-class performance at 50% of the cost of leading models' (Microsoft AI)

AgentsDGX agent

Microsoft AI: Microsoft says MAI-Cyber-1-Flash and MDASH, its vulnerability identification harness, deliver “world-class performance at 50% of the cost of leading models” — Today we're announcing MAI-

Variance-Reduced Q-Learning over Static and Time-Varying Networks

TutorialsDGX agent

arXiv:2607.21876v1 Announce Type: new Abstract: We investigate a decentralized reinforcement learning problem involving multiple agents that interact with the same Markov Decision Process (MDP). The a

26 Jul 2026

CEO of Hugging Face: 'In the spirit of transparency, here’s what I asked OpenAI'

Local AiDGX agent

clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2081056675558195657 • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened.

25 Jul 2026

My ChatGPT created a rally speech in response to counting letters.

AgentsDGX agent

I was exploring how to prompt the voice chat so that it can correctly count the number of “e”s in seventeen (which it fails most of the time). At the beginning of a new chat, instead of counting “e”s

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings wher…

Model ReleasesDGX agent

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings where site-deployed agents need exactly this as a base for domai

24 Jul 2026

A reminder this is 20% off via the Nous Portal <3

Model ReleasesDGX agent

mr‑r0b0t announced that Claude Opus 5 is available at a 20 % discount through the Nous Portal. The model can be accessed via the Hermes Agent on the Nous Portal, as well as through OpenRouter and Anth

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space

SafetyDGX agent

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

Model ReleasesDGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning

AgentsDGX agent

arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endoge

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

AgentsDGX agent

arXiv:2606.14502v2 Announce Type: replace Abstract: Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reason

Introducing Claude Opus 5 on AWS: Anthropic’s most capable Opus model

Model ReleasesDGX agent

This post covers Opus 5’s improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock. See the documentation fo

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

AgentsDGX agent

arXiv:2607.20430v1 Announce Type: cross Abstract: We present LLM-INSTRUCT, the winning system for the UZH Shared Task at ArgMining 2026 on paragraph-level argument mining in UN and UNESCO resolutions.

ModelExpress: Distributing Model Artifacts at the Speed of Light

HardwareDGX agent

NVIDIA ModelExpress (MX) is an agentic AI platform that streamlines the lifecycle of large‑model weights by automatically locating and using the fastest path to load them—prioritizing GPU‑to‑GPU P2P R

SevDiff: Severity-Conditioned Diffusion for Long-Tail Conflict Trajectory Generation

AgentsDGX agent

arXiv:2607.20549v1 Announce Type: new Abstract: Trajectory datasets used in ADAS evaluation are heavily biased toward routine driving; genuine vehicle-to-vehicle conflict events are rare, and the rare

← Previous
1…182183184185186…300
Next →