AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Model Releases

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

DGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

model-releasesarxiv-cs-cl
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them.

DGX agent

i doubt that a single person who has criticized me understands this. since i have been perfectly clear about this, that’s on them. @gabriberton Hmm, @ylecun always clarified he was talking about autor

agentsgary-marcus--x
3 Aug 2026
Agents

Overcoming the Weakest-Link Effect in LLM-Driven Program Optimization via Heterogeneous Edit Recombination

DGX agent

arXiv:2607.28947v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to solve complex problems by searching over program space, offering a general paradigm for scientific

agentsarxiv-cs-lg
3 Aug 2026
Agents

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusias…

DGX agent

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusiastic projections around Anthropic’s Q3. (Author below leaves

agentsgary-marcus--x
3 Aug 2026
Agents

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But…

DGX agent

It's not time to slow down but to accelerate! The recent AI-powered cyberattacks have everyone talking about the risks of AI. We should. But let's not lose sight of the bigger picture! If we work hard

agentsclem-delangue--x
2 Aug 2026
Model Releases

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the mode…

DGX agent

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the model. Same intelligence, wildly different accuracy and cost, de

model-releasesitamar-friedman--x
1 Aug 2026
Model Releases

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stoppe…

DGX agent

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stopped using it are wrong - not wrong forever, just giving up on

model-releasesswyx--x
1 Aug 2026
Local Ai

AI DOOMERS BE LIKE: 'GLM 5.1 WILL WIPE OUT HUMANS IN 2030'

DGX agent

I swear some AI doomers have never actually used a local model. They watched one flashy keynote, one YouTube thumbnail with a guy making this face 😱, read three headlines, and suddenly civilization is

local-air-ollama
31 Jul 2026
Safety

AI Security Priorities: A Field-Wide Agenda

DGX agent

arXiv:2607.26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI securit

safetyarxiv-cs-ai
31 Jul 2026
Model Releases

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

DGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

DGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

DGX agent

arXiv:2607.28568v1 Announce Type: new Abstract: Recursive self-improvement (RSI) requires AI systems that improve the process of building AI (i.e., AI4AI); machine learning engineering (MLE) offers a

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Towards Autonomous Aircraft Surveillance from Nanosatellites through On-Board Inference and Generative Data Augmentation

DGX agent

arXiv:2607.28470v1 Announce Type: cross Abstract: Airborne surveillance from low Earth orbit is hindered by two interconnected bottlenecks: nanosatellites have a limited downlink budget, yet the conve

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

DGX agent

arXiv:2607.26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benc

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

DGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

model-releasesjerry-liu--x
31 Jul 2026
Model Releases

GPT-Red: Automated Red Teaming via Self-Play at Scale

DGX agent

arXiv:2607.26115v1 Announce Type: cross Abstract: We introduce extbf{GPT-Red}, an automated red-teaming agent that is trained to discover novel prompt injection attacks against frontier LLMs. The goal

model-releasesarxiv-cs-cl
30 Jul 2026
Agents

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also ac…

DGX agent

Interesting data from the OpenRouter Kimi K3 dashboard: @togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt cache hit rates. Low prici

agentstogether-ai--x
30 Jul 2026
Agents

Practice Makes Policies: Bootstrapping and Consolidating Robotic Capabilities from Zero Human Demonstrations

DGX agent

arXiv:2607.26809v1 Announce Type: new Abstract: General-purpose robotic manipulation requires robots to perform diverse tasks in open-world environments while improving their skills over time. Despite

agentsarxiv-cs-ro
30 Jul 2026
Agents

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a require…

DGX agent

AI can write more code than any team can review by hand, and a pull request can look fine while hiding a security issue or missing a requirement. In our new short course, AI Code Review, built in coll

agentsitamar-friedman--x
29 Jul 2026
Agents

Authenticate with Private Key JWT using Amazon Bedrock AgentCore Identity

DGX agent

This post explains how Private Key JWT client authentication works in AgentCore Identity and reviews the supported grant flows. We then walk through creating an AWS KMS signing key, registering its pu

agentsaws-ml-blog
29 Jul 2026
Agents

Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines

DGX agent

arXiv:2607.25620v1 Announce Type: new Abstract: Quattrociocchi and colleagues warn that the fluent outputs of large language models may allow linguistic plausibility to substitute for epistemic evalua

agentsarxiv-cs-ai
29 Jul 2026
Agents

F(AI)2R: Who Did What, and Who Checked? Verifiable AI Provenance as an Executable Skill

DGX agent

arXiv:2607.25637v1 Announce Type: cross Abstract: F(AI)2R is FAIR research with AI in the loop, twice: an AI-assisted authoring pass and a machine-readable audit pass over every artefact. AI systems n

agentsarxiv-cs-ai
29 Jul 2026
Agents

Machine Intelligence on the Edge: Interpretable Cardiac Pattern Localisation Using Reinforcement Learning

DGX agent

arXiv:2508.21652v2 Announce Type: replace-cross Abstract: Matched filters are widely used to localise signal patterns due to their high efficiency and interpretability. However, their effectiveness de

agentsarxiv-cs-lg
29 Jul 2026
Safety

OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis

DGX agent

arXiv:2607.25108v1 Announce Type: cross Abstract: Biomedical image analysis spans diverse modalities and tasks, yet real-world deployment is hindered by severe distribution shifts across scanners, pro

safetyarxiv-cs-ai
29 Jul 2026
Agents

What Gets Lost When Memory Becomes Media? Evaluating AI-Generated Oral History Visualization

DGX agent

arXiv:2607.24756v1 Announce Type: cross Abstract: What gets lost when memory becomes media? Diaspora oral-history interviews require a double transformation; first-person recollection to third-person

agentsarxiv-cs-ai
29 Jul 2026
Agents

DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport

DGX agent

arXiv:2603.08111v2 Announce Type: replace Abstract: Generalizing decentralized multi-robot cooperative transport across objects with diverse shapes and physical properties remains a fundamental challe

agentsarxiv-cs-ro
28 Jul 2026
Agents

Distributed Coordination for Resilient Multi-UAV Remote Sensing: A Photovoltaic Inspection Case Study

DGX agent

arXiv:2607.24482v1 Announce Type: new Abstract: Deploying multiple UAVs for remote sensing enables proportional reductions in mission time, but realizing these benefits requires the fleet to coordinat

agentsarxiv-cs-ro
28 Jul 2026
Safety

Do LLM Debates Repeat Arguments Differently Across Languages?

DGX agent

arXiv:2607.23442v1 Announce Type: new Abstract: LLM debate is usually evaluated by final answers, but transcripts also reveal whether later turns develop new argumentative content or return to earlier

safetyarxiv-cs-cl
28 Jul 2026
Agents

HydroAgent: Formalizing Forecaster Expertise into Skill-Orchestrated Flood Forecasting Workflows

DGX agent

arXiv:2607.23983v1 Announce Type: cross Abstract: Operational flood forecasting depends on tacit forecaster expertise that is difficult to formalize, audit, and transfer. Although artificial intellige

agentsarxiv-cs-lg
28 Jul 2026
Agents

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-sou…

DGX agent

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-source model that we have tested to date. Try it today in Cosmo

agentsfireworks-ai--x
28 Jul 2026
Model Releases

PeopleSearchBench: A Multi-Dimensional Benchmark for Evaluating AI-Powered People Search Platforms

DGX agent

arXiv:2603.27476v2 Announce Type: replace Abstract: AI-powered people search platforms are increasingly used in recruiting, sales prospecting, and professional networking, yet no widely accepted bench

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

Schema-Aware Localisation (SAL): Live Schema Grounding and Hallucination Validation for Oracle NL2SQL

DGX agent

arXiv:2607.22572v1 Announce Type: new Abstract: Large language models can generate fluent SQL from natural language, but on real enterprise Oracle databases they frequently fail at execution time: col

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

DGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Active few-shot segmentation by reinforcing data selection

DGX agent

arXiv:2607.22371v1 Announce Type: new Abstract: Few-shot learning enables medical image segmentation models to adapt to new tasks using only a small number of labelled examples. However, adaptation pe

agentsarxiv-cs-cv
27 Jul 2026
Local Ai

Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

DGX agent

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

local-air-ollama
27 Jul 2026
Agents

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day!

DGX agent

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day! Kimi K3 is now live on Together AI. We’re proud to be a Day 0 launch pa

agentstogether-ai--x
27 Jul 2026
Agents

Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration

DGX agent

arXiv:2607.22153v1 Announce Type: cross Abstract: Industrial health management increasingly relies on heterogeneous information sources, including condition monitoring systems, supervisory control and

agentsarxiv-cs-lg
27 Jul 2026
Agents

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes.

DGX agent

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes. .@Kimi_Moonshot K3 from Moonshot AI is now live on DigitalOcean

agentskimi-moonshot--x
27 Jul 2026
Agents

Microsoft says MAI-Cyber-1-Flash and MDASH, its vulnerability identification harness, deliver 'world-class performance at 50% of the cost of leading models' (Microsoft AI)

DGX agent

Microsoft AI: Microsoft says MAI-Cyber-1-Flash and MDASH, its vulnerability identification harness, deliver “world-class performance at 50% of the cost of leading models” — Today we're announcing MAI-

agentstechmeme
27 Jul 2026
Tutorials

Variance-Reduced Q-Learning over Static and Time-Varying Networks

DGX agent

arXiv:2607.21876v1 Announce Type: new Abstract: We investigate a decentralized reinforcement learning problem involving multiple agents that interact with the same Markov Decision Process (MDP). The a

tutorialsarxiv-cs-lg
27 Jul 2026
Local Ai

CEO of Hugging Face: 'In the spirit of transparency, here’s what I asked OpenAI'

DGX agent

clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2081056675558195657 • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened.

local-air-localllama
26 Jul 2026
Agents

My ChatGPT created a rally speech in response to counting letters.

DGX agent

I was exploring how to prompt the voice chat so that it can correctly count the number of “e”s in seventeen (which it fails most of the time). At the beginning of a new chat, instead of counting “e”s

agentsr-chatgpt
25 Jul 2026
Model Releases

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings wher…

DGX agent

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings where site-deployed agents need exactly this as a base for domai

model-releasesyann-lecun--x
25 Jul 2026
Model Releases

A reminder this is 20% off via the Nous Portal <3

DGX agent

mr‑r0b0t announced that Claude Opus 5 is available at a 20 % discount through the Nous Portal. The model can be accessed via the Hermes Agent on the Nous Portal, as well as through OpenRouter and Anth

model-releasesnous-research--x
24 Jul 2026
Safety

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space

DGX agent

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

DGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning

DGX agent

arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endoge

agentsarxiv-cs-lg
24 Jul 2026
Agents

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

DGX agent

arXiv:2606.14502v2 Announce Type: replace Abstract: Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reason

agentsarxiv-cs-ai
24 Jul 2026
← Previous
1…228229230231232…375
Next →