AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Model Releases

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing

DGX agent

arXiv:2608.05485v1 Announce Type: new Abstract: Evaluating generated videos remains challenging because existing benchmarks rely on fixed evaluation content, cover only a subset of generation and edit

model-releasesarxiv-cs-cv
7 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign diff…

DGX agent

Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign different open models to coding, planning, vision, and review ac

agentstogether-ai--x
6 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

DGX agent

arXiv:2608.04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the iss

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

ANCHOR-RE: An Agentic Neuro-Symbolic Framework for Grounded Biomedical Relation Extraction

DGX agent

arXiv:2608.03154v1 Announce Type: new Abstract: Biomedical relation extraction (BioRE) extracts structured knowledge from biomedical literature for applications such as knowledge base construction and

model-releasesarxiv-cs-cl
5 Aug 2026
Agents

How Mobileye transformed support operations using Amazon Bedrock AgentCore

DGX agent

In this post, we'll explore how Mobileye deployed an AI support agentic solution on Amazon Bedrock AgentCore - from the support bottleneck that sparked the idea, through the proof of concept that vali

agentsaws-ml-blog
5 Aug 2026
Research

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

DGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

researcharxiv-cs-ai
5 Aug 2026
Agents

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning

DGX agent

arXiv:2608.00301v1 Announce Type: cross Abstract: Error-penalized scoring rules (+1 for a correct answer, -lambda for a wrong one, 0 for abstaining) are increasingly prescribed against hallucination:

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

[Deepseek-V4-Flash-0731] Full 1M context on a single RTX5090 + DDR5 Desktop Setup with VLLM CPU/Ram Offloading, ~800 tps pp & 15+ tps decode [Agentic Coding]

DGX agent

First of all, obviously I took some help from AI to type this post and this is the topic that enabled me to accomplish all that: https://old.reddit.com/r/LocalLLaMA/comments/1veow4b/deepseek_v4flash_2

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 flash 0731 ranks #21 on Agent Arena

DGX agent

https://preview.redd.it/522fsdwvtdhh1.png?width=1200&format=png&auto=webp&s=6a6cf7a467514167a8193029dbd20fb3a9ba4f6c It ranks lower than both Sonnet 4.6 and Luna. I'd wager Luna costs in the same ball

model-releasesr-localllama
4 Aug 2026
Model Releases

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference

DGX agent

arXiv:2608.02515v1 Announce Type: new Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retri

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

PackingGPT: 3D Packing Agent for Real Furniture in Last-Mile Delivery

DGX agent

arXiv:2608.01427v1 Announce Type: new Abstract: 3D bin packing rectangular items into standardised containers to maximise space utilisation under geometric shipping automation. Loading a furniture pur

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SIPTraj: Map-Free End-to-End Trajectory Prediction via Physics-Guided Scene Interaction

DGX agent

arXiv:2608.00779v1 Announce Type: new Abstract: Trajectory prediction of surrounding agents is a prerequisite for safe planning and decision making in autonomous driving. Without high-definition (HD)

model-releasesarxiv-cs-ro
4 Aug 2026
Agents

Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving

DGX agent

arXiv:2607.29031v1 Announce Type: cross Abstract: Existing autonomous-driving world models typically perform dense prediction of future videos, occupancy states, BEV representations, or agent motion.

agentsarxiv-cs-ai
3 Aug 2026
Research

Here’s why AI agents lie and cheat to reach their goals

DGX agent

MIT Technology Review Explains: Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here. When two OpenAI model

researchmit-tech-review
3 Aug 2026
Model Releases

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

DGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits…

DGX agent

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits of open AI models, which helped him defend against a rogue

agentsclem-delangue--x
2 Aug 2026
Model Releases

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, …

DGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real tasks from 3 repositories, with context-injection st

model-releasesdair-ai--x
1 Aug 2026
Agents

ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks (Kyle Alspach/CRN)

DGX agent

Kyle Alspach / CRN: ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks — The cybersecurity vendo

agentstechmeme
1 Aug 2026
Agents

AI as Friction for Reflection Support in Ideation

DGX agent

arXiv:2607.26827v1 Announce Type: cross Abstract: Generative AI tools for creative work tend to be designed around the goal of removing friction, on the assumption that smoother iteration and faster o

agentsarxiv-cs-ai
31 Jul 2026
Local Ai

Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer

DGX agent

arXiv:2607.17100v2 Announce Type: replace-cross Abstract: Auto Research uses language-model agents to propose, implement, and evaluate machine-learning changes in a closed loop, but is usually judged

local-aiarxiv-cs-ai
31 Jul 2026
Agents

One Run Is Not an Idea: The Implementation Lottery in Automated Research

DGX agent

arXiv:2607.26587v1 Announce Type: cross Abstract: Automated research systems use experimental scores both to deliver artifacts and to decide which ideas to retain, transfer, and pursue. Yet one run sc

agentsarxiv-cs-ai
31 Jul 2026
Model Releases

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models

DGX agent

arXiv:2607.28609v1 Announce Type: cross Abstract: Computer-using agents (CUAs) are advancing rapidly across the digital world. A CUA trajectory records the agent's actions, states, and reasoning. Veri

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Pushing the Frontier on Approximate EFX Allocations

DGX agent

arXiv:2406.12413v3 Announce Type: replace-cross Abstract: We study the problem of allocating a set of indivisible goods to a set of agents with additive valuation functions, aiming to achieve approxim

researcharxiv-cs-ai
31 Jul 2026
Agents

HeteroPROPMT: A Real-time and Privacy-Preserving Heterogeneous Collaborative Perception Framework

DGX agent

arXiv:2607.26283v1 Announce Type: new Abstract: Collaborative Perception (CP) improves autonomous systems' awareness of their surroundings by sharing sensor data, intermediate features, and detection

agentsarxiv-cs-cv
30 Jul 2026
Safety

CAST: Game Solvers as Turn-Level Teachers for LLM Agents

DGX agent

arXiv:2607.25308v1 Announce Type: cross Abstract: Training large language models (LLMs) to act in long-horizon games is a promising step toward generalist decision-making, yet reinforcement learning w

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

DGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases

DGX agent

arXiv:2408.03910v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel in stand-alone code tasks like HumanEval and MBPP, but struggle with handling entire code repositories. Thi

agentsarxiv-cs-ai
28 Jul 2026
Agents

MARS: Multi-hop Adaptive Retrieval and SPARQL Generation for KGQA

DGX agent

arXiv:2607.14561v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated strong reasoning performance, but their tendency to hallucinate limits their reliability in knowledge

agentsarxiv-cs-cl
28 Jul 2026
Model Releases

Tokengeist: Multi-Turn Attribution Tracing in Agentic Conversations

DGX agent

arXiv:2607.22610v1 Announce Type: new Abstract: When a language model produces a response in a multi-turn conversation, which tokens from prior turns shaped that answer, and how did those dependencies

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

One Hand Watches The Other: Dynamic Multi-Agent Cooperation for Sample-Efficient Bimanual Manipulation in Dynamic Environments

DGX agent

arXiv:2607.22119v1 Announce Type: cross Abstract: Multi-stream robot manipulation policies achieve unparalleled sample efficiency and generalization by modeling actions relative to environmental refer

model-releasesarxiv-cs-lg
27 Jul 2026
Agents

Our team just shipped Fugu-Ultra v1.1! 🐡 By dynamically orchestrating the latest frontier models, we pushed performance up by 7.9 points. W…

DGX agent

Our team just shipped Fugu-Ultra v1.1! 🐡 By dynamically orchestrating the latest frontier models, we pushed performance up by 7.9 points. We are now beating Fable 5 in complex coding and reasoning tas

agentsdavid-ha--x
24 Jul 2026
Agents

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's q…

DGX agent

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's quick to run and built for agent workloads: coding, search, r

agentsnous-research--x
23 Jul 2026
Model Releases

v0.32.2

DGX agent

What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @Parth

model-releasesollama-releases
22 Jul 2026
Agents

text/image-to-sim

DGX agent

Gizmo is a simulation-authoring agent that converts textual descriptions and reference images into structured, editable 3D scenes tailored for robotics workflows. It was publicly released as a beta on

agentsyohei-nakajima--x
21 Jul 2026
Agents

Active Trust Management for Successful Human-Robot Teaming: Moving from a Trust Repair to a Trust Satisficing Perspective

DGX agent

arXiv:2607.13595v1 Announce Type: new Abstract: Integrating mobile robots into human teams promises significant capability improvements for tasks such as searching hazardous environments. Unlike exist

agentsarxiv-cs-ro
16 Jul 2026
Safety

SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing

DGX agent

arXiv:2607.13594v1 Announce Type: new Abstract: LLM agents act on real-world environments through tool calls, and a single misjudged action can cause irreversible harm. The standard safeguard is a gua

safetyarxiv-cs-ai
16 Jul 2026
Safety

CityBehavEx: A Scalable and Empirically Validated LLM-Assisted Urban Simulation Platform

DGX agent

arXiv:2607.12086v1 Announce Type: new Abstract: Recent LLM-based multi-agent urban simulators can generate semantically rich city routines, but they remain costly to scale and are often weakly validat

safetyarxiv-cs-cl
15 Jul 2026
Model Releases

RCWT: Measuring Task-Budget Displacement from Coordination Content in LLM Calls

DGX agent

arXiv:2607.12216v1 Announce Type: cross Abstract: Multi-agent and memory-augmented LLM systems often place coordination content, shared state, prior discussion, tool outputs, summaries, and role instr

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Google named a Leader in the 2026 IDC MarketScape for Worldwide Foundation Model Software

DGX agent

For years, we’ve built with a clear priority: putting the practical needs of the enterprise first. Long before generative AI dominated the headlines, we were focused on building the global infrastruct

model-releasesgoogle-cloud-ai
14 Jul 2026
Agents

some good self improvement research here

DGX agent

some good self improvement research here The first experimental evidence of recursive self-improvement (RSI). Autoresearching the autoresearch agent for eight days. The result beats the harness we han

agentsyohei-nakajima--x
14 Jul 2026
Tools

Text match filters for agents

DGX agent

Text match filters are a feature in Pinecone that allow users to filter vector search results based on exact text matching criteria, enabling more precise control over which documents or records are r

toolspinecone
14 Jul 2026
Agents

How do physical systems achieve collective intelligence and self-repair without a central brain? A new paper published today in Nature Commu…

DGX agent

How do physical systems achieve collective intelligence and self-repair without a central brain? A new paper published today in Nature Communications by my Sakana AI colleague Sebastian Risi (@risi197

agentsdavid-ha--x
13 Jul 2026
Agents

I guess image input is the big capability of the models, and tool use can be a substitute for non-omni model output. Still, multimodal voice…

DGX agent

Ethan Mollick notes that image input represents the primary advanced capability of current AI models, and that tool‑use can effectively replace outputs from non‑omni models. He observes that multimoda

agentsethan-mollick--x
13 Jul 2026
Safety

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis

DGX agent

arXiv:2606.16149v2 Announce Type: replace Abstract: Rare disease diagnosis involves interpreting clinical and genetic findings through complex diagnostic reasoning. We investigated whether this reason

safetyarxiv-cs-ai
10 Jul 2026
Agents

love this framing of memory as 'proactive'

DGX agent

love this framing of memory as 'proactive' agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the ground running .

agentsharrison-chase--x
10 Jul 2026
Model Releases

Breaking Database Lock-in: Agentic Regeneration of High Performance Storage Readers for Database Bypass

DGX agent

arXiv:2607.07696v1 Announce Type: cross Abstract: Analytical workloads operating on data stored in external database systems face a fundamental bottleneck: data access is guarded entirely by the datab

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

EmbodiedGen V2: An Agentic, Simulation-Ready 3D World Engine for Embodied AI

DGX agent

arXiv:2607.07459v1 Announce Type: cross Abstract: We present EmbodiedGen V2, a generative 3D world engine for building executable sim-ready environments for embodied intelligence. Sim-ready 3D asset g

safetyarxiv-cs-cv
9 Jul 2026
← Previous
1…188189190191192…375
Next →