AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,950 results
Model Releases

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

DGX agent

arXiv:2608.09316v1 Announce Type: new Abstract: Offline context optimization improves an agent by revising its instructions and examples while keeping the model frozen. This approach learns from rollo

model-releasesarxiv-cs-cv
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Multi-modal Interactive Control of Robotic Arm based on Offline Large Language Models

DGX agent

arXiv:2608.08183v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly revolutionized the modern society with numerous advanced interactions between humans and AI agents, wher

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases

Muse-Glimmer 30B Hits ~280 t/s in Real Production Coding

DGX agent

These numbers were captured during a real feature implementation task in Next.js and Nest.js (adding a theme switching system across components). The structural predictability of UI/state refactoring

model-releasesr-localllama
11 Aug 2026
Model Releases

OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories

DGX agent

arXiv:2608.08557v1 Announce Type: new Abstract: Visual tool use has emerged as a fundamental capability for multimodal agents to actively acquire evidence beyond a fixed image encoding. The prevailing

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Other active promotions: - Free models: Solar Pro 4 (1 week), Hy3, Step 3.7 Flash, Laguna S and XS - 90% off DeepSeek V4 Flash for ~2 more d…

DGX agent

Nous Research has extended its 20 % discount on all models—including high‑end frontier options—throughout the Nous Portal for an additional two weeks (until the end of April). Free model trials such a

model-releasesnous-research--x
11 Aug 2026
Model Releases

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

DGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation

DGX agent

arXiv:2608.09467v1 Announce Type: cross Abstract: Unmanned aerial vehicle vision-language navigation (UAV-VLN) requires agents to translate visual observations and language instructions into reliable

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance

DGX agent

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO)

safetyarxiv-cs-lg
11 Aug 2026
Research

SC^{2}-WM: A Self-Correcting World Model with Closed-Loop Feedback for Vision-and-Language Navigation in Continuous Environments

DGX agent

arXiv:2608.07548v1 Announce Type: cross Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to make fine-grained navigation decisions under partial observabili

researcharxiv-cs-cv
11 Aug 2026
Applications

Search over the Visual World: Persistent Visual Memory, Layered Indexes, and Source-Grounded Evidence

DGX agent

arXiv:2608.08075v1 Announce Type: cross Abstract: Most video-retrieval systems assume a bounded corpus and return ranked files or timestamps. Agents operating over cameras, screens, streams, and archi

applicationsarxiv-cs-cv
11 Aug 2026
Model Releases

SIMMER: Benchmarking Latent Failures in LLM Executable Planning with a World Model

DGX agent

arXiv:2606.14574v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as planners for autonomous agents in household environments. While existing benchmarks

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding

DGX agent

arXiv:2608.07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Th

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Tested in Coding: BF16 Muse Glimmer vs BF16 Qwen3.6 27B

DGX agent

I'm guessing that many people have been waiting for this comparison. For clarity, both models are running at full FP16 KV-cache. Due to VRAM limitations, Muse Glimmer is running full 262,144 context,

model-releasesr-localllama
11 Aug 2026
Model Releases

VTO: Visual Tool Orchestration for Video Anomaly Detection

DGX agent

arXiv:2608.08219v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is a critical yet challenging task due to the complex and diverse nature of real-world scenarios. Traditional deep learn

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation

DGX agent

arXiv:2505.11146v3 Announce Type: replace-cross Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition challenge due to the significant

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Automated item evaluation: Predicting item acceptance and rejection using LLM-generated critiques

DGX agent

arXiv:2608.06609v1 Announce Type: new Abstract: Automated item evaluation (AIE) refers to the use of computational methods to assess item quality without requiring manual expert review or field testin

safetyarxiv-cs-ai
10 Aug 2026
Safety

Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps

DGX agent

arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine judges provide roughly the effective inform

safetyarxiv-cs-ai
10 Aug 2026
Safety

Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving

DGX agent

arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges

safetyarxiv-cs-lg
10 Aug 2026
Model Releases

Comparing how Cline, Kilo, and Qwen Code handle long-task context/state (and why context loops keep happening)

DGX agent

I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside the conversation that gets reinjected on a c

model-releasesr-localllama
10 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

DGX agent

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

model-releasesr-localllama
10 Aug 2026
Local Ai

How to prevent LLM to act like a robot/assistant?

DGX agent

I'm playing with a conversational agent I made using either api/generate or api/chats. In both case I do ask him to not ask follow up question, to not act like an assistant, etc. Either from a system

local-air-ollama
10 Aug 2026
Model Releases

I Seek You in Videos: Identity-Conditioned Queries for Person-Centric Video Reasoning

DGX agent

arXiv:2608.07417v1 Announce Type: cross Abstract: Real-world video reasoning often involves multimodal, multi-source inputs, whereas existing video reasoning tasks typically assume a simplified video-

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Introducing Muse Glimmer

DGX agent

Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to

model-releasessimon-willison
10 Aug 2026
Model Releases

Muse Glimmer on 1/2 AMD v620

DGX agent

Hey. Just tried it on my old ass gpus 😄 Surprisingly Tensor Split is working on 2 gpus almost doubling PP (wonder how it will work with 4 gpus) Q6 — 1 GPU llama-server --model <MODEL_DIR>/Muse-Glimmer

model-releasesr-localllama
10 Aug 2026
Model Releases

Please Share Your Experience About Muse Glimmer

DGX agent

I have a classic test for local LLM's. I asked for 8 ball pool game with only one HTML file and Muse Glimmer spend 21k Token(I m using full context so 128k) and only created a 220 lines of HTML and sa

model-releasesr-localllama
10 Aug 2026
Hardware

Scalable High-Fidelity Macromolecular Docking for GPU-Accelerated Supercomputers

DGX agent

arXiv:2608.07078v1 Announce Type: cross Abstract: Flexible macromolecular docking offers high-fidelity predictions of biomolecular interactions, but remains prohibitively expensive at scale. Among exi

hardwarearxiv-cs-ai
10 Aug 2026
Model Releases

TradeVerse: A Longitudinal Benchmark of Political Negotiation in International Trade

DGX agent

arXiv:2608.06549v1 Announce Type: cross Abstract: LLMs are increasingly being applied to tasks involving institutional and political texts, but existing benchmarks evaluate them on isolated documents

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

v0.32.7

DGX agent

Muse Glimmer Note: Muse Glimmer is currently available via initial support via Ollama's MLX engine on Apple Silicon. Additional support and optimizations for Apple Silicon, NVIDIA, AMD, and other plat

model-releasesollama-releases
10 Aug 2026
Model Releases

Now we have a timeline of the OpenAI accidental attack against Hugging Face

DGX agent

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

model-releasessimon-willison
8 Aug 2026
Model Releases

Tesla V100 Qwen3.6 27B Performance

DGX agent

Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi coding agent llama.cpp model preset: [*] spec-default =

model-releasesr-localllama
8 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

DGX agent

arXiv:2608.06012v1 Announce Type: new Abstract: Search-agent rewards mix answer quality, citation grounding, tool cost, and anti-hacking terms; a high score therefore need not imply that cited evidenc

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

How Google Cloud detects, contains, and protects against emerging threats

DGX agent

At Google Cloud, securing your data and business systems is our foundational commitment. We empower our customers with the tools, governance, and infrastructure needed to securely deploy workloads and

model-releasesgoogle-cloud-ai
7 Aug 2026
Model Releases

Improving the Realism of Synthetic Clinical Benchmarks Under Utility Constraints

DGX agent

arXiv:2608.06265v1 Announce Type: new Abstract: Synthetic clinical benchmarks for enterprise AI agents can pass existing utility checks and still remain structurally unrealistic, especially in privacy

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models

DGX agent

arXiv:2608.05706v1 Announce Type: new Abstract: World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied int

safetyarxiv-cs-cv
7 Aug 2026
Model Releases

My issue with Artificial Analysis's 'intelligence index'

DGX agent

I swear AA is not the bipartisan they so claim. An open source mode (Qwen 3.8 max) was number 1 on the agentic index, then they just so happen to launch 'v4.1.1' of their index in which they just adju

model-releasesr-localllama
7 Aug 2026
Model Releases

Project2Task: Graph-Guided Project-Level Planning for Autonomous Research

DGX agent

arXiv:2608.05225v1 Announce Type: new Abstract: Research agents can increasingly search literature, propose hypotheses, generate code, run experiments, and draft manuscripts from a single topic. Howev

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized Views

DGX agent

arXiv:2607.21910v2 Announce Type: replace Abstract: World models let agents plan against predicted physical state, but that state drifts; re-observation is costly and delayed, and repair can fail. We

local-aiarxiv-cs-ai
7 Aug 2026
Applications

VLMs for Videogame Data Annotation

DGX agent

arXiv:2608.05949v1 Announce Type: new Abstract: Vision Language Models (VLMs) and Artificial Intelligence (AI) agents have revolutionized how engineers approach complex problems in real-world applicat

applicationsarxiv-cs-ai
7 Aug 2026
Model Releases

When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories

DGX agent

arXiv:2608.06057v1 Announce Type: new Abstract: Tool-calling agents infer task state from accumulated dialogue and tool traces. In persistent interactions, however, historical traces may remain struct

model-releasesarxiv-cs-ai
7 Aug 2026
Tutorials

Build visibility for Codex on Amazon Bedrock with OpenTelemetry and Amazon CloudWatch

DGX agent

As engineering teams adopt coding agents like Codex, leaders need visibility into adoption, consumption, and reliability. This post shows how to route Codex OpenTelemetry metrics through a local colle

tutorialsaws-ml-blog
6 Aug 2026
Tutorials

Configure rate limits for AI traffic on AgentCore gateway

DGX agent

Learn how to configure rate limits on Amazon Bedrock AgentCore gateway to enforce per-user and per-target traffic controls. Define request, token, and connection limits scoped by JWT claims or IAM ide

tutorialsaws-ml-blog
6 Aug 2026
Safety

Corrigibility Transformation: Constructing Goals That Accept Updates

DGX agent

arXiv:2510.15395v2 Announce Type: replace Abstract: An AI agent will learn a desired goal more effectively if it does not resist the training process, but many partially learned goals incentivize an A

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

DGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

model-releasesgoogle-cloud-ai
6 Aug 2026
Research

Emergence of Hierarchical Emotion Organization in Large Language Models

DGX agent

arXiv:2507.10599v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly power conversational agents, understanding how they model users' emotional states is critical for

researcharxiv-cs-ai
6 Aug 2026
Research

MemFly: On-the-Fly Memory Optimization via Information Bottleneck

DGX agent

arXiv:2602.07885v2 Announce Type: replace Abstract: Long-term memory enables large language model agents to tackle complex tasks through historical interactions. However, existing frameworks encounter

researcharxiv-cs-ai
6 Aug 2026
Safety

Overcoming Statistical Bias in Action-Controllable World Models

DGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

safetyarxiv-cs-cv
6 Aug 2026
Safety

Reward Structure Shapes the Interaction Between Episodic Exploration and Neural Memory in Reinforcement Learning

DGX agent

arXiv:2608.05111v1 Announce Type: new Abstract: In partially observable reinforcement learning, agents face a dual bottleneck: they must explore to encounter rewarding states and retain that experienc

safetyarxiv-cs-lg
6 Aug 2026
← Previous
1…282283284285286…374
Next →