AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,737 results
Safety

Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation

DGX agent

arXiv:2608.04794v1 Announce Type: new Abstract: Self-distillation (SD) has emerged as a compute-efficient alternative to reinforcement learning with verifiable rewards: a self-teacher, conditioned on

safetyarxiv-cs-ai
6 Aug 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Reinforcement Learning and Consumption-Savings Behavior

DGX agent

arXiv:2510.20748v2 Announce Type: replace-cross Abstract: This paper demonstrates how reinforcement learning can explain two puzzling empirical patterns in household consumption behavior during econom

researcharxiv-cs-ai
6 Aug 2026
Model Releases

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights,…

DGX agent

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights, fully composable. paste these four images into claude and a

model-releasesharrison-chase--x
6 Aug 2026
Local Ai

Risk Is Not the Target: A Monotonic Framework for Evaluating Wildfire Operational Risk Signals

DGX agent

arXiv:2607.21597v2 Announce Type: replace Abstract: Evaluating wildfire risk systems using standard machine-learning metrics such as F1-score or IoU is fundamentally flawed: these metrics assess event

local-aiarxiv-cs-ai
6 Aug 2026
Safety

Structured LLM Reasoning for Zero-Shot Human--Robot Coordination Under Hidden Goals

DGX agent

arXiv:2608.04309v1 Announce Type: new Abstract: We present a structured large-language-model (LLM) architecture for zero-shot human--robot coordination in a cooperative construction task with private

safetyarxiv-cs-ro
6 Aug 2026
Applications

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that w…

DGX agent

This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that was 'merely' good at hacking under human instructions. Initia

applicationsethan-mollick--x
6 Aug 2026
Applications

Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media work…

DGX agent

Together Serverless Inference gives developers a managed production path for bringing FLUX 3 into creative products and automated media workflows. Start building: https://www.together.ai/models/flux-3

applicationstogether-ai--x
6 Aug 2026
Research

Uncertainty-aware Predict-Then-Optimize Framework for Equitable Post-Disaster Power Restoration

DGX agent

arXiv:2508.04780v2 Announce Type: replace-cross Abstract: The increasing frequency of extreme weather events, such as hurricanes, highlights the urgent need for efficient and equitable power system re

researcharxiv-cs-ai
6 Aug 2026
Safety

A Security-Oriented Lifecycle Model for Large Language Model Systems

DGX agent

arXiv:2608.03626v1 Announce Type: cross Abstract: Large language models are being integrated into critical infrastructure and enterprise workflows at unprecedented scale,yet the lifecycle frameworks g

safetyarxiv-cs-ai
5 Aug 2026
Applications

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear c…

DGX agent

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear communication about incidents that is neither hyped up nor hi

applicationsethan-mollick--x
5 Aug 2026
Model Releases

Anyone interested in building a harness-only benchmark?

DGX agent

There are a lot of LLM benchmarks but few, if any, harness benchmarks. I am thinking this would be a really good community project to build one. End goal: a leaderboard of harness performance (multipl

model-releasesr-localllama
5 Aug 2026
Model Releases

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

DGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

DGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets diffe…

DGX agent

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets differently, we are screwed. Anthropic's Mythos created fake iden

safetygary-marcus--x
5 Aug 2026
Model Releases

CUDA MPC: A GPU-Native Solver for Model Predictive Control

DGX agent

arXiv:2608.03051v1 Announce Type: new Abstract: Model Predictive Control (MPC) delivers constraint-aware control, but its reliance on online optimization limits its use on systems with fast dynamics,

model-releasesarxiv-cs-ro
5 Aug 2026
Model Releases

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

DGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

model-releasesollama--x
5 Aug 2026
Model Releases

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

DGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

model-releasesr-localllama
5 Aug 2026
Safety

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

DGX agent

arXiv:2608.03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneo

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all…

DGX agent

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all document processing solutions - just screenshot the page an

model-releasesjerry-liu--x
5 Aug 2026
Applications

Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems

DGX agent

arXiv:2608.03413v1 Announce Type: new Abstract: As artificial intelligence (AI) continues to evolve and mature, recent AI practices have moved beyond large language models (LLMs) and text or image gen

applicationsarxiv-cs-ai
5 Aug 2026
Safety

Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction

DGX agent

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often r

safetyarxiv-cs-lg
5 Aug 2026
Applications

Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Translation Gap

DGX agent

arXiv:2608.02699v1 Announce Type: new Abstract: When algorithms make or influence consequential decisions---about loan eligibility, hiring, or healthcare---EU law grants affected individuals a Right t

applicationsarxiv-cs-ai
5 Aug 2026
Model Releases

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

DGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

DGX agent

arXiv:2608.03330v1 Announce Type: new Abstract: This thesis addresses fundamental challenges in traffic scene prediction for autonomous driving by introducing robust and computationally efficient mode

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Mistral introduces Shieldstral to provide lightweight policy-aware moderation for AI models

DGX agent

French artificial intelligence startup Mistral AI SAS today introduced a lightweight multimodal safety artificial intelligence open-weight model that can classify outputs for AI models that outperform

model-releasessiliconangle
5 Aug 2026
Model Releases

MultiGlobeQA: A Multilingual and Globally Diverse Benchmark for Geospatial Reasoning

DGX agent

arXiv:2608.03882v1 Announce Type: cross Abstract: Geospatial reasoning, i.e., computing distances, containment, and other spatial relations over real-world entities, is central to navigation and logis

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Optimal Liability Design for Medical AI

DGX agent

arXiv:2608.03114v1 Announce Type: cross Abstract: Artificial intelligence (AI) is increasingly integrated into medical decision-making, yet its liability implications remain complex, particularly when

safetyarxiv-cs-ai
5 Aug 2026
Safety

PULSE: An Executable Contract Language for Spatiotemporal Knowledge Graph Engineering

DGX agent

arXiv:2608.02630v1 Announce Type: new Abstract: Knowledge graph engineering often distributes accepted state, observations, constraints, processes, and hypothetical scenarios across artifacts whose co

safetyarxiv-cs-ai
5 Aug 2026
Applications

Representing Random Utility Choice Models with Neural Networks

DGX agent

arXiv:2207.12877v3 Announce Type: replace Abstract: Motivated by the successes of deep learning, we propose a class of neural network-based discrete choice models, called RUMnets, inspired by the rand

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

The production of meaning in the processing of natural language

DGX agent

arXiv:2603.20381v2 Announce Type: replace-cross Abstract: Understanding the fundamental mechanisms governing the production of meaning in the processing of natural language is critical for designing s

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series

DGX agent

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, fin

model-releasesarxiv-cs-lg
5 Aug 2026
Research

Towards a new paradigm of scientific discovery with socialized artificial intelligence

DGX agent

arXiv:2608.02775v1 Announce Type: new Abstract: Scientific discovery has advanced through successive transformations in the organization of knowledge. Observation and experimentation established the e

researcharxiv-cs-ai
5 Aug 2026
Safety

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

DGX agent

arXiv:2608.04007v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) enables LLMs to solve complex tasks through iterative tool interactions. However, existing reinforcement learning meth

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos …

DGX agent

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos 5 pursued its mission (fake identities, social engineering,

safetyethan-mollick--x
5 Aug 2026
Model Releases

A Triple-Robustness Analysis of Retrieval-Augmented Generation for Multi-Hop Requirements Traceability

DGX agent

arXiv:2608.00705v1 Announce Type: cross Abstract: Reported verdicts on GraphRAG versus vector RAG disagree, and the evidence is typically tied to a single corpus, embedder, and judge -- and, we show,

model-releasesarxiv-cs-cl
4 Aug 2026
Hardware

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory…

DGX agent

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory compute closer to the stars. The next chapter of AI infrastr

hardwareelon-musk--x
4 Aug 2026
Model Releases

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

DGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

model-releasesnous-research--x
4 Aug 2026
Safety

ARMOR: A Robust Self-Supervised Framework for Root Cause Analysis in Microservices under Missing Modality

DGX agent

arXiv:2603.25538v3 Announce Type: replace Abstract: Automated incident management is critical for microservice reliability. While recent unified frameworks leverage multimodal data for joint optimizat

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Auditable Release Control for Pedagogical Leakage in LLM Tutors

DGX agent

arXiv:2608.00515v1 Announce Type: cross Abstract: Large language model tutors can be correct and helpful yet disclose an answer or decisive reasoning before that disclosure is authorized. We formalize

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Belief-Contraction-Driven Active Inverse Source Localization and Characterization

DGX agent

arXiv:2501.13084v2 Announce Type: replace Abstract: Active inverse source localization and characterization (ISLC) in dynamic fields requires sequential decision making under partial observability, wh

local-aiarxiv-cs-lg
4 Aug 2026
Hardware

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

DGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

hardwarearxiv-cs-cl
4 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

DGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

model-releasesollama--x
4 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 (Q4) now reaches 1,328 tok/s prefill and ~29 tok/s decode on one RTX PRO 6000

DGX agent

I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB. These are timing-disabled internal Krasis r

model-releasesr-localllama
4 Aug 2026
Safety

Domain-Generalized Adaptive Semantic Communication for Collaborative Perception

DGX agent

arXiv:2608.00056v1 Announce Type: cross Abstract: We propose RSTA, a domain-generalized semantic communication framework enabling source-free V2X collaborative perception under both observation-domain

safetyarxiv-cs-lg
4 Aug 2026
Applications

Douyin Multimodal Embedding Model Technical Report

DGX agent

arXiv:2608.02148v1 Announce Type: cross Abstract: Multimodal representation learning is a cornerstone of modern AI. By encoding multimodal queries and targets into vectors, it powers industrial search

applicationsarxiv-cs-cl
4 Aug 2026
Safety

Entity-Aware Sequence Transduction for Player-Centric Ball Action Spotting

DGX agent

arXiv:2608.01696v1 Announce Type: new Abstract: Player-centric ball action spotting requires temporally precise event detection together with actor attribution in crowded, partially observed multi-age

safetyarxiv-cs-cv
4 Aug 2026
← Previous
1…328329330331332…370
Next →