AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,420 results
Agents

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

DGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

agentsarxiv-cs-ai
5 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

LDU-Bench: Multimodal LLM Evaluation for Lithography Defect Understanding under Layout-Varying Circuit Backgrounds

DGX agent

arXiv:2608.03078v1 Announce Type: new Abstract: Multimodal large language models have demonstrated strong defect recognition capability in industrial anomaly detection. However, in lithography review,

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

DGX agent

TensorSharp's MoE CPU-offload feature has been merged into main. Here is the parameters description of this feature: Mixture-of-Experts CPU offload: --n-cpu-moe <N> | -ncmoe <N> Keep the routed MoE ex

model-releasesr-localllama
5 Aug 2026
Safety

MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory

DGX agent

arXiv:2608.02843v1 Announce Type: cross Abstract: Persistent agent memory must adapt as later outcomes change earlier evidence, yet mutable retrieval weights create an attribution problem: reviewers m

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

One-shotting a Raccoon Heist game using Claude Fable 5

DGX agent

Back in 2024 I tweeted screenshots of a game concept generated by GPT-3 and some concept 'art' created using DALL-E. Today, on the fourth anniversary of that tweet, I decided to see if Claude Fable 5

model-releasessimon-willison
5 Aug 2026
Research

Open-Linguistic Concept Unified Learning for Cross-Site Interpretable Dermatology Image Diagnosis

DGX agent

arXiv:2608.03225v1 Announce Type: new Abstract: Human-interpretable computer-aided diagnosis is crucial for clinical decision making. Concept-based models excel by providing transparent reasoning and

researcharxiv-cs-cv
5 Aug 2026
Safety

PhyAI: Real-Time Physical AI at the Edge, Scalable Rollouts in the Cloud

DGX agent

arXiv:2608.03682v1 Announce Type: new Abstract: Physical AI policies require inference throughout their lifecycle, including model evaluation, cloud reinforcement learning rollout, edge GPU serving, a

safetyarxiv-cs-ai
5 Aug 2026
Agents

Principles of Robot Autonomy

DGX agent

arXiv:2608.03496v1 Announce Type: cross Abstract: Autonomous robots are moving rapidly from research labs into everyday life - on roads, in the air, in warehouses, and in space. Robot autonomy is no l

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

DGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

DGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Safety

Some people are surprised that APIs (aka what Anthropic, OpenAI, and others provide) are treated differently than open weights in the new AI…

DGX agent

Some people are surprised that APIs (aka what Anthropic, OpenAI, and others provide) are treated differently than open weights in the new AI model framework. I'm not surprised at all, and it's actuall

safetyclem-delangue--x
5 Aug 2026
Model Releases

Stuck on 'A': Diagnosing and Repairing Interface Injury in Attention-to-KDA Linearization of a 0.6B Language Model

DGX agent

arXiv:2608.02689v1 Announce Type: new Abstract: We convert 21 of 28 full-attention layers of Qwen3-0.6B-Base into KDA (Kimi Delta Attention) linear-attention layers on a single consumer-grade GPU budg

model-releasesarxiv-cs-cl
5 Aug 2026
Safety

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

DGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

DGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense

DGX agent

arXiv:2608.00583v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring is meant to catch the reward hacks that look clean in the actions and betray themselves only in the reasoning. We sh

agentsarxiv-cs-cl
4 Aug 2026
Safety

Beyond Random Partitioning: Unsupervised Spatio-Temporal Stratification for Cohort Balancing in Longitudinal Medical Imaging

DGX agent

arXiv:2608.00073v1 Announce Type: new Abstract: Rigorous dataset partitioning is a foundational, yet frequently overlooked, prerequisite for reliable deep learning in longitudinal medical imaging. Nai

safetyarxiv-cs-cv
4 Aug 2026
Local Ai

bFaaaP: An Inclusive, Head-Angle Piano-Pedal Interaction that Quantitatively Reproduces a Pianist's Intended Pedalling -- Foot-Free, for Acoustic and Electronic Pianos

DGX agent

arXiv:2608.00633v1 Announce Type: cross Abstract: Expressive piano performance depends on the sustain (damper) pedal, operated by foot, excluding players who cannot readily use their feet: wheelchair

local-aiarxiv-cs-ro
4 Aug 2026
Hardware

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

DGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

hardwarearxiv-cs-cl
4 Aug 2026
Model Releases

Conditional Deep Levy Models for Exotic Derivatives: History-Aware Path Generation and P-Q Payoff Diagnostics

DGX agent

arXiv:2509.13374v2 Announce Type: replace-cross Abstract: We develop and audit a history-aware financial path generator based on Denoising Levy Probabilistic Models (DLPMs) for conditional equity-inde

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

DGX agent

arXiv:2608.00538v1 Announce Type: new Abstract: Recent advancements of zero-shot Named Entity Recognition (NER) establish strong baselines by formulating sequence labeling into question answering wher

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

DecoupleGS: Interactive 3D Gaussian Splatting for End-to-End Autonomous Driving Testing

DGX agent

arXiv:2608.01761v1 Announce Type: new Abstract: End-to-end (E2E) autonomous driving algorithms require rigorous closed-loop validation in simulation environments offering high visual fidelity, strong

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

Deepseek V4 Flash 2-bit quant is the first model I can run locally that achieves 100% in this SQL benchmark

DGX agent

I really like to use this one SQL benchmark when testing new models. I had another post some time ago with my benchmarks, but I decided to post a new one because of how well Deepseek did. I like the b

model-releasesr-localllama
4 Aug 2026
Research

Fast Discovery of Inclusion Dependencies with Desbordante

DGX agent

arXiv:2608.02213v1 Announce Type: cross Abstract: Inclusion dependency is a relation between attributes of tables that indicates possible Primary Key-Foreign Key references. Automatic discovery of inc

researcharxiv-cs-lg
4 Aug 2026
Research

FAU at ImageCLEF 2026 Task on Multimodal Reasoning Robust Candidate Scoring and Concise Multilingual Visual Answering

DGX agent

arXiv:2608.01664v1 Announce Type: new Abstract: We present our ImageCLEF 2026 Multimodal Reasoning system for the Visual Multiple Choice Question Answering (Visual MCQ) and Visual Open Question Answer

researcharxiv-cs-cv
4 Aug 2026
Research

From fragmented data to actionable design: Physics-calibrated learning for plastic upcycling

DGX agent

arXiv:2608.02402v1 Announce Type: new Abstract: Thermochemical upgrading of plastic waste is a key upcycling pathway, yet the experimental literature is fragmented by heterogeneous conditions and inco

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Humans Are More Diverse: Frontier LLMs Show Extreme Policies in Idealised AI Development Races

DGX agent

arXiv:2608.01193v1 Announce Type: cross Abstract: An AI development race creates a multi-agent safety dilemma. Each company can develop slowly and safely, or move faster while taking a risk that may r

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation

DGX agent

arXiv:2608.01804v1 Announce Type: new Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

DGX agent

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more …

DGX agent

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more observable, easier to steer, easier to understand we want to

model-releaseslinus-lee--x
4 Aug 2026
Model Releases

OpenAI and Anthropic have both just posted about an overlapping cyber incident involving GPT-5.6-Sol and Mythos 5 during an evaluation by UK…

DGX agent

OpenAI and Anthropic have both just posted about an overlapping cyber incident involving GPT-5.6-Sol and Mythos 5 during an evaluation by UKAISI. I will quote: 'In the most serious case, an agent trie

model-releasesgary-marcus--x
4 Aug 2026
Safety

OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs

DGX agent

arXiv:2603.11804v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) adapted to remote sensing rely heavily on domain-specific image-text supervision, yet high-quality annotations for sat

safetyarxiv-cs-cv
4 Aug 2026
Tutorials

Perception-and-action system for humanoid robot task execution in construction

DGX agent

arXiv:2608.01600v1 Announce Type: new Abstract: Humanoid robots, with their human-like shape and multi-tasking capabilities, are well-aligned with human-dominated workplaces, like those in civil and c

tutorialsarxiv-cs-ro
4 Aug 2026
Agents

Practical Online KV Cache Compaction for LLM Agents: An Empirical Study

DGX agent

arXiv:2608.00902v1 Announce Type: new Abstract: LLM agents accumulate long trajectories of reasoning steps, tool calls, and environment feedback, making the KV cache a major inference bottleneck. KV c

agentsarxiv-cs-cl
4 Aug 2026
Research

Predicting Startup Exit from Textual Descriptors - A Computational Linguistics Framework

DGX agent

arXiv:2608.00045v1 Announce Type: new Abstract: This study shows that textual descriptors alone can predict early-stage startup success, defined as Exit, without relying on contextual, financial, or h

researcharxiv-cs-cl
4 Aug 2026
Research

Pretraining on Call Graphs: When Binary Analysis Tasks Profit From Context

DGX agent

arXiv:2608.02084v1 Announce Type: cross Abstract: Binary function embedding models are trained to encode the semantics of binary code in such a way that they can be generalized to a variety of reverse

researcharxiv-cs-lg
4 Aug 2026
Research

ReBRAC-v2: The Return of the King

DGX agent

arXiv:2608.01205v1 Announce Type: new Abstract: Recent offline reinforcement learning methods increasingly rely on expressive generative policies and specialized value-guidance mechanisms. We ask whet

researcharxiv-cs-lg
4 Aug 2026
Applications

Scene2Sound: Auditory-Grounded Soundscape Generation for 3D Gaussian Worlds

DGX agent

arXiv:2608.00463v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) turns captured or generated imagery into photorealistic 3D world simulations that users can freely explore, yet these world

applicationsarxiv-cs-cv
4 Aug 2026
Local Ai

TELLER: Non-intrusive Cross-Layer Root-Cause Analysis for LLM Inference

DGX agent

arXiv:2608.01975v1 Announce Type: cross Abstract: Large language model (LLM) inference has evolved from an offline workload into a continuously operated software service, yet root-cause analysis remai

local-aiarxiv-cs-cl
4 Aug 2026
Industry

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing th…

DGX agent

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing the paradigm… describe the molecule you want, and the model de

industrysonya-huang--x
4 Aug 2026
Local Ai

TS-MAMP: A Remanufactured Agricultural Robot Powered by Second-Life EV Components and NMS-Free On-Device Weed Detection

DGX agent

arXiv:2608.02270v1 Announce Type: new Abstract: Agriculture 4.0 robotic systems improve field efficiency yet remain too capital-intensive for the fragmented smallholdings that dominate global agricult

local-aiarxiv-cs-ro
4 Aug 2026
Local Ai

VertiAKD: Adaptive Off-Road Kinodynamics on Vertically Challenging Terrain

DGX agent

arXiv:2608.00945v1 Announce Type: new Abstract: Off-road mobility requires autonomous mobile robots to generalize across heterogeneous vehicle fleets and continuously changing terrain conditions. Exis

local-aiarxiv-cs-ro
4 Aug 2026
Model Releases

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding

DGX agent

arXiv:2608.00036v1 Announce Type: new Abstract: Real-world document tasks often ask professionals to answer questions from annual reports, regulations, clinical guidelines, and technical manuals that

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

A Model-Driven Approach for Developing Families of Reinforcement Learning Environments

DGX agent

arXiv:2606.20324v2 Announce Type: replace-cross Abstract: Virtual training environments are software-intensive systems in which reinforcement learning (RL) agents learn, adapt, and demonstrate meaning

local-aiarxiv-cs-lg
3 Aug 2026
Tutorials

A user's guide to PINNs in geometric analysis: lessons from the asymptotic Plateau problem

DGX agent

arXiv:2607.28733v1 Announce Type: cross Abstract: This proceedings contribution elaborates on the findings of arXiv:2605.26234v2: a joint work with Marco Usula, where we introduced a machine learning

tutorialsarxiv-cs-ai
3 Aug 2026
Agents

Behind the scenes: How we build, test, and scale Google Agent Skills

DGX agent

AI agents are only as good as the instructions and context you give them. When we launched Google Agent Skills, our goal was simple: encode Google Cloud domain knowledge into structured, open-source i

agentsgoogle-cloud-ai
3 Aug 2026
Research

COntExt: Towards Context-Aware Ontology Extension from Operational Metrics

DGX agent

arXiv:2607.29553v1 Announce Type: new Abstract: Organizations increasingly define operational metrics in structured, machine-readable formats to monitor systems, processes, and compliance. These metri

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Cortex Framework v7 is GA: Build agentic workflows without disrupting SAP operations

DGX agent

Businesses want to quickly and safely deploy AI agents to drive revenue, mitigate risk, and optimize capital, all without disrupting mission-critical ERP systems. And to power AI agents, you need more

model-releasesgoogle-cloud-ai
3 Aug 2026
← Previous
1…6061626364…113
Next →