AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
11 Aug 2026

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding

Local AiDGX agent

arXiv:2608.08832v1 Announce Type: new Abstract: Distributed deployment of large vision foundation models often partitions a ViT backbone and exchanges intermediate token features between computing nod

VTO: Visual Tool Orchestration for Video Anomaly Detection

Model ReleasesDGX agent

arXiv:2608.08219v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is a critical yet challenging task due to the complex and diverse nature of real-world scenarios. Traditional deep learn

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

SafetyDGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

Model ReleasesDGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

Who Bridges Safety? Identifying and Targeting Cross-Lingual Shared Safety Pathways

Local AiDGX agent

arXiv:2608.09095v1 Announce Type: new Abstract: Uncovering the internal mechanisms underlying the safety capabilities of large language models (LLMs) is crucial for developing trustworthy artificial i

X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation

Model ReleasesDGX agent

arXiv:2505.11146v3 Announce Type: replace-cross Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition challenge due to the significant

XFeat Revisited: Reproducibility and Evaluation of a Lightweight Image Matcher

Model ReleasesDGX agent

arXiv:2608.09519v1 Announce Type: new Abstract: We present a reproducibility study of XFeat, a lightweight local feature extractor and matcher designed to identify corresponding points across images e

10 Aug 2026

An AI4AI Framework for Visual Token Pruning

SafetyDGX agent

arXiv:2608.07193v1 Announce Type: cross Abstract: Visual-token pruning can substantially reduce the inference cost of multimodal large language models (MLLMs), yet existing methods largely rely on fix

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration

SafetyDGX agent

arXiv:2608.07419v1 Announce Type: new Abstract: Preference alignment often makes large language models (LLMs) overconfident and poorly calibrated. Traditional post-hoc temperature scaling is inherentl

CADSpotting: Robust Panoptic Symbol Spotting on Large-Scale CAD Drawings

Model ReleasesDGX agent

arXiv:2412.07377v5 Announce Type: replace Abstract: We introduce CADSpotting, an effective method for panoptic symbol spotting in large-scale architectural CAD drawings. Existing approaches often stru

Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests

SafetyDGX agent

arXiv:2608.06908v1 Announce Type: cross Abstract: We propose Zero-phase Component Analysis (ZCA) whitening as a geometric pre-processing step for the Word Embedding Association Test (WEAT). WEAT is a

CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems

SafetyDGX agent

arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-driven interactions that produce emergent

CubicQuant: Parametric Non-Uniform Codebooks for High-Throughput LLM Inference with 1-8-Bit Weights

HardwareDGX agent

arXiv:2608.06763v1 Announce Type: new Abstract: Weight quantization for large-language-model inference must balance adaptive reconstruction levels with representations regular enough for efficient GPU

Degradation-Aware Prompt Learning with Cross-Modal Compensation for Adverse Weather Removal

SafetyDGX agent

arXiv:2608.06939v1 Announce Type: new Abstract: Adverse weather causes diverse and complex image degradations, severely compromising the reliability of computer vision systems. Existing all-in-one res

EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision

Model ReleasesDGX agent

arXiv:2608.07196v1 Announce Type: new Abstract: Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system

Equivariant Sparse Autoencoders: Mechanistic Interpretability of Neural Networks on Symmetric Data

ApplicationsDGX agent

arXiv:2511.09432v2 Announce Type: replace Abstract: Machine learning (ML) models achieve remarkable performance but remain hard to interpret due to their scale and complexity. In particular, their act

Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2608.06381v1 Announce Type: cross Abstract: Explainable AI (XAI) has shown promise for human-agent collaboration, yet results rely on hand-crafted policies in custom environments, limiting gener

GeoDistill-Refine: Silhouette-First Geometry Distillation for Annotation-Free Spacecraft Segmentation

ResearchDGX agent

arXiv:2608.07405v1 Announce Type: cross Abstract: Foundation segmentation models can provide supervision for spacecraft imagery without manual training masks, but their predictions vary with textual p

Georeferencing Non-Gazetteered Place Names using Biological Specimen Records

Model ReleasesDGX agent

arXiv:2608.06884v1 Announce Type: cross Abstract: Biological specimen records collected by natural history institutions constitute a rich source of temporal geographic knowledge, capturing biodiversit

GPT 5.6 Sol High and X-High (Web Chat) feels severely nerfed since 08/06/2026 update

Model ReleasesDGX agent

GPT-5.6 Sol High and X-High (Web Chat) feels severely nerfed since 08/06/2026 update (https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt) I use it for a pretty complex Unreal Engine 5 project (

Grok Imagine is focused on professional usefulness, fun for consumers and overall ease-of-use

Model ReleasesDGX agent

Grok Imagine is focused on professional usefulness, fun for consumers and overall ease-of-use Grok 4.5 just ranked #1 on Design Arena for Daily Usage - measuring the average unique users actually inte

Hidden Gauge Controls Feature Specialization in ReLU Networks

Model ReleasesDGX agent

arXiv:2608.06766v1 Announce Type: cross Abstract: Training changes a network's predictions while allocating task-relevant structure across its internal units. In an overparameterized ReLU network, sev

Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design

Local AiDGX agent

arXiv:2608.07091v1 Announce Type: cross Abstract: Edge Artificial Intelligence (Edge AI) enables the deployment of AI models directly on local edge devices, while such deployments are subject to stric

InsertFuse: A Unified Framework for Multi-Category Reference-Guided Image Insertion

Model ReleasesDGX agent

arXiv:2608.06490v1 Announce Type: new Abstract: We present InsertFuse, a unified framework for multi-category reference-guided image insertion. Its key idea is to decouple category-specific expertise

Introducing the Developer Device Platform for agentic mobile app development

Model ReleasesDGX agent

Most enterprises connect with their customers through a device. Whether it’s using a mobile app to order a product, contact customer service, view content, or manage their account, the customer experi

Investigating Quantum-Embedded Transformers on Classical Datasets for Cross-Modality Classification

ResearchDGX agent

arXiv:2608.06846v1 Announce Type: cross Abstract: We test whether a parameterized quantum circuit (PQC) improves a hybrid quantum-classical model's performance on classical datasets, using an interfac

Iterative Training of Physics-Informed Neural Networks with Fourier-enhanced Features

Model ReleasesDGX agent

arXiv:2510.19399v2 Announce Type: replace Abstract: Spectral bias, the tendency of neural networks to learn low-frequency features first, is a well-known issue with many training algorithms for physic

Literally for years I have been warning everyone: @GaryMarcus, December, 2022: “GPT-4 will still, like its predecessors, be a bull in a chin…

Model ReleasesDGX agent

Literally for years I have been warning everyone: @GaryMarcus, December, 2022: “GPT-4 will still, like its predecessors, be a bull in a china shop, reckless and hard to control.” @GaryMarcus, November

Local Epistemic Uncertainty Guided Active Sampling for Plug-and-play Diffusive Image Restoration

ResearchDGX agent

arXiv:2608.06981v1 Announce Type: new Abstract: Diffusion models have demonstrated remarkable effectiveness in image restoration tasks. However, when guiding image reconstruction, existing Diffusion M

Measuring Concept Content in Text from LLM Activations: ESG Evidence from Concept Vectors and Linear Probes

ResearchDGX agent

arXiv:2608.07208v1 Announce Type: cross Abstract: Existing measures of how much a text is about a concept read the surface of the text: dictionary word shares, topic proportions, embedding similaritie

Multi-Perspective Triad Interaction Graph Neural Network for Cognitive Distortion Detection

SafetyDGX agent

arXiv:2608.06785v1 Announce Type: new Abstract: Cognitive distortion detection is a key task in computational mental health, yet existing approaches often overlook the psychological structure of disto

Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity

Model ReleasesDGX agent

arXiv:2608.07120v1 Announce Type: new Abstract: Many flow-based video frame interpolation (VFI) methods synthesize an intermediate frame by estimating optical flow fields, warping the two input frames

PACE: Primitive-Aware Code Evolution for Automated Algorithm Design

ResearchDGX agent

arXiv:2608.07395v1 Announce Type: cross Abstract: Large Language Model (LLM)-based automated algorithm design typically evolves algorithms as complete, indivisible programs. While this whole-program p

Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers

Model ReleasesDGX agent

arXiv:2608.07436v1 Announce Type: new Abstract: Under the standard split, Muon gets hidden matrices and AdamW embeddings/output head. Muon groks modular addition faster, but its solutions do not hold.

ReGraph: Learning to Generate Recipe Graphs from Food Images

ResearchDGX agent

arXiv:2608.06917v1 Announce Type: new Abstract: Recent Large Multimodal Models (LMMs) have achieved impressive performance in recipe generation from food images.However, cooking is a structured transf

Retrieval-Constrained Policy Optimization for Attack Technique Extraction from Cyber Threat Intelligence

Model ReleasesDGX agent

arXiv:2608.06778v1 Announce Type: cross Abstract: Mapping cyber threat intelligence (CTI) text to MITRE ATT&CK techniques is essential for structured threat analysis, yet manual annotation is costly a

SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization

Model ReleasesDGX agent

arXiv:2602.12187v2 Announce Type: replace-cross Abstract: Search-Augmented Generative Engines (SAGE) have emerged as a new paradigm for information access, bridging web-scale retrieval with generative

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents,…

Model ReleasesDGX agent

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents, edit videos, work with 3D/CAD, and more. From multimodal mod

SkySeaLand: A Wide-Format Satellite Transportation Benchmark with an Ultra-Lightweight Detection Baseline

Model ReleasesDGX agent

arXiv:2608.07382v1 Announce Type: new Abstract: Satellite object detection is challenged by small targets and wide-format scenes that lose detail under standard square-input resizing. We introduce Sky

Stream Learning: Partition-Fair Gossip Learning Without Tokens

Local AiDGX agent

arXiv:2608.06946v1 Announce Type: cross Abstract: In gossip learning, a network of nodes trains a shared model collaboratively, without a central coordinator, by repeatedly exchanging parts of their l

Transformers are famously bad at arithmetic, so I set one's weights by hand (no training) and it multiplies with 100% accuracy [P]

Model ReleasesDGX agent

Obviously nobody needs a transformer that's good at multiplication. I wanted to know whether a stock transformer could do exact arithmetic if I chose its weights directly. I implemented the grade-scho

Uncovering expert objectives in production planning via inverse optimization: An industrial case study

TutorialsDGX agent

arXiv:2608.07398v1 Announce Type: cross Abstract: Production planning in the manufacturing industry often relies on the use of optimization models, but defining an appropriate objective function can b

WaveFreqAnchor: Wave-Structural Anchoring and Frequency Correction Diffusion for Training-Free Face Restoration

ApplicationsDGX agent

arXiv:2608.06717v1 Announce Type: new Abstract: Diffusion-based face restoration that adjusts the sampling trajectory of pre-trained diffusion models has achieved remarkable progress. However, existin

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

Model ReleasesDGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

When GNNs Fail: Quantifying and Overcoming Temporal Correlation Volatility in Time Series

ApplicationsDGX agent

arXiv:2608.07333v1 Announce Type: new Abstract: Modeling multivariate time series by representing them as graphs, where individual series act as nodes and pairwise temporal corre- lations serve as edg

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

Model ReleasesDGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

You can now try offline computer use with cua-driver + Muse Glimmer, through our friends at @ollama 🦙 How-to: https://cua.ai/docs/how-to-gu…

Local AiDGX agent

Francesco @francedot reported that the first fully offline computer‑based LLM experience was achieved by running Muse Glimmer 30B (from @AIatMeta) locally on macOS, using Cua Driver to control native

9 Aug 2026

not wrong!

AgentsDGX agent

In a recent Twitter exchange, Luca Ambrogioni stated that large‑language models (LLMs) likely possess fundamental limitations that may be obscured by progress in reasoning and agentic pipelines, thoug

The future of FDE work seems closely related with all work around evals/posttraining/RL envs. FDEs are effectively responsible for the follo…

Model ReleasesDGX agent

The future of FDE work seems closely related with all work around evals/posttraining/RL envs. FDEs are effectively responsible for the following: 1. Define the business problem. 2. Codify the business

8 Aug 2026

DeepSeek V4 Flash 0731 appreciation post

Model ReleasesDGX agent

I’m running DSV4F 0731 on dual spark, and honestly… wow. It’s an absolute workhorse, and the benchmarks are real. Everyday tasks with Hermes agent? Effortless. Coding tasks with OpenCode? I’m genuinel

IDE with Locall LLMs?

Local AiDGX agent

What IDE are you using. its another problem area for me . I usually use VSCode , but with local llms I have not found an extension which works optimally VSCode CoPilot chat with Ollama: CoPilot bloats

Is anyone else finding DeepSeek-V4-Flash unreliable for non-coding tasks?

Model ReleasesDGX agent

(I am not a native speaker, written by myself, so please bear with me) I really want to like DeepSeek-V4-Flash-0731. But it has serious flaws that don't align with the high score on intelligence bench

PSA for anyone with multiple V620's or other gfx1030 cards having problems making llama.cpp tensor split work -- set '-ub 384' and -b to a multiple of that depending on number of GPUs

Model ReleasesDGX agent

Basically what the title says. For me, it would always crash and burn trying to use tensor split. Apparently, there's some bug where GPU memory gets corrupted with the default microbatch (512) or high

7 Aug 2026

100% Local RAG Without Internet and Without Ollama

Model ReleasesDGX agent

Build a 100% offline fast Retrieval Augmented Generation (RAG) system that runs without an internet connection, without cloud APIs, without OpenAI/Ollama Published a video where you can build a fully

A llama.cpp PR makes Q2_0 3.0–3.6x faster on x86 CPUs, 8B decode goes 2.39 → 8.20 tok/s

Model ReleasesDGX agent

I was going through the current llama.cpp CPU PRs and #26348 stood out because this isn't the usual +5% kernel optimization. It adds an x86 VNNI implementation for the Q2_0 × Q8_0 dot product, and the

A Unified Risk View of Uncertainty: Posterior Risk for Disentanglement and Evaluation Beyond Proxies

Model ReleasesDGX agent

arXiv:2608.05995v1 Announce Type: new Abstract: Reliable uncertainty estimates are critical in safety-sensitive applications, where understanding the sources of predictive uncertainty is essential. Th

AMD Acquires Taalas to Advance Compute Solutions for Rapidly Growing AI Inference Market

Local AiDGX agent

Press My earlier prediction that Tesla would buy them completely missed the mark. With AMD focusing heavily on the enterprise side, the idea of consumer-facing hot-swappable AI model chips looks prett

Automatic Detection of Deaths from Social Networking Sites

ResearchDGX agent

arXiv:2608.05183v1 Announce Type: cross Abstract: This dissertation analysed and discussed the differences in linguistic characteristics between pre-mortem and post-mortem social media content, and re

Autoscaling peaky LLM inference workloads is completely different than autoscaling something like a web service. I wrote a deepdive covering…

ToolsDGX agent

Zain (@zainhas) published a detailed article on August 7, 2026 explaining that autoscaling for highly peaky large‑language‑model (LLM) inference is fundamentally different from autoscaling conventiona

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

Model ReleasesDGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

← Previous
1…506507508509510…1062
Next →