AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training

DGX agent

arXiv:2608.02391v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents produce long, multi-turn trajectories, making gradient-based post-training memory-intensive. Evolution st

model-releasesarxiv-cs-lg
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

CrossProjection: Geometric Grounding Beyond Viewpoint Change in Architectural Drawings

DGX agent

arXiv:2608.00473v1 Announce Type: cross Abstract: Architectural drawings violate the usual assumption behind multi-view reasoning: plans and sections are cuts, while elevations are facade projections,

model-releasesarxiv-cs-cl
4 Aug 2026
Tutorials

Does Machine 'know' interpersonal pragmatics? Evidence from MARBERT's learning of emoji pragmatics in Arabic digital discourse

DGX agent

arXiv:2608.01174v1 Announce Type: new Abstract: This study examines Transformer-based models' ability to learn emoji pragmatics in Arabic digital discourse (ADD), providing evidence from MARBERT's beh

tutorialsarxiv-cs-cl
4 Aug 2026
Tutorials

Entity-Faithful Repair of Synthetic Supervision for Zero-Shot Image Captioning

DGX agent

arXiv:2608.00994v1 Announce Type: cross Abstract: Zero-shot image captioning aims to generate image descriptions without annotated image-text pairs. Recent approaches exploit text-to-image models to s

tutorialsarxiv-cs-cl
4 Aug 2026
Model Releases

Ethyca launches Astralis to govern enterprise AI agents in real time

DGX agent

Data privacy engineering company Ethyca Inc. today launched Astralis, a platform that governs how enterprise artificial intelligence models and agents use company data in real time. The company is pit

model-releasessiliconangle
4 Aug 2026
Model Releases

EulerLoRA: Rank-Driven Jump Dynamics for Calibrated Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2608.01142v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning, but standard LoRA produces a single deterministic model and does not directly suppor

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

FL-OA: A Byzantine-Robust Federated Learning Framework with Outsourced Auditing for Intelligent Devices

DGX agent

arXiv:2608.01095v1 Announce Type: new Abstract: Federated learning (FL) enables multiple intelligent devices to collaboratively train a high-accuracy model without sharing raw data. However, due to it

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Geometric Analysis of Token Selection in Multi-Head Attention

DGX agent

arXiv:2602.01893v2 Announce Type: replace-cross Abstract: We present a geometric framework for analysing multi-head attention in large language models (LLMs). Without altering the mechanism, we view s

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Grounding and Explaining Visual Evidence for AI-Generated Image Detection in Human-Centric Scenes

DGX agent

arXiv:2608.01988v1 Announce Type: new Abstract: Rapid advances in image generation models call for interpretable AI-generated image detection methods that not only determine authenticity but also prov

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

How Benchmarks and Evaluation Protocols Shape Conclusions in Provenance-Based Intrusion Detection

DGX agent

arXiv:2608.01454v1 Announce Type: cross Abstract: Provenance-based intrusion detection systems (PIDS) frequently report strong performance, but the conclusions drawn from these results can be highly s

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

inclusionAI/Ling-3.0-flash · Hugging Face

DGX agent

The Ling-3.0-flash MoE is now open-weighted at 124B A5B params. I know the original announcements were before the Kimi K3, DeepSeek-V4-Flash and Qwen3.8 hype, but this model might still have a good ni

model-releasesr-localllama
4 Aug 2026
Model Releases

LexisNexis opens customer innovation lab driven by AI to change the future of legal work

DGX agent

LexisNexis Legal & Professional, a division of RELX plc, today announced the opening of its Customer Innovation Lab in New York City, which will deliver a new model for how legal artificial intelligen

model-releasessiliconangle
4 Aug 2026
Model Releases

Linear Multi-Timescale Retention as a Memory-Efficient Vision-Language Bridge

DGX agent

arXiv:2608.01614v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face a critical computational bottleneck when processing high-resolution imagery due to the O(N^2) memory complexity of So

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Loggia dei Lanzi: AI Thermography Enhancement Comparisons through 3D Photogrammetry

DGX agent

arXiv:2608.02404v1 Announce Type: new Abstract: The Loggia dei Lanzi in the Piazza della Signoria is one of Florence's most prominent structures visited by millions every year. Its construction histor

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

MBO Scheme for Local Chan--Vese Segmentation

DGX agent

arXiv:2608.00893v1 Announce Type: new Abstract: Robust to intensity inhomogeneity, the local Chan--Vese (LCV) model extends the classical Chan--Vese (CV) image segmentation method by incorporating loc

local-aiarxiv-cs-cv
4 Aug 2026
Model Releases

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

DGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

model-releasesarxiv-cs-cv
4 Aug 2026
Tutorials

Noise-Robust Conditional Flow Matching: Generating Clean Samples from Noisy Datasets

DGX agent

arXiv:2608.00064v1 Announce Type: new Abstract: Generative models learn the statistical properties of their training data, so high-quality generation depends on clean and representative datasets. In s

tutorialsarxiv-cs-cv
4 Aug 2026
Model Releases

PhysAgent: A Multi-Agent Framework for Reliable Remote Heart Rate Estimation

DGX agent

arXiv:2608.00066v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact heart-rate estimation from facial videos, but its weak physiological signal is easily corrupted b

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

PipeNetwork/minimax-h3-mlx

DGX agent

PipeNetwork/minimax-h3-mlx MiniMax released MiniMax-H3 two days ago - they describe it as a 'a general-purpose, omni-modal generative system', which in practice means it accepts text, images, audio an

model-releasessimon-willison
4 Aug 2026
Research

Pretraining on Call Graphs: When Binary Analysis Tasks Profit From Context

DGX agent

arXiv:2608.02084v1 Announce Type: cross Abstract: Binary function embedding models are trained to encode the semantics of binary code in such a way that they can be generalized to a variety of reverse

researcharxiv-cs-lg
4 Aug 2026
Research

Pruned BPE: Post-training Visibility Pruning and Token Reallocation for Byte Pair Encoding

DGX agent

arXiv:2608.00837v1 Announce Type: new Abstract: Byte Pair Encoding (BPE) is widely used for subword tokenization, but standard BPE exposes every learned merge token to the downstream model, including

researcharxiv-cs-cl
4 Aug 2026
Research

QR-Erase: Efficient Subspace-Based Machine Unlearning with Layer Localization

DGX agent

arXiv:2608.01422v1 Announce Type: new Abstract: Machine unlearning seeks to remove targeted information from trained models without requiring costly retraining. Existing optimization-based methods oft

researcharxiv-cs-cl
4 Aug 2026
Model Releases

QuerySplat: Decoupling Geometry and Appearance Representations in 3DGS Prediction

DGX agent

arXiv:2608.01186v1 Announce Type: new Abstract: While feed-forward 3D Gaussian Splatting (3DGS) enables efficient 3D reconstruction, achieving high-fidelity rendering remains challenging. Existing pix

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Qwen3.8-Max, better and cheaper. Try it out. 👀

DGX agent

Qwen3.8-Max, better and cheaper. Try it out. 👀 We ran a test between the new Qwen3.8-Max, Opus 5 and GPT-5.6 Sol. 3 models. same prompt. one-shot with the /design command. Reviewed gameplay features,

model-releasesqwen--x
4 Aug 2026
Research

ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection

DGX agent

arXiv:2608.01793v1 Announce Type: new Abstract: Unified anomaly detection requires modeling highly heterogeneous normal data without access to anomalous samples. While foundation models like DINOv2 pr

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Residual-Based Adaptive Kalman Filtering for Legged Robot State Estimation

DGX agent

arXiv:2608.02316v1 Announce Type: new Abstract: State estimation is a key component in model-based control of walking robots and, more broadly, applicable wherever hidden variables must be inferred. T

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforcement Learning

DGX agent

arXiv:2608.01418v1 Announce Type: cross Abstract: Autoregressive rollout generation is a major computational cost in reinforcement learning for large language models. Reusing each rollout batch for ad

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Riemannian Attention Mechanisms for Transformers: A Theoretical Framework and Architecture Design

DGX agent

arXiv:2608.01283v1 Announce Type: new Abstract: All Transformer-based large language models compute attention via the Euclidean inner product, an architectural choice that Dong et al. (2021) proved ca

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs

DGX agent

arXiv:2608.01473v1 Announce Type: cross Abstract: Multimodal large language models (MLLM) for surgical scene understanding typically inject hundreds of dense visual tokens into a language model, leadi

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

ST-LoRA: Single Trajectory LoRA Ensemble for Uncertainty Aware Agricultural Segmentation

DGX agent

arXiv:2608.01530v1 Announce Type: new Abstract: Reliable decision-support in digital agriculture requires accurate predictions and well-calibrated uncertainty estimates, particularly for dense predict

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

TBSG-Net: Temporal Bipartite Scene Graph Network for Fine-Grained Video Moment Retrieval

DGX agent

arXiv:2608.02056v1 Announce Type: new Abstract: Recent advances in proposal-free Video Moment Retrieval (VMR) have highlighted the effectiveness of Static Scene Graphs (SSGs). By modeling objects and

local-aiarxiv-cs-cv
4 Aug 2026
Local Ai

TELLER: Non-intrusive Cross-Layer Root-Cause Analysis for LLM Inference

DGX agent

arXiv:2608.01975v1 Announce Type: cross Abstract: Large language model (LLM) inference has evolved from an offline workload into a continuously operated software service, yet root-cause analysis remai

local-aiarxiv-cs-cl
4 Aug 2026
Model Releases

Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping

DGX agent

arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generali

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

UCBound-Net: Uncertainty-Guided Boundary-Aware Continual Learning for Domain-Incremental Ultrasound Segmentation

DGX agent

arXiv:2608.01518v1 Announce Type: new Abstract: Continual learning in clinical imaging faces a dual challenge: a model must assimilate knowledge from new anatomical domains while retaining representat

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding

DGX agent

arXiv:2608.00036v1 Announce Type: new Abstract: Real-world document tasks often ask professionals to answer questions from annual reports, regulations, clinical guidelines, and technical manuals that

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

XSPA: Crafting Imperceptible X-Shaped Sparse Adversarial Perturbations for Transferable Attacks on VLMs

DGX agent

arXiv:2603.28568v2 Announce Type: replace Abstract: Vision-language models (VLMs) share visual-textual representations across zero-shot classification, image captioning, and visual question answering

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

An Ontology-Guided, Deduplication-Aware Extraction Layer for Knowledge Graph Construction from Heterogeneous Documents

DGX agent

arXiv:2607.28662v1 Announce Type: new Abstract: Large language models extract entities and relationships from unstructured documents fluently but inconsistently: type vocabularies fracture across docu

local-aiarxiv-cs-ai
3 Aug 2026
Agents

AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance

DGX agent

arXiv:2512.05131v2 Announce Type: replace-cross Abstract: Active 3D reconstruction enables an agent to autonomously select viewpoints to efficiently obtain accurate and complete scene geometry, rather

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

b10238

DGX agent

model: MTP support for Qwen3-Next (#25589) mtp for qwen3nex fix for python type-check Fix to compute num_mtp from directly mtp layer define opt_num_mtp_layers in _QwenMtpMixin and fix some comments Fi

model-releasesllama-cpp-releases
3 Aug 2026
Model Releases

b10242

DGX agent

CUDA: Add backend sampler for penalties sampler (#25262) sampling: enhance penalty handling in common_sampler_init Set default value for penalty_last_n based on model context if not specified. Ensure

model-releasesllama-cpp-releases
3 Aug 2026
Model Releases

b10244

DGX agent

model: M3: Move MSA into a new memory implementation (#26338) Move MSA logic from llama-kv-cache into llama-kv-cache-msa cont : minor cont : ws fix Co-authored-by: Georgi Gerganov ggerganov@gmail.com

model-releasesllama-cpp-releases
3 Aug 2026
Model Releases

BWM: A Low-Cost High-Fidelity World Simulator for Robot Learning

DGX agent

arXiv:2607.29302v1 Announce Type: cross Abstract: Reliable robot learning requires a world simulator that can predict action consequences before execution on physical hardware, including risky and fai

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents

DGX agent

arXiv:2607.19338v2 Announce Type: replace Abstract: Coding agents increasingly operate in executable environments where a failed attempt produces actionable feedback rather than merely an incorrect an

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation

DGX agent

arXiv:2604.05150v2 Announce Type: replace-cross Abstract: We study compiled AI, a paradigm in which large language models generate executable code artifacts during a compilation phase, after which wor

safetyarxiv-cs-ai
3 Aug 2026
Research

Dense Temporal Contrast Synthesis via Conditioned Latent Transport

DGX agent

arXiv:2607.29394v1 Announce Type: cross Abstract: Dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI) is essential for breast cancer management, but reliance on gadolinium-based contrast ag

researcharxiv-cs-ai
3 Aug 2026
Agents

Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning

DGX agent

arXiv:2511.22226v2 Announce Type: replace Abstract: The standard theory of model-free reinforcement learning assumes that the environment dynamics are stationary and that agents are decoupled from the

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

I compared MinerU, Granite-Docling, and PaddleOCR-VL on 12 PDF-parsing capabilities using 6 document types

DGX agent

I tested them by sending the 6 documents, each meant to represent a different document type, through my own webapp and comparing every output against the source. All ran on the same L4 GPU. The docume

model-releasesr-localllama
3 Aug 2026
Local Ai

I got tired of ad-filled mobile wrappers for Ollama, so I built PocketLLM Lite an open-source, offline Android client (Local GGUF, SKILL.md plugins, local RAG)

DGX agent

Hey, Like a lot of people here, I use local models via Ollama on my desktop/server and wanted a mobile client that actually felt responsive, worked offline, and respected privacy. Most apps on the Pla

local-air-ollama
3 Aug 2026
← Previous
1…495496497498499…1359
Next →