AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,987 results
19 Aug 2026

MCTS-KBQA: Monte Carlo Tree Search with Information Gain Rewards for Knowledge Base Question Answering

TutorialsDGX agent

arXiv:2502.13428v2 Announce Type: replace-cross Abstract: This work investigates how to improve large language model (LLM)-based reasoning for knowledge base question answering (KBQA) via Monte Carlo

PACE: Policy-Attested Contract Execution for Safe AI Agents in Decentralized Finance

Model ReleasesDGX agent

arXiv:2608.17220v1 Announce Type: cross Abstract: Autonomous AI agents are emerging as interfaces for decentralized finance (DeFi) actions such as swaps, lending operations, and yield management. Beca

PathoArgus: Advancing Evidence-Grounded Long-Context Visual Reasoning across Gigapixel Whole-Slide and Multi-Slide Case Contexts

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.17607v1 Announce Type: new Abstract: Whole-slide pathology reasoning requires models to integrate gigapixel-scale visual evidence across complete case-linked slides, yet current question-an

PROBE: Manipulation-Grounded Visual Question Answering with VLM Agents

AgentsDGX agent

arXiv:2608.17129v1 Announce Type: new Abstract: Vision-language Models (VLMs) excel at 2D grounding, spatial reasoning and agentic tool-based planning in static scenes. However, consider asking a home

Qwen 3.5 cybersecurity

Model ReleasesDGX agent

Hi guys , this is a fine tune of qwen 3.5 4b on cybersecurity , specifically malware. This model has been trained on 2.6k rows of simple but effective malware code , though its still significantly lim

Rapid Debris-Volume Estimation from Post-Hurricane Aerial Imagery

Model ReleasesDGX agent

arXiv:2608.17165v1 Announce Type: new Abstract: Hurricane debris removal is planned, contracted, and federally reimbursed on the basis of volume estimates, yet operational practice still relies on par

Recirculation

ResearchDGX agent

arXiv:2608.17981v1 Announce Type: new Abstract: We describe an inference-time architectural enhancement for off-the-shelf foundation models that markedly reduces perplexity and boosts accuracy across

Reflex-Guard: A Low-Latency Guardrail for LLM Prompt Safety Using Dense Semantic Embeddings

Model ReleasesDGX agent

arXiv:2608.17556v1 Announce Type: cross Abstract: Large Language Models (LLMs) in real-world applications often face the risks of specially crafted prompts designed to bypass the safety controls. Exis

RoBell-RVFL: A Robust Generalized Bell Random Vector Functional Link Network

Model ReleasesDGX agent

arXiv:2608.16965v1 Announce Type: new Abstract: The dominance of majority classes in real-world datasets poses a fundamental challenge to randomized neural networks, often biasing decision boundaries

Scanline-Aware Animatable Gaussian Avatars from Rolling-Shutter Videos

Model ReleasesDGX agent

arXiv:2608.17314v1 Announce Type: new Abstract: Animatable human avatars are routinely reconstructed from multi-view video under a silent assumption: that every pixel of a frame observes the same inst

Serverless Apache Spark on Google Cloud: Architecture Choices & AI Troubleshooting

Model ReleasesDGX agent

In modern enterprise data engineering, Apache Spark remains a cornerstone framework for processing massive datasets at scale. However, managing infrastructure such as provisioning clusters, tuning YAR

SIGMA: SHAP-Guided Implicit-Trajectory Generation for Metadata-Free LLM-Based AutoFE

Model ReleasesDGX agent

arXiv:2608.17948v1 Announce Type: cross Abstract: Recent research has leveraged Large Language Models (LLMs) to enhance Automated Feature Engineering (AutoFE) through semantic descriptions and traject

SPSA Hyperparameter Tuning for Variational Quantum Natural Language Inference

Model ReleasesDGX agent

arXiv:2608.16939v1 Announce Type: cross Abstract: Training variational quantum models requires choosing between parameter-shift gradients, which are exact but cost O(P) forward evaluations, and simult

We quantized the new Ornith 1.5 9B and 35B-A3B

Model ReleasesDGX agent

ornith lab dropped new ornith 1.5 today, a 9B dense with vision and a 35B-A3B MoE, both MIT, trained on a loop that generates its own tasks. in addition there was giant 397b model, but we didn't quant

When will there be a frontier level llm with updateable engrams or fixed engrams?

Model ReleasesDGX agent

DeepSeek released on a paper on pretrained engrams in January. I’m surprised ds didn’t release engrams with v4 pro. When will ds or another lab release a fixed engram model? Fixed engrams will be the

18 Aug 2026

A Policy Algebra for Trust-Preserving Agentic AI Execution

Model ReleasesDGX agent

arXiv:2608.16402v1 Announce Type: new Abstract: Large language model-based agentic frameworks primarily optimize capability: whether an agent can reason, retrieve information, call tools, delegate wor

A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains

ResearchDGX agent

arXiv:2608.15809v1 Announce Type: new Abstract: Behavioral accuracy, linear decodability, and successful activation interventions do not by themselves show that a model carries an operation-level stru

A survey of AI-generated voices and their detection

Model ReleasesDGX agent

arXiv:2608.15411v1 Announce Type: new Abstract: The ability of artificial intelligence (AI) models to generate highly realistic human voices has advanced rapidly. These technologies power accessibilit

A Unified Mamba--MoE Surrogate for Closed-Loop Simulation and Measurement-Window Forecasting of Inverter Transients

ResearchDGX agent

arXiv:2608.15051v1 Announce Type: new Abstract: This paper proposes a Mamba surrogate model with mixture-of-experts (MoE) routing to represent the transient dynamics of inverter-based resources. A Mam

BabelSteering: Multilingual Safety Alignment via English Steering Vectors

Model ReleasesDGX agent

arXiv:2608.16577v1 Announce Type: new Abstract: Large language models (LLMs) are deployed globally in high-stakes settings, yet most safety research and alignment efforts remain concentrated on Englis

Beam-Wise Statistical Background Subtraction for Static Roadside LiDAR: A Cross-Sensor Benchmark Study

Model ReleasesDGX agent

arXiv:2608.14868v1 Announce Type: new Abstract: Background subtraction is a key preprocessing step for infrastructure-based LiDAR perception, enabling efficient isolation of dynamic traffic participan

Beyond Thresholds: A Quality-Aware Decision Intelligence Framework for Cold Chain IoT Systems

Model ReleasesDGX agent

arXiv:2608.15082v1 Announce Type: new Abstract: Cold chain logistics has advanced technologically, yet most deployed systems remain reactive monitors, not decision-making agents: thresholds trigger al

Bounded Semantic Planning and Deterministic Compilation for Reliable Enterprise Text-to-SQL

Model ReleasesDGX agent

arXiv:2608.16663v1 Announce Type: cross Abstract: Direct text-to-SQL asks a language model to do two jobs: interpret the business question and construct the complete relational query. In enterprise sc

Brex’s summer benchmark shows how quickly the infrastructure behind the agent economy is growing. 14 of the 25 fastest-growing software vend…

Model ReleasesDGX agent

Brex’s summer benchmark shows how quickly the infrastructure behind the agent economy is growing. 14 of the 25 fastest-growing software vendors serve teams building AI products, and Together AI ranks

Can Neural Networks Learn by Experimenting on Themselves? Self-Interventional Learning from Functional Consequences to Predictive Self-Knowledge

SafetyDGX agent

arXiv:2608.14894v1 Announce Type: new Abstract: Machine-learning systems usually model external data, while their internal functional organization is analyzed by external observers. This work introduc

Chameleon: An Adaptive AI-Driven Honeypot Architecture Using Threat-Calibrated Particle Swarm Optimization and Semantic Deception Rapidly-Exploring Random Trees

Model ReleasesDGX agent

arXiv:2608.15407v1 Announce Type: cross Abstract: An invariant behavioral profile is the defining vulnerability of traditional honeypot installations: a skilled adversary can confirm the presence of a

ClawGym II: Exploring Black-Box RL on Agent Harness

Model ReleasesDGX agent

arXiv:2608.16798v1 Announce Type: cross Abstract: Agent harnesses have substantially improved performance on long-horizon tasks by coordinating agent interactions with the environment. However, reinfo

Comprehensive Benchmarking of Deep Learning Architectures for Lung Cancer Histopathology

Model ReleasesDGX agent

arXiv:2608.15915v1 Announce Type: cross Abstract: Lung cancer remains the leading cause of cancer-related mortality worldwide, while histopathological diagnosis is often affected by inter-observer var

Data-knowledge dual-driven intelligent framework for full-chain, experiment-efficient synthesis of 2D dendrites

Model ReleasesDGX agent

arXiv:2603.16959v2 Announce Type: replace-cross Abstract: Exemplified by the chemical vapor deposition growth of two-dimensional dendrites, which has potential applications in catalysis and presents a

DeltaLog: Deferred Materialization of Recurrent States for Linear Attention Decoding

ResearchDGX agent

arXiv:2608.15533v1 Announce Type: cross Abstract: Linear attention models eliminate the quadratic prefix computation and context-growing KV cache of softmax attention by replacing pairwise token inter

DFlash 2 available for Qwen 3.8 27B and Muse Glimmer

Model ReleasesDGX agent

Apparently a second version of DFlash from the original authors of DFlash GGUF quants are already made available with an accompanying llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/27342 sub

Do Visual Grounding Decoders Need Feed-Forward Networks? A Controlled Study over Frozen Vision-Language Features

Model ReleasesDGX agent

arXiv:2608.15061v1 Announce Type: new Abstract: Do feed-forward networks (FFNs) in visual grounding decoders add essential computation once a pretrained vision-language model has already encoded image

Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

Model ReleasesDGX agent

arXiv:2608.16889v1 Announce Type: cross Abstract: Long-horizon robot manipulation chains many contact-rich skills into one multi-stage task. Vision-language-action (VLA) models increasingly master the

FabriMAE I Trust Myself? Self-Evaluating VLA Action Generation with Markov Attention Entropy

Model ReleasesDGX agent

arXiv:2608.16697v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) integrate visual perception, language instruction, and action generation into end-to-end policies across heterogene

FloodReasonBench: Benchmarking VLM Reasoning Segmentation for Embodied Flood Response at the Edge

Model ReleasesDGX agent

arXiv:2608.15410v1 Announce Type: cross Abstract: Reasoning segmentation enables vision-language models (VLMs) to translate mission-relevant language requests into pixel-level visual grounding, offeri

From Generalist to Specialist: A Context-Fusion Framework for Endoscopic Polyp Reporting with a Frozen VLM

Model ReleasesDGX agent

arXiv:2608.15580v1 Announce Type: new Abstract: Reliable endoscopic polyp reporting requires integrating quantitative lesion sizing, standardized Paris classification, and clinically meaningful morpho

Graph Machine Learning: An Opportunity for Power Systems

Model ReleasesDGX agent

arXiv:2608.16494v1 Announce Type: cross Abstract: Modern power systems face growing operational complexity driven by the integration of renewable energy sources, decentralization, and the need for rea

Handoff-H1: An Orchestrated Vision-Agent System for Material Quantity Takeoff from Construction Blueprints

Model ReleasesDGX agent

arXiv:2608.15032v1 Announce Type: cross Abstract: Converting a set of architectural blueprints into a complete material quantity takeoff requires visual perception across drawing sheets, dimensional a

Honeyquest for LLMs: Rethinking Cyber Deception for AI Attackers

Model ReleasesDGX agent

arXiv:2606.21037v2 Announce Type: replace-cross Abstract: The empirical foundation of cyber deception relies on human-centered hypotheses, but the rapid emergence of autonomous, AI-enabled attackers c

I pushed Qwen3.8-27B to 124 tps on a single request on a RTX 3090

Model ReleasesDGX agent

Two days ago I released a hyper-optimized Qwen3.8-27B inference engine for an RTX 3090 (82 tps single request, 672 peak) - yesterday's update took that to 99 tps single-user / ~1,000 tps at 64 concurr

I tested DFlash2 for Qwen3.8 27B on a 5090

Model ReleasesDGX agent

Here's the DFlash2 announcement, and I was pretty excited for this after trying out DSpark on llama.cpp a few days ago and being somewhat disappointed that it wasn't really working. Anyways, I spent a

iFuzz-Meta: An Interpretable Fuzzy Learning Framework Bridging Top-Down and Bottom-Up Knowledge Integration

Model ReleasesDGX agent

arXiv:2608.14646v1 Announce Type: cross Abstract: Interpretable representation learning remains a key challenge in modern neural computation, particularly when models are expected not only to perform

LangSmith’s new Tuned Evaluators look pretty interesting. Perceived Error can flag agent mistakes in production by picking up on user correc…

Model ReleasesDGX agent

LangSmith’s new Tuned Evaluators look pretty interesting. Perceived Error can flag agent mistakes in production by picking up on user corrections and unresolved conversations. What’s interesting is it

Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization

Model ReleasesDGX agent

arXiv:2608.16072v1 Announce Type: cross Abstract: Reinforcement learning (RL) with group-relative advantages has become the de facto standard for post-training language model reasoners. However, when

Learning to Unlearn: Machine Unlearning via Learning the Unlearning Behaviors

SafetyDGX agent

arXiv:2608.16700v1 Announce Type: cross Abstract: Various machine unlearning techniques have been developed in response to privacy legislation requirements, enabling individuals to exercise their lega

LLMs Get Smarter from Targeted Synthetic Multilingual Data

TutorialsDGX agent

arXiv:2608.15964v1 Announce Type: cross Abstract: Language-specific competency (LSC) is the phenomenon of a language model performing better or worse depending on the language of the prompt. In other

LORA-CRAFT: Cross-layer Rank Adaptation via Frozen Tucker Decomposition of Pre-trained Attention Weights

Model ReleasesDGX agent

arXiv:2602.17510v2 Announce Type: replace-cross Abstract: We introduce LoRA-CRAFT (extbf{C}ross-layer extbf{R}ank extbf{A}daptation via extbf{F}rozen extbf{T}ucker), abbreviated CRAFT throughout, an e

Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2608.14684v1 Announce Type: cross Abstract: LLM judges increasingly evaluate responses against fine-grained rubric checklists. When a sample requires multiple rubrics, current methods typically

New Tuned Evaluators from @LangChain Versioned judges you just 'turn on' in a tracing project The 'Perceived Error' Judge flags conversation…

Model ReleasesDGX agent

New Tuned Evaluators from @LangChain Versioned judges you just 'turn on' in a tracing project The 'Perceived Error' Judge flags conversations where the agent probably messed up (proxy for user feedbac

No Task Fails Every Time: Why One-Shot Audits Are Structurally Blind to Agent Damage

AgentsDGX agent

arXiv:2608.15286v1 Announce Type: cross Abstract: We introduce AgentRelBench, an environment-agnostic reliability instrument that computes ground-truth, severity-priced damage from database state diff

OpenAI lays out new security changes after its AI hacked Hugging Face

Model ReleasesDGX agent

OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments

OvDSGG: End-to-End Open-Vocabulary Dynamic Scene Graph Generation

Model ReleasesDGX agent

arXiv:2608.14835v1 Announce Type: new Abstract: Dynamic scene graphs (DSGs) capture spatio-temporal interactions across videos as langlesubject, predicate, objectrangle triplets, and underpin downstre

Really interesting paper. I recommend it to anyone interested in training agents using existing harnesses. (bookmark it) ClawGym II runs RL …

Model ReleasesDGX agent

Really interesting paper. I recommend it to anyone interested in training agents using existing harnesses. (bookmark it) ClawGym II runs RL through OpenClaw and Claude Code as opaque boxes. A serving

Reasoning-Based Personalized Generation for Users with Sparse Data

Model ReleasesDGX agent

arXiv:2602.21219v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) personalization holds great promise for tailoring responses by leveraging personal context and history. However, re

Representation Is Not Enough: Body-Localized Thermal Evidence for Contactless Stress and Craving Sensing in Opioid Use Disorder

Model ReleasesDGX agent

arXiv:2608.16087v1 Announce Type: new Abstract: Removing wearables from physiological monitoring also removes their supervision: the signal indicating where and when a stress response occurred. Contac

Risk-Adaptive Edge--Cloud Visual Reasoning for Communication-Efficient Autonomous Driving

Local AiDGX agent

arXiv:2608.14991v1 Announce Type: new Abstract: Cloud-hosted vision-language models (VLMs) offer greater contextual reasoning capabilities than smaller onboard models, but frequent visual uploads incr

SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning

Model ReleasesDGX agent

arXiv:2608.16249v1 Announce Type: new Abstract: Machine unlearning in Large Language Models (LLMs) faces a critical trade-off between erasing target knowledge and preserving general utility. We propos

STAR-FL: Secure Federated Learning with Spatial-Temporal Analysis and Robust Aggregation

Model ReleasesDGX agent

arXiv:2608.14861v1 Announce Type: cross Abstract: Data poisoning attacks pose serious security threats to Federated Learning (FL) systems in Computer Vision. Despite growing research attention, two ke

tencent/UI-Mate-27B · Hugging Face

Model ReleasesDGX agent

Overview UI-Mate-27B is an open-weight foundation GUI agent for long-horizon work across applications and operating systems. It observes live screenshots, reasons over the visible state, and produces

THESIS-MoE: Trainable Hierarchical Extraction and SteerIng of Sycophancy in Mixture-of-Experts

Local AiDGX agent

arXiv:2608.15687v1 Announce Type: new Abstract: Sycophancy, the tendency of a language model to change its answer to match a user's stated belief, is a common alignment failure. Existing activation st

← Previous
1…373374375376377…1050
Next →