AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlog
88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
Model Releases

Position: Profiling Game Worlds by Transition Complexity

DGX agent

arXiv:2608.18079v1 Announce Type: new Abstract: Game world modeling (GWM) and reinforcement learning (RL) are often confounded because research papers rarely quantify how difficult the underlying tran

model-releasesarxiv-cs-ai
20 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Same Facts, Different Updates: Inference Setup Shapes LLM Behavior in Medical Allocation

DGX agent

arXiv:2608.18108v1 Announce Type: cross Abstract: Large language models are being incorporated into sensitive and important decision-making processes across nearly all fields. While prior work studies

safetyarxiv-cs-ai
20 Aug 2026
Model Releases

Task-Conditioned Least-Privilege Learning for Executable Terminal and MCP Agents

DGX agent

arXiv:2608.18351v1 Announce Type: cross Abstract: Tool-using large language-model agents can complete a task while exercising authority that the user did not grant or the task does not need, causing e

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

VTONQA: A Multi-Dimensional Quality Assessment Dataset for Virtual Try-on

DGX agent

arXiv:2601.02945v2 Announce Type: replace Abstract: With the rapid development of e-commerce and digital fashion, image-based virtual try-on (VTON) has attracted increasing attention. However, existin

model-releasesarxiv-cs-cv
20 Aug 2026
Safety

Certified but Private: Scalable Zero-Knowledge Proofs for Neural Network Guarantees

DGX agent

arXiv:2608.17070v1 Announce Type: new Abstract: With the growing deployment of machine learning models, formal guarantees of the robustness and fairness of these models have become increasingly import

safetyarxiv-cs-lg
19 Aug 2026
Model Releases

CKAA: Cross-subspace Knowledge Alignment and Aggregation for Robust Continual Learning

DGX agent

arXiv:2507.09471v2 Announce Type: replace Abstract: Continual Learning (CL) empowers AI models to continuously learn from sequential task streams. Recently, parameter-efficient fine-tuning (PEFT)-base

model-releasesarxiv-cs-cv
19 Aug 2026
Research

Encoded but Not Actionable: Auditing the Decode-Generate-Steer Gap in Frozen LLMs for Geometric Constraints

DGX agent

arXiv:2608.17843v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance on structured reasoning tasks, but what they encode and whether it informs model beh

researcharxiv-cs-ai
19 Aug 2026
Model Releases

Foundation Agents Meet Agentic Deep Research: Evidence-Grounded Clinical Code Forecasting

DGX agent

arXiv:2608.17075v1 Announce Type: cross Abstract: Next-encounter ICD forecasting predicts which standardized diagnosis codes will be documented at a future visit from the longitudinal record available

model-releasesarxiv-cs-ai
19 Aug 2026
Local Ai

Post-Train NVIDIA Cosmos 3 Edge for On-Device Robot Control

DGX agent

NVIDIA’s Cosmos 3 Edge is a 4‑billion‑parameter omni‑model (with a 2‑billion‑parameter Nemotron reasoner) that can run on an NVIDIA Jetson Thor, providing on‑device policy inference for robot manipula

local-ainvidia-developer
19 Aug 2026
Local Ai

Predicting Male Domestic Violence Using Explainable Ensemble Learning and Exploratory Data Analysis

DGX agent

arXiv:2403.15594v4 Announce Type: replace-cross Abstract: Domestic violence is commonly viewed as a gendered issue that primarily affects women, which tends to leave male victims largely overlooked. T

local-aiarxiv-cs-lg
19 Aug 2026
Model Releases

Replit Free Mode, powered by @OpenAI GPT-5.6 Luna. Let’s make intelligence accessible to everyone.

DGX agent

Replit has introduced a free mode powered by OpenAI’s GPT‑5.6 Luna model, announced on 19 August 2026. The service offers real‑time AI assistance for coding and collaboration at no cost, aiming to bro

model-releasesopenai--x
19 Aug 2026
Model Releases

SE-MoLoRA: Shared-Expert LoRA Adapters for Domain-Specific Photographic Assessment

DGX agent

arXiv:2608.17514v1 Announce Type: new Abstract: Vision-language models can describe images fluently, but they often fail to provide actionable photographic critique because semantic content and aesthe

model-releasesarxiv-cs-cv
19 Aug 2026
Local Ai

Seeing is Free, Speaking is Not: Uncovering the True Energy Bottleneck in Edge VLM Inference

DGX agent

arXiv:2607.09520v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are the perceptual backbone of embodied AI, but their energy footprint on edge hardware remains poorly understoo

local-aiarxiv-cs-ai
19 Aug 2026
Model Releases

StartupBench: Benchmarking General-Purpose Agents on Market-Validated End-to-End Workflows

DGX agent

arXiv:2608.17800v1 Announce Type: new Abstract: Recent advances in Large Language Models(LLMs) and agents have substantially improved the ability of AI systems to execute complex tasks. Yet existing b

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

Stop Anthropomorphisizing Intermediate Tokens: Qwen3.8 doesn't 'overthink'

DGX agent

Intermediate tokens, called 'thinking' or 'reasoning' actually are nothing like it. Humans do step-by-step reasoning leading to the conclusion. LLMs use intermediate traces to augment their prompt. Th

model-releasesr-localllama
19 Aug 2026
Applications

TabNSM: Neural Sparse Mixer for Tabular Regression

DGX agent

arXiv:2608.18026v1 Announce Type: new Abstract: Large-scale, high-dimensional tabular regression remains challenging: tree-based models are robust but lack end-to-end representation learning, while de

applicationsarxiv-cs-lg
19 Aug 2026
Safety

The Emergence of Lab-Driven Alignment Signatures: A Psychometric Framework for Auditing Latent Bias and Compounding Risk in Generative AI

DGX agent

arXiv:2602.17127v2 Announce Type: replace Abstract: Large language models increasingly serve as reasoning layers in multi-agent systems, where one provider's models may generate, judge, and summarize

safetyarxiv-cs-cl
19 Aug 2026
Model Releases

AA is the reason for Qwen3.8 27B shipped with xhigh

DGX agent

I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment does

model-releasesr-localllama
18 Aug 2026
Model Releases

ChainSpace: A Chained-Reasoning Paradigm for Spatial Intelligence

DGX agent

arXiv:2608.15788v1 Announce Type: new Abstract: Spatial intelligence requires foundation models to maintain coherent spatial state across interactions with the physical world. However, existing data-c

model-releasesarxiv-cs-cv
18 Aug 2026
Research

Do Uncertainty Signals Help? A Systematic Study of Uncertainty-Aware Decoding with Rollback Mechanisms

DGX agent

arXiv:2608.14653v1 Announce Type: cross Abstract: Prediction uncertainty is a widely adopted metric for quantifying model confidence, with downstream applications spanning model explanation, data sele

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Does a Tool Result Carry More Authority Than Plain Text? Three Prospective Studies of False-Claim Adoption in a Synthetic Assignment Task with Claude Opus 5

DGX agent

arXiv:2608.14992v1 Announce Type: new Abstract: Language-model systems increasingly read from stores they also write to, so a claim that was merely written earlier can return looking retrieved. We tes

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Does the LM Head Create a Harmful Gradient Bottleneck? A Causal Test

DGX agent

arXiv:2608.16671v1 Announce Type: new Abstract: The language-model head maps a hidden state of width D to a vocabulary of size V, so its transpose can return at most D independent directions to the Tr

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Don't ignore llama.cpp RPC with old hardware. Results of a 5070 Ti and 1080 Ti over gigabit ethernet: it's actually functional.

DGX agent

Results up front: I had to prioritize prefill or token generation - there was no happy medium. Using UD-Q4_K_XL, q8 kv cache, and 96k max context: focus on generation (MTP = 2): 350 pp and 36 tg @ 12k

model-releasesr-localllama
18 Aug 2026
Model Releases

Enhancing the Non-Functional Quality Compliance of LLM-Generated Code through Quality-Aware Preference Learning

DGX agent

arXiv:2503.09020v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been widely adopted in commercial code completion engines, significantly enhancing coding efficiency and pro

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction

DGX agent

arXiv:2608.15727v1 Announce Type: cross Abstract: Diffusion models have recently shown strong potential for multivariate time-series anomaly detection by learning the distribution of normal data throu

model-releasesarxiv-cs-ai
18 Aug 2026
Research

Forward Pass Domain Adaptation (Without Cross-Layer Backpropagation)

DGX agent

arXiv:2608.14563v1 Announce Type: cross Abstract: Forward-Pass-Only MLP training (FPO) adapts large language models without a backward pass through the model body, achieving 2.7--3.2x the throughput o

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Gathered, Not Admitted: How Attention Brings a Latent Variable into Verbalizable Form

DGX agent

arXiv:2608.15022v1 Announce Type: new Abstract: Language models hold latent quantities in a form they can report on, and more of a quantity is present in that form when the task requires reusing it fl

model-releasesarxiv-cs-ai
18 Aug 2026
Tutorials

Human Pose Estimation in Trampoline Gymnastics: How to Improve Performance on Extreme Poses

DGX agent

arXiv:2604.01322v2 Announce Type: replace Abstract: Trampoline gymnastics involves extreme human poses and uncommon viewpoints, on which state-of-the art pose estimation models tend to under-perform.

tutorialsarxiv-cs-cv
18 Aug 2026
Model Releases

HyMem: Hierarchical Context Management for Long-Horizon Agents via Information Isolation

DGX agent

arXiv:2608.15703v1 Announce Type: new Abstract: Large language model (LLM) agents often perform poorly on complex, long-horizon tasks because their context becomes increasingly cluttered over time. As

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities

DGX agent

arXiv:2501.12147v2 Announce Type: replace-cross Abstract: Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong ca

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Ling-3.0 (BailingMoE3) lands in llama.cpp mainline - Quick benchmarks on Intel Arc B580

DGX agent

Finally llama.cpp now officially supports Ling-3.0! (Starting from build b10472+) If you want to run them locally, bartowski has already released the GGUF imatrix quantizations for both models: - Ling

model-releasesr-localllama
18 Aug 2026
Model Releases

Listen, Reason, and Segment: Aligning LALMs with Editorial Judgment for Media Chapterization

DGX agent

arXiv:2608.16539v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) have made rapid progress on standardized benchmarks, yet their deployment in practical media workflows, curation,

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

LLMs for Zero-Shot Threat Detection via Structured Risk Indicators

DGX agent

arXiv:2608.16508v1 Announce Type: cross Abstract: We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from het

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

Local Qwen 3.8 27B vs GPT‑5.6 Terra vs Grok 4.6

DGX agent

I gave three AI models the same brief: build a premium Three.js fragrance launch site from the same Git baseline, independently and with no collaboration. Three very different results. Here’s the full

model-releasesr-ollama
18 Aug 2026
Model Releases

MetaReason: Precise Interleaved Multimodal Reasoning via Editing Meta Information for Solving Geometry Problems

DGX agent

arXiv:2608.15006v1 Announce Type: cross Abstract: Although visual reasoning is crucial for solving complex geometry tasks, existing vision-language models rely heavily on text-only reasoning. Some rec

model-releasesarxiv-cs-ai
18 Aug 2026
Safety

MLLM-Guided Semantic Correction for Text-to-Video Generation

DGX agent

arXiv:2608.16513v1 Announce Type: cross Abstract: Recent advances in diffusion models and Transformer architectures have led to significant progress in text-to-video generation. However, these models

safetyarxiv-cs-ai
18 Aug 2026
Model Releases

Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off

DGX agent

arXiv:2608.15459v1 Announce Type: cross Abstract: Attention mechanisms have driven machine learning for a decade, from neural machine translation to language models that do general-purpose reasoning.

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

PL-Guard: Probabilistic Logic Reasoning for LLM Guardrails

DGX agent

arXiv:2608.15673v1 Announce Type: cross Abstract: Large language model guardrails can be viewed as policy-consistency problems: a system must determine which policy-relevant facts hold in a prompt-res

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers

DGX agent

arXiv:2608.15062v1 Announce Type: new Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

SchurQuant: Groupwise Discrete Optimization for Layer-Wise LLM Quantization

DGX agent

arXiv:2608.15567v1 Announce Type: new Abstract: Weight-only post-training quantization (PTQ) enables the deployment of large language models under tight memory budgets, but accuracy often collapses at

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression

DGX agent

arXiv:2604.04120v2 Announce Type: replace Abstract: Long chain-of-thought (Long-CoT) reasoning models have motivated a growing body of work on compressing reasoning traces to reduce inference cost, ye

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

SkillCommit: Evolving Agent Skills through Behaviorally Validated Scope Expansion

DGX agent

arXiv:2608.15165v1 Announce Type: new Abstract: Large language model (LLM) agents can continually improve without parameter updates by converting historical experience into reusable procedural knowled

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues

DGX agent

arXiv:2602.12381v2 Announce Type: replace Abstract: Recent generative models produce near-photorealistic images, challenging the trustworthiness of photographs. Synthetic image detection (SID) methods

model-releasesarxiv-cs-cv
18 Aug 2026
Safety

TwinGridShield: Consequence-Aware Runtime Authorization for LLM Grid-Agent Actions

DGX agent

arXiv:2608.15391v1 Announce Type: new Abstract: Large language model (LLM)-assisted energy-management tools can translate natural-language context into structured grid commands, but syntactic validity

safetyarxiv-cs-ai
18 Aug 2026
Model Releases

When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry

DGX agent

arXiv:2608.14680v1 Announce Type: new Abstract: Reliability in LLM-based agentic systems is a property of the whole execution (its tool calls, model calls, guardrails, and inter-agent messages), not o

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Agentic Transaction: Towards ACID-Compliant Agent Systems

DGX agent

arXiv:2608.13900v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from conversational assistants into autonomous systems that execute long-horizon tasks through reasonin

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding

DGX agent

arXiv:2608.14385v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have been widely adopted in real-time interactive applications such as coding assistants, real-time audio-video intera

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Generating Benchmark Health Data Using a Tabular Diffusion Transformer

DGX agent

arXiv:2608.14496v1 Announce Type: cross Abstract: Cross-Tabular Data Generation (CTDG) seeks to learn a generative model from multiple heterogeneous tables and produce new synthetic tabular datasets.

model-releasesarxiv-cs-ai
17 Aug 2026
← Previous
1…377378379380381…1324
Next →