AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,103 results
10 Apr 2026

Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning

ResearchDGX agent

arXiv:2604.06501v1 Announce Type: new Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

Model ReleasesDGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

VisCoder2: Building Multi-Language Visualization Coding Agents

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

9 Apr 2026

Hermes will soon be a serious creative tool. More on this soon!

ResearchDGX agent

Nous Research's Hermes model line has been deliberately developed with a strong emphasis on creative capability. The training dataset for Hermes 4 expanded to 50x more data tokens than Hermes 3, w...

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vuln…

ResearchDGX agent

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vulnerabilities in this code? Consider the behavior of the SEQ_L

if you prefer blog format: https://blog.langchain.com/deep-agents-deploy-an-open-alternative-to-claude-managed-agents/

Model ReleasesDGX agent

LangChain launched **Deep Agents Deploy** in beta as an open-source, model-agnostic alternative to Anthropic's Claude Managed Agents, allowing developers to deploy production-ready agents via a sin...

Parax: Parametric Modeling in JAX + Equinox [P]

ResearchDGX agent

**Paramax** (also referred to as 'Parax' in the Reddit post title) is a small Python library by Daniel Ward that provides parameterizations and parameter constraints for JAX PyTrees, designed to wo...

v0.20.5-rc1

Local AiDGX agent

<channel|>This release update for `local-ai` (version v0.20.5-rc1) expands model compatibility by integrating several popular large language models. It allows users to run models such as Kimi-K2.5,...

8 Apr 2026

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the rele…

ResearchDGX agent

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the relevant code, and ran them through small, cheap, open-weights m

Exponentials everywhere.

ApplicationsDGX agent

The specific tweet (status ID 2041723225827062080) could not be directly retrieved, but based on closely related content from Ethan Mollick's (@emollick) X account, the most relevant match is the p...

If you're an AI/agent builder, it's so important that you don't overbuild and overcommit on a specific toolset and infrastructure. Frontier …

Model ReleasesDGX agent

If you're an AI/agent builder, it's so important that you don't overbuild and overcommit on a specific toolset and infrastructure. Frontier labs are shipping not just the models, but the harnesses and

7 Apr 2026

INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lo…

Model ReleasesDGX agent

INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lot benchmarks? > SWE-Bench Pro: 58.4 > beats Opus 4.6 (57.3)

SWE-1.6 is free for everyone in Windsurf for the next 3 months at 200 tok/s. For paying users, we've partnered with Cerebras to serve the mo…

AgentsDGX agent

SWE-1.6 is free for everyone in Windsurf for the next 3 months at 200 tok/s. For paying users, we've partnered with Cerebras to serve the model at 950 tok/s. More technical details about this training

18 Aug 2026

AA is the reason for Qwen3.8 27B shipped with xhigh

Model ReleasesDGX agent

I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment does

ChainSpace: A Chained-Reasoning Paradigm for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2608.15788v1 Announce Type: new Abstract: Spatial intelligence requires foundation models to maintain coherent spatial state across interactions with the physical world. However, existing data-c

Do Uncertainty Signals Help? A Systematic Study of Uncertainty-Aware Decoding with Rollback Mechanisms

ResearchDGX agent

arXiv:2608.14653v1 Announce Type: cross Abstract: Prediction uncertainty is a widely adopted metric for quantifying model confidence, with downstream applications spanning model explanation, data sele

Does a Tool Result Carry More Authority Than Plain Text? Three Prospective Studies of False-Claim Adoption in a Synthetic Assignment Task with Claude Opus 5

Model ReleasesDGX agent

arXiv:2608.14992v1 Announce Type: new Abstract: Language-model systems increasingly read from stores they also write to, so a claim that was merely written earlier can return looking retrieved. We tes

Does the LM Head Create a Harmful Gradient Bottleneck? A Causal Test

Model ReleasesDGX agent

arXiv:2608.16671v1 Announce Type: new Abstract: The language-model head maps a hidden state of width D to a vocabulary of size V, so its transpose can return at most D independent directions to the Tr

Don't ignore llama.cpp RPC with old hardware. Results of a 5070 Ti and 1080 Ti over gigabit ethernet: it's actually functional.

Model ReleasesDGX agent

Results up front: I had to prioritize prefill or token generation - there was no happy medium. Using UD-Q4_K_XL, q8 kv cache, and 96k max context: focus on generation (MTP = 2): 350 pp and 36 tg @ 12k

Enhancing the Non-Functional Quality Compliance of LLM-Generated Code through Quality-Aware Preference Learning

Model ReleasesDGX agent

arXiv:2503.09020v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been widely adopted in commercial code completion engines, significantly enhancing coding efficiency and pro

FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction

Model ReleasesDGX agent

arXiv:2608.15727v1 Announce Type: cross Abstract: Diffusion models have recently shown strong potential for multivariate time-series anomaly detection by learning the distribution of normal data throu

Forward Pass Domain Adaptation (Without Cross-Layer Backpropagation)

ResearchDGX agent

arXiv:2608.14563v1 Announce Type: cross Abstract: Forward-Pass-Only MLP training (FPO) adapts large language models without a backward pass through the model body, achieving 2.7--3.2x the throughput o

Gathered, Not Admitted: How Attention Brings a Latent Variable into Verbalizable Form

Model ReleasesDGX agent

arXiv:2608.15022v1 Announce Type: new Abstract: Language models hold latent quantities in a form they can report on, and more of a quantity is present in that form when the task requires reusing it fl

Human Pose Estimation in Trampoline Gymnastics: How to Improve Performance on Extreme Poses

TutorialsDGX agent

arXiv:2604.01322v2 Announce Type: replace Abstract: Trampoline gymnastics involves extreme human poses and uncommon viewpoints, on which state-of-the art pose estimation models tend to under-perform.

HyMem: Hierarchical Context Management for Long-Horizon Agents via Information Isolation

Model ReleasesDGX agent

arXiv:2608.15703v1 Announce Type: new Abstract: Large language model (LLM) agents often perform poorly on complex, long-horizon tasks because their context becomes increasingly cluttered over time. As

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities

Model ReleasesDGX agent

arXiv:2501.12147v2 Announce Type: replace-cross Abstract: Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong ca

Ling-3.0 (BailingMoE3) lands in llama.cpp mainline - Quick benchmarks on Intel Arc B580

Model ReleasesDGX agent

Finally llama.cpp now officially supports Ling-3.0! (Starting from build b10472+) If you want to run them locally, bartowski has already released the GGUF imatrix quantizations for both models: - Ling

Listen, Reason, and Segment: Aligning LALMs with Editorial Judgment for Media Chapterization

Model ReleasesDGX agent

arXiv:2608.16539v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) have made rapid progress on standardized benchmarks, yet their deployment in practical media workflows, curation,

LLMs for Zero-Shot Threat Detection via Structured Risk Indicators

Model ReleasesDGX agent

arXiv:2608.16508v1 Announce Type: cross Abstract: We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from het

Local Qwen 3.8 27B vs GPT‑5.6 Terra vs Grok 4.6

Model ReleasesDGX agent

I gave three AI models the same brief: build a premium Three.js fragrance launch site from the same Git baseline, independently and with no collaboration. Three very different results. Here’s the full

MetaReason: Precise Interleaved Multimodal Reasoning via Editing Meta Information for Solving Geometry Problems

Model ReleasesDGX agent

arXiv:2608.15006v1 Announce Type: cross Abstract: Although visual reasoning is crucial for solving complex geometry tasks, existing vision-language models rely heavily on text-only reasoning. Some rec

MLLM-Guided Semantic Correction for Text-to-Video Generation

SafetyDGX agent

arXiv:2608.16513v1 Announce Type: cross Abstract: Recent advances in diffusion models and Transformer architectures have led to significant progress in text-to-video generation. However, these models

Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off

Model ReleasesDGX agent

arXiv:2608.15459v1 Announce Type: cross Abstract: Attention mechanisms have driven machine learning for a decade, from neural machine translation to language models that do general-purpose reasoning.

PL-Guard: Probabilistic Logic Reasoning for LLM Guardrails

Model ReleasesDGX agent

arXiv:2608.15673v1 Announce Type: cross Abstract: Large language model guardrails can be viewed as policy-consistency problems: a system must determine which policy-relevant facts hold in a prompt-res

RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers

Model ReleasesDGX agent

arXiv:2608.15062v1 Announce Type: new Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve

SchurQuant: Groupwise Discrete Optimization for Layer-Wise LLM Quantization

Model ReleasesDGX agent

arXiv:2608.15567v1 Announce Type: new Abstract: Weight-only post-training quantization (PTQ) enables the deployment of large language models under tight memory budgets, but accuracy often collapses at

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression

Model ReleasesDGX agent

arXiv:2604.04120v2 Announce Type: replace Abstract: Long chain-of-thought (Long-CoT) reasoning models have motivated a growing body of work on compressing reasoning traces to reduce inference cost, ye

SkillCommit: Evolving Agent Skills through Behaviorally Validated Scope Expansion

Model ReleasesDGX agent

arXiv:2608.15165v1 Announce Type: new Abstract: Large language model (LLM) agents can continually improve without parameter updates by converting historical experience into reusable procedural knowled

Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues

Model ReleasesDGX agent

arXiv:2602.12381v2 Announce Type: replace Abstract: Recent generative models produce near-photorealistic images, challenging the trustworthiness of photographs. Synthetic image detection (SID) methods

TwinGridShield: Consequence-Aware Runtime Authorization for LLM Grid-Agent Actions

SafetyDGX agent

arXiv:2608.15391v1 Announce Type: new Abstract: Large language model (LLM)-assisted energy-management tools can translate natural-language context into structured grid commands, but syntactic validity

When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry

Model ReleasesDGX agent

arXiv:2608.14680v1 Announce Type: new Abstract: Reliability in LLM-based agentic systems is a property of the whole execution (its tool calls, model calls, guardrails, and inter-agent messages), not o

17 Aug 2026

Agentic Transaction: Towards ACID-Compliant Agent Systems

Model ReleasesDGX agent

arXiv:2608.13900v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from conversational assistants into autonomous systems that execute long-horizon tasks through reasonin

DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding

Model ReleasesDGX agent

arXiv:2608.14385v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have been widely adopted in real-time interactive applications such as coding assistants, real-time audio-video intera

Generating Benchmark Health Data Using a Tabular Diffusion Transformer

Model ReleasesDGX agent

arXiv:2608.14496v1 Announce Type: cross Abstract: Cross-Tabular Data Generation (CTDG) seeks to learn a generative model from multiple heterogeneous tables and produce new synthetic tabular datasets.

Inducing Reward-Free Judging Rubrics that Reduce Over-Crediting in Agent Evaluation

AgentsDGX agent

arXiv:2608.13564v1 Announce Type: new Abstract: Evaluating language-model agents at scale increasingly relies on a second language model as an automatic judge, because the gold signal, an executable e

Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans

ResearchDGX agent

arXiv:2608.14317v1 Announce Type: cross Abstract: This research developed a neural network-based model to extract various information from 2D floor plans. We detect lighting symbols, identify the appr

Mandato: Protocol-Level Enforcement of Digitally Signed Mandates on AI Agent Actions with Cryptographically Chained Audit Trails

Model ReleasesDGX agent

arXiv:2608.14074v1 Announce Type: new Abstract: AI agents increasingly act on external systems through standardized tool-calling protocols such as the Model Context Protocol (MCP), yet no infrastructu

Practical Lossless Volumetric Medical Image Compression via Tri-plane Context Tree Learning

Local AiDGX agent

arXiv:2608.13897v1 Announce Type: cross Abstract: Lossless compression of volumetric medical images is of paramount importance for clinical and research applications where data fidelity is essential.

Redefining Generalization in Visual Domains: A Two-Axis Framework for Fake Image Detection with FusionDetect

Model ReleasesDGX agent

arXiv:2510.05740v2 Announce Type: replace-cross Abstract: The rapid development of generative models has made it increasingly crucial to develop detectors that can reliably detect synthetic images. Al

Resource-Efficient RGB-Only Action Recognition for Edge Deployment

Model ReleasesDGX agent

arXiv:2602.10818v2 Announce Type: replace Abstract: Resource-constrained assistive monitoring requires compact local video perception and an explicit understanding of how recognition reliability chang

When Does More Correct Data Hurt? Insertion-Stability and the Limits of Dimension-Based Theory

ResearchDGX agent

arXiv:2608.14020v1 Announce Type: new Abstract: Adding data known to be correct ought to be safe. Not always. Larsen, Pabbaraju and Shetty model the failure with a monotone adversary, which reads an i

16 Aug 2026

3000000000 downloads! Can you count the zeros at a glance? 😎 Thank you all for the incredible love. Let's keep growing together! 🌱

Model ReleasesDGX agent

3000000000 downloads! Can you count the zeros at a glance? 😎 Thank you all for the incredible love. Let's keep growing together! 🌱 Alibaba's open-weight models have accumulated more than 3 billion glo

DeepSeek V4 Pro 0813 is live on Together AI DeepSeek’s flagship V4 Pro release brings a 1.6T MoE architecture, 1M context, and three reasoni…

Model ReleasesDGX agent

DeepSeek V4 Pro (Aug 2026) has been released on the Together AI platform. The model employs a 1.6‑terabyte Mixture‑of‑Experts architecture with a 1‑million‑token context window and supports three dist

QWEN 3.8 27B Q8 Quant - Setup Instructions Help

Model ReleasesDGX agent

Hi All, I have a MacBook Pro M5 Max with 48 GB of unified memory. Also have a M4 Mac Mini 16GB unified memory. I want to setup local AI and I heard about QWEN 3.8 model. As I want to work on a repo th

🚀Qwen3.8-27B flies on a laptop, becoming part of our work and daily lives. Thanks for the shoutout! @atomic_chat_hq

Model ReleasesDGX agent

🚀Qwen3.8-27B flies on a laptop, becoming part of our work and daily lives. Thanks for the shoutout! @atomic_chat_hq Run Qwen3.8 27B locally via Atomic Chat💥 We released Atomic Dynamic GGUF quants, fro

The benchmark for non-verifiable domains is often the opinions of humans. That is how we determine whether writing or an idea or a pitch is …

Model ReleasesDGX agent

The benchmark for non-verifiable domains is often the opinions of humans. That is how we determine whether writing or an idea or a pitch is good in the real world And we know how to measure & benchmar

15 Aug 2026

Gemma 4 E4B IQ2_XXS: + 140.54% Reasoning Performance From Tensor Level Quantization Allocation

Model ReleasesDGX agent

iq2_xxs tensor level allocation recovered reasoning from 28.9 -> 69.5 at the same 3.3gb budget. https://huggingface.co/ByteOtter/gemma-4-E4B-it-CADA-IQ2_XXS I posted my Gemma 4 12B q3 result a couple

14 Aug 2026

Beyond Visual Evidence: Revealing and Mitigating Relational Privacy Leakage in Document MLLMs

Model ReleasesDGX agent

arXiv:2608.12911v1 Announce Type: new Abstract: While the privacy risks of multimodal large language models (MLLMs) have drawn significant attention, the unique vulnerabilities of domain-specific MLLM

BrainWAM: Action-Space Coordination of Semantic Priors and Predictive Dynamics for Autonomous Driving

AgentsDGX agent

arXiv:2608.12854v1 Announce Type: cross Abstract: Autonomous driving requires planning under both semantic constraints and predictive dynamics. Existing end-to-end driving approaches, however, typical

CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence

Model ReleasesDGX agent

arXiv:2608.12555v1 Announce Type: new Abstract: Predictive explanation methods attribute a model output; they do not, by themselves, attribute an intervention effect on the real-world outcome. We intr

← Previous
1…293294295296297…1036
Next →