AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlog
88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
Safety

Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning

DGX agent

arXiv:2608.01743v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central paradigm for large language model (LLM) post-training, but optimization toward new objectives can deg

safetyarxiv-cs-cl
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

TrimMoE A communication aware and adaptive depth framework for distributed edge inference

DGX agent

arXiv:2608.00573v1 Announce Type: cross Abstract: Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission. The ex

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Two-Stage Bengali Sentiment Classification: Domain Adaptation Through Continual Learning and Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2608.01471v1 Announce Type: new Abstract: Understanding sentiment in low-resource languages remains a key challenge for Natural Language Processing (NLP), particularly when domain-specific data

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

v0.32.6

DGX agent

What's Changed Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically /v1/chat/completions streaming now matches OpenAI's wire format: rol

model-releasesollama-releases
4 Aug 2026
Research

A Human-Centered Validation of the Explainability-Performance Coefficient

DGX agent

arXiv:2607.29614v1 Announce Type: cross Abstract: The rapid adoption of deep learning models in high-risk domains has intensified the need for trustworthy Explainable Artificial Intelligence (XAI). Ho

researcharxiv-cs-ai
3 Aug 2026
Model Releases

ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation

DGX agent

arXiv:2607.29539v1 Announce Type: cross Abstract: Standard AI-text detection benchmarks compare human-written text against text generated directly by large language models (LLMs). While prior work has

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Ask anything, anonymously. Qwen3.8-Max has landed on Venice. Give it a try!

DGX agent

Qwen from Alibaba has released the Qwen 3.8‑Max model on the Venice platform, enabling users to ask questions anonymously. The announcement encourages users to try the new functionality immediately. T

model-releasesqwen--x
3 Aug 2026
Research

BLADE: Boundary-Expanded and Layer-Adaptive Dynamic Exit for Efficient LLM Reasoning

DGX agent

arXiv:2607.28966v1 Announce Type: new Abstract: Large language models often improve task performance by generating long reasoning traces, but the resulting computation is frequently wasted on redundan

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Communication-Efficient Secure Aggregation in Decentralized Learning

DGX agent

arXiv:2405.07708v3 Announce Type: replace Abstract: Decentralized learning (DL) enables participants to collaboratively train models without a central server, yet it faces significant scalability chal

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Cortex Framework v7 is GA: Build agentic workflows without disrupting SAP operations

DGX agent

Businesses want to quickly and safely deploy AI agents to drive revenue, mitigate risk, and optimize capital, all without disrupting mission-critical ERP systems. And to power AI agents, you need more

model-releasesgoogle-cloud-ai
3 Aug 2026
Model Releases

DASH-OPD: Discrepancy-Aware Switching with Hysteresis for On-Policy Distillation

DGX agent

arXiv:2607.29078v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models on their own rollouts to reduce exposure bias. However, in multi-turn agent scenarios, early student

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

DeepSeek V4-Flash (284B MoE) at 33 tok/s single / 68 tok/s aggregate on 2× RTX 3090 + a used quad-Xeon DDR4 server — full config

DGX agent

Ran DeepSeek V4-Flash-0731 — the full official checkpoint, not a re-quant — on commodity used hardware. Sharing because I couldn't find anyone else publishing Ampere results for this engine. Why bothe

model-releasesr-localllama
3 Aug 2026
Model Releases

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

DGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Inference-time Trajectory Optimization for Structure-Preserving Manga Image Editing

DGX agent

arXiv:2603.27790v2 Announce Type: replace Abstract: We present a lightweight, training-free trajectory correction method that adapts a pretrained image editing model to each input manga image using on

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Matterhorn: Masked Time-to-First-Spike Encoding by Reassigning the Silent State for Sparse and Energy-Efficient Spiking Transformers

DGX agent

arXiv:2601.22876v2 Announce Type: replace Abstract: Spiking neural networks (SNNs) promise energy-efficient inference for large language models (LLMs), yet most reported savings rely on compute-operat

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

MoRoute: Dynamic Routing for In-Context Multimodal Video Generation

DGX agent

arXiv:2607.29545v1 Announce Type: new Abstract: Multimodal video generation aims to generate and edit videos conditioned on arbitrary combinations of text, images, and videos within a single model, al

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

NousResearch keeps doing things on hermes

DGX agent

Has anyone followed nousresearch work on Hermes? I mean we are Q3 2026. We have some crazy models trickling down from HGX territory to multi gpu workstation. And we have nousresearch deploying the 0.2

model-releasesr-localllama
3 Aug 2026
Model Releases

PluRel-to-RDB-PFN: Schema-Guided Synthetic Relational Pretraining

DGX agent

arXiv:2607.29129v1 Announce Type: new Abstract: Relational Foundation Models (RFMs) require large-scale synthetic relational databases for pretraining, but existing approaches tightly couple data gene

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

[RELEASE] SupraBrain-50M-v0.1

DGX agent

Hey there! So today we're releasing SupraBrain-50M, a hybrid language model that combines Gated DeltaNet linear recurrence with Sliding-Window Attention and Surprise-Gated update mechanisms to deliver

model-releasesr-localllama
3 Aug 2026
Model Releases

Retrieval-Driven Training-Free AI-Generated Video Attribution

DGX agent

arXiv:2607.28955v1 Announce Type: cross Abstract: AI-generated videos are becoming increasingly realistic and difficult to distinguish from authentic ones, which facilitates malicious misuse and poses

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing

DGX agent

arXiv:2607.28814v1 Announce Type: cross Abstract: In Motivational Interviewing (MI), a client's sustain talk (arguments for the status quo) calls for the counselor to roll with resistance, a move that

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

The Parts Are Greater Than the Sum: Automated Task Sequencing for Efficient Training of Multi-Policy LLMs

DGX agent

arXiv:2607.29601v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) commonly adapts large language models using a single shared Low-Rank Adapter (LoRA). This shared optimization spa

model-releasesarxiv-cs-lg
3 Aug 2026
Research

Tokenizer Transplantation: Mitigating Autoregressive Collapse in Edge-Efficient Bengali ASR

DGX agent

arXiv:2607.09598v2 Announce Type: replace Abstract: Lightweight speech recognition models are critical for edge deployment, yet highly optimized architectures like Moonshine often fail on morphologica

researcharxiv-cs-cl
3 Aug 2026
Safety

Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents

DGX agent

arXiv:2607.29254v1 Announce Type: new Abstract: AI agents extend large language models (LLMs) with external tools, enabling them to perform complex tasks and translate model outputs into consequential

safetyarxiv-cs-ai
3 Aug 2026
Research

Towards the Holographic Characteristic of LLMs for Efficient Short-text Generation

DGX agent

arXiv:2601.22546v2 Announce Type: replace-cross Abstract: The recent advancements in Large Language Models (LLMs) have attracted interest in exploring their in-context learning abilities and chain-of-

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

DGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Comfyui VRAM tracker

DGX agent

Hello! VRAM tracker is a node that track the full memory lifecycle of a comfyui run: when each weight is reserved, paged into VRAM, computed on, evicted, and freed. It renders it as an interactive HTM

model-releasesr-stablediffusion
2 Aug 2026
Model Releases

DSpark Benchmark Result on Deepseek v4 Flash 0731

DGX agent

TensorSharp supports DSpark on Deepseek v4 Flash 0731 now. Here is the benchmark result on 4x Nvidia A40 GPUs, cuda 12.8 with/without DSpark: Model: DeepSeek-V4-Flash-0731-UD-Q8_K_XL from https://hugg

model-releasesr-localllama
2 Aug 2026
Hardware

Hugging Face CEO Clément Delague says “AI is actually an opportunity to fix a lot of the cybersecurity problems” because his company used Nv…

DGX agent

Hugging Face CEO Clément Delague says “AI is actually an opportunity to fix a lot of the cybersecurity problems” because his company used Nvidia’s version of a Chinese open model to defend itself agai

hardwareclem-delangue--x
2 Aug 2026
Model Releases

Deepseek v4 flash 0731 still not holding up.

DGX agent

The biggest issue with preview was its inability to follow rules prompts and skills. It seems like no matter what you do it ignores them. I've tried first person and second person. I've tried Chinese

model-releasesr-localllama
1 Aug 2026
Tools

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep u…

DGX agent

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie

toolsfireworks-ai--x
1 Aug 2026
Model Releases

ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine

DGX agent

arXiv:2607.28625v1 Announce Type: new Abstract: Embodied intelligence faces a fundamental data bottleneck. Models must capture how first-person perception, whole-body motion, dexterous manipulation, o

model-releasesarxiv-cs-cv
31 Jul 2026
Local Ai

AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining

DGX agent

arXiv:2607.26642v1 Announce Type: new Abstract: Automated alpha mining has increasingly adopted large language model (LLM) agents for factor generation and iterative discovery. However, existing LLM-b

local-aiarxiv-cs-ai
31 Jul 2026
Model Releases

Beyond Geometric Complementarity: Coherent Overlap in Sparse Mixture-of-Experts Routing

DGX agent

arXiv:2607.28308v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) language models route each token to multiple experts, suggesting a geometric account of their benefit: co-selected exper

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Bridging AI and Energy Forecasting: An Autonomous Workflow with Customized Toolkit

DGX agent

arXiv:2307.07191v3 Announce Type: replace Abstract: Energy forecasting is crucial for the power grid, but fundamentally different from general time series analysis: it highly relies on covariates like

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction

DGX agent

arXiv:2607.26385v1 Announce Type: cross Abstract: Empirical work on algorithmic collusion asks one question of the data: are prices supracompetitive? We show this can be answered 'no' by a conspiracy

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Efficient LLMs with AMP: Attention Heads and MLP Pruning

DGX agent

arXiv:2504.21174v2 Announce Type: replace Abstract: Deep learning drives a new wave in computing systems and triggers the automation of increasingly complex problems. In particular, Large Language Mod

model-releasesarxiv-cs-lg
31 Jul 2026
Agents

FaithEyes: Towards Faithful Tool Use via Multi-Agent Process-Image Verification

DGX agent

arXiv:2607.28225v1 Announce Type: new Abstract: Agentic vision-language models (VLMs), which interleave textual reasoning with explicit tool calls such as cropping and code-based image manipulation, h

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

Flat Score, Amplified Failures: How the Error Budget Masks Damage in Quantized LLM Agents

DGX agent

arXiv:2607.27275v1 Announce Type: new Abstract: Post-training quantization to 4-bit weights is widely reported to be nearly lossless. We test this claim for multi-turn, tool-calling agents, where it n

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Generalization and Trade-off in Adversarial Training: An RKHS Perspective via Kernel Integral Operators

DGX agent

arXiv:2607.27995v1 Announce Type: cross Abstract: Adversarial training has emerged as a powerful approach for protecting models against adversarial attacks in a broad range of real-world applications.

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Gradient-free Task-Conditioned Retrieval for On-Device In-Context Learning

DGX agent

arXiv:2607.27766v1 Announce Type: new Abstract: On-device in-context learning (ICL) relies on pre-inference retrieval to select demonstrations for useful context before downstream model inference. Thi

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Hallucinations and Truth: A Comprehensive Accuracy Evaluation of RAG, LoRA and DoRA

DGX agent

arXiv:2502.10497v2 Announce Type: replace Abstract: Recent advancements in Generative AI have significantly improved the efficiency and adaptability of natural language processing (NLP) systems, parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

IFCMemoryBench: Evaluating Long-Term Memory of LLM-Based Agents in BIM Information Retrieval

DGX agent

arXiv:2607.26072v1 Announce Type: cross Abstract: Long-term memory is becoming a core capability of LLM-based agents, but existing evaluations largely test conversational recall in open-domain or pers

model-releasesarxiv-cs-ai
31 Jul 2026
Research

Metaphor Tracer: A Theory-Informed Analysis of Hidden States

DGX agent

arXiv:2607.28434v1 Announce Type: cross Abstract: What do a language model's hidden states say about the organization of a single text? From one forward pass, without training, we score every token po

researcharxiv-cs-cl
31 Jul 2026
Research

Meteosat Third Generation imagery improves CNN-based SSI retrieval

DGX agent

arXiv:2607.28093v1 Announce Type: cross Abstract: Accurate Surface Solar Irradiance (SSI) estimation is increasingly important for photovoltaic energy monitoring and forecasting. The recently introduc

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?

DGX agent

Laguna 2.1 at NVFP4 Deepseek v4 at Q2 Inkling-Small at IQ3 Which models you guys running now ? How it compares to 122b? Upcoming in few days : Ling 3.0 124B (Could be new king) LongCat 69B A3B ( very

model-releasesr-localllama
31 Jul 2026
Safety

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

DGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

safetyarxiv-cs-cv
31 Jul 2026
← Previous
1…423424425426427…1327
Next →