AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
Agents

From Prompt to Harness: Coderlet from Scratch

DGX agent

arXiv:2608.09480v1 Announce Type: new Abstract: A model alone does not determine how a programming agent acts. What the model sees, how actions enter the environment, how feedback returns, and how one

agentsarxiv-cs-ai
11 Aug 2026
Tutorials

Generative Models: Principles, Architectures, and Applications

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2608.08101v1 Announce Type: new Abstract: Generative AI has emerged as one of the most transformative forces in modern artificial intelligence, reshaping how we create, imagine, and interact wit

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

Gradient Under Microscope: Benchmarking Resource Utilization of Memory-Efficient Gradient Computation Methods

DGX agent

arXiv:2608.08961v1 Announce Type: new Abstract: AI training's rising resource intensity is straining electricity supplies and carbon budgets, motivating systematic study of memory-efficient training o

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Impact of Dataset Composition on Embedded Real-Time UAV Wildfire Detection Using Compact YOLO Models

DGX agent

arXiv:2608.07554v1 Announce Type: new Abstract: The development of vision-based wildfire detection systems for unmanned aerial vehicles is constrained by the limited availability of diverse real-world

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

LogicIF: Towards Complex Logic Instruction Following

DGX agent

arXiv:2508.09125v4 Announce Type: replace Abstract: Instruction following has catalyzed the recent era of Large Language Models (LLMs) and is the foundational skill underpinning more advanced capabili

model-releasesarxiv-cs-cl
11 Aug 2026
Applications

Mask-aware inference with State-Space Models

DGX agent

arXiv:2603.04568v2 Announce Type: replace Abstract: Many real-world computer vision tasks, such as depth completion, must handle inputs with arbitrarily shaped regions of missing or invalid data. For

applicationsarxiv-cs-cv
11 Aug 2026
Safety

MultiShadow: Multi-Object Shadow Generation for Image Compositing via Diffusion Model

DGX agent

arXiv:2603.02743v4 Announce Type: replace Abstract: Realistic shadow generation is crucial for achieving seamless image compositing, yet existing methods primarily focus on single-object insertion and

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI

DGX agent

As AI shifts from chatbots to autonomous agents, open models are serving market demands for full control over where AI runs and how it’s deployed and evolves. Today, NVIDIA is expanding its Nemotron 3

model-releasesnvidia-blog
11 Aug 2026
Model Releases

Nvidia releases Nemotron 3.5 Lightning and NeMo Switchyard to give enterprise AI capability options

DGX agent

Artificial intelligence silicon and software giant Nvidia Corp. today announced two new services: a highly customizable Nemotron model and an agentic AI model router named NeMo Switchyard. As enterpri

model-releasessiliconangle
11 Aug 2026
Model Releases

Out-of-Distribution Federated Distillation with Domain-Aware Proxy

DGX agent

arXiv:2608.08525v1 Announce Type: new Abstract: Federated Learning is a distributed machine learning paradigm that trains a global model by aggregating local clients without sharing private data of ea

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update

DGX agent

arXiv:2607.11505v2 Announce Type: replace-cross Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behav

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Real Data Closes Synthetic-to-Real Gap in Optical Chemical Structure Recognition

DGX agent

arXiv:2608.09100v1 Announce Type: cross Abstract: Millions of chemical structures appear in patents and papers only as drawings, and using that information at scale requires reading the drawings. OCSR

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Probability Aggregation and Logic-Representation Editing

DGX agent

arXiv:2608.08514v1 Announce Type: new Abstract: We independently reproduce two recent methods for making large language model (LLM) reasoning more reliable, and stress-test them across domains and mod

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Scale-to-Dialogue: Low-Burden Elicitation of Daily Premenstrual Symptom Ratings with Small Language Models

DGX agent

arXiv:2608.08746v1 Announce Type: new Abstract: Prospective daily symptom tracking is central to premenstrual health assessment, but repeated ordinal forms impose substantial response burden. We formu

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

SHE: Trajectory-driven Safety Harness Evolution for LLM Agents

DGX agent

arXiv:2608.09885v1 Announce Type: new Abstract: The safety of large language model (LLM) agents depends not only on model weights but also on the agent harness that manages context, memory, tools, per

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents

DGX agent

arXiv:2601.18077v3 Announce Type: replace Abstract: Cooperative reasoning under incomplete information remains challenging for both humans and multi-agent systems. The card game Hanabi embodies this c

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Agents

STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in LLMs

DGX agent

arXiv:2608.08164v1 Announce Type: cross Abstract: Knowledge Distillation is a widely adopted technique in the training and fine-tuning of large language models (LLMs) enabling transfer of structured i

agentsarxiv-cs-ai
11 Aug 2026
Research

SUMI: Scalable Unified Model for 3D Point Cloud Inference

DGX agent

arXiv:2608.08115v1 Announce Type: new Abstract: Point cloud completion commonly follows a coarse-to-fine paradigm, where a low-density coarse shape is first predicted and then upsampled to the target

researcharxiv-cs-cv
11 Aug 2026
Safety

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

DGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

Tested in Coding: BF16 Muse Glimmer vs BF16 Qwen3.6 27B

DGX agent

I'm guessing that many people have been waiting for this comparison. For clarity, both models are running at full FP16 KV-cache. Due to VRAM limitations, Muse Glimmer is running full 262,144 context,

model-releasesr-localllama
11 Aug 2026
Model Releases

Think Deep, Speak Once: Relit, A Recursive Latent Implicit Transformer Framework

DGX agent

arXiv:2608.08113v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has become the dominant paradigm for eliciting reasoning in Large Language Models (LLMs), yet it creates substantial co

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

To Memorize or to Retrieve: Scaling the Interaction Between Pretraining and Retrieval

DGX agent

arXiv:2604.00715v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) improves language model (LM) performance by providing relevant context at test time for knowledge-intensi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Wiener Representation Filtering for VLM Hallucination Suppression

DGX agent

arXiv:2608.08167v1 Announce Type: new Abstract: Vision-language models (VLMs) excel at open-ended captioning and visual QA but often describe objects, attributes, or relations absent from the image, a

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers

DGX agent

arXiv:2608.06694v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics extends polymer simulation beyond the scales accessible to all-atom (AA) methods, but bottom-up CG modeling is la

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Does More Retrieved Evidence Help Visual Retrieval-Augmented Generation with Diffusion Language Models?

DGX agent

arXiv:2608.07006v1 Announce Type: new Abstract: Visual retrieval-augmented generation (RAG) commonly expands the retrieved evidence set to improve answer-page coverage, implicitly assuming that all av

researcharxiv-cs-cl
10 Aug 2026
Research

Free Denoising Diffusion Models

DGX agent

arXiv:2510.22778v3 Announce Type: replace-cross Abstract: We develop a free-probabilistic framework for denoising diffusion, in which the data is a self-adjoint operator and its law a spectral distrib

researcharxiv-cs-lg
10 Aug 2026
Model Releases

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

DGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Latent Fact-Checking: Detecting Misinformation through Activation Engineering

DGX agent

arXiv:2608.06417v1 Announce Type: cross Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level ling

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Need real world ML problems to evaluate my educational ML tools

DGX agent

I'm a retired platform engineer, coding mainly in Rust, and involved with a ML study group. I developed a ML programming language (alternative to Python, Colab) to help me learn (and teach) ML concept

model-releasesr-localllama
10 Aug 2026
Local Ai

omlab/VLX-Seek-1.5-10B · Hugging Face

DGX agent

VLX-Seek-1.5-10B VLX-Seek-1.5-10B is the open-source 10B model in the VLX-Seek 1.5 family, designed for fine-grained perception and visual grounding in embodied scenarios. It targets practical setting

local-air-localllama
10 Aug 2026
Model Releases

Test-Time Adaptation with Online Personalized Energy-Based Cache for Fine-Grained Video Expression Recognition

DGX agent

arXiv:2608.06467v1 Announce Type: new Abstract: Facial expression recognition (FER) in videos is challenging because models must identify subtle, temporally evolving affective states that vary across

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

UniREditBench: A Unified Reasoning-based Image Editing Benchmark

DGX agent

arXiv:2511.01295v3 Announce Type: replace Abstract: Recent advances in multi-modal generative models have driven substantial improvements in image editing. However, current generative models still str

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

v0.32.7

DGX agent

Muse Glimmer Note: Muse Glimmer is currently available via initial support via Ollama's MLX engine on Apple Silicon. Additional support and optimizations for Apple Silicon, NVIDIA, AMD, and other plat

model-releasesollama-releases
10 Aug 2026
Research

A Foundational EDM2-Based Generative Model for High-Resolution Synthetic Fetal Ultrasound Imaging from Open Datasets

DGX agent

arXiv:2608.05471v1 Announce Type: cross Abstract: Prenatal ultrasound imaging is key for assessing fetal health, but AI progress is limited by scarce, privacy-restricted, and hard-to-annotate datasets

researcharxiv-cs-cv
7 Aug 2026
Model Releases

A Two-Tier Perspective on Inference-Time Parallelism in Multi-Agent LLM Systems

DGX agent

arXiv:2608.05791v1 Announce Type: cross Abstract: Large language model (LLM)-driven multi-agent systems typically require multiple model invocations and complex coordination during inference, and thei

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

Agentic Software Issue Resolution with Large Language Models: A Survey

DGX agent

arXiv:2512.22256v2 Announce Type: replace-cross Abstract: Software issue resolution aims to address real-world issues in software repositories based on natural language descriptions provided by users,

agentsarxiv-cs-ai
7 Aug 2026
Safety

Deterministic World Models for Closed-loop Reachability Analysis of End-to-End Vision-based Control

DGX agent

arXiv:2512.08991v3 Announce Type: replace Abstract: End-to-end image controllers that map raw camera frames directly to control actions are increasingly deployed in safety-critical systems. However, f

safetyarxiv-cs-cv
7 Aug 2026
Model Releases

Learning When to Trust via Selective Context Preference Optimization

DGX agent

arXiv:2608.06377v1 Announce Type: cross Abstract: Language models increasingly condition their answers on external signals, and a single misleading one can turn a correct answer wrong. The obvious rem

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Otter: A Time-Aware, History-Conditioned Human Chess AI

DGX agent

arXiv:2608.05206v1 Announce Type: new Abstract: Otter is a 15.3M-parameter human chess AI that predicts human move selection by modeling play as a time-aware, sequential process rather than treating e

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Reducing belief in conspiracy theories as they unfold using large language models

DGX agent

arXiv:2608.06151v1 Announce Type: cross Abstract: The emergence of conspiracy theories in the wake of major events is a significant societal challenge. Here we test whether conversational dialogues wi

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Runtime Observability for Heterogeneous Attention Memory

DGX agent

arXiv:2608.05863v1 Announce Type: new Abstract: Modern models no longer keep a plain KV cache: latent caches, learned sparse selectors and recurrent states each carry the model's memory in a different

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations

DGX agent

arXiv:2507.03631v4 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experi

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation

DGX agent

arXiv:2608.05970v1 Announce Type: cross Abstract: Embodied visuomotor models, including Diffusion Policy (DP) and Vision-Language-Action (VLA) models, have demonstrated promising performance on roboti

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

TS-RAG: Retrieval Augmented Generation for Time Series Forecasting

DGX agent

arXiv:2608.06223v1 Announce Type: new Abstract: While deep learning models, particularly transformer-based architectures, have shown impressive performance in time series forecasting, the application

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Beyond Boundary Frames: Talking-Head Inbetweening via Context-Aware Motion Modeling

DGX agent

arXiv:2512.03590v3 Announce Type: replace Abstract: Existing talking-head generation methods primarily target open-ended generation rather than bridging two existing video segments. In this paper, we

researcharxiv-cs-cv
6 Aug 2026
Model Releases

Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks

DGX agent

arXiv:2608.04286v1 Announce Type: new Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations

model-releasesarxiv-cs-cl
6 Aug 2026
Research

Image Classification Using CNN-QNN Hybrid Model with Optimized Correlated Features

DGX agent

arXiv:2608.04379v1 Announce Type: cross Abstract: We propose a method to optimize the correlation among convolutional neural network (CNN) features that are used as inputs to quantum neural network (Q

researcharxiv-cs-ai
6 Aug 2026
← Previous
1…313314315316317…1302
Next →