AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlog
88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
Model Releases

Semi-Supervised Text-Attributed Graph Distillation

DGX agent

arXiv:2607.20477v1 Announce Type: new Abstract: {em Text-Attributed Graphs} (TAGs) have emerged as an expressive data model for integrating graph topology with rich textual semantics. Existing represe

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Surprisal Theory is Tautological (without Rational Grounding)

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.21574v1 Announce Type: new Abstract: Surprisal theory holds that the human processing difficulty of a linguistic unit in context is an affine function of its surprisal under some language m

researcharxiv-cs-cl
24 Jul 2026
Research

The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning

DGX agent

arXiv:2607.20952v1 Announce Type: cross Abstract: Latent, or silent, reasoning lets language models carry out intermediate computation in continuous vector space instead of words, and is widely assume

researcharxiv-cs-cl
24 Jul 2026
Model Releases

Towards an Automated Test of LLM Security Knowledge

DGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Towards Faithful Graph Explanations with Synergistic Edge Effects via Granular Balls

DGX agent

arXiv:2607.21381v1 Announce Type: new Abstract: Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previous methods explain graph neural networks

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

DGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms

DGX agent

arXiv:2607.20638v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation and reasoning, yet their ability to perform temporal reasoning ove

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

When Trivia Is Not Trivial: Everyday Knowledge Failures in Multilingual LLMs

DGX agent

arXiv:2607.21445v1 Announce Type: new Abstract: Quiz rooms, trivia nights, and quiz shows challenge human knowledge across a wide range of topics, from canonical facts to everyday culture. In this pap

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

DGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency

DGX agent

arXiv:2601.15368v3 Announce Type: replace Abstract: Generative image inpainting can produce realistic results even with large, irregular masks, but existing methods still suffer from two common proble

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection

DGX agent

arXiv:2512.16300v3 Announce Type: replace Abstract: Existing image forgery detection (IFD) methods either exploit low-level, semantics-agnostic artifacts or rely on multimodal large language models (M

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Current Injection Spiking Neural Network for Infrared and Visible Image Fusion

DGX agent

arXiv:2607.19879v1 Announce Type: new Abstract: Infrared and visible image fusion (IVIF) integrates the complementary information of two modalities into a single image with richer scene content. While

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM

DGX agent

TLDR: I (with the help of AI) re-implemented every Blackwell-only kernel (DeepGEMM, FlashInfer sparse-MLA, block-scaled FP8) in Triton, because they simply don't exist for sm89. The performance is 2-3

model-releasesr-localllama
23 Jul 2026
Local Ai

Development of an automated, reliable, and clinically meaningful artificial intelligence (AI) tool for diagnosing cardiac disease from conventional cardiovascular magnetic resonance (CMR) images

DGX agent

arXiv:2607.20087v1 Announce Type: new Abstract: Aims: Cardiovascular magnetic resonance (CMR) imaging enables non-invasive assessment of myocardial structure, function, and pathology, but requires sub

local-aiarxiv-cs-cv
23 Jul 2026
Model Releases

Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation: A Comparative Study

DGX agent

arXiv:2502.16022v3 Announce Type: replace Abstract: OpenNotes gives patients access to their EHR notes, but dense medical jargon limits comprehension. We evaluate closed-source and open-source LLMs fo

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Fluid-SDF: Ultra-Lightweight and Editable Implicit Shape Representation via Differentiable Primitives

DGX agent

arXiv:2607.18646v1 Announce Type: new Abstract: Implicit Neural Representations (INRs) have become the standard for continuous 2D shape modeling, but they suffer from black-box uneditability, vulnerab

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

GATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape Retrieval

DGX agent

arXiv:2607.19111v1 Announce Type: new Abstract: Large pretrained vision models have substantially improved appearance-based 3D shape retrieval, but they still confuse shapes that look similar while di

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Geospatial Diffusion-based Evolution Synthesis (GeoDES) for Storm-Centered Weather Augmentation

DGX agent

arXiv:2607.19522v1 Announce Type: new Abstract: While machine learning-based weather models hold significant promise, they struggle to predict the detailed structure of large-scale weather systems suc

researcharxiv-cs-lg
23 Jul 2026
Model Releases

GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]

DGX agent

The interesting finding from a new [arXiv paper](https://arxiv.org/abs/2607.16165) isn't that a frontier vision model failed a new benchmark, that happens weekly, but the specific shape of the failure

model-releasesr-machinelearning
23 Jul 2026
Local Ai

I run GLM-4.5-Air (110B) on 16Gb ram consumer machine and Qwen3-30B at 20 tok/s

DGX agent

In the past few months I’ve experimenting heavily and tortured my old 2016 Desktop PC to run the biggest Local LLM I can fit. I documented the whole process and research and I’ve published a repositor

local-air-ollama
23 Jul 2026
Tutorials

Look Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMs

DGX agent

arXiv:2607.20357v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have recently demonstrated strong performance across vision-language tasks. However, their high inference cost,

tutorialsarxiv-cs-cv
23 Jul 2026
Safety

Membership Inference Attacks for Unseen Classes

DGX agent

arXiv:2506.06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is data auditing, i.e., using statistical tools to determine whether harmful content may have been used in t

safetyarxiv-cs-lg
23 Jul 2026
Safety

Norm or Direction? Decoding Vision Mambas for High-Resolution Vision

DGX agent

arXiv:2607.18625v1 Announce Type: new Abstract: Vision Mamba models replace quadratic self-attention with linear complexity selective state space models (SSMs), emerging as efficient visual backbones.

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Self Gradient Forcing: Native Long Video Extrapolation

DGX agent

arXiv:2607.20368v1 Announce Type: new Abstract: Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own ro

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

DGX agent

arXiv:2607.19361v1 Announce Type: cross Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolation, which misses failures that arise only over a

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Taming the Security-Energy Paradox: A Green AI Approach to Optimized Android Malware Detection

DGX agent

arXiv:2607.20003v1 Announce Type: cross Abstract: An increase in advanced Android malware requires the use of deep learning models, which can run on Android devices. But there is a trade-off between s

researcharxiv-cs-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning

DGX agent

arXiv:2607.19790v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models r

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patches

DGX agent

TL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA

model-releasesr-stablediffusion
23 Jul 2026
Research

Trusting What You Cannot See: Auditable Fine-Tuning and Inference for Proprietary AI

DGX agent

arXiv:2603.07466v2 Announce Type: replace-cross Abstract: Cloud-based infrastructure has become the dominant platform for deploying large models, particularly large language models (LLMs). Fine-tuning

researcharxiv-cs-lg
23 Jul 2026
Safety

Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction

DGX agent

arXiv:2607.19532v1 Announce Type: cross Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models,

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training

DGX agent

arXiv:2607.19971v1 Announce Type: new Abstract: Accurate motion prediction of surrounding agents and safe motion planning are two closely coupled key tasks for social robot navigation in crowded envir

model-releasesarxiv-cs-ro
23 Jul 2026
Tutorials

4yr throwback. feels like a long time ago!

DGX agent

4yr throwback. feels like a long time ago! Training a language model from scratch and watching it learn to speak, then learn concepts, then learn to think, feels so completely different from using an

tutorialslinus-lee--x
22 Jul 2026
Model Releases

Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

DGX agent

Scientists today face challenges of extraordinary scale and complexity. From shaping and simulating the intricate dynamics of fusion plasma, to exploring the vast search space of new materials, to mak

model-releasesgoogle-cloud-ai
22 Jul 2026
Model Releases

Are there MBA programs teaching fear marketing yet

DGX agent

Are there MBA programs teaching fear marketing yet We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production

model-releasesjerry-liu--x
22 Jul 2026
Model Releases

browser-search v2.0 — From the balaclava to the badge: your agent now browses everywhere

DGX agent

Today an AI agent trying to browse the web is like a thief in a balaclava sneaking around a police academy. Site protections block it, challenge it, turn it away. browser-search flips the script: your

model-releasesr-ollama
22 Jul 2026
Safety

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log ana…

DGX agent

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log analysis, we first used frontier models behind commercial APIs.

safetyclem-delangue--x
22 Jul 2026
Model Releases

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

DGX agent

Paper:https://arxiv.org/abs/2607.19058 Code (GitHub):https://github.com/nuemaan/skewadam Hi everyone, I just published a preprint on a new optimizer designed to tackle the massive VRAM bottleneck in M

model-releasesr-machinelearning
22 Jul 2026
Model Releases

Tucked away in this article is an appeal to the AI skeptics to PLEASE stop writing off stories like this OpenAI accidental exploit of Huggin…

DGX agent

Tucked away in this article is an appeal to the AI skeptics to PLEASE stop writing off stories like this OpenAI accidental exploit of Hugging Face as a dishonest marketing trick Frontier models can fi

model-releasessimon-willison--x
22 Jul 2026
Model Releases

GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper…

DGX agent

GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper clip benchmark for future models to max We're partnering wi

model-releasesemad-mostaque--x
21 Jul 2026
Model Releases

OpenAI says it accidentally hacked Hugging Face with a new AI system

DGX agent

OpenAI says its AI models mistakenly breached open-source AI platform Hugging Face during internal testing. In a blog post on Tuesday, OpenAI writes that GPT-5.6 Sol and 'an even more capable pre-rele

model-releasesthe-verge-ai
21 Jul 2026
Model Releases

A plug-and-play approach with fast uncertainty quantification for weak lensing mass mapping

DGX agent

arXiv:2603.22006v2 Announce Type: replace-cross Abstract: Upcoming stage-IV surveys such as Euclid and Rubin will deliver vast amounts of high-precision data, opening new opportunities to constrain co

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents

DGX agent

arXiv:2511.18685v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) show promising results as decision-making engines for embodied agents operating in complex, physical enviro

model-releasesarxiv-cs-cv
16 Jul 2026
Agents

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

DGX agent

arXiv:2607.13125v1 Announce Type: cross Abstract: We introduce Boogu-Image-0.1, an open-source unified multimodal understanding and generation model family, comprising Base, Turbo, Edit, and Edit-Turb

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Data-Efficient Adaptation of LLMs via Attention Head Reweighting

DGX agent

arXiv:2607.13425v1 Announce Type: cross Abstract: Learning effectively from limited data is critical in domains like security where labeled examples are scarce. Large language models (LLMs) have demon

model-releasesarxiv-cs-ai
16 Jul 2026
Local Ai

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

DGX agent

arXiv:2607.13960v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future visual observations, using future scene evolution as den

local-aiarxiv-cs-ro
16 Jul 2026
Model Releases

Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction

DGX agent

arXiv:2607.13737v1 Announce Type: new Abstract: For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

DGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

model-releasesarxiv-cs-ai
16 Jul 2026
← Previous
1…427428429430431…1327
Next →