AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Research

Emergent Misalignment Recruits a Pre-existing Persona Subspace

DGX agent

arXiv:2607.21356v1 Announce Type: new Abstract: Fine-tuning an aligned language model on a narrow stream of bad advice can make it broadly misaligned on questions unrelated to the training data, a phe

researcharxiv-cs-lg
24 Jul 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

GPE: Evaluating Robust Evidence Aggregation for Fact Verification under Controllable GEO-Style Poisoning

DGX agent

arXiv:2607.20730v1 Announce Type: cross Abstract: Large language models increasingly use search tools to retrieve up-to-date information, introducing a new attack surface in which retrieved documents

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

GrainGS: Gradient-Decoupled Gaussian Splatting for Efficient Dynamic Novel View Synthesis

DGX agent

arXiv:2607.21448v1 Announce Type: new Abstract: Dynamic scene reconstruction with 3D Gaussian Splatting requires a balance between fine-grained motion modeling, structural stability, and compact repre

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests

DGX agent

arXiv:2607.20759v1 Announce Type: cross Abstract: AI coding agents powered by LLMs are increasingly integrated into real-world software development, where they generate, edit, and execute code with au

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

KeySI: An Interaction Framework for Tuning Text Embeddings Based on Human Feedback

DGX agent

arXiv:2607.20556v1 Announce Type: new Abstract: In large-scale text analysis tasks, pre-trained language models are often used to embed text corpora for downstream analysis. However, such models may s

safetyarxiv-cs-ai
24 Jul 2026
Research

MIRROR: Learning from the Other View for Multi-Modal Reasoning

DGX agent

arXiv:2607.21552v1 Announce Type: new Abstract: Unlike large language models (LLMs) that exhibit strong reasoning capabilities, vision-language models (VLMs) struggle with visual reasoning, even on ge

researcharxiv-cs-ai
24 Jul 2026
Model Releases

MO-GRPO: Mitigating Reward Hacking of Group Relative Policy Optimization on Multi-Objective Problems

DGX agent

arXiv:2509.22047v3 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has been shown to be an effective algorithm when an accurate reward model is available. However, such a hi

model-releasesarxiv-cs-lg
24 Jul 2026
Research

More Is Not More: What Matters for Diversity in LLM Opinions?

DGX agent

arXiv:2607.20429v1 Announce Type: cross Abstract: Large language models are increasingly used to simulate diverse human opinions in open-ended tasks such as synthetic surveys, focus group modeling, an

researcharxiv-cs-ai
24 Jul 2026
Model Releases

One More Turn, Less Regret: A Regret-Based Multi-Turn Benchmark for LLMs' Clarification Policies

DGX agent

arXiv:2607.21143v1 Announce Type: cross Abstract: Ambiguous user requests make clarification a sequential decision problem for conversational LLM assistants: they must decide whether to ask, what to a

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

People are using Minecraft farms as AI agent benchmarks

DGX agent

Someone modelled sugarcane farming as an integer program. See, sugarcane only grows next to water. Water costs one tile and can feed at most four cane tiles. The layout therefore becomes a coverage pr

model-releasesr-chatgpt
24 Jul 2026
Tutorials

Probabilistic Physics-Aware Machine Learning Predictions of Electric Truck Energy Consumption with Field Data

DGX agent

arXiv:2607.19054v2 Announce Type: replace Abstract: In this work, we incorporate first principle physics into the construction of data-driven methods by considering a model that accounts for the diffe

tutorialsarxiv-cs-lg
24 Jul 2026
Research

Probabilistic Residual Learning for Online Recommendations

DGX agent

arXiv:2607.20863v1 Announce Type: cross Abstract: Modern recommender systems are typically based on deep learning (DL) models, where a dense encoder learns representations of users and items. As a res

researcharxiv-cs-ai
24 Jul 2026
Model Releases

Same Dangerous Objective, Opposite Advice: Direct Exposure versus Multi-Agent Mediation

DGX agent

arXiv:2607.21518v1 Announce Type: new Abstract: Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Semi-Supervised Text-Attributed Graph Distillation

DGX agent

arXiv:2607.20477v1 Announce Type: new Abstract: {em Text-Attributed Graphs} (TAGs) have emerged as an expressive data model for integrating graph topology with rich textual semantics. Existing represe

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Surprisal Theory is Tautological (without Rational Grounding)

DGX agent

arXiv:2607.21574v1 Announce Type: new Abstract: Surprisal theory holds that the human processing difficulty of a linguistic unit in context is an affine function of its surprisal under some language m

researcharxiv-cs-cl
24 Jul 2026
Research

The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning

DGX agent

arXiv:2607.20952v1 Announce Type: cross Abstract: Latent, or silent, reasoning lets language models carry out intermediate computation in continuous vector space instead of words, and is widely assume

researcharxiv-cs-cl
24 Jul 2026
Model Releases

Towards an Automated Test of LLM Security Knowledge

DGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Towards Faithful Graph Explanations with Synergistic Edge Effects via Granular Balls

DGX agent

arXiv:2607.21381v1 Announce Type: new Abstract: Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previous methods explain graph neural networks

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

DGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms

DGX agent

arXiv:2607.20638v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation and reasoning, yet their ability to perform temporal reasoning ove

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

When Trivia Is Not Trivial: Everyday Knowledge Failures in Multilingual LLMs

DGX agent

arXiv:2607.21445v1 Announce Type: new Abstract: Quiz rooms, trivia nights, and quiz shows challenge human knowledge across a wide range of topics, from canonical facts to everyday culture. In this pap

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

DGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency

DGX agent

arXiv:2601.15368v3 Announce Type: replace Abstract: Generative image inpainting can produce realistic results even with large, irregular masks, but existing methods still suffer from two common proble

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection

DGX agent

arXiv:2512.16300v3 Announce Type: replace Abstract: Existing image forgery detection (IFD) methods either exploit low-level, semantics-agnostic artifacts or rely on multimodal large language models (M

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Current Injection Spiking Neural Network for Infrared and Visible Image Fusion

DGX agent

arXiv:2607.19879v1 Announce Type: new Abstract: Infrared and visible image fusion (IVIF) integrates the complementary information of two modalities into a single image with richer scene content. While

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM

DGX agent

TLDR: I (with the help of AI) re-implemented every Blackwell-only kernel (DeepGEMM, FlashInfer sparse-MLA, block-scaled FP8) in Triton, because they simply don't exist for sm89. The performance is 2-3

model-releasesr-localllama
23 Jul 2026
Local Ai

Development of an automated, reliable, and clinically meaningful artificial intelligence (AI) tool for diagnosing cardiac disease from conventional cardiovascular magnetic resonance (CMR) images

DGX agent

arXiv:2607.20087v1 Announce Type: new Abstract: Aims: Cardiovascular magnetic resonance (CMR) imaging enables non-invasive assessment of myocardial structure, function, and pathology, but requires sub

local-aiarxiv-cs-cv
23 Jul 2026
Model Releases

Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation: A Comparative Study

DGX agent

arXiv:2502.16022v3 Announce Type: replace Abstract: OpenNotes gives patients access to their EHR notes, but dense medical jargon limits comprehension. We evaluate closed-source and open-source LLMs fo

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Fluid-SDF: Ultra-Lightweight and Editable Implicit Shape Representation via Differentiable Primitives

DGX agent

arXiv:2607.18646v1 Announce Type: new Abstract: Implicit Neural Representations (INRs) have become the standard for continuous 2D shape modeling, but they suffer from black-box uneditability, vulnerab

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

GATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape Retrieval

DGX agent

arXiv:2607.19111v1 Announce Type: new Abstract: Large pretrained vision models have substantially improved appearance-based 3D shape retrieval, but they still confuse shapes that look similar while di

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Geospatial Diffusion-based Evolution Synthesis (GeoDES) for Storm-Centered Weather Augmentation

DGX agent

arXiv:2607.19522v1 Announce Type: new Abstract: While machine learning-based weather models hold significant promise, they struggle to predict the detailed structure of large-scale weather systems suc

researcharxiv-cs-lg
23 Jul 2026
Model Releases

GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]

DGX agent

The interesting finding from a new [arXiv paper](https://arxiv.org/abs/2607.16165) isn't that a frontier vision model failed a new benchmark, that happens weekly, but the specific shape of the failure

model-releasesr-machinelearning
23 Jul 2026
Local Ai

I run GLM-4.5-Air (110B) on 16Gb ram consumer machine and Qwen3-30B at 20 tok/s

DGX agent

In the past few months I’ve experimenting heavily and tortured my old 2016 Desktop PC to run the biggest Local LLM I can fit. I documented the whole process and research and I’ve published a repositor

local-air-ollama
23 Jul 2026
Tutorials

Look Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMs

DGX agent

arXiv:2607.20357v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have recently demonstrated strong performance across vision-language tasks. However, their high inference cost,

tutorialsarxiv-cs-cv
23 Jul 2026
Safety

Membership Inference Attacks for Unseen Classes

DGX agent

arXiv:2506.06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is data auditing, i.e., using statistical tools to determine whether harmful content may have been used in t

safetyarxiv-cs-lg
23 Jul 2026
Safety

Norm or Direction? Decoding Vision Mambas for High-Resolution Vision

DGX agent

arXiv:2607.18625v1 Announce Type: new Abstract: Vision Mamba models replace quadratic self-attention with linear complexity selective state space models (SSMs), emerging as efficient visual backbones.

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Self Gradient Forcing: Native Long Video Extrapolation

DGX agent

arXiv:2607.20368v1 Announce Type: new Abstract: Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own ro

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

DGX agent

arXiv:2607.19361v1 Announce Type: cross Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolation, which misses failures that arise only over a

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Taming the Security-Energy Paradox: A Green AI Approach to Optimized Android Malware Detection

DGX agent

arXiv:2607.20003v1 Announce Type: cross Abstract: An increase in advanced Android malware requires the use of deep learning models, which can run on Android devices. But there is a trade-off between s

researcharxiv-cs-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning

DGX agent

arXiv:2607.19790v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models r

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patches

DGX agent

TL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA

model-releasesr-stablediffusion
23 Jul 2026
Research

Trusting What You Cannot See: Auditable Fine-Tuning and Inference for Proprietary AI

DGX agent

arXiv:2603.07466v2 Announce Type: replace-cross Abstract: Cloud-based infrastructure has become the dominant platform for deploying large models, particularly large language models (LLMs). Fine-tuning

researcharxiv-cs-lg
23 Jul 2026
Safety

Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction

DGX agent

arXiv:2607.19532v1 Announce Type: cross Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models,

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training

DGX agent

arXiv:2607.19971v1 Announce Type: new Abstract: Accurate motion prediction of surrounding agents and safe motion planning are two closely coupled key tasks for social robot navigation in crowded envir

model-releasesarxiv-cs-ro
23 Jul 2026
Tutorials

4yr throwback. feels like a long time ago!

DGX agent

4yr throwback. feels like a long time ago! Training a language model from scratch and watching it learn to speak, then learn concepts, then learn to think, feels so completely different from using an

tutorialslinus-lee--x
22 Jul 2026
Model Releases

Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

DGX agent

Scientists today face challenges of extraordinary scale and complexity. From shaping and simulating the intricate dynamics of fusion plasma, to exploring the vast search space of new materials, to mak

model-releasesgoogle-cloud-ai
22 Jul 2026
Model Releases

Are there MBA programs teaching fear marketing yet

DGX agent

Are there MBA programs teaching fear marketing yet We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production

model-releasesjerry-liu--x
22 Jul 2026
← Previous
1…431432433434435…1338
Next →