AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robo…

DGX agent

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robots — Gemini 3.5 Flash-Lite is our fastest, most cost-effecti

model-releasesgoogle-ai--x
31 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

DGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

model-releasesarxiv-cs-lg
31 Jul 2026
Research

LAST: The Last Query Token Guides Visual Token Pruning for Edge-Cloud Collaborative MLLM Inference

DGX agent

arXiv:2607.27952v1 Announce Type: new Abstract: Multimodal foundation models are reshaping edge-cloud visual intelligence from task-specific feature pipelines into token-based interfaces, where edge d

researcharxiv-cs-cv
31 Jul 2026
Model Releases

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts

DGX agent

arXiv:2607.27787v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) for mathematical reasoning suffers from a structural blind spot: on 'cliff' prompts-those on which

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

DGX agent

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageabl

model-releasesr-localllama
31 Jul 2026
Model Releases

MMHBench: A Multi-Perspective Benchmark for Mental Health Understanding in Long-Form Videos

DGX agent

arXiv:2607.27895v1 Announce Type: cross Abstract: Mental health understanding in long-form videos requires nuanced reasoning over observable behavior, interpersonal context, and latent psychological s

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding

DGX agent

arXiv:2511.12449v3 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) have significantly advanced e-commerce product understanding. However, they still face three challen

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

PCAP-LM: An LLM-Native Text Representation for TLS Bulk Traffic Analysis

DGX agent

arXiv:2607.28100v1 Announce Type: cross Abstract: Large language models (LLMs) offer powerful reasoning capabilities for network traffic analysis, but standard capture formats and their textual equiva

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Policy Gradient Steering: Interventions from Behavioral Objectives

DGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

Private Face Recognition Training Dataset Publication via Identity-Decoupled and Geometry-Preserving Face Distillation

DGX agent

arXiv:2607.27764v1 Announce Type: new Abstract: Publishing private face recognition~(FR) training datasets is privacy-sensitive because faces expose identity information. Private FR training dataset p

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

RedFlow: Redirect Failure into Action-Level Corrections for Flow-matching VLA Policy

DGX agent

arXiv:2607.27782v1 Announce Type: new Abstract: Flow-matching Vision-Language-Action (VLA) policies have shown strong potential for robotic manipulation but often suffer from compounding errors caused

model-releasesarxiv-cs-ro
31 Jul 2026
Research

S-Avatar: Diffusion-Guided Gaussian Head Avatars from a Single Image

DGX agent

arXiv:2607.28164v1 Announce Type: new Abstract: We propose S-Avatar, a novel method for generating photorealistic 3D head avatars from a single image using a diffusion-guided 3D model generation modul

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Shared Symbolic Backbones for Physically Consistent Multi-Output Symbolic Regression

DGX agent

arXiv:2607.26528v1 Announce Type: cross Abstract: Symbolic regression provides analytical expressions, but it is usually applied one output at a time. This is limiting in process systems, where state

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Some deepseek-v4-flash 20260731 opinion review

DGX agent

First of all, I want to apologize if it's off-topic or in the wrong format. Having tried Deepseek Flash with reasoning high on a conceptually difficult task, involving Machine Learning classifiers and

model-releasesr-localllama
31 Jul 2026
Research

Theia: Large-Scale Multimodal Captioning and Automated Validation of the Incidents1M Dataset for Data-Free Distillation

DGX agent

arXiv:2607.28269v1 Announce Type: new Abstract: The deployment of Vision-Language Models (VLMs) in critical domains like disaster management requires high-quality multimodal datasets, especially for t

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Thinking Once Is Enough: Intermediate-Layer Evidence Routing for High-Resolution VQA

DGX agent

arXiv:2607.27830v1 Announce Type: new Abstract: High-resolution visual question answering (HR-VQA) is often treated as a problem of insufficient evidence acquisition, where failing multimodal large la

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

This will happen frequently as AI becomes smarter and more agentic

DGX agent

This will happen frequently as AI becomes smarter and more agentic In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or wh

model-releaseselon-musk--x
31 Jul 2026
Agents

ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

DGX agent

arXiv:2607.27528v1 Announce Type: cross Abstract: Threat modeling is essential for secure software development, yet manual analysis of cloud-native architectures is slow and demands scarce security ex

agentsarxiv-cs-cl
31 Jul 2026
Tutorials

Towards Robust Monocular Depth Estimation in Non-Lambertian Surfaces

DGX agent

arXiv:2408.06083v2 Announce Type: replace Abstract: In the field of monocular depth estimation (MDE), many models with excellent zero-shot performance in general scenes emerge recently. However, these

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

DGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

model-releasesarxiv-cs-cv
31 Jul 2026
Agents

Training Skills Like Parameters via Self-Supervised Semantic Diffusion

DGX agent

arXiv:2607.27557v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable general instruction-following capabilities, they often fall short of human experts in highly s

agentsarxiv-cs-cl
31 Jul 2026
Model Releases

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improve…

DGX agent

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improvement a concrete, executable testbed. OpenMLE is an open full

model-releasesdair-ai--x
31 Jul 2026
Research

VETO: Towards Protecting Images From Frontier AI Editing

DGX agent

arXiv:2607.27292v1 Announce Type: new Abstract: The rise of powerful, accessible image-editing models such as FLUX.2 has brought high-fidelity editing within broad reach. Their capabilities now extend

researcharxiv-cs-cv
31 Jul 2026
Model Releases

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

DGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

model-releasesjerry-liu--x
31 Jul 2026
Model Releases

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

DGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

What's your local AI coding setup on a MacBook Pro M4?

DGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

model-releasesr-ollama
31 Jul 2026
Model Releases

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

DGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

model-releasesperplexity--x
30 Jul 2026
Model Releases

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

DGX agent

arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last b

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

DGX agent

arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization

DGX agent

arXiv:2607.26232v1 Announce Type: new Abstract: Background manipulation is a practical but under-specified image-forensics setting: the manipulated evidence can sit outside the salient foreground obje

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility

DGX agent

arXiv:2607.26828v1 Announce Type: new Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adapti

model-releasesarxiv-cs-lg
30 Jul 2026
Research

Dense Local Dependencies Induce Attention-Logit Explosion and Training Instability During Long-Sequence Transformer Training

DGX agent

arXiv:2505.15548v2 Announce Type: replace Abstract: Autoregressive transformer language models frequently exhibit training instability when trained on long sequences, particularly under low-precision

researcharxiv-cs-lg
30 Jul 2026
Research

Dissecting Sensitivity to Training Language in Self-Supervised Speech Learning Using Neural Audio Codec Tokens

DGX agent

arXiv:2607.26350v1 Announce Type: cross Abstract: Neural audio codecs (NACs) have become popular for obtaining speech representations as discrete tokens. Beyond compression, discrete tokens can be use

researcharxiv-cs-cl
30 Jul 2026
Research

DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues

DGX agent

arXiv:2607.26178v1 Announce Type: new Abstract: Turn-taking is a central component of full-duplex interaction. Which turn-taking behaviors are appropriate varies with the scenario, yet current models

researcharxiv-cs-cl
30 Jul 2026
Model Releases

Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces

DGX agent

arXiv:2505.16035v3 Announce Type: replace Abstract: We introduce Equivariant Neural Eikonal Solvers, a novel framework that integrates Equivariant Neural Fields (ENFs) with Neural Eikonal Solvers. Our

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making th…

DGX agent

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @Goog

model-releasesgoogle-ai--x
30 Jul 2026
Research

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥

DGX agent

Inkling-small. 2 weeks after inkling Nearly as good as Inkling but 4x smaller. We're just getting started...🔥 Today, we are releasing Inkling-Small. Inkling-Small achieves comparable performance to In

researchsoumith-chintala--x
30 Jul 2026
Local Ai

Is it possible to have multiple concept in one LORA?

DGX agent

I have question. I am trying to train a LORA, and my concept is for Indian wedding and tradional wardrobe based on Regions. I was planning to train a model which understand each region clothing style

local-air-stablediffusion
30 Jul 2026
Model Releases

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/207051816739969…

DGX agent

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/2070518167399698490?s=20 'I tried to analyze why Claude Code uses more toke

model-releasessebastian-raschka--x
30 Jul 2026
Model Releases

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of…

DGX agent

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we ar

model-releasesopenai--x
30 Jul 2026
Model Releases

MediaWiki Code2Code Search: Neural Retrieval for the Semantic Discovery of Open-Source Software Entities

DGX agent

arXiv:2607.26766v1 Announce Type: cross Abstract: Code search in large-scale ecosystems is often hindered by the lexical gap between user queries and implementation details, alongside the trade-off be

model-releasesarxiv-cs-cl
30 Jul 2026
Safety

Misalignment Has a Personality: A Big Five Account of Emergent Misalignment

DGX agent

arXiv:2607.26389v1 Announce Type: new Abstract: Fine-tuning a language model on data containing a narrow flaw, such as insecure code or incorrect mathematical answers, can cause broad misalignment thr

safetyarxiv-cs-cl
30 Jul 2026
Model Releases

OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding

DGX agent

arXiv:2607.27155v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing benchmarks provide limited support

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Progressive Multimodal Alignment for Continual Instruction Tuning

DGX agent

arXiv:2607.26947v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) rely on a projector to align visual representations with the language embedding space, making it central to cro

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV

DGX agent

arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent co

model-releasesarxiv-cs-lg
30 Jul 2026
Research

Representation Trajectories Matters: Complementary Evidence for OOD Detection and Image Classification

DGX agent

arXiv:2607.26565v1 Announce Type: new Abstract: Vision models do not form a representation at once; each block revises it. We ask whether the resulting computation path contains evidence that the fina

researcharxiv-cs-cv
30 Jul 2026
Research

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes

DGX agent

arXiv:2607.26627v1 Announce Type: new Abstract: Speculative Decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose tokens that are subsequently verif

researcharxiv-cs-cl
30 Jul 2026
Model Releases

Semantic-Aware Temporal Adaptation for UAV Anti-UAV Tracking

DGX agent

arXiv:2607.26511v1 Announce Type: new Abstract: UAV Anti-UAV tracking is an emerging low-altitude security task for localizing an adversarial UAV using the onboard camera of a moving observer UAV. It

model-releasesarxiv-cs-cv
30 Jul 2026
← Previous
1…497498499500501…1359
Next →