AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,588 results
31 Jul 2026

Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory

HardwareDGX agent

arXiv:2607.28263v1 Announce Type: new Abstract: Transformer depth is not used uniformly: lower and middle layers build semantic representations, while upper layers increasingly specialize them for pre

Understanding Submodular Information Measure Based Objectives for Representation Learning: A Variance and Separation Perspective

ResearchDGX agent

arXiv:2607.27660v1 Announce Type: cross Abstract: Submodular Information Measures (SIMs) have recently emerged as a powerful framework for representation learning and multimodal learning. In particula

UniCross: Unified Cross-Skill Dexterous Manipulation Synthesis

SafetyDGX agent

arXiv:2607.28198v1 Announce Type: cross Abstract: Many dexterous manipulation tasks require the object to remain securely held throughout the interaction. From the perspective of hand-object relationa

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Unifying Adversarially Robust Model Experts in Vision-Language Models

SafetyDGX agent

arXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.

UrbanDS: A Graph-Guided LLM Multi-Agent System for Data-Intensive Urban Tasks

Model ReleasesDGX agent

arXiv:2607.26724v1 Announce Type: new Abstract: Large language model (LLM) agents have been widely applied in automating data science tasks. However, existing methods typically rely on a limited set o

Using an AMD V620 workstation card for ComfyUI - success

Model ReleasesDGX agent

A few weeks ago I posted about if it was worth using a V620 for Comfyui, and was told it likely wouldn't work, at least in Windows 11. And if it did, it would be far too slow and unusable. I decided t

Using Large Language Models for Idea Generation in Innovation

Model ReleasesDGX agent

arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for ne

VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation

SafetyDGX agent

arXiv:2607.28590v1 Announce Type: cross Abstract: Multimodal on-policy distillation (OPD) transfers fine-grained visual knowledge by supervising student-generated trajectories with a privileged-view t

Variance-Aware Baselines and Adaptive Learning Rates for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2511.23310v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the de

VCP-DCN: Beyond Visual Concealed Property via Depth Collaborative Network for Camouflaged Object Detection

Local AiDGX agent

arXiv:2607.27843v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to identify and segment camouflaged objects in complex environments, which are often concealed because their col

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that comm…

AgentsDGX agent

verbalizing one of those aha moments i had that seems retroactively pretty obvious: if you prioritize pretrain data quality enough that commoncrawl isn't good enough for you, you have to build a Whole

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improve…

Model ReleasesDGX agent

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improvement a concrete, executable testbed. OpenMLE is an open full

VESTIGE: A Knowledge-Guided Masking Strategy for Corruption-Aware Fine-Tuning of Genomic Transformers, Validated on Ancient DNA Reconstruction

Model ReleasesDGX agent

arXiv:2607.27712v1 Announce Type: new Abstract: Standard masked-language-model fine-tuning applies a uniform masking probability across every token position, assuming reconstruction difficulty is posi

VETO: Towards Protecting Images From Frontier AI Editing

ResearchDGX agent

arXiv:2607.27292v1 Announce Type: new Abstract: The rise of powerful, accessible image-editing models such as FLUX.2 has brought high-fidelity editing within broad reach. Their capabilities now extend

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System

AgentsDGX agent

arXiv:2607.27380v1 Announce Type: new Abstract: Text-to-video models have achieved remarkable visual quality, yet they still struggle to generate physically consistent dynamics because the temporal ev

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA

ApplicationsDGX agent

arXiv:2607.28442v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) and vision-language models (VLMs) have enabled new possibilities for 3D question answering (3D-QA), a ke

ViP-Rig: Visual-Prompted Controllable Rigging

ResearchDGX agent

arXiv:2607.27982v1 Announce Type: new Abstract: Rigging is inherently task-dependent because the same mesh may require different skeletons and deformation behaviors across animation tasks. In practice

VisualRouter: Query-Grounded Visual Sampling for Long Video Understanding

ResearchDGX agent

arXiv:2607.28463v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have achieved significant progress in video understanding, yet understanding long videos remains challenging due to

We applied BitNet-style ternary quantization to a super-resolution transformer. The whole model is 668 KB gzipped and runs in the browser.

Local AiDGX agent

Everyone's been doing 1.58-bit for LLMs, so we tried it on a vision transformer: Swin2SR (lightweight ×2 variant, 1.01M params), quantized so every weight is −1, 0, or +1 with a small per-group scale

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages a…

Model ReleasesDGX agent

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages are full of native text, which can be directly handled with Li

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

Model ReleasesDGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

Weather Emulators at the Frontier of Heat Extremes Predictability

ResearchDGX agent

arXiv:2607.28220v1 Announce Type: cross Abstract: Atmospheric predictability declines rapidly beyond the next ten days, such that forecasts at longer lead times primarily convey large-scale trends rat

We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

Model ReleasesDGX agent

I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B. On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is

What am I doing wrong in my LoRA training?

Local AiDGX agent

Hi everyone, this is my very first time training a LoRA, so I might be missing something basic! I trained an art style LoRA using noobaiXLNAIXL_vPred10Version as the base model, with 21 images, 8 epoc

What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation

Model ReleasesDGX agent

arXiv:2607.26935v1 Announce Type: new Abstract: Bot detectors deployed at scale treat traffic as binary: human or bot. This assumption breaks when AI agents browse the web through browser automation,

What Is The Performance Ceiling of My Classifier? Utilizing Category-Wise Influence Functions for Pareto Frontier Analysis

ResearchDGX agent

arXiv:2510.03950v2 Announce Type: replace Abstract: Data-centric learning seeks to improve model performance from the perspective of data quality, and has been drawing increasing attention in the mach

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

Model ReleasesDGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

What Makes Graph Unified? Principles and Generative Sliding-Window Transformer for Graph Foundation Models

TutorialsDGX agent

arXiv:2607.27966v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have recently emerged as a promising paradigm for general-purpose graph learning, aiming to learn reusable knowledge that

What to Remove, What to Preserve: Dual-Ambiguity Rectification for All-in-One Image Restoration

ResearchDGX agent

arXiv:2607.28526v1 Announce Type: new Abstract: All-in-one image restoration aims to handle diverse degradations within a unified framework. Existing methods commonly encode heterogeneous degradation

What’s new in AI infrastructure and orchestration this month

Model ReleasesDGX agent

At Google, AI is a soup-to-nuts endeavor. Obviously, we make leading AI models like Gemini and Nano Banana. We incorporate AI into the tools you use every day (think Gmail, BigQuery, AlloyDB, Google C

What's your local AI coding setup on a MacBook Pro M4?

Model ReleasesDGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

When Does Explicit View Routing Work? A Controlled Study of Multi-View Graph-Text Alignment

SafetyDGX agent

arXiv:2607.27530v1 Announce Type: new Abstract: Graph-text retrieval typically maps a graph and its description to a single embedding, even when a query concerns only one semantic aspect, such as a cl

When Does Muon Help Agentic Reinforcement Learning?

Model ReleasesDGX agent

arXiv:2607.16169v3 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its operating regime in reinforcement-learning post-training remains unclear. We map

When Robots Exchange Meaning: A Demo of Goal-Oriented Semantic Communications for Collaborative Robotics

ResearchDGX agent

arXiv:2607.28256v1 Announce Type: new Abstract: Collaborative robotics is a representative task-oriented 6G use-case, where communication quality should be reflected in mission execution, environment

When Should AI Follow? Task Structure and Joint Adaptation by Human and AI Agents

Local AiDGX agent

arXiv:2504.20903v4 Announce Type: replace-cross Abstract: How should organizations divide and sequence decision tasks between human and artificial agents? We develop a computational model of joint seq

When unlearning is free: leveraging low influence points to reduce computational costs

ApplicationsDGX agent

arXiv:2512.05254v2 Announce Type: replace Abstract: As concerns around data privacy in machine learning grow, the ability to unlearn, or remove, specific data points from trained models becomes increa

Where and When to Commit: Candidate-Aware Decoding for Diffusion Language Models

ResearchDGX agent

arXiv:2607.28166v1 Announce Type: new Abstract: Diffusion language models (DLMs) expose a provisional prediction at every denoising step, creating an opportunity for generation-time early exit that st

WhisperRec: Latent Reasoning for Efficient Foundation Recommendation Models

Model ReleasesDGX agent

arXiv:2607.26621v2 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong reasoning capabilities, motivating their adoption as backbones for foundation recommendation mod

Why are AI model tests always the same generic prompts?

Model ReleasesDGX agent

Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are 'make a car game,' 'make a website,' or something equally generic, usually from a prompt that's barely a lin

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

Model ReleasesDGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

WIDE: Boosting Adaptive LLM Inference via Token-level Dynamic Width Pruning

ApplicationsDGX agent

arXiv:2607.28418v1 Announce Type: cross Abstract: Pruning is a promising approach for improving the efficiency of LLMs. Existing static structured pruning methods are hardware-friendly and can deliver

Will ollama upgrade Deepseek V4 Flash on cloud?

Model ReleasesDGX agent

https://preview.redd.it/vxl4zslewigh1.png?width=1435&format=png&auto=webp&s=5a419870ca0cb13076be6c9ff4ef33d8177d5eda New version is 25% better than previous one and is near GLM-5.2 quality submitted b

Windowed thinning and query complexity for the bouncy particle and Zigzag samplers

ResearchDGX agent

arXiv:2607.28413v1 Announce Type: cross Abstract: Let mu(d x)propto e^{-U(x)} d x on R^d, where U is m-strongly convex and L-smooth, and denote by kappa=L/m the condition number. We consider windowed

Wiring diagram extraction and gluing: a case study in classifying figure skating jumps using 3D dataset

ApplicationsDGX agent

arXiv:2607.27598v1 Announce Type: cross Abstract: Hasse clustering is an algorithm that extracts common patterns in sequential data and represents them in graphical forms. As the number of expected cl

With release of Deepseek V4 I wanted see how the model sizes are trending over time. The trend is that by this time next year, we probably will have Opus 4.5 level models on consumer grade laptops!

Model ReleasesDGX agent

I was surprised to see that Deepseek V4 Flash is extremely smart and small enough to fit in setup that can be built with < $50,000. Expensive, but not a datacenter. So I wanted to see the trend over t

Witness Evidence Portfolios: Single-Prefill Risk Detection for Closed Multimodal Answers

ResearchDGX agent

arXiv:2607.27667v1 Announce Type: new Abstract: Reliable deployment of multimodal large language models (MLLMs) requires deciding whether a confident visual answer should be trusted, reviewed, or rout

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models

SafetyDGX agent

arXiv:2607.27599v1 Announce Type: cross Abstract: Building generalizable agents for diverse applications remains a fundamental challenge. While imitation learning-based policies succeed in specific tr

Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning

Model ReleasesDGX agent

arXiv:2607.28478v1 Announce Type: new Abstract: As large language models (LLMs) continue to advance in complex reasoning tasks, they have learned to heavily prioritize explicit conditions provided in

Write-Safe Flow Field Mapping under Ambiguous Onboard Sensing and Localization Drift

Local AiDGX agent

arXiv:2607.27713v1 Announce Type: new Abstract: Mobile robots can infer local flow structure from onboard sensing, but a locally plausible estimate is not always safe to write into a global map. Simil

X-NavDP: Generalizing Navigation Diffusion Policy to Novel Behavior and Embodiments with Group Q-score Reweighted Matching

Local AiDGX agent

arXiv:2607.28560v1 Announce Type: new Abstract: Pretraining navigation diffusion policies rely on large-scale expert demonstrations. These data are typically generated by a fully-informed oracle plann

You Only Look Omni Gradient Backpropagation for Moving Infrared Small Target Detection

ResearchDGX agent

arXiv:2511.13013v2 Announce Type: replace Abstract: Moving infrared small target detection is a key component of infrared search and tracking systems, yet it remains extremely challenging due to low s

ZAPs: A Reward Attribution Framework for DeFi Ecosystems with Adversarial-Robust Scoring via Parallel Anomaly Ensemble Detection

ApplicationsDGX agent

arXiv:2607.27859v1 Announce Type: cross Abstract: Incentive programs are central to user acquisition in decentralized finance, but many reward systems rely on raw volume, transaction count, and wallet

Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model

ResearchDGX agent

arXiv:2607.26742v1 Announce Type: cross Abstract: Zero-shot text-to-speech (TTS) clones a voice from a short audio prompt, but this reliance on reference audio is a barrier when only visual informatio

ZMIS-SAM: Segment Anything Model Enhanced with Wavelet Transform for Zooplankton Microscopy Image Instance Segmentation

ResearchDGX agent

arXiv:2607.27585v1 Announce Type: new Abstract: As primary consumers in the marine food chain, zooplankton play a crucial role in maintaining marine ecological balance. However, the Segment Anything M

ZUNA1.1: A more flexible EEG foundation model for Denoising and Super-resolution

Model ReleasesDGX agent

arXiv:2607.27308v1 Announce Type: new Abstract: We introduce ZUNA1.1, a 380M-parameter diffusion autoencoder for flexible EEG signal reconstruction. ZUNA1.1 is capable of reconstructing variable lengt

30 Jul 2026

2 images + 1 prompt > expected output

Local AiDGX agent

Hi, I'm trying to replicate a thing locally, that I can do on ChatGPT. What I want is to give a local AI two reference images (a face and a background item) and a prompt about the composition of the p

2× Radeon R9700 for Local AI Was Choosing AMD Instead of NVIDIA a Mistake Without CUDA?

Local AiDGX agent

Hello together I decided to go with 2× Radeon AI PRO R9700 GPUs (64 GB total VRAM) for my local AI server. However, I keep reading that AMD/ROCm is still not as mature as NVIDIA/CUDA when it comes to

3DGBGS: 3D Granular Ball Gaussian Splatting for Compact Novel View Synthesis

Local AiDGX agent

arXiv:2607.26578v1 Announce Type: new Abstract: Three-dimensional Gaussian Splatting (3DGS) enables high-quality real-time novel-view synthesis through explicit Gaussian primitives and differentiable

4090 + 5060 Ti + 64GB RAM: 206 t/s on a 35B-A3B, and a 122B at 37 t/s

Model ReleasesDGX agent

I've been benchmarking a two-card box for a few weeks and I still can't quite get over some of these numbers, so I'm dumping them here. Box: RTX 4090 (24GB) + RTX 5060 Ti (16GB), i9-13900K, 64GB DDR5.

A Closer Look at Dynamic Scene Graph Generation In the Era of Multimodal Large Language Models

ResearchDGX agent

arXiv:2503.15846v2 Announce Type: replace Abstract: Dynamic Scene Graph Generation (DSGG) aims to capture objects and their evolving relations in videos. Despite recent progress, the practicality and

← Previous
1…151152153154155…1410
Next →