AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlog
88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
Model Releases

ScaleResfusion: Residual Rectified Flow based on Residual Vector Field

DGX agent

arXiv:2607.25275v1 Announce Type: cross Abstract: Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Although recent diffusion-based

model-releasesarxiv-cs-ai
29 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenA…

DGX agent

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenAI. @_arohan_ was a pre-training lead on Gemini after years at

model-releasessonya-huang--x
29 Jul 2026
Local Ai

Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast

DGX agent

Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders

local-air-localllama
29 Jul 2026
Model Releases

Agent Team Work Zone: An Automated, Persistent Workspace for Long-Lived Coding Agent Teams

DGX agent

arXiv:2607.22917v1 Announce Type: new Abstract: Large Language Model (LLM) agents have significantly improved coding and programming workflows. Claude Code, in particular, is one of the most powerful

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Are Prompt Optimizers Blind? Cross-Modal Visual Feedback for Automatic Prompt Optimization

DGX agent

arXiv:2607.24354v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has been widely adopted to adapt vision-language models (VLMs) to downstream tasks without weight updates, yielding

researcharxiv-cs-ai
28 Jul 2026
Model Releases

AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use

DGX agent

arXiv:2505.12650v2 Announce Type: replace-cross Abstract: Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Biggest ever MCP update brings metadata, cybersecurity enhancements

DGX agent

The developers of the Model Context Protocol, an open-source technology that underpins many artificial intelligence applications, today released a new version of the software. The release is described

model-releasessiliconangle
28 Jul 2026
Safety

Breaking the Synthetic-Real Domain Shortcut for Training-Free Generative Replay-based Class Incremental Learning

DGX agent

arXiv:2607.22994v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to continuously acquire new knowledge while avoiding catastrophic forgetting. While exemplar replay is

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

CameraAnything: Refilming Videos with Arbitrary Camera Control

DGX agent

arXiv:2607.24591v1 Announce Type: new Abstract: We introduce CameraAnything, the first unified framework for camera controlled video editing that enables joint control of both intrinsic and extrinsic

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLook: Deeper Thinking with Lookahead

DGX agent

arXiv:2607.22602v1 Announce Type: new Abstract: Inference-time scaling has emerged as a powerful paradigm for improving large language model reasoning, often delivering larger gains on difficult reaso

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Diffusion-Guided Search via Exponential Tilting (DiffTilt): An Application to Falsification of Safety-Critical Systems

DGX agent

arXiv:2607.23134v1 Announce Type: new Abstract: Discovering rare safety-critical failures in autonomous and cyber-physical systems is a fundamental challenge in verification and validation. Existing f

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

DraftExpert: Expansion-Aware Self-Speculative Decoding for End-Device MoE Inference

DGX agent

arXiv:2607.24434v1 Announce Type: cross Abstract: Large Mixture-of-Experts (MoE) language models are attractive for end-device deployment because only a small subset of experts is active per token, bu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

EditCLEVR: A Paired-Scene Intervention Benchmark for Compositional Faithfulness of Object-Centric Representations

DGX agent

arXiv:2607.22705v1 Announce Type: new Abstract: Object-centric learning aims to represent scenes as objects whose properties can be reused in new combinations. Existing evaluations usually score segme

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Exclusive: Dymium introduces single gateway to govern enterprise AI use

DGX agent

Secure artificial intelligence infrastructure startup Dymium Inc. today introduced GhostAI, a gateway designed to apply security and governance policies across the models, data, context and tools used

model-releasessiliconangle
28 Jul 2026
Research

Explaining BiomedCLIP with Weighted Banzhaf Interactions Supported by Tree-Gram Parsing

DGX agent

arXiv:2607.23368v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are demonstrating significant capabilities in medical tasks like radiology analysis, yet providing faithful and interpre

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Fairness Interventions in Classification: A Study on AI Explainability

DGX agent

arXiv:2407.14766v4 Announce Type: replace-cross Abstract: This paper presents a philosophical and experimental study of fairness interventions in AI classification, centered on the explainability and

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Flash-CNNCap: Capacitance Extraction via Image Mapping

DGX agent

arXiv:2607.23877v1 Announce Type: new Abstract: We present Flash-CNNCap, a CNN-based capacitance extractor that reformulates full-matrix capacitance prediction as image-to-image regression over spatia

model-releasesarxiv-cs-lg
28 Jul 2026
Research

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis

DGX agent

arXiv:2607.24459v1 Announce Type: new Abstract: Large language models increasingly solve scientific-computing tasks, but executable feedback from one problem rarely becomes durable capability on subse

researcharxiv-cs-ai
28 Jul 2026
Model Releases

GaitFace: A Multimodal Dataset for Long-Range Person Identification

DGX agent

arXiv:2607.23542v1 Announce Type: new Abstract: Efficient border control is becoming a significant global challenge, mainly due to severe congestion and extended passenger waiting times. To mitigate t

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

GRAPE: Graduated Routing for Articulated Portrait mesh Estimation

DGX agent

arXiv:2607.23657v1 Announce Type: new Abstract: Articulated portrait mesh estimation is fundamental to 3D understanding, avatar generation, and immersive interaction. Existing approaches primarily rel

model-releasesarxiv-cs-cv
28 Jul 2026
Agents

Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning

DGX agent

arXiv:2607.23605v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) now act as agents in interactive environments, where success requires coherent reasoning and decision-making across

agentsarxiv-cs-ai
28 Jul 2026
Applications

Investigating the Visual Cues of CNNs for Vascular Segmentation: A Case Study in Microscopy and Fundus Imaging

DGX agent

arXiv:2607.23371v1 Announce Type: cross Abstract: Vascular segmentation is a standard procedure for clinical diagnosis, yet the specific visual features determining model decisions remain poorly under

applicationsarxiv-cs-cv
28 Jul 2026
Hardware

Kalypso: Relational LLM Serving

DGX agent

arXiv:2607.23815v1 Announce Type: cross Abstract: Large language models are increasingly used as semantic operators for filtering, extracting, ranking, joining, and transforming unstructured data. Exi

hardwarearxiv-cs-ai
28 Jul 2026
Research

LanteRn: Latent Visual Structured Reasoning

DGX agent

arXiv:2603.25629v2 Announce Type: replace Abstract: While language reasoning models excel in many tasks, visual reasoning remains challenging for current large multimodal models (LMMs). As a result, m

researcharxiv-cs-cv
28 Jul 2026
Model Releases

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

DGX agent

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with

model-releasesr-localllama
28 Jul 2026
Model Releases

MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation

DGX agent

arXiv:2601.17039v2 Announce Type: replace-cross Abstract: Mangroves are critical for climate-change mitigation, requiring reliable monitoring for effective conservation. While deep learning has emerge

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Manifold-Constrained Noise Optimization for Diverse Diffusion Sampling

DGX agent

arXiv:2607.23937v1 Announce Type: new Abstract: Few-step distilled diffusion models generate high-quality images quickly, but often lose per-prompt diversity, producing near-identical samples across r

researcharxiv-cs-cv
28 Jul 2026
Local Ai

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

DGX agent

Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s most powerful on-device foundation model. This work presents th

local-aiapple-ml-research
28 Jul 2026
Local Ai

Multimodal Surface EMG Hand Gesture Recognition Using Query-Based Transformers for Prosthetic Control

DGX agent

arXiv:2607.22779v1 Announce Type: cross Abstract: Hand gesture recognition via surface electromyography (sEMG) is fundamental to prosthetic control. In this field, deep learning approaches have become

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Out-of-Length Scene Text Recognition: A Two-Axis Diagnosis and a Training-Free Fix

DGX agent

arXiv:2607.23194v1 Announce Type: new Abstract: Scene Text Recognition (STR) models are trained almost exclusively on word crops of at most 25 characters, yet real deployments (signage, product labels

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

DGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

model-releasesr-localllama
28 Jul 2026
Model Releases

ParasGB: A Graph Benchmark Suite for Parasitic Estimation on AMS Circuits

DGX agent

arXiv:2607.23225v1 Announce Type: new Abstract: As chip manufacturing processes advance to deep submicron nodes, parasitic interconnect effects increasingly dominate the performance of analog and mixe

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis

DGX agent

arXiv:2607.23794v1 Announce Type: cross Abstract: Pathological diagnosis is inherently multi-scale, requiring the integration of global tissue architecture at low magnification with cellular morpholog

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Perturbation-Aware Diffusion-Guided Hybrid Segmentation for Robust and Annotation-Efficient Plant Stress Phenotyping

DGX agent

arXiv:2607.23680v1 Announce Type: new Abstract: Semantic segmentation in agricultural imagery is often evaluated under in-domain protocols, yet practical deployment requires robustness to appearance p

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

PriSAR: 3D Geometric-Prior-Guided Diffusion for Parameter-Controlled SAR Image Generation

DGX agent

arXiv:2607.22963v1 Announce Type: cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains difficu

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?

DGX agent

arXiv:2607.23440v1 Announce Type: cross Abstract: In this paper, we push the boundary of LLM reasoning by testing them in a Chinese language game, xiehouyu, with novel xiehouyu created by linguists th

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SEGRA: Structured Experience-Guided Graph Reasoning Agent for Gremlin Based Question Answering

DGX agent

arXiv:2607.22713v1 Announce Type: new Abstract: Enterprise IT support knowledge graphs capture rich relationships among cases, users, devices, symptoms, taxonomic categories, root causes, and historic

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Sheaf-Laplacian Obstruction and Projection Hardness for Cross-Modal Compatibility on a Modality-Independent Site

DGX agent

arXiv:2604.07632v2 Announce Type: replace-cross Abstract: Cross-modal representations vary in how easily they can be aligned, and compatibility is generally non-transitive: two modalities may align th

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

StanceBench: A Benchmark for Audio LLM-Based Interpersonal Stance Evaluation from Speech

DGX agent

arXiv:2607.22658v1 Announce Type: new Abstract: Speech-to-speech dialogue models increasingly depend on prosody and interactional nuance to convey social intent, yet benchmarks for these cues remain l

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents

DGX agent

arXiv:2607.22798v1 Announce Type: cross Abstract: Computer-use agents are usually improved by strengthening perception: better models for reading a screenshot and choosing where to click. Yet a screen

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation

DGX agent

arXiv:2607.24054v1 Announce Type: new Abstract: A correct answer can conceal why an agent succeeded. Once agents change their information state during evaluation, correctness no longer distinguishes i

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

DGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

safetyarxiv-cs-ai
28 Jul 2026
Applications

The Illusion of Secure LLM Code: Closing the Security Gap via Iterative Reprompting

DGX agent

arXiv:2607.23710v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly integrated into software development workflows, yet their ability to autonomously generate secure authen

applicationsarxiv-cs-ai
28 Jul 2026
Research

TreeAdapter: Hierarchical Taxonomy-Guided Adapter Composition for Fine-Grained Species Image Generation

DGX agent

arXiv:2607.24215v1 Announce Type: new Abstract: Although general text-to-image models excel in open-domain generation, their performance degrades significantly in specialized downstream domains, parti

researcharxiv-cs-cv
28 Jul 2026
Research

UMI3D: Robust 3D Generation on Unconstrained Multi-Image Inputs via Simultaneous Focus Cross-Attention Routing

DGX agent

arXiv:2607.24298v1 Announce Type: new Abstract: Recent 3D foundation models can generate high-quality assets from a single image, but degrade markedly on unconstrained multi-image inputs, often produc

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training

DGX agent

arXiv:2512.24125v3 Announce Type: replace-cross Abstract: General-purpose robotic systems operating in open-world environments must achieve both broad generalization and high-precision action executio

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

VlogReward: Learning Multi-Dimensional Evaluation for Vlog Editing

DGX agent

arXiv:2607.22632v1 Announce Type: new Abstract: The rapid rise of vlogs as a personalized storytelling medium has created a demand for automated systems to evaluate and refine vlog editing plans. Howe

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…425426427428429…1327
Next →