AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Research

From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios

DGX agent

arXiv:2607.25546v1 Announce Type: new Abstract: Given a model that is already trained, which features does it rely on causally versus spuriously? Existing methods require access to the training proced

researcharxiv-cs-ai
29 Jul 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generative Distributionally Robust Optimization

DGX agent

arXiv:2607.24983v1 Announce Type: cross Abstract: Generative models are increasingly adopted in distributionally robust optimization (DRO), but existing approaches trade off model compatibility and ad

researcharxiv-cs-ai
29 Jul 2026
Model Releases

I pre-trained a 700m on 18B tokens optimized for Python and Wikitext | TheOneWhoWill/Shibai-700M-Base · Hugging Face

DGX agent

I know this is the 1000000th new sub billion parameter model out there and probably isn't as good as Qwen 3 0.6B or Qwen 3.5 0.8B but it still packs a decent punch. My intention to to continuously pre

model-releasesr-localllama
29 Jul 2026
Model Releases

IMPRINT: Image-Conditioned Query Enrichment for Long-Tail Object Goal Navigation

DGX agent

arXiv:2607.25106v1 Announce Type: new Abstract: Embodied AI increasingly relies on queryable semantic maps built from pre-trained vision-language models to enable zero-shot Object Goal Navigation (Obj

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Med-R^3: Enhancing Medical Retrieval-Augmented Reasoning of LLMs via Progressive Reinforcement Learning

DGX agent

arXiv:2507.23541v5 Announce Type: replace Abstract: In medical scenarios, effectively retrieving external knowledge and leveraging it for rigorous logical reasoning is of significant importance. Despi

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing

DGX agent

arXiv:2607.25300v1 Announce Type: new Abstract: Video editing is fundamentally message-driven: even from the same source footage, the selected shots change depending on the narrative the editor wishes

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

OmniPhys: Knowledge-Graph-Driven Benchmarking and Collective Optimization for Physical Commonsense in Text-to-Image Generation

DGX agent

arXiv:2607.25641v1 Announce Type: cross Abstract: While text-to-image models exhibit remarkable visual fidelity, they frequently violate fundamental physical commonsense. Existing benchmarks often rel

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Reinforcement Learning for Code Optimization

DGX agent

arXiv:2607.25970v1 Announce Type: cross Abstract: RL for code correctness is now established: have the model generate a program, run it against hidden test cases, and reward solutions that pass. Exten

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Rethinking CD: A Reproducibility Study and Extension on the Ineffectiveness of Contrastive Decoding at Mitigating Object Hallucinations in MLLMs

DGX agent

arXiv:2607.25196v1 Announce Type: new Abstract: Contrastive decoding (CD) has been proposed as a training-free strategy for mitigating object hallucinations in multimodal large language models (MLLMs)

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

ScaleResfusion: Residual Rectified Flow based on Residual Vector Field

DGX agent

arXiv:2607.25275v1 Announce Type: cross Abstract: Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Although recent diffusion-based

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenA…

DGX agent

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenAI. @_arohan_ was a pre-training lead on Gemini after years at

model-releasessonya-huang--x
29 Jul 2026
Local Ai

Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast

DGX agent

Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders

local-air-localllama
29 Jul 2026
Model Releases

Agent Team Work Zone: An Automated, Persistent Workspace for Long-Lived Coding Agent Teams

DGX agent

arXiv:2607.22917v1 Announce Type: new Abstract: Large Language Model (LLM) agents have significantly improved coding and programming workflows. Claude Code, in particular, is one of the most powerful

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Are Prompt Optimizers Blind? Cross-Modal Visual Feedback for Automatic Prompt Optimization

DGX agent

arXiv:2607.24354v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has been widely adopted to adapt vision-language models (VLMs) to downstream tasks without weight updates, yielding

researcharxiv-cs-ai
28 Jul 2026
Model Releases

AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use

DGX agent

arXiv:2505.12650v2 Announce Type: replace-cross Abstract: Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Biggest ever MCP update brings metadata, cybersecurity enhancements

DGX agent

The developers of the Model Context Protocol, an open-source technology that underpins many artificial intelligence applications, today released a new version of the software. The release is described

model-releasessiliconangle
28 Jul 2026
Safety

Breaking the Synthetic-Real Domain Shortcut for Training-Free Generative Replay-based Class Incremental Learning

DGX agent

arXiv:2607.22994v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to continuously acquire new knowledge while avoiding catastrophic forgetting. While exemplar replay is

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

CameraAnything: Refilming Videos with Arbitrary Camera Control

DGX agent

arXiv:2607.24591v1 Announce Type: new Abstract: We introduce CameraAnything, the first unified framework for camera controlled video editing that enables joint control of both intrinsic and extrinsic

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLook: Deeper Thinking with Lookahead

DGX agent

arXiv:2607.22602v1 Announce Type: new Abstract: Inference-time scaling has emerged as a powerful paradigm for improving large language model reasoning, often delivering larger gains on difficult reaso

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Diffusion-Guided Search via Exponential Tilting (DiffTilt): An Application to Falsification of Safety-Critical Systems

DGX agent

arXiv:2607.23134v1 Announce Type: new Abstract: Discovering rare safety-critical failures in autonomous and cyber-physical systems is a fundamental challenge in verification and validation. Existing f

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

DraftExpert: Expansion-Aware Self-Speculative Decoding for End-Device MoE Inference

DGX agent

arXiv:2607.24434v1 Announce Type: cross Abstract: Large Mixture-of-Experts (MoE) language models are attractive for end-device deployment because only a small subset of experts is active per token, bu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

EditCLEVR: A Paired-Scene Intervention Benchmark for Compositional Faithfulness of Object-Centric Representations

DGX agent

arXiv:2607.22705v1 Announce Type: new Abstract: Object-centric learning aims to represent scenes as objects whose properties can be reused in new combinations. Existing evaluations usually score segme

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Exclusive: Dymium introduces single gateway to govern enterprise AI use

DGX agent

Secure artificial intelligence infrastructure startup Dymium Inc. today introduced GhostAI, a gateway designed to apply security and governance policies across the models, data, context and tools used

model-releasessiliconangle
28 Jul 2026
Research

Explaining BiomedCLIP with Weighted Banzhaf Interactions Supported by Tree-Gram Parsing

DGX agent

arXiv:2607.23368v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are demonstrating significant capabilities in medical tasks like radiology analysis, yet providing faithful and interpre

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Fairness Interventions in Classification: A Study on AI Explainability

DGX agent

arXiv:2407.14766v4 Announce Type: replace-cross Abstract: This paper presents a philosophical and experimental study of fairness interventions in AI classification, centered on the explainability and

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Flash-CNNCap: Capacitance Extraction via Image Mapping

DGX agent

arXiv:2607.23877v1 Announce Type: new Abstract: We present Flash-CNNCap, a CNN-based capacitance extractor that reformulates full-matrix capacitance prediction as image-to-image regression over spatia

model-releasesarxiv-cs-lg
28 Jul 2026
Research

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis

DGX agent

arXiv:2607.24459v1 Announce Type: new Abstract: Large language models increasingly solve scientific-computing tasks, but executable feedback from one problem rarely becomes durable capability on subse

researcharxiv-cs-ai
28 Jul 2026
Model Releases

GaitFace: A Multimodal Dataset for Long-Range Person Identification

DGX agent

arXiv:2607.23542v1 Announce Type: new Abstract: Efficient border control is becoming a significant global challenge, mainly due to severe congestion and extended passenger waiting times. To mitigate t

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

GRAPE: Graduated Routing for Articulated Portrait mesh Estimation

DGX agent

arXiv:2607.23657v1 Announce Type: new Abstract: Articulated portrait mesh estimation is fundamental to 3D understanding, avatar generation, and immersive interaction. Existing approaches primarily rel

model-releasesarxiv-cs-cv
28 Jul 2026
Agents

Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning

DGX agent

arXiv:2607.23605v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) now act as agents in interactive environments, where success requires coherent reasoning and decision-making across

agentsarxiv-cs-ai
28 Jul 2026
Applications

Investigating the Visual Cues of CNNs for Vascular Segmentation: A Case Study in Microscopy and Fundus Imaging

DGX agent

arXiv:2607.23371v1 Announce Type: cross Abstract: Vascular segmentation is a standard procedure for clinical diagnosis, yet the specific visual features determining model decisions remain poorly under

applicationsarxiv-cs-cv
28 Jul 2026
Hardware

Kalypso: Relational LLM Serving

DGX agent

arXiv:2607.23815v1 Announce Type: cross Abstract: Large language models are increasingly used as semantic operators for filtering, extracting, ranking, joining, and transforming unstructured data. Exi

hardwarearxiv-cs-ai
28 Jul 2026
Research

LanteRn: Latent Visual Structured Reasoning

DGX agent

arXiv:2603.25629v2 Announce Type: replace Abstract: While language reasoning models excel in many tasks, visual reasoning remains challenging for current large multimodal models (LMMs). As a result, m

researcharxiv-cs-cv
28 Jul 2026
Model Releases

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

DGX agent

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with

model-releasesr-localllama
28 Jul 2026
Model Releases

MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation

DGX agent

arXiv:2601.17039v2 Announce Type: replace-cross Abstract: Mangroves are critical for climate-change mitigation, requiring reliable monitoring for effective conservation. While deep learning has emerge

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Manifold-Constrained Noise Optimization for Diverse Diffusion Sampling

DGX agent

arXiv:2607.23937v1 Announce Type: new Abstract: Few-step distilled diffusion models generate high-quality images quickly, but often lose per-prompt diversity, producing near-identical samples across r

researcharxiv-cs-cv
28 Jul 2026
Local Ai

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

DGX agent

Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s most powerful on-device foundation model. This work presents th

local-aiapple-ml-research
28 Jul 2026
Local Ai

Multimodal Surface EMG Hand Gesture Recognition Using Query-Based Transformers for Prosthetic Control

DGX agent

arXiv:2607.22779v1 Announce Type: cross Abstract: Hand gesture recognition via surface electromyography (sEMG) is fundamental to prosthetic control. In this field, deep learning approaches have become

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Out-of-Length Scene Text Recognition: A Two-Axis Diagnosis and a Training-Free Fix

DGX agent

arXiv:2607.23194v1 Announce Type: new Abstract: Scene Text Recognition (STR) models are trained almost exclusively on word crops of at most 25 characters, yet real deployments (signage, product labels

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

DGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

model-releasesr-localllama
28 Jul 2026
Model Releases

ParasGB: A Graph Benchmark Suite for Parasitic Estimation on AMS Circuits

DGX agent

arXiv:2607.23225v1 Announce Type: new Abstract: As chip manufacturing processes advance to deep submicron nodes, parasitic interconnect effects increasingly dominate the performance of analog and mixe

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis

DGX agent

arXiv:2607.23794v1 Announce Type: cross Abstract: Pathological diagnosis is inherently multi-scale, requiring the integration of global tissue architecture at low magnification with cellular morpholog

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Perturbation-Aware Diffusion-Guided Hybrid Segmentation for Robust and Annotation-Efficient Plant Stress Phenotyping

DGX agent

arXiv:2607.23680v1 Announce Type: new Abstract: Semantic segmentation in agricultural imagery is often evaluated under in-domain protocols, yet practical deployment requires robustness to appearance p

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

PriSAR: 3D Geometric-Prior-Guided Diffusion for Parameter-Controlled SAR Image Generation

DGX agent

arXiv:2607.22963v1 Announce Type: cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains difficu

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?

DGX agent

arXiv:2607.23440v1 Announce Type: cross Abstract: In this paper, we push the boundary of LLM reasoning by testing them in a Chinese language game, xiehouyu, with novel xiehouyu created by linguists th

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SEGRA: Structured Experience-Guided Graph Reasoning Agent for Gremlin Based Question Answering

DGX agent

arXiv:2607.22713v1 Announce Type: new Abstract: Enterprise IT support knowledge graphs capture rich relationships among cases, users, devices, symptoms, taxonomic categories, root causes, and historic

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Sheaf-Laplacian Obstruction and Projection Hardness for Cross-Modal Compatibility on a Modality-Independent Site

DGX agent

arXiv:2604.07632v2 Announce Type: replace-cross Abstract: Cross-modal representations vary in how easily they can be aligned, and compatibility is generally non-transitive: two modalities may align th

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…441442443444445…1371
Next →