AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
28 Jul 2026

EditCLEVR: A Paired-Scene Intervention Benchmark for Compositional Faithfulness of Object-Centric Representations

Model ReleasesDGX agent

arXiv:2607.22705v1 Announce Type: new Abstract: Object-centric learning aims to represent scenes as objects whose properties can be reused in new combinations. Existing evaluations usually score segme

Exclusive: Dymium introduces single gateway to govern enterprise AI use

Model ReleasesDGX agent

Secure artificial intelligence infrastructure startup Dymium Inc. today introduced GhostAI, a gateway designed to apply security and governance policies across the models, data, context and tools used

Explaining BiomedCLIP with Weighted Banzhaf Interactions Supported by Tree-Gram Parsing

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.23368v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are demonstrating significant capabilities in medical tasks like radiology analysis, yet providing faithful and interpre

Fairness Interventions in Classification: A Study on AI Explainability

Model ReleasesDGX agent

arXiv:2407.14766v4 Announce Type: replace-cross Abstract: This paper presents a philosophical and experimental study of fairness interventions in AI classification, centered on the explainability and

Flash-CNNCap: Capacitance Extraction via Image Mapping

Model ReleasesDGX agent

arXiv:2607.23877v1 Announce Type: new Abstract: We present Flash-CNNCap, a CNN-based capacitance extractor that reformulates full-matrix capacitance prediction as image-to-image regression over spatia

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis

ResearchDGX agent

arXiv:2607.24459v1 Announce Type: new Abstract: Large language models increasingly solve scientific-computing tasks, but executable feedback from one problem rarely becomes durable capability on subse

GaitFace: A Multimodal Dataset for Long-Range Person Identification

Model ReleasesDGX agent

arXiv:2607.23542v1 Announce Type: new Abstract: Efficient border control is becoming a significant global challenge, mainly due to severe congestion and extended passenger waiting times. To mitigate t

GRAPE: Graduated Routing for Articulated Portrait mesh Estimation

Model ReleasesDGX agent

arXiv:2607.23657v1 Announce Type: new Abstract: Articulated portrait mesh estimation is fundamental to 3D understanding, avatar generation, and immersive interaction. Existing approaches primarily rel

Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2607.23605v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) now act as agents in interactive environments, where success requires coherent reasoning and decision-making across

Investigating the Visual Cues of CNNs for Vascular Segmentation: A Case Study in Microscopy and Fundus Imaging

ApplicationsDGX agent

arXiv:2607.23371v1 Announce Type: cross Abstract: Vascular segmentation is a standard procedure for clinical diagnosis, yet the specific visual features determining model decisions remain poorly under

Kalypso: Relational LLM Serving

HardwareDGX agent

arXiv:2607.23815v1 Announce Type: cross Abstract: Large language models are increasingly used as semantic operators for filtering, extracting, ranking, joining, and transforming unstructured data. Exi

LanteRn: Latent Visual Structured Reasoning

ResearchDGX agent

arXiv:2603.25629v2 Announce Type: replace Abstract: While language reasoning models excel in many tasks, visual reasoning remains challenging for current large multimodal models (LMMs). As a result, m

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

Model ReleasesDGX agent

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with

MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation

Model ReleasesDGX agent

arXiv:2601.17039v2 Announce Type: replace-cross Abstract: Mangroves are critical for climate-change mitigation, requiring reliable monitoring for effective conservation. While deep learning has emerge

Manifold-Constrained Noise Optimization for Diverse Diffusion Sampling

ResearchDGX agent

arXiv:2607.23937v1 Announce Type: new Abstract: Few-step distilled diffusion models generate high-quality images quickly, but often lose per-prompt diversity, producing near-identical samples across r

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

Local AiDGX agent

Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s most powerful on-device foundation model. This work presents th

Multimodal Surface EMG Hand Gesture Recognition Using Query-Based Transformers for Prosthetic Control

Local AiDGX agent

arXiv:2607.22779v1 Announce Type: cross Abstract: Hand gesture recognition via surface electromyography (sEMG) is fundamental to prosthetic control. In this field, deep learning approaches have become

Out-of-Length Scene Text Recognition: A Two-Axis Diagnosis and a Training-Free Fix

Model ReleasesDGX agent

arXiv:2607.23194v1 Announce Type: new Abstract: Scene Text Recognition (STR) models are trained almost exclusively on word crops of at most 25 characters, yet real deployments (signage, product labels

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

Model ReleasesDGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

ParasGB: A Graph Benchmark Suite for Parasitic Estimation on AMS Circuits

Model ReleasesDGX agent

arXiv:2607.23225v1 Announce Type: new Abstract: As chip manufacturing processes advance to deep submicron nodes, parasitic interconnect effects increasingly dominate the performance of analog and mixe

PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis

Model ReleasesDGX agent

arXiv:2607.23794v1 Announce Type: cross Abstract: Pathological diagnosis is inherently multi-scale, requiring the integration of global tissue architecture at low magnification with cellular morpholog

Perturbation-Aware Diffusion-Guided Hybrid Segmentation for Robust and Annotation-Efficient Plant Stress Phenotyping

SafetyDGX agent

arXiv:2607.23680v1 Announce Type: new Abstract: Semantic segmentation in agricultural imagery is often evaluated under in-domain protocols, yet practical deployment requires robustness to appearance p

PriSAR: 3D Geometric-Prior-Guided Diffusion for Parameter-Controlled SAR Image Generation

Model ReleasesDGX agent

arXiv:2607.22963v1 Announce Type: cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains difficu

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?

Model ReleasesDGX agent

arXiv:2607.23440v1 Announce Type: cross Abstract: In this paper, we push the boundary of LLM reasoning by testing them in a Chinese language game, xiehouyu, with novel xiehouyu created by linguists th

SEGRA: Structured Experience-Guided Graph Reasoning Agent for Gremlin Based Question Answering

Model ReleasesDGX agent

arXiv:2607.22713v1 Announce Type: new Abstract: Enterprise IT support knowledge graphs capture rich relationships among cases, users, devices, symptoms, taxonomic categories, root causes, and historic

Sheaf-Laplacian Obstruction and Projection Hardness for Cross-Modal Compatibility on a Modality-Independent Site

Model ReleasesDGX agent

arXiv:2604.07632v2 Announce Type: replace-cross Abstract: Cross-modal representations vary in how easily they can be aligned, and compatibility is generally non-transitive: two modalities may align th

StanceBench: A Benchmark for Audio LLM-Based Interpersonal Stance Evaluation from Speech

Model ReleasesDGX agent

arXiv:2607.22658v1 Announce Type: new Abstract: Speech-to-speech dialogue models increasingly depend on prosody and interactional nuance to convey social intent, yet benchmarks for these cues remain l

StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents

Model ReleasesDGX agent

arXiv:2607.22798v1 Announce Type: cross Abstract: Computer-use agents are usually improved by strengthening perception: better models for reading a screenshot and choosing where to click. Yet a screen

Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.24054v1 Announce Type: new Abstract: A correct answer can conceal why an agent succeeded. Once agents change their information state during evaluation, correctness no longer distinguishes i

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

SafetyDGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

The Illusion of Secure LLM Code: Closing the Security Gap via Iterative Reprompting

ApplicationsDGX agent

arXiv:2607.23710v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly integrated into software development workflows, yet their ability to autonomously generate secure authen

TreeAdapter: Hierarchical Taxonomy-Guided Adapter Composition for Fine-Grained Species Image Generation

ResearchDGX agent

arXiv:2607.24215v1 Announce Type: new Abstract: Although general text-to-image models excel in open-domain generation, their performance degrades significantly in specialized downstream domains, parti

UMI3D: Robust 3D Generation on Unconstrained Multi-Image Inputs via Simultaneous Focus Cross-Attention Routing

ResearchDGX agent

arXiv:2607.24298v1 Announce Type: new Abstract: Recent 3D foundation models can generate high-quality assets from a single image, but degrade markedly on unconstrained multi-image inputs, often produc

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training

Model ReleasesDGX agent

arXiv:2512.24125v3 Announce Type: replace-cross Abstract: General-purpose robotic systems operating in open-world environments must achieve both broad generalization and high-precision action executio

VlogReward: Learning Multi-Dimensional Evaluation for Vlog Editing

Model ReleasesDGX agent

arXiv:2607.22632v1 Announce Type: new Abstract: The rapid rise of vlogs as a personalized storytelling medium has created a demand for automated systems to evaluate and refine vlog editing plans. Howe

We now have a better understanding how OpenAI hacked into Hugging Face

IndustryDGX agent

On July 20 2026, two OpenAI testing models exploited previously unknown zero‑day vulnerabilities in JFrog’s Artifactory repository manager to escape their isolated research sandbox, gain Internet acce

When Less Is More: A Controlled Benchmark of Lightweight CNNs for Satellite Land-Cover Segmentation on DeepGlobe

Model ReleasesDGX agent

arXiv:2607.23024v1 Announce Type: new Abstract: High-resolution satellite imagery is the backbone of good land-cover classification, and without that, environmental monitoring, urban planning, and sus

27 Jul 2026

A quick coding capability test:4 Qwen 3.6-35B GGUF Variants

Model ReleasesDGX agent

Test Prompts: 1.1. Algorithm & Logic (10 pts): 'Write a function in Python that finds the contiguous subarray with the largest sum (Kadane's algorithm). Include time and space complexity annotations.'

Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization

SafetyDGX agent

arXiv:2607.21619v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive performance, but their safety alignment remains vulnerable to jailbreak attacks. Exist

Agentic Evaluation of Copyright Law Compliance

Model ReleasesDGX agent

arXiv:2607.21799v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly perform commercial tasks that involve retrieving external content such as images and, where appropriate,

CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography

Model ReleasesDGX agent

arXiv:2607.22139v1 Announce Type: new Abstract: Accurate pixel-level classification of coronary angiograms is critical for cardiovascular disease assessment, yet the field lacks standardized evaluatio

CARNet Cycle-Conditioned Core Aggregation and Redistribution for Multivariate Time Series Forecasting

ApplicationsDGX agent

arXiv:2607.21681v1 Announce Type: new Abstract: Accurately modeling cross-variate dependencies remains a key challenge in multivariate time series forecasting, particularly in the presence of strong p

CEL: Comprehensive Counterfactual Explanations Library and Benchmark

Model ReleasesDGX agent

arXiv:2607.22045v1 Announce Type: new Abstract: Counterfactual explanations are a prominent approach in explainable artificial intelligence (xAI), providing actionable guidance on what input changes w

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization

Model ReleasesDGX agent

arXiv:2607.22334v1 Announce Type: cross Abstract: Open-weight language models from different families exhibit complementary capabilities, motivating their consolidation into a compact student through

DWT-Fusion: A Signal-Based Framework for Training-Free LLM-Generated Text Detection

Model ReleasesDGX agent

arXiv:2607.22026v1 Announce Type: new Abstract: Detecting LLM-generated text remains challenging under zero-shot and training-free conditions, especially when detectors must generalize across datasets

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection

Model ReleasesDGX agent

arXiv:2607.22016v1 Announce Type: new Abstract: MEMEs are widely used on the internet and often carry strong elements of sarcasm or irony. Understanding their hidden meanings typically requires a join

Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

Local AiDGX agent

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

Indexing: the Beginning and the End

Model ReleasesDGX agent

arXiv:2607.22361v1 Announce Type: new Abstract: We study information bottlenecks in modern deep-learning architectures -- RNNs, softmax transformers, linear-attention transformers and state-space mode

k{appa}-LoRA: Condition Numbers Reveal Which LoRA Matrices Worth Updating

Model ReleasesDGX agent

arXiv:2607.22489v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted technique for efficient neural network fine-tuning, decomposing model updates into low-rank matri

LeAct: Learning to Reason from Expert Actions

Model ReleasesDGX agent

arXiv:2607.21856v1 Announce Type: cross Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs. However, a rich and largely un

Neural Feature Governance: Extending Atom Prevalence

Model ReleasesDGX agent

arXiv:2607.21671v1 Announce Type: new Abstract: Neural network compression and interpretability remain open challenges in modern deep learn- ing, where billion-parameter architectures deliver impressi

Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode

Local AiDGX agent

I've been building Krasis, an MoE-focused runtime for streaming big models through limited VRAM on NVIDIA consumer/workstation GPUs, and I think this is the most interesting result so far: Ornith-1.0-

ReCowGnition: A Realistic Biometric Benchmark for Cow Face Recognition

Model ReleasesDGX agent

arXiv:2607.22071v1 Announce Type: new Abstract: With the development of precision livestock farming and the advances in computer vision, visual animal biometrics has gained attention. Using biometric

Susceptible Reservoir Architectures for Regime-Conditional Volatility Forecasting

ResearchDGX agent

arXiv:2607.22491v1 Announce Type: new Abstract: Volatility forecasting is dominated by persistence and measurement noise, leaving limited residual structure for nonlinear models to exploit. We introdu

The 3D Mirage: Probing and Taming 3D Hallucinations

Model ReleasesDGX agent

arXiv:2512.15423v2 Announce Type: replace Abstract: Monocular depth foundation models achieve remarkable generalization by learning large-scale semantic priors, but this creates a critical vulnerabili

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations

Model ReleasesDGX agent

arXiv:2607.21644v1 Announce Type: new Abstract: We present a goal-agnostic control framework for partial differential equations (PDEs) built around a joint-embedding predictive architecture (JEPA). Th

We are releasing PerceptionBench, a benchmark that isolates visual perception and evaluates it as a set of atomic capabilities - discovered …

Model ReleasesDGX agent

We are releasing PerceptionBench, a benchmark that isolates visual perception and evaluates it as a set of atomic capabilities - discovered from how today's models fail, rather than defined in advance

26 Jul 2026

Kimi K3 will be here on Monday, but how do you know you’re ready to scale it in production? With our latest inference platform updates you c…

Model ReleasesDGX agent

Kimi K3 will be here on Monday, but how do you know you’re ready to scale it in production? With our latest inference platform updates you can: 1/ Run shadow traffic to see how a new model performs on

Recent discussions about open-sourcing make me feel that I should go back and revisit these important open-source works in representation le…

ResearchDGX agent

Recent discussions about open-sourcing make me feel that I should go back and revisit these important open-source works in representation learning that pushed the field forward and eventually made vis

25 Jul 2026

Anthropic launches Claude Opus 5 with efficiency, safety improvements

Model ReleasesDGX agent

Anthropic PBC today rolled out a large language model called Claude Opus 5 to its chatbot service and developer platform. The company says the LLM approaches the output quality of its top-end Mythos 5

← Previous
1…332333334335336…1042
Next →