AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Local Ai

World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation

DGX agent

arXiv:2608.05369v1 Announce Type: cross Abstract: Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in r

local-aiarxiv-cs-cv
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Chained Recursive Language Models for Multi-Iteration Reasoning

DGX agent

arXiv:2608.05124v1 Announce Type: cross Abstract: Long context reasoning in large language models (LLMs) is usually constrained by the fact that a single inference trajectory has to simultaneously exp

researcharxiv-cs-ai
6 Aug 2026
Research

EdgeLM: Edge Demonstrations for Language Models' Table Understanding

DGX agent

arXiv:2608.04390v1 Announce Type: new Abstract: Large language models (LLMs) perform table-centric prediction through in-context learning, making demonstration selection critical to performance. Exist

researcharxiv-cs-cl
6 Aug 2026
Model Releases

Large Language Models for Low-Resource Languages: A Conceptual Framework for an Electronic Explanatory Dictionary of the Tajik Language

DGX agent

arXiv:2608.04186v1 Announce Type: new Abstract: This paper presents a conceptual framework for developing an electronic explanatory dictionary of the Tajik language using large language models (LLMs).

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

DGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight

DGX agent

arXiv:2608.04657v1 Announce Type: new Abstract: World action models (WAMs) built on video generation backbones are a rising recipe for robot learning, yet remain confined to tabletop manipulation. Mob

model-releasesarxiv-cs-cv
6 Aug 2026
Local Ai

MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation

DGX agent

arXiv:2608.05076v1 Announce Type: cross Abstract: Recent advances in machine learning have enabled training of wireless foundation models, which aim to support tasks such as channel estimation, beam p

local-aiarxiv-cs-ai
6 Aug 2026
Safety

Overcoming Statistical Bias in Action-Controllable World Models

DGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

Persistent Object Narratives for Token-Efficient Video Language Models

DGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

DGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

model-releasesarxiv-cs-ai
6 Aug 2026
Tutorials

Unforgettable Generalization in Language Models

DGX agent

arXiv:2409.02228v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained to forget (or 'unlearn'') a skill, how precisely does their behavior change? We study the behavior of t

tutorialsarxiv-cs-cl
6 Aug 2026
Model Releases

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models

DGX agent

arXiv:2608.03112v1 Announce Type: cross Abstract: Vision-language models excel at image and video understanding but suffer from high inference latency due to the need to process thousands of tokens pe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

DGX agent

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation m

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

BanglaWild: An In-the-Wild Bengali Scene Text Recognition Benchmark for OCR and Vision-Language Models

DGX agent

arXiv:2608.03884v1 Announce Type: cross Abstract: In-the-wild Bengali scene text recognition is largely unmeasured: existing resources target handwritten documents or constrained sign-board parsing, r

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

DGX agent

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, a

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Disentangling Language Modeling and Boundaries

DGX agent

arXiv:2608.03599v1 Announce Type: new Abstract: Byte-level language models are usually argued for on the grounds of robustness, multilingual fairness, and character-level skills. We point to a differe

safetyarxiv-cs-cl
5 Aug 2026
Research

Federated generative event models for tokenized electronic health records

DGX agent

arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer.

researcharxiv-cs-lg
5 Aug 2026
Research

FOUND-AF: Benchmarking ECG Foundation Models for Atrial Fibrillation Detection

DGX agent

arXiv:2608.03597v1 Announce Type: new Abstract: Atrial fibrillation (AF) is the most common sustained cardiac arrhythmia and is associated with increased risks of stroke, heart failure, and mortality.

researcharxiv-cs-ai
5 Aug 2026
Model Releases

HUKUKBERT: Domain-Specific Language Model for Turkish Law

DGX agent

arXiv:2604.04790v2 Announce Type: replace Abstract: Natural language processing (NLP) advances have powered a generation of LegalTech systems, but Turkish law remains under-served by domain-specific d

model-releasesarxiv-cs-cl
5 Aug 2026
Local Ai

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

DGX agent

arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsib

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Modeling Long-Term Memory and Temporal Attention Shifts for Video Salient Object Ranking with a New Benchmark

DGX agent

arXiv:2203.17257v2 Announce Type: replace Abstract: Salient Object Ranking (SOR) aims to estimate the relative saliency order among multiple salient objects. While SOR has been extensively studied in

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

DGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

DGX agent

arXiv:2608.03810v1 Announce Type: cross Abstract: Large language models routinely describe socially salient targets, including political figures, countries, religions, organizations, historical events

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Attention-Steered Vision-Language Models for Sign Language Translation

DGX agent

arXiv:2608.00235v1 Announce Type: new Abstract: Vision-language models (VLMs) have emerged as a powerful framework for multimodal video understanding. However, they remain limited in the sign language

researcharxiv-cs-cv
4 Aug 2026
Research

CopyCat: Improving Fine-Grained Subject Consistency in Subject-to-Image Models within Seconds

DGX agent

arXiv:2608.00674v1 Announce Type: new Abstract: Recent subject-to-image models have achieved impressive progress in personalized image generation, yet they still struggle to preserve fine-grained subj

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Disentangling Visuo-Tactile Foresight: Oracle-Guided Interface Discovery for World Action Models

DGX agent

arXiv:2608.00547v1 Announce Type: new Abstract: Contact-rich manipulation remains challenging because successful control depends on physical interaction cues that are often weakly observable from visi

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

DynImmune-BERT: Dynamic Immune Repertoire Modeling with Neural ODE Driven Continuous Transformers

DGX agent

arXiv:2607.17244v2 Announce Type: replace Abstract: Longitudinal T cell receptor repertoires contain signals of clonal expansion, contraction, disappearance, and reappearance after immune perturbation

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Evolutionary Curriculum Learning Improves Biological Sequence Modeling

DGX agent

arXiv:2608.00697v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) trained on multiple sequence alignments (MSAs) have emerged as powerful generative models for biological sequences, wi

safetyarxiv-cs-lg
4 Aug 2026
Safety

Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models

DGX agent

arXiv:2604.01622v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit

safetyarxiv-cs-cl
4 Aug 2026
Research

HAFI-VLM: A Frequency Perspective for Diagnosing and Enhancing Visual Perception in Vision-Language Models

DGX agent

arXiv:2608.02124v1 Announce Type: cross Abstract: Vision-language models (VLMs) remain unreliable when predictions require fine-grained visual evidence. We identify a previously overlooked cause: spec

researcharxiv-cs-cl
4 Aug 2026
Research

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models

DGX agent

arXiv:2607.15732v2 Announce Type: replace Abstract: Visual grounding with multimodal large language models is commonly formulated as autoregressive coordinate generation, where a model outputs boundin

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Kilobyte Models: Neural Networks as a Seed and a Quantized Latent

DGX agent

arXiv:2608.00860v1 Announce Type: new Abstract: The cost of storing and transmitting a trained neural network scales with its parameter count, a bottleneck for over-the-air updates, on-device librarie

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Leak It: A Probabilistic Approach to Training-Data Extraction from Black-Box Language Models

DGX agent

arXiv:2608.00144v1 Announce Type: cross Abstract: Membership inference (MIA) on language models is usually summarised by an aggregate ROC-AUC, but such evaluations are confounded: model-free blind bas

researcharxiv-cs-cl
4 Aug 2026
Model Releases

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

DGX agent

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

model-releasesarxiv-cs-cv
4 Aug 2026
Applications

Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP

DGX agent

arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score.

applicationsarxiv-cs-lg
4 Aug 2026
Model Releases

Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Stealthy Backdoor

DGX agent

arXiv:2608.00543v1 Announce Type: cross Abstract: Although semantic watermarking is considered a promising safeguard for images generated by Latent Diffusion Models (LDMs), the reliance of the waterma

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

RSVideo: Are Your Vision-Language Models Ready for Remote Sensing Videos?

DGX agent

arXiv:2608.02039v1 Announce Type: new Abstract: Remote-sensing videos enable real-time observation of changes in target attributes, short-term activities, and scene evolution. They record motion, acti

model-releasesarxiv-cs-cv
4 Aug 2026
Applications

SCALP: Semi-Supervised Statistical Shape Modeling from Imperfect 3D Photogrammetry via Landmark-Anchored Spectral Warp

DGX agent

arXiv:2608.00187v1 Announce Type: new Abstract: Correspondence-based statistical shape modeling (SSM) is vital for population-level morphometric analysis, but conventional pipelines assume clean, full

applicationsarxiv-cs-cv
4 Aug 2026
Safety

SelfWAM: A Self-Grounded Unified World Action Model for Fast Robot Control

DGX agent

arXiv:2608.00725v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future observations. However, conditioning future prediction on

safetyarxiv-cs-ro
4 Aug 2026
Safety

SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space

DGX agent

arXiv:2608.01397v1 Announce Type: cross Abstract: World Action Models (WAMs) couple action generation with prediction of future states. Their effectiveness depends on whether future dynamics are model

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

SPECTRA: Band-Routed Embedding and Stage-Wise LoRA for Cross-Sensor Fine-Tuning of Geospatial Foundation Models

DGX agent

arXiv:2608.01751v1 Announce Type: new Abstract: Geospatial foundation models (GeoFMs), pretrained on large-scale geospatial data such as Earth observation (EO), climate, and weather data, have shown p

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection

DGX agent

arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache th

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Understanding Synergistic Interactions among Pathology Foundation Models via Adaptive Fusion

DGX agent

arXiv:2608.01370v1 Announce Type: new Abstract: Pathology foundation models (PFMs) provide strong tile-level representations via self-supervised pre-training on large-scale pathology images. Yet, PFMs

researcharxiv-cs-cv
4 Aug 2026
Model Releases

CAER: Conflict-Aware Evidence Routing with Dual Prefix Experts for Multimodal Large Language Models

DGX agent

arXiv:2607.28991v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable capabilities in multimodal understanding and generation. However, when textual inp

model-releasesarxiv-cs-cv
3 Aug 2026
Research

Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation

DGX agent

arXiv:2510.12953v4 Announce Type: replace-cross Abstract: Recent medical vision-language models have shown promise on tasks such as VQA, report generation, and anomaly detection. However, most are ada

researcharxiv-cs-ai
3 Aug 2026
Safety

Faster but Different: Diagnosing and Controlling Content Drift in Accelerated Multimodal Diffusion Language Models

DGX agent

arXiv:2607.29079v1 Announce Type: new Abstract: Training-free acceleration makes diffusion-based multimodal large language models (dMLLMs) more deployable, but it may silently change generated content

safetyarxiv-cs-cl
3 Aug 2026
Safety

In-situ Autoguidance: Eliciting Self-Correction in Diffusion Models

DGX agent

arXiv:2510.17136v2 Announce Type: replace Abstract: The generation of high-quality, diverse, and prompt-aligned images is a central goal in image-generating diffusion models. The popular classifier-fr

safetyarxiv-cs-lg
3 Aug 2026
Local Ai

MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

DGX agent

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

local-aiarxiv-cs-ai
3 Aug 2026
← Previous
1…9899100101102…1030
Next →