AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

SLMs as Multi-Agent Routers: A Progressive SFT and Reinforcement Learning Approach

DGX agent

arXiv:2608.00030v1 Announce Type: new Abstract: Specialised retrieval agents typically surface higher quality results than general-purpose search, but selecting the optimal agent for a given query rem

model-releasesarxiv-cs-cl
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SoM-1K: A Thousand-Problem Benchmark Dataset for Strength of Materials

DGX agent

arXiv:2509.21079v2 Announce Type: replace Abstract: Foundation models have shown remarkable capabilities in various domains, but their performance on complex, multimodal engineering problems remains l

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

SoniSpeech: A Large-Scale Open-Vocabulary Tri-Modal Dataset for Wearable Silent Speech Interfaces

DGX agent

arXiv:2608.00803v1 Announce Type: cross Abstract: Wearable silent speech interfaces (SSIs) are limited to small, closed vocabularies. Approaches achieving larger vocabularies require obtrusive hardwar

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

SPARC-Rad: A Multimodal Benchmark Dataset and Evaluation Pipeline for Spatial and Anatomical Reasoning in Radiology Vision-Language Models

DGX agent

arXiv:2608.00100v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly being evaluated for medical imaging, but many available benchmarks emphasize disease classification, repo

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SPARE: Structural Parameter-Free Affinity Regularization for Flow Matching

DGX agent

arXiv:2608.01990v1 Announce Type: new Abstract: Denoising diffusion transformers achieve strong generation quality but converge slowly during training. Regularizing their internal representations has

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SpatialQuery: Benchmarking Geometry-Grounded Multi-Instance Spatial Reasoning in Vision-Language Models

DGX agent

arXiv:2608.01709v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong semantic understanding but remain unreliable in metric spatial reasoning, particularly when queries require

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models

DGX agent

arXiv:2608.01899v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) perform well on commonsense reasoning tasks but struggle with visual spatial reasoning. Most existing solutions introduc

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

SPECTRA: Band-Routed Embedding and Stage-Wise LoRA for Cross-Sensor Fine-Tuning of Geospatial Foundation Models

DGX agent

arXiv:2608.01751v1 Announce Type: new Abstract: Geospatial foundation models (GeoFMs), pretrained on large-scale geospatial data such as Earth observation (EO), climate, and weather data, have shown p

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SphereVideo: Prototype-anchored Hyperspherical Boundary for Continual AI-generated Video Detection

DGX agent

arXiv:2608.01334v1 Announce Type: new Abstract: AI-generated video (AIGV) detection aims to distinguish real videos from AI-generated ones. In practice, detectors trained on existing data often fail t

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

ST-LoRA: Single Trajectory LoRA Ensemble for Uncertainty Aware Agricultural Segmentation

DGX agent

arXiv:2608.01530v1 Announce Type: new Abstract: Reliable decision-support in digital agriculture requires accurate predictions and well-calibrated uncertainty estimates, particularly for dense predict

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Staged Multi-Agent Training (SMAT) for Hip Exoskeletons: Metabolic and Biomechanical Validation of a Simulation-Trained Co-Adaptive Controller

DGX agent

arXiv:2608.00715v1 Announce Type: cross Abstract: Learning-based controllers can deliver exoskeleton assistance after training entirely in physics-based simulation, yet few controllers that address hu

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection

DGX agent

arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache th

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Structured Proxy Features for Multimodal NSCLC Survival Prediction from Pretreatment CT

DGX agent

arXiv:2608.00446v1 Announce Type: new Abstract: Lung cancer results in roughly 1.8 million fatalities annually worldwide, with non-small cell lung cancer (NSCLC) comprising the majority of cases. Desp

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Style Wins, Substance Loses: A Diagnosis of LLM-as-Judge in Idea Generation

DGX agent

arXiv:2608.01666v1 Announce Type: new Abstract: However, whether these judges truly evaluate the scientific substance of ideas or are influenced by superficial stylistic presentation remains an open q

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

SVGEval: A Vision-Grounded Framework for Perceptual-Quality Benchmarking and Evaluation in Text-to-SVG Generation

DGX agent

arXiv:2608.01977v1 Announce Type: new Abstract: Multimodal large models are increasingly used to generate scalable vector graphics (SVG), but reliable evaluation remains underexplored. Existing protoc

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SyncPlan: Long-Horizon LLM Coordination with Explicit Synchronization and Adaptive Correction

DGX agent

arXiv:2608.01652v1 Announce Type: new Abstract: LLM-based multi-agent coordination faces a fundamental trade-off between efficiency and adaptivity in dynamic environments. Existing approaches typicall

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

T^2exture: Sparsely Perturbed Thermal-to-Texture Imaging

DGX agent

arXiv:2608.02192v1 Announce Type: new Abstract: Thermal imaging remains effective under adverse illumination, yet passive long-wave infrared (LWIR) measurements often lack fine texture. Existing therm

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

TAB-PO: Preference Optimization with a Token-Level Adaptive Barrier for Token-Critical Structured Generation

DGX agent

arXiv:2603.00025v3 Announce Type: replace Abstract: Direct Preference Optimization (DPO) is effective for offline alignment but poorly matched to ontology-driven structured prediction, where preferred

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction

DGX agent

arXiv:2608.01400v1 Announce Type: new Abstract: Tabular foundation models, driven by in-context learning, have rapidly grown in quality and popularity. However, recent approaches with either cell-base

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Test-Time Curriculum for Open-Set AIGC Detection

DGX agent

arXiv:2608.00559v1 Announce Type: new Abstract: AI-generated image detectors deployed in open-world environments inevitably face distribution shifts as new and stronger generative models continue to e

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Tevatron Meets Megatron: Expert-Parallel LLM Reranker Training on an Academic Budget

DGX agent

arXiv:2608.00916v1 Announce Type: cross Abstract: Modern reranking recipes---billion-scale cross-encoders, mixture-of-experts (MoE) backbones, and distillation against strong teachers---have outpaced

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

TextNCA: Neural Cellular Automata for Language Modeling via Hierarchical Local Attention

DGX agent

arXiv:2608.02050v1 Announce Type: new Abstract: Can a strictly local, iterated, weight-shared computation primitive support language modelling, and which of those three properties actually drives the

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

The Condition-Number Barrier in Sparse Least Squares

DGX agent

arXiv:2608.02588v1 Announce Type: cross Abstract: In [AS21], Axiotis and Sviridenko conjectured that the linear dependence on the restricted condition number in sparse convex optimization cannot be im

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

The Learning Objective Governs Perceptual Narrowing: A Cross-Lingual, Layer-Wise, Ten-Seed Study of Self-Supervised Speech Encoders

DGX agent

arXiv:2608.00507v1 Announce Type: new Abstract: Perceptual narrowing---the developmental loss of non-native phoneme discrimination in the first year of life itep{werker1984}---is a canonical developme

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

The Push-Forward Transform for Continuous and Robust Comparison of Dynamic Shapes

DGX agent

arXiv:2608.02306v1 Announce Type: new Abstract: We introduce a mathematical framework for shape comparison based on mapping functions from the shape domain to a common reference domain. This Push-Forw

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

The Role of Disfluencies in Speech Translation

DGX agent

arXiv:2608.02138v1 Announce Type: new Abstract: Current speech translation systems, including SpeechLLMs, are trained on cleaned text and tend to strip disfluencies like filled pauses and false starts

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Towards Anomaly Detection on Relational Data

DGX agent

arXiv:2606.18621v2 Announce Type: replace Abstract: Relational databases are widely used for managing structured data in real-world systems. Detecting anomalies from such relational data is crucial fo

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Training Deep Morphological Neural Networks as Universal Approximators

DGX agent

arXiv:2505.09710v4 Announce Type: replace Abstract: We investigate deep morphological neural networks (DMNNs), studying how changes in algebraic structure affect the expressivity and trainability of d

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Training nGPT

DGX agent

arXiv:2608.01284v1 Announce Type: new Abstract: The normalized Transformer (nGPT) realizes hyperspherical representation learning by constraining model parameter vectors and activation vectors to the

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs

DGX agent

arXiv:2608.00640v1 Announce Type: new Abstract: Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

TrimMoE A communication aware and adaptive depth framework for distributed edge inference

DGX agent

arXiv:2608.00573v1 Announce Type: cross Abstract: Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission. The ex

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Trustworthiness Costs of Domain Adaptation in Small Language Models:A Cross-Architecture Empirical Study

DGX agent

arXiv:2608.00042v1 Announce Type: new Abstract: Domain adaptation of small language models (SLMs) has emerged as a practical strategy for deploying capable NLP systems in resource-constrained, high-st

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping

DGX agent

arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generali

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Two-Stage Bengali Sentiment Classification: Domain Adaptation Through Continual Learning and Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2608.01471v1 Announce Type: new Abstract: Understanding sentiment in low-resource languages remains a key challenge for Natural Language Processing (NLP), particularly when domain-specific data

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

UCBound-Net: Uncertainty-Guided Boundary-Aware Continual Learning for Domain-Incremental Ultrasound Segmentation

DGX agent

arXiv:2608.01518v1 Announce Type: new Abstract: Continual learning in clinical imaging faces a dual challenge: a model must assimilate knowledge from new anatomical domains while retaining representat

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Uncertainty Is Not Enough: Value-of-Information Routing for Mixtures of LoRA Experts

DGX agent

arXiv:2608.02528v1 Announce Type: new Abstract: Mixtures of low-rank adaptation experts increase parameter-efficient capacity by routing each input through a subset of adapters. Recent dynamic routers

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation

DGX agent

arXiv:2608.00915v1 Announce Type: new Abstract: Uplift modeling (conditional-average-treatment-effect estimation) drives personalized targeting, yet published uplift benchmarks frequently disagree on

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh

DGX agent

arXiv:2608.00094v1 Announce Type: new Abstract: Pretrained video diffusion models can act as renderers when the desired scene state is already specified by an animated mesh, a camera trajectory, and a

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

VR3D: View-Robust 3D Representation Learning for Aerial-Ground Person Re-Identification

DGX agent

arXiv:2608.02598v1 Announce Type: new Abstract: Aerial-ground person re-identification is a challenging task due to cross-platform viewpoint variations, which cause severe occlusion and geometric defo

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

What Carries the Signal in Pathology Foundation-Model Atlases? A Patient-Level Controlled Benchmark in Breast Cancer

DGX agent

arXiv:2608.00105v1 Announce Type: new Abstract: Pathology foundation models are reported to encode molecular programmes in tissue morphology, but the evidence is usually a cohort-wide ranked gene list

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

What Makes Position Zero Special? A Mechanistic Study of Position Zero Attention Sinks in LLMs

DGX agent

arXiv:2603.06591v2 Announce Type: replace-cross Abstract: Transformers frequently allocate disproportionate attention to specific tokens, a phenomenon known as attention sinks. Causal large language m

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs

DGX agent

arXiv:2608.00013v1 Announce Type: new Abstract: Choosing the right large language model (LLM) backbone is the most consequential decision when building a vision-language model (VLM), yet it remains fu

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

When Measurement Conventions Masquerade as Calibration Gains in Cardiac Digital Twins

DGX agent

arXiv:2608.01602v1 Announce Type: new Abstract: Cardiac digital twins convert clinical images into physiological measurements through observation operators, yet calibration studies often assume a fixe

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

When Retrieval Helps and Distracts: Evaluating Evidence-Generating LLMs for Biomedical Claim Verification

DGX agent

arXiv:2608.01409v1 Announce Type: new Abstract: Biomedical fact-checking systems must do more than predict whether a claim is supported, contradicted, or unaddressed: they should also produce evidence

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Who Belongs in the Eval Set? A Capability-Taxonomy-Driven Pipeline for Curating Regression Eval Sets in Agent-Extensibility Platforms

DGX agent

arXiv:2608.01004v1 Announce Type: new Abstract: Platform teams hosting agent-extensibility surfaces face a regression-economics paradox: every onboarding customer ships an evaluation set tuned to thei

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety

DGX agent

arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequenc

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Why Large Language Models Fail at Tabular Prediction

DGX agent

arXiv:2608.02412v1 Announce Type: new Abstract: Large language models (LLMs) have become the default tool for a remarkable range of tasks, yet they have had conspicuously little success at one of the

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…3536373839…357
Next →