AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

DGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

model-releasesarxiv-cs-ai
6 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training

DGX agent

arXiv:2608.04407v1 Announce Type: cross Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct appl

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments

DGX agent

arXiv:2608.04933v1 Announce Type: new Abstract: Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap

DGX agent

arXiv:2608.04160v1 Announce Type: new Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so th

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2608.04633v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight

DGX agent

arXiv:2608.04657v1 Announce Type: new Abstract: World action models (WAMs) built on video generation backbones are a rising recipe for robot learning, yet remain confined to tabletop manipulation. Mob

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements

DGX agent

arXiv:2506.02260v4 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental c

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding

DGX agent

arXiv:2608.04054v1 Announce Type: cross Abstract: Multimodal intent recognition requires understanding not only what textual, acoustic, and visual signals share, but also how they disagree. Such disag

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Monte Carlo Tree Search for Table-to-Multimodal Report Generation

DGX agent

arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical c

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MOON3.0: Reasoning-aware Multimodal Representation Learning for E-commerce Product Understanding

DGX agent

arXiv:2604.00513v3 Announce Type: replace-cross Abstract: With the rapid growth of e-commerce, exploring general representations rather than task-specific ones has attracted increasing attention. Alth

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Neural Diversity Regularizes Hallucinations in Language Models

DGX agent

arXiv:2510.20690v3 Announce Type: replace-cross Abstract: Language models continue to hallucinate despite increases in parameters, compute, and data. We propose neural diversity -- decorrelated parall

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

DGX agent

arXiv:2608.04397v1 Announce Type: new Abstract: We introduce NOLLI, a procedurally generated English-Korean puzzle benchmark designed to diagnose where Korean performance gaps arise. It comprises 15 p

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics

DGX agent

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization,

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

DGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

DGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

DGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Films

DGX agent

arXiv:2608.04224v1 Announce Type: new Abstract: Historical films suffer from co-occurring visual and audio degradations---blur, noise, flicker, hiss, clipping, and dropout---yet existing methods resto

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

On the Effectiveness of Adaptation Strategies for VLM-Based Federated Learning in Remote Sensing

DGX agent

arXiv:2608.04791v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative training of deep learning models across decentralized image archives without requiring data centralization

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP

DGX agent

arXiv:2505.19840v3 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve f

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

PADFormer: Pose-agnostic Anomaly Detection from Sparse View Images

DGX agent

arXiv:2608.04210v1 Announce Type: new Abstract: Pose-agnostic Anomaly Detection (PAD) remains challenging as anomalies can appear under arbitrary viewpoints, requiring methods to handle significant po

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Persistent Object Narratives for Token-Efficient Video Language Models

DGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

DGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Physics-informed reduced-order modelling with equivariant spectral submanifolds

DGX agent

arXiv:2608.04239v1 Announce Type: new Abstract: Spectral submanifold (SSM) reduction has emerged as a mathematically principled route to reliable nonlinear reduced-order models, capturing dynamics bey

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

DGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

DGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

DGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs

DGX agent

arXiv:2608.04501v1 Announce Type: new Abstract: Video surveillance in public safety, healthcare, and smart environments has made continuous human monitoring routine, raising real risks to personal ide

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Protoreasoning in Tiny Transformers

DGX agent

arXiv:2608.04980v1 Announce Type: cross Abstract: We show that tiny transformers can profitably employ a simple form of Chain of Thought, which we call protoreasoning, allowing us to study step-by-ste

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Radar4D-VLM: Proposal-Grounded Temporal 4D Radar Reasoning Across Frozen Language Models

DGX agent

arXiv:2608.04130v1 Announce Type: new Abstract: Vision-language models for autonomous driving primarily rely on cameras and LiDAR, leaving 4D radar largely unexplored as a standalone perceptual modali

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

DGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination

DGX agent

arXiv:2608.04385v1 Announce Type: new Abstract: Vision-Language Models (VLMs) often lose visual grounding during multi-step reasoning: as reasoning chains grow longer, later inference steps rely incre

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

DGX agent

arXiv:2608.04569v1 Announce Type: new Abstract: Hard prompt compression reduces long-context inference cost by independently scoring tokens, sentences, or chunks and retaining the highest-scoring unit

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

RepairFormer: Automated Repair of Structured Inputs Using Transformers

DGX agent

arXiv:2608.05060v1 Announce Type: cross Abstract: Structured input files such as JSON, DOT, OBJ, INI, S-expression, and TinyC are widely used in software systems, but small corruptions can cause parse

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

DGX agent

arXiv:2608.04783v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale ass

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

RESPClinBench: Benchmarking Multimodal Clinical Decision-Making and Longitudinal Disease Management in Respiratory Specialty Care

DGX agent

arXiv:2608.04514v1 Announce Type: new Abstract: Background: Respiratory specialty care requires multimodal interpretation, longitudinal risk assessment, guideline-concordant intervention, and whole-co

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

ResPlan: A Large-Scale Vector-Graph Dataset of 17,000 Residential Floor Plans

DGX agent

arXiv:2508.14006v2 Announce Type: replace Abstract: We introduce ResPlan, a dataset of 17,000 residential floor plans with vector geometry, room-connectivity graphs, and metric-scale coordinates. Each

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Retrieve in Time, Correct in Frequency

DGX agent

arXiv:2608.04527v1 Announce Type: new Abstract: Frozen vision-language-action (VLA) policies generate temporally extended action chunks, but long-horizon manipulation remains vulnerable to accumulated

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition

DGX agent

arXiv:2608.05115v1 Announce Type: cross Abstract: Can computer vision help make classrooms safer? In this pilot study, we investigate privacy-aware and computationally efficient classroom incident rec

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Robust and Personalized Federated Learning for Aircraft-Engine Prognostics under Benign and Adversarial Client Heterogeneity

DGX agent

arXiv:2608.04045v1 Announce Type: cross Abstract: Federated learning (FL) enables aircraft fleet operators to jointly train remaining-useful-life (RUL) models from engine sensor telemetry without shar

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Robust Control under Stationary Ambiguity

DGX agent

arXiv:2608.04832v1 Announce Type: new Abstract: Control policies optimized in simulation can perform poorly in the real system when the parameters x of the simulator are estimated from limited data bu

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Robustness Emerges Early in Training Dynamics, but Is Not Preserved

DGX agent

arXiv:2608.04442v1 Announce Type: cross Abstract: Robustness to natural corruptions remains a fundamental challenge for deep neural networks. In this paper, we identify a robustness fading phenomenon

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis

DGX agent

arXiv:2602.11506v4 Announce Type: replace-cross Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characteri

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Same Formulas, Different Semantics: Do Language Models Follow Modal Logic Specifications?

DGX agent

arXiv:2608.05097v1 Announce Type: new Abstract: Reasoning about necessity and possibility depends on assumptions about accessibility between worlds and about which objects exist at each one. The same

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

SciCode-Verified: How Benchmark Defects Underestimated the Scientific-Coding Ability of Language Models

DGX agent

arXiv:2608.04975v1 Announce Type: cross Abstract: SciCode is the standard measure of the scientific-coding ability of language models: research-level problems that demand both frontier scientific theo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

DGX agent

arXiv:2608.04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the iss

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

SEAR: Simple and Efficient Adaptation of Visual Geometric Transformers for Unpaired RGB+Thermal 3D Reconstruction

DGX agent

arXiv:2603.18774v2 Announce Type: replace Abstract: Foundational feed-forward visual geometry models enable accurate and efficient camera pose estimation and scene reconstruction by learning strong sc

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Semantic Frame Interpolation

DGX agent

arXiv:2507.05173v2 Announce Type: replace Abstract: Generating intermediate video content of varying lengths based on given first and last frames, along with text prompt information, offers significan

model-releasesarxiv-cs-cv
6 Aug 2026
← Previous
1…2324252627…357
Next →