AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

SSMNBench: Diagnosing Image-based Cross-View Human-Object Understanding via Single-View Sufficiency and Multi-View Necessity

DGX agent

arXiv:2606.25634v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in single-image perception, yet their ability to reason about complex cross-view

model-releasesarxiv-cs-cv
25 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Stable-Shift: Biologically Structured Prediction of Transcriptional Responses to Unseen Gene Perturbations

DGX agent

arXiv:2606.24940v1 Announce Type: cross Abstract: Predicting transcriptional responses to genetic perturbations could reduce the experimental burden of functional genomics, but extrapolation to genes

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Stage-Aware and Roughness-Constrained Diffusion Policy for Multi-Stage Robotic Polishing

DGX agent

arXiv:2606.25754v1 Announce Type: new Abstract: Polishing is a critical finishing process in high-end manufacturing fields such as aerospace, where surface quality directly affects the service perform

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents

DGX agent

arXiv:2606.25632v1 Announce Type: new Abstract: Recent LLM role-playing systems build character agents from novels by extracting characters, scenes, and relations. Yet long-narrative role-playing suff

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

STEB: A Speech-to-Speech Translation Expressiveness Benchmark for Evaluating Beyond Translation Fidelity

DGX agent

arXiv:2606.25529v1 Announce Type: cross Abstract: Speech-to-speech translation (S2ST) should preserve not only lexical meaning, but also expressive attributes: emotion, scenario style (e.g., news repo

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Steering Vision-Language Models with Joint Sparse Autoencoders

DGX agent

arXiv:2606.25657v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have shown promise for analyzing language models, but applying them to vision-language models (VLMs) often yields representat

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Story Operators: Decomposing the Original o Sequel Transformation in Embedding Space

DGX agent

arXiv:2606.25379v1 Announce Type: new Abstract: I treat a book as a point in a sentence-embedding space and a literary transformation as an operation on points. Given an original novel and its sequel,

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SurgAtlas: A Large-Scale Surgical Video-Language Dataset with 2,391 Hours of Open and Minimally Invasive Surgery

DGX agent

arXiv:2606.25905v1 Announce Type: new Abstract: We introduce SurgAtlas, the largest surgical video-language dataset to date, comprising 15,291 videos (2,391 hours) spanning 18 surgical specialties and

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Swazure: Swarm Measurement of Pose for Flying Light Specks

DGX agent

arXiv:2606.25222v1 Announce Type: new Abstract: One may construct a 3D multimedia display using miniature drones configured with light sources, Flying Light Specks (FLSs). Swarms of FLSs localize to i

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning

DGX agent

arXiv:2507.16518v3 Announce Type: replace-cross Abstract: Recent advances in multimodal large language models (MLLMs) have shown impressive reasoning capabilities. However, further enhancing existing

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

TACO: Towards Task-Consistent Open-Vocabulary Adaptation in Video Recognition

DGX agent

arXiv:2606.25478v1 Announce Type: new Abstract: Adapting CLIP for open-vocabulary video recognition necessitates a delicate balance between newly acquired video knowledge and the pretrained generaliza

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

TacVerse: A Multi-Sensor Dataset and Benchmark for Cross-Sensor Vision-Based Tactile Perception

DGX agent

arXiv:2606.25877v1 Announce Type: new Abstract: Vision-based tactile sensors (VBTSs) enable robots to infer contact geometry and force-related cues by imaging deformation through an internal camera, y

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Tensorion: A Tensor-Aware Generalization of the Muon Optimizer

DGX agent

arXiv:2606.25975v1 Announce Type: cross Abstract: Common first-order optimizers, such as Adam, implicitly treat each parameter block as an unstructured vector, which disregards the multilinear weight

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

The 4/elta Bound: Designing Predictable LLM-Verifier Systems for Formal Method Guarantee

DGX agent

arXiv:2512.02080v3 Announce Type: replace-cross Abstract: The integration of Formal Verification tools with Large Language Models (LLMs) offers a path to scale software verification beyond manual work

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Three Buddhist Vocabularies: Computational Stylometry of the English Pali Canon across Sutta, Vinaya, and Abhidhamma

DGX agent

arXiv:2606.25372v1 Announce Type: new Abstract: We present a computational stylometric analysis of the Tipitaka across all three Pitakas in English translation, extending earlier work on the Sutta Pit

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

TokenMinds: Pretrained User Tokens and Embeddings for User Understanding in Large Recommender Systems

DGX agent

arXiv:2606.25147v1 Announce Type: cross Abstract: User modeling in industrial recommender systems typically produces dense embeddings, which suffer from representational constraints inherent to fixed-

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

TopoCast: A Topological Fidelity Framework for Evaluating Transformer-Based Time Series Forecasting

DGX agent

arXiv:2606.25439v1 Announce Type: new Abstract: Deep learning-based models have achieved state-of-the-art performance in Time Series Forecasting (TSF), yet their evaluation remains dominated by pointw

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Toten: A Knowledge-Based System For Structure-Preserving Representation Of Physical Quantities And Technical Notation In Brazilian Portuguese

DGX agent

arXiv:2606.19626v2 Announce Type: replace-cross Abstract: AI pipelines that reason quantitatively over technical text depend on input where physical quantities, numbers, units, and symbolic expression

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding

DGX agent

arXiv:2606.25160v1 Announce Type: cross Abstract: The rapid rise of Vision-Language Models (VLMs) in egocentric visual understanding has made low-latency inference in human-robot collaborative (HRC) t

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Training for the Model You Return: Improving Optimization for Iterate-Averaged Language Models

DGX agent

arXiv:2606.25086v1 Announce Type: new Abstract: Many modern Language Model (LM) pipelines return an averaged model, such as an exponential moving average of the training iterates, rather than the fina

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

TriViewBench: Controlled Complexity Scaling for Multi-View Structural Reasoning in MLLMs

DGX agent

arXiv:2606.26029v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) demonstrate strong performance on standard visual question answering benchmarks, yet their scalability under co

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Two-dimensional Hyperbolic RNN Neural Quantum State

DGX agent

arXiv:2606.25600v1 Announce Type: cross Abstract: In the first part of this work, we construct the first type of two-dimensional (2D) hyperbolic neural quantum state (NQS) in the form of the Lorentz 2

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Type Checking Project Haystack Grids using JSON Schema and Pydantic

DGX agent

arXiv:2606.24891v1 Announce Type: cross Abstract: Ontologies enable scalable energy services in buildings by supporting interoperability and automation. Project Haystack is a building ontology that is

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Models and GUI Grounding Datasets

DGX agent

arXiv:2606.25760v1 Announce Type: cross Abstract: Computer-use agents turn vision-language model (VLM) predictions into executable GUI clicks, so reliable uncertainty estimates are essential for rejec

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

USS: Unified Spatial-Semantic Prompts for Embodied Visual Tracking with Latent Dynamics Learning

DGX agent

arXiv:2606.25880v1 Announce Type: new Abstract: Embodied Visual Tracking (EVT) requires an agent to continuously follow a specified target while actively moving through dynamic environments. However,

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning

DGX agent

arXiv:2606.25319v1 Announce Type: new Abstract: Fine-grained visual reasoning requires multimodal large language models (MLLMs) to identify task-relevant visual evidence and ground their reasoning in

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

DGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks

DGX agent

arXiv:2606.25592v1 Announce Type: new Abstract: Recent advancements in Image-to-Video (I2V) generation have transformed input images from simple appearance references into interactive control interfac

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

What Actually Works for Spacecraft Fault-Tolerant Control: An Honest Settled-Gate Benchmark of Learned and Classical Methods

DGX agent

arXiv:2606.25374v1 Announce Type: new Abstract: Recent learned fault-tolerant-control (FTC) work reports high success on spacecraft actuator faults, but often in simulation, on narrow fault sets, and

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

What Do Language Priors Contribute to Darcy-Flow Inversion? A Mechanistic Audit

DGX agent

arXiv:2606.24967v1 Announce Type: new Abstract: In ill-posed inverse problems, the recovered solution depends as much on the prior as on the data, yet much of the engineering knowledge that could serv

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

What Does the Brain See? Multiview Neural Representations to Demystify the Brain-Visual Alignment

DGX agent

arXiv:2606.25718v1 Announce Type: new Abstract: Zero-shot visual decoding from electroencephalography (EEG) aims to infer visual semantics from non-invasive neural recordings, but remains challenging

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

DGX agent

arXiv:2606.25182v1 Announce Type: new Abstract: Jailbreak attacks reveal a persistent weakness in aligned Large Language Models: carefully crafted prompts can elicit policy-violating responses despite

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

When Multi-Sensor Fusion Fails to Generalize: Cattle Posture Classification Under Animal-Level and Temporal Distribution Shift

DGX agent

arXiv:2606.24986v1 Announce Type: new Abstract: Automated cattle posture-classification systems frequently report near-perfect accuracy, yet their robustness under realistic deployment conditions rema

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

WOLF-VLA: Whole-Body Humanoid Optimal Locomotion Framework for Vision-Language-Action Learning

DGX agent

arXiv:2606.25591v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently demonstrated strong generalization in robotic manipulation, yet their applicability to whole-body, con

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

3DCarGen: Scalable 3D Car Generation via 3D-consistent Multi-view Synthesis

DGX agent

arXiv:2606.24257v1 Announce Type: new Abstract: High-quality 3D vehicle assets are essential for autonomous driving simulation. Although multi-view diffusion-based paradigms enable controllable single

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy

DGX agent

arXiv:2606.24115v1 Announce Type: cross Abstract: Vision-language models (VLMs) are prone to hallucination, which remains a major barrier to their safe deployment in clinical practice. To date, most h

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR

DGX agent

arXiv:2604.00725v2 Announce Type: replace Abstract: End-to-end OCR for historical newspapers remains challenging, as models must handle long text sequences, degraded print quality, and complex layouts

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

A Hybrid Quantum-Classical Approach for Melt Pool Prediction in Laser Powder Bed Fusion

DGX agent

arXiv:2606.23719v1 Announce Type: cross Abstract: Laser powder bed fusion (LPBF) is a promising additive manufacturing technique that suffers from quality assurance concerns. Predicting melt pools fro

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

A Pairwise Human-Human Interaction Detection and Recognition Framework for Mobile Service Robots

DGX agent

arXiv:2602.22346v2 Announce Type: replace Abstract: Autonomous mobile service robots, such as lawnmowers or cleaning robots, operating in human-populated environments need to reason about human-human

model-releasesarxiv-cs-ro
24 Jun 2026
Model Releases

A Paninian Foundation for Indic Language Processing

DGX agent

arXiv:2606.24172v1 Announce Type: cross Abstract: More than a billion people communicate in Indic languages, yet the natural language processing infrastructure serving them remains fragmented and unde

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Physics-Informed Fourier-Wavelet Transformer for Multiscale Computational Fluid Dynamics Surrogate Modeling

DGX agent

arXiv:2606.24696v1 Announce Type: cross Abstract: Physics-informed surrogate models can accelerate computational fluid dynamics simulations. However, many existing methods reproduce global flow patter

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial

DGX agent

arXiv:2606.24510v1 Announce Type: new Abstract: Rare diseases affect millions of individuals worldwide, yet timely diagnosis remains a major public health challenge due to scarcity of specialized clin

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Synthetic Reliability-Aware PINN Benchmark for Offshore Wind Turbine Support-Structure Monitoring with Bayesian Inverse Identification

DGX agent

arXiv:2606.24176v1 Announce Type: new Abstract: Reliable structural health monitoring (SHM) of offshore wind turbine (OWT) support structures requires fast state estimation from sparse measurements. R

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ABACUS: Adapting Unified Foundation Model for Bridging Image Count Understanding and Generation

DGX agent

arXiv:2606.23835v1 Announce Type: new Abstract: ABACUS is a unified vision-language model that handles object counting, crowd counting, referring-expression counting, and count-faithful image generati

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Acquisition state behaves as a structured, measurable variable governing lung-nodule AI: kernel-driven measurement instability and noise-driven detection fragility, invisible to DICOM metadata

DGX agent

arXiv:2606.12824v2 Announce Type: replace-cross Abstract: AI governance for medical imaging is formalizing: the 2026 ACR-SIIM Practice Parameter recommends local acceptance testing and ongoing drift m

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

DGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
← Previous
1…120121122123124…361
Next →