AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics

DGX agent

arXiv:2512.13660v3 Announce Type: replace-cross Abstract: Spatial tracing, as a fundamental embodied interaction ability for robots, is inherently challenging as it requires multi-step metric-grounded

model-releasesarxiv-cs-cv
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

DGX agent

arXiv:2606.30560v1 Announce Type: cross Abstract: Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge r

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

DGX agent

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Translating Natural Language to Strategic Temporal Specifications via LLMs

DGX agent

arXiv:2606.30441v1 Announce Type: cross Abstract: A rigorous formalization of system requirements is a fundamental prerequisite for the verification of Multi-Agent Systems (MAS). However, writing corr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

DGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

DGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

DGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Tumor-aware augmentation with task-guided attention analysis improves rectal cancer segmentation from magnetic resonance images

DGX agent

arXiv:2605.05522v2 Announce Type: replace-cross Abstract: Although self-supervised pretraining is expected to learn broadly transferable representations, its effectiveness across imaging modalities su

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Two kinds of robustness are not the same: disentangling fault tolerance and low-SNR robustness in multi-domain event detection on real data

DGX agent

arXiv:2606.29339v1 Announce Type: cross Abstract: Reliable event detection underpins induced-seismicity monitoring for Carbon dioxide Capture and Storage (CCS) and geothermal operations, distributed a

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

UniCA: Bi-directional Cross-Attention with Positive Similarity Loss for Robust Multi-Modal Retrieval

DGX agent

arXiv:2606.28350v1 Announce Type: cross Abstract: Multi-modal retrieval has become increasingly critical for handling the growing volume of integrated visual-textual data in real-world applications, b

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

DGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Unlocking the Visual Record of Materials Science: A Large-Scale Multimodal Dataset from Scientific Literature

DGX agent

arXiv:2606.29667v1 Announce Type: cross Abstract: The materials science literature encodes decades of experimental knowledge in figures, yet this visual record remains locked away and inaccessible to

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

UrbanCDNet: Appearance-Robust and Boundary-Aware Bitemporal Change Detection for Korean Urban Building Monitoring

DGX agent

arXiv:2606.29781v1 Announce Type: new Abstract: Urban building change detection from bi-temporal aerial imagery is important for redevelopment monitoring, infrastructure management, and unauthorized-c

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

DGX agent

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

DGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

DGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

DGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

What Drives the Inlier-Memorization Effect? A Theory of Outlier Detection via Early Training Dynamics

DGX agent

arXiv:2606.29791v1 Announce Type: cross Abstract: Outlier detection (OD) aims to identify anomalous instances by learning the underlying structure of normal data (inliers), and is particularly challen

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When AI Reviews Its Own Code: Recursive Self-Training Collapse in Code LLMs

DGX agent

arXiv:2606.28438v1 Announce Type: cross Abstract: Recursive self-training can degrade neural generative models when generated data is reused without fresh human data or external quality control. We st

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries

DGX agent

arXiv:2606.28332v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical and health-related questions, yet their safety in high-risk medical scenarios remains p

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling

DGX agent

arXiv:2606.28661v1 Announce Type: cross Abstract: People overthink; language models over-sample, and the extra effort can talk both into a worse answer. Reasoning systems answer a hard question by sam

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

DGX agent

arXiv:2606.30383v1 Announce Type: new Abstract: A rapidly growing class of LLM agents is multi-party: the agent acts for a principal (who briefs it, sends follow-ups, and receives results) while also

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Why Do We Need Warm-up? A Theoretical Perspective

DGX agent

arXiv:2510.03164v2 Announce Type: replace Abstract: Learning rate warm-up -- increasing the learning rate at the beginning of training -- has become a ubiquitous heuristic in modern deep learning, yet

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression

DGX agent

arXiv:2606.29712v1 Announce Type: new Abstract: Large language models achieve high reasoning performance via explicit chain-of-thought and reinforcement learning, but require long output sequences and

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation

DGX agent

arXiv:2412.15529v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) synergizes the retrieval of pertinent data with the generative capabilities of Large Language Models (LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

XYZ-IBD: Benchmarking Robust 6D Object Pose Estimation under Real-World Industrial Complexity

DGX agent

arXiv:2506.00599v3 Announce Type: replace Abstract: While current 6D pose estimation benchmarks have reached near-saturation on household objects, they often fail to capture the stochastic and optical

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

DGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Zero-Gated Language-conditioned Human Motion Prediction

DGX agent

arXiv:2606.29208v1 Announce Type: new Abstract: Pose histories provide the core kinematic evidence for 3D human motion prediction, but they lack explicit high-level semantic guidance. This paper intro

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Zero-Shot Depth from Defocus

DGX agent

arXiv:2603.26658v2 Announce Type: replace Abstract: Depth from Defocus (DfD) is the task of estimating a dense metric depth map from a focus stack. Unlike previous works overfitting to a certain datas

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

DGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

A Multi-Attribute Latent Space for Visual Analysis of Watches

DGX agent

arXiv:2606.27897v1 Announce Type: new Abstract: We present a design rationale, embedding model, and interactive visual-analysis system for exploring large wristwatch collections through heterogeneous

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

DGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

A Unified Framework for Vision Transformers Equivariant to Discrete Subgroups of O(2)

DGX agent

arXiv:2606.27864v1 Announce Type: new Abstract: Vision transformers have become a dominant architecture for visual recognition. However, standard models do not explicitly encode the planar symmetries

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Accelerating Attention with Basis Decomposition

DGX agent

arXiv:2510.01718v2 Announce Type: replace Abstract: Attention is a core operation in large language models (LLMs). We present BD Attention (BDA), a lossless algorithmic reformulation of attention. BDA

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Adaptive Momentum and Nonlinear Damping for Neural Network Training

DGX agent

arXiv:2602.00334v2 Announce Type: replace Abstract: Momentum Stochastic Gradient Descent (mSGD) relies on a fixed momentum coefficient shared across all parameters, failing to account for the heteroge

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Advancing Speaker-Based Vocal Effort Classification with WavLM and Data Augmentation in Naturalistic Non-Calibrated Speech Recordings

DGX agent

arXiv:2606.27543v1 Announce Type: cross Abstract: The variations in vocal effort range (e.g. whisper, soft, neutral, loud, shout) alter production and speech acoustics, reducing intelligibility and li

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Agentic Hardware Design as Repository-Level Code Evolution

DGX agent

arXiv:2606.28279v1 Announce Type: cross Abstract: We present HORIZON, a self-evolving agent framework that treats hardware design as repository-level code evolution. A Markdown harness is compiled int

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

AirGroundBench: Probing Spatial Intelligence in Multimodal Large Models under Heterogeneous Multi-View Embodied Collaboration

DGX agent

arXiv:2606.28049v1 Announce Type: new Abstract: In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometric

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Aloe-Vision: Robust Vision-Language Models for Healthcare

DGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

An Empirical Analysis of Factual Errors in Human-Written Text and its Application

DGX agent

arXiv:2606.27959v1 Announce Type: new Abstract: Factual Error Detection (FED), which is the task of identifying factually incorrect spans in a given text, has long been recognized as an important rese

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Applicability of memorization indicators for early spotting of overfitting while recalibrating sEMG-decoders on low sample sizes

DGX agent

arXiv:2606.27855v1 Announce Type: cross Abstract: Deep learning models for surface electromyography (sEMG) can benefit substantially from subject-specific (re-)calibration, since no sufficiently large

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Aurora: A Leverage-Aware Spectral Optimizer

DGX agent

arXiv:2606.27715v1 Announce Type: new Abstract: We show that for tall matrix parameters, like projection matrices in the MLP layers, the Muon update can have row norms that are arbitrarily non-uniform

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Benchmarking Multi-Modal Graph-based Social Media Popularity Prediction

DGX agent

arXiv:2606.27539v1 Announce Type: cross Abstract: Social media popularity prediction aims to forecast the future reach or influence of online content from early-stage observations. Accurate prediction

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Benchmarking on Tasks That Matter: Dataset Selection for Preserving Model Rankings

DGX agent

arXiv:2606.27997v1 Announce Type: new Abstract: Benchmarks of machine learning models often include many datasets, making evaluation expensive. For efficiency, it is preferable to perform evaluations

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Bridging Ab Initio Symmetries and Global Nuclear Masses with Interpretable Neural Networks

DGX agent

arXiv:2606.28287v1 Announce Type: cross Abstract: Ab initio modeling has established Wigner's SU(4) and Elliott's SU(3) as dominant symmetries of the nuclear force in light and intermediate-mass nucle

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence Cloud-Native Simulation Infrastructure for Embodied Intelligence Training, Evaluation, and Data Collection

DGX agent

arXiv:2606.27962v1 Announce Type: new Abstract: This paper presents a cloud-native simulation infrastructure framework for embodied intelligence that supports large-scale training, standardized evalua

model-releasesarxiv-cs-ro
29 Jun 2026
← Previous
1…110111112113114…361
Next →