AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Markovian Circuit Tracing for Transformer State Dynamic

DGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

model-releasesarxiv-cs-lg
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs

DGX agent

arXiv:2605.20410v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases.

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MedCRP-CL: Continual Medical Image Segmentation via Bayesian Nonparametric Semantic Modality Discovery

DGX agent

arXiv:2605.20297v1 Announce Type: new Abstract: Medical image segmentation faces a fundamental challenge in continual learning: data arrives sequentially from heterogeneous sources, yet effective cont

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

DGX agent

arXiv:2605.20197v1 Announce Type: new Abstract: Medical concept extraction from electronic health records underpins many downstream applications, yet remains challenging because medically meaningful c

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MemGym: a Long-Horizon Memory Environment for LLM Agents

DGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Memory-Efficient Partitioned DNN Inference on Resource-Constrained Android Crowds

DGX agent

arXiv:2605.20723v1 Announce Type: new Abstract: Deploying large deep neural networks on memory-constrained mobile devices is a central challenge in edge ML. While compression, pruning, and quantizatio

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Memory Grafting: Scaling Language Model Pre-training via Offline Conditional Memory

DGX agent

arXiv:2605.20948v1 Announce Type: new Abstract: Scaling conditional memory offers a promising way to increase language-model capacity, but existing methods such as Engram learn large memory tables fro

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset

DGX agent

arXiv:2605.21272v1 Announce Type: new Abstract: Training large text-to-image models requires high-quality, curated datasets with diverse content and detailed captions. Yet the cost and complexity of c

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

DGX agent

arXiv:2605.20729v1 Announce Type: new Abstract: Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Multimodal Optimal Transport for Training-free Temporal Segmentation in Surgical Robotics

DGX agent

arXiv:2602.24138v2 Announce Type: replace Abstract: Automated recognition of surgical phases and steps is a fundamental capability for intraoperative decision support, workflow automation, and skill a

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Neural Negative Binomial Regression for Weekly Seismicity Forecasting: Per-Cell Dispersion Estimation and Tail Risk Assessment

DGX agent

arXiv:2605.21437v1 Announce Type: cross Abstract: Standard approaches to forecasting the weekly number of earthquakes on a spatial grid rely on the Poisson distribution with a single global dispersion

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding

DGX agent

arXiv:2605.20525v1 Announce Type: cross Abstract: We present NeuroQA, a large-scale benchmark for visual question answering in 3D brain magnetic resonance imaging (MRI), with 56,953 QA pairs from 12,9

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Nonparametric Learning and Earning with One-Point Feedback under Nonstationarity

DGX agent

arXiv:2605.21263v1 Announce Type: new Abstract: Firms increasingly rely on dynamic pricing to respond to evolving customer demand, yet in many applications they observe only the revenue generated by a

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists

DGX agent

arXiv:2605.20668v1 Announce Type: new Abstract: With the advancement of AI capabilities, AI reviewers are beginning to be deployed in scientific peer review, yet their capability and credibility remai

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

On the Suboptimality of GP-UCB under Polynomial Effective Optimism

DGX agent

arXiv:2312.01386v2 Announce Type: replace Abstract: Gaussian process upper confidence bound (GP-UCB) is widely used for sequential optimization of expensive black-box functions. Although many upper bo

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Optimization Hyper-parameter Laws for Large Language Models

DGX agent

arXiv:2409.04777v4 Announce Type: replace Abstract: Large Language Models have driven significant AI advancements, yet their training is resource-intensive and highly sensitive to hyper-parameter sele

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Parameters as Experts: Adapting Vision Models with Dynamic Parameter Routing

DGX agent

arXiv:2602.06862v2 Announce Type: replace Abstract: Adapting pre-trained vision models using parameter-efficient fine-tuning (PEFT) remains challenging, as it aims to achieve performance comparable to

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

PGC: Peak-Guided Calibration for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2605.21207v1 Announce Type: new Abstract: The rapid evolution of generative AI, from GANs to modern diffusion models, has resulted in increasingly subtle discriminative clues. These fine-grained

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models

DGX agent

arXiv:2605.20873v1 Announce Type: cross Abstract: Planning is a fundamental capability for large language models (LLMs) because such complex tasks require models to coordinate goals, constraints, reso

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Point Cloud Sequence Encoding for Material-conditioned Graph Network Simulators

DGX agent

arXiv:2605.20978v1 Announce Type: new Abstract: Graph Network Simulators (GNSs) have emerged as powerful surrogates for complex physics-based simulation, offering inherent differentiability and orders

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Post-Hoc Understanding of Metaphor Processing in Decoder-Only Language Models via Conditional Scale Entropy

DGX agent

arXiv:2605.21391v1 Announce Type: new Abstract: Metaphor requires a language model to resolve a token whose contextual meaning diverges from its basic literal sense. Understanding how transformer mode

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Preserve, Reveal, Expand: Faithful 4D Video Editing with Region-Aware Conditioning

DGX agent

arXiv:2605.20961v1 Announce Type: new Abstract: Existing 4D-driven video diffusion models primarily target plausible generation, but faithful 4D editing requires preserving source-observed regions whi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Pseudo-Formalization for Automatic Proof Verification

DGX agent

arXiv:2605.20531v1 Announce Type: cross Abstract: Reliable verification of proofs remains a bottleneck for training and evaluating AI systems on hard mathematical reasoning. Fully formal proofs, in la

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Q-DiT4SR: Exploration of Detail-Preserving Diffusion Transformer Quantization for Real-World Image Super-Resolution

DGX agent

arXiv:2602.01273v4 Announce Type: replace Abstract: Recently, Diffusion Transformers (DiTs) have emerged in Real-World Image Super-Resolution (Real-ISR) to generate high-quality textures, yet their he

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Quantifying Hyperparameter Transfer and the Importance of Embedding Layer Learning Rate

DGX agent

arXiv:2605.21486v1 Announce Type: new Abstract: Hyperparameter transfer allows extrapolating optimal optimization hyperparameters from small to large scales, making it critical for training large lang

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Quantum reservoir computing in Jaynes-Cummings models: Nonlinear memory and time-series prediction

DGX agent

arXiv:2510.00171v2 Announce Type: replace-cross Abstract: We investigate quantum reservoir computing (QRC) using a hybrid qubit-boson system described by the Jaynes-Cummings (JC) Hamiltonian and its d

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments

DGX agent

arXiv:2512.09447v2 Announce Type: replace-cross Abstract: Structurally repetitive environments produce visually plausible but aliased LiDAR loop candidates that can destabilize pose-graph optimization

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs

DGX agent

arXiv:2605.20584v1 Announce Type: new Abstract: Mobile app marketplaces require developers to disclose standardized content rating descriptors (CRDs) to inform users about potentially sensitive or res

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RadProPoser: Probabilistic Radar Tensor Human Pose Estimation That Knows Its Limits

DGX agent

arXiv:2508.03578v2 Announce Type: replace Abstract: Radar-based human pose estimation enables privacy-preserving motion tracking for ambient intelligence, yet the noisy nature of radar sensing makes u

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution

DGX agent

arXiv:2605.21195v1 Announce Type: new Abstract: Discrete autoregressive (AR) text-to-image (T2I) models pair a VQ tokenizer with an AR policy, and current post-training pipelines optimize only the pol

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RCGDet3D: Rethinking 4D Radar-Camera Fusion-based 3D Object Detection with Enhanced Radar Feature Encoding

DGX agent

arXiv:2605.21112v1 Announce Type: new Abstract: 4D automotive radar is indispensable for autonomous driving due to its low cost and robustness, yet its point cloud sparsity challenges 3D object detect

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Refining and Reusing Annotation Guidelines for LLM Annotation

DGX agent

arXiv:2605.20809v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable performance on zero-shot annotation tasks, they often struggle with the specialized convention

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Reinforcing Human Behavior Simulation via Verbal Feedback

DGX agent

arXiv:2605.20506v1 Announce Type: cross Abstract: Humans learn social norms and behaviors from verbal feedback (e.g., a parent saying 'that was rude' or a friend explaining 'here's why that hurt'). Ye

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing

DGX agent

arXiv:2605.20262v1 Announce Type: new Abstract: We study selective refusal editing as a three-way control problem: induce non-refusal on designated edit prompts while preserving benign behavior and ha

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Resolving Long-Tail Ambiguity in Unsupervised 3D Point Cloud Segmentation with Language Priors

DGX agent

arXiv:2605.20737v1 Announce Type: new Abstract: Existing approaches for unsupervised 3D point cloud segmentation predominantly rely on a purely visual similarity-based learning-by-clustering paradigm,

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Retrieval-Augmented Long-Context Translation for Cultural Image Captioning: Gators submission for AmericasNLP 2026 shared task

DGX agent

arXiv:2605.20626v1 Announce Type: new Abstract: We present the University of Florida Gators submission to the AmericasNLP 2026 shared task on cultural image captioning for Indigenous languages. Our tw

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Robust Personalized Recommendation under Hidden Confounding in MNAR

DGX agent

arXiv:2605.21066v1 Announce Type: new Abstract: Recommender systems often rely on observational user--item interaction data, which is prone to selection bias due to users' selective interactions with

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

roto 2.0: The Robot Tactile Olympiad

DGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Runtime-Certified Bounded-Error Quantized Attention

DGX agent

arXiv:2605.20868v1 Announce Type: new Abstract: KV cache quantization reduces the memory cost of long-context LLM inference, but introduces approximation error that is typically validated only empiric

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Safety-Critical Control for Smoothed Implicit Contact Dynamics

DGX agent

arXiv:2605.21138v1 Announce Type: new Abstract: Smoothed implicit contact dynamics enables gradient-based planning and control for contact-rich tasks without predefined mode sequences. However, safety

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

Seeing Through Fog: Towards Fog-Invariant Action Recognition

DGX agent

arXiv:2605.20645v1 Announce Type: new Abstract: Foggy conditions are commonly encountered in real-world applications; however, existing action recognition approaches typically assume favorable weather

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Semiparametric Efficient Bilevel Gradient Estimation

DGX agent

arXiv:2605.21341v1 Announce Type: cross Abstract: Functional bilevel methods estimate a lower-level function and plug it into a hypergradient, but this plug-in gradient can retain first-order bias whe

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Sequential Data Augmentation for Generative Recommendation

DGX agent

arXiv:2509.13648v3 Announce Type: replace Abstract: Generative recommendation plays a crucial role in personalized systems, predicting users' future interactions from their historical behavior sequenc

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society

DGX agent

arXiv:2605.20510v1 Announce Type: new Abstract: Urban heat exposure is becoming an increasingly critical challenge due to the intensifying urban heat island effect. Fine-grained shade patterns, especi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

ShapeBench: A Scalable Benchmark and Diagnostic Suite for Standardized Evaluation in Aerodynamic Shape Optimization

DGX agent

arXiv:2605.20763v1 Announce Type: new Abstract: Rapid progress in aerodynamic shape optimization (ASO) has outpaced currently-available standardized evaluation frameworks. Fair comparison requires a u

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass

DGX agent

arXiv:2602.06358v2 Announce Type: replace Abstract: We propose SHINE (Scalable Hyper In-context NEtwork), a scalable hypernetwork that can map diverse meaningful contexts into high-quality LoRA adapte

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

SMoA: Spectrum Modulation Adapter for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2605.21147v1 Announce Type: cross Abstract: As the number of model parameters increases, parameter-efficient fine-tuning (PEFT) has become the go-to choice for tailoring pre-trained large langua

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

DGX agent

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

model-releasesarxiv-cs-lg
21 May 2026
← Previous
1…214215216217218…361
Next →