AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms

DGX agent

arXiv:2605.11117v1 Announce Type: new Abstract: Scientific discovery can be modeled as a sequence of probabilistic decisions that map physical problems to numerical solutions. Recent agentic AI system

model-releasesarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Grid Games: The Power of Multiple Grids for Quantizing Large Language Models

DGX agent

arXiv:2605.12327v1 Announce Type: new Abstract: A major recent advance in quantization is given by microscaled 4-bit formats such as NVFP4 and MXFP4, quantizing values into small groups sharing a scal

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Grokking or Glitching? How Low-Precision Drives Slingshot Loss Spikes

DGX agent

arXiv:2605.06152v2 Announce Type: replace-cross Abstract: Deep neural networks exhibit periodic loss spikes during unregularized long-term training, a phenomenon known as the 'Slingshot Mechanism.' Ex

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

GRP: Goal-Reversed Prompting for Zero-Shot Evaluation with LLMs

DGX agent

arXiv:2503.06139v2 Announce Type: replace Abstract: Pairwise LLM-as-a-judge evaluation asks the judge to identify the better of two candidate answers. We study a one-line modification that asks for th

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

gym-invmgmt: An Open Benchmarking Framework for Inventory Management Methods

DGX agent

arXiv:2605.11355v1 Announce Type: new Abstract: Inventory-policy comparisons are often difficult to interpret because performance depends on the evaluation contract as much as on the policy itself. Di

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

HE-SNR: Uncovering Latent Logic via Entropy for Guiding Mid-Training on SWE-bench

DGX agent

arXiv:2601.20255v2 Announce Type: replace-cross Abstract: SWE-bench has emerged as the premier benchmark for evaluating Large Language Models on complex software engineering tasks. While these capabil

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model

DGX agent

arXiv:2605.11255v1 Announce Type: new Abstract: We present Hebatron, a Hebrew-specialized open-weight large language model built on the NVIDIA Nemotron-3 sparse Mixture-of-Experts architecture. Traini

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation

DGX agent

arXiv:2605.11208v1 Announce Type: new Abstract: Automated, clinician-grade assessment reports for surgical procedures could reduce documentation burden and provide objective feedback, yet remain chall

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

DGX agent

arXiv:2605.11061v1 Announce Type: new Abstract: The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Holder Policy Optimisation

DGX agent

arXiv:2605.12058v1 Announce Type: new Abstract: Group Relative Policy Optimisation (GRPO) enhances large language models by estimating advantages across a group of sampled trajectories. However, mappi

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Human-Grounded Multimodal Benchmark with 900K-Scale Aggregated Student Response Distributions from Japan's National Assessment of Academic Ability

DGX agent

arXiv:2605.11663v1 Announce Type: new Abstract: Authentic school examinations provide a high-validity test bed for evaluating multimodal large language models (MLLMs), yet benchmarks grounded in Japan

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference

DGX agent

arXiv:2505.13770v3 Announce Type: replace-cross Abstract: Reliable causal inference is essential for making decisions in high-stakes areas like medicine, economics, and public policy. However, it rema

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

DGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Investigating simple target-covariate relationships for Chronos-2 and TabPFN-TS

DGX agent

arXiv:2605.12200v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently achieved state-of-the-art performance, often outperforming supervised models in zero-shot settings.

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

KAN-CL: Per-Knot Importance Regularization for Continual Learning with Kolmogorov-Arnold Networks

DGX agent

arXiv:2605.12306v1 Announce Type: cross Abstract: Catastrophic forgetting remains the central obstacle in continual learning (CL): parameters shared across tasks interfere with one another, and existi

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Keeping Score: Efficiency Improvements in Neural Likelihood Surrogate Training via Score-Augmented Loss Functions

DGX agent

arXiv:2605.12118v1 Announce Type: cross Abstract: For stochastic process models, parameter inference is often severely bottlenecked by computationally expensive likelihood functions. Simulation-based

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

DGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Large-Small Model Collaboration for Farmland Semantic Change Detection

DGX agent

arXiv:2605.12282v1 Announce Type: new Abstract: Farmland Semantic Change Detection (SCD) is essential for cultivated land protection, yet existing benchmarks and models remain insufficient for fine-gr

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Latent Causal Void: Explicit Missing-Context Reconstruction for Misinformation Detection

DGX agent

arXiv:2605.12156v1 Announce Type: new Abstract: Automatic misinformation detection performs well when deception is visible in what an article explicitly states. However, some misinformation articles r

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR

DGX agent

arXiv:2605.11115v1 Announce Type: new Abstract: High Dynamic Range (HDR) generation remains challenging for generative models, which are largely limited to low dynamic range outputs. Recent diffusionb

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Learnable Multi-level Discrete Wavelet Transforms for 3D Gaussian Splatting Frequency Modulation

DGX agent

arXiv:2602.14199v2 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful approach for novel view synthesis. However, the number of Gaussian primitives often gro

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Learning Compact Boolean Networks

DGX agent

arXiv:2602.05830v2 Announce Type: replace-cross Abstract: Floating-point neural networks dominate modern machine learning but incur substantial inference costs, motivating emerging interest in Boolean

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Learning, Fast and Slow: Towards LLMs That Adapt Continually

DGX agent

arXiv:2605.12484v1 Announce Type: new Abstract: Large language models (LLMs) are trained for downstream tasks by updating their parameters (e.g., via RL). However, updating parameters forces them to a

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Learning Subspace-Preserving Sparse Attention Graphs from Heterogeneous Multiview Data

DGX agent

arXiv:2605.11881v1 Announce Type: new Abstract: The high-dimensional features extracted from large-scale unlabeled data via various pretrained models with diverse architectures are referred to as hete

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation

DGX agent

arXiv:2605.11739v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, existing studies largely attribute t

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Learning What Matters: Adaptive Information-Theoretic Objectives for Robot Exploration

DGX agent

arXiv:2605.12084v1 Announce Type: cross Abstract: Designing learnable information-theoretic objectives for robot exploration remains challenging. Such objectives aim to guide exploration toward data t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

LiBrA-Net: Lie-Algebraic Bilateral Affine Fields for Real-Time 4K Video Dehazing

DGX agent

arXiv:2605.11508v1 Announce Type: new Abstract: Currently, there is a gap in the field of ultra-high-definition (UHD) video dehazing due to the lack of a benchmark for evaluation. Furthermore, existin

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Lite3R: A Model-Agnostic Framework for Efficient Feed-Forward 3D Reconstruction

DGX agent

arXiv:2605.11354v1 Announce Type: new Abstract: Transformer-based 3D reconstruction has emerged as a powerful paradigm for recovering geometry and appearance from multi-view observations, offering str

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

LOFT: Low-Rank Orthogonal Fine-Tuning via Task-Aware Support Selection

DGX agent

arXiv:2605.11872v1 Announce Type: new Abstract: Orthogonal parameter-efficient fine-tuning (PEFT) adapts pretrained weights through structure-preserving multiplicative transformations, but existing me

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs

DGX agent

arXiv:2509.19207v2 Announce Type: replace Abstract: Contrastive vision-language models (VLMs) have made significant progress in binding visual and textual information, yet understanding long, composit

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues

DGX agent

arXiv:2605.12493v1 Announce Type: new Abstract: Long-term memory is crucial for agents in specialized web environments, where success depends on recalling interface affordances, state dynamics, workfl

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Measuring Five-Nines Reliability: Sample-Efficient LLM Evaluation in Saturated Benchmarks

DGX agent

arXiv:2605.11209v1 Announce Type: new Abstract: While existing benchmarks demonstrate the near-perfect performance of large language models (LLMs) on various tasks, this apparent saturation often obsc

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering

DGX agent

arXiv:2605.12361v1 Announce Type: new Abstract: Evaluating large language models (LLMs) in the biomedical domain requires benchmarks that can distinguish reasoning from pattern matching and remain dis

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

MEME: Multi-entity & Evolving Memory Evaluation

DGX agent

arXiv:2605.12477v1 Announce Type: cross Abstract: LLM-based agents increasingly operate in persistent environments where they must store, update, and reason over information across many sessions. Whil

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Mitigating Context-Memory Conflicts in LLMs through Dynamic Cognitive Reconciliation Decoding

DGX agent

arXiv:2605.12185v1 Announce Type: new Abstract: Large language models accumulate extensive parametric knowledge through pre-training. However, knowledge conflicts occur when outdated or incorrect para

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Modality-Inconsistent Continual Learning of Multimodal Large Language Models

DGX agent

arXiv:2412.13050v2 Announce Type: replace-cross Abstract: In this paper, we introduce Modality-Inconsistent Continual Learning (MICL), a new continual learning scenario for Multimodal Large Language M

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing

DGX agent

arXiv:2605.11836v1 Announce Type: cross Abstract: Lifelong Model Editing aims to continuously update evolving facts in Large Language Models while preserving unrelated knowledge and general capabiliti

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models

DGX agent

arXiv:2501.02955v2 Announce Type: replace Abstract: In recent years, vision language models (VLMs) have made significant advancements in video understanding. However, a crucial capability - fine-grain

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

MULTI: Disentangling Camera Lens, Sensor, View, and Domain for Novel Image Generation

DGX agent

arXiv:2605.12134v1 Announce Type: new Abstract: Recent text-to-image models produce high-quality images, yet text ambiguity hinders precise control when specific styles or objects are required. There

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Multi-Narrow Transformation as a Single-Model Ensemble: Boundary Conditions, Mechanisms, and Failure Modes

DGX agent

arXiv:2605.11530v1 Announce Type: new Abstract: Single-model ensembles (SMEs) have attracted attention as a way to approximate some of the benefits of deep ensembles within a single network. However,

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Multi-Task Representation Learning for Conservative Linear Bandits

DGX agent

arXiv:2605.12176v1 Announce Type: new Abstract: This paper presents the Constrained Multi-Task Representation Learning (CMTRL) framework for linear bandits. We consider T linear bandit tasks in a d di

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

MuonQ: Enhancing Low-Bit Muon Quantization via Directional Fidelity Optimization

DGX agent

arXiv:2605.11396v1 Announce Type: new Abstract: The Muon optimizer has emerged as a compelling alternative to Adam for training large language models, achieving remarkable computational savings throug

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Nautilus: From One Prompt to Plug-and-Play Robot Learning

DGX agent

arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

NavOL: Navigation Policy with Online Imitation Learning

DGX agent

arXiv:2605.11762v1 Announce Type: new Abstract: Learning robust navigation policies remains a core challenge in robotics. Offline imitation learning suffers from distribution shift and compounding err

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Neural ARFIMA model for forecasting BRIC exchange rates with long memory

DGX agent

arXiv:2509.06697v2 Announce Type: replace-cross Abstract: Accurate forecasting of exchange rates remains a persistent challenge, particularly for emerging economies such as Brazil, Russia, India, and

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

No More, No Less: Task Alignment in Terminal Agents

DGX agent

arXiv:2605.12233v1 Announce Type: new Abstract: Terminal agents are increasingly capable of executing complex, long-horizon tasks autonomously from a single user prompt. To do so, they must interpret

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Not How Many, But Which: Parameter Placement in Low-Rank Adaptation

DGX agent

arXiv:2605.12207v1 Announce Type: cross Abstract: We study the extit{parameter placement problem}: given a fixed budget of k trainable entries within the B matrix of a LoRA adapter (A frozen), does th

model-releasesarxiv-cs-cl
13 May 2026
← Previous
1…247248249250251…361
Next →