AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

DGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

model-releasesarxiv-cs-lg
10 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

NOVA: Symbolic Regression Discovery of Interpretable Car-Following and Lane-Change Models with Driver Heterogeneity

DGX agent

arXiv:2606.10583v1 Announce Type: cross Abstract: We present NOVA, an autonomous symbolic regression framework that identifies interpretable car-following and lane-change structures from raw trajector

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

DGX agent

Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma to run even faster across NVIDIA GeForce R

model-releasesnvidia-blog
10 Jun 2026
Model Releases

OncoTraj: a public benchmark for longitudinal resistance prediction in EGFR-mutant non-small-cell lung cancer on osimertinib

DGX agent

arXiv:2606.11144v1 Announce Type: new Abstract: Resistance to first-line osimertinib in EGFR-mutant non-small-cell lung cancer (NSCLC) is the canonical example of predictable clonal evolution under th

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

DGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

DGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

DGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Optimal Post-Training Quantization Scales and Where to Find Them

DGX agent

arXiv:2606.10890v1 Announce Type: cross Abstract: Post-training quantization (PTQ) compresses large language models by mapping weights to low-bit representations. The scaling factor that defines the q

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Optimization-based Online Conformal Prediction for Multi-step Forecasting

DGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Overcoming Rank Collapse in Feedback Alignment

DGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning

DGX agent

arXiv:2606.11152v1 Announce Type: new Abstract: Multimodal large language models can write code to produce complex programs as well as use programs to do 3D modeling, which opens up a new avenue for 3

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Parallel Causal Associative Fields: Gated Sparse Memory for Long-Context Language Modeling

DGX agent

arXiv:2606.10435v1 Announce Type: cross Abstract: Transformers achieve strong language modeling performance by providing direct token-to-token communication paths, but causal self-attention scales qua

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Parametric Knowledge is Not All You Need: Toward Honest Large Language Models via Retrieval of Pretraining Data

DGX agent

arXiv:2601.21218v2 Announce Type: replace Abstract: Large language models (LLMs) are highly capable of answering questions, but they are often unaware of their own knowledge boundary, i.e., knowing wh

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

PhantomBench: Benchmarking the Non-existential Threat of Language Models

DGX agent

arXiv:2606.11105v1 Announce Type: cross Abstract: Hallucinations, where language models (LMs) generate factually ungrounded responses, pose serious risks, as users tend to blindly rely on them. This i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Piper: A Programmable Distributed Training System

DGX agent

arXiv:2606.11169v1 Announce Type: cross Abstract: Large-scale model training increasingly relies on composing multiple parallelism strategies, such as data, pipeline, and expert parallelism, together

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

POPSICLE: Benchmark Datasets for Segmentation and Localization in CryoET

DGX agent

arXiv:2606.10255v1 Announce Type: cross Abstract: Cryo-electron tomography (cryoET) has emerged as a powerful tool in structural and cellular biology by enabling direct visualization of macromolecular

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

PRC-linked influence operations are targeting AI debates in the US

DGX agent

OpenAI reported that People's Republic of China (PRC)-linked actors are conducting coordinated influence operations aimed at shaping artificial intelligence policy debates in the United States. These

model-releasesopenai
10 Jun 2026
Model Releases

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

DGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ProbeLLM: Automating Principled Diagnosis of LLM Failures

DGX agent

arXiv:2602.12966v2 Announce Type: replace Abstract: Understanding how and why large language models (LLMs) fail is becoming a central challenge as models rapidly evolve and static evaluations fall beh

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Quality Is Not a Safety Proxy Under Quantization

DGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Quantifying Uncertainty in AI Visibility: A Statistical Framework for Generative Search Measurement

DGX agent

arXiv:2603.08924v2 Announce Type: replace-cross Abstract: AI-powered answer engines are inherently non-deterministic: identical queries submitted at different times can produce different responses and

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

DGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Quoting Jeremy Howard

DGX agent

Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de

model-releasessimon-willison
10 Jun 2026
Model Releases

Rank Collapse, Fixed Points, and the Renormalization Group Structure of MLP Residual Networks

DGX agent

arXiv:2606.10324v1 Announce Type: new Abstract: The analogy between deep neural network forward passes and renormalization group (RG) flows has been repeatedly noted in the literature, but existing tr

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

RAT: Reference-Augmented Training for ASV Anti-Spoofing

DGX agent

arXiv:2606.10908v1 Announce Type: cross Abstract: We introduce a spoofing countermeasure architecture conditioned on speaker-reference recordings, but observe that it converges to a solution that effe

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

READER: Robust Evidence-based Authorship Decoding via Extracted Representations

DGX agent

arXiv:2606.10794v1 Announce Type: new Abstract: As agentic applications increasingly route user tasks through official and third-party LLM APIs, provenance becomes an operational question: which model

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

DGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info ne…

DGX agent

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info needed to train a frontier model, down to hparams. I also thou

model-releasesyann-lecun--x
10 Jun 2026
Model Releases

RealMath-Eval: Why SOTA Judges Struggle with Real Human Reasoning

DGX agent

arXiv:2606.10254v1 Announce Type: new Abstract: While Large Language Models (LLMs) have achieved near-perfect performance in solving high-school mathematics, their ability to evaluate the diverse reas

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

DGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

DGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Recoverable but Not Stationary:Local Linear Structures in Weights and Activations

DGX agent

arXiv:2606.10929v1 Announce Type: cross Abstract: Task vectors, LoRA, activation steering, and random search around pretrained weights all suggest that learned behaviour can be controlled by linear di

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

DGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

DGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

DGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

DGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

DGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

DGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

DGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

DGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

DGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

DGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

SCOPE: Sequential Causal Optimization of Process Interventions

DGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPE: Coalition-Aware Expert Pruning for Sparse Mixture-of-Experts LLMs

DGX agent

arXiv:2606.09886v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) large language models achieve strong quality with low per-token compute, yet their deployment is often limited by the

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

DGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

DGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

DGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

DGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

model-releasesarxiv-cs-ai
10 Jun 2026
← Previous
1…191192193194195…472
Next →