AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
10 Jun 2026

mlr3mbo: Bayesian Optimization in R

Model ReleasesDGX agent

arXiv:2603.29730v2 Announce Type: replace-cross Abstract: We present mlr3mbo, a modular toolbox for Bayesian optimization in R. mlr3mbo supports single- and multi-objective optimization, multi-point p

MMClima: A Framework for Multimodal Climate Science Data and Evaluation

Model ReleasesDGX agent

arXiv:2606.10194v1 Announce Type: cross Abstract: Climate change research increasingly requires AI systems that reason across text, dynamic visual content, and scientific figures, yet existing climate

Modeling Complex Behaviors: Multi-Personality Composition and Dynamic Switching in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.11074v1 Announce Type: cross Abstract: With the widespread deployment of Multimodal Large Language Models (MLLMs) in social interaction, understanding and controlling their behavior under c


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MoE Enhanced Federated Learning for Spatiotemporal Prediction

Model ReleasesDGX agent

arXiv:2606.10499v1 Announce Type: cross Abstract: Traffic prediction is fundamental to intelligent transportation systems and urban computing, yet many cities continue to suffer from traffic data scar

Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football

Model ReleasesDGX agent

arXiv:2606.11120v1 Announce Type: new Abstract: We recast pass evaluation in football (soccer) as a Monte Carlo Tree Search (MCTS)-like evaluation problem whose components mostly exist in the literatu

Moonshine: An Autonomous Mathematical Research Agent Centered on Conjecture Generation

Model ReleasesDGX agent

arXiv:2606.10806v1 Announce Type: new Abstract: Moonshine is an autonomous agent whose central objective is to generate mathematical conjectures. Its core capability is to extract structure from class

MV-Actor: Aligning Multi-View Semantics and Spatial Awareness for Bimanual Manipulation

Model ReleasesDGX agent

arXiv:2606.10899v1 Announce Type: new Abstract: Robotic manipulation has been widely applied in industrial scenarios. Compared with single-arm manipulation, bimanual manipulation is equipped with mult

N-GRPO: Embedding-Level Neighbor Mixing for Enhanced Policy Optimization

Model ReleasesDGX agent

arXiv:2606.10768v1 Announce Type: cross Abstract: The success of Large Language Models in mathematical reasoning relies heavily on the generation of diverse and valid solution paths during the rollout

nCMD: Benign-Anchored Feature Selection for Imbalanced Network Intrusion Detection

Model ReleasesDGX agent

arXiv:2606.09934v1 Announce Type: new Abstract: Feature selection is critical for network intrusion detection systems (NIDS) operating under high-dimensional, highly imbalanced traffic, as found in op

Next Forcing: Causal World Modeling with Multi-Chunk Prediction

Model ReleasesDGX agent

arXiv:2606.11187v1 Announce Type: new Abstract: Autoregressive video generation has emerged as a powerful paradigm for World Action Models (WAMs). However, existing approaches suffer from slow trainin

Non-Parametric Structural Priors for Geometry Theorem Prediction

Model ReleasesDGX agent

arXiv:2603.04852v2 Announce Type: replace Abstract: Multi-step theorem prediction is a central challenge in geometry problem solving. Existing neural-symbolic approaches rely heavily on supervised par

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

Model ReleasesDGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

NOVA: Symbolic Regression Discovery of Interpretable Car-Following and Lane-Change Models with Driver Heterogeneity

Model ReleasesDGX agent

arXiv:2606.10583v1 Announce Type: cross Abstract: We present NOVA, an autonomous symbolic regression framework that identifies interpretable car-following and lane-change structures from raw trajector

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

Model ReleasesDGX agent

Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma to run even faster across NVIDIA GeForce R

OncoTraj: a public benchmark for longitudinal resistance prediction in EGFR-mutant non-small-cell lung cancer on osimertinib

Model ReleasesDGX agent

arXiv:2606.11144v1 Announce Type: new Abstract: Resistance to first-line osimertinib in EGFR-mutant non-small-cell lung cancer (NSCLC) is the canonical example of predictable clonal evolution under th

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

Model ReleasesDGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

Model ReleasesDGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

Model ReleasesDGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

Optimal Post-Training Quantization Scales and Where to Find Them

Model ReleasesDGX agent

arXiv:2606.10890v1 Announce Type: cross Abstract: Post-training quantization (PTQ) compresses large language models by mapping weights to low-bit representations. The scaling factor that defines the q

Optimization-based Online Conformal Prediction for Multi-step Forecasting

Model ReleasesDGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

Overcoming Rank Collapse in Feedback Alignment

Model ReleasesDGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning

Model ReleasesDGX agent

arXiv:2606.11152v1 Announce Type: new Abstract: Multimodal large language models can write code to produce complex programs as well as use programs to do 3D modeling, which opens up a new avenue for 3

Parallel Causal Associative Fields: Gated Sparse Memory for Long-Context Language Modeling

Model ReleasesDGX agent

arXiv:2606.10435v1 Announce Type: cross Abstract: Transformers achieve strong language modeling performance by providing direct token-to-token communication paths, but causal self-attention scales qua

Parametric Knowledge is Not All You Need: Toward Honest Large Language Models via Retrieval of Pretraining Data

Model ReleasesDGX agent

arXiv:2601.21218v2 Announce Type: replace Abstract: Large language models (LLMs) are highly capable of answering questions, but they are often unaware of their own knowledge boundary, i.e., knowing wh

PhantomBench: Benchmarking the Non-existential Threat of Language Models

Model ReleasesDGX agent

arXiv:2606.11105v1 Announce Type: cross Abstract: Hallucinations, where language models (LMs) generate factually ungrounded responses, pose serious risks, as users tend to blindly rely on them. This i

Piper: A Programmable Distributed Training System

Model ReleasesDGX agent

arXiv:2606.11169v1 Announce Type: cross Abstract: Large-scale model training increasingly relies on composing multiple parallelism strategies, such as data, pipeline, and expert parallelism, together

POPSICLE: Benchmark Datasets for Segmentation and Localization in CryoET

Model ReleasesDGX agent

arXiv:2606.10255v1 Announce Type: cross Abstract: Cryo-electron tomography (cryoET) has emerged as a powerful tool in structural and cellular biology by enabling direct visualization of macromolecular

PRC-linked influence operations are targeting AI debates in the US

Model ReleasesDGX agent

OpenAI reported that People's Republic of China (PRC)-linked actors are conducting coordinated influence operations aimed at shaping artificial intelligence policy debates in the United States. These

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

Model ReleasesDGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

ProbeLLM: Automating Principled Diagnosis of LLM Failures

Model ReleasesDGX agent

arXiv:2602.12966v2 Announce Type: replace Abstract: Understanding how and why large language models (LLMs) fail is becoming a central challenge as models rapidly evolve and static evaluations fall beh

Quality Is Not a Safety Proxy Under Quantization

Model ReleasesDGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

Quantifying Uncertainty in AI Visibility: A Statistical Framework for Generative Search Measurement

Model ReleasesDGX agent

arXiv:2603.08924v2 Announce Type: replace-cross Abstract: AI-powered answer engines are inherently non-deterministic: identical queries submitted at different times can produce different responses and

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

Model ReleasesDGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

Quoting Jeremy Howard

Model ReleasesDGX agent

Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de

Rank Collapse, Fixed Points, and the Renormalization Group Structure of MLP Residual Networks

Model ReleasesDGX agent

arXiv:2606.10324v1 Announce Type: new Abstract: The analogy between deep neural network forward passes and renormalization group (RG) flows has been repeatedly noted in the literature, but existing tr

RAT: Reference-Augmented Training for ASV Anti-Spoofing

Model ReleasesDGX agent

arXiv:2606.10908v1 Announce Type: cross Abstract: We introduce a spoofing countermeasure architecture conditioned on speaker-reference recordings, but observe that it converges to a solution that effe

READER: Robust Evidence-based Authorship Decoding via Extracted Representations

Model ReleasesDGX agent

arXiv:2606.10794v1 Announce Type: new Abstract: As agentic applications increasingly route user tasks through official and third-party LLM APIs, provenance becomes an operational question: which model

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

Model ReleasesDGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info ne…

Model ReleasesDGX agent

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info needed to train a frontier model, down to hparams. I also thou

RealMath-Eval: Why SOTA Judges Struggle with Real Human Reasoning

Model ReleasesDGX agent

arXiv:2606.10254v1 Announce Type: new Abstract: While Large Language Models (LLMs) have achieved near-perfect performance in solving high-school mathematics, their ability to evaluate the diverse reas

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

Model ReleasesDGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

Model ReleasesDGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

Recoverable but Not Stationary:Local Linear Structures in Weights and Activations

Model ReleasesDGX agent

arXiv:2606.10929v1 Announce Type: cross Abstract: Task vectors, LoRA, activation steering, and random search around pretrained weights all suggest that learned behaviour can be controlled by linear di

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

Model ReleasesDGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

Model ReleasesDGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

Model ReleasesDGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

Model ReleasesDGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

Model ReleasesDGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

Model ReleasesDGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

SCOPE: Sequential Causal Optimization of Process Interventions

Model ReleasesDGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

SHAPE: Coalition-Aware Expert Pruning for Sparse Mixture-of-Experts LLMs

Model ReleasesDGX agent

arXiv:2606.09886v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) large language models achieve strong quality with low per-token compute, yet their deployment is often limited by the

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

Model ReleasesDGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

Model ReleasesDGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

Model ReleasesDGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

Model ReleasesDGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

Small Data, Big Noise: Adversarial Training for Robust Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10610v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become essential for adapting foundation models to downstream NLP tasks. However, current PEFT methods often

← Previous
1…152153154155156…377
Next →