AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents

DGX agent

arXiv:2602.13372v2 Announce Type: replace-cross Abstract: Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection

model-releasesarxiv-cs-lg
23 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and ru…

DGX agent

oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and run a small experiment I would like leveraging this, and expla

model-releasesyohei-nakajima--x
23 May 2026
Model Releases

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

DGX agent

arXiv:2602.12506v3 Announce Type: replace Abstract: Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivatin

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

DGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Posterior Collapse as Automatic Spectral Pruning

DGX agent

arXiv:2605.22691v1 Announce Type: new Abstract: We show that posterior collapse in eta-VAEs implements automatic spectral pruning. A latent mode collapses if its contribution to reconstruction is belo

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Prior Knowledge-enhanced Spatio-temporal Epidemic Forecasting

DGX agent

arXiv:2602.22270v2 Announce Type: replace Abstract: Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

DGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

DGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

DGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

DGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

DGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

DGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching

DGX agent

arXiv:2605.22083v1 Announce Type: cross Abstract: While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issu

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

DGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

RWKV-7 G1g is here: the world's best pure RNN LLM, and a competitive LLM in general. Try https://huggingface.co/spaces/BlinkDL/RWKV-Gradio-2…

DGX agent

RWKV-7 G1g is here: the world's best pure RNN LLM, and a competitive LLM in general. Try https://huggingface.co/spaces/BlinkDL/RWKV-Gradio-2 for bsz16 7B inference. G1h in June 🙂 p.s. const 15000+tps

model-releasesjeremy-howard--x
23 May 2026
Model Releases

Scalable On-Policy Reinforcement Learning via Adaptive Batch Scaling

DGX agent

arXiv:2605.21557v1 Announce Type: cross Abstract: Conventional wisdom holds that large-batch training is fundamentally incompatible with Reinforcement Learning (RL) - beyond a modest threshold, increa

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

DGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

SeqLoRA: Bilevel Orthogonal Adaptation for Continual Multi-Concept Generation

DGX agent

arXiv:2605.22743v1 Announce Type: new Abstract: Parameter-efficient fine-tuning enables fast personalization of text-to-image diffusion models, but composing multiple custom concepts remains challengi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the …

DGX agent

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the frame is not the framer. Models climb whatever benchmark we

model-releasesitamar-friedman--x
23 May 2026
Model Releases

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

DGX agent

arXiv:2605.22142v1 Announce Type: new Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly mode

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

DGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

DGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Symbolic Density Estimation for Discrete Distributions

DGX agent

arXiv:2605.21813v1 Announce Type: new Abstract: Discrete probability laws underpin statistical modeling, yet the catalog of interpretable distributions has expanded only gradually through centuries of

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Tabular foundation models for robust calibration of near-infrared chemical sensing data

DGX agent

arXiv:2605.21544v1 Announce Type: new Abstract: Near-infrared spectroscopy is increasingly used as a rapid, non-destructive chemical sensing technology for the analysis of food, pharmaceutical, biolog

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally ever…

DGX agent

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally everything it says—a textbook example of propaganda: - I didn’t

model-releasesgary-marcus--x
23 May 2026
Model Releases

The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation

DGX agent

arXiv:2605.21856v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated impressive reasoning abilities across a wide range of tasks, but data contamination undermines the object

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Secretary Problem with a Stochastic Precursor

DGX agent

arXiv:2605.22653v1 Announce Type: cross Abstract: In learning-augmented online algorithms, predictions are usually valued for what they say: a value estimate, a solution, or an algorithmic recommendat

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Volterra signature

DGX agent

arXiv:2603.04525v2 Announce Type: replace-cross Abstract: Modern approaches for learning from non-Markovian time series, such as recurrent neural networks, neural controlled differential equations or

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, es…

DGX agent

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, especially in the last 3 version bumps of flicker and gerperte

model-releasesjeremy-howard--x
23 May 2026
Model Releases

this is even easier

DGX agent

this is even easier oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and run a small experiment I would like levera

model-releasesyohei-nakajima--x
23 May 2026
Model Releases

Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

DGX agent

Nemotron-Labs Diffusion Language Models represent NVIDIA's approach to achieving faster text generation through diffusion-based architectures, potentially offering significant speed improvements over

model-releaseshugging-face
23 May 2026
Model Releases

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

DGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

UNAD+: An Explainable Hybrid Framework for Unknown Network Attack Detection

DGX agent

arXiv:2605.22621v1 Announce Type: cross Abstract: The detection of previously unseen network attacks remains a major challenge for intrusion detection systems. Although supervised learning methods oft

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation

DGX agent

arXiv:2605.22368v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed for software engineering, constructing high-quality benchmarks is crucial for evaluating not j

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning

DGX agent

arXiv:2605.21606v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own rollouts using a privileged teacher, but its standard objective weights all generated tok

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Why SGD is not Brownian Motion: A New Perspective on Stochastic Dynamics

DGX agent

arXiv:2605.22644v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is commonly modeled as a Langevin process, assuming that minibatch noise acts as Brownian motion. However, this approx

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Wordle 1,798 5/6 ⬛⬛🟨⬛⬛ 🟨⬛⬛⬛🟨 ⬛🟩⬛🟩🟩 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This is a Wordle game result post from Anthropic's X (Twitter) account showing the solution found in 5 of 6 attempts, with the color-coded emoji grid indicating which letters were correct, misplaced,

model-releasesanthropic--x
23 May 2026
Model Releases

3D LULC classification using multispectral LiDAR and deep learning: current and prospective schemes

DGX agent

arXiv:2605.22328v1 Announce Type: new Abstract: Land Use Land Cover (LULC) classification is essential for national 3D mapping, geospatial analysis, and sustainable planning. Multispectral (MS) LiDAR

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

A Task-Agnostic Algebraic Integrity Metric for Event-Camera Streams Toward SOTIF-Compliant Perception using Pearson Correlation Coefficient

DGX agent

arXiv:2605.21500v1 Announce Type: cross Abstract: Event cameras have emerged as a high-bandwidth, low-latency sensing modality for safety-critical perception in automated driving systems (ADS), offeri

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AesFormer: Transform Everyday Photos into Beautiful Memories

DGX agent

arXiv:2605.22126v1 Announce Type: new Abstract: In everyday photography, aesthetically appealing moments are often captured with structural flaws (e.g., composition, camera viewpoint, or pose) that ex

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows

DGX agent

arXiv:2605.20425v1 Announce Type: new Abstract: Designing multi-agent workflows is especially difficult in open-ended scientific settings where tasks lack curated training sets, reliable scalar evalua

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

DGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding

DGX agent

arXiv:2605.22034v1 Announce Type: new Abstract: Visual grounding, the task of localizing objects described by natural-language expressions, is a foundational capability for agricultural AI systems, en

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment

DGX agent

arXiv:2512.20538v2 Announce Type: replace Abstract: Single-view RGB model-based object pose estimation methods achieve strong generalization but are fundamentally limited by depth ambiguity, clutter,

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AMEL: Accumulated Message Effects on LLM Judgments

DGX agent

arXiv:2605.22714v1 Announce Type: cross Abstract: Large language models are routinely used as automated evaluators: to review code, moderate content, or score outputs, often with many items passing th

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Anthropic says Claude Mythos Preview has been used to find more than 10,000 high- or critical-severity vulnerabilities since the launch of Project Glasswing (Anthropic)

DGX agent

Anthropic: Anthropic says Claude Mythos Preview has been used to find more than 10,000 high- or critical-severity vulnerabilities since the launch of Project Glasswing — Last month, we launched Projec

model-releasestechmeme
22 May 2026
Model Releases

ArabDiscrim: A Decade-Long Arabic Facebook Corpus on Racism and Discrimination

DGX agent

arXiv:2605.22081v1 Announce Type: new Abstract: We present ArabDiscrim, a decade-long lexical resource and corpus of 293K public Arabic Facebook posts (2014--2024) discussing racism and discrimination

model-releasesarxiv-cs-cl
22 May 2026
← Previous
1…275276277278279…472
Next →