AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

It was a very good day https://www.spacex.com/launches/starship-flight-12

DGX agent

SpaceX's Starship Flight 12 achieved significant milestones in its development program, likely including successful booster catch, stage separation, or other critical test objectives. Elon Musk's cele

model-releaseselon-musk--x
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Learn anything with our new /lesson-generator skill

DGX agent

Learn anything with our new /lesson-generator skill Just released my new /lesson-generator skill. Use it with your agent to learn anything: - generate lessons/courses on any topic - include nano-banan

model-releasesdair-ai--x
23 May 2026
Model Releases

Lumberjack: Better Differentially Private Random Forests through Heavy Hitter Detection in Trees

DGX agent

arXiv:2605.22756v1 Announce Type: new Abstract: Random forests are widely used in fields involving sensitive tabular data, but existing approaches to enforcing differential privacy (DP) typically degr

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

MapTab: Are MLLMs Ready for Multi-Criteria Route Planning in Heterogeneous Graphs?

DGX agent

arXiv:2602.18600v3 Announce Type: replace Abstract: Systematic evaluation of Multimodal Large Language Models (MLLMs) is crucial for advancing Artificial General Intelligence (AGI). However, existing

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability

DGX agent

arXiv:2605.22168v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) map complex visual inputs to semantic spaces, but interpreting the cross-modal reasoning of VLMs currently relies on pos

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Memory-Efficient LLM Pretraining via Minimalist Optimizer Design

DGX agent

arXiv:2506.16659v3 Announce Type: replace Abstract: Training large language models (LLMs) relies on adaptive optimizers such as Adam, which introduce extra operations and require significantly more me

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Memory-R2: Fair Credit Assignment for Long-Horizon Memory-Augmented LLM Agents

DGX agent

arXiv:2605.21768v1 Announce Type: new Abstract: Memory-augmented LLM agents enable interactions that extend beyond finite context windows by storing, updating, and reusing information across sessions.

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization

DGX agent

arXiv:2605.21751v1 Announce Type: new Abstract: Text-to-optimization requires two separable capabilities: modeling -- choosing the right optimization structure -- and binding -- grounding every coeffi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents

DGX agent

arXiv:2602.13372v2 Announce Type: replace-cross Abstract: Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and ru…

DGX agent

oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and run a small experiment I would like leveraging this, and expla

model-releasesyohei-nakajima--x
23 May 2026
Model Releases

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

DGX agent

arXiv:2602.12506v3 Announce Type: replace Abstract: Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivatin

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

DGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Posterior Collapse as Automatic Spectral Pruning

DGX agent

arXiv:2605.22691v1 Announce Type: new Abstract: We show that posterior collapse in eta-VAEs implements automatic spectral pruning. A latent mode collapses if its contribution to reconstruction is belo

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Prior Knowledge-enhanced Spatio-temporal Epidemic Forecasting

DGX agent

arXiv:2602.22270v2 Announce Type: replace Abstract: Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

DGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

DGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

DGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

DGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

DGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

DGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching

DGX agent

arXiv:2605.22083v1 Announce Type: cross Abstract: While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issu

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

DGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

RWKV-7 G1g is here: the world's best pure RNN LLM, and a competitive LLM in general. Try https://huggingface.co/spaces/BlinkDL/RWKV-Gradio-2…

DGX agent

RWKV-7 G1g is here: the world's best pure RNN LLM, and a competitive LLM in general. Try https://huggingface.co/spaces/BlinkDL/RWKV-Gradio-2 for bsz16 7B inference. G1h in June 🙂 p.s. const 15000+tps

model-releasesjeremy-howard--x
23 May 2026
Model Releases

Scalable On-Policy Reinforcement Learning via Adaptive Batch Scaling

DGX agent

arXiv:2605.21557v1 Announce Type: cross Abstract: Conventional wisdom holds that large-batch training is fundamentally incompatible with Reinforcement Learning (RL) - beyond a modest threshold, increa

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

DGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

SeqLoRA: Bilevel Orthogonal Adaptation for Continual Multi-Concept Generation

DGX agent

arXiv:2605.22743v1 Announce Type: new Abstract: Parameter-efficient fine-tuning enables fast personalization of text-to-image diffusion models, but composing multiple custom concepts remains challengi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the …

DGX agent

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the frame is not the framer. Models climb whatever benchmark we

model-releasesitamar-friedman--x
23 May 2026
Model Releases

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

DGX agent

arXiv:2605.22142v1 Announce Type: new Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly mode

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

DGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

DGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Symbolic Density Estimation for Discrete Distributions

DGX agent

arXiv:2605.21813v1 Announce Type: new Abstract: Discrete probability laws underpin statistical modeling, yet the catalog of interpretable distributions has expanded only gradually through centuries of

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Tabular foundation models for robust calibration of near-infrared chemical sensing data

DGX agent

arXiv:2605.21544v1 Announce Type: new Abstract: Near-infrared spectroscopy is increasingly used as a rapid, non-destructive chemical sensing technology for the analysis of food, pharmaceutical, biolog

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally ever…

DGX agent

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally everything it says—a textbook example of propaganda: - I didn’t

model-releasesgary-marcus--x
23 May 2026
Model Releases

The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation

DGX agent

arXiv:2605.21856v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated impressive reasoning abilities across a wide range of tasks, but data contamination undermines the object

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Secretary Problem with a Stochastic Precursor

DGX agent

arXiv:2605.22653v1 Announce Type: cross Abstract: In learning-augmented online algorithms, predictions are usually valued for what they say: a value estimate, a solution, or an algorithmic recommendat

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Volterra signature

DGX agent

arXiv:2603.04525v2 Announce Type: replace-cross Abstract: Modern approaches for learning from non-Markovian time series, such as recurrent neural networks, neural controlled differential equations or

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, es…

DGX agent

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, especially in the last 3 version bumps of flicker and gerperte

model-releasesjeremy-howard--x
23 May 2026
Model Releases

this is even easier

DGX agent

this is even easier oh even better, since you can test the runtime without an LLM, you can post this into Claude: “look up http://activegraph.ai, install and run a small experiment I would like levera

model-releasesyohei-nakajima--x
23 May 2026
Model Releases

Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

DGX agent

Nemotron-Labs Diffusion Language Models represent NVIDIA's approach to achieving faster text generation through diffusion-based architectures, potentially offering significant speed improvements over

model-releaseshugging-face
23 May 2026
Model Releases

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

DGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

UNAD+: An Explainable Hybrid Framework for Unknown Network Attack Detection

DGX agent

arXiv:2605.22621v1 Announce Type: cross Abstract: The detection of previously unseen network attacks remains a major challenge for intrusion detection systems. Although supervised learning methods oft

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation

DGX agent

arXiv:2605.22368v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed for software engineering, constructing high-quality benchmarks is crucial for evaluating not j

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning

DGX agent

arXiv:2605.21606v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own rollouts using a privileged teacher, but its standard objective weights all generated tok

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Why SGD is not Brownian Motion: A New Perspective on Stochastic Dynamics

DGX agent

arXiv:2605.22644v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is commonly modeled as a Langevin process, assuming that minibatch noise acts as Brownian motion. However, this approx

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Wordle 1,798 5/6 ⬛⬛🟨⬛⬛ 🟨⬛⬛⬛🟨 ⬛🟩⬛🟩🟩 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This is a Wordle game result post from Anthropic's X (Twitter) account showing the solution found in 5 of 6 attempts, with the color-coded emoji grid indicating which letters were correct, misplaced,

model-releasesanthropic--x
23 May 2026
Model Releases

3D LULC classification using multispectral LiDAR and deep learning: current and prospective schemes

DGX agent

arXiv:2605.22328v1 Announce Type: new Abstract: Land Use Land Cover (LULC) classification is essential for national 3D mapping, geospatial analysis, and sustainable planning. Multispectral (MS) LiDAR

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

A Task-Agnostic Algebraic Integrity Metric for Event-Camera Streams Toward SOTIF-Compliant Perception using Pearson Correlation Coefficient

DGX agent

arXiv:2605.21500v1 Announce Type: cross Abstract: Event cameras have emerged as a high-bandwidth, low-latency sensing modality for safety-critical perception in automated driving systems (ADS), offeri

model-releasesarxiv-cs-cv
22 May 2026
← Previous
1…275276277278279…472
Next →