AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

SAP SAPPHIRE 2026: Google Cloud unveils unified agentic vision and massive compute scaling

DGX agent

In today's hyper-connected market, an enterprise's most valuable asset — mission-critical data — often remains trapped in legacy silos. For years, leadership teams have navigated a data pipeline dilem

model-releasesgoogle-cloud-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SayNext-Bench: Why Do LLMs Struggle with Next-Utterance Anticipation?

DGX agent

arXiv:2602.00327v2 Announce Type: replace Abstract: We explore the use of large language models (LLMs) for next-utterance anticipation in human dialogue. Despite recent advances in LLMs demonstrating

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Scalable Gaussian process inference via neural feature maps

DGX agent

arXiv:2605.10285v1 Announce Type: cross Abstract: We present a theoretically grounded Gaussian process framework that leverages neural feature maps to construct expressive kernels. We show that the le

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SCALAR: A Neurosymbolic Framework for Automated Conjecture and Reasoning in Quantum Circuit Analysis

DGX agent

arXiv:2605.10327v1 Announce Type: cross Abstract: In this paper, we present SCALAR (Symbolic Conjecture and LLM-Assisted Reasoning), a neurosymbolic framework for automated conjecture generation in qu

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Scaling Limits of Long-Context Transformers

DGX agent

arXiv:2605.08505v1 Announce Type: cross Abstract: We study the long-context limit of softmax self-attention with a fixed query and a random context of n i.i.d. keys on the sphere, viewing the inverse

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Scaling the Memory of Balanced Adam

DGX agent

arXiv:2605.10119v1 Announce Type: new Abstract: Recent evidence suggests that Adam performs robustly when its momentum parameters are tied, eta_1=eta_2, reducing the optimizer to a single remaining pa

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Scaling Vision Models Does Not Consistently Improve Localisation-Based Explanation Quality

DGX agent

arXiv:2605.10142v1 Announce Type: cross Abstract: Artificial intelligence models are increasingly scaled to improve predictive accuracy, yet it remains unclear whether scale improves the quality of po

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs

DGX agent

arXiv:2509.02372v3 Announce Type: replace-cross Abstract: Large Language Models have become critical to modern software development, but their reliance on uncurated web-scale datasets for training int

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SciIntegrity-Bench: A Benchmark for Evaluating Academic Integrity in AI Scientist Systems

DGX agent

arXiv:2605.10246v1 Announce Type: new Abstract: AI scientist systems are increasingly deployed for autonomous research, yet their academic integrity has never been systematically evaluated. We introdu

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation

DGX agent

arXiv:2605.10187v1 Announce Type: new Abstract: Scientific reasoning is a key aspect of human intelligence, requiring the integration of multimodal inputs, domain expertise, and multi-step inference a

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness

DGX agent

arXiv:2603.14889v2 Announce Type: replace-cross Abstract: The rapid evolution of end-to-end spoken dialogue systems demands transcending mere textual semantics to incorporate paralinguistic nuances an

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

SeBA: Semi-supervised few-shot learning via Separated-at-Birth Alignment for tabular data

DGX agent

arXiv:2605.08519v1 Announce Type: new Abstract: Learning from scarce labeled data with a larger pool of unlabeled samples, known as semi-supervised few-shot learning (SS-FSL), remains critical for app

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Seed Hijacking of LLM Sampling and Quantum Random Number Defense

DGX agent

arXiv:2605.08313v1 Announce Type: cross Abstract: Large language models (LLMs) rely on deterministic pseudorandom number generators (PRNGs) for autoregressive sampling, creating a critical supply-chai

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning

DGX agent

arXiv:2605.09266v1 Announce Type: new Abstract: We introduce SeePhys Pro, a fine-grained modality transfer benchmark that studies whether models preserve the same reasoning capability when critical in

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind

DGX agent

arXiv:2603.26089v2 Announce Type: replace-cross Abstract: The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Selective LoRA for Visual Tokens and Attention Heads

DGX agent

arXiv:2512.19219v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is widely used for parameter-efficient fine-tuning, but its standard all-token, all-head design ignores the heterog

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SEMASIA: A Large-Scale Dataset of Semantically Structured Latent Representations

DGX agent

arXiv:2605.09485v1 Announce Type: new Abstract: Latent representations learned by neural networks often exhibit semantic structure, where concept similarity is reflected by geometric proximity in embe

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Semi-Supervised Neural Super-Resolution for Mesh-Based Simulations

DGX agent

arXiv:2605.09284v1 Announce Type: cross Abstract: Mesh-based simulations provide high-fidelity solutions to partial differential equations (PDEs), but achieving such accuracy typically requires fine m

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Sens-VisualNews: A Benchmark Dataset for Sensational Image Detection

DGX agent

arXiv:2605.10394v1 Announce Type: new Abstract: The detection of sensational content in media items can be a critical filtering mechanism for identifying check-worthy content and flagging potential di

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models

DGX agent

arXiv:2605.10576v1 Announce Type: cross Abstract: Low-level visual perception underpins reliable remote sensing (RS) image analysis, yet current image quality assessment (IQA) methods output uninterpr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought

DGX agent

arXiv:2605.09906v1 Announce Type: new Abstract: Audio and vision provide complementary evidence for audio-visual question answering, yet current audio-visual large language models may suffer from cros

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Sequential Causal Discovery with Noisy Language Model Priors

DGX agent

arXiv:2506.16234v2 Announce Type: replace Abstract: Causal discovery from observational data typically assumes access to complete data and availability of perfect domain experts. In practice, data oft

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Sequential Feature Selection for Efficient Landslide Segmentation from Multi-Spectral Data

DGX agent

arXiv:2605.09746v1 Announce Type: cross Abstract: Landslide detection from satellite imagery has advanced through deep learning, yet most models rely on large, highly correlated spectral-topographic i

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Set Prediction for Next-Day Active Fire Forecasting

DGX agent

arXiv:2605.10298v1 Announce Type: new Abstract: Accurate next-day active fire forecasts can support early warning, disaster response, forest risk assessment, and downstream estimation of fire-related

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

simpleposter: a simple baseline for product poster generation

DGX agent

arXiv:2605.08784v1 Announce Type: new Abstract: Product poster generation poses distinct challenges beyond general poster design, requiring both faithful preservation of product appearance and precise

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data

DGX agent

arXiv:2605.10498v1 Announce Type: cross Abstract: Long-tailed distributions in class-imbalanced data present a fundamental challenge for deep learning models, which tend to be biased toward majority c

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success

DGX agent

arXiv:2605.09070v1 Announce Type: cross Abstract: Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders

DGX agent

arXiv:2605.08731v1 Announce Type: cross Abstract: JPEG decode is routine ML infrastructure, but Python decoder choices are often justified by single-process, single-thread microbenchmarks. We audit th

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Sinkhorn Treatment Effects: A Causal Optimal Transport Measure

DGX agent

arXiv:2605.08485v1 Announce Type: cross Abstract: We introduce the Sinkhorn treatment effect, an entropic optimal transport measure of divergence between counterfactual distributions. Unlike classical

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching

DGX agent

arXiv:2605.08658v1 Announce Type: cross Abstract: SKETCHVERIFY is a within-tier cost-performance policy, not a universal accuracy improvement. The operational question: a practitioner stuck with a sma

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SLAM: Structural Linguistic Activation Marking for Language Models

DGX agent

arXiv:2605.05443v2 Announce Type: replace-cross Abstract: LLM watermarks must be detectable without compromising text quality, yet most existing schemes bias the next-token distribution and pay for de

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SleepWalk: A Three-Tier Benchmark for Stress-Testing Instruction-Guided Vision-Language Navigation

DGX agent

arXiv:2605.10376v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have advanced rapidly in multimodal perception and language understanding, yet it remains unclear whether they can reliabl

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SLIM: Sparse Latent Steering for Interpretable and Property-Directed LLM-Based Molecular Editing

DGX agent

arXiv:2605.10831v1 Announce Type: cross Abstract: Large language models possess strong chemical reasoning capabilities, making them effective molecular editors. However, property-relevant information

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

DGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SMIXAE: Towards Unsupervised Manifold Discovery in Language Models

DGX agent

arXiv:2605.09224v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have been used widely to decompose and interpret neural network activations, especially those of transformer language models.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy

DGX agent

arXiv:2605.09598v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently shown strong potential in soccer video understanding. However, given the high complexity of soccer videos du

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

DGX agent

arXiv:2605.09063v1 Announce Type: new Abstract: Following the recent achievement of gold-medal performance on the IMO by frontier LLMs, the community is searching for the next meaningful and challengi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Source or It Didn't Happen: A Multi-Agent Framework for Citation Hallucination Detection

DGX agent

arXiv:2605.08583v1 Announce Type: new Abstract: Large language models are increasingly used in scientific writing, yet they can fabricate citation-shaped references that appear plausible but fail bibl

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Sources: Anthropic is in talks to raise between 30B and 50B in a funding round that would value it at up to $950B (Mike Isaac/New York Times)

DGX agent

Mike Isaac / New York Times: Sources: Anthropic is in talks to raise between 30B and 50B in a funding round that would value it at up to 950B — The start-up, which recently released a powerful A.I. mo

model-releasestechmeme
12 May 2026
Model Releases

SpaceX is officially my fav art account now

DGX agent

SpaceX's official social media account is being praised for its engaging visual content and aesthetic presentation, earning recognition as a favorite among followers for its artistic approach to shari

model-releaseselon-musk--x
12 May 2026
Model Releases

Sparsity Moves Computation: How FFN Architecture Reshapes Attention in Small Transformers

DGX agent

arXiv:2605.09403v1 Announce Type: cross Abstract: Architectural choices inside the Transformer feedforward network (FFN) block do not merely affect the block itself; they reshape the computations lear

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?

DGX agent

arXiv:2602.03916v3 Announce Type: replace-cross Abstract: Spatial reasoning is a fundamental aspect of human cognition, yet it remains a major challenge for contemporary vision-language models (VLMs).

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

SPDEBench: An Extensive Benchmark for Learning Stochastic PDEs

DGX agent

arXiv:2505.18511v2 Announce Type: replace Abstract: Stochastic Partial Differential Equations (SPDEs) driven by random noise play a central role in modeling physical processes with rough spatio-tempor

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Spectral Characterization and Mitigation of Sequential Knowledge Editing Collapse

DGX agent

arXiv:2601.11042v2 Announce Type: replace-cross Abstract: Sequential knowledge editing in large language models often causes catastrophic collapse of the model's general abilities, especially for para

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SpectraLLM: Uncovering the Ability of LLMs for Molecular Structure Elucidation from Multi-Spectral Data

DGX agent

arXiv:2508.08441v3 Announce Type: replace-cross Abstract: Automated molecular structure elucidation remains challenging, as existing approaches often depend on pre-compiled databases or restrict thems

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

DGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Statistical Model Checking of the Keynes+Schumpeter Model: A Transient Sensitivity Analysis of a Macroeconomic ABM

DGX agent

arXiv:2605.10447v1 Announce Type: cross Abstract: Agent-based models (ABMs) are increasingly used in macroeconomics, but their analysis still often relies on ad hoc Monte Carlo campaigns with heteroge

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols

DGX agent

arXiv:2605.09618v1 Announce Type: new Abstract: When should a language model answer directly, sample and vote, or engage in multi-agent debate? Recent work shows voting often explains much of the gain

model-releasesarxiv-cs-cl
12 May 2026
← Previous
1…335336337338339…471
Next →