AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Sequential Feature Selection for Efficient Landslide Segmentation from Multi-Spectral Data

DGX agent

arXiv:2605.09746v1 Announce Type: cross Abstract: Landslide detection from satellite imagery has advanced through deep learning, yet most models rely on large, highly correlated spectral-topographic i

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Set Prediction for Next-Day Active Fire Forecasting

DGX agent

arXiv:2605.10298v1 Announce Type: new Abstract: Accurate next-day active fire forecasts can support early warning, disaster response, forest risk assessment, and downstream estimation of fire-related

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

simpleposter: a simple baseline for product poster generation

DGX agent

arXiv:2605.08784v1 Announce Type: new Abstract: Product poster generation poses distinct challenges beyond general poster design, requiring both faithful preservation of product appearance and precise

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data

DGX agent

arXiv:2605.10498v1 Announce Type: cross Abstract: Long-tailed distributions in class-imbalanced data present a fundamental challenge for deep learning models, which tend to be biased toward majority c

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success

DGX agent

arXiv:2605.09070v1 Announce Type: cross Abstract: Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders

DGX agent

arXiv:2605.08731v1 Announce Type: cross Abstract: JPEG decode is routine ML infrastructure, but Python decoder choices are often justified by single-process, single-thread microbenchmarks. We audit th

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Sinkhorn Treatment Effects: A Causal Optimal Transport Measure

DGX agent

arXiv:2605.08485v1 Announce Type: cross Abstract: We introduce the Sinkhorn treatment effect, an entropic optimal transport measure of divergence between counterfactual distributions. Unlike classical

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching

DGX agent

arXiv:2605.08658v1 Announce Type: cross Abstract: SKETCHVERIFY is a within-tier cost-performance policy, not a universal accuracy improvement. The operational question: a practitioner stuck with a sma

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SLAM: Structural Linguistic Activation Marking for Language Models

DGX agent

arXiv:2605.05443v2 Announce Type: replace-cross Abstract: LLM watermarks must be detectable without compromising text quality, yet most existing schemes bias the next-token distribution and pay for de

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SleepWalk: A Three-Tier Benchmark for Stress-Testing Instruction-Guided Vision-Language Navigation

DGX agent

arXiv:2605.10376v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have advanced rapidly in multimodal perception and language understanding, yet it remains unclear whether they can reliabl

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SLIM: Sparse Latent Steering for Interpretable and Property-Directed LLM-Based Molecular Editing

DGX agent

arXiv:2605.10831v1 Announce Type: cross Abstract: Large language models possess strong chemical reasoning capabilities, making them effective molecular editors. However, property-relevant information

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

DGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SMIXAE: Towards Unsupervised Manifold Discovery in Language Models

DGX agent

arXiv:2605.09224v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have been used widely to decompose and interpret neural network activations, especially those of transformer language models.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy

DGX agent

arXiv:2605.09598v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently shown strong potential in soccer video understanding. However, given the high complexity of soccer videos du

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

DGX agent

arXiv:2605.09063v1 Announce Type: new Abstract: Following the recent achievement of gold-medal performance on the IMO by frontier LLMs, the community is searching for the next meaningful and challengi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Source or It Didn't Happen: A Multi-Agent Framework for Citation Hallucination Detection

DGX agent

arXiv:2605.08583v1 Announce Type: new Abstract: Large language models are increasingly used in scientific writing, yet they can fabricate citation-shaped references that appear plausible but fail bibl

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Sparsity Moves Computation: How FFN Architecture Reshapes Attention in Small Transformers

DGX agent

arXiv:2605.09403v1 Announce Type: cross Abstract: Architectural choices inside the Transformer feedforward network (FFN) block do not merely affect the block itself; they reshape the computations lear

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?

DGX agent

arXiv:2602.03916v3 Announce Type: replace-cross Abstract: Spatial reasoning is a fundamental aspect of human cognition, yet it remains a major challenge for contemporary vision-language models (VLMs).

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

SPDEBench: An Extensive Benchmark for Learning Stochastic PDEs

DGX agent

arXiv:2505.18511v2 Announce Type: replace Abstract: Stochastic Partial Differential Equations (SPDEs) driven by random noise play a central role in modeling physical processes with rough spatio-tempor

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Spectral Characterization and Mitigation of Sequential Knowledge Editing Collapse

DGX agent

arXiv:2601.11042v2 Announce Type: replace-cross Abstract: Sequential knowledge editing in large language models often causes catastrophic collapse of the model's general abilities, especially for para

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SpectraLLM: Uncovering the Ability of LLMs for Molecular Structure Elucidation from Multi-Spectral Data

DGX agent

arXiv:2508.08441v3 Announce Type: replace-cross Abstract: Automated molecular structure elucidation remains challenging, as existing approaches often depend on pre-compiled databases or restrict thems

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

DGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Statistical Model Checking of the Keynes+Schumpeter Model: A Transient Sensitivity Analysis of a Macroeconomic ABM

DGX agent

arXiv:2605.10447v1 Announce Type: cross Abstract: Agent-based models (ABMs) are increasingly used in macroeconomics, but their analysis still often relies on ad hoc Monte Carlo campaigns with heteroge

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols

DGX agent

arXiv:2605.09618v1 Announce Type: new Abstract: When should a language model answer directly, sample and vote, or engage in multi-agent debate? Recent work shows voting often explains much of the gain

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Steerable but Not Decodable: Function Vectors Operate Beyond the Logit Lens

DGX agent

arXiv:2604.02608v2 Announce Type: replace Abstract: Activation steering presupposes that task-relevant behaviors correspond to linear directions in activation space -- directions that should both stee

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Step Rejection Fine-Tuning: A Practical Distillation Recipe

DGX agent

arXiv:2605.10674v1 Announce Type: cross Abstract: Rejection Fine-Tuning (RFT) is a standard method for training LLM agents, where unsuccessful trajectories are discarded from the training set. In the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Strategic commitments shape collective cybersecurity under AI inequality

DGX agent

arXiv:2605.09415v1 Announce Type: new Abstract: The growing integration of AI into cybersecurity is reshaping the balance between attackers and defenders. When access to advanced AI-enabled defence to

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

DGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Structure-Preserving Reconstruction of Convex Lipschitz Functionals on Hilbert Spaces from Finite Samples

DGX agent

arXiv:2605.08559v1 Announce Type: cross Abstract: Convex functionals are ubiquitous in applied analysis, appearing as value functions, risk measures, super-hedging prices, and loss functionals in mach

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SWE Atlas: Benchmarking Coding Agents Beyond Issue Resolution

DGX agent

arXiv:2605.08366v1 Announce Type: new Abstract: We introduce SWE Atlas, a benchmark suite for coding agents spanning three professional software engineering workflows: Codebase Q&A (124 tasks), Test W

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding

DGX agent

arXiv:2605.08412v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made rapid progress in single-video understanding, yet their ability to reason across multiple independent

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Synthetic Pre-Pre-Training Improves Language Model Robustness to Noisy Pre-Training Data

DGX agent

arXiv:2605.10129v1 Announce Type: new Abstract: Large language models (LLMs) rely on web-scale corpora for pre-training. The noise inherent in these datasets tends to obscure meaningful patterns and u

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Tabular Foundation Model for Generative Modelling

DGX agent

arXiv:2605.09424v1 Announce Type: new Abstract: Generative modelling is a demanding test of foundation models, because it requires robust, holistic representation learning for a given data modality, r

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.09539v1 Announce Type: new Abstract: Multi-agent systems (MAS) have emerged as a promising paradigm for solving complex tasks. Recent work has explored self-evolving MAS that automatically

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation

DGX agent

arXiv:2505.11604v5 Announce Type: replace Abstract: Editing presentation slides is a frequent yet tedious task, ranging from creative layout design to repetitive text maintenance. While recent GUI-bas

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning

DGX agent

arXiv:2605.10247v1 Announce Type: new Abstract: Using Large Language Models (LLMs) to process graph-structured data is an active research area, yet current state-of-the-art approaches typically rely o

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

DGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Test-Time Speculation

DGX agent

arXiv:2605.09329v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a fast draft model to generate tokens and a more accurate target model to verify them. Its perfo

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Text-Guided Multi-Scale Frequency Representation Adaptation

DGX agent

arXiv:2605.08181v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods introduce a small number of training parameters, enabling pre-trained models to adapt rapidly to new data dist

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

TFM-Retouche: A Lightweight Input-Space Adapter for Tabular Foundation Models

DGX agent

arXiv:2605.06047v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs), such as TabPFN-2.6, TabICLv2, ConTextTab, Mitra, LimiX, and TabDPT, achieve strong zero-shot performance thr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The Alpha Blending Hypothesis: Compositing Shortcut in Deepfake Detection

DGX agent

arXiv:2605.10334v1 Announce Type: new Abstract: Recent deepfake detection methods demonstrate improved cross-dataset generalization, yet the underlying mechanisms remain underexplored. We introduce th

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

DGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The autoPET3 Challenge: Automated Lesion Segmentation in Whole-Body PET/CT nicode{x2013} Multitracer Multicenter Generalization

DGX agent

arXiv:2605.05775v2 Announce Type: replace-cross Abstract: We report the design and results of the third autoPET challenge (MICCAI 2024), which benchmarked automated lesion segmentation in whole-body P

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The Differences Between Direct Alignment Algorithms are a Blur

DGX agent

arXiv:2502.01237v3 Announce Type: replace Abstract: Direct Alignment Algorithms (DAAs) simplify LLM alignment by directly optimizing policies, bypassing reward modeling and RL. While DAAs differ in th

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

DGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The Extrapolation Cliff in On-Policy Distillation of Near-Deterministic Structured Outputs

DGX agent

arXiv:2605.08737v1 Announce Type: cross Abstract: On-policy distillation (OPD) is widely used for LLM post-training. When pushed with a reward-extrapolation coefficient lambda > 1, the student can lif

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

The Geometric Wall: Manifold Structure Predicts Layerwise Sparse Autoencoder Scaling Laws

DGX agent

arXiv:2605.09887v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) operationalise the linear representation hypothesis: they reconstruct model activations as sparse linear combinations of in

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations

DGX agent

arXiv:2605.09195v1 Announce Type: new Abstract: Large language models confidently produce outdated answers, and no existing method can detect them. We show this is not an engineering failure but a str

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…260261262263264…361
Next →