AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

GraphFlow: A Graph-Based Workflow Management for Efficient LLM-Agent Serving

DGX agent

arXiv:2605.22566v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents demonstrate strong reasoning and execution capabilities on complex tasks when guided by structured instructions,

model-releasesarxiv-cs-lg
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

DGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Holomorphic Neural ODEs with Kolmogorov-Arnold Networks for Interpretable Discovery of Complex Dynamics

DGX agent

arXiv:2605.22235v1 Announce Type: new Abstract: Complex dynamical systems governed by holomorphic maps such as z^2 + c exhibit fractal boundaries with extreme sensitivity to initial conditions. Accura

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Hybrid Kolmogorov-Arnold Network and XGBoost Framework for Week-Ahead Price Forecasting in Australia's National Electricity Market

DGX agent

arXiv:2605.22387v1 Announce Type: new Abstract: Accurate electricity price forecasting (EPF) is essential for market participants to support operational planning and risk management, yet remains chall

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Hyperparameter Transfer with Mixture-of-Expert Layers

DGX agent

arXiv:2601.20205v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers have emerged as an important tool in scaling up modern neural networks by decoupling total trainable parameters from

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

I-SAFE: Wasserstein Coherence Metrics for Structural Auditing of Scientific AI Models

DGX agent

arXiv:2605.21731v1 Announce Type: new Abstract: Deep learning models are increasingly used in scientific prediction tasks where strong benchmark performance is often interpreted as evidence of scienti

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

IKNO: Infinite-order Kernel Neural Operators

DGX agent

arXiv:2605.22182v1 Announce Type: new Abstract: Neural operators have achieved significant success in modern scientific computing due to their flexibility and strong generalization capabilities. Exist

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Integrable Elasticity via Neural Demand Potentials

DGX agent

arXiv:2605.22820v1 Announce Type: new Abstract: We propose the Integrable Context-Dependent Demand Network (ICDN), a demand-first neural model for multiproduct retail demand. The model learns log-dema

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Lumberjack: Better Differentially Private Random Forests through Heavy Hitter Detection in Trees

DGX agent

arXiv:2605.22756v1 Announce Type: new Abstract: Random forests are widely used in fields involving sensitive tabular data, but existing approaches to enforcing differential privacy (DP) typically degr

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

MapTab: Are MLLMs Ready for Multi-Criteria Route Planning in Heterogeneous Graphs?

DGX agent

arXiv:2602.18600v3 Announce Type: replace Abstract: Systematic evaluation of Multimodal Large Language Models (MLLMs) is crucial for advancing Artificial General Intelligence (AGI). However, existing

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability

DGX agent

arXiv:2605.22168v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) map complex visual inputs to semantic spaces, but interpreting the cross-modal reasoning of VLMs currently relies on pos

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Memory-Efficient LLM Pretraining via Minimalist Optimizer Design

DGX agent

arXiv:2506.16659v3 Announce Type: replace Abstract: Training large language models (LLMs) relies on adaptive optimizers such as Adam, which introduce extra operations and require significantly more me

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Memory-R2: Fair Credit Assignment for Long-Horizon Memory-Augmented LLM Agents

DGX agent

arXiv:2605.21768v1 Announce Type: new Abstract: Memory-augmented LLM agents enable interactions that extend beyond finite context windows by storing, updating, and reusing information across sessions.

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization

DGX agent

arXiv:2605.21751v1 Announce Type: new Abstract: Text-to-optimization requires two separable capabilities: modeling -- choosing the right optimization structure -- and binding -- grounding every coeffi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents

DGX agent

arXiv:2602.13372v2 Announce Type: replace-cross Abstract: Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

DGX agent

arXiv:2602.12506v3 Announce Type: replace Abstract: Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivatin

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

DGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Posterior Collapse as Automatic Spectral Pruning

DGX agent

arXiv:2605.22691v1 Announce Type: new Abstract: We show that posterior collapse in eta-VAEs implements automatic spectral pruning. A latent mode collapses if its contribution to reconstruction is belo

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Prior Knowledge-enhanced Spatio-temporal Epidemic Forecasting

DGX agent

arXiv:2602.22270v2 Announce Type: replace Abstract: Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

DGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

DGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

DGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

DGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

DGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

DGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching

DGX agent

arXiv:2605.22083v1 Announce Type: cross Abstract: While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issu

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

DGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Scalable On-Policy Reinforcement Learning via Adaptive Batch Scaling

DGX agent

arXiv:2605.21557v1 Announce Type: cross Abstract: Conventional wisdom holds that large-batch training is fundamentally incompatible with Reinforcement Learning (RL) - beyond a modest threshold, increa

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

DGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

SeqLoRA: Bilevel Orthogonal Adaptation for Continual Multi-Concept Generation

DGX agent

arXiv:2605.22743v1 Announce Type: new Abstract: Parameter-efficient fine-tuning enables fast personalization of text-to-image diffusion models, but composing multiple custom concepts remains challengi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

DGX agent

arXiv:2605.22142v1 Announce Type: new Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly mode

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

DGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

DGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Symbolic Density Estimation for Discrete Distributions

DGX agent

arXiv:2605.21813v1 Announce Type: new Abstract: Discrete probability laws underpin statistical modeling, yet the catalog of interpretable distributions has expanded only gradually through centuries of

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Tabular foundation models for robust calibration of near-infrared chemical sensing data

DGX agent

arXiv:2605.21544v1 Announce Type: new Abstract: Near-infrared spectroscopy is increasingly used as a rapid, non-destructive chemical sensing technology for the analysis of food, pharmaceutical, biolog

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation

DGX agent

arXiv:2605.21856v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated impressive reasoning abilities across a wide range of tasks, but data contamination undermines the object

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Secretary Problem with a Stochastic Precursor

DGX agent

arXiv:2605.22653v1 Announce Type: cross Abstract: In learning-augmented online algorithms, predictions are usually valued for what they say: a value estimate, a solution, or an algorithmic recommendat

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

The Volterra signature

DGX agent

arXiv:2603.04525v2 Announce Type: replace-cross Abstract: Modern approaches for learning from non-Markovian time series, such as recurrent neural networks, neural controlled differential equations or

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

DGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

UNAD+: An Explainable Hybrid Framework for Unknown Network Attack Detection

DGX agent

arXiv:2605.22621v1 Announce Type: cross Abstract: The detection of previously unseen network attacks remains a major challenge for intrusion detection systems. Although supervised learning methods oft

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation

DGX agent

arXiv:2605.22368v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed for software engineering, constructing high-quality benchmarks is crucial for evaluating not j

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning

DGX agent

arXiv:2605.21606v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own rollouts using a privileged teacher, but its standard objective weights all generated tok

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Why SGD is not Brownian Motion: A New Perspective on Stochastic Dynamics

DGX agent

arXiv:2605.22644v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is commonly modeled as a Langevin process, assuming that minibatch noise acts as Brownian motion. However, this approx

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

3D LULC classification using multispectral LiDAR and deep learning: current and prospective schemes

DGX agent

arXiv:2605.22328v1 Announce Type: new Abstract: Land Use Land Cover (LULC) classification is essential for national 3D mapping, geospatial analysis, and sustainable planning. Multispectral (MS) LiDAR

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

A Task-Agnostic Algebraic Integrity Metric for Event-Camera Streams Toward SOTIF-Compliant Perception using Pearson Correlation Coefficient

DGX agent

arXiv:2605.21500v1 Announce Type: cross Abstract: Event cameras have emerged as a high-bandwidth, low-latency sensing modality for safety-critical perception in automated driving systems (ADS), offeri

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AesFormer: Transform Everyday Photos into Beautiful Memories

DGX agent

arXiv:2605.22126v1 Announce Type: new Abstract: In everyday photography, aesthetically appealing moments are often captured with structural flaws (e.g., composition, camera viewpoint, or pose) that ex

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows

DGX agent

arXiv:2605.20425v1 Announce Type: new Abstract: Designing multi-agent workflows is especially difficult in open-ended scientific settings where tasks lack curated training sets, reliable scalar evalua

model-releasesarxiv-cs-ai
22 May 2026
← Previous
1…208209210211212…361
Next →