AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Is Grep All You Need? How Agent Harnesses Reshape Agentic Search

DGX agent

arXiv:2605.15184v1 Announce Type: new Abstract: Recent advances in Large Language Model (LLM) agents have enabled complex agentic workflows where models autonomously retrieve information, call tools,

model-releasesarxiv-cs-cl
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation

DGX agent

arXiv:2512.12772v2 Announce Type: replace-cross Abstract: Understanding videos inherently requires reasoning over both visual and auditory information. To properly evaluate Omni-Large Language Models

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

K-Models: a Flexible and Interpretable Method for Ordinal Clustering with Application to Antigen-Antibody Interaction Profiles

DGX agent

arXiv:2605.14828v1 Announce Type: cross Abstract: Existing clustering methods for functional data often prioritize partitioning accuracy over interpretability, making it challenging to extract meaning

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

DGX agent

arXiv:2509.25826v3 Announce Type: replace Abstract: Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generaliz

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Kolmogorov-Arnold Chemical Reaction Neural Networks for learning pressure-dependent kinetic rate laws

DGX agent

arXiv:2511.07686v2 Announce Type: replace-cross Abstract: Chemical Reaction Neural Networks (CRNNs) have emerged as an interpretable machine learning framework for discovering reaction kinetics direct

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Krause Synchronization Transformers

DGX agent

arXiv:2602.11534v3 Announce Type: replace-cross Abstract: Self-attention in Transformers relies on globally normalized softmax weights, causing all tokens to compete for influence at every layer. When

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

L2R: Low-Rank and Lipschitz-Controlled Routing for Mixture-of-Experts

DGX agent

arXiv:2601.21349v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models scale neural networks by conditionally activating a small subset of experts, where the router plays a central

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Latency-Quality Routing for Functionally Equivalent Tools in LLM Agents

DGX agent

arXiv:2605.14241v1 Announce Type: new Abstract: Tool-augmented LLM agents increasingly access the same tool type through multiple functionally equivalent providers, such as web-search APIs, retrievers

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Learning Cross-Coupled and Regime Dependent Dynamics for Aerial Manipulation

DGX agent

arXiv:2605.14805v1 Announce Type: new Abstract: Accurate dynamics models are critical for aerial manipulators operating under complex tasks such as payload transport. However, modeling these systems r

model-releasesarxiv-cs-ro
15 May 2026
Model Releases

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling

DGX agent

arXiv:2605.14186v1 Announce Type: new Abstract: Large language models (LLMs) often expose useful signals of self-monitoring: before solving a problem, they can estimate whether they are likely to succ

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

LLMs Should Express Uncertainty Explicitly

DGX agent

arXiv:2604.05306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident yet incorrect answers, which can lead to risky failures in real-world applications. We st

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

LoMETab: Beyond Rank-1 Ensembles for Tabular Deep Learning

DGX agent

arXiv:2605.14365v1 Announce Type: cross Abstract: Recent tabular learning benchmarks increasingly show a tight performance cluster rather than a clear hierarchy among leading methods, spanning gradien

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning

DGX agent

arXiv:2508.06202v2 Announce Type: replace-cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models (MLLMs) to incrementally learn new tasks over time. Howeve

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

MahaVar: OOD Detection via Class-wise Mahalanobis Distance Variance under Neural Collapse

DGX agent

arXiv:2605.14413v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection is a critical component for ensuring the reliability of deep neural networks in safety-critical applications. In t

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Masked Next-Scale Prediction for Self-supervised Scene Text Recognition

DGX agent

arXiv:2605.14885v1 Announce Type: new Abstract: Scene Text Recognition requires modeling visual structures that evolve from coarse layouts to fine-grained character strokes. Training such models relie

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MathAtlas: A Benchmark for Autoformalization in the Wild

DGX agent

arXiv:2605.14061v1 Announce Type: new Abstract: Current autoformalization benchmarks are largely focused on olympiad or undergraduate mathematics, while graduate and research-level mathematics remains

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

MD-PNOP: Equation-Recast Neural Operators for Minimal-Data Extrapolation and PDE Solver Acceleration

DGX agent

arXiv:2509.01416v2 Announce Type: replace Abstract: The computational overhead of traditional numerical solvers for partial differential equations (PDEs) remains a critical bottleneck for large-scale

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Mechanistic Interpretability of EEG Foundation Models via Sparse Autoencoders

DGX agent

arXiv:2605.13930v1 Announce Type: new Abstract: EEG foundation models achieve state-of-the-art clinical performance, yet the internal computations driving their predictions remain opaque: a barrier to

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

MechVerse: Evaluating Physical Motion Consistency in Video Generation Models

DGX agent

arXiv:2605.14843v1 Announce Type: new Abstract: Text- and image-conditioned video generation models have achieved strong visual fidelity and temporal coherence, but they often fail to generate motion

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory

DGX agent

arXiv:2605.15128v1 Announce Type: cross Abstract: Long-term agent memory is increasingly multimodal, yet existing evaluations rarely test whether agents preserve the visual evidence needed for later r

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models

DGX agent

arXiv:2605.14906v1 Announce Type: new Abstract: Memory is essential for large vision-language models (LVLMs) to handle long, multimodal interactions, with two method directions providing this capabili

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval

DGX agent

arXiv:2605.06132v2 Announce Type: replace Abstract: In agent memory systems, the reranking model serves as the critical bridge connecting user queries with long-term memory. Most systems adopt the 're

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Merging Methods for Multilingual Knowledge Editing for Large Language Models: An Empirical Odyssey

DGX agent

arXiv:2605.13919v1 Announce Type: new Abstract: Multilingual knowledge editing (MKE) remains challenging because language-specific edits interfere with one another, even when locate-then-edit methods

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Mini-JEPA Foundation Model Fleet Enables Agentic Hydrologic Intelligence

DGX agent

arXiv:2605.14120v1 Announce Type: cross Abstract: Geospatial foundation models compress multispectral observations into dense embeddings increasingly used in natural-language environmental reasoning s

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Minimal-Intervention KV Retention: A Design-Space Study and a Diversity-Penalty Survivor

DGX agent

arXiv:2605.14292v1 Announce Type: cross Abstract: KV-cache compression at small budgets is a crowded design space spanning cache representation, head-wise routing, compression cadence, decoding behavi

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Mining Subscenario Refactoring Opportunities in Behaviour-Driven Software Test Suites: ML Classifiers and LLM-Judge Baselines

DGX agent

arXiv:2605.14568v1 Announce Type: cross Abstract: Context. Behaviour-Driven Development (BDD) software test suites accumulate duplicated step subsequences. Three published refactoring patterns are ava

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Mixed Integer Goal Programming for Personalized Meal Optimization with User-Defined Serving Granularity

DGX agent

arXiv:2605.13849v1 Announce Type: new Abstract: Determining what to eat to satisfy nutritional requirements is one of the oldest optimization problems in operations research, yet existing formulations

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ML-Embed: Inclusive and Efficient Embeddings for a Multilingual World

DGX agent

arXiv:2605.15081v1 Announce Type: cross Abstract: The development of high-quality text embeddings is increasingly drifting toward an exclusionary future, defined by three critical barriers: prohibitiv

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

MMTutorBench: The First Multimodal Benchmark for AI Math Tutoring

DGX agent

arXiv:2510.23477v2 Announce Type: replace Abstract: Effective math tutoring requires not only solving problems but also diagnosing students' difficulties and guiding them step by step. While multimoda

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models

DGX agent

arXiv:2511.08565v3 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly operate in social contexts, motivating analysis of how they express and shift moral judgments. In th

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation

DGX agent

arXiv:2605.13857v1 Announce Type: cross Abstract: The creation of cinematic-quality animal effects necessitates the precise modeling of muscle and fur dynamics, a process that remains both labor-inten

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MultiEmo-Bench: Multi-label Visual Emotion Analysis for Multi-modal Large Language Models

DGX agent

arXiv:2605.14635v1 Announce Type: cross Abstract: This paper introduces a multi-label visual emotion analysis benchmark dataset for comprehensively evaluating the ability of multimodal large language

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

MUON+: Towards More Effective Muon via One Additional Normalization Step for LLM Pre-training

DGX agent

arXiv:2602.21545v3 Announce Type: replace Abstract: Muon has recently emerged as a strong optimizer for large language model pre-training, orthogonalizing the momentum matrix via Newton--Schulz polar

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

nASR: An End-to-End Trainable Neural Layer for Channel-Level EEG Artifact Subspace Reconstruction in Real-Time BCI

DGX agent

arXiv:2605.14941v1 Announce Type: cross Abstract: Electroencephalogram (EEG) signals are highly susceptible to artifacts, resulting in a low signal-to-noise ratio which makes extraction of meaningful

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Near-Miss: Latent Policy Failure Detection in Agentic Workflows

DGX agent

arXiv:2603.29665v2 Announce Type: replace Abstract: Agentic systems for business process automation often require compliance with policies governing conditional updates to the system state. Evaluation

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Network-Aware Bilinear Tokenization for Brain Functional Connectivity Representation Learning

DGX agent

arXiv:2605.14048v1 Announce Type: new Abstract: Masked autoencoders (MAEs) have recently shown promise for self-supervised representation learning of resting-state brain functional connectivity (FC).

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Neural Fields for NV-Center Inverse Sensing

DGX agent

arXiv:2605.13988v1 Announce Type: new Abstract: Inverse problems in scientific sensing are often solved with either hand-designed regularizers or supervised networks trained on simulated labels, yet b

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Neural Signals Generate Clinical Notes in the Wild

DGX agent

arXiv:2601.22197v3 Announce Type: replace-cross Abstract: Generating clinical reports that summarize abnormal patterns, diagnostic findings, and clinical interpretations from long-term EEG recordings

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

NeuroAtlas: Benchmarking Foundation Models for Clinical EEG and Brain-Computer Interfaces

DGX agent

arXiv:2605.14698v1 Announce Type: cross Abstract: Foundation models (FMs) promise to extract unified representations that generalize across downstream tasks. They have emerged across fields, including

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

NeuroMambaLLM: Dynamic Graph Learning of fMRI Functional Connectivity in Autistic Brains Using Mamba and Language Model Reasoning

DGX agent

arXiv:2602.13770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong semantic reasoning across multimodal domains. However, their integration with graph-base

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

NodeSynth: Socially Aligned Synthetic Data for AI Evaluation

DGX agent

arXiv:2605.14381v1 Announce Type: cross Abstract: Recent advancements in generative AI facilitate large-scale synthetic data generation for model evaluation. However, without targeted approaches, thes

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Octopus: History-Free Gradient Orthogonalization for Continual Learning in Multimodal Large Language Models

DGX agent

arXiv:2605.14938v1 Announce Type: cross Abstract: Continual learning in multimodal large language models (MLLMs) aims to sequentially acquire knowledge while mitigating catastrophic forgetting, yet ex

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

On the Cultural Anachronism and Temporal Reasoning in Vision Language Models

DGX agent

arXiv:2605.15071v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied to cultural heritage materials, from digital archives to educational platforms. This work ident

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

OpenDeepThink: Parallel Reasoning via Bradley--Terry Aggregation

DGX agent

arXiv:2605.15177v1 Announce Type: new Abstract: Test-time compute scaling is a primary axis for improving LLM reasoning. Existing methods primarily scale depth by extending a single reasoning trace. S

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling

DGX agent

arXiv:2601.19924v2 Announce Type: replace-cross Abstract: We investigate the capabilities and scalability of Large Language Models (LLMs) in optimization modeling, a domain requiring structured reason

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems

DGX agent

arXiv:2511.16964v2 Announce Type: replace-cross Abstract: Maximizing performance on available GPU hardware is an ongoing challenge for modern AI inference systems. Traditional approaches include writi

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

PaAno: Patch-Based Representation Learning for Time-Series Anomaly Detection

DGX agent

arXiv:2602.01359v2 Announce Type: replace-cross Abstract: Although recent studies on time-series anomaly detection have increasingly adopted ever-larger neural network architectures such as transforme

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts

DGX agent

arXiv:2605.14055v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) is widely used for adapting Large Language Models (LLMs) for various tasks. Recently, there has been an increas

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…238239240241242…361
Next →