AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

WLNO: Wavelet-Laplace Neural Operator for Solving Partial Differential Equations

DGX agent

arXiv:2605.24658v1 Announce Type: new Abstract: This work introduces the Wavelet-Laplace Neural Operator (WLNO), a novel neural operator that fuses Haar wavelet multi-scale spatial decomposition with

model-releasesarxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

World-State Transformations for Neuro-symbolic Interactive Storytelling

DGX agent

arXiv:2605.24719v1 Announce Type: cross Abstract: Large Language Models (LLMs) have changed the possibilities of Interactive Storytelling systems that process free-text user input. However, as more of

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

WorldGUI: An Interactive Benchmark for Desktop GUI Automation from Any Starting Point

DGX agent

arXiv:2502.08047v5 Announce Type: replace Abstract: Recent progress in GUI agents has substantially improved visual grounding, yet robust planning remains challenging, particularly when the environmen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

DGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

A European Multi-Center Breast Cancer MRI Dataset

DGX agent

arXiv:2506.00474v3 Announce Type: replace-cross Abstract: Early detection of breast cancer is critical for improving patient outcomes. While mammography remains the primary screening modality, magneti

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

DGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

A Reproducible Universal Dependencies-Style Pipeline for Katharevousa Greek Parliamentary Text

DGX agent

arXiv:2605.22978v1 Announce Type: new Abstract: Katharevousa Greek remains poorly served by contemporary NLP pipelines despite its importance for legal, administrative, and parliamentary archives. We

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

A Systematic Evaluation of Co-folding Model Representations for Small-Molecule Learning

DGX agent

arXiv:2602.13249v2 Announce Type: replace-cross Abstract: Small-molecule foundation models are typically pretrained on standalone molecular data, unlike vision and language models that often benefit f

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Agentic Proving for Program Verification

DGX agent

arXiv:2605.23772v1 Announce Type: new Abstract: Agentic systems have recently emerged as state-of-the-art approaches for automated theorem proving in formal mathematics. To assess how far these capabi

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

AI Evaluation Should Require Standardized Item-Level Data Releases

DGX agent

arXiv:2604.03244v2 Announce Type: replace Abstract: This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluatio

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Any2Any: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking

DGX agent

arXiv:2605.23733v1 Announce Type: cross Abstract: Whole-body tracking (WBT) models have become a key foundation for humanoid robots, enabling them to imitate diverse motions with high fidelity. Traini

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Anytime Training with Schedule-Free Spectral Optimization

DGX agent

arXiv:2605.23061v1 Announce Type: cross Abstract: Standard neural network training relies on learning-rate schedules tied to a fixed horizon, leading to strong path dependence and costly re-tuning as

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Approaching I/O-optimality for Approximate Attention

DGX agent

arXiv:2605.23751v1 Announce Type: new Abstract: We revisit the I/O complexity of attention in large language models. Given query-key-value matrices Q,K,VinR^{nimes d}, and a machine with fast memory s

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse

DGX agent

arXiv:2605.23325v1 Announce Type: new Abstract: Social media has become a crucial arena for shaping public narratives during armed conflicts, providing space for both harmful and constructive communic

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Archimedean Copula Inference via Taylor-Mode AD

DGX agent

arXiv:2605.23134v1 Announce Type: new Abstract: No existing nested Archimedean copula tool handles all three of (a) arbitrary per-variable (right-)censoring in survival analysis, (b) arbitrary nesting

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from Dual-Mode Vulnerability Benchmarks

DGX agent

arXiv:2605.23243v1 Announce Type: cross Abstract: We evaluate whether frontier LLMs are ready for cybersecurity through a dual-mode benchmark: white-box function-level vulnerability detection (VulnLLM

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

As X, Do Y: How Persona and Task Combine in Instruction-Tuned LLMs

DGX agent

arXiv:2605.23147v1 Announce Type: cross Abstract: Role prompts of the form As X, do Y admit a clean linear decomposition at one specific site in the residual stream: the prompt-to-answer transition --

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

DGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Atom-level Protein Representation Learning Improves Protein Structure Prediction

DGX agent

arXiv:2605.22133v2 Announce Type: replace-cross Abstract: Recent advances in generative modeling show that pretrained representations can improve generation as conditioning features or alignment targe

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

DGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Benchmarking and Enhancing VLM for Compressed Image Understanding

DGX agent

arXiv:2512.20901v2 Announce Type: replace Abstract: With the rapid development of Vision-Language Models (VLMs) and the growing demand for their applications, efficient compression of the image inputs

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems

DGX agent

arXiv:2605.23618v1 Announce Type: new Abstract: We benchmark Google Embeddings (GE2), a Vertex-AI-hosted bi-encoder with 2,048-token context and explicit task-type conditioning, against five open-sour

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

BOHM: Zero-Cost Hierarchical Attribution for Compound AI Systems

DGX agent

arXiv:2605.22866v1 Announce Type: new Abstract: Compound AI systems route tasks through hierarchies of specialised components. Attribution is dominated by Shapley-based methods (SHAP), which decompose

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Brain-LLM Alignment Tracks Training Data, Not Typology

DGX agent

arXiv:2605.23032v1 Announce Type: cross Abstract: Brain-LLM alignment is well established in English, yet the brain's language network is neuroanatomically universal across languages. Does alignment a

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

DGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

DGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

CARE: Class-Adaptive Expert Consensus for Reliable Learning with Long-Tailed Noisy Labels

DGX agent

arXiv:2605.23254v1 Announce Type: new Abstract: Learning from real-world data is frequently hindered by the compound challenge of long-tailed class distributions and noisy annotations. Existing method

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering

DGX agent

arXiv:2605.23216v1 Announce Type: new Abstract: Cause-and-effect reasoning in video is a significant challenge for Vision-Language Models (VLMs), as it requires going beyond surface-level perception t

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models

DGX agent

arXiv:2605.23694v1 Announce Type: new Abstract: Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

CHRONOS: Temporally-Aware Multi-Agent Coordination for Evolving Data Marketplaces

DGX agent

arXiv:2605.23887v1 Announce Type: cross Abstract: Temporal knowledge-graph data marketplaces face three coupled failures in static designs: stale hybrid index shortcuts reduce recall as edges evolve,

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Complete-muE: Optimal Hyperparameter Transfer and Scaling for MoE Models

DGX agent

arXiv:2605.23893v1 Announce Type: new Abstract: We propose Complete-muE, a framework which targets hyperparameter transfer across dense FFN and any Mixture-of-Experts (MoE) setups in transformer block

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Computable Fairness: Boltzmann-Softmax Control for AI Resource Allocation

DGX agent

arXiv:2605.22827v1 Announce Type: cross Abstract: In large-scale AI systems, allocating scarce resources such as GPU compute time and bandwidth among multiple agents is a critical challenge. Conventio

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Convex Optimization for Alignment and Preference Learning on a Single GPU

DGX agent

arXiv:2605.23244v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) to align with human preferences has driven the success of systems such as Gemini and ChatGPT. However, approach

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Cost-Effective Model Evaluation with Meta-Learning

DGX agent

arXiv:2605.23595v1 Announce Type: cross Abstract: The rapid growth of machine learning has produced an ever-expanding ecosystem of models, making it increasingly challenging to verify the reliability

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization

DGX agent

arXiv:2605.23391v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) for coupled multiphysics systems suffer systematic accuracy degradation as inter-equation coupling strengthens.

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models

DGX agent

arXiv:2605.23699v1 Announce Type: new Abstract: Video prediction is increasingly viewed as a path toward generalizable world models, yet it remains unclear whether these systems learn underlying causa

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Cultural Adaptation in Large Language Models for Political Discourse

DGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception

DGX agent

arXiv:2605.23655v1 Announce Type: cross Abstract: High-resolution (HR) image perception presents a key bottleneck for multimodal large language models (MLLMs). While visual search offers a promising s

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

D2 Actor Critic: Diffusion Actor Meets Distributional Critic

DGX agent

arXiv:2510.03508v3 Announce Type: replace Abstract: We introduce D2AC, a new model-free reinforcement learning (RL) algorithm designed to train expressive diffusion policies online effectively. At its

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures

DGX agent

arXiv:2511.15503v2 Announce Type: replace-cross Abstract: High-performance Host processors can integrate Processing-In-Memory (PIM) devices, which can accelerate memory-intensive kernels of Machine Le

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs

DGX agent

arXiv:2605.23629v1 Announce Type: new Abstract: Medical diagnosis is not a single prediction from a fully specified vignette. It is a sequential workup: clinicians decide what evidence to obtain, revi

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Decomposing and Measuring Evaluation Awareness

DGX agent

arXiv:2605.23055v1 Announce Type: cross Abstract: Frontier language models sometimes recognize that they are being evaluated and adjust their behavior, undermining validity of benchmark results. Yet t

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

DGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Decomposition-Based Modular Conformal Prediction for Two-Stage Modeling

DGX agent

arXiv:2510.04406v2 Announce Type: replace-cross Abstract: Conformal prediction offers finite-sample coverage guarantees under minimal assumptions. However, existing methods treat the entire modeling p

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

DGX agent

arXiv:2605.23355v1 Announce Type: new Abstract: Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professiona

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

DepthAgent: Towards Better Universal Depth Estimation via Sample-wise Expert Selection

DGX agent

arXiv:2605.23281v1 Announce Type: new Abstract: Monocular metric depth estimation has achieved strong progress with large-scale training and universal-camera modeling, yet robust deployment across div

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
← Previous
1…204205206207208…361
Next →