AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

DGX agent

arXiv:2604.22271v1 Announce Type: new Abstract: Large language models can detect their own errors and sometimes correct them without external feedback, but the underlying mechanisms remain unknown. We

model-releasesarxiv-cs-lg
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

HubRouter: A Pluggable Sub-Quadratic Routing Primitive for Hybrid Sequence Models

DGX agent

arXiv:2604.22442v1 Announce Type: new Abstract: We introduce HubRouter, a pluggable module that replaces O(n^2) attention layers with O(nM) hub-mediated routing, where M =20 shows increasing seed sens

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!

DGX agent

arXiv:2507.03014v2 Announce Type: replace-cross Abstract: Large language models (LLMs) face significant copyright and intellectual property challenges as the cost of training increases and model reuse

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Knowledge Visualization: A Benchmark and Method for Knowledge-Intensive Text-to-Image Generation

DGX agent

arXiv:2604.22302v1 Announce Type: new Abstract: Recent text-to-image (T2I) models have demonstrated impressive capabilities in photorealistic synthesis and instruction following. However, their reliab

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

KuaiLive: A Real-time Interactive Dataset for Live Streaming Recommendation

DGX agent

arXiv:2508.05633v2 Announce Type: replace-cross Abstract: Live streaming platforms have become a dominant form of online content consumption, offering dynamically evolving content, real-time interacti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Language Specific Knowledge: Do Models Know Better in X than in English?

DGX agent

arXiv:2505.14990v3 Announce Type: replace Abstract: Often, multilingual language models are trained with the objective to map semantically similar content (in different languages) in the same latent s

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Learning Coverage- and Power-Optimal Transmitter Placement from Building Maps: A Comparative Study of Direct and Indirect Neural Approaches

DGX agent

arXiv:2604.22056v1 Announce Type: new Abstract: Optimal wireless transmitter placement is a central task in radio-network planning, yet exhaustive search becomes prohibitively expensive at scale. This

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Learning from Natural Language Feedback for Personalized Question Answering

DGX agent

arXiv:2508.10695v2 Announce Type: replace-cross Abstract: Personalization is crucial for enhancing both the effectiveness and user satisfaction of language technologies, particularly in information-se

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines

DGX agent

arXiv:2411.08027v3 Announce Type: replace-cross Abstract: Most learning-based approaches to complex physical reasoning sidestep the crucial problem of parameter identification (e.g., mass, friction) t

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

LLMs as Assessors: Right for the Right Reason?

DGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Long-tail Internet photo reconstruction

DGX agent

arXiv:2604.22714v1 Announce Type: new Abstract: Internet photo collections exhibit an extremely long-tailed distribution: a few famous landmarks are densely photographed and easily reconstructed in 3D

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

LTBs-KAN: Linear-Time B-splines Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.22034v1 Announce Type: cross Abstract: Kolmogorov-Arnold Networks (KANs) are a recent neural network architecture offering an alternative to Multilayer Perceptrons (MLPs) with improved expl

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

MacrOData: New Benchmarks of Thousands of Datasets for Tabular Outlier Detection

DGX agent

arXiv:2602.09329v2 Announce Type: replace Abstract: Quality benchmarks are essential for fairly and accurately tracking scientific progress and enabling practitioners to make informed methodological c

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Math Takes Two: A test for emergent mathematical reasoning in communication

DGX agent

arXiv:2604.21935v1 Announce Type: new Abstract: Although language models demonstrate remarkable proficiency on mathematical benchmarks, it remains unclear whether this reflects true mathematical reaso

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

MolClaw: An Autonomous Agent with Hierarchical Skills for Drug Molecule Evaluation, Screening, and Optimization

DGX agent

arXiv:2604.21937v1 Announce Type: new Abstract: Computational drug discovery, particularly the complex workflows of drug molecule screening and optimization, requires orchestrating dozens of specializ

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models

DGX agent

arXiv:2604.22492v1 Announce Type: cross Abstract: Understanding social dominance in animal behavior is critical for neuroscience and behavioral studies. In this work, we explore the capability of Mult

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

DGX agent

arXiv:2604.22548v1 Announce Type: cross Abstract: Problem definition: Data-driven models in machine learning have enabled efficient management of production systems. However, a majority of machine lea

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Multi-Task Optimization over Networks of Tasks

DGX agent

arXiv:2604.21991v1 Announce Type: cross Abstract: Multi-task optimization is a powerful approach for solving a large number of tasks in parallel. However, existing algorithms face distinct limitations

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Navigating Large-Scale Document Collections: MuDABench for Multi-Document Analytical QA

DGX agent

arXiv:2604.22239v1 Announce Type: cross Abstract: This paper introduces the task of analytical question answering over large, semi-structured document collections. We present MuDABench, a benchmark fo

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

DGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

On Benchmark Hacking in ML Contests: Modeling, Insights and Design

DGX agent

arXiv:2604.22230v1 Announce Type: cross Abstract: Benchmark hacking refers to tuning a machine learning model to score highly on certain evaluation criteria without improving true generalization or fa

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Optimal Question Selection from a Large Question Bank for Clinical Field Recovery in Conversational Psychiatric Intake

DGX agent

arXiv:2604.22067v1 Announce Type: cross Abstract: Psychiatric intake is a sequential, high-stakes information-gathering process in which clinicians must decide what to ask, in what order, and how to i

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Optimal sequential decision-making for error propagation mitigation in digital twins

DGX agent

arXiv:2604.22168v1 Announce Type: new Abstract: Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study th

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Parameter-Efficient Conditioning for Material Generalization in Graph-Based Simulators

DGX agent

arXiv:2511.05456v2 Announce Type: replace Abstract: Graph network-based simulators (GNS) have demonstrated strong potential for learning particle-based physics (such as fluids, deformable solids, and

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

PL-MTEB: Polish Massive Text Embedding Benchmark

DGX agent

arXiv:2405.10138v2 Announce Type: replace Abstract: In this paper, we introduce the Polish Massive Text Embedding Benchmark (PL-MTEB), a comprehensive benchmark for text embeddings in the Polish langu

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

DGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Railway Artificial Intelligence Learning Benchmark (RAIL-BENCH): A Benchmark Suite for Perception in the Railway Domain

DGX agent

arXiv:2604.22507v1 Announce Type: new Abstract: Automated train operation on existing railway infrastructure requires robust camera-based perception, yet the railway domain lacks public benchmark suit

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Recent Advances in Multi-Agent Human Trajectory Prediction: A Comprehensive Review

DGX agent

arXiv:2506.14831v3 Announce Type: replace Abstract: With the emergence of powerful data-driven methods in human trajectory prediction (HTP), gaining a finer understanding of multi-agent interactions l

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

DGX agent

arXiv:2604.22390v1 Announce Type: new Abstract: Visual Place Recognition (VPR) determines a query image's geographic location by matching it against geotagged databases. However, existing methods stru

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Regularized Meta-Learning for Improved Generalization

DGX agent

arXiv:2602.12469v2 Announce Type: replace Abstract: Deep ensemble methods often improve predictive performance, yet they suffer from three practical limitations: redundancy among base models that infl

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Relaxation-Informed Training of Neural Network Surrogate Models

DGX agent

arXiv:2604.22746v1 Announce Type: cross Abstract: ReLU neural networks trained as surrogate models can be embedded exactly in mixed-integer linear programs (MILPs), enabling global optimization over t

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores

DGX agent

arXiv:2604.22063v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in clinical reasoning and risk assessment. However, their interpretive reliability in critical

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

DGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Publication: A Certification Framework for AI-Enabled Research

DGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problem

DGX agent

arXiv:2602.18734v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has demonstrated strong effectiveness in knowledge-intensive tasks by grounding language generation in ex

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Score-based Membership Inference on Diffusion Models

DGX agent

arXiv:2509.25003v2 Announce Type: replace-cross Abstract: Membership inference attacks (MIAs) against Diffusion Models (DMs) raise pressing privacy concerns by revealing whether a sample was part of t

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Selective Depthwise Separable Convolution for Lightweight Joint Source-Channel Coding in Wireless Image Transmission

DGX agent

arXiv:2604.22338v1 Announce Type: cross Abstract: Depthwise separable convolutional (DSConv) layers have been successfully applied to deep learning (DL)-based joint source-channel coding (JSCC) scheme

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Shaken or Stirred? An Analysis of MetaFormer's Token Mixing for Medical Imaging

DGX agent

arXiv:2510.05971v3 Announce Type: replace Abstract: The generalization of the Transformer architecture via MetaFormer has reshaped our understanding of its success in computer vision. By replacing sel

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Sovereign Agentic Loops: Decoupling AI Reasoning from Execution in Real-World Systems

DGX agent

arXiv:2604.22136v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly issue API calls that mutate real systems, yet many current architectures pass stochastic model outputs

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

SpaMEM: Benchmarking Dynamic Spatial Reasoning via Perception-Memory Integration in Embodied Environments

DGX agent

arXiv:2604.22409v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced static visual--spatial reasoning, yet they often fail to preserve long-horizon spatial coherence

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection

DGX agent

arXiv:2604.22753v1 Announce Type: new Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, asse

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

DGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models

DGX agent

arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Time-Localized Parametric Decomposition of Respiratory Airflow for Sub-Breath Analysis

DGX agent

arXiv:2604.22695v1 Announce Type: cross Abstract: Respiratory airflow signals provide critical insight into breathing mechanics, yet conventional analysis methods remain limited in their ability to ch

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Toward Automated Robustness Evaluation of Mathematical Reasoning

DGX agent

arXiv:2506.05038v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in various reasoning-intensive tasks. However, these models exhibit unexpecte

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution

DGX agent

arXiv:2604.22464v1 Announce Type: new Abstract: Continual Model Merging (CMM) sequentially integrates task-specific models into a unified architecture without intensive retraining. However, existing C

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

DGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

model-releasesarxiv-cs-cv
27 Apr 2026
← Previous
1…299300301302303…357
Next →