AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives

DGX agent

arXiv:2605.18866v1 Announce Type: cross Abstract: Reconstructing continuous flow fields from sparse surface-mounted sensors is central to aerodynamic design, flow control, and digital-twin instrumenta

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FLUXtrapolation: A benchmark on extrapolating ecosystem fluxes

DGX agent

arXiv:2605.19812v1 Announce Type: cross Abstract: We introduce FLUXtrapolation, a benchmark for extrapolating ecosystem fluxes under progressively harder distribution shifts. Ecosystem fluxes are cent

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

From Llama to Cria: Scaling Down Neural Networks via Neuron-Level Spectral Structural Importance Evaluation

DGX agent

arXiv:2605.18860v1 Announce Type: cross Abstract: This paper proposes a neuron pruning framework based on neuron-level spectral structural importance evaluation. Given a trained neural network, we rec

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

DGX agent

arXiv:2605.19824v1 Announce Type: new Abstract: Recent attempts to support high-level scene interpretation and planning in Autonomous Vehicles (AVs) using ensembles of Large Language Models (LLMs) and

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

From SGD to Muon: Adaptive Optimization via Schatten-p Norms

DGX agent

arXiv:2605.19781v1 Announce Type: new Abstract: Modern optimizers, like Muon, impose matrix-wise geometry constraints on their updates. These matrix-wise constraints can be unified under Linear Minimi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

From Simple to Complex: Curriculum-Guided Physics-Informed Neural Networks via Gaussian Mixture Models

DGX agent

arXiv:2605.19263v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) offer a mesh-free framework for solving partial differential equations (PDEs), yet training often suffers from

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation

DGX agent

arXiv:2605.18865v1 Announce Type: cross Abstract: Self-attention serves as the core foundation of large-scale transformer pretraining, but its quadratic token interaction cost makes inference expensiv

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

General Lower Bounds for Differentially Private Federated Learning with Arbitrary Public-Transcript Interactions

DGX agent

arXiv:2605.19813v1 Announce Type: new Abstract: We prove a general lower bound for differentially private federated learning protocols with arbitrary public-transcript interactions. The protocol may u

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Generalization Bounds of Surrogate Policies for Combinatorial Optimization Problems

DGX agent

arXiv:2407.17200v3 Announce Type: replace-cross Abstract: Many real-world decision problems require solving, again and again, combinatorial optimization instances drawn from a common distribution. A r

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

DGX agent

arXiv:2605.20006v1 Announce Type: new Abstract: Geospatial reasoning requires solving image-grounded problems over the complex spatial structure of a scene. However, developing this capability is hind

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

GoLongRL: Capability-Oriented Long Context Reinforcement Learning with Multitask Alignment

DGX agent

arXiv:2605.19577v1 Announce Type: new Abstract: We present GoLongRL, a fully open-source, capability-oriented post-training recipe for long-context reinforcement learning with verifiable rewards (RLVR

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

DGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

GRAB: A Risk Taxonomy--Grounded Benchmark for Unsupervised Topic Discovery in Financial Disclosures

DGX agent

arXiv:2509.21698v2 Announce Type: replace Abstract: Risk categorization in 10-K risk disclosures matters for oversight and investment, yet no public benchmark evaluates unsupervised topic models for t

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Graph Neural Networks for Community Detection in Graph Signal Analysis

DGX agent

arXiv:2605.19733v1 Announce Type: cross Abstract: Community detection is a central problem in graph analysis, with applications ranging from network science to graph signal processing. In recent years

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

GroupAffect-4: A Multimodal Dataset of Four-Person Collaborative Interaction

DGX agent

arXiv:2605.19765v1 Announce Type: new Abstract: Existing affective-computing, social-signal-processing, and meeting corpora capture important parts of human interaction, but they rarely support analys

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Hallucination as Exploit: Evidence-Carrying Multimodal Agents

DGX agent

arXiv:2605.19192v1 Announce Type: new Abstract: Multimodal agents use screenshots, documents, and webpages to choose tool calls. When a false visual claim triggers a click, email, extraction, or trans

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models

DGX agent

arXiv:2605.19341v1 Announce Type: cross Abstract: Hallucination remains a central failure mode of large language models, but existing benchmarks operationalize it inconsistently across summarization,

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

DGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models

DGX agent

arXiv:2605.18795v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) dominates parameter-efficient fine-tuning of large language models, yet most variants target dense architectures. Mixture-o

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

How do LLMs Compute Verbal Confidence

DGX agent

arXiv:2603.17839v3 Announce Type: replace-cross Abstract: Verbal confidence -- prompting LLMs to state their confidence as a number or category -- is widely used to extract uncertainty estimates from

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

How Faithful Is Trajectory-Based Data Attribution? Error Sources, Remedies, and Practical Guidelines

DGX agent

arXiv:2605.18814v1 Announce Type: new Abstract: Trajectory-based data attribution methods estimate the influence of training samples on model predictions by unrolling the training trajectory. They are

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

How Far Are We From True Auto-Research?

DGX agent

arXiv:2605.19156v1 Announce Type: new Abstract: Recent auto-research systems can produce complete papers, but feasibility is not the same as quality, and the field still lacks a systematic study of ho

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

DGX agent

arXiv:2605.18822v1 Announce Type: cross Abstract: Post-training has become essential for adapting large language models (LLMs) to complex downstream behaviors, including instruction following, prefere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

DGX agent

arXiv:2605.19301v1 Announce Type: new Abstract: Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastroph

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis

DGX agent

arXiv:2605.19346v1 Announce Type: cross Abstract: We present IMLJD, an open dataset of 3,613 Indian court judgments covering matrimonial disputes under IPC Section 498A, the Protection of Women from D

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

In-Context Learning Operates as Concept Subspace Learning

DGX agent

arXiv:2605.18830v1 Announce Type: new Abstract: Regression and Bayesian accounts of in-context learning (ICL) explain how demonstrations can induce predictors, while mechanistic analyses often identif

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Information Processing Capacity of Stationary Physical Systems: Theory, Data-efficient Estimation Methods, and Photonic Demonstration

DGX agent

arXiv:2605.19152v1 Announce Type: cross Abstract: Physical computing systems provide a promising route toward hardware-native machine learning, but their computational capabilities remain difficult to

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

INSHAPE: Instance-Level Shapelets for Interpretable Time-Series Classification

DGX agent

arXiv:2605.20088v1 Announce Type: cross Abstract: Discovering shapelets -- i.e., discriminative temporal patterns within time series -- has been widely studied to address the inherent complexity of ti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

K-Quantization and its Impact on Output Performance

DGX agent

arXiv:2605.19645v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have shown their remarkable capacities in many NLP tasks. However, their substantial size often pres

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

KappaPlace: Learning Hyperspherical Uncertainty for Visual Place Recognition via Prototype-Anchored Supervision

DGX agent

arXiv:2605.19435v1 Announce Type: cross Abstract: Visual Place Recognition (VPR) is critical for autonomous navigation, yet state-of-the-art methods lack well-calibrated uncertainty estimation. Standa

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiency

DGX agent

arXiv:2605.19008v1 Announce Type: new Abstract: Modern language-model training is increasingly exposed to instability, degraded runs, and wasted compute, especially under aggressive learning-rate, sca

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learning-Accelerated Optimization-based Trajectory Planning for Cooperative Aerial-Ground Handover Missions

DGX agent

arXiv:2605.19562v1 Announce Type: cross Abstract: This paper presents a learning-augmented trajectory planning framework for cooperative unmanned aerial vehicle (UAV) and unmanned ground vehicle (UGV)

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Learning Efficient Guardrails for Compliance

DGX agent

arXiv:2510.03485v2 Announce Type: replace Abstract: Autonomous web agents are increasingly deployed for long-horizon tasks, yet their ability to adhere to real-world policies remains critically undere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learning When to Adapt

DGX agent

arXiv:2605.19028v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) is a widely used parameter-efficient fine-tuning method, yet its learned correction is static: the same low-rank update is ap

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition

DGX agent

arXiv:2605.19578v1 Announce Type: cross Abstract: RGB camera-based surveillance systems enable human action recognition for public safety and healthcare, yet raise serious privacy concerns. Existing m

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Less Back-and-Forth: A Comparative Study of Structured Prompting

DGX agent

arXiv:2605.20149v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompts can lead to low-quality answers and additional interacti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

DGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

LIFT and PLACE: A Simple, Stable, and Effective Knowledge Distillation Framework for Lightweight Diffusion Models

DGX agent

arXiv:2605.19729v1 Announce Type: cross Abstract: We demonstrate that in knowledge distillation for diffusion models, the teacher network's highly complex denoising process - stemming from its substan

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Lightweight and Fast Backdoor Model Detection

DGX agent

arXiv:2605.18907v1 Announce Type: cross Abstract: Deep neural networks (DNN), despite their remarkable performance, are highly vulnerable to backdoor attacks. Existing defenses mainly rely on activati

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

LLM Benchmark Datasets Should Be Contamination-Resistant

DGX agent

arXiv:2605.19999v1 Announce Type: cross Abstract: Benchmark datasets are critical for reproducible, reliable, and discriminative evaluation of LLMs. However, recent studies reveal that many benchmark

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening

DGX agent

arXiv:2605.19597v1 Announce Type: new Abstract: Evaluating large language models (LLMs) on natural-language logical reasoning is essential because rule-governed tasks require conclusions to follow str

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

LMM-Track4D: Eliciting 4D Dynamic Reasoning in LMMs via Trajectory-Grounded Dialogue

DGX agent

arXiv:2605.19390v1 Announce Type: new Abstract: Recent large multimodal models (LMMs) have become increasingly capable on image and video understanding, yet still struggle to sustain 4D continuous spa

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling

DGX agent

arXiv:2605.18838v1 Announce Type: cross Abstract: Scaling laws predict loss from compute but not how capabilities interact. We measure the coupling between reasoning and truthfulness across 63 base mo

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection

DGX agent

arXiv:2411.08982v3 Announce Type: replace Abstract: Selective parameter activation provided by Mixture-of-Expert (MoE) models have made them a popular choice in modern foundational models. However, Mo

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

m3BERT: A Modern, Multi-lingual, Matryoshka Bidirectional Encoder

DGX agent

arXiv:2605.19568v1 Announce Type: new Abstract: Embedding models are pivotal in industrial information retrieval systems like search and advertising. However, existing pretrained models often exhibit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification

DGX agent

arXiv:2605.19359v1 Announce Type: new Abstract: Deep learning methods have demonstrated promising results in predicting BI-RADS scores from mammography images. However, the interpretation of these ima

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning

DGX agent

arXiv:2605.19080v1 Announce Type: cross Abstract: In Online Continual Learning (OCL), a neural network sequentially learns from a non-stationary data stream in a single-pass with access only to a limi

model-releasesarxiv-cs-ai
20 May 2026
← Previous
1…218219220221222…361
Next →