AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

SMA: Submodular Modality Aligner For Data Efficient Multimodal Learning

DGX agent

arXiv:2605.12872v1 Announce Type: new Abstract: Despite the recent success of Multimodal Foundation Models (FMs), their reliance on massive paired datasets limits their applicability in low-data and r

model-releasesarxiv-cs-lg
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SpatialReward: Bridging the Perception Gap in Online RL for Image Editing via Explicit Spatial Reasoning

DGX agent

arXiv:2602.07458v4 Announce Type: replace Abstract: Online Reinforcement Learning (RL) offers a promising avenue for complex image editing but is currently constrained by the scarcity of reliable and

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

State-Space NTK Collapse Near Bifurcations

DGX agent

arXiv:2605.12763v1 Announce Type: new Abstract: Rich feature learning in tasks that unfold over time often requires the model to pass through bifurcations, constituting qualitative changes in the unde

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism

DGX agent

arXiv:2605.12524v1 Announce Type: cross Abstract: We introduce ProofGrid, a benchmark suite for evaluating LLM reasoning through machine-checkable proofs rather than final answers alone. ProofGrid con

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SynCABEL: Synthetic Contextualized Augmentation for Biomedical Entity Linking

DGX agent

arXiv:2601.19667v2 Announce Type: replace-cross Abstract: We present SynCABEL (Synthetic Contextualized Augmentation for Biomedical Entity Linking), a framework that addresses a central bottleneck in

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm

DGX agent

arXiv:2507.18553v4 Announce Type: replace Abstract: Quantizing the weights of large language models (LLMs) from 16-bit to lower bitwidth is the de facto approach to deploy massive transformers onto mo

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

TokaMind for Power Grid: Cross-Domain Transfer from Fusion Plasma

DGX agent

arXiv:2605.11033v1 Announce Type: cross Abstract: TokaMind is a multi-modal transformer (MMT) foundation model pre-trained on tokamak plasma diagnostics data from MAST, where it was shown to outperfor

model-releasesarxiv-cs-ai
14 May 2026
Agents

TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints

DGX agent

arXiv:2605.13414v1 Announce Type: new Abstract: Deploying language models as autonomous agents requires more than per-task accuracy: when an agent faces a queue of problems under a finite token budget

agentsarxiv-cs-ai
14 May 2026
Model Releases

Unlocking Patch-Level Features for CLIP-Based Class-Incremental Learning

DGX agent

arXiv:2605.13835v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) enables models to continuously integrate new knowledge while mitigating catastrophic forgetting. Driven by the remarkab

model-releasesarxiv-cs-cv
14 May 2026
Research

WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data

DGX agent

arXiv:2605.13846v1 Announce Type: cross Abstract: This paper introduces WARDEN, an early language model system capable of transcribing and translating Wardaman, an endangered Australian indigenous lan

researcharxiv-cs-ai
14 May 2026
Research

Weakly Supervised Segmentation as Semantic-Based Regularization

DGX agent

arXiv:2605.13674v1 Announce Type: cross Abstract: Weakly supervised semantic segmentation (WSSS) trains dense pixel-level segmentation models from partial or coarse annotations such as bounding boxes,

researcharxiv-cs-ai
14 May 2026
Model Releases

Ada-MK: Adaptive MegaKernel Optimization via Automated DAG-based Search for LLM Inference

DGX agent

arXiv:2605.11581v1 Announce Type: new Abstract: When large language models (LLMs) serve real-time inference in commercial online advertising systems, end-to-end latency must be strictly bounded to the

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Approximation Theory of Laplacian-Based Neural Operators for Reaction-Diffusion System

DGX agent

arXiv:2605.12025v1 Announce Type: new Abstract: Neural operators provide a framework for learning solution operators of partial differential equations (PDEs), enabling efficient surrogate modeling for

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

BEExformer: A Fast Inferencing Binarized Transformer with Early Exits

DGX agent

arXiv:2412.05225v3 Announce Type: replace Abstract: Large Language Models (LLMs) based on transformers achieve cutting-edge results on a variety of applications. However, their enormous size and proce

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images

DGX agent

arXiv:2605.12413v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) show strong visual perception, yet remain limited in reasoning about space under changing viewpoints. We study

model-releasesarxiv-cs-cv
13 May 2026
Research

BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking

DGX agent

arXiv:2602.00767v2 Announce Type: replace Abstract: Emergent misalignment can arise when a language model is fine-tuned on a narrowly scoped supervised objective: the model learns the target behavior,

researcharxiv-cs-lg
13 May 2026
Safety

BSO: Safety Alignment Is Density Ratio Matching

DGX agent

arXiv:2605.12339v1 Announce Type: new Abstract: Aligning language models for both helpfulness and safety typically requires complex pipelines-separate reward and cost models, online reinforcement lear

safetyarxiv-cs-lg
13 May 2026
Safety

CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography

DGX agent

arXiv:2605.11304v1 Announce Type: new Abstract: Chest radiograph interpretation requires temporal reasoning over prior and current studies, yet most vision-language models are trained on static image-

safetyarxiv-cs-cv
13 May 2026
Model Releases

Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification

DGX agent

arXiv:2603.28488v2 Announce Type: replace Abstract: Large language models (LLMs) remain unreliable for high-stakes claim verification due to hallucinations and shallow reasoning. While retrieval-augme

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Covering Human Action Space for Computer Use: Data Synthesis and Benchmark

DGX agent

arXiv:2605.12501v1 Announce Type: new Abstract: Computer-use agents (CUAs) automate on-screen work, as illustrated by GPT-5.4 and Claude. Yet their reliability on complex, low-frequency interactions i

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

CTFusion: A CTF-based Benchmark for LLM Agent Evaluation

DGX agent

arXiv:2605.11504v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have enabled agentic systems for complex, multi-step tasks; cybersecurity is emerging as a prominent app

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Deploying Self-Supervised Learning for Real Seismic Data Denoising

DGX agent

arXiv:2605.11109v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as a promising approach to seismic data denoising as it does not require clean reference data. In this work

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Detecting Data Contamination in LLMs via In-Context Learning

DGX agent

arXiv:2510.27055v2 Announce Type: replace Abstract: We present Contamination Detection via Context (CoDeC), a practical and accurate method to detect and quantify training data contamination in large

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

DiffScore: Text Evaluation Beyond Autoregressive Likelihood

DGX agent

arXiv:2605.11601v1 Announce Type: new Abstract: Autoregressive language models are widely used for text evaluation, however, their left-to-right factorization introduces positional bias, i.e., early t

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

DisagMoE: Computation-Communication overlapped MoE Training via Disaggregated AF-Pipe Parallelism

DGX agent

arXiv:2605.11005v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures enable trillion-parameter LLMs with sparsely activated experts. Expert parallelism (EP) is a widely adopted MoE t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness

DGX agent

arXiv:2605.11019v1 Announce Type: new Abstract: Although large language models rely on chain-of-thought for complex reasoning, the overthinking phenomenon severely degrades inference efficiency. Exist

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Exact Stiefel Optimization for Probabilistic PLS: Closed-Form Updates, Error Bounds, and Calibrated Uncertainty

DGX agent

arXiv:2605.11607v1 Announce Type: cross Abstract: Probabilistic partial least squares (PPLS) is a central likelihood-based model for two-view learning when one needs both interpretable latent factors

model-releasesarxiv-cs-lg
13 May 2026
Safety

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation

DGX agent

arXiv:2605.11613v1 Announce Type: new Abstract: On-policy self-distillation has emerged as a promising paradigm for post-training language models, in which the model conditions on environment feedback

safetyarxiv-cs-lg
13 May 2026
Model Releases

GeoR-Bench: Evaluating Geoscience Visual Reasoning

DGX agent

arXiv:2605.11541v1 Announce Type: new Abstract: Geoscience intelligence is expected to understand, reason about, and predict earth system changes to support human decision-making in critical domains s

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization

DGX agent

arXiv:2605.12369v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models aim for general robot learning by aligning action as a modality within powerful Vision-Language Models (VLMs). Exist

tutorialsarxiv-cs-ro
13 May 2026
Model Releases

HE-SNR: Uncovering Latent Logic via Entropy for Guiding Mid-Training on SWE-bench

DGX agent

arXiv:2601.20255v2 Announce Type: replace-cross Abstract: SWE-bench has emerged as the premier benchmark for evaluating Large Language Models on complex software engineering tasks. While these capabil

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Human-Grounded Multimodal Benchmark with 900K-Scale Aggregated Student Response Distributions from Japan's National Assessment of Academic Ability

DGX agent

arXiv:2605.11663v1 Announce Type: new Abstract: Authentic school examinations provide a high-validity test bed for evaluating multimodal large language models (MLLMs), yet benchmarks grounded in Japan

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference

DGX agent

arXiv:2505.13770v3 Announce Type: replace-cross Abstract: Reliable causal inference is essential for making decisions in high-stakes areas like medicine, economics, and public policy. However, it rema

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

DGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

model-releasesarxiv-cs-cl
13 May 2026
Safety

Metaphor Is Not All Attention Needs

DGX agent

arXiv:2605.12128v1 Announce Type: new Abstract: Large language models are increasingly deployed in safety-critical applications, where their ability to resist harmful instructions is essential. Althou

safetyarxiv-cs-cl
13 May 2026
Model Releases

MuonQ: Enhancing Low-Bit Muon Quantization via Directional Fidelity Optimization

DGX agent

arXiv:2605.11396v1 Announce Type: new Abstract: The Muon optimizer has emerged as a compelling alternative to Adam for training large language models, achieving remarkable computational savings throug

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Predicting Psychological Well-Being from Spontaneous Speech using LLMs

DGX agent

arXiv:2605.11303v1 Announce Type: new Abstract: We investigate the use of Large Language Models (LLMs) for zero-shot prediction of Ryff Psychological Well-Being (PWB) scores from spontaneous speech. U

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Reconsidering the energy efficiency of spiking neural networks

DGX agent

arXiv:2409.08290v4 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their eve

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Reviving In-domain Fine-tuning Methods for Source-Free Cross-domain Few-shot Learning

DGX agent

arXiv:2605.11659v1 Announce Type: new Abstract: Cross-Domain Few-Shot Learning (CDFSL) aims to adapt large-scale pretrained models to specialized target domains with limited samples, yet the few-shot

model-releasesarxiv-cs-cv
13 May 2026
Safety

Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter

DGX agent

arXiv:2605.11685v1 Announce Type: new Abstract: Large language model (LLM) unlearning aims to remove specific data influences from pre-trained model without costly retraining, addressing privacy, copy

safetyarxiv-cs-cl
13 May 2026
Model Releases

Robust Promptable Video Object Segmentation

DGX agent

arXiv:2605.12006v1 Announce Type: new Abstract: The performance of promptable video object segmentation (PVOS) models substantially degrades under input corruptions, which prevents PVOS deployment in

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems

DGX agent

arXiv:2605.11800v1 Announce Type: cross Abstract: Large language models (LLMs) with mixture-of-experts (MoE) architectures achieve remarkable scalability by sparsely activating a subset of experts per

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Slicing and Dicing: Configuring Optimal Mixtures of Experts

DGX agent

arXiv:2605.11689v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures have become standard in large language models, yet many of their core design choices - expert count, granularit

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

DGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

model-releasesarxiv-cs-lg
13 May 2026
Hardware

To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation

DGX agent

arXiv:2412.14461v4 Announce Type: replace Abstract: Unstructured text data annotation is foundational to management research. LLMs offer a cost-effective and scalable alternative to human annotation,

hardwarearxiv-cs-cl
13 May 2026
Safety

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching

DGX agent

arXiv:2605.12288v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used RL-free method for aligning language models from pairwise preferences, but it models preferences o

safetyarxiv-cs-cl
13 May 2026
Safety

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

DGX agent

arXiv:2605.11974v1 Announce Type: new Abstract: Large Language Models (LLMs) suffer from order bias, where their performance is affected by the arrangement order of input elements. This unfairness lim

safetyarxiv-cs-lg
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
← Previous
1…370371372373374…1074
Next →