AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

Parameter-Efficient Adaptation of SAM3 for Prompt-Driven Surgical Concept Segmentation

DGX agent

arXiv:2607.23694v1 Announce Type: new Abstract: Efficient surgical segmentation empowers clinical diagnosis, intraoperative monitoring, and downstream robotic pipelines for reconstruction and simulati

model-releasesarxiv-cs-cv
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation

DGX agent

arXiv:2607.22588v1 Announce Type: new Abstract: Modern compute-intensive software must migrate across a changing ecosystem of accelerators, programming APIs, compiler stacks, and portability layers, i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Phenology-based learning framework for yield estimation and harvest forecasting of raspberry fruits

DGX agent

arXiv:2411.00967v2 Announce Type: replace Abstract: The future of agriculture is intertwined with automation. Accurate fruit detection, yield estimation, and harvest time prediction are crucial for ef

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Poison to Detect: Detection of Targeted Overfitting in Federated Learning

DGX agent

arXiv:2509.11974v3 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training among clients without centralising data, making it a widely adopted privacy-enhan

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios

DGX agent

arXiv:2607.23088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely used for code generation, yet their security behavior in realistic development workflows remains underexplored

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Random Forest-Based Prediction of Bone Volume Fraction and Fracture Position from S-Parameters

DGX agent

arXiv:2607.23563v1 Announce Type: new Abstract: In this paper, we propose a method for predicting bone volume fraction (BVF) and fracture position by constructing a random forest model based on multic

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy

DGX agent

arXiv:2607.22554v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong accuracy on benchmarks, yet it remains unclear how reliably they apply this knowledge when the same qu

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

DGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Source-Free Controlled Adaptation of Teachers for Continual Test-Time Adaptation

DGX agent

arXiv:2607.23735v1 Announce Type: cross Abstract: In many real-world scenarios, encountering continual shifts in domain during inference is very common. Consequently, continual test-time adaptation (C

model-releasesarxiv-cs-cv
28 Jul 2026
Hardware

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

DGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

hardwarearxiv-cs-ai
28 Jul 2026
Model Releases

TextRich: A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2

DGX agent

arXiv:2606.19259v2 Announce Type: replace-cross Abstract: Text-rich images often contain privacy-sensitive, transactional, or decision-relevant information. As recent multimodal image generation model

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

TokenMem: Faithful Knowledge Injection for Frozen LLMs

DGX agent

arXiv:2607.22625v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) with external knowledge, but suffers from knowledge conflicts: when retrieved

model-releasesarxiv-cs-ai
28 Jul 2026
Research

VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference

DGX agent

arXiv:2512.01031v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) are becoming increasingly capable across diverse robotic tasks. However, these models are typically deplo

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Weighted Low-Rank Matrix Approximation: Acceleration and Applications

DGX agent

arXiv:2109.11057v2 Announce Type: replace-cross Abstract: Weighted low-rank matrix approximation (WLRMA) generalizes classical low-rank approximation and matrix completion by allowing arbitrary elemen

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

XGRVFL-MV: Residual-Coupled Graph-Embedded Multi-View Random Vector Functional Link Network with FleXi Guardian Loss

DGX agent

arXiv:2607.23149v1 Announce Type: new Abstract: Random Vector Functional Link (RVFL) networks provide an efficient randomized learning framework for classification. Existing multi-view RVFL methods ut

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

DGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Do emulated quantum circuits change what CNNs look at? Performance and explainability comparison in medical image classification

DGX agent

arXiv:2607.21186v1 Announce Type: cross Abstract: Numerous studies have analyzed the use of hybrid quantum-classical convolutional neural networks as a promising alternative to classical deep learning

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

DGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

DGX agent

arXiv:2607.22508v1 Announce Type: new Abstract: Electroencephalography (EEG) is widely used to diagnose neurological conditions, but its analysis usually relies on either predefined spectral features

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Learning What Matters: Supervising Sparse Attention Routing with Causal Evidence Sets

DGX agent

arXiv:2607.21692v1 Announce Type: cross Abstract: Sparse attention reduces the cost of long contexts by allowing each query to read only selected parts of the input. These selectors are often trained

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Multi-Horizon Consistency as Geometry: When Latent Dynamics Contract, and When They Do Not

DGX agent

arXiv:2607.21645v1 Announce Type: new Abstract: Multi-horizon latent consistency is a common training knob in video predictors and world models, but practitioners rarely know what it does to transitio

researcharxiv-cs-lg
27 Jul 2026
Model Releases

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures

DGX agent

arXiv:2607.21612v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like LoRA have become the default for adapting large language models, succeeding across instruction following,

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

SceneActBench: Can Agents Act on the 3D Scenes They See?

DGX agent

arXiv:2607.22393v1 Announce Type: cross Abstract: Vision-language model (VLM) agents increasingly use tools to act on 3D scenes rather than only describe them. Existing 3D benchmarks score textual res

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Spatially-Enhanced Temporal Fusion Transformer: Interpretable Multi-Output Prediction for Parametric Dynamical Systems with Time-Varying Inputs

DGX agent

arXiv:2505.00473v2 Announce Type: replace Abstract: We explore the promising performance of a transformer model in predicting outputs of parametric dynamical systems with external time-varying input s

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Variational Low-rank Tensor Decomposition for Multisubject Spatiotemporal Data Analysis

DGX agent

arXiv:2607.22262v1 Announce Type: cross Abstract: Modeling shared and subject-specific structure in multisubject spatiotemporal data remains challenging, particularly in neuroimaging, where both spati

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs

DGX agent

arXiv:2607.21042v1 Announce Type: new Abstract: Autoregressive text-to-speech models achieve strong naturalness but suffer from slow inference due to sequential token generation, limiting their deploy

model-releasesarxiv-cs-ai
24 Jul 2026
Tutorials

Gumbel Distillation for Parallel Text Generation

DGX agent

arXiv:2603.22216v2 Announce Type: replace Abstract: The slow, sequential nature of autoregressive (AR) language models has driven the adoption of parallel decoding methods. However, these non-AR model

tutorialsarxiv-cs-cl
24 Jul 2026
Research

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

DGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

researcharxiv-cs-ai
24 Jul 2026
Research

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

DGX agent

arXiv:2607.20426v1 Announce Type: cross Abstract: Existing LLM hallucination mitigation methods, including prompt engineering and model optimization, either hardly alter models'internal knowledge or h

researcharxiv-cs-ai
24 Jul 2026
Model Releases

ODeform: Learning Continuous 4D Motion for Shape Deformation with Neural ODEs

DGX agent

arXiv:2607.20670v1 Announce Type: new Abstract: Modeling continuous object deformation is important for many computer vision and robotics tasks, such as manipulation and simulation. Existing approache

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

OPOD: On-Policy Omni Distillation

DGX agent

arXiv:2607.20918v1 Announce Type: new Abstract: Omni-modal models can handle text, images, and audio in one system, but improving all of these abilities together remains difficult. Training a single m

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning

DGX agent

arXiv:2607.21022v1 Announce Type: new Abstract: Improving video captioning quality typically demands retraining large vision-language models, an expensive and often impractical requirement. Existing t

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Refusal-Gated Decoding: Preserving Refusal Behavior Under High-Temperature Sampling

DGX agent

arXiv:2607.20791v1 Announce Type: new Abstract: High-temperature sampling is one of the primary mechanisms for increasing diversity in LLMs. Recent advances in truncation-based sampling techniques hav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts

DGX agent

arXiv:2607.09999v2 Announce Type: replace Abstract: We show that post-training quantization can silently alter how large language models reason even when task accuracy is preserved. Using a six-catego

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion

DGX agent

arXiv:2607.20543v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study thi

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Adaptive Visual Autoregressive Acceleration via Dual-Linkage Entropy Analysis

DGX agent

arXiv:2602.01345v2 Announce Type: replace Abstract: Visual AutoRegressive modeling (VAR) suffers from substantial computational cost due to the massive token count involved. Failing to account for the

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Benchmarking Confidential GPU Inference on NVIDIA H100 under Intel TDX

DGX agent

arXiv:2607.19353v1 Announce Type: new Abstract: Confidential computing is becoming a practical deployment requirement for AI inference workloads that process sensitive inputs or protect proprietary mo

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

CEO-Bench: Can Agents Play the Long Game?

DGX agent

arXiv:2606.18543v2 Announce Type: replace Abstract: Language model agents are becoming proficient executors at isolated, short-horizon tasks such as software engineering and customer service. Yet real

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Continual Video-MLLM Adaptation over Evolving Domains

DGX agent

arXiv:2607.18716v1 Announce Type: new Abstract: Video multimodal large language models have shown strong capability in video understanding, yet their adaptation to sequentially evolving domains remain

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

ECoNGS: Efficient Compressive Neural Gaussian Splats for Volume Visualization

DGX agent

arXiv:2607.18466v1 Announce Type: new Abstract: Recent advances in differentiable Gaussian splatting have highlighted the potential of primitive-based approaches as alternative scene representations f

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

ExpertVerse: A General-Purpose Benchmark for Expert-Level Reasoning in Knowledge-Intensive Visual Synthesis

DGX agent

arXiv:2607.19341v1 Announce Type: new Abstract: Recent advances in multimodal generative models have enabled instruction-based image generation to move beyond semantic manipulation to knowledge-driven

model-releasesarxiv-cs-cv
23 Jul 2026
Applications

FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads

DGX agent

arXiv:2607.19349v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as always-on online services, making efficient LLM serving a critical systems challenge. Achievin

applicationsarxiv-cs-ai
23 Jul 2026
Model Releases

HalluTruthQA: A Fine-Grained Benchmark for Hallucination Detection, Localization, and Explanation in Arabic Question Answering

DGX agent

arXiv:2607.20219v1 Announce Type: new Abstract: Large language models (LLMs) can generate fluent Arabic answers, yet factual errors remain difficult to detect, localize, explain, and verify. Existing

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Multi-stage Dynamic Selection for Cross-Project Defect Prediction

DGX agent

arXiv:2607.20151v1 Announce Type: cross Abstract: Cross-Project Defect Prediction (CPDP) involves building models using data from external projects, called training projects, to predict modules from t

model-releasesarxiv-cs-lg
23 Jul 2026
Research

Pixel-Space Diffusion Transformers

DGX agent

arXiv:2607.17585v2 Announce Type: replace Abstract: Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual

researcharxiv-cs-cv
23 Jul 2026
Model Releases

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

DGX agent

arXiv:2607.19385v1 Announce Type: new Abstract: This paper tackles the problem of stock ranking and portfolio construction under realistic investment settings by jointly modeling temporal dynamics and

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability

DGX agent

arXiv:2607.20301v1 Announce Type: cross Abstract: Fine-tuning has been widely used to adapt large language models (LLMs) for domain-specific tasks. Parameter efficient fine-tuning (PEFT) methods such

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

When Does Knowledge Distillation Hurt? Reliability-Aware Distillation for Low-Resource Language Summarization

DGX agent

arXiv:2607.19956v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a standard approach for compressing sequence-to-sequence models, but its per-sample effects are rarely examined. On the

model-releasesarxiv-cs-ai
23 Jul 2026
← Previous
1…306307308309310…1058
Next →